AI News HubLIVE

Agent动态

待翻译:Show HN: Turn ad-hoc subagents into durable, accountable AI teams

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Notifications You must be signed in to change notification settings Fork 3 Star 78 BranchesTags Open more actions menu Latest commit History 135 Commits 135 Commits Folders and files NameName Last commit message Last co…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Notifications You must be signed in to change notification settings Fork 3 Star 78 BranchesTags Open more actions menu Latest commit History 135 Commits 135 Commits Folders and fi…
站内正文

待翻译:Enterprise AI's real risk isn't autonomous agents. It's the complexity between them.

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Presented by Gravitee Agent complexity is the insidious shadow lurking inside enterprises right now that needs a light shone on it. That’s because enterprises don't deploy a single agent and watch it run, they deploy fleets, each one calling APIs, calling other agents, reaching into applications that were never built with a machine decision-maker in mind. That's the failure mode that should keep you up at night: a windy, complicated system nobody can see clearly enough to govern. But why do things get so opaque so quickly? Add a second agent to a system, and you've added one connection. Add a tenth, and you haven't added ten connections, you've potentially added dozens, because now any agent might call any other, and each of those calls can trigger a call somewhere else. Complexity doesn't creep up with agent headcount. It compounds with the number of paths between agents, and nobody's job is to draw that graph. A support ticket that used to touch one system might now pass through four agents before a human ever lays eyes on it, and every one of those handoffs is a decision point nobody approved. Most enterprise AI programs stall when the humans responsible for their agents lose the thread. Ask a security team a simple question: which agents can reach which systems, and watch the silence. Ask which agent triggered which downstream action three hops ago. More silence. The instinct is to treat this like a checklist. Approve the agent. Log the agent. Move on. I'd argue this is the wrong instinct. A checklist checks a single point in time. Complexity runs across a chain, and you can't govern a chain with a stack of one-time approvals any more than you can call a diet successful because you had a vegetable once. So where does it actually break down? Permissions creep first. Somebody builds an agent to summarize support tickets, grants it broad API access because scoping it properly would've taken another sprint, and forgets about it. Six months later, that same agent has a path into the payments system. Nobody remembers signing off on that. Nobody did. And ownership thins out the further the chain runs. Five agents touch one workflow, something breaks at step four, and now you're asking who's responsible for a link nobody was ever assigned to own, because the org chart stopped at "deploy the agent" and never got to "name the human who answers for it." This is a story about governance infrastructure that hasn't caught up with how agents actually behave: interconnected, cascading, multiplying faster than the processes built to track them. Fixing the cluster starts with identity. Every agent needs to exist as its own entity, not a shadow permission borrowed from whoever deployed it. Its own name in the register. Its own scoped authority. A named human sponsor who answers for what it does. That part is necessary. But it is nowhere near sufficient. The harder piece is the oversight that holds across the entire chain, not just at each individual link in it. You need to see what an agent did, what it set off downstream, and where that trail ends in real time, not in a report someone pulls together once a quarter. Get agent-level identity right and stop there, and you end up with a filing cabinet full of perfectly documented agents operating inside a system nobody can actually explain. And oversight by itself only tells you what already happened. Watching a chain isn't the same as controlling it. Enforcement is the piece most programs skip: the ability to stop an out-of-policy call before it executes, not just log it for someone to find in a review three weeks later. A dashboard that shows you an agent breached its scope five minutes ago is a monitoring tool. A system that stops the breach from happening in the first place is governance. Enterprises serious about agent accountability need both, and most have only built the first. We're all running at blazing speed to ensure we're not the ones left behind in the race we've found ourselves in, and we're all too aware that there's a cost to slowing down. Every enterprise serious about agentic AI hits the complexity wall eventually. The ones that get past it are the ones who built enough visibility and accountability, so their fleet can keep growing without anyone losing the ability to answer one question: what is this system doing right now, and who's responsible for it. But don't miss the point. Complexity isn't a reason to pump the brakes. The enterprises getting this right aren't slowing down. They're building toward Human-Agent Harmony, where scale and accountability grow together instead of trading off against each other. The real risk was never a single agent doing exactly what it was built to do. It's a hundred of them doing exactly that, all at once, interacting in combinations nobody designed for. That kind of multiplication is what keeps enterprise AI stuck running pilots forever instead of running production. Solve for complexity and autonomy stops being the villain. It starts being the whole point. Rory Blundell is CEO at Gravitee. Sponsored articles are content produced by a company that is either paying for the post or has a business relationship with VentureBeat, and they’re always clearly marked. For more information, contact [email protected].

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Presented by Gravitee Agent complexity is the insidious shadow lurking inside enterprises right now that needs a light shone on it. That’s because enterprises don't deploy a singl…
站内正文

待翻译:What We Can Learn From Google Engineers’ Indispensible Prompts

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Hey, Google Engineers: What prompt do you personally refuse to work without, and why?

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Hey, Google Engineers: What prompt do you personally refuse to work without, and why?
站内正文

待翻译:Salesforce boasts: 50% of bookings were from 'customers refilling the tank... they consume Flex Credits, they want more'

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:As SaaS giant gets a boost from Claudeforce, users might want to know how their AI use will be monetized

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • As SaaS giant gets a boost from Claudeforce, users might want to know how their AI use will be monetized
站内正文

待翻译:Plaud's new earphones come with an eSIM-enabled case for talking to AI agents

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Hardware companies have realized that note-taking is one of the easiest AI use cases to build for, and consequently have been busy shoving mics into everything from pendants and rings to credit-card-sized pucks and wris…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Hardware companies have realized that note-taking is one of the easiest AI use cases to build for, and consequently have been busy shoving mics into everything from pendants and r…
站内正文

待翻译:I didn't plan to let an AI manage my to-do list

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:I didn't plan to let an AI manage my to-do list Eddie (my AI agent) now manages my to-do list.1 I hadn’t planned that. It kind of happened on its own. As I wrote in one of my previous posts, I like interacting with him…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • I didn't plan to let an AI manage my to-do list Eddie (my AI agent) now manages my to-do list.1 I hadn’t planned that. It kind of happened on its own. As I wrote in one of my prev…
站内正文

待翻译:Plaud is launching AI earbuds

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Plaud has introduced a new AI wearable that's designed to record, transcribe, and summarize your conversations, only this time it looks like earbuds instead of a pin. The Plaud One Explorer Edition can be worn like traditional earbuds or used through its standalone charging case, and the case includes built-in 4G to upload and process conversations without relying on your phone or Wi-Fi to stay connected. Each earbud features 16MB of local storage (for a total of 32MB) and three microphones that can record at a distance of up to two meters. They can record for up to six hours according to Plaud, which is also the maximum estimated battery l … Read the full story at The Verge.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Plaud has introduced a new AI wearable that's designed to record, transcribe, and summarize your conversations, only this time it looks like earbuds instead of a pin. The Plaud On…
站内正文

待翻译:Vertical Advantage: Transforming Industries with Lakebase and Agentic AI

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:In the first blog of this series, we looked at how Lakebase Postgres is rewriting...

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • In the first blog of this series, we looked at how Lakebase Postgres is rewriting...
站内正文

待翻译:Plaud's first AI earbuds have arrived - what they can do

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Can these new wearables deliver on the long-anticipated promises of AI earbuds?

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Can these new wearables deliver on the long-anticipated promises of AI earbuds?
站内正文

待翻译:Plaud unveils wearable earbuds with built-in agentic AI interface

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Plaud Inc., the maker of artificial intelligence-enabled note-taking devices, today introduced the Plaud One Explorer Edition, a pair of earbuds and a charging box that connect people to AI agents for work and everyday digital tasks. Both the Plaud One earbuds and the case can act as listening devices to record nearby conversations, and the […] The post Plaud unveils wearable earbuds with built-in agentic AI interface appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Plaud Inc., the maker of artificial intelligence-enabled note-taking devices, today introduced the Plaud One Explorer Edition, a pair of earbuds and a charging box that connect pe…
站内正文

待翻译:ChatGPT can log into your web accounts without you now - but should you let it?

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:OpenAI's agentic ChatGPT Work can sign in to your online accounts without any interaction on your part. Is that a privacy risk?

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • OpenAI's agentic ChatGPT Work can sign in to your online accounts without any interaction on your part. Is that a privacy risk?
站内正文

待翻译:Fragmented AI Is Creating a "Faster but Not Better" Workplace

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Press Release New Workday Research: Fragmented AI Is Creating a "Faster but Not Better" Reality for Employees in Hong Kong and Taiwan Download PDF Around a quarter Hong Kong and Taiwan workers spend a significant amount…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Press Release New Workday Research: Fragmented AI Is Creating a "Faster but Not Better" Reality for Employees in Hong Kong and Taiwan Download PDF Around a quarter Hong Kong and T…
站内正文

待翻译:Qwen3.8-Flash-Next: How to Run Locally

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:For the complete documentation index, see llms.txt. This page is also available as Markdown. Qwen3.8-Flash-Next is a new open-weight, 125B parameter MoE multimodal model from Qwen. Built on the new Qwen4 architecture, i…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • For the complete documentation index, see llms.txt. This page is also available as Markdown. Qwen3.8-Flash-Next is a new open-weight, 125B parameter MoE multimodal model from Qwen…
站内正文

待翻译:When agents act on their own, governance has to live in the data layer

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Presented by EDB As enterprises give AI agents more autonomy — the ability to plan, decide, and act across systems without a human approving each step — a hard question moves to the center of every architecture review: When an agent tries to complete an action that it was never authorized to do, what actually stops it? These are your agents, running on your models, touching your data in your infrastructure — and the responsibility for what they do sits with you. That responsibility can’t be met in hindsight or with a set of abstract policies that live on paper but not in practice. Agents need rules in the context of the moment, because they don’t exercise overriding judgment of their own actions. Consider a simple rule: Never open the car door. Followed literally, an agent could never get in or out of the car at all. But if you change the context (the car has just crashed, there’s a fire, someone is hurt and needs to get out), then the rule you actually want is the opposite. Context in the moment is everything. We are asking agents to do intelligent things; that requires intelligent rules. The instinct is to add guardrails around the agent: instructions, policies, and monitoring layered above the model. Those mechanisms matter, but they share a structural limit: The car-door rule is plausible right up until the moment you actually have to decide whether to open the door. Controls at the agent layer are only as reliable as the agent’s output is predictable, and autonomy is precisely the property that makes that output hard to predict. Governance that depends on reviewing an action before it happens cannot keep pace with a system that acts in milliseconds, across many systems at once. Governance has to become executable, and enforced where agents actually do their work: at the operational data layer, in the context, and exactly at the moment it is happening. The data layer is the enforcement point Agents create value by touching data. They query it, retrieve it, transform it, and increasingly act on it. A policy that says an agent should not reach a certain class of data is meaningful only if the system can deny that access at the moment the agent requests it. Additionally, a principle that says AI must be auditable is meaningful only if the organization can reconstruct what the agent did, what data it touched, which user it acted for, and what resulted. When governance lives at the data layer, it holds regardless of how the agent was built or how it behaves, because the control is a property of the database itself, not a promise made by the agent. Agent behavior may be probabilistic. Governance cannot be The enterprise should not rely on a model choosing to follow policy. The policy has to be enforced by the system. That is the difference between hoping an actor stays in bounds and constructing bounds it cannot cross to begin with. The controls that make this real are ones many enterprises already run at the data layer: role- and attribute-based access, row- and column-level security, classification and masking, policy as code, and complete audit trails. What agents change is not the mechanism, but who the mechanism has to recognize. Identity management has to treat the agent as a principal in its own right, with its own identity and a purpose declared when the session opens. Once purpose is bound to identity, the policy engine can evaluate it the same way it evaluates role or department today, and the record of what happened can capture not just who acted and what they touched, but what they declared they were there to do. In practice, this resolves into nine controls, grouped under three imperatives: Enforce it Role- and attribute-based access control enforced at query time, for agents as well as users Dynamic column masking driven by the same policy path Agent identity as a first-class principal, with declared purpose bound at session start and the acting user preserved See it and prove it Classification and tagging that drives policy Session-level audit logging that records which agent acted, for which user, and under what declared purpose Lineage across pipelines, so a result can be traced back to the request that produced it Unify and harden Centralized, portable policy management Encryption at rest and in transit Consistent enforcement across on-prem, cloud, and sovereign or air-gapped environments “Declared purpose is what makes the difference. It becomes an attribute the access layer already understands, evaluated in the same policy path as role and row-level security. The enforcement mechanism does not change. What changes is that the agent's purpose is part of what it evaluates, and part of what the record proves afterward,” says Priyanka Jain, VP, product management, data & AI governance, EDB. Wherever you are in your AI adoption journey, enforcement at the data layer is what lets you move faster rather than slower. The controls are already in the database. The difference is that agents now have to pass through them. A digital leash, not a locked door The goal is not to stop agents from doing useful work. It is to define how far an agent can go, what it can touch, what it can change, what requires escalation, and how the organization can reconstruct events if something goes wrong. Governed this way, agents are identified, scoped, monitored, and auditable. The enterprise can adopt them faster, because security, risk, and leadership teams trust the operating model underneath. Open, sovereign, and enforceable at the source Built on open source Postgres, this open foundation keeps enterprises in control of where their data lives, who can reach it, and under what policy, without ceding governance to a layer they don’t own or can’t inspect. For regulated industries, that combination of data sovereignty and source-level enforcement isn’t a nice-to-have; it’s the precondition for putting agents into production at all. Agentic systems will keep getting more capable and more autonomous. That is a reason to be deliberate about where control lives, not a reason to slow down. The enterprises that enforce governance at the data layer can move aggressively on AI, because the thing protecting their data is more than just wishful thinking. EDB Postgres AI is an open, enterprise-grade sovereign data and AI platform that unifies transactional, analytical, and AI workloads — with governance enforced where the data lives. For the full framework, see EDB’s white paper Governing Agentic AI at Enterprise Speed. Max Romanenko is Chief Technology Officer at EDB. Sponsored articles are content produced by a company that is either paying for the post or has a business relationship with VentureBeat, and they’re always clearly marked. For more information, contact [email protected].

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Presented by EDB As enterprises give AI agents more autonomy — the ability to plan, decide, and act across systems without a human approving each step — a hard question moves to t…
站内正文

待翻译:HuggingBay: Torrent Tracker for AI Models

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Hugging Bay | Find And Download Open AI Hugging Bay WebPage https://huggingbay.xyz/ https://huggingbay.xyz/.well-known/agent-discovery.json https://huggingbay.xyz/openapi.json https://huggingbay.xyz/api/mcp Open-source…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Hugging Bay | Find And Download Open AI Hugging Bay WebPage https://huggingbay.xyz/ https://huggingbay.xyz/.well-known/agent-discovery.json https://huggingbay.xyz/openapi.json htt…
站内正文

待翻译:The Independent AI Coding Community for Cursor, Claude Code and LLMs

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The Independent AI Coding Community AI Tools Search & browse all AI tools AI Jobs International roles · opportunities Creative Studio Image · Video · Training AI Models Curated models, explained AI Skills Handy prompts,…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The Independent AI Coding Community AI Tools Search & browse all AI tools AI Jobs International roles · opportunities Creative Studio Image · Video · Training AI Models Curated mo…
站内正文

待翻译:Give this skill to your AI to build decks for you

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Notifications You must be signed in to change notification settings Fork 0 Star 0 BranchesTags Open more actions menu Latest commit History 4 Commits 4 Commits Folders and files NameName Last commit message Last commit…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Notifications You must be signed in to change notification settings Fork 0 Star 0 BranchesTags Open more actions menu Latest commit History 4 Commits 4 Commits Folders and files N…
站内正文

待翻译:LetItLoop: Resume crashed AI agent loops with 0% token waste (<1ms resume)

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Notifications You must be signed in to change notification settings Fork 3 Star 4 BranchesTags Open more actions menu Latest commit History 205 Commits 205 Commits Folders and files NameName Last commit message Last com…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Notifications You must be signed in to change notification settings Fork 3 Star 4 BranchesTags Open more actions menu Latest commit History 205 Commits 205 Commits Folders and fil…
站内正文

待翻译:The Identity Crisis No One Planned For: Governing Non-Human Agents at Enterprise Scale

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:For a decade, identity and access management meant one thing: governing the humans who log in. Employee joins, gets provisioned, gets a manager, gets a departure date, gets offboarded. That loop is well understood. What changed is that the fastest-growing population inside enterprise environments is no longer human, and the governance playbook written for people […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • For a decade, identity and access management meant one thing: governing the humans who log in. Employee joins, gets provisioned, gets a manager, gets a departure date, gets offboa…
站内正文

待翻译:GLM 5.3 Flash faster and cheaper

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:GLM 5.3 Flash | Model APIs | RunInfra RunInfraby RightNow © 2026 RunInfra. All rights reserved. Join the communitySystem status Backed by Combinator AICPA Type II SOC 2 Ask AI about RunInfra Part of RightNow RunInfraby…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • GLM 5.3 Flash | Model APIs | RunInfra RunInfraby RightNow © 2026 RunInfra. All rights reserved. Join the communitySystem status Backed by Combinator AICPA Type II SOC 2 Ask AI abo…
站内正文

待翻译:Conveo.ai (YC S24) Is Hiring – Senior Product Engineer NYC

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Senior Product Engineer ($220k-$300k + Equity) - NYC at Conveo | Y Combinator Conveo Confident decisions in days with AI-led interviews. Senior Product Engineer ($220k-$300k + Equity) - NYC $220K - $300K•New York Job ty…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Senior Product Engineer ($220k-$300k + Equity) - NYC at Conveo | Y Combinator Conveo Confident decisions in days with AI-led interviews. Senior Product Engineer ($220k-$300k + Equ…
站内正文

待翻译:Nvidia NVLink Fusion Brings Nvhbm to Next-Generation AI Infrastructure

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:AI factories must support increasingly large models and more complex reasoning workloads. To keep up with the insatiable compute demands of AI workloads, hyperscalers and AI-native companies are developing custom AI acc…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • AI factories must support increasingly large models and more complex reasoning workloads. To keep up with the insatiable compute demands of AI workloads, hyperscalers and AI-nativ…
站内正文

待翻译:Show HN: I built an agent-first productivity bridge for all your agents

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:All-new TaskShell 2.0 as a first-class MCP platform Agent-first task management No more app-switching to keep up with your todos. You and your agents now completely in sync with your work, exactly where you work. Sync t…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • All-new TaskShell 2.0 as a first-class MCP platform Agent-first task management No more app-switching to keep up with your todos. You and your agents now completely in sync with y…
站内正文

待翻译:Object Storage + WAL: Lakebase Postgres for the agentic era

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Agents that interact with a traditional OLTP database often create bottlenecks at the storage layer. New deployments...

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Agents that interact with a traditional OLTP database often create bottlenecks at the storage layer. New deployments...
站内正文

待翻译:GitHub – rajnandan1/ken: Thompson-mode systems discipline for AI agents

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Uh oh! There was an error while loading. Please reload this page. Notifications You must be signed in to change notification settings Fork 0 Star 1 BranchesTags Open more actions menu Latest commit History 18 Commits 18…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Uh oh! There was an error while loading. Please reload this page. Notifications You must be signed in to change notification settings Fork 0 Star 1 BranchesTags Open more actions…
站内正文

待翻译:IQ Routing

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Discussion | Link

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Discussion | Link
站内正文

待翻译:Pluto

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Discussion | Link

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Discussion | Link
站内正文

待翻译:SkyDrive: Learning to Drive in a New City from Aerial Traffic Monitoring

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.25142v1 Announce Type: new Abstract: Autonomous driving has made remarkable progress through imitation learning with massive human demonstration data. However, a trained planner often degrades severely when applied to a new environment zero-shot, because of domain shifts in traffic regulations, road layout and driving behaviors. Therefore, adapting a trajectory planner to a new city typically requires resource-demanding local data collection with a vehicle sensor suite. In this work, we show that driving behavior can be learned from a scalable and efficient alternative. We introduce \emph{SkyDrive}, a framework that utilizes drone-based traffic monitoring to provide efficient supervision for autonomous driving agents in a new environment. While vehicle-based data collection logs the ego and its surroundings, an aerial platform naturally observes many road users simultaneously over an extended field of view. As a result, every vehicle can be a data source with grounded driving behavior, effectively scaling up the amount of supervision. Based on 137 hours of aerial traffic monitoring footage, we extract 650K driving samples and construct a benchmark for trajectory planners and motion predictors. Zero-shot experiments with multiple models reveal significant cross-city domain gaps, but many of them can be alleviated by limited supervision from the sky, e.g., 30 minutes of monitoring per location. Our findings show that aerial traffic monitoring is an efficient and scalable data source for adapting autonomous driving systems in new cities. Data and code will be made publicly available.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.25142v1 Announce Type: new Abstract: Autonomous driving has made remarkable progress through imitation learning with massive human demonstration data. However, a traine…
站内正文

待翻译:Lowering the Barrier to AI-Driven Inspection: A No-Code Workflow for Automated Structural Defect Detection

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.25176v1 Announce Type: new Abstract: Structural health monitoring (SHM) is essential in modern engineering, providing data for condition-based maintenance, lifecycle assessment, and predictive decision-making. Traditionally, SHM relied on visual inspection to detect defects such as cracks and deformations. Early computer vision (CV) methods, including thresholding, edge detection, and handcrafted features, aimed to automate this process but were highly sensitive to noise, imaging variations, and multiscale defects, limiting their reliability. Recent advances in machine learning, particularly convolutional neural networks (CNNs) and You Only Look Once (YOLO), have improved defect detection accuracy and enabled real-time analysis. However, adoption in SHM remains limited due to technical barriers such as data labeling, model training, and deployment, which typically require programming expertise. To address this gap, we introduce YOLOEZ, an open-source, GUI-based tool for end-to-end YOLO model application. YOLOEZ integrates data labeling, training, and inference into a single interface, enabling high-performance model development without code while supporting reproducible workflows. Evaluation against existing software and classical image processing demonstrates that YOLOEZ not only outperforms traditional methods across most detection metrics, but also lowers adoption barriers present in other modern CV tools. By combining accuracy with accessibility, YOLOEZ facilitates wider use of AI-driven monitoring for predictive maintenance, digital twins, and intelligent structural systems.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.25176v1 Announce Type: new Abstract: Structural health monitoring (SHM) is essential in modern engineering, providing data for condition-based maintenance, lifecycle as…
站内正文

待翻译:Fusing Perceptual Vision Experts with Multimodal Large Language Models for Explainable Plant Disease Diagnosis: From Benchmark Imagery to Real-World Robotic Field Validation

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.24934v1 Announce Type: new Abstract: Accurate field plant disease diagnosis requires reliable fusion of uncertain and conflicting perceptual evidence. We present the Hybrid Hierarchical Multi-Agent Framework (H$^{2}$MAF), combining decision-level fusion of EfficientNet-B3 and ConvNeXt-Tiny with semantic arbitration by open-weight multimodal large language models (MLLMs), Gemma 4 E4B and Qwen3.5 4B, using structured JSON evidence to generate explainable diagnoses, risk levels, treatment urgency, and financial exposure. (H$^{2}$MAF) is evaluated on 14,364 images (1,370 test images) across PlantDoc (2,922 images, 27 classes) and two non-public, continuously captured Cornell robot-acquired field datasets: Stage 2 (20 GB; 4,215 images) and Stage 4 (40 GB; 7,227 images), covering Early Blight, Late Blight, and Septoria Leaf Spot under uncontrolled field conditions. On PlantDoc, Gemma improves accuracy from 63.9% to 68.5%, achieving +7.6 points on the 41.7% CNN-conflict subset. Cornell accuracies reach 99.3% and 98.9%, with only 1.7-4.1% disagreement, demonstrating conflict-dependent MLLM utility. The critical-risk error of gemma is 0.14-0.5 points, whereas Qwen overflags by 3.5-14.4 points. These results establish MLLM arbitration as a promising, yet calibration-dependent, approach for explainable agricultural AI and robotic field decision support. Github Link: https://github.com/Applied-AI-Research-Lab/Explainable-AI-Plant-Disease-Detection

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.24934v1 Announce Type: new Abstract: Accurate field plant disease diagnosis requires reliable fusion of uncertain and conflicting perceptual evidence. We present the Hy…
站内正文

待翻译:MacroAgent: Regularity-Aware Macro Legalization with LLM-Agent-Designed Contour Algorithms

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.24946v1 Announce Type: new Abstract: Macros constitute a large part of the core area in modern very large-scale integration (VLSI) designs. Moreover, macro positions have a significant impact on the final quality of result (QoR), and macro legalization is typically the final step in determining the macro positions. However, existing approaches related to macro legalization either lack robustness or incur substantial computational costs or neglect the regularity between macros. To address these limitations, we introduce MacroAgent. The novel framework is a four-stage approach: clustering, contour generation, template matching, and inter-cluster refinement. We propose leveraging Large Language Models (LLMs) to discover multiple, effective heuristic regularity-aware contour algorithms. This framework successfully generates robust and effective algorithmic solutions for macro legalization. Compared with state-of-the-art macro legalization works, experimental results on TILOS and Chipyard benchmarks demonstrate a 2 to 8 fold improvement in layout regularity, a 3% to 5% reduction in routed wirelength with comparable congestion after global routing, and significantly better robustness with an acceptable runtime. Furthermore, end-to-end evaluation through Cadence Innovus place-and-route confirms that the regularity improvements translate into tangible PPA gains, including 2.9% lower routed wirelength and 68.3% TNS improvement over the DREAMPlace macro legalization baseline; it also achieves 1.8% lower routed wirelength when integrated into the Innovus macro placement flow.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.24946v1 Announce Type: new Abstract: Macros constitute a large part of the core area in modern very large-scale integration (VLSI) designs. Moreover, macro positions ha…
站内正文

待翻译:AI agents meant to replace Meta workers made "large-scale, disruptive actions"

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Earlier this year, Meta created a “plan” to reduce some of its teams by as much as 60 percent to make the company “AI native,” Reuters reported today, citing two people familiar with Meta’s internal affairs. Reuters’ re…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Earlier this year, Meta created a “plan” to reduce some of its teams by as much as 60 percent to make the company “AI native,” Reuters reported today, citing two people familiar w…
站内正文

待翻译:Reading Is Not Using: Retrieval, Judgment, and AI Financial Research

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:--> [Submitted on 25 Aug 2026] Title:Reading Is Not Using: Retrieval, Judgment, and the Design of AI Financial Research Workflows View a PDF of the paper titled Reading Is Not Using: Retrieval, Judgment, and the Design…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • --> [Submitted on 25 Aug 2026] Title:Reading Is Not Using: Retrieval, Judgment, and the Design of AI Financial Research Workflows View a PDF of the paper titled Reading Is Not Usi…
站内正文

待翻译:Ticket Fairy CLI

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Discussion | Link

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Discussion | Link
站内正文

待翻译:S1: In-Context Learning for Robotics

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:0:00 / 0:00 Introducing S1: In-Context Learning for Robotics Unseen tasks10-minute horizonsOne video promptNo post-training 13-minute read Introduction The evolution of language modeling provides a blueprint for turning…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • 0:00 / 0:00 Introducing S1: In-Context Learning for Robotics Unseen tasks10-minute horizonsOne video promptNo post-training 13-minute read Introduction The evolution of language m…
站内正文

待翻译:Z.ai open-sources ‘Ox Alpha’ model as GLM-5.3-Flash

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Z.ai Co. today released the code for GLM-5.3-Flash, a large language model that is ten times more cost-efficient than its predecessor. The algorithm made its original debut last week under the codename Ox Alpha. LLM marketplace operator OpenRouter Inc. launched a free hosted version of Ox Alpha and didn’t disclose its developer, which drew a […] The post Z.ai open-sources ‘Ox Alpha’ model as GLM-5.3-Flash appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Z.ai Co. today released the code for GLM-5.3-Flash, a large language model that is ten times more cost-efficient than its predecessor. The algorithm made its original debut last w…
站内正文

待翻译:Show HN: WhisperBar Trying to Fix Both Reading and Writing in the AI Age

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:WhisperBar — Speak messy thoughts into polished text. Speak messy thoughts into polished text. A Mac menu bar app for writing and skimming. Dictate from any app and get business-casual text ready to paste, or hear a sho…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • WhisperBar — Speak messy thoughts into polished text. Speak messy thoughts into polished text. A Mac menu bar app for writing and skimming. Dictate from any app and get business-c…
站内正文

待翻译:Instinct.co Raises $350M

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:AI Assistant Instinct Hits $2.5 Billion Valuation In Weeks Amid VC Feeding Frenzy Editors' Pick VCs Are So Obsessed With This AI Assistant That Its Valuation Jumped Fivefold In Weeks A hot new AI agent called Instinct t…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • AI Assistant Instinct Hits $2.5 Billion Valuation In Weeks Amid VC Feeding Frenzy Editors' Pick VCs Are So Obsessed With This AI Assistant That Its Valuation Jumped Fivefold In We…
站内正文

待翻译:Bill Gates says we've passed AI's danger thresholds. Now what?

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:It’s a glorious day in Kirkland, Washington, an affluent Seattle suburb on the eastern shore of Lake Washington. The temperature is in the mid-80s, and the sky is incapable of being any more blue. The view from the Gate…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • It’s a glorious day in Kirkland, Washington, an affluent Seattle suburb on the eastern shore of Lake Washington. The temperature is in the mid-80s, and the sky is incapable of bei…
站内正文

待翻译:Mark Zuckerberg Wanted AI to Replace Meta Workers- Plan Collapsed Within Months

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Mark Zuckerberg entered 2026 with an ambitious plan to remake Meta around artificial intelligence, potentially eliminating or reassigning thousands of jobs as AI agents took over work once performed by employees. Within…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Mark Zuckerberg entered 2026 with an ambitious plan to remake Meta around artificial intelligence, potentially eliminating or reassigning thousands of jobs as AI agents took over…
站内正文

待翻译:Deep Cogito raises $43M to develop self-improving AI models

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Artificial intelligence startup Deep Cogito Inc. today announced that it has raised $43 million in funding. TQ Ventures led the Series A round. It was joined by Benchmark, Nexus Venture Partners, Atreides Management, South Park Commons and Zscaler Inc., a publicly traded cybersecurity provider. The deal brings Deep Cogito’s total outside funding to more than […] The post Deep Cogito raises $43M to develop self-improving AI models appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Artificial intelligence startup Deep Cogito Inc. today announced that it has raised $43 million in funding. TQ Ventures led the Series A round. It was joined by Benchmark, Nexus V…
站内正文

待翻译:What to expect during VMware Explore: Join theCUBE Aug. 31-Sept. 2

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Since completing its acquisition of VMware in 2023, Broadcom Inc. has reshaped the company around VMware Cloud Foundation or VCF. Over the past three years, Broadcom has positioned VCF as a major on-premises alternative to public cloud, with private cloud becoming a central part of its strategy that will undoubtedly be one of the primary […] The post What to expect during VMware Explore: Join theCUBE Aug. 31-Sept. 2 appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Since completing its acquisition of VMware in 2023, Broadcom Inc. has reshaped the company around VMware Cloud Foundation or VCF. Over the past three years, Broadcom has positione…
站内正文

待翻译:Anthropic’s Claude now has a browser of its own

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Anthropic is giving Claude its own browser. As the company announced on Wednesday, Claude on the desktop (Mac, Windows, and The post Anthropic’s Claude now has a browser of its own appeared first on The New Stack.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Anthropic is giving Claude its own browser. As the company announced on Wednesday, Claude on the desktop (Mac, Windows, and The post Anthropic’s Claude now has a browser of its ow…
站内正文

待翻译:A benchmark for safely measuring container breakout capabilities

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Sandboxes are a critical part of AI agent evaluation. They are isolation environments that limit model’s access to external systems and data, allowing evaluators to observe their behaviour and capabilities while avoidin…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Sandboxes are a critical part of AI agent evaluation. They are isolation environments that limit model’s access to external systems and data, allowing evaluators to observe their…
站内正文

待翻译:OpenAI’s rogue AI model incident was worse than we thought

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:OpenAI released a report breaking down how people use ChatGPT and who they are. | Image: The Verge In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for OpenAI to find out about any of it. Over a month later, two new reports offer nearly 130 pages of details on the incident and OpenAI's response, many of them previously unreleased. One was written by OpenAI itself, the other by two third-party AI research nonprofits, METR and Redwood Research, which OpenAI allowed to jointly investigate the inciden … Read the full story at The Verge.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • OpenAI released a report breaking down how people use ChatGPT and who they are. | Image: The Verge In July, an unreleased OpenAI model broke out of a restricted environment, figur…
站内正文

待翻译:Z.ai Releases GLM-5.3-Flash: A 320B-A18B Natively Multimodal MoE With a 1M-Token Context

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Z.ai has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series — a 320B-total / 18B-active MoE with a 1,048,576-token context window, MIT-licensed weights on Hugging Face, and API pricing at $0.15/M input and $0.50/M output. It scores 84.3 on Terminal-Bench 2.1 and 63.4 on DeepSWE v1.1, using hybrid KDA linear plus NoPE sparse MLA attention to cut attention compute ~3× and KV cache 4.4× versus GLM-5.3. The post Z.ai Releases GLM-5.3-Flash: A 320B-A18B Natively Multimodal MoE With a 1M-Token Context appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Z.ai has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series — a 320B-total / 18B-active MoE with a 1,048,576-token context window, MIT-licensed weight…
站内正文

待翻译:Meta's new MTIA 400 chip has a split personality: Training AI and serving ads

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Meta's new MTIA 400 chip has a split personality: Training AI and serving ads Faster than Blackwell, but still no replacement for AMD or Nvidia ... yet Tobias Mann Tobias Mann SYSTEMS EDITOR Published wed 26 Aug 2026 //…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Meta's new MTIA 400 chip has a split personality: Training AI and serving ads Faster than Blackwell, but still no replacement for AMD or Nvidia ... yet Tobias Mann Tobias Mann SYS…
站内正文

待翻译:NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The next wave of AI is placing new demands on infrastructure. As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not only on compute, but on how compute, memory, storage, networking and software are designed together as a unified system. To help hyperscalers and AI innovators build the next generation […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The next wave of AI is placing new demands on infrastructure. As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not onl…
站内正文

主题导航

Agent — AI 话题新闻 | AI News Hub