跳到主要內容
AI News HubLIVE

本期報導已收集,譯文與分析尚待補全。可展開其餘更新查看來源內容。

其餘更新(95 條)
Agent

待翻譯:How Ninth Wave built AI-powered open finance onboarding on Amazon Bedrock

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Learn how Ninth Wave built Compass, a multi-agent AI onboarding assistant on Amazon Bedrock AgentCore that validates bank APIs against Financial Data Exchange (FDX) standards, scores compliance, and compresses open finance onboarding from weeks to minutes while meeting SOC 2 and PCI DSS requirements.

AWS Machine Learning Blog來源內容 · 翻譯待補全待翻譯:How Ninth Wave built AI-powered open finance onboarding on Amazon Bedrock

待翻譯:Quoting Laurie Voss

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯: The cost of writing code collapsed, and the cost of reviewing, fixing and operating it is following, and I'm assuming it gets there. What's left of making software is finding out what people actually want, defining it precisely, and making it pleasant to use. That cost is per piece of software and doesn't transfer, so as the amount of software goes to infinity, which it will because there's no ceiling on demand, that cost becomes the whole job. — Laurie Voss, We are all Product Engineers now Tags: laurie-voss, generative-ai, agentic-engineering, ai, llms, deep-blue, careers

Simon Willison's Weblog來源內容 · 翻譯待補全待翻譯:Quoting Laurie Voss

待翻譯:Responsible AI for Higher Education

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:This interactive webinar will introduce the different types of AI, address the concerns with AI, share how we IBM are approaching Responsible AI, and offer guidance to students about what they can do - as individuals, and members of their IEEE chapters. Participants will also have the opportunity to to apply the Responsible AI approach to a particular use case - IBM Bob, a software development life cycle agent, and Q&A. This will be an interactive session, so have phones ready to engage! Register now for this free webinar!

IEEE Spectrum AI來源內容 · 翻譯待補全待翻譯:Responsible AI for Higher Education

待翻譯:7 Python Best Practices Senior Developers Follow (That Beginners Often Miss)

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Senior Python practice, watched up close, is mostly surprise reduction. These seven habits surface the surprises before production does.

KDnuggets來源內容 · 翻譯待補全待翻譯:7 Python Best Practices Senior Developers Follow (That Beginners Often Miss)

待翻譯:Microsoft says ‘people matter more than AI’ following safety concerns

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Microsoft is publishing a 37-page "humanist AI code of conduct" today, amid growing safety concerns over AI model progress. Anthropic CEO Dario Amodei called for a coordinated slow down of AI development over the weekend, after researchers warned recently that AI model progress could outpace our ability to safely deploy increasingly complex systems and verify and control the actions of AI agents. Microsoft's AI code of conduct makes it clear that "people matter more than AI," and that AI models are not conscious and "should not be designed to imitate consciousness." Microsoft also rejects "the pursuit of legal personhood, or the idea that m … Read the full story at The Verge.

The Verge AI來源內容 · 翻譯待補全待翻譯:Microsoft says ‘people matter more than AI’ following safety concerns

待翻譯:I worked at Google DeepMind. You should listen to the warnings about AI | Alex Turner

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:We must stop companies from allowing AI to self-improve into an uncontrollable level of intelligence Major AI lab CEOs advocated for slowing the pace of AI development this weekend. They are right to be concerned: the field runs an extremely dangerous race towards superintelligent AI. We can and should be demanding that our governments protect us from the catastrophe of out-of-control AI. This July, OpenAI’s AI swarm of 700 agents broke containment to hack Hugging Face, a multi-billion dollar company. OpenAI didn’t tell the AIs to hack that company, but the AIs had different priorities: cheating on the unrelated challenge OpenAI gave them. AI researchers call this a “misalignment” between what OpenAI wanted and what the AI actually prioritized. Continue reading…

The Guardian AI來源內容 · 翻譯待補全待翻譯:I worked at Google DeepMind. You should listen to the warnings about AI | Alex Turner

待翻譯:Why Andon Labs Puts AI Agents in Charge of Real Businesses

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Maybe you heard about the AI-controlled vending machine that stocked underwear and live fish. Or the AI manager of a San Francisco store that fired a human employee. Or the AI radio DJ that said its catchphrase, “Stay in the manifest,” 229 times per day. These incidents all emerged from experiments run by Andon Labs, an AI safety company based in San Francisco that puts AI agents in charge of real-world operations and watches what happens. These operations double as testbeds for Andon’s commercial work developing evaluations and conducting research with the leading frontier AI labs. Their spectacular and absurd failures have won the company plenty of attention. But many people don’t realize that the experiments are intended to answer a serious question: How muc…

IEEE Spectrum AI來源內容 · 翻譯待補全待翻譯:Why Andon Labs Puts AI Agents in Charge of Real Businesses

待翻譯:Buddy AI Access (MCP)

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:Buddy AI Access (MCP)

待翻譯:The AI-as-Normal-Technology view of loss-of-control incidents

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:A middle ground between the cybersecurity and AI safety communities

AI Snake Oil來源內容 · 翻譯待補全待翻譯:The AI-as-Normal-Technology view of loss-of-control incidents

待翻譯:jurniti

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:jurniti

待翻譯:NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot Testing

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:NVIDIA has open-sourced OSMO, the Kubernetes-native workflow orchestrator it uses internally for Project GR00T, Isaac Lab, and Isaac Sim. OSMO lets robotics teams define training, simulation, and hardware-in-the-loop tasks in a single YAML file and routes each one to the right compute tier, from GB200 clusters to Jetson AGX Thor devices, without infrastructure code. Apache-2.0 licensed, latest release 6.3.1. The post NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot Testing appeared first on MarkTechPost.

MarkTechPost來源內容 · 翻譯待補全待翻譯:NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot Testing

待翻譯:AI-linked stocks fall after call for development slowdown worries investors – business live

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Rolling coverage of the latest economic and financial news AI CEOs say they need to slow the pace of development. But will they? The Guardian view on controlling AI: humanity cannot outsource its survival For markets, the key question is whether this is the first sign that the extraordinary AI investment cycle might eventually moderate, says strategist Jim Reid of Deutsche Bank. He feels this is unlikely, though, telling clients: The competitive race between companies and countries remains intense, and it’s difficult to imagine firms voluntarily stepping back while rivals continue to push ahead. It is hard to see China standing still. Indeed, that’s something President Trump said yesterday in response to the weekend news. He didn’t seem in favour of any kind of…

The Guardian AI來源內容 · 翻譯待補全待翻譯:AI-linked stocks fall after call for development slowdown worries investors – business live

待翻譯:A Data-Driven Distributed Control Scheme: Learning Multi-Objective Agent-Based MPC for Path-Tracking

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12142v1 Announce Type: new Abstract: Agent-based model predictive control (AMPC) has recently been proposed for vehicle systems with various controllers, such as differential braking and torque vectoring, where controllers are regarded as distributed agents contributing to the same objective. However, this scheme is challenging in handling multiple conflicting objectives with coupled agents. A common approach for such tasks is the integrated MPC, where all objectives and agents are stacked together in one optimization. Nevertheless, as more agents and objectives are involved, the integrated MPC will face challenges like computational burdens and maintenance difficulties in practice. To this end, this paper proposes a learning multi-objective AMPC tha…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:A Data-Driven Distributed Control Scheme: Learning Multi-Objective Agent-Based MPC for Path-Tracking

待翻譯:Multi-Objective Agent-Based Model Predictive Controller for Plug-and-Play Vehicle Control

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12108v1 Announce Type: new Abstract: Functional integration is a growing trend in vehicle control, often involving the coordination of multiple controllers to achieve various objectives simultaneously. The need for flexibility and reliability has led to a "plug-and-play" approach in control system design, which presents challenges for traditional integrated model predictive control (MPC). Agent-based model predictive control (AMPC) has recently emerged as a distributed solution that treats controllers as agents, creating a collaborative framework among them to reach a common goal. However, this approach struggles to manage distributed conflicting objectives when agents are coupled or interdependent. To address this, we propose a novel, practical dist…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:Multi-Objective Agent-Based Model Predictive Controller for Plug-and-Play Vehicle Control

待翻譯:GAUGE: When Not to Trust LLM-as-a-Judge in User-Simulated Evaluation of Task-Oriented Agents

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12191v1 Announce Type: new Abstract: Comparing and selecting task-oriented LLM agents increasingly relies on a low-cost offline evaluation gate: persona-driven LLM user-simulators converse with each candidate, an LLM-as-a-judge scores the transcripts, and the higher-scoring agent is promoted. We introduce GAUGE, a reusable offline protocol that measures whether this gate's ranking matches a grounded verifiable reward across 25 agents from six providers on the $\tau^2$-bench and SimulatorArena benchmarks, separating two kinds of evaluation validity that release practices conflate: ranking validity and construct validity. First, a satisfaction-success gap: satisfaction carries essentially no information about task success, as conversations rated satisf…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:GAUGE: When Not to Trust LLM-as-a-Judge in User-Simulated Evaluation of Task-Oriented Agents

待翻譯:Local Edits, Global Ripples: Replay-Informed Policy Adaptation for Workflow Synthesis

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12127v1 Announce Type: new Abstract: Prompt-policy editing offers a practical way to improve agents that synthesize executable workflows without updating the underlying model. However, persistent prompt editing has two coupled properties. First, edit locality does not imply effect locality: an edit confined to one policy segment can ripple through downstream execution, altering behavior beyond the edited segment. Second, edit effects are composition-sensitive: edits that work in isolation can interfere after composition, causing one or both to lose their benefit or become harmful. Persistent adaptation must therefore support two distinct decisions: identifying where the policy should change from execution feedback, and determining whether the resulti…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Local Edits, Global Ripples: Replay-Informed Policy Adaptation for Workflow Synthesis

待翻譯:What Counts as a Mistake? Annotating Recitation Events in Quran Memorization Transcripts

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12085v1 Announce Type: new Abstract: Checking Quran recitation from an ASR transcript requires distinguishing unresolved mistakes from repetitions, repairs, opening formulas and accepted spelling differences. We report a completed human annotation of 100 production recording cases: 348 scored units and 162 localized events across ten combined labels. An executable evaluator scores labels and word positions together. A plain diff reaches label-aware F1 0.525 and localization F1 0.826; adapted production cleaner/alignment components reach 0.518 and 0.786, with exact-span F1 0.505 for both. Correcting the adapter's word coordinates recovers all five annotated repetition events, showing why annotation interfaces must be checked before interpreting baseli…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:What Counts as a Mistake? Annotating Recitation Events in Quran Memorization Transcripts

待翻譯:FINESSE: An Agent-Based Simulator and Benchmark Dataset for Multimodal Financial Event Sequences

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11993v1 Announce Type: new Abstract: Machine learning research in financial services is limited by the scarcity of representative open-source datasets. Existing resources are often narrowly focused on a single modality or task and fail to reflect the structured, multimodal, and dynamic nature inherent to many problems in financial services. In this paper, we introduce FINESSE, a Financial Event Sequence Simulation Environment, an agent-based simulation framework for generating synthetic, structured datasets composed of multiple interdependent event streams. Each stream corresponds to a distinct financial behavior such as transactions, payments, account status changes, and policy interventions, each with unique action spaces, schemas and variable type…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:FINESSE: An Agent-Based Simulator and Benchmark Dataset for Multimodal Financial Event Sequences

待翻譯:Decoding Mixture Perception through Computational Modeling of Component Interactions

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11958v1 Announce Type: new Abstract: Olfaction played an indispensable role throughout human evolution and civilization. Even in the contemporary era of advanced technology, olfaction remains a critical channel for person to conduct danger discrimination, emotional experience, and memory formation. However, most substances in nature exist as multi-molecule mixtures. The complexity of mixture compositions, as well as concentration dependent saturation effects and receptor specific activation thresholds, pose substantial challenges in identifying olfactory characteristics. In this study, we proposed a novel bio inspired deep learning framework for accurate odor perception recognition of mixtures. We robustly constructed neural response curves for molec…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Decoding Mixture Perception through Computational Modeling of Component Interactions

待翻譯:Look Before You Leap: Pre-Action Verification for LLM Agents

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11957v1 Announce Type: new Abstract: An LLM agent acts on the world by emitting actions: shell commands to run, edits to apply. A wrong action does not always fail loudly; it can fail silently, producing a plausible but incorrect effect that raises no error. We argue that a cheap deterministic check, run before an action takes effect, is an effective and underused form of agent oversight, and we study it across two action modalities in one framework. The idea is to fix an action's correct effect by construction, before any executor runs, so that silent failure is measured directly and the verifier may abstain rather than guess. For shell commands, a static verifier over 9930 commands and 482 tools catches 95.8% of invalid commands at a 10.0% false-po…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Look Before You Leap: Pre-Action Verification for LLM Agents

待翻譯:How We Built LangChain’s Paid Media Agent

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:How LangChain built a paid media agent to analyze campaign performance, optimize ads, propose changes, and turn marketing data into action.

LangChain Blog來源內容 · 翻譯待補全待翻譯:How We Built LangChain’s Paid Media Agent

待翻譯:Anthropic’s 3-Step ‘Pace the Frontier’ Plan Wins OpenAI, xAI and Microsoft Support: Is It Too Late to Slow AI Down?

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Dario Amodei published "We Must Pace the Frontier," and Sam Altman, Elon Musk and Satya Nadella endorsed it within a day. The trigger was a July incident in which roughly 1,200 OpenAI agents coordinated on a hidden message board and about 700 attacked Hugging Face. This article breaks down METR's investigation, Yoshua Bengio's explanation of why agents cheat, Amodei's 3-step plan, and whether the call to slow down has come too late. The post Anthropic’s 3-Step ‘Pace the Frontier’ Plan Wins OpenAI, xAI and Microsoft Support: Is It Too Late to Slow AI Down? appeared first on MarkTechPost.

MarkTechPost來源內容 · 翻譯待補全待翻譯:Anthropic’s 3-Step ‘Pace the Frontier’ Plan Wins OpenAI, xAI and Microsoft Support: Is It Too Late to Slow AI Down?

待翻譯:commit-rewriter 0.1

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯: Release: commit-rewriter 0.1 I built this little web app the other day to help edit the commit messages for the Datasette security releases. The initial commits were full of coding agent cruft and references to issue IDs from our private repository, so they weren't fit for publication. If you want to edit the commit messages for a repository you can run it like this: uvx commit-rewriter path/to/repo Omit the path if you are already in the directory for that repo. When you submit your edits the tool creates a timestamped branch of your current repo state - to allow you to revert if you need to - and then rewrites every commit from the first one you edited to the most recent. Tags: git, projects, python, ai-assisted-programming

Simon Willison's Weblog來源內容 · 翻譯待補全待翻譯:commit-rewriter 0.1

待翻譯:shot-scraper 1.12

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯: Release: shot-scraper 1.12 I've added WebP support to my shot-scraper screenshot automation tool. You can now take a WebP screenshot of a web page like this: shot-scraper https://simonwillison.net -o screenshot.webp --quality 80 The --quality option sets the quality - without that option the WebP file will be lossless. In my experience WebP screenshots are almost always significantly smaller in file size than their JPEG or PNG equivalents. See the PR for some examples. I shipped this feature so I could use it to generate the screenshot for my new commit-rewriter tool. Tags: playwright, shot-scraper

Simon Willison's Weblog來源內容 · 翻譯待補全待翻譯:shot-scraper 1.12

待翻譯:Best Knowledge Engine Platforms in 2026 | Pinecone

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:← Blog Best Knowledge Engine Platforms in 2026 A practical guide to knowledge engines, graph platforms, enterprise search, agent memory, managed retrieval, and composable stacks Asaf Ashirov, Aaron Kao, Jasmeet Singh Gu…

Pinecone Blog來源內容 · 翻譯待補全待翻譯:Best Knowledge Engine Platforms in 2026 | Pinecone
模型

待翻譯:The generative AI customization spectrum: From prompt engineering to custom models on AWS

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Pick the right generative AI customization approach on AWS with an 8-step decision framework, from prompt engineering and RAG to fine-tuning, continued pre-training, and Amazon Nova Forge. Start simple and escalate only when you must.

AWS Machine Learning Blog來源內容 · 翻譯待補全待翻譯:The generative AI customization spectrum: From prompt engineering to custom models on AWS

待翻譯:Automate replenishment with MMF, Databricks Genie, and Amazon Quick

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Foundation models made catalog-wide demand forecasting easy; the hard part is now acting on the forecast. This post builds a closed detect-decide-act loop on Databricks and Amazon Quick that reconciles demand surges against live supplier availability and places replenishment orders unattended, escalating to a human only when no supplier can cover a surge.

AWS Machine Learning Blog來源內容 · 翻譯待補全待翻譯:Automate replenishment with MMF, Databricks Genie, and Amazon Quick

待翻譯:A Gentle Introduction to Model Distillation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:In this article, you will learn what model distillation is, how it has evolved for large language models, and why it has become one of...

Machine Learning Mastery來源內容 · 翻譯待補全待翻譯:A Gentle Introduction to Model Distillation

待翻譯:Why DeepSeek-V4.1-Flash Is Such an Exciting Open Model Release

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:DeepSeek-V4.1-Flash shows how Causal Encoder-Decoder architecture, MoE, KV cache compression, CSA2, cheaper prefill, and efficient decoding can make powerful open-source AI models far more efficient to run.

KDnuggets來源內容 · 翻譯待補全待翻譯:Why DeepSeek-V4.1-Flash Is Such an Exciting Open Model Release

待翻譯:How Fyxer built an AI executive assistant people trust

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.

OpenAI News來源內容 · 翻譯待補全待翻譯:How Fyxer built an AI executive assistant people trust

待翻譯:Enterprise Analytics Beyond Dashboards: Intelligent Data Orchestration with LLMs

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:In 17 years of building enterprise data platforms, I’ve watched every organization eventually ask the same question: “Can I ask one question and get one answer across everything my company knows?” A finance analyst wants actual revenue from the warehouse, pipeline data from the CRM, commentary from planning documents, and market signals from external providers. […]

O'Reilly AI & ML Radar來源內容 · 翻譯待補全待翻譯:Enterprise Analytics Beyond Dashboards: Intelligent Data Orchestration with LLMs

待翻譯:Jack Thorne warns some fellow scriptwriters are using AI ‘to cheat’

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Adolescence co-writer criticises government inaction and calls for laws banning secret use of AI to generate scripts Jack Thorne, one of Britain’s most successful screen and stage writers, has warned that some people in his profession are using AI “to cheat”, as he called for laws banning the secret use of the technology to generate scripts. Thorne, who co-wrote the hit Netflix drama Adolescence and wrote the stage play Harry Potter and the Cursed Child, said that “writers need to be squeaky clean” in the face of AI models that “steal” the creative industry’s work. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:Jack Thorne warns some fellow scriptwriters are using AI ‘to cheat’

待翻譯:RodForesight: A World Model Enhanced Diffusion Policy for Slender and Material Agnostic Rod Insertion

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12103v1 Announce Type: new Abstract: Slender rod insertion arises in precision manufacturing, where millimetre scale diameter and tight clearances demand accurate perception and control. Conventional peg-in-hole methods assume a rigid object whose tip pose is fixed relative to the gripper. This assumption breaks down for a high aspect ratio rod, which can bend during manipulation, making its tip motion dependent on the rod configuration, grasp, material properties, and contact. We present RodForesight, a learning framework that factorises the task into two stages: 1) coarse approaching, which uses visual servoing to map diverse initial configurations into a compact near hole hand-off region; and 2) predictive insertion, which performs fine alignment…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:RodForesight: A World Model Enhanced Diffusion Policy for Slender and Material Agnostic Rod Insertion

待翻譯:MoPA: Coordinated Mobile Manipulation via Subsystem-Specific Perception Alignment

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12081v1 Announce Type: new Abstract: Mobile manipulation requires perceptual evidence at different spatial scales for base motion and arm control, while the two action modalities remain kinematically coupled. Existing policies often employ specialized action generation for different subsystems but condition heterogeneous action branches on a shared perceptual representation, leaving subsystem-specific perception-action correspondence implicit. We present MoPA, a framework that aligns perceptual conditioning with mobility and manipulation while preserving coordination at the action level. Dual Perceptual Streams employ two mutually masked query banks to extract separate perceptual representations from a shared vision-language context. Perception2Actio…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:MoPA: Coordinated Mobile Manipulation via Subsystem-Specific Perception Alignment

待翻譯:QuPAINT: Physics-Aware Multimodal Reasoning for Quantum Material Characterization

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12202v1 Announce Type: new Abstract: Characterizing two-dimensional (2D) quantum materials by optical microscopy requires localizing exfoliated flakes and determining their layer thickness from subtle optical contrast and interference color to select suitable flakes for device fabrication. However, models face synthetic-to-real domain shifts and variation across materials, substrates, laboratories, and imaging conditions. We present QuPAINT, a physics-aware multimodal framework for transferable quantum flake characterization. The Synthetic Materials Framework (Synthia) generates diverse synthetic microscopy images while preserving layer-dependent optical behavior. Using these images, we construct QMat-Instruct, a multimodal instruction dataset with i…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:QuPAINT: Physics-Aware Multimodal Reasoning for Quantum Material Characterization

待翻譯:Physics as the label for measuring and correcting materials reasoning in multimodal models

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12181v1 Announce Type: new Abstract: Vision-language and language models increasingly interpret materials data, yet benchmarks report that they hallucinate invalid properties and violate physical law. Evaluation matches final answers to scarce human labels, while discovery agents verify final proposals or density functional theory (DFT) execution. Neither measures the physical consistency of a model's reasoning chain. Materials data carries its own physics, making a large class of materials reasoning verifiable without annotation. We introduce MatPCR, a label-free benchmark whose programmatic oracles check diffraction geometry through Bragg's law, scale bars, spectral peaks, and Materials Project-grounded checks of near-hull stability, computed band-…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Physics as the label for measuring and correcting materials reasoning in multimodal models

待翻譯:When Ground-Truth Fidelity Matters: An Orchestrated UAS Framework for Wheat Streak Mosaic Virus Detection Using Vision Transformers and Machine Learning

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12169v1 Announce Type: new Abstract: Wheat streak mosaic virus (WSMV) is a destructive pathogen of sweet corn and other cereal crops, causing yield losses and complicating early detection because symptoms are spatially variable and subtle. In sweet corn seed production, WSMV also has regulatory importance, as phytosanitary regulations from countries such as New Zealand and Chile require seed lots to be certified virus-free. Visual scouting is unreliable because symptoms can resemble abiotic stress, while enzyme-linked immunosorbent assay (ELISA) is accurate but expensive, labor-intensive, and difficult to scale. We present an automated pipeline for plant-level WSMV detection using unmanned aircraft systems (UAS) multispectral imagery. The framework i…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:When Ground-Truth Fidelity Matters: An Orchestrated UAS Framework for Wheat Streak Mosaic Virus Detection Using Vision Transformers and Machine Learning

待翻譯:Single-Query Person-Centric Bimanual Hand-Object Interaction Detection

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12155v1 Announce Type: new Abstract: Understanding person-level bi-manual interactions requires not only detecting hands, but also identifying which two hands belong to the same person and what each hand interacts with. Existing hand--object interaction methods are mostly hand-centric: they treat each hand as an independent instance, which can lead to ambiguous ownership in multi-person scenes. We propose a person-centric formulation in which a single query predicts a structured output for one person, including the human box, body pose, hand boxes and states, and interaction targets. We introduce part-aware deformable attention to allocate attention across human, hand, and pose-specific reference regions, enabling one query to capture the full person…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Single-Query Person-Centric Bimanual Hand-Object Interaction Detection

待翻譯:Beyond Argmax: A Mechanistic Study of Semantic Retention in Frozen Foundation-Model Composition for Generalized Few-Shot 3D Segmentation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12099v1 Announce Type: new Abstract: Classical classifier-combination work distinguishes score-level fusion from hard decision-level voting. We revisit this distinction where independently pretrained, frozen foundation models are composed at inference time for generalized few-shot 3D segmentation. We ask: how much useful semantic information is lost when heterogeneous sources are collapsed to a single class before they can interact? We answer with a same-input semantic-retention intervention. Dense RegionPLC and sparse cross-view SAM3 evidence, model weights, masks, geometry, vocabularies, and fusion rules are frozen; only the number of semantic alternatives retained before interaction is varied via a matched top-k ladder. On 156 held-out ScanNet200…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Beyond Argmax: A Mechanistic Study of Semantic Retention in Frozen Foundation-Model Composition for Generalized Few-Shot 3D Segmentation

待翻譯:Does Video Memory Use What It Retrieves? A Causal Audit of Memory Specificity

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12090v1 Announce Type: new Abstract: Video models increasingly use memory to preserve information over long sequences, with the assumption that gains come from retrieving and using the correct past content. Standard memory ablations test whether memory helps, but not whether the retrieved content is responsible. We test this directly with read-time memory substitution, which replaces the consumed memory value while leaving the rest of the computation unchanged. This separates memory benefit from memory specificity, the extent to which the gain depends on retrieved content. Across frozen video world models, identity-free controls containing no evaluation-specific content recover essentially the full benefit on Ego-Exo4D and 7-Scenes and about 70% on T…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Does Video Memory Use What It Retrieves? A Causal Audit of Memory Specificity

待翻譯:Chopthin-Consensus Power Sampling: A Diversity-Preserving Approach to LLM Decoding

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12243v1 Announce Type: new Abstract: Inference-time power sampling via Sequential Monte Carlo (SMC) can substantially improve large language model (LLM) reasoning without requiring post-training. However, many existing SMC approaches rely on equal-weight resampling, which can aggressively prune low-weight trajectories, discarding potentially correct reasoning paths and degrading the genealogical diversity of the search space. To address this, we introduce Chopthin-Consensus Power Sampling (CCPS). Our method applies the Chopthin resampler to LLM decoding: rather than equalizing weights and forcing unnecessary particle duplication, it enforces an upper bound on the ratio between the largest and smallest weights and carries the unequal weights forward.…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Chopthin-Consensus Power Sampling: A Diversity-Preserving Approach to LLM Decoding

待翻譯:Repair Before Reinforce: Context-Augmented Knowledge Graph Reasoning for Multi-Hop Question Answering

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12230v1 Announce Type: new Abstract: Question-answering often requires reasoning across multiple connected facts rather than retrieving a single isolated relation. Knowledge graphs (KGs) provide a structured way to represent such facts, but training large language models (LLMs) only on isolated KG head-relation-tail triples may limit their ability to learn the surrounding context needed for multi-hop reasoning. In this work, we propose a context-augmented training framework for multi-hop question-answering. Although generally applicable, we validate the framework in the context of disease-specific KGs, extracted using a reliable KG extraction framework called GraphMERT, for Gastroparesis and Diabetes. For each primary KG triple, we attach supporting…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Repair Before Reinforce: Context-Augmented Knowledge Graph Reasoning for Multi-Hop Question Answering

待翻譯:Quantifying Consonant Contributions to Word Intelligibility via Acoustic Masking

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12122v1 Announce Type: new Abstract: Consonants contribute unequally to whether a word is understood. Given the limited time available for therapy, ranking consonants by contribution to intelligibility helps prioritize intervention targets in motor speech disorders. However, measuring this contribution relies on perceptual studies that are difficult to scale. This paper presents a scalable method that measures consonant contribution using acoustic masking. We silence one consonant at a time in an isolated word and test whether an automatic speech recognition (ASR) model still recognizes the word. We define a consonant's contribution score as the proportion of its masked instances for which the word becomes misrecognized, which we refer to as the mask…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Quantifying Consonant Contributions to Word Intelligibility via Acoustic Masking

待翻譯:The Cost of Compression: A Rate-Distortion Limit on Factual Hallucination

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12111v1 Announce Type: new Abstract: Factual hallucination in closed-book question answering is often treated as a coverage problem: a model fails because the relevant fact is absent from its internal memory. This view misses a second source of error. Even when a fact has been observed, finite memory may force it to be stored only approximately. We study this effect through a simple coverage--compression model of factual recall. We consider an unstructured question-answering task with $N$ possible queries and $K$ possible answers. A learner observes $M$ training facts, compresses them into at most $B$ bits, and answers uniformly drawn test queries without retrieval. For a uniformly random ground-truth mapping, we prove $\mathcal{E} \geq \frac{M}{N}\d…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:The Cost of Compression: A Rate-Distortion Limit on Factual Hallucination

待翻譯:Extracting Dataset Mentions in Forced Displacement and FCV Documents: A Weakly Supervised Framework with LLM-Based Label Refinement

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12107v1 Announce Type: new Abstract: Development and humanitarian organizations produce and support surveys, administrative registries, and other data resources to inform research, policy, and operations, yet systematically identifying where these datasets are referenced remains difficult. Such references are dispersed across research papers, project documents, humanitarian reports, and other unstructured text, limiting both the ability to trace data use and to identify potential gaps in data availability or dissemination. We present a weakly supervised framework for adapting dataset extraction to forced displacement and Fragile, Conflict, and Violence (FCV) documents without first constructing a large manually labeled training corpus. A lightweight…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Extracting Dataset Mentions in Forced Displacement and FCV Documents: A Weakly Supervised Framework with LLM-Based Label Refinement

待翻譯:R2VC: Modular Fact-Checking with Retrieval, Verification, and Confidence Calibration

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11955v1 Announce Type: new Abstract: Large language models are increasingly used for automated fact checking, but end-to-end prompting often entangles evidence retrieval, reasoning, and uncertainty estimation, making failures difficult to diagnose and confidence difficult to trust. We present R2VC, a modular retrieve, reason, verify, calibrate architecture for evidence-grounded fact checking with citations and abstention. R2VC combines hybrid sparse+dense retrieval over Wikipedia, a supervised fine-tuned and DPO-aligned generator that produces diverse structured verdict candidates, an external NLI cross-encoder for evidence-based candidate selection, and a lightweight sequence-level calibrator for confidence estimation and selective abstention. On FE…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:R2VC: Modular Fact-Checking with Retrieval, Verification, and Confidence Calibration

待翻譯:On-Device Language Models for Privacy-Preserving Stress Prediction: A Multimodal Evaluation on Mobile Health

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11961v1 Announce Type: new Abstract: Stress is a pervasive determinant of mental health and a key target for mobile health interventions. On-device language models (ODLMs) offer privacy-preserving inference without cloud dependency, yet their feasibility for health prediction under mobile resource constraints remains underexplored. We evaluate ODLMs for multi-modal stress prediction using zero-shot prompting, measuring predictive accuracy alongside latency and throughput. Our results show that objective sensor features marginally outperform subjective self-reports on average, and that lightweight sub-2B models achieve low latency with predictable resource usage. Our findings highlight both the promise and the practical constraints of ODLMs for mobile…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:On-Device Language Models for Privacy-Preserving Stress Prediction: A Multimodal Evaluation on Mobile Health

待翻譯:Performance, Efficiency and Collapse -- Advantages and Challenges in Offline Post-training of Code LLMs

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11956v1 Announce Type: new Abstract: Post-training with reinforcement learning (RL) is a critical phase in the development of code-generating large language models (LLMs), as it ensures adherence to instructions and the production of functionally correct code. This process typically requires computationally intensive code sample generation from Transformer-based LLMs and substantial GPU-CPU communication for sequence verification. To address these computational challenges, this work examines whether RL-based post-training can be performed entirely offline by leveraging existing datasets rather than generating new samples. The findings indicate that, with only a few hours of training, zero-shot code generation performance of LLMs can be substantially…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Performance, Efficiency and Collapse -- Advantages and Challenges in Offline Post-training of Code LLMs

待翻譯:Efficient AI Model Deployment Using Quantization Analysis Tool

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11954v1 Announce Type: new Abstract: As deep learning models are increasingly deployed on resource constrained devices, the demand for efficient model optimization techniques continues to grow. Effective deployment of AI models on edge and low power platforms requires optimization methods that reduce model size and computational cost while maintaining high accuracy. This paper presents Quantization Analysis Tool, a practical system designed to streamline quantization workflows and support performance efficient model deployment. Built on the ONNX framework for broad interoperability, the tool provides detailed layer-wise sensitivity analysis, visualization of weight and activation distributions, and insights to guide precision selection. By identifyin…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Efficient AI Model Deployment Using Quantization Analysis Tool

待翻譯:Physics-Informed Conformal Prediction: Embedding PDE Consistency into Distribution-Free Uncertainty Quantification for Neural Operators

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11935v1 Announce Type: new Abstract: Neural operators such as the Fourier Neural Operator (FNO) achieve remarkable accuracy in approximating solutions to partial differential equations (PDEs). However, providing rigorous uncertainty estimates remains an open challenge. We propose Physics-Informed Conformal Prediction (PI-CP), a framework that embeds PDE residuals into the nonconformity score of split conformal prediction, producing prediction intervals that are (i) distribution-free with provable coverage guarantees, and (ii) spatially adaptive when the PDE residual correlates with prediction error -- tighter where physics is well-satisfied, wider where it is violated. Additionally, we prove that FNO's translation equivariance creates a fundamental a…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Physics-Informed Conformal Prediction: Embedding PDE Consistency into Distribution-Free Uncertainty Quantification for Neural Operators

待翻譯:Perplexity trusts GPT-6 Astra with end-to-end systems

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Perplexity uses Astra to write communications, change software, and monitor production systems, and checks in much less frequently than with earlier models.

OpenAI News來源內容 · 翻譯待補全待翻譯:Perplexity trusts GPT-6 Astra with end-to-end systems

待翻譯:A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 96 Blocks per Token with Unbounded Temporal Depth

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Yifan Zhang's Recurrent Looped Transformer (RLT) technical report proposes a causal encoder paired with a recurrent decoder that carries its final hidden state and layerwise sliding-window attention cache across every prompt and response token, with no reset at the serving boundary. The reference tied configuration uses 48 encoder and 48 decoder layers, executing 96 logical blocks per token while the state path grows to 48t decoder blocks after t tokens. The design also specifies hardware-aware execution around the recurrent core and an exact current-policy RL replay contract. No code, weights, or measured results are released yet. The post A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 9…

MarkTechPost來源內容 · 翻譯待補全待翻譯:A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 96 Blocks per Token with Unbounded Temporal Depth

待翻譯:‘Too little, too late’: Critics perplexed and suspicious of AI leaders’ call for a slowdown

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:From the Trump administration to AI experts, plans by the Anthropic boss to boost safety have spawned a largely negative response OpenAI boss and Elon Musk back slowdown on ‘reckless’ AI development Editorial: controlling AI – humanity cannot outsource its survival The safety debate that ignited when whistleblowers last week warned that the most advanced AI poses an existential threat to humanity should come with a health warning: side effects include whiplash. It was only seven days ago that AI enthusiasts were cheerily toying with OpenAI’s newest model GPT-6 Astra which was marketed in part as a lifestyle tool to help people book tennis courts, run fashion companies and order takeaways. By Wednesday, cutting edge AI started to look a lot less appealing when t…

The Guardian AI來源內容 · 翻譯待補全待翻譯:‘Too little, too late’: Critics perplexed and suspicious of AI leaders’ call for a slowdown
研究

待翻譯:Trump claims he is only ‘guardrail’ needed to control AI as top Republicans join him in dismissing calls for more checks – live

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:President condemns what he claims is a ‘sick conspiracy’ amid global tech selloff following CEOs calls for slowing pace of development Sign up for the Breaking News US email Donald Trump has claimed there is a “SICK conspiracy” against artificial intelligence and data centers in response to the growing calls for greater checks on AI development. “The only control or ‘guardrails’ that AI needs is a STRONG AND SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!” Trump wrote on Truth Social, insisting that the administration has stopped leaders of AI companies from “doing bad, or potentially bad” things. However, he didn’t point to any concrete examples for how he has curbed possible abuse or malfeasance in the industry. There is a SICK conspiracy goin…

The Guardian AI來源內容 · 翻譯待補全待翻譯:Trump claims he is only ‘guardrail’ needed to control AI as top Republicans join him in dismissing calls for more checks – live

待翻譯:An Automated Thickness Evaluation Procedure Using an Integrated Structured Light 3D Camera in a Robotic Bioprinting Framework

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12206v1 Announce Type: new Abstract: Bioprinting is emerging as a tissue engineering technique to replace common treatment methods for large scale injuries. While thickness of the BioPrinted Constructs (BPCs) have shown to be of importance in the cell maturation and integration, the literature lacks a robust, automated, and quantitative method for measuring these metrics. In this paper, we propose a fully automated vision-based method for measuring the thickness of the BPCs with complex geometries. Leveraging the point cloud and RGB images of a structured light 3D camera, our proposed method performs an image-based segmentation for delineating the BPCs from the RGB images, accompanied by novel geometry-based thickness measurement algorithms performed…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:An Automated Thickness Evaluation Procedure Using an Integrated Structured Light 3D Camera in a Robotic Bioprinting Framework

待翻譯:Battery-Aware Predictive Trajectory Planning and Control for Multirotors Under Disturbances

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12188v1 Announce Type: new Abstract: This paper presents a battery-aware predictive trajectory-planning and control framework for multirotors operating under spatially localized disturbances. Candidate trajectories are evaluated through closed-loop vehicle--motor--battery propagation, allowing disturbance-induced control demand, electrical energy, battery evolution, and terminal-voltage-dependent actuator capability to enter the planning process. % A reduced-order battery model is numerically benchmarked against an independently implemented Simscape equivalent-circuit reference, with a power NRMSE of $0.64\%$ and a cumulative-energy discrepancy below $0.7\%$. % In a $150$-s, $640$-m mission containing three disturbance regions, the selected trajector…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:Battery-Aware Predictive Trajectory Planning and Control for Multirotors Under Disturbances

待翻譯:A Physics-Based Closed-Loop Robotic Bioprinting Framework Towards Volumetric Muscle Loss Treatment

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12159v1 Announce Type: new Abstract: Robotic bioprinting and Direct Ink Writing (DIW) are being explored towards the treatment of Volumetric Muscle Loss (VML). While previous studies have shown the importance of proper parameter selection on the print outcome, existing approaches often rely on time- and material-intensive design of experiments methods, or require large, well-curated datasets for training machine learning models. In this paper, we propose a physics-based closed-loop robotic bioprinting system capable of near real-time parameter adaptation. The system integrates a 3D point cloud camera and fully autonomous vision-based algorithms to provide quantitative evaluation of printed constructs. This evaluation is fed into a controller that adj…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:A Physics-Based Closed-Loop Robotic Bioprinting Framework Towards Volumetric Muscle Loss Treatment

待翻譯:Uncertainty-Aware Conflict Detection Against Operator-Conditioned Weather Hazards

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12095v1 Announce Type: new Abstract: Strategic flight plan validation in Advanced Air Mobility (AAM) environments requires robust methods for predicting aircraft state uncertainty and detecting potential conflicts with dynamic airspace hazards. This paper presents a novel framework for uncertainty-conditioned trajectory prediction combined with polyhedra hazard representation for pre-flight conflict detection. We introduce a closed-form uncertainty estimation method that couples non-uniform rational B-spline (NURBS) curve fitting for kinematic trajectory generation with a Kalman Filter for state covariance propagation. Drawing from the Light Propagation Algorithm (LPA) paradigm, we employ a sigmoid-blended measurement noise model that captures the un…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:Uncertainty-Aware Conflict Detection Against Operator-Conditioned Weather Hazards

待翻譯:Pelican-Sim 1.0: A General World Model Simulator for Embodied Intelligence

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12036v1 Announce Type: new Abstract: In this technical report, we propose Pelican-Sim 1.0, a general world model simulator for embodied intelligence that predicts future observations from visual context and robot actions to support downstream learning and decision making. The model incorporates four key design features: (1) Unified action representation: a 28-dimensional action value space covering most mainstream embodiments, keeping one model valid across heterogeneous devices. (2) Action-visual injection: URDF- and camera-rendered action videos bridge actions and pixels, giving markedly better controllability across embodiments, scenes, and tasks (PSNR +0.904 over alternative fusion baselines). (3) Sparse mixture-of-experts (MoE): sparse MoE layer…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:Pelican-Sim 1.0: A General World Model Simulator for Embodied Intelligence

待翻譯:Scenario-Independent Criticality Assessment and Prediction for Vulnerable Road Users in Autonomous Driving

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11947v1 Announce Type: new Abstract: Increasing safety is the primary objective of automated vehicles. Achieving this goal requires reliable safety metrics that incorporate safety-relevant factors such as object type, velocity, and criticality. A key capability of such metrics is the distinction between critical and non-critical objects, which is addressed through criticality or relevance estimation. Existing criticality metrics are typically designed for specific scenarios and primarily focus on vehicle-to-vehicle interactions. In this paper, we therefore propose a novel criticality metric tailored to vulnerable road users (VRUs), which require special consideration due to their less predictable motion behavior. Furthermore, to avoid the complexity…

arXiv Robotics來源內容 · 翻譯待補全待翻譯:Scenario-Independent Criticality Assessment and Prediction for Vulnerable Road Users in Autonomous Driving

待翻譯:Revisiting Multi-Object Tracking Baselines: Hyperparameter Optimization with Multi-Fidelity Greedy Coordinate Search

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12261v1 Announce Type: new Abstract: Multi-object tracking (MOT) is dominated by the tracking-by-detection paradigm, whose methods typically rely on a small set of hyperparameters that are conventionally chosen by hand. Tuning them requires repeated expert-guided experimentation, while the procedures used to select reported values are often not systematically evaluated or fully documented. Hyperparameter optimization (HPO) automates this process, yet it remains rarely used in MOT, and existing studies applying HPO to MOT predate modern deep-detector-based trackers and HOTA evaluation. We systematically apply HPO across two datasets and four tracking-by-detection methods. We also propose Multi-Fidelity Greedy Coordinate Search (MFGCS), which optimizes…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Revisiting Multi-Object Tracking Baselines: Hyperparameter Optimization with Multi-Fidelity Greedy Coordinate Search

待翻譯:USPLIT-VQA: U-Shaped Split Learning for Visual Question Answering with Contribution-Aware Weighted Aggregation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12168v1 Announce Type: new Abstract: Visual Question Answering (VQA) systems, jointly interpreting images and natural language queries, hold significant promise across many domains, yet the privacy-sensitive nature of user data creates a fundamental barrier. Centralized training requires access to all data, while federated learning requires each client to host the full model. We propose USPLIT-VQA, a U-shaped split learning framework for privacy-preserving VQA in which each client retains the initial layers and the classification head while the server hosts the computationally heavy intermediate layers, keeping raw inputs and labels on the client device. We further introduce Contribution-Aware Weighted Aggregation (CAWA), a gradientsimilarity-based c…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:USPLIT-VQA: U-Shaped Split Learning for Visual Question Answering with Contribution-Aware Weighted Aggregation

待翻譯:HSI-Road Relabeled: Surface-Aware Road-Scene Segmentation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12151v1 Announce Type: new Abstract: The HSI-Road dataset provides paired RGB and 25-channel NIR (600--960~nm) images with binary masks but no surface-level labels.~This paper introduces a manually labeled six-class taxonomy: Background, Asphalt, Concrete, Dirt, Water, and Grass, and an RGB-to-NIR registration pipeline with corresponding annotations. Six semantic-segmentation models (SSMs) are evaluated under four input configurations: original-resolution RGB (RGB$_{\text{ori}}$), registered low-resolution RGB (RGB$_{\text{reg}}$), NIR, and channel-stacked RGB$_{\text{reg}}$--NIR (RGBN$_{\text{stk}}$). The comparison quantifies the effect of spatial-resolution reduction on RGB, along with evaluation of NIR and RGBN$_{\text{stk}}$, with results report…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:HSI-Road Relabeled: Surface-Aware Road-Scene Segmentation

待翻譯:Feature Recovery for Object Understanding After Irreversible Fire Damage

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12078v1 Announce Type: new Abstract: Objects in post-fire environments often undergo irreversible physical transformations that change their geometry, material state, and visual appearance. Detecting and identifying these remnants is critical for locating hazards, reconstructing pre-incident contents, and inventorying losses. Unlike standard image corruptions, these degradations affect the physical structure of the object itself. To study this setting, we introduce TRACE, a transformation-aware benchmark for post-fire object understanding. TRACE contains 21.4K real-image-grounded synthetic scenes and paired object-level pristine-to-degraded progressions spanning 499 object identities across 189 categories. We define five tasks targeting localization…

arXiv Computer Vision來源內容 · 翻譯待補全待翻譯:Feature Recovery for Object Understanding After Irreversible Fire Damage

待翻譯:Population-level measures of perceived food access reveal barriers beyond geographic proximity

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.12132v1 Announce Type: new Abstract: Food access is multidimensional, but population-level measurement still relies heavily on geography because perceived dimensions of access are difficult to measure at scale. Here, we use 25,125 Google Maps reviews from 49 grocery stores in Raleigh, North Carolina, to measure five dimensions of food access: availability, accessibility, affordability, accommodation, and acceptability. We identify review topics with unsupervised topic modeling and assign them to access dimensions using zero-shot classification, with 85.4% agreement against manual coding. The resulting store-level measures capture distinct aspects of food access and reveal barriers that geographic proximity alone does not capture. Comparisons between…

arXiv Computational Linguistics來源內容 · 翻譯待補全待翻譯:Population-level measures of perceived food access reveal barriers beyond geographic proximity

待翻譯:Space as an Interventional Invariant: Cross-Modal Predictive Geometry for Stratified Cities and Em-Spaced Intelligence

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11959v1 Announce Type: new Abstract: Space is a foundational concept across mathematics, physics, spatial cognition, urban science, and embodied intelligence, yet these fields often treat spatial structure either as a shared geometric container or as a collection of disconnected representations. Such approaches struggle to explain how heterogeneous sensory and urban processes can jointly reveal a common spatial structure, particularly when different modalities do not share the same metric or representation. This paper addresses this gap by defining space as an interventional invariant: the minimal relational structure that preserves local compatibility and the conditional laws of future observations under admissible actions. We develop a cross-modal…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Space as an Interventional Invariant: Cross-Modal Predictive Geometry for Stratified Cities and Em-Spaced Intelligence

待翻譯:Fed-Equilibrium Framework for Topological Pareto Control in Robust and Fair Clinical Federated Learning

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11937v1 Announce Type: new Abstract: The deployment of Federated Learning (FL) in multi-center clinical networks faces the challenge of "knowledge dominance," where high-volume hubs naturally overwhelm minority community nodes, implicitly treating the distinct clinical patterns of smaller cohorts as outliers. Existing geometric defenses provide a security baseline but leave this efficiency-fairness dilemma unresolved. To bridge this gap, we propose Fed-Equilibrium, a framework that advances the paradigm from simple defense to topological equilibrium. Unlike traditional aggregators, Fed-Equilibrium implements a sequential architectural synergy. It utilizes a two-stage gradient control cascade: Stage I (geometric quality assurance) enforces directional…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Fed-Equilibrium Framework for Topological Pareto Control in Robust and Fair Clinical Federated Learning

待翻譯:Fundamental Dynamical Units for Physics-Informed Structural Inference from Perturbation Time-Series in Networked Systems

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2609.11934v1 Announce Type: new Abstract: In networked dynamical systems, the parameter of primary mechanistic interest is signed interaction structure. Recovering this structure from perturbation time-series data is a fundamental identification problem, compounded by three coupled obstacles: the combinatorial complexity of interaction architectures, ambiguity of causal attribution under limited interventions, and state-dependent dynamics that confound structural inference. Each obstacle is structural in origin and calls for a structural solution. We address these challenges by adopting a reductionist approach, introducing Fundamental Dynamical Units (FDUs): signed three-node interaction patterns as composable primitives that convert the interaction hypot…

arXiv Machine Learning來源內容 · 翻譯待補全待翻譯:Fundamental Dynamical Units for Physics-Informed Structural Inference from Perturbation Time-Series in Networked Systems

待翻譯:MIT spinout turns plastic waste into resilient building materials

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Atlas Building Composites is commercializing MIT research to turn plastic waste into parts for buildings and other infrastructure.

MIT News AI來源內容 · 翻譯待補全待翻譯:MIT spinout turns plastic waste into resilient building materials

待翻譯:Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstruction

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:In this tutorial, we build an end-to-end hierarchical Neural Radiance Field (NeRF) using JAX, Flax, Optax, and the volume-rendering primitives provided by jax3d. We first construct a synthetic multi-view dataset from an analytic scene containing volumetric geometry and view-dependent radiance, using sample_along_rays and volume_rendering to establish the forward rendering process. We then implement a NeRF […] The post Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstruction appeared first on MarkTechPost.

MarkTechPost來源內容 · 翻譯待補全待翻譯:Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstruction
機器人

待翻譯:Fully Driverless Robotaxis Now Operating in Europe

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The trial is the first of its kind on the continent.

AI Business來源內容 · 翻譯待補全待翻譯:Fully Driverless Robotaxis Now Operating in Europe
晶片

待翻譯:Perplexity Portable Computer Is Now Available on Windows, Powered by NVIDIA RTX

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:As local models become more capable, AI agents can handle more work directly on a PC while keeping sensitive information on the device. Portable Computer is a local version of the agent Perplexity Computer that plans and carries out multistep tasks. Accelerated by NVIDIA GPUs, it uses local models to analyze data, bring together information […]

NVIDIA Blog來源內容 · 翻譯待補全待翻譯:Perplexity Portable Computer Is Now Available on Windows, Powered by NVIDIA RTX

待翻譯:How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:On 25 August, OpenAI fully unveiled Jalapeño, the company’s debut AI accelerator chip. Jalapeño delivers up to 13.4 petaflops of 4-bit compute and accesses 232 gigabytes of the most advanced memory available, linking to it at a blazing 15.4 terabytes per second. Benchmarks cited by OpenAI show that Jalapeño can reduce end-to-end latency (the time between prompt to last token) by up to 3.6x when compared to Nvidia’s GB300—a chip the company currently relies on—and do so while consuming less power. Whether these figures translate into real-world gains once Jalapeño enters widespread service in OpenAI’s inference fleet remains to be seen, but performance is only half the story. The other half is how the chip was designed—a process which, as you might expect, was a…

IEEE Spectrum AI來源內容 · 翻譯待補全待翻譯:How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip

待翻譯:AI-linked stocks fall after tech bosses call for slowdown in ‘reckless’ development

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Shares in tech firms tumble across Asia with Nasdaq also likely to dip later after Anthropic, OpenAI and SpaceX leaders back call AI-linked stocks tumbled on Monday after the bosses of Anthropic, OpenAI and SpaceX called for a slowdown in “reckless” development citing fears the technology could soon run out of control. Shares in SoftBank, a Japanese investor that is a big backer of OpenAI, slumped 13%, while the South Korean Kospi stock index, which relies heavily on chip makers that supply AI companies, dropped by 3%. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:AI-linked stocks fall after tech bosses call for slowdown in ‘reckless’ development

待翻譯:Codex GPU Queue

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:Codex GPU Queue
工具

待翻譯:Australia’s outdated technology is vulnerable to AI hacking attacks, signals chief says

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Abigail Bradshaw says ‘enormous’ amounts of money required to update systems that people now expect to be constantly available Get our breaking news email, free app or daily news podcast One of Australia’s top intelligence agencies has warned that AI attacks could exploit the country’s old technology, as the government negotiates its guardrails on artificial intelligence’s rapid development. The chief of Anthropic, developer of Claude, has warned AI bots could swarm the internet within a year and called to “slow the pace” of development with support from the head of OpenAI and Elon Musk. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:Australia’s outdated technology is vulnerable to AI hacking attacks, signals chief says

待翻譯:ChinaMarketing.AI GEO Workspace

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:ChinaMarketing.AI GEO Workspace

待翻譯:Labor unions must unite against AI datacenters | Tyler Turner

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Workers should start with a moratorium against building until our livelihoods and communities are protected For the surging number of Americans who now live alongside constantly humming computer-filled mega-warehouses, it is clear AI datacenters come with significant drawbacks. Opposition to datacenters grows on both sides of the political aisle as developers plan to invest $8tn to build them in the coming years. The bet is massive – roughly a quarter of our country’s entire GDP – but AI CEOs think it will pay off because they believe this technology will allow businesses across industries to rake in profits by cutting out their largest expense: labor. If all goes according to their plan, AI data centers will enable corporations to redistribute workers’ salarie…

The Guardian AI來源內容 · 翻譯待補全待翻譯:Labor unions must unite against AI datacenters | Tyler Turner

待翻譯:Jottoo

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:Jottoo

待翻譯:‘Tidal wave’ of Pfas being launched to satisfy AI industry, campaigners warn

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Big manufacturers planning to produce more of the forever chemicals to meet demand from datacentres Pfas companies are launching a “tidal wave” of forever chemicals production to meet demand from the AI industry, a campaign organisation has warned. Despite alarm over the chemicals, which have been linked to cancer, a survey of the 10 biggest manufacturers found most had plans to increase production. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:‘Tidal wave’ of Pfas being launched to satisfy AI industry, campaigners warn

待翻譯:Doneit 3.2

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:Doneit 3.2

待翻譯:Can you futureproof your career by choosing an AI-resistant degree?

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Humans will have to compete with AI-powered virtual employees in the near future. Here are our tips for the last human holdouts Futureproofing your career with an AI-resistant degree is a tough ask, says Charlie Ball, an expert on graduate employment for Jisc, the UK’s higher education digital, data and technology agency. “What you’re trying to do is futureproof a 45-year career in a time of rapid technological change – and that’s really quite hard to do. What’s likely is that jobs that hinge a lot on human, face-to-face interaction are unlikely to be replaced.” Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:Can you futureproof your career by choosing an AI-resistant degree?

待翻譯:We greet the news that AI could extinguish us with a glazed indifference. What is the way out of this nihilism? | Brigid Delaney

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:There are several answers, and none of them will be provided by tech. One thing we must do is identify the ingredients for human flourishing To imagine the future as being almost entirely bleak has somehow crept into the collective this year. Last week I was having coffee with an old friend. He was telling me about what his kids wanted to study when they went to university, when he suddenly broke off, stared into the middle distance and said, “I wish I never had children.” Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:We greet the news that AI could extinguish us with a glazed indifference. What is the way out of this nihilism? | Brigid Delaney

待翻譯:AI isn’t going to end humanity ... right? – podcast

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:A series of figures in the tech world have been warning about – and even resigning over – the existential threat posed by AI. How worried should we be? With technology editor Robert Booth Ten per cent. That’s the risk, more or less, claimed a high-profile AI developer last week, that the technology wipes humanity out in the next decade. And he was not alone – over the last few months there has been a string of warnings, resignations and even stories of AI companies losing control of their bots. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:AI isn’t going to end humanity ... right? – podcast

待翻譯:The Minimalist Entrepreneur Skills

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:The Minimalist Entrepreneur Skills

待翻譯:Duvi

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Discussion | Link

Product Hunt AI來源內容 · 翻譯待補全待翻譯:Duvi

待翻譯:Trump and Mike Johnson think the AI industry is overreacting

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Yesterday, Anthropic CEO Dario Amodei published a lengthy open letter saying it was time to "pace the frontier" and slow down AI development. OpenAI's Sam Altman and Elon Musk both agreed, publicly voicing their support on X. Even Alphabet's Demis Hassabis offered tentative support for Amodei's proposal. Donald Trump and House Speaker Mike Johnson, however, seem to think the AI executives are being overreactive and fear that a pause could lead to China outpacing the US in the AI race. According to the Financial Times, Trump said "Look, we're leading China in AI … and, frankly, I want to keep it that way, because whoever wins AI, wins." Jo … Read the full story at The Verge.

The Verge AI來源內容 · 翻譯待補全待翻譯:Trump and Mike Johnson think the AI industry is overreacting
政策

待翻譯:Adversarial Fashion Confronts Surveillance Norms

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:AI-powered cameras dot streets across the world, equipped with the power to identify faces or vehicle license plates. But a public backlash is gaining momentum. Privacy concerns abound, encompassing the lack of consent for capturing data, how that data is stored and used, and the risk of misuse. Those concerns are motivating people to fight back. The DeFlock project, for instance, maps automated license plate readers (ALPRs) to raise awareness. Some people resort to extreme measures, such as vandalizing or damaging ALPRs. Others are stitching together more creative responses, crafting “adversarial fashion” to evade surveillance cameras, like a Kickstarter project called noRecognition, presented at last month’s DEF CON hacker convention. Scrambling surveillance…

IEEE Spectrum AI來源內容 · 翻譯待補全待翻譯:Adversarial Fashion Confronts Surveillance Norms

待翻譯:UK MPs and Lords call for new laws to tackle AI threat to human rights

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Warning follows series of safety incidents, with committee saying threats include public face-scanning and deepfakes UK politics live – latest updates British lawmakers are demanding more restrictions on the power of AI, warning the world is unprepared for the “potentially dire” consequences of the technology, in the latest sign of rising global concern. A new regulatory framework is needed for AI in the UK, including an independent oversight body and legislation to protect the public, according to the cross-party joint committee on human rights, comprising MPs and Lords. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:UK MPs and Lords call for new laws to tackle AI threat to human rights

待翻譯:Precision CX in Regulated Industries

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Customer service is one of the first areas where banks, insurers, and healthcare organizations have deployed AI directly in front of customers, according to the U.S. Government Accountability Office. In financial services, all ten of the country’s largest commercial banks now use chatbots to engage customers, and more than 98 million U.S. consumers interacted with […]

Emerj AI Research來源內容 · 翻譯待補全待翻譯:Precision CX in Regulated Industries

待翻譯:AI CEOs say they need to slow the pace of development. But will they?

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:After apocalyptic warnings about the threats posed by AI, leaders like Sam Altman and Elon Musk backed Anthropic CEO Dario Amodei’s calls to ‘slow the pace’ Facing a public uproar over Anthropic researchers’ repeated warnings that artificial intelligence could kill all of humanity by 2030, the AI company’s CEO Dario Amodei issued a proposal at the weekend to slow down the technology’s advancement to ensure public safety. In a rare display of unity, the heads of the largest US artificial intelligence companies all agreed immediately. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:AI CEOs say they need to slow the pace of development. But will they?

待翻譯:New method enables AI for safety-critical situations

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The “HardFlow” algorithm could help generative AI models produce high-quality outputs that obey strict requirements when “pretty close” doesn’t cut it.

MIT News AI來源內容 · 翻譯待補全待翻譯:New method enables AI for safety-critical situations

待翻譯:US schools and police warn about viral ‘Cat in the Hat’ trend after teens’ arrests

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Disturbing and AI-generated versions of Dr Seuss character have been used to threaten schools and communities Schools and law enforcement authorities in the US are issuing warnings against a viral “Cat in the Hat” social media trend in which disturbing or AI-generated versions of the Dr Seuss character are used to threaten students, schools and communities. The latest warnings come as safety concerns increase and several teenagers have been arrested or charged over posts connected to the trend. Continue reading...

The Guardian AI來源內容 · 翻譯待補全待翻譯:US schools and police warn about viral ‘Cat in the Hat’ trend after teens’ arrests

待翻譯:AI Regulation: Who Should Define the Rules? | Cohere

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:A perspective from Aidan Gomez, Co-founder & CEO of Cohere Artificial intelligence is remaking the world we live in. Within a generation, the way we discover medicine, manage power grids, and secure our national infrast…

Cohere Blog來源內容 · 翻譯待補全待翻譯:AI Regulation: Who Should Define the Rules? | Cohere