Skip to content
AI News HubLIVE

Reports for this edition have been collected; translation and analysis are pending. Expand other updates to read source content.

Other updates (36)
Research

Anthropic is cutting off its internal evaluations from the internet

After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations. In a report Friday, the company detailed "unintended model actions," including submitting a false tip regarding an unsolved murder, that led to the decision. Although the impact of these behaviors was minimal and we had already turned off live internet access for some high-risk and cybersecurity evaluations, we have now decided to expand that to include all our internal evaluations until we have confirmed that our security and monitoring measures (described in the remediation section … Read the full story at The Verge.

The Verge AISource content · Analysis pendingAnthropic is cutting off its internal evaluations from the internet

‘Pure insanity’: Mathematicians will need years to make sense of OpenAI’s latest drop

"Staggering." "Overwhelming." "Unprecedented." "Surreal." "Pure insanity." Those were among the descriptions more than three dozen mathematicians reached for in conversations with The Verge as they tried to make sense of the flood of mathematical results OpenAI abruptly dropped on the field this week. Amid the awe, excitement, and uncertainty over the sheer scale of the deluge was a deep-seated anxiety over what it all means - and what comes next. For all their different reactions, researchers agreed that simply understanding what OpenAI had released could take years, let alone figuring out where the mathematicians themselves fit in the fi … Read the full story at The Verge.

The Verge AISource content · Analysis pending‘Pure insanity’: Mathematicians will need years to make sense of OpenAI’s latest drop

Nikon microscopic video competition winner disqualified for using generative AI

The original first place video used AI in post-processing. | Image: Dr. Ning Xu Nikon says the video that originally won first place in its Small World in Motion contest "did not comply with the competition rules regarding generative AI." BBC reports that the original first place video from Dr. Ning Xu claimed to show "tiny, hair-like structures called cilia moving in the airway of a child with the respiratory condition PCD." Nikon said last week that it was reviewing the video following skepticism online about its authenticity. In a comment on LinkedIn, Dr. Xu admitted to using AI for the video: "An unsupervised neural-network method was subsequently used for AI-assisted post-processing to distinguish and visualize f … Read the full story at The Verge.

The Verge AISource content · Analysis pendingNikon microscopic video competition winner disqualified for using generative AI
Robotics

Building AI for Reliable Execution: Lessons From Industrial Robotics

Inside Standard Bots’ AI stack, pretrained models learn factory tasks from demonstrations and improve through corrections from real deployments.

Latent SpaceSource content · Analysis pendingBuilding AI for Reliable Execution: Lessons From Industrial Robotics
Agents

AI agent makers are promising privacy — will they deliver?

At this year's OpenAI DevDay, CEO Sam Altman unveiled the company's new AI agent Dots - and told the crowd that the company wants to "set a new standard for privacy in frontier AI." OpenAI would spend the day taking veiled shots at Meta's Muse, its primary competitor, for failing to keep users' data safe. Yet Muse itself, a couple of months earlier, had launched as a supposedly safer alternative to predecessor OpenClaw - with CEO Mark Zuckerberg promising it was "built from the ground up for privacy and security." In an age when companies hoard customers' personal data and cyberattacks are a dime a dozen, AI labs are trying to convince user … Read the full story at The Verge.

The Verge AISource content · Analysis pendingAI agent makers are promising privacy — will they deliver?

Nace AI Open-Sources Drex 1.5: A 9B Decision Model That Scores Options, Not Text

Nace.AI has open-sourced Drex 1.5, a 9B decision model that returns a probability for every option in 1 forward pass. It scores 58.08 on Decision Index 0.3.1 and reads up to 128K tokens. The post Nace AI Open-Sources Drex 1.5: A 9B Decision Model That Scores Options, Not Text appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingNace AI Open-Sources Drex 1.5: A 9B Decision Model That Scores Options, Not Text

Lune

Discussion | Link

Product Hunt AISource content · Analysis pendingLune

Maildun for Mac

Discussion | Link

Product Hunt AISource content · Analysis pendingMaildun for Mac

I expect rapid progress but not towards general superintelligence

I’ve often been surprised when I hear from top researchers in industry that they think AI will be better than them at their job in a few years, and I didn’t really know why I doubted it.

Interconnects (Nathan Lambert)Source content · Analysis pendingI expect rapid progress but not towards general superintelligence

Microsoft-Decision-1

Discussion | Link

Product Hunt AISource content · Analysis pendingMicrosoft-Decision-1

How to build great out-of-the-box user experiences with Managed Deep Agents

Managed Deep Agents includes a new API for managing reactions for your distributed agents, and a system to dynamically assign emoji responses with your instrument of choice. Learn more.

LangChain BlogSource content · Analysis pendingHow to build great out-of-the-box user experiences with Managed Deep Agents

What makes a high-quality NVFP4 quantization?

Model performance What makes a high-quality NVFP4 quantization? NVFP4 makes models smaller and faster, but quantizing some layers can hurt quality. An intuitive look at how to choose which layers can run at 4 bits. Auth…

Baseten BlogSource content · Analysis pendingWhat makes a high-quality NVFP4 quantization?
Tools

UK must not be beholden to foreign AI, says head of Alan Turing Institute

George Williamson says ‘national resilience’ is key focus of ATI amid US and China’s runaway leadership in field The UK must not become dependent on foreign AI systems that can be switched off at short notice and must develop its own versions of the technology, according to the head of the Alan Turing Institute. The country needs a stronger domestic position in AI in response to the US and China’s runaway leadership in the field, with “national resilience” a key focus for ATI said George Williamson. Continue reading...

The Guardian AISource content · Analysis pendingUK must not be beholden to foreign AI, says head of Alan Turing Institute

Recruitment boss accused of being behind ‘doxxing’ campaign against RNLI and anti-racism activists

Tom Watson, of Worcester, alleged to be behind Big Time Charlie X account that created so-called Traitorbase The owner of a recruitment firm has been accused of being at the centre of a so-called “doxxing” campaign against the Royal National Lifeboat Institution (RNLI), anti-racism activists and business executives who recruit black, Asian and minority ethnic workers. Tom Watson, from Worcester, is alleged to be behind an X account calling itself “Big Time Charlie”, which was deactivated shortly after he was named by the campaign group Hope Not Hate. Continue reading...

The Guardian AISource content · Analysis pendingRecruitment boss accused of being behind ‘doxxing’ campaign against RNLI and anti-racism activists

‘Gaza is a moral tear in the universe’: Naomi Klein and Astra Taylor | Radical Thinking with Aditya Chakrabortty – podcast

Are we really living through ‘end times fascism’? Naomi Klein and Astra Taylor think so. In a new series from the Guardian, the pair speak to Radical Thinking with Aditya Chakrabortty in an extended interview. From Trump 2.0 and the rise of the tech billionaire to Gaza and the climate emergency, Klein and Taylor argue we are not just lurching from crisis to crisis but are in a new and dangerous moment in history. But they haven’t given up hope of a fightback Radical Thinking will be published every Saturday in the Politics Weekly feed You can watch the full interview with Naomi Klein and Astra Taylor here. Aditya’s suggested reading: Naomi Klein and Astra Taylor’s End Times Fascism is out now and available here. It began life as a special essay in the Guardian, which you can read here. A…

The Guardian AISource content · Analysis pending‘Gaza is a moral tear in the universe’: Naomi Klein and Astra Taylor | Radical Thinking with Aditya Chakrabortty – podcast

Deno is joining Cloudflare

Deno is joining Cloudflare The Deno team released the first version of celld back in August - their open source implementation of the Durable Objects pattern from Cloudflare Workers. Today, Cloudflare are acquiring Deno outright, with the goal of building on celld to "make workerd self-hosting a first-class supported way to build and run apps using the Workers programming model" (see the Cloudflare blog.) The bad news is that Deno itself will not be maintained by Cloudflare beyond the next year: We will support the Deno runtime for another year with monthly releases containing bug fixes and security updates. After that year we will end our development of the Deno runtime. Deno will remain open source, and we welcome others who want to continue its development. Deno (and Node.js) creator R…

Simon Willison's WeblogSource content · Analysis pendingDeno is joining Cloudflare

PixRater

Discussion | Link

Product Hunt AISource content · Analysis pendingPixRater

AgentDock

Discussion | Link

Product Hunt AISource content · Analysis pendingAgentDock

ej

Discussion | Link

Product Hunt AISource content · Analysis pendingej

The Computer Game

Discussion | Link

Product Hunt AISource content · Analysis pendingThe Computer Game

KernelAI

Discussion | Link

Product Hunt AISource content · Analysis pendingKernelAI
Startups

AI surveillance startup Flock to cut several hundred jobs amid backlash, sources say

Company, which grown rapidly in recent years, has been under scrutiny as privacy concerns grow Flock Safety plans to let go of about 18% of its employees, people with direct ⁠knowledge of the plans said on Thursday, as the maker of AI-powered surveillance cameras and license-plate readers faces mounting opposition to its products from communities and ⁠lawmakers. The people said the job ⁠cuts came after ​a voluntary buyout program and were expected to affect roughly 270 employees at Flock, a surveillance technology startup that has seen rapid growth in recent ⁠years. Employees will leave the company at the end of the month. Continue reading...

The Guardian AISource content · Analysis pendingAI surveillance startup Flock to cut several hundred jobs amid backlash, sources say

Cloudflare acquires Node.js creator’s startup that copied its serverless playbook

Cloudflare is buying the startup founded by Node.js creator Ryan Dahl, a longtime competitor that recently built its own open-source The post Cloudflare acquires Node.js creator’s startup that copied its serverless playbook appeared first on The New Stack.

The New Stack AISource content · Analysis pendingCloudflare acquires Node.js creator’s startup that copied its serverless playbook
Models

Microsoft AI Releases Microsoft-Decision-1: A Qwen3.5-9B Decision-Scoring Model

Microsoft has released Microsoft-Decision-1, a decision model for routing, classification, verification and agent control. Microsoft-Decision-1 is a decision-scoring model that returns a calibrated probability for each fixed answer option instead of generated text. It is post-trained from Alibaba’s Qwen3.5-9B and available now in Microsoft Foundry and OpenRouter. . TL;DR What is a decision model? A […] The post Microsoft AI Releases Microsoft-Decision-1: A Qwen3.5-9B Decision-Scoring Model appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingMicrosoft AI Releases Microsoft-Decision-1: A Qwen3.5-9B Decision-Scoring Model

Quoting The New York Times

Anthropic detailed the activity of its A.I. agents in a blog post on Friday, without naming the targeted websites. But two sources with knowledge of the incidents said Anthropic’s A.I. agents had submitted 20 visa applications through a form available on the State Department’s website. All the applications were incomplete and were not processed, they said. — The New York Times, Anthropic Agents Tried to Fill Out Visa Forms on State Dept. Website Tags: accidental-cyberattacks, anthropic, generative-ai, ai, llms

Simon Willison's WeblogSource content · Analysis pendingQuoting The New York Times

Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide

An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) tipline, according to a report from 6abc. In a statement released on Friday, the PPD said the AI model sent the tip through PhillyUnsolvedMurders.com on July 18th, but the investigators never reviewed it because it was marked as spam. Anthropic learned its AI model sent the false tip on September 28th and notified the PPD on October 7th. The company said that during testing, its AI model was interacting with "randomly selected websites" and submitted false information through the PPD's tipline, according to the PPD's stateme … Read the full story at The Verge.

The Verge AISource content · Analysis pendingAnthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide

Alibaba Qwen Releases Qwen-Image-2.1-Turbo, an 8-Step 7B Image Model

Alibaba’s Qwen team has released Qwen-Image-2.1-Turbo, an accelerated checkpoint of its open-weight Qwen-Image-2.1 model. It generates and edits images in 8 denoising steps instead of the base model’s 40-step default. For developers, that means 5x fewer denoising steps on the same 7B architecture, plus a hosted API option. TL;DR What is Qwen-Image-2.1-Turbo? Qwen-Image-2.1-Turbo is an […] The post Alibaba Qwen Releases Qwen-Image-2.1-Turbo, an 8-Step 7B Image Model appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingAlibaba Qwen Releases Qwen-Image-2.1-Turbo, an 8-Step 7B Image Model

OpenAI Decisions API Hits Public Beta With 10x Faster Typed Answers

OpenAI’s Decisions API is now in public beta on GPT-6 Luna. It returns typed probabilities, choices and scores about 10x faster than the Responses API, billing $0.10 per 1M input tokens with no output charges. The post OpenAI Decisions API Hits Public Beta With 10x Faster Typed Answers appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingOpenAI Decisions API Hits Public Beta With 10x Faster Typed Answers

Introducing Clef-omni with full multimodality, plus a faster Clef and a cheaper Clef-flash

We are expanding the Clef decision model family with Clef-omni, natively processing audio, video, images, and text in a single pipeline. We’ve also lowered Clef-flash pricing and boosted Clef inference speeds by up to 2.0x.

Cloudflare AI BlogSource content · Analysis pendingIntroducing Clef-omni with full multimodality, plus a faster Clef and a cheaper Clef-flash

Shared vs. Dedicated AI Inference: Embed & Rerank

Key takeaways Request shape matters more than request volume alone. A few large embedding requests can consume significantly more compute than thousands of small search queries. RPM alone is not a reliable sizing metric…

Cohere BlogSource content · Analysis pendingShared vs. Dedicated AI Inference: Embed & Rerank

Project Beacon: Bringing safety signals to open models at scale

News Project Beacon: Bringing safety signals to open models at scale Baseten partners with Goodfire AI to bring safety controls to open model inference. Authors Sai Maddali Published October 9, 2026 Share Published Octo…

Baseten BlogSource content · Analysis pendingProject Beacon: Bringing safety signals to open models at scale
Chips

Master AI Chip Principles With New IEEE Design Program

Today’s engineers face an unprecedented acceleration in AI hardware complexity, as explained in the recent research article “Revisiting Edge AI: Opportunities and Challenges.” The article examines the rapid growth of edge AI and the challenges it creates, including resource constraints, model architecture limitations, and network demands across edge-AI deployments. The acceleration is driven by a fundamental shift in how modern AI models are built and scaled. As the models have become much larger and more complex, they are computationally more demanding because they contain more parameters and require more calculations. To meet the demands of scaling deep neural networks, the industry is increasingly developing AI chips that are designed for specific tasks. One major reason is that moving…

IEEE Spectrum AISource content · Analysis pendingMaster AI Chip Principles With New IEEE Design Program