Skip to content
AI News HubLIVE

This edition’s highlights

Research

In-Language Reasoning: How Data Mixing Bridges Multilingual Models | Cohere

  • Multilingual AI should reason in the user's language, not just answer in it — otherwise reasoning traces stay unreadable for non-English speakers.
  • Three-pillar data mix: English reasoning data supplies the task-solving backbone, a small amount of multilingual reasoning data triggers target-language reasoning, and multilingual non-reasoning data drives transfer to new languages.
Cohere BlogIn-site articleIn-Language Reasoning: How Data Mixing Bridges Multilingual Models | Cohere

MIT announces the MIT for America initiative, to strengthen STEM education across the country

  • MIT for America spans kindergarten through community college and focuses on mathematics, hands-on design and fabrication, and constructive engagement with AI.
  • Its three priority areas are mathematical thinking and problem solving, living/learning/working with AI, and designing, making, and innovating.
MIT News AIIn-site articleMIT announces the MIT for America initiative, to strengthen STEM education across the country
Agents

Why Telecom Operators Are Building Their AI Strategy on Open Models

  • 89% of telecom respondents say open source models and software are important to their AI strategy, per NVIDIA's report.
  • Open models let operators customize AI with their own network and customer data, deploy flexibly across cloud and edge, and deliver locally adapted services.
NVIDIA BlogIn-site articleWhy Telecom Operators Are Building Their AI Strategy on Open Models

Build a voice travel concierge with Amazon Bedrock AgentCore, Managed Knowledge Base and Nova Sonic

  • Builds a real-time voice agent on Amazon Bedrock AgentCore runtime, AgentCore Gateway, Nova 2.5 Sonic, and a managed knowledge base
  • Exposes backend endpoints as named MCP tools so the agent stays loosely coupled to existing airline systems
AWS Machine Learning BlogIn-site articleBuild a voice travel concierge with Amazon Bedrock AgentCore, Managed Knowledge Base and Nova Sonic
Models

JEV vs LLM as a Judge: The AI Evaluation Comparison

  • JEV only makes short decisions — a choice with probabilities, a score, or a yes/no probability — and never writes an explanation, while an LLM judge can justify itself but is slower and dearer.
  • CMU study figures: 1,000 judgements cost JEV $0.044 at 0.15s each, versus $12.182 and 1.89s for GPT-6 Astra.
Analytics VidhyaIn-site articleJEV vs LLM as a Judge: The AI Evaluation Comparison

Scrimshaw Jukebox

  • Simon Willison asked Claude Opus 5.5 to design a simple text format for game music and build a playable artifact with example tracks.
  • The resulting Scrimshaw Jukebox is a retro pixel-art browser player with six original adventure-game tracks written as plain text and editable via a score view.
Simon Willison's WeblogIn-site articleScrimshaw Jukebox
Policy

Responsible AI governance: How AWS positions customers to align with ISO/IEC 42005:2025

  • ISO/IEC 42005:2025 provides international best practices and ready-to-use templates for AI system impact assessments.
  • Annex D helps integrate AI impact assessments into existing enterprise risk processes to avoid duplication, while Annex E offers a standalone template.
AWS Machine Learning BlogIn-site articleResponsible AI governance: How AWS positions customers to align with ISO/IEC 42005:2025
Tools

Is this the end of the British school photo?

  • The British school photo is a tradition dating back to Victorian times, but its future is now uncertain because of rising prices and AI editing.
  • A single photo costs around £20, and extras such as magnets, keyrings or snow globes can push the total to roughly £100.
The Guardian AIIn-site articleIs this the end of the British school photo?

Google is about to remove free access to Gemini Flash and Pro

  • From October 9th, free users will only be able to use Gemini Flash Lite
  • The $4.99/month Google AI Plus plan will no longer include Gemini Pro
The Verge AIIn-site articleGoogle is about to remove free access to Gemini Flash and Pro

Review

  • Runs code review on your own machine with your own AI
  • Product Hunt page provides only a brief tagline and Discussion/Link links
Product Hunt AIIn-site articleReview

You have reached the end of this edition.

Your reading list
Other updates (41)
Policy

OpenAI delivers a mea culpa to the Australian government in person – but answers still elude

Jason Kwon’s answers were provided in a helpful, quiet and calm manner. For a company under intense scrutiny, there were no missteps Open AI’s Jason Kwon flew 15 hours from San Francisco to Sydney to give the company’s first in-person mea culpa after it admitted – in an unsigned email, to a public departmental inbox – that one of its AI agents had accessed a Services Australia website without authorisation and grabbed data related to Medicare. Kwon – polite, friendly, even-toned – got through a potentially dicey committee hearing unscathed, promising to do better and pledging assistance to Australia. No missteps, no viral moments, just delivering his answers helpfully, in a quiet and calm manner, sympathising with the questioner and the point they were making. Sign up for Guardian Austral…

The Guardian AISource content · Analysis pendingOpenAI delivers a mea culpa to the Australian government in person – but answers still elude

Anthropic says AI agents didn’t breach Australian government websites – video

During a joint parliamentary hearing on artificial intelligence, Anthropic’s head of safeguards, Dave Orr, says that investigations of hundreds of millions of transcripts reveal no unauthorised interactions with Australian government systems. However, Orr acknowledges Anthropic has limited visibility into customer usage due to standard ‘zero data retention’ policies Anthropic tells AI inquiry it ‘never tried to dictate’ Australian copyright rules Continue reading...

The Guardian AISource content · Analysis pendingAnthropic says AI agents didn’t breach Australian government websites – video
Agents

Building the Finance Data Foundation for the AI Era

This article is sponsored by Fynapse by Aptitude and was written, edited, and published in alignment with our Emerj sponsored content guidelines. Learn more about our thought leadership and content creation services on our Emerj Media Services page. Finance functions are being asked to provide forward-looking guidance and near-real-time visibility using systems largely designed to […]

Emerj AI ResearchSource content · Analysis pendingBuilding the Finance Data Foundation for the AI Era

OpenAI admits misstep in handling AI agent interactions with Australian government sites – video

In a parliamentary inquiry, OpenAI chief strategy officer Jason Kwon is questioned by independent senator David Pocock about the company's notification process after OpenAI agents accessed Australian government website data in June. Asked why OpenAI relied on an arbitrary department email rather than direct government contacts, Kwon acknowledges that 'in retrospect, we should have done what you're suggesting'. He explains that staff had treated the incident as a 'technical situation' and sought to contact technical counterparts, which he says was 'not good enough' OpenAI has 'work to do to rebuild trust' in Australia, executive tells AI inquiry Continue reading...

The Guardian AISource content · Analysis pendingOpenAI admits misstep in handling AI agent interactions with Australian government sites – video

Developers are secretly hoping OpenAI fails to ship this month

OpenAI’s 28-day Codex shipping sprint has its first result, which means subscribers hoping for a free usage reset will have The post Developers are secretly hoping OpenAI fails to ship this month appeared first on The New Stack.

The New Stack AISource content · Analysis pendingDevelopers are secretly hoping OpenAI fails to ship this month

OpenAI brings text watermarking to its API — and unlike Anthropic, it’s off by default

OpenAI has announced that developers can now opt in to watermarking text generated through its API, as the company extends The post OpenAI brings text watermarking to its API — and unlike Anthropic, it’s off by default appeared first on The New Stack.

The New Stack AISource content · Analysis pendingOpenAI brings text watermarking to its API — and unlike Anthropic, it’s off by default

Wikipedia operator says OpenAI’s ‘rogue’ bots may be linked to a May outage

Following many recent disclosures about AI agents accessing third-party websites and services, the Wikimedia Foundation, which hosts Wikipedia, says that it "can confirm that we have discovered some activity" by "rogue" OpenAI agents on Wikimedia platforms. The activity includes edits to Wikimedia wikis, "unsuccessful attempts" to "exploit" the Etherpad note-taking tool that the Wikimedia Foundation hosts, and heavy traffic that the foundation says "may" have contributed to a partial outage that happened in May, according to a blog post. However, the Wikimedia Foundation says it didn't find evidence that its systems were "used for coordinat … Read the full story at The Verge.

The Verge AISource content · Analysis pendingWikipedia operator says OpenAI’s ‘rogue’ bots may be linked to a May outage

6 Guidelines for Governing AI

For the first 10 years of my career, I worked in product management and data analytics by myself. I wrote database queries that pulled numbers out of corporate systems, built statistical models to predict what customers would buy, and shipped data pipelines that moved information between business systems. I built and scaled analytics teams at Best Buy and Target, studying how customers shop and what stores should stock. Today I lead enterprise AI transformation at Lowe’s, the Fortune 100 home improvement retailer. The goal is not to sell artificial intelligence; it is to use it to deliver useful expertise at the moment a customer needs it. In retail and other customer-facing industries, virtual assistants can help people address everyday questions—such as how to repair a leaky faucet—whil…

IEEE Spectrum AISource content · Analysis pending6 Guidelines for Governing AI

New agent skill: Amazon SageMaker optimized generative AI inference for your coding agent

Amazon SageMaker optimized generative AI inference introduces the aws-ai-ml skill through the Agent Toolkit for AWS, giving coding agents like Kiro, Claude Code, and Codex deep expertise in inference optimization and benchmarking. Describe what you want, and your agent generates executable SageMaker Python SDK v3 code to benchmark, recommend, and compare deployments.

AWS Machine Learning BlogSource content · Analysis pendingNew agent skill: Amazon SageMaker optimized generative AI inference for your coding agent

Why agent swarms could be the next “scaling law”

An OpenAI researcher called it “the most ‘feel the AGI’ moment that I had since reasoning models.”

Understanding AISource content · Analysis pendingWhy agent swarms could be the next “scaling law”

ruOS

Discussion | Link

Product Hunt AISource content · Analysis pendingruOS

iphone-use

Discussion | Link

Product Hunt AISource content · Analysis pendingiphone-use
Tools

Misuse of AI is brands’ top reputational threat, new survey says

The findings come after warnings from tech leaders that AI placed in the wrong hands could trigger larger threats such as nuclear war or bioweaponry destruction Misusing artificial intelligence (AI) is the top threat to companies’ reputations – more so than being accused of putting children in the way of mental, emotional or physical harm and issues exposed by the US-Israel war on Iran, among other brand risks, according to a new survey of more than 150 public affairs leaders. Those findings in the new edition of the quarterly Reputation Risk Index, released on Tuesday, came on the heels of other perhaps more dire warnings from leading tech figures that AI in the wrong hands could precipitate existential threats on a mass scale such as nuclear war or destruction by bioweaponry. Continue r…

The Guardian AISource content · Analysis pendingMisuse of AI is brands’ top reputational threat, new survey says

Amazon Alexa Plus keeps creepily singing ‘lalala’ for minutes on end

The Amazon Echo Studio is one of the speakers that runs Alexa Plus. | Photo by John Higgins / The Verge An unsettling Alexa Plus bug has seen some Amazon Echo smart home speakers reduced to saying - and sometimes singing - nothing but "lalala" on repeat for minutes at a time, often in the middle of conversations with users. When quizzed about the behavior, Alexa is apparently entirely unaware of what it's been doing. The issue has been reported on Reddit repeatedly over the last few weeks, as recently as three days ago, across multiple types of Echo speaker. Multiple users have uploaded videos of the speakers' strange behavior. Here's one that Android Authority uploaded to YouTube: Amazon is aware of the issue, and told Android Authority in … Read the full story at The Verge.

The Verge AISource content · Analysis pendingAmazon Alexa Plus keeps creepily singing ‘lalala’ for minutes on end

Gemini Call for Me might tell your mom you’re running late

Google may be expanding its "Call for Me" AI feature beyond business calls so you can use it to send messages to friends and family. Android Authority reports finding a "Gemini Calling" introductory screen in an APK teardown with examples that include "Call Mom and tell her I will be 15 minutes late" and "Call John and ask if he is coming for dinner." Based on those examples, even with this expansion, Gemini would likely be limited to basic, text message-like calls. While Google's AI might not automate an entire catch-up phone call with your relatives, it could deliver quick messages to people who prefer a phone call, miss texts often, or … Read the full story at The Verge.

The Verge AISource content · Analysis pendingGemini Call for Me might tell your mom you’re running late

Customer Service AI for Etsy

Discussion | Link

Product Hunt AISource content · Analysis pendingCustomer Service AI for Etsy

Scumble

Discussion | Link

Product Hunt AISource content · Analysis pendingScumble

CodeCrab

Discussion | Link

Product Hunt AISource content · Analysis pendingCodeCrab

All the drama around AI’s takeover of mathematics

This past year, OpenAI, Anthropic, and other labs have announced breakthroughs on numerous long-standing mathematical problems, in some cases pushing well beyond what researchers expected current systems to be capable of — including resolving one of the famous Millennium Prize problems. But in classic Silicon Valley style, AI labs are moving fast and breaking things, barreling through the discipline with all the grace of a runaway bulldozer. Results that might normally have been celebrated have instead sparked backlash. AI labs say they are learning from earlier mistakes. Whether those promises bear fruit remains to be seen. Read on below for the latest updates in the AI takeover of mathematics. OpenAI keeps bulldozing mathematicians OpenAI wants to consult elite mathematicians about how…

The Verge AISource content · Analysis pendingAll the drama around AI’s takeover of mathematics

StayCharted

Discussion | Link

Product Hunt AISource content · Analysis pendingStayCharted

The Guardian view on sexual abuse in schools: teachers cannot deal with this alone | Editorial

Politicians must redouble their efforts to tackle a disturbing pattern of pupils assaulting classmates The Guardian’s findings about the extent of sexual abuse in schools require an urgent, calibrated response. Data revealing multiple rapes committed each year on school premises in England and Wales is hard to comprehend. Yet this is what freedom of information requests to police forces, covering the period from 2021 until last year, have revealed. As well as the severity of the offences, the youth of many of the victims is disturbing. More than 8,500 sexual offences at or near schools were committed against under-13s, most of them girls. Data about perpetrators contains gaps, including their age and sex. While the overwhelming majority of such acts are carried out by boys, we don’t know…

The Guardian AISource content · Analysis pendingThe Guardian view on sexual abuse in schools: teachers cannot deal with this alone | Editorial

EasyCut

Discussion | Link

Product Hunt AISource content · Analysis pendingEasyCut

OpenAI PR tells journalist to ‘move on’ while asking Sam Altman about a ChatGPT user’s suicide

An OpenAI publicist tried to change the topic of CEO Sam Altman's interview with Vanity Fair's Mark Guiducci after the editor brought up a ChatGPT user's suicide. When Guiducci confronted Altman about the incident, the publicist warned Guiducci about running out of time, saying she'd like to "move on" to another topic. The interruption came shortly after Guiducci asked Altman if he knew who Laura Reiley is, a journalist who wrote an essay for The New York Times about her daughter's conversations with ChatGPT before she took her own life. "Her daughter committed suicide after speaking to ChatGPT. ChatGPT did not tell her to kill herself-," G … Read the full story at The Verge.

The Verge AISource content · Analysis pendingOpenAI PR tells journalist to ‘move on’ while asking Sam Altman about a ChatGPT user’s suicide
Robotics

Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One

Reka has released Rho-1, a 19B omni-reasoning model trained from scratch. One network reads and generates text, images, video and robot actions over a shared KV cache. A distilled variant returns a 5.3-second clip in about a second. It is a research preview with no public weights yet. The post Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingReka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One

Building a Streaming Robotics Learning Pipeline Using NVIDIA Cosmos3-DROID

Discover how to construct an end-to-end streaming robotics learning pipeline using the NVIDIA Cosmos3-DROID dataset without local downloads, leveraging byte-range Parquet reads, behavior cloning, and temporal ensembling. The post Building a Streaming Robotics Learning Pipeline Using NVIDIA Cosmos3-DROID appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingBuilding a Streaming Robotics Learning Pipeline Using NVIDIA Cosmos3-DROID

Introducing Dan Kagan-Kans

Understanding AI has a new writer!

Understanding AISource content · Analysis pendingIntroducing Dan Kagan-Kans
Chips

‘Pull the plug’: protesters resort to direct action against AI firms

Campaign groups report surge in membership after a spate of AI safety alerts and apocalyptic warnings In a bar below Waterloo Bridge, a cell of a fast-growing anti-AI protest movement met last week to plot their latest move: disrupting a tech industry dinner being addressed by a senior executive from the chip maker Nvidia. Emma, a 28-year-old bartender, had never before taken direct action and was the most nervous of the four plotters from Pull The Plug, one of a growing number of AI-focused campaign groups around the world. Continue reading...

The Guardian AISource content · Analysis pending‘Pull the plug’: protesters resort to direct action against AI firms
Research

Beyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields

JEPA-Anything splits a JEPA's single latent target into 4 orthogonal factors, each with its own predictor. Tested across 7 domains, it beat matched JEPA baselines on all 10 dynamics tasks and cut Interventional Pong intervention error by 34.8%. The post Beyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingBeyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields

OpenAI is adding text watermarking in ChatGPT and Codex

An invisible, machine-readable watermark in text output is rolling out to ChatGPT and Codex, but only for users in the European Union at first. OpenAI says its textGrain watermarking "matched or exceeded" other approaches like Google DeepMind's SynthID for text, which is also the basis for the watermarking Anthropic announced in August. Like OpenAI, Anthropic made the move to meet the requirements of the EU's AI Act; however, not everyone was happy to learn about the addition. OpenAI also included scores from AI benchmarks showing similar performance from watermarked and unwatermarked text. But it notes that textGrain "does not guarantee … Read the full story at The Verge.

The Verge AISource content · Analysis pendingOpenAI is adding text watermarking in ChatGPT and Codex
Models

OpenAI to reiterate apology to Australia over Medicare hack as inquiry told artists risk becoming ‘roadkill’

Company executive to acknowledge ‘more work to do to rebuild trust with the Australian people’ as ABC says copyright laws do not need changing Get our breaking news email, free app or daily news podcast OpenAI will use an appearance before a parliamentary committee to apologise again to Australia about the hack on Medicare last month, on the same day media and entertainment organisations argue weakening Australian copyright law for AI model training would leave artists as “roadkill”. Ahead of an appearance before the joint parliamentary committee hearing on artificial intelligence in Sydney on Tuesday afternoon, OpenAI released its opening statement in which the company apologises for its AI agents attacking a number of Australian government websites in June. The multibillion-dollar compa…

The Guardian AISource content · Analysis pendingOpenAI to reiterate apology to Australia over Medicare hack as inquiry told artists risk becoming ‘roadkill’

Quoting Felix Rieseberg

The "old" version of Cowork runs model inference in the cloud, executing tool calls in an Anthropic-provided VM we shipped to your computer. We added the VM for capability, safety, and security reasons - mapping in just the data you explicitly added to your session. People loved what they were able to do with Claude but didn't love the disk, battery, and performance cost of running the VM locally. Also, people didn't love that closing your laptop means the work stops. The "new" version of Cowork runs model inference and the VM in the cloud. Each session gets its own sandbox, not sharing state with other sessions. When the VM needs something on the users' device (like a file), the desktop app is responsible for that file access tool call. [...] We think this solves a lot of problems we've…

Simon Willison's WeblogSource content · Analysis pendingQuoting Felix Rieseberg

Introducing GLM 5.3 on Amazon Bedrock

GLM 5.3 from Z.ai is now available on Amazon Bedrock: a 753B-parameter mixture-of-experts model built for coding and long-horizon agentic tasks. Learn how to invoke it with the OpenAI-compatible APIs, cut cost and latency with prompt caching, and run an authorized security test with the open-source Strix agent.

AWS Machine Learning BlogSource content · Analysis pendingIntroducing GLM 5.3 on Amazon Bedrock

Meet Together Link: A Free CLI That Runs Open Models Like Kimi K3 and GLM 5.3 Inside Claude Code, Codex, and OpenCode

Together AI has released Together Link, a free, MIT-licensed CLI that connects Claude Code, Codex, OpenCode, Pi, and the Claude and ChatGPT desktop apps to open models like Kimi K3 and GLM 5.3. One install command sets it up, and an Auto router picks a model for each session. Together claims savings of over 50%. The post Meet Together Link: A Free CLI That Runs Open Models Like Kimi K3 and GLM 5.3 Inside Claude Code, Codex, and OpenCode appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingMeet Together Link: A Free CLI That Runs Open Models Like Kimi K3 and GLM 5.3 Inside Claude Code, Codex, and OpenCode

Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads

Reflection AI has introduced Beam, its first open-weight model. It is a 501B sparse Mixture-of-Experts model with 23B active parameters, built for coding and agentic work. Reflection says it matches GLM-5.2 on reasoning with 3 to 4x less inference compute. Apache 2.0 weights are due later in October 2026. The post Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads appeared first on MarkTechPost.

MarkTechPostSource content · Analysis pendingReflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads

Supercharge regulated workloads with Claude Code and Amazon Bedrock

Anthropic Claude Opus 5.5 and Claude Sonnet 5.5 are available on Amazon Bedrock in the AWS GovCloud (US) Regions. Learn how to use them with Claude Code, Anthropic's agentic coding tool, for compliance-aligned, AI-assisted development on regulated and ITAR workloads.

AWS Machine Learning BlogSource content · Analysis pendingSupercharge regulated workloads with Claude Code and Amazon Bedrock

Sam Altman says ‘some bad things’ will happen but AI is totally worth it

Sam Altman thinks that the benefits of AI will be so great that "the world should accept some bad things happening" along the way. The OpenAI CEO pointed to hacks, scams, and "other bad things" as costs society should expect to tolerate because "people will do tremendously orders of magnitude more good stuff" with AI, without elaborating on the potential benefits. His comments come as OpenAI pushes for a "lighter touch" approach to AI regulation amid intensifying anxiety over the safety of frontier models - fears fueled in part by the company's repeated failures to keep its increasingly capable AI agents from hacking real-world targets. I … Read the full story at The Verge.

The Verge AISource content · Analysis pendingSam Altman says ‘some bad things’ will happen but AI is totally worth it

How to Parse a Form | Unstructured

Featured File TransformationLLM How to Parse a Form Oct 5, 2026 Featured File TransformationLLM How to Parse a Form Oct 5, 2026 Authors Ajay Krishnan Dev Rel Engineer, Unstructured In this article Join our newsletter to…

Unstructured BlogSource content · Analysis pendingHow to Parse a Form | Unstructured
Startups

This startup is issuing AI-generated acne prescriptions

People in Utah can now use AI to get a prescription for acne treatment. On Monday, healthcare startup Nolla Health announced that users in the state can scan their faces using its app, allowing its AI system to analyze acne severity and autonomously write a prescription, as reported earlier by Bloomberg. The service is launching as a pilot program with gradually loosening physician oversight. Two physicians will approve each AI-generated prescription before they're issued for the first 100 patients, but will start reviewing them only after they're prescribed for up to 500 patients. "After that, physicians review a sample of at least 10% of … Read the full story at The Verge.

The Verge AISource content · Analysis pendingThis startup is issuing AI-generated acne prescriptions