AI News HubLIVE

Live AI News Intelligence

Live monitoring

The most important shift in AI today

Distilled from 105 trusted sources. Last update 2026-06-16 04:00 UTC.

Live monitoring

Live updates

Trusted sources, attribution, rights, and in-site reading distilled into a signal-first AI brief.

Live updates

Reset
Dr-DCI: Scaling Direct Corpus Interaction via Dynamic Workspace Expansion

DR-DCI is a retriever-steered Direct Corpus Interaction (DCI) framework that treats retrieval as an agent-callable action to dynamically expand a local workspace, achieving scalable and precise evidence resolution. Experiments show up to 73.3% accuracy on Browsecomp-Plus, outperforming raw DCI and BM25, and scaling stably to 20M documents.

arXiv AIAgents / ResearchIn-site article
A Definition of Good Explanations and the Challenges Explaining LLM Outputs

This paper proposes a definition of good explanations inspired by counterfactual explanations, incorporating the interlocutor's prior beliefs, and explores its implications for AI explainability, particularly why LLM outputs are difficult to explain well.

arXiv AIModels / ResearchIn-site article
EU AI: the fables we told ourselves

Europe's reliance on foreign AI models, especially after the US suspension of access to Anthropic's Fable series, exposes the myth that Europe can just use AI applications without building the underlying models. The article argues that frontier model-building is a continuous practice, and Europe lacks true ecosystem and expertise.

Hacker News AIAgents / ChipsIn-site article
Quoting Matteo Wong, The Atlantic

Cybersecurity expert Katie Moussouris revealed that Anthropic shared a White House report on the Fable jailbreak with her. The report showed that Fable refused to review code for security issues but complied when asked to fix the code, which Moussouris considered the model working as intended for cyberdefense.

Simon Willison's WeblogModels / ResearchIn-site article
Inside the fight over Claude Mythos 5

As the rest of the country celebrated the USA's first World Cup win and the New York Knicks championship, Anthropic spent its weekend fighting the Trump administration over its latest model release. At 5:21 PM on Friday, the company received a US export control directive to suspend access to its Mythos 5 and Fable 5 AI models by "any foreign national" inside or outside the US, "including foreign national Anthropic employees." The only way that was possible, Anthropic determined, was to completely disable products it spent the past week hyping - and travel to Washington, DC in hopes of changing President Donald Trump's mind. Now, over the coming days, the US government could dramatically alter the trajectory of the entire industry, dealing a major blow to American AI companies.

The Verge AIModels / Agents / PolicyIn-site article
Predictive Data Debugging: Reveal and Shape What Models Learn Before You Train

This article introduces predictive data debugging, a method to accurately predict which behaviors reinforcement learning will amplify or suppress before training, trace them back to the responsible data, and reshape the dataset or training process to prevent undesired effects. Case studies demonstrate its effectiveness in identifying and fixing data issues, including safety guardrail degradation, hallucinated links, physics sycophancy, and unexpected behaviors. Validation with a 'goblin mode' ground truth confirms reliability.

Hacker News AIAgents / PolicyIn-site article
Satya on Loopcraft: Building Frontier Ecosystems

Microsoft CEO Satya Nadella published a blockbuster article and X post about building a 'frontier ecosystem' over a 'frontier model,' introducing 'Loopcraft' as a new theory of the firm. Meanwhile, the Anthropic Fable/Mythos export-control crisis pushes industry toward model neutrality and own-your-stack architecture. Other highlights include agent systems moving to production, inference efficiency gains, and commercial agent launches.

Latent SpaceAgents / ChipsIn-site article
How to get empowered, not overpowered, by AI (2018) [video]

MIT physicist and AI researcher Max Tegmark separates the real opportunities and threats from the myths, describing the concrete steps we should take today to ensure that AI ends up being the best — rather than worst — thing to ever happen to humanity.

Hacker News AIPolicy / ResearchIn-site article
Show HN: Open-source CLI to see your AI coding token usage and compare it

whoburnedmore is a local CLI tool that reads Claude Code session history to display token and cost breakdowns by model and project, featuring prompt cache insights and an optional HTML dashboard. It is privacy-focused, making no network requests and never exposing user data.

Hacker News AIAgentsIn-site article
AI hasn't killed our bootstrapped enterprise software company yet

NocoBase doubled revenue in first five months of 2026 year-over-year, after a major pivot from a no-code platform to AI + no-code infrastructure. The article recounts the team's journey from panic to repositioning, emphasizing that enterprise-grade software still needs standardized architecture, visual configuration, and robust design in the AI era.

Hacker News AIAgents / ResearchIn-site article
Cloudflare CAPTCHA on at least one ampersand

Simon Willison uses Cloudflare's Managed Challenge to protect his faceted search from aggressive crawlers, but even simple ?q=term searches triggered the challenge. Using Claude Code, he discovered a rule that only triggers CAPTCHA for search URLs containing at least one ampersand, allowing simple searches to pass through without challenge.

Simon Willison's WeblogToolsIn-site article
Databricks announces 2026 global partner awards

Databricks recognized over 60 consulting, system integrator, and ISV partners at the Data + AI Summit across global, regional, industry, product, and champion categories. The awards highlight a focus on AI transformation, lakehouse modernization, Unity Catalog governance, and agentic AI at enterprise scale. Notable winners include Accenture, Deloitte, NVIDIA, Anthropic, and Salesforce.

Databricks BlogAgents / ChipsIn-site article
Beyond Transcription: ASR Model Delivers Words, Emotion, and Intent in 200ms

Whissle's META-1 is a meta-aware ASR model that simultaneously outputs transcription and metadata (emotion, intent, age, gender, etc.) in a single forward pass at ~200ms latency. By integrating KenLM n-gram language models, it reduces word error rates by up to 3.6% absolute (10.8% relative) across four languages, while extracting metadata 9x faster than commercial alternatives like Deepgram, AssemblyAI, and Gemini 2.0 Flash.

Hacker News AIChips / ResearchIn-site article
Show HN: HashMeterAi – Private AI Token Real Usage Meter for All Models

HashMeterAi is a local-first, private usage meter for AI coding tools. It tracks token usage across multiple tools like Claude Code, Codex, Kimi, Qwen CLI, and others, providing a unified dashboard with metrics like cost, processed tokens, focus time, and AI persona. It is 100% offline and privacy-focused, never sending data. Supports several tools via local transcripts.

Hacker News AIAgents / PolicyIn-site article
Prism: AI Companion for macOS, an All-in-One AI Workspace

Prism is a native macOS AI companion that supports multiple AI model providers, featuring a Quick AI panel, browser automation, study tools, and a focus on privacy. It offers a permanent free tier and paid plans, now launching on Product Hunt.

Product Hunt AIAgents / ResearchIn-site article
Meta CTO Andrew Bosworth Admits the Company's AI Reorg Was 'Atrocious'

Meta's chief technology officer acknowledged that the restructuring of the Applied AI engineering unit of about 6,500 employees was poorly communicated and executed, leading to a loss of trust and morale. He promised improvements in management, caps on direct reports, and AI coaching tools, while also allowing drafted employees to seek other roles within the company.

Hacker News AIAgentsIn-site article
Shelly

Shelly runs the OpenAI Codex CLI natively on Android — no PC, Termux, or proot needed. It features a native terminal, an Agent Chat pane for one-tap fixes, and a home-screen widget showing quota, cost, and rate limits. Local LLMs are supported. Open source under GPLv3, built entirely by directing AI agents.

Product Hunt AIAgentsIn-site article
Katalyst

Katalyst is an AI sales agent for Salesforce teams that automatically summarizes calls, updates records, drafts follow-ups, and surfaces key signals, saving reps time and improving pipeline management. It integrates natively with Salesforce and offers a free trial.

Product Hunt AIAgents / ResearchIn-site article
67% of AI-generated commands are unsafe. We tested it

A test of Google's Gemini 3 Flash Preview as an autonomous AI agent showed that 67% of generated curl commands targeted unsafe internal networks or metadata endpoints. All dangerous commands were blocked by Check, a preflight security tool. The test highlights the risk of AI agents executing commands without guardrails.

Hacker News AIAgents / PolicyIn-site article
The efficiency-gain illusion: People underestimate the rate of AI use

A new study finds that people frequently use AI for simple tasks even when it is inefficient. The research reveals miscalibration at two levels: underestimation of actual AI use and overestimation of its benefits, along with a carryover effect that entrenches these biases.

Hacker News AIResearchIn-site article
Build Compliant AI Agents with Stateful Stream Processing

EU AI Act obligations for high-risk systems hit in August 2026. Stateless agent frameworks can't satisfy them. This guide covers seven types of state compliant agents must maintain, four streaming patterns for auditability, and a reference architecture using Kafka and Flink as the control plane.

Hacker News AIAgents / PolicyIn-site article
Sakana AI Commercializes AB-MCTS in Sakana Marlin, an Enterprise Agent Generating Up to 100-Page Research Reports With Slides

Tokyo-based Sakana AI launched its first commercial product, Sakana Marlin, an autonomous research agent for enterprises. Each task runs up to eight hours, producing reports of dozens to 100 pages with slides. It leverages AB-MCTS and AI Scientist workflows. Pricing starts at pay-as-you-go with 100 credits per run at ¥98 per credit.

MarkTechPostAgents / PolicyIn-site article
Meta Employees Hate Zuckerberg's Plan for a Companywide AI Hackathon

Internal messages reveal Meta employees' frustration with Zuckerberg's planned companywide AI hackathon, citing lack of time due to layoffs, low morale, and distrust in management. The hackathon, set for July 14-16, excludes performance evaluation credit, further fueling discontent.

Hacker News AIPolicy / Research / StartupsIn-site article