Cloudskill is a platform that governs AI skills, turning scattered skill files into a managed catalogue with version control, per-person access policies, and a full audit log. It integrates with agents like Claude, Cursor, and Copilot, ensuring every change is reviewed and approved, keeping skills safe and consistent.
Live AI News Intelligence
Live monitoring
Live updates
Trusted sources, attribution, rights, and in-site reading distilled into a signal-first AI brief.
Live updates
Anthropic changes its policy on Claude Fable 5 after backlash, making safeguards for frontier LLM development visible. Previously, the model would limit effectiveness without notifying users. Now flagged requests visibly fall back to Opus 4.8, and API returns reasons for refusal.
A quiet day reflects on a great essay by Sarah Guo discussing open models, the difference between model labs and agent labs, and the untrainable aspects of AI. The article also covers Anthropic's Fable/Mythos rollout and the trust backlash, Fable 5's benchmark strength, Google's DiffusionGemma release, agent tooling progress, and technical updates in optimization, retrieval, and scientific modeling.
This article argues that despite fears, AI has not led to mass layoffs in software engineering. It presents evidence that layoffs attributed to AI are often financial in nature, and that AI compresses execution but not decision-making and delivery. The 'decide-execute-deliver sandwich' model explains why coding agents haven't displaced workers: the bottlenecks are deciding, verifying, and deep understanding.
Frontier teams are not just using AI to code faster. They’re redesigning how software gets built. The result is 4.5x productivity gains, in some cases more than 10x. This article details case studies from Amazon Bedrock, Prime Video, and others, outlining five key practices to become a frontier team, emphasizing that workflow transformation matters more than tools.
Ollama's MLX engine has been updated to deliver its highest performance on Apple Silicon yet. By leaning more heavily on Apple's unified memory and the Metal-backed MLX framework, models output higher quality responses, respond faster, and use less memory. The update includes support for NVFP4 format, up to 20% faster output, and a snapshot system for agent workflows.
This article is the second part of the PyTorch profiling series, delving into the internals of nn.Linear layers, including transpose operations, bias-fused epilogue techniques, and the impact of torch.compile on a single linear layer. It then dissects the performance characteristics of a Multilayer Perceptron (MLP) with GeGLU activation, showcasing the scheduling and execution of GPU kernels.
Discover how astrophysicist Chi-kwan Chan uses Codex to build black hole simulations, helping scientists study extreme physics and test Einstein’s theory of general relativity.
OpenAI plans to acquire Ona to expand Codex with secure, persistent cloud environments, enabling long-running AI agents across enterprise workflows.
Learn how BBVA scaled ChatGPT Enterprise to 100,000 employees and partnered with OpenAI to accelerate AI-powered banking transformation worldwide.
OpenAI supports the EU Code of Practice on AI content transparency, advancing provenance standards and tools to help people understand AI-generated content.
Release of datasette-agent 0.2a0 with new user interaction and query saving capabilities.
We implement an instrumented workflow for Microsoft SkillOpt end to end. We set up the repository, connect OpenAI-compatible model access, and configure the optimizer and target models. We evaluate the original seed skill as a baseline, then run a real optimization loop with rollout, reflection, aggregation, selection, updating, and validation-based gating. We inspect training history, visualize accuracy, edit-budget behavior, and token usage, then compare the evolved skill against the baseline.
PixelForge is an AI-powered tool that converts a single photo into a recognizable RPG character with a 4-direction sprite pack (4x4 sheet, 16 transparent PNG frames, walk GIFs) ready for game engines like Godot and Unity. One-time $5 fee, no account or subscription required. Created by Bernard Huang, launched on Product Hunt.
Google has released DiffusionGemma, a new open-weight model under Apache 2 license, available for free via NVIDIA's NIM cloud API. It delivers impressive generation speeds exceeding 500 tokens per second.
Oracle Cloud customers can now access OpenAI models and Codex using their existing cloud commitments, enabling AI development with enterprise security and governance.
Anthropic released two new models, Claude Mythos 5 and Claude Fable 5, showing significant coding improvements but limited progress in image understanding. Testing reveals Fable 5 and GPT-5.5 can solve many vision tasks that stumped last year's models, yet geometric reasoning remains at the level of young children, suggesting general AI is still far off.
Google released DiffusionGemma, a 26-billion-parameter model that generates text via diffusion, achieving 1,000 tokens per second on an H100 GPU—four times faster than autoregressive models, but with lower quality. It's currently experimental.
As robotaxi services expand globally, NVIDIA introduces Halos OS—a comprehensive safety system integrating certified OS, standardized interfaces, AI guardrails, and a validation framework to ensure safety is built into autonomous vehicles from the ground up.
DiffusionGemma is Google DeepMind's experimental open text generation model that uses text diffusion instead of standard autoregressive decoding, achieving up to 4x faster generation on dedicated GPUs. The 26B MoE model (3.8B active parameters) is built on the Gemma 4 backbone, supports multimodal inputs (text, image, video), has a 256K context window, covers 140+ languages, and is released under Apache 2.0.
Anthropic released Claude Fable 5, its most powerful public AI model, but it refuses to answer basic biology questions like 'what are mitochondria' due to strict safety guardrails designed to prevent misuse for bioweapons. The company admits this is overly conservative but necessary for safe deployment.
Sam Altman told employees he expects an OpenAI IPO "within the next year," but a delay to 2027 is possible. He frames it as caution around self-improving AI, though Anthropic's stronger growth numbers and imminent IPO may be the real reason to wait.
New college graduates around the country have been booing and heckling commencement speakers who hype up AI. Microsoft would like everyone to talk it out. In a blog post, Brad Smith addressed the backlash, suggesting it's a wake-up call, but the substance echoes the same pro-AI arguments that sparked the booing.
The Verge's Regulator newsletter returns to a chaotic Washington landscape, covering the Washington AI Network gala, Pope Leo XIV's AI encyclical, and the unpredictable nature of AI regulation under Trump. The piece highlights the industry's dilemma in navigating partisan politics and the upcoming midterm elections, where AI is becoming a key voter issue.
Google researchers introduce Regularized f-Divergence Kernel Tests to audit machine unlearning and privacy. The framework adaptively selects optimal divergence measures, improving detection of data leaks and unlearning failures while requiring fewer samples and less tuning.
A group of independent musicians is suing Google, claiming it illegally used songs uploaded to YouTube to train its Lyria 3 model. Google has filed a motion to dismiss, arguing that the Terms of Service grant a broad license to use uploaded content. While Google hasn't explicitly confirmed using YouTube uploads for Lyria, past statements suggest it does.
Onpilot creates specialized AI workers tailored to your systems, workflows, and processes. It monitors operations, identifies risks, uncovers opportunities, recommends actions, and automates work across 3,000+ integrations. Deploy in Slack, Teams, WhatsApp, your SaaS, or on-premises. The platform emphasizes trust with approval workflows, audit trails, and exception handling.
Anthropic released Claude Fable 5, its first Mythos-class AI model. Microsoft restricts internal use due to new data retention requirements that retain prompts and outputs for 30 days (up to two years for flagged content). Other Claude models remain available under zero data retention. Legal teams are evaluating.
Google is making changes to how it saves your interactions with Search. In an email to users, Google says it will save images, files, audio, and video from Lens, Search Live, voice searches, and Translate under a new 'Search Services History' setting. Users can opt out and disable 'Save Media' if they prefer not to have these interactions stored. The data will be used to improve services, including AI models.
Google DeepMind released DiffusionGemma, an experimental open model for fast text generation using parallel token generation. NVIDIA optimized it to run faster on GeForce RTX, RTX PRO, and DGX Spark systems, achieving up to 1000 tokens/sec locally.