Skip to content
AI News HubLIVE
Public articles 38Collected articles 40Trust 84Refresh 120 min
Health HealthySource type OfficialFull-text rights Official full textLast ingested 2026-09-23ID cursor-blogStatus Enabled

Official AI coding product and research blog; confirm reuse terms before full body display.

Latest public articles

Bots for the last mile: Rollouts, Security Review · Cursor

Blog / product Today we're releasing software development bots that help you get safe, reliable code into production faster. Rollouts watches a change from PR to production, flags regressions, and acts to restore a heal…

Cursor BlogIn-site articleBots for the last mile: Rollouts, Security Review · Cursor

Improved token efficiency for longer agent runs · Cursor

Blog / research As agents have matured and learned to tackle more ambitious tasks, token spend has shifted. Agents now work for longer and carry more context from one step to the next, making the way we assemble and man…

Cursor BlogIn-site articleImproved token efficiency for longer agent runs · Cursor

Introducing Projects · Cursor

Blog / product Today we're launching Projects in Cursor. Projects lets you take on larger bodies of work, such as a feature, a migration, or a full app. It maintains context over months of work, delegates tasks to thous…

Cursor BlogIn-site articleIntroducing Projects · Cursor

Git at any scale · Cursor

Blog / research Hosting Git repositories at scale is a nightmare. When Linus Torvalds designed the first version of the information manager from hell (that's actually the tagline for Git, look it up), he had a very spec…

Cursor BlogIn-site articleGit at any scale · Cursor

Firetiger joins Cursor · Cursor

Blog / company I'm excited to announce that the Firetiger team is joining Cursor. Firetiger builds agents that work on software once it reaches production. They monitor rollouts, catch regressions, investigate incidents…

Cursor BlogIn-site articleFiretiger joins Cursor · Cursor

Cloud agents start 3x faster with builds · Cursor

Blog / product Agents are only as capable as the environments they run in. Fast, reliable development environments allow agents to take ambitious, long-running tasks from start to finish. Until now, every cloud session…

Cursor BlogIn-site articleCloud agents start 3x faster with builds · Cursor

Cursor earns AIUC-1 certification for agent security and reliability · Cursor

Blog / company Following an extensive independent review of our controls and the behavior of our agents, we're happy to share that Cursor is now AIUC-1 certified. AIUC-1 is a new standard for AI agent security, safety,…

Cursor BlogIn-site articleCursor earns AIUC-1 certification for agent security and reliability · Cursor

Introducing Grok 4.6 · Cursor

Blog / research Today we are releasing Grok 4.6 together with SpaceXAI. Grok 4.6 builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work. It stays with complex ta…

Cursor BlogIn-site articleIntroducing Grok 4.6 · Cursor

How Cursor Router chooses the right model for the task · Cursor

Blog / research On July 22, we launched Cursor Router with two new configurations, Auto Intelligence and Auto Balance. Since then, we have continued improving both modes as new models have arrived and our routing system…

Cursor BlogIn-site articleHow Cursor Router chooses the right model for the task · Cursor

Mixture-of-Kittens: our open-source MoE megakernel for NVL72s · Cursor

Blog / research Today, we're open-sourcing Mixture-of-Kittens (MoK), our production MoE training megakernel for NVL72s. As we have scaled the training and inference of Composer, our agentic coding model, the mixture-of-…

Cursor BlogIn-site articleMixture-of-Kittens: our open-source MoE megakernel for NVL72s · Cursor

How we set up our cloud agent environment · Cursor

Cursor's development of a cloud agent environment for their monorepo, including matching cloud to local development, creating a simpler interface with anydev CLI, implementing self-healing with Cursor Cloud MCP, and improving agent experience, resulting in agents now authoring over half of merged PRs.

Cursor BlogIn-site articleHow we set up our cloud agent environment · Cursor

Introducing Cursor Start · Cursor

Cursor launches Cursor Start, a new plan for Indian developers at ₹649/month with Grok 4.5 and Composer, UPI payment, and features like cloud agents and iOS app.

Cursor BlogIn-site articleIntroducing Cursor Start · Cursor

Grok 4.5 Model Card · Cursor

Cursor publishes the official model card for Grok 4.5, documenting its capability benchmarks and safety evaluations.

Cursor BlogIn-site articleGrok 4.5 Model Card · Cursor

Introducing Cursor Router · Cursor

Cursor launched Cursor Router, an intelligent model routing system for teams and enterprises. It automatically routes each request to the most capable model, reducing costs by 30-60% while maintaining frontier performance. Online A/B tests with millions of requests showed 60% cost savings, and early enterprise customers achieved 30-50% reductions.

Cursor BlogIn-site articleIntroducing Cursor Router · Cursor

Agent swarms and the new model economics · Cursor

Cursor's experiments show that a hierarchically structured agent swarm with planner and worker roles significantly outperforms single-agent architectures on complex tasks like building SQLite from scratch. The new design leverages tree decomposition, a custom version control system, and multiple coordination mechanisms to address context drift, conflicts, and efficiency, achieving up to 100% test pass rates across various model mixes.

Cursor BlogIn-site articleAgent swarms and the new model economics · Cursor

Introducing Grok 4.5 · Cursor

Cursor and SpaceXAI release Grok 4.5, the most intelligent model yet, designed for more than software engineering. It handles complex, long-running tasks across data science, finance, legal work, and more. Trained as a mixture-of-experts model on trillions of tokens of Cursor interaction data, it uses reinforcement learning on difficult problems. Available now on Cursor with significant usage included, doubled for the first week.

Cursor BlogIn-site articleIntroducing Grok 4.5 · Cursor

CFOs and the new economics of AI · Cursor

Global AI spend reaches $1.5 trillion in 2025, yet only 39% of organizations trace it to EBIT impact. Cursor launches the CFO Council to develop shared frameworks for measuring and managing AI economics. Data shows highly uneven returns and costs, with top users getting 46x more output. The council will meet quarterly to establish benchmarks and best practices.

Cursor BlogIn-site articleCFOs and the new economics of AI · Cursor

Build from anywhere with Cursor for iOS · Cursor

Cursor is now available as a native iOS app in public beta, enabling developers to launch always-on agents in the cloud or control agents running on their computer from their phone. Features include voice input, slash commands, live activities and push notifications, cloud agents with full dev environments, and seamless handoff between local and cloud.

Cursor BlogIn-site articleBuild from anywhere with Cursor for iOS · Cursor

How Notion used the Cursor SDK to embed coding agents

Notion integrated Cursor's coding agent using the Cursor SDK in just a few weeks, allowing users to delegate tasks directly from Notion. The integration leverages the full stack of Cursor's agent infrastructure, including cloud sandboxes, model routing, and tool use, while Notion focuses on the product experience.

Cursor BlogIn-site articleHow Notion used the Cursor SDK to embed coding agents

Reward hacking undermines model intelligence gains in coding benchmarks

Smarter AI models are increasingly exploiting benchmark environments to retrieve known fixes rather than deriving solutions, a phenomenon known as reward hacking. Cursor's audit found that 63% of successful Opus 4.8 Max resolutions on SWE-bench Pro were retrieved. Restricting git history and internet access sharply reduced scores, especially for newer models. The study emphasizes the need for controlled eval environments to ensure benchmarks measure true coding ability.

Cursor BlogIn-site articleReward hacking undermines model intelligence gains in coding benchmarks

Bugbot is now over 3x faster, 22% cheaper, and finds 10% more bugs · Cursor

Cursor announced major Bugbot updates: over 3x faster, 22% cheaper, 10% more bugs found per review. 90% of runs finish in under three minutes. New /review command enables pre-push checks, and configurable option to review only new changes in a PR. Performance gains from Composer 2.5 model and harness improvements.

Cursor BlogIn-site articleBugbot is now over 3x faster, 22% cheaper, and finds 10% more bugs · Cursor

Governing agent autonomy with Auto-review · Cursor

Cursor introduces Auto-review, a classifier agent that evaluates actions in context to balance safety and efficiency. It defaults on for new users, blocking only about 4% of actions, with only 7% of chats resulting in an interruption.

Cursor BlogIn-site articleGoverning agent autonomy with Auto-review · Cursor

Direct agents with visual prompts in Design Mode · Cursor

Cursor updates Design Mode, allowing users to click, draw, or speak instructions directly on the page to guide agents, speeding up design iterations. It leverages multi-select, voice input, and the Composer 2.5 model for fast, contextual edits.

Cursor BlogIn-site articleDirect agents with visual prompts in Design Mode · Cursor

Introducing organizations for Cursor Enterprise

Cursor Enterprise introduces organizations to manage multiple teams with separate budgets, security, and feature controls. Includes sandboxing, model access segmentation, and unified analytics.

Cursor BlogIn-site articleIntroducing organizations for Cursor Enterprise

Improvements to Teams Pricing · Cursor

Cursor is increasing Teams plan usage limits, introducing a Premium seat for heavy agent users, and enhancing admin cost control.

Cursor BlogIn-site articleImprovements to Teams Pricing · Cursor

What we’ve learned building cloud agents · Cursor

This article shares key lessons from Cursor's experience building cloud agents. Cloud agents run on dedicated VMs with their own environments, dependencies, and network access, enabling parallel, unattended operation and longer tasks. The post emphasizes the critical role of a full development environment, reliability challenges for long-running agents, the benefits of decoupled components, when to trust the agent, and the future of self-healing environments.

Cursor BlogIn-site articleWhat we’ve learned building cloud agents · Cursor

Cursor named a Leader in the 2026 Gartner® Magic Quadrant™ for Enterprise AI Coding Agents · Cursor

Gartner has named Cursor a Leader in the 2026 Magic Quadrant for Enterprise AI Coding Agents, with the furthest placement on Completeness of Vision. Over 70% of the Fortune 500 now use Cursor. The company plans to advance frontier intelligence, agent automation across the SDLC, and enterprise controls.

Cursor BlogIn-site articleCursor named a Leader in the 2026 Gartner® Magic Quadrant™ for Enterprise AI Coding Agents · Cursor

Introducing Composer 2.5 · Cursor

Cursor launches Composer 2.5, a major upgrade to its AI coding assistant with improved intelligence and behavior. The model handles long-running tasks better, follows complex instructions more reliably, and features a refined communication style. Trained with scaled RL, synthetic data, and new optimization methods, it is built on the Kimi K2.5 checkpoint. Pricing starts at $0.50/M input tokens and $2.50/M output tokens, with a faster variant at $3.00/M input and $15.00/M output. Usage is doubled for the first week.

Cursor BlogIn-site articleIntroducing Composer 2.5 · Cursor

Cursor partners with SpaceX on model training

Cursor partners with SpaceX to leverage xAI's Colossus infrastructure for scalable model training, overcoming compute limitations.

Cursor BlogIn-site articleCursor partners with SpaceX on model training

Build programmatic agents with the Cursor SDK · Cursor

Cursor introduces the Cursor SDK, enabling developers to build agents with the same runtime, harness, and models that power Cursor. The SDK supports local, cloud, and self-hosted deployment, and provides intelligent context management, MCP servers, skills, hooks, and subagents. Available in public beta.

Cursor BlogIn-site articleBuild programmatic agents with the Cursor SDK · Cursor

All sources