Announced at Code with Claude London on May 19, self-hosted sandboxes let enterprises run agent tool execution on their own infrastructure. MCP tunnels let agents reach private servers without opening firewall ports.
Cursor shipped 3.4 on May 13 and 3.5 on May 20, adding full-screen agent tabs, Dockerfile-based cloud development environments, expanded Automations with multi-repo support, and a native Jira integration.
Cursor released Composer 2.5 on May 18, built on Moonshot's Kimi K2.5 checkpoint with 85% of compute spent on Cursor's own post-training pipeline. It scores 63.2% on CursorBench v3.1, edging out both Opus 4.7 and GPT-5.5 at a fraction of the inference cost.
Gemini Spark launched at I/O 2026 as a proactive cloud agent that handles tasks across Gmail, Docs, and Slides while you're offline. The detail buried in the announcement: custom sub-agents and a local browser are coming. Here's what shipped.
Google added Managed Agents to the Gemini API today. A single API call provisions an agent with reasoning, tool use, and code execution inside an isolated Linux environment, powered by the Antigravity harness and Gemini 3.5 Flash.
The largest maintenance release in recent weeks adds /resume support so background sessions appear alongside interactive ones, per-session model switching, and a fix for the startup hang that could freeze Claude Code for 75 seconds behind a VPN or captive portal.
Two May 13 updates: JetBrains IDE users can now delegate tasks to a locally running Copilot CLI agent with worktree or workspace isolation, while a new REST API lets teams trigger cloud agent tasks programmatically.
OpenAI added Codex access to the ChatGPT iOS and Android apps on May 14. Your phone connects to a Codex session running on your Mac via QR code, letting you review outputs, approve commands, and start new tasks from anywhere.
xAI launched an early beta of Grok Build on May 15, a terminal-based agentic coding CLI powered by Grok 4.3. It's available now to SuperGrok Heavy subscribers at an introductory price, with plans for broader access.
Windsurf added Claude Opus 4.7 in fast mode to the editor on May 12, offering Opus-level intelligence at roughly 2.5x the output speed. A week earlier, version 2.2.17 extended Devin Review to all Pro, Max, and Teams subscribers with a two-week free trial.
GitHub's new native desktop app lets developers start agentic coding sessions directly from issues and pull requests, with each session running in an isolated branch and workspace. The app shipped v0.2.4 on May 15 with queued messages and collapsible tool-call panels.
Claude Code 2.1.140 shipped on May 12 with smarter Agent tool subagent_type matching, an updated color palette for the agent view added in 2.1.139, and fixes for the /goal command, settings hot-reload, background service startup on enterprise machines, and a Windows event-loop stall.
Claude Code 2.1.139 shipped May 11 with agent view — a unified dashboard for every running, waiting, and finished session — and a /goal command that keeps Claude working until a condition you define is met.
OpenAI released Codex CLI 0.130.0 on May 8 with a new codex remote-control command that starts a headless, remotely controllable app-server — and a GitHub discussion thread confirms ChatGPT mobile is the intended controller.
Zed shipped version 1.0 on April 29, reaching the milestone with cross-platform support, a new Agent Client Protocol backed by Google and JetBrains, and a Zed for Business tier.
Anthropic's 2026 Agentic Coding Trends Report, published in late April, identifies eight shifts reshaping software engineering. The most telling number: engineers use AI in roughly 60% of their work but say they can fully delegate only 0–20% of tasks.
ServiceNow made Build Agent generally available at Knowledge 2026, extending it into the four major AI coding tools so developers can build ServiceNow apps without leaving their preferred IDE.
Coder Technologies released Coder Agents in beta on May 6 — a self-hosted, model-agnostic AI coding agent platform for enterprises that need to keep code and prompts inside their own infrastructure.
AWS and OpenAI expanded their partnership on April 28, putting Codex, GPT-5.5, and a new Managed Agents service into Amazon Bedrock for enterprise teams that want OpenAI's models inside AWS infrastructure.
Novee researchers disclosed a high-severity RCE vulnerability in Cursor IDE on April 28. A crafted Git repository with a hidden pre-commit hook can trigger arbitrary code execution when Cursor's AI agent runs routine git operations. Cursor patched it in February 2026.
Amazon announced on April 30 that Amazon Q Developer IDE plugins and paid subscriptions will reach end of support on April 30, 2027. New signups are blocked starting May 15, 2026. Kiro, AWS's spec-driven agentic coding environment, is the intended replacement.
Moonshot AI released Kimi K2.6 on April 20. The open-weight model scores 80.2% on SWE-Bench Verified and 66.7% on Terminal-Bench 2.0, with pricing of $0.60 per million input tokens on the official API. It's a meaningful upgrade over K2.5 on every benchmark that matters for coding agents.
Cursor Security Review enters beta on Teams and Enterprise plans with two always-on agents: a Security Reviewer that comments on every PR, and a Vulnerability Scanner that runs scheduled codebase sweeps.
Microsoft's Agent 365 went generally available today, May 1, as part of the new Microsoft 365 E7 tier. It's a central control plane for registering, governing, and auditing AI agents across Microsoft, Copilot, and third-party platforms.
ESC
Start typing to search across tools, news articles, and reviews
No results found
Optional analytics and social embeds help us improve the site. It works fully without them.
Privacy
Privacy Settings
Choose which optional features you want to enable.
Your browser sent a global privacy opt-out signal. We respect that by default.