Meta Superintelligence Labs shipped a terminal coding agent in beta on August 5. It uses persistent background agents, a local event log for crash recovery, and a co-trained model called Muse Spark 1.2. Pricing starts at $1.25 per million tokens.
The first joint product from the SpaceX-Cursor merger puts autonomous agents on your Mac, iPhone, and PC. They get their own computer, sign into your tools, and keep working while you're doing something else.
Moonshot AI's 2.8-trillion-parameter open-weight model is rolling out across Copilot's Pro, Pro+, Max, Business, and Enterprise plans with usage-based billing at competitive rates.
Starting with models released after August 2, 2026, Anthropic embeds invisible text watermarks and C2PA metadata into all Claude outputs — worldwide, not just in the EU.
VS Code 1.132 released August 5 with Agent Host Protocol support, a dedicated Agents Window for running multiple concurrent agent sessions, side chats via /btw, and multilingual dictation using an on-device model.
Anthropic's self-hosted environments for Claude Code entered public beta on August 6. Team and Enterprise customers can now run coding agent sessions on their own infrastructure, keeping source code and credentials off Anthropic's servers.
Novee Security showed at Black Hat USA 2026 how a single malicious GitHub issue can trigger remote code execution through Claude Code, credential theft through Gemini CLI (CVSS 10.0), and persistent instruction poisoning in Codex.
Anthropic's introductory rate of $2/$10 per million tokens for Claude Sonnet 5 expires at the end of the month. Standard pricing of $3/$15 takes effect September 1, and the new tokenizer adds another layer of cost that most estimates undercount.
Zed's first stable release on the 1.14 branch lands with OS-level sandboxing for agent tools and a default keymap change that will trip up existing VSCode-style users: cmd-i and F5 no longer do what they used to.
Inference hooks, now in beta for Claude Enterprise, route every employee prompt through your organization's security server before Claude processes anything. Netskope, Palo Alto, Proofpoint, and Zscaler integrations ship on day one.
xAI is releasing Grok 4.6 around August 7, with Grok 4.7 (the actual 2.1T model) following weeks later. Both matter for Cursor users now that SpaceX owns Anysphere.
GitHub stopped accepting new Spark users on August 4 and will fully retire the AI app builder by month's end. The shutdown follows GitHub Models going dark July 30, and lands the same week Copilot got comment-triggered automations and reasoning-level controls.
Released as part of Cloudflare Agents Week on August 3, @cloudflare/computer is an open-source npm package that gives AI coding agents their own computer — a virtual filesystem backed by SQLite, with shell access, isolates, and full Linux containers on demand.
Cursor released Google Workspace plugins on August 3, giving coding agents direct access to Gmail, Drive, Calendar, Docs, and Sheets without leaving the editor.
The latest Claude Code release adds a Focus view that hides tool noise behind per-turn summaries, fixes two permission-check bypass vulnerabilities in Bash and PowerShell, and adds sandboxed credential masking on Linux and WSL.
A new public preview rolling out to most enterprise customers on August 3 lets AI admins grant specific Copilot models to individual teams rather than setting one policy for the whole organization. The Copilot Billing Preview app also retired today.
ChatGPT Atlas, OpenAI's standalone AI browser, goes offline in six days. Users need to manually export their bookmarks and data before the cutoff. The browser capabilities are being folded into the ChatGPT desktop app and Codex.
Alibaba's biggest model yet went from preview to general availability on August 3. Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts model with a 1M-token context window and strong agent benchmarks — though independent evaluations haven't caught up yet.
Zed's latest preview release restricts what the AI agent can do in the terminal and on the network using native OS sandbox mechanisms — Seatbelt on macOS, Bubblewrap on Linux. Also: Skip Hooks, undo/redo for file ops, and adaptive thinking toggles.
Supabase Evals runs Claude Code, Codex, and OpenCode against actual Supabase tasks — building schemas, fixing Edge Functions, debugging RLS policies — in real environments with real scoring.
DeepSeek V4-Flash-0731 went official on July 31 with a completely retrained post-training pass. The result: it now outperforms V4-Pro-Preview on all nine published agent and coding benchmarks — at one-third the price.
GitHub's July 2026 update to Copilot in Visual Studio ships a new SDK-based agent with shorter, more actionable responses, built-in .NET and Azure skills from Microsoft's own teams, a right-click code review, and organization-wide custom instructions.
VS Code 1.131 brings offline voice dictation across chat, editors, and terminals; live subagent monitoring showing model, elapsed time, and active tool; and a new hybrid Markdown editor inside the Agents window.
VS Code 1.130 extends the agent host architecture to Claude and Codex agents, adds AI-assisted tool approvals, compact multi-file diffs in the Agents window, credit usage visibility for Copilot Business users, and bundles TypeScript 7.
ESC
Start typing to search across tools, news articles, and reviews
No results found
Optional analytics and social embeds help us improve the site. It works fully without them.
Privacy
Privacy Settings
Choose which optional features you want to enable.
Your browser sent a global privacy opt-out signal. We respect that by default.