Moonshot AI's 2.8-trillion-parameter open-weight model is rolling out across Copilot's Pro, Pro+, Max, Business, and Enterprise plans with usage-based billing at competitive rates.
xAI is releasing Grok 4.6 around August 7, with Grok 4.7 (the actual 2.1T model) following weeks later. Both matter for Cursor users now that SpaceX owns Anysphere.
Alibaba's biggest model yet went from preview to general availability on August 3. Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts model with a 1M-token context window and strong agent benchmarks — though independent evaluations haven't caught up yet.
Supabase Evals runs Claude Code, Codex, and OpenCode against actual Supabase tasks — building schemas, fixing Edge Functions, debugging RLS policies — in real environments with real scoring.
DeepSeek V4-Flash-0731 went official on July 31 with a completely retrained post-training pass. The result: it now outperforms V4-Pro-Preview on all nine published agent and coding benchmarks — at one-third the price.
Responding to accusations of opposing open-source AI, Anthropic's CEO clarified the company's actual position: open-weight models are a public good, with one serious exception.
Claude Opus 5, Anthropic's newest flagship model, is now available in GitHub Copilot on Pro+, Max, Business, and Enterprise plans. It targets agentic workflows where careful reasoning and multi-step execution matter, and carries usage-based pricing at the provider's list rate.
Moonshot AI released the full open weights for Kimi K3 on July 26-27. At 1.4TB in MXFP4 format, it's the largest open-weight model ever published. Together AI and Modal launched day-0 hosted access for teams without Blackwell hardware.
Claude Opus 5 arrived July 24 with the same pricing as Opus 4.8, a 1 million-token context window, and benchmark numbers that push past Claude Fable 5 on agentic and computer-use tasks. It scores 96% on SWE-bench Verified, solves every problem on IMO 2026, and includes a new effort dial that lets you trade cost for capability per request.
Three new Gemini models landed on July 21, led by Gemini 3.6 Flash — cheaper, faster, and more accurate than its predecessor. Meanwhile Google confirms Gemini 4 pretraining has started, and 3.5 Pro is still stuck in partner testing.
Gemini 2.5 Pro and Gemini 3 Flash are being retired from all Copilot experiences on July 31. Kimi K2.7 Code, the first open-weight model in Copilot, is now available for Business and Enterprise plans.
Thinking Machines Lab released Inkling, a 975B-parameter open-weight model trained on 45 trillion tokens across text, image, audio, and video. Former OpenAI CTO Mira Murati is betting enterprises want AI they can customize, not just rent.
Kimi K3 has 2.8 trillion total parameters, a 1M-token context window, and native vision. It launched July 16 on the Kimi app and API. Open weights arrive by July 27.
Weeks after the $60 billion acquisition closed, SpaceXAI and Cursor shipped Grok 4.5 — a model built for software engineering, legal work, and finance that's now available across all Cursor plans.
DeepSeek V4 brings two model tiers, 1M-token context by default, and a hard cutoff on July 24 for legacy API names that millions of developers still use.
Google DeepMind delayed Gemini 3.5 Pro from June to July 17 after abandoning the 2.5 Pro foundation model. The rebuilt version adds a 2M token context window, Deep Think reasoning, and improved math.
Moonshot AI's open-weight Kimi K2.7 Code model expanded to Copilot Business and Enterprise plans on July 7, but it's off by default and requires admin action to enable.
Moonshot AI's Kimi K2.7 Code reached general availability in GitHub Copilot on July 1, marking the first open-weight model available in the platform's model picker.
OpenAI's next model family breaks into three tiers: Sol (flagship), Terra (balanced), and Luna (cheap and fast). A US government request limited the initial rollout to around 20 vetted partners. General availability is expected in coming weeks.
Microsoft's 5B-parameter in-house coding model MAI-Code-1-Flash reached general availability for GitHub Copilot Business and Copilot Enterprise on June 26, with admin policy controls required before users can access it.
The 13-day free window for Claude Fable 5 on paid plans closed on June 22. Starting June 23, all Claude subscribers need usage credits to access the model. Here's what that means in practice and what Anthropic has said about restoring it.
Two model changes hit GitHub Copilot on June 18: Microsoft's MAI-Code-1-Flash expanded from VS Code to eight more surfaces, and GitHub announced Opus 4.6 (fast) will be deprecated across all Copilot experiences on June 29.
ESC
Start typing to search across tools, news articles, and reviews
No results found
Optional analytics and social embeds help us improve the site. It works fully without them.
Privacy
Privacy Settings
Choose which optional features you want to enable.
Your browser sent a global privacy opt-out signal. We respect that by default.