xAI, Anthropic, OpenAI, Moonshot AI Release 4 Frontier Models in July

Key Takeaways
  • xAI, Anthropic, OpenAI, and Moonshot AI each released frontier AI models in July 2026 targeting multi-hour task completion.
  • Grok 4.5 costs $2 input and $6 output per million tokens, more than 60% below Claude Opus 4.8 and GPT-5.5.
  • Moonshot AI committed to publishing Kimi K3 full open weights by July 27 after releasing the 2.8 trillion parameter model on July 16.

xAI, Anthropic, OpenAI, and Moonshot AI released four frontier AI models within three weeks in July 2026. xAI shipped Grok 4.5 on July 8, Anthropic released Claude Opus 5 on July 24, OpenAI moved GPT-5.6 to general availability on July 9, and Moonshot AI published Kimi K3 on July 16. Each model targets multi-hour task completion—from research to coding to structured reporting—without losing track of execution plans. All four carry context windows of 500,000 tokens or more, with three reaching 1 million tokens, enabling single sessions to process full code repositories or document collections.

xAI Releases Grok 4.5 With Cursor Training Data

xAI released Grok 4.5 on July 8. The model runs on a 1.5 trillion parameter base and was trained in part on real usage data from Cursor, the coding platform SpaceXAI acquired earlier this year. Pricing sits at $2 per million input tokens and $6 per million output tokens, with a 500,000 token context window. Elon Musk described Grok 4.5 as an Opus-class model that runs faster and at lower cost. Independent trackers place it fourth on the Artificial Analysis Intelligence Index, ahead of every open weight model, at pricing more than 60% below Claude Opus 4.8 and GPT-5.5. On Terminal-Bench 2.1, a test that scores how well a model completes command-line engineering tasks, Grok 4.5 scored 83.3%. xAI says the model needs roughly a fifth of the output tokens Opus 4.8 required for comparable tasks.

Elon X post screenshot.

OpenAI Launches GPT-5.6 in Three-Tier Pricing Structure

OpenAI moved GPT-5.6 to general availability on July 9 after a two-week preview limited to roughly 20 organizations vetted by the U.S. government, following an executive order tied to frontier model safety review. The family ships as three tiers: Sol, the flagship, priced at $5 input and $30 output per million tokens; Terra, a mid-tier model at $2.50 and $15 that OpenAI says matches GPT-5.5 at half the cost; and Luna, a fast, low-cost tier at $1 and $6. OpenAI reports Sol leads the Artificial Analysis Coding Agent Index and hits 88.8% on Terminal-Bench 2.1, rising to 91.9% when the model runs four sub-agents in parallel under its new ultra mode. All three tiers carry OpenAI's highest internal risk rating for cyber and biological misuse potential.

OpenAI X post screenshot.

Anthropic Ships Claude Opus 5 as Default Max Model

Anthropic released Claude Opus 5 on July 24, positioning it as a model that reaches close to the performance of Claude Fable 5 at half of Fable's $10 input and $50 output pricing. Opus 5 carries the same $5 input and $25 output pricing as its predecessor, Opus 4.8. The model ships with a 1 million token context window, 128,000 max output tokens, and an adjustable reasoning effort setting that ranges from low to a new xhigh mode. Anthropic says Opus 5 sets new marks on Frontier-Bench and GDPval-AA, two coding and knowledge work evaluations, though it trails the restricted Claude Mythos 5 model on cybersecurity tasks. Opus 5 is now the default model on Claude Max and the strongest option on Claude Pro.

Claude X post screenshot.

Moonshot AI Publishes Kimi K3 as Largest Open-Weight Model

Moonshot AI released Kimi K3 on July 16, a 2.8 trillion parameter mixture of experts model with 896 total experts and 16 active per task. The model carries a 1 million token context window and native multimodal input, with API pricing at $3 input and $15 output per million tokens. Moonshot has committed to publishing full open weights by July 27. K3 is the largest open-weight model released to date, roughly 75% bigger than the previous largest widely used open model. Independent trackers place K3 fourth among current frontier systems, behind Claude Fable 5 and GPT-5.6 Sol but ahead of Claude Opus 4.8.

Kimi Moonshot X post screenshot.

Extended Context Windows Enable Full-Repository Processing

Context windows below 200,000 tokens once forced developers to break large codebases or research packets into fragments. Every model in this group now runs at 500,000 tokens or beyond, with three of the four at 1 million, letting a single session hold a full repository or a stack of primary source documents. Systems from the 2023 and 2024 period often lost their plan after a handful of tool calls. The models released this month are built to sustain dozens of coordinated steps and recover when a tool call returns bad data instead of stalling out. Effort controls in Opus 5, tiered pricing in GPT-5.6, and the efficiency claims behind Grok 4.5 point toward the same goal: letting teams choose how much compute a task deserves instead of paying flagship prices for every request. The government-gated rollout of GPT-5.6 signals that oversight is now built into release schedules for the largest models. Anthropic's brief, government-directed suspension of Fable and Mythos access in June was an earlier version of the same pattern.

FAQ

What do the four AI models released in July 2026 have in common?

Grok 4.5, Claude Opus 5, GPT-5.6, and Kimi K3 all target multi-hour task completion—from research to coding to structured reporting—without losing track of execution plans. Each model carries a context window of at least 500,000 tokens, with three reaching 1 million tokens, enabling single sessions to process full code repositories or document collections without fragmentation.

Why did OpenAI release GPT-5.6 through a government-gated preview?

OpenAI moved GPT-5.6 to general availability on July 9 after a two-week preview limited to roughly 20 organizations vetted by the U.S. government, following an executive order tied to frontier model safety review. All three GPT-5.6 tiers carry OpenAI's highest internal risk rating for cyber and biological misuse potential, which triggered the added review.

How does Grok 4.5 pricing compare to Claude Opus 4.8?

Grok 4.5 is priced at $2 per million input tokens and $6 per million output tokens, more than 60% below Claude Opus 4.8 and GPT-5.5. xAI says the model needs roughly a fifth of the output tokens Opus 4.8 required for comparable tasks, which lowers the cost of running long agent sessions.

Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments