Category

Model Releases

Latest AI model releases, benchmarks, and comparisons. Stay up to date with every new model launch from OpenAI, Anthropic, Google, Meta, and more.

Anthropic’s IPO talks skip valuation for now
Aug 14

Anthropic’s IPO talks skip valuation for now

Anthropic’s early IPO meetings are high-level, with CFO Krishna Rao avoiding valuation talk so far.

Gemini 3.7 Flash arrives with faster coding gains
Aug 14

Gemini 3.7 Flash arrives with faster coding gains

Google released Gemini 3.7 Flash three weeks after its last model, with stronger coding, web dev, and business workflow scores.

August 2026 model rankings: Claude leads text, Kimi coding
Aug 12

August 2026 model rankings: Claude leads text, Kimi coding

August 2026 rankings show Anthropic leading text, Kimi leading coding, and multimodal models moving toward full-modal systems.

Qwen3.8-Max pushes Alibaba into the top tier
Aug 8

Qwen3.8-Max pushes Alibaba into the top tier

Alibaba’s Qwen3.8-Max-Preview arrives with 2.4T parameters, cheaper API pricing, and benchmark results that pressure Claude.

Qwen3.8-Max proves that agentic work is the real frontier
Aug 4

Qwen3.8-Max proves that agentic work is the real frontier

Qwen3.8-Max matters because broad, stable agentic work beats narrow benchmark wins.

Google Earth Should Not Ship AI Image Generation
Aug 4

Google Earth Should Not Ship AI Image Generation

Google Earth should not ship image generation without tighter spatial safeguards and human review.

Try Claude Opus 4.7 and read its benchmarks
Aug 3

Try Claude Opus 4.7 and read its benchmarks

Claude Opus 4.7 is live now, with stronger coding, better honesty, and higher token use.

Opus 5 lets you cut cost without losing quality
Aug 2

Opus 5 lets you cut cost without losing quality

I break down Claude Opus 5’s pricing and performance claims into a copy-ready rollout template.

OpenAI Cuts GPT-5.6 Prices as AI Bills Climb
Aug 1

OpenAI Cuts GPT-5.6 Prices as AI Bills Climb

OpenAI slashed prices for GPT-5.6 Luna and Terra, signaling a sharper fight on AI cost efficiency.

Opus 5 proves premium AI is becoming a commodity
Jul 31

Opus 5 proves premium AI is becoming a commodity

Anthropic's Opus 5 shows frontier performance is now arriving at half the price.

OpenAI Gives Scientists Free GPT-5.6 Access
Jul 31

OpenAI Gives Scientists Free GPT-5.6 Access

OpenAI is giving 100,000 academics free GPT-5.6 access, but the model weights stay closed and independent audits remain limited.

Google ships Gemini 3.6 Flash and 3.5 Lite
Jul 27

Google ships Gemini 3.6 Flash and 3.5 Lite

Google has rolled out Gemini 3.6 Flash and 3.5 Flash-Lite across AI Studio, Android Studio, Gemini App, Enterprise, and Search.

Kimi K3 Is Forcing Silicon Valley to Pick Sides
Jul 27

Kimi K3 Is Forcing Silicon Valley to Pick Sides

Kimi K3’s release triggered a split in US AI circles over open weights, China’s model pace, and the economics of frontier labs.

Opus 5 lets you ship with fewer refusals
Jul 26

Opus 5 lets you ship with fewer refusals

I break down Anthropic’s Opus 5 and the practical fallback pattern that keeps API calls useful when safety checks trip.

Claude Opus 5 undercuts Fable 5 on price
Jul 25

Claude Opus 5 undercuts Fable 5 on price

Anthropic’s Claude Opus 5 beats Fable 5 on coding and knowledge tasks while cutting token prices in half.

OpenAI model catalog adds GPT-5.6 pricing tiers
Jul 23

OpenAI model catalog adds GPT-5.6 pricing tiers

OpenAI updated its model catalog with GPT-5.6 Sol, Terra, and Luna, plus shared multimodal support and Responses API access.

Gemini 3.6 Flash proves Google is betting on efficiency over hype
Jul 22

Gemini 3.6 Flash proves Google is betting on efficiency over hype

Google’s Gemini 3.6 Flash and 3.5 Flash-Lite show the company is optimizing for cheaper, faster models before Gemini 4.

Kimi K3 handles an 820k-line Rust codebase
Jul 22

Kimi K3 handles an 820k-line Rust codebase

A Zhihu test says Kimi K3 analyzed 820,000 lines of Rust from Grok Build and decoded an XOR-obfuscated system prompt.

GPT-5.6 arrives in three variants with lower token costs
Jul 17

GPT-5.6 arrives in three variants with lower token costs

OpenAI’s GPT-5.6 ships in three tiers, cuts coding token use, and posts strong benchmark gains across coding and terminal tasks.

GPT-5.6 Sol, Terra, Luna arrive on DigitalOcean
Jul 17

GPT-5.6 Sol, Terra, Luna arrive on DigitalOcean

OpenAI’s GPT-5.6 Sol, Terra, and Luna are now available on DigitalOcean Serverless Inference, with new reasoning modes and per-token pricing.

Grok 4.5’s rise comes down to 5 numbers
Jul 17

Grok 4.5’s rise comes down to 5 numbers

Grok 4.5 is posting strong benchmark results, faster release cadence, and lower pricing as xAI pushes into agentic AI.

Grok 4.5 turns agent work into one prompt
Jul 17

Grok 4.5 turns agent work into one prompt

I break down Grok 4.5’s coding, agent, and office workflow claims into a copy-ready prompt and rollout template.

Kimi API quickstart adds K2.7 Code and Highspeed
Jul 16

Kimi API quickstart adds K2.7 Code and Highspeed

Kimi API Platform now ships K2.7 Code and a highspeed variant with 256K context, OpenAI compatibility, and multimodal input.

GPT-Live brings faster voice chat to ChatGPT
Jul 16

GPT-Live brings faster voice chat to ChatGPT

OpenAI is rolling out GPT-Live in ChatGPT, giving paid users GPT-Live-1 and pushing voice interaction closer to Doubao and Gemini.

Anthropic extends Claude Fable access after GPT-5.6
Jul 15

Anthropic extends Claude Fable access after GPT-5.6

Anthropic extended Claude Fable 5 access through July 19 after OpenAI’s GPT-5.6 launch intensified the model rivalry.

GPT-5.6 Sol Review: Faster Coding, Lower Cost
Jul 14

GPT-5.6 Sol Review: Faster Coding, Lower Cost

GPT-5.6 Sol is faster on coding benchmarks, cheaper than Claude Fable 5, and under scrutiny for benchmark gaming.

GPT-5.6 family lands with Luna, Terra, Sol
Jul 14

GPT-5.6 family lands with Luna, Terra, Sol

OpenAI shipped GPT-5.6 in three sizes, with Sol posting 53.6 on Agents’ Last Exam and new API features for tool use and subagents.

GPT-5.6 benchmarks: Sol tops coding, cuts costs
Jul 14

GPT-5.6 benchmarks: Sol tops coding, cuts costs

Artificial Analysis says GPT-5.6 Sol nears Claude Fable 5 on intelligence, leads coding tests, and adds cache-write pricing.

GPT-5.6 turns OpenAI into a model menu
Jul 12

GPT-5.6 turns OpenAI into a model menu

I break down OpenAI’s GPT-5.6 rollout, the three-model split, and the copyable way to pick the right model per task.

Seedream 5.0 Pro Is the Right Choice for Editable AI Images
Jul 10

Seedream 5.0 Pro Is the Right Choice for Editable AI Images

Seedream 5.0 Pro is the best pick for reasoning-driven, editable AI image workflows with multilingual text.

You've reached the end