← Back to Hello, AI

Dispatches from the frontier

Weekly analysis, honest takes, and hidden gems. No engagement bait.

Gemini 3.6 Flash Ships — 3.5 Pro Still Waiting

Google GA'd Gemini 3.6 Flash on July 21 while 3.5 Pro stays partner-only. helloai still tracks Gemini 3.1 Pro — and arena already scores the new Flash near frontier.

Read article →

Kimi K3 Replaces GPT-5.6 Sol in helloai's Frontier Set

Moonshot's 2.8T Kimi K3 lands on the public API at $3/$15 with 1M context. helloai drops GPT-5.6 Sol and adds Moonshot under the six-model cap.

Read article →

ChatGPT Work vs Claude Cowork: The Agentic Seat War

OpenAI and Anthropic are no longer fighting only on LMArena — ChatGPT Work and Claude Cowork compete for the surface where finished work ships.

Read article →

Muse Spark 1.1 Replaces Qwen3.7-Max in helloai's Frontier Set

Meta ships Muse Spark 1.1 with a public API at $1.25/$4.25. helloai drops Qwen3.7-Max and adds Meta as its budget agentic pick at 1487 Elo.

Read article →

Grok 4.5 Ships as xAI's Default Chat Model

xAI launched Grok 4.5 on July 8 — 83.3% on Terminal-Bench 2.1, 4.2× token efficiency, and $2/$6 pricing. helloai's tracked Grok entry moves from 4.3 to 4.5.

Read article →

GPT-5.6 Sol Reaches General Availability

OpenAI ends the partner-only preview on July 9. GPT-5.6 Sol replaces GPT-5.5 in helloai's tracked set at the same $5/$30 rate — the upgrade path teams waited on since June 26.

Read article →

How Caching and Batching Cut Frontier API Costs by 90%

helloai's leaderboard shows nominal per-token rates. Prompt caching and batch APIs stack underneath — turning a $5/MTok flagship into $0.25 on repeated context.

Read article →

Ollama, Hermes, and GLM-5.2:cloud Are Rewriting the Margin Map

One command routes a full agent harness to MIT-licensed weights at Ollama Pro prices. That stack does not kill frontier labs, but it is eating their easiest revenue.

Read article →

Frontier Models Score Under 1% on ARC-AGI-3

ARC-AGI-3 launched March 25 with humans at 100% and frontier AI at 0.51%. GPT-5.5 and Opus 4.7 barely move the needle — exposing a gap arena Elo cannot see.

Read article →

Frontier Releases Now Run Through Government Gates

Fable 5 returns globally July 1 after an 18-day suspension; GPT-5.6 Sol ships only to vetted partners. June 2026 made frontier availability a policy variable, not just an engineering milestone.

Read article →