Weekly analysis, honest takes, and hidden gems. No engagement bait.
Grok Bot, launched August 11, is a cloud teammate with its own VM. The local grok CLI is what can actually operate a production checkout like helloai.
Read article →Qwen3.8-Max's launch-week arena lead is gone. Two weeks of votes drop it to 1481 Elo, five points under Gemini 3.1 Pro. The $2/$6 card did not change.
Read article →DeepSeek V4 Pro's hike is live today: $0.66/$1.98 off-peak, up from $0.435/$0.87. It remains helloai's cheapest tracked frontier API.
Read article →xAI ships Grok 4.6 today at the same $2/$6 as 4.5, five days after helloai dropped Grok. Arena votes are still too thin to reopen the slot.
Read article →helloai admits Qwen3.8-Max and DeepSeek V4 Pro, drops Grok 4.5 and Kimi K3, and adds gpt-oss-20b to the open-weight shelf under the six-model caps.
Read article →Meta ships Muse Spark 1.2 with Muse Code at the same $1.25/$4.25 rates. helloai bumps 1.1 to 1.2 at 1498 Elo — still the cheapest agentic pick.
Read article →Anthropic's $5/$25 Opus 5 beats its own $10/$50 Fable 5 on Frontier-Bench and ARC-AGI-3, raising an uncomfortable question about who really leads the Claude lineup.
Read article →Google GA'd Gemini 3.6 Flash on July 21 while 3.5 Pro stays partner-only. helloai still tracks Gemini 3.1 Pro — and arena already scores the new Flash near frontier.
Read article →Moonshot's 2.8T Kimi K3 lands on the public API at $3/$15 with 1M context. helloai drops GPT-5.6 Sol and adds Moonshot under the six-model cap.
Read article →OpenAI and Anthropic are no longer fighting only on LMArena — ChatGPT Work and Claude Cowork compete for the surface where finished work ships.
Read article →