AI, actually tested · weekly

I test AI tools so you don't have to read the docs.

Recent updates

updated 5 days agoread all →
Jul 9 · Models · the big one

Meta ships Muse Spark 1.1 and opens its first paid proprietary API

A 1M-context multimodal agentic model on a new OpenAI-compatible Meta Model API (US preview), reported near $1.25/$4.25 per 1M with $20 free credits. Point your OpenAI SDK base URL at it to A/B test.

read the source →

Changing AI Trends

read all →
// the big one

Verification is the bottleneck now.

Code-gen speed is no longer the constraint; human and pipeline review capacity is. "Agents don't work for us" now means "our verification pipeline can't absorb the volume."

02
Cost arbitrage graduated to core infra. Custom model routers are everywhere: Warp routers, Vercel AI Gateway, and Product Hunt winners. Cross-provider cost routing is now a funded category, not a hack.
03
The agent governance layer is maturing. Copilot MDM and OTel, the Codex "writes" approval mode, and reviewable diffs in Devin all race to make autonomous agents auditable and enterprise-controllable.
04
Model releases now reach app builders in hours, not months. GPT-5.6 and Grok 4.5 were usable in Cursor, Figma, and Codex the same week they shipped, propagation is near-instant.

The AI that actually mattered this week, in five minutes.

No hype. No 12-part funnels. Just the tools I've actually run.

one email a week · unsubscribe in one tap

From the channel

watch all on YouTube →
GPT-5.6 Just Beat Claude & Gemini (You Can't Use It)1:08WTF is AI Poisioning ?0:54You're Using Claude WRONG - 5 Free Skills1:10Anthropic's Fable 5 is Back : But with a Catch1:04The Future of AI work : Loop Engineering0:49
~/subscribe $