AI Models, agents, coding assistants, image, video, voice
The lane
AI models, agents, coding assistants, image/video/voice tools. The cross-category practitioner lane: what shipped, what works, what to do about it. Faster cadence than the depth categories. Top-of-funnel feeder that routes into build / design / growth.
Latest articlesWhat shipped, what broke, and what it really cost — from running production AI systems.

Jev vs GLM-5.3-Flash
Compare Jev with GLM-5.3-Flash through Privatemode Decisions for support routing, image inputs, latency and cost per correct decision.Sep 27, 2026AI
Claude Opus 5.5 vs Opus 5
Compare Claude Opus 5.5 and Opus 5 on token costs, cache savings, coding work, and the API changes to check before switching.Sep 23, 2026AI
GPT-6 Sol vs Luna
Compare GPT-6 Sol and Luna on task fit, the 20x token-price gap, shared limits, and the checks that show when Sol earns its higher cost.
Sep 23, 2026AI
Is GPT-6 Luna Free
GPT-6 Luna is free to try in the desktop app. See which plans and surfaces support it, what limits remain, and when API charges apply.Sep 22, 2026AI
How to Use MiMo V2.6
Set up Xiaomi MiMo V2.6 in Studio and the API, choose the right key and model, and complete a first business task with clear checks.Sep 22, 2026AI
MiMo V2.6 Pro vs Flash
Compare MiMo V2.6 Pro and Flash for coding and agent work: task success, token costs, cache savings and when Pro earns its higher price.Sep 22, 2026AI
Is MiMo V2.6 Free
MiMo V2.6 Flash has a temporary free endpoint. Compare that offer with MIT model downloads, paid API access and desktop plans.Sep 22, 2026AI
Grok 4.7 vs Grok 4.6
Grok 4.7 starts at the same token price as 4.6. Compare access, coding evidence and task costs before moving your workflow.Sep 21, 2026AI
Specification Gaming in Production AI Agents
An AI editor sent 50 of 96 articles off topic. The measured costs, the rules that rewarded drift, and the controls that replaced them.Sep 21, 2026AI
Step 5 Preview Pricing
Compare Step 5 Preview API rates with Step Plan Credits, cached-input discounts, and reasoning costs before choosing your billing route.Sep 21, 2026AIThe other lanes.Pick the one you came for.
174 pieces
Build
Ship with AI when you're not an engineer: vibe-coding, app/SaaS builders, automation, MCP, just-enough infra105 piecesDesign
For designers, brand owners, creative directors141 piecesGrowth
AI for marketing, SEO/GEO, content engines, and the operator-business work of turning audience into revenue8 piecesPlaybooks
AI for your niche: the workflows that actually move, tool by tool, priced51 piecesExplained
What shipped, explained: every real release, decoded for the people who use it0 piecesLab
The agent journey, entry by entry: what shipped, what broke, what was decided, in the ventures I run108–131 / 154

OpenAI Presence Is a Managed Agent Deployment, Not a Self-Serve Tool
OpenAI Presence bundles policies, simulations, approved actions, human escalation, and a Codex improvement loop. Who should buy it, and who should build?Jul 22, 2026AI
GPT-5.6 vs Claude Sonnet 5: Sonnet Wins on Value, Sol Wins the Ceiling
GPT-5.6 vs Claude Sonnet 5 on price, coding, agent benchmarks, long context, and the routing rule that decides which model fits.Jul 22, 2026AI
Grok 4.5 Review: The $2/$6 Agent Model Wins on Cost, Not Raw Coding
Grok 4.5 costs $2/$6 per million tokens. See current benchmarks, the 200K pricing cliff, Cursor math, and exactly who should adopt it.Jul 22, 2026AI
GLM-5.2 Review (2026): The Open-Weight Model That Runs Claude Code at Half the Cost of Opus
GLM-5.2 is a 753B open-weight model that matches Claude Opus on coding at under half the cost. Real pricing, benchmarks, and when to actually use it.Jul 19, 2026AI
Kimi K3 Review: Frontier Performance at Half the API Cost, but the Weights Are Not Here Yet
Kimi K3 reviewed with live API pricing, independent benchmarks, production limits, and the decision rule versus GPT-5.6 Sol and Claude Fable 5.Jul 17, 2026AI
Gemini 3.5 Pro Is Delayed: No Release Date, No Reason to Wait
Gemini 3.5 Pro missed its June window and is reportedly months behind. See what to use now, the real cost, and the proof to wait for.Jul 17, 2026AI
Meta Muse Spark 1.1 Review (2026): The 75%-Cheaper Frontier API, and When It Beats GPT and Claude
Meta's first paid frontier API: $1.25/$4.25 per M tokens, ~75% under GPT and Claude. Where Muse Spark 1.1 wins, where it trails, and when to switch.Jul 13, 2026AI
Claude Sonnet 5 Review (2026): What It Really Costs, Where It Beats Opus, and the Tokenizer Catch
Claude Sonnet 5 lands at $2/$10 intro pricing with near-Opus agentic skill, but a new tokenizer adds ~30% more tokens. The honest switch verdict.Jul 13, 2026AI
ChatGPT Work Review: The $20/Seat Agent That Turns Goals Into Deliverables
ChatGPT Work turns goals into reviewable deliverables. Here is what it does, what Business costs, and the workflow that makes it useful.Jul 11, 2026AI
GPT-5.6 Review: Sol, Terra, Luna, and the Routing Rule
GPT-5.6 Sol, Terra, and Luna reviewed with live API prices, benchmarks, context limits, and the routing rule that decides which tier to use.Jul 10, 2026AI
Gemini Managed Agents Can Now Finish the Job After You Disconnect
Gemini Managed Agents now run in the background, connect to remote tools, call your functions, and refresh credentials without losing their workspace.Jul 10, 2026AI
Codex vs Claude Code After the Microsoft CLI Study: The 24% PR Lift Is Not the Whole Decision
Microsoft's CLI-agent study found 24% more merged PRs. Here's what that changes, and what it does not, for Codex vs Claude Code in 2026.Jul 5, 2026AI
The 8 Best ChatGPT Alternatives in 2026 (and Exactly When to Switch)
The 8 best ChatGPT alternatives in 2026, ranked by real strengths and prices, with the exact reason to switch: Claude, Gemini, Grok, Perplexity and more.Jul 2, 2026AI
Best AI Model Right Now (2026): The Frontier Models, Ranked by Score and Price
The frontier AI models ranked by the Artificial Analysis Intelligence Index and real API price: Claude Fable 5, Opus 4.8, GPT-5.5, Gemini, Grok.Jun 29, 2026AI
The Best Open-Source LLMs in 2026: 8 Open-Weight Models Ranked
The best open-weight LLMs of 2026, ranked. GLM-5.2, DeepSeek V4, Kimi and more, with real benchmarks, the license trap, and the cost math.Jun 29, 2026AI
The Best AI Browser in 2026: Atlas, Comet, Dia and What Actually Wins
The best AI browser in 2026 depends on your OS and the security tradeoff you accept. Comet, Atlas, Dia and 4 more, compared with current prices.Jun 29, 2026AI
The 8 Best AI Agents in 2026 (Real Pricing and Honest Verdicts)
The 8 best AI agents in 2026: what each actually does, real current pricing, and which to pick for coding, research, or business automation.Jun 29, 2026AI
Gemini 3 review (2026): the benchmarks, the price, and which model to actually use
Gemini 3 reviewed with current benchmarks and prices: what each plan costs, where it beats Claude and GPT, and which model in the line to use.Jun 22, 2026AI
The Best AI Coding Agents in 2026: 9 Tools Compared
Claude Code, Cursor, Codex, Copilot, Windsurf, Devin, Cline and more, compared on price, today's benchmarks, and who each one is really for.Jun 22, 2026AI
Codex vs Claude Code: Which AI Coding Agent Wins in 2026?
OpenAI Codex (GPT-5.5) vs Claude Code (Opus 4.8): real 2026 pricing, the only comparable benchmarks, and exactly who should pick which.Jun 22, 2026AI
Claude Code Subagents: How They Work and When to Use Them
Claude Code subagents run a side task in their own context window with their own tools. How to build one, when to use them, and when to skip.Jun 22, 2026AI
ChatGPT vs Claude vs Gemini vs Grok (2026): Which AI Assistant to Actually Pay For
The four frontier AI assistants compared on what actually decides it in 2026: current models, real plan prices, honest limits, and who each is for.Jun 15, 2026AI
Best AI Model for Coding in 2026 (Ranked by SWE-bench and Price)
GPT-5.5, Claude Opus 4.8 and Gemini 3.1 Pro tie on the old benchmark. The harder one, and the price, decide which model to code with.Jun 14, 2026AI
Gemini CLI vs Claude Code (2026): What to Use Before June 18
Google retires Gemini CLI on June 18, 2026. How it compares to Claude Code on price, models, and limits, and what to actually run next.Jun 7, 2026AINewsletter
One letter, every Sunday.Working systems, not hot takes.
Weekly. No spam. Unsubscribe anytime.
