- Blog
- AI
AI
AI models, agents, coding assistants, image/video/voice tools. The cross-category practitioner lane: what shipped, what works, what to do about it. Faster cadence than the depth categories. Top-of-funnel feeder that routes into build / design / growth.
All Articles

Gemini 3.6 Flash Review: Adopt It for Agents, Not Every Task
Gemini 3.6 Flash cuts output price to $7.50/M and output use by 17%. See the benchmarks, migration risks, cost math, and adoption verdict.

OpenAI Presence Is a Managed Agent Deployment, Not a Self-Serve Tool
OpenAI Presence bundles policies, simulations, approved actions, human escalation, and a Codex improvement loop. Who should buy it, and who should build?

GPT-5.6 vs Claude Sonnet 5: Sonnet Wins on Value, Sol Wins the Ceiling
GPT-5.6 vs Claude Sonnet 5 on price, coding, agent benchmarks, long context, and the routing rule that decides which model fits.

Grok 4.5 Review: The $2/$6 Agent Model Wins on Cost, Not Raw Coding
Grok 4.5 costs $2/$6 per million tokens. See current benchmarks, the 200K pricing cliff, Cursor math, and exactly who should adopt it.

GLM-5.2 Review (2026): The Open-Weight Model That Runs Claude Code at Half the Cost of Opus
GLM-5.2 is a 753B open-weight model that matches Claude Opus on coding at under half the cost. Real pricing, benchmarks, and when to actually use it.

Kimi K3 Review: Frontier Performance at Half the API Cost, but the Weights Are Not Here Yet
Kimi K3 reviewed with live API pricing, independent benchmarks, production limits, and the decision rule versus GPT-5.6 Sol and Claude Fable 5.

Gemini 3.5 Pro Is Delayed: No Release Date, No Reason to Wait
Gemini 3.5 Pro missed its June window and is reportedly months behind. See what to use now, the real cost, and the proof to wait for.

Meta Muse Spark 1.1 Review (2026): The 75%-Cheaper Frontier API, and When It Beats GPT and Claude
Meta's first paid frontier API: $1.25/$4.25 per M tokens, ~75% under GPT and Claude. Where Muse Spark 1.1 wins, where it trails, and when to switch.

Claude Sonnet 5 Review (2026): What It Really Costs, Where It Beats Opus, and the Tokenizer Catch
Claude Sonnet 5 lands at $2/$10 intro pricing with near-Opus agentic skill, but a new tokenizer adds ~30% more tokens. The honest switch verdict.

ChatGPT Work Review: The $20/Seat Agent That Turns Goals Into Deliverables
ChatGPT Work turns goals into reviewable deliverables. Here is what it does, what Business costs, and the workflow that makes it useful.

GPT-5.6 Review: Sol, Terra, Luna, and the Routing Rule
GPT-5.6 Sol, Terra, and Luna reviewed with live API prices, benchmarks, context limits, and the routing rule that decides which tier to use.

Gemini Managed Agents Can Now Finish the Job After You Disconnect
Gemini Managed Agents now run in the background, connect to remote tools, call your functions, and refresh credentials without losing their workspace.

Codex vs Claude Code After the Microsoft CLI Study: The 24% PR Lift Is Not the Whole Decision
Microsoft's CLI-agent study found 24% more merged PRs. Here's what that changes, and what it does not, for Codex vs Claude Code in 2026.

The 8 Best ChatGPT Alternatives in 2026 (and Exactly When to Switch)
The 8 best ChatGPT alternatives in 2026, ranked by real strengths and prices, with the exact reason to switch: Claude, Gemini, Grok, Perplexity and more.

Best AI Model Right Now (2026): The Frontier Models, Ranked by Score and Price
The frontier AI models ranked by the Artificial Analysis Intelligence Index and real API price: Claude Fable 5, Opus 4.8, GPT-5.5, Gemini, Grok.

The Best Open-Source LLMs in 2026: 8 Open-Weight Models Ranked
The best open-weight LLMs of 2026, ranked. GLM-5.2, DeepSeek V4, Kimi and more, with real benchmarks, the license trap, and the cost math.

The Best AI Browser in 2026: Atlas, Comet, Dia and What Actually Wins
The best AI browser in 2026 depends on your OS and the security tradeoff you accept. Comet, Atlas, Dia and 4 more, compared with current prices.

The 8 Best AI Agents in 2026 (Real Pricing and Honest Verdicts)
The 8 best AI agents in 2026: what each actually does, real current pricing, and which to pick for coding, research, or business automation.

Gemini 3 review (2026): the benchmarks, the price, and which model to actually use
Gemini 3 reviewed with current benchmarks and prices: what each plan costs, where it beats Claude and GPT, and which model in the line to use.

The Best AI Coding Agents in 2026: 9 Tools Compared
Claude Code, Cursor, Codex, Copilot, Windsurf, Devin, Cline and more, compared on price, today's benchmarks, and who each one is really for.

Codex vs Claude Code: Which AI Coding Agent Wins in 2026?
OpenAI Codex (GPT-5.5) vs Claude Code (Opus 4.8): real 2026 pricing, the only comparable benchmarks, and exactly who should pick which.

Claude Code Subagents: How They Work and When to Use Them
Claude Code subagents run a side task in their own context window with their own tools. How to build one, when to use them, and when to skip.

Nano Banana, Explained: Google's Gemini Image Models (Pro, 2, and the Original)
Nano Banana is Google's Gemini image model line. What Pro, 2, and the original each do, real per-image pricing, how to use them, and which to pick.

ChatGPT vs Claude vs Gemini vs Grok (2026): Which AI Assistant to Actually Pay For
The four frontier AI assistants compared on what actually decides it in 2026: current models, real plan prices, honest limits, and who each is for.

Best AI Model for Coding in 2026 (Ranked by SWE-bench and Price)
GPT-5.5, Claude Opus 4.8 and Gemini 3.1 Pro tie on the old benchmark. The harder one, and the price, decide which model to code with.

Gemini CLI vs Claude Code (2026): What to Use Before June 18
Google retires Gemini CLI on June 18, 2026. How it compares to Claude Code on price, models, and limits, and what to actually run next.

Codex vs Claude Code vs Cursor: Which AI Coding Agent to Use in 2026
Codex, Claude Code, and Cursor compared for 2026: pair, delegate, or dispatch. Real pricing, current models, and the one rule that picks one.

Cursor vs GitHub Copilot (2026): Which AI Coding Tool to Pay For
Cursor is the standalone AI editor with its own fast Composer model. GitHub Copilot is the multi-model plugin in your IDE. Which to pay for in 2026.

Claude vs ChatGPT vs Gemini (2026): Which $20 Plan to Actually Pay For
Claude Pro, ChatGPT Plus, and Google AI Pro all cost about $20. Which wins for coding, writing, and research, and the rule that picks for you.

NotebookLM Alternatives (2026): What to Actually Use, Matched to the Job
NotebookLM's podcasts are unmatched, but its 50-source cap and lock-in push people out. The 5 best alternatives in 2026, matched to what you do.

Perplexity vs ChatGPT (2026): Which to Pay For
Both cost $20/mo and Perplexity runs the same GPT and Claude models as ChatGPT. So the real call is interface, not IQ. Here's which to pay for.

Grok vs ChatGPT: Which AI Is Actually Better in 2026?
A current, price-honest head-to-head of Grok 4.3 and ChatGPT: real 2026 pricing, the axis that decides, and which to pay for.

10 Best AI Coding Assistants in 2026 (Real Pricing)
The 10 best AI coding assistants in 2026, priced from live pages: Cursor, Claude Code, Copilot, Windsurf and more, plus the ones to avoid.

Claude Code vs Cursor (2026): Terminal Agent or AI IDE, and Which to Use
Claude Code vs Cursor in 2026: both run the same models, so code quality is a wash. The real split is terminal autonomy vs IDE control, plus the cost math.

Gemini vs ChatGPT (2026): Which One to Pay For
The model race is basically a tie. After Google's May 2026 pricing reset, the deciding factor is what you do and where your data lives.

9 Best AI Agent Platforms in 2026 (Real Pricing, and the Ones to Avoid)
The 9 best AI agent platforms in 2026, priced from their live pages: n8n, Lindy, Gumloop, Relay and more, with honest cons and the ones to avoid.

Claude vs ChatGPT (2026): Which One to Pay For
Claude Opus 4.8 vs GPT-5.5, head to head: coding, writing, pricing, images. The clear verdict on which AI to actually pay for in 2026.

Codex folded into the ChatGPT org: my roadmap-risk test
OpenAI folded Codex, ChatGPT and the API into one org under Brockman. Codex is now a super-app wedge, not a standalone tool. My roadmap-risk test.

Opus 4.7 +27%, GPT-5.5 2x: The Harness That Held
Three vendors quietly raised AI cost in one week. Which mechanism hits your stack, and the cache, router and spend-cap architecture that absorbs it.

Grok Build: cheapest model, $300/mo door — the operator math
Grok Build's cheap model is gated behind a $300/mo tier: the subscription-vs-API math for an operator already on Claude Max, and why it waits.

Claude Agent SDK on a Separate Meter June 15: The Token Math
On May 13 Anthropic announced that starting June 15, 2026, programmatic Claude usage (Agent SDK, `claude -p`, Claude Code GitHub Actions, third-party apps) draws from a separate credit pool instead of subscription limits. Pro gets $20,…

PwC Certifies 30,000 on Claude: 18-Month Arbitrage Window
PwC and Anthropic announced a multi-year alliance expansion on May 14, 2026 that puts Claude Code and Cowork into hundreds of thousands of PwC professionals' hands, with 30,000 US staff certified inside the year. The release publishes…

Codex Mobile Preview: Async Coding as the Solo-Op Default
On May 14, 2026, OpenAI shipped a preview of Codex inside the ChatGPT mobile app on iOS, iPad, and Android, available on all plans including Free and Go. The phone becomes a thin controller over a Codex desktop session running on a paired…

Codex in ChatGPT mobile: async coding for solo operators
OpenAI shipped Codex inside ChatGPT mobile on May 14, free tier included. The phone is now the agent control plane for solo Mac operators.

Notion Developer Platform: the August 11 credit cliff
On May 13, 2026, Notion launched its Developer Platform — Workers (hosted code runtime), External Agents with partner integrations for Claude Code, Cursor, Codex and Decagon, an External Agent API for custom agents, a CLI called `ntn`,…

Anthropic Buying Stainless: What to Do in Your Codebase Now
On May 12, The Information reported Anthropic is in advanced talks to acquire Stainless for at least $300 million, a deal that would hand the Claude maker control of the SDK-generation pipeline that produces the official client libraries…

Claude Code +50% Weekly Cap: Migration Math for July 13
On May 13, 2026, Anthropic raised Claude Code weekly limits 50 percent through July 13 for Pro, Max, Team, and seat-based Enterprise accounts. The same day, OpenAI offered two free months of Codex to teams who migrate inside 30 days. The…

Cursor Cloud Agents May 2026: 2 gaps it still can't close
Cursor's May 13 release moves cloud agent environments from demo infra to production tooling. The release is real, the numbers are real, and the SERP is reading it wrong. This piece argues the real choice isn't managed-versus-DIY, it is…

Claude for Small Business: What It Kills in a $50K Stack
On May 13 Anthropic launched Claude for Small Business, a Cowork add-on with eight native connectors and fifteen ready-to-run agentic workflows priced at zero incremental cost on top of an existing Claude license. This piece is the…
One letter, every Sunday. Working systems, not hot takes.
Build logs, working systems, and field notes from running a portfolio of AI ventures.