The Models I Actually Use in 2026
2026-07-21aiinmydraft

The Models I Actually Use in 2026

There are many strong models in 2026. But the ones I reach for daily are not the strongest on benchmarks. They are the ones that fit the shape of my work.

Every few months a new model tops the benchmarks. It is tempting to chase the leaderboard. But the model I reach for daily is not the strongest on paper. It is the one that fits the shape of the work I actually do.

Most of my work is small edits inside an existing codebase. A model that is fast, cheap, and good at continuation beats a model that is brilliant at one-shot party tricks.

What shapes my choice

  • Speed of the editing loop
  • Cost across a long session
  • Quality of continuation, not just generation
  • How well it reads existing structure

Pick the model that fits the shape of your work. Benchmarks measure something, but not the thing that matters most at the keyboard.

More Updates

Search Experience Acceptance Criteria That Test Real Behavior
general2026-10-05

Search Experience Acceptance Criteria That Test Real Behavior

Teams usually find gaps in search experience when an exception exposes an unclear decision. A better starting point is to connect the query box to helpful ranking, filters, result context, empty states, and recovery. This matters because a technically fast…

search experienceacceptance-criteriapractical guide
Read
Homepage Structure Acceptance Criteria That Test Real Behavior
general2026-10-04

Homepage Structure Acceptance Criteria That Test Real Behavior

A credible homepage structure implementation has a narrow promise: lead with an evidence-backed promise, audience, primary action, proof, and a clear route to detail. Treat a collection of slogans forces visitors to infer what the product does and why they…

homepage structureacceptance-criteriapractical guide
Read
Design System Acceptance Criteria That Test Real Behavior
general2026-10-04

Design System Acceptance Criteria That Test Real Behavior

The hard part of design system is not adding another tool or screen. It is deciding how to standardize the repeated tokens, components, states, accessibility rules, versions, and exceptions already in use, while accounting for one concrete failure: a…

design systemacceptance-criteriapractical guide
Read
Back to updates