GPT-5.5 Instant
The everyday ChatGPT default: fast enough for routine questions, drafting, rewriting and low-consequence transformations.
- Useful clue
- Quick explanations and first drafts
- Access
- ChatGPT
The Lab · 13 current model entries
Names change faster than people can learn them. The Lab keeps official positions, independent observations and Human Bit tests separate, then lets you compare only the fields relevant to your job.
Coverage rule
We include a model when it is currently accessible, materially different from another tier and relevant to a real user decision. The catalogue order is not a rank.
OpenAI · Anthropic · Google · xAI · Moonshot AIFilter the bench
Start with the work surface or provider you can actually use. Select up to three entries to compare their stated fit, limits, evidence coverage and review dates side by side.
The everyday ChatGPT default: fast enough for routine questions, drafting, rewriting and low-consequence transformations.
OpenAI’s main reasoning model for complex knowledge work, research, coding and tasks that need more deliberate analysis.
The highest-capability GPT-5.6 option for difficult and longer-running work where a standard reasoning pass is not enough.
Anthropic’s current Sonnet balance for everyday professional work, coding and agentic tasks, with adaptive thinking by default.
Anthropic’s current model for complex agentic coding and enterprise work, and the model it points Opus 4.8 users at. Thinking is on by default, so the effort setting, not a thinking switch, is how you control depth and cost.
A high-capability Claude model for complex agentic coding and enterprise work, now superseded by Claude Opus 5 but still available on every platform that carried it.
Anthropic’s most capable widely released model for demanding reasoning and long-horizon agentic work, with stricter safety classifiers than Mythos.
Google’s fast current model for multimodal, coding and agentic work, deployed broadly across consumer, developer and enterprise products.
A preview model for precise multi-step reasoning, software engineering and tool use across text, images, video, audio and PDFs.
Google’s lower-latency, lower-cost model for high-volume extraction and lightweight multimodal jobs.
xAI’s current flagship for coding, agentic tasks and knowledge work, with configurable reasoning and an emphasis on speed.
Kimi’s hosted multimodal flagship for long-horizon coding, knowledge work and tool-using tasks, with a one-million-token context window.
Kimi’s dedicated coding model for long-context, tool-using software work, with a separate high-speed endpoint for the same model.
Each entry needs an official record, a practical decision, an evidence state and a review owner. Missing independent evidence remains visible rather than being filled with a guess.