Skip to content

Catalog · August 2026

Find the AI model that fits the job.

Search and filter 26 models from major labs. Answer a short questionnaire and we will rank matches with reasons — using official specs where they exist, and clearly labeling everything else.

7 models · Speed

Low-latency chat and lightweight coding

  • speed
  • chat
  • coding
  • Fastest comparative latency in the Claude lineup
  • Extended thinking
  • 64K max output tokens
Official specs
Details

Everyday coding and chat with fast latency

  • chat
  • coding
  • reasoning
  • speed
  • 1M-token context
  • Adaptive thinking
  • 128K max output tokens
Official specs
Details

Cheap, fast coding and chat

  • speed
  • coding
  • chat
  • Very low API price (reported)
  • 1M-token context (reported)
Reported
Details

High tokens-per-second chat and cheap inference

  • speed
  • chat
  • Very high output speed (reported)
  • 1M-token context (reported)
  • Introductory Flash pricing
Reported
Details

High-volume, latency-sensitive chat

  • speed
  • chat
  • Lowest-cost GPT-5.6 tier
  • Built for speed
Reported
Details

Daily ChatGPT-class work at mid-tier price

  • chat
  • coding
  • reasoning
  • speed
  • Mid-tier GPT-5.6
  • ChatGPT default family
  • API access
Reported
Details

Chat, coding, and agentic work with configurable reasoning

  • chat
  • reasoning
  • coding
  • research
  • speed
  • 500K-token context window
  • Configurable reasoning
  • Agentic tool calling
Official specs
Details