| Note | Current flagship; 2M-token context window (largest production context of any model family), strong coding/reasoning/multimodal. GA February 2026. | Current closed-weights flagship; Intelligence Index v4.0 score 56.6 (top Chinese model); lowest reported hallucination rate (22.9%). $2.50/$7.50 per 1M tokens. | First publicly available Mythos-class model (above Opus); Mythos 5 weights with safety classifiers active. 1M context, $10/$50, highest benchmarks of any GA model. Refusals return stop_reason refusal; mandatory 30-day retention. |
|---|