Compare models
Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.
4 is the maximum. Remove one to add another.
| Attribute | Kimi K3kimi-k3 | GPT-6 Astragpt-6-astra | GPT-6 Astra Progpt-6-astra-pro | Hy4 previewhy4-preview |
|---|---|---|---|---|
| Pricing | ||||
| Input | $3.00 / 1M | $10.00 / 1M | $10.00 / 1M | $0.834 / 1M |
| Output | $15.00 / 1M | $50.00 / 1M | $50.00 / 1M | $2.50 / 1M |
| Cache Write (5m) | $3.00 / 1M | $10.00 / 1M | $10.00 / 1M | $0.834 / 1M |
| Cache Write (1h) | $3.00 / 1M | $10.00 / 1M | $10.00 / 1M | $0.834 / 1M |
| Cache Read | $3.00 / 1M | $10.00 / 1M | $10.00 / 1M | $0.834 / 1M |
| Web Search | $0 / 1M | $0 / 1M | $0 / 1M | $0 / 1M |
| Context | ||||
| Max context | 1M | 1M | 1M | 1M |
| Max output | N/A | N/A | N/A | N/A |
| Capabilities | ||||
| Vision | Yes | Yes | Yes | Yes |
| Function Calling | Yes | Yes | Yes | Yes |
| JSON Mode | Yes | Yes | Yes | Yes |
| Streaming | Yes | Yes | Yes | Yes |
| Catalogue | ||||
| Provider | MoonShot AI | OpenAI | OpenAI | Tencent |
| Category | chat | chat | chat | chat |
| Charge type | Pay As You Go | Pay As You Go | Pay As You Go | Pay As You Go |
| Released | — | — | — | — |
| Description | ||||
| Summary | Kimi K3 is Moonshot AI's 2.8T-parameter open-weight multimodal reasoning model, designed for complex coding, knowledge work, and long-horizon agentic workflows. It excels at repository-scale development, tool use, debugging, and iterative problem solving across text, images, logs, tests, and runtime feedback. Built with KDA and Attention Residuals for improved computational efficiency, Kimi K3 delivers strong performance on advanced engineering and multimodal reasoning tasks, making it well suited for autonomous coding agents and large-scale production workflows. | GPT-6 Astra is OpenAI's flagship model for demanding end-to-end professional work, designed for advanced analysis, software engineering, deep research, scientific tasks, and document creation. It is particularly strong in long-horizon agentic workflows, including tasks that require sustained reasoning, tool orchestration, and computer and browser use, making it well suited for complex autonomous workflows and production-grade knowledge work. | GPT-6 Astra Pro uses the same underlying model as GPT-6 Astra, but runs with reasoning.mode set to pro for higher-quality responses on complex tasks. Optimized for deeper reasoning, greater accuracy, and more reliable multi-step execution, it is well suited for demanding coding, analysis, and agentic workflows where solution quality takes priority over speed and cost. | Tencent Hy4 Preview is a Mixture-of-Experts (MoE) model from Tencent, featuring 770B total parameters with 49B activated per token. It is designed for coding agents, complex tool-driven workflows, and professional productivity tasks that require strong planning and reliable execution. Optimized for context continuity and sustained multi-step work, Hy4 Preview is well suited for long-horizon coding, agentic automation, tool orchestration, and complex real-world workflows. |