Compare models
Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.
| Attribute | Claude Fable 5.1claude-fable-5.1 | Nemotron 3 Nano Omni (Free)nemotron-3-nano-omni-30b-a3b-reasoning:free |
|---|---|---|
| Pricing | ||
| Input | $10.00 / 1M | $0 / 1M |
| Output | $50.00 / 1M | $0 / 1M |
| Cache Write (5m) | $12.50 / 1M | — |
| Cache Write (1h) | $20.00 / 1M | — |
| Cache Read | $1.00 / 1M | $0 / 1M |
| Web Search | $0 / 1M | — |
| Cache Write | — | $0 / 1M |
| Context | ||
| Max context | 1M | 256K |
| Max output | N/A | N/A |
| Capabilities | ||
| Vision | Yes | Yes |
| Function Calling | Yes | Yes |
| JSON Mode | Yes | Yes |
| Streaming | Yes | Yes |
| Catalogue | ||
| Provider | Anthropic | NVIDIA |
| Category | chat | chat |
| Charge type | Pay As You Go | Free |
| Released | — | — |
| Description | ||
| Summary | Claude Fable 5.1 is an upgraded version of Fable 5, delivering broad improvements with particularly strong gains in agentic coding, long-running workflows, and professional knowledge work. It excels at large code refactors, front-end and visual code generation, financial analysis, and complex analytical tasks. Compared with Fable 5, it also produces more concise plans and summaries while maintaining strong performance across extended tasks, making it a natural upgrade for existing Fable workflows and a strong option alongside Opus 5 for reasoning-intensive applications. | NVIDIA Nemotron 3 Nano Omni is an open 30B-A3B multimodal model designed as a perception and context sub-agent for enterprise agent systems. It supports text, image, video, and audio inputs with text output, enabling unified multimodal reasoning within a single inference loop. Built on a hybrid MoE Transformer–Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers significantly improved efficiency for video reasoning—achieving ~2× higher throughput and 2.5× lower compute compared to separate pipelines. With up to 300K context length and extended thinking support, it is well suited for scalable, multimodal agent workflows. |