Skip to content

Compare models

Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.

  1. Nemotron 3 Nano Omni (Free)NVIDIARemove
  2. Claude Fable 5.1AnthropicRemove
nemotron-3-nano-omni-30b-a3b-reasoning:free vs claude-fable-5.1
AttributeNemotron 3 Nano Omni (Free)nemotron-3-nano-omni-30b-a3b-reasoning:freeClaude Fable 5.1claude-fable-5.1
Pricing
Input$0 / 1M$10.00 / 1M
Output$0 / 1M$50.00 / 1M
Cache Write$0 / 1M
Cache Read$0 / 1M$1.00 / 1M
Cache Write (5m)$12.50 / 1M
Cache Write (1h)$20.00 / 1M
Web Search$0 / 1M
Context
Max context256K1M
Max outputN/AN/A
Capabilities
VisionYesYes
Function CallingYesYes
JSON ModeYesYes
StreamingYesYes
Catalogue
ProviderNVIDIAAnthropic
Categorychatchat
Charge typeFreePay As You Go
Released
Description
SummaryNVIDIA Nemotron 3 Nano Omni is an open 30B-A3B multimodal model designed as a perception and context sub-agent for enterprise agent systems. It supports text, image, video, and audio inputs with text output, enabling unified multimodal reasoning within a single inference loop. Built on a hybrid MoE Transformer–Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers significantly improved efficiency for video reasoning—achieving ~2× higher throughput and 2.5× lower compute compared to separate pipelines. With up to 300K context length and extended thinking support, it is well suited for scalable, multimodal agent workflows.Claude Fable 5.1 is an upgraded version of Fable 5, delivering broad improvements with particularly strong gains in agentic coding, long-running workflows, and professional knowledge work. It excels at large code refactors, front-end and visual code generation, financial analysis, and complex analytical tasks. Compared with Fable 5, it also produces more concise plans and summaries while maintaining strong performance across extended tasks, making it a natural upgrade for existing Fable workflows and a strong option alongside Opus 5 for reasoning-intensive applications.