Skip to content

Compare models

Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.

  1. Nemotron 3 Nano Omni (Free)NVIDIARemove
  2. Muse Spark 1.3MetaRemove
nemotron-3-nano-omni-30b-a3b-reasoning:free vs muse-spark-1.3
AttributeNemotron 3 Nano Omni (Free)nemotron-3-nano-omni-30b-a3b-reasoning:freeMuse Spark 1.3muse-spark-1.3
Pricing
Input$0 / 1M$1.25 / 1M
Output$0 / 1M$4.25 / 1M
Cache Write$0 / 1M
Cache Read$0 / 1M$1.25 / 1M
Cache Write (5m)$1.25 / 1M
Cache Write (1h)$1.25 / 1M
Web Search$0 / 1M
Context
Max context256K1M
Max outputN/AN/A
Capabilities
VisionYesYes
Function CallingYesYes
JSON ModeYesYes
StreamingYesYes
Catalogue
ProviderNVIDIAMeta
Categorychatchat
Charge typeFreePay As You Go
Released
Description
SummaryNVIDIA Nemotron 3 Nano Omni is an open 30B-A3B multimodal model designed as a perception and context sub-agent for enterprise agent systems. It supports text, image, video, and audio inputs with text output, enabling unified multimodal reasoning within a single inference loop. Built on a hybrid MoE Transformer–Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers significantly improved efficiency for video reasoning—achieving ~2× higher throughput and 2.5× lower compute compared to separate pipelines. With up to 300K context length and extended thinking support, it is well suited for scalable, multimodal agent workflows.Muse Spark 1.3 is Meta's multimodal reasoning model designed for long-running agentic, multi-agent, and coding workflows. It maintains context and information across extended tasks, enabling reliable execution in complex, multi-step environments. The model is optimized to resolve conflicting information, seek clarification or confirmation when necessary, and execute concisely, making it well suited for autonomous agents, collaborative multi-agent systems, and long-horizon software engineering workflows.