Kimi K3 Model Serving
Announcing native availability of NVIDIA Nemotron 3 Nano, NVIDIA’s latest reasoning model

Announcing native availability of NVIDIA Nemotron 3 Nano, NVIDIA’s latest reasoning model

12/15/2025

What this post added

Introduces native availability of NVIDIA Nemotron 3 Nano, a hybrid Mamba-Transformer + sparse MoE reasoning model with ~3B active parameters and 1M-token context. Highlights its optimization on Together AI for high throughput and cost-efficiency, making it suitable for agentic systems, coding assistants, scientific agents, tool-using planners, and enterprise context applications. Details performance, reliability, and cost-efficiency benefits of running Nemotron 3 Nano on Together AI, including an OpenAI-compatible interface for easy adoption.

Read the original post ↗