Monday, 21 September 2026 | التحديث اليومي نظرة ثاقبة للذكاء الاصطناعي، مكتوبة للبناة

NVIDIA Nemotron 3 Nano Omni

NVIDIA Nemotron 3 Nano Omni — المواصفات

Compiled by Mustafa Ihsan from the vendor’s published documentation · Last updated

المطوّر NVIDIA
النوع متعدد الوسائط (شامل)
النمط نص، صورة، صوت، فيديو → نص
عدد المعايير (البارامترات) إجمالي 30 مليار معامل / ~3 مليارات معامل نشطة (بنمط خليط الخبراء MoE)
نافذة السياق 256 ألف رمز
الترخيص اتفاقية NVIDIA للنماذج المفتوحة
أوزان مفتوحة المصدر نعم
أُطلِقَ 2026
مقدمو واجهات برمجة التطبيقات (API) Hugging Face، OpenRouter، NVIDIA NIM

تشغيله محليًّا

الذاكرة المخصصة لوحدة معالجة الرسومات (VRAM) (بتعميق 16 بت FP16/ BF16) ~62 جيجابايت
ذاكرة الفيديو (VRAM) بتنسيق 4 بت ~21 جيجابايت (باستخدام تنسيق NVFP4)
أقل وحدة معالجة رسومية (GPU) مطلوبة RTX 5090 بسعة 32 جيجابايت (باستخدام تنسيق NVFP4) أو H100 بسعة 80 جيجابايت (باستخدام تنسيق BF16)

المقاييس المرجعية

OCRBench V2 67.04
Video-MME 72.2
أو إس وورلد 47.4
Speech IF 89.39

الصفحة الرسمية →

What is NVIDIA Nemotron 3 Nano Omni?

NVIDIA Nemotron 3 Nano Omni is an open omni-modal model: it sees, hears, watches and reads
— text, image, audio and video in, text out — from a single 30B-A3B mixture-of-experts that
activates only about 3B parameters per token. It is a Mamba-Transformer hybrid with a 256K
context, released under the NVIDIA Open Model Agreement, which permits commercial use. It
scores 67.04 on OCRBench V2, 72.2 on Video-MME, 47.4 on OSWorld and 89.39 on Speech IF.

Four input modalities in one 30B model that runs on a single RTX 5090 at NVFP4 (about
21 GB) is the notable part. The usual way to build an application that handles audio, video,
images and text is to chain three or four specialist models, each with its own deployment,
latency budget and failure mode. Nemotron 3 Nano Omni collapses that into one, which is why
its 47.4 on OSWorld — a computer-use benchmark that requires seeing a screen and acting on it
— is more interesting than the raw number suggests. There is no public per-token API price;
the cost is the hardware, which for a single-GPU deployment is unusually approachable for a
model of this breadth.

NVIDIA Nemotron 3 Nano Omni pricing

NVIDIA Nemotron 3 Nano Omni has no public per-token API price. It is open-weight, so the only cost is the hardware you run it on — see the local-hardware table above and size a GPU with the حاسبة الذاكرة VRAM.

انتقل إلى الأعلى