Serenities AI™
ConnectorsMarketplacePricing
DocsArticles
Sign InGet Started

Product

  • Pricing
  • Demo & Showcase
  • Connectors
  • Builder Challenges
  • AI Models

Resources

  • Guide
  • Flows Guide
  • Quick Start
  • Seller Guide
  • Affiliate Program

Company

  • About
  • Editorial Policy
  • Privacy Policy
  • Terms of Service

Connect

  • Email Support

© 2026 Serenities AI™. All rights reserved.

  • Rankings
  • All Models
  • Providers
  • Compare
  • Models Pricing
  • Benchmarks
Providers/Nvidia

Nvidia

Nvidia serves 104 of the models we track. We do not support bring-your-own-key for this provider yet.

text — 94 models

ModelMade byInput / 1MOutput / 1MContext
GPT-OSS 120BOpenAI$0$0128K
GPT-OSS 20BOpenAI$0$0131K
Qwen3.5 122B-A10BAlibaba$0$0262K
Gemma 4 31B ITGoogle$0$0256K
Nemotron Nano 12B v2 VLNVIDIA$0$0128K
Llama 3.3 Nemotron Super 49B v1NVIDIA$0$0131K
Muse Glimmer 30BNVIDIA$0$0131K
Phi-4 MiniMicrosoft$0$0131K
GLM-5.2Mistral$0$01.0M
Mistral Medium 3Mistral$0$0131K
DeepSeek V4 Flash 0731Alibaba$0$01.0M
Qwen3-Next 80B-A3B InstructAlibaba$0$0262K
Qwen3.5 397B-A17BAlibaba$0$0262K
Gemma 2 2b ItNVIDIA$0$0128K
Qwen3-Coder 480B-A35B InstructAlibaba$0$0262K
GLM-5.3-FlashZhipu AI$0$01.0M
Kimi K2.6Moonshot AI$0$0262K
Kimi K3Moonshot AI$0$01.0M
MiniMax-M3MiniMax$0$01.0M
MiniMax-M2.7MiniMax$0$0205K
Step 3.5 FlashStepFun$0$0256K
Qwen2.5 Coder 32b InstructNVIDIA$0$0128K
DeepSeek V4 Pro 0813NVIDIA$0$01.0M
Laguna XS 2.1NVIDIA$0$0262K
Mistral Large 3 675B Instruct 2512NVIDIA$0$0262K
Ministral 3 14B Instruct 2512NVIDIA$0$0262K
mistral-small-4-119b-2603NVIDIA$0$0128K
Mistral: Mixtral 8x7B InstructNVIDIA$0$033K
Mistral-7B-Instruct-v0.3NVIDIA$0$066K
Magistral Small 2506NVIDIA$0$033K
streampetrNVIDIA$0$0128K
Llama 3.3 Nemotron Super 49B v1.5NVIDIA$0$0131K
usdcodeNVIDIA$0$0128K
Nemotron 3 Nano OmniNVIDIA$0$0256K
nemotron-voicechatNVIDIA$0$0128K
studiovoiceNVIDIA$0$0128K
nemotron-3-content-safetyNVIDIA$0$0128K
bevformerNVIDIA$0$0128K
Llama 3.1 Nemotron Nano VL 8B v1NVIDIA$0$033K
Mistral: Mixtral 8x22B InstructNVIDIA$0$066K
llama-nemotron-embed-vl-1b-v2NVIDIA$0$033K
synthetic-video-detectorNVIDIA$0$00K
llama-nemotron-rerank-vl-1b-v2NVIDIA$0$0128K
usdvalidateNVIDIA$0$00K
Active Speaker DetectionNVIDIA$0$00K
Llama 3.1 Nemotron Ultra 253BNVIDIA$0$0128K
llama-3_2-nemoretriever-300m-embed-v1NVIDIA$0$033K
nv-embedcode-7b-v1NVIDIA$0$033K
llama-3.1-nemotron-safety-guard-8b-v3NVIDIA$0$0128K
nemotron-mini-4b-instructNVIDIA$0$0128K
nemotron-content-safety-reasoning-4bNVIDIA$0$0128K
riva-translate-4b-instruct-v1_1NVIDIA$0$0128K
BGE M3NVIDIA$0$08K
sparsedriveNVIDIA$0$0128K
gliner-piiNVIDIA$0$0128K
Llama 3.1 Nemotron Nano 8B v1NVIDIA$0$0131K
nv-embed-v1NVIDIA$0$033K
rerank-qa-mistral-4bNVIDIA$0$0128K
Nemotron 3.5 Lightning 30B A3BNVIDIA$0$0262K
nemotron-3-nano-30b-a3bNVIDIA$0$0131K
paligemmaNVIDIA$0$0128K
InklingNVIDIA$0$01.0M
nvidia-nemotron-nano-9b-v2NVIDIA$0$0131K
Cosmos Reason2 8BNVIDIA$0$0131K
Gemma 3 4B ITNVIDIA$0$0131K
Gemma 3n E2b ItNVIDIA$0$0128K
Gemma 3n E4b ItNVIDIA$0$0128K
Llama 3.1 8B InstructNVIDIA$0$016K
Llama Guard 4 12BNVIDIA$0$0128K
Llama-3.2-90B-Vision-InstructNVIDIA$0$0128K
Llama 3.2 1b InstructNVIDIA$0$0128K
esmfoldNVIDIA$0$0128K
Llama 3.2 11b Vision InstructNVIDIA$0$0128K
Llama 3.1 70b InstructNVIDIA$0$0128K
Llama 3.3 70b InstructNVIDIA$0$0128K
ByteDance-Seed/Seed-OSS-36B-InstructNVIDIA$0$0262K
sarvam-mNVIDIA$0$0128K
Phi 4 MultimodalNVIDIA$0$0128K
Llama 3.1 Nemotron 70B InstructNVIDIA$0$0128K
dracarys-llama-3.1-70b-instructNVIDIA$0$0128K
Whisper Large v3NVIDIA$0$00K
solar-10.7b-instructNVIDIA$0$0128K
Gemma 3 12B ITNVIDIA$0$0131K
Llama 3.2 3B InstructNVIDIA$0$033K
Llama 4 Maverick 17b 128e InstructNVIDIA$0$0128K
esm2-650mNVIDIA$0$0128K
Kimi K2 0905NVIDIA$0$0262K
Mistral Medium 3.5Mistral$0$0262K
Step 3.7 FlashStepFun$0$0256K
mistral-nemotronNVIDIA$0$0128K
DeepSeek V4 FlashDeepSeek$0.14$0.281.0M
Nemotron 3 SuperNVIDIA$0.2$0.8262K
DeepSeek V4 ProDeepSeek$0.435$0.871.0M
Nemotron 3 Ultra 550B A55BNVIDIA$0.5$2.51.0M

image — 6 models

ModelMade byInput / 1MOutput / 1MContext
Qwen ImageNVIDIA$0$00K
Qwen Image EditNVIDIA$0$00K
FLUX.1-schnellNVIDIA$0$00K
FLUX.2 Klein 4BNVIDIA$0$041K

video — 3 models

ModelMade byInput / 1MOutput / 1MContext
cosmos-transfer1-7bNVIDIA$0$00K
cosmos-transfer2.5-2bNVIDIA$0$00K
cosmos-predict1-5bNVIDIA$0$00K

audio — 1 model

ModelMade byInput / 1MOutput / 1MContext
magpie-tts-zeroshotNVIDIA$0$00K

Nvidia — Frequently Asked Questions

How many AI models does Nvidia serve?

Nvidia serves 104 of the models we track, across 4 model types.

What is the cheapest model on Nvidia?

GPT-OSS 120B at $0 per million input tokens.

Can I use my own Nvidia API key?

Not yet — Nvidia is not currently one of the providers Serenities supports for bring-your-own-key.

FLUX.1-dev
NVIDIA
$0
$0
4K
FLUX.1-Kontext-devNVIDIA$0$041K