NVIDIA NIM

NVIDIA NIM (Inference Microservices) provides optimized cloud and self-hosted API endpoints for enterprise and open-source models with 1,000 free inference credits on registration.

Showing 40 of 40 active models
NVIDIA NIM

GPT-OSS 20b

OpenAI's fast 20B open-weights model on NVIDIA NIM for code and agent loops.

NVIDIA NIM

GPT-OSS 120b

OpenAI's 120B open-weights model hosted on NVIDIA NIM with 1M context.

NVIDIA NIM

DeepSeek V4 Flash

DeepSeek V4 Flash 284B MoE (0731 release) on NVIDIA NIM — ultra-fast coding & tool calling with 1M context.

NVIDIA NIM

DeepSeek V4 Pro

DeepSeek V4 Pro 1.65T MoE flagship (0813 release) via NVIDIA NIM — 1M context, 1,000 free credits.

NVIDIA NIM

BEVFormer Bird's-Eye-View Perception

NVIDIA BEVFormer bird's-eye-view camera fusion perception model on NIM.

NVIDIA NIM

SparseDrive End-to-End Perception

NVIDIA SparseDrive end-to-end perception and motion forecasting on NIM.

NVIDIA NIM

StreamPETR 3D Perception

NVIDIA StreamPETR real-time 3D object detection microservice on NIM.

NVIDIA NIM

Synthetic Video Detector (SVD)

NVIDIA AI synthetic deepfake video detection microservice on NIM.

NVIDIA NIM

Cosmos Transfer 1.0 7B

NVIDIA Cosmos Transfer 1.0 7B high-fidelity physical video transfer on NIM.

NVIDIA NIM

Cosmos Transfer 2.5 2B

NVIDIA Cosmos Transfer 2.5 2B physical video style transfer on NIM.

NVIDIA NIM

Cosmos 3 Nano Reasoner

NVIDIA Cosmos 3 Nano physical spatial dynamics reasoning model on NIM.

NVIDIA NIM

Cosmos 3 Nano World Model

NVIDIA Cosmos 3 Nano physical AI world foundation model on NIM.

NVIDIA NIM

PaliGemma 3B Vision-Language

Google's 3B vision-language model for image captioning and OCR on NIM.

NVIDIA NIM

DiffusionGemma 26B-A4B IT

Google's Diffusion-Transformer hybrid image synthesis model on NIM.

NVIDIA NIM

Nemotron 3 Embed 1B

NVIDIA's 1B dense text embedding model with top MTEB retrieval performance.

NVIDIA NIM

Riva Translate 4B Instruct v1.1

NVIDIA Riva 4B high-throughput machine translation v1.1 model on NIM.

NVIDIA NIM

Riva Translate 4B Instruct v2

NVIDIA Riva 4B neural machine translation v2 model on NIM.

NVIDIA NIM

NVIDIA Active Speaker Detection

NVIDIA audio-visual neural active speaker detection microservice.

NVIDIA NIM

NVIDIA Background Noise Removal

NVIDIA real-time background noise removal and echo cancellation.

NVIDIA NIM

NVIDIA Studio Voice

NVIDIA studio-quality voice enhancement and audio restoration microservice.

NVIDIA NIM

Nemotron VoiceChat Duplex

NVIDIA's full-duplex sub-300ms real-time conversational voice agent on NIM.

NVIDIA NIM

Magpie TTS Zero-Shot

NVIDIA's neural zero-shot voice cloning and TTS model on NIM.

NVIDIA NIM

Llama Guard 4 12B

Meta's 4th gen multimodal safety & content moderation model on NIM.

NVIDIA NIM

Llama 3.1 Nemotron Safety Guard 8B v3

NVIDIA 8B safety guard model for prompt injection & jailbreak detection.

NVIDIA NIM

Nemotron 3.5 Content Safety

NVIDIA NeMo Guardrails content safety & moderation model on NIM.

NVIDIA NIM

Ising Calibration 1.0 35B-A3B

NVIDIA 35B-A3B MoE reasoning calibration model on NIM.

NVIDIA NIM

Ising Calibration 1.5 31B

NVIDIA 31B reasoning calibration and verification model on NIM.

NVIDIA NIM

Laguna XS 2.1

Poolside's specialized coding intelligence model on NVIDIA NIM.

NVIDIA NIM

Llama 3.2 90B Vision Instruct

Meta's high-capacity 90B multimodal vision model on NVIDIA NIM.

NVIDIA NIM

Llama 3.2 11B Vision Instruct

Meta's lightweight 11B multimodal vision model on NVIDIA NIM.

NVIDIA NIM

Gemma 4 31B IT

Google's 2026 Gemma 4 31B instruction-tuned model on NVIDIA NIM.

NVIDIA NIM

Mistral Nemotron

Mistral AI and NVIDIA jointly optimized model for structured JSON and reasoning.

NVIDIA NIM

Nemotron 3 Nano Omni

NVIDIA 30B-A3B omni multimodal vision-language reasoning model.

NVIDIA NIM

Nemotron 3 Nano 30B

NVIDIA ultra-efficient 30B MoE (3B active) low-latency coding model.

NVIDIA NIM

Nemotron 3 Super 120B

NVIDIA 120B MoE (12B active) high-throughput reasoning model on NIM.

NVIDIA NIM

Nemotron 3 Ultra 550B

NVIDIA flagship 550B MoE (55B active) reasoning & code synthesis model via NIM.

NVIDIA NIM

Meta Muse Glimmer 30B

Meta's 30B multimodal reasoning model on NVIDIA NIM — text & image input with tool calling and separate thinking.

NVIDIA NIM

Nemotron 3.5 Lightning

NVIDIA's 30B-A3B Mamba-2/MoE hybrid architecture with NVFP4 quantization and millisecond latency.

NVIDIA NIM

Kimi K3

Moonshot AI 2.8T Kimi K3 MoE API via NVIDIA NIM — 1M context, native vision, free endpoint with 40 RPM limit.

NVIDIA NIM

MiniMax M3

MiniMax M3 1M MSA Sparse Attention model on NVIDIA NIM — 59% SWE-bench Pro for autonomous coding agents.

Try it in your browser

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Frequently Asked Questions about NVIDIA NIM

Community Feedback

Comments are tied to your GitHub account — sign in to join the discussion.

Explore Other Free AI Providers