All free API endpoints

144 results found
AMD Radeon Cloud

MiniCPM 5 1B

Ultra-fast 1B parameter on-device language model served with high throughput on AMD GPU Cloud.

AMD Radeon Cloud

Qwen 3.8 Flash Next

AMD Radeon Cloud high-speed endpoint for Qwen 3.8 Flash Next with 256K context.

AMD Radeon Cloud

DeepSeek V4 Flash Vision (Exp)

Official AMD endpoint for DeepSeek V4 Flash Vision (Exp) multimodal reasoning with 1M context.

AMD Radeon Cloud

DeepSeek V4 Flash

AMD Radeon Cloud Token Factory free API for DeepSeek V4 Flash with 1M context and dual OpenAI/Anthropic protocol support.

Vercel AI Gateway

DeepSeek V4 Flash

Access deepseek-v4-flash free through Vercel's $5 monthly credit.

NVIDIA NIM

GPT-OSS 20b

OpenAI's fast 20B open-weights model on NVIDIA NIM for code and agent loops.

NVIDIA NIM

GPT-OSS 120b

OpenAI's 120B open-weights model hosted on NVIDIA NIM with 1M context.

NVIDIA NIM

DeepSeek V4 Flash

DeepSeek V4 Flash 284B MoE (0731 release) on NVIDIA NIM — ultra-fast coding & tool calling with 1M context.

NVIDIA NIM

DeepSeek V4 Pro

DeepSeek V4 Pro 1.65T MoE flagship (0813 release) via NVIDIA NIM — 1M context, 1,000 free credits.

GMI Cloud

MiniMax M3

MiniMax M3 1M context coding flagship via GMI Cloud — 0-cost direct API access.

DeepSeek

DeepSeek V4 Pro

DeepSeek official V4 Pro API with free registration token grants, Context Caching, and direct OpenAI SDK compatibility.

Higgsfield

Seedance 2.0

ByteDance Seedance 2.0 HD video generation via Higgsfield platform.