All free API endpoints
144 results foundMiniCPM 5 1B
Ultra-fast 1B parameter on-device language model served with high throughput on AMD GPU Cloud.
Qwen 3.8 Flash Next
AMD Radeon Cloud high-speed endpoint for Qwen 3.8 Flash Next with 256K context.
DeepSeek V4 Flash Vision (Exp)
Official AMD endpoint for DeepSeek V4 Flash Vision (Exp) multimodal reasoning with 1M context.
DeepSeek V4 Flash
AMD Radeon Cloud Token Factory free API for DeepSeek V4 Flash with 1M context and dual OpenAI/Anthropic protocol support.
DeepSeek V4 Flash
Access deepseek-v4-flash free through Vercel's $5 monthly credit.
GPT-OSS 20b
OpenAI's fast 20B open-weights model on NVIDIA NIM for code and agent loops.
GPT-OSS 120b
OpenAI's 120B open-weights model hosted on NVIDIA NIM with 1M context.
DeepSeek V4 Flash
DeepSeek V4 Flash 284B MoE (0731 release) on NVIDIA NIM — ultra-fast coding & tool calling with 1M context.
DeepSeek V4 Pro
DeepSeek V4 Pro 1.65T MoE flagship (0813 release) via NVIDIA NIM — 1M context, 1,000 free credits.
MiniMax M3
MiniMax M3 1M context coding flagship via GMI Cloud — 0-cost direct API access.
DeepSeek V4 Pro
DeepSeek official V4 Pro API with free registration token grants, Context Caching, and direct OpenAI SDK compatibility.
Seedance 2.0
ByteDance Seedance 2.0 HD video generation via Higgsfield platform.