Vultr
Vultr 是全球顶尖的独立云基础设施服务商,提供高性能 Cloud GPU 与原生兼容 OpenAI 接口标准的 Serverless Inference 无服务器推理平台。新用户通过专属推荐可获 $300 试用额度(30天有效),支持免部署零门槛直调 DeepSeek V4.1 Flash、Qwen 3.8 等前沿大模型。
Showing 3 of 3 active models
Vultr
GLM-5.3-Flash
Vultr Serverless Inference endpoint for GLM 5.3 Flash ($0.10 in, $0.35 out).
Vultr
Qwen 3.8 Flash Next
Vultr Serverless Inference endpoint for Qwen 3.8 Flash Next ($0.10 in, $0.20 out).
Vultr
DeepSeek V4.1 Flash
Official Vultr Serverless Inference endpoint for DeepSeek V4.1 Flash ($0.15 in, $0.60 out, payable with $300 trial credit).
Try it in your browser
Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.
Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.
Frequently Asked Questions about Vultr
Explore Other Free AI Providers
AMD Radeon Cloud
AMD 官方面向开发者推出的 Radeon Cloud Token Factory 算力平台,提供开箱即用的免配置 AI 模型 API(涵盖 DeepSeek V4 Flash 1M、Qwen 3.8 Flash、MiniCPM 等),原生兼容 OpenAI 与 Anthropic 双协议。
NVIDIA NIM
NVIDIA NIM (Inference Microservices) provides optimized cloud and self-hosted API endpoints for enterprise and open-source models with 1,000 free inference credits on registration.
Zhipu AI
智谱 AI 开放平台(BigModel),提供 GLM-5.3-Flash、GLM-4 等全系列前沿大模型与 GLM Coding Plan 开发者订阅。
Empero
Empero is an independent AI research lab based in Germany, dedicated to advancing open-weights models. It provides a 100% free, OpenAI-compatible community endpoint for models like GLM-5.3-Flash with no registration or API key required.
GMI Cloud
GMI Cloud 是专为高性能 AI 与 GPU 算力打造的云服务平台。提供 MiniMax M3、Speech 2.8 与 Music 3.0 等核心旗舰模型的 0 元直接 API 调用,免绑卡、免预充值。
OpenCode Go
OpenCode Go 是专为 Coding Agent 和高频开发者打造的超高性价比订阅计划(Coding Plan)。首月仅 $5(续费 $10/月),提供每月 $60 基础额度池、Hy3 独享 $480/月 算力、GPT-5.6 Sol 5 折以及 Muse Spark 1.2 Contributor 近乎无限额度,量大管饱,可直接生成 API Key 供 Cursor、Claude Code、Aider 等各种工具调用。
Community Feedback
Comments are tied to your GitHub account — sign in to join the discussion.