Qwen 3.8 Flash Next (Vultr)
Vultr Serverless Inference endpoint for Qwen 3.8 Flash Next ($0.10 in, $0.20 out).
更新于 2026/9/160 个代码示例
关于该模型
Why Choose Qwen 3.8 Flash Next?
Qwen 3.8 Flash Next is Alibaba's next-generation lightweight high-speed LLM, optimized for complex reasoning, mathematics, and agentic workflows, hosted on AMD GPU Cloud.
Key Strengths & Capabilities
- 256K Extended Context: Accommodates large source documents and multi-turn interactive conversations.
- Enhanced Reasoning: Outstanding logic and structured tool-calling capabilities.
- Sub-Second TTFT: Minimized time-to-first-token for conversational bots and IDE integration.
Free Tier Basis
Free public endpoint on AMD Radeon Cloud Token Factory with zero billing.
如何免费接入 (通过 Vultr)
暂无详细的接入说明。
在浏览器里试用
选个模型、粘贴你自己的免费 Key,即可运行。你的 Key 只用于调用一次该提供商,绝不存储在我们的服务器上。
选个模型、粘贴你自己的免费 Key,即可运行。你的 Key 只用于调用一次该提供商,绝不存储在我们的服务器上。
调用代码示例
暂未提供代码示例。
社区反馈
评论与你的 GitHub 账号绑定,登录后即可参与讨论。