GLM 5.3 Flash (Vultr)
Vultr Serverless Inference endpoint for GLM 5.3 Flash ($0.10 in, $0.35 out).
Updated 9/16/20260 code examples
About the Model
为什么选择 GLM-5.3-Flash?
GLM-5.3-Flash(原代号 Ox Alpha)是智谱 AI 发布的轻量旗舰模型,采用 320B 总参数、18B 激活的混合注意力 MoE 架构,支持 100 万 Token 上下文与亚秒级响应,官方提供终身免费调用额度。
How to Access for Free (via Vultr)
No detailed access description available.
Try it in your browser
Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.
Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.
Code Examples
No examples provided yet.
Community Feedback
Comments are tied to your GitHub account — sign in to join the discussion.