GLM 5.3 Flash (Vultr)

Vultr Serverless Inference endpoint for GLM 5.3 Flash ($0.10 in, $0.35 out).

Updated 9/16/20260 code examples

About the Model

为什么选择 GLM-5.3-Flash?

GLM-5.3-Flash(原代号 Ox Alpha)是智谱 AI 发布的轻量旗舰模型,采用 320B 总参数、18B 激活的混合注意力 MoE 架构,支持 100 万 Token 上下文与亚秒级响应,官方提供终身免费调用额度。

How to Access for Free (via Vultr)

No detailed access description available.

Try it in your browser

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Code Examples

No examples provided yet.

Community Feedback

Comments are tied to your GitHub account — sign in to join the discussion.