- Free Tools
- LLM Cost Estimator
- Vultr
- GLM-5.3-Flash

GLM-5.3-Flash via Vultr
Specifications
Context Window
1,048,576 tokens
Release Date
2026-08-26
Capabilities
AttachmentsReasoningTool callingStructured outputTemperatureImage inputVideo input
Availability
Open Weights
Model Overview
Vultr is a cloud infrastructure provider offering AI model inference through their serverless GPU platform, providing simple and affordable access to popular models.
GLM-5.3-Flash is a glm-flash-family model by Vultr with a 1.0M token context window and up to 131k output tokens. It is priced at $0.1000/1M input tokens and $0.3500/1M output tokens.
Key capabilities include: attachments, reasoning, tool calling, structured output, temperature, image input, video input. It supports advanced reasoning for complex multi-step tasks. It can call external tools and functions for agentic workflows.






