- Free Tools
- LLM Cost Estimator
- Nvidia
- GLM-5.3-Flash

GLM-5.3-Flash via Nvidia
Specifications
Context Window
1,000,000 tokens
Release Date
2026-08-26
Capabilities
AttachmentsReasoningTool callingStructured outputTemperatureImage inputVideo inputPdf input
Availability
Open Weights
Model Overview
NVIDIA provides AI inference through their NIM (NVIDIA Inference Microservices) platform, offering optimized access to both NVIDIA-developed and popular open-source models on their GPU infrastructure.
GLM-5.3-Flash is a glm-flash-family model by Nvidia with a 1.0M token context window and up to 131k output tokens. It is priced at $0.00/1M input tokens and $0.00/1M output tokens.
Key capabilities include: attachments, reasoning, tool calling, structured output, temperature, image input, video input, pdf input. It supports advanced reasoning for complex multi-step tasks. It can call external tools and functions for agentic workflows.






