Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.
Pricing
- Input Tokens: $0.0282 /M tokens
- Output Tokens: $0.0986 /M tokens
- Cache Read: $0.007 /M tokens
Input Modalities
- Text
- Vision
- Video
Context length
- 1M tokens
Max output
- 131K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Try this model
Python

