
is now available on AI Gateway.GLM 5.3 FlashX GLM 5.3 FlashX is a high-speed serving option for Z.ai's multimodal coding model, delivering inference at ~200 tokens per second for faster streamed responses. The higher serving speed is useful for coding agents, tool loops, and interactive…
No discussion yet. Be the first to share your thoughts!