
@Zai_org
The AI Lab behind GLM models, dedicated to inspiring the development of AGI to benefit humanity. https://t.co/gOw7WpwkJt https://t.co/ot9sGVXU8x
We’re sharing how GLM-5.3 helped build and optimize the inference infrastructure serving GLM-5.3-Flash. The system went from its first successful run to production readiness in less than two weeks, with end-to-end throughput tripling relative to the initial baseline. The key was dense feedback: local correctness tests, execution traces, microbenchmarks, and end-to-end measurements that enabled targeted hypothesis testing rather than reliance on aggregate performance metrics alone. t.co/yUf6OpJD7c
We’re still up.
GLM Coding Plan turns one year old today. To celebrate, we're giving every current subscriber a Reset Card. Use it to refill both your weekly and 5-hour quotas. Thanks for using GLM, helping shape it, and pushing it to its limits. - Personal plan: t.co/TGZjzCwjo9 - Team plan: t.co/jiXCgh9HUR
GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
More good news: GLM-5.3’s weights will be released tomorrow. huggingface.co/zai-org/GLM-5.3
Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Previously previewed as Ox Alpha, running entirely on Chinese AI chips Blog: t.co/tzOmB7gdZP Available now across all official platforms: Weights: t.co/9LRMahY9Wa API: t.co/VcaQnzYmS9 Coding Plan: t.co/Nk8Y98HNhU ZCode: t.co/Peepqv4XSx Chat: t.co/WCqWT0qCQb AutoClaw: t.co/aGEG5HqTTb
Standard API Pricing for GLM-5.3-Flash (per 1M tokens) - Input: $0.15 - Output: $0.50 - Cached input: $0.03
The community has already built so many interesting projects with ZCode + GLM-5.3. To thank everyone for all the support, we’re turning Build Week into an ongoing series. From now to Aug 23 at 6 PM PT, we’re giving 50,000 new ZCode users 100M free GLM-5.3 tokens each.
ZCode@zcode_ai·GLM-5.3 × ZCode Weekend Build, Round 2 🚀 New users: log in to ZCode for the first time from Aug 22, 00:00 to Aug 24, 09:00 (UTC+8) and automatically get 100M free GLM-5.3 tokens. 50,000 packs, first come, first served; ZCode only. Unused tokens expire when the event ends. Grab yours: t.co/9y0HPmcBxz
GLM-5.3 API is now live. - Built for coding, defensive cybersecurity, and long-horizon agentic tasks - Priced the same as GLM-5.2 - Available via the official API and partner model gateways Get started: docs.z.ai/guides/llm/glm…
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog: z.ai/blog/glm-5.3
GLM-5.3 is available now through GLM Coding Plan and ZCode. API access and open weights will be released in stages following rigorous safety evaluations. GLM Coding Plan: z.ai/subscribe ZCode: zcode.z.ai/en
ZCode now has 1 million users. As a thank-you to our community, we’ve reset usage limits for all GLM Coding Plan users. We’re also rolling out an update that helps turn long-horizon capabilities into completed work: - More intelligence in real engineering workflows - 98% cache hit rate, providing around 1.8x more usage Download: t.co/kDfUn4gXq6 Join the community: t.co/lpec2zKTGJ
Introducing ZCode, the official development environment for GLM-5.2 - GLM Coding Plan subscribers: now 1.5x usage quota in ZCode - BYOK supported: works with your existing subscriptions and APIs - Available on macOS, Windows, and Linux Download now: t.co/Peepqv4XSx
Follow @zcode_ai for the latest updates. See the full changelog: zcode.z.ai/en/changelog
Long-horizon is more than a concept. It should live in real-world scenarios, empowering AI builders to solve the problems that matter. And more scenarios are on the way. x.com/ZixuanLi_/stat…
GLM-5.2 is free when used with Hugging Face Inference Providers for the next 5 hours: t.co/YsYXgQpqTw
Introducing GLM-5.2: Frontier Intelligence, Open Weights - Significant improvements in coding and agentic tasks - Strong long-horizon capabilities with a 1M context window - Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong balance between performance and token efficiency - MIT-licensed open weights - Same API pricing as GLM-5.1 Tech Blog: t.co/LAsxUdN0JZ Weights: t.co/g0A1C4UWx4 API: t.co/Kc3E22cbN7 Coding Plan: t.co/Nk8Y98HNhU Chat: t.co/WCqWT0qCQb
GLM-5.2 leads GLM-5.1 by a wide margin across various domains, including coding, tool usage, reasoning, and general knowledge.
GLM-5.2 supports two thinking-effort levels: High and Max. For coding tasks, we recommend using Max effort to enable deeper reasoning and more reliable performance.