GLM-5.3-Flash by @Zai_org, aka Ox Alpha, is now live in LM Studio Bionic!
This model surpasses GLM 5.2 in performance while being 9-10x cheaper. It also supports image input.
In Bionic, this model is served from US-based servers with ZDR enabled by default.
Happy building! ⚡️
Introducing GLM-5.3-Flash
- Leading capabilities at a highly competitive price
- Natively multimodal with a 1M-token context window
- A 320B-A18B model released under the MIT License
- Previously previewed as Ox Alpha, running entirely on Chinese AI chips
Blog:
Bionic now supports agent skills.
Skills are reusable shortcuts that can capture a task, specific knowledge, or custom instructions.
Use, install, and create skills - just ask Bionic to do it for you, or use the @ syntax in the composer.
DeepSeek V4 Pro 0813 is now live in Bionic!
This model shows strong performance across agentic coding and automation tasks at a very compelling cost.
Hosted on US-based servers 🇺🇸, with ZDR enabled by default.
We’re launching DeepSeek-V4-Pro today! 🚀
🔷 Major Agent upgrades with strong production gains!
🔷 Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, max for complex tasks.
🔷 Native OpenAI Responses API support, optimized for
We promised open weights for Qwen3.8. Now, time to meet them! 🎉
⚡ Qwen3.8-27B:
- A native multimodal dense model. With just 27B parameters, it outperforms Qwen3.7-Plus overall and shines in real-world coding & office workflows.
- 262K native context, easily extendable to 1M