GLM-5.3 Open Weights Come With a $10B Catch
Z.ai is letting developers self-host its flagship AI model while reserving a security gate for the largest model-service operators.
Related reporting, guides, and analysis from Zeniteq.
Z.ai is letting developers self-host its flagship AI model while reserving a security gate for the largest model-service operators.
Business Insider described ongoing talks, but the transaction has not been publicly announced or closed.
The 320B-A18B open-weight model pairs native multimodality with a one-million-token context window and unusually low API prices.
Qwen3.8-Flash-Next and GLM-5.3-Flash bring recent Claude-class performance to downloadable models, but “local” still means server-grade hardware.
The 27B dense model favors practical deployment, while the 2.4-trillion-parameter MoE brings Qwen’s largest foundation model to self-hosted infrastructure.
Z.ai’s 743B-class LLM improves coding and exploit-chain performance without a new base model, while its API arrives before the open weights.
Alibaba's Qwen3.7-Plus unifies vision and language into one agent foundation that perceives, reasons, codes, and acts across GUI and CLI environments.
Bonsai Image 4B compresses a 4B diffusion transformer by up to 8.3x, making on-device image generation a practical reality for the first…