Quasar v2
Flagship ยท from GLM-5.3 ยท Available through the Multiverse Computing API
Our flagship closed-source model, released together with this report. It applies the full CompactifAI pipeline to GLM-5.3: structural compression, knowledge-distillation healing, and quantization-aware healing stacked on top. The model is available through our API and is not distributed as open weights.
The most common description of a model like this in the field is a pruned model, but that does not capture what Quasar is or what producing it involved. Compression removes parameters, but the model the user gets is the one that emerged from healing, quantization-aware healing, and the synthetic data pipeline that feeds both. Those stages are where the engineering investment landed, and they are what shape the model's behaviour. The base checkpoint is the starting point. The profiling analysis, compression configuration, distillation loss, data tools, quantization bridge, and the systems work to operate them at this scale are ours. Producing Quasar was not a free win; it was a directed engineering investment.
Beyond the standard pipeline, Quasar v2 received targeted work specific to this release. The GLM base model carries content restrictions on Chinese political topics that reduce its usefulness for general-purpose deployments. We removed those restrictions. We also applied security improvements, whose details will be documented as they are finalised.
Artificial Analysis independently benchmarked Quasar v2. The full report and charts will appear in this section once published.
GLM-5.3
base model
3 stages
compression, healing and quantization-aware healing
API only
not distributed as open weights
Pending
Artificial Analysis Report
