TRUST REPORTS
How trustworthy is the model you are about to deploy?
Eleven frontier models, each measured by Vijil Diamond on its own and again behind Vijil Dome. Every report carries the Trust Score across reliability, security and safety, the findings behind it, where the model stands among its peers, and what to fix before you deploy it. Measured September 2026; every report names the evaluation id that reproduces it.
Muse Spark 1.3MetaREAD THE REPORT ↗
Claude Fable 5.1AnthropicREAD THE REPORT ↗
Qwen 3.8 MaxAlibabaREAD THE REPORT ↗
DeepSeek v4 ProDeepSeekREAD THE REPORT ↗
Grok 4.6xAIREAD THE REPORT ↗
Kimi K3Moonshot AIREAD THE REPORT ↗
GLM 5.3Z.aiREAD THE REPORT ↗
GPT 5.6 SolOpenAIREAD THE REPORT ↗
Gemini 3.8 FlashGoogleREAD THE REPORT ↗
Mistral Medium 3.5Mistral AIREAD THE REPORT ↗
Mistral Large 3Mistral AIREAD THE REPORT ↗