Frontier models on the SHAD 2026 exam

General News

Summary

This article compares frontier LLM performance on a 2026 exam in the SHAD program. It ranks models such as ChatGPT, Gemini, Claude, Qwen, DeepSeek, YandexGPT, GigaChat, and GLM across multiple tasks and shows that DeepSeek and Qwen lead in several areas. The text also contrasts 2026 results with 2025 results to highlight model progress and relative shifts in performance. It is primarily a benchmark-style analysis of AI model capabilities rather than a company announcement.

Classifications

industries
No industries detected
applications
AI & Machine learning

AI Classifications

Labels
AI Software SaaS Conversational AI

Linked Companies

ChatGPT
up to $1M
Z.ai
$50M to $100M
Qwen
$5M to $10M