Find and compare
Benchmarks and prices
Releases and labs
Learn and sources
Model Directory
Tracked models by provider, price, and status
Leaderboard
Frontier, value, speed, and evaluated views
Release Desk
Newly tracked launches and briefs
Benchmark Library
External eval sources and caveats
Methodology
How AIRH scores and evidence states work
Reliability Floor
Everyday trust basics
News Desk
Latest captured AI signal
Updates
Data freshness and site changes
Changelog
Recent hub activity
LLM Compare
Prices, speed, context, and quality
Head-to-Head
Compare two models directly
Best Value
Quality per blended cost
Guides
Practical explainers and buying decisions
AI Academy
Structured learning path
Prompt Library
Reusable prompt patterns
Dossier / live reference cards
Hourly source check is green.
AI Resource Hub
# route boot: tags / reasoning
airh > since-last-visit
now Jobs market snapshot refreshed
now Recomputed benchmark-weighted quality scores
now Synced Chatbot Arena benchmark track
now Updated speed measurements
airh > open /tags/reasoning/
Models with advanced logical and mathematical reasoning
19 resources tagged
Claude 3.7 Sonnet
Anthropic
DeepSeek R1
DeepSeek
R1 0528
Gemini 2.5 Pro
Google
GPT-5
OpenAI
GPT-5.2
Grok 4
xAI
o1
o1-mini
O3
o3 Mini
O3 Pro
O4 Mini
Phi-4 Reasoning
Microsoft
QwQ 32B
Alibaba
AIME 2025
reasoning — American Invitational Mathematics Examination 2025
ARC Challenge
reasoning — AI2 Reasoning Challenge — grade school science
GPQA Diamond
reasoning — Graduate-level science questions, expert-validated
MATH-500
reasoning — Competition-level mathematics problems