Read the transcript
The cost of using AI is collapsing, even as the valuations of AI companies — and the money they spend on infrastructure — keep soaring. Chinese providers now offer genuinely capable models and coding plans at a fraction of Western prices, raising a real question: can trillion-dollar valuations built on frontier-model scarcity actually last?
The likely answer is that value shifts away from owning a model, toward controlling compute, distribution, and the cheapest capable supply at scale. The practical strategy is simple — use inexpensive models for routine work, and reserve the premium frontier systems for the tasks that genuinely need them.
This site is built by Bo Shang, an independent builder working where cybersecurity, applied AI, and Chinese technology meet — behind Trenchwork, Vigil, Women Who Defend, and Erosolar — looking to join an AI engineering team ready for hard problems, technical and human.
Every frontier model,
ranked by what it can actually do.
Benchmarks and real-world user reports for the world's leading AI models — Claude Fable 5, GPT-5.5, Gemini 3.1 Pro, Grok 4.3, DeepSeek V4-Pro and the best Chinese open weights — web-verified and refreshed daily.
Price-Adjusted Ranking — intelligence per dollar
Ranked by measured intelligence per dollar. Switch pricing scenarios; the full methodology, component math and data sources are published below.
Capability leaderboard
Ranked by aggregate frontier capability. Toggle geography, sort any column (including full in/out pricing), search by name. Bars are normalized within each column. Input/Output prices baked for all models.
Benchmark boards
State-of-the-art per evaluation. Top models on each board with current scores.
What users actually report
Aggregated sentiment from LMArena, r/LocalLLaMA, Hacker News and developer leaderboards — the gap between benchmark scores and lived experience.
Release & events timeline
Dated model launches and field-shaping events, newest first.
Fable 5 / Mythos 5 watch
Full analysis of Anthropic's June 2026 Mythos-class launch and the US-government suspension directive.
China Catch-Up Watch
US vs China frontier gap, lab profiles, DeepSeek details, catch-up forecast, hardware constraints (NVIDIA vs Huawei), plus integrated usChina scoreboard, military AI / autonomy programs, and cyber deep-dive analysis — all from the single main dataset and served on this Firebase Hosting deployment.