Listen — the 60-second summary
Where AI cost, value & competition are heading · narrated
Read the transcript

The cost of using AI is collapsing, even as the valuations of AI companies — and the money they spend on infrastructure — keep soaring. Chinese providers now offer genuinely capable models and coding plans at a fraction of Western prices, raising a real question: can trillion-dollar valuations built on frontier-model scarcity actually last?

The likely answer is that value shifts away from owning a model, toward controlling compute, distribution, and the cheapest capable supply at scale. The practical strategy is simple — use inexpensive models for routine work, and reserve the premium frontier systems for the tasks that genuinely need them.

This site is built by Bo Shang, an independent builder working where cybersecurity, applied AI, and Chinese technology meet — behind Trenchwork, Vigil, Women Who Defend, and Erosolar — looking to join an AI engineering team ready for hard problems, technical and human.

Live capability tracker · updated daily

Every frontier model,
ranked by what it can actually do.

Benchmarks and real-world user reports for the world's leading AI models — Claude Fable 5, GPT-5.5, Gemini 3.1 Pro, Grok 4.3, DeepSeek V4-Pro and the best Chinese open weights — web-verified and refreshed daily.

loading… · models tracked data sources Artificial Analysis + provider docs as of
00

Price-Adjusted Ranking — intelligence per dollar

Ranked by measured intelligence per dollar. Switch pricing scenarios; the full methodology, component math and data sources are published below.

Capability leader
Best open-weight
Top community rating
Headline event
01

Capability leaderboard

Ranked by aggregate frontier capability. Toggle geography, sort any column (including full in/out pricing), search by name. Bars are normalized within each column. Input/Output prices baked for all models.

Loading leaderboard…
02

Benchmark boards

State-of-the-art per evaluation. Top models on each board with current scores.

03

What users actually report

Aggregated sentiment from LMArena, r/LocalLLaMA, Hacker News and developer leaderboards — the gap between benchmark scores and lived experience.

04

Release & events timeline

Dated model launches and field-shaping events, newest first.

05

Fable 5 / Mythos 5 watch

Full analysis of Anthropic's June 2026 Mythos-class launch and the US-government suspension directive.

06

China Catch-Up Watch

US vs China frontier gap, lab profiles, DeepSeek details, catch-up forecast, hardware constraints (NVIDIA vs Huawei), plus integrated usChina scoreboard, military AI / autonomy programs, and cyber deep-dive analysis — all from the single main dataset and served on this Firebase Hosting deployment.

Summary & Gap

Chinese Labs vs US Frontier

DeepSeek V4-Pro Profile

Catch-up Forecast

Hardware & Constraints (NVIDIA vs Huawei)

US–China Scoreboard & Dimensions

Military AI / Autonomy Programs (China)

Cyber / Technical Deep-Dive Highlights

All China analysis, usChina, militaryAI, deepDive, hardware and related data live in the single unified seed + Firestore dataset and are served exclusively from this main site (one Firebase Hosting deployment). No separate China site or hosting.