# "GPU demand" — X 热门讨论 (2026-10-01 14:48 UTC)
## @starmexxx (starmex) · 10-01 08:55 · ♥30 ↻3 💬6 JENSEN HUANG IS WORRIED ABOUT THE WRONG THING
everyone thinks the threat to nvidia is a cheaper chip out of china. it isn't. it's a model that refuses to write a single word
sam altman is building stargate. elon is stacking gpus inside colossus. dario keeps signing compute deals bigger than most countries' budgets. all of it rests on one assumption: every ai call needs a giant model generating tokens
jev just broke that assumption
it's the model the guy who invented rlhf built after he walked out of openai. it doesn't generate, it decides. classify, route, score, approve, block. 70 milliseconds, output tokens free, nothing to hallucinate
and here's the part wall street hasn't priced in yet:
→ pull the logs from any company running agents and 70-80% of the calls are decisions, not writing → every one of those calls currently runs on a frontier model, on nvidia silicon, billed by the token → jev makes the same call on roughly 1/200th of the compute → move that 80% over and most of the gpu demand behind it quietly evaporates
the data centers don't disappear. they just turn out to be massively oversized for what the work actually needed
i moved my own stack last month. 2.1M calls a day, 1.7M of them were decisions. my bill on those dropped from $11,400 to $63
multiply that by every company running agents and you get the most expensive misunderstanding in tech: we built trillion-dollar infrastructure to answer yes or no
the labs will keep the writing. jev is quietly taking everything else
bookmark this. in 12 months it's either the dumbest take you read this year or the one you wish you'd acted on > 引用 @leopardracer: How to Build an AI Market Analyst on Kimi K3 and GPT-6 Astra: The Full Architecture https://x.com/starmexxx/status/2105582267074544022