Underclassic
Central AI concentrates chips, power, and metered APIs. Local AI puts usable models on machines you already own. Efficiency cuts the cost of both—and decides how often you must rent.
Date: 2026-08-05
Updated: 2026-08-05 17:23
Status: LIVE
Axis
local [####################----------] central
Score: 65
central weather
+3 toward central this week
Texas grid freeze and hyperscaler capex keep central AI in the lead. Local silicon and open weights still cut edge costs, but power and cluster scale remain the binding facts.
Forces
| Force | Score | Pull | Note |
| CHIP / CAPEX | 73 | central | Hyperscaler capex and cluster announcements remain the gravity well. |
| ENERGY | 81 | central | Texas grid moratorium shows power queues now limit central buildout speed. |
| EFFICIENCY | 42 | local | List token prices keep falling; on-device runtimes keep eating rented work. |
Token prices (USD per 1M tokens; as of 2026-08-05; blended = (3*input + 1*output) / 4)
| Model | Tier | In | Out | Blended |
| Claude Sonnet 5 | workhorse | $2.00 | $10.00 | $4.00 |
| Introductory rate through 2026-08-31; standard later listed at $3/$15. | ||||
| GPT-5.5 | flagship | $5.00 | $30.00 | $11.25 |
| Standard short-context list rates. | ||||
| DeepSeek V4 Flash | commodity | $0.14 | $0.28 | $0.17 |
| Cache-miss input / output. Cache hits much lower. | ||||
List API prices only. Not a subsidy meter — labs do not publish fully loaded token COGS. Gap between cheap open/API tiers and flagship output rates is the useful public signal.
Broadcast
Texas froze new data-center grid connections after requests hit 474 GW—five times peak demand. Hyperscalers booked a $2.3 trillion cloud backlog that now funds the next round of central clusters. On the edge, M4 Max decode throughput passed Nvidia’s GB10 and AMD’s Strix Halo, and the White House dropped security review for open-weight models. Token list prices continue to fall, but the power queue, not the model list, sets the pace of central buildout.
Watch
Texas halts data center connections to power grid amid overwhelming demand
State freezes new connections after 474 GW of requests exceed peak record demand by five times. 1,800 projects now subject to audit and moratorium.
Power queues now bind how fast central AI capacity can land, tilting the axis toward central.
2026-08-04 · Tom's Hardware · energy, chip · central
Big Tech's cloud backlog just hit $2.3 trillion
Backlog feeds AI capex plans. $760 billion in 2026 spend expected across hyperscalers.
Central AI capex is locked in by booked revenue, not speculative demand.
2026-08-03 · Yahoo Finance · chip · central
Exploring Apple Silicon’s local AI performance with the Mac Studio and M4 Max
M4 Max beats GB10 and Strix Halo in decode throughput. Memory bandwidth remains the limiter for larger models.
Local silicon keeps improving decode speed, widening the set of workloads that no longer need rented tokens.
2026-07-30 · Tom's Hardware · efficiency · local
White House exempts open-weight AI models from new government security review
Policy change reduces regulatory friction for open weights. China’s open-weight lead cited as competitive factor.
Lower regulatory cost for open weights tilts efficiency gains toward local deployment.
2026-08-04 · qz.com · efficiency · local
If useful, support the broadcast.
bc1qfs3yw5qlq8sxs50crzh3g84ug27gvww9npu85n