Underclassic
Central AI concentrates chips, power, and metered APIs. Local AI puts usable models on machines you already own. Efficiency cuts the cost of both — and decides how often you must rent.

Date: 2026-08-04
Updated: 2026-08-04 21:30
Status: LIVE


Axis
local [###################-----------] central
Score: 62   central weather   +3 toward central this week
Buildout and power deals still dominate the tape. Local silicon and open weights keep cutting costs at the edge, but the loudest facts remain campus-scale.

Forces

ForceScorePullNote
CHIP / CAPEX 71 central Hyperscaler capex and cluster announcements remain the gravity well.
ENERGY 78 central Grid queues and multi-GW power hunts bind how fast central capacity can actually land.
EFFICIENCY 44 local List token prices keep falling across tiers; on-device runtimes keep eating rented work.

Token prices (USD per 1M tokens; as of 2026-08-04; blended = (3*input + 1*output) / 4)

ModelTierInOutBlended
Claude Sonnet 5 workhorse $2.00 $10.00 $4.00
Introductory rate through 2026-08-31; standard later listed at $3/$15.
GPT-5.5 flagship $5.00 $30.00 $11.25
Standard short-context list rates.
DeepSeek V4 Flash commodity $0.14 $0.28 $0.17
Cache-miss input / output. Cache hits much lower.

List API prices only. Not a subsidy meter — labs do not publish fully loaded token COGS. Gap between cheap open/API tiers and flagship output rates is the useful public signal.


Broadcast

Do not confuse announcement GW with delivered intelligence. Blueprints print faster than substations pour concrete; quantizations print faster than procurement notices them. Three dials: who locks long power, who drives $/M tokens down at usable quality, and who can finish real work on the machine already on the table. There is no honest public 'subsidy %' for tokens — only list prices, quality-held deflation, and the rent gap versus running weights yourself.

Watch


U.S. data-center permits point to a sharp jump in electricity hunger
A Business Insider permit analysis puts potential annual load from recently permitted sites in the hundreds of TWh if projects complete — and flags 2026 hyperscaler capex above $600B, mostly into buildout.
Central capacity is not abstract. It files for power. Completed campuses harden rented inference as default infrastructure.
2026-06-07 · Business Insider · energy, chip · central

Paper GW vs real electrons: buildout meets grid and politics
Forbes coverage of the 2026 buildout stresses power/water limits and notes industry skepticism that announced megaproject pipelines are fully real — one builder estimate put only about one in four projects as likely.
Central ambitions exceed the grid's metabolism. Delay is not decentralization, but it is time for local runtime and cheaper tokens to take share.
2026-05-19 · Forbes · energy, chip · mixed

On-device runtimes keep turning laptops into small foundries
2026 local-LLM guides still center Apple silicon unified memory and Neural Engine class NPUs: practical 14B–70B class work on high-RAM Macs, with MLX/Ollama-style stacks as the common toolkit.
When a desk can host serious inference, the question shifts from which API to rent toward what still must stay rented.
2026-03-18 · AI Magicx / local LLM guides · efficiency, chip · local

Reality check from the workshop: local quality rises, harness still decides
Practitioners on high-RAM Apple silicon report strong local coding/STEM performance from large MoE/open weights, while admitting frontier always-on agent polish still favors cloud harnesses — for now.
Local wins in pieces: weights and watts-per-token first, then tools and habits. Central products still sell the integrated workflow.
2026-03-08 · r/LocalLLM · efficiency · local


If useful, support the broadcast.
bc1qfs3yw5qlq8sxs50crzh3g84ug27gvww9npu85n
BTC QR

data