Long-form 23 items
Thu, 24 Sept 2026
Long-form 05:25 ET
Alpha Exchange — Amanda Lynam: GS credit on AI capex financing
GS Chief Credit Strategist: hyperscaler IG $250bn/'26→$400bn/'27; $6tn capex '26–'30; hyperscalers 40% of AI issuance; little crowding-out; IG absorb.
asr Alpha Exchange · Goldman Sachs · Meta · Bloomberg
Long-form 05:20 ET
SemiAnalysis — ClusterMAX 3.0: Nebius platinum, rankings, financing
Ep.033: 77 providers ranked; Nebius joins CoreWeave platinum; Google gold; Azure/AWS down; GPU-hour backwardation; NVDA backstop ~$588bn→$2tn; SLAs.
asr SemiAnalysis Weekly · Nebius · CoreWeave · Oracle
Wed, 23 Sept 2026
Long-form 05:15 ET
Latent Space — Diogo Almeida / Jev: System One for prod, not God
InstructGPT coauthor: frontier chat/RLHF APIs wrong for software; Jev as code-consumed System One; >1T tokens/day machine traffic; dark data + agents.
asr Latent Space · TypeSafe · Jev · OpenAI
Tue, 22 Sept 2026
Long-form 05:40 ET
ILTB — Gabe Stengel / Rogo: investing superintelligence, harness, last mile
Rogo CEO: o1 Pro→Opus 4.5 unlocked junior-analyst work; next 2–5y is firm reinvention; harness/compliance/last-mile beat raw models for buy-side.
asr Invest Like the Best · Rogo · Jane Street · Goldman Sachs
Long-form 05:30 ET
All-In — Naveen Rao: AI energy wall, 4D computing, 1000x efficiency bet
Unconventional AI CEO: Google-scale token energy already ~12GW; ~50% of token cost is power; aims 1000x efficiency in ~3.5y via dynamical chips.
asr All-In Summit · Unconventional AI · Nervana · Intel
Mon, 21 Sept 2026
Long-form 05:30 ET
Excess Returns — Jason Hsu: China AI gap, capex arms race, S&P seven
Rayliant CIO: China models on-par/open-source; energy grid edge; hardware rents until overcapacity; Mag7 CapEx arms race; S&P is one-tree, not diversifier.
asr Excess Returns · Rayliant Global Advisors · Research Affiliates · DeepSeek
Long-form 05:25 ET
Odds on Open — Lihong Wang: ex-IMC semis quant, AI stack portfolio
Ex-IMC semis options MM on flow/V, NVDA–AMD relative vol, DeepSeek corr blowups; 50-name AI stack book at ~2×; models-beat-S&P claim needs harness.
asr Odds on Open · IMC · NVIDIA · AMD
Sun, 20 Sept 2026
Long-form 05:20 ET
MiB — Glen Kacher: AI boom is catch-up, not overbuild
Light Street CIO on AI5 semis concentration, NVDA ~85% share, demand ahead of supply, 10–20y stack cycle, agents→~5× tokens; DC politics as education risk.
asr Masters in Business · Light Street Capital · NVIDIA · AMD
Sat, 19 Sept 2026
Long-form 05:20 ET
a16z — Ali Ghodsi: enterprise stall is context, not IQ
Databricks CEO on pacing PR vs cyber risk, four-test RSI bar, ontology/Genie as the adoption bind, Uni Gateway cost control + GLM shift.
asr The a16z Show · Databricks · OpenAI · Hugging Face
Long-form 05:20 ET
No Priors — Ermon/Inception: diffusion wins inference parallelism
Stefano Ermon on Mercury ≈ Haiku/Flash/mini speed tier, ~10× decode vs AR at GPT-2 scale, OpenCall leaving Cerebras for NVDA GPUs, 20–30% latency wedge.
asr No Priors · Inception · Mercury · OpenAI
Fri, 18 Sept 2026
Long-form 05:20 ET
Dwarkesh — Noam Brown: agent swarms, RSI speedup, alignment bind
OpenAI's Noam Brown on 10k-agent Navier-Stokes solve (130B tokens/88h), Ultra Mode multi-agent, Codex $7–8k/day internal, RSI ≠ 100x overnight.
transcript Dwarkesh Podcast · OpenAI · Hugging Face · Astra
Thu, 17 Sept 2026
Long-form 05:20 ET
Latent Space — AIUC: trust/liability as the agent adoption bind
Rune Kvist (ex-Anthropic) on $40M Series A: AIUC-1 quarterly agent standard, Lloyd's-backed policies, Waymo/Air Canada liability — eval+insurance stack.
transcript Latent Space · AIUC · Anthropic · Cursor
Long-form 05:15 ET
All-In — Gerstner: no AI bubble; semis = ~70% of Nasdaq return
Altimeter's Brad Gerstner on All-In: earnings-driven tape, offtake must fund Mag5 capex, Dylan 43GW too hot (~25GW), lab RR as takeoff switch. [asr]
asr All-In Podcast · NVIDIA · Anthropic · OpenAI
Wed, 16 Sept 2026
Long-form 05:20 ET
SemiAnalysis Ep.031 — pacing may eat more compute, not less
Emergency ep on Amodei pacing: OpenAI CoT monitoring ~20% of rollup compute; safety spend likely raises, not cuts, infra demand; HF as shot across bow. [asr]
asr SemiAnalysis Weekly · Anthropic · OpenAI · Hugging Face
Long-form 05:15 ET
All-In — Satya: pace with common sense; MSFT builds, leases, rents
Nadella on All-In: broad diffusion over mystical slowdown; ~30m enterprise Copilot users of ~250–300m TAM; kit ~60% of cost; Quincy DC ~400–500 MW. [asr]
asr All-In Podcast · Microsoft · OpenAI · Anthropic
Tue, 15 Sept 2026
Long-form 07:30 ET
Elon Musk & Gwynne Shotwell — AI Peer Review, Starship, Terafab, SpaceX/Tesla Merger (All-In)
SpaceX President Gwynne Shotwell says SpaceX is as much an AI business as a space business by revenue, with compute rental 'a heck of a business' and Starlink at ~1.5–2% penetration; Elon Musk joins from Memphis to push cross-lab model peer review, handicap Starship ship-catch at ~50–60%, frame Terafab as build-or-fail-to-scale, and non-deny a Tesla–SpaceX combination.
asr All-In Podcast · SpaceX · Tesla · xAI
Long-form 07:30 ET
All-In — Jensen: doomer math fails; open models carry apps
Huang on All-In: extinction %s unscientific; ~$400bn AI VC ~80% open-model; NVIDIA goes "as deep as needed." Trump brands DC opposition a hoax. [asr]
asr All-In Podcast · NVIDIA · Anthropic · OpenAI
Mon, 14 Sept 2026
Long-form 20:50 ET
Jensen Huang — Nvidia's Future, Physical AI, Rise of the Agent, Inference Explosion (All-In)
NVIDIA CEO Jensen Huang tells the All-In hosts that agentic workloads drove a ~10,000x compute step in two years, that a higher-capex Vera Rubin factory can still deliver the lowest token cost via ~10x throughput, and that Physical AI is already a near-$10bn NVIDIA line while open-weight agents redefine the desktop OS — with China licenses restarting and consensus growth paths rejected as undersized.
asr All-In Podcast · NVIDIA · Groq · Anthropic
Long-form 20:30 ET
Gavin Baker — Why AI Demand Is Outrunning Compute Supply (a16z Show)
Atreides CIO Gavin Baker tells David George that AI fundamentals accelerated through July–August while related equities drew down; argues sub-one-year compute paybacks and thin heavy-user penetration make undersupply through 2028 the base case, with NVIDIA’s financeable stack and hybrid open-source routers as the durable structure.
asr The a16z Show · NVIDIA · OpenAI · Anthropic
Long-form 17:29 ET
TBPN: The AI Slowdown Debate
Metadata-only: TBPN's Sep 14 episode (full + Diet cut) titled The AI Slowdown Debate, with guests including Nico Wittenborn, Scott Keogh, Mitchell Green, David Rosenthal, Ben Gilbert, and Faraj Aalaei.
metadata-only TBPN
Long-form 13:00 ET
SemiAnalysis Weekly: why 4-Hi HBM may win on inference economics
Metadata-only capture of SemiAnalysis Weekly Ep. 030: Myron Xie and Jordan Nanos on Rubin Ultra shipping 192GB HBM versus a 1TB preview, supply-driven decontenting, and why less memory per chip can still be the right call.
metadata-only SemiAnalysis Weekly · NVIDIA · SemiAnalysis
Long-form 12:06 ET
Eisman Playbook: Big Short partners on rates, AI, gold
Metadata-only: Steve Eisman with Vincent Daniel and Porter Collins on bonds, Treasury buybacks, OpenAI risk, gold, shorting mechanics, and two live short ideas.
metadata-only The Real Eisman Playbook · OpenAI
Long-form 12:04 ET
Latent Space: Richard Socher on recursive self-improvement
Metadata-only: Latent Space interviews Richard Socher (Recursive / You.com) on recursive self-improvement as the next major AI step; show notes flag AI×finance conference promo.
metadata-only Latent Space · You.com · Recursive
Long-form · Tue, 22 Sept 2026 · 05:30 ET

All-In — Naveen Rao: AI energy wall, 4D computing, 1000x efficiency bet

Unconventional AI CEO: Google-scale token energy already ~12GW; ~50% of token cost is power; aims 1000x efficiency in ~3.5y via dynamical chips.

asr Naveen RaoChamath Palihapitiya Unconventional AINervanaIntelMosaicMLDatabricksGoogleNVIDIA Source ↗
Venue: All-In SummitHost: Chamath PalihapitiyaDuration: ~25mPublished: Mon, 21 Sept 2026 · 17:03 ET

Opening

Energy — not floor space or GPUs — is the binding DC constraint: Google-class token loads already imply ~12GW at ~10J/token, ~50% of serving cost is power, and Rao's Unconventional AI is pitching a dynamical/"4D" chip stack for ~1000× efficiency within ~3.5 years. All-In Summit stage interview with Naveen Rao (Nervana→Intel AI group; MosaicML→Databricks; now Unconventional AI CEO). Ground covered: public energy/token math, biology-vs-GPU bit-movement, first physical dynamical prototype (Jan–Jun tape-out), product path as a rack/tokens-in-tokens-out data-center system (~2 years), porting at model layer not CUDA ops. YouTube auto-captions (asr — "Nervana" as "Nirvana," Mosaic/Databricks figures unverified). Watch.

Key takes

DC planning has flipped to energy-first: floor space → networking → GPUs → power contracts, and operators must "monetize every watt." Rao: US data centers ~40GW today; world under ~100GW; one public Google figure of 3.2 quadrillion tokens/month × ~10J/token (his lower-end assumption) ≈ 12GW for that company's AI services alone — so larger models + rising demand hit an energy wall in "~3 years" on his estimate. [asr]

About half of token serving cost is energy — so efficiency is the business case, not a side metric. Frame: gap between exponentially growing AI market (he sketches ~trillion-dollar by 2030) and roughly linear energy supply is the problem Unconventional claims to close by monetizing watts "~1000× better." [asr]

Root inefficiency is bit movement, not arithmetic: cortex ~16B bits/sec vs high-end GPU ~30T bits in/out of memory per second (plus 10–100× more on-chip). Biology runs on ~eight milliwatts for a "squirrel brain" class system; synthetic stacks burn energy shuttling state. Thesis: collapse memory/compute into dynamical elements so you stop paying the von Neumann tax. [asr]

Unconventional's bet is a non–von Neumann "dynamical computer" / "4D computing" (time + 3D die stack) with a claimed first physical prototype taped out Jun 1, built in ~5 months from a Jan start. Claims: images from the chip at ~500 nanojoules/image vs GPU-class millijoules; sparsity that both cuts n² connections and improves trainability; goal revised from 5y to ~3.5y to hit ~1000× power efficiency / approach 2D-lithography limits; "beat biology" over a decade; eventual shift from gigawatt campuses to many small local DCs + robot forms. [asr]

Product path: ~2 years to a full data-center rack product (tokens in/out over the network; different guts); existing models work via model-layer port, not op-level CUDA clone — "fair bit of compute" to transition. Software bridge described as Python libraries for time-varying/stochastic elements, not a CUDA equivalent. Chamath presses ecosystem/fab path; Rao leans on making the efficiency gain large enough that migration pain is worth it. [asr]

Track record cited as credibility: Nervana sold "way too early" to Intel (ran Intel AI group); MosaicML ~$20M → ~$700–800M revenue scale then Databricks deal; claims Mosaic became "~a quarter of total revenue" at Databricks. Used to underwrite that he has shipped AI infra before — not proof the dynamical substrate works at scale. [asr]

Key math

Google ~3.2 quadrillion tokens/month (asr — Rao citing public Google) — scale anchor for energy. [asr]

~10 joules/token (asr — Rao, "lower end") → ~12GW for that Google AI load — one-company energy claim. [asr]

US DC energy ~40GW; world DC energy under ~100GW (asr) — capacity ceiling framing. [asr]

~50% of token serving cost is energy (asr) — cost stack. [asr]

Cortex ~16B bits/sec vs GPU ~30T bits/sec memory traffic (asr; on-chip "10, 100×" more) — bit-movement gap. [asr]

Efficiency goal: ~1000× in ~3.5 years (was 5) (asr) — company target. [asr]

Prototype: ~500 nJ/image vs GPU-order mJ (asr) — early efficiency claim. [asr]

MosaicML ~$20M → ~$700–800M before Databricks; "~quarter" of Databricks revenue (asr) — prior-company scale. [asr]

Quotes

"The US puts about 40 gigawatts of energy into data centers today… 12 gigawatts is going into one company just for AI services." — Naveen Rao [asr]

"About 50% of the cost of serving a token… is energy." — Naveen Rao [asr]

"Today, it's about energy. First, you think about energy." — Naveen Rao [asr]

"We're going to run out of energy pretty fast, in like 3 years or so is my estimate." — Naveen Rao [asr]

"This is actually the first physical dynamical computer ever built." — Naveen Rao [asr]

"If you make something 1/1000 the price, you'll consume more than 1/1000 of it." — Naveen Rao [asr]

Variant perception

Priced in — Power is the new DC bottleneck; hyperscaler GW deals dominate the tape; Jevons/efficiency → more demand is a familiar bull frame; NVDA von Neumann GPU stack is the incumbent.

What's new — Concrete Google-token → GW arithmetic stated on stage; ~50% energy share of token cost as a cost-stack claim; public claim of a working dynamical prototype with nJ/image figures and a 2-year rack product path; explicit model-layer (not CUDA) port strategy.

Bear case — Prototype ≠ production rack at GW scale; 1000× / 3.5y may be fundraising physics; model-layer port friction could strand the product; if energy wall is solved by nuclear/gas interconnects instead, unconventional substrate stays niche; ASR may garble joule and revenue figures.

Discount — Founder launching Unconventional AI at All-In Summit — maximum book-talking. Prior exits (Nervana, Mosaic) buy credibility but also a pattern of selling into platforms rather than owning the long substrate cycle. Anti-doomer venue selection bias. No independent third-party validation of the chip results on stage.

Positioning

AI capex durability — STRENGTHENS near-term, WEAKENS long-duration if thesis lands. Near-term: energy-first DC math and "~3 years" wall reinforce continued power/build spend. Long-duration: if ~1000× efficiency arrives, GW-campus intensity and GPU-watt monetization soften — spend shifts from raw megawatts toward new substrate/racks (Jevons still grows token demand).

Inference margin inversion — STRENGTHENS (soft). Energy as ~half of token cost means serving GM is power-bound; any real joules-per-token collapse is the cost side of the inversion — without lab GM disclosure.

HBM supply binds — NEUTRAL / soft WEAKENS if dynamical memory-compute lands. Thesis attacks von Neumann memory traffic; does not claim HBM relief on today's GPU path — only that a different architecture could unbind bit-movement. No near-term HBM volume evidence.

Enterprise agent stall — NEUTRAL. No enterprise deployment evidence; robotics/local-DC vision is forward narrative only.

The Open/Close  ·  Research commentary, not investment advice. Positions may be held in securities mentioned.