Long-form 23 items
Thu, 24 Sept 2026
Long-form 05:25 ET
Alpha Exchange — Amanda Lynam: GS credit on AI capex financing
GS Chief Credit Strategist: hyperscaler IG $250bn/'26→$400bn/'27; $6tn capex '26–'30; hyperscalers 40% of AI issuance; little crowding-out; IG absorb.
asr Alpha Exchange · Goldman Sachs · Meta · Bloomberg
Long-form 05:20 ET
SemiAnalysis — ClusterMAX 3.0: Nebius platinum, rankings, financing
Ep.033: 77 providers ranked; Nebius joins CoreWeave platinum; Google gold; Azure/AWS down; GPU-hour backwardation; NVDA backstop ~$588bn→$2tn; SLAs.
asr SemiAnalysis Weekly · Nebius · CoreWeave · Oracle
Wed, 23 Sept 2026
Long-form 05:15 ET
Latent Space — Diogo Almeida / Jev: System One for prod, not God
InstructGPT coauthor: frontier chat/RLHF APIs wrong for software; Jev as code-consumed System One; >1T tokens/day machine traffic; dark data + agents.
asr Latent Space · TypeSafe · Jev · OpenAI
Tue, 22 Sept 2026
Long-form 05:40 ET
ILTB — Gabe Stengel / Rogo: investing superintelligence, harness, last mile
Rogo CEO: o1 Pro→Opus 4.5 unlocked junior-analyst work; next 2–5y is firm reinvention; harness/compliance/last-mile beat raw models for buy-side.
asr Invest Like the Best · Rogo · Jane Street · Goldman Sachs
Long-form 05:30 ET
All-In — Naveen Rao: AI energy wall, 4D computing, 1000x efficiency bet
Unconventional AI CEO: Google-scale token energy already ~12GW; ~50% of token cost is power; aims 1000x efficiency in ~3.5y via dynamical chips.
asr All-In Summit · Unconventional AI · Nervana · Intel
Mon, 21 Sept 2026
Long-form 05:30 ET
Excess Returns — Jason Hsu: China AI gap, capex arms race, S&P seven
Rayliant CIO: China models on-par/open-source; energy grid edge; hardware rents until overcapacity; Mag7 CapEx arms race; S&P is one-tree, not diversifier.
asr Excess Returns · Rayliant Global Advisors · Research Affiliates · DeepSeek
Long-form 05:25 ET
Odds on Open — Lihong Wang: ex-IMC semis quant, AI stack portfolio
Ex-IMC semis options MM on flow/V, NVDA–AMD relative vol, DeepSeek corr blowups; 50-name AI stack book at ~2×; models-beat-S&P claim needs harness.
asr Odds on Open · IMC · NVIDIA · AMD
Sun, 20 Sept 2026
Long-form 05:20 ET
MiB — Glen Kacher: AI boom is catch-up, not overbuild
Light Street CIO on AI5 semis concentration, NVDA ~85% share, demand ahead of supply, 10–20y stack cycle, agents→~5× tokens; DC politics as education risk.
asr Masters in Business · Light Street Capital · NVIDIA · AMD
Sat, 19 Sept 2026
Long-form 05:20 ET
a16z — Ali Ghodsi: enterprise stall is context, not IQ
Databricks CEO on pacing PR vs cyber risk, four-test RSI bar, ontology/Genie as the adoption bind, Uni Gateway cost control + GLM shift.
asr The a16z Show · Databricks · OpenAI · Hugging Face
Long-form 05:20 ET
No Priors — Ermon/Inception: diffusion wins inference parallelism
Stefano Ermon on Mercury ≈ Haiku/Flash/mini speed tier, ~10× decode vs AR at GPT-2 scale, OpenCall leaving Cerebras for NVDA GPUs, 20–30% latency wedge.
asr No Priors · Inception · Mercury · OpenAI
Fri, 18 Sept 2026
Long-form 05:20 ET
Dwarkesh — Noam Brown: agent swarms, RSI speedup, alignment bind
OpenAI's Noam Brown on 10k-agent Navier-Stokes solve (130B tokens/88h), Ultra Mode multi-agent, Codex $7–8k/day internal, RSI ≠ 100x overnight.
transcript Dwarkesh Podcast · OpenAI · Hugging Face · Astra
Thu, 17 Sept 2026
Long-form 05:20 ET
Latent Space — AIUC: trust/liability as the agent adoption bind
Rune Kvist (ex-Anthropic) on $40M Series A: AIUC-1 quarterly agent standard, Lloyd's-backed policies, Waymo/Air Canada liability — eval+insurance stack.
transcript Latent Space · AIUC · Anthropic · Cursor
Long-form 05:15 ET
All-In — Gerstner: no AI bubble; semis = ~70% of Nasdaq return
Altimeter's Brad Gerstner on All-In: earnings-driven tape, offtake must fund Mag5 capex, Dylan 43GW too hot (~25GW), lab RR as takeoff switch. [asr]
asr All-In Podcast · NVIDIA · Anthropic · OpenAI
Wed, 16 Sept 2026
Long-form 05:20 ET
SemiAnalysis Ep.031 — pacing may eat more compute, not less
Emergency ep on Amodei pacing: OpenAI CoT monitoring ~20% of rollup compute; safety spend likely raises, not cuts, infra demand; HF as shot across bow. [asr]
asr SemiAnalysis Weekly · Anthropic · OpenAI · Hugging Face
Long-form 05:15 ET
All-In — Satya: pace with common sense; MSFT builds, leases, rents
Nadella on All-In: broad diffusion over mystical slowdown; ~30m enterprise Copilot users of ~250–300m TAM; kit ~60% of cost; Quincy DC ~400–500 MW. [asr]
asr All-In Podcast · Microsoft · OpenAI · Anthropic
Tue, 15 Sept 2026
Long-form 07:30 ET
Elon Musk & Gwynne Shotwell — AI Peer Review, Starship, Terafab, SpaceX/Tesla Merger (All-In)
SpaceX President Gwynne Shotwell says SpaceX is as much an AI business as a space business by revenue, with compute rental 'a heck of a business' and Starlink at ~1.5–2% penetration; Elon Musk joins from Memphis to push cross-lab model peer review, handicap Starship ship-catch at ~50–60%, frame Terafab as build-or-fail-to-scale, and non-deny a Tesla–SpaceX combination.
asr All-In Podcast · SpaceX · Tesla · xAI
Long-form 07:30 ET
All-In — Jensen: doomer math fails; open models carry apps
Huang on All-In: extinction %s unscientific; ~$400bn AI VC ~80% open-model; NVIDIA goes "as deep as needed." Trump brands DC opposition a hoax. [asr]
asr All-In Podcast · NVIDIA · Anthropic · OpenAI
Mon, 14 Sept 2026
Long-form 20:50 ET
Jensen Huang — Nvidia's Future, Physical AI, Rise of the Agent, Inference Explosion (All-In)
NVIDIA CEO Jensen Huang tells the All-In hosts that agentic workloads drove a ~10,000x compute step in two years, that a higher-capex Vera Rubin factory can still deliver the lowest token cost via ~10x throughput, and that Physical AI is already a near-$10bn NVIDIA line while open-weight agents redefine the desktop OS — with China licenses restarting and consensus growth paths rejected as undersized.
asr All-In Podcast · NVIDIA · Groq · Anthropic
Long-form 20:30 ET
Gavin Baker — Why AI Demand Is Outrunning Compute Supply (a16z Show)
Atreides CIO Gavin Baker tells David George that AI fundamentals accelerated through July–August while related equities drew down; argues sub-one-year compute paybacks and thin heavy-user penetration make undersupply through 2028 the base case, with NVIDIA’s financeable stack and hybrid open-source routers as the durable structure.
asr The a16z Show · NVIDIA · OpenAI · Anthropic
Long-form 17:29 ET
TBPN: The AI Slowdown Debate
Metadata-only: TBPN's Sep 14 episode (full + Diet cut) titled The AI Slowdown Debate, with guests including Nico Wittenborn, Scott Keogh, Mitchell Green, David Rosenthal, Ben Gilbert, and Faraj Aalaei.
metadata-only TBPN
Long-form 13:00 ET
SemiAnalysis Weekly: why 4-Hi HBM may win on inference economics
Metadata-only capture of SemiAnalysis Weekly Ep. 030: Myron Xie and Jordan Nanos on Rubin Ultra shipping 192GB HBM versus a 1TB preview, supply-driven decontenting, and why less memory per chip can still be the right call.
metadata-only SemiAnalysis Weekly · NVIDIA · SemiAnalysis
Long-form 12:06 ET
Eisman Playbook: Big Short partners on rates, AI, gold
Metadata-only: Steve Eisman with Vincent Daniel and Porter Collins on bonds, Treasury buybacks, OpenAI risk, gold, shorting mechanics, and two live short ideas.
metadata-only The Real Eisman Playbook · OpenAI
Long-form 12:04 ET
Latent Space: Richard Socher on recursive self-improvement
Metadata-only: Latent Space interviews Richard Socher (Recursive / You.com) on recursive self-improvement as the next major AI step; show notes flag AI×finance conference promo.
metadata-only Latent Space · You.com · Recursive
Long-form · Mon, 14 Sept 2026 · 20:50 ET

Jensen Huang — Nvidia's Future, Physical AI, Rise of the Agent, Inference Explosion (All-In)

NVIDIA CEO Jensen Huang tells the All-In hosts that agentic workloads drove a ~10,000x compute step in two years, that a higher-capex Vera Rubin factory can still deliver the lowest token cost via ~10x throughput, and that Physical AI is already a near-$10bn NVIDIA line while open-weight agents redefine the desktop OS — with China licenses restarting and consensus growth paths rejected as undersized.

asr Jensen HuangChamath PalihapitiyaJason CalacanisDavid SacksDavid Friedberg NVIDIAGroqAnthropicOpenAIMetaAWSGoogleAMDTeslaUberBYDOpenClawClaude CodeDynamoVera RubinBlackwellOmniverseCUDAMellanoxBluefieldSynopsysCadenceBittensor Source ↗
Venue: All-In PodcastHost: Chamath Palihapitiya, Jason Calacanis, David Sacks, David FriedbergDuration: 66mPublished: Thu, 19 Mar 2026 · 14:27 ET

Abstract

Jensen Huang, CEO of NVIDIA (accelerators, networking, and full-stack AI factory systems), sits with Chamath Palihapitiya, Jason Calacanis, David Sacks, and David Friedberg for a ~66-minute special All-In episode (YouTube: Jensen Huang: Nvidia's Future, Physical AI, Rise of the Agent, Inference Explosion, AI PR Crisis). Ground covered: Groq acquisition and disaggregated inference inside Dynamo / Vera Rubin; factory capex vs token cost; Physical AI and digital-biology timelines; OpenClaw as an open agentic OS; doomerism and Anthropic messaging; internal token spend norms; open vs proprietary models; China license restart and Taiwan/Middle East supply posture; customer custom silicon; analyst consensus growth; space compute; healthcare and robotics; model-company revenue and application-layer moats.

Theses

Factory list price is the wrong unit — a ~$50bn NVIDIA inference factory can still produce the lowest-cost tokens if throughput is ~10x alternatives whose shell/power/networking costs are largely shared. Huang: ~$20bn of a $50bn build is 'land power and shell'; GPU half-price does not cut the factory from $50bn to $30bn; even free chips are 'not cheap enough' if they lag the stack. [book]

Agentic processing, not chatbot Q&A, is the demand driver: generative→reasoning→agentic stacked ~100x then ~100x compute (~10,000x in two years), with people paying for work done rather than answers. Huang: agents beat storage, mix large/small/diffusion/autoregressive models, and 'get work done'; consumption ~100x with scaling 'not even started'; 'absolutely at a million X.' [book]

NVIDIA's rack TAM expanded as the company moved from one-rack GPU seller to multi-rack AI factory (Groq + Bluefield + CPUs + networking), with Groq targeted at ~25% of Vera Rubin deployments. Huang: TAM 'call it… 33%, 50% higher'; 'add Groq to about 25% of the Vera Rubins in the data center.' [book]

Physical AI is already a large, inflecting NVIDIA P&L line against a claimed ~$50T addressable industrial base, with digital biology framed as near a ChatGPT-scale moment on a multi-year clock. Huang: Physical AI 'close to $10 billion a year now' and 'growing exponentially' after a ~10-year build; biology representation of genes/proteins/cells 'two, three, five years' then healthcare inflection. [book]

OpenClaw (and Claude Code before it) is treated as the cultural and architectural proof that agents are a full computer — memory, scheduling, I/O, skills — and the open blueprint of modern computing, with governance as the hard constraint. Huang: four elements 'fundamentally define a computer'; agents should get two of three of sensitive data / code execution / external comms, not all three; NVIDIA engineers working with Peter Steinberger on security. [book]

China sales are restarting from a stated 95%→0% share collapse in the second-largest market via Lutnick-approved licenses and purchase orders, while the preferred end-state is American stack share (~90% aspiration), not universal American models. Huang: licenses approved, Chinese firms 'have given us purchase orders,' supply chain 'cranking up'; solar/rare-earth/telecom outcomes are the national-security anti-pattern. [book]

Consensus multi-year growth (Sacks cites ~30% / ~20% / ~7% into 2029) understates breadth: AI is not only the top-five hyperscalers, ~40% of NVIDIA needs full CUDA/AI-factory stack, and AWS is named as buying ~1m chips over 'the next couple of years.' Huang: gaining share via Anthropic/Meta/open models plus enterprise/edge; system difficulty, not chip ASICs, is the binding problem. [book]

Enterprise software is not destroyed by agents — seat-limited tools get ~100x agent load as Synopsys/Cadence/Blender remain the human control/ground-truth surface; application moat is deep vertical specialization, not horizontal model ownership. Huang: 'butts and seats' → agents 'banging on those tools'; moat is 'deep specialization' and specialized sub-agents trained in-house. [book]

Key math

~$50bn NVIDIA inference factory vs chatter of $25–30bn ASIC/AMD alternatives; ~$20bn of the $50bn is land/power/shell; illustrative GPU half-price moves factory $50bn→$40bn, not to $30bn, against claimed ~10x throughput (asr) [book] Token-cost framing — list factory price ≠ cost per token.

Groq on ~25% of Vera Rubin data-center deployments; NVIDIA TAM +~33–50% from expanding one rack to ~five (storage/Bluefield, Groq, CPUs, networking) (asr) [book] Disaggregated agentic factory — heterogeneous silicon mix.

Physical AI ~$10bn/year NVIDIA revenue, growing exponentially, after ~10-year investment into a stated ~$50T industry (asr) [book] Long-tail P&L — already material, not optionality.

Telecom base stations framed as part of a ~$2T industry to be absorbed into AI edge infrastructure (asr) [book] Third computer — edge/robotics/radio.

Generative→reasoning ~100x compute; reasoning→agentic another ~100x → ~10,000x in ~two years; consumption ~100x with 'million X' still the working frame (asr) [book] Demand step-function — agentic work vs chatbot tokens.

NVIDIA ~43,000 employees, 38,000 engineers; thought experiment: $500k engineer should spend ≥$250k/year on tokens (not ~$5k) (asr) [book] Internal consumption norm — tokens as CAD for knowledge work.

~40% of NVIDIA business requires full CUDA / AI-factory stack (customers 'don't know what to do with' chip-only offers); AWS ~1m chips over next couple of years on top of prior buys (asr) [book] Share/mix — system vs ASIC; named hyperscaler pull.

China: from ~95% share to ~0% in second-largest market; licenses approved (Lutnick), POs in hand (asr) [book] Geopolitical volume restart — stated, not quantified shipments.

Quotes

"You should not equate the price of the factory and the price of the tokens." — Jensen Huang [asr] [book]

"Even when the chips are free, it's not cheap enough." — Jensen Huang [asr] [book]

"If that $500,000 engineer did not consume at least $250,000 worth of tokens, I am going to be deeply alarmed." — Jensen Huang [asr] [book]

"Warning is good, scaring is less good." — Jensen Huang [asr]

"Nvidia gave up a 95% market share in the second largest market in the world and we're at zero percent." — Jensen Huang [asr] [book]

"Models is a technology, not a product." — Jensen Huang [asr] [book]

Variant perception

Priced in Inference superseding training as the scarce resource; NVIDIA as full-stack AI infrastructure vendor; customer custom silicon (TPU/Trainium) as a known competitive overhang; Physical AI / robotics as a multi-year story; open vs closed models as coexistence; China export-policy overhang on NVIDIA.

What's new The load-bearing arithmetic is token cost from factory throughput, not accelerator ASP vs ASIC: shared shell means a dearer GPU rack can still win on $/token at ~10x efficiency. Agentic demand is quantified as stacked ~100x steps (~10,000x compute in two years) with internal NVIDIA token norms (~50% of eng comp as a floor). Groq is operationalized as ~25% of Vera Rubin mix inside a multi-rack TAM expansion. OpenClaw is elevated from demo to 'operating system of modern computing,' with an explicit two-of-three agent permission rule. China is framed as restart-in-progress (licenses + POs), not permanent zero. Anthropic/OpenAI revenue paths are argued up via enterprise-software VAR distribution, not only direct seats. Analyst 2027–29 deceleration is rejected as hyperscaler-narrow.

The bear case From the conversation's own material: customer ASIC programs (Google, Amazon named) persist even if Huang claims share gains; China POs are not shipped revenue and policy can reverse; helium and Taiwan concentration remain acknowledged risks; robotics still needs 'two, three cycles' (~3–5 years) and depends on Chinese motors/magnets/rare earth; space data centers remain exploratory with radiation-cooling cost; doomerism/popularity (Sacks cites ~17% US AI popularity) can still produce data-center moratoria; Chamath's ~$350bn+ revenue / ~$200bn FCF line is host assertion in the room, not Jensen guidance; 'million X' and ~10x factory throughput are CEO claims without disclosed customer unit economics here.

Discount Huang is CEO of NVIDIA talking NVIDIA product mix, TAM, share, China restart, and factory economics — all [book]. Hosts are All-In principals with disclosed AI/tech investment exposure and a summit-adjacent live audience; Sacks holds a White House AI/crypto czar role in this window and steers policy framing. Anthropic/OpenAI revenue optimism and 'enterprise software as VAR' expand the demand narrative that fills NVIDIA factories. Banter lines (revenue-while-sitting-here) are explicitly walked back as non-guidance.

Positioning read

ai-capex-durability — STRENGTHENS. Agentic ~10,000x compute step, 'million X' inference frame, Vera Rubin factory throughput argument, named AWS million-chip pull, and rejection of sharp consensus deceleration all treat sustained build as the base case rather than a digest-the-capex pause.

enterprise-agent-stall — WEAKENS. Huang and Friedberg describe agents already doing software and research work (Claude Code / OpenClaw, internal token floors, Friedberg's 90-minute stack replacement); the stall thesis's 'pilot only' frame is contested by production-flavored anecdotes, though no Fortune 500 names a $ or headcount figure here.

Frameworks

Factory price ≠ token cost (2026-03-19, Huang on All-In). Shared land/power/shell/networking means accelerator ASP gaps compress at the factory level; winning metric is tokens per watt/dollar at system throughput, not chip list price. [book]

Three computers (2026-03-19, Huang on All-In). Training system; evaluation/simulation (Omniverse, physics-faithful gym); edge/robotics/telecom radio. Physical AI needs all three. [book]

Agent permission triad (2026-03-19, Huang on All-In). Sensitive data, code execution, external communication — allow any two, not all three simultaneously — as the governance pattern for open agentic runtimes.

The Open/Close  ·  Research commentary, not investment advice. Positions may be held in securities mentioned.