Discussions 17 items
Wed, 23 Sept 2026
Discussion 07:28 ET
Is 16Hi HBM really pushed to 2029 — and does that extend 12Hi tight?
jukan05: JPM HBM model now sees 16Hi adoption 2029 earliest. Replies split on earlier slip vs prolonged 8/12Hi scarcity; no JPM primary in-thread.
top-replies JPMorgan · SK Hynix · Samsung
Tue, 22 Sept 2026
Discussion 07:26 ET
Is EMIB taking real CoWoS share — or just relocating the HBM queue?
SemiAnalysis ChipBook: Korea→Malaysia HBM +467% YoY, ~$3B/2mo via Penang. Replies split: customers fleeing CoWoS scarcity vs queue transfer / qualify risk.
top-replies Intel · TSMC
Thu, 17 Sept 2026
Discussion 15:40 ET
Are AI-infra vendor leads durable when every roadmap hits FFWD?
Nick Dorsey from AI Infra Summit: vendor roadmaps overlap, Astera/CXL example — today's technical leads feel wobbly under industry-wide acceleration.
top-replies Astera Labs
Discussion 15:40 ET Updated
Is AMD MI355X closing agentic-inference TCO vs GB300 — or only on some stacks?
SemiAnalysis says MI355X is quickly closing perf/TCO vs GB300 in agentic inference via SGLang/MoRI/UMBP; replies split on InferenceX $/M-token snapshots and software vs silicon.
top-replies AMD · NVIDIA · SemiAnalysis
Discussion 07:28 ET
Does CPO kill M8/M9 CCL upgrades — or is that scale-out/scale-up confusion?
Market rumor: CPO cuts ultralow-loss CCL demand. Jeff Pu says upgrade stays — CPO not on compute board soon; NVSwitch still needs high-spec PTFE CCL.
top-replies NVIDIA
Discussion 07:25 ET
Did OpenAI flip OpenRouter share vs Anthropic from 20/80 to 50/50?
Baker cites OpenRouter: OpenAI 20%→50% vs Anthropic since June. Talia notes token vs revenue share diverging; Baker agrees OpenAI is discounting.
top-replies OpenAI · Anthropic · OpenRouter
Wed, 16 Sept 2026
Discussion 07:22 ET
Is the Hynix–Intel story really a hyperscaler memory JV — or just talks?
jukan05 says the load-bearing angle is a possible Hynix–Intel JV with hyperscalers hungry for memory; replies stress prior denials, cycle cost, and that talks ≠ wafers this quarter.
top-replies SK Hynix · Intel
Discussion 07:12 ET Updated
Does "pacing" mean more alignment compute and lower lab margins — not less spend?
Gavin Baker argues Anthropic/OpenAI "pacing" is more compute on alignment, monitoring, and evals at the cost of slightly lower margins — not a capex cut. Roon frames pacing as asymmetric margin compression; xEBITDA argues safety spend can raise total semiconductor demand.
top-replies Anthropic · OpenAI
Tue, 15 Sept 2026
Discussion 15:30 ET
Did Anthropic's pacing call already cost frontier enterprise spend share?
Ramp's Ara Kharazian says Astra takes 13% of enterprise AI spend vs Fable at 8%; he frames Anthropic's pace-the-frontier call as already losing adoption. Replies split on whether Ramp is the right sample.
top-replies OpenAI · Anthropic · Ramp
Discussion 07:30 ET
If GPT-6 Astra is looped depth, does that bend HBM vs FLOP intensity?
SemiAnalysis posts that GPT-6 Astra is "basically confirmed" to use loop transformers — deeper passes over layers without growing parameter count. Visible replies split on whether that is HBM-sparing or sequential-compute-heavy. Video body unread; root text truncates mid-quote.
top-replies OpenAI · SemiAnalysis
Discussion 05:00 ET
Is Samsung Taylor's early ramp a real AI foundry recovery — or pilot optics?
jukan05 relays Korean media that Taylor utilization hit ~30% (from ~20% last month), pulled forward vs a November plan on Tesla AI5 demand, with some expecting full utilization by year-end. Aju Press same week frames pilot lines and mass production early next year — a utilization-vs-pilot split.
top-replies Samsung · Tesla · TSMC
Discussion 04:02 ET
Does Samsung dual-sourcing HBM base dies with TSMC ease the HBM bind — or just meet customers?
jukan05 and ZDNet Korea report Samsung will use both Samsung Foundry and TSMC for custom HBM base dies by customer request, with Memory Division owning design when TSMC fabs the die. Replies frame it as avoiding turnkey isolation; wafer-capex allocation to TSMC remains the open bind.
top-replies Samsung · TSMC · NVIDIA
Discussion 03:54 ET
Do Broadcom's FY27/FY28 AI revenue targets rebut Monday's "cycle end" tape?
After Monday's chip de-rating, JP Insights weighs Hock Tan's reaffirmed ~$115B FY27 / ~$230B FY28 AI semiconductor outlook — and Anthropic as largest custom silicon customer in 2027 — against weekly end-of-cycle takes. Visible replies favor operator numbers; demand vs delivery remains open.
top-replies Broadcom · Anthropic · OpenAI
Mon, 14 Sept 2026
Discussion 18:48 ET
Can Anthropic finance a 1/5/10 GW Broadcom TPU path — and top Google as XPU customer?
Tanay Jaipuria relays Anthropic's Broadcom TPU roadmap — 1 GW Ironwood in 2026, 5 GW TPU v8i in 2027, line of sight to 10 more GW in 2028 — with Anthropic as Broadcom's largest XPU customer next year. Thin replies challenge cash generation and call the Google overtake "wild."
top-replies Anthropic · Broadcom · Google
Discussion 18:33 ET
Is a ~90% hike probability the right Wednesday call — or is the room still 50/50?
Eric Balchunas flags a 90% vs 38% juxtaposition on rate-hike odds into the Fed meeting. Replies split between trusting economist/market hike pricing after a hot inflation print and a personal 50/50 that Wednesday can still go either way.
top-replies Federal Reserve · CME
Discussion 17:12 ET
Why do three US labs take ~70% of OpenRouter spend but only ~27% of tokens?
OpenRouter's Peter Walker shows three major American labs attracting about 70% of spend while accounting for only 27% of tokens. Thin reply chain; the chart is the claim. Google's weight depends on whether the metric is spend or tokens.
top-replies OpenRouter · OpenAI · Anthropic
Discussion 08:36 ET
Is datacenter HBM stuck above 4-Hi, or is shorter stack the inference optimum?
jukan05 cites TrendForce that 4-Hi HBM is not enough and doubts suppliers would make it anyway; Jeff Pu concurs that datacenter de-spec to 4-Hi is unlikely. SemiAnalysis the same day argues 4-Hi maximizes tokens per HBM wafer for inference — a live supply-vs-architecture split.
top-replies NVIDIA · TrendForce · SemiAnalysis
‹ All posts Discussions /Discussion ARCHIVE
Discussion · Tue, 15 Sept 2026 · 15:30 ET

Did Anthropic's pacing call already cost frontier enterprise spend share?

Ramp's Ara Kharazian says Astra takes 13% of enterprise AI spend vs Fable at 8%; he frames Anthropic's pace-the-frontier call as already losing adoption. Replies split on whether Ramp is the right sample.

top-replies Ara Kharaziantae kimPeter Walker OpenAIAnthropicRampOpenRouter Source ↗
Review after: 2026-10-15

Opening

Ramp's spend split favors reading Anthropic's pacing call as already expensive for frontier share — with an open sample caveat. Ara Kharazian (Ramp) posts Astra at 13% of enterprise AI spend vs Fable at 8% this week, and argues Anthropic's frontier model has already fallen behind on adoption after the pace call; OpenAI growth is shifts from Sol and some Anthropic models plus net-new. Visible replies push back that Ramp may miss Global-2000 / regulated pipelines and that OpenRouter shows a parallel OpenAI spend lead. Tail unread beyond top replies (top-replies).

Sides

FOR — pacing already shows up as lost frontier spend

  • @arakharazian [named] [principal] — Astra 13% vs Fable 8% on Ramp; pace call was a "big risk"; OpenAI gaining from Sol + some Anthropic shifts and net-new
  • @firstadopter [named] — relays the Ramp framing that Anthropic's frontier model has already fallen behind on adoption
  • @trmcdonald [anon] — five-point spend gap turns pacing from principle into strategy

AGAINST — wrong sample / wrong buyer

  • @bmlascolea [named] — ETR-linked claim Anthropic still gains among Global 2000 even as it slips among smaller orgs
  • @reachmaheshkmb [anon] — pacing is aimed at 12-month procurement / audit buyers; this week's spend share is the developer market
  • @patrickdonohoe [anon] — Anthropic still dominates highest-margin Fable spend; OpenAI owns the rest of the Pareto; compute rationing may explain Ant prioritization

THIRD — same spend signal, different venue

  • @PeterJ_Walker [named] [principal] — OpenRouter: users spent more on OpenAI than Anthropic last week for first time in >2.5 years; Astra top by spend, Luna top by tokens
  • @victoria_neiman [anon] — Ramp = enterprise, OpenRouter = developer; Claude Code still #2 app by volume while routes go elsewhere

Room: Ramp principal on card-spend data; OpenRouter principal on router spend; Global-2000 counterclaim is secondhand via reply quote.

Quotes

OpenAI is winning enterprise spend at the frontier. As of this week, Astra takes 13% of enterprise AI spend vs. Fable (8%) per Ramp data. — Ara Kharazian (@arakharazian)

Anthropic took a big risk in its recent call to pace the frontier. It's frontier model has already fallen behind on adoption. — Ara Kharazian (@arakharazian)

OpenAI's growth is primarily coming from shifts from Sol and some Anthropic models, net-new usage too. — Ara Kharazian (@arakharazian)

OpenRouter users spent more on OpenAI models than on Anthropic models last week. This hasn't happened for more than 2.5 years. — tae kim (@firstadopter)

Astra = top model by spend last week / Luna = top model by tokens (by a LOT) — Peter Walker (@PeterJ_Walker)

Interesting data - we're finding that Anthropic continues to gain among the largest organizations (Global 2000 shown below). — Brad LaScolea (@bmlascolea)

What would settle it

Ramp AI Index (or comparable card-spend panel) publishes a multi-week series of frontier model spend shares with Anthropic vs OpenAI broken out Knowable — next Ramp public drops — through 2026-10-15

OpenRouter weekly spend share OpenAI vs Anthropic stays inverted for ≥4 consecutive weeks (or reverts) Knowable — OpenRouter insights — through 2026-10-15

ETR / Global-2000 (or similar enterprise survey) prints Anthropic vs OpenAI share among largest orgs for the same window as Ramp Knowable — next ETR / survey release — through 2026-10-31

Connects to

Pacing as mix, not cliff — Baker's mechanism that pacing reallocates compute rather than kills spend (discussion).

Spend vs tokens — OpenRouter's prior ~70% spend / ~27% tokens split at three US labs (discussion).

Astra as product — overnight loop-transformer claim on GPT-6 Astra (discussion).

The Open/Close  ·  Research commentary, not investment advice. Positions may be held in securities mentioned.

Prior documents3 documents

Gavin Baker argues Anthropic/OpenAI "pacing" is more compute on alignment, monitoring, and evals at the cost of slightly lower margins — not a capex cut. Roon frames pacing as asymmetric margin compression; xEBITDA argues safety spend can raise total semiconductor demand.

Wed, 16 Sept 2026 · 07:12 ET

OpenRouter's Peter Walker shows three major American labs attracting about 70% of spend while accounting for only 27% of tokens. Thin reply chain; the chart is the claim. Google's weight depends on whether the metric is spend or tokens.

Mon, 14 Sept 2026 · 17:12 ET

SemiAnalysis posts that GPT-6 Astra is "basically confirmed" to use loop transformers — deeper passes over layers without growing parameter count. Visible replies split on whether that is HBM-sparing or sequential-compute-heavy. Video body unread; root text truncates mid-quote.

Tue, 15 Sept 2026 · 07:30 ET