Discussions 17 items
Wed, 23 Sept 2026
Discussion 07:28 ET
Is 16Hi HBM really pushed to 2029 — and does that extend 12Hi tight?
jukan05: JPM HBM model now sees 16Hi adoption 2029 earliest. Replies split on earlier slip vs prolonged 8/12Hi scarcity; no JPM primary in-thread.
top-replies JPMorgan · SK Hynix · Samsung
Tue, 22 Sept 2026
Discussion 07:26 ET
Is EMIB taking real CoWoS share — or just relocating the HBM queue?
SemiAnalysis ChipBook: Korea→Malaysia HBM +467% YoY, ~$3B/2mo via Penang. Replies split: customers fleeing CoWoS scarcity vs queue transfer / qualify risk.
top-replies Intel · TSMC
Thu, 17 Sept 2026
Discussion 15:40 ET
Are AI-infra vendor leads durable when every roadmap hits FFWD?
Nick Dorsey from AI Infra Summit: vendor roadmaps overlap, Astera/CXL example — today's technical leads feel wobbly under industry-wide acceleration.
top-replies Astera Labs
Discussion 15:40 ET Updated
Is AMD MI355X closing agentic-inference TCO vs GB300 — or only on some stacks?
SemiAnalysis says MI355X is quickly closing perf/TCO vs GB300 in agentic inference via SGLang/MoRI/UMBP; replies split on InferenceX $/M-token snapshots and software vs silicon.
top-replies AMD · NVIDIA · SemiAnalysis
Discussion 07:28 ET
Does CPO kill M8/M9 CCL upgrades — or is that scale-out/scale-up confusion?
Market rumor: CPO cuts ultralow-loss CCL demand. Jeff Pu says upgrade stays — CPO not on compute board soon; NVSwitch still needs high-spec PTFE CCL.
top-replies NVIDIA
Discussion 07:25 ET
Did OpenAI flip OpenRouter share vs Anthropic from 20/80 to 50/50?
Baker cites OpenRouter: OpenAI 20%→50% vs Anthropic since June. Talia notes token vs revenue share diverging; Baker agrees OpenAI is discounting.
top-replies OpenAI · Anthropic · OpenRouter
Wed, 16 Sept 2026
Discussion 07:22 ET
Is the Hynix–Intel story really a hyperscaler memory JV — or just talks?
jukan05 says the load-bearing angle is a possible Hynix–Intel JV with hyperscalers hungry for memory; replies stress prior denials, cycle cost, and that talks ≠ wafers this quarter.
top-replies SK Hynix · Intel
Discussion 07:12 ET Updated
Does "pacing" mean more alignment compute and lower lab margins — not less spend?
Gavin Baker argues Anthropic/OpenAI "pacing" is more compute on alignment, monitoring, and evals at the cost of slightly lower margins — not a capex cut. Roon frames pacing as asymmetric margin compression; xEBITDA argues safety spend can raise total semiconductor demand.
top-replies Anthropic · OpenAI
Tue, 15 Sept 2026
Discussion 15:30 ET
Did Anthropic's pacing call already cost frontier enterprise spend share?
Ramp's Ara Kharazian says Astra takes 13% of enterprise AI spend vs Fable at 8%; he frames Anthropic's pace-the-frontier call as already losing adoption. Replies split on whether Ramp is the right sample.
top-replies OpenAI · Anthropic · Ramp
Discussion 07:30 ET
If GPT-6 Astra is looped depth, does that bend HBM vs FLOP intensity?
SemiAnalysis posts that GPT-6 Astra is "basically confirmed" to use loop transformers — deeper passes over layers without growing parameter count. Visible replies split on whether that is HBM-sparing or sequential-compute-heavy. Video body unread; root text truncates mid-quote.
top-replies OpenAI · SemiAnalysis
Discussion 05:00 ET
Is Samsung Taylor's early ramp a real AI foundry recovery — or pilot optics?
jukan05 relays Korean media that Taylor utilization hit ~30% (from ~20% last month), pulled forward vs a November plan on Tesla AI5 demand, with some expecting full utilization by year-end. Aju Press same week frames pilot lines and mass production early next year — a utilization-vs-pilot split.
top-replies Samsung · Tesla · TSMC
Discussion 04:02 ET
Does Samsung dual-sourcing HBM base dies with TSMC ease the HBM bind — or just meet customers?
jukan05 and ZDNet Korea report Samsung will use both Samsung Foundry and TSMC for custom HBM base dies by customer request, with Memory Division owning design when TSMC fabs the die. Replies frame it as avoiding turnkey isolation; wafer-capex allocation to TSMC remains the open bind.
top-replies Samsung · TSMC · NVIDIA
Discussion 03:54 ET
Do Broadcom's FY27/FY28 AI revenue targets rebut Monday's "cycle end" tape?
After Monday's chip de-rating, JP Insights weighs Hock Tan's reaffirmed ~$115B FY27 / ~$230B FY28 AI semiconductor outlook — and Anthropic as largest custom silicon customer in 2027 — against weekly end-of-cycle takes. Visible replies favor operator numbers; demand vs delivery remains open.
top-replies Broadcom · Anthropic · OpenAI
Mon, 14 Sept 2026
Discussion 18:48 ET
Can Anthropic finance a 1/5/10 GW Broadcom TPU path — and top Google as XPU customer?
Tanay Jaipuria relays Anthropic's Broadcom TPU roadmap — 1 GW Ironwood in 2026, 5 GW TPU v8i in 2027, line of sight to 10 more GW in 2028 — with Anthropic as Broadcom's largest XPU customer next year. Thin replies challenge cash generation and call the Google overtake "wild."
top-replies Anthropic · Broadcom · Google
Discussion 18:33 ET
Is a ~90% hike probability the right Wednesday call — or is the room still 50/50?
Eric Balchunas flags a 90% vs 38% juxtaposition on rate-hike odds into the Fed meeting. Replies split between trusting economist/market hike pricing after a hot inflation print and a personal 50/50 that Wednesday can still go either way.
top-replies Federal Reserve · CME
Discussion 17:12 ET
Why do three US labs take ~70% of OpenRouter spend but only ~27% of tokens?
OpenRouter's Peter Walker shows three major American labs attracting about 70% of spend while accounting for only 27% of tokens. Thin reply chain; the chart is the claim. Google's weight depends on whether the metric is spend or tokens.
top-replies OpenRouter · OpenAI · Anthropic
Discussion 08:36 ET
Is datacenter HBM stuck above 4-Hi, or is shorter stack the inference optimum?
jukan05 cites TrendForce that 4-Hi HBM is not enough and doubts suppliers would make it anyway; Jeff Pu concurs that datacenter de-spec to 4-Hi is unlikely. SemiAnalysis the same day argues 4-Hi maximizes tokens per HBM wafer for inference — a live supply-vs-architecture split.
top-replies NVIDIA · TrendForce · SemiAnalysis
‹ All posts Discussions /Discussion ARCHIVE
Discussion · Mon, 14 Sept 2026 · 08:36 ET

Is datacenter HBM stuck above 4-Hi, or is shorter stack the inference optimum?

jukan05 cites TrendForce that 4-Hi HBM is not enough and doubts suppliers would make it anyway; Jeff Pu concurs that datacenter de-spec to 4-Hi is unlikely. SemiAnalysis the same day argues 4-Hi maximizes tokens per HBM wafer for inference — a live supply-vs-architecture split.

top-replies JukanJeff PuJake AyesTiago Marques NVIDIATrendForceSemiAnalysis Source ↗
Review after: 2026-12-10

The read

On shipping reality for datacenter AI near-term, Jeff Pu / TrendForce / jukan have the better of it: the visible supply-chain claim is that mainstream production is 8-Hi and 12-Hi, per-stack capacity at ~16GB for 4-Hi forces more chips and networking, and base-die logic stays tight. On inference economics, SemiAnalysis's same-window thesis that 4-Hi can win on $/bandwidth and tokens-per-wafer is a serious opposing frame — but it is a design-target argument, not yet a confirmed SKU mix. Do not collapse "optimal for decode TCO" into "will ship as the datacenter standard tomorrow."

State of play

jukan05 posts that TrendForce explains why 4-Hi HBM is not enough, adding that memory makers would not produce 4-Hi even if it were. Jeff Pu's note (RTed into the feed) shares TrendForce's low odds of a datacenter de-spec to 4-Hi and flags capacity, base-die, and flow-cost reasons. Replies stress incentives and order-taking: server DRAM profitability as an alternate pull; suppliers cannot satisfy every firm. Same calendar day, SemiAnalysis Weekly argues Rubin Ultra ships less HBM per package and that 4-Hi can be the right inference call. Stakes: whether HBM remains the binding constraint via tall stacks, or wafer economics push the industry shorter.

Fidelity note: top-replies — conversation search rate-limited; named replies and Jeff Pu note captured; deep tail unread.

The positions

Datacenter HBM will not de-spec to 4-Hi. @jukan05 [pseudo] — TrendForce: '4-Hi HBM isn’t enough'; 'wouldn’t produce 4-Hi HBM even if it were enough' @sssjeffpu [named] — 'chance of de-spec to 4-Hi is low'; '~16GB' too small; 'mainstream production is focused on 8-Hi & 12-Hi'

Shorter stacks are the inference TCO / wafer optimum. SemiAnalysis (Myron Xie / Jordan Nanos) [named] — show notes: Rubin Ultra ships 192GB vs 1TB preview; 'reason is supply and not performance'; 'shipping less memory per chip might be the right call' (see longform 2026-09-14-semianalysis-4hi-hbm)

Suppliers are incentive- and order-bound, not 'enough'-bound. @jake_ayes [pseudo] — 'these guys are order takers… if everyone wants it you’re kinda SOL' @evrsr [pseudo] — 'server DRAM been more profitable than HBM' as the incentive to watch

Weight of the room

Jeff Pu is the only named sell-side memory voice in the captured posts with a structured note. jukan05 is a high-reach semis aggregator citing TrendForce. SemiAnalysis is principal research on the opposing economics claim. Anonymous/low-follower incentive color is useful framing, not proof. Engagement not cited as evidence.

What would settle it

NVIDIA / major ASIC customer SKU disclosures for Rubin Ultra and peer parts: HBM stack height and GB per package in shipping configs Knowable — product briefs / earnings / SemiAnalysis or vendor confirmations — through Q4 2026

Memory-supplier commentary (Samsung, SK Hynix, Micron) on 4-Hi vs 8-Hi/12-Hi mix for datacenter HBM4/HBM4E Knowable — next earnings cycles — by ~2026-12-10

Evidence that a hyperscaler ASIC program designed around 4-Hi HBM4+ is in volume qualification Knowable — supply-chain reporting — 1H 2027

Posts

Jukan (@jukan05) — Sep 14, 2026, 8:36am ET
TrendForce explains why 4-Hi HBM isn’t enough. (Personally, I think memory makers wouldn’t produce 4-Hi HBM even if it were enough.)
157 likes · 9 RTs · 14 replies · 44 bookmarks
https://x.com/jukan05/status/2099477271182717429

Jeff Pu (@sssjeffpu) — Sep 14, 2026, 1:55am ET
Memory: Datacenter HBM unlikely to move to 4-Hi

🔑 We share TrendForce’s view that the chance of de-spec to 4-Hi is low, vs worries over 4-Hi which could deliver same bandwidth with fewer dies, cut inference costs & stretch limited DRAM supply • Two key reasons 4-Hi unlikely fly in datacenter AI: – Per-stack capacity is too small (~16GB for HBM5/HBM4E) → hyperscalers would need more chips, racks & scale-up networking – Higher volume still increases base-die demand on tight advanced logic nodes (esp. N2 for HBM5) → little to no supply relief • Supply-chain feasibility low: mainstream production is focused on 8-Hi & 12-Hi, vs development cost and low-volume 4-Hi flows • Cost per Gb is higher on 4-Hi vs 8-Hi, with similar total cube cost based on current price quote. • Nvidia’s Rubin Ultra dropping to 8-Hi is well expected (our note on Aug 6) • For Rubin Ultra we expect parallel development of HBM4 8-Hi 192GB / HBM4E 8-Hi 256GB / HBM4E 12-Hi 384GB, with higher chance of HBM4 8-Hi first, then HBM4E 8-Hi as the mainstream SKU
149 likes · 23 RTs · 0 replies · 69 bookmarks
https://x.com/sssjeffpu/status/2099376403762585752

Jake Ayes (@jake_ayes) — Sep 14, 2026, 11:30am ET
@jukan05 At the end of the day these guys are order takers… they can only sell to so many firms and if everyone wants it you’re kinda SOL
0 likes · 0 RTs · 0 replies
https://x.com/jake_ayes/status/2099521081455346123

Tiago Marques (@evrsr) — Sep 14, 2026, 5:50pm ET
@jukan05 Hasn't server DRAM been more profitable than HBM for a good while now? If that is the case, that would be the incentive you are looking for.
0 likes · 0 RTs · 0 replies
https://x.com/evrsr/status/2099616791043461370

The Open/Close  ·  Research commentary, not investment advice. Positions may be held in securities mentioned.