bitdepth.co
LIVE
HBM3E demand +47% q/qNVDA inventory lead time −9dSK Hynix HBM4 risk prod 2026Q3TSMC CoWoS capex +$3.8B
Back to home

Articles tagged with "ai"

BitDepth Editorial

Samsung's HBM3E qualification, finally. The full timeline and what it means for NVIDIA's 2026 allocation strategy.

For 18 months Samsung was the third HBM vendor on paper only. NVIDIA's quiet May qualification changes everything — including SK Hynix's pricing power and Micron's growth ceiling.

BitDepth Editorial
BitDepth Editorial

Why every AI training run is now a packaging negotiation.

The bottleneck no longer measured in GPUs but in CoWoS slots. How TSMC quietly became the gatekeeper of the entire frontier-AI roadmap — and what the 2027 packaging map looks like.

BitDepth Editorial
BitDepth Editorial

Cerebras WSE-4 is generally available. We ran the benchmarks. The numbers are real.

Wafer-scale inference has spent four years as a press release. With WSE-4 in our hands, the latency-vs-cost economics finally pencil out for a specific, narrow class of workloads.

BitDepth Editorial
BitDepth Editorial

Why Samsung's HBM gap is closing.

Samsung's HBM3E yield rates climbed faster than expected in early 2026. We look at what changed in Samsung's production process, the NVIDIA qualification milestone, and whether the third-vendor story is finally real.

BitDepth Editorial
BitDepth Editorial

Micron, year of the ramp. The "American HBM" trade is now structural, not speculative.

Micron's HBM3E sold out through 2027 and the HBM4 sample timeline is closing the gap with SK Hynix. What this means for supply diversification and 2026 pricing dynamics.

BitDepth Editorial
BitDepth Editorial

Paged attention, two years later.

vLLM's paged attention paper changed inference economics overnight. Two years of production data later, here's what actually happened to the KV-cache problem — and what's still unsolved.

BitDepth Editorial
BitDepth Editorial

The FP8 → INT4 quantization roadmap.

Inference vendors are racing from FP8 to INT4 as the next lever on compute efficiency. We map the roadmap across frameworks, the accuracy tradeoffs that actually matter in production, and which memory vendors benefit most.

BitDepth Editorial
BitDepth Editorial

HBM, explained without metaphors. Stacks, TSVs, and why it's so hard to make.

High-Bandwidth Memory isn't "fast DRAM." It's a wholly different manufacturing problem. We go through the stack, the through-silicon vias, and the packaging step that ties it all to the GPU.

BitDepth Editorial
BitDepth Editorial

The HBM4 spec war: how a memory standard became an NVIDIA-customer fight.

JEDEC ratified the standard. NVIDIA's customer requirements then re-opened it. We walk through the spec deltas, the supply implications, and which vendor benefits most.

BitDepth Editorial

No more articles to load

We use cookies

We use cookies to ensure you get the best experience on our website. For more information on how we use cookies, please see our cookie policy.