home.social

#cxl — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #cxl, aggregated by home.social.

fetched live
  1. Wird der Speicher #controller in Zeiten von #CXL und #InMemoryCompute der neue zentrale Baustein für Rechner Architekturen,und löst damit die #CPU als core-compute-dsp ab? Und könnte #ESMC in #Europa eine Rolle dabei übernehmen?

  2. XCENA MX1 CXL Computational Memory Device at Hot Chips 2026 with Samsung

    Hot Chips 2026 XCENA MX1 CXL Slide 2 (Memory Xcelerator) XCENA is presenting its MX1 CXL computational memory…
    #EuropeSays #Korea #KR #Samsung #CXL #SamsungGroup #XCENA
    europesays.com/korea/132151/

  3. At Hot Chips 2026, XCENA showed off how the MX1 combines a CXL memory controller with 3072 RISC-V cores and tiering to SSDs, and Samsung showed how it scales#CXL #samsung #XCENA
    XCENA MX1 CXL Computational Memory Device at Hot Chips 2026 with Samsung
  4. Петабайт на двух дисках и технология CXL — индустрия хранения подстраивается под аппетиты ИИ

    4–6 августа в Санта-Кларе прошла выставка Flash Memory Summit, на которой был представлен ряд интересных новинок. Контекст 2026 таков, что

    habr.com/ru/companies/selectel

    #selectel #ssd #накопители #cxl #fms #pcie #nvme #kioxia #инфраструктура_ии #synopsys

  5. Петабайт на двух дисках и технология CXL — индустрия хранения подстраивается под аппетиты ИИ

    4–6 августа в Санта-Кларе прошла выставка Flash Memory Summit, на которой был представлен ряд интересных новинок. Контекст 2026 таков, что

    habr.com/ru/companies/selectel

    #selectel #ssd #накопители #cxl #fms #pcie #nvme #kioxia #инфраструктура_ии #synopsys

  6. Петабайт на двух дисках и технология CXL — индустрия хранения подстраивается под аппетиты ИИ

    4–6 августа в Санта-Кларе прошла выставка Flash Memory Summit, на которой был представлен ряд интересных новинок. Контекст 2026 таков, что

    habr.com/ru/companies/selectel

    #selectel #ssd #накопители #cxl #fms #pcie #nvme #kioxia #инфраструктура_ии #synopsys

  7. AI Data Centers Overheating: Semiconductor, Cooling, and Power Industries Battle Thermal Runaway — BigGo Finance

    As the artificial intelligence (AI) data center market undergoes explosive growth, “thermal management” has emerged as a core…
    #EuropeSays #Korea #KR #SamsungElectronics #AIdatacenter #CXL #FMS2026 #HBM #HDHyundaiOilbank #HyosungHeavyIndustries #immersioncooling #LGElectronics #LSElectric #NVIDIA #S-Oil #Samsung #SKEnmove #SKhynix
    europesays.com/korea/112576/

  8. 靠快閃記憶體紓緩 DRAM 不足 KIOXIA 推 XL1 記憶體擴充模組
    日本記憶體大廠 KIOXIA 於 8 月 3 日宣布推出全新記憶體擴充模組「KIOXIA XL1」系列,將原本 […]
    #人工智能 #AI運算 #CXL #DRAM
    unwire.hk/2026/08/05/kioxia-xl

  9. 靠快閃記憶體紓緩 DRAM 不足 KIOXIA 推 XL1 記憶體擴充模組
    日本記憶體大廠 KIOXIA 於 8 月 3 日宣布推出全新記憶體擴充模組「KIOXIA XL1」系列,將原本 […]
    #人工智能 #AI運算 #CXL #DRAM
    unwire.hk/2026/08/05/kioxia-xl

  10. 靠快閃記憶體紓緩 DRAM 不足 KIOXIA 推 XL1 記憶體擴充模組
    日本記憶體大廠 KIOXIA 於 8 月 3 日宣布推出全新記憶體擴充模組「KIOXIA XL1」系列,將原本 […]
    #人工智能 #AI運算 #CXL #DRAM
    unwire.hk/2026/08/05/kioxia-xl

  11. Samsung, SK Push PIM and CXL as US, China, Japan Challenge HBM

    Samsung Electronics’ CXL Memory Module (CMM-D) product. Photo courtesy of Samsung Electronics Samsung Electronics (005930.KS) and SK hynix…
    #EuropeSays #Korea #KR #SamsungElectronics #CXL #HBM #LPDDR5X #Micron #next-generationmemory #PIM #Processing-in-Memory #Samsung #SKhynix
    europesays.com/korea/109816/

  12. #CXL で低速メモリに繋ぐなら、フラッシュメモリより羅生門したメモリを積める方が嬉しいのでは?

    キオクシア、メモリ不足をフラッシュメモリで補う拡張モジュール - PC Watch pc.watch.impress.co.jp/docs/ne?

  13. #CXL で低速メモリに繋ぐなら、フラッシュメモリより羅生門したメモリを積める方が嬉しいのでは?

    キオクシア、メモリ不足をフラッシュメモリで補う拡張モジュール - PC Watch pc.watch.impress.co.jp/docs/ne?

  14. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  15. Most CXL coverage collapses three separate things into one, which is why the technology reads as both imminent and permanently delayed.

    Expansion gives a single host extra CXL-attached memory as a second tier. It is in production at hyperscale — 768 GB of local DDR5-6400 alongside 256 GB of reclaimed DDR4-2400 per node, at 0.13 the cost per gigabyte of local DRAM, because those DIMMs were already bought.

    Pooling gives several hosts exclusive slices of a shared device, and does not require a switch: a multi-port device pools across whatever hosts plug into it. NSDI '26 measurements put multi-port paths at 260–300 ns against 490–600 ns switched.

    Sharing — many hosts on one region, coherent in hardware — is defined in CXL 3.0 and absent from shipping devices.

    The economics are the open question. Google researchers put switched break-even around 24 nodes. The same NSDI work finds sparse multi-port topologies save 3% to 5.4%. Topology decides.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Disaggregation #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure

    buysellram.com/blog/cxl-memory

  16. Most CXL coverage collapses three separate things into one, which is why the technology reads as both imminent and permanently delayed.

    Expansion gives a single host extra CXL-attached memory as a second tier. It is in production at hyperscale — 768 GB of local DDR5-6400 alongside 256 GB of reclaimed DDR4-2400 per node, at 0.13 the cost per gigabyte of local DRAM, because those DIMMs were already bought.

    Pooling gives several hosts exclusive slices of a shared device, and does not require a switch: a multi-port device pools across whatever hosts plug into it. NSDI '26 measurements put multi-port paths at 260–300 ns against 490–600 ns switched.

    Sharing — many hosts on one region, coherent in hardware — is defined in CXL 3.0 and absent from shipping devices.

    The economics are the open question. Google researchers put switched break-even around 24 nodes. The same NSDI work finds sparse multi-port topologies save 3% to 5.4%. Topology decides.

    buysellram.com/blog/cxl-memory

  17. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  18. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  19. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  20. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  21. Meta measures 43.7% of its servers as memory-capacity bound. They run out of RAM before they run out of cores, network, or storage.

    CXL is the standard meant to fix that, and it is further along than most coverage suggests in one direction and further behind in another. Expansion runs in production today. Pooling works on real hardware.

    #CXL #ComputeExpressLink #MemoryPooling #ServerMemory #DataCenter #Interconnect #DRAM #DDR4 #DDR5 #AIInfrastructure #tech

    buysellram.com/blog/cxl-memory

  22. Samsung, SK Hynix advance new memory layer to tackle AI scaling limits

    Samsung Electronics and SK hynix pushed to commercialize a new class of memory as artificial intelligence outgrew high-bandwidth…
    #EuropeSays #Korea #KR #SKHynix #AImemory #ComputeExpressLink #CXL #DRAM #high-bandwidthmemory #memorycapacity #memorymanagementtools #Samsung #SamsungElectronics #SK #SKhynix
    europesays.com/korea/92218/

  23. Samsung, SK Hynix Race to Mass-Produce CXL 3.2 Memory This Year, Ushering in the ‘Next Memory’ Era — BigGo Finance

    Samsung Electronics (005930) and SK Hynix (000660) are transitioning to mass production this year to seize the initiative…
    #EuropeSays #Korea #KR #SKHynix #AImemory #CMM-D3.0 #CXL #CXL3.2 #GoldmanSachs #HBM #Micron #NVIDIA #SamsungElectronics #SK #SKhynix
    europesays.com/korea/91164/

  24. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling

  25. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling

  26. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors

  27. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors

  28. SK hynix's parent group says it may offer "Memory as a Service" — RAM capacity sold like a cloud subscription rather than a chip in a box. The idea is early, but the hardware underneath it is not.

    This explainer walks through the full stack: why memory gets stranded inside servers, how CXL adds the cache coherency that PCIe never had, what memory appliances and CXL 3.0 switches actually do, and why tiering keeps local DRAM first. It also draws a line many articles blur — memory expansion (which Meta already runs across millions of servers, reusing retired DDR4) versus rack-level pooling, which is only now reaching real silicon.

    The business angle matters too: subscriptions would smooth the boom-and-bust cycle memory makers have never escaped. And for everyone else, the takeaway is simpler — memory is becoming an asset with value independent of the server it shipped in.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling #DDR4 #ServerHardware #Semiconductors #AIHardware #ITAD #technology

  29. SK hynix's parent group says it may offer "Memory as a Service" — RAM capacity sold like a cloud subscription rather than a chip in a box. The idea is early, but the hardware underneath it is not.

    This explainer walks through the full stack: why memory gets stranded inside servers, how CXL adds the cache coherency that PCIe never had, what memory appliances and CXL 3.0 switches actually do, and why tiering keeps local DRAM first. It also draws a line many articles blur — memory expansion (which Meta already runs across millions of servers, reusing retired DDR4) versus rack-level pooling, which is only now reaching real silicon.

    The business angle matters too: subscriptions would smooth the boom-and-bust cycle memory makers have never escaped. And for everyone else, the takeaway is simpler — memory is becoming an asset with value independent of the server it shipped in.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling #DDR4 #ServerHardware #Semiconductors #AIHardware #ITAD #technology

  30. SK hynix's parent group says it may offer "Memory as a Service" — RAM capacity sold like a cloud subscription rather than a chip in a box. The idea is early, but the hardware underneath it is not.

    This explainer walks through the full stack: why memory gets stranded inside servers, how CXL adds the cache coherency that PCIe never had, what memory appliances and CXL 3.0 switches actually do, and why tiering keeps local DRAM first. It also draws a line many articles blur — memory expansion (which Meta already runs across millions of servers, reusing retired DDR4) versus rack-level pooling, which is only now reaching real silicon.

    The business angle matters too: subscriptions would smooth the boom-and-bust cycle memory makers have never escaped. And for everyone else, the takeaway is simpler — memory is becoming an asset with value independent of the server it shipped in.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling #DDR4 #ServerHardware #Semiconductors #AIHardware #ITAD #technology

  31. SK hynix's parent group says it may offer "Memory as a Service" — RAM capacity sold like a cloud subscription rather than a chip in a box. The idea is early, but the hardware underneath it is not.

    This explainer walks through the full stack: why memory gets stranded inside servers, how CXL adds the cache coherency that PCIe never had, what memory appliances and CXL 3.0 switches actually do, and why tiering keeps local DRAM first. It also draws a line many articles blur — memory expansion (which Meta already runs across millions of servers, reusing retired DDR4) versus rack-level pooling, which is only now reaching real silicon.

    The business angle matters too: subscriptions would smooth the boom-and-bust cycle memory makers have never escaped. And for everyone else, the takeaway is simpler — memory is becoming an asset with value independent of the server it shipped in.

    buysellram.com/blog/memory-as-
    #MemoryAsAService #CXL #DRAM #DataCenter #AIInfrastructure #SKhynix #MemoryPooling #DDR4 #ServerHardware #Semiconductors #AIHardware #ITAD #technology

  32. SK hynix's parent group says it may offer "Memory as a Service" — RAM capacity sold like a cloud subscription rather than a chip in a box. The idea is early, but the hardware underneath it is not.

    This explainer walks through the full stack: why memory gets stranded inside servers, how CXL adds the cache coherency that PCIe never had, what memory appliances and CXL 3.0 switches actually do, and why tiering keeps local DRAM first. It also draws a line many articles blur — memory expansion (which Meta already runs across millions of servers, reusing retired DDR4) versus rack-level pooling, which is only now reaching real silicon.

    The business angle matters too: subscriptions would smooth the boom-and-bust cycle memory makers have never escaped. And for everyone else, the takeaway is simpler — memory is becoming an asset with value independent of the server it shipped in.

    buysellram.com/blog/memory-as-

  33. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors #tech

  34. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors #tech

  35. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors #tech

  36. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors #tech

  37. RAM has always been trapped inside the server it shipped in — one machine's shortage next to another's idle gigabytes. SK hynix now says memory could be sold as a subscription instead. This explainer covers the technology that would carry it: CXL, memory appliances, rack-level switches, and tiering — plus Meta's production proof that pooled thinking already pays, right down to reused DDR4.

    buysellram.com/blog/memory-as-

    #MemoryAsAService #CXL #DataCenter #AIInfrastructure #DRAM #Semiconductors #tech

  38. Dentro la gerarchia della memoria delle GPU: come i server AI spostano i dati dagli SSD alla HBM

    Ti sei mai chiesto come i modelli di intelligenza artificiale trasferiscono i dati alla GPU? Questo articolo spiega in modo semplice i diversi livelli di memoria all'interno di un server AI, perché la memoria della GPU è diventata un collo di bottiglia e quali nuove tecnologie potrebbero migliorare le prestazioni dell'intelligenza artificiale.

    buysellram.com/blog/inside-the…

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD

  39. Dentro la gerarchia della memoria delle GPU: come i server AI spostano i dati dagli SSD alla HBM

    Ti sei mai chiesto come i modelli di intelligenza artificiale trasferiscono i dati alla GPU? Questo articolo spiega in modo semplice i diversi livelli di memoria all'interno di un server AI, perché la memoria della GPU è diventata un collo di bottiglia e quali nuove tecnologie potrebbero migliorare le prestazioni dell'intelligenza artificiale.

    buysellram.com/blog/inside-the…

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD

  40. Dentro la gerarchia della memoria delle GPU: come i server AI spostano i dati dagli SSD alla HBM

    Ti sei mai chiesto come i modelli di intelligenza artificiale trasferiscono i dati alla GPU? Questo articolo spiega in modo semplice i diversi livelli di memoria all'interno di un server AI, perché la memoria della GPU è diventata un collo di bottiglia e quali nuove tecnologie potrebbero migliorare le prestazioni dell'intelligenza artificiale.

    buysellram.com/blog/inside-the…

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD

  41. Dentro la gerarchia della memoria delle GPU: come i server AI spostano i dati dagli SSD alla HBM

    Ti sei mai chiesto come i modelli di intelligenza artificiale trasferiscono i dati alla GPU? Questo articolo spiega in modo semplice i diversi livelli di memoria all'interno di un server AI, perché la memoria della GPU è diventata un collo di bottiglia e quali nuove tecnologie potrebbero migliorare le prestazioni dell'intelligenza artificiale.

    buysellram.com/blog/inside-the…

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD

  42. Dentro la gerarchia della memoria delle GPU: come i server AI spostano i dati dagli SSD alla HBM

    Ti sei mai chiesto come i modelli di intelligenza artificiale trasferiscono i dati alla GPU? Questo articolo spiega in modo semplice i diversi livelli di memoria all'interno di un server AI, perché la memoria della GPU è diventata un collo di bottiglia e quali nuove tecnologie potrebbero migliorare le prestazioni dell'intelligenza artificiale.

    buysellram.com/blog/inside-the…

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD

  43. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #buysellram

  44. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #buysellram

  45. Why can't a GPU just carry more HBM? Interposers max out in size, stacking 12 to 16 DRAM dies compounds yield losses, and every stack sits beside a kilowatt-class package that hates sharing heat. So capacity climbs in careful steps — 80 GB on the H100, 141 on the H200, 192 on the B200, 288 on Blackwell Ultra — while KV caches for long-context inference balloon past 40 GB per request.

    That gap between what models demand and what packaging permits is reshaping server design. NVIDIA's Rubin platform treats CPU memory and HBM as one coherent pool. SanDisk and SK hynix are standardizing High Bandwidth Flash as a capacity tier under HBM, with first samples due this half. CXL 4.0 pooling hardware is landing in racks now.

    This article walks the whole memory hierarchy, from on-chip SRAM to NVMe, and explains what each emerging technology actually solves — and what it doesn't.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #technology

  46. Why can't a GPU just carry more HBM? Interposers max out in size, stacking 12 to 16 DRAM dies compounds yield losses, and every stack sits beside a kilowatt-class package that hates sharing heat. So capacity climbs in careful steps — 80 GB on the H100, 141 on the H200, 192 on the B200, 288 on Blackwell Ultra — while KV caches for long-context inference balloon past 40 GB per request.

    That gap between what models demand and what packaging permits is reshaping server design. NVIDIA's Rubin platform treats CPU memory and HBM as one coherent pool. SanDisk and SK hynix are standardizing High Bandwidth Flash as a capacity tier under HBM, with first samples due this half. CXL 4.0 pooling hardware is landing in racks now.

    This article walks the whole memory hierarchy, from on-chip SRAM to NVMe, and explains what each emerging technology actually solves — and what it doesn't.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #technology

  47. Why can't a GPU just carry more HBM? Interposers max out in size, stacking 12 to 16 DRAM dies compounds yield losses, and every stack sits beside a kilowatt-class package that hates sharing heat. So capacity climbs in careful steps — 80 GB on the H100, 141 on the H200, 192 on the B200, 288 on Blackwell Ultra — while KV caches for long-context inference balloon past 40 GB per request.

    That gap between what models demand and what packaging permits is reshaping server design. NVIDIA's Rubin platform treats CPU memory and HBM as one coherent pool. SanDisk and SK hynix are standardizing High Bandwidth Flash as a capacity tier under HBM, with first samples due this half. CXL 4.0 pooling hardware is landing in racks now.

    This article walks the whole memory hierarchy, from on-chip SRAM to NVMe, and explains what each emerging technology actually solves — and what it doesn't.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #technology

  48. Why can't a GPU just carry more HBM? Interposers max out in size, stacking 12 to 16 DRAM dies compounds yield losses, and every stack sits beside a kilowatt-class package that hates sharing heat. So capacity climbs in careful steps — 80 GB on the H100, 141 on the H200, 192 on the B200, 288 on Blackwell Ultra — while KV caches for long-context inference balloon past 40 GB per request.

    That gap between what models demand and what packaging permits is reshaping server design. NVIDIA's Rubin platform treats CPU memory and HBM as one coherent pool. SanDisk and SK hynix are standardizing High Bandwidth Flash as a capacity tier under HBM, with first samples due this half. CXL 4.0 pooling hardware is landing in racks now.

    This article walks the whole memory hierarchy, from on-chip SRAM to NVMe, and explains what each emerging technology actually solves — and what it doesn't.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #ITAD #technology

  49. Why can't a GPU just carry more HBM? Interposers max out in size, stacking 12 to 16 DRAM dies compounds yield losses, and every stack sits beside a kilowatt-class package that hates sharing heat. So capacity climbs in careful steps — 80 GB on the H100, 141 on the H200, 192 on the B200, 288 on Blackwell Ultra — while KV caches for long-context inference balloon past 40 GB per request.

    That gap between what models demand and what packaging permits is reshaping server design. NVIDIA's Rubin platform treats CPU memory and HBM as one coherent pool. SanDisk and SK hynix are standardizing High Bandwidth Flash as a capacity tier under HBM, with first samples due this half. CXL 4.0 pooling hardware is landing in racks now.

    This article walks the whole memory hierarchy, from on-chip SRAM to NVMe, and explains what each emerging technology actually solves — and what it doesn't.

    buysellram.com/blog/inside-the

  50. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #tech

  51. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #tech

  52. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #tech

  53. A modern AI server runs five layers of memory, from nanosecond on-chip SRAM to petabyte-scale SSDs, and keeping the GPU fed is the whole engineering game. This piece walks the full hierarchy: why HBM became the bottleneck, why adding more isn't simple, and where HBF, CXL, and PIM fit into the next generation.

    buysellram.com/blog/inside-the

    #HBM #HBM4 #GPUMemory #AIInfrastructure #DataCenter #CXL #HighBandwidthFlash #AIHardware #MemoryHierarchy #AIInference #NVMe #tech