home.social

#nvlink — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #nvlink, aggregated by home.social.

fetched live
  1. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (flexibel, bitte, damit sie von 2 bis 4 Steckplätzen funktioniert). Auf eBay kosten sie bereits über 1.000 US-Dollar – da kann man richtig Geld verdienen, Eric Hartford (@QuixiAI). Es gibt 2-Steckplatz-Modelle auf eBay, aber nur wenn man 2-Steckplätze (chinesische Modifikationen) oder wassergekühlte RTX 3090-Grafikkarten hat. Die 3-Steckplatz- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen unerschwinglich geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #HardwareModding #NVLink #RTX3090 #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  2. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (flexibel, bitte, damit sie von 2 bis 4 Steckplätzen funktioniert). Auf eBay kosten sie bereits über 1.000 US-Dollar – da kann man richtig Geld verdienen, Eric Hartford (@QuixiAI). Es gibt 2-Steckplatz-Modelle auf eBay, aber nur wenn man 2-Steckplätze (chinesische Modifikationen) oder wassergekühlte RTX 3090-Grafikkarten hat. Die 3-Steckplatz- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen unerschwinglich geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #HardwareModding #NVLink #RTX3090 #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  3. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (flexibel, bitte, damit sie von 2 bis 4 Steckplätzen funktioniert). Auf eBay kosten sie bereits über 1.000 US-Dollar – da kann man richtig Geld verdienen, Eric Hartford (@QuixiAI). Es gibt 2-Steckplatz-Modelle auf eBay, aber nur wenn man 2-Steckplätze (chinesische Modifikationen) oder wassergekühlte RTX 3090-Grafikkarten hat. Die 3-Steckplatz- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen unerschwinglich geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #HardwareModding #NVLink #RTX3090 #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  4. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (bitte flexibel, damit sie von 2 bis 4 Steckplätzen funktionieren). Auf eBay kosten sie bereits über 1.000 Dollar, da kann man richtig Geld verdienen. Eric Hartford (@QuixiAI) sagt, es gibt 2-Steckplatz-Modelle auf eBay, aber das gilt nur, wenn man 2-Steckplatz-Grafikkarten (chinesische Modifikationen) oder wassergekühlte RTX 3090er hat. Die 3- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen zu teuer geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware.) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #Hardware #NVLink #PCBuilding #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  5. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (bitte flexibel, damit sie von 2 bis 4 Steckplätzen funktionieren). Auf eBay kosten sie bereits über 1.000 Dollar, da kann man richtig Geld verdienen. Eric Hartford (@QuixiAI) sagt, es gibt 2-Steckplatz-Modelle auf eBay, aber das gilt nur, wenn man 2-Steckplatz-Grafikkarten (chinesische Modifikationen) oder wassergekühlte RTX 3090er hat. Die 3- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen zu teuer geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware.) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #Hardware #NVLink #PCBuilding #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  6. RT @QuixiAI: Wer ist Vibe-Hardware? Die Welt braucht NVLink-Brücken (bitte flexibel, damit sie von 2 bis 4 Steckplätzen funktionieren). Auf eBay kosten sie bereits über 1.000 Dollar, da kann man richtig Geld verdienen. Eric Hartford (@QuixiAI) sagt, es gibt 2-Steckplatz-Modelle auf eBay, aber das gilt nur, wenn man 2-Steckplatz-Grafikkarten (chinesische Modifikationen) oder wassergekühlte RTX 3090er hat. Die 3- und 4-Steckplatz-Modelle (für Retail-GPUs) scheinen inzwischen zu teuer geworden zu sein. (Vielleicht ist es Zeit für Vibe-Hardware.) — nitter.net/QuixiAI/status/2073

    mehr auf Arint.info

    #Grafikkarten #Hardware #NVLink #PCBuilding #TechNews #VibeHardware #arint_info

    https://x.com/QuixiAI/status/2073529939333144939#m

  7. Нейро сети для самых маленьких. Часть первая (которая после нулевой). Удобство в прокрустовом ложе оптимизации

    Это первая (после нулевой) статья из серии Нейро сети для самых маленьких , в которой мы разбираем инфраструктуру для запуска нейронных сетей. Для обучения и инференса нейросетей и для любых видов High Performance Computing используются специализированные технологии: GPU/TPU, RDMA, Kernel bypass, NVLink, InfiniBand, RoCE и другие. Про некоторые из них большинство только что-то слышали, но сталкиваться с ними не приходилось. Нельзя просто взять ванильный стек Linux, воткнуть в него 400Gb Ethernet+IP и получить рабочее решение. Почему? Потому что общее решение на масштабе в большинстве случаев проигрывает специализированным как в скорости, так и в стоимости. Как бы странно последнее ни звучало.

    habr.com/ru/companies/yandex/a

    #rdma #gpudirect_rdma #infiniband #tcp #ethernet #zero_copy #roce #nvlink #nvidia #gpu

  8. Нейро сети для самых маленьких. Часть первая (которая после нулевой). Удобство в прокрустовом ложе оптимизации

    Это первая (после нулевой) статья из серии Нейро сети для самых маленьких , в которой мы разбираем инфраструктуру для запуска нейронных сетей. Для обучения и инференса нейросетей и для любых видов High Performance Computing используются специализированные технологии: GPU/TPU, RDMA, Kernel bypass, NVLink, InfiniBand, RoCE и другие. Про некоторые из них большинство только что-то слышали, но сталкиваться с ними не приходилось. Нельзя просто взять ванильный стек Linux, воткнуть в него 400Gb Ethernet+IP и получить рабочее решение. Почему? Потому что общее решение на масштабе в большинстве случаев проигрывает специализированным как в скорости, так и в стоимости. Как бы странно последнее ни звучало.

    habr.com/ru/companies/yandex/a

    #rdma #gpudirect_rdma #infiniband #tcp #ethernet #zero_copy #roce #nvlink #nvidia #gpu

  9. Нейро сети для самых маленьких. Часть первая (которая после нулевой). Удобство в прокрустовом ложе оптимизации

    Это первая (после нулевой) статья из серии Нейро сети для самых маленьких , в которой мы разбираем инфраструктуру для запуска нейронных сетей. Для обучения и инференса нейросетей и для любых видов High Performance Computing используются специализированные технологии: GPU/TPU, RDMA, Kernel bypass, NVLink, InfiniBand, RoCE и другие. Про некоторые из них большинство только что-то слышали, но сталкиваться с ними не приходилось. Нельзя просто взять ванильный стек Linux, воткнуть в него 400Gb Ethernet+IP и получить рабочее решение. Почему? Потому что общее решение на масштабе в большинстве случаев проигрывает специализированным как в скорости, так и в стоимости. Как бы странно последнее ни звучало.

    habr.com/ru/companies/yandex/a

    #rdma #gpudirect_rdma #infiniband #tcp #ethernet #zero_copy #roce #nvlink #nvidia #gpu

  10. How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost As organizations move from AI pilots to production AI factories, infrastructure decisions have shifted from peak chip specificat...

    #AI #Infrastructure #Hardware #Networking #Software #CUDA #Dynamo #Inference #NVIDIA #Blackwell #NVLink

    Origin | Interest | Match
  11. NVIDIA’s new Vera Rubin platform brings together specialized chips (Vera CPUs, Rubin GPUs, Groq LPUs, and BlueField-4 DPUs) into coordinated, rack-scale systems designed for real-time AI.

    The big shift: AI isn’t just about training models anymore — it’s about orchestrating entire systems to power intelligent, autonomous agents in real time.
    buysellram.com/blog/the-agenti
    #NVIDIAGTC #AgenticAI #VeraRubin #DataCenter #GPU #InferenceFactory #AIInfrastructure #Groq #NVIDIA #NVLink #AIHardware #technology

  12. NVIDIA’s new Vera Rubin platform brings together specialized chips (Vera CPUs, Rubin GPUs, Groq LPUs, and BlueField-4 DPUs) into coordinated, rack-scale systems designed for real-time AI.

    The big shift: AI isn’t just about training models anymore — it’s about orchestrating entire systems to power intelligent, autonomous agents in real time.
    buysellram.com/blog/the-agenti
    #NVIDIAGTC #AgenticAI #VeraRubin #DataCenter #GPU #InferenceFactory #AIInfrastructure #Groq #NVIDIA #NVLink #AIHardware #technology

  13. NVIDIA’s new Vera Rubin platform brings together specialized chips (Vera CPUs, Rubin GPUs, Groq LPUs, and BlueField-4 DPUs) into coordinated, rack-scale systems designed for real-time AI.

    The big shift: AI isn’t just about training models anymore — it’s about orchestrating entire systems to power intelligent, autonomous agents in real time.
    buysellram.com/blog/the-agenti
    #NVIDIAGTC #AgenticAI #VeraRubin #DataCenter #GPU #InferenceFactory #AIInfrastructure #Groq #NVIDIA #NVLink #AIHardware #technology

  14. NVIDIA’s new Vera Rubin platform brings together specialized chips (Vera CPUs, Rubin GPUs, Groq LPUs, and BlueField-4 DPUs) into coordinated, rack-scale systems designed for real-time AI.

    The big shift: AI isn’t just about training models anymore — it’s about orchestrating entire systems to power intelligent, autonomous agents in real time.
    buysellram.com/blog/the-agenti
    #NVIDIAGTC #AgenticAI #VeraRubin #DataCenter #GPU #InferenceFactory #AIInfrastructure #Groq #NVIDIA #NVLink #AIHardware #technology

  15. NVIDIA’s new Vera Rubin platform brings together specialized chips (Vera CPUs, Rubin GPUs, Groq LPUs, and BlueField-4 DPUs) into coordinated, rack-scale systems designed for real-time AI.

    The big shift: AI isn’t just about training models anymore — it’s about orchestrating entire systems to power intelligent, autonomous agents in real time.
    buysellram.com/blog/the-agenti
    #NVIDIAGTC #AgenticAI #VeraRubin #DataCenter #GPU #InferenceFactory #AIInfrastructure #Groq #NVIDIA #NVLink #AIHardware #technology

  16. New SemiAnalysis InferenceX Data Shows NVIDIA Blackwell Ultra Delivers up to 50x Better Performance and 35x Lower Costs for Agentic AI

    The NVIDIA Blackwell platform has been widely adopted by leading inference providers such as Baseten, DeepInfra, Fireworks AI…
    #NewsBeep #News #Artificialintelligence #agenticai #AI #ArtificialIntelligence #AU #Australia #Dynamo #inference #NvidiaBlackwell #NVIDIARubin #NVLink #Technology #TensorRT #ThinkSMART
    newsbeep.com/au/486156/

  17. Một người dùng muốn dùng 2 card RTX 3090 kết nối qua 2 Oculink x4 (PCIe 4.0) để kích hoạt NVLink phục vụ AI/render. Họ hỏi: NVLink có hoạt động ổn? Bandwidth có đủ không? Đã có ai thử chưa? #GPU #NVLINK #AI #Hardware #ThiếtBịCôngNghệ #TríTuệNhânTạo

    reddit.com/r/LocalLLaMA/commen

  18. Fúzionál az NVLINK-kel a SiFive

    Az NVIDIA az előző évi Computexen jelentette be az NVLink Fusiont, amely lehetővé teszi a partnerek számára, hogy egyedileg…
    #Hungary #HU #Europe #Europa #EU #AI #fusion #hír #hungary #infrastruktura #licenc #Magyarország #Nvidia #nvlink #RISC-V #sifive #szerver #teszt
    europesays.com/2717211/

  19. NVIDIA V100 SXM2 gặp sự cố liên kết NVLink không hoạt động. Chủ đề này được thảo luận trên Reddit, nơi người dùng LeastExperience1579 đang tìm kiếm giải pháp sau khi mua máy chủ Supermicro từ nước ngoài. #nvidia #techsupport #NVLink #V100SXM2 #gpu #server #trợgiúptech

    reddit.com/r/LocalLLaMA/commen

  20. Cập nhật thử nghiệm mô hình MiniMax-M2 Q3_K_M với 4 GPU V100 32GB qua llama.cpp và NVLink. Khi dùng "--split-mode layer", tốc độ xử lý tăng từ 20 lên 38 tok/s so với "row", đạt 1683 tok/s khi khởi tạo. Tuy NVLink chưa tối ưu cho inference, nhưng combo V100 16GB SXM2 giá ~$100 + adapter ($50) vẫn đáng cân nhắc cho các dự án DIY. #AI #LLM #llamaCPP #NVLink #V100 #DOITech

    reddit.com/r/LocalLLaMA/commen

  21. Mixture of Experts Powers the Most Intelligent Frontier AI Models, Runs 10x Faster to Deliver 1/10 the Token Cost on NVIDIA Blackwell NVL72 The top 10 most intelligent open-source models all use a ...

    #Data #Center #Artificial #Intelligence #Dynamo #Inference #NVIDIA #Blackwell #NVLink #Open #Source

    Origin | Interest | Match
  22. Mixture of Experts Powers the Most Intelligent Frontier AI Models, Runs 10x Faster to Deliver 1/10 the Token Cost on NVIDIA Blackwell NVL72 The top 10 most intelligent open-source models all use a ...

    #Data #Center #Artificial #Intelligence #Dynamo #Inference #NVIDIA #Blackwell #NVLink #Open #Source

    Origin | Interest | Match
  23. Mixture of Experts Powers the Most Intelligent Frontier AI Models, Runs 10x Faster on NVIDIA Blackwell NVL72 The top 10 most intelligent open-source models all use a mixture-of-experts architecture...

    #Data #Center #Artificial #Intelligence #Dynamo #Inference #NVIDIA #Blackwell #NVLink #Open #Source

    Origin | Interest | Match
  24. #AWS announced #Trainium3, a new #AItrainingchip with significant performance and energy efficiency improvements. #Trainium4, already in development, will offer even better performance and support #Nvidia’s #NVLink Fusion technology, potentially attracting more AI applications to AWS. techcrunch.com/2025/12/02/amaz #tech #media #news

  25. #AWS announced #Trainium3, a new #AItrainingchip with significant performance and energy efficiency improvements. #Trainium4, already in development, will offer even better performance and support #Nvidia’s #NVLink Fusion technology, potentially attracting more AI applications to AWS. techcrunch.com/2025/12/02/amaz #tech #media #news

  26. #AWS announced #Trainium3, a new #AItrainingchip with significant performance and energy efficiency improvements. #Trainium4, already in development, will offer even better performance and support #Nvidia’s #NVLink Fusion technology, potentially attracting more AI applications to AWS. techcrunch.com/2025/12/02/amaz #tech #media #news

  27. #AWS announced #Trainium3, a new #AItrainingchip with significant performance and energy efficiency improvements. #Trainium4, already in development, will offer even better performance and support #Nvidia’s #NVLink Fusion technology, potentially attracting more AI applications to AWS. techcrunch.com/2025/12/02/amaz #tech #media #news

  28. #AWS announced #Trainium3, a new #AItrainingchip with significant performance and energy efficiency improvements. #Trainium4, already in development, will offer even better performance and support #Nvidia’s #NVLink Fusion technology, potentially attracting more AI applications to AWS. techcrunch.com/2025/12/02/amaz #tech #media #news

  29. #Arm and #Nvidia are #partnering to integrate Arm-based #Neoverse #CPUs with Nvidia’s #GPUs using Nvidia’s #NVLink Fusion technology. This #collaboration will benefit customers, particularly #hyperscalers, who prefer custom infrastructure setups. The partnership highlights Nvidia’s strategy of collaborating with major tech companies to expand its influence in the #AIindustry. cnbc.com/2025/11/17/arm-nvidia #tech #media #news

  30. #Arm and #Nvidia are #partnering to integrate Arm-based #Neoverse #CPUs with Nvidia’s #GPUs using Nvidia’s #NVLink Fusion technology. This #collaboration will benefit customers, particularly #hyperscalers, who prefer custom infrastructure setups. The partnership highlights Nvidia’s strategy of collaborating with major tech companies to expand its influence in the #AIindustry. cnbc.com/2025/11/17/arm-nvidia #tech #media #news

  31. #Arm and #Nvidia are #partnering to integrate Arm-based #Neoverse #CPUs with Nvidia’s #GPUs using Nvidia’s #NVLink Fusion technology. This #collaboration will benefit customers, particularly #hyperscalers, who prefer custom infrastructure setups. The partnership highlights Nvidia’s strategy of collaborating with major tech companies to expand its influence in the #AIindustry. cnbc.com/2025/11/17/arm-nvidia #tech #media #news

  32. #Arm and #Nvidia are #partnering to integrate Arm-based #Neoverse #CPUs with Nvidia’s #GPUs using Nvidia’s #NVLink Fusion technology. This #collaboration will benefit customers, particularly #hyperscalers, who prefer custom infrastructure setups. The partnership highlights Nvidia’s strategy of collaborating with major tech companies to expand its influence in the #AIindustry. cnbc.com/2025/11/17/arm-nvidia #tech #media #news

  33. #Arm and #Nvidia are #partnering to integrate Arm-based #Neoverse #CPUs with Nvidia’s #GPUs using Nvidia’s #NVLink Fusion technology. This #collaboration will benefit customers, particularly #hyperscalers, who prefer custom infrastructure setups. The partnership highlights Nvidia’s strategy of collaborating with major tech companies to expand its influence in the #AIindustry. cnbc.com/2025/11/17/arm-nvidia #tech #media #news

  34. 8x AMD Instinct #MI355X (288GB @8TB/s) take back the lead over 8x Nvidia #B200 (180GB @8TB/s) in #FluidX3D #CFD, achieving 362k MLUPs/s (vs. 219k MLUPs/s). Thanks to Jon Stevens from Hot Aisle to run the benchmarks! 🖖😊

    In single-GPU, both perform about the same, but in 8x #GPU config, MI355X is 65% faster. The difference comes from PCIe bandwidth - MI355X does 55GB/s, B200 only 14GB/s. #Nvidia leaves a lot of perf on the table by not exposing #NVLink P2P to #OpenCL.

    github.com/ProjectPhysX/FluidX

  35. 8x AMD Instinct #MI355X (288GB @8TB/s) take back the lead over 8x Nvidia #B200 (180GB @8TB/s) in #FluidX3D #CFD, achieving 362k MLUPs/s (vs. 219k MLUPs/s). Thanks to Jon Stevens from Hot Aisle to run the benchmarks! 🖖😊

    In single-GPU, both perform about the same, but in 8x #GPU config, MI355X is 65% faster. The difference comes from PCIe bandwidth - MI355X does 55GB/s, B200 only 14GB/s. #Nvidia leaves a lot of perf on the table by not exposing #NVLink P2P to #OpenCL.

    github.com/ProjectPhysX/FluidX

  36. 8x AMD Instinct #MI355X (288GB @8TB/s) take back the lead over 8x Nvidia #B200 (180GB @8TB/s) in #FluidX3D #CFD, achieving 362k MLUPs/s (vs. 219k MLUPs/s). Thanks to Jon Stevens from Hot Aisle to run the benchmarks! 🖖😊

    In single-GPU, both perform about the same, but in 8x #GPU config, MI355X is 65% faster. The difference comes from PCIe bandwidth - MI355X does 55GB/s, B200 only 14GB/s. #Nvidia leaves a lot of perf on the table by not exposing #NVLink P2P to #OpenCL.

    github.com/ProjectPhysX/FluidX

  37. 8x AMD Instinct #MI355X (288GB @8TB/s) take back the lead over 8x Nvidia #B200 (180GB @8TB/s) in #FluidX3D #CFD, achieving 362k MLUPs/s (vs. 219k MLUPs/s). Thanks to Jon Stevens from Hot Aisle to run the benchmarks! 🖖😊

    In single-GPU, both perform about the same, but in 8x #GPU config, MI355X is 65% faster. The difference comes from PCIe bandwidth - MI355X does 55GB/s, B200 only 14GB/s. #Nvidia leaves a lot of perf on the table by not exposing #NVLink P2P to #OpenCL.

    github.com/ProjectPhysX/FluidX

  38. 8x AMD Instinct #MI355X (288GB @8TB/s) take back the lead over 8x Nvidia #B200 (180GB @8TB/s) in #FluidX3D #CFD, achieving 362k MLUPs/s (vs. 219k MLUPs/s). Thanks to Jon Stevens from Hot Aisle to run the benchmarks! 🖖😊

    In single-GPU, both perform about the same, but in 8x #GPU config, MI355X is 65% faster. The difference comes from PCIe bandwidth - MI355X does 55GB/s, B200 only 14GB/s. #Nvidia leaves a lot of perf on the table by not exposing #NVLink P2P to #OpenCL.

    github.com/ProjectPhysX/FluidX

  39. Felkarolja az NVLink Fusiont a Samsung bérgyártó részlege

    Hirdetés Az NVIDIA az idei Computexen jelentette be az NVLink Fusiont, amely elviekben lehetővé teszi a partnerek számára, hogy…
    #Hungary #HU #Europe #Europa #EU #bérgyártó #egyéb #fusion #global #hír #hungary #Magyarország #Nvidia #nvlink #OCP #Samsung #summit #teszt
    europesays.com/2495362/

  40. L.SpringBootApplication xây dựng cluster GPU: Tư vấn thêm card GPU RTX 3090 cho hệ thống hiện tại. Hệ thống hiện tại: Ryzen 5950X, bo mạch chủ X570 Dark Hero III, dual RTX 3090 NVLink.dereferenceLOCAL Cần thêm 2 card RTX 3090 để tạo cluster 4 GPU. Các lựa chọn: sử dụng card bifurcation PCIe và Oculink, lo lắng về công suất nguồn và hiệu suất. Cân nhắc xây dựng hệ thống mới cho tương lai. #GPU #PCBuilding #RTX3090 #NVLink #Hardware

    reddit.com/r/LocalLLaMA/commen