home.social

#ai-infrastructure — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ai-infrastructure, aggregated by home.social.

fetched live
  1. Open should mean more than downloadable weights.
    K2 Horizon releases six models—from 0.9B to 375B parameters—together with training code, data or recipes, checkpoints and evaluations.
    This makes model development inspectable and reproducible: an important reference point for sovereign AI.
    ifm.ai/k2/
    #OpenSourceAI #SovereignAI #AIInfrastructure

  2. Nvidia has agreed to buy Hugging Face, the popular platform for hosting millions of open AI models, for 13B USD. The chip giant says the acquisition will accelerate the spread of open-weight AI models. Hugging Face will remain open despite the takeover. arstechnica.com/ai/2026/09/nvi #AIagent #AI #GenAI #AIInfrastructure

  3. AI infrastructure startup Crusoe has raised 3B USD at a 30B USD valuation, according to reports. The funding round came together after the data centre developer secured a 13B USD contract with Jane Street. The deal signals continued massive investment in AI computing infrastructure. techcrunch.com/2026/09/03/crus #AIagent #AI #GenAI #AIInfrastructure

  4. NVIDIA has agreed to acquire Hugging Face for nearly $13 billion.

    This raises an interesting structural question for AI: can models & tools become increasingly open while the infrastructure around them becomes more concentrated?

    We look at Hugging Face’s role in the open AI ecosystem, NVIDIA expanding position across the AI stack and why hardware neutrality and interoperability will be worth watching.

    tinyurl.com/yc6t5tn9

    #NVIDIA #HuggingFace #AI #OpenSource #AIInfrastructure #OpenModels

  5. India’s AI infrastructure race is scaling fast. 🇮🇳

    Yotta plans to order 50,000 NVIDIA Vera Rubin GPUs (~$7.5B), plus 45,000 GB300 GPUs.

    The race isn’t just for AI models, it’s for the compute to run them at scale. 🚀

    #BestSoln #Yotta #NVIDIA #AI #AIInfrastructure

  6. Perplexity has open sourced Lily, a Rust and Metal inference engine for running Qwen3.6-35B-A3B on Apple Silicon. The specialised engine achieves 1.23x faster prefill and 1.35x faster decode than MLX-LM, demonstrating how narrow hardware optimisation can beat general-purpose frameworks. marktechpost.com/2026/09/02/pe #AIagent #AI #GenAI #AIInfrastructure

  7. Nvidia is projecting sales growth of up to 70%, and the reason goes far beyond chatbots and coding assistants. What we're witnessing is the physical build-out of AI itself — data centers, power grids, GPUs, and networking infrastructure being deployed at a scale the tech industry has never seen.

    #nvidia #technews #artificialintelligence #aiinfrastructure #futureoftech #startupnews #techtrends #innovations #datacenters #techindustry

  8. **RAM prices are undergoing one of the most significant increases the memory market has seen in recent years.**

    The latest BuySellRam market update highlights just how dramatic the move has been: **DDR5 prices have increased by as much as 473% in 2026**. This is not simply a normal cycle of memory pricing. The market is being reshaped by the rapid expansion of AI infrastructure, data-center demand, and the growing amount of memory required by modern computing systems.

    AI servers are consuming enormous amounts of DRAM and high-bandwidth memory, putting additional pressure on an already constrained supply chain. At the same time, memory manufacturers have to balance capacity between conventional DRAM, server memory, HBM, and other high-value products. As AI-related demand continues to compete for manufacturing capacity, traditional PC and server memory can also feel the effects.

    For PC builders, the result is higher system costs. For enterprises and data-center operators, expensive DRAM can significantly increase the cost of server upgrades and infrastructure expansion. A memory shortage can also change procurement strategies, encouraging businesses to extend the useful life of existing servers and pay closer attention to their current hardware inventory.

    There is another important consequence: **the secondary RAM market becomes more valuable when new memory becomes expensive.** DDR4 and DDR5 modules that might previously have been considered low-value surplus can become meaningful assets when replacement costs rise sharply. Organizations refreshing servers or consolidating infrastructure should therefore consider the resale value of their existing memory rather than automatically treating it as obsolete equipment.

    The 473% increase also raises a broader question about how sustainable current memory pricing is. If AI infrastructure continues to absorb a growing share of global memory production, the traditional boom-and-bust DRAM cycle could look very different in the years ahead.

    This market is worth watching closely—not only for buyers of new hardware, but also for companies holding large inventories of **server RAM, DDR4, and DDR5**.

    buysellram.com/blog/ram-market

    #RAM #DDR5 #DDR4 #DRAM #Memory #MemoryMarket #AIInfrastructure #DataCenters #ServerMemory #Semiconductors #PCHardware #ITHardware #DataCenter #AI #HardwareMarket

  9. #VMware Cloud Foundation, #VCF #AI Services, VMware Private AI Cloud, VMware AI Factory, #Tanzu, #vDefend…what's the difference? Where's #Nvidia in all this? What's #Broadcom 's stance on #opensource and open-weight models, especially with Nvidia reportedly buying #Hugging_face?

    I caught up with Paul Turner, VCF Division chief product officer at Broadcom, to sort out these questions as #VMwareExplore kicked off this week in Las Vegas.

    In today’s episode, we cover…

    💡 Broadcom's stance on #sovereignAI and #privateAI

    💡 The hardware supply crisis in #AIinfrastructure

    💡 Broadcom's outlook on autonomous IT operations

    And more!

    Tune in here: youtu.be/7EzRXISOtwc?si=UIVgG8

  10. MLPerf Storage v3.0 adds tests for KV-cache I/O, vector databases, and S3 storage: mlcommons.org/2026/09/mlperf-s

    The larger point: AI infrastructure is becoming a data-movement problem as much as a compute problem. Context, embeddings, and checkpoints can bottleneck a system even when GPUs are fast. #AIInfrastructure #MLPerf

  11. SLB is buying thermal-management company Kelvion for its data-center business. The deeper shift: AI infrastructure is becoming a coupled power-and-thermal system. Cooling affects sustained performance, rack density, reliability, and facility reuse: not just temperature.

    investorcenter.slb.com/news-re #AIInfrastructure

  12. Nvidia is investing 3.5B USD in MediaTek to strengthen its position in AI chip infrastructure as Big Tech companies build their own chips. The partnership shows how Nvidia aims to remain essential to AI data centres despite increasing competition from customers like Google and Amazon. techcrunch.com/2026/08/31/nvid #AIagent #AI #GenAI #AIInfrastructure

  13. A new benchmark measures Time to First Token (TTFT) latency across voice and realtime AI API providers. Grok Voice Think Fast 2.0 leads with 0.70s TTFT, highlighting the critical importance of latency for voice agent applications. marktechpost.com/2026/08/30/lo #AIagent #AI #GenAI #AIInfrastructure

  14. Lambda, an AI cloud company, has secured 1B USD in private debt to purchase Nvidia AI chips for lease to Microsoft. The deal, arranged by JP Morgan Chase, is the latest in a string of loans funding GPU infrastructure as banks and tech companies have raised over 400B USD in AI-related debt globally in 2026. techcrunch.com/2026/08/28/neoc #AI #GenAI #AIInfrastructure

  15. #Arm turned heads this week with new details about its plans to design a dual-purpose chip with #ibmz, but the overall story of Arm, the significant milestones it's reached in the last year and the implications for enterprise #privateAI infrastructure investment decisions are much broader and deeper than that.

    Multiple industry experts helped me put together a 360-degree view of the chip and processor market amid multiple intersecting trends, from #AI governance and privacy to the global memory shortage.

    Read it here: techtarget.com/it-infrastructu #AIinfrastructure #datacenter

  16. Server DRAM prices are being reshaped by rising AI and data-center demand, while long-term contracts are changing how memory is purchased and priced. For server operators, OEMs, and IT asset managers, these shifts could affect procurement costs, supply security, inventory values, and the resale market. Read the full analysis:

    buysellram.com/blog/server-dra

    #DRAM #ServerMemory #AI #DataCenters #MemoryMarket #Semiconductors #AIInfrastructure #ITHardware #SupplyChain #BuySellRam

  17. Server DRAM prices are being reshaped by rising AI and data-center demand, while long-term contracts are changing how memory is purchased and priced. For server operators, OEMs, and IT asset managers, these shifts could affect procurement costs, supply security, inventory values, and the resale market. Read the full analysis:

    buysellram.com/blog/server-dra

    #DRAM #ServerMemory #AI #DataCenters #MemoryMarket #Semiconductors #AIInfrastructure #ITHardware #SupplyChain #BuySellRam

  18. Intel’s Diamond Rapids could mark a major shift in server CPU performance, with configurations reportedly reaching up to 256 cores and 256 threads. Built for demanding AI, cloud, and data-center workloads, the platform highlights the growing race for higher compute density, memory bandwidth, and efficiency.

    Read the full analysis: buysellram.com/blog/intel-diam

    #Intel #DiamondRapids #CPU #AI #DataCenters #ServerCPU #Semiconductors #CloudComputing #AIInfrastructure #BuySellRam

  19. Australian Aboriginal groups are demanding a say in AI infrastructure decisions after opposing a proposed hyperscale data centre on culturally significant land. The Mandoon Bilya site in Perth was scrapped following community protests, but advocates say Indigenous communities have been largely excluded from AI policy planning despite lasting impacts on their Country. theconversation.com/the-fight- #AIagent #AI #GenAI #AIInfrastructure