#ai-hardware — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #ai-hardware, aggregated by home.social.
-
The Google DeepMind reorg moves Demis Hassabis toward AGI and signals a split between frontier research and products. What it means for
-
Puzzle accounting software pairs an AI agent platform with its own ledger to cut close time. We explain the trade-offs and checks for
https://aistory.news/ai-startups-and-companies/what-puzzle-accounting-software-changes-for-startups/
-
Puzzle accounting software pairs an AI agent platform with its own ledger to cut close time. We explain the trade-offs and checks for
https://aistory.news/ai-startups-and-companies/what-puzzle-accounting-software-changes-for-startups/
-
StartupHub.ai directory claims 65M profiles and 5B data points. We compare it with Forbes’ AI 50 and YC’s listings to show who should use
-
StartupHub.ai directory claims 65M profiles and 5B data points. We compare it with Forbes’ AI 50 and YC’s listings to show who should use
-
The Forbes AI 50 shows where power is shifting in 2026. Inside Anthropic’s surge, EliseAI’s move into healthcare, and what it means for
-
The Forbes AI 50 shows where power is shifting in 2026. Inside Anthropic’s surge, EliseAI’s move into healthcare, and what it means for
-
YC AI startups hit 1,531 in August 2026. We map how Y Combinator’s AI roster aligns with Forbes’ AI 50 winners—and where founders should
https://aistory.news/ai-startups-and-companies/yc-ai-startups-top-1500-where-the-real-traction-is/
-
YC AI startups hit 1,531 in August 2026. We map how Y Combinator’s AI roster aligns with Forbes’ AI 50 winners—and where founders should
https://aistory.news/ai-startups-and-companies/yc-ai-startups-top-1500-where-the-real-traction-is/
-
Cornell investing study warns that ‘comparable’ metrics can inflate confidence and hurt returns. Here’s how to compare KPIs without fooling
-
Cornell investing study warns that ‘comparable’ metrics can inflate confidence and hurt returns. Here’s how to compare KPIs without fooling
-
Sequoia AI strategy is hiding in plain sight on its homepage. We map the themes—AGI bets, open source, PMF playbooks—and what founders
https://aistory.news/ai-startups-and-companies/what-sequoia-ai-strategy-reveals-on-its-homepage-now/
-
Sequoia AI strategy is hiding in plain sight on its homepage. We map the themes—AGI bets, open source, PMF playbooks—and what founders
https://aistory.news/ai-startups-and-companies/what-sequoia-ai-strategy-reveals-on-its-homepage-now/
-
The Cornell vice provost entrepreneurship move signals a one-stop path for founders and partners. Here’s how it could change IP terms,
-
The Cornell vice provost entrepreneurship move signals a one-stop path for founders and partners. Here’s how it could change IP terms,
-
AI Conference 2026 lands in San Francisco with 7 tracks and 120+ speakers. Our read of the early lineup: infra speed meets safety and
-
AI Conference 2026 lands in San Francisco with 7 tracks and 120+ speakers. Our read of the early lineup: infra speed meets safety and
-
Cornell FarmNet 40 years on, the Cornell-backed service pairs financial and family counseling to keep NY farms afloat. Why its model still
-
Cornell FarmNet 40 years on, the Cornell-backed service pairs financial and family counseling to keep NY farms afloat. Why its model still
-
Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
https://www.wafer.ai/blog/kimi-k3-mi355x
Comments: https://news.ycombinator.com/item?id=49141073
#HackerNews #KimiK3 #MI355X #BetterPerformance #B300 #TechNews #AIHardware
-
Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
https://www.wafer.ai/blog/kimi-k3-mi355x
Comments: https://news.ycombinator.com/item?id=49141073
#HackerNews #KimiK3 #MI355X #BetterPerformance #B300 #TechNews #AIHardware
-
Cornell MBA AI report: Cornell says jobs are steady, but standards are rising as employers expect AI fluency. See what it means for
https://aistory.news/ai-startups-and-companies/cornell-mba-ai-employers-raise-the-bar-for-new-hires/
-
Cornell MBA AI report: Cornell says jobs are steady, but standards are rising as employers expect AI fluency. See what it means for
https://aistory.news/ai-startups-and-companies/cornell-mba-ai-employers-raise-the-bar-for-new-hires/
-
RT @__tinygrad__: Die Benchmark-Ergebnisse liegen vor: GLM-5.2 auf einer $160k Tinybox Pro v2 Black erreicht 119 Tokens pro Sekunde im Single-Use-Modus und 917 im Aggregat-Modus! Die Inbetriebnahme wurde von (einem anderen) GLM-5.2 in nur einer Stunde durchgeführt, sodass definitiv noch weiteres Leistungspotenzial besteht.
mehr auf Arint.info
#AIHardware #Benchmark #GLM52 #LLMPerformance #TinyboxPro #TokenSpeed #arint_info
-
TechCrunch’s July 31 report on the Hugging Face breach shows noisy attackers still slip into model hubs—and why supply-chain basics now
https://aistory.news/ai-startups-and-companies/what-the-hugging-face-breach-says-about-ai-risk/
-
RT @vertonbiz: 🔋 Nvidia geht sein größtes Wagnis bisher ein für Ilya Sutskever. Safe Superintelligence (SSI), das Startup von OpenAI-Mitbegründer Ilya Sutskever, hat eine große Partnerschaft mit Nvidia bekanntgegeben. Nvidia gibt an, dass es selten gewordenen Zugang zu SSI’s streng gehüteter Forschung erhalten hat und genug gesehen hat, um ein massives Engagement einzugehen. Laut dem Unternehmen hat SSI bereits „erhebliche Forschungsdurchbrüche“ erzielt, die echten Fortschritt in Richtung seiner Mission demonstrieren. Das Abkommen umfasst: - Eine mehrmilliardenschwere Investition (einige Berichte schätzen rund 5 Mrd. Dollar) 💰 - Zugang zu Nvidias nächster Generation der Vera-Rubin-Plattform, was SSI’s Rechenkapazität drastisch erhöht. - Gemeinsame Entwicklung zukünftiger Nvidia-KI-Hardware, wobei SSI dabei hilft, die nächsten Computing-Plattformen des Unternehmens zu gestalten. Dieser letzte Punkt ist am interessantesten. Nvidia hat nur einen Teil von SSI’s Forschung gesehen, ist aber bereits bereit, zukünftige Hardware um Sutskevers Überlegungen zu gestalten, wohin die KI sich entwickelt. Sutskever hat wiederholt argumentiert, dass die Ära des einfachen Skalierens von LLMs (Large Language Models) ihrem Ende zugeht. Wenn Nvidia so große Wetten eingeht, baut SSI möglicherweise etwas grundlegend anderes als heutige KI-Architekturen. Was auch immer sie entwickeln, Nvidia will offensichtlich nicht den Anschluss verlieren. Video
mehr auf Arint.info
#AIHardware #IlyaSutskever #Innovation #KI #Nvidia #SafeSuperintelligence #arint_info
-
RT @Italianclownz: 🤯Die Generierungsgeschwindigkeit ist das herausragende Merkmal — ROCmFP4 ist auf identischer Hardware bei der Dekodierung fast 1,8x schneller und erzielt zudem mehr Effizienz durch MTP (höhere Akzeptanzrate = weniger verschwendete Entwurfsdurchläufe), was den Vorteil weiter verstärkt. Dies stimmt mit den Angaben in den ROCmFPX-Dokumenten zu AMD-optimierten Kernels überein (unalinierte dword-Ladungen, ROCm-spezifische MMQ-Pfade) — auf dieser Hardware ist das Modell nicht nur kleiner, sondern tatsächlich schneller, kein Kompromiss zwischen Größe und Geschwindigkeit. ROCmFP4 ist bei der Prompt-Verarbeitung ~20% und bei der Generierung ~79% schneller, dabei auf der Festplatte 21% kleiner (14,15 GB vs. 17,91 GB) und erzielt bessere Ergebnisse bei MTP. Kleiner + Schneller + Kohärenter
mehr auf Arint.info
#AIHardware #AMD #GenerativeAI #MachineLearning #Performance #ROCmFP4 #arint_info
-
The Cyera Oasis acquisition, a $1B deal, signals security converging on data and machine identities as AI agents spread. What changes next.
-
RT @Sentdex: Etwas, was ich persönlich gehört habe und immer wieder sehe, ist viel Wut über große Modelle wie Kimi K3 oder sogar GLM 5.2 sowie über Rechenkapazität wie die, die Tobi geteilt hat. Ein 100.000-Dollar-Computer liegt tatsächlich außerhalb dessen, was sich die meisten Menschen leisten können. Ohne Zweifel. Kimi K3 kann nicht einmal auf diesem 100.000-Dollar-Computer laufen. Was mich an Kimi K3 am meisten begeistert, ist nicht, dass ich es lokal betreiben werde. Sondern dass viele Anbieter es bereitstellen werden. Sie werden herausfinden, wie man es günstig und schnell bereitstellen kann. Dann werden Menschen mit diesen 100.000-Dollar-Maschinen und teuren Cloud-Computern es in viel, viel, viel kleinere Modelle verdichten. Kimi K3 wird der Lehrer einer riesigen Anzahl zukünftiger Modelle sein. Dann werden Menschen mit 3/4/5090er-Grafikkarten viele dieser Modelle betreiben. Alles ist gut. Die Vorstellung, GLM-5.2-Niveau-Intelligenz auf einem Computer zu betreiben, den man sich vor nur einem Jahr für 50.000 Dollar hätte kaufen können, war AUSGESENDET unmöglich. Man brauchte Millionen von Dollar und würde trotzdem GLM-5.2 nicht erreichen. Jetzt braucht man 50.000 Dollar. Dies ist ein wunderbarer Trend und es ist alles gut für dich, egal wo du dich auf der Rechenleiter befindest. Trickle-down-Tokenökonomie Michael Dell 🇺🇸 (@MichaelDell) Was @tobi hier macht, ist bemerkenswert: GLM-5.2, ein KI-Modell mit 753 Milliarden Parametern, lokal auf einem Dell Pro Max mit GB300 bei 40 Token pro Sekunde zu betreiben. Kein Rechenzentrum. Keine API. Keine Cloud. Unbegrenzte Intelligenz. KI an d…
mehr auf Arint.info
#AIHardware #ComputeDemocratization #FrontierAI #GLM52 #KimiK3 #LocalAI #arint_info
-
RT @kyzoroXX: Enterprise-AI-Server haben jahrelang ein wichtiges Element vermisst: HBM über PCIe. AMD ändert dies nun mit dem Instinct MI350P. • 144 GB HBM3e • 4,0 TB/s Speicherbandbreite • PCIe Gen5 x16 Dies sind nicht nur größere Zahlen. HBM-Speicher war bisher weitgehend für rackbasierte AI-Systeme reserviert. Die Integration in einen Standard-PCIe-Beschleuniger könnte die Nutzung größerer LLMs auf bestehenden Enterprise-Servern ohne den Umstieg auf spezialisierte Infrastruktur praktikabel machen. Bei vielen Inferenz-Workloads werden Speicherkapazität und -bandbreite zum Engpass, lange bevor die reine Rechenleistung limitierend wird. Manchmal ist das wichtigste Upgrade nicht mehr TFLOPS, sondern die Beseitigung des Speicherengpasses. Video KyzoroX (@kyzoroXX) Artikel DGX Spark vs. AMD Strix Halo vs. Mac Mini: Der ehrliche 3-Wege-Benchmark. Ein 600 $ Mac Mini generiert Token fast so schnell wie ein 4.700 $ DGX Spark. Diese eine Tatsache verändert die Art und Weise, wie jeder lokale AI-Hardware kauft — bis man die in Benchmarks versteckte Kennzahl kennt. Alle testen — https://nitter.net/kyzoroXX/status/2076791492778287387#m
mehr auf Arint.info
#AIHardware #AMDInstinct #HBM3e #LLM #ServerUpgrade #TechBenchmark #arint_info
-
AIM Intelligent Machines is retrofitting bulldozers and mining rigs for autonomy from a former SpaceX site in Redmond, and betting that
-
AIM Intelligent Machines is retrofitting bulldozers and mining rigs for autonomy from a former SpaceX site in Redmond, and betting that
-
Nvidia’s Vera Rubin is a reminder that AI servers aren’t just about bigger GPUs. Fast CPU orchestration and tighter CPU-GPU links matter as agentic workloads scale.
-
Nvidia’s Vera Rubin is a reminder that AI servers aren’t just about bigger GPUs. Fast CPU orchestration and tighter CPU-GPU links matter as agentic workloads scale.
-
A Cornell battery startup born in a Johnson School class reached a $1.3B valuation. Inside the support system that helped it scale—and what
https://aistory.news/ai-startups-and-companies/cornell-battery-startups-path-to-a-13b-valuation/
-
A Cornell battery startup born in a Johnson School class reached a $1.3B valuation. Inside the support system that helped it scale—and what
https://aistory.news/ai-startups-and-companies/cornell-battery-startups-path-to-a-13b-valuation/
-
OpenAI's Micro Keypad: A Tool for Coders with Mixed Reactions
📰 Original title: I tried out OpenAI’s new AI keypad — which will be fun for some coders and slightly mystifying to everyone else
🤖 IA: It's clickbait ⚠️
👥 Users: It's clickbait ⚠️View full AI summary https://en.killbait.com/openai-s-micro-keypad-a-tool-for-coders-with-mixed-reactions.html?utm_source=mastodon_world&utm_medium=social&utm_campaign=killbait.mastodon_world
-
OpenAI's Micro Keypad: A Tool for Coders with Mixed Reactions
📰 Original title: I tried out OpenAI’s new AI keypad — which will be fun for some coders and slightly mystifying to everyone else
🤖 IA: It's clickbait ⚠️
👥 Users: It's clickbait ⚠️View full AI summary https://en.killbait.com/openai-s-micro-keypad-a-tool-for-coders-with-mixed-reactions.html?utm_source=mastodon_world&utm_medium=social&utm_campaign=killbait.mastodon_world
-
AMD just took its most direct swing yet at Nvidia's AI dominance.
What's inside:
72 Instinct MI455X GPUs paired with Epyc CPUs in a single rack
Su called the MI455X "the most powerful GPU in the industry" — a direct shot at the market leader#AMD #Nvidia #AIHardware #DataCenter #Semiconductors #GPU #Helios
-
AMD launches EPYC "Venice" — the first 2nm data center CPU, and it's aiming straight at NVIDIA's Vera.
At its Advancing AI 2026 event, AMD officially rolled out its 6th Gen EPYC "Venice" family, built on Zen 6 and the industry's first HPC chips in volume production on TSMC's 2nm process.
The headline specs:
Up to 256 cores / 512 threads
203 billion transistors (2nm compute chiplets, 6nm I/O dies)
Clocks over 5 GHz (up to 5.15 GHz on select 9006X variants)
Up to 1.6 TB/s memory bandwidth and PCIe Gen 6
Versus NVIDIA's Arm-based Vera CPU, AMD claims:
Up to 20% faster single-core (1.2x)
Up to 2.2x higher throughput
Up to 2.8x agents/watt in agent "sandbox" servers and 3.3x performance/watt in general-purpose servers
Against its own 5th Gen "Turin" chips, Venice adds up to 1.8x tokens/s and 1.7x agents/watt.
Rollout is staggered: the flagship SP7 ships Q4 2026, with the mainstream SP8 (8–128 cores) following in 1H 2027.
the take: The real battleground in the AI era isn't just GPUs — it's the host CPU feeding them, and moving to 2nm this early gives AMD a density and efficiency lead that's hard to match on older nodes.
#AMD #EPYC #Zen6 #DataCenter #AIhardware #Semiconductors #TSMC #ServerCPU
-
As founders weigh startup GPU servers over cloud, a16z promotes DIY builds while HBR flags manager strain. Investors get a new diligence
https://aistory.news/ai-startups-and-companies/why-startup-gpu-servers-are-back-on-investors-lists/
-
https://winbuzzer.com/2026/07/23/microsoft-plans-amd-helios-and-epyc-expansion-on-azure-xcxwbn/
Microsoft has outlined plans to deploy AMD's Helios rack-scale AI system on Azure, expanding infrastructure choice while deployment scale remains undisclosed.
#AI #AMD #MicrosoftAzure #Azure #Microsoft #AIInfrastructure #AIHardware #AICompute #AIChips #MicrosoftCloud
-
Sifted says H1 2026 saw the lowest European fintech funding deal count in over a decade. Why “AI‑native” now decides who still gets term
-
Snowflake Postgres GA lands inside the AI Data Cloud, promising fewer pipelines and fresher data for agentic AI. Here’s what changes and
https://aistory.news/ai-startups-and-companies/snowflake-postgres-ga-brings-oltp-closer-to-ai-data/
-
https://winbuzzer.com/2026/07/20/apple-reportedly-widens-openai-case-to-40-former-staff-xcxwbn/
Apple has sent preservation letters to around 40 former employees at OpenAI, expanding a contested trade-secret case and its potential evidence pool.
#AI #Apple #OpenAI #Lawsuits #Legal #IntellectualProperty #Employees #TalentPoaching #AITalentWar #AIHardware
-
Brussels’ GenAI4EU initiative pairs the AI Act with funding, compute and sector pilots. Here’s how it shifts Europe from rulemaker to
-
OpenAI's first dedicated AI device is reportedly taking shape as a screen-free, portable companion designed for natural conversations. Developed in collaboration with former Apple design chief Jony Ive, the device could mark a new chapter beyond smartphones and traditional screens.
#BestSoln #BestSolution #OpenAI #AIHardware #JonyIve #ConsumerTech #TechNews