#gpu — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #gpu, aggregated by home.social.
-
توّ نرجعوا مع خبر ل
#KDE #Plasma
ال 6.6 مش تكون
#LTS
https://pointieststick.com/2026/08/13/what-a-real-lts-looks-like-kubuntu-26-04/
في
#Linux
كان واحد يحبّ يعمل حاجة ما تخدم كان على
windows
ينجّم يستعمل يا
#VM
يا حاجة كيمة
#Winboat
أما كانت عندها حدود خاصّة مع الحاجات إلّي يستعمل ال
#GPU
أما هذي مش تجي ل
#Linux
https://blog.getutm.app/2026/introducing-triton-directx-11-driver-for-qemu/
https://github.com/winboat-org/winboat/tree/gpu-accel
#LinuxMint
يخدم على أكثر أدوات
https://blog.linuxmint.com/?p=5050 -
توّ نرجعوا مع خبر ل
#KDE #Plasma
ال 6.6 مش تكون
#LTS
https://pointieststick.com/2026/08/13/what-a-real-lts-looks-like-kubuntu-26-04/
في
#Linux
كان واحد يحبّ يعمل حاجة ما تخدم كان على
windows
ينجّم يستعمل يا
#VM
يا حاجة كيمة
#Winboat
أما كانت عندها حدود خاصّة مع الحاجات إلّي يستعمل ال
#GPU
أما هذي مش تجي ل
#Linux
https://blog.getutm.app/2026/introducing-triton-directx-11-driver-for-qemu/
https://github.com/winboat-org/winboat/tree/gpu-accel
#LinuxMint
يخدم على أكثر أدوات
https://blog.linuxmint.com/?p=5050 -
ИИ‑прогноз погоды на своём железе: запускаем AIFS 2.0 от ECMWF
Открытые веса модели мало помогают, если для её запуска нужен конкретный GPU и непростое окружение. Разберём, как запустить AIFS Single 2.0 от ECMWF через Hugging Face Jobs или локально, обойти зависимость от flash-attn и получить собственный прогноз погоды даже на CPU или Apple MPS. Запустить AIFS
https://habr.com/ru/companies/otus/articles/1069812/
#AIFS_20 #ECMWF #прогнозирование_погоды #машинное_обучение #нейросети #инференс #Hugging_Face #локальный_запуск #GPU #временные_ряды
-
RT @LottoLabs: Man, man sieht die verrücktesten Sachen auf localmaxxing 112 Concurrency 🤣 localmaxxing.com/en/runs/cms… Link Qwen3.6-35B-A3B — 6001.4 tok/s auf Radeon AI Pro R9700 · 32 GB ×4 Qwen3.6-35B-A3B erreichte 6001.4 tok/s Ausgabegeschwindigkeit auf Radeon AI Pro R9700 · 32 GB ×4 mit hipfire · MQ4R. Siehe den vollständigen Benchmark auf LocalMaxxing. localmaxxing.com
mehr auf Arint.info
-
RE: https://mastodon.social/@aleksandarilic/116153387777550859
it bears repeating, in light of #nvidia powering russian (and probably other's) rockets.
#ai #computers #tech #technology #politics #eu #europe #games #gaming #PC #war #death #antifa #corruption #world #genocide #gpu #economy #llm #it #RAM #oligarchy
-
RE: https://mastodon.social/@aleksandarilic/116153387777550859
it bears repeating, in light of #nvidia powering russian (and probably other's) rockets.
#ai #computers #tech #technology #politics #eu #europe #games #gaming #PC #war #death #antifa #corruption #world #genocide #gpu #economy #llm #it #RAM #oligarchy
-
RT @TeksEdge: 🔥 Deine 24GB-GPU kann nicht nur Modelle LAUFEN lassen, sondern sie auch lokal TRAINIEREN! Erlebe mit diesem Qwen3.8-27B! @Unsloth hat das lokale Fine-Tuning für Meta Muse Glimmer 30B hinzugefügt und sagt, dass du es auf nur 24GB VRAM machen kannst. Und dies beschränkt sich nicht nur auf Standard-Fine-Tuning: 🧠 30B-Modell 🎯 Lokal feinabstimmen 🏋️ GRPO-Verstärkungslernen ⚡ 1,5× schnelleres Training 💾 ~50% weniger VRAM im Vergleich zu FlashAttention-2-Setups 🆓 Kostenlose Trainings-Notebooks verfügbar Ein einzelner Consumer-GPU kann nun ein ernstzunehmendes 30B-Agentenmodell nehmen und ihm deine eigenen Fähigkeiten beibringen: 🤖 Agentenverhalten 🛠️ Werkzeugnutzung 💻 Coding-Workflows 👁️ Multimodale Aufgaben 📄 Spezialisierte Dokumente 🎯 Benutzerdefinierte Belohnungsfunktionen Dies ist kein Pretraining eines 30B-Modells von Grund auf. Es ist etwas viel Praktischeres. Ein bereits fähiges Modell nehmen und es dir eigen machen. Lokale KI bedeutete früher „Ich kann die Gewichte ausführen.“ Dann wurde es zu „Ich kann die Gewichte quantisieren.“ Jetzt wird es mehr wie 🔥 „Ich kann die Gewichte auch TRAINIEREN.“ Das fühlt sich an wie die nächste Phase des Localmaxxings.
mehr auf Arint.info
#FineTuning #GPU #LocalAI #MachineLearning #OpenSource #Unsloth #arint_info
-
RT @thdxr: Deepseek ist bei der Inferenz wahnsinnig gut – sie erreichen ein Cache-Verhältnis von 96,56 %, während unser zweitbester Anbieter nur 91,60 % schafft. Das mag nicht viel erscheinen, bedeutet aber, dass sie etwa die Hälfte der GPU-Zeit weniger benötigen.
mehr auf Arint.info
#CacheRatio #Deepseek #GPU #Inference #TechEfficiency #arint_info
-
RT @LottoLabs: Man, du siehst die verrücktesten Sachen auf LocalMaxxing 112 Concurrency 🤣 localmaxxing.com/en/runs/cms… Link Qwen3.6-35B-A3B — 6001.4 tok/s auf Radeon AI Pro R9700 · 32 GB ×4 Qwen3.6-35B-A3B erreichte eine Ausgabegeschwindigkeit von 6001.4 tok/s auf Radeon AI Pro R9700 · 32 GB ×4 mit hipfire · MQ4R. Siehe den vollständigen Benchmark auf LocalMaxxing. localmaxxing.com
mehr auf Arint.info
-
RT @LottoLabs: Man, du siehst die verrücktesten Sachen auf LocalMaxxing 112 Concurrency 🤣 localmaxxing.com/en/runs/cms… Link Qwen3.6-35B-A3B — 6001.4 tok/s auf Radeon AI Pro R9700 · 32 GB ×4 Qwen3.6-35B-A3B erreichte eine Ausgabegeschwindigkeit von 6001.4 tok/s auf Radeon AI Pro R9700 · 32 GB ×4 mit hipfire · MQ4R. Siehe den vollständigen Benchmark auf LocalMaxxing. localmaxxing.com
mehr auf Arint.info
-
Nvidia is personally guaranteeing 25% of the resale value on the GPUs backing its $500B financing deal. You don't insure an asset's value unless you're worried it might not have one.
https://blog.ppb1701.com/nvidias-25-bet-against-its-own-bubble
#nvidia #ai #aibubble #bigtech #circularfinancing #jensenhuang #gpu #datacenters #blog
-
Nvidia is personally guaranteeing 25% of the resale value on the GPUs backing its $500B financing deal. You don't insure an asset's value unless you're worried it might not have one.
https://blog.ppb1701.com/nvidias-25-bet-against-its-own-bubble
#nvidia #ai #aibubble #bigtech #circularfinancing #jensenhuang #gpu #datacenters #blog
-
🎉🎨 So, apparently, Gaussian Splatting is the new #yoga for GPUs, and #Julia is the instructor. 🤓🤔 Who knew that splatting 6 million Gaussians was the path to #enlightenment in Kyiv? But hey, at least your MacBook can now join the #GPU party—it's like being invited to a LAN party in 1999! 😂✨
https://pxl-th.github.io/blog/better-gs-julia/ #GaussianSplatting #MacBookParty #HackerNews #ngated -
🎉🎨 So, apparently, Gaussian Splatting is the new #yoga for GPUs, and #Julia is the instructor. 🤓🤔 Who knew that splatting 6 million Gaussians was the path to #enlightenment in Kyiv? But hey, at least your MacBook can now join the #GPU party—it's like being invited to a LAN party in 1999! 😂✨
https://pxl-th.github.io/blog/better-gs-julia/ #GaussianSplatting #MacBookParty #HackerNews #ngated -
Update your system, they say. It's important, they say.
So, running an absolute fine system and updating it leads to: #nvidia #gpu stopped working.
#GamesOnLinux running in 2 to 4 fps on an rtx 5070 (speaking of "surviving the aftermath", not a very "heavy" game). Before update, it runs at an insane smoothness, now, you can watch every frame.
Update was of kind "nvidia-something" and now it's all crap. #linuxSry guys, one thing, I cannot hold back longer: This never happend on windows.
-
Update your system, they say. It's important, they say.
So, running an absolute fine system and updating it leads to: #nvidia #gpu stopped working.
#GamesOnLinux running in 2 to 4 fps on an rtx 5070 (speaking of "surviving the aftermath", not a very "heavy" game). Before update, it runs at an insane smoothness, now, you can watch every frame.
Update was of kind "nvidia-something" and now it's all crap. #linuxSry guys, one thing, I cannot hold back longer: This never happend on windows.
-
Adding a secondary screen to your Qt-based product is often expensive and complex. What if you could turn your users' smartphones, tablets, or laptops into an interface for your application instead?
In this video, Christoph Sterz demonstrates how to extend your existing Qt application with #WebRTC streaming, allowing it to be accessed from any device with a modern web browser. This solution functions like a high-performance live stream, but with a critical advantage: full bidirectional interaction.
Using #GStreamer, Christoph handles video capture and encoding directly on the #GPU, ensuring efficiency while maintaining low latency. We also utilize a specialized backchannel to send touch, keyboard, and mouse inputs from the browser back to your app, meaning users can operate your software remotely.
This video covers:
- High-Performance Streaming: Why we use GStreamer to keep video processing on the GPU.
- Full #Interactivity: How to implement the WebRTC backchannel for remote control.
- Implementation Guide: A walkthrough of the code and the streaming pipeline.
- Real-World Use Cases: From headless industrial #IoT devices to remote maintenance and offsite quality assurance.Watch the full video: https://www.youtube.com/watch?v=BhJCvV8Z5-I
-
Adding a secondary screen to your Qt-based product is often expensive and complex. What if you could turn your users' smartphones, tablets, or laptops into an interface for your application instead?
In this video, Christoph Sterz demonstrates how to extend your existing Qt application with #WebRTC streaming, allowing it to be accessed from any device with a modern web browser. This solution functions like a high-performance live stream, but with a critical advantage: full bidirectional interaction.
Using #GStreamer, Christoph handles video capture and encoding directly on the #GPU, ensuring efficiency while maintaining low latency. We also utilize a specialized backchannel to send touch, keyboard, and mouse inputs from the browser back to your app, meaning users can operate your software remotely.
This video covers:
- High-Performance Streaming: Why we use GStreamer to keep video processing on the GPU.
- Full #Interactivity: How to implement the WebRTC backchannel for remote control.
- Implementation Guide: A walkthrough of the code and the streaming pipeline.
- Real-World Use Cases: From headless industrial #IoT devices to remote maintenance and offsite quality assurance.Watch the full video: https://www.youtube.com/watch?v=BhJCvV8Z5-I
-
RT @thdxr: DeepSeek ist bei der Inferenz wahnsinnig gut – sie erreichen ein Cache-Verhältnis von 96,56 %, während unser zweitbester Anbieter nur 91,60 % schafft. Das mag nicht viel erscheinen, bedeutet aber, dass sie etwa die Hälfte der GPU-Zeit weniger benötigen.
mehr auf Arint.info
#AI #Cache #DeepSeek #GPU #Inference #Performance #arint_info
-
RT @thdxr: DeepSeek ist bei der Inferenz wahnsinnig gut – sie erreichen ein Cache-Verhältnis von 96,56 %, während unser zweitbester Anbieter nur 91,60 % schafft. Das mag nicht viel erscheinen, bedeutet aber, dass sie etwa die Hälfte der GPU-Zeit weniger benötigen.
mehr auf Arint.info
#AI #Cache #DeepSeek #GPU #Inference #Performance #arint_info
-
En un contexto global donde los recursos de agua dulce aptos para el consumo humano son extremadamente limitados y están distribuidos de forma desigual, se están destinando millones de esos litros de agua dulce para refrigerar servidores y generar electricidad para el funcionamiento de los datacenters que abastecen sistemas de IA e IA generativa ⚠️
#AI #genAI #generativeAI #water #tech #technology #footprint #energy #datacenter #GPU #environment
-
En un contexto global donde los recursos de agua dulce aptos para el consumo humano son extremadamente limitados y están distribuidos de forma desigual, se están destinando millones de esos litros de agua dulce para refrigerar servidores y generar electricidad para el funcionamiento de los datacenters que abastecen sistemas de IA e IA generativa ⚠️
#AI #genAI #generativeAI #water #tech #technology #footprint #energy #datacenter #GPU #environment
-
GPU driven rendering in AnKi https://anki3d.org/gpu-driven-rendering-in-anki/
-
MoE на 5090 недобирает полтора раза, и дело не в маршрутизации
RTX 5090, одна и та же сборка llama.cpp, один драйвер, батч 1. Плотная Qwen3.5-9B выбирает 72% пропускной способности памяти карты. MoE Qwen3.5-35B-A3B на той же карте выбирает 41%. Маршрутизация экспертов тут почти ни при чём, я её измерил: ядра top-k занимают 7% времени. CUDA-графы захватываются, дыр между запусками нет, занятость SM 97%. Карта работает почти всё время, просто делает работу медленнее, чем позволяет память. Причина оказалась в том, сколько байт читает одно ядро за запуск. Я написал программу на ggml, которая гоняет тот же mul_mat_vec_q на холодных весах, и снял кривую: на 1 МБ ядро берёт 23% полосы, на 4.2 МБ уже 53%, на 33.6 МБ 91%. У MoE на одно ядро приходится 4.6 МБ, у плотной модели 22.2 МБ. Вот и весь разрыв. В статье: разбивка всех 437 матвеков по трассе nsys, четыре способа померить это неправильно и посидеть в каждой ловушке по очереди, проверка модели на другой архитектуре с ошибкой 3.7% и на четырёх глубинах контекста, цена сэмплера и всей обвязки llama-server. Последнее оказалось крупнее всего остального: пользователь видит 225 токенов в секунду там, где бенчмарк показывает 317.
https://habr.com/ru/articles/1069574/
#llamacpp #moe #rtx_5090 #пропускная_способность_памяти #cuda #gpu #инференс_ллм #бенчмарк #профилирование #qwen_35
-
Как оптимизировать расходы на инфраструктуру для искусственного интеллекта. Выбираем железо и модели под разные задачи
В отчете The State of AI 2025 агентство McKinsey заявляет, что 88% компаний уже используют ИИ хотя бы в одном бизнес-процессе. Grand View Research отмечает, что по их прогнозам глобальный рынок вырастет до 3,5 триллионов долларов к 2033. Уже сейчас мы можем отследить эту динамику. Однако AI-сервисы не бесплатны. Компании, которые начали с токенизированных API, часто обнаруживают, что при реальной нагрузке их счета растут быстрее, чем польза от ИИ. Поэтому возникает вопрос: как развернуть искусственный интеллект на собственной инфраструктуре и не переплатить. Привет! Привет! Меня зовут Сергей Ковалёв, я менеджер продукта
https://habr.com/ru/companies/selectel/articles/1069388/
#selectel #искусственный_интеллект #серверы #оптимизация #llm #gpu #ml #itинфраструктура #внедрение_ии
-
Linux GPU Control Application LACT: v0.10.0 released https://playingtux.com/en/articles/2026/08/linux-gpu-control-application-lact-v0100-released/ #Linux #Gaming #LinuxGaming #GPU #Overclocking #OpenSource
-
Linux GPU Control Application LACT: v0.10.0 released https://playingtux.com/en/articles/2026/08/linux-gpu-control-application-lact-v0100-released/ #Linux #Gaming #LinuxGaming #GPU #Overclocking #OpenSource
-
Para los que argumentan que el coste energético de los modelos de IA generativa es igual al de usar internet o enviar un email, sepan que no. Los modelos como GPT-4 se entrenan y se implementan en servidores que consumen 100 veces más que un modelo simple de Inteligencia Artificial (a secas).
#AI #genAI #ChatGPT #Datacenter #GPU #water #energy #tech #technology #environment #generativeAI #OpenAI #stopgenAI
-
Para los que argumentan que el coste energético de los modelos de IA generativa es igual al de usar internet o enviar un email, sepan que no. Los modelos como GPT-4 se entrenan y se implementan en servidores que consumen 100 veces más que un modelo simple de Inteligencia Artificial (a secas).
#AI #genAI #ChatGPT #Datacenter #GPU #water #energy #tech #technology #environment #generativeAI #OpenAI #stopgenAI
-
Паравиртуализация при отладке графических приложений KasperskyOS в QEMU
Привет! Меня зовут Денис Молодяков, я — тимлид команды графики в KasperskyOS . Мы отвечаем за разработку графического стека полного цикла для микроядерной ОС: от создания низкоуровневых графических драйверов до всего необходимого для фреймворков. Современная реальность разработки диктует новые правила: удаленка и распределенные команды стали нормой. Моя команда не исключение — ребята базируются в разных регионах. При этом каждому разработчику требуется доступ к аппаратным платформам для сборки, запуска ОС, написания и отладки драйверов. Обеспечение каждого сотрудника полным набором плат сопряжено с серьезными логистическими и финансовыми затратами. В этой статье я хочу рассказать, как мы искали оптимальное решение этой проблемы и почему пришли к использованию методики паравиртуализации.
https://habr.com/ru/companies/kaspersky/articles/1067226/
#gpu #virtio #ci #системное_программирование #паравиртуализация #микроядро #wayland #kasperskyos #qemu #виртуализация
-
TVB 進軍 AI 算力中心 提供人工智能算力服務 擬向外部租用 GPU
無綫集團宣布擬與基滙資本合組公司,於將軍澳電視城園區興建算力設施,採購 GPU 及 CPU 等硬件,向外部客戶 […]
#人工智能 #新科創業 #智慧城市 #GPU
https://unwire.hk/2026/08/11/tvb-gaw-capital-ai-computing/unwire-space/?utm_source=rss&utm_medium=rss&utm_campaign=tvb-gaw-capital-ai-computing -
New AI infrastructure for data-sovereign research in Germany: The new #GPU cluster combines #Fraunhofer IGD’s #AI expertise with the energy-efficient data center infrastructure of GSI/FAIR’s #GreenIT Cube. The initial pilot applications come from very different research fields: https://www.gsi.de/en/start/news/details/2026/08/11/ki-cluster-fraunhofer-igd
-
New AI infrastructure for data-sovereign research in Germany: The new #GPU cluster combines #Fraunhofer IGD’s #AI expertise with the energy-efficient data center infrastructure of GSI/FAIR’s #GreenIT Cube. The initial pilot applications come from very different research fields: https://www.gsi.de/en/start/news/details/2026/08/11/ki-cluster-fraunhofer-igd
-
In SuperMUC-NG 2 am LRZ arbeiten #GPU von Intel. In einem Bootcamp können Forschende der #lifescience #biologie Bioinformatik und Materialwissenschaften das System am 17. September kennenlernen:
• Umgang mit dem #HPC -System
• skalierbare #Workflows aufbauen,
• mit Forschenden vernetzen.
Bis zum 31. August läuft die Bewerbungsfrist für das Bootcamp. es ist für Masterstudierende, Doktorand:innen und Postdocs gemacht_ https://app1.edoobox.com/en/LRZ/Bootcamps%20and%20Hackathons/Workshop.ed.8d3ac6031182_10473187807.Life%20and%20Material%20Science%20Bootcamp%20on%20SuperMUC-NG%20Phase%202%20apply%20via%20e-mail -
In SuperMUC-NG 2 am LRZ arbeiten #GPU von Intel. In einem Bootcamp können Forschende der #lifescience #biologie Bioinformatik und Materialwissenschaften das System am 17. September kennenlernen:
• Umgang mit dem #HPC -System
• skalierbare #Workflows aufbauen,
• mit Forschenden vernetzen.
Bis zum 31. August läuft die Bewerbungsfrist für das Bootcamp. es ist für Masterstudierende, Doktorand:innen und Postdocs gemacht_ https://app1.edoobox.com/en/LRZ/Bootcamps%20and%20Hackathons/Workshop.ed.8d3ac6031182_10473187807.Life%20and%20Material%20Science%20Bootcamp%20on%20SuperMUC-NG%20Phase%202%20apply%20via%20e-mail -
The AI boom is driving up demand for HBM, DRAM, SSDs, and server CPUs, pushing manufacturers to prioritize high-margin AI infrastructure over traditional consumer hardware. The result: memory and storage are becoming strategic bottlenecks, while CPUs are also seeing stronger demand as AI workloads scale.
AI is no longer just a GPU story. It is reshaping the economics of compute, memory, and storage.
https://www.buysellram.com/blog/ai-is-repricing-memory-storage-and-cpus-not-just-gpus/
#AI #Memory #DRAM #HBM #SSD #CPU #GPU #DataCenter #technology
-
The AI boom is driving up demand for HBM, DRAM, SSDs, and server CPUs, pushing manufacturers to prioritize high-margin AI infrastructure over traditional consumer hardware. The result: memory and storage are becoming strategic bottlenecks, while CPUs are also seeing stronger demand as AI workloads scale.
AI is no longer just a GPU story. It is reshaping the economics of compute, memory, and storage.
https://www.buysellram.com/blog/ai-is-repricing-memory-storage-and-cpus-not-just-gpus/
#AI #Memory #DRAM #HBM #SSD #CPU #GPU #DataCenter #technology
-
NVIDIA sold the RTX 5060 Ti with 8GB or 16GB. Same chip. Launch price difference: $50.
Today the market difference is $191.
The 4060 Ti ran the same experiment two years earlier and landed within 50 cents per gigabyte.https://www.buysellram.com/blog/nvidia-consumer-gpu-price-report-august-2026/
#NVIDIA #GPUPrices #RTX50 #RTX5090 #VRAM #GDDR7 #DRAM #GPU #PCHardware #ITAD #ITAssetDisposition #DataCenter #ResaleValue #technology