home.social

#local-llms — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #local-llms, aggregated by home.social.

fetched live
  1. Learn how KV cache spilling hurts LLM inference, how to detect the cliff, and how to estimate it using real VRAM and throughput signals. hackernoon.com/how-to-find-the #localllms

  2. GLM-4.7-Flash is a 30B model with 3B active params, so why did it crawl on my mini PC? Not the weights, not the quant. The attention design, and the runtime. hackernoon.com/a-30b-model-cra #localllms

  3. Your local LLM runs fine until it doesn't. A look at KV cache spilling from VRAM into shared memory, and why it happens silently on Windows. hackernoon.com/why-local-llms- #localllms

  4. J'ai pris un moment pour jouer avec les données de #LLM de @betagouv

    👉 comparia.beta.gouv.fr/ranking

    Un modèle local capable consomme 40 à 120x moins d'énergie par token qu'un frontier!!!

    Pour mon cas, utiliser les modèles locaux pour le "quotidien" (rédaction, brainsto, recherches) et garder les API frontier seulement pour le dev, c'est ❗une emprunte carbone divisée par 80❗

    + la confidentialité des données :))

    Note : la conso comparIA est un ordre de grandeur

    #data #AI #localLLMs

  5. NVIDIA's RTX Spark brings CUDA and 128GB unified memory to mainstream Windows PCs this fall. The move reshapes local-AI hardware decisions: the question shifts from 'buy the biggest card you can afford' to 'which constraint fails first—memory, bandwidth, or model quality.' implicator.ai/nvidias-rtx-spar #AI #Hardware #LocalLLMs

  6. Mein #arbeitgeber labert grade in so nem #MicrosoftTeams Call für alle Mitarbeiter was von #digitalesouveranitat und dann soll ich MEHR mit #microsoft #github #copilot machen. Und selbstverständlich wird ALLES #ai. Sogar unsere TLD wechselt von .net auf .ai.
    Wir sollen ganz explizit doch bitte #ki in die tägliche #Arbeit einbinden, der Vertrieb soll sich beim. Aber bloß "unsere" nutzen, wegen den Daten. Muss mich gleich mal informieren, ob #localLLMs erlaubt sind.
    hessen.social/@Moonstone2487/1

  7. Just a note for parallel universe me: flash attention is bad for #LocalLLMs

  8. So I have been trying the new #Gemma4 models on my M1 macbook pro, specifically the gemma4:26b which is 17gb in size.

    Obviously not the most challenging coding challenge and tasks but...

    Much much faster response times than local models 6-12 months ago. Previously qwen, deepseek, and even Gemm3 simply took too long to be practical.

    I find it incredible this can run on just my 5.5 year old laptop.

    #ai #llm #ollama #localllms #llms

  9. Just so we are clear: #LocalLLMs are an asset if trained and used well. But please be aware that many projects are pretending to be open source but their releases contain closed source components where it's not transparent what is going on.

    Go to the source. Llama.cpp, PyTorch, etc.

  10. If you are running #LocalLLMs you may be using LM Studio. Just a fair warning.... While this is practical, it's also proxying everything through their infrastructure. It's a privacy nightmare.

    #lmstudio

  11. Find out which AI models your machine can actually run.

    CanIRun.ai — Can your machine run AI models? canirun.ai/

    ht @researchbuzz.bsky.social

    #LocalLLMs

  12. #Eurollm #europeanai

    mastodon.social/@silentexcepti

    👍

    Is there any initiative to pool the various initiatives across European academia and public research to come up with a common #ecosystem of #europeandata and #LocalLLMs ?

  13. I spent some time playing around with Local LLMs on Apple Silicon. Here's how that played out.

    Local LLMs on Spare Apple Silicon: A Cautionary Tale
    macadminmusings.com/blog/2026/

  14. MiniMax M2 & Agent: Ingenious in Simplicity MiniMax M2 & Agent: Ingenious in Simplicity MiniMax M2 was released on Monday 27th October by MiniMax, a Chinese AI lab founded in December 2021....

    #ai #generative-ai #local-llms #llms #llm #llm-pricing #pelican-riding-a-bicycle #llm-release #ai-in-china #minimax

    Origin | Interest | Match
  15. 🚨 ALERT: Local LLMs, the supposed guardians of your digital fortress, are apparently about as secure as a wet paper bag. 🤦‍♂️ This "groundbreaking" #research reveals that local models are easily tricked, making them the cyber equivalent of a friendly puppy that wags its tail at everyone, including burglars. 🐶🔓
    quesma.com/blog/local-llms-sec #LocalLLMs #CyberSecurity #DigitalFortress #Vulnerability #AIModels #HackerNews #ngated

  16. 🚀 Take control of your AI usage! With LiteLLM + OpenWebUI you can unify cloud & local models, set real budgets, and never get surprise bills. Perfect for home labs and small teams. 🧑‍💻💡

    #LiteLLM #OpenWebUI #Docker #AItools #HomeLab #LocalLLMs #APIGateway #AIbudget #TechBlog #SmallBusinessAI

    victornava.dev/2025/09/02/lite