home.social

#llamafile — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #llamafile, aggregated by home.social.

fetched live
  1. Finally (maybe) a usable local LLM that works on normal (and old) CPUs, not like a gimmick. This one on a 9th gen Intel, 16GB RAM with built-in GPUs (and with zram active). This specific Gemma model seems to be doing quite well in terms of speed, is not too dumb, and it apparently uses the Vulkan API out of the box. It also doesn't seem to spike the CPU at 100%, although it has crashed when the context was long. Web UI has improved. Still needs lots of testing though...

    #ai #llm #llamafile

  2. #Mozilla #AI Releases #Llamafile 0.10.4, their solution for easy-to-use #LLM as a single file that work across hardware and operating systems. With Llamafile 0.10.4 is now #Transcribefile, as a new piece built off their recently announced Transcribe.cpp project
    Llamafile 0.10.4 also updates against its Llama.cpp upstream build, brings a few improvements to its #Vulkan #API and #AMD #ROCm acceleration handling, HTTPS download support, and pledge/SECCOMP sandboxing support
    phoronix.com/news/Llamafile-0.

  3. What if the soon to be released @silex Desktop shipped a way to fine-tune your own small, open AI locally?

    A ~1B open model could be fine-tuned right inside Silex to learn how the user uses Silex and imitate via MCP. Runs on modest hardware, your data never leaves your machine

    The stack: #Luciole, the model by #OpenLLMFrance / @LINAGORA, #llamafile by @mozilla

    I'd be happy to find Silex users who want to co-build this, I have no idea how to make it happen 😅

    In? 👇

    #foss #LocalAI #llamafile

  4. (more Linux and FOSS news in previous posts of thread)

    Mozilla AI Releases Llamafile 0.10.4 With New Transcribefile Built On Transcribe.cpp:
    phoronix.com/news/Llamafile-0.

    Home Assistant Matter Server 9.1.0 adds support for Matter 1.6.0:
    alternativeto.net/news/2026/7/

    Zed Editor 1.11.3 Brings Dedicated Git Views, Sonnet 5 Support, and 60+ Bug Fixes:
    linuxcompatible.org/story/zed-

    Design tool Penpot brings background blur, better WebGL rendering & design token upgrades:
    alternativeto.net/news/2026/7/

    GitLab 19.2 brings Duo CLI, custom AI flows, and scheduled policies GA:
    alternativeto.net/news/2026/7/

    Gitea 1.27.0 patches major security flaws and enhances Actions:
    alternativeto.net/news/2026/7/

    Forgejo 16.0 adds SSRF protection, granular notifications, and refined review tools:
    alternativeto.net/news/2026/7/

    PostgreSQL 19 Beta 2 Released: Native Graph Queries, Unified REPACK, and Feature Freeze:
    linuxcompatible.org/story/post

    Rust 1.97.1 Drops Exactly One Week After 1.97.0 to Fix Critical Miscompilation:
    linuxcompatible.org/story/rust

    Python 3.15 Beta 4 Ships With Built-in frozendict, Lazy Imports, and UTF-8 Defaults:
    linuxcompatible.org/story/pyth

    OpenClaw v2026.7.1 released with Control UI overhaul, easier setup, new GPT, Codex, Tencent, Claude, etc. models available:
    docs2.openclaw.ai/releases/202

    FosseryWeb and page-builders update:
    I added article page generation support to page-builders, and ordered list, strikethrough text conversion support to its converter component. Also regenerated the existing article pages, to fix inconsistencies in the code, plus included the title and date of the article in the downloadable ODT and Markdown versions.
    fosseryweb.codeberg.page/artic
    fosseryweb-min.codeberg.page/a
    codeberg.org/fosseryweb/page-b

    (more FOSS news in comment)

    #WeeklyNews #OpenSource #FOSSNews #OpenSourceNews #FOSS #News #Llamafile #AI #HomeAssistant #Zed #Penpot #GitLab #Gitea #Forgejo #PostgreSQL #Rust #Python #OpenClaw #FosseryWeb #FosseryTech

  5. I bet that in 2030 we'll be using claude or gpt only to setup local models with llamafile, optimized for specific tasks on a specific laptop or phone

    > The only thing Internet Explorer is good for is downloading Firefox

    @mozilla #llamafile is a #foss alternative to VC funded #ollama

    #ai #llm #slm #fantasticfour #enshitification #OpenSource

  6. Neste vídeo, mostro como criar uma IA local portátil no Linux utilizando o Llamafile e um modelo no formato GGUF, armazenados em um pendrive ou SSD externo.

    Link: youtu.be/UELV_vluOPQ

    #llamafile #ibmgranite #ialocal #tech #ferramentaia #iagratis #InovacaoTech #IAPortatil #linuxmint #linux

  7. CW: #AI #agents

    Searching the Web with agents: aittalam.github.io/posts/2026-

    Or: build your own search agent with Pi, , and

    I think open, (relatively) small, custom agents can be quite good at searching. To me it feels as mindblowing as AI-assisted coding, but (1) potentially useful to more people, (2) running on cheaper compute, (3) easier to verify (4) providing you with more knowledge than what you started with, instead of atrophying the one you had. (Examples available in Part 2)

  8. Encántame este vídeo de Mancomún sobre Llamafile! Aprende a executar modelos de IA de xeito sinxelo no teu propio ordenador, con explicacións claras e ligazón ao artigo. Perfecto para quen quere probar IA local e software libre. Anímate a experimentar! #Llamafile #IA #InteligenciaArtificial #SoftwareLibre #Mancomun #Galego #PeerTube
    v.eurorede.com/videos/watch/22

  9. (Linux news in previous posts)

    FOSS NEWS

    Mozilla is packing new features into Firefox:
    betanews.com/article/mozilla-i

    Is Firefox getting a new logo? Mozilla’s socials suggest so…:
    omgubuntu.co.uk/2026/03/is-fir

    Signal introduces group member labels, allowing you to specify your role or responsibility:
    alternativeto.net/news/2026/3/

    Blender 5.1 Open-Source 3D Graphics Software Released with Many New Features:
    9to5linux.com/blender-5-1-open

    LibreOffice 26.8 To Add A Donation Banner To Its Start Center:
    phoronix.com/news/LibreOffice-

    Mozilla Releases Llamafile 0.10 To Enhance Their AI Offering For Easy-To-Use LLMs:
    phoronix.com/news/Mozilla-AI-L

    Ollama integration brings all models to OpenClaw platform:
    alternativeto.net/news/2026/3/

    OpenShot 3.5 is (yet again) the ‘biggest’ and ‘fastest’ release ever:
    omgubuntu.co.uk/2026/03/opensh

    Bitwarden Send improves security with email verification for paid users:
    alternativeto.net/news/2026/3/

    SuperTux 0.7.0 Arcade Game Is Out with Complete Level Design, Revamped Graphics:
    9to5linux.com/supertux-0-7-0-a

    PlayStation 3 emulator RPCS3 gets easier to use with Steam:
    gamingonlinux.com/2026/03/play

    OpenTTD devs clarify store changes with Transport Tycoon Deluxe re-release as Atari contribute server funding:
    gamingonlinux.com/2026/03/open

    GNUnet 0.27 Released For Those With "Some Reasonable Pain Tolerance":
    phoronix.com/news/GNUnet-0.27-

    Immich 2.6 improves map side panel, asset viewer, shared link slugs & presets, and more:
    alternativeto.net/news/2026/3/

    (more FOSS news in comments)

    #WeeklyNews #OpenSource #FOSSNews #FOSS #OpenSourceNews #News #Firefox #Signal #Blender #LibreOffice #Llamafile #Ollama #OpenShot #Bitwarden #BitwardenSend #SuperTux #RPCS3 #OpenTTD #GNUnet #Immich #Browser #WebBrowser #VideoEditor #VideoEditing #ContentCreation #Gaming #FOSSGame #FOSSGaming #OpenSourceGame #PlayStation #PlayStation3 #FosseryTech

  10. #Mozilla Releases #Llamafile 0.10 To Enhance Their #AI Offering For Easy-To-Use LLMs
    Llamafile is a Mozilla.ai project to distribute and run large language models as a single file. With a single Llamafile you can run the #LLM across platforms and with varying hardware support. Their intentions are on making LLMs more accessible and convenient to both developers and end-users.
    phoronix.com/news/Mozilla-AI-L

  11. Mozilla Releases Llamafile 0.10 To Enhance Their AI Offering For Easy-To-Use LLMs - Phoronix

    「 Llamafile is a Mozilla.ai project to distribute and run large language models as a single file. With a single Llamafile you can run the LLM across platforms and with varying hardware support. Their intentions are on making LLMs more accessible and convenient to both developers and end-users 」

    phoronix.com/news/Mozilla-AI-L

    #ai #llamafile #llm #mozilla

  12. Mozilla rilancia Llamafile con la versione 0.10: supporto immagini, GPU Metal e CUDA, Whisper e Stable Diffusion integrati. Un passo avanti per rendere i modelli linguistici più accessibili. #Mozilla #Llamafile #Linux #OpenSource

    linuxeasy.org/mozilla-rilascia

  13. 🦙 llamafile returns!

    Mozilla.ai is adopting llamafile to push forward open, local, privacy-first AI.

    We’re refreshing the codebase and rebuilding the roadmap with feedback from the community.

    Share your ideas on GitHub, Discord, or Hacker News—and help shape the next phase of llamafile.

    🔗 blog.mozilla.ai/llamafile-retu

    #MozillaAI #llamafile #opensource #LocalAI #AIcommunity

  14. CW: AI, LLMs

    blog.mozilla.ai/llamafile-retu

    has been instrumental in my first projects relying on language models (I purposely left “large” aside, as the one I used for is tiny).

    While looking for support for more recent models I tried other applications, but I never felt “at home” with them as I did with llamafile. This is why I am so happy of this move: breathing new life into it is, for me, giving ppl a new chance to better understand (what ppl now call) “#AI” while tinkering with it.

  15. At DjangoCon US 2025 in Chicago, more than one person shared the workflow of dictating their articles or slide notes to a template using mobile apps 🎙️

    I was experimenting with huggingface.co/Mozilla/whisper, and it seems to work well on my PC 🔴

    It occurred to me that it could be used to add live captioning to meetups or small conferences that can't afford live captioners as good as the one we had at DjangoCon 💡

    Have any of you done any experiments?

  16. @naturzukunft

    There aren't that many choices to manage exploding information complexity so something must happen.

    Even if it is not the semantic web as we knew it, it must solve similar problems (and hence look similar - as form follows function).

    Removing frictions is key imho. See e.g. how using complex LLM models is now down to just one #llamafile

    Semantic web tooling (at least the public / open source part) never reached a sweetspot for broad adoption, always somewhat aloof

    @rzeta0

  17. In , we use to calculate *sentence embeddings*.
    If you don’t know what embeddings are, just think about them as numerical descriptors of your Mastodon statuses, which are closer (as in two cities’ coordinates being close on a map) the more semantically similar their respective statuses are. We’ll get back later to this with a more visual description. If you are interested in embeddings and wanna delve deeper, see vickiboykis.com/what_are_embed by @vicki.

  18. is a Mozilla tool that packages a language model in a single executable file that will run on most platforms. It is 100% local and has been optimized to run on slower hardware, from Raspberry Pis to my 8yo laptop. It is based on llama.cpp which supports a plethora of models, not just LLMs: I chose all-minilm because it’s tiny (50MB) and has open code, research papers, and datasets. github.com/Mozilla-Ocho/llamaf huggingface.co/sentence-transf

  19. Второе пришествие мейнфреймов. Всё больше компаний хотят запускать ИИ у себя в офисе

    Мейнфрейм IBM z16 во время лабораторных тестов в 2022 г, источник Приложения ИИ находят применение в бизнесе. Но есть проблема: корпоративные данные и документация представляют коммерческую тайну. Их нельзя передавать на сторону, тем более в облачную систему машинного обучения. Кроме того, что сама передача небезопасна, так ещё и публичная модель будет обучаться на наших секретах , а потом помогать конкурентам. В общем, у коммерческих компаний остаётся один вариант: поднимать собственный сервер или вычислительный кластер с ИИ. Таким образом, из эпохи облачных вычислений мы возвращаемся к старому доброму самохостингу, только сейчас это самохостинг GPU , серверы и мейнфреймы.

    habr.com/ru/companies/ruvds/ar

    #IBM #мейнфреймы #Telum_II #AnythingLLM #ChatGPT #Meta #обман #человечество #Mount_Diablo #IBM_z16 #самохостинг #ЦОД #датацентр #WatsonX #Copilot #Twinny #llamafile #Ollama #GPT4All #FraudGPT #WormGPT #Khoj #LocalAI #ruvds_статьи

  20. Llamafile by Mozilla

    Hi #fedifriends

    has any of you experience with #llamafile by #Mozilla?
    I would like to use a local #LLM #AI to get some help to create #XLST styles to run against #XLS and #CSV files with the goal of create #XML files without being an #XSLT champion...

    I would like to do it locally, not depending stricly on #Linux and #Nvidia, with some #ethical software and looks like Llamafile check marks all those voices.

    Thanks... 🙏

    #fedihelp #fediask

  21. Experimenting with , mastodon API, and for my "Build Your Own Timeline Algorithm" project. 100% local, except those 11 seconds it took me to download my home timeline.

    I have been pitching this idea around for almost two years now, and while I felt bad for not pushing more for it, I am glad that in the meantime new cool tech became available allowing me to do something even better than I imagined.

  22. I also put some Llama 3.2 llamafiles onto the Docker hub: hub.docker.com/?namespace=mara . Quite nice for CI. #llm #llamafile

  23. If you're into #llm stuff, I've created a small repository with scripts to create a #llamafile from a GGUF model.

    In plain English, it makes it easy to download a language model and turn it into a file you can run directly. Super fun, super easy. github.com/maragudk/llamafile