home.social

#edgeai — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #edgeai, aggregated by home.social.

fetched live
  1. Muse Glimmer puts a 30B open-weight agent on consumer hardware. Local execution keeps files off cloud APIs and shifts cost from tokens to hardware.
    huggingface.co/meta-models/Mus
    #EdgeAI #DigitalSovereignty

  2. Muse Glimmer puts a 30B open-weight agent on consumer hardware. Local execution keeps files off cloud APIs and shifts cost from tokens to hardware.
    huggingface.co/meta-models/Mus
    #EdgeAI #DigitalSovereignty

  3. Thanks to all the AI haters, many AI providers are cutting prices, making heavy AI users like me smile everyday.

    Besides, these AI haters are actually driving AI research and development out of the US into countries like China, India, and even Vietnam. They are also speeding up the development of much smaller #agenticAI and #edgeAI that can even be run on PCs and very soon all smart phones.

    AI will replace AI haters sooner, not later.

    #AI #LLM

  4. Thanks to all the AI haters, many AI providers are cutting prices, making heavy AI users like me smile everyday.

    Besides, these AI haters are actually driving AI research and development out of the US into countries like China, India, and even Vietnam. They are also speeding up the development of much smaller #agenticAI and #edgeAI that can even be run on PCs and very soon all smart phones.

    AI will replace AI haters sooner, not later.

    #AI #LLM

  5. 🤖 #LiquidAI released #LFM2_5 2.6B, an #agentic model that runs entirely on-device: planning, tool calling & multi-step tasks without any cloud API #AI #LLM #EdgeAI #opensource
    🧵👇

    ⚡ Decodes 220 tokens/s on an M5 Max CPU, 113 tokens/s on a Ryzen AI Max+ 395 and 30 tokens/s on a phone, staying under 2.5 GB memory

  6. 🤖 #LiquidAI released #LFM2_5 2.6B, an #agentic model that runs entirely on-device: planning, tool calling & multi-step tasks without any cloud API #AI #LLM #EdgeAI #opensource
    🧵👇

    ⚡ Decodes 220 tokens/s on an M5 Max CPU, 113 tokens/s on a Ryzen AI Max+ 395 and 30 tokens/s on a phone, staying under 2.5 GB memory

  7. Banana PI BPI-AI2N public sale
    The first #Renesas & #BananaPI open source SOM
    SoC: RZ/V2N
    Renesas #AI Accelerator: DRP-AI, supporting up to 15TOPS AI performance
    OS: #Yocto, #Armbian

    Check out the official docs & specs here:

    👉 docs.banana-pi.org/en/BPI-AI2N

    #EdgeAI #SingleBoardComputer #Maker #TechNews
    #SOM

  8. #Jetson Orin Nano #Ubuntu is very unfriendly with no good native apps and is relatively slow running most programs. So, it is only good for doing some small LLM inferencing and #edgeAI stuff like #LiDAR which I don't care cause I will be using #esp32 for that purpose.

    But the good news is, I can use Pi Apps which uses #Flatpak, allowing some of my favorite apps. like #Obsidian notes, to run on this Jetson board that runs on an ARM 64 version of Ubuntu.

    #AI #LLM

  9. #Jetson Orin Nano #Ubuntu is very unfriendly with no good native apps and is relatively slow running most programs. So, it is only good for doing some small LLM inferencing and #edgeAI stuff like #LiDAR which I don't care cause I will be using #esp32 for that purpose.

    But the good news is, I can use Pi Apps which uses #Flatpak, allowing some of my favorite apps. like #Obsidian notes, to run on this Jetson board that runs on an ARM 64 version of Ubuntu.

    #AI #LLM

  10. Every AI API call is a copy of your data leaving the building.

    The vivibit E-series keeps it in-house: an E1001 hub plus up to four NVIDIA DGX Spark nodes (1 PFLOPS each) — a private cluster running DeepSeek & Qwen on hardware you own.

    No tokens metered. No data shipped out. Just your models, on your rack.

    #AI #EdgeAI #SelfHosted #LLM #Privacy

  11. Introducing the vivibit E-series — AI compute you own, from desktop to cluster.

    Start with one E1001 node: 8-core ARMv9, 48GB LPDDR5, up to 366TB storage — runs 14B models locally at ≤100W.

    Scale to a full cluster (E1001 + up to 4 compute nodes) over built-in 50GbE, no data-center switch:
    ▪ up to 4 PFLOPS AI compute
    ▪ 512GB unified memory
    ▪ up to 366TB storage
    ▪ runs DeepSeek & Qwen on-prem

    Your models. Your data. Your rack.

    #AI #EdgeAI #SelfHosted #LLM