home.social

#mixtureofexperts — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #mixtureofexperts, aggregated by home.social.

fetched live
  1. 🍹 Introducing Liquid AI's latest cocktail: an "Even Better onDevice MixtureofExperts" with a splash of #buzzwords and a twist of jargon 🍹 Just what you needed! Now you can "unlock" your business potential by drowning it in a sea of acronyms and overpriced "solutions" that claim to be the world's "most efficient" 🙄 Cheers to that! 🎉
    liquid.ai/blog/lfm2-5-8b-a1b #LiquidAI #Cocktail #BusinessSolutions #MixtureofExperts #OverpricedSolutions #HackerNews #ngated

  2. Nemotron 3 Super pushes the frontier with 40 M supervised & alignment samples, leveraging a Mamba‑Transformer backbone and Mixture‑of‑Experts scaling. The model shows stronger agent reasoning, RL‑based fine‑tuning, and tighter AI alignment. Dive into the details to see how this LLM reshapes open‑source AI. #Nemotron3 #MixtureOfExperts #AIAlignment #SupervisedFineTuning

    🔗 aidailypost.com/news/nemotron-

  3. Alibaba just released the Qwen‑3.5‑Medium model as open‑source, delivering Sonnet 4.5‑level performance on a single GPU. It uses a Mixture‑of‑Experts architecture and a new “Thinking Mode” to boost AI inference efficiency while staying lightweight. Dive into the details and see how this could reshape open‑source LLM development. #Qwen3_5 #OpenSourceLLM #MixtureOfExperts #ModelEfficiency

    🔗 aidailypost.com/news/alibaba-o

  4. NVIDIA’s new co‑design with Sarvam AI slashes time‑to‑first‑token to under a second for LLM inference. By marrying Mixture‑of‑Experts models with GPU acceleration, they boost throughput while trimming latency. This hardware‑software synergy could reshape how we deploy large language models at scale. Read more to see the numbers and tech behind the breakthrough. #NVIDIA #SarvamAI #MixtureOfExperts #TTFT

    🔗 aidailypost.com/news/nvidia-co

  5. Alibaba's new Qwen 3.5 397B-A17 outperforms even larger rivals by using multi-token prediction and a sparse mixture-of-experts architecture. It cuts inference cost while keeping top-tier performance, hinting at a new era for multimodal AI. Curious how 397 billion parameters can be cheaper? Read the full story. #Qwen3_5 #AlibabaAI #MixtureOfExperts #MultiTokenPrediction

    🔗 aidailypost.com/news/alibabas-

  6. MiniMax's new M2.5 model slashes costs to 1/20 of Claude Opus while handling 30% of HQ tasks. Built on a Mixture‑of‑Experts sparse architecture, it delivers strong code‑generation and LLM performance—all open‑source. Discover how this AI agent could boost productivity in your projects. #MiniMaxM2_5 #MixtureOfExperts #OpenSourceAI #AIProductivity

    🔗 aidailypost.com/news/minimaxs-