home.social

#linearattention — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #linearattention, aggregated by home.social.

fetched live
  1. Kimi Linear is a really interesting way to circumvent the KV cache bottleneck for long-horizon tasks. Basically, it's a hybrid model that blends Linear Attention, State Space Modelling and MLA with NoPE encoding to learn a constant state per head instead of growing a cache.

    The gating and rank-1 delta-rule correction ensure the memory doesn't accumulate forever and we can have both weight decay and correction

    #Kimi #MoonshotAI #TransformerResearch #LinearAttention

  2. Kimi Linear is a really interesting way to circumvent the KV cache bottleneck for long-horizon tasks. Basically, it's a hybrid model that blends Linear Attention, State Space Modelling and MLA with NoPE encoding to learn a constant state per head instead of growing a cache.

    The gating and rank-1 delta-rule correction ensure the memory doesn't accumulate forever and we can have both weight decay and correction

    #Kimi #MoonshotAI #TransformerResearch #LinearAttention

  3. 🔥 Alibaba Qwen3-Next: 10x effizienter, 90% weeniger Trainingskosten!

    ▶️ Entdecke Hybrid-MoE nun
    ▶️ Aktiviere 262K Kontext!
    ▶️ Starte SGLang Turbo nun

    #ai #ki #artificialintelligence #qwen3next #alibaba #largelanguagemodels #mixtureofexperts #linearattention

    🔥 Jetzt KLICKEN & KOMMENTIEREN! 💭

    kinews24.de/qwen3-next-alibaba

  4. 🔥 Alibaba Qwen3-Next: 10x effizienter, 90% weeniger Trainingskosten!

    ▶️ Entdecke Hybrid-MoE nun
    ▶️ Aktiviere 262K Kontext!
    ▶️ Starte SGLang Turbo nun

    #ai #ki #artificialintelligence #qwen3next #alibaba #largelanguagemodels #mixtureofexperts #linearattention

    🔥 Jetzt KLICKEN & KOMMENTIEREN! 💭

    kinews24.de/qwen3-next-alibaba