home.social

#qwen3 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #qwen3, aggregated by home.social.

  1. RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀

    mehr auf Arint.info

    #AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info

    https://x.com/imnotchalk/status/2095637567979114654

  2. RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀

    mehr auf Arint.info

    #AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info

    https://x.com/imnotchalk/status/2095637567979114654

  3. RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀

    mehr auf Arint.info

    #AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info

    https://x.com/imnotchalk/status/2095637567979114654

  4. RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀

    mehr auf Arint.info

    #AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info

    https://x.com/imnotchalk/status/2095637567979114654

  5. RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀

    mehr auf Arint.info

    #AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info

    https://x.com/imnotchalk/status/2095637567979114654

  6. RT @UnslothAI: Wir veröffentlichen neue Qwen3.8-27B GGUFs mit 10 % höherer Genauigkeit. Unsloth Dynamic V3 schlägt andere um 10 % auf Div-300, KLD und weiteren Benchmarks. Zudem veröffentlichen wir 1-Bit-Quantisierungen, die 77 % der Genauigkeit beibehalten. Läuft auf 8 GB RAM. Blog: unsloth.ai/docs/basics/dynam… GGUF: huggingface.co/unsloth/Qwen3…

    mehr auf Arint.info

    #AI #GGUF #MachineLearning #OpenSource #Qwen3 #Unsloth #arint_info

    https://x.com/UnslothAI/status/2090103470015828184

  7. 🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
    piszczek.pl/blog/qwen38-27b-25 #breakthrough #AIexperiment #technews #HackerNews #ngated

  8. 🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
    piszczek.pl/blog/qwen38-27b-25 #breakthrough #AIexperiment #technews #HackerNews #ngated

  9. 🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
    piszczek.pl/blog/qwen38-27b-25 #breakthrough #AIexperiment #technews #HackerNews #ngated

  10. 🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
    piszczek.pl/blog/qwen38-27b-25 #breakthrough #AIexperiment #technews #HackerNews #ngated

  11. 🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
    piszczek.pl/blog/qwen38-27b-25 #breakthrough #AIexperiment #technews #HackerNews #ngated

  12. Not a fan of Qwen 3.8 27B.

    Overthinks too much => wasted tokens.

    Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?

    #AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI

  13. [Перевод] Как я объединил четыре Mac Studio в один компьютер с 1,5 ТБ – и запустил ИИ, который не влезает ни в одну видеокарту

    Представляю вашему вниманию довольно необычный эксперимент: объединение четырех компьютеров Mac Studio в единый кластер с 1,5 ТБ общей памяти с помощью новой технологии RDMA через Thunderbolt 5. Это нужно, чтобы запускать локально гигантские языковые модели уровня DeepSeek V3 и Kimi K2 Thinking на 1 триллион параметров. Автор делится результатами тестов производительности, сравнивает систему с решениями от Nvidia и AMD, а также подробно рассказывает о трудностях настройки и проблемах стабильности, с которыми пришлось столкнуться в процессе.

    habr.com/ru/companies/timeweb/

    #Mac_Studio #локальные_LLM #тензорный_параллелизм #Thunderbolt #DeepSeek #Kimi_K2 #Qwen3 #timeweb_статьи_перевод

  14. Сжатие декодерных эмбеддеров: как ужать 8B до продакшена без потери recall

    Декодерный эмбеддер 7–8B дает качество, но платит за него памятью, latency и деньгами. Разбираем все оси сжатия - int8, int4, binary + rescoring, PQ, MRL-усечение - на реальных замерах recall@10: где деградация мягкая, а где обрыв. С воспроизводимым кодом и Colab-ноутбуком под Qwen3

    habr.com/ru/articles/1054930/

    #сжатие_эмбеддингов #квантизация #эмбеддинги #embeddings #RAG #Qdrant #Qwen3 #binary_quantization #Matryoshka #retrieval

  15. Qwen 3b lokal auf meinem Laptop... Ihr könnt alle einpacken, ihr Internet-basierten KI's!

    #KI #qwen3 #lokal #computer #AI

  16. 🎯 Supported models include #GPT-OSS-120B, #GPT-OSS-20B, #Llama4 Maverick, #Llama4 Scout, #Llama33-70B, #Llama31-8B, #KimiK2, #Qwen3-32B

    🔧 Key features: deterministic inference for faster tool-using agents, cost-effective scaling, approved tool use with clear allowlists, seamless migration capability

    📋 Ready-to-use cookbook tutorials with #BrowserBase #MCP, #BrowserUse #MCP, #Exa #MCP, #Firecrawl #MCP, #HuggingFace #MCP, #Parallel #MCP, #Stripe #MCP, #Tavily #MCP

  17. Alibaba has launched Qwen 3, an advanced AI model with hybrid reasoning, aiming to rival DeepSeek and Baidu in China’s escalating AI race. The model balances quick responses with deeper, self-checking logic, targeting developers with a smarter, more flexible platform.

    #Alibaba #Qwen3 #AIinChina #ArtificialIntelligence #Baidu #DeepSeek #HybridAI #TechNews #TECHi

    Read Full Article Here :- techi.com/alibaba-debuts-sophi