#qwen3 — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #qwen3, aggregated by home.social.
-
RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀
mehr auf Arint.info
#AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info
-
RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀
mehr auf Arint.info
#AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info
-
RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀
mehr auf Arint.info
#AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info
-
RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀
mehr auf Arint.info
#AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info
-
RT @imnotchalk: Qwen3.8 27B ist jetzt im @cerebras Shared Tier verfügbar mit einer Geschwindigkeit von ~1500 Token pro Sekunde. 🚀
mehr auf Arint.info
#AI #Cerebras #LLM #MachineLearning #Qwen3 #TechNews #arint_info
-
RT @UnslothAI: Wir veröffentlichen neue Qwen3.8-27B GGUFs mit 10 % höherer Genauigkeit. Unsloth Dynamic V3 schlägt andere um 10 % auf Div-300, KLD und weiteren Benchmarks. Zudem veröffentlichen wir 1-Bit-Quantisierungen, die 77 % der Genauigkeit beibehalten. Läuft auf 8 GB RAM. Blog: unsloth.ai/docs/basics/dynam… GGUF: huggingface.co/unsloth/Qwen3…
mehr auf Arint.info
#AI #GGUF #MachineLearning #OpenSource #Qwen3 #Unsloth #arint_info
-
🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
https://piszczek.pl/blog/qwen38-27b-256k-50-tps-24gb-gpu #breakthrough #AIexperiment #technews #HackerNews #ngated -
🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
https://piszczek.pl/blog/qwen38-27b-256k-50-tps-24gb-gpu #breakthrough #AIexperiment #technews #HackerNews #ngated -
🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
https://piszczek.pl/blog/qwen38-27b-256k-50-tps-24gb-gpu #breakthrough #AIexperiment #technews #HackerNews #ngated -
🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
https://piszczek.pl/blog/qwen38-27b-256k-50-tps-24gb-gpu #breakthrough #AIexperiment #technews #HackerNews #ngated -
🚀🎉 Wow, it's truly groundbreaking news that #Qwen3.8 can churn out a whole 37 tokens per second on a #GPU you'd need a #loan to buy! 💸 Apparently, this "experiment" showed that throwing every kitchen sink at a model doesn't exactly equal genius-level #AI. 🙄
https://piszczek.pl/blog/qwen38-27b-256k-50-tps-24gb-gpu #breakthrough #AIexperiment #technews #HackerNews #ngated -
Not a fan of Qwen 3.8 27B.
Overthinks too much => wasted tokens.
Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?
#AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI
-
Qwen3.8-Max: A New Bar for Coding and Cowork
https://qwen.ai/blog?id=qwen3.8
-
[Перевод] Как я объединил четыре Mac Studio в один компьютер с 1,5 ТБ – и запустил ИИ, который не влезает ни в одну видеокарту
Представляю вашему вниманию довольно необычный эксперимент: объединение четырех компьютеров Mac Studio в единый кластер с 1,5 ТБ общей памяти с помощью новой технологии RDMA через Thunderbolt 5. Это нужно, чтобы запускать локально гигантские языковые модели уровня DeepSeek V3 и Kimi K2 Thinking на 1 триллион параметров. Автор делится результатами тестов производительности, сравнивает систему с решениями от Nvidia и AMD, а также подробно рассказывает о трудностях настройки и проблемах стабильности, с которыми пришлось столкнуться в процессе.
https://habr.com/ru/companies/timeweb/articles/1061842/
#Mac_Studio #локальные_LLM #тензорный_параллелизм #Thunderbolt #DeepSeek #Kimi_K2 #Qwen3 #timeweb_статьи_перевод
-
Сжатие декодерных эмбеддеров: как ужать 8B до продакшена без потери recall
Декодерный эмбеддер 7–8B дает качество, но платит за него памятью, latency и деньгами. Разбираем все оси сжатия - int8, int4, binary + rescoring, PQ, MRL-усечение - на реальных замерах recall@10: где деградация мягкая, а где обрыв. С воспроизводимым кодом и Colab-ноутбуком под Qwen3
https://habr.com/ru/articles/1054930/
#сжатие_эмбеддингов #квантизация #эмбеддинги #embeddings #RAG #Qdrant #Qwen3 #binary_quantization #Matryoshka #retrieval
-
🎯 Supported models include #GPT-OSS-120B, #GPT-OSS-20B, #Llama4 Maverick, #Llama4 Scout, #Llama33-70B, #Llama31-8B, #KimiK2, #Qwen3-32B
🔧 Key features: deterministic inference for faster tool-using agents, cost-effective scaling, approved tool use with clear allowlists, seamless migration capability
📋 Ready-to-use cookbook tutorials with #BrowserBase #MCP, #BrowserUse #MCP, #Exa #MCP, #Firecrawl #MCP, #HuggingFace #MCP, #Parallel #MCP, #Stripe #MCP, #Tavily #MCP
-
Alibaba has launched Qwen 3, an advanced AI model with hybrid reasoning, aiming to rival DeepSeek and Baidu in China’s escalating AI race. The model balances quick responses with deeper, self-checking logic, targeting developers with a smarter, more flexible platform.
#Alibaba #Qwen3 #AIinChina #ArtificialIntelligence #Baidu #DeepSeek #HybridAI #TechNews #TECHi
Read Full Article Here :- https://www.techi.com/alibaba-debuts-sophisticated-qwen-3ai-escalting-tech-rivalry/