home.social

#qwen3 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #qwen3, aggregated by home.social.

  1. Not a fan of Qwen 3.8 27B.

    Overthinks too much => wasted tokens.

    Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?

    #AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI

  2. Not a fan of Qwen 3.8 27B.

    Overthinks too much => wasted tokens.

    Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?

    #AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI

  3. Not a fan of Qwen 3.8 27B.

    Overthinks too much => wasted tokens.

    Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?

    #AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI

  4. Not a fan of Qwen 3.8 27B.

    Overthinks too much => wasted tokens.

    Has someone ran it without too much thinking? Does it perform well with that knob down, or letting it think all the way is required to be good?

    #AI #ArtificialIntelligence #LLM #Qwen #Qwen38 #Qwen3 #Ollama #LlamaCCP #vLLM #LargeLanguageModels #LocalAI

  5. Qwen3-next đã đạt tốc độ xử lý 103 token/giây trên CPU Intel Xeon v4 (22 lõi, 64GB RAM) mà KHÔNG CẦN GPU! Đây là một bước tiến đáng kể, cho phép chạy các mô hình ngôn ngữ lớn (LLM) hiệu quả hơn trên phần cứng CPU thông thường, đặc biệt với kiến trúc MoE và biên dịch Llama tối ưu.
    #AI #LLM #CPU #Xeon #Qwen3 #NoGPU #Performance #TechNews #TríTuệNhânTạo #MôHìnhNgônNgữLớn #HiệuNăng #CôngNghệ

    reddit.com/r/ollama/comments/1