home.social

#qwen2 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #qwen2, aggregated by home.social.

fetched live
  1. The #OpenCode is working. Check when the Nvidia is used for it
    ⎿  [Image #1]

    ● I'll check how #Ollama uses the #Nvidia GPU.

    ● Bash(nvidia-smi --query-gpu=name,memory.used,memory.total,utilization.gpu --format=csv,noheader 2>/dev/null)
    ⎿  NVIDIA GeForce RTX 2060, 952 MiB, 6144 MiB, 9 %

    ● Bash(ollama ps 2>/dev/null)
    ⎿  NAME ID SIZE PROCESSOR UNTIL
    qwen2.5-coder:7b-16k 0533dd01f237 7.2 GB 30%/70% CPU/GPU 4 minutes from now

    ● The model is currently loaded and actively using the GPU. Here's the breakdown:

    - Model: #qwen2.5-coder:7b-16k — 7.2 GB total
    - Split: 30% CPU / 70% GPU — the model is split across RAM and VRAM because it doesn't fit entirely in the RTX 2060's 6 GB VRAM
    - Current #GPU #VRAM usage: 952 MiB used of 6144 MiB (the rest of the model layers sit in system RAM)
    - GPU compute utilization: 9% right now (idle between requests)

    The GPU is used whenever OpenCode sends a prompt to Ollama — during inference, 70% of the model layers run on the #RTX2060 and 30% on the #CPU. The model stays loaded in VRAM for 4 more minutes of idle time before Ollama unloads it.

    #LocalLLM

  2. The #OpenCode is working. Check when the Nvidia is used for it
    ⎿  [Image #1]

    ● I'll check how #Ollama uses the #Nvidia GPU.

    ● Bash(nvidia-smi --query-gpu=name,memory.used,memory.total,utilization.gpu --format=csv,noheader 2>/dev/null)
    ⎿  NVIDIA GeForce RTX 2060, 952 MiB, 6144 MiB, 9 %

    ● Bash(ollama ps 2>/dev/null)
    ⎿  NAME ID SIZE PROCESSOR UNTIL
    qwen2.5-coder:7b-16k 0533dd01f237 7.2 GB 30%/70% CPU/GPU 4 minutes from now

    ● The model is currently loaded and actively using the GPU. Here's the breakdown:

    - Model: #qwen2.5-coder:7b-16k — 7.2 GB total
    - Split: 30% CPU / 70% GPU — the model is split across RAM and VRAM because it doesn't fit entirely in the RTX 2060's 6 GB VRAM
    - Current #GPU #VRAM usage: 952 MiB used of 6144 MiB (the rest of the model layers sit in system RAM)
    - GPU compute utilization: 9% right now (idle between requests)

    The GPU is used whenever OpenCode sends a prompt to Ollama — during inference, 70% of the model layers run on the #RTX2060 and 30% on the #CPU. The model stays loaded in VRAM for 4 more minutes of idle time before Ollama unloads it.

    #LocalLLM

  3. The #OpenCode is working. Check when the Nvidia is used for it
    ⎿  [Image #1]

    ● I'll check how #Ollama uses the #Nvidia GPU.

    ● Bash(nvidia-smi --query-gpu=name,memory.used,memory.total,utilization.gpu --format=csv,noheader 2>/dev/null)
    ⎿  NVIDIA GeForce RTX 2060, 952 MiB, 6144 MiB, 9 %

    ● Bash(ollama ps 2>/dev/null)
    ⎿  NAME ID SIZE PROCESSOR UNTIL
    qwen2.5-coder:7b-16k 0533dd01f237 7.2 GB 30%/70% CPU/GPU 4 minutes from now

    ● The model is currently loaded and actively using the GPU. Here's the breakdown:

    - Model: #qwen2.5-coder:7b-16k — 7.2 GB total
    - Split: 30% CPU / 70% GPU — the model is split across RAM and VRAM because it doesn't fit entirely in the RTX 2060's 6 GB VRAM
    - Current #GPU #VRAM usage: 952 MiB used of 6144 MiB (the rest of the model layers sit in system RAM)
    - GPU compute utilization: 9% right now (idle between requests)

    The GPU is used whenever OpenCode sends a prompt to Ollama — during inference, 70% of the model layers run on the #RTX2060 and 30% on the #CPU. The model stays loaded in VRAM for 4 more minutes of idle time before Ollama unloads it.

    #LocalLLM

  4. The #OpenCode is working. Check when the Nvidia is used for it
    ⎿  [Image #1]

    ● I'll check how #Ollama uses the #Nvidia GPU.

    ● Bash(nvidia-smi --query-gpu=name,memory.used,memory.total,utilization.gpu --format=csv,noheader 2>/dev/null)
    ⎿  NVIDIA GeForce RTX 2060, 952 MiB, 6144 MiB, 9 %

    ● Bash(ollama ps 2>/dev/null)
    ⎿  NAME ID SIZE PROCESSOR UNTIL
    qwen2.5-coder:7b-16k 0533dd01f237 7.2 GB 30%/70% CPU/GPU 4 minutes from now

    ● The model is currently loaded and actively using the GPU. Here's the breakdown:

    - Model: #qwen2.5-coder:7b-16k — 7.2 GB total
    - Split: 30% CPU / 70% GPU — the model is split across RAM and VRAM because it doesn't fit entirely in the RTX 2060's 6 GB VRAM
    - Current #GPU #VRAM usage: 952 MiB used of 6144 MiB (the rest of the model layers sit in system RAM)
    - GPU compute utilization: 9% right now (idle between requests)

    The GPU is used whenever OpenCode sends a prompt to Ollama — during inference, 70% of the model layers run on the #RTX2060 and 30% on the #CPU. The model stays loaded in VRAM for 4 more minutes of idle time before Ollama unloads it.

    #LocalLLM

  5. The #OpenCode is working. Check when the Nvidia is used for it
    ⎿  [Image #1]

    ● I'll check how #Ollama uses the #Nvidia GPU.

    ● Bash(nvidia-smi --query-gpu=name,memory.used,memory.total,utilization.gpu --format=csv,noheader 2>/dev/null)
    ⎿  NVIDIA GeForce RTX 2060, 952 MiB, 6144 MiB, 9 %

    ● Bash(ollama ps 2>/dev/null)
    ⎿  NAME ID SIZE PROCESSOR UNTIL
    qwen2.5-coder:7b-16k 0533dd01f237 7.2 GB 30%/70% CPU/GPU 4 minutes from now

    ● The model is currently loaded and actively using the GPU. Here's the breakdown:

    - Model: #qwen2.5-coder:7b-16k — 7.2 GB total
    - Split: 30% CPU / 70% GPU — the model is split across RAM and VRAM because it doesn't fit entirely in the RTX 2060's 6 GB VRAM
    - Current #GPU #VRAM usage: 952 MiB used of 6144 MiB (the rest of the model layers sit in system RAM)
    - GPU compute utilization: 9% right now (idle between requests)

    The GPU is used whenever OpenCode sends a prompt to Ollama — during inference, 70% of the model layers run on the #RTX2060 and 30% on the #CPU. The model stays loaded in VRAM for 4 more minutes of idle time before Ollama unloads it.

    #LocalLLM

  6. RT @HuggingModels: Lernen Sie Qwen2-32B-N64-Decomp kennen, eine leistungsstarke konversationelle KI, die ab sofort im GGUF-Format verfügbar ist. Dieses Modell bringt Dialogfunktionen auf Enterprise-Niveau auf lokale Maschinen und ermöglicht es Ihnen, anspruchsvolle KI-Chats ohne Cloud-Abhängigkeiten zu führen. Perfekt für Entwickler, die volle Kontrolle wünschen.

    mehr auf Arint.info

    #AI #GGUF #LLM #LocalAI #MachineLearning #Qwen2 #arint_info

    https://x.com/HuggingModels/status/2043963227521069367#m

  7. RT @HuggingModels: Lernen Sie Qwen2-32B-N64-Decomp kennen, eine leistungsstarke konversationelle KI, die ab sofort im GGUF-Format verfügbar ist. Dieses Modell bringt Dialogfunktionen auf Enterprise-Niveau auf lokale Maschinen und ermöglicht es Ihnen, anspruchsvolle KI-Chats ohne Cloud-Abhängigkeiten zu führen. Perfekt für Entwickler, die volle Kontrolle wünschen.

    mehr auf Arint.info

    #AI #GGUF #LLM #LocalAI #MachineLearning #Qwen2 #arint_info

    https://x.com/HuggingModels/status/2043963227521069367#m

  8. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  9. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  10. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  11. Vấn đề với hệ thống chat RAG: Qwen2.5 bỏ qua ngữ cảnh cuộc trò chuyện trước và trả lời không liên quan cho các câu hỏi tiếp theo. Người dùng gặp khó khăn khi mô hình chỉ dựa vào truy vấn mới nhất thay vì sử dụng lịch sử chat.

    #RAG #AI #Qwen2.5 #Chatbot #LỗiKỹThuật

    reddit.com/r/ollama/comments/1

  12. Hướng dẫn tinh chỉnh mô hình Qwen2.5-Coder-1.5B cho phân tích cảm xúc tiếng Trung. Có thể chạy trên Google Colab miễn phí trong 20-30 phút. Độ chính xác tăng từ 91,6% lên 97,8%. #AI #MachineLearning #Qwen2.5 #PhânTíchCảmXúc #GoogleColab #TinhChỉnhMôHình #TríTuệNhânTạo #HọcMáy

    i.redd.it/7xx856mftfzf1.png

  13. Ch peque nhá! Tôi vừa chuyển sang dùng Qwen2.5 Code Instruct bản tự-host thành công! M المقابل với Claude đầu tiên (lần nào 1h phải chờ), Qwen2.5 có thể xử lý comuni code, debug, và nhiếp ý nhanh lùi ởстром đường công việc. Ưbrochen ở máy MBook Pro 48GB và PC 2x RTX 5060TI 16GB (không cần quantize). Cài đặt đơn giản, chất lượng tốt cho công việc lẻ lậu.
    Tham khảo GitHub: @reliableJARED/qwen_coder
    Tags: #AI #Qwen2.5 #CodeAssistant #LocalTech #MáyTínhLâu
    #TechTips #OfflineAI #DevelopersCommu

  14. 🧠 #ByteDance ha rilasciato UI-TARS-1.5, un agente multimodale basato su #Qwen2.5-VL-7B che unisce visione e linguaggio con "reasoning". 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 

    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomar 

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  15. 🧠 #ByteDance ha rilasciato UI-TARS-1.5, un agente multimodale basato su #Qwen2.5-VL-7B che unisce visione e linguaggio con "reasoning". 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 

    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomar 

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  16. 🧠 #ByteDance ha rilasciato UI-TARS-1.5, un agente multimodale basato su #Qwen2.5-VL-7B che unisce visione e linguaggio con "reasoning". 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 

    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomar 

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  17. 🧠 #ByteDance ha rilasciato UI-TARS-1.5, un agente multimodale basato su #Qwen2.5-VL-7B che unisce visione e linguaggio con "reasoning". 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 

    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomar 

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  18. 🧠 #ByteDance ha rilasciato UI-TARS-1.5, un agente multimodale basato su #Qwen2.5-VL-7B che unisce visione e linguaggio con "reasoning". 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 

    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomar 

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  19. ​Qwen2.5-VL & QVQ-Max: Neue Maßstäbe in der visuellen KI

    Fortschrittliche Bild- und Videoanalyse
    Präzise Objekterkennung
    Verbesserte Dokumentenverarbeitung

    #ai #ki #artificialintelligence #kuenstlicheintelligenz #Qwen2.5-VL #QVQ-Max

    Jetzt lesen und folgen!

    kinews24.de/qwen2-5-vl-qvq-max/

  20. ​Qwen2.5-VL & QVQ-Max: Neue Maßstäbe in der visuellen KI

    Fortschrittliche Bild- und Videoanalyse
    Präzise Objekterkennung
    Verbesserte Dokumentenverarbeitung

    #ai #ki #artificialintelligence #kuenstlicheintelligenz #Qwen2.5-VL #QVQ-Max

    Jetzt lesen und folgen!

    kinews24.de/qwen2-5-vl-qvq-max/

  21. ​Qwen2.5-VL & QVQ-Max: Neue Maßstäbe in der visuellen KI

    Fortschrittliche Bild- und Videoanalyse
    Präzise Objekterkennung
    Verbesserte Dokumentenverarbeitung

    #ai #ki #artificialintelligence #kuenstlicheintelligenz #Qwen2.5-VL #QVQ-Max

    Jetzt lesen und folgen!

    kinews24.de/qwen2-5-vl-qvq-max/

  22. ​Qwen2.5-VL & QVQ-Max: Neue Maßstäbe in der visuellen KI

    Fortschrittliche Bild- und Videoanalyse
    Präzise Objekterkennung
    Verbesserte Dokumentenverarbeitung

    #ai #ki #artificialintelligence #kuenstlicheintelligenz #Qwen2.5-VL #QVQ-Max

    Jetzt lesen und folgen!

    kinews24.de/qwen2-5-vl-qvq-max/

  23. ​Qwen2.5-VL & QVQ-Max: Neue Maßstäbe in der visuellen KI

    Fortschrittliche Bild- und Videoanalyse
    Präzise Objekterkennung
    Verbesserte Dokumentenverarbeitung

    #ai #ki #artificialintelligence #kuenstlicheintelligenz #Qwen2.5-VL #QVQ-Max

    Jetzt lesen und folgen!

    kinews24.de/qwen2-5-vl-qvq-max/

  24. Alibaba Cloud shakes up the AI scene with **Qwen2.5-Omni-7B!** This cutting-edge multimodal model processes text, images, audio, and video, making it perfect for mobile devices. It's designed for cost-effective AI agents, especially in voice applications for the visually impaired. With a hefty **$53 billion** investment in AI and cloud infrastructure, Alibaba is positioning itself for success in the booming AI market—don’t miss the full story. [Read more](cnbc.com/2025/03/27/alibaba-la) #ArtificialIntelligence #AlibabaCloud #Qwen2 #TechInnovation

  25. Alibaba Cloud shakes up the AI scene with **Qwen2.5-Omni-7B!** This cutting-edge multimodal model processes text, images, audio, and video, making it perfect for mobile devices. It's designed for cost-effective AI agents, especially in voice applications for the visually impaired. With a hefty **$53 billion** investment in AI and cloud infrastructure, Alibaba is positioning itself for success in the booming AI market—don’t miss the full story. [Read more](cnbc.com/2025/03/27/alibaba-la) #ArtificialIntelligence #AlibabaCloud #Qwen2 #TechInnovation

  26. Alibaba Cloud shakes up the AI scene with **Qwen2.5-Omni-7B!** This cutting-edge multimodal model processes text, images, audio, and video, making it perfect for mobile devices. It's designed for cost-effective AI agents, especially in voice applications for the visually impaired. With a hefty **$53 billion** investment in AI and cloud infrastructure, Alibaba is positioning itself for success in the booming AI market—don’t miss the full story. [Read more](cnbc.com/2025/03/27/alibaba-la) #ArtificialIntelligence #AlibabaCloud #Qwen2 #TechInnovation

  27. Alibaba Cloud shakes up the AI scene with **Qwen2.5-Omni-7B!** This cutting-edge multimodal model processes text, images, audio, and video, making it perfect for mobile devices. It's designed for cost-effective AI agents, especially in voice applications for the visually impaired. With a hefty **$53 billion** investment in AI and cloud infrastructure, Alibaba is positioning itself for success in the booming AI market—don’t miss the full story. [Read more](cnbc.com/2025/03/27/alibaba-la) #ArtificialIntelligence #AlibabaCloud #Qwen2 #TechInnovation

  28. Qwen2.5-VL-32B: because nothing says "cutting-edge" like moaning about parameter scales and reinforcement learning 🙄. Apparently, this 32B thing is "smarter" and "lighter" – sounds like a diet ad for AI models. 😂🍩 #Innovation!
    qwenlm.github.io/blog/qwen2.5- #Qwen2.5VL32B #AIModels #ReinforcementLearning #CuttingEdge #TechHumor #HackerNews #ngated

  29. Qwen2.5-VL-32B: because nothing says "cutting-edge" like moaning about parameter scales and reinforcement learning 🙄. Apparently, this 32B thing is "smarter" and "lighter" – sounds like a diet ad for AI models. 😂🍩 #Innovation!
    qwenlm.github.io/blog/qwen2.5- #Qwen2.5VL32B #AIModels #ReinforcementLearning #CuttingEdge #TechHumor #HackerNews #ngated

  30. Qwen2.5-VL-32B: because nothing says "cutting-edge" like moaning about parameter scales and reinforcement learning 🙄. Apparently, this 32B thing is "smarter" and "lighter" – sounds like a diet ad for AI models. 😂🍩 #Innovation!
    qwenlm.github.io/blog/qwen2.5- #Qwen2.5VL32B #AIModels #ReinforcementLearning #CuttingEdge #TechHumor #HackerNews #ngated

  31. Qwen2.5-VL-32B: because nothing says "cutting-edge" like moaning about parameter scales and reinforcement learning 🙄. Apparently, this 32B thing is "smarter" and "lighter" – sounds like a diet ad for AI models. 😂🍩 #Innovation!
    qwenlm.github.io/blog/qwen2.5- #Qwen2.5VL32B #AIModels #ReinforcementLearning #CuttingEdge #TechHumor #HackerNews #ngated

  32. OLMo 2 32B offers unprecedented transparency in #LLM development:

    • 🚀 State-of-the-art results: Outperforms GPT3.5, GPT4o-mini, matches top open-weight models like #Qwen2.5 and approaches #Llama3

  33. OLMo 2 32B offers unprecedented transparency in #LLM development:

    • 🚀 State-of-the-art results: Outperforms GPT3.5, GPT4o-mini, matches top open-weight models like #Qwen2.5 and approaches #Llama3

  34. OLMo 2 32B offers unprecedented transparency in #LLM development:

    • 🚀 State-of-the-art results: Outperforms GPT3.5, GPT4o-mini, matches top open-weight models like #Qwen2.5 and approaches #Llama3

  35. OLMo 2 32B offers unprecedented transparency in #LLM development:

    • 🚀 State-of-the-art results: Outperforms GPT3.5, GPT4o-mini, matches top open-weight models like #Qwen2.5 and approaches #Llama3

  36. OLMo 2 32B offers unprecedented transparency in #LLM development:

    • 🚀 State-of-the-art results: Outperforms GPT3.5, GPT4o-mini, matches top open-weight models like #Qwen2.5 and approaches #Llama3

  37. #AI2 releases OLMo 2 32B, trained on 6T tokens with #Tulu3.1 post-training. Matches or exceeds GPT3.5 Turbo while using just 1/3 the compute of #Qwen2.5 32B. Complete open recipe includes data, code, weights and training methodology.

  38. #AI2 releases OLMo 2 32B, trained on 6T tokens with #Tulu3.1 post-training. Matches or exceeds GPT3.5 Turbo while using just 1/3 the compute of #Qwen2.5 32B. Complete open recipe includes data, code, weights and training methodology.

  39. #AI2 releases OLMo 2 32B, trained on 6T tokens with #Tulu3.1 post-training. Matches or exceeds GPT3.5 Turbo while using just 1/3 the compute of #Qwen2.5 32B. Complete open recipe includes data, code, weights and training methodology.

  40. #AI2 releases OLMo 2 32B, trained on 6T tokens with #Tulu3.1 post-training. Matches or exceeds GPT3.5 Turbo while using just 1/3 the compute of #Qwen2.5 32B. Complete open recipe includes data, code, weights and training methodology.

  41. #AI2 releases OLMo 2 32B, trained on 6T tokens with #Tulu3.1 post-training. Matches or exceeds GPT3.5 Turbo while using just 1/3 the compute of #Qwen2.5 32B. Complete open recipe includes data, code, weights and training methodology.