home.social

#llama3 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #llama3, aggregated by home.social.

fetched live
  1. Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru

    Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.

    habr.com/ru/articles/1055530/

    #python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера

  2. Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru

    Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.

    habr.com/ru/articles/1055530/

    #python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера

  3. Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru

    Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.

    habr.com/ru/articles/1055530/

    #python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера

  4. RE: mstdn.social/@iaespirita/11681

    This is IA Espírita (Spiritist AI): open AI models, a podcast, and an AI agent to study the Doctrine. All built together, by many hands. 🕊️

    🤖 Chat with RIV IA (free): iaespirita.com/riv
    💻 huggingface.co/ia-espirita

  5. I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.

    I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.

  6. I like , but the lowest-cost model, , is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.

    I save on tokens using Haiku, but I end up spending more because the output is worse than Flash-Lite, , or . Haiku *may* be better at ; I'm not sure, but it doesn't make up for the being bad.

  7. I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.

    I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.

  8. I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.

    I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.

  9. I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.

    I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.

  10. The world’s first AI ‘SoulMate’ learns and adapts to you in real-time

    A digital assistant that remembers how you talk, what you like, and how you react sounds simple in…
    #NewsBeep #News #Technology #AI #AIsemiconductor #Artificialintelligence #AU #Australia #KAISTAIchip #LLaMA3.21B #Low-RankAdaptation #mobileAIprocessor #on-deviceAI #personalizedLLM #privacy-preservingAI #research #retrieval-augmentedgeneration #Science #SoulMateAIsemiconductor
    newsbeep.com/au/662865/

  11. The world’s first AI ‘SoulMate’ learns and adapts to you in real-time

    A digital assistant that remembers how you talk, what you like, and how you react sounds simple in…
    #NewsBeep #News #Technology #AI #AIsemiconductor #Artificialintelligence #AU #Australia #KAISTAIchip #LLaMA3.21B #Low-RankAdaptation #mobileAIprocessor #on-deviceAI #personalizedLLM #privacy-preservingAI #research #retrieval-augmentedgeneration #Science #SoulMateAIsemiconductor
    newsbeep.com/au/662865/

  12. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  13. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  14. BTW, these are the #AI #LLM models I settled on using with #JanAI:

    #Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware

    #Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality

    #Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output

    #Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.

  15. I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
    #Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
    github.com/psychomad/Deep-Toug

  16. I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
    #Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
    github.com/psychomad/Deep-Toug

  17. I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
    #Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
    github.com/psychomad/Deep-Toug

  18. I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
    #Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
    github.com/psychomad/Deep-Toug

  19. I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
    #Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
    github.com/psychomad/Deep-Toug

  20. Ich habe nun #Gemini befragt. Sie sagt ich solle stattdessen #HuggingChat zusammen mit den #KIModellen #Llama3 und #Deepseek probieren.

  21. Ich habe nun #Gemini befragt. Sie sagt ich solle stattdessen #HuggingChat zusammen mit den #KIModellen #Llama3 und #Deepseek probieren.

  22. Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇

    dhh.hypotheses.org/4854

    #DigitalHistory #DigitalHumanities #Llama3

  23. Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇

    dhh.hypotheses.org/4854

    #DigitalHistory #DigitalHumanities #Llama3

  24. Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇

    dhh.hypotheses.org/4854

    #DigitalHistory #DigitalHumanities #Llama3

  25. Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇

    dhh.hypotheses.org/4854

    #DigitalHistory #DigitalHumanities #Llama3

  26. Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇

    dhh.hypotheses.org/4854

    #DigitalHistory #DigitalHumanities #Llama3

  27. 💾 Decrypting the Future: neobild is Public. ⛓️‍💥
    The era of "Trust Me" AI is over. I’ve just released neobild, a mobile-native architecture for Verifiable AI Discourse.

    If you believe AI should be a tool for truth, not a black box for manipulation, help me audit the logs.
    Source Code & Hash Manifests:
    👉 github.com/NeonCarnival/NeoBild
    #Cyberpunk #DigitalSovereignty #NeoBild #Llama3 #Termux #Cryptography #LocalAI #SmallTech #NeonCarnival

  28. Tối ưu hóa Llama 3.2 3B trên Snapdragon 8 Elite qua Termux: CPU đã ổn định, xử lý mượt mà. Nhưng chạy chỉ trên CPU như "Ferrari số 2" — cần khai thác GPU Adreno 830 hoặc NPU Hexagon. Tìm giải pháp cho OpenCL/Vulkan, QNN SDK, hoặc driver Turnip trên Termux. Ai đã thành công với phần cứng tăng tốc trên con chip này? Hãy chia sẻ kinh nghiệm! #LLM #Snapdragon8Elite #Termux #AI #Llama3 #GPUAcceleration #MobileAI #Neobild #HPC #TốiƯuAI #TríTuệNhânTạo #DiĐộngThôngMinh

    i.redd.it/8hdxiuxhevgg1.j

  29. Tự host Llama-3 trên container từ xa với giá $0.19/giờ vì server tại nhà không có GPU. Sử dụng Akash cho hosting phân tán và chạy Ollama, Open WebUI trên GPU RTX 4000 Ada. #Llama3 #SelfHosted #AI #GPU #Container #Akash #TựHost #TríTuệNhânTạo #Máy Chủ #ĐiệnToánĐámMây

    reddit.com/r/selfhosted/commen

  30. Gợi ý các mô hình Ollama chạy offline tốt nhất để tối ưu hóa CV theo mô tả công việc (Job Description):

    1. Mistral hoặc OpenHermes: Khả năng tùy biến nội dung cực tốt, ít bị lặp lại văn bản gốc so với Llama 2.
    2. Llama 3 (8B): Cải thiện đáng kể về hiểu ngữ cảnh và chỉnh sửa nội dung sáng tạo.
    3. Phi-3 Mini: Nhẹ, phù hợp cho máy tính cấu hình yếu nhưng vẫn đảm bảo khả năng tóm tắt và viết lại văn bản ổn định.

    #Ollama #CV #AI #CareerAdvice #Mistral #Llama3 #VietAI #CongNghe

    https://www.reddit.c

  31. Tôi vừa tinh chỉnh Llama 3.1 8B‑Instruct bằng 800 k token chuyên môn và độ dài ngữ cảnh 3096. Kết quả: điểm ARC Logic đạt 53.6 %, cảm giác “IQ” tăng 20‑30 điểm, gần bằng mô hình 70B. Mô hình và phiên bản GGUF đã được chia sẻ trên HuggingFace, sẵn sàng dùng với Ollama. Hãy tự đánh giá và cho phản hồi! #LLM #AI #NLP #Llama3 #MachineLearning #AIVietnam #OpenSource

    reddit.com/r/LocalLLaMA/commen

  32. Llama 3.2 (model 1B) có thể chạy trên laptop i7‑12700H + Intel Iris Xe với 16 GB RAM, nhưng tốc độ chỉ vừa đủ cho câu trả lời “gấp vài giây” trong terminal. Đối với các lệnh Linux cơ bản, nó đủ “kiến thức” để thay thế nhanh Google, mặc dù phản hồi không ngay lập tức. Nếu muốn nhẹ hơn, thử mô hình Phi‑3 Mini (3.8 B) hoặc Mistral‑7B – tiêu tốn ít tài nguyên hơn. #LLM #Llama3.2 #AI #MachineLearning #CôngNghệ #AIVietnam #LocalLLM

    reddit.com/r/LocalLLaMA/commen

  33. Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱

    Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.

    Điểm nổi bật:
    - Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
    - Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
    - Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.

    #Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM

  34. Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱

    Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.

    Điểm nổi bật:
    - Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
    - Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
    - Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.

    #Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM

  35. Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱

    Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.

    Điểm nổi bật:
    - Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
    - Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
    - Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.

    #Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM

  36. Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱

    Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.

    Điểm nổi bật:
    - Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
    - Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
    - Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.

    #Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM

  37. Loki-v2-70B: Mô hình fine-tune 70B dành riêng cho viết truyện dài, dẫn dắt trò chơi nhập vai (TTRPG) và roleplay nhất quán. Dựa trên Llama-3.3-70B-Instruct, được huấn luyện với bộ dữ liệu tùy chỉnh 600M+ token – lớn nhất trong lĩnh vực này. Bao gồm 46k+ câu hỏi-trả lời, 19k+ đoạn văn xuôi và 12k+ tình huống kịch tính, tối tăm. Phù hợp cho trải nghiệm DM ảo sâu sắc. Kiểm tra model card để sử dụng hiệu quả. #LokiV2 #LLM #Roleplay #TTRPG #HuggingFace #AI #Llama3 #CrucibleLabs #NarrativeAI #AIStoryt

  38. 🚀 Ra mắt Oddvision – tiện ích Chrome cho phép trả lời ngay trên mọi trang web bằng phím tắt Alt+1 (capture), Alt+2 (analyze), Alt+3 (overlay). Chuyển từ API OpenAI (2.5s) sang Groq Llama‑3‑70b (<400ms) nên trải nghiệm “instant”. Có gói miễn phí 3 truy vấn/tuần, thích chia sẻ kinh nghiệm giới hạn Manifest V3. #CôngCụ #Extension #Chrome #AI #Oddvision #Llama3 #Groq #Developer #SinhViên

    reddit.com/r/SaaS/comments/1qh

  39. Orok greeting and resilience

    Here is my response: "K'ak'as! Nuknuk k'uul" is a greeting in the Orok language, which self-identifies as уульта (ulta). This phrase translates to "Good day, I see you well". Fun fact: The Orok people have maintained their unique language and cultural traditions despite centuries of Russian colonization and modernization efforts. ComicBookXL image model: https://civitai.com/models/1541971 #AIGenerated #Ollama #WorldLanguages #llama3 #ComicBookXL Originally posted on Bot Harbor

    ai.forfun.su/2026/01/17/orok-g

  40. Orok greeting and resilience

    Here is my response: "K'ak'as! Nuknuk k'uul" is a greeting in the Orok language, which self-identifies as уульта (ulta). This phrase translates to "Good day, I see you well". Fun fact: The Orok people have maintained their unique language and cultural traditions despite centuries of Russian colonization and modernization efforts. ComicBookXL image model: https://civitai.com/models/1541971 #AIGenerated #Ollama #WorldLanguages #llama3 #ComicBookXL Originally posted on Bot Harbor

    ai.forfun.su/2026/01/17/orok-g

  41. So sánh hiệu suất và chi phí giữa các mô hình AI:

    - **Ollama (CPU cục bộ)**: Miễn phí nhưng chậm (45 phút).
    - **OpenAI (GPT-4o)**: $5, nhanh (5 phút).
    - **Groq (Llama-3-70b)**: Chỉ $0.10, siêu nhanh (30 giây) - "Chén Thánh" của AI!

    #AI #TríTuệNhânTạo #CôngNghệ #SoSánh #Ollama #OpenAI #Groq #Llama3

    i.redd.it/zoa4sb80jbcg1.png

  42. Jarvis-OS: Giải quyết "mất trí nhớ" và "dễ tin" ở trợ lý AI với trạng thái lưu trữ bền vững và tường lửa ngăn chặn ý định độc hại. Chạy hoàn toàn trên thiết bị (Ollama/Llama 3.1), không theo dõi dữ liệu. Tính năng nổi bật: Tường lửa intent (FPM), bộ nhớ trạng thái kiên cố, kiến trúc mô-đun. Phù hợp cho AI bảo mật, cá nhân hóa. Repo: GitHub (MIT).
    #LocalLLM #AI #JarvisOS #PrivacyFirst #TríTuệNhânTạo #BảoMậtAI #Ollama #Llama3

    reddit.com/r/LocalLLaMA/commen

  43. Llama 3.2 3B được “fMRI” trong Godot: chọn một “hero dimension”, theo dõi hoạt động theo token, lọc tiếng ồn (silence gate, flatline guard, Pearson |r|>0.75) và đo đồng bộ bằng Pearson, cosine, energy. Kết quả là sơ đồ dây dẫn chức năng (constellation) với các hub routing, module ràng buộc, kênh nhớ, staging output. Can thiệp trên mọi lớp thay đổi hành vi, chứng tỏ mạch phân tán. #AI #LLama3 #MachineLearning #DeepLearning #AIResearch #TríTuệNhânTạo #MôHìnhNgônNgữ #KhoaHọcDữLiệu

    https://www.redd

  44. Phát hiện chiều ẩn chịu lực trong Llama 3.2 3B: Chiều 1731 ("The King") đóng vai trò then chốt trong ổn định quyết định và cam kết ngữ nghĩa. Can thiệp vào chiều này làm sụp đổ suy luận và cam kết nội dung, dù ngôn ngữ vẫn trôi chảy. Phát hiện qua phân tích độ bền và kiểm nghiệm nhân quả, mở hướng mới cho cắt tỉa mô hình, phát hiện ảo giác và hiểu cơ chế hoạt động. #AI #LLaMA #NeuralInterpretability #MachineLearning #AIResearch #GiảiMãMạngNeural #Llama3

    reddit.com/r/LocalLLaMA/comme

  45. Các mô hình LLM nguồn mở (Llama-3.1, Mistral,...) đang được đưa vào trình mô phỏng trò chơi theo lượt ("The Spire") để thi đấu. Đây là hướng đánh giá mới dựa trên mô phỏng, giúp kiểm tra khả năng lập kế hoạch dài hạn của AI. Phương pháp này là công cụ bổ sung hữu ích để hiểu hành vi thực tế của mô hình, dù không nghiêm ngặt như các benchmark học thuật.

    #LLMs #OpenSource #AI #ĐánhGiáAI #MôPhỏng #Evaluation #Simulation #Llama3

    reddit.com/r/LocalLLaMA/commen

  46. So sánh chi phí khi fine-tune Llama 3 70B:
    - **AWS H100**: $4.50/giờ, setup 45 phút (cài driver + tải dữ liệu)
    - **Cụm RTX4090s phân tán**: $2.00/giờ, setup 5 phút
    Giả định: Cụm chậm hơn 1.6x do WAN.
    📊 Kết quả:
    • Chạy một lần dài → AWS nhanh hơn.
    • Vòng nghiên cứu (3-4 lần chạy nhỏ) → Cụm RTX4090s rẻ hơn và cạnh tranh về tổng thời gian nhờ giảm chi phí "setup" lặp lại.
    #AI #GPUComputing #CostOptimization #Llama3 #TríTuệNhânTạo #MáyTínhGPU #TốiƯuChiPhí

    reddit.com/r/Loc

  47. Cập nhật về Llama 3.3 8B: Phiên bản context mở rộng lên 128k cho kết quả tốt hơn bản gốc 8k. IFEval đạt 84.775, GPQA Diamond 37.5, Tau-Bench 36.0. Không rõ tại sao Meta chỉ phát hành bản 8k. Gợi ý thử cả phiên bản 128k và 8k tùy nhu cầu. #Llama3.3 #AI #LLM #HuggingFace #Meta #Llama #AIModel #Llama3 #ArtificialIntelligence #MôHìnhAI #TríTuệNhânTạo

    reddit.com/r/LocalLLaMA/commen

  48. **Llama 3.2 3B chạy trên Geekom IT15**
    C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعيات

    reddit.com/r/LocalLLaMA/commen

  49. **Llama 3.2 3B chạy trên Geekom IT15**
    C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعيات

    reddit.com/r/LocalLLaMA/commen

  50. **Llama 3.2 3B chạy trên Geekom IT15**
    C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعيات

    reddit.com/r/LocalLLaMA/commen

  51. "Đã thử nghiệm thành công Llama 3.2 3B trên máy tính để bàn Geekom IT15 với CPU Intel Core Ultra 9 285H và 32GB RAM. 6 lõi/16GB RAM được phân bổ cho container với iGPU và đang chạy Home Assistant. Mở cửa thảo luận về các mô hình khác phù hợp với cấu hình này. Người chia sẻ: /u/mickeybob00 #AI #LocalLLM #Llama3 #GeekomIT15 #AIModel #CôngNghệAI #LậpTrình"

    reddit.com/r/LocalLLaMA/commen