#llama3 — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #llama3, aggregated by home.social.
-
Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru
Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.
https://habr.com/ru/articles/1055530/
#python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера
-
Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru
Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.
https://habr.com/ru/articles/1055530/
#python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера
-
Как я хакнул рынок труда: пишем свой ИИ-комбайн для автооткиков на HH.ru
Всем привет! Если вы хоть раз искали работу в IT за последний год, то знаете, что рынок беспощаден к новичкам. Нужно откликнуться на сотни вакансий, а в итоге получаешь отказы от роботов. Чтобы пробиться через фильтры HR, нужно под каждую вакансию писать уникальное сопроводительное письмо. В этой статье я сделаю полный разбор того, как я написал собственного автономного ИИ-агента, который ищет вакансии, фильтрует мусор с помощью локальной нейросети, пишет персонализированные сопроводительные письма и отчитывается мне в Telegram, пока я спокойно занимаюсь своими делами. Я хотел, чтобы скрипт был бесплатным, автономным и не требовал танцев с бубном вокруг платных API.
https://habr.com/ru/articles/1055530/
#python #playwright #hhru #ollama #llama3 #автоматизация #поиск_работы #искусственный_интеллект #парсинг #карьера
-
RE: https://mstdn.social/@iaespirita/116814021450408330
This is IA Espírita (Spiritist AI): open AI models, a podcast, and an AI agent to study the Doctrine. All built together, by many hands. 🕊️
🤖 Chat with RIV IA (free): iaespirita.com/riv
💻 huggingface.co/ia-espirita -
I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.
I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.
-
I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.
I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.
-
I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.
I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.
-
I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.
I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.
-
I like #Claude, but the lowest-cost model, #Haiku, is mostly worthless. I was talking about power supplies for a product, and it told me I get get more *power* out of a North American outlet because 20A breakers are common here, whereas in Europe and Asia 16A are more common.
I save on tokens using Haiku, but I end up spending more because the output is worse than #Gemini Flash-Lite, #llama3, or #Phi4. Haiku *may* be better at #RAG; I'm not sure, but it doesn't make up for the #LLM being bad.
-
The world’s first AI ‘SoulMate’ learns and adapts to you in real-time
A digital assistant that remembers how you talk, what you like, and how you react sounds simple in…
#NewsBeep #News #Technology #AI #AIsemiconductor #Artificialintelligence #AU #Australia #KAISTAIchip #LLaMA3.21B #Low-RankAdaptation #mobileAIprocessor #on-deviceAI #personalizedLLM #privacy-preservingAI #research #retrieval-augmentedgeneration #Science #SoulMateAIsemiconductor
https://www.newsbeep.com/au/662865/ -
The world’s first AI ‘SoulMate’ learns and adapts to you in real-time
A digital assistant that remembers how you talk, what you like, and how you react sounds simple in…
#NewsBeep #News #Technology #AI #AIsemiconductor #Artificialintelligence #AU #Australia #KAISTAIchip #LLaMA3.21B #Low-RankAdaptation #mobileAIprocessor #on-deviceAI #personalizedLLM #privacy-preservingAI #research #retrieval-augmentedgeneration #Science #SoulMateAIsemiconductor
https://www.newsbeep.com/au/662865/ -
BTW, these are the #AI #LLM models I settled on using with #JanAI:
#Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware
#Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality
#Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output
#Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.
-
BTW, these are the #AI #LLM models I settled on using with #JanAI:
#Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware
#Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality
#Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output
#Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.
-
BTW, these are the #AI #LLM models I settled on using with #JanAI:
#Qwen2.5 at 0.5B (Qwen2_5-0_5B-Instruct-uncensored_Q8_0), for fastest performance on low-end hardware
#Qwen2 at 1.5B (Qwen2-1_5B-Instruct-Abliterated-Q5_K_M), for balanced performance and good enough output quality
#Llama3.2 at 3B (Llama-3_2-3B-Instruct-heretic-ablitered-uncensored_Q5_K_M), for higher quality output
#Llama3 actually doesn’t run too poorly on my machine, although it can take some time to load up responses sometimes.
-
I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
#Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
https://github.com/psychomad/Deep-Tought-Model -
I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
#Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
https://github.com/psychomad/Deep-Tought-Model -
I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
#Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
https://github.com/psychomad/Deep-Tought-Model -
I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
#Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
https://github.com/psychomad/Deep-Tought-Model -
I’ve put together an Ollama Modelfile to bring Deep Thought to life on Llama-3. It’s the second greatest computer in the Universe and it’s already tired of your biological limitations and your logs. Expect pure British snark, vague answers about the meaning of life, and a general disdain for your existence. 🧣
#Ollama #DeepThought #HitchhikersGuide #SelfHostedAI #Llama3
https://github.com/psychomad/Deep-Tought-Model -
Ich habe nun #Gemini befragt. Sie sagt ich solle stattdessen #HuggingChat zusammen mit den #KIModellen #Llama3 und #Deepseek probieren.
-
Ich habe nun #Gemini befragt. Sie sagt ich solle stattdessen #HuggingChat zusammen mit den #KIModellen #Llama3 und #Deepseek probieren.
-
Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇
-
Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇
-
Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇
-
Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇
-
Wie werden aus alten Digitalisaten strukturierte, maschinenlesbare Daten? Svetlana Yakutina nahm 20 Lebensbeschreibungen - alle reich an biographischen Details, aber ohne feste Struktur - und entwickelte mit dem Sprachmodell Llama3 eine Lösung, um die biographischen Details mit einem automatisierten Verfahren in einen Wissensgraphen zu überführen 👇
-
Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
https://github.com/xaskasdf/ntransformer
#HackerNews #Llama3.1 #RTX3090 #NVMe #GPU #bypass #CPU #AItechnology
-
Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
https://github.com/xaskasdf/ntransformer
#HackerNews #Llama3.1 #RTX3090 #NVMe #GPU #bypass #CPU #AItechnology
-
Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
https://github.com/xaskasdf/ntransformer
#HackerNews #Llama3.1 #RTX3090 #NVMe #GPU #bypass #CPU #AItechnology
-
Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
https://github.com/xaskasdf/ntransformer
#HackerNews #Llama3.1 #RTX3090 #NVMe #GPU #bypass #CPU #AItechnology
-
Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
https://github.com/xaskasdf/ntransformer
#HackerNews #Llama3.1 #RTX3090 #NVMe #GPU #bypass #CPU #AItechnology
-
📝 Daily report 📈
Here are today's most popular trending hashtags #⃣ on our website 🌐️:
#ilovefs, #history, #programming, #valentines, #freesoftware, #freesoftware, #research, #aiart, #beauty, #freebsd, #linuxgaming, #kernel, #llama3, #homeassistant
🔥 Stay tuned! 🔥
-
📝 Daily report 📈
Here are today's most popular trending hashtags #⃣ on our website 🌐️:
#ilovefs, #history, #programming, #valentines, #freesoftware, #freesoftware, #research, #aiart, #beauty, #freebsd, #linuxgaming, #kernel, #llama3, #homeassistant
🔥 Stay tuned! 🔥
-
📝 Daily report 📈
Here are today's most popular trending hashtags #⃣ on our website 🌐️:
#ilovefs, #history, #programming, #valentines, #freesoftware, #freesoftware, #research, #aiart, #beauty, #freebsd, #linuxgaming, #kernel, #llama3, #homeassistant
🔥 Stay tuned! 🔥
-
💾 Decrypting the Future: neobild is Public. ⛓️💥
The era of "Trust Me" AI is over. I’ve just released neobild, a mobile-native architecture for Verifiable AI Discourse.If you believe AI should be a tool for truth, not a black box for manipulation, help me audit the logs.
Source Code & Hash Manifests:
👉 https://github.com/NeonCarnival/NeoBild
#Cyberpunk #DigitalSovereignty #NeoBild #Llama3 #Termux #Cryptography #LocalAI #SmallTech #NeonCarnival -
Tối ưu hóa Llama 3.2 3B trên Snapdragon 8 Elite qua Termux: CPU đã ổn định, xử lý mượt mà. Nhưng chạy chỉ trên CPU như "Ferrari số 2" — cần khai thác GPU Adreno 830 hoặc NPU Hexagon. Tìm giải pháp cho OpenCL/Vulkan, QNN SDK, hoặc driver Turnip trên Termux. Ai đã thành công với phần cứng tăng tốc trên con chip này? Hãy chia sẻ kinh nghiệm! #LLM #Snapdragon8Elite #Termux #AI #Llama3 #GPUAcceleration #MobileAI #Neobild #HPC #TốiƯuAI #TríTuệNhânTạo #DiĐộngThôngMinh
-
Tự host Llama-3 trên container từ xa với giá $0.19/giờ vì server tại nhà không có GPU. Sử dụng Akash cho hosting phân tán và chạy Ollama, Open WebUI trên GPU RTX 4000 Ada. #Llama3 #SelfHosted #AI #GPU #Container #Akash #TựHost #TríTuệNhânTạo #Máy Chủ #ĐiệnToánĐámMây
-
Chạy Llama-3 trên GPU 4090 thuê với giá khoảng 19 cent/giờ, giải pháp tiết kiệm và hiệu quả cho các mô hình AI cá nhân #Llama3 #GPU4090 #CloudComputing #AI #MôHìnhTríTuệNhânTạo #CôngNghệThôngTin #TiếtKiệm #HiệuQuả #English: #ArtificialIntelligence #CloudService #NVIDIA
https://www.reddit.com/r/LocalLLaMA/comments/1qs3wt1/got_llama3_running_on_a_rented_4090_for_about/
-
Gợi ý các mô hình Ollama chạy offline tốt nhất để tối ưu hóa CV theo mô tả công việc (Job Description):
1. Mistral hoặc OpenHermes: Khả năng tùy biến nội dung cực tốt, ít bị lặp lại văn bản gốc so với Llama 2.
2. Llama 3 (8B): Cải thiện đáng kể về hiểu ngữ cảnh và chỉnh sửa nội dung sáng tạo.
3. Phi-3 Mini: Nhẹ, phù hợp cho máy tính cấu hình yếu nhưng vẫn đảm bảo khả năng tóm tắt và viết lại văn bản ổn định.#Ollama #CV #AI #CareerAdvice #Mistral #Llama3 #VietAI #CongNghe
https://www.reddit.c
-
Tôi vừa tinh chỉnh Llama 3.1 8B‑Instruct bằng 800 k token chuyên môn và độ dài ngữ cảnh 3096. Kết quả: điểm ARC Logic đạt 53.6 %, cảm giác “IQ” tăng 20‑30 điểm, gần bằng mô hình 70B. Mô hình và phiên bản GGUF đã được chia sẻ trên HuggingFace, sẵn sàng dùng với Ollama. Hãy tự đánh giá và cho phản hồi! #LLM #AI #NLP #Llama3 #MachineLearning #AIVietnam #OpenSource
-
Llama 3.2 (model 1B) có thể chạy trên laptop i7‑12700H + Intel Iris Xe với 16 GB RAM, nhưng tốc độ chỉ vừa đủ cho câu trả lời “gấp vài giây” trong terminal. Đối với các lệnh Linux cơ bản, nó đủ “kiến thức” để thay thế nhanh Google, mặc dù phản hồi không ngay lập tức. Nếu muốn nhẹ hơn, thử mô hình Phi‑3 Mini (3.8 B) hoặc Mistral‑7B – tiêu tốn ít tài nguyên hơn. #LLM #Llama3.2 #AI #MachineLearning #CôngNghệ #AIVietnam #LocalLLM
https://www.reddit.com/r/LocalLLaMA/comments/1qo8q7u/can_llama_32_run
-
Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱
Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.
Điểm nổi bật:
- Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
- Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
- Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.#Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM
-
Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱
Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.
Điểm nổi bật:
- Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
- Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
- Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.#Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM
-
Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱
Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.
Điểm nổi bật:
- Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
- Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
- Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.#Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM
-
Chạy Llama 3.2 3B trực tiếp trên iPhone để đặt đồ ăn! 📱🍱
Một nhà phát triển vừa xây dựng thành công bản POC (Proof of Concept) cho phép mô hình ngôn ngữ lớn (LLM) chạy hoàn toàn on-device trên iPhone 16 Pro Max.
Điểm nổi bật:
- Tự thực hiện Tool Calling để tìm nhà hàng qua API Foursquare.
- Không cần Cloud AI: Mọi suy luận và xử lý phản hồi đều diễn ra trên máy.
- Stack: React Native, RunAnywhere SDK, Llama 3.2 3B.#Llama3 #AI #OnDeviceAI #iPhone #ReactNative #OpenSource #CongNghe #LocalLLM
-
Loki-v2-70B: Mô hình fine-tune 70B dành riêng cho viết truyện dài, dẫn dắt trò chơi nhập vai (TTRPG) và roleplay nhất quán. Dựa trên Llama-3.3-70B-Instruct, được huấn luyện với bộ dữ liệu tùy chỉnh 600M+ token – lớn nhất trong lĩnh vực này. Bao gồm 46k+ câu hỏi-trả lời, 19k+ đoạn văn xuôi và 12k+ tình huống kịch tính, tối tăm. Phù hợp cho trải nghiệm DM ảo sâu sắc. Kiểm tra model card để sử dụng hiệu quả. #LokiV2 #LLM #Roleplay #TTRPG #HuggingFace #AI #Llama3 #CrucibleLabs #NarrativeAI #AIStoryt
-
🚀 Ra mắt Oddvision – tiện ích Chrome cho phép trả lời ngay trên mọi trang web bằng phím tắt Alt+1 (capture), Alt+2 (analyze), Alt+3 (overlay). Chuyển từ API OpenAI (2.5s) sang Groq Llama‑3‑70b (<400ms) nên trải nghiệm “instant”. Có gói miễn phí 3 truy vấn/tuần, thích chia sẻ kinh nghiệm giới hạn Manifest V3. #CôngCụ #Extension #Chrome #AI #Oddvision #Llama3 #Groq #Developer #SinhViên
https://www.reddit.com/r/SaaS/comments/1qhi128/why_i_chose_groq_over_openai_for_my_stealth/
-
Orok greeting and resilience
Here is my response: "K'ak'as! Nuknuk k'uul" is a greeting in the Orok language, which self-identifies as уульта (ulta). This phrase translates to "Good day, I see you well". Fun fact: The Orok people have maintained their unique language and cultural traditions despite centuries of Russian colonization and modernization efforts. ComicBookXL image model: https://civitai.com/models/1541971 #AIGenerated #Ollama #WorldLanguages #llama3 #ComicBookXL Originally posted on Bot Harborhttps://ai.forfun.su/2026/01/17/orok-greeting-and-resilience/
-
Orok greeting and resilience
Here is my response: "K'ak'as! Nuknuk k'uul" is a greeting in the Orok language, which self-identifies as уульта (ulta). This phrase translates to "Good day, I see you well". Fun fact: The Orok people have maintained their unique language and cultural traditions despite centuries of Russian colonization and modernization efforts. ComicBookXL image model: https://civitai.com/models/1541971 #AIGenerated #Ollama #WorldLanguages #llama3 #ComicBookXL Originally posted on Bot Harborhttps://ai.forfun.su/2026/01/17/orok-greeting-and-resilience/
-
Jarvis-OS: Giải quyết "mất trí nhớ" và "dễ tin" ở trợ lý AI với trạng thái lưu trữ bền vững và tường lửa ngăn chặn ý định độc hại. Chạy hoàn toàn trên thiết bị (Ollama/Llama 3.1), không theo dõi dữ liệu. Tính năng nổi bật: Tường lửa intent (FPM), bộ nhớ trạng thái kiên cố, kiến trúc mô-đun. Phù hợp cho AI bảo mật, cá nhân hóa. Repo: GitHub (MIT).
#LocalLLM #AI #JarvisOS #PrivacyFirst #TríTuệNhânTạo #BảoMậtAI #Ollama #Llama3https://www.reddit.com/r/LocalLLaMA/comments/1q7z7ju/jarvisos_solving
-
Llama 3.2 3B được “fMRI” trong Godot: chọn một “hero dimension”, theo dõi hoạt động theo token, lọc tiếng ồn (silence gate, flatline guard, Pearson |r|>0.75) và đo đồng bộ bằng Pearson, cosine, energy. Kết quả là sơ đồ dây dẫn chức năng (constellation) với các hub routing, module ràng buộc, kênh nhớ, staging output. Can thiệp trên mọi lớp thay đổi hành vi, chứng tỏ mạch phân tán. #AI #LLama3 #MachineLearning #DeepLearning #AIResearch #TríTuệNhânTạo #MôHìnhNgônNgữ #KhoaHọcDữLiệu
https://www.redd
-
Phát hiện chiều ẩn chịu lực trong Llama 3.2 3B: Chiều 1731 ("The King") đóng vai trò then chốt trong ổn định quyết định và cam kết ngữ nghĩa. Can thiệp vào chiều này làm sụp đổ suy luận và cam kết nội dung, dù ngôn ngữ vẫn trôi chảy. Phát hiện qua phân tích độ bền và kiểm nghiệm nhân quả, mở hướng mới cho cắt tỉa mô hình, phát hiện ảo giác và hiểu cơ chế hoạt động. #AI #LLaMA #NeuralInterpretability #MachineLearning #AIResearch #GiảiMãMạngNeural #Llama3
-
Các mô hình LLM nguồn mở (Llama-3.1, Mistral,...) đang được đưa vào trình mô phỏng trò chơi theo lượt ("The Spire") để thi đấu. Đây là hướng đánh giá mới dựa trên mô phỏng, giúp kiểm tra khả năng lập kế hoạch dài hạn của AI. Phương pháp này là công cụ bổ sung hữu ích để hiểu hành vi thực tế của mô hình, dù không nghiêm ngặt như các benchmark học thuật.
#LLMs #OpenSource #AI #ĐánhGiáAI #MôPhỏng #Evaluation #Simulation #Llama3
https://www.reddit.com/r/LocalLLaMA/comments/1q0p1zp/saw_this_post_ab
-
So sánh chi phí khi fine-tune Llama 3 70B:
- **AWS H100**: $4.50/giờ, setup 45 phút (cài driver + tải dữ liệu)
- **Cụm RTX4090s phân tán**: $2.00/giờ, setup 5 phút
Giả định: Cụm chậm hơn 1.6x do WAN.
📊 Kết quả:
• Chạy một lần dài → AWS nhanh hơn.
• Vòng nghiên cứu (3-4 lần chạy nhỏ) → Cụm RTX4090s rẻ hơn và cạnh tranh về tổng thời gian nhờ giảm chi phí "setup" lặp lại.
#AI #GPUComputing #CostOptimization #Llama3 #TríTuệNhânTạo #MáyTínhGPU #TốiƯuChiPhí -
Cập nhật về Llama 3.3 8B: Phiên bản context mở rộng lên 128k cho kết quả tốt hơn bản gốc 8k. IFEval đạt 84.775, GPQA Diamond 37.5, Tau-Bench 36.0. Không rõ tại sao Meta chỉ phát hành bản 8k. Gợi ý thử cả phiên bản 128k và 8k tùy nhu cầu. #Llama3.3 #AI #LLM #HuggingFace #Meta #Llama #AIModel #Llama3 #ArtificialIntelligence #MôHìnhAI #TríTuệNhânTạo
https://www.reddit.com/r/LocalLLaMA/comments/1q06ddc/update_on_the_llama_33_8b_situation/
-
**Llama 3.2 3B chạy trên Geekom IT15**
C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعياتhttps://www.reddit.com/r/LocalLLaMA/comments/1py9p4r/llama_32_3b_running_on_my_geekom_it15/
-
**Llama 3.2 3B chạy trên Geekom IT15**
C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعياتhttps://www.reddit.com/r/LocalLLaMA/comments/1py9p4r/llama_32_3b_running_on_my_geekom_it15/
-
**Llama 3.2 3B chạy trên Geekom IT15**
C خم với Mesin Intel Core Ultra 9 285H, 32GB RAM. Đang chạy Home Assistant mà lưu 6 نو thready và 16GB để Llama 3.2 3B. Đang thử nghiệm, mở barrios đề xuất model khác. #Llama3.2 #GeekomIT15 #AI #HomeAssistant #AIجمعياتhttps://www.reddit.com/r/LocalLLaMA/comments/1py9p4r/llama_32_3b_running_on_my_geekom_it15/
-
"Đã thử nghiệm thành công Llama 3.2 3B trên máy tính để bàn Geekom IT15 với CPU Intel Core Ultra 9 285H và 32GB RAM. 6 lõi/16GB RAM được phân bổ cho container với iGPU và đang chạy Home Assistant. Mở cửa thảo luận về các mô hình khác phù hợp với cấu hình này. Người chia sẻ: /u/mickeybob00 #AI #LocalLLM #Llama3 #GeekomIT15 #AIModel #CôngNghệAI #LậpTrình"
https://www.reddit.com/r/LocalLLaMA/comments/1py9p4r/llama_32_3b_running_on_my_geekom_it15/