home.social

#ollama — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ollama, aggregated by home.social.

  1. Локальный ассистент для зумов, часть 2: как граф встреч становится памятью и почему ему можно верить

    В первой части я собрал локального ассистента для созвонов: диаризация, стенограмма, граф знаний в Obsidian, ноль облаков. За полтора месяца ежедневной работы оказалось, что записать встречу — меньшая половина дела. Вторая часть — про то, как граф живёт неделями и почему ему можно верить: факты с датами «с» и «по», провенанс цитат из стенограммы, досье как этаж между поиском и графом, ночной прогон, песочница для облака, поиск по блокам и скрипт «забыть встречу целиком». Кода мало, устройства много. Читать дальше

    habr.com/ru/articles/1078828/

    #локальные_llm #ассистент_встреч #ollama #диаризация #граф_знаний #graphrag #obsidian #speechtotext #zoom #приватность

  2. Локальный ассистент для зумов, часть 2: как граф встреч становится памятью и почему ему можно верить

    В первой части я собрал локального ассистента для созвонов: диаризация, стенограмма, граф знаний в Obsidian, ноль облаков. За полтора месяца ежедневной работы оказалось, что записать встречу — меньшая половина дела. Вторая часть — про то, как граф живёт неделями и почему ему можно верить: факты с датами «с» и «по», провенанс цитат из стенограммы, досье как этаж между поиском и графом, ночной прогон, песочница для облака, поиск по блокам и скрипт «забыть встречу целиком». Кода мало, устройства много. Читать дальше

    habr.com/ru/articles/1078828/

    #локальные_llm #ассистент_встреч #ollama #диаризация #граф_знаний #graphrag #obsidian #speechtotext #zoom #приватность

  3. Локальный ассистент для зумов, часть 2: как граф встреч становится памятью и почему ему можно верить

    В первой части я собрал локального ассистента для созвонов: диаризация, стенограмма, граф знаний в Obsidian, ноль облаков. За полтора месяца ежедневной работы оказалось, что записать встречу — меньшая половина дела. Вторая часть — про то, как граф живёт неделями и почему ему можно верить: факты с датами «с» и «по», провенанс цитат из стенограммы, досье как этаж между поиском и графом, ночной прогон, песочница для облака, поиск по блокам и скрипт «забыть встречу целиком». Кода мало, устройства много. Читать дальше

    habr.com/ru/articles/1078828/

    #локальные_llm #ассистент_встреч #ollama #диаризация #граф_знаний #graphrag #obsidian #speechtotext #zoom #приватность

  4. @Madic noch darauf Bezug nehmen:

    Bei #ollama Cloud bekommst du in pro abo für 20$ echtes Geld 60$ token Guthaben... Also Faktor 3 - dann ist es ja dauerhaft 66% reduziert dort

  5. @Madic noch darauf Bezug nehmen:

    Bei #ollama Cloud bekommst du in pro abo für 20$ echtes Geld 60$ token Guthaben... Also Faktor 3 - dann ist es ja dauerhaft 66% reduziert dort

  6. @Madic noch darauf Bezug nehmen:

    Bei #ollama Cloud bekommst du in pro abo für 20$ echtes Geld 60$ token Guthaben... Also Faktor 3 - dann ist es ja dauerhaft 66% reduziert dort

  7. @Madic noch darauf Bezug nehmen:

    Bei Cloud bekommst du in pro abo für 20$ echtes Geld 60$ token Guthaben... Also Faktor 3 - dann ist es ja dauerhaft 66% reduziert dort

  8. Learn when to migrate from Ollama to vLLM. Migration signals, planning steps, Docker Compose setup, and a practical checklist for moving your local LLM server.

    #Ollama #vLLM #LLM #AI #Self-Hosting #Docker #API #DevOps

    glukhov.org/llm-hosting/compar

  9. Learn when to migrate from Ollama to vLLM. Migration signals, planning steps, Docker Compose setup, and a practical checklist for moving your local LLM server.

    #Ollama #vLLM #LLM #AI #Self-Hosting #Docker #API #DevOps

    glukhov.org/llm-hosting/compar

  10. Learn when to migrate from Ollama to vLLM. Migration signals, planning steps, Docker Compose setup, and a practical checklist for moving your local LLM server.

    #Ollama #vLLM #LLM #AI #Self-Hosting #Docker #API #DevOps

    glukhov.org/llm-hosting/compar

  11. Learn when to migrate from Ollama to vLLM. Migration signals, planning steps, Docker Compose setup, and a practical checklist for moving your local LLM server.

    #Ollama #vLLM #LLM #AI #Self-Hosting #Docker #API #DevOps

    glukhov.org/llm-hosting/compar

  12. Learn when to migrate from Ollama to vLLM. Migration signals, planning steps, Docker Compose setup, and a practical checklist for moving your local LLM server.

    -Hosting

    glukhov.org/llm-hosting/compar

  13. 📝 Daily report 📈

    Here are today's most popular trending hashtags #⃣ on our website 🌐️:

    #openai, #gnu, #sysadmin, #ollama

    🔥 Stay tuned! 🔥

  14. 📝 Daily report 📈

    Here are today's most popular trending hashtags #⃣ on our website 🌐️:

    #openai, #gnu, #sysadmin, #ollama

    🔥 Stay tuned! 🔥

  15. 📝 Daily report 📈

    Here are today's most popular trending hashtags #⃣ on our website 🌐️:

    #openai, #gnu, #sysadmin, #ollama

    🔥 Stay tuned! 🔥

  16. Most developers still think #Java + #AI means calling APIs from Spring Boot. The ecosystem goes deeper—from RAG frameworks to local inference with #Ollama or GPU workloads on #JVM.
    @ArturSkowronski maps the #GenAI tooling iceberg for modern #Java: javapro.io/2026/06/03/the-gen-

    @ollama

  17. Most developers still think #Java + #AI means calling APIs from Spring Boot. The ecosystem goes deeper—from RAG frameworks to local inference with #Ollama or GPU workloads on #JVM.
    @ArturSkowronski maps the #GenAI tooling iceberg for modern #Java: javapro.io/2026/06/03/the-gen-

    @ollama

  18. Most developers still think #Java + #AI means calling APIs from Spring Boot. The ecosystem goes deeper—from RAG frameworks to local inference with #Ollama or GPU workloads on #JVM.
    @ArturSkowronski maps the #GenAI tooling iceberg for modern #Java: javapro.io/2026/06/03/the-gen-

    @ollama

  19. Most developers still think #Java + #AI means calling APIs from Spring Boot. The ecosystem goes deeper—from RAG frameworks to local inference with #Ollama or GPU workloads on #JVM.
    @ArturSkowronski maps the #GenAI tooling iceberg for modern #Java: javapro.io/2026/06/03/the-gen-

    @ollama

  20. Most developers still think #Java + #AI means calling APIs from Spring Boot. The ecosystem goes deeper—from RAG frameworks to local inference with #Ollama or GPU workloads on #JVM.
    @ArturSkowronski maps the #GenAI tooling iceberg for modern #Java: javapro.io/2026/06/03/the-gen-

    @ollama

  21. أطلق Raycast تحديثه v2، الذي يعيد دعم "Bring Your Own Model" (BYOM) لربط مزودين متوافقين مع OpenAI، أو تشغيل نماذج محلية عبر Ollama، أو استخدام OpenRouter API ضمن Raycast AI. هذه الميزات تتطلب الآن اشتراك Raycast Pro. كما يوسع التحديث إدارة النوافذ بأوامر دقيقة لتغيير الحجم والتحريك، وإمكانية إنشاء تخطيطات مخصصة من النوافذ الحالية. ويشمل التحسينات دعم النص من اليمين إلى اليسار في Raycast AI، وتحسينات في Quick AI، وتعديلات على ترتيب البحث عن الملفات.

    #Raycast #AI #Ollama

  22. أطلق Raycast تحديثه v2، الذي يعيد دعم "Bring Your Own Model" (BYOM) لربط مزودين متوافقين مع OpenAI، أو تشغيل نماذج محلية عبر Ollama، أو استخدام OpenRouter API ضمن Raycast AI. هذه الميزات تتطلب الآن اشتراك Raycast Pro. كما يوسع التحديث إدارة النوافذ بأوامر دقيقة لتغيير الحجم والتحريك، وإمكانية إنشاء تخطيطات مخصصة من النوافذ الحالية. ويشمل التحسينات دعم النص من اليمين إلى اليسار في Raycast AI، وتحسينات في Quick AI، وتعديلات على ترتيب البحث عن الملفات.

    #Raycast #AI #Ollama

  23. A short recap of my recent experiments of running an LLM.

    Would be great to hear suggestions from others.
    - Where else can we get decent hardware?
    - What models do you run for your teams?
    - How to reduce complexity of the setup?

    linkedin.com/pulse/recap-my-jo

    #AI #LLM #ollama #openwebui #hetzner #trooperai #selfhosted

  24. A short recap of my recent experiments of running an LLM.

    Would be great to hear suggestions from others.
    - Where else can we get decent hardware?
    - What models do you run for your teams?
    - How to reduce complexity of the setup?

    linkedin.com/pulse/recap-my-jo

    #AI #LLM #ollama #openwebui #hetzner #trooperai #selfhosted

  25. A short recap of my recent experiments of running an LLM.

    Would be great to hear suggestions from others.
    - Where else can we get decent hardware?
    - What models do you run for your teams?
    - How to reduce complexity of the setup?

    linkedin.com/pulse/recap-my-jo

    #AI #LLM #ollama #openwebui #hetzner #trooperai #selfhosted

  26. A short recap of my recent experiments of running an LLM.

    Would be great to hear suggestions from others.
    - Where else can we get decent hardware?
    - What models do you run for your teams?
    - How to reduce complexity of the setup?

    linkedin.com/pulse/recap-my-jo

    #AI #LLM #ollama #openwebui #hetzner #trooperai #selfhosted

  27. A short recap of my recent experiments of running an LLM.

    Would be great to hear suggestions from others.
    - Where else can we get decent hardware?
    - What models do you run for your teams?
    - How to reduce complexity of the setup?

    linkedin.com/pulse/recap-my-jo

    #AI #LLM #ollama #openwebui #hetzner #trooperai #selfhosted

  28. Я хотел просто навести порядок в Obsidian. В итоге написал два индекса, semantic search и RAG

    Я начинал с простого AI-аудита заметок, а в итоге Vault Audit AI вырос в систему с двумя индексами, поиском по смыслу, Similar Notes, semantic duplicates и RAG по собственному хранилищу. В статье разбираю, как всё это устроено, что ломалось по дороге и почему почти 700 тестов всё равно не спасли от сюрпризов в реальном Obsidian.

    habr.com/ru/articles/1078328/

    #Obsidian #Vault_Audit_AI #RAG #semantic_search #embeddings #LLM #TypeScript #vector_search #knowledge_management #Ollama

  29. Я хотел просто навести порядок в Obsidian. В итоге написал два индекса, semantic search и RAG

    Я начинал с простого AI-аудита заметок, а в итоге Vault Audit AI вырос в систему с двумя индексами, поиском по смыслу, Similar Notes, semantic duplicates и RAG по собственному хранилищу. В статье разбираю, как всё это устроено, что ломалось по дороге и почему почти 700 тестов всё равно не спасли от сюрпризов в реальном Obsidian.

    habr.com/ru/articles/1078328/

    #Obsidian #Vault_Audit_AI #RAG #semantic_search #embeddings #LLM #TypeScript #vector_search #knowledge_management #Ollama

  30. Я хотел просто навести порядок в Obsidian. В итоге написал два индекса, semantic search и RAG

    Я начинал с простого AI-аудита заметок, а в итоге Vault Audit AI вырос в систему с двумя индексами, поиском по смыслу, Similar Notes, semantic duplicates и RAG по собственному хранилищу. В статье разбираю, как всё это устроено, что ломалось по дороге и почему почти 700 тестов всё равно не спасли от сюрпризов в реальном Obsidian.

    habr.com/ru/articles/1078328/

    #Obsidian #Vault_Audit_AI #RAG #semantic_search #embeddings #LLM #TypeScript #vector_search #knowledge_management #Ollama

  31. Агент написал себе навык и соврал, что тот работает

    Я делаю агента, который умеет дописывать себе навыки: упёрся в то, чего не умеет — сгенерировал код, собрал, установил, пользуется. Первый раз это сработало ровно так, как я мечтал. Агент сказал «у меня нет такого навыка», спросил разрешения, минуту пыхтел компилятором и отчитался: готово. Метода не было. То есть он был — в манифесте, в списке, который скилл про себя рассказывает. А в коде, который этот вызов должен обрабатывать, его не было. Модель объявила метод, забыла реализовать и об этом не узнала. Валидация прошла. Установка прошла. Проблема оказалась не в том, что модель ошиблась — это скучно и ожидаемо. Проблема в том, что я не смог поймать ошибку, потому что спрашивал у того же, кто её сделал. Скилл заявляет о себе сам, проверить заявление можно только у него же, а он в этом вопросе — заинтересованная сторона. Дальше — про то, как это чинится отрицательным контролем на девять строк, почему у проверки должно быть три исхода вместо двух, и зачем на самом деле нужен контейнер. Как я его поймал

    habr.com/ru/articles/1074880/

    #rust #llm #aiагенты #валидация #плагины #ollama #open_source #selfimprovement

  32. 📊 Requires #PHP 8.3+ and #Laravel 12, built on laravel/ai with Anthropic as default provider (#OpenAI, #Gemini, #Groq, #Ollama also work). Memory: none, file or database. MIT licensed. #devtools

    🌐 github.com/JordanDalton/larave

  33. От стримов к вебсокетам: как я боролся с буферизацией и наконец победил

    Привет. Меня зовут Николай Пискунов, я руководитель направления Big Data и эксперт курса Cloud DevSecOps по безопасной разработке от Академии вАЙТИ

    habr.com/ru/companies/beeline_

    #spring_ai #spring_boot #java #websocket #stomp #serversent_events #sse #ollama #llm #typescript

  34. Confused by the exploding number of #AI tools in the #JVM ecosystem? Teams mix #SpringAI, #LangChain4j, MCP & #Ollama without understanding the layers underneath. Artur Skowronski explains what each part of the #Java AI stack is actually for: javapro.io/2026/06/03/the-gen-

    @langchain4j

  35. Chat memory gets fuzzy fast once the UI hides what LangChain4j is actually retaining.

    I wrote a Quarkus tutorial that makes retained-memory pressure visible with `TokenWindowChatMemory`, Ollama request counts, a turn ledger, and OpenTelemetry attributes. The useful split is simple: your app-level eviction budget is not the model context limit. the-main-thread.com/p/quarkus- #Java #Quarkus #LangChain4j #Ollama #OpenTelemetry

  36. Local AI gets risky when the first confident answer becomes the system answer.

    I wrote a Quarkus tutorial that sends the same text to two Ollama models, uses Quarkus Signals to escalate only on disagreement, and keeps `UNCERTAIN` separate from `FAILED`. the-main-thread.com/p/quarkus- #Java #Quarkus #LangChain4j #Ollama

  37. Эволюция клиента для Ollama: от PostgreSQL к MongoDB

    Привет. Меня зовут Николай Пискунов, я руководитель направления Big Data и эксперт курса Cloud DevSecOps по безопасной разработке от Академии вАЙТИ

    habr.com/ru/companies/beeline_

    #java #spring_boot #postgresql #ollama #llm #artificial_intelligence #react #typescript #code_review #code_review_ai

  38. Confused by the exploding number of #AI tools in the #JVM ecosystem? Teams mix #SpringAI, #LangChain4j, MCP & #Ollama without understanding the layers underneath. Artur Skowronski explains what each part of the #Java AI stack is actually for: javapro.io/2026/06/03/the-gen-

    @langchain4j

  39. Cheap questions should not burn the same local model as real debugging work.

    I wrote a Quarkus + LangChain4j tutorial that classifies prompts, routes them between two Ollama models, and keeps the decision observable with CDI events and tests. the-main-thread.com/p/quarkus- #Java #Quarkus #LangChain4j #Ollama

  40. Running local AI for Ruby development with Ollama, Aider, and VSCode.

    No cloud APIs.
    No telemetry.
    No proprietary code leaving your machine.

    A practical look at local LLM workflows while working on Ruby-LibGD.

    rubystacknews.com/2026/05/28/r

    #ruby #rubylang #rails #ollama #ai #opensource

  41. Once a tool-calling assistant grows from 5 tools to 50, the problem stops being “prompting” and starts being context geometry.

    This walkthrough builds a Quarkus + LangChain4j + Ollama example and shows what tool search actually changes: smaller working sets, visible search rounds, and more prompt headroom even when local latency is messy.

    the-main-thread.com/p/langchai

    #Quarkus #LangChain4j #Ollama #Java

  42. RAG: Как собрать свой ретривер для особых случаев

    С опытом у RAG-инженера накапливается солидный багаж эвристик и инструментов, которые в определенных задачах превосходят по качеству или скорости стандартные. Фраза «а для этого у меня есть собственный ретривер» звучит с некоторым снобизмом, но добавляет к профессионализму несколько пойнтов. Хотите в свою коллекцию ретривер, который умеет работать с терминами, плохо различимыми в векторном пространстве эмбеддинга, в частности с именами и названиями? Тогда давайте перейдём от снобизма к практике. Начнём с обработки текста и сегментируем его на фрагменты - «чанки». Далее сделаем TFIDF модель, добавим поиск и обернём всё это в ретривер LangChain. Наконец сравним наш ретривер с двумя-тремя стандартными решениями. А Ollama поможет с вопросами для бенчмарка.

    habr.com/ru/articles/1022244/

    #rag #rag_pipeline #text_mining #text_generation #retrieval #ollama #gensim #langchain

  43. Do we have any owners of one of those Ryzen AI Max+ 395 128GB UMA boxes here that operate them on the daily for at least a few months as a claude LLM coding server and are capable of giving a comparative run down on their performance vs the OG claude and its collection of formal prose generators?

    Also: Especially curious to hear any numbers that came out of a watt meter in daily consumption and base/peak numbers. Same with the used models, their size, their respective achieved tok/s and response times.

    And should you have had the opportunity of comparing this against non-UMA beefy dGPUs on the above parameters that'd also be quote interesting.

    #claude #aicoding #AIAsssisted #ollama #onprem #selfhosing #StrixPoint #ryzenaimaxplus395 #ryzenAiMax #powerconsumption #costefficiency #uma