home.social

#ollama — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ollama, aggregated by home.social.

  1. Compare 12 self-hosted Deep Research systems: GPT Researcher, Onyx, Open WebUI, Khoj, Vane and more. Architectures, local LLM, RAG, licenses.

    Agents -Hosting Source

    glukhov.org/ai-systems/compari

  2. I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:

    It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!

    My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.

    Apple is cooking damn.

    #llm #ai #localAi #apple #appleSilicon #ollama #LocalLLaMA

  3. I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:

    It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!

    My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.

    Apple is cooking damn.

    #llm #ai #localAi #apple #appleSilicon #ollama #LocalLLaMA

  4. I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:

    It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!

    My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.

    Apple is cooking damn.

    #llm #ai #localAi #apple #appleSilicon #ollama #LocalLLaMA

  5. I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:

    It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!

    My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.

    Apple is cooking damn.

    #llm #ai #localAi #apple #appleSilicon #ollama #LocalLLaMA

  6. I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:

    It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!

    My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.

    Apple is cooking damn.

    #llm #ai #localAi #apple #appleSilicon #ollama #LocalLLaMA

  7. Собираем приватный RAG-стек на pgvector

    Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.

    habr.com/ru/articles/1085468/

    #RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность

  8. Собираем приватный RAG-стек на pgvector

    Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.

    habr.com/ru/articles/1085468/

    #RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность

  9. Собираем приватный RAG-стек на pgvector

    Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.

    habr.com/ru/articles/1085468/

    #RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность