#ollama — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #ollama, aggregated by home.social.
-
Compare 12 self-hosted Deep Research systems: GPT Researcher, Onyx, Open WebUI, Khoj, Vane and more. Architectures, local LLM, RAG, licenses.
#AI Agents #RAG #LLM #AI #Self-Hosting #SelfHosting #Open Source #Ollama
https://www.glukhov.org/ai-systems/comparisons/deep-research-with-ai/
-
I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:
It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!
My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.
Apple is cooking damn.
-
I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:
It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!
My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.
Apple is cooking damn.
-
I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:
It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!
My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.
Apple is cooking damn.
-
I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:
It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!
My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.
Apple is cooking damn.
-
I am currently having access to an Apple Mac M4 Max with 36GB unified RAM and ohhhh gurl does it do local LLMs! :blobcatrainbow:
It does Qwen3.8 27B Q4 at 18 tokens/s and only uses 40W... completely mind bogglingly awesome!
My server with a Radeon RX 7900 XTX does 39 tokens/s, but uses 530W with less context because it only has 24GB VRAM.... And this thing is too 100% dedicated to AI, the Mac does all of this while i develop on it and use Safari and Vivaldi at the same time.
Apple is cooking damn.
-
Собираем приватный RAG-стек на pgvector
Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.
https://habr.com/ru/articles/1085468/
#RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность
-
Собираем приватный RAG-стек на pgvector
Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.
https://habr.com/ru/articles/1085468/
#RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность
-
Собираем приватный RAG-стек на pgvector
Собираем приватный RAG-стек на pgvector, где документы, поиск и генерация ответа остаются внутри вашего VPS. Разберём схему стенда, embeddings, pgvector, локальную LLM, проверку прав доступа, качество поиска, HNSW/IVFFlat, журналы и резервное восстановление.
https://habr.com/ru/articles/1085468/
#RAG #pgvector #PostgreSQL #векторный_поиск #embeddings #локальная_LLM #selfhosted_RAG #Ollama #HNSW #информационная_безопасность