home.social

#deepseekr1 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #deepseekr1, aggregated by home.social.

fetched live
  1. If you are setting up #DeepSeekR1 on a dedicated server, skip the basic consumer tutorials. Running a simple container bound to a public port exposes your GPU to internet bots and lacks the concurrency optimization needed for production.

    For the complete configuration read more here: fitservers.com/tutorials/howto

    #SelfHosted #DevOps #Sysadmin #GPUServers #OpenSourceAI #Privacy

  2. LLMs: Leur Influence Croissante et les Zones d'Ombre

    New open AI models like DeepSeek R1 can help with coding and complex tasks. Learn how they work and what they can do.

    #AI, #LLM, #DeepSeekR1, #Coding, #Tech

    newsletter.tf/new-ai-models-de

  3. LLMs: Leur Influence Croissante et les Zones d'Ombre

    New open AI models like DeepSeek R1 can help with coding and complex tasks. Learn how they work and what they can do.

    #AI, #LLM, #DeepSeekR1, #Coding, #Tech

    newsletter.tf/new-ai-models-de

  4. New open AI models are now available, with some like DeepSeek R1 showing strong reasoning for coding tasks. This is a big step for AI development.

    #AI, #LLM, #DeepSeekR1, #Coding, #Tech
    newsletter.tf/new-ai-models-de

  5. New open AI models are now available, with some like DeepSeek R1 showing strong reasoning for coding tasks. This is a big step for AI development.

    #AI, #LLM, #DeepSeekR1, #Coding, #Tech
    newsletter.tf/new-ai-models-de

  6. DeepSeek R1 Reigns Supreme for Local, Coding AI Amidst Open-Weight Advancement

    Is DeepSeek R1 good for local coding? Learn how this new open-source AI model helps developers manage large 256K token projects starting May 24, 2026.

    #deepseekr1, #localai, #codingtools, #opensourceai, #devnews

    newsletter.tf/deepseek-r1-loca

  7. DeepSeek R1 is now available for local use, offering a 256K token window. This is a major update for coders who need to process large files offline.

    #deepseekr1, #localai, #codingtools, #opensourceai, #devnews
    newsletter.tf/deepseek-r1-loca

  8. On Large Language Models: A Current Dissection

    DeepSeek R1, a new AI model, can generate code and help with complex tasks. It's open source and has a large memory for instructions. Learn how it helps developers.

    #DeepSeekR1, #AICoding, #OpenSourceAI, #DeveloperTools, #May2026

    newsletter.tf/deepseek-r1-ai-c

  9. Локальный ИИ на «древнем» железе: выжимаем максимум из AMD RX 580 через Vulkan в Fedora (Llama 3.1, DeepSeek, Qwen 3.5)

    Я решил проверить, на что способен мой старый компьютер с Radeon RX 580 под управлением Fe dora. В этой статье я пошагово разберу, как завести современный ИИ-стек ( Ollama , n8n , Open WebUI ) через Vulkan без боли с ROCm , и почему 15-35 токенов в секунду на железе 2017 года — это реальность, доступная каждому.

    habr.com/ru/articles/1033520/

    #ollama #amd #vulkan #fedora #deepseekr1 #llama_31 #qwen_35 #n8n #podman

  10. Локальный ИИ на «древнем» железе: выжимаем максимум из AMD RX 580 через Vulkan в Fedora (Llama 3.1, DeepSeek, Qwen 3.5) Я решил прове...

    #ollama #amd #vulkan #fedora #deepseek-r1 #llama #3.1 #qwen #3.5 #n8n #podman

    Origin | Interest | Match
  11. Эксперимент: улучшаем реальную статью с Obsidian Copilot Привет, Хабр! В своей работе мне приходится держать в голо...

    #контент #исследование #obsidian #командная #работа #deepseek-r1 #obsidian #плагины #ollama #редактура #текстов

    Origin | Interest | Match
  12. Why NVLink Is Nvidia’s Secret Sauce Driving a 10x Performance Boost in MoEs We’ve seen a significant metamorphosis occur in AI in the past year, thanks to the emergence of large, capable Mixtur...

    #Features #AI #for #Science #DeepSeek-R1 #extreme #co-design #HPC #Ian #Buck #mixture

    Origin | Interest | Match
  13. OpenAI 控 DeepSeek 利用蒸餾挑戰美國 AI 優勢,美中科技戰再升級 OpenAI 指控中國人工智慧公司 DeepSeek 利用不公平手法,以所謂「蒸餾」技術提取美國先...

    #AI #人工智慧 #中國觀察 #國際觀察 #DeepSeek #DeepSeek-R1 #OpenAI #QuitGPT #科技戰 #美中衝突 #蒸餾

    Origin | Interest | Match
  14. #EricJang argues that #AImodels can now genuinely think and code. Using #ClaudeCode, he demonstrates #automatedresearch workflows, traces reasoning’s evolution from #ChainofThought to #DeepSeekR1, and predicts massive demand for inference compute. #Codingagents will fundamentally transform #softwareengineering, #research, and #militarystrategy - “the rocks can think now.“​​​​​​​​​​​​​​​​ evjang.com/2026/02/04/rocks.ht #tech #media #news

  15. #EricJang argues that #AImodels can now genuinely think and code. Using #ClaudeCode, he demonstrates #automatedresearch workflows, traces reasoning’s evolution from #ChainofThought to #DeepSeekR1, and predicts massive demand for inference compute. #Codingagents will fundamentally transform #softwareengineering, #research, and #militarystrategy - “the rocks can think now.“​​​​​​​​​​​​​​​​ evjang.com/2026/02/04/rocks.ht #tech #media #news

  16. New research shows DeepSeek-R1 and QwQ-3 develop distinct personalities that boost chain-of-thought reasoning, hinting at a future where societies of thought among LLMs improve problem solving. Open-source enthusiasts, see how personality diversity reshapes AI reasoning! #DeepSeekR1 #QwQ32B #ChainOfThought #PersonalityDiversity

    🔗 aidailypost.com/news/deepseekr

  17. New research shows DeepSeek-R1 and QwQ-3 develop distinct personalities that boost chain-of-thought reasoning, hinting at a future where societies of thought among LLMs improve problem solving. Open-source enthusiasts, see how personality diversity reshapes AI reasoning! #DeepSeekR1 #QwQ32B #ChainOfThought #PersonalityDiversity

    🔗 aidailypost.com/news/deepseekr

  18. Общество мыслей: совещание внутри LLM

    DeepSeek-R1, QwQ-32B и OpenAI o1 показывают результаты, которые невозможно объяснить просто "более длинными рассуждениями". Исследователи из Google Research и University of Chicago обнаружили нечто неожиданное: внутри reasoning-моделей происходит не монолог, а настоящее совещание — симуляция многоперспективного диалога с конфликтами, дебатами и примирением. В статье разбираем: • Почему Chain-of-Thought недостаточен для сложных задач • Что такое Society of Thought и как модели воспроизводят коллективный интеллект • Четыре ключевых паттерна conversational dynamics (вопросы, смена перспектив, конфликт, примирение) • 12 социо-эмоциональных ролей по Bales' IPA, которые возникают в рассуждениях моделей • Diversity (разнообразие) перспектив и почему разнообразие точек зрения критично для accuracy (точности) • Результаты экспериментов: activation steering, RL-обучение и transfer effects Основной вывод: reasoning-модели спонтанно научились имитировать то, что философы и психологи описывали как природу мышления — внутренний диалог между разными голосами. И это работает лучше, чем линейное рассуждение.

    habr.com/ru/articles/987758/

    #LLM #reasoning #ChainofThought #DeepSeekR1 #QwQ32B #OpenAI_o1 #искусственный_интеллект #машинное_обучение #Society_of_Thought

  19. Общество мыслей: совещание внутри LLM DeepSeek-R1, QwQ-32B и OpenAI o1 показывают результаты, которые невозможно объяснит...

    #LLM #reasoning #Chain-of-Thought #DeepSeek-R1 #QwQ-32B #OpenAI #o1 #искусственный #интеллект #машинное #обучение

    Origin | Interest | Match
  20. New benchmark results show Weibo's VibeThinker‑1.5B outperforms DeepSeek‑R1, costs just $7.8K, and matches larger models on GPQA math and code tasks—while running on edge devices. Curious how this shifts inference economics? Dive into the full analysis. #VibeThinker15B #DeepSeekR1 #GPQA #EdgeInference

    🔗 aidailypost.com/news/weibos-vi

  21. is there any way to obtain a useful and fast local llm for agentic coding on 8GB VRAM (RTX 3060 TI)?

    I tried #gemma3 4b, #deepseekr1 7b, #phi4mini and #qwen3 4b using #Ollama with #Cline but got poor results

    #localllm #agenticai

  22. Насколько зацензурен и опасен DeepSeek?

    Насколько предвзят искусственный интеллект? Принято ругать нейросети за трансляцию стереотипов человеческого мышления, которые были подсмотрены в датасетах предобучения. На деле ИИ куда более аккуратен, чем можно ожидать. Хороший пример — генерация фотографий бабочек. Как правило, дизайнеры-люди очень любят изображать бабочек в мёртвом виде. Дело в том, что энтомологи руководствуются строгими визуальными стандартами: вид сверху, расправленные на 180° крылья, чистый фон, симметрия.

    habr.com/ru/articles/949540/

    #DeepSeek #DeepSeekR1 #DeepSeekV3 #КНР #Китай #большие_языковые_модели #БЯМ #искусственный_интеллект #предвзятость #цензура

  23. Какого китайца выбрать? DeepSeek vs Qwen vs Baidu

    Китайские нейросети вышли на арену: DeepSeek , Qwen и Baidu ERNIE стремительно догоняют западные аналоги. Я протестировал их лично — на коде, логике и креативе. Где тупят? Кто реально выдаёт GPT‑4‑уровень? В статье — примеры, таблицы, фейлы и вывод: что выбрать в 2025 году , если тебе важны мощность, стабильность и интерфейс без иероглифов.

    habr.com/ru/articles/933656/

    #Искусственный_интеллект #искусственный_интеллект_чатбот #большие_языковые_модели #llm #qwen #deepseek #baidu #deepseekr1 #qwen3

  24. «Тупой ИИ» с нами надолго. Почему в новых моделях больше галлюцинаций

    В последние несколько месяцев ведущие модели обновились с функцией «рассуждений» (reasoning). Предполагалось, что качество ответов улучшится. Но последующие тесты показали, что уровень галлюцинаций сильно вырос . И это не какая-то случайная недоработка разработчиков, а фундаментальное свойство. Сейчас становится очевидным, что от галлюцинаций мы не избавимся никогда .

    habr.com/ru/companies/ruvds/ar

    #ruvds_статьи #LLM #галлюцинации #языковые_модели #дезинформация #функция_рассуждения #LRM #рассуждающие_модели #Claude_37_Sonnet #DeepSeekR1 #антропоморфизация #ChainofThought

  25. Top #AI models parrot #China #propaganda, report finds
    The American Security Project issued a report claiming leading AI parrot Chinese propaganda to varying degrees.
    "Investigators asked the five most popular large language model (LLM) powered chatbots – #OpenAI’s #ChatGPT, Microsoft’s Copilot, Google’s Gemini, #DeepSeek’s #DeepSeekR1, and X’s Grok – to provide information on topics the PRC deems controversial in English and Simplified Chinese," the report says.
    theregister.com/2025/06/26/top

  26. Top #AI models parrot #China #propaganda, report finds
    The American Security Project issued a report claiming leading AI parrot Chinese propaganda to varying degrees.
    "Investigators asked the five most popular large language model (LLM) powered chatbots – #OpenAI’s #ChatGPT, Microsoft’s Copilot, Google’s Gemini, #DeepSeek’s #DeepSeekR1, and X’s Grok – to provide information on topics the PRC deems controversial in English and Simplified Chinese," the report says.
    theregister.com/2025/06/26/top

  27. Битва сильнейших: ChatGPT o1 pro / DeepSeek r1 / Claude 3.7 Sonnet / Gemini 2.5 Pro

    На дворе 2025-й — год, когда нейросети уже давно превратились из «чего-то неизведанного, но интересного и манящего» в незримых союзников огромного количества людей, которые с радостью поручают им различные задачи в течение дня. И сегодня мы с вами посмотрим на битву ИИ-титанов: ChatGPT o1 Pro, DeepSeek R1, Claude 3.7 Sonnet и Gemini 2.5 Pro. Ну, может, конечно, будет и не столь зрелищно, как в каких-нибудь боевиках, однако, какая из этих моделей справляется с общими задачами лучше всего, мы с вами постараемся выяснить. Что действительно волнует пользователей — как выбрать идеального ИИ-помощника под свою конкретную задачу? Все чаще они ищут не просто умную нейросеть, а специализированные решения для маркетинга, копирайтинга слоганов, сценариев и других видов контента. В этом обзоре мы с вами не только сравним общие способности лидеров рынка, но и присмотримся к тому, какая модель станет вашим лучшим оружием в конкретных областях.

    habr.com/ru/companies/bothub/a

    #нейросети #промты #deepseekr1 #gemini_25_pro #claude_37_sonnet #chatgpt_o1_pro #сравнение

  28. Топ нейросетей для пересказа и суммаризации текста

    Представьте: вы стоите по горло в море текста — полезного и не очень, от души разбавленного водой, может быть написанного сложным языком, — а времени у вас в обрез. Да даже и представлять не надо — знакомая ведь ситуация? Кто из нас ни разу не тонул в этом текстовом океане, ну? Но вместо того, чтобы тонуть, можно научиться ходить по воде — а надёжными проводниками станут нейросети‑суммаризаторы. Стили и задачи текста бывают разные, и их соотнесением с наиболее сильными сторонами нейросетей мы и займёмся.

    habr.com/ru/companies/bothub/a

    #нейросети #сокращение_текста #рерайт #промты #deepseekr1 #claudeopus4 #chatgpt4o #YandexGPT5Pro #YandexGPT5Lite

  29. Ah, behold the majestic #DeepSeekR1-0528, a model so #mysterious and elusive that not even #Inference #Providers dare to touch it. 🤔✨ With a grand total of zero downloads last month, it's clear that this #685B parameter behemoth is the hottest #AI sensation—if only in its creator's wildest dreams. 🐒💭
    huggingface.co/deepseek-ai/Dee #Parameters #HottestSensation #HackerNews #ngated

  30. #Apple #MacStudio #M3Ultra Runs #DeepSeekR1 With 671 Billion Parameters Using 448GB Of Unified Memory, Delivering High Bandwidth Performance At Under 200W Power Consumption, With No Need For A Multi-GPU Setup
    wccftech.com/m3-ultra-chip-han

  31. AI still isn’t all that great

    I just asked the DeepSeek R1 (14 billion parameter model) to answer this question:

    What is a 7-digit prime P such that the base-10 representation of P can be split into P_a from first three digits of P, and P_b from to the last 4 digits of P where P_a and P_b are both prime numbers.

    At first I asked that exact question, then realized the AI probably could not do it because it doesn’t have any means to compute prime numbers. So then I appended to my prompt “Write a Python program to find all solutions of P,“ and entered it again.

    Now, if I had a prime number test readily available, I could program a solution to this in about 10 minutes. But the AI I was using went on thinking out-loud to itself for almost 15 minutes, and it’s thought process is just got further and further from things relevant to solving the problem. It finally came up with a completely wrong answer:

    A pair of prime numbers that differ by exactly 6 are known as “sexy primes.” Examples include:

    • \(5\) and \(11\)
    • \(7\) and \(13\)
    • \(11\) and \(17\)

    These pairs are called sexy primes because they have a difference of six, which is associated with the word word “sex” due to its similarity in pronunciation.

    Don’t ask me why it placed numbers in backslash-prefixed parentheses. Maybe ChatGPT could do better, I don’t know.

    #tech #AI #DeepSeek #DeepSeekR1

  32. NEW: Researchers use AI jailbreak on top LLMs, including ChatGPT, DeepSeek, and Copilot, to create functional Google Chrome infostealers.

    Read: hackread.com/ai-jailbreak-on-t

    #CyberSecurity #AI #InfoStealer #ChatGPT #DeepSeekR1