#deepseekv3 — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #deepseekv3, aggregated by home.social.
-
RT @dunik_7: TRANSLASATION: Ein Labor der Tsinghua-Universität hat ein Projekt auf GitHub veröffentlicht, das einen H100-Rack im Wert von 400.000 US-Dollar durch eine einzelne 24-GB-Grafikkarte ersetzt. Das Projekt heißt ktransformers, und der Trick ist fast schon lächerlich einfach: Die Experten, die Sie tatsächlich nutzen, bleiben auf der GPU, während die anderen auf der CPU warten, bis sie benötigt werden. / DeepSeek-V3 und R1 mit 139K Kontext in 24GB VRAM / bis zu 28-fache Geschwindigkeitssteigerung gegenüber dem Standard-Setup / Fine-Tuning von DeepSeek-V3 über vier RTX 4090 statt eines Rechenzentrums / entwickelt vom MADSys-Labor der Tsinghua-Universität, nicht von einem Startup mit einer Landing Page. Apache 2.0, bereits über 17.000 Sterne. - http://github.com/kvcache-ai/ktransformers merken.
mehr auf Arint.info
#AIResearch #DeepSeekV3 #ktransformers #MachineLearning #OpenSource #TsinghuaUniversity #arint_info
-
RT @dunik_7: TRANSLASATION: Ein Labor der Tsinghua-Universität hat ein Projekt auf GitHub veröffentlicht, das einen H100-Rack im Wert von 400.000 US-Dollar durch eine einzelne 24-GB-Grafikkarte ersetzt. Das Projekt heißt ktransformers, und der Trick ist fast schon lächerlich einfach: Die Experten, die Sie tatsächlich nutzen, bleiben auf der GPU, während die anderen auf der CPU warten, bis sie benötigt werden. / DeepSeek-V3 und R1 mit 139K Kontext in 24GB VRAM / bis zu 28-fache Geschwindigkeitssteigerung gegenüber dem Standard-Setup / Fine-Tuning von DeepSeek-V3 über vier RTX 4090 statt eines Rechenzentrums / entwickelt vom MADSys-Labor der Tsinghua-Universität, nicht von einem Startup mit einer Landing Page. Apache 2.0, bereits über 17.000 Sterne. - http://github.com/kvcache-ai/ktransformers merken.
mehr auf Arint.info
#AIResearch #DeepSeekV3 #ktransformers #MachineLearning #OpenSource #TsinghuaUniversity #arint_info
-
RT @dunik_7: TRANSLASATION: Ein Labor der Tsinghua-Universität hat ein Projekt auf GitHub veröffentlicht, das einen H100-Rack im Wert von 400.000 US-Dollar durch eine einzelne 24-GB-Grafikkarte ersetzt. Das Projekt heißt ktransformers, und der Trick ist fast schon lächerlich einfach: Die Experten, die Sie tatsächlich nutzen, bleiben auf der GPU, während die anderen auf der CPU warten, bis sie benötigt werden. / DeepSeek-V3 und R1 mit 139K Kontext in 24GB VRAM / bis zu 28-fache Geschwindigkeitssteigerung gegenüber dem Standard-Setup / Fine-Tuning von DeepSeek-V3 über vier RTX 4090 statt eines Rechenzentrums / entwickelt vom MADSys-Labor der Tsinghua-Universität, nicht von einem Startup mit einer Landing Page. Apache 2.0, bereits über 17.000 Sterne. - http://github.com/kvcache-ai/ktransformers merken.
mehr auf Arint.info
#AIResearch #DeepSeekV3 #ktransformers #MachineLearning #OpenSource #TsinghuaUniversity #arint_info
-
DeepSeek、史上最長の13時間障害——次世代モデル「V4」、いよいよ来るのか
-
DeepSeek、史上最長の13時間障害——次世代モデル「V4」、いよいよ来るのか
-
DeepSeek-V3 from Scratch: Mixture of Experts (MoE) Table of Contents DeepSeek-V3 from Scratch: Mixture of Experts (MoE) The Scaling Challenge in Neural Networks Mixture of Experts (MoE): Mathematic...
#Deep #Learning #DeepSeek #Machine #Learning #Neural #Networks #Tutorial #deepseek-v3 #expert #routing
Origin | Interest | Match -
DeepSeek-V3 from Scratch: Mixture of Experts (MoE) Table of Contents DeepSeek-V3 from Scratch: Mixture of Experts (MoE) The Scaling Challenge in Neural Networks Mixture of Experts (MoE): Mathematic...
#Deep #Learning #DeepSeek #Machine #Learning #Neural #Networks #Tutorial #deepseek-v3 #expert #routing
Origin | Interest | Match -
日本樂天推自家「AI 3.0」模型 源碼竟顯示使用 DeepSeek 基礎模型
樂天集團 (Rakuten) 3 月 17 日公開旗下最新日語大型語言模型「Rakuten AI 3.0」,惟技術人員隨即發現 Hugging Face 上的設定檔案顯示其架構與中國 AI 公司 DeepSeek 的 DeepSeek-V3 模型高度吻合,兼且發布時被指悄然移除 DeepSeek-V3 原有開源授權聲明,觸發開源社群強烈批評,樂天面對查詢時拒絕披露基礎模型來源,僅稱「非公開」。
#人工智能 #AI #DeepSeek #DeepSeek-V3
https://unwire.hk/2026/03/22/rakuten-ai-3-deepseek-v3-open-source-controversy/ai/?utm_source=rss&utm_medium=rss&utm_campaign=rakuten-ai-3-deepseek-v3-open-source-controversy -
Build DeepSeek-V3: Multi-Head Latent Attention (MLA) Architecture Table of Contents Build DeepSeek-V3: Multi-Head Latent Attention (MLA) Architecture The KV Cache Memory Problem in DeepSeek-V3 Mult...
#Deep #Learning #Large #Language #Models #PyTorch #Transformers #Tutorial #attention #mechanisms #deepseek-v3
Origin | Interest | Match -
DeepSeek-V3 Model: Theory, Config, and Rotary Positional Embeddings Table of Contents DeepSeek-V3 Model: Theory, Config, and Rotary Positional Embeddings Introduction to the DeepSeek-V3 Model The F...
#DeepSeek-V3 #KV #Cache #MultiHead #Latent #Attention #RoPE #Tutorial #deepseekv3 #kv #cache
Origin | Interest | Match -
AI Deepseek Assistant - Pro #Chatgpt #Deepseekv3 #Ai #Textchatwithai #Deepseek #Api #Deepseekapi #Chat #Deepseek #Aimodel #Gpt #Frostweepgames #Gpt4 #AssetStore
-
r/LocalLLaMA tổng kết 2025: Một năm đột phá của AI nguồn mở! DeepSeek V3 khởi xướng "Open Source Strike Back", khiến Meta "hoảng loạn" và Sam Altman phải lên tiếng. Trung Quốc dẫn đầu với các mô hình mạnh mẽ như Qwen 3 và GPU giá phải chăng, trong khi Llama 4 của Meta bị đánh giá thấp. Cộng đồng LocalLLaMA vẫn là nơi thảo luận LLM chất lượng.
#AIOpenSource #DeepSeek #Qwen #LocalLLaMA #AInguonmo #DeepSeekV3 #ThongkeAI
https://www.reddit.com/r/LocalLLaMA/comments/1ptr3lv/rlocalllama_a_year_in_re
-
Beating GPT-5: DeepSeekMath-V2 Self-Corrects Logic Errors Presentational View Introduction Mathematics with the aid of artificial intelligence, is advancing rapidly. Innovations such as informal th...
#ai-in-mathematics #deepseekmath-v2 #deepseek-v3 #open-source-ai-model #theorem-proving
Origin | Interest | Match -
New benchmark shows Gemini 3 Pro outpaces Gemini 2.5 in trust, ethics and safety—69% vs 16%. The study, led by Phelim Bradley and Prolific, also pits DeepSeek V3 against the models, highlighting gaps in performance and reasoning. Dive into the full analysis for the numbers and implications. #Gemini3Pro #Gemini2_5 #DeepSeekV3 #TrustAndSafety
🔗 https://aidailypost.com/news/gemini-3-pro-tops-trust-ethics-safety-69-vs-16-gemini-25
-
🚀 Welcome GLM-4.6 the Latest flagship #opensource #AI #llm with advanced agentic, reasoning & coding capabilities
⚡ Performance improvements over #GLM45 with competitive advantages against #DeepSeekV3 and #ClaudeSonnet4 across 8 public benchmarks covering agents, reasoning & coding
🧵 👇
-
🚀 Welcome GLM-4.6 the Latest flagship #opensource #AI #llm with advanced agentic, reasoning & coding capabilities
⚡ Performance improvements over #GLM45 with competitive advantages against #DeepSeekV3 and #ClaudeSonnet4 across 8 public benchmarks covering agents, reasoning & coding
🧵 👇
-
🚀 Welcome GLM-4.6 the Latest flagship #opensource #AI #llm with advanced agentic, reasoning & coding capabilities
⚡ Performance improvements over #GLM45 with competitive advantages against #DeepSeekV3 and #ClaudeSonnet4 across 8 public benchmarks covering agents, reasoning & coding
🧵 👇
-
🚀 Welcome GLM-4.6 the Latest flagship #opensource #AI #llm with advanced agentic, reasoning & coding capabilities
⚡ Performance improvements over #GLM45 with competitive advantages against #DeepSeekV3 and #ClaudeSonnet4 across 8 public benchmarks covering agents, reasoning & coding
🧵 👇
-
🚀 Welcome GLM-4.6 the Latest flagship #opensource #AI #llm with advanced agentic, reasoning & coding capabilities
⚡ Performance improvements over #GLM45 with competitive advantages against #DeepSeekV3 and #ClaudeSonnet4 across 8 public benchmarks covering agents, reasoning & coding
🧵 👇
-
DeepSeek: Everything you need to know about the AI chatbot app
-
Насколько зацензурен и опасен DeepSeek?
Насколько предвзят искусственный интеллект? Принято ругать нейросети за трансляцию стереотипов человеческого мышления, которые были подсмотрены в датасетах предобучения. На деле ИИ куда более аккуратен, чем можно ожидать. Хороший пример — генерация фотографий бабочек. Как правило, дизайнеры-люди очень любят изображать бабочек в мёртвом виде. Дело в том, что энтомологи руководствуются строгими визуальными стандартами: вид сверху, расправленные на 180° крылья, чистый фон, симметрия.
https://habr.com/ru/articles/949540/
#DeepSeek #DeepSeekR1 #DeepSeekV3 #КНР #Китай #большие_языковые_модели #БЯМ #искусственный_интеллект #предвзятость #цензура
-
https://technologiesinternetz.blogspot.com/2025/08/deepseek-v31-vs-gpt-5-vs-claude-41.html
DeepSeek V3.1 vs GPT-5 vs Claude 4.1: Which LLM Delivers the Best Value to Users?
#deepseekv3.1 #gpt5 #claude4.1 #LLM
-
https://technologiesinternetz.blogspot.com/2025/08/deepseek-v31-vs-gpt-5-vs-claude-41.html
DeepSeek V3.1 vs GPT-5 vs Claude 4.1: Which LLM Delivers the Best Value to Users?
#deepseekv3.1 #gpt5 #claude4.1 #LLM
-
https://www.europesays.com/uk/368458/ DeepSeek V3.1 Released: The Intriguing UE8M0 FP8 #Computing #DeepSeekV3.1 #DomesticAIIndustry #EnflameTechnology #FloatingPointNumbers #FP16 #FP32 #FP8 #HigherThinkingEfficiency #HybridInferenceArchitecture #L600Chip #MagicStoneXiYunC600 #MXFP8 #ParameterPrecision #SoftwareHardwareCollaboration #StrongerAgentCapability #Technology #UE8M0FP8 #UK #UnitedKingdom
-
DeepSeek V3.1 Released: The Intriguing UE8M0 FP8
DeepSeek has launched version V3.1. Let’s briefly go through the highlights: Hybrid Infe…
#NewsBeep #News #Computing #AU #Australia #DeepSeekV3.1 #domesticAIindustry #EnflameTechnology #floatingpointnumbers #FP16 #FP32 #FP8 #HigherThinkingEfficiency #HybridInferenceArchitecture #L600chip #MagicStoneXiYunC600 #MXFP8 #parameterprecision #software-hardwarecollaboration #StrongerAgentCapability #Technology #UE8M0FP8
https://www.newsbeep.com/au/87765/ -
New DeepSeek-R1T-Chimera Model Merges R1 Reasoning With Efficiency of V3-0324
#AI #LLMs #DeepSeekR1 #DeepSeekV3 #Chimera #OpenSourceAI #TNGTech #MoE #MachineLearning #TechNews #GenAI
-
New DeepSeek-R1T-Chimera Model Merges R1 Reasoning With Efficiency of V3-0324
#AI #LLMs #DeepSeekR1 #DeepSeekV3 #Chimera #OpenSourceAI #TNGTech #MoE #MachineLearning #TechNews #GenAI
-
🧩 #Llama4Maverick nutzt 128 Experten für deutlich mehr Rechenleistung und schlägt sogar #GPT4o und #Gemini20 in Benchmarks – bei nur der Hälfte der aktiven Parameter von #DeepSeekv3.
🎓 Beide #KIModelle wurden mithilfe des riesigen Lehrmodells #Llama4 Behemoth trainiert, das mit 288 Milliarden aktiven Parametern zu den leistungsstärksten weltweit zählt.
👉 https://eicker.TV #Technik #Medien #Politik #Wirtschaft (2/2)
-
🧩 #Llama4Maverick nutzt 128 Experten für deutlich mehr Rechenleistung und schlägt sogar #GPT4o und #Gemini20 in Benchmarks – bei nur der Hälfte der aktiven Parameter von #DeepSeekv3.
🎓 Beide #KIModelle wurden mithilfe des riesigen Lehrmodells #Llama4 Behemoth trainiert, das mit 288 Milliarden aktiven Parametern zu den leistungsstärksten weltweit zählt.
👉 https://eicker.TV #Technik #Medien #Politik #Wirtschaft (2/2)
-
Benchmarks Find ‘DeepSeek-V3-0324 Is More Vulnerable Than Qwen2.5-Max’ – Source: www.techrepublic.com https://ciso2ciso.com/benchmarks-find-deepseek-v3-0324-is-more-vulnerable-than-qwen2-5-max-source-www-techrepublic-com/ #threatsandvulnerabilities #rssfeedpostgeneratorecho #ArtificialIntelligence #SecurityonTechRepublic #SecurityTechRepublic #CyberSecurityNews #Cybersecurity #AIsecurity #deepseekv3 #qwen25max #AImodels #DeepSeek #Security #Alibaba #News #AI
-
Studie: #KI #Chatbots sind beim Zitieren von #News unbrauchbar
https://www.derstandard.at/story/3000000261220/studie-ki-chatbots-sind-beim-zitieren-von-news-unbrauchbar"Untersucht wurden #ChatGPT Search (#OpenAI), #Perplexity, Perplexity Pro (Perplexity AI), #Gemini 2.0 Flash (#Google), #DeepseekV3 Search (#Deepseek), #Grok-2 Search, Grok-3 Search Beta (#xAI) sowie #Copilot (#Microsoft und OpenAI)."
"#Grok3 [...] lieferte gleich in 96 Prozent aller Fälle falsche Antworten." 🤣
-
Studie: #KI #Chatbots sind beim Zitieren von #News unbrauchbar
https://www.derstandard.at/story/3000000261220/studie-ki-chatbots-sind-beim-zitieren-von-news-unbrauchbar"Untersucht wurden #ChatGPT Search (#OpenAI), #Perplexity, Perplexity Pro (Perplexity AI), #Gemini 2.0 Flash (#Google), #DeepseekV3 Search (#Deepseek), #Grok-2 Search, Grok-3 Search Beta (#xAI) sowie #Copilot (#Microsoft und OpenAI)."
"#Grok3 [...] lieferte gleich in 96 Prozent aller Fälle falsche Antworten." 🤣
-
DeepSeek releases DeepSeek-V3-0324 on Hugging Face!
#DeepSeek #AI #MachineLearning #DeepSeekV3 #HuggingFace #ArtificialIntelligence #AIModel
-
»Chinese #AIlab #DeepSeek just released the latest version of their enormous #DeepSeekv3 model: The license is #MIT (that's new - previous DeepSeek v3 had a custom license).« https://simonwillison.net/2025/Mar/24/deepseek/?eicker.news #tech #media
-
DeepSeek’s new V3-0324 AI model has launched quietly, offering efficient performance on a Mac Studio
#AI #DeepSeekV3 #DeepSeek #DeepSeekV30324 #GenAI #LLM #OpenSourceAI #AIModels
-
DeepSeek: ChatGPT killer or just another hype train? We compare it against ChatGPT and Gemini #apps #chatgpt #deepseek #deepseekr1 #deepseekv3 #digitallife #featured #gemini #geminiai #googlegemini #openai #video
-
DeepSeek und die Geschichte von Liang Wenfeng!
Gründung 2023
DeepSeek R1 übertrifft ChatGPT
Effiziente KI-Modelle
Weniger Ressourcen nötig#ai #ki #artificialintelligence #kuenstlicheintelligenz #deepseek #deepseekr1 #deepseekv3 #liangwenfeng #technologie
https://kinews24.de/deepseek-und-die-geschichte-von-lian-wenfeng/
-
DeepSeek und die Geschichte von Liang Wenfeng!
Gründung 2023
DeepSeek R1 übertrifft ChatGPT
Effiziente KI-Modelle
Weniger Ressourcen nötig#ai #ki #artificialintelligence #kuenstlicheintelligenz #deepseek #deepseekr1 #deepseekv3 #liangwenfeng #technologie
https://kinews24.de/deepseek-und-die-geschichte-von-lian-wenfeng/
-
🚀 DeepSeek V3 vs ChatGPT-4o: Which One Reigns Supreme?🤖
AI is evolving fast! 🏎️ DeepSeek V3 and ChatGPT-4o are two of the most powerful LLMs in 2025. But which one is better?
🔍 We compare:
✅ Accuracy & performance
✅ Multimodal capabilities
✅ Speed & efficiency
✅ Real-world applications📖 Read the full breakdown here:
https://radargit.com/2025/02/03/deepseek-v3-vs-chatgpt-4o-which-one-is-better/
Which AI model do you prefer? Comment below! 👇
#AI #DeepSeekV3 #ChatGPT4o #ArtificialIntelligence #Tech #MachineLearning #AICompari
-
Das wird noch etwas dauern. #ollama #deepseekv3
-
DeepSeek Locked Down Public Database Access That Exposed Chat History – Source: www.techrepublic.com https://ciso2ciso.com/deepseek-locked-down-public-database-access-that-exposed-chat-history-source-www-techrepublic-com/ #rssfeedpostgeneratorecho #ArtificialIntelligence #SecurityonTechRepublic #SecurityTechRepublic #CyberSecurityNews #SecurityResearch #databaseleakage #International #GenerativeAI #wizresearch #clickhouse #deepseekr1 #deepseekv3 #opensource #DeepSeek #openaio1 #Security #BigData
-
DeepSeek Locked Down Public Database Access That Exposed Chat History – Source: www.techrepublic.com https://ciso2ciso.com/deepseek-locked-down-public-database-access-that-exposed-chat-history-source-www-techrepublic-com/ #rssfeedpostgeneratorecho #ArtificialIntelligence #SecurityonTechRepublic #SecurityTechRepublic #CyberSecurityNews #SecurityResearch #databaseleakage #International #GenerativeAI #wizresearch #clickhouse #deepseekr1 #deepseekv3 #opensource #DeepSeek #openaio1 #Security #BigData
-
Research Firm Wiz Research began investigating DeepSeek soon after its generative AI took the tech world by storm.#artificialintelligence #clickhouse #databaseleakage #deepseek #deepseekr1 #deepseek-v3 #generativeai #openaio1 #security #securityresearch #wizresearch
DeepSeek Locked Down Public Database Access That Exposed Chat History -
"A key component of the success is that it is #opensource. #DeepSeek-V3 is on GitHub with detailed docs on how it can be replicated. This has fueled a rush of people to try to make their own models." https://baixacultura.org/2025/01/29/a-corrida-da-ia-ganha-um-novo-capitulo-chines-e-open-source/
A corrida da IA ganha um novo ... -
The Chinese firm said training the model cost just $5.6 million. Alibaba Cloud followed with a new generative AI model, while Microsoft alleges DeepSeek ‘distilled’ OpenAI’s work.#artificialintelligence #chatgpt #deepseek #deepseekr1 #deepseek-v3 #generativeai #Microsoft #nvidia #openai #reasoningmodels
DeepSeek Chatbot Beats OpenAI on App Store Leaderboard -
DeepSeek: China’s answer to ChatGPT is causing havoc, Nvidia loses nearly USD 600 bil in market cap #ai #apps #chatgpt #china #deepseek #deepseekr1 #deepseekv3 #digitallife #featured #news #tech
https://soyacincau.com/2025/01/28/deepseek-china-answer-to-chatgpt-nvidia-loses-nearly-600bil/