#quantization — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #quantization, aggregated by home.social.
-
What the Nielsen DoubleVerify deal means for AI ad verification, brand safety, and CTV measurement—with concrete steps marketers should
https://aistory.news/machine-learning/nielsen-doubleverify-deal-shifts-ai-ad-verification-power/
-
RT @jun_song: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einer Spitzenleistung von 122 Token pro Sekunde. Die Implementierung erfolgte basierend auf dem Rezept von @MiaAIlab, um DFlash MTP zu nutzen. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Das Modell wurde außerdem abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach erstaunlich.
mehr auf Arint.info
#AI #MachineLearning #Performance #Quantization #SuperDeepseek #arint_info
-
RT @jun_song: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einer Spitzenleistung von 122 Token pro Sekunde. Die Implementierung erfolgte basierend auf dem Rezept von @MiaAIlab, um DFlash MTP zu nutzen. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Das Modell wurde außerdem abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach erstaunlich.
mehr auf Arint.info
#AI #MachineLearning #Performance #Quantization #SuperDeepseek #arint_info
-
Google DeepMind reorganization moves Demis Hassabis to Alphabet Chief Scientist. Here’s what it signals for Gemini, developers, and AI
https://aistory.news/machine-learning/why-google-deepmind-reorganization-puts-agi-first-fast/
-
RT @jun_song: TRANSLASATION: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einem Spitzenwert von 122 Token pro Sekunde. Die Arbeit basiert auf dem Rezept von @MiaAIlab zur Ausführung mit DFlash MTP. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Außerdem wurde es abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach unglaublich.
mehr auf Arint.info
#AI #DeepSeek #MachineLearning #Performance #Quantization #arint_info
-
RT @jun_song: TRANSLASATION: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einem Spitzenwert von 122 Token pro Sekunde. Die Arbeit basiert auf dem Rezept von @MiaAIlab zur Ausführung mit DFlash MTP. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Außerdem wurde es abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach unglaublich.
mehr auf Arint.info
#AI #DeepSeek #MachineLearning #Performance #Quantization #arint_info
-
Hackster.io maker trends point to offline-first builds, desktop PCB tools, and tiny on-device AI. Here’s what that means for makers and
https://aistory.news/machine-learning/hacksterio-maker-trends-on-device-ai-privacy-desktop-fab/
-
NeurIPS 2026 program expands with evaluation, reproducibility and creative AI. What the 40th edition means for authors, reviewers, and
https://aistory.news/machine-learning/what-the-neurips-2026-program-signals-for-ml-in-2027/
-
TurboQuant compression attacks KV cache bloat and vector search overhead. We explain how it could cut serving costs and extend context
https://aistory.news/generative-ai/turboquant-compression-shifts-llm-costs-to-the-kv-cache/
-
Good article about quantized models, "The Model You Audit Is Not the Model You Ship".
https://www.techpolicy.press/the-model-you-audit-is-not-the-model-you-ship/
#AI #models #quantization #LLM #Audit -
Good article about quantized models, "The Model You Audit Is Not the Model You Ship".
https://www.techpolicy.press/the-model-you-audit-is-not-the-model-you-ship/
#AI #models #quantization #LLM #Audit -
Red Hat, NVIDIA and IBM back a project as AI policy as code shifts from slides to CI/CD. What it means for audits, speed, and spend in 2026.
https://aistory.news/machine-learning/ai-policy-as-code-moves-from-talk-to-tooling-for-growth/
-
Break Through Tech AI Program blends a Cornell ML certificate, mentor-led studios, and micro-internships so early undergrads can build
https://aistory.news/machine-learning/how-break-through-tech-ai-program-builds-hire-ready-talent/
-
OpenAI’s move to align with the EU AI Act GPAI Code resets enterprise AI buying. Use our GPAI Code compliance checklist to cut risk and
https://aistory.news/machine-learning/what-eu-ai-act-gpai-code-means-for-enterprise-ai-buys/
-
MIT JARVIS Challenge shows AI copilots can speed parts of jet engine design, but reveal limits in safety-critical steps. What it means for
https://aistory.news/machine-learning/mit-jarvis-challenge-ai-copilots-help-build-a-jet-engine/
-
Yahoo’s Amazon Bedrock retargeting rollout and AWS safety and infra updates show genAI moving into ad tech production. Here’s what it
https://aistory.news/machine-learning/yahoo-taps-amazon-bedrock-retargeting-in-aws-ai-push/
-
OpenAI coding agents are tied to faster science builds, per AI News. Here’s how teams can measure gains, reduce risk, and turn speed into
https://aistory.news/machine-learning/openai-coding-agents-speed-builds-what-it-means-for-rd/
-
NeurIPS 2026 satellites will run in Atlanta and Paris on December 9–13 alongside Sydney’s Dec 6–12 main site. Here’s what the tri-site plan
https://aistory.news/machine-learning/neurips-2026-satellites-set-for-atlanta-and-paris-dec-9-13/
-
Google aims to cut enterprise agent token costs with Gemini 3.6, per Artificial Intelligence News. See why per‑task pricing could reset AI
https://aistory.news/machine-learning/google-targets-enterprise-agent-token-costs-with-gemini-36/
-
Meta, Microsoft, Nvidia, and IBM back an open-weight AI alliance. Here’s how it could reshape model procurement, reduce lock-in, and speed
https://aistory.news/machine-learning/why-the-open-weight-ai-alliance-could-cut-enterprise-costs/