home.social

#quantization — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #quantization, aggregated by home.social.

fetched live
  1. What the Nielsen DoubleVerify deal means for AI ad verification, brand safety, and CTV measurement—with concrete steps marketers should

    aistory.news/machine-learning/

  2. RT @jun_song: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einer Spitzenleistung von 122 Token pro Sekunde. Die Implementierung erfolgte basierend auf dem Rezept von @MiaAIlab, um DFlash MTP zu nutzen. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Das Modell wurde außerdem abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach erstaunlich.

    mehr auf Arint.info

    #AI #MachineLearning #Performance #Quantization #SuperDeepseek #arint_info

    https://x.com/jun_song/status/2087208879021301975#m

  3. RT @jun_song: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einer Spitzenleistung von 122 Token pro Sekunde. Die Implementierung erfolgte basierend auf dem Rezept von @MiaAIlab, um DFlash MTP zu nutzen. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Das Modell wurde außerdem abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach erstaunlich.

    mehr auf Arint.info

    #AI #MachineLearning #Performance #Quantization #SuperDeepseek #arint_info

    https://x.com/jun_song/status/2087208879021301975#m

  4. Google DeepMind reorganization moves Demis Hassabis to Alphabet Chief Scientist. Here’s what it signals for Gemini, developers, and AI

    aistory.news/machine-learning/

  5. RT @jun_song: TRANSLASATION: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einem Spitzenwert von 122 Token pro Sekunde. Die Arbeit basiert auf dem Rezept von @MiaAIlab zur Ausführung mit DFlash MTP. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Außerdem wurde es abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach unglaublich.

    mehr auf Arint.info

    #AI #DeepSeek #MachineLearning #Performance #Quantization #arint_info

    https://x.com/jun_song/status/2087208879021301975#m

  6. RT @jun_song: TRANSLASATION: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einem Spitzenwert von 122 Token pro Sekunde. Die Arbeit basiert auf dem Rezept von @MiaAIlab zur Ausführung mit DFlash MTP. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Außerdem wurde es abgeleitet und die Qualität wurde behoben, um die Gesamtleistung zu verbessern. Das ist einfach unglaublich.

    mehr auf Arint.info

    #AI #DeepSeek #MachineLearning #Performance #Quantization #arint_info

    https://x.com/jun_song/status/2087208879021301975#m

  7. Hackster.io maker trends point to offline-first builds, desktop PCB tools, and tiny on-device AI. Here’s what that means for makers and

    aistory.news/machine-learning/

  8. NeurIPS 2026 program expands with evaluation, reproducibility and creative AI. What the 40th edition means for authors, reviewers, and

    aistory.news/machine-learning/

  9. TurboQuant compression attacks KV cache bloat and vector search overhead. We explain how it could cut serving costs and extend context

    aistory.news/generative-ai/tur

  10. Red Hat, NVIDIA and IBM back a project as AI policy as code shifts from slides to CI/CD. What it means for audits, speed, and spend in 2026.

    aistory.news/machine-learning/

  11. OpenAI’s move to align with the EU AI Act GPAI Code resets enterprise AI buying. Use our GPAI Code compliance checklist to cut risk and

    aistory.news/machine-learning/

  12. MIT JARVIS Challenge shows AI copilots can speed parts of jet engine design, but reveal limits in safety-critical steps. What it means for

    aistory.news/machine-learning/

  13. Yahoo’s Amazon Bedrock retargeting rollout and AWS safety and infra updates show genAI moving into ad tech production. Here’s what it

    aistory.news/machine-learning/

  14. OpenAI coding agents are tied to faster science builds, per AI News. Here’s how teams can measure gains, reduce risk, and turn speed into

    aistory.news/machine-learning/

  15. NeurIPS 2026 satellites will run in Atlanta and Paris on December 9–13 alongside Sydney’s Dec 6–12 main site. Here’s what the tri-site plan

    aistory.news/machine-learning/

  16. Google aims to cut enterprise agent token costs with Gemini 3.6, per Artificial Intelligence News. See why per‑task pricing could reset AI

    aistory.news/machine-learning/

  17. Meta, Microsoft, Nvidia, and IBM back an open-weight AI alliance. Here’s how it could reshape model procurement, reduce lock-in, and speed

    aistory.news/machine-learning/