home.social

#diffusiongemma — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #diffusiongemma, aggregated by home.social.

fetched live
  1. “Now that organisations have been weaned off earlier 'all you can eat' #subscription plans and onto 'pay-as-you-go' metered #token consumption, they're all in various stages of sticker shock.

    Several talks at the conference discussed managing token costs, such as AJ Fisher's exploration of 'diffusion' models. Analogous to the diffusers used to generate images, they generate text at lighting speed, making them cheaper to operate while also being less accurate than the pricey and slower “autoregressive” #FrontierModels.

    Fisher's solution? Use a low-quality model and make it iterate on a problem (that new classic, the #RalphWiggumLoop) until it gets a satisfactory solution. This approach delivers the same result as a full-fat model, for anywhere from one half to one tenth the spend. #Google released its #DiffusionGemma model, which produces text at prodigious speed, just days after Fisher's talk, giving everyone the ability to try this approach.” — #MarkPesce

    #AI / #ArtificialIntelligence / #developers / #software / #RalphWiggens / #Simpsons <theregister.com/columnists/202>

  2. “Now that organisations have been weaned off earlier 'all you can eat' #subscription plans and onto 'pay-as-you-go' metered #token consumption, they're all in various stages of sticker shock.

    Several talks at the conference discussed managing token costs, such as AJ Fisher's exploration of 'diffusion' models. Analogous to the diffusers used to generate images, they generate text at lighting speed, making them cheaper to operate while also being less accurate than the pricey and slower “autoregressive” #FrontierModels.

    Fisher's solution? Use a low-quality model and make it iterate on a problem (that new classic, the #RalphWiggumLoop) until it gets a satisfactory solution. This approach delivers the same result as a full-fat model, for anywhere from one half to one tenth the spend. #Google released its #DiffusionGemma model, which produces text at prodigious speed, just days after Fisher's talk, giving everyone the ability to try this approach.” — #MarkPesce

    #AI / #ArtificialIntelligence / #developers / #software / #RalphWiggens / #Simpsons <theregister.com/columnists/202>

  3. 🧠 #Google ha presentato #DiffusionGemma, un nuovo modello open source sperimentale che esplora un approccio diverso alla generazione del testo. 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 
    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomaro

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  4. 🧠 #Google ha presentato #DiffusionGemma, un nuovo modello open source sperimentale che esplora un approccio diverso alla generazione del testo. 

    👉 I dettagli: linkedin.com/posts/alessiopoma

    ___ 
    ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶𝗼𝗿𝗻𝗮𝘁𝗼/𝗮 𝘀𝘂 𝗾𝘂𝗲𝘀𝘁𝗲 𝘁𝗲𝗺𝗮𝘁𝗶𝗰𝗵𝗲, 𝗶𝘀𝗰𝗿𝗶𝘃𝗶𝘁𝗶 𝗮𝗹𝗹𝗮 𝗺𝗶𝗮 𝗻𝗲𝘄𝘀𝗹𝗲𝘁𝘁𝗲𝗿: bit.ly/newsletter-alessiopomaro

    #AI #GenAI #GenerativeAI #IntelligenzaArtificiale #LLM 

  5. RT @LottoLabs: DiffusionGemma 26B-A4B mit llama.cpp-Fork. Dies ist ein gutes Beispiel dafür, wie Diffusionsmodelle einen Textblock parallel im Gegensatz zum nächsten Token generieren. Allerdings muss ich auf bessere Server-Unterstützung für llama.cpp warten oder zu vllm oder ktransformers wechseln, um tatsächliche Auswertungen etc. durchzuführen. Video.

    mehr auf Arint.info

    #AI #DiffusionGemma #DiffusionModels #ktransformers #llama #vllm #arint_info

    https://x.com/LottoLabs/status/2064920298206728560#m

  6. RT @LottoLabs: DiffusionGemma 26B-A4B mit llama.cpp-Fork. Dies ist ein gutes Beispiel dafür, wie Diffusionsmodelle einen Textblock parallel im Gegensatz zum nächsten Token generieren. Allerdings muss ich auf bessere Server-Unterstützung für llama.cpp warten oder zu vllm oder ktransformers wechseln, um tatsächliche Auswertungen etc. durchzuführen. Video.

    mehr auf Arint.info

    #AI #DiffusionGemma #DiffusionModels #ktransformers #llama #vllm #arint_info

    https://x.com/LottoLabs/status/2064920298206728560#m

  7. 👀 DiffusionGemma: Google lancia un nuovo modello open source per esecuzione in locale che elabora 256 token in parallelo, usa attention bidirezionale e si auto-corregge in tempo reale.
    gomoot.com/diffusiongemma-il-n

    #DiffusionGemma #geminidiffusion #google #news

  8. Google、ローカルAIが4倍速くなるテキスト生成モデル「DiffusionGemma」を実験的に発表、逐次ではなく一括で生成/「GeForce RTX 5090」で700トークン/秒超を達成
    forest.watch.impress.co.jp/doc

    #forest_watch_impress #Gemma #Google_DeepMind #Gemma_4 #DiffusionGemma #genai #文章生成 #AIコーディング #Gemini