#voicesynthesis — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #voicesynthesis, aggregated by home.social.
-
Released my first Godot Asset! A retro voice synthesizer :mycomputer:
Asset: https://store.godotengine.org/asset/eibriel/tts-079/
Try it on Itchio: https://eibriel.itch.io/scp-079-voice
-
Released my first Godot Asset! A retro voice synthesizer :mycomputer:
Asset: https://store.godotengine.org/asset/eibriel/tts-079/
Try it on Itchio: https://eibriel.itch.io/scp-079-voice
-
StepFun Unveils "StepAudio 2.5 Realtime," Promising End-to-End Voice Synthesis
StepFun launches StepAudio 2.5 Realtime, an end-to-end voice model for roleplaying. Learn about its features and how developers can use GELab-Zero-4B-preview.
#StepAudio2.5, #VoiceSynthesis, #AIforGaming, #RealtimeAI, #GELabZero
https://newsletter.tf/stepfun-stepaudio-2-5-realtime-voice-model-release/
-
StepFun's new StepAudio 2.5 Realtime model can generate speech instantly, making roleplaying games more immersive. It uses the GELab-Zero-4B-preview model.
#StepAudio2.5, #VoiceSynthesis, #AIforGaming, #RealtimeAI, #GELabZero
https://newsletter.tf/stepfun-stepaudio-2-5-realtime-voice-model-release/ -
Success!! I've completed a first reimplementation of Dr. Sbaitso voice synthesis in Godot!
Try it out at Itchio: https://eibriel.itch.io/scp-079-voice
Ive configure the accessibility settings so the app properly describes the input fields and buttons, but for some reason is not working for me (tested on Linux with Orca). Let me know if it works for you!
#TTS #DrSbaitso #VoiceSynthesis #TextToSpeech #079 #SCP079 #SCP #Godot
-
Some improvements to the concatenation, prosody is still missing.
Here is a well known phrase by SCP 079.
The audio contains the same phrase first performed by Dr. Sbaitso TTS and the by Godot reimplementation.
#TTS #DrSbaitso #VoiceSynthesis #TextToSpeech #079 #SCP079 #SCP #Godot
-
Some improvements to the concatenation, prosody is still missing.
Here is a well known phrase by SCP 079.
The audio contains the same phrase first performed by Dr. Sbaitso TTS and the by Godot reimplementation.
#TTS #DrSbaitso #VoiceSynthesis #TextToSpeech #079 #SCP079 #SCP #Godot
-
Dr. Sbaitso compared to my reimplementation in Godot (Sbaitso first) :computer_explorer: :pc_color:
Implemented: basic waveform concatenation
Missing: Interpolation, pitch control, prosody, text to phonemesIm very happy with the progress, will be great to be able to run the voice without needing emulation.
-
Dr. Sbaitso compared to my reimplementation in Godot (Sbaitso first) :computer_explorer: :pc_color:
Implemented: basic waveform concatenation
Missing: Interpolation, pitch control, prosody, text to phonemesIm very happy with the progress, will be great to be able to run the voice without needing emulation.
-
What I've learned so far while reverse engineering Dr Sbaitso's voice:
- Reverse engineering is hardAlso, the voice was made by very clever people. It's optimized to sound as good as possible, while consuming very few resources.
Progress after 5 days: 10%
-
What I've learned so far while reverse engineering Dr Sbaitso's voice:
- Reverse engineering is hardAlso, the voice was made by very clever people. It's optimized to sound as good as possible, while consuming very few resources.
Progress after 5 days: 10%
-
Primera prueba del sintetizador de voz por difonos hecho en Godot.
Tiene un millón de problemas, grabé la voz así nomás.
-
Building a Diphone TTS engine in Godot for no reason at all.
-
Building a Diphone TTS engine in Godot for no reason at all.
-
Germany gets a new AI call assistant from Deutsche Telekom that works straight from the cellular network—no app required. Powered by ElevenLabs’ voice synthesis, it can translate languages on the fly. Unveiled at Mobile World Congress, it shows how open‑source‑friendly AI can reshape everyday calls. Curious how it works? #MagentaAI #DeutscheTelekom #ElevenLabs #VoiceSynthesis
🔗 https://aidailypost.com/news/magenta-ai-call-assistant-launches-germany-no-app-needed
-
Germany gets a new AI call assistant from Deutsche Telekom that works straight from the cellular network—no app required. Powered by ElevenLabs’ voice synthesis, it can translate languages on the fly. Unveiled at Mobile World Congress, it shows how open‑source‑friendly AI can reshape everyday calls. Curious how it works? #MagentaAI #DeutscheTelekom #ElevenLabs #VoiceSynthesis
🔗 https://aidailypost.com/news/magenta-ai-call-assistant-launches-germany-no-app-needed
-
ElevenLabs appoints Karthik Rajaram as India Country Head to accelerate AI voice growth. His leadership will boost multilingual audio, voice synthesis and conversational AI for creators and brands across the Indian market. Discover how this move could reshape digital content creation. #AIvoice #VoiceSynthesis #ElevenLabs #MultilingualAudio
🔗 https://aidailypost.com/news/elevenlabs-names-karthik-rajaram-india-country-head-power-ai-voice
-
Designing with the KT142C voice chip? Remember, its user-accessible memory is 320KB. If your audio files are larger, you'll need a different solution.
The KT142F chip allows for an external Flash memory, letting you scale your voice storage capacity as needed. The trade-off is a slight increase in component cost and board space, but you gain total flexibility.
https://www.linkedin.com/pulse/built-in-320kbyte-memory-kt142c-voice-chip-other-solutions-tsui-vbexc
#OpenHardware #Embedded #Electronics #VoiceSynthesis #PCBDesign #KT142 #Maker
-
Ever wondered how AI voices are becoming so human-like? We've moved beyond simple "voice packs" (pre-recorded clips) to true AI "voice clones" that generate new speech from text.
The magic is in the details: AI models learn a voice's unique pitch and cadence. The secret sauce? "Emotional tuning," which adds happiness, sadness, or empathy to the performance. It's a game-changer for accessibility and content creation. #AIVoice #VoiceSynthesis #Tech
-
Project update: LingoFreq #Mandarin
Generates themed example sentences from language word-frequency list (Chinese HSK). Builds a printable book & (soon) voice audio files. This theme is "technology".
We'll be building this for #English & #Spanish.
Tech: #Python, OpenAI ChatGPT, ElevenLabs #VoiceSynthesis (coming soon), SQLite database, Jinja templates
Source code:
- Sentence builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
- Book builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/build_book.py -
Project update: LingoFreq #Mandarin
Generates themed example sentences from language word-frequency list (Chinese HSK). Builds a printable book & (soon) voice audio files. This theme is "technology".
We'll be building this for #English & #Spanish.
Tech: #Python, OpenAI ChatGPT, ElevenLabs #VoiceSynthesis (coming soon), SQLite database, Jinja templates
Source code:
- Sentence builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
- Book builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/build_book.py -
Project update: LingoFreq - language learning tool
Goal: Create voice audio files of high-frequency Mandarin words & example sentences
Loads high-frequency #Mandarin #Chinese words from HSK CSV file into an #SQLite #database, using #Python. Then we loop over all the Mandarin words, and call to OpenAI ChatGPT to create simple example sentences. Next, we'll use ElevenLabs #VoiceSynthesis text-to-speech to vocalize the example sentences as audio files.
Code: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
-
Project update: LingoFreq - language learning tool
Goal: Create voice audio files of high-frequency Mandarin words & example sentences
Loads high-frequency #Mandarin #Chinese words from HSK CSV file into an #SQLite #database, using #Python. Then we loop over all the Mandarin words, and call to OpenAI ChatGPT to create simple example sentences. Next, we'll use ElevenLabs #VoiceSynthesis text-to-speech to vocalize the example sentences as audio files.
Code: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
-
#Google, your automated service already told me that you're an automated service confirming our business hours. Don't patronize me by pretending to be a human. Why the bleep are you inserting artificial "um" and "sorry" and the like into your sentences? It doesn't fool me or make me feel better. #ai #voiceSynthesis
-
#Google, your automated service already told me that you're an automated service confirming our business hours. Don't patronize me by pretending to be a human. Why the bleep are you inserting artificial "um" and "sorry" and the like into your sentences? It doesn't fool me or make me feel better. #ai #voiceSynthesis
-
@mark I've been doing digging into #openSource text to speech lately. The best overall I have found is Piper (https://github.com/OHF-Voice/piper1-gpl), but being honest - the voices are not great.
To those in the know - Is there a ranked repository of better voices somewhere? How easy is it to create your own voices? - Using this for building language learning tools (#English, #Spanish, #Mandarin). Other TTS tools to consider?
-
@mark I've been doing digging into #openSource text to speech lately. The best overall I have found is Piper (https://github.com/OHF-Voice/piper1-gpl), but being honest - the voices are not great.
To those in the know - Is there a ranked repository of better voices somewhere? How easy is it to create your own voices? - Using this for building language learning tools (#English, #Spanish, #Mandarin). Other TTS tools to consider?
-
#OpenInvite #OpenSource #CodeJam @ #Medein
Universidad #EAFIT / Bloque 20 pavilion / Behind Laboratorio de #Cafe. Next to Bloque 19 / #MakerSpace. We are learning & building with #Python, #FastAPi, #VoiceSynthesis.Come by. Bring long power extension cables if you got'm!
-
#OpenInvite #OpenSource #CodeJam @ #Medein
Universidad #EAFIT / Bloque 20 pavilion / Behind Laboratorio de #Cafe. Next to Bloque 19 / #MakerSpace. We are learning & building with #Python, #FastAPi, #VoiceSynthesis.Come by. Bring long power extension cables if you got'm!
-
Scarlett Johansson & OpenAI: AI Voice Ethics & The 'Her' Paradox Exposed https://aiorbit.app/scarlett-johansson-openai-ai-voice-ethics-the-her-paradox-exposed/ #AIethics
#ScarlettJohansson
#VoiceSynthesis
#CreativeRights -
Llasa: Llama-Based Speech Synthesis
https://llasatts.github.io/llasatts/
#HackerNews #Llasa #Llama-Based #Speech #Synthesis #SpeechTechnology #AIInnovation #VoiceSynthesis
-
Llasa: Llama-Based Speech Synthesis
https://llasatts.github.io/llasatts/
#HackerNews #Llasa #Llama-Based #Speech #Synthesis #SpeechTechnology #AIInnovation #VoiceSynthesis
-
Google has integrated its Chirp 3 HD voice model into Vertex AI enhancing speech synthesis capabilities with customizable and lifelike voice features
#AI #GoogleAI #VertexAI #Chirp3 #VoiceSynthesis #AIVoices #TextToSpeech #GenAI #CustomAIVoices #Alphabet
https://winbuzzer.com/2025/03/17/google-expands-vertex-ai-with-chirp-3-hd-voice-model-xcxwbn/
-
Google has integrated its Chirp 3 HD voice model into Vertex AI enhancing speech synthesis capabilities with customizable and lifelike voice features
#AI #GoogleAI #VertexAI #Chirp3 #VoiceSynthesis #AIVoices #TextToSpeech #GenAI #CustomAIVoices #Alphabet
https://winbuzzer.com/2025/03/17/google-expands-vertex-ai-with-chirp-3-hd-voice-model-xcxwbn/
-
Eerily realistic AI voice demo sparks amazement and discomfort online - In late 2013, the Spike Jonze film Her imagined a future where people woul... - https://arstechnica.com/ai/2025/03/users-report-emotional-bonds-with-startlingly-realistic-ai-voice-demo/ #machinelearning #voicesynthesis #aiassistants #emotionalai #chatbots #chatgpt #chatgtp #biz #sesame #tech #ai
-
Eerily realistic AI voice demo sparks amazement and discomfort online - In late 2013, the Spike Jonze film Her imagined a future where people woul... - https://arstechnica.com/ai/2025/03/users-report-emotional-bonds-with-startlingly-realistic-ai-voice-demo/ #machinelearning #voicesynthesis #aiassistants #emotionalai #chatbots #chatgpt #chatgtp #biz #sesame #tech #ai
-
OK, this is probably a rather long shot, but does anyone know of a voice synthesis model that is open-weights or ideally even open-source, and capable of producing intonation (specifically, something like rap lyrics)? #text2speech #voicesynthesis #ai
-
OK, this is probably a rather long shot, but does anyone know of a voice synthesis model that is open-weights or ideally even open-source, and capable of producing intonation (specifically, something like rap lyrics)? #text2speech #voicesynthesis #ai
-
Your AI clone could target your family, but there’s a simple defense - On Tuesday, the US Federal Bureau of Investigation advised Americans to sh... - https://arstechnica.com/ai/2024/12/your-ai-clone-could-target-your-family-but-theres-a-simple-defense/ #machinelearning #audiosynthesis #imagesynthesis #voicesynthesis #voicecloning #aipassword #secretword #asaranear #deepfakes #aiethics #safeword #aicrime #aifraud #twitter #biz #fbi #ai
-
Your AI clone could target your family, but there’s a simple defense - On Tuesday, the US Federal Bureau of Investigation advised Americans to sh... - https://arstechnica.com/ai/2024/12/your-ai-clone-could-target-your-family-but-theres-a-simple-defense/ #machinelearning #audiosynthesis #imagesynthesis #voicesynthesis #voicecloning #aipassword #secretword #asaranear #deepfakes #aiethics #safeword #aicrime #aifraud #twitter #biz #fbi #ai
-
Next-gen voice synthesis is transforming human-AI interaction! 🤖🎤 Discover how enhanced speech tech is making conversations with AI more natural, realistic, and engaging. Ready to explore the future of communication? 👇 #AI #VoiceSynthesis #TechInnovation #FutureOfAI
https://pupuweb.com/how-does-next-generation-voice-synthesis-enhance-human-ai-interaction/
-
Man tricks OpenAI’s voice bot into duet of The Beatles’ “Eleanor Rigby” - Enlarge / A screen capture of AJ Smith doing his Eleanor Rigby duet wit... - https://arstechnica.com/?p=2052995 #largelanguagemodels #advancedvoicemode #aipromptinjection #machinelearning #promptinjection #audiosynthesis #musicsynthesis #voicesynthesis #paulmccartney #eleanorrigby #aicopyright #thebeatles #aifairuse #copyright #ajsmith #fairuse #biz #openai #ai
-
Man tricks OpenAI’s voice bot into duet of The Beatles’ “Eleanor Rigby” - Enlarge / A screen capture of AJ Smith doing his Eleanor Rigby duet wit... - https://arstechnica.com/?p=2052995 #largelanguagemodels #advancedvoicemode #aipromptinjection #machinelearning #promptinjection #audiosynthesis #musicsynthesis #voicesynthesis #paulmccartney #eleanorrigby #aicopyright #thebeatles #aifairuse #copyright #ajsmith #fairuse #biz #openai #ai
-
Due to AI fakes, the “deep doubt” era is here - Enlarge (credit: Memento | Aurich Lawson)
Given the flood of p... - https://arstechnica.com/?p=2042584 #daniellek.citron #machinelearning #imagesynthesis #musicsynthesis #videosynthesis #voicesynthesis #liarsdividend #medialiteracy #robertchesney #textsynthesis #kamalaharris #donaldtrump #socialmedia #deepdoubt #deepfakes #features #facebook #joebiden #politics #history #biz #doubt #ai #x
-
Due to AI fakes, the “deep doubt” era is here - Enlarge (credit: Memento | Aurich Lawson)
Given the flood of p... - https://arstechnica.com/?p=2042584 #daniellek.citron #machinelearning #imagesynthesis #musicsynthesis #videosynthesis #voicesynthesis #liarsdividend #medialiteracy #robertchesney #textsynthesis #kamalaharris #donaldtrump #socialmedia #deepdoubt #deepfakes #features #facebook #joebiden #politics #history #biz #doubt #ai #x
-
Google beats OpenAI to wide release of interruptible AI voice chat mode - Enlarge / The Google Gemini logo. (credit: Google)
On Thursday... - https://arstechnica.com/?p=2049707 #advancedvoicemode #machinelearning #voicesynthesis #googlegemini #aivoicechat #geminilive #chatgpt #chatgtp #biz #google #openai #ai
-
Google beats OpenAI to wide release of interruptible AI voice chat mode - Enlarge / The Google Gemini logo. (credit: Google)
On Thursday... - https://arstechnica.com/?p=2049707 #advancedvoicemode #machinelearning #voicesynthesis #googlegemini #aivoicechat #geminilive #chatgpt #chatgtp #biz #google #openai #ai
-
New Episode: hpr4188 :: Re: HPR4172 Comment by Ken Fallon
Hosted by Archer72 on 2024-08-21 is flagged as Clean and is released under a CC-BY-SA license.
Tags: #tts, #TextToSpeech, #VoiceSynthesis, #accessibility.
-
New Episode: hpr4188 :: Re: HPR4172 Comment by Ken Fallon
Hosted by Archer72 on 2024-08-21 is flagged as Clean and is released under a CC-BY-SA license.
Tags: #tts, #TextToSpeech, #VoiceSynthesis, #accessibility.
-
ChatGPT Advanced Voice Mode impresses testers with sound effects, catching its breath - Enlarge / A stock photo of a robot whispering to a man. (credit: Andrey... - https://arstechnica.com/?p=2040213 #largelanguagemodels #advancedvoicemode #aivoicegenerators #machinelearning #audiosynthesis #voicesynthesis #textsynthesis #chatgpt #chatgtp #biz #openai #ai
-
ChatGPT Advanced Voice Mode impresses testers with sound effects, catching its breath - Enlarge / A stock photo of a robot whispering to a man. (credit: Andrey... - https://arstechnica.com/?p=2040213 #largelanguagemodels #advancedvoicemode #aivoicegenerators #machinelearning #audiosynthesis #voicesynthesis #textsynthesis #chatgpt #chatgtp #biz #openai #ai
-
New Episode: hpr4172 :: Re: hpr4072 Piper voice synthesis
Hosted by Archer72 on 2024-07-30 is flagged as Clean and is released under a CC-BY-SA license.
Tags: #tts, #TextToSpeech, #VoiceSynthesis, #accessibility.
-
AI-generated Al Michaels to provide daily recaps during 2024 Summer Olympics - Enlarge / Al Michaels looks on prior to the game between the Minnesota ... - https://arstechnica.com/?p=2033917 #2024summerolympics #sportscommentators #machinelearning #speechsynthesis #summerolympics #voicesynthesis #olympicgames #almichaels #deepfakes #olympics #chatgpt #chatgtp #biz #sports #nbc #nfl #ai
-
Music industry giants allege mass copyright violation by AI firms - Enlarge / Michael Jackson in concert, 1986. Sony Music owns a large por... - https://arstechnica.com/?p=2033128 #univeralmusicgroup #machinelearning #audiosynthesis #michaeljackson #musicsynthesis #voicesynthesis #warnerrecords #generativeai #ailawsuit #microsoft #sonymusic #reuters #biz #policy #google #openai #suno #udio #ai