#voicesynthesis — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #voicesynthesis, aggregated by home.social.
-
Released my first Godot Asset! A retro voice synthesizer :mycomputer:
Asset: https://store.godotengine.org/asset/eibriel/tts-079/
Try it on Itchio: https://eibriel.itch.io/scp-079-voice
-
Some improvements to the concatenation, prosody is still missing.
Here is a well known phrase by SCP 079.
The audio contains the same phrase first performed by Dr. Sbaitso TTS and the by Godot reimplementation.
#TTS #DrSbaitso #VoiceSynthesis #TextToSpeech #079 #SCP079 #SCP #Godot
-
Dr. Sbaitso compared to my reimplementation in Godot (Sbaitso first) :computer_explorer: :pc_color:
Implemented: basic waveform concatenation
Missing: Interpolation, pitch control, prosody, text to phonemesIm very happy with the progress, will be great to be able to run the voice without needing emulation.
-
What I've learned so far while reverse engineering Dr Sbaitso's voice:
- Reverse engineering is hardAlso, the voice was made by very clever people. It's optimized to sound as good as possible, while consuming very few resources.
Progress after 5 days: 10%
-
Building a Diphone TTS engine in Godot for no reason at all.
-
Germany gets a new AI call assistant from Deutsche Telekom that works straight from the cellular network—no app required. Powered by ElevenLabs’ voice synthesis, it can translate languages on the fly. Unveiled at Mobile World Congress, it shows how open‑source‑friendly AI can reshape everyday calls. Curious how it works? #MagentaAI #DeutscheTelekom #ElevenLabs #VoiceSynthesis
🔗 https://aidailypost.com/news/magenta-ai-call-assistant-launches-germany-no-app-needed
-
Designing with the KT142C voice chip? Remember, its user-accessible memory is 320KB. If your audio files are larger, you'll need a different solution.
The KT142F chip allows for an external Flash memory, letting you scale your voice storage capacity as needed. The trade-off is a slight increase in component cost and board space, but you gain total flexibility.
https://www.linkedin.com/pulse/built-in-320kbyte-memory-kt142c-voice-chip-other-solutions-tsui-vbexc
#OpenHardware #Embedded #Electronics #VoiceSynthesis #PCBDesign #KT142 #Maker
-
Ever wondered how AI voices are becoming so human-like? We've moved beyond simple "voice packs" (pre-recorded clips) to true AI "voice clones" that generate new speech from text.
The magic is in the details: AI models learn a voice's unique pitch and cadence. The secret sauce? "Emotional tuning," which adds happiness, sadness, or empathy to the performance. It's a game-changer for accessibility and content creation. #AIVoice #VoiceSynthesis #Tech
-
Project update: LingoFreq #Mandarin
Generates themed example sentences from language word-frequency list (Chinese HSK). Builds a printable book & (soon) voice audio files. This theme is "technology".
We'll be building this for #English & #Spanish.
Tech: #Python, OpenAI ChatGPT, ElevenLabs #VoiceSynthesis (coming soon), SQLite database, Jinja templates
Source code:
- Sentence builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
- Book builder: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/build_book.py -
Project update: LingoFreq - language learning tool
Goal: Create voice audio files of high-frequency Mandarin words & example sentences
Loads high-frequency #Mandarin #Chinese words from HSK CSV file into an #SQLite #database, using #Python. Then we loop over all the Mandarin words, and call to OpenAI ChatGPT to create simple example sentences. Next, we'll use ElevenLabs #VoiceSynthesis text-to-speech to vocalize the example sentences as audio files.
Code: https://codeberg.org/jro/LingoFreq-app/src/branch/main/apps/generate_example_sentences.py
-
#Google, your automated service already told me that you're an automated service confirming our business hours. Don't patronize me by pretending to be a human. Why the bleep are you inserting artificial "um" and "sorry" and the like into your sentences? It doesn't fool me or make me feel better. #ai #voiceSynthesis
-
@mark I've been doing digging into #openSource text to speech lately. The best overall I have found is Piper (https://github.com/OHF-Voice/piper1-gpl), but being honest - the voices are not great.
To those in the know - Is there a ranked repository of better voices somewhere? How easy is it to create your own voices? - Using this for building language learning tools (#English, #Spanish, #Mandarin). Other TTS tools to consider?
-
#OpenInvite #OpenSource #CodeJam @ #Medein
Universidad #EAFIT / Bloque 20 pavilion / Behind Laboratorio de #Cafe. Next to Bloque 19 / #MakerSpace. We are learning & building with #Python, #FastAPi, #VoiceSynthesis.Come by. Bring long power extension cables if you got'm!
-
Scarlett Johansson & OpenAI: AI Voice Ethics & The 'Her' Paradox Exposed https://aiorbit.app/scarlett-johansson-openai-ai-voice-ethics-the-her-paradox-exposed/ #AIethics
#ScarlettJohansson
#VoiceSynthesis
#CreativeRights -
Llasa: Llama-Based Speech Synthesis
https://llasatts.github.io/llasatts/
#HackerNews #Llasa #Llama-Based #Speech #Synthesis #SpeechTechnology #AIInnovation #VoiceSynthesis
-
Google has integrated its Chirp 3 HD voice model into Vertex AI enhancing speech synthesis capabilities with customizable and lifelike voice features
#AI #GoogleAI #VertexAI #Chirp3 #VoiceSynthesis #AIVoices #TextToSpeech #GenAI #CustomAIVoices #Alphabet
https://winbuzzer.com/2025/03/17/google-expands-vertex-ai-with-chirp-3-hd-voice-model-xcxwbn/
-
Eerily realistic AI voice demo sparks amazement and discomfort online - In late 2013, the Spike Jonze film Her imagined a future where people woul... - https://arstechnica.com/ai/2025/03/users-report-emotional-bonds-with-startlingly-realistic-ai-voice-demo/ #machinelearning #voicesynthesis #aiassistants #emotionalai #chatbots #chatgpt #chatgtp #biz #sesame #tech #ai
-
OK, this is probably a rather long shot, but does anyone know of a voice synthesis model that is open-weights or ideally even open-source, and capable of producing intonation (specifically, something like rap lyrics)? #text2speech #voicesynthesis #ai
-
Your AI clone could target your family, but there’s a simple defense - On Tuesday, the US Federal Bureau of Investigation advised Americans to sh... - https://arstechnica.com/ai/2024/12/your-ai-clone-could-target-your-family-but-theres-a-simple-defense/ #machinelearning #audiosynthesis #imagesynthesis #voicesynthesis #voicecloning #aipassword #secretword #asaranear #deepfakes #aiethics #safeword #aicrime #aifraud #twitter #biz #fbi #ai
-
Man tricks OpenAI’s voice bot into duet of The Beatles’ “Eleanor Rigby” - Enlarge / A screen capture of AJ Smith doing his Eleanor Rigby duet wit... - https://arstechnica.com/?p=2052995 #largelanguagemodels #advancedvoicemode #aipromptinjection #machinelearning #promptinjection #audiosynthesis #musicsynthesis #voicesynthesis #paulmccartney #eleanorrigby #aicopyright #thebeatles #aifairuse #copyright #ajsmith #fairuse #biz #openai #ai
-
Due to AI fakes, the “deep doubt” era is here - Enlarge (credit: Memento | Aurich Lawson)
Given the flood of p... - https://arstechnica.com/?p=2042584 #daniellek.citron #machinelearning #imagesynthesis #musicsynthesis #videosynthesis #voicesynthesis #liarsdividend #medialiteracy #robertchesney #textsynthesis #kamalaharris #donaldtrump #socialmedia #deepdoubt #deepfakes #features #facebook #joebiden #politics #history #biz #doubt #ai #x
-
Google beats OpenAI to wide release of interruptible AI voice chat mode - Enlarge / The Google Gemini logo. (credit: Google)
On Thursday... - https://arstechnica.com/?p=2049707 #advancedvoicemode #machinelearning #voicesynthesis #googlegemini #aivoicechat #geminilive #chatgpt #chatgtp #biz #google #openai #ai
-
New Episode: hpr4188 :: Re: HPR4172 Comment by Ken Fallon
Hosted by Archer72 on 2024-08-21 is flagged as Clean and is released under a CC-BY-SA license.
Tags: #tts, #TextToSpeech, #VoiceSynthesis, #accessibility.
-
ChatGPT Advanced Voice Mode impresses testers with sound effects, catching its breath - Enlarge / A stock photo of a robot whispering to a man. (credit: Andrey... - https://arstechnica.com/?p=2040213 #largelanguagemodels #advancedvoicemode #aivoicegenerators #machinelearning #audiosynthesis #voicesynthesis #textsynthesis #chatgpt #chatgtp #biz #openai #ai
-
New Episode: hpr4172 :: Re: hpr4072 Piper voice synthesis
Hosted by Archer72 on 2024-07-30 is flagged as Clean and is released under a CC-BY-SA license.
Tags: #tts, #TextToSpeech, #VoiceSynthesis, #accessibility.
-
AI-generated Al Michaels to provide daily recaps during 2024 Summer Olympics - Enlarge / Al Michaels looks on prior to the game between the Minnesota ... - https://arstechnica.com/?p=2033917 #2024summerolympics #sportscommentators #machinelearning #speechsynthesis #summerolympics #voicesynthesis #olympicgames #almichaels #deepfakes #olympics #chatgpt #chatgtp #biz #sports #nbc #nfl #ai