home.social

#voicetotext — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #voicetotext, aggregated by home.social.

fetched live
  1. Trasforma la voce in testo, note e comandi direttamente dal desktop con OpenWhispr, un progetto open source attento alla privacy e disponibile anche per Linux. #Linux #OpenSource #OpenWhispr #AI #Whisper #VoiceToText linuxeasy.org/openwhispr-porta

  2. Trasforma la voce in testo, note e comandi direttamente dal desktop con OpenWhispr, un progetto open source attento alla privacy e disponibile anche per Linux. #Linux #OpenSource #OpenWhispr #AI #Whisper #VoiceToText linuxeasy.org/openwhispr-porta

  3. Trasforma la voce in testo, note e comandi direttamente dal desktop con OpenWhispr, un progetto open source attento alla privacy e disponibile anche per Linux. #Linux #OpenSource #OpenWhispr #AI #Whisper #VoiceToText linuxeasy.org/openwhispr-porta

  4. Trasforma la voce in testo, note e comandi direttamente dal desktop con OpenWhispr, un progetto open source attento alla privacy e disponibile anche per Linux. #Linux #OpenSource #OpenWhispr #AI #Whisper #VoiceToText linuxeasy.org/openwhispr-porta

  5. Trasforma la voce in testo, note e comandi direttamente dal desktop con OpenWhispr, un progetto open source attento alla privacy e disponibile anche per Linux. #Linux #OpenSource #OpenWhispr #AI #Whisper #VoiceToText linuxeasy.org/openwhispr-porta

  6. OpenWhispr is a privacy-first, open-source voice-to-text app for Windows, macOS, and Linux. It can run transcription locally with Whisper or Parakeet, keeping your audio on your device.

    It also supports meeting transcription, AI agents, notes, translation, and local or cloud AI models.

    MIT licensed • Open source • Local AI • Cross-platform

    More details & alternatives: digitalescapetools.com/tools/t

    #OpenSource #FOSS #Privacy #LocalAI #VoiceToText

  7. OpenWhispr is a privacy-first, open-source voice-to-text app for Windows, macOS, and Linux. It can run transcription locally with Whisper or Parakeet, keeping your audio on your device.

    It also supports meeting transcription, AI agents, notes, translation, and local or cloud AI models.

    MIT licensed • Open source • Local AI • Cross-platform

    More details & alternatives: digitalescapetools.com/tools/t

    #OpenSource #FOSS #Privacy #LocalAI #VoiceToText

  8. OpenWhispr is a privacy-first, open-source voice-to-text app for Windows, macOS, and Linux. It can run transcription locally with Whisper or Parakeet, keeping your audio on your device.

    It also supports meeting transcription, AI agents, notes, translation, and local or cloud AI models.

    MIT licensed • Open source • Local AI • Cross-platform

    More details & alternatives: digitalescapetools.com/tools/t

    #OpenSource #FOSS #Privacy #LocalAI #VoiceToText

  9. OpenWhispr is a privacy-first, open-source voice-to-text app for Windows, macOS, and Linux. It can run transcription locally with Whisper or Parakeet, keeping your audio on your device.

    It also supports meeting transcription, AI agents, notes, translation, and local or cloud AI models.

    MIT licensed • Open source • Local AI • Cross-platform

    More details & alternatives: digitalescapetools.com/tools/t

    #OpenSource #FOSS #Privacy #LocalAI #VoiceToText

  10. OpenWhispr is a privacy-first, open-source voice-to-text app for Windows, macOS, and Linux. It can run transcription locally with Whisper or Parakeet, keeping your audio on your device.

    It also supports meeting transcription, AI agents, notes, translation, and local or cloud AI models.

    MIT licensed • Open source • Local AI • Cross-platform

    More details & alternatives: digitalescapetools.com/tools/t

    #OpenSource #FOSS #Privacy #LocalAI #VoiceToText

  11. Я ошибался: бенчмарк 23 ASR-нейросетей для русской айтишной диктовки

    В марте я советовал Whisper Large v3 для голосовой диктовки и ошибался дважды: в методе (выбирал модель «на ощупь») и в самой модели. Пересобрал всё по-научному: 23 ASR-модели в 60+ конфигурациях, больше 100 000 прогонов на русско-английском IT-корпусе, 120+ часов инференса на одной RTX 5070 Ti. Мерил не только WER, но и сохранение английских терминов латиницей и пунктуацию. Внутри — полный лидерборд, разбор каждого сюрприза, вклад тюнинга и промптов, cloud-против-open по деньгам и инструкция, как поставить победителя за 10 минут.

    habr.com/ru/articles/1066528/

    #whisper #breezeasr #распознавание_речи #ASR #voicetotext #whisper_large_v3 #whisper_turbo #gigaam_v3 #faster_whisper #codeswitching

  12. Я ошибался: бенчмарк 23 ASR-нейросетей для русской айтишной диктовки

    В марте я советовал Whisper Large v3 для голосовой диктовки и ошибался дважды: в методе (выбирал модель «на ощупь») и в самой модели. Пересобрал всё по-научному: 23 ASR-модели в 60+ конфигурациях, больше 100 000 прогонов на русско-английском IT-корпусе, 120+ часов инференса на одной RTX 5070 Ti. Мерил не только WER, но и сохранение английских терминов латиницей и пунктуацию. Внутри — полный лидерборд, разбор каждого сюрприза, вклад тюнинга и промптов, cloud-против-open по деньгам и инструкция, как поставить победителя за 10 минут.

    habr.com/ru/articles/1066528/

    #whisper #breezeasr #распознавание_речи #ASR #voicetotext #whisper_large_v3 #whisper_turbo #gigaam_v3 #faster_whisper #codeswitching

  13. Я ошибался: бенчмарк 23 ASR-нейросетей для русской айтишной диктовки

    В марте я советовал Whisper Large v3 для голосовой диктовки и ошибался дважды: в методе (выбирал модель «на ощупь») и в самой модели. Пересобрал всё по-научному: 23 ASR-модели в 60+ конфигурациях, больше 100 000 прогонов на русско-английском IT-корпусе, 120+ часов инференса на одной RTX 5070 Ti. Мерил не только WER, но и сохранение английских терминов латиницей и пунктуацию. Внутри — полный лидерборд, разбор каждого сюрприза, вклад тюнинга и промптов, cloud-против-open по деньгам и инструкция, как поставить победителя за 10 минут.

    habr.com/ru/articles/1066528/

    #whisper #breezeasr #распознавание_речи #ASR #voicetotext #whisper_large_v3 #whisper_turbo #gigaam_v3 #faster_whisper #codeswitching

  14. 𝗛𝗮𝗻𝗱𝘆 𝗖𝗼𝗺𝗽𝘂𝘁𝗲𝗿:

    #VoiceToText #OpenSource

    thewhale.cc/posts/handy-comput

    Speak into any text field, the free and open source app for speech to text.

  15. 𝗛𝗮𝗻𝗱𝘆 𝗖𝗼𝗺𝗽𝘂𝘁𝗲𝗿:

    #VoiceToText #OpenSource

    thewhale.cc/posts/handy-comput

    Speak into any text field, the free and open source app for speech to text.

  16. 𝗛𝗮𝗻𝗱𝘆 𝗖𝗼𝗺𝗽𝘂𝘁𝗲𝗿:

    #VoiceToText #OpenSource

    thewhale.cc/posts/handy-comput

    Speak into any text field, the free and open source app for speech to text.

  17. Google Unleashes 'AI Edge Eloquent', A Free, Offline Dictation Tool

    Google launched AI Edge Eloquent, a free dictation app that works offline. It keeps your voice data private by processing it on your phone. Available now on iOS.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews

    newsletter.tf/google-ai-edge-e

  18. Google Unleashes 'AI Edge Eloquent', A Free, Offline Dictation Tool

    Google launched AI Edge Eloquent, a free dictation app that works offline. It keeps your voice data private by processing it on your phone. Available now on iOS.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews

    newsletter.tf/google-ai-edge-e

  19. Google Unleashes 'AI Edge Eloquent', A Free, Offline Dictation Tool

    Google launched AI Edge Eloquent, a free dictation app that works offline. It keeps your voice data private by processing it on your phone. Available now on iOS.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews

    newsletter.tf/google-ai-edge-e

  20. Google Unleashes 'AI Edge Eloquent', A Free, Offline Dictation Tool

    Google launched AI Edge Eloquent, a free dictation app that works offline. It keeps your voice data private by processing it on your phone. Available now on iOS.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews

    newsletter.tf/google-ai-edge-e

  21. Google's new free dictation app, AI Edge Eloquent, works offline and keeps your voice data private. This is a new option for users worried about cloud storage.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews
    newsletter.tf/google-ai-edge-e

  22. Google's new free dictation app, AI Edge Eloquent, works offline and keeps your voice data private. This is a new option for users worried about cloud storage.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews
    newsletter.tf/google-ai-edge-e

  23. Google's new free dictation app, AI Edge Eloquent, works offline and keeps your voice data private. This is a new option for users worried about cloud storage.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews
    newsletter.tf/google-ai-edge-e

  24. Google's new free dictation app, AI Edge Eloquent, works offline and keeps your voice data private. This is a new option for users worried about cloud storage.

    #AIDictation, #GoogleApp, #OfflinePrivacy, #VoiceToText, #TechNews
    newsletter.tf/google-ai-edge-e

  25. Локальный голосовой ввод: Whisper + Ollama на Python

    Мне нужен был голосовой ввод. Не диктовка в Google Docs, не облачный API — а простая штука: зажал клавишу, сказал, отпустил, текст появился в активном окне. Локально, без отправки аудио куда-либо. Готовых решений, которые бы устроили, сходу не нашёл. Сделал свое. Может, кому будет полезно.

    habr.com/ru/articles/1009538/

    #whisper #ollama #speechtotext #voicetotext #pushtotalk #голосовой_ввод #python #localfirst #privacy

  26. Локальный голосовой ввод: Whisper + Ollama на Python

    Мне нужен был голосовой ввод. Не диктовка в Google Docs, не облачный API — а простая штука: зажал клавишу, сказал, отпустил, текст появился в активном окне. Локально, без отправки аудио куда-либо. Готовых решений, которые бы устроили, сходу не нашёл. Сделал свое. Может, кому будет полезно.

    habr.com/ru/articles/1009538/

    #whisper #ollama #speechtotext #voicetotext #pushtotalk #голосовой_ввод #python #localfirst #privacy

  27. Локальный голосовой ввод: Whisper + Ollama на Python

    Мне нужен был голосовой ввод. Не диктовка в Google Docs, не облачный API — а простая штука: зажал клавишу, сказал, отпустил, текст появился в активном окне. Локально, без отправки аудио куда-либо. Готовых решений, которые бы устроили, сходу не нашёл. Сделал свое. Может, кому будет полезно.

    habr.com/ru/articles/1009538/

    #whisper #ollama #speechtotext #voicetotext #pushtotalk #голосовой_ввод #python #localfirst #privacy

  28. People have been raving to me about how good voice-to-text has become, so I'm trying Wispr Flow. But the user experience feels a bit off: I have to stop recording to get it to type out the text? Not really "hands-free", is it?
    Is that just how it works, or is there a continuous dictation mode I'm missing? Are there other LLM-based tools that behave more like classic dictation?
    Looking forward to your recommendations!

    #VoiceToText #AI #Productivity

  29. People have been raving to me about how good voice-to-text has become, so I'm trying Wispr Flow. But the user experience feels a bit off: I have to stop recording to get it to type out the text? Not really "hands-free", is it?
    Is that just how it works, or is there a continuous dictation mode I'm missing? Are there other LLM-based tools that behave more like classic dictation?
    Looking forward to your recommendations!

    #VoiceToText #AI #Productivity

  30. People have been raving to me about how good voice-to-text has become, so I'm trying Wispr Flow. But the user experience feels a bit off: I have to stop recording to get it to type out the text? Not really "hands-free", is it?
    Is that just how it works, or is there a continuous dictation mode I'm missing? Are there other LLM-based tools that behave more like classic dictation?
    Looking forward to your recommendations!

    #VoiceToText #AI #Productivity

  31. People have been raving to me about how good voice-to-text has become, so I'm trying Wispr Flow. But the user experience feels a bit off: I have to stop recording to get it to type out the text? Not really "hands-free", is it?
    Is that just how it works, or is there a continuous dictation mode I'm missing? Are there other LLM-based tools that behave more like classic dictation?
    Looking forward to your recommendations!

    #VoiceToText #AI #Productivity

  32. People have been raving to me about how good voice-to-text has become, so I'm trying Wispr Flow. But the user experience feels a bit off: I have to stop recording to get it to type out the text? Not really "hands-free", is it?
    Is that just how it works, or is there a continuous dictation mode I'm missing? Are there other LLM-based tools that behave more like classic dictation?
    Looking forward to your recommendations!

    #VoiceToText #AI #Productivity

  33. SoundVibes 0.2.0 is out!

    - More people than me that needed to handle multiple languages without the delay of the autodetection.
    - Now possible to toggle specific languages to transcribe.

    Still modest amount of users, but ⭐ are rising as well as downloads.

    #voicetotext made easy on #linux!

    github.com/kejne/soundvibes/re

  34. SoundVibes 0.2.0 is out!

    - More people than me that needed to handle multiple languages without the delay of the autodetection.
    - Now possible to toggle specific languages to transcribe.

    Still modest amount of users, but ⭐ are rising as well as downloads.

    made easy on !

    github.com/kejne/soundvibes/re

  35. 🎤✨ Wispr Flow is coming to Android and I'm HYPED! 🚀

    Say goodbye to typing and hello to pure voice magic ✍️➡️🎤 This game-changing AI tool is launching Feb 12 and honestly, I can't wait!

    Move up the waitlist and get early access: wisprflow.ai/waitlist?ADONTAI1

    Drop your referral link with mine and we'll both climb the ranks together! 💪

    #WisperFlow #Android #VoiceToText #AI #Tech #Innovation #Productivity #MustHave #EarlyAccess #JoinTheWaitlist #androidapp #aitools #productivity #VoiceTech

  36. 🎤✨ Wispr Flow is coming to Android and I'm HYPED! 🚀

    Say goodbye to typing and hello to pure voice magic ✍️➡️🎤 This game-changing AI tool is launching Feb 12 and honestly, I can't wait!

    Move up the waitlist and get early access: wisprflow.ai/waitlist?ADONTAI1

    Drop your referral link with mine and we'll both climb the ranks together! 💪

    #WisperFlow #Android #VoiceToText #AI #Tech #Innovation #Productivity #MustHave #EarlyAccess #JoinTheWaitlist #androidapp #aitools #productivity #VoiceTech

  37. 🎤✨ Wispr Flow is coming to Android and I'm HYPED! 🚀

    Say goodbye to typing and hello to pure voice magic ✍️➡️🎤 This game-changing AI tool is launching Feb 12 and honestly, I can't wait!

    Move up the waitlist and get early access: wisprflow.ai/waitlist?ADONTAI1

    Drop your referral link with mine and we'll both climb the ranks together! 💪

    #WisperFlow #Android #VoiceToText #AI #Tech #Innovation #Productivity #MustHave #EarlyAccess #JoinTheWaitlist #androidapp #aitools #productivity #VoiceTech

  38. 🎤✨ Wispr Flow is coming to Android and I'm HYPED! 🚀

    Say goodbye to typing and hello to pure voice magic ✍️➡️🎤 This game-changing AI tool is launching Feb 12 and honestly, I can't wait!

    Move up the waitlist and get early access: wisprflow.ai/waitlist?ADONTAI1

    Drop your referral link with mine and we'll both climb the ranks together! 💪

    #WisperFlow #Android #VoiceToText #AI #Tech #Innovation #Productivity #MustHave #EarlyAccess #JoinTheWaitlist #androidapp #aitools #productivity #VoiceTech

  39. 🎤✨ Wispr Flow is coming to Android and I'm HYPED! 🚀

    Say goodbye to typing and hello to pure voice magic ✍️➡️🎤 This game-changing AI tool is launching Feb 12 and honestly, I can't wait!

    Move up the waitlist and get early access: wisprflow.ai/waitlist?ADONTAI1

    Drop your referral link with mine and we'll both climb the ranks together! 💪

    #WisperFlow #Android #VoiceToText #AI #Tech #Innovation #Productivity #MustHave #EarlyAccess #JoinTheWaitlist #androidapp #aitools #productivity #VoiceTech

  40. Happy to see some issues filed already! 😀

    I managed to publish #soundvibes just before getting a horrible cold, just to "get back" to two useful feature requests and a bug report (which someone already seems to be drafting a PR for).

    Grateful to see #opensource contributions so fast and happy that my tool fills a need! 🙌

    github.com/kejne/soundvibes

    #voicetotext #linux

  41. Happy to see some issues filed already! 😀

    I managed to publish #soundvibes just before getting a horrible cold, just to "get back" to two useful feature requests and a bug report (which someone already seems to be drafting a PR for).

    Grateful to see #opensource contributions so fast and happy that my tool fills a need! 🙌

    github.com/kejne/soundvibes

    #voicetotext #linux

  42. Happy to see some issues filed already! 😀

    I managed to publish just before getting a horrible cold, just to "get back" to two useful feature requests and a bug report (which someone already seems to be drafting a PR for).

    Grateful to see contributions so fast and happy that my tool fills a need! 🙌

    github.com/kejne/soundvibes

  43. Happy to see some issues filed already! 😀

    I managed to publish #soundvibes just before getting a horrible cold, just to "get back" to two useful feature requests and a bug report (which someone already seems to be drafting a PR for).

    Grateful to see #opensource contributions so fast and happy that my tool fills a need! 🙌

    github.com/kejne/soundvibes

    #voicetotext #linux

  44. Happy to see some issues filed already! 😀

    I managed to publish #soundvibes just before getting a horrible cold, just to "get back" to two useful feature requests and a bug report (which someone already seems to be drafting a PR for).

    Grateful to see #opensource contributions so fast and happy that my tool fills a need! 🙌

    github.com/kejne/soundvibes

    #voicetotext #linux

  45. Remember me posting about running my agent while marking footballs during the weekend?

    Well, now it's time to share the results!

    An open source voice-to-text application for Linux which enables you to hotkey speech capture to input your voice wherever your cursor is!

    I was annoyed by the complexity of the tools that were available so I created one which comes as a single binary, written in Rust.

    Check it out:
    soundvibes.teashaped.dev/

    Wrote a blog post about the creation of it:
    teashaped.dev/blog/soundvibes-

    #linuxvtt #voicetotext #vibecoding #opensource #FAAFO

  46. Remember me posting about running my agent while marking footballs during the weekend?

    Well, now it's time to share the results!

    An open source voice-to-text application for Linux which enables you to hotkey speech capture to input your voice wherever your cursor is!

    I was annoyed by the complexity of the tools that were available so I created one which comes as a single binary, written in Rust.

    Check it out:
    soundvibes.teashaped.dev/

    Wrote a blog post about the creation of it:
    teashaped.dev/blog/soundvibes-

  47. DeepFlo – công cụ viết & chỉnh sửa văn bản bằng giọng nói, hoạt động trên mọi nền tảng (email, Slack, VS Code,…). Đang tìm beta tester macOS để “roast” và đưa ra phản hồi thẳng thắn. Cần trả lời: giá trị ngay lập tức? Khác biệt so với dictation cơ bản? Đối tượng người dùng thực sự là ai? Nếu muốn thử, hãy để lại ý kiến!

    #BetaTest #VoiceToText #SaaS #DeepFlo #CôngNghệ #KiểmThử #AI #ỨngDụng #TríTuệNhânTạo

    reddit.com/r/SaaS/comments/1qo

  48. 🗣️ Công cụ mới: "Unchained Vibes for Claude" – extension Chrome cho phép nói chuyện với Claude bằng giọng nói, tự động chuyển sang văn bản. Tính năng: tạm dừng để suy nghĩ, chụp màn hình khi nói lệnh, kích hoạt bằng cụm từ đặc biệt. Tiện lợi cho ai mệt mỏi khi gõ! #AI #Claude #VoiceToText #ChromeExtension #CôngNghệ #TiếngNói #TríTuệNhânTạo

    reddit.com/r/SideProject/comme

  49. 🚀 Đang xây dựng app chuyển giọng nói thành To‑Do, ghi chú, nhật ký ngay lập tức, sắp xếp theo thư mục. Không còn đống văn bản thô, mỗi câu nói sẽ thành nhiệm vụ có thể đánh dấu hoàn thành. Muốn thử sớm? Tham gia danh sách chờ! #VoiceToText #ToDo #Notes #Journal #Productivity #CôngNghệ #GhiChú #CôngViệc

    reddit.com/r/SideProject/comme

  50. Một founder chia sẻ thất bại: sau 1 tháng phát triển công cụ chuyển giọng nói sang email thông minh, phải nghỉ du lịch, đối thủ ra mắt và bị WisprFlow mua. Thêm rào cản CASA của Google, chi phí $500‑$1800. Hiện cân nhắc bán sản phẩm hoặc từ bỏ. Các bạn có lời khuyên gì? 🙏 #SaaS #Startup #MicroSaaS #VoiceToText #EmailAutomation #Entrepreneur #KhởiNghiệp #CôngNghệ

    reddit.com/r/SaaS/comments/1qm

  51. 🚀 Ứng dụng mã nguồn mở chuyển giọng nói thành văn bản, chạy trên Linux & Windows! Dùng sherpa‑onnx + liteLLM, hỗ trợ từ vựng tùy chỉnh, xử lý thông minh qua LLM (loại bỏ "um", sửa ngữ pháp) và chạy mô hình Whisper hoặc Nvidia Parakeet. Mã nguồn và bản phát hành có trên GitHub. #OpenSource #VoiceToText #Linux #Windows #AI #LLM #sherpa #liteLLM

    reddit.com/r/LocalLLaMA/commen