home.social

#demucs — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #demucs, aggregated by home.social.

fetched live
  1. PVAD: как научить ИИ слышать нужного человека в шумной комнате

    Представьте переговорную, где одновременно говорят несколько человек, а системе нужно понять, когда говорит целевой спикер, и выделить именно его голос. Задача нетривиальная, но именно её нам и нужно было решить. Над этим мы с коллегами работали в проекте PVAD (Personal Voice Activity Detection) — технологии для определения того, когда в общем аудиопотоке говорит конкретный человек. Короткий простой обзор на один из кейсов нашей команды.

    habr.com/ru/articles/1069410/

    #машинное_обучение #выделение_голоса #шумоподавление #Demucs #xvector #ResNet #PyTorch #TensorFlow #LibriSpeech #VoxCeleb2

  2. It'd be nice if there was a #Demucs model that could separate laugh tracks from sitcom episodes. I know an #AI laugh track remover exists already, but to be honest, I wasn't impressed at all by the demo. It sounds like it just turned the episode all the way down when a laugh track came in. Unfortunately, I think the reason it can't happen easily yet is because there aren't many public domain croud sounds out there that you can just train AI on if any, or at least, not to my knowledge. #ML

  3. Well damn. By using '-d mps' on my M1 Max, I got #Demucs to run at about 21 seconds per-second.

  4. Well damn. By using '-d mps' on my M1 Max, I got #Demucs to run at about 21 seconds per-second.

  5. So, we now have GPU support from PyTorch on M1, but how can one utilize this for #Demucs if at all? I am not able to find any references!

  6. So, we now have GPU support from PyTorch on M1, but how can one utilize this for #Demucs if at all? I am not able to find any references!