You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
I often want to dictate over long periods of time with my headphones on, and there's sometimes TV or music or other transient speech in the background. The mic picks all of those up and FluidVoice transcribes it, so transcription only reliably works in silent settings.
I think if we could specify a minimum volume floor/silence threshold for FluidVoice manually, we could make it work in noisier settings too. I have done some work with local transcription / diarization pipelines on my homeserver, my guess is we can expose the VAD threshold to the user without too much difficulty? What do you think - i can take a crack at a PR.
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
I often want to dictate over long periods of time with my headphones on, and there's sometimes TV or music or other transient speech in the background. The mic picks all of those up and FluidVoice transcribes it, so transcription only reliably works in silent settings.
I think if we could specify a minimum volume floor/silence threshold for FluidVoice manually, we could make it work in noisier settings too. I have done some work with local transcription / diarization pipelines on my homeserver, my guess is we can expose the VAD threshold to the user without too much difficulty? What do you think - i can take a crack at a PR.
All reactions