r/selfhosted 15h ago

Software Development Running multiple speech models in parallel in real time

https://nanosamur.ai/blog/posts/qwen-vs-whisper/

In this comparison, I am streaming the same audio stream to different asr models (whisper vs qwen) in parallel and observing the results in the real time.

All the code used is OSS and linked in github.

7 Upvotes

2 comments sorted by

u/asimovs-auditor 15h ago edited 15h ago

Expand the replies to this comment to learn how AI was used in this post/project.

→ More replies (1)