Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
matcha-video
7 months ago
|
parent
|
context
|
favorite
| on:
Show HN: OWhisper – Ollama for realtime speech-to-...
Question for folks who work a lot with STT models - What is your favorite model that supports word-level timestamps, has good dysfluency detection (whisper isn't great), and is also supported by transformers.js?
williamsss
7 months ago
[–]
A few months ago I had to work around a problem like this and the best out there was WhisperX. Not sure about transformer.js support.
Link to the repo -
https://github.com/m-bain/whisperX
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: