Speech to text: Difference between revisions

Jump to navigation Jump to search
mNo edit summary
 
Line 45: Line 45:
[https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova]
[https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova]
* Input file: Audio files
* Input file: Audio files
* Support Language: English
* Support Language: English {{exclaim}} Even if the audio file contains a mix of two languages (e.g. Mandarin Chinese/English), the final output will only be in English. There is also no prompt available to modify this behavior.
* Speaker identification: No {{exclaim}}
* Speaker identification: No {{exclaim}}
* Output file format: TXT or JSON (contains timestamp info.)
* Output file format: TXT or JSON (contains timestamp info.)
Line 97: Line 97:
* Price: Free or Pro plan
* Price: Free or Pro plan
* Output file format: TXT, DOCX, SRT, VTT, JSON and more
* Output file format: TXT, DOCX, SRT, VTT, JSON and more
* Notes:  
* Notes:


== Speech to text API ==
== Speech to text API ==

Navigation menu