Speech to text: Difference between revisions

Jump to navigation Jump to search
No edit summary
 
(One intermediate revision by the same user not shown)
Line 1: Line 1:
Comparison of speech to text (transcription) software
Comparison of speech to text (transcription) software
{{Template:Generative AI Tool}}


== Speech to text software ==
== Speech to text software ==
Line 43: Line 45:
[https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova]
[https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova]
* Input file: Audio files
* Input file: Audio files
* Support Language: English
* Support Language: English {{exclaim}} Even if the audio file contains a mix of two languages (e.g. Mandarin Chinese/English), the final output will only be in English. There is also no prompt available to modify this behavior.
* Speaker identification: No {{exclaim}}
* Speaker identification: No {{exclaim}}
* Output file format: TXT or JSON (contains timestamp info.)
* Output file format: TXT or JSON (contains timestamp info.)
Line 95: Line 97:
* Price: Free or Pro plan
* Price: Free or Pro plan
* Output file format: TXT, DOCX, SRT, VTT, JSON and more
* Output file format: TXT, DOCX, SRT, VTT, JSON and more
* Notes:  
* Notes:


== Speech to text API ==
== Speech to text API ==

Navigation menu