15,049
edits
No edit summary |
|||
| (One intermediate revision by the same user not shown) | |||
| Line 1: | Line 1: | ||
Comparison of speech to text (transcription) software | Comparison of speech to text (transcription) software | ||
{{Template:Generative AI Tool}} | |||
== Speech to text software == | == Speech to text software == | ||
| Line 43: | Line 45: | ||
[https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova] | [https://huggingface.co/spaces/Xenova/whisper-web Whisper Web - a Hugging Face Space by Xenova] | ||
* Input file: Audio files | * Input file: Audio files | ||
* Support Language: English | * Support Language: English {{exclaim}} Even if the audio file contains a mix of two languages (e.g. Mandarin Chinese/English), the final output will only be in English. There is also no prompt available to modify this behavior. | ||
* Speaker identification: No {{exclaim}} | * Speaker identification: No {{exclaim}} | ||
* Output file format: TXT or JSON (contains timestamp info.) | * Output file format: TXT or JSON (contains timestamp info.) | ||
| Line 95: | Line 97: | ||
* Price: Free or Pro plan | * Price: Free or Pro plan | ||
* Output file format: TXT, DOCX, SRT, VTT, JSON and more | * Output file format: TXT, DOCX, SRT, VTT, JSON and more | ||
* Notes: | * Notes: | ||
== Speech to text API == | == Speech to text API == | ||