Models
By default, Whisper produces per-sentence timestamp segmentation.
whisper-timestamped gives per-word timestamps.
Example
Supported formats
mp3wav
Additional parameters
Each Whisper variant supports parameters likelanguage, task (transcribe vs. translate), and more. Check the model’s documentation page for the full list: