Skip to main content
This node transcribes audio into text using the Fish Audio speech-to-text service. It automatically detects the language of the audio and can optionally return word-level timestamped segments as JSON.

Inputs

Note: The language parameter is only a hint — the language is always auto-detected from the audio. When precise_timestamps is false (the default), word-level timestamps are not returned; when true, the output segments include word-level timestamps.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): eaf1c9a9d2b90ec962a408615cc417b552864354c3f272144b8e239b23961920