Skip to content

Audio Node Catalog

Audio category

Generated from 12 catalog nodes in AI/Generative/Audio.

AI/Generative/Audio/Options

Nodes in this category

Showing 12 of 12 generated node docs.

Local Speech to Text

AI/Generative/Audio

Transcribes audio locally with an installed any-speech-to-text model bit. Decodes WAV, MP3, FLAC, OGG (Vorbis/Opus), WebM/Opus, M4A/MP4 (AAC) and PCM, including browser MediaRecorder output (Chrome WebM/Opus, Safari MP4/AAC).

Local Text to Speech

AI/Generative/Audio

Generates WAV speech locally with an installed any-tts model bit.

Speech to Text

AI/Generative/Audio

Transcribes or translates audio with an existing provider Bit.

Text to Speech

AI/Generative/Audio

Generates speech audio with an existing provider Bit and writes it to FlowPath.

Google STT Options

AI/Generative/Audio/Options

Creates typed speech-to-text options for Gemini and Vertex audio transcription.

Google TTS Options

AI/Generative/Audio/Options

Creates typed text-to-speech options for Gemini and Vertex speech models.

Hugging Face TTS Options

AI/Generative/Audio/Options

Creates typed text-to-speech options for Hugging Face speech models.

Mistral TTS Options

AI/Generative/Audio/Options

Creates typed text-to-speech options for Mistral speech models.

OpenAI-Compatible STT Options

AI/Generative/Audio/Options

Creates typed speech-to-text options for OpenAI-compatible providers.

OpenAI-Compatible TTS Options

AI/Generative/Audio/Options

Creates typed text-to-speech options for OpenAI-compatible providers.

xAI STT Options

AI/Generative/Audio/Options

Creates typed speech-to-text options for xAI transcription models.

xAI TTS Options

AI/Generative/Audio/Options

Creates typed text-to-speech options for xAI speech models.