Audio Node Catalog
Audio category
Generated from 12 catalog nodes in AI/Generative/Audio.
Nodes in this category
Showing 12 of 12 generated node docs.
Local Speech to Text
AI/Generative/AudioTranscribes audio locally with an installed any-speech-to-text model bit. Decodes WAV, MP3, FLAC, OGG (Vorbis/Opus), WebM/Opus, M4A/MP4 (AAC) and PCM, including browser MediaRecorder output (Chrome WebM/Opus, Safari MP4/AAC).
Local Text to Speech
AI/Generative/AudioGenerates WAV speech locally with an installed any-tts model bit.
Speech to Text
AI/Generative/AudioTranscribes or translates audio with an existing provider Bit.
Text to Speech
AI/Generative/AudioGenerates speech audio with an existing provider Bit and writes it to FlowPath.
Google STT Options
AI/Generative/Audio/OptionsCreates typed speech-to-text options for Gemini and Vertex audio transcription.
Google TTS Options
AI/Generative/Audio/OptionsCreates typed text-to-speech options for Gemini and Vertex speech models.
Hugging Face TTS Options
AI/Generative/Audio/OptionsCreates typed text-to-speech options for Hugging Face speech models.
Mistral TTS Options
AI/Generative/Audio/OptionsCreates typed text-to-speech options for Mistral speech models.
OpenAI-Compatible STT Options
AI/Generative/Audio/OptionsCreates typed speech-to-text options for OpenAI-compatible providers.
OpenAI-Compatible TTS Options
AI/Generative/Audio/OptionsCreates typed text-to-speech options for OpenAI-compatible providers.
xAI STT Options
AI/Generative/Audio/OptionsCreates typed speech-to-text options for xAI transcription models.
xAI TTS Options
AI/Generative/Audio/OptionsCreates typed text-to-speech options for xAI speech models.