★ 0 · Updated 2026-09-21
Routes NVIDIA Nemotron Speech (Riva) NIM tasks for ASR, TTS, and NMT across cloud-hosted and self-hosted environments.
Browse skills that share this tag.
★ 0 · Updated 2026-09-21
Routes NVIDIA Nemotron Speech (Riva) NIM tasks for ASR, TTS, and NMT across cloud-hosted and self-hosted environments.
★ 96 · Updated 2026-09-14
Turn YouTube URLs or local videos into Chinese-localized videos with subtitles, dubbing, and audio mixing.
★ 3 · Updated 2026-09-13
Provides Python client library for real-time and batch speech-to-text transcription with timestamps and speaker diarization.
★ 281 · Updated 2026-09-12
Transcribe audio files to text with optional speaker diarization and known-speaker hints.
★ 0 · Updated 2026-09-09
Transcribes inbound audio clips using local faster-whisper to generate text transcripts, bulleted summaries, and action items.
★ 0 · Updated 2026-09-09
Transcribe audio files into text or JSON using the OpenAI Audio Transcriptions API.
★ 0 · Updated 2026-09-09
Transcribes audio files via OpenAI Audio Transcriptions API endpoint.
★ 18 · Updated 2026-09-09
Fixes speech-to-text misrecognitions, spacing, and hyphenation in dictated content related to GitHub Agentic Workflows.
★ 16 · Updated 2026-05-28
Develop real-time voice applications with OpenAI Realtime, Vapi, Deepgram, ElevenLabs, LiveKit, and WebRTC
★ 33 · Updated 2026-03-22
Text-to-speech, speech-to-text, voice conversion, and audio processing via EachLabs API
★ 816 · Updated 2026-03-21
High-performance audio library for Apple Silicon with text-to-speech and speech-to-text
★ 54 · Updated 2026-02-13
Transcribes audio files to text locally using faster-whisper in Docker containers
★ 1 · Updated 2026-02-12
Transcribe audio files using OpenAI's Whisper API