LogoClawIndex
CasesSkillsAbout
LogoClawIndex

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.

Skills tagged: speech-to-text

Browse skills that share this tag.

  • nemotron-speech - Deploy, run, and test NVIDIA Nemotron Speech NIMs.
    nvidianemotron-speechrivanim

    ★ 0 · Updated 2026-09-21

    Routes NVIDIA Nemotron Speech (Riva) NIM tasks for ASR, TTS, and NMT across cloud-hosted and self-hosted environments.

    ⚙ Deploy ASR TTS and NMT NIMs⚙ Run cloud or self hosted inference⚙ Convert custom NeMo models to NIM
  • youtube-chinese-localizer - Localize YouTube or foreign-language videos into Chinese
    video-localizationyoutubesubtitlesdubbing

    ★ 96 · Updated 2026-09-14

    Turn YouTube URLs or local videos into Chinese-localized videos with subtitles, dubbing, and audio mixing.

    ⚙ Download or ingest source videos⚙ Transcribe speech to timestamped segments⚙ Translate lines into Simplified Chinese
  • azure-ai-transcription-py - Azure AI Transcription SDK for Python
    azurespeech-to-textpythontranscription

    ★ 3 · Updated 2026-09-13

    Provides Python client library for real-time and batch speech-to-text transcription with timestamps and speaker diarization.

    ⚙ Perform batch speech transcription⚙ Perform real-time stream transcription⚙ List transcription jobs
  • transcribe - Audio Transcription and Speaker Diarization
    audiotranscriptiondiarizationspeech-to-text

    ★ 281 · Updated 2026-09-12

    Transcribe audio files to text with optional speaker diarization and known-speaker hints.

    ⚙ Transcribe audio files⚙ Perform speaker diarization⚙ Label known speakers
  • audio-transcription - Transcribe audio clips to text using local Whisper
    audio-transcriptionwhisperspeech-to-textfaster-whisper

    ★ 0 · Updated 2026-09-09

    Transcribes inbound audio clips using local faster-whisper to generate text transcripts, bulleted summaries, and action items.

    ⚙ Transcribe inbound audio files⚙ Generate bulleted text summaries⚙ Extract action items and questions
  • openai-whisper-api - Transcribe audio via OpenAI Whisper API
    audio-transcriptionopenaiwhisperspeech-to-text

    ★ 0 · Updated 2026-09-09

    Transcribe audio files into text or JSON using the OpenAI Audio Transcriptions API.

    ⚙ Transcribe audio files⚙ Export transcription as JSON⚙ Set transcription language
  • openai-whisper-api - Transcribe audio files using OpenAI Whisper API
    audio-transcriptionspeech-to-textopenaiwhisper

    ★ 0 · Updated 2026-09-09

    Transcribes audio files via OpenAI Audio Transcriptions API endpoint.

    ⚙ Transcribe audio files⚙ Specify transcription language⚙ Set transcription prompt
  • Dictation Instructions - Fix speech-to-text errors in dictated content
    dictationspeech-to-textgithub-agentic-workflows

    ★ 18 · Updated 2026-09-09

    Fixes speech-to-text misrecognitions, spacing, and hyphenation in dictated content related to GitHub Agentic Workflows.

    ⚙ Fix speech-to-text errors⚙ Correct spacing and hyphenation⚙ Apply project-specific terms
  • voice-ai-development - Build Real-Time Voice AI Applications
    voice-aireal-timespeech-recognitiontext-to-speech

    ★ 16 · Updated 2026-05-28

    Develop real-time voice applications with OpenAI Realtime, Vapi, Deepgram, ElevenLabs, LiveKit, and WebRTC

    ⚙ create real-time session⚙ listen to events⚙ convert speech to text
  • eachlabs-voice-audio - EachLabs Voice & Audio Processing
    text-to-speechspeech-to-textvoice conversionaudio processing

    ★ 33 · Updated 2026-03-22

    Text-to-speech, speech-to-text, voice conversion, and audio processing via EachLabs API

    ⚙ convert text to speech⚙ transcribe speech to text⚙ convert voice audio
  • qwen3-audio - High-performance audio library for Apple Silicon
    text-to-speechspeech-to-textvoice cloningaudio processing

    ★ 816 · Updated 2026-03-21

    High-performance audio library for Apple Silicon with text-to-speech and speech-to-text

    ⚙ convert text to speech⚙ clone voice from audio⚙ create voice from text
  • whisper-transcribe-docker - Local Speech-to-Text Transcription with Docker
    speech-to-texttranscriptiondockerlocal-processing

    ★ 54 · Updated 2026-02-13

    Transcribes audio files to text locally using faster-whisper in Docker containers

    ⚙ transcribe audio files⚙ generate timestamped transcripts⚙ output text or JSON formats
  • openai-whisper-api - OpenAI Whisper API Audio Transcription
    audiotranscriptionspeech-to-textopenai

    ★ 1 · Updated 2026-02-12

    Transcribe audio files using OpenAI's Whisper API

    ⚙ transcribe audio files⚙ output text transcripts⚙ output JSON transcripts