★ 740 · Updated 2026-09-28
Generates and processes speech, subtitles, music, images, and videos through local AI workflows, with audio analysis and video editing.
Browse skills that use this input.
★ 740 · Updated 2026-09-28
Generates and processes speech, subtitles, music, images, and videos through local AI workflows, with audio analysis and video editing.
★ 1 · Updated 2026-09-24
Aligns audio_map2 monthly JSON timestamps to millisecond precision based on Word text and HTML question boundaries using SRT and FunASR.
★ 1 · Updated 2026-09-18
Transcribes audio into text and translates speech into English across 99 languages using the Whisper model.
★ 3 · Updated 2026-09-16
A multi-featured summarization tool that automatically detects content types, selects summary depth, and processes text, PDFs, audio, images, and videos.
★ 0 · Updated 2026-09-09
Generates spectrograms and feature-panel visualizations from audio using the songsee CLI.
★ 0 · Updated 2026-09-09
Provides CLI commands to upload, update, delete, like, unlike, and download tracks on plyr.fm.
★ 650 · Updated 2026-09-09
Guides users to perform mutation actions on plyr.fm via CLI, such as uploading, updating, deleting, liking, and downloading tracks.
★ 33 · Updated 2026-03-24
Separate speakers from audio files, compare voice samples, and process audio segments using Demucs, pyannote, and Resemblyzer.
★ 25 · Updated 2026-03-10
Unity engine fundamentals and workflows for 3D asset integration
★ 22 · Updated 2026-02-15
Fine-tunes text-to-speech models using Unsloth's optimized framework for voice cloning and speech synthesis.
★ 276 · Updated 2026-02-11
Generate professional storyboard prompts and create videos through Jimeng API with automatic download