Transcribe or translate audio files using OpenAI Whisper. Supports all common audio formats, multiple languages, and outputs text/json/srt/vtt. Use for speech-to-text, meeting transcription, subtitle generation, and audio translation.
Transcribe or translate audio files using OpenAI Whisper. Supports all common audio formats, multiple languages, and outputs text/json/srt/vtt. Use for speech-to-text, meeting transcription, subtitle generation, and audio translation.
使用 OpenAI Whisper 转录或翻译音频文件。支持所有常见音频格式、多种语言,输出包括 text/json/srt/vtt。适用于语音转文本、会议转录、字幕生成和音频翻译等场景。
Category: media-generate (媒体生成) · Author: internalkernel · Version: @main
该 Skill 暂无文档文件。
Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).
Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
Extract frames or short clips from videos using ffmpeg.
Search GIF providers with CLI/TUI, download results, and extract stills/sheets.
Category:media-generate