High-quality text-to-speech using the ElevenLabs API. Like macOS `say` but with modern AI voices. Supports streaming, file output, voice selection, speed control, and multiple models (v3, v2, v2.5 Flash/Turbo).
High-quality text-to-speech using the ElevenLabs API. Like macOS `say` but with modern AI voices. Supports streaming, file output, voice selection, speed control, and multiple models (v3, v2, v2.5 Flash/Turbo).
使用 ElevenLabs API 的高质量文本转语音。类似 macOS 的 `say`,但采用现代 AI 语音。支持流式播放、文件输出、语音选择、速度控制和多种模型(v3、v2、v2.5 Flash/Turbo)。
Category: media-generate (媒体生成) · Author: internalkernel · Version: @main
该 Skill 暂无文档文件。
Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).
Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
Extract frames or short clips from videos using ffmpeg.
Search GIF providers with CLI/TUI, download results, and extract stills/sheets.
Category:media-generate