Convert text to natural speech with multiple Chinese voices and emotion support.
将文字转换为自然语音,支持多种中文音色和情绪
FAL_KEY:fal.ai API Keypip install -r requirements.txt
python run.py --text "你好世界" --voice sweet_lady --output ./output.mp3
python run.py --text "你好,世界!" --voice sweet_lady
python run.py --text "重要公告" --voice executive --speed 0.9 --emotion serious
python run.py --list-voices
python run.py --test-all-voices --output-dir ./voice_test
Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).
Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
Extract frames or short clips from videos using ffmpeg.
Search GIF providers with CLI/TUI, download results, and extract stills/sheets.
Category:media-generate