podcast-generation

1.8k
openclawopenclaw

Generate AI-powered podcast-style audio narratives using Azure OpenAI's GPT Realtime Mini model via WebSocket. Use when building text-to-speech features, audio narrative generation, podcast creation from content, or integrating with Azure OpenAI Realtime API for real audio output. Covers full-stack implementation from React frontend to Python FastAPI backend with WebSocket streaming.

192 days ago

podcast-generation

1.6k
microsoftmicrosoft

Generate AI-powered podcast-style audio narratives using Azure OpenAI's GPT Realtime Mini model via WebSocket. Use when building text-to-speech features, audio narrative generation, podcast creation from content, or integrating with Azure OpenAI Realtime API for real audio output. Covers full-stack implementation from React frontend to Python FastAPI backend with WebSocket streaming.

191 days ago

generating-sounds-with-ai

293
raphaelsalajaraphaelsalaja

Audit Web Audio API code for sound synthesis best practices. Use when reviewing procedural audio, implementing UI sounds, or checking audio parameter quality. Outputs file:line findings.

192 days ago

heygen

48
heygen-comheygen-com

HeyGen AI video creation API. Use when: (1) Using Video Agent for one-shot prompt-to-video generation, (2) Generating AI avatar videos with /v2/video/generate, (3) Working with HeyGen avatars, voices, backgrounds, or captions, (4) Creating transparent WebM videos for compositing, (5) Polling video status or handling webhooks, (6) Integrating HeyGen with Remotion for programmatic video, (7) Translating or dubbing existing videos, (8) Generating standalone TTS audio with the Starfish model via /v1/audio.

192 days ago

win31-audio-design

43
curiositechcuriositech

Expert in Windows 3.1 era sound vocabulary for modern web/mobile apps. Creates satisfying retro UI sounds using CC-licensed 8-bit audio, Web Audio API, and haptic coordination. Activate on 'win31 sounds', 'retro audio', '90s sound effects', 'chimes', 'tada', 'ding', 'satisfying UI sounds'. NOT for modern flat UI sounds, voice synthesis, or music composition.

audioretrowindows+2
192 days ago

2000s-visualization-expert

43
curiositechcuriositech

Expert in 2000s-era music visualization (Milkdrop, AVS, Geiss) and modern WebGL implementations. Specializes in Butterchurn integration, Web Audio API AnalyserNode FFT data, GLSL shaders for audio-reactive visuals, and psychedelic generative art. Activate on "Milkdrop", "music visualization", "WebGL visualizer", "Butterchurn", "audio reactive", "FFT visualization", "spectrum analyzer". NOT for simple bar charts/waveforms (use basic canvas), video editing, or non-audio visuals.

audiowebglvisualization+2
192 days ago

web-audio-api

25
martinholovskymartinholovsky

Web Audio API for JARVIS audio feedback and voice processing

192 days ago

Audio Producer

17
daffy0208daffy0208

Expert in web audio, audio processing, and interactive sound design

audioweb-audio-apisound-design+2
192 days ago

videodb

17
video-dbvideo-db

What you get A single API-first video stack for agents. Ingest anything, process server-side, and ship playable streams without FFmpeg glue. • Core capabilities Ingest + Transcode - Accept any format - Change codec, bitrate, FPS, resolution - Output a playable stream for your app (CDN + hosting) Scene-level Search Engine - Build a searchable index of your media, scene by scene - Find exact moments and auto-create clips - "Indexes as code": describe what you need, programmatically - Manage 1000s of hours of footage cleanly Generate + Compose - Generate assets: image, audio, video, text - Overlay text/images, branding, motion captions - Dub videos, translate captions, add animations Real-time RTSP - Connect live streams and create understanding in real time - Define events and set up alerts - Ideal for security cams and monitoring workflows Desktop Perception - Capture screen, mic, and system audio for real-time context - Stream your desktop live - Define alerts and triggers from what's happening on screen - Store episodic memory and semantic-search sessions - Record local sessions for QA and review • Try it now "Ingest this file and give me a playable web stream link" "Generate subtitles, burn them in, and add light background music" "Index this folder and find every scene with people" "Connect this RTSP URL and alert when a person enters the zone" "start recording and give me a actionable summary when it ends"

191 days ago

audio-analysis

7
Bbeierle12Bbeierle12

Audio analysis with Tone.js and Web Audio API including FFT, frequency data extraction, amplitude measurement, and waveform analysis. Use when extracting audio data for visualizations, beat detection, or any audio-reactive features.

192 days ago

ui-audio-theme

6
b-open-iob-open-io

Generate cohesive UI audio themes with subtle, minimal sound effects for applications. This skill should be used when users want to create a set of coordinated interface sounds for wallet apps, dashboards, or web applications - generating sounds mapped to UI interaction constants like button clicks, notifications, and navigation transitions using ElevenLabs API.

192 days ago

threejs

4
hoadhhoadh

Build 3D web apps with Three.js (WebGL/WebGPU). 556 searchable examples, 60 API classes, 20 use cases. Actions: create 3D scene, load model, add animation, implement physics, build VR/XR. Topics: GLTF loader, PBR materials, particle effects, shadows, post-processing, compute shaders, TSL. Integrations: WebGPU, physics engines, spatial audio.

191 days ago

threejs

1
hoadhhoadh

Build 3D web apps with Three.js (WebGL/WebGPU). 556 searchable examples, 60 API classes, 20 use cases. Actions: create 3D scene, load model, add animation, implement physics, build VR/XR. Topics: GLTF loader, PBR materials, particle effects, shadows, post-processing, compute shaders, TSL. Integrations: WebGPU, physics engines, spatial audio.

191 days ago

podcast generation

haniakrim21haniakrim21

Generate AI-powered podcast-style audio narratives using Azure OpenAI's GPT Realtime Mini model via WebSocket. Use this when building text-to-speech features, generating audio narratives, creating podcasts from content, or integrating with the Azure OpenAI Realtime API for real audio output. Covers full-stack implementation from a React frontend to a Python FastAPI backend with WebSocket streaming.

191 days ago

generating-sounds-with-ai

JuanJoseGonGiJuanJoseGonGi

Audit Web Audio API code for sound synthesis best practices. Use when reviewing procedural audio, implementing UI sounds, or checking audio parameter quality. Outputs file:line findings.

192 days ago

citedy-content-ingestion

CitedyCitedy

Turn any URL into structured content — YouTube videos (via Gemini Video API), web articles, PDFs, and audio files. Extract transcripts, summaries, and metadata for use in any LLM pipeline. Powered by Citedy.

content-ingestionyoutubetranscription+4
192 days ago