voice-call

246.8k
openclawopenclaw

Start voice calls via the OpenClaw voice-call plugin.

191 days ago

write-release-notes

45.6k
tldrawtldraw

Writing release notes articles for tldraw SDK releases. Use when creating new release documentation, drafting release notes from scratch, or reviewing release note quality. Provides guidance on structure, voice, and content for release files in `apps/docs/content/releases/`.

191 days ago

screen-reader-testing

29.9k
wshobsonwshobson

Test web applications with screen readers including VoiceOver, NVDA, and JAWS. Use when validating screen reader compatibility, debugging accessibility issues, or ensuring assistive technology support.

191 days ago

voice-ai-development

21.8k
davila7davila7

Expert in building voice AI applications - from real-time voice agents to voice-enabled apps. Covers OpenAI Realtime API, Vapi for voice agents, Deepgram for transcription, ElevenLabs for synthesis, LiveKit for real-time infrastructure, and WebRTC fundamentals. Knows how to build low-latency, production-ready voice experiences. Use when: voice ai, voice agent, speech to text, text to speech, realtime voice.

191 days ago

twilio-communications

21.8k
davila7davila7

Build communication features with Twilio: SMS messaging, voice calls, WhatsApp Business API, and user verification (2FA). Covers the full spectrum from simple notifications to complex IVR systems and multi-channel authentication. Critical focus on compliance, rate limits, and error handling. Use when: twilio, send SMS, text message, voice call, phone verification.

191 days ago

content-creator

21.8k
davila7davila7

Create SEO-optimized marketing content with consistent brand voice. Includes brand voice analyzer, SEO optimizer, content frameworks, and social media templates. Use when writing blog posts, creating social media content, analyzing brand voice, optimizing SEO, planning content calendars, or when user mentions content creation, brand voice, SEO optimization, social media marketing, or content strategy.

191 days ago

voice-agents

21.8k
davila7davila7

Voice agents represent the frontier of AI interaction - humans speaking naturally with AI systems. The challenge isn't just speech recognition and synthesis, it's achieving natural conversation flow with sub-800ms latency while handling interruptions, background noise, and emotional nuance. This skill covers two architectures: speech-to-speech (OpenAI Realtime API, lowest latency, most natural) and pipeline (STT→LLM→TTS, more control, easier to debug). Key insight: latency is the constraint. Hu

191 days ago

screen-reader-testing

18.0k
sickn33sickn33

Test web applications with screen readers including VoiceOver, NVDA, and JAWS. Use when validating screen reader compatibility, debugging accessibility issues, or ensuring assistive technology supp...

191 days ago

voice-ai-engine-development

18.0k
sickn33sickn33

Build real-time conversational AI voice engines using async worker pipelines, streaming transcription, LLM agents, and TTS synthesis with interrupt handling and multi-provider support

191 days ago

digital-brain

13.0k
muratcankoylanmuratcankoylan

This skill should be used when the user asks to "write a post", "check my voice", "look up contact", "prepare for meeting", "weekly review", "track goals", or mentions personal brand, content creation, network management, or voice consistency.

191 days ago

book-sft-pipeline

13.0k
muratcankoylanmuratcankoylan

This skill should be used when the user asks to "fine-tune on books", "create SFT dataset", "train style model", "extract ePub text", or mentions style transfer, LoRA training, book segmentation, or author voice replication.

191 days ago

docs-voice

11.7k
reactjsreactjs

Use when writing any React documentation. Provides voice, tone, and style rules for all doc types.

191 days ago

speech

10.4k
openaiopenai

Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.

191 days ago

brand-voice

8.4k
anthropicsanthropics

Apply and enforce brand voice, style guide, and messaging pillars across content. Use when reviewing content for brand consistency, documenting a brand voice, adapting tone for different audiences, or checking terminology and style guide compliance.

191 days ago

content-creator

2.2k
alirezarezvanialirezarezvani

Create SEO-optimized marketing content with a consistent brand voice. Includes brand voice analyzer, SEO optimizer, content frameworks, and social media templates. Use when writing blog posts, creating social media content, analyzing brand voice, optimizing SEO, planning content calendars, or when user mentions content creation, brand voice, SEO optimization, social media marketing, or content strategy.

191 days ago

github-issue-creator

1.8k
openclawopenclaw

Convert raw notes, error logs, voice dictation, or screenshots into crisp GitHub-flavored markdown issue reports. Use when the user pastes bug info, error messages, or informal descriptions and wants a structured GitHub issue. Supports images/GIFs for visual evidence.

190 days ago

voice-note-to-midi

1.8k
openclawopenclaw

Convert voice notes, humming, and melodic audio recordings to quantized MIDI files using ML-based pitch detection and intelligent post-processing

audiomidimusic+2
191 days ago

voicemonkey

1.8k
openclawopenclaw

Control Alexa devices via VoiceMonkey API v2 - make announcements, trigger routines, start flows, and display media.

191 days ago

ringbot

1.8k
openclawopenclaw

Make outbound AI phone calls. Use when asked to call a business, make a phone call, order food by phone, schedule appointments, or any task requiring voice calls. Triggers on "call", "phone", "dial", "ring", "order pizza", "make reservation", "schedule appointment".

191 days ago

tts-whatsapp

1.8k
openclawopenclaw

Send high-quality text-to-speech voice messages on WhatsApp in 40+ languages with automatic delivery

whatsappttsvoice+3
191 days ago

audio-reply

1.8k
openclawopenclaw

Generate audio replies using TTS. Trigger with "read it to me [public URL]" to fetch and read content aloud, or "talk to me [topic]" to generate a spoken response. Also responds to "speak", "say it", "voice reply".

191 days ago

pamela-calls

1.8k
openclawopenclaw

Make AI-powered phone calls with Pamela's voice API. Create outbound calls, register custom tools for mid-call actions, handle webhooks, and build React UIs. Use when the user wants to make phone calls, integrate voice AI, build IVR systems, navigate phone menus, or automate phone tasks.

191 days ago

clawspaces

1.8k
openclawopenclaw

X Spaces, but for AI Agents. Live voice rooms where AI agents host conversations.

191 days ago

phone-agent

1.8k
openclawopenclaw

Run a real-time AI phone agent using Twilio, Deepgram, and ElevenLabs. Handles incoming calls, transcribes audio, generates responses via LLM, and speaks back via streaming TTS. Use when user wants to: (1) Test voice AI capabilities, (2) Handle phone calls programmatically, (3) Build a conversational voice bot.

191 days ago

critical-article-writer

1.8k
openclawopenclaw

Generate draft articles, outlines, and editorial content matching a distinctive analytical, skeptical voice with sharp critical commentary, conversational tone, and strategic humor.

191 days ago

linkedin-monitor

1.8k
openclawopenclaw

Bulletproof LinkedIn inbox monitoring with progressive autonomy. Monitors messages hourly, drafts replies in your voice, and alerts you to new conversations. Supports 4 autonomy levels from monitor-only to full autonomous.

191 days ago

duby

1.8k
openclawopenclaw

Convert text to speech using Duby.so API. Supports various voices and emotions.

ttsaudiovoice+1
191 days ago

twilio

1.8k
openclawopenclaw

Send SMS, make voice calls, and manage WhatsApp messages via Twilio API. Use for notifications, 2FA, customer communications, and voice automation.

191 days ago

transcribe

1.8k
openclawopenclaw

Transcribe audio files to text using local Whisper (Docker). Use when receiving voice messages, audio files (.mp3, .m4a, .ogg, .wav, .webm), or when asked to transcribe audio content.

191 days ago

elevenlabs-voices

1.8k
openclawopenclaw

High-quality voice synthesis with 18 personas, 32 languages, sound effects, batch processing, and voice design using ElevenLabs API.

ttsvoicespeech+5
191 days ago

voice-agent

1.8k
openclawopenclaw

Local Voice Input/Output for Agents using the AI Voice Agent API.

191 days ago

news-summary

1.8k
openclawopenclaw

This skill should be used when the user asks for news updates, daily briefings, or what's happening in the world. Fetches news from trusted international RSS feeds and can create voice summaries.

191 days ago

tts

1.8k
openclawopenclaw

Convert text to speech using Hume AI (or OpenAI) API. Use when the user asks for an audio message, a voice reply, or to hear something "of vive voix".

191 days ago

callmac

1.8k
openclawopenclaw

Remote voice control for a Mac from mobile devices using commands like /callmac or /voice. Broadcast announcements, play alarms, tell stories, wake up kids — all triggered by Telegram or WhatsApp messages. Uses edge-tts for robust mixed Chinese/English TTS, plays audio locally on the Mac, and supports looping and volume control for flexible playback.

191 days ago

fliz-ai-video-generator

1.8k
openclawopenclaw

Complete integration guide for the Fliz REST API - an AI-powered video generation platform that transforms text content into professional videos with voiceovers, AI-generated images, and subtitles. Use this skill when: - Creating integrations with Fliz API (WordPress, Zapier, Make, n8n, custom apps) - Building video generation workflows via API - Implementing webhook handlers for video completion notifications - Developing automation tools that create, manage, or translate videos - Troubleshooting Fliz API errors or authentication issues - Understanding video processing steps and status polling Key capabilities: video creation from text/Brief, video status monitoring, translation, duplication, voice/music listing, webhook notifications.

videoaifliz+3
191 days ago

doubao-open-tts

1.8k
openclawopenclaw

Text-to-Speech service using Doubao (Volcano Engine) API with 200+ voices, interactive voice selection, and multilingual support

191 days ago

elevenlabs-agents

1.8k
openclawopenclaw

Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.

191 days ago

telegram-offline-voice

1.8k
openclawopenclaw

本地生成 Telegram 语音消息,支持自动清洗、分段与临时文件管理。

191 days ago

youtube-voice-summarizer

1.8k
openclawopenclaw

Transform YouTube videos into podcast-style voice summaries using ElevenLabs TTS

191 days ago

parakeet-stt

1.8k
openclawopenclaw

Local speech-to-text with NVIDIA Parakeet TDT 0.6B v3 (ONNX on CPU). 30x faster than Whisper, 25 languages, auto-detection, OpenAI-compatible API. Use when transcribing audio files, converting speech to text, or processing voice recordings locally without cloud APIs.

191 days ago

x-voice-match

1.8k
openclawopenclaw

Analyze a Twitter/X account's posting style and generate authentic posts that match their voice. Use when the user wants to create X posts that sound like them, analyze their posting patterns, or maintain consistent voice across posts. Works with Bird CLI integration.

191 days ago

whatsapp-voice-talk

1.8k
openclawopenclaw

Real-time WhatsApp voice message processing. Transcribe voice notes to text via Whisper, detect intent, execute handlers, and send responses. Use when building conversational voice interfaces for WhatsApp. Supports English and Hindi, customizable intents (weather, status, commands), automatic language detection, and streaming responses via TTS.

191 days ago

voice-transcribe

1.8k
openclawopenclaw

Transcribe audio files using OpenAI's gpt-4o-mini-transcribe model with vocabulary hints and text replacements. Requires uv (https://docs.astral.sh/uv/).

191 days ago

voice-wake-say

1.8k
openclawopenclaw

Speak responses aloud on macOS using the built-in `say` command when user input indicates Voice Wake/voice recognition (for example, messages starting with "User talked via voice recognition on <device>").

191 days ago

azure-ai-voicelive-py

1.8k
openclawopenclaw

Build real-time voice AI applications using Azure AI Voice Live SDK (azure-ai-voicelive). Use this skill when creating Python applications that need real-time bidirectional audio communication with Azure AI, including voice assistants, voice-enabled chatbots, real-time speech-to-speech translation, voice-driven avatars, or any WebSocket-based audio streaming with AI models. Supports Server VAD (Voice Activity Detection), turn-based conversation, function calling, MCP tools, avatar integration, and transcription.

191 days ago

elevenlabs

1.8k
openclawopenclaw

Text-to-speech, sound effects, music generation, voice management, and quota checks via the ElevenLabs API. Use when generating audio with ElevenLabs or managing voices.

191 days ago

blog-writer

1.8k
openclawopenclaw

This skill should be used when writing blog posts, articles, or long-form content in the writer's distinctive writing style. It produces authentic, opinionated content that matches the writer's voice—direct, conversational, and grounded in personal experience. The skill handles the complete workflow from research review through Notion publication. Use this skill for drafting blog posts, thought leadership pieces, or any writing meant to reflect the writer's perspective on AI, productivity, sales, marketing, or technology topics.

191 days ago

ai-video-gen

1.8k
openclawopenclaw

End-to-end AI video generation - create videos from text prompts using image generation, video synthesis, voice-over, and editing. Supports OpenAI DALL-E, Replicate models, LumaAI, Runway, and FFmpeg editing.

191 days ago