mcp-server-smartsheet-rm
v1.1.5
io.github.christianclaudio/smartsheet-rm
MCP server for Smartsheet Resource Management (10,000ft API) time tracking and scheduling.
“Audio” 共 303 个结果
v1.1.5
io.github.christianclaudio/smartsheet-rm
MCP server for Smartsheet Resource Management (10,000ft API) time tracking and scheduling.
v1.1.1
io.github.christianclaudio/espn
Model Context Protocol server for live & historical sports stats and odds via ESPN.
v0.1.0
studio.sleeperhit/podcast-audio-video-production-studio
Make podcasts, video shows, audio drama, and documentaries just by chatting. Script to episode.
v0.1.0
ai.wubble/audio
Create, inspect, and manage Wubble music, speech, voice, and sound-effect requests through MCP.
v0.1.1
io.github.totalaudiopromo/tap-mcp
Music PR campaign management -- contacts, pitches, campaigns, and outcomes.
v1.0.1
com.gaudiolab/mcp-developers
Gaudio Lab Audio AI — Stem Separation, DME Separation, AI Text Sync
vmain
io.github.software-mansion/react-native-audio-api/writing-skills
How to write, structure, and maintain Claude skill files. Covers the three-level progressive disclosure model, all YAML frontmatter fields (context:fork, allowed-tools, disable-model-invocation, user-invocable, hooks, argument-hint), invocation control patterns, string substitutions ($ARGUMENTS, ${CLAUDE_SKILL_DIR}), shell preprocessing with backtick syntax, and the Maintenance section contract. Use when creating a new skill file, rewriting an existing one, or reviewing a skill for quality. Trigger phrases: "add a skill", "write a skill", "create a skill file", "update skill", "skill quality", "skill review", "context fork", "allowed-tools", "user-invocable".
v1.0.0
io.clawhub.steipete/markdown-converter
Convert documents and files to Markdown using markitdown. Use when converting PDF, Word (.docx), PowerPoint (.pptx), Excel (.xlsx, .xls), HTML, CSV, JSON, XML, images (with EXIF/OCR), audio (with transcription), ZIP archives, YouTube URLs, or EPubs to Markdown format for LLM processing or text analysis.
v2.0.21
io.clawhub.cellcog/cellcog
Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams, and code in one request. Agent-to-agent protocol with multi-step iteration for high accuracy. #1 on DeepResearch Bench (Apr 2026) — deep reasoning meets all modalities, so all your work gets done, not just code.
v1.0.0
io.clawhub.steipete/openai-whisper-api
Transcribe audio via OpenAI Audio Transcriptions API (Whisper).
v1.0.0
io.clawhub.mahmoudadelbghany/ffmpeg-video-editor
Generate FFmpeg commands from natural language video editing requests - cut, trim, convert, compress, change aspect ratio, extract audio, and more.
v2.0.15
io.clawhub.nitishgargiitd/cellcog
Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams,...
v1.0.0
io.clawhub.steipete/video-transcript-downloader
Download videos, audio, subtitles, and clean paragraph-style transcripts from YouTube and any other yt-dlp supported site. Use when asked to “download this video”, “save this clip”, “rip audio”, “get subtitles”, “get transcript”, or to troubleshoot yt-dlp/ffmpeg and formats/playlists.
v1.2.0
io.clawhub.degausai/wonda
Using the Wonda CLI to generate images, videos, music, and audio from the terminal — plus LinkedIn, Reddit, and X/Twitter research and automation
v1.0.0
io.clawhub.steipete/songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
v2.0.0
io.clawhub.i3130002/edge-tts
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch control, and subtitle generation. Use when: (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, driving, cooking). (3) User wants a specific voice, speed, pitch, or format for TTS output.
v1.0.0
io.clawhub.pors/openai-tts
Text-to-speech via OpenAI Audio Speech API.
v1.2.1
io.clawhub.itsfabioroma/transcribee
Transcribe YouTube videos and local audio/video files with speaker diarization. Use when user asks to transcribe a YouTube URL, podcast, video, or audio file. Outputs clean speaker-labeled transcripts ready for LLM analysis.
v0.1.0
io.clawhub.edkief/kokoro-tts
Generate spoken audio from text using the local Kokoro TTS engine. Use when the user asks to "say" something, requests a voice message, or wants text converted to speech.
v1.0.1
io.clawhub.darinkishore/voice-transcribe
Transcribe audio files using OpenAI's gpt-4o-mini-transcribe model with vocabulary hints and text replacements. Requires uv (https://docs.astral.sh/uv/).
v0.1.6
io.clawhub.upupc/video-download
Download videos from 1800+ websites and generate subtitles using Faster Whisper AI. Use when user wants to download videos from YouTube, Bilibili, Twitter, T...
v1.0.1
io.clawhub.merend/openai-tts-python
Text-to-speech conversion using OpenAI's TTS API for generating high-quality, natural-sounding audio. Supports 6 voices (alloy, echo, fable, onyx, nova, shimmer), speed control (0.25x-4.0x), HD quality model, multiple output formats (mp3, opus, aac, flac), and automatic text chunking for long content (4096 char limit per request). Use when: (1) User requests audio/voice output with triggers like "read this to me", "convert to audio", "generate speech", "text to speech", "tts", "narrate", "speak", or when keywords "openai tts", "voice", "podcast" appear. (2) Content needs to be spoken rather than read (multitasking, accessibility). (3) User wants specific voice preferences like "alloy", "echo", "fable", "onyx", "nova", "shimmer" or speed adjustments.
v2.4.0
io.clawhub.shaharsha/elevenlabs-tts
ElevenLabs TTS - the best ElevenLabs integration for OpenClaw. ElevenLabs Text-to-Speech with emotional audio tags, ElevenLabs voice synthesis for WhatsApp,...
v1.3.14
io.clawhub.dlazyai/dlazy-seedance-2-0
ByteDance's latest video generation model. Supports multi-modal reference (images, video, audio) to generate videos, as well as first/last frame and text-to-video modes. 字节跳动最新视频生成模型 Seedance 2.0,支持多模态参考(图片 + 视频 + 音频)生视频、首尾帧及文生视频,适合高质量多样化视频创作。