Audio
v1.0.1
io.clawhub.ivangdavila/audio
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
“Audio” 共 302 个结果
v1.0.1
io.clawhub.ivangdavila/audio
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
v1.0.0
io.clawhub.aktheknight/audio-transcribe
Auto-transcribe voice messages locally using faster-whisper with selectable Whisper models, no API key required.
v0.2.2
io.clawhub.guoqiao/mlx-audio-server
Local 24x7 OpenAI-compatible API server for STT/TTS, powered by MLX on your Mac.
v1.1.0
io.clawhub.matrixy/audio-reply-skill
Generate audio replies using TTS. Trigger with "read it to me [public URL]" to fetch and read content aloud, or "talk to me [topic]" to generate a spoken res...
v1.3.14
io.clawhub.dlazyai/dlazy-vidu-audio-clone
Clone voice and generate new text reading audio with one click using Vidu Audio Clone. 使用 Vidu 声音克隆技术,通过参考音频一键复制音色并生成新文本的朗读音频。
v1.0.17
io.clawhub.cellcog/audio-generation-cellcog
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Three voice providers (OpenAI, ElevenLabs, MiniMax). Professional audio production from text prompts.
v1.3.8
io.clawhub.dlazyai/dlazy-kling-audio-clone
Generate customized speech that highly restores the timbre by uploading reference audio using Kling Audio Clone. 使用可灵 (Kling) 声音克隆模型,通过上传参考音频,生成高度还原该音色的定制语音。
v1.0.0
io.clawhub.obviyus/openrouter-transcribe
Transcribe audio files via OpenRouter using audio-capable models (Gemini, GPT-4o-audio, etc).
v1.3.16
io.clawhub.dlazyai/dlazy-audio-generate
Audio generation skill. Automatically selects the best dlazy CLI audio/TTS model based on the prompt. 音频生成技能。根据提示词自动选择最佳的 dlazy CLI 音频/TTS 模型。
v1.2.3
io.clawhub.rakesh1002/audiopod
Use AudioPod AI's API for audio processing tasks including AI music generation (text-to-music, text-to-rap, instrumentals, samples, vocals), stem separation, text-to-speech, noise reduction, speech-to-text transcription, speaker separation, and media extraction. Use when the user needs to generate music/songs/rap from text, split a song into stems/vocals/instruments, generate speech from text, clean up noisy audio, transcribe audio/video, or extract audio from YouTube/URLs. Requires AUDIOPOD_API_KEY env var or pass api_key directly.
v1.2.0
io.clawhub.brokemac79/webchat-audio-notifications
Add browser audio notifications to Moltbot/Clawdbot webchat with 5 intensity levels - from whisper to impossible-to-miss (only when tab is backgrounded).
v1.0.0
io.clawhub.udiedrichsen/audio-gen
Generate audiobooks, podcasts, or educational audio content on demand. User provides an idea or topic, Claude AI writes a script, and ElevenLabs converts it to high-quality audio. Supports multiple formats (audiobook, podcast, educational), custom lengths, and voice effects. Use when asked to create audio content, make a podcast, generate an audiobook, or produce educational audio. Returns MP3 audio file via MEDIA token.
v1.0.12
io.clawhub.nitishgargiitd/audio-cog
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Thre...
v1.0.0
ru.transkriba/transcription
Transcribe Russian audio and video with timestamps and optional speaker labels from files or links.
v1.0.1
ru.vibe2text/vibe2text
Audio & video to text, Russian-first: diarization, timestamps, summary, action items, subtitles.
v0.4.0
io.github.ni-c/audiobookshelf-mcp
Browse your Audiobookshelf libraries and keep listening progress, bookmarks and playlists in sync
v6.1.0
com.audioeye/testing-sdk-mcp
Scans live pages with the AudioEye rules engine and maps accessibility issues to JSX source.
v0.1.0
io.github.yasib48/audiomade
Generate four real game-audio candidates, audition and select one, then install it into a project.
v0.5.0
pro.magicmaster/mastering
Audio mastering for AI agents: LUFS/True Peak targets, Suno/Udio AI-fingerprint removal.
v0.2.0-beta.1
io.github.daredoole/audio-calibration-mcp
Evidence-bound REW measurement, analysis, conservative EQ, DSP, and listening-test tools.
v1.2.0
io.github.ripunjay-kashyap/audio-sonic-mcp
Local-first audio analysis: BPM, musical key, production profile, and CLAP vibe embeddings.
v0.1.2
io.github.nirholas/audio-mcp
Text-to-speech, speech-to-text, audio-to-face lipsync, and motion-capture clips for 3D agents.
v1.0.10
io.github.CSOAI-ORG/voice-audio-mcp
Voice Audio MCP server. Tools: text to speech, list voices, transcribe. Built by MEOK AI Labs.
v2.0.0
io.github.Evozim/audio-watermarking-detector-mcp
Premium agentic endpoint for audio-watermarking-detector-mcp.