Audio
v1.0.1
io.clawhub.ivangdavila/audio
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
“Audio Generation” 共 393 个结果
v1.0.1
io.clawhub.ivangdavila/audio
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
v0.3.12
io.clawhub.rainer-liao/pexoai-agent
AI video generation skill with auto model selection across Seedance 2, Kling 3.0, HappyHorse, and 10+ models. Produces finished multi-shot videos (5–120s) fr...
v1.0.17
io.clawhub.cellcog/image-generation-cellcog
AI image generation and photo editing powered by CellCog. Text-to-image, image-to-image, consistent characters, product photography, reference-based generation, style transfer, sets of images, social media visuals, brand assets, stickers, comics, GIFs. Professional image creation with multiple AI models.
v1.0.2
io.clawhub.evolinkai/best-image-generation
Best quality AI image generation (~$0.12-0.20/image). Text-to-image, image-to-image, and image editing via the EvoLink API.
v1.0.3
io.clawhub.ivangdavila/image-generation
Create AI images with GPT Image, Gemini Nano Banana, FLUX, Imagen, and top providers using prompt engineering, style control, and smart editing.
v1.3.16
io.clawhub.dlazyai/dlazy-vidu-audio-clone
Clone voice and generate new text reading audio with one click using Vidu Audio Clone. 使用 Vidu 声音克隆技术,通过参考音频一键复制音色并生成新文本的朗读音频。
v1.0.2
io.clawhub.dpaluy/reve-ai
Generate, edit, and remix images using the Reve AI API. Use when creating images from text prompts, editing existing images with instructions, or combining/remixing multiple reference images. Requires REVE_API_KEY or REVE_AI_API_KEY environment variable.
v1.0.0
io.clawhub.aktheknight/audio-transcribe
Auto-transcribe voice messages locally using faster-whisper with selectable Whisper models, no API key required.
v1.0.18
io.clawhub.cellcog/game-asset-generation-cellcog
AI game asset generation and game development powered by CellCog. Character-consistent art, sprites, tilesets, music, UI, 3D models, GDDs, level design, game prototypes. Cohesive game assets across every modality from a single prompt.
v1.0.11
io.clawhub.mogens9/ai-podcast
Generate AI podcast episodes from PDFs, text, notes, and links using MagicPodcast in OpenClaw. Creates natural two-person dialogue audio, supports custom lan...
v1.0.19
io.clawhub.cellcog/podcast-generation-cellcog
AI podcast generation and production powered by CellCog. Full podcast episodes from a single prompt — multi-voice dialogue with up to 10 distinct speakers, structured episodes with cold opens and segment stingers, music beds ducked under speech, broadcast loudness mastering, finished MP3 plus chapter markers. Episode scripts, show notes, interview prep, audiograms.
v1.0.18
io.clawhub.cellcog/video-generation-cellcog
AI video generation and production powered by CellCog. Marketing videos, product demos, explainers, educational content, lipsync spokesperson videos, UGC, news reports, training materials, cinematic short films, social media reels, YouTube content. Up to 4-minute videos — scripted, voiced, scored, and edited from a single prompt.
v1.2.52
io.clawhub.luruibu/beauty-generation-api
AI portrait image generation with 140+ nationalities, diverse styles, professional headshots, character design, and fashion visualization. Fast generation (<20 seconds), You can use this API service for free or donation
v0.2.2
io.clawhub.guoqiao/mlx-audio-server
Local 24x7 OpenAI-compatible API server for STT/TTS, powered by MLX on your Mac.
v1.1.0
io.clawhub.matrixy/audio-reply-skill
Generate audio replies using TTS. Trigger with "read it to me [public URL]" to fetch and read content aloud, or "talk to me [topic]" to generate a spoken res...
v1.0.2
io.clawhub.evolinkai/cheapest-image-generation
Possibly the cheapest AI image generation (~$0.0036/image). Text-to-image via the EvoLink API.
v1.0.15
io.clawhub.cellcog/music-generation-cellcog
AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros, game soundtracks, ambient soundscapes, jingles, lo-fi beats, orchestral compositions, songs with lyrics. Royalty-free.
v2.2.0
io.clawhub.atyachin/lead-generation
Lead Generation — Find high-intent buyers in live Twitter, Instagram, and Reddit conversations. Auto-researches your product, generates targeted search queries, and discovers people actively looking for solutions you offer. Social selling and prospecting powered by 1.5B+ indexed posts via Xpoz MCP.
v1.0.17
io.clawhub.cellcog/audio-generation-cellcog
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Three voice providers (OpenAI, ElevenLabs, MiniMax). Professional audio production from text prompts.
v1.0.1
io.clawhub.apify/apify-lead-generation
Generates B2B/B2C leads by scraping Google Maps, websites, Instagram, TikTok, Facebook, LinkedIn, YouTube, and Google Search. Use when user asks to find leads, prospects, businesses, build lead lists, enrich contacts, or scrape profiles for sales outreach.
v1.3.8
io.clawhub.dlazyai/dlazy-kling-audio-clone
Generate customized speech that highly restores the timbre by uploading reference audio using Kling Audio Clone. 使用可灵 (Kling) 声音克隆模型,通过上传参考音频,生成高度还原该音色的定制语音。
v1.0.0
io.clawhub.tobisamaa/content-generation
Generate high-quality content across multiple formats. Create articles, reports, social media posts, marketing copy, and other content types with professiona...
v1.0.0
io.clawhub.26medias/runware
Generate images and videos via Runware API. Access to FLUX, Stable Diffusion, Kling AI, and other top models. Supports text-to-image, image-to-image, upscaling, text-to-video, and image-to-video. Use when generating images, creating videos from prompts or images, upscaling images, or doing AI image transformation.
v1.0.0
io.clawhub.ivangdavila/music-generation
Generate AI music with optimized prompts, style control, and production-ready audio output.