QR Code Generator
v1.1.0
io.clawhub.claudiodrusus/qr-gen
Generate QR codes from text, URLs, WiFi credentials, vCards, or any data. Use when the user wants to create a QR code, share a link as a scannable code, gene...
“Audio Generation” 共 263 个结果
v1.1.0
io.clawhub.claudiodrusus/qr-gen
Generate QR codes from text, URLs, WiFi credentials, vCards, or any data. Use when the user wants to create a QR code, share a link as a scannable code, gene...
v1.0.14
io.clawhub.elestirelbilinc-sketch/vap-multimedia-generation
VAP Media API skill for image, video, music, and media editing through VAP. Uses VAP product keys with current Media API endpoints.
v1.0.0
io.clawhub.steipete/markdown-converter
Convert documents and files to Markdown using markitdown. Use when converting PDF, Word (.docx), PowerPoint (.pptx), Excel (.xlsx, .xls), HTML, CSV, JSON, XML, images (with EXIF/OCR), audio (with transcription), ZIP archives, YouTube URLs, or EPubs to Markdown format for LLM processing or text analysis.
vlatest
io.clawhub.emilioacc/atxp
Access ATXP paid API tools for web search, AI image generation, music creation, video generation, X/Twitter search, email, and agent account management. Use...
v2.0.21
io.clawhub.cellcog/cellcog
Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams, and code in one request. Agent-to-agent protocol with multi-step iteration for high accuracy. #1 on DeepResearch Bench (Apr 2026) — deep reasoning meets all modalities, so all your work gets done, not just code.
v1.0.0
io.clawhub.steipete/openai-whisper-api
Transcribe audio via OpenAI Audio Transcriptions API (Whisper).
v1.0.1
io.clawhub.steipete/nano-banana-pro
Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.
v1.0.0
io.clawhub.mahmoudadelbghany/ffmpeg-video-editor
Generate FFmpeg commands from natural language video editing requests - cut, trim, convert, compress, change aspect ratio, extract audio, and more.
v1.0.0
io.clawhub.olliewazza/larry
Automate TikTok slideshow marketing for any app or product. Researches competitors, generates AI images, adds text overlays, posts via Postiz, tracks analyti...
v2.0.15
io.clawhub.nitishgargiitd/cellcog
Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams,...
v2.0.0
io.clawhub.ipedrax/antigravity-image-gen
Generate images using the internal Google Antigravity API (Gemini 3 Pro Image). High quality, native generation without browser automation.
v1.0.0
io.clawhub.steipete/video-transcript-downloader
Download videos, audio, subtitles, and clean paragraph-style transcripts from YouTube and any other yt-dlp supported site. Use when asked to “download this video”, “save this clip”, “rip audio”, “get subtitles”, “get transcript”, or to troubleshoot yt-dlp/ffmpeg and formats/playlists.
v1.0.12
io.clawhub.nitishgargiitd/image-cog
AI image generation and photo editing powered by CellCog. Text-to-image, image-to-image, consistent characters, product photography, reference-based generati...
v1.2.0
io.clawhub.degausai/wonda
Using the Wonda CLI to generate images, videos, music, and audio from the terminal — plus LinkedIn, Reddit, and X/Twitter research and automation
v1.0.0
io.clawhub.steipete/songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
v1.4.0
io.clawhub.shreefentsar/remotion-video-toolkit
Complete toolkit for programmatic video creation with Remotion + React. Covers animations, timing, rendering (CLI/Node.js/Lambda/Cloud Run), captions, 3D, charts, text effects, transitions, and media handling. Use when writing Remotion code, building video generation pipelines, or creating data-driven video templates.
v2.0.0
io.clawhub.i3130002/edge-tts
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch control, and subtitle generation. Use when: (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, driving, cooking). (3) User wants a specific voice, speed, pitch, or format for TTS output.
v1.0.0
io.clawhub.pors/openai-tts
Text-to-speech via OpenAI Audio Speech API.
v1.2.1
io.clawhub.itsfabioroma/transcribee
Transcribe YouTube videos and local audio/video files with speaker diarization. Use when user asks to transcribe a YouTube URL, podcast, video, or audio file. Outputs clean speaker-labeled transcripts ready for LLM analysis.
v1.0.13
io.clawhub.nitishgargiitd/video-cog
AI video generation and production powered by CellCog. Marketing videos, product demos, explainers, educational content, lipsync spokesperson videos, UGC, ne...
v0.1.0
io.clawhub.edkief/kokoro-tts
Generate spoken audio from text using the local Kokoro TTS engine. Use when the user asks to "say" something, requests a voice message, or wants text converted to speech.
v1.3.0
io.clawhub.buddyh/veo
Generate video using Google Veo (Veo 3.1 / Veo 3.0).
v1.3.15
io.clawhub.dlazyai/dlazy-suno-music
Suno music generation model. Supports inspiration mode (auto lyrics) and custom mode (manual lyrics), generating music with or without vocals. Suno 音乐生成模型。支持灵感模式(自动作词)和自定义模式(手动填词),可生成包含人声或纯器乐的音乐。
v2.1.1
io.clawhub.jonisjongithub/venice-ai
Complete Venice AI platform — text generation, vision/image analysis, web search, X/Twitter search, embeddings, TTS, speech-to-text, image generation, backgr...