pexoai/pexo-skills▌
6 approved skills in this repository
seedance-2.0-prompter
Productivity
This skill transforms a user's scattered multimodal assets (images, videos, audio) and ambiguous creative intent into a structured, executable prompt for the Seedance 2.0 video generation model. It acts as an expert prompt engineer, ensuring the highest quality output from the underlying model.
seedance-prompter
Productivity
This skill transforms a user's scattered multimodal assets (images, videos, audio) and ambiguous creative intent into a structured, executable prompt for the Seedance 2.0 video generation model. It acts as an expert prompt engineer, ensuring the highest quality output from the underlying model.
pexo-agent
Productivity
Conversational AI video creation agent that plans, generates, and delivers finished videos from natural language descriptions. \n \n Supports short-form video output (5–60 seconds) in three aspect ratios: 16:9, 9:16, and 1:1, suitable for YouTube, TikTok, Instagram, and other platforms \n Accepts reference materials including product photos, brand assets, style examples, and audio files to guide creative direction and visual consistency \n Engages in multi-turn dialogue, asking clarifying questi
videoagent-audio-studio
Video
Unified audio generation dispatcher routing TTS, music, sound effects, and voice cloning to optimal models. \n \n Routes requests to ElevenLabs (TTS, voice cloning, SFX) or fal.ai (music) based on request type, with latencies ranging from <1s to ~15s \n Supports five audio capabilities: multilingual text-to-speech with voice selection, low-latency turbo TTS, background music composition, sound effect generation (up to 22 seconds), and voice cloning from audio samples \n Requires only ELEVEN
videoagent-image-studio
Video
Unified access to 8 AI image generation models with automatic model selection and zero API key setup. \n \n Supports Midjourney, Flux (Pro/Dev/Schnell), Ideogram, Recraft, SDXL, and Nano Banana with automatic model routing based on user intent \n Handles Midjourney's async polling transparently; all models return consistent output format with image URLs \n Includes Midjourney actions (upscale, variation, reroll) and reference image support for style consistency \n All requests routed through hos
videoagent-video-studio
Video
Generate short AI videos from text or images using 7 backend models with zero API key setup. \n \n Supports three generation modes: text-to-video, image-to-video, and reference-based generation for consistent output \n Seven models available (minimax, kling, veo, hunyuan, grok, seedance, pixverse) with automatic selection or manual override via --model flag \n Configurable duration (4–12 seconds), aspect ratios (16:9, 9:16, 1:1, 4:3, 3:4), and automatic prompt enhancement for better results \n S