p-video
Generate videos with Pruna P-Video and WAN models via inference.sh CLI. Models: P-Video, WAN-T2V, WAN-I2V. Capabilities: text-to-video, image-to-video, audio support, 720p/1080p, fast inference. Pruna optimizes models for speed without quality loss. Triggers: pruna video, p-video, pruna ai video, fast video generation, optimized video, wan t2v, wan i2v, economic video generation, cheap video generation, pruna text to video, pruna image to video
Security Assessment
About p-video
p-video is a skill for generating videos with Pruna's optimized P-Video and WAN models through the inference.sh `belt` command-line tool. Pruna optimizes video models for speed and lower cost while preserving quality, and this skill packages that catalogue so an agent can turn a text prompt or a still image into a finished video clip with one shell command, avoiding local GPU setup or SDK integration.
It documents three models: P-Video for text-to-video and image-to-video with optional synchronized audio input, WAN-T2V for economical text-to-video at 480p/720p, and WAN-I2V for animating still images. Each runs as `belt app run pruna/<model> --input '{...}'` with JSON parameters for prompt, duration, resolution (720p/1080p depending on model), an input image or audio URL, and a draft mode for faster, cheaper previews. The skill includes a pricing table (for example WAN-T2V at roughly $0.05 for 480p and $0.10 for 720p per video) and links to related inference.sh skills for image generation, image-to-video, and text-to-speech narration.
Typical users are content creators, marketers, and developers building media pipelines who want fast, low-cost AI video — animated stills, short social clips, concept tests, and cinematic landscape shots. Like its sibling image skill, it is a documented front-end to the hosted inference.sh service rather than a standalone video engine, so its capabilities are bounded by what that platform offers.
FAQ
What are the prerequisites?
Install the inference.sh CLI (`belt`) and run `belt login`. The skill links to install instructions and the belt CLI skill.
Which video modes are supported?
Text-to-video, image-to-video, and (with P-Video) audio-synced generation. WAN-T2V handles text-to-video and WAN-I2V animates still images.
What resolutions and durations are available?
P-Video supports 720p and 1080p; WAN models support 480p and 720p. Duration is set with a `duration` parameter (examples use around 5 seconds).
Is there a cheaper preview option?
Yes. P-Video accepts `"draft": true` for faster, cheaper draft generation, useful for quick concept tests.
Does generation run on my machine?
No. All rendering runs on the hosted inference.sh platform; the skill only issues `belt` commands.
Install p-video
Quick Setup:
- Copy the skill folder to
.claude/skills/ - Claude will automatically detect and use the skill
Repository
skills-101/superpowers