Back to Skills

p-video

Generate videos with Pruna P-Video and WAN models via inference.sh CLI. Models: P-Video, WAN-T2V, WAN-I2V. Capabilities: text-to-video, image-to-video, audio support, 720p/1080p, fast inference. Pruna optimizes models for speed without quality loss. Triggers: pruna video, p-video, pruna ai video, fast video generation, optimized video, wan t2v, wan i2v, economic video generation, cheap video generation, pruna text to video, pruna image to video

7starsUpdated 8/14/2026

Security Assessment

Safe(95/100)
Security Score95/100

About p-video

p-video is a skill for generating videos with Pruna's optimized P-Video and WAN models through the inference.sh `belt` command-line tool. Pruna optimizes video models for speed and lower cost while preserving quality, and this skill packages that catalogue so an agent can turn a text prompt or a still image into a finished video clip with one shell command, avoiding local GPU setup or SDK integration.

It documents three models: P-Video for text-to-video and image-to-video with optional synchronized audio input, WAN-T2V for economical text-to-video at 480p/720p, and WAN-I2V for animating still images. Each runs as `belt app run pruna/<model> --input '{...}'` with JSON parameters for prompt, duration, resolution (720p/1080p depending on model), an input image or audio URL, and a draft mode for faster, cheaper previews. The skill includes a pricing table (for example WAN-T2V at roughly $0.05 for 480p and $0.10 for 720p per video) and links to related inference.sh skills for image generation, image-to-video, and text-to-speech narration.

Typical users are content creators, marketers, and developers building media pipelines who want fast, low-cost AI video — animated stills, short social clips, concept tests, and cinematic landscape shots. Like its sibling image skill, it is a documented front-end to the hosted inference.sh service rather than a standalone video engine, so its capabilities are bounded by what that platform offers.

FAQ

What are the prerequisites?

Install the inference.sh CLI (`belt`) and run `belt login`. The skill links to install instructions and the belt CLI skill.

Which video modes are supported?

Text-to-video, image-to-video, and (with P-Video) audio-synced generation. WAN-T2V handles text-to-video and WAN-I2V animates still images.

What resolutions and durations are available?

P-Video supports 720p and 1080p; WAN models support 480p and 720p. Duration is set with a `duration` parameter (examples use around 5 seconds).

Is there a cheaper preview option?

Yes. P-Video accepts `"draft": true` for faster, cheaper draft generation, useful for quick concept tests.

Does generation run on my machine?

No. All rendering runs on the hosted inference.sh platform; the skill only issues `belt` commands.

Install p-video

Download and extract the skill files to your .claude/skills/ directory.

Quick Setup:

  1. Copy the skill folder to .claude/skills/
  2. Claude will automatically detect and use the skill