ai-video/skills/ai-video-prompt-enhancer/SKILL.md
Use when the user wants to generate a single AI video clip and needs to turn a rough idea into a detailed, cinematic prompt that modern AI video models can execute well. For multi-shot videos (TikTok Reels, Ads, Explainers), use ai-video-storyboard instead.
npx skillsauth add bernatmv/ai-rules ai-video-prompt-enhancerInstall this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.
3 of 9 scanners reported clean
Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.
Most people write AI video prompts the same way they'd describe a scene to a friend: "a horse running in a field." Modern AI video models can technically render that, but the output looks generic because the prompt doesn't specify any of the things cinematographers actually care about — composition, camera movement, lighting, lens character, color grading, subject detail, or motion quality.
This skill turns a rough idea into a production-grade cinematic prompt that leverages the full vocabulary of filmmaking. Instead of "horse running in a field," you get something like:
A majestic brown stallion galloping across a golden wheat field at sunset, cinematic wide tracking shot moving parallel to the horse, warm golden-hour backlight with visible lens flare, shallow depth of field at f/2.8, muted earth-tone color grade, 35mm anamorphic look with subtle film grain, slow motion 60fps feel, 8 seconds, 16:9, cinematic 1080p
The enhanced prompt specifies the visual language modern AI video models need to produce professional output. The difference in quality is usually the difference between "interesting AI clip" and "usable B-roll."
Do NOT use this skill for:
Ask the user one combined question (not a multi-turn interrogation):
What scene do you want to generate? Tell me: (1) the subject, (2) what it's doing, (3) the mood or vibe, (4) the platform/aspect ratio (TikTok 9:16, YouTube 16:9, feed 1:1), and (5) the duration you're targeting (4s / 5s / 8s / 10s).
Accept whatever the user gives you, even partial answers. Fill in sensible defaults for missing fields:
| Missing field | Default | |---|---| | Duration | 5 seconds | | Aspect ratio | 9:16 vertical | | Mood | "cinematic" |
Based on the subject + mood, decide on:
Shot type. Extreme close-up (ECU) for intimacy/detail, close-up (CU) for emotion, medium shot (MS) for action, medium wide (MWS) for subject-in-environment, wide shot (WS) for landscape/scale.
Camera movement. Locked-off for poise, slow dolly-in for building tension, tracking for parallel motion, crane-up for reveal, handheld for raw energy.
Lighting. Golden hour for warmth, overcast for mood, neon for urban/nightlife, rim light for drama, motivated window light for natural interiors, practical light sources for grounded realism.
Lens character. Shallow DOF for focus pulls and bokeh, deep focus for landscape/documentary, wide-angle for scale and mild distortion, telephoto for compressed backgrounds.
Color grade. Warm gold for nostalgia, cool teal for tech/modern, muted earth tones for documentary, high contrast for drama, desaturated for bleak mood.
Film look. Clean digital for contemporary, 16mm grain for analog warmth, 35mm anamorphic for cinematic scale, VHS/tape for retro.
Motion quality. Slow motion (60fps feel) for impact, normal speed for documentary, time-lapse for transitions.
Compose the prompt as a single paragraph, 40–80 words, following this structure:
[Shot type] [camera movement] of [detailed subject] [action] [environment],
[lighting description], [lens character and aperture], [color palette],
[film look], [technical style notes], [duration]s, [aspect ratio], cinematic 1080p
Concrete over abstract. Replace every generic noun with specific detail:
| Abstract | Concrete | |---|---| | "a woman" | "a woman in her late twenties with wavy auburn hair, wearing a denim jacket" | | "a car" | "a vintage 1970 Ford Mustang in deep navy blue with chrome trim" | | "a forest" | "an old-growth redwood forest with dense ferns and shafts of morning sunlight" | | "a city" | "a rain-slicked Tokyo street in Shibuya, neon signs reflecting on the pavement" | | "fast motion" | "slow-motion 60fps feel with motion blur on the limbs" |
Always end with: [duration]s, [aspect ratio], cinematic 1080p
Present the user with:
This teaches the user the vocabulary over time, and builds trust in the output. Without the explanation, enhanced prompts feel like black-box magic. With the explanation, users start writing better prompts themselves.
If the user wants alternatives, generate 2–3 variants that change one axis at a time:
This is how cinematographers explore coverage on a real set: same scene, multiple angles.
User input:
"I want a video of a cat jumping off a kitchen counter"
Enhanced prompt:
Slow-motion medium close-up of a sleek black-and-white tuxedo cat mid-jump from a white marble kitchen counter toward the floor, captured at 120fps slow-motion with subtle motion blur on the fur and paws, warm morning light streaming through a window on the left casting long shadows, shallow depth of field on the cat with blurred kitchen background, clean contemporary digital look with muted warm tones, 5 seconds, 9:16 vertical, cinematic 1080p
Why these choices: Slow-motion was the obvious pick for a physical action shot — it turns a half-second real event into five seconds of visual content, and motion blur signals cinematic intent. Warm morning light grounds the scene in a specific time of day (which helps the model commit to consistent lighting), and the medium close-up shot keeps the cat as the subject while still showing enough environment to read as a kitchen.
MIT — use freely, commercial or personal.
development
Keyword research and validation with real search-demand data — never ship keywords from intuition alone. Probes Google Autocomplete per language (free, no account) to prove demand and discover the exact phrasing people type, checks SERPs for winnability, and uses Keyword Planner/Ahrefs/Semrush exports when the user has access. Activates when: choosing or reviewing SEO keywords, meta keywords, page titles, article topics or slugs, landing page copy targeting search, App Store/ASO keyword fields, multilingual keyword sets, 'what should we rank for', 'keyword analysis', or auditing why a page doesn't rank. Also invoke it as a validation pass whenever another skill or task produces a keyword list.
development
Automate YooAsset hot-update and asset bundles — build bundles, run Editor simulate builds, manage Collector groups, analyze BuildReport, and validate runtime. Use when building or simulating YooAsset bundles, configuring collectors, or validating hot-update assets, even if the user just says "热更" or "打AB包". 自动化 YooAsset 热更新与资源包(构建 bundle、编辑器模拟构建、管理 Collector 分组、分析 BuildReport、运行时校验);当用户要构建或模拟 YooAsset 资源包、配置 collector、或校验热更资源时使用。
development
Source-anchored design rules for YooAsset v2.3.18 — initialization, default-package shortcuts, play modes, asset handles, loading, updates, filesystem, build, and pitfalls. Use when writing or reviewing YooAsset code, initializing packages, loading assets via handles, setting up hot-update/download, or choosing a play mode, even if the user just says "热更" or "资源包". 为 YooAsset v2.3.18 提供源码锚定的设计规则(初始化、默认包快捷方式、运行模式、资源句柄、加载、更新、文件系统、构建、陷阱);当用户要编写或审查 YooAsset 代码、初始化 package、用句柄加载资源、配置热更/下载、或选择运行模式时使用。
data-ai
Last-resort guidance for safely hand-editing Unity serialized YAML (.unity/.prefab/.asset/.meta/ProjectSettings) — reference/fileID repair, GUID safety, and merge-conflict fixes. Use when REST cannot reach the change and YAML must be hand-edited — fixing m_Script GUIDs, broken fileID references, .meta files, or merge conflicts, even if the user just says "场景文件打不开" or "引用丢了". 安全手编 Unity 序列化 YAML(.unity/.prefab/.asset/.meta/ProjectSettings)的最后手段(引用/fileID 修复、GUID 安全、合并冲突修复);当 REST 无法触达、必须手编 YAML 时使用——修复 m_Script GUID、断裂 fileID 引用、.meta 文件或合并冲突。