Adoption

Agent Skills are supported by leading AI development tools.

VS Code Gemini CLI GitHub Goose Amp Cursor Claude Code Letta OpenCode Claude OpenAI Codex Factory VS Code Gemini CLI GitHub Goose Amp Cursor Claude Code Letta OpenCode Claude OpenAI Codex Factory

nuva-lab/skills/chunk-process

Name: skills/chunk-process
Author: nuva-lab

skills/chunk-process/SKILL.md

npx skillsauth add nuva-lab/vibecut skills/chunk-process

Clean

TrivyContainer and dependency vulnerability scanner

Clean

SemgrepStatic code analysis for vulnerabilities

Clean

mcp-scan (Snyk)Model Context Protocol security validation

Skipped

Snyk (dep)Open source security scanning

Skipped

Socket.devSupply chain security analysis

Skipped

VirusTotalMulti-engine malware detection

Skipped

CrowdStrikeAdvanced threat intelligence

Skipped

OSV-ScannerOpen Source Vulnerability database check

Skipped

OWASP Dep-Check

Chunk Process Skill

Smart video chunking and MLX-accelerated transcription for long-form content.

Problem Solved

Raw footage too long for single Gemini upload (~47 min = 5GB+)
Need word-level timestamps for precise cutting
Fixed-length chunks break mid-sentence

Smart Chunking

Instead of fixed 5-minute segments, smart_chunk.py finds natural break points:

python skills/chunk-process/smart_chunk.py raw_footage.mp4 -o chunks/

How it works:

Detect silence regions (>500ms gaps)
Target 2.5-3.5 minute chunks
Split at natural pauses, not mid-sentence
Output: chunk_001.mp4, chunk_002.mp4, ...

Options:

--min-chunk 150: Minimum chunk length (seconds)
--max-chunk 210: Maximum chunk length (seconds)
--silence-thresh -40: Silence detection threshold (dB)

MLX Transcription

mlx_transcribe.py uses MLX-accelerated Qwen3-ASR for fast transcription on Mac:

# Single file
python skills/chunk-process/mlx_transcribe.py audio.wav -o transcript.json

# Batch process chunks
python skills/chunk-process/mlx_transcribe.py chunks/ --batch --word-timestamps

Features:

3-5x faster than CPU on Apple Silicon
Auto language detection (Chinese/English/Mixed)
Word-level timestamps (~30ms precision)
Outputs: transcript.json with segments + words

Combined Pipeline

# 1. Smart chunk the video
python skills/chunk-process/smart_chunk.py raw.mp4 -o chunks/

# 2. Transcribe all chunks
python skills/chunk-process/mlx_transcribe.py chunks/ --batch --word-timestamps

# Output: chunks/transcript.json (merged from all chunks)

Output Format

{
  "language": "English",
  "segments": [
    {"text": "First sentence.", "start": 0.0, "end": 2.5},
    {"text": "Second sentence.", "start": 2.5, "end": 5.0}
  ],
  "words": [
    {"text": "First", "start": 0.0, "end": 0.3},
    {"text": "sentence", "start": 0.3, "end": 0.8}
  ],
  "full_text": "First sentence. Second sentence..."
}

Why Smart Chunking?

| Fixed Chunks | Smart Chunks | |--------------|--------------| | Breaks mid-word | Breaks at pauses | | 5 min arbitrary | 2.5-3.5 min natural | | Hard cuts | Clean transitions | | Timestamp gaps | Continuous timeline |

nuva-lab/skills/chunk-process

skills/chunk-process/SKILL.md

# Chunk Process Skill Smart video chunking and MLX-accelerated transcription for long-form content. ## Problem Solved - Raw footage too long for single Gemini upload (~47 min = 5GB+) - Need word-level timestamps for precise cutting - Fixed-length chunks break mid-sentence ## Smart Chunking Instead of fixed 5-minute segments, `smart_chunk.py` finds natural break points: ```bash python skills/chunk-process/smart_chunk.py raw_footage.mp4 -o chunks/ ``` **How it works:** 1. Detect silence regi

5 stars

development

Updated Apr 9, 2026

$ install --global

skillsauth

npx skillsauth add nuva-lab/vibecut skills/chunk-process

Install this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.

Security Scan Results

3 of 9 scanners reported clean

Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.

Scanners Passed

Scanners in report

Clean

TrivyContainer and dependency vulnerability scanner

95%

Clean

SemgrepStatic code analysis for vulnerabilities

95%

Clean

mcp-scan (Snyk)Model Context Protocol security validation

95%

Skipped

Snyk (dep)Open source security scanning

50%

Skipped

Socket.devSupply chain security analysis

50%

Skipped

VirusTotalMulti-engine malware detection

50%

Skipped

CrowdStrikeAdvanced threat intelligence

50%

Skipped

OSV-ScannerOpen Source Vulnerability database check

50%

Skipped

OWASP Dep-Check

50%

Last scanned: Apr 9, 2026, 2:20 AM7.0s5 files scanned

SKILL.md

Chunk Process Skill

Smart video chunking and MLX-accelerated transcription for long-form content.

Problem Solved

Raw footage too long for single Gemini upload (~47 min = 5GB+)
Need word-level timestamps for precise cutting
Fixed-length chunks break mid-sentence

Smart Chunking

Instead of fixed 5-minute segments, smart_chunk.py finds natural break points:

python skills/chunk-process/smart_chunk.py raw_footage.mp4 -o chunks/

How it works:

Detect silence regions (>500ms gaps)
Target 2.5-3.5 minute chunks
Split at natural pauses, not mid-sentence
Output: chunk_001.mp4, chunk_002.mp4, ...

Options:

--min-chunk 150: Minimum chunk length (seconds)
--max-chunk 210: Maximum chunk length (seconds)
--silence-thresh -40: Silence detection threshold (dB)

MLX Transcription

mlx_transcribe.py uses MLX-accelerated Qwen3-ASR for fast transcription on Mac:

# Single file
python skills/chunk-process/mlx_transcribe.py audio.wav -o transcript.json

# Batch process chunks
python skills/chunk-process/mlx_transcribe.py chunks/ --batch --word-timestamps

Features:

3-5x faster than CPU on Apple Silicon
Auto language detection (Chinese/English/Mixed)
Word-level timestamps (~30ms precision)
Outputs: transcript.json with segments + words

Combined Pipeline

# 1. Smart chunk the video
python skills/chunk-process/smart_chunk.py raw.mp4 -o chunks/

# 2. Transcribe all chunks
python skills/chunk-process/mlx_transcribe.py chunks/ --batch --word-timestamps

# Output: chunks/transcript.json (merged from all chunks)

Output Format

{
  "language": "English",
  "segments": [
    {"text": "First sentence.", "start": 0.0, "end": 2.5},
    {"text": "Second sentence.", "start": 2.5, "end": 5.0}
  ],
  "words": [
    {"text": "First", "start": 0.0, "end": 0.3},
    {"text": "sentence", "start": 0.3, "end": 0.8}
  ],
  "full_text": "First sentence. Second sentence..."
}

Why Smart Chunking?

Related Skills

nuva-lab/write-script

tools

VerifiedTrustedCommunity

Generate voiceover scripts in Joyce's style for video clips

5SKILL.mdUpdated Apr 9, 2026

nuva-lab/write-script

nuva-lab/voice-clone

tools

VerifiedTrustedCommunity

Clone a voice using qwen3-tts and generate speech from text

5SKILL.mdUpdated Apr 9, 2026

nuva-lab/skills/validate-media

development

VerifiedTrustedCommunity

# Validate Media Skill Pre-flight media validation and diagnostics using ffprobe. ## Purpose Check video/audio files for common issues before rendering: - Duration mismatches between video and audio tracks - Missing audio tracks - Codec compatibility - Volume levels - Potential freeze points ## Usage ```bash python skills/validate-media/validate.py <video_file> [--verbose] ``` ## Output JSON report with issues and recommendations: ```json { "file": "video.mp4", "video_duration": 35.1

5SKILL.mdUpdated Apr 9, 2026

nuva-lab/skills/validate-media

nuva-lab/transcribe-clip

tools

VerifiedTrustedCommunity

Transcribe a video clip using Gemini to get timestamped segments for captions

5SKILL.mdUpdated Apr 9, 2026

nuva-lab/transcribe-clip

Download

For Claude Desktop. Download once, then upload the file in the app — no terminal needed.

Need help? View full Cowork setup guide →

Install manually

Choose your platform

# Clone the repo
git clone https://github.com/nuva-lab/vibecut.git

# Copy into Claude Code skills folder (global)
cp -r vibecut/skills/chunk-process ~/.claude/skills/

Claude Code Skills — official skills path docs.

Repository

nuva-lab/vibecut

5 stars

Compatible with

Claude Code

OpenAI Codex CLI

ChatGPT