video-editing
Video editing with ffmpeg
What this skill does
# Video Editing Video manipulation and conversion using ffmpeg. ## Get video info ```bash ffprobe -v quiet -print_format json -show_format -show_streams input.mp4 ``` ## Trim video ```bash # Trim from 00:01:30 for 60 seconds ffmpeg -i input.mp4 -ss 00:01:30 -t 60 -c copy trimmed.mp4 # Trim from start to specific end time ffmpeg -i input.mp4 -ss 00:00:00 -to 00:02:00 -c copy trimmed.mp4 ``` ## Convert format ```bash # MP4 to WebM ffmpeg -i input.mp4 -c:v libvpx-vp9 -c:a libopus output.webm # MOV to MP4 ffmpeg -i input.mov -c:v libx264 -c:a aac output.mp4 # AVI to MP4 ffmpeg -i input.avi -c:v libx264 -c:a aac -movflags +faststart output.mp4 ``` ## Extract audio ```bash # Extract audio as MP3 ffmpeg -i input.mp4 -vn -acodec libmp3lame -q:a 2 audio.mp3 # Extract audio as WAV ffmpeg -i input.mp4 -vn -acodec pcm_s16le audio.wav # Extract audio as AAC ffmpeg -i input.mp4 -vn -acodec aac audio.m4a ``` ## Create thumbnail ```bash # Extract a single frame at 10 seconds ffmpeg -i input.mp4 -ss 00:00:10 -frames:v 1 thumbnail.jpg # Create a thumbnail grid (4x4) ffmpeg -i input.mp4 -frames 1 -vf "select=not(mod(n\,100)),scale=320:240,tile=4x4" grid.jpg ``` ## Resize video ```bash # Resize to 1280x720 ffmpeg -i input.mp4 -vf scale=1280:720 -c:a copy resized.mp4 # Resize keeping aspect ratio (width 1280, auto height) ffmpeg -i input.mp4 -vf scale=1280:-2 -c:a copy resized.mp4 # Resize to 50% ffmpeg -i input.mp4 -vf scale=iw/2:ih/2 -c:a copy half.mp4 ``` ## Add watermark ```bash # Add image watermark (bottom-right corner) ffmpeg -i input.mp4 -i watermark.png -filter_complex "overlay=W-w-10:H-h-10" watermarked.mp4 # Add text watermark ffmpeg -i input.mp4 -vf "drawtext=text='Copyright 2025':fontsize=24:fontcolor=white:x=W-tw-10:y=H-th-10" watermarked.mp4 ``` ## Concatenate videos ```bash # Create a file list printf "file '%s'\n" part1.mp4 part2.mp4 part3.mp4 > filelist.txt # Concatenate using the file list ffmpeg -f concat -safe 0 -i filelist.txt -c copy output.mp4 ``` ## Extract frames ```bash # Extract all frames as images ffmpeg -i input.mp4 frames/frame_%04d.png # Extract 1 frame per second ffmpeg -i input.mp4 -vf fps=1 frames/frame_%04d.png # Extract frames at specific interval (every 5 seconds) ffmpeg -i input.mp4 -vf fps=1/5 frames/frame_%04d.png ``` ## Compress video ```bash # Compress with CRF (lower = better quality, 18-28 typical) ffmpeg -i input.mp4 -c:v libx264 -crf 23 -preset medium -c:a aac -b:a 128k compressed.mp4 # Aggressive compression for smaller file ffmpeg -i input.mp4 -c:v libx264 -crf 28 -preset slow -c:a aac -b:a 96k small.mp4 # Two-pass encoding for target file size ffmpeg -i input.mp4 -c:v libx264 -b:v 1M -pass 1 -f null /dev/null ffmpeg -i input.mp4 -c:v libx264 -b:v 1M -pass 2 -c:a aac -b:a 128k output.mp4 ```
Related in Image & Video
watch
IncludedWatch a video (URL or local path). Downloads with yt-dlp, extracts auto-scaled frames with ffmpeg, pulls the transcript from captions (or Whisper API fallback), and hands the result to Claude so it can answer questions about what's in the video.
physical-ai-defect-image-generation
IncludedUse when the user wants to orchestrate defect image generation, run associated setup, or handle outputs on OSMO. The Day 0 path handles cold-start with USD-to-ROI, image-edit augmentation, and AnomalyGen to create initial PCBA datasets. The Day 1 path performs inference and labeling on real images. This skill helps with first-time asset setup, creation of finetuning checkpoints, and configuring deployment. Trigger keywords: defect image generation, dig workflow, dig pipeline, defect image detection workflow, aoi pipeline, aoi anomalygen, usd2roi anomalygen, day 0 pcba, day 1 pcba, day 1 real-photo alignment, day 1 manual roi, metal surface anomaly, glass defect, anomalygen finetune, setup_pcb, setup_metal, setup_glass, setup_pretrained, dig setup, dig datasets, dig pretrained checkpoint, dig image-edit endpoint.
accelint-react-best-practices
IncludedReact performance optimization and best practices. ALWAYS use this skill when working with any React code - writing components, hooks, JSX; refactoring; optimizing re-renders, memoization, state management; reviewing for performance; fixing hydration mismatches; debugging infinite re-renders, stale closures, input focus loss, animations restarting; preventing remounting; implementing transitions, lazy initialization, effect dependencies. Even simple React tasks benefit from these patterns. Covers React 19+ (useEffectEvent, Activity, ref props). Triggers - useEffect, useState, useMemo, useCallback, memo, inline components, nested components, components inside components, re-render, performance, hydration, SSR, Next.js, useDeferredValue, combined hooks.
elevenlabs-agents
IncludedBuild conversational AI voice agents with ElevenLabs Platform using React, JavaScript, React Native, or Swift SDKs. Configure agents, tools (client/server/MCP), RAG knowledge bases, multi-voice, and Scribe real-time STT. Use when: building voice chat interfaces, implementing AI phone agents with Twilio, configuring agent workflows or tools, adding RAG knowledge bases, testing with CLI "agents as code", or troubleshooting deprecated @11labs packages, Android audio cutoff, CSP violations, dynamic variables, or WebRTC config. Keywords: ElevenLabs Agents, ElevenLabs voice agents, AI voice agents, conversational AI, @elevenlabs/react, @elevenlabs/client, @elevenlabs/react-native, @elevenlabs/elevenlabs-js, @elevenlabs/agents-cli, elevenlabs SDK, voice AI, TTS, text-to-speech, ASR, speech recognition, turn-taking model, WebRTC voice, WebSocket voice, ElevenLabs conversation, agent system prompt, agent tools, agent knowledge base, RAG voice agents, multi-voice agents, pronunciation dictionary, voice speed control, elevenlabs scribe, @11labs deprecated, Android audio cutoff, CSP violation elevenlabs, dynamic variables elevenlabs, case-sensitive tool names, webhook authentication
humanizer
IncludedHumanize AI-generated text by detecting and removing patterns typical of LLM output. Rewrites text to sound natural, specific, and human. Uses 28 pattern detectors, 560+ AI vocabulary terms across 3 tiers, and statistical analysis (burstiness, type-token ratio, readability) for comprehensive detection. Use when asked to humanize text, de-AI writing, make content sound more natural/human, review writing for AI patterns, score text for AI detection, or improve AI-generated drafts. Covers content, language, style, communication, and filler categories.
generating-mermaid-diagrams
IncludedSalesforce architecture diagrams using Mermaid with ASCII fallback. Use this skill when generating text-based diagrams for Salesforce architecture, OAuth flows, ERDs, integration sequences, or Agentforce structure. TRIGGER when: user says "diagram", "visualize", "ERD", or asks for sequence diagrams, flowcharts, class diagrams, or architecture visualizations in Mermaid. DO NOT TRIGGER when: user wants PNG/SVG image output (use generating-visual-diagrams), or asks about non-Salesforce systems.