ai-video-prompting
Générez des prompts optimisés pour chaque modèle de génération vidéo IA (Veo 3, Runway Gen-3, Kling 2.6, Pika), en exploitant leurs forces spécifiques. Use when: **Animer des frames de storyboard** - Transformer des images fixes en vidéo; **Choisir le bon modèle** - Sélectionner Veo, Runway, Kling ou Pika selon le besoin; **Optimiser la qualité de génération** - Prompts structurés pour meilleurs résultats; **Créer des transitions fluides** - Scene extension, first/last frame; **Utiliser le mo...
What this skill does
# AI Video Prompting > Générez des prompts optimisés pour chaque modèle de génération vidéo IA (Veo 3, Runway Gen-3, Kling 2.6, Pika), en exploitant leurs forces spécifiques. ## When to Use This Skill - **Animer des frames de storyboard** - Transformer des images fixes en vidéo - **Choisir le bon modèle** - Sélectionner Veo, Runway, Kling ou Pika selon le besoin - **Optimiser la qualité de génération** - Prompts structurés pour meilleurs résultats - **Créer des transitions fluides** - Scene extension, first/last frame - **Utiliser le motion control** - Face-driving et performance transfer avec Kling ## Methodology Foundation **Source**: Documentation officielle Veo 3.1, Runway Gen-3, Kling 2.6, Pika 2.2 + PJ Ace workflow **Core Principle**: "Chaque modèle a ses forces. Veo excelle en audio natif, Runway en contrôle caméra, Kling en motion control, Pika en accessibilité. Choisissez le bon outil pour chaque shot." **Why This Matters**: Un prompt générique donne des résultats génériques. En structurant vos prompts selon les spécificités de chaque modèle, vous obtenez des vidéos de qualité professionnelle. ## What Claude Does vs What You Decide | Claude Does | You Decide | |-------------|------------| | Structures production workflow | Final creative direction | | Suggests technical approaches | Equipment and tool choices | | Creates templates and checklists | Quality standards | | Identifies best practices | Brand/voice decisions | | Generates script outlines | Final script approval | ## What This Skill Does 1. **Sélectionne le modèle optimal** - Selon le type de shot et les besoins 2. **Structure les prompts par modèle** - Formats optimisés pour chaque plateforme 3. **Configure les paramètres techniques** - Durée, résolution, aspect ratio 4. **Guide le workflow image-to-video** - Upload et animation des frames 5. **Planifie les extensions de scène** - Clips continus au-delà de 8 secondes ## How to Use ### Animer un storyboard complet ``` J'ai ces frames de storyboard: [description]. Génère les prompts d'animation pour Veo 3.1. ``` ### Choisir le bon modèle ``` J'ai besoin de [décrire le shot]. Quel modèle utiliser et comment structurer le prompt? ``` ### Créer une séquence continue ``` J'ai un clip de 8 secondes et je veux l'étendre à 30 secondes. Guide-moi avec scene extension. ``` ## Instructions ### Step 1: Sélectionner le modèle optimal ``` ## Guide de Sélection | Besoin | Modèle recommandé | Pourquoi | |--------|-------------------|----------| | Audio synchronisé natif | **Veo 3.1** | Seul modèle avec génération audio intégrée | | Contrôle caméra précis | **Runway Gen-3** | Camera control directionnel | | Face-driving / Performance | **Kling 2.6** | Motion control depuis vidéo référence | | Transitions image-image | **Pika 2.2** | Pikaframes optimisé pour ça | | Budget limité | **Pika 2.2** | Freemium généreux | | Qualité maximale | **Veo 3.1** | 1080p natif, meilleure cohérence | | Expérimentation rapide | **Pika 2.2** | Itération rapide | | Personnage récurrent | **Kling 2.6** | Character reference | ### Matrice de décision **Q1: Avez-vous besoin d'audio synchronisé?** - Oui → Veo 3.1 (dialogue, sound effects) - Non → Continuer **Q2: Avez-vous une vidéo de performance à transférer?** - Oui → Kling 2.6 (motion control) - Non → Continuer **Q3: Avez-vous besoin de contrôle caméra précis?** - Oui → Runway Gen-3 (camera control) - Non → Continuer **Q4: Avez-vous un budget limité?** - Oui → Pika 2.2 - Non → Veo 3.1 (qualité par défaut) ``` --- ### Step 2: Prompts Veo 3.1 (Google) ``` ## Veo 3.1 Prompt Structure ### Format de base [SHOT TYPE] of [SUBJECT] [ACTION] in [SETTING]. [CAMERA MOVEMENT]. [LIGHTING]. [STYLE/MOOD]. [AUDIO DESCRIPTION if needed]. ### Paramètres - Durée: 4-8 secondes/génération - Résolution: 1080p - Aspect: 16:9 ou 9:16 - Audio: Activé par défaut (dialogue, SFX, ambiance) ### Exemple complet ``` Medium shot of a confident businesswoman walking through a modern glass office lobby. Slow dolly following her from the side. Warm morning sunlight streaming through windows, creating lens flares. Cinematic corporate style, shallow depth of field. Sound of heels clicking on marble floor, ambient office murmur in background. ``` ### Tips Veo 3.1 - Décrire explicitement l'audio souhaité - Utiliser des termes cinématographiques (dolly, tracking, pan) - Spécifier l'éclairage en détail - Inclure l'ambiance sonore pour meilleur matching ### Scene Extension (clips longs) 1. Générer clip initial (8s) 2. Télécharger 3. Dans Flow/Gemini: "Extend this scene with [continuation]" 4. Le modèle utilise la dernière seconde comme référence 5. Répéter jusqu'à durée souhaitée ``` --- ### Step 3: Prompts Runway Gen-3 Alpha ``` ## Runway Gen-3 Prompt Structure ### Format de base [DESCRIPTIVE SCENE]. [SUBJECT ACTION]. [CAMERA INSTRUCTION]. [LIGHTING AND ATMOSPHERE]. [STYLE REFERENCE]. ### Paramètres - Durée: 5 ou 10 secondes - Mode: Text-to-Video ou Image-to-Video - Camera Control: Direction + Intensité - First/Last Frame: Définir début et/ou fin ### Camera Control Options | Direction | Effet | |-----------|-------| | Pan Left/Right | Rotation horizontale | | Tilt Up/Down | Rotation verticale | | Dolly In/Out | Mouvement avant/arrière | | Truck Left/Right | Mouvement latéral | | Pedestal Up/Down | Mouvement vertical | | Zoom In/Out | Changement focal | | Static | Pas de mouvement | Intensité: Low / Medium / High ### Exemple avec Camera Control ``` Prompt: A mysterious forest at twilight, fog rolling between ancient trees, fireflies beginning to glow. Ethereal fantasy atmosphere, cinematic color grading. Camera: Dolly In, Medium intensity First Frame: [Upload image du storyboard] Duration: 10 seconds ``` ### First/Last Frame Technique - **First Frame**: L'image devient le point de départ - **Last Frame**: L'image devient la destination - **Both**: Transition animée entre deux images ### Workflow Image-to-Video 1. Upload frame upscalée du storyboard 2. Sélectionner "First Frame" ou "Last Frame" 3. Écrire prompt décrivant le mouvement souhaité 4. Configurer Camera Control 5. Générer (10 crédits/sec standard) ``` --- ### Step 4: Prompts Kling 2.6 Motion Control ``` ## Kling 2.6 Prompt Structure ### Concept clé: Motion Transfer Kling 2.6 permet de "conduire" un personnage généré avec une vidéo de performance réelle. ### Inputs requis 1. **Reference Image**: Le personnage/look souhaité 2. **Driving Video**: La vidéo de performance (mouvements, expressions) 3. **Text Prompt**: Instructions supplémentaires ### Format de prompt [SUBJECT DESCRIPTION matching reference image]. [SCENE/ENVIRONMENT]. [STYLE]. Motion: Transfer expressions and movements from driving video. ### Exemple Motion Control ``` Reference Image: [Portrait d'un astronaute stylisé] Driving Video: [Vous parlant à la caméra, 10 secondes] Prompt: An astronaut in a futuristic space station, speaking directly to camera. Sci-fi movie lighting, volumetric fog, control panels glowing in background. Transfer facial expressions and lip movements from driving video. Duration: 10 seconds Resolution: 1080p ``` ### Best Practices Motion Control - **Vidéo driving**: Fond simple, contraste élevé - **Silhouette claire**: Le sujet doit être distinct - **Mouvements**: Fonctionne pour expressions, lips, corps entier - **Limite**: Éviter les actions trop rapides ou complexes ### Use Cases | Scénario | Setup | |----------|-------| | 1 acteur → N personnages | 1 driving video, N reference images | | Lip-sync dialogue | Acteur parle, personnage anime | | Performance dance | Vidéo danse → personnage stylisé | | Deep fake éthique | Avec autorisation du sujet | ``` --- ### Step 5: Prompts Pika 2.2 ``` ## Pika 2.2 Prompt Structure ### Format de base [SCENE DESCRIPTION]. [ACTION]. [STYLE]. ### Paramètres - Durée: Jusqu'à 10 secondes - Résolution: 1080p - Outils spéciaux: Pikaframes, Pikaswaps, Pikadditions ### Pikaframes (Transitions) Anime une transition fluide entre
Related in Image & Video
watch
IncludedWatch a video (URL or local path). Downloads with yt-dlp, extracts auto-scaled frames with ffmpeg, pulls the transcript from captions (or Whisper API fallback), and hands the result to Claude so it can answer questions about what's in the video.
physical-ai-defect-image-generation
IncludedUse when the user wants to orchestrate defect image generation, run associated setup, or handle outputs on OSMO. The Day 0 path handles cold-start with USD-to-ROI, image-edit augmentation, and AnomalyGen to create initial PCBA datasets. The Day 1 path performs inference and labeling on real images. This skill helps with first-time asset setup, creation of finetuning checkpoints, and configuring deployment. Trigger keywords: defect image generation, dig workflow, dig pipeline, defect image detection workflow, aoi pipeline, aoi anomalygen, usd2roi anomalygen, day 0 pcba, day 1 pcba, day 1 real-photo alignment, day 1 manual roi, metal surface anomaly, glass defect, anomalygen finetune, setup_pcb, setup_metal, setup_glass, setup_pretrained, dig setup, dig datasets, dig pretrained checkpoint, dig image-edit endpoint.
accelint-react-best-practices
IncludedReact performance optimization and best practices. ALWAYS use this skill when working with any React code - writing components, hooks, JSX; refactoring; optimizing re-renders, memoization, state management; reviewing for performance; fixing hydration mismatches; debugging infinite re-renders, stale closures, input focus loss, animations restarting; preventing remounting; implementing transitions, lazy initialization, effect dependencies. Even simple React tasks benefit from these patterns. Covers React 19+ (useEffectEvent, Activity, ref props). Triggers - useEffect, useState, useMemo, useCallback, memo, inline components, nested components, components inside components, re-render, performance, hydration, SSR, Next.js, useDeferredValue, combined hooks.
elevenlabs-agents
IncludedBuild conversational AI voice agents with ElevenLabs Platform using React, JavaScript, React Native, or Swift SDKs. Configure agents, tools (client/server/MCP), RAG knowledge bases, multi-voice, and Scribe real-time STT. Use when: building voice chat interfaces, implementing AI phone agents with Twilio, configuring agent workflows or tools, adding RAG knowledge bases, testing with CLI "agents as code", or troubleshooting deprecated @11labs packages, Android audio cutoff, CSP violations, dynamic variables, or WebRTC config. Keywords: ElevenLabs Agents, ElevenLabs voice agents, AI voice agents, conversational AI, @elevenlabs/react, @elevenlabs/client, @elevenlabs/react-native, @elevenlabs/elevenlabs-js, @elevenlabs/agents-cli, elevenlabs SDK, voice AI, TTS, text-to-speech, ASR, speech recognition, turn-taking model, WebRTC voice, WebSocket voice, ElevenLabs conversation, agent system prompt, agent tools, agent knowledge base, RAG voice agents, multi-voice agents, pronunciation dictionary, voice speed control, elevenlabs scribe, @11labs deprecated, Android audio cutoff, CSP violation elevenlabs, dynamic variables elevenlabs, case-sensitive tool names, webhook authentication
humanizer
IncludedHumanize AI-generated text by detecting and removing patterns typical of LLM output. Rewrites text to sound natural, specific, and human. Uses 28 pattern detectors, 560+ AI vocabulary terms across 3 tiers, and statistical analysis (burstiness, type-token ratio, readability) for comprehensive detection. Use when asked to humanize text, de-AI writing, make content sound more natural/human, review writing for AI patterns, score text for AI detection, or improve AI-generated drafts. Covers content, language, style, communication, and filler categories.
generating-mermaid-diagrams
IncludedSalesforce architecture diagrams using Mermaid with ASCII fallback. Use this skill when generating text-based diagrams for Salesforce architecture, OAuth flows, ERDs, integration sequences, or Agentforce structure. TRIGGER when: user says "diagram", "visualize", "ERD", or asks for sequence diagrams, flowcharts, class diagrams, or architecture visualizations in Mermaid. DO NOT TRIGGER when: user wants PNG/SVG image output (use generating-visual-diagrams), or asks about non-Salesforce systems.