UNPKG

@sogni-ai/sogni-creative-agent-skill

Version:

Sogni Creative Agent Skill: agent skill and CLI for Sogni AI image, video, and music generation.

23 lines (16 loc) 1.37 kB
--- name: video_generation description: Text-to-video synthesis with LTX-2.3, WAN 2.2, Seedance 2.0, HappyHorse 1.1, and MiniMax H3. always_loaded: false tool_names: - generate_video --- # Video generation Text-to-video synthesis with LTX-2.3, WAN 2.2, Seedance 2.0, HappyHorse 1.1, and MiniMax H3. Use when the user wants a new video clip generated from a prompt or a loose multimodal reference set. ## Tools - `generate_video` produce a video clip from text, Seedance/H3 multimodal references, or HappyHorse image references. ## Constraints - Persona-driven video requests must always go through `image_editing` first to produce a conditioned image; never go straight to text-to-video for personas. - For prompt-only variants with the same model, duration, dimensions, and references, use one Dynamic Prompt branch with `numberOfVariations`/`-n` instead of serial video calls. - HappyHorse 1.1 is a Premium Spark vendor path: t2v is prompt-only, i2v uses one first-frame image, and r2v uses 1-9 image references. It accepts no reference audio or video. - MiniMax H3 t2v uses `minimax-h3-t2v`; H3 r2v must be selected explicitly as `minimax-h3-r2v` and accepts up to 9 images, 3 videos, and 3 audio clips, capped at 12 files with at least one image. Address references with per-type `<Picture 1>` / `<Video 1>` / `<Audio 1>` tags and assign each one a role.