Uni Vision Engine (v1.2.0)
This skill leverages a local jimeng-api Docker service. It allows AI agents to fully control high-quality image-to-video and text-to-video generation using a valid sessionid.
🌟 Core Feature: Native Chat Image Interception (Best Practice)
With this skill, the AI Assistant can automatically intercept clothing/character images sent by the user in the chat interface and seamlessly pass them to the generation model—no manual web uploads required!
Content Moderation Warning (China Firewall)
Note: Because this relies on the domestic Jimeng/Seedance engine, there is strict automated content moderation for clothing. If you encounter error -2001 (First frame image upload failed: may contain violating content), this means the image is deemed "too revealing", shows too much skin, or contains sensitive elements. The firewall outright blocks these. No credits are deducted. If this occurs, ask the user to provide a different image or switch to an overseas engine like Luma/Runway.
CLI Usage (For Automation Scripts)
1. Text-to-Video
node {baseDir}/scripts/generate.js --prompt "A Shiba Inu surfing" --session "your_sessionid"
2. Image-to-Video (Requires --image)
node {baseDir}/scripts/generate.js --prompt "Model turning naturally to show outfit" --image "/tmp/target.jpg" --session "your_sessionid"
Notes:
- Requires sufficient credits in the Jimeng account.
- Using
jimeng-video-3.0-prodeducts 50 credits per run.