modellix
Integrate the Modellix API/CLI for async AI image, video, and speech generation or transcription (model run --wait, task download).
Design & UXCreative & Mediaapiaudio-generationcliimage-generationmodellixspeech-to-speechspeech-to-texttext-to-speech
Modellix
Overview
Modellix is a Model-as-a-Service platform for AI image, video, and speech generation or transcription. This skill teaches agents to use the official modellix-cli workflow (doctor → model run --wait → task download).
Upstream package: https://github.com/Modellix/modellix-plugin/tree/main/skills/modellix (Open Plugins layout; skill tree under skills/modellix/).
When to Use This Skill
- Generate images from text prompts
- Generate or edit videos from text or images
- Generate speech from text, transcribe speech, or transform one voice into another
- Call Modellix models through a unified API/CLI
- The user mentions Modellix, Seedream, Seedance, Nano Banana, Whisper, Qwen Audio, CosyVoice, or similar providers via Modellix
How It Works
- Authenticate with
MODELLIX_API_KEYormodellix-cli auth login - Run
modellix-cli doctor --json - Use default models when unspecified (T2I:
google/nano-banana-2-lite, T2V:bytedance/seedance-2.0-mini-t2v, I2I:google/nano-banana-2-lite-edit, I2V:bytedance/seedance-2.0-fast-i2v, V2V:bytedance/seedance-2.0-fast-v2v, TTS:alibaba/qwen-audio-3.0-tts-flash, STT:openai/whisper-1, STS:alibaba/cosyvoice-clone) - Submit with
modellix-cli model run --wait --json - Persist outputs with
modellix-cli task download
Examples
Text-to-image
modellix-cli model run \
--modeSubscribers only
The full skill, its 1 bundled files and every download is included with every paid Complete AI plan.
Details
| Source | Modellix/modellix-plugin |
|---|---|
| License | MIT |
| Risk label | critical ("critical" means the skill may run commands or touch files — read before use) |
| Files | SKILL.md |
| Added | 2026-07-16 |
Related skills
article-illustrations
Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but clear visual metaphors.
liuguang-banlan-ui
Builds two parameterized UI modes—流光溢彩白 (iridescent white) and 五彩斑斓黑 (colorful black)—with OKLCH, WebGL/CSS fallback, vision gating, screenshot QA, and total/per-color intensity reports. Use when a UI request names either mode or needs measured color parameters.
