Complete AI TrainingYourJobSkills for your job

Skills / creative

modellix

Integrate the Modellix API/CLI for async AI image, video, and speech generation or transcription (model run --wait, task download).

Design & UXCreative & Mediaapiaudio-generationcliimage-generationmodellixspeech-to-speechspeech-to-texttext-to-speech

Modellix

Overview

Modellix is a Model-as-a-Service platform for AI image, video, and speech generation or transcription. This skill teaches agents to use the official modellix-cli workflow (doctor → model run --wait → task download).

Upstream package: https://github.com/Modellix/modellix-plugin/tree/main/skills/modellix (Open Plugins layout; skill tree under skills/modellix/).

When to Use This Skill

  • Generate images from text prompts
  • Generate or edit videos from text or images
  • Generate speech from text, transcribe speech, or transform one voice into another
  • Call Modellix models through a unified API/CLI
  • The user mentions Modellix, Seedream, Seedance, Nano Banana, Whisper, Qwen Audio, CosyVoice, or similar providers via Modellix

How It Works

  1. Authenticate with MODELLIX_API_KEY or modellix-cli auth login
  2. Run modellix-cli doctor --json
  3. Use default models when unspecified (T2I: google/nano-banana-2-lite, T2V: bytedance/seedance-2.0-mini-t2v, I2I: google/nano-banana-2-lite-edit, I2V: bytedance/seedance-2.0-fast-i2v, V2V: bytedance/seedance-2.0-fast-v2v, TTS: alibaba/qwen-audio-3.0-tts-flash, STT: openai/whisper-1, STS: alibaba/cosyvoice-clone)
  4. Submit with modellix-cli model run --wait --json
  5. Persist outputs with modellix-cli task download

Examples

Text-to-image

modellix-cli model run \
  --mode

Subscribers only

The full skill, its 1 bundled files and every download is included with every paid Complete AI plan.

Details

SourceModellix/modellix-plugin
LicenseMIT
Risk labelcritical ("critical" means the skill may run commands or touch files — read before use)
FilesSKILL.md
Added2026-07-16

Related skills

article-illustrations

Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but clear visual metaphors.

liuguang-banlan-ui

Builds two parameterized UI modes—流光溢彩白 (iridescent white) and 五彩斑斓黑 (colorful black)—with OKLCH, WebGL/CSS fallback, vision gating, screenshot QA, and total/per-color intensity reports. Use when a UI request names either mode or needs measured color parameters.