AI models
Every model worth using.
You never have to choose one.
Misu routes each shot to the image or video model that fits it — then holds your product and your cast steady across all of them. Here is what each one is for.
01
Video models
Image-to-video, reference-to-video and text-to-video — for product motion, campaign films and social cuts.
- ByteDanceSeedance 2.5ByteDance's longest single-shot model. Thirty seconds in one take, with the sound made in the same pass as the picture.
- ByteDanceSeedance 2.0Cinematic motion and sound for clips up to 15 seconds. The Fast tier runs the same engine for quick iteration.
- KuaishouKling 3.0The product motion model. Physical realism for spins, detail shots and garments, in three tiers.
- GoogleVeo 3Google's cinematic video model. The pick when light, lens and mood carry the shot.
- MiniMax (Hailuo)MiniMax H3Hailuo 3.0. Clips of 5 to 15 seconds at 2K, with stereo sound on every generation.
- ShengShuViduSeveral subjects, one shot. Give Vidu up to seven images and it composes them into a single clip.
02
Image models
Stills: product photography, on-model imagery, typography and edits.
- GoogleNano Banana ProGoogle's instruction-following image model. Say exactly what to change, and it changes that.
- OpenAIGPT ImageOpenAI's image models, from everyday shots to campaign finals. Flare for speed, Sunburst for control.
- ByteDanceSeedreamByteDance's image family. One model for dense, text-heavy heroes, two for keeping a face across a campaign.
- Black Forest LabsFlux.2Photorealism and close prompt adherence from Black Forest Labs. Kontext Max edits what you already have.
Get started
One brief.
The right model per shot.
Character Lock · Product Lock · Full commercial rights
Brands don't need an account with every lab. Misu keeps the face and the product consistent while the model underneath changes.
PRIVATE BETA · ONBOARDED PERSONALLY · FOUNDING PRICING LOCKED