AI Models
131+ video and image models, each with a live API playground, schemas, and code examples.
Sorted by newest
Gemini 3.8 Flash
studioGenerate text from a prompt and optional image_urls. Supports max_tokens and system_prompt.
GPT-Image 2.5
OpenAIOpenAI's latest image model. Two tiers: Flare for fast, high-quality output and Sunburst for maximum fidelity and instruction adherence.
Sound Effects
studioGenerate short sound effects from a text description — impacts, whooshes, fire, explosions, UI clicks, and other foley.
Seed Audio 1.0
studioGenerate layered scene audio from a text description — room tone, ambience, and a sequence of events in one pass, optionally guided by up to three reference clips. Use this instead of Sound Effects when you want an atmosphere rather than a single isolated hit.
Image Enhance
studioRepair photos at their existing resolution: sharpen out-of-focus and motion-blurred shots, clean high-ISO noise, restore damaged or degraded images, and correct exposure, white balance and colour.
Video Enhance
studioRepair and retime video without enlarging it: stabilize shaky footage, remove noise and motion blur, interpolate to a higher frame rate or slow motion, colorize black-and-white sources, and convert SDR to HDR.
Aion 3.0 Mini
studioGenerate text. Supports temperature, max_tokens, and system_prompt.
Gemma 4 Uncensored
studioGenerate text. Supports temperature, max_tokens, and system_prompt.
Qwen 3.6 Plus Uncensored
studioGenerate text. Supports temperature, max_tokens, and system_prompt.
LTX 2.5 Fast
studioFast LTX 2.5 video generation with native audio, up to 4K and 20 seconds
LTX 2.5 Pro
studioHigh-fidelity LTX 2.5 video generation with native audio, up to 1080p and 10 seconds
Wan 3
AlibabaWan 3 Prime, the faster Wan 3 endpoint. Hyper-real video up to 30 seconds with native audio, first/last frame control, and reference-to-video from up to 10 images, 5 clips, and 5 voice references.
Minimax Music 3
studioGenerate complete songs up to five minutes long from a music description and lyrics. Structure lyrics with tags like [verse] and [chorus] on their own lines.
Flux 3
Black Forest LabsNative-audio video with first/last frame control, up to 20 seconds and 1080p.
Filler Word Removal
studioAutomatically cut filler words ("um", "uh", "you know") and overlong silences out of a finished talking video. The result is a tighter re-cut of the same footage — no re-shoot, no re-render, and the original audio and framing are preserved between cuts.
MiniMax H3
MiniMaxFast 768P video with native stereo audio, strong prompt adherence, and text / image / reference-to-video modes.
World
studioTurn an image or a prompt into a navigable 3D world. Returns a Gaussian splat scene you can fly a camera through — generate a location once, find the exact angle, and capture the frame as a reference for video generation.
Browse by capability
Guide
Which AI model should I use?
Pick by what you are trying to create — you do not need to read every model card.
| Use case | Recommended model type | Best next page |
|---|---|---|
| AI influencer videos | Video model with reference / talking-head support | AI influencer generator |
| UGC product ads | Video model with a product workflow | AI UGC |
| Product photos | High-realism image model | Ecommerce visuals |
| Social ad variants | Fast/budget model first, premium model for winners | AI video ads |
| Talking heads | Lip-sync / audio-capable model | Talking head video |
Start free with credits and test any model
Every model runs on the same credits and the same API — switch between them without changing your integration.
Join Influencer Studio Today
Start creating amazing AI-generated content for your brand







