---
source: https://autoadify.com/ai-models
title: "AI models in Autoadify"
---

# AI models in Autoadify

> Autoadify is the social media management platform that lets you choose the AI model for every generation. 76 selectable models from 16 labs across four modalities, each named, each tagged with what it is good at.

The model is chosen per generation, not per account: a single post can use Claude Sonnet 5 for the caption, Nano Banana Pro for the image and Veo 3.1 for the reel. Model choice is included in every paid plan rather than metered.

## Do you need your own API keys or a plan at each lab?

No. All 76 models across 16 labs are reached through one Autoadify subscription. There is nothing to create at OpenAI, Google, Anthropic, xAI or ElevenLabs, no key to paste or rotate, and no separate plan at any of them. Autoadify is free to start, then Pro at $29/mo, billed monthly per workspace, not per seat or per channel.

## Which model for which job

- **Best AI model for captions and hooks** — GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna, GPT-5.4 Mini, Claude Sonnet 5, Claude Sonnet 4.6, Claude Haiku 4.5, Gemini 3 Flash, Grok 4.3, Grok 4.20, Mistral Medium 3.5.
- **Best AI model for LinkedIn posts and threads** — GPT-5.6 Sol, GPT-5.6 Terra, Claude Opus 4.8, Claude Sonnet 5, Claude Opus 4.6, Claude Sonnet 4.6, Gemini 3.1 Pro, Grok 4.20 Multi-Agent, Kimi K3, Qwen 3.8 2.4T (MoE), DeepSeek V4 Pro.
- **Best AI model for product photography** — Nano Banana 2, Nano Banana Pro, Seedream 4.5, Seedream 5 Lite, Wan 2.7 Image Pro, Krea 2 Medium, Seedream 5 Pro.
- **Best AI model for text inside an image** — Nano Banana 2, Nano Banana Pro, GPT Image 2, Qwen Image 3, Nano Banana Pro Edit, GPT Image 2 Edit, Qwen Image 3 Edit.
- **Best AI model for illustration and graphic styles** — GPT Image 2, Grok Imagine, Qwen Image 3, Krea 2 Medium, Krea 2 Medium Turbo.
- **Best AI model for editing an existing photo** — Nano Banana Pro Edit, Nano Banana 2 Edit, GPT Image 2 Edit, Seedream 4.5 Edit, Seedream 5 Lite Edit, Grok Imagine Edit, Qwen Image 3 Edit, Krea 2 Medium Edit, Krea 2 Medium Turbo Edit, Seedream 5 Pro Edit.
- **Best AI model for talking-head video** — Veo 3.1, Veo 3.1 Fast, Veo 3.1 i2v, Veo 3.1 Fast i2v.
- **Best AI model for B-roll and product motion** — Veo 3.1, Veo 3.1 Fast, Veo 3.1 Lite, Grok Imagine Video, Grok Imagine i2v, MiniMax H3, Happyhorse Video, MiniMax H3 i2v, Happyhorse i2v, Kling 3.0 Pro, Kling 3.0 Standard, Kling 3.0 4K, Seedance 2.0, Seedance 2.0 Fast, Veo 3.1 i2v, Veo 3.1 Fast i2v, Veo 3.1 Lite i2v, Seedance 2.0 i2v, Seedance 2.0 Fast i2v, Runway Gen-4.5, Runway Gen-4.5 i2v, Seedance 2.0 Mini, Seedance 2.0 Mini i2v, FLUX.3 Video, FLUX.3 Video i2v, Kling 3.0 Pro i2v, Kling 3.0 Standard i2v, Kling 3.0 4K i2v.
- **Cheapest AI models for high-volume work** — GPT-5.6 Luna, GPT-5.4 Mini, GPT-5.4 Nano, Claude Haiku 4.5, Gemini 3 Flash, Gemini 3.1 Flash Lite, Qwen 3.7 Flash, DeepSeek V4 Pro, DeepSeek V4 Flash, MiniMax M3, Mistral Medium 3.5, Mistral Small, Seedream 5 Lite, Krea 2 Medium Turbo, Seedream 5 Lite Edit, Krea 2 Medium Turbo Edit, Veo 3.1 Lite, Happyhorse Video, Happyhorse i2v, Kling 3.0 Standard, Seedance 2.0 Fast, Veo 3.1 Lite i2v, Seedance 2.0 Fast i2v, Seedance 2.0 Mini, Seedance 2.0 Mini i2v, Kling 3.0 Standard i2v.

## Every model

### Text models (24)

Captions, hooks, threads, replies and long-form posts. The spread here is mostly about register and cost — a caption does not need a flagship, a LinkedIn essay usually does.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| GPT-5.6 Sol | OpenAI | OpenAI's latest flagship — 1M context | Long-form, Short copy | Mid |
| GPT-5.6 Terra | OpenAI | GPT-5.6 all-rounder — cheaper than Sol | Long-form, Short copy | Mid |
| GPT-5.6 Luna | OpenAI | Fast, low-cost GPT-5.6 — 1M context | Short copy, Budget | Low |
| GPT-5.4 Mini | OpenAI | Fast, lower-cost GPT-5.4 | Short copy, Budget | Low |
| GPT-5.4 Nano | OpenAI | Cheapest GPT — high-volume tasks | Budget | Low |
| Claude Opus 4.8 | Anthropic | Anthropic's most powerful model | Long-form | High |
| Claude Sonnet 5 | Anthropic | Fast and highly capable — great default | Long-form, Short copy | Mid |
| Claude Opus 4.6 | Anthropic | Anthropic's most powerful model | Long-form | High |
| Claude Sonnet 4.6 | Anthropic | Fast and highly capable | Long-form, Short copy | High |
| Claude Haiku 4.5 | Anthropic | Fast, cheapest Claude | Short copy, Budget | Mid |
| Gemini 3.1 Pro | Google | Google's top reasoning model | Long-form | Mid |
| Gemini 3 Flash | Google | Fast multimodal Gemini | Short copy, Budget | Low |
| Gemini 3.1 Flash Lite | Google | Cheapest Gemini for high-volume tasks | Budget | Low |
| Grok 4.3 | xAI | xAI reasoning + real-time web | Short copy | Low |
| Grok 4.20 | xAI | xAI's latest flagship | Short copy | Low |
| Grok 4.20 Multi-Agent | xAI | Multi-agent reasoning — best for complex tasks | Long-form | Low |
| Kimi K3 | Moonshot AI | Moonshot AI — 1M context | Long-form | High |
| Qwen 3.8 2.4T (MoE) | Alibaba | Alibaba's flagship — 2.4T mixture-of-experts, 95B active | Long-form | Mid |
| Qwen 3.7 Flash | Alibaba | Cheapest model in the catalog — high-volume tasks | Budget | Low |
| DeepSeek V4 Pro | DeepSeek | DeepSeek's flagship — strong reasoning, low cost | Long-form, Budget | Low |
| DeepSeek V4 Flash | DeepSeek | Fast DeepSeek — 1.3M context | Budget | Low |
| MiniMax M3 | MiniMax | MiniMax — 1M context, low cost | Budget | Low |
| Mistral Medium 3.5 | Mistral | Mistral's balanced model — 262K context | Short copy, Budget | Mid |
| Mistral Small | Mistral | Small, very low cost — high-volume tasks | Budget | Low |

### Image models — text to image (11)

Product stills, lifestyle scenes, carousels and thumbnails, generated from a written prompt. Split on whether you need photoreal product truth or a stylised look.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| Nano Banana 2 | Google | Google Gemini 3.1 Flash image | Photoreal, Text in image | Mid |
| Nano Banana Pro | Google | Google DeepMind — 2K sharp imagery | Photoreal, Text in image | High |
| GPT Image 2 | OpenAI | OpenAI GPT Image 2 — text to image | Text in image, Illustration | Mid |
| Seedream 4.5 | ByteDance | ByteDance Seedream 4.5 — text to image | Photoreal | Mid |
| Seedream 5 Lite | ByteDance | ByteDance Seedream 5 Lite — fast text to image | Photoreal, Budget | Mid |
| Wan 2.7 Image Pro | Alibaba | Alibaba Wan 2.7 — pro-tier image generation | Photoreal | High |
| Grok Imagine | xAI | xAI multimodal image generation | Illustration | Low |
| Qwen Image 3 | Alibaba | Alibaba Qwen — sharp text rendering down to 10px | Text in image, Illustration | Mid |
| Krea 2 Medium | Krea | Krea's balanced model — stable, consistent generations | Photoreal, Illustration | Mid |
| Krea 2 Medium Turbo | Krea | Distilled Krea 2 Medium — fastest iteration | Illustration, Budget | Low |
| Seedream 5 Pro | ByteDance | ByteDance Seedream 5 Pro — precise editing control, lifelike scenes | Photoreal | Mid |

### Image models — editing an existing image (11)

Point these at a photo you already have: swap a background, restyle it, extend the frame, or clean it up. This is the tier that turns one product shot into a month of posts.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| Nano Banana Pro Edit | Google | Edit existing images — HD quality | Photo editing, Text in image | High |
| Nano Banana 2 Edit | Google | Edit existing images — fast | Photo editing | Mid |
| GPT Image 2 Edit | OpenAI | OpenAI GPT Image 2 — image to image | Photo editing, Text in image | Low |
| Seedream 4.5 Edit | ByteDance | ByteDance Seedream 4.5 — image to image | Photo editing | Low |
| Seedream 5 Lite Edit | ByteDance | ByteDance Seedream 5 Lite — fast image to image | Photo editing, Budget | Low |
| Background Remover | fal | Remove background from any image | — | Low |
| Grok Imagine Edit | xAI | xAI Grok Imagine — image to image | Photo editing | Low |
| Qwen Image 3 Edit | Alibaba | Qwen Image 3 — edit with up to 4 reference images | Photo editing, Text in image | Low |
| Krea 2 Medium Edit | Krea | Krea 2 Medium — restyle from one reference image | Photo editing | Mid |
| Krea 2 Medium Turbo Edit | Krea | Distilled Krea 2 Medium — fast restyle from a reference | Photo editing, Budget | Low |
| Seedream 5 Pro Edit | ByteDance | Seedream 5 Pro — edit with up to 14 reference images | Photo editing | Mid |

### Video models — text to video (14)

Reels, Shorts and TikToks from a written prompt. Several generate audio in the same pass as the picture, so the clip arrives finished rather than silent.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| Veo 3.1 | Google | Google Veo 3.1 — flagship cinematic AI video | Talking video, B-roll | High |
| Veo 3.1 Fast | Google | Veo 3.1 at faster speed — strong quality, cheaper | Talking video, B-roll | Mid |
| Veo 3.1 Lite | Google | Veo 3.1 Lite — cheapest tier, high-volume friendly | B-roll, Budget | Low |
| Grok Imagine Video | xAI | xAI text-to-video | B-roll | Low |
| MiniMax H3 | MiniMax | MiniMax H3 — 2K text to video, strong text & brand rendering | B-roll | Mid |
| Happyhorse Video | Alibaba | Alibaba Happyhorse — text to video | B-roll, Budget | Mid |
| Kling 3.0 Pro | Kling | Kling 3.0 — text to video at 1080p | B-roll | Mid |
| Kling 3.0 Standard | Kling | Kling 3.0 — text to video at 720p (cheaper, faster) | B-roll, Budget | Mid |
| Kling 3.0 4K | Kling | Kling 3.0 — text to video at 4K (premium) | B-roll | High |
| Seedance 2.0 | ByteDance | Bytedance Seedance 2.0 — text to video, up to 1080p | B-roll | High |
| Seedance 2.0 Fast | ByteDance | Seedance 2.0 Fast — quicker, 720p cap | B-roll, Budget | High |
| Runway Gen-4.5 | Runway | Runway Gen-4.5 — text to video, 720p, up to 10s | B-roll | Mid |
| Seedance 2.0 Mini | ByteDance | Seedance 2.0 Mini — cheap text to video, 720p | B-roll, Budget | Mid |
| FLUX.3 Video | Black Forest Labs | Black Forest Labs FLUX.3 — text to video at 1080p | B-roll | High |

### Video models — animate a still (14)

Start from an image you control — your own product photography, or a still you just generated — and animate it. The reliable way to keep a product on-model in motion.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| Grok Imagine i2v | xAI | xAI Grok Imagine — animate a still image | B-roll | Low |
| MiniMax H3 i2v | MiniMax | MiniMax H3 — animate a still image at 2K | B-roll | Mid |
| Happyhorse i2v | Alibaba | Alibaba Happyhorse — animate a still image | B-roll, Budget | Mid |
| Veo 3.1 i2v | Google | Google Veo 3.1 — animate a still image | Talking video, B-roll | High |
| Veo 3.1 Fast i2v | Google | Veo 3.1 Fast — animate a still image | Talking video, B-roll | Mid |
| Veo 3.1 Lite i2v | Google | Veo 3.1 Lite — animate a still image, cheap | B-roll, Budget | Low |
| Seedance 2.0 i2v | ByteDance | Seedance 2.0 — animate a still image, up to 1080p | B-roll | High |
| Seedance 2.0 Fast i2v | ByteDance | Seedance 2.0 Fast — animate a still image, 720p cap | B-roll, Budget | High |
| Runway Gen-4.5 i2v | Runway | Runway Gen-4.5 — animate a still image at 720p | B-roll | Mid |
| Seedance 2.0 Mini i2v | ByteDance | Seedance 2.0 Mini — animate a still image at 720p | B-roll, Budget | Mid |
| FLUX.3 Video i2v | Black Forest Labs | FLUX.3 — animate a still image at 1080p | B-roll | High |
| Kling 3.0 Pro i2v | Kling | Kling 3.0 — animate a still image at 1080p | B-roll | Mid |
| Kling 3.0 Standard i2v | Kling | Kling 3.0 — image to video, 720p (cheaper, faster) | B-roll, Budget | Mid |
| Kling 3.0 4K i2v | Kling | Kling 3.0 — image to video at 4K (premium) | B-roll | High |

### Music (1)

Original, licence-clean backing tracks, generated from a text description.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| ElevenLabs Music | ElevenLabs | Generate songs from text — natural, high quality | — | High |

### Voiceover (1)

Spoken narration for reels and explainers, in multiple languages.

| Model | Lab | What it does | Best for | Cost band |
| --- | --- | --- | --- | --- |
| ElevenLabs Dialogue v3 | ElevenLabs | ElevenLabs Dialogue v3 — natural multi-language TTS | — | High |

## Model disclosure across the category

Buffer, Hootsuite, Later and Ocoya each run AI features without naming the underlying model, and none allow the user to choose one. Later additionally meters AI by credits (5–100 per month by tier). Autoadify names all 76 models and includes model choice in every paid plan rather than metering it.

---

_Source: https://autoadify.com/ai-models_
