Back

AI model overview

See which platforms expose which models, what they cost through the main public route, and where the wrapper/platform choice matters more than the model name.

Updated August 2026: model names and prices moved fast

Sora is now legacy context, OpenAI image work should be read as GPT Image rather than DALL-E, Veo is the clearest native-audio video lane, and Ideogram plus Flux now deserve more weight in image workflows. Pricing is shown by access route because the same model can be sold as API usage, plan credits, bundled subscription access, or self-hosted infrastructure.

Start with the live buyer guides feeding this page

If you arrived here from a model question, anchor it in a real buying workflow first. For the live analyst branch, keep the order literal: use /compare first when the packaged shortlist is down to real finalists, keep the PDF-analysis guide as the active feeder that pressure-tests those document-heavy finalists, use the broader data-analysis guide only as the wider support branch, and open /stack only after that when the blocker is which model access route or wrapper depth actually differs. This branch should stay separate from the colder note-taking and research detour on /recommend.

Stable Diffusion 3.5
Image
by Stability AI
9 platforms

Fully open-source image generation model offering maximum customization and fine-tuning capabilities.

4K · open weights
Self-host at infra cost; Stability API credits are $0.01 each
open-source
Self-host or Stability API via Stability AI / self-host

Stable Diffusion 3.5 pricing depends on whether you self-host or use a managed endpoint.

Source · verified 2026-08-06
FLUX.1 Kontext
Image
by Black Forest Labs
8 platforms

Black Forest Labs' image generation and editing model family for text-to-image, image-to-image, and instruction-based edits with strong character and style consistency.

4K · open weights
$0.04 per Kontext Pro image edit; Kontext Max text-to-image $0.08/image
per-image
fal FLUX Kontext Pro via fal / Black Forest Labs

BFL also sells service credits directly; wrapper prices vary by provider and resolution.

Source · verified 2026-08-06
Llama 4 Maverick
Model
by Meta
5 platforms

Llama 4 Maverick is the previous-generation Meta workhorse and still one of the cheapest capable open-weight models on hosted APIs. It remains a common self-hosting baseline.

open weights
≈$0.24 / 1M input · ≈$0.97 / 1M output
per-token
Hosted inference providers via Meta

Weights are free; the quoted rate is a representative hosted price and varies by inference provider.

Source · verified 2026-08-13
Llama 5
Model
by Meta
5 platforms

Llama 5 is Meta's 600B-parameter mixture-of-experts flagship with native video and audio input. Its 5M-token context window is the largest of any publicly available model, open or closed, and the weights are downloadable.

open weights
Free to self-host · hosted rates vary by provider
open-source
Open weights / hosted inference providers via Meta

Released under the Llama 5 Community License, which is not OSI-approved. Review the commercial-use terms before deploying.

Source · verified 2026-08-13
DeepSeek V4
Model
by DeepSeek
4 platforms

DeepSeek V4 pairs a 1M-token context window with open weights and unusually aggressive API pricing. It is the standard cost-sensitive pick when you want frontier-adjacent quality without frontier pricing.

open weights
≈$0.44 / 1M input · ≈$0.87 / 1M output
per-token
DeepSeek API via DeepSeek

1M context under an MIT license. Cache-hit input bills substantially lower than the cache-miss rate quoted here.

Source · verified 2026-08-13
Kimi K2.6
Model
by Moonshot AI
4 platforms

Kimi K2.6 is the multimodal agent tier, combining text, image, and video input with tool use at roughly a third of K3's price. Context caps at 256K, well below the K3 flagship.

open weights
≈$0.95 / 1M input · ≈$4.00 / 1M output
per-token
Moonshot API via Moonshot AI

Roughly a third of K3's cost, with multimodal input but a 256K context ceiling.

Source · verified 2026-08-13
Kimi K3
Model
by Moonshot AI
4 platforms

Kimi K3 is Moonshot's 2.8-trillion-parameter flagship and the largest open-weight model in the comparison set. It is priced like a frontier closed model rather than like the cheaper open-weight lane.

open weights
$3.00 / 1M input · $15.00 / 1M output
per-token
Moonshot API via Moonshot AI

Priced like a closed frontier model despite shipping open weights — the largest open-weight model here at 2.8T parameters.

Source · verified 2026-08-13
Mistral Small 4
Model
by Mistral AI
4 platforms

Mistral Small 4 is the Apache 2.0 workhorse of the Mistral line — small enough to self-host on modest hardware, cheap enough to run at volume, and unrestricted for commercial use.

open weights
$0.15 / 1M input · $0.60 / 1M output
per-token
Mistral La Plateforme via Mistral AI

Apache 2.0 weights with no commercial restrictions, and small enough to self-host on modest hardware.

Source · verified 2026-08-13
Qwen3.5 Plus
Model
by Alibaba
4 platforms

Qwen3.5 Plus is the mid-tier Qwen model, trading some of the Flash line's cost advantage for stronger reasoning while keeping the 1M-token context window and Apache 2.0 license.

open weights
$0.40 / 1M input · $2.40 / 1M output
per-token
Alibaba Cloud Model Studio via Alibaba

Mid-tier Qwen model, still Apache 2.0 and still a 1M-token context window.

Source · verified 2026-08-13
Qwen3.7 Flash
Model
by Alibaba
4 platforms

Qwen3.7 Flash is the cheapest 1M-context API in the comparison set by a wide margin. Apache 2.0 licensing means the weights can also be self-hosted without a bespoke commercial agreement.

open weights
$0.03 / 1M input · $0.13 / 1M output
per-token
Alibaba Cloud Model Studio via Alibaba

Cheapest per-token rate in this comparison set. Apache 2.0 weights also allow self-hosting.

Source · verified 2026-08-13
DeepSeek V4 Flash
Model
by DeepSeek
3 platforms

DeepSeek V4 Flash is the low-latency V4 variant and one of the cheapest 1M-context models available anywhere, at a fraction of the flagship rates from OpenAI, Anthropic, and Google.

open weights
$0.14 / 1M input · $0.28 / 1M output
per-token
DeepSeek API via DeepSeek

Quoted rate is for cache-miss input. One of the cheapest 1M-context APIs available.

Source · verified 2026-08-13
GLM-5.2
Model
by Zhipu AI
3 platforms

GLM-5.2 is the long-context coding specialist of the open-weight field, consistently ranked against DeepSeek V4 and Qwen on repository-scale code tasks.

open weights
≈$1.40 / 1M input · ≈$4.40 / 1M output
per-token
Zhipu Open Platform via Zhipu AI

Positioned as the long-context coding pick among open-weight models. MIT-licensed weights.

Source · verified 2026-08-13
GPT-5.6 Sol
Model
by OpenAI
3 platforms

GPT-5.6 Sol is OpenAI's flagship reasoning tier, the top of the three-tier GPT-5.6 family. It carries the largest published context window of any major closed model at 1.05M tokens.

$5.00 / 1M input · $30.00 / 1M output
per-token
OpenAI API via OpenAI

Flagship reasoning tier. Output pricing is the highest of the mainstream closed flagships at this input rate.

Source · verified 2026-08-13
Mistral Large 3
Model
by Mistral AI
3 platforms

Mistral Large 3 is the flagship of the European open-weight route, with EU data residency and on-prem deployment options that matter more to regulated buyers than raw benchmark position.

open weights
$0.50 / 1M input · $1.50 / 1M output
per-token
Mistral La Plateforme via Mistral AI

EU data residency and on-prem deployment are available, which is usually the deciding factor over raw price.

Source · verified 2026-08-13
Claude Opus 5
Model
by Anthropic
2 platforms

Claude Opus 5 is Anthropic's model for complex agentic coding and long-horizon enterprise work. Thinking is on by default, effort is tunable from low through max, and it ships a 1M-token context window at standard pricing.

$5.00 / 1M input · $25.00 / 1M output
per-token
Claude API via Anthropic

1M context at standard pricing with no long-context premium. Prompt caching reads bill at roughly 0.1x.

Source · verified 2026-08-13
Gemini 3.1 Pro
Model
by Google
2 platforms

Gemini 3.1 Pro is Google's flagship reasoning model. Pricing is tiered by prompt length: rates roughly double once a request passes 200K tokens, which matters more here than on flat-rate competitors.

$2.00 / 1M input · $12.00 / 1M output (≤200K context)
per-token
Gemini API / Vertex AI via Google

Above 200K context the rate steps to $4.00 input / $18.00 output. Context caching cuts repeat-prompt cost by up to ~90%.

Source · verified 2026-08-13
GPT Image 2
Image
by OpenAI
2 platforms

OpenAI's current state-of-the-art image generation and editing model, replacing GPT Image 1 as the safer default for new OpenAI image workflows.

flexible
$5/M text input, $8/M image input, $30/M image output tokens
per-token
OpenAI image API via OpenAI

Batch pricing is lower; image cost depends on generated token count, quality, and size.

Source · verified 2026-08-06
GPT-5.6 Terra
Model
by OpenAI
2 platforms

GPT-5.6 Terra is the balanced production tier of the GPT-5.6 family, priced between Sol and Luna and intended as the general-purpose default for most API workloads.

$2.00 / 1M input · $12.00 / 1M output
per-token
OpenAI API via OpenAI

Balanced production tier, priced level with Gemini 3.1 Pro's sub-200K rate.

Source · verified 2026-08-13
Adobe Firefly Image 3
Image
by Adobe
1 platform

Adobe's commercially oriented image model for brand-safe generation and Creative Cloud workflows.

4K
Free credits; paid Firefly and Creative Cloud plans bundle generative credits
credits
Firefly / Creative Cloud generative credits via Adobe

Commercial buyers should compare Firefly credits inside the specific Adobe plan they already own.

Source · verified 2026-08-06
Adobe Firefly Video Model
Video
by Adobe
1 platform

Adobe's commercially oriented video generation model for Firefly text-to-video and image-to-video workflows, especially useful when Creative Cloud integration and IP-safe production controls matter.

5s max · 1080p
Firefly and Creative Cloud plans bundle generative credits for video
credits
Firefly / Creative Cloud generative credits via Adobe

Commercial buyers should compare included credits inside the Adobe plan they already own.

Source · verified 2026-08-13
AIVA Opus
Music
by AIVA
1 platform

AI composer specializing in classical, cinematic, and orchestral music with MIDI export.

600s max
Free personal tier; Standard about $11/mo; Pro about $33/mo
subscription
AIVA plans via AIVA

AIVA sells composition/download rights by plan rather than per-token generation.

Source · verified 2026-08-06
Claude Fable 5
Model
by Anthropic
1 platform

Claude Fable 5 is Anthropic's most capable widely released model, priced above the Opus tier and aimed at the most demanding reasoning and long-horizon agentic work. Thinking is always on and cannot be disabled.

$10.00 / 1M input · $50.00 / 1M output
per-token
Claude API via Anthropic

Priced above the Opus tier. Requires 30-day data retention and is not available under zero-data-retention terms.

Source · verified 2026-08-13
Claude Haiku 4.5
Model
by Anthropic
1 platform

Claude Haiku 4.5 is Anthropic's fastest and cheapest model, aimed at classification, extraction, and high-volume simple tasks. It is the only current Claude model that caps below 1M context.

$1.00 / 1M input · $5.00 / 1M output
per-token
Claude API via Anthropic

Cheapest Claude tier. Context caps at 200K rather than the 1M offered by the Opus and Sonnet tiers.

Source · verified 2026-08-13
Claude Sonnet 5
Model
by Anthropic
1 platform

Claude Sonnet 5 is the balanced Anthropic tier, reaching near-Opus quality on coding and agentic work at roughly half the price. Adaptive thinking is on by default and the context window matches the Opus tier at 1M tokens.

$3.00 / 1M input · $15.00 / 1M output
per-token
Claude API via Anthropic

Introductory rate of $2.00 / $10.00 applies through 2026-08-31, after which standard pricing takes effect.

Source · verified 2026-08-13
ElevenLabs v3
Voice
by ElevenLabs
1 platform

State-of-the-art text-to-speech and voice cloning model with 29 language support and emotional nuance.

$0.10 per 1,000 characters for Multilingual v2/v3 API TTS
per-token
ElevenAPI via ElevenLabs

Website generation uses credits; Eleven v3 is 1 credit per character in the app.

Source · verified 2026-08-06
FLUX 3 Video
Video
by Black Forest Labs
1 platform

FLUX 3 Video is Black Forest Labs' multimodal video, audio, image, and action-prediction model line, adding a video-generation lane to a brand previously tracked on CompareGen only for images.

20s max · native audio
Route-specific model/API pricing
custom
BFL or hosted API route via Black Forest Labs / hosted providers

Keep separate from Flux image pricing because video and audio generation costs will not normalize to per-image rates.

Source · verified 2026-08-13
Gemini 2.5 Pro
Model
by Google
1 platform

Gemini 2.5 Pro is the previous-generation Google flagship, still available and still competitive on coding. It remains the cheapest Pro-tier entry point in the Gemini family.

$1.25 / 1M input · $10.00 / 1M output (≤200K context)
per-token
Gemini API / Vertex AI via Google

Above 200K context the rate steps to $2.50 input / $15.00 output.

Source · verified 2026-08-13
Gemini 3.6 Flash
Model
by Google
1 platform

Gemini 3.6 Flash is the speed-first Gemini tier, positioned as the highest-throughput option in the family before 3.7 arrived. Pricing is promotional through the end of 2026.

$0.75 / 1M input · $3.75 / 1M output
per-token
Gemini API / Vertex AI via Google

Promotional pricing through 2026-12-31. Some third-party trackers still list the earlier $1.50 / $7.50 rate.

Source · verified 2026-08-13
Gemini 3.7 Flash
Model
by Google
1 platform

Gemini 3.7 Flash is Google's most capable Flash model, tuned for agentic workflows. Its listed rate is promotional through December 31, 2026 and increases in January 2027.

$0.75 / 1M input · $3.75 / 1M output
per-token
Gemini API / Vertex AI via Google

Promotional pricing through 2026-12-31; rates increase 2027-01-01. Budget accordingly on annual contracts.

Source · verified 2026-08-13
Gemini Omni Flash
Video
by Google
1 platform

Google's Gemini Omni Flash is a top current video-generation model on independent text-to-video, image-to-video, and video-editing leaderboards, and should be compared separately from the Veo product route.

native audio
Google route-specific pricing; compare separately from Veo plans
custom
Google model route via Google

Gemini Omni Flash is tracked separately because benchmark naming and access route are distinct from the Veo product page.

Source · verified 2026-08-13
Gen-4
Video
by Runway
1 platform

Runway's fourth-generation image-to-video model offering cinematic motion quality, prompt adherence, and visual fidelity for 5- or 10-second generations.

10s max · 4K
12 credits/sec; API credits $0.01 each
credits
Runway API / credit plans via Runway

A 5-second Gen-4 generation is 60 credits; web plans bundle credits.

Source · verified 2026-08-06
Gen-4.5
Video
by Runway
1 platform

Runway's frontier video model for stronger motion quality, visual fidelity, and prompt adherence. Treat it as the premium Runway lane rather than a separate workflow category.

10s max · 4K
12 credits/sec; API credits $0.01 each
credits
Runway API / credit plans via Runway

Runway's Max plan also frames 9,500 credits as 791s of Gen-4.5 or Gen-4.

Source · verified 2026-08-06
GPT-5.4
Model
by OpenAI
1 platform

GPT-5.4 is the previous-generation OpenAI mid-tier, still widely deployed and still available on the API. Input above 272K tokens is billed at double the standard rate.

$2.50 / 1M input · $15.00 / 1M output
per-token
OpenAI API via OpenAI

Input above 272K tokens bills at 2x. Cached input receives a 75-90% discount.

Source · verified 2026-08-13
GPT-5.6 Luna
Model
by OpenAI
1 platform

GPT-5.6 Luna is the cost-optimized tier of the GPT-5.6 family, competing directly with Gemini Flash and the cheaper open-weight APIs on high-volume work.

$0.20 / 1M input · $1.20 / 1M output
per-token
OpenAI API via OpenAI

Cost-optimized tier competing with Gemini Flash and the hosted open-weight APIs.

Source · verified 2026-08-13
Grok 4.3
Model
by xAI
1 platform

Grok 4.3 is xAI's value tier and its widest-context option at 1M tokens. Output pricing is notably cheaper than the 4.5/4.6 flagship line, which makes it the better fit for output-heavy work.

$1.25 / 1M input · $2.50 / 1M output
per-token
xAI API via xAI

Cheapest xAI general-purpose tier and the only 1M-context option in the family.

Source · verified 2026-08-13
Grok 4.6
Model
by xAI
1 platform

Grok 4.6 is xAI's current flagship, aimed at long-running agents, coding, and research. Its 500K context is smaller than the Grok 4.3 line, and prompts past 200K tokens bill the entire request at double rate.

$2.00 / 1M input · $6.00 / 1M output (<200K prompt)
per-token
xAI API via xAI

Once a prompt reaches 200K tokens the entire request bills at $4.00 input / $12.00 output. Cached input is $0.50 / 1M.

Source · verified 2026-08-13
Grok Imagine Video 1.5
Video
by xAI
1 platform

xAI's Grok Imagine Video 1.5 is the current Grok video lane for text or image to video generation, with strong leaderboard performance in image-to-video and native-audio comparisons.

1080p · native audio
Bundled with Grok/X subscription access; no stable standalone per-video tariff
subscription
Grok/X subscription via xAI

Compare as subscription access unless xAI publishes a clean standalone video API price.

Source · verified 2026-08-13
Hailuo (MiniMax)
Video
by MiniMax
1 platform

MiniMax's Hailuo video model focused on lifelike motion and emotion, with text and image to video generation.

Credit/subscription pricing; Standard listed at $6.99/mo on Hailuo guides
credits
Hailuo plans via MiniMax / Hailuo

Hailuo publishes many plan and model guides; verify the in-app checkout for exact current credits.

Source · verified 2026-08-06
Ideogram 4.0
Image
by Ideogram
1 platform

Ideogram's current visual intelligence model for photorealistic images, legible text, precise style control, editing, API workflows, MCP, and open-weight deployment.

4K · open weights
Free + paid credit plans; API pricing lives on Ideogram's live pricing page
credits
Ideogram plans/API via Ideogram

Ideogram docs explicitly route current plan and API prices to the live pricing page.

Source · verified 2026-08-06
Imagen 4
Image
by Google
1 platform

Google's latest text-to-image model delivering stunning high-quality visuals with excellent precision and realism. Best-in-class for professional integration.

4K
Gemini image output examples: about $0.067-$0.24/image by model/resolution
per-token
Vertex AI image models via Google

Google's current public image table prices image output by tokens and resolution.

Source · verified 2026-08-06
Kling 3.0
Video
by Kuaishou
1 platform

Kuaishou's current Kling 3.0 model series improves consistency, photorealism, 3-15 second short-form output, native multilingual audio, and newer reference-control workflows.

15s max · 4K · native audio
Standard video example: 9 credits/sec
credits
Kling web plans via Kling AI

Credit cost changes by model, mode, duration, and quality; route budgets through the live credit table.

Source · verified 2026-08-06
Leonardo Phoenix
Image
by Leonardo AI
1 platform

Leonardo's foundational image model for prompt adherence, coherent text rendering, and iterative brand or concept-art workflows.

Free daily tokens; paid plans from about $12/mo
credits
Leonardo token plans via Leonardo.Ai

Phoenix pricing is effective token cost inside Leonardo rather than a simple public per-image rate.

Source · verified 2026-08-06
LTX-2.5
Video
by Lightricks
1 platform

LTX-2.5 is the current Lightricks open video model for synchronized audio-video generation, multi-shot scenes, footage editing, and local or API-backed production workflows.

4K · native audio · open weights
Open weights; pay infra or hosted provider usage
open-source
Self-host or hosted API via Lightricks / hosted providers

LTX is a model-stack lane: cost depends on local hardware, cloud GPU, or the wrapper provider.

Source · verified 2026-08-13
Luma Dream Machine 2.0
Video
by Luma AI
1 platform

Luma's multimodal creative model for generating images and videos from text or reference images in one workflow.

Lite $9.99/mo; Plus $29.99/mo; Unlimited $94.99/mo web plans
credits
Dream Machine plans via Luma AI

Use the same Luma plan budget as Ray unless a specific API route is selected.

Source · verified 2026-08-06
Midjourney v7
Image
by Midjourney
1 platform

Midjourney's V7 image model, released in April 2025, with stronger prompt precision, richer texture detail, Draft Mode, and Omni Reference for consistent characters and objects.

4K
$10/$30/$60/$120 per month by Basic/Standard/Pro/Mega plan
subscription
Midjourney subscription via Midjourney

Midjourney prices access as GPU time and relaxed/fast modes, not per image.

Source · verified 2026-08-06
Midjourney Video V1
Video
by Midjourney
1 platform

Midjourney Video V1 turns still images into 5-second image-to-video clips with optional motion prompting, making Midjourney relevant to creator-video comparisons beyond static image generation.

21s max
Included in Midjourney plans, consuming plan generation capacity
subscription
Midjourney subscription via Midjourney

Video is image-to-video inside Midjourney's subscription workflow rather than a separate public API.

Source · verified 2026-08-13
MiniMax H3
Video
by MiniMax
1 platform

MiniMax H3 is an open-weights omni-modal video model that reads text, images, video, and audio as unified context and generates up to 15-second 2K clips with native stereo audio.

15s max · 2K · native audio · open weights
Credit/subscription pricing; public H3 estimates around $0.073-$0.120/sec
credits
Hailuo plans via MiniMax / Hailuo

Use this as a directional MiniMax H3/Hailuo route, not a guaranteed API tariff.

Source · verified 2026-08-06
Murf AI Voice
Voice
by Murf
1 platform

Professional voiceover model with 120+ AI voices across 20 languages for e-learning and corporate content.

Free trial minutes; Basic about $23/mo; Pro about $39/mo
per-minute
Murf plans via Murf

Murf budgets are best compared as included voiceover minutes per seat.

Source · verified 2026-08-06
Muse Video
Videoannounced
by Meta
1 platform

Meta's Muse Video is the media-generation sibling to Muse Image, previewed with exceptional visual fidelity and native audio support. Track it separately from Muse Spark, which is the reasoning and agent model.

native audio
No stable public buyer route yet
unavailable
Announced model via Meta

Track Muse Video as model-stack context and keep it out of default buyer picks until access is clearer.

Source · verified 2026-08-13
Pika 2.5
Video
by Pika Labs
1 platform

Pika's idea-to-video model focused on creative expression and quick content generation for social media.

10s max · 1080p
Free test tier; paid creator plans are credit-based
subscription
Pika web plans via Pika

Public per-second API pricing is not cleanly exposed, so compare effective plan credits.

Source · verified 2026-08-06
Ray3.14
Video
by Luma AI
1 platform

World's first reasoning video model that can think, plan, and create studio-grade content. Native 1080p HD, 4x faster generation, and 3x cheaper per-second pricing.

120s max · 1080p · native audio
Lite $9.99/mo; Plus $29.99/mo; Unlimited $94.99/mo web plans
credits
Dream Machine plans via Luma AI

Luma bundles Ray access into plan credits; API pricing is exposed through Luma API billing.

Source · verified 2026-08-06
Recraft V3
Image
by Recraft
1 platform

Recraft V3 is a design-focused image generation model for brand visuals, illustrations, and editable creative assets.

Free + paid credit plans; API priced by current Recraft billing route
credits
Recraft plans/API via Recraft

Design workflows should compare included generations, export rights, and team seats.

Source · verified 2026-08-06
Resemble V3
Voice
by Resemble AI
1 platform

Real-time voice cloning model with cross-language localization and emotion control.

Free trial minutes; paid voice plans from about $29/mo
per-minute
Resemble plans/API via Resemble AI

Resemble pricing depends on voice cloning, localization, API, and minute bundle.

Source · verified 2026-08-06
Reve Image
Image
by Reve
1 platform

Reve's image model is strongest when prompt adherence and detailed composition accuracy matter more than ecosystem breadth.

Free tier; Pro about $15/mo; Unlimited about $39/mo
subscription
Reve plans via Reve

Treat as app-plan pricing until Reve exposes a stable public per-image API tariff.

Source · verified 2026-08-06
Seedance 2.5
Video
by ByteDance
1 platform

ByteDance Seed's current Seedance model is built for 30-second audio-video storytelling, precise reference control, powerful editing, and multimodal text, image, audio, and video inputs.

30s max · 1080p · native audio
Route-specific API or partner pricing; verify the live buyer route before budgeting
custom
Seed / partner API route via ByteDance Seed

Seedance access and pricing are less standardized than Runway or Vertex, so CompareGen should avoid a fake normalized price.

Source · verified 2026-08-13
Sora 2 (legacy)
Videodeprecated
by OpenAI
1 platform

OpenAI's deprecated video and audio generation model. It still matters historically for narrative coherence and synchronized audio, but it is no longer a safe default for new production video workflows.

25s max · 1080p · native audio
$0.10/sec Sora 2 720p; $0.30-$0.70/sec Sora 2 Pro by resolution
per-second
OpenAI video API via OpenAI

Keep as legacy context on CompareGen even though OpenAI still lists API prices.

Source · verified 2026-08-06
Suno v5
Music
by Suno
1 platform

Latest Suno model for AI music generation with improved vocal quality and longer song support.

240s max
Free 50 credits/day; Pro $8/mo yearly with 2,500 credits; Premier $24/mo yearly
credits
Suno plans via Suno

Suno sells song generation as subscription credits, not a direct per-token API price.

Source · verified 2026-08-06
Udio 2.0
Music
by Udio
1 platform

Professional-grade AI music model producing near-indistinguishable audio quality with superior instrument separation.

300s max
Standard $10/mo; Pro $30/mo; paid accounts use monthly credit pools
credits
Udio plans via Udio

Udio credits are consumed per generation set and plan quota rather than a single model tariff.

Source · verified 2026-08-06
Veo 3.1
Video
by Google
1 platform

Google's flagship video model line with native audio, stronger prompt adherence, richer audiovisual quality, and 8-second 720p/1080p outputs in Flow, Gemini, and Vertex/AI Studio routes.

8s max · 1080p · native audio
$0.20/sec video, $0.40/sec video+audio; Fast from $0.08/sec
per-second
Vertex AI / Agent Platform via Google

Veo 3.1 pricing varies by Fast/Lite tier, 720p/1080p/4K, and whether audio is generated.

Source · verified 2026-08-06
Use cases:Product Demos
Wan 2.7
Video
by Alibaba
1 platform

Alibaba's Wan video family is the important open video-generation lane for teams that want text-to-video and image-to-video workflows with local or API-hosted access instead of a closed creative app.

720p · open weights
Open model; pay infra or hosted provider usage
open-source
Self-host or hosted API via Alibaba / hosted providers

Wan cost depends on whether the buyer self-hosts or uses fal/Replicate/other hosted routes.

Source · verified 2026-08-13

Go back down-funnel from the model layer

Once you understand the stack, jump back into the workflow-first pages that actually turn model knowledge into platform decisions.

Explore Related Categories