Dynamic Model Registry
A Living Catalogue, Not a Hard-Coded List.
Every card on this page is rendered from a central registry — provider, model ID, version, modalities, pricing, latency, quality rating, status, endpoints and sync dates. When a provider ships a new model, it appears here without touching the interface.
Add providers & models
Register a new provider or add a model manually with full metadata.
Auto-sync catalogues
Where supported, available models are retrieved automatically from each provider’s API.
Edit model information
Update descriptions, capabilities, scores and endpoints at any time.
Update pricing
Keep input, output and per-generation costs current across providers.
Change rankings
Adjust Recommended, Fastest, Highest Quality and Lowest Cost per modality.
Brand approval
Mark models as brand approved and restrict models by workspace.
Deprecate & disable
Disable a model or mark it deprecated without breaking existing workflows.
Test API connections
Verify credentials and latency for every connected endpoint.
Comparison Mode
Compare Up to Four Models Side by Side.
Select any four models from the registry to compare quality, speed, cost, capabilities and API status — here shown for video generation.
Veo 3.1
Google Veo
Quality
9.6
Speed
5.5
Cost efficiency
4.5
Est. cost
$0.40 / second
Inputs
Text, images
Outputs
Video with audio
API status
Production
Best for
Cinematic video, hero films, product launches
Sora 2
OpenAI
Quality
9.5
Speed
6
Cost efficiency
5
Est. cost
$0.90 / 5s clip
Inputs
Text, images, video
Outputs
Video with audio
API status
Production
Best for
Narrative spots, social video, concept films
Ray 3
Luma AI
Quality
9.2
Speed
7
Cost efficiency
6.5
Est. cost
$0.22 / second
Inputs
Text, images, keyframes
Outputs
Video
API status
Production
Best for
Storyboards, cinematic drafts, brand films
Kling 2.5 Turbo
Kling AI
Quality
8.5
Speed
9
Cost efficiency
9
Est. cost
$0.07 / second
Inputs
Text, images
Outputs
Video
API status
Production
Best for
Social video, quick drafts, volume production
Runs your prompt through every selected model and returns the outputs side by side.
Smart Model Routing
Or Let Auto Select Route It for You.
With Auto Select on, DreamCache scores every eligible model against your task before routing — no model is universally best, so the decision is recalculated every time.
Auto Select enabled
Required modality
User objective
Desired quality
Maximum cost
Required speed
Brand restrictions
Output format
Model availability
Routing weighs required modality, your objective, desired quality, maximum cost, required speed, brand restrictions, output format, commercial-use requirements, your previous ratings, workflow history and live model availability.
Filter by modality
Claude Opus 4.5
Anthropic Claude ·
claude-opus-4-5
Production
Reasoning and Strategy
Recommended, Highest Quality, Best for Storytelling
Deep reasoning model with careful multi-step strategy work and nuanced brand-safe writing.
Best for: Strategic research, creative ideation, storytelling
9.6
Quality
6
Speed
5
Cost Eff.
Connected
View model →
Veo 3.1
Google Veo ·
veo-3-1
Production
Video Generation
Highest Quality, Best for Video, Best for Photorealism
Cinematic video with native audio, physics-accurate motion and strong prompt control.
Best for: Cinematic video, hero films, product launches
9.6
Quality
5.5
Speed
4.5
Cost Eff.
Connected
View model →
GPT-5.2
OpenAI ·
gpt-5.2
Production
Text and Language
Recommended, Highest Quality, Best for Enterprise
OpenAI's flagship general model with strong writing, planning and instruction-following across long documents.
Best for: Copywriting, long-form writing, campaign concepts
9.5
Quality
7.5
Speed
6
Cost Eff.
Connected
View model →
FLUX 2 Pro
Black Forest Labs FLUX ·
flux-2-pro
Production
Image Generation
Recommended, Highest Quality, Best for Photorealism
Photorealistic image model with exceptional detail, lighting and skin rendering.
Best for: Image campaigns, product photography, fashion imagery
9.5
Quality
7
Speed
6.5
Cost Eff.
Connected
View model →
Eleven v4
ElevenLabs ·
eleven-v4
Production
Voice and Text-to-Speech
Recommended, Highest Quality
Expressive voice synthesis with emotional direction, cloning and 70+ languages.
Best for: Voiceover, audiobooks, character voices
9.5
Quality
8
Speed
6.5
Cost Eff.
Connected
View model →
Sora 2
OpenAI ·
sora-2
Production
Video Generation
Highest Quality
OpenAI's flagship video model with synced audio, strong physics and multi-shot narrative consistency.
Best for: Narrative spots, social video, concept films
9.5
Quality
6
Speed
5
Cost Eff.
Connected
View model →
Uni-1
Luma AI ·
uni-1
Production
Multimodal Understanding
Highest Quality
Luma AI's unified multimodal model. Luma AI provides advanced image, video and multimodal capabilities that can be accessed through the DreamCache Model Hub, subject to API availability.
Best for: Cross-modal reasoning, creative direction, unified generation
9.5
Quality
8
Speed
7
Cost Eff.
Connected
View model →
Imagen 4 Ultra
Google Imagen ·
imagen-4-ultra
Production
Image Generation
Highest Quality, Best for Brands
Google's highest-fidelity image model with excellent text rendering and prompt adherence.
Best for: Hero visuals, art direction, typography-heavy images
9.4
Quality
6.5
Speed
6
Cost Eff.
Connected
View model →
GPT-5.2-Codex
OpenAI ·
gpt-5.2-codex
Production
Coding and Prototyping
Recommended, Highest Quality
Agentic coding model for building sites, prototypes and production tools from briefs.
Best for: Websites, mobile prototypes, e-commerce builds
9.4
Quality
7
Speed
6
Cost Eff.
Connected
View model →
Lyria 3
Google Lyria ·
lyria-3
Preview
Music Generation
Highest Quality, Preview, Best for Brands
High-fidelity instrumental music with fine-grained control over mood, tempo and instrumentation.
Best for: Brand scores, ad music, licensed-safe soundtracks
9.3
Quality
6.5
Speed
6
Cost Eff.
Connected
View model →
Gemini 3 Pro
Google Gemini ·
gemini-3-pro
Production
Multimodal Understanding
Recommended, Best for Enterprise
Long-context multimodal model that reads video, audio, images and documents in a single request.
Best for: Audience insight, asset analysis, research synthesis
9.2
Quality
7
Speed
6.5
Cost Eff.
Connected
View model →
FLUX.2 Kontext
Black Forest Labs FLUX ·
flux-2-kontext
Production
Image Editing
Recommended, Best for Character Consistency
Instruction-based image editing that preserves characters, products and scenes across edits.
Best for: Product swaps, character consistency, retouching
9.2
Quality
7.5
Speed
7
Cost Eff.
Connected
View model →
Ray 3
Luma AI ·
ray-3
Production
Video Generation
Recommended, Best for Storytelling
Luma's video model sharing the Uni-1 foundation that powers DreamCache — natural motion, HDR and reasoning-guided direction.
Best for: Storyboards, cinematic drafts, brand films
9.2
Quality
7
Speed
6.5
Cost Eff.
Connected
View model →
Suno v5
Suno ·
suno-v5
Production
Music Generation
Recommended
Full-song generation with vocals, structure control and radio-ready mixing.
Best for: Campaign music, jingles, soundtrack drafts
9.2
Quality
7.5
Speed
7.5
Cost Eff.
Connected
View model →
Recraft V4
Recraft ·
recraft-v4
Production
Graphic Design and Typography
Highest Quality, Best for Brands, Best for Agencies
Vector-native design model producing brand-consistent SVG, icons and layout systems.
Best for: Brand systems, icons, vector illustration
9.1
Quality
7
Speed
7
Cost Eff.
Connected
View model →
Gen-4.5 Aleph
Runway ·
gen-4-5-aleph
Production
Video Editing
Recommended, Highest Quality, Best for Agencies
In-context video editing — restyle, relight, remove and extend existing footage with prompts.
Best for: Video editing, VFX, post-production for agencies
9.1
Quality
6.5
Speed
5.5
Cost Eff.
Connected
View model →
Ideogram 3.0
Ideogram ·
ideogram-3
Production
Graphic Design and Typography
Recommended, Best for Typography
Best-in-class text rendering inside images — posters, packaging, lockups and layouts.
Best for: Typography, posters, packaging mockups
9
Quality
7.5
Speed
7.5
Cost Eff.
Connected
View model →
Avatar IV
HeyGen ·
avatar-iv
Production
Avatar and Digital Humans
Recommended, Best for Social Content
Photoreal talking avatars with natural gesture and 175+ languages from a single photo.
Best for: Presenter videos, localised spokespeople, training
9
Quality
7.5
Speed
7
Cost Eff.
Connected
View model →
GPT Image 2
OpenAI ·
gpt-image-2
Production
Image Editing
Recommended
Instruction-following image generation and editing with best-in-class text rendering and layout awareness.
Best for: Ad variations, packshots, text-in-image design
9
Quality
7
Speed
6.5
Cost Eff.
Connected
View model →
Nano Banana Pro
Google DeepMind ·
gemini-3-image
Production
Image Editing
Recommended
Gemini-powered image editing with uncanny character consistency, scene blending and conversational refinement.
Best for: Character consistency, product staging, iterative edits
9
Quality
8
Speed
7
Cost Eff.
Connected
View model →
Scribe v2
ElevenLabs ·
scribe-v2
Production
Speech-to-Text
Recommended
State-of-the-art speech-to-text with word-level timestamps, diarization and 99-language coverage.
Best for: Subtitling, interview transcription, voiceover QA
9
Quality
8.5
Speed
8
Cost Eff.
Connected
View model →
Sonar Pro
Perplexity ·
sonar-pro
Production
Research and Search
Recommended
Search-grounded answers with citations, tuned for factual research and source-backed briefs.
Best for: Strategic research, competitive analysis, fact-checking
8.8
Quality
8
Speed
7.5
Cost Eff.
Connected
View model →
DeepSeek V4
DeepSeek ·
deepseek-v4
Production
Reasoning and Strategy
Lowest Cost, New
Open reasoning model with near-frontier chain-of-thought quality at a fraction of the cost.
Best for: Analysis, planning, budget reasoning workloads
8.8
Quality
7
Speed
9.8
Cost Eff.
Connected
View model →
Nova 3
Deepgram ·
nova-3
Production
Speech-to-Text
Recommended, Fastest, Lowest Cost
Fast, accurate transcription with diarisation and keyword boosting at very low cost.
Best for: Transcription, subtitles, meeting notes
8.8
Quality
9.5
Speed
9.5
Cost Eff.
Connected
View model →
Meshy 6
Meshy ·
meshy-6
Production
3D Generation
Recommended
Text- and image-to-3D with clean topology, PBR textures and game-ready export.
Best for: 3D assets, AR experiences, product visualisation
8.8
Quality
7.5
Speed
8
Cost Eff.
Connected
View model →
Kimi K2.5
Moonshot AI Kimi ·
kimi-k2-5
Beta
Agentic and Tool-Using Models
Recommended, New
Agentic model tuned for long tool-use chains, browsing and multi-step task execution.
Best for: Automation agents, tool orchestration, workflows
8.7
Quality
7
Speed
8.5
Cost Eff.
Connected
View model →
Qwen3-Max
Alibaba Qwen ·
qwen3-max
Production
Translation and Localisation
Recommended, Lowest Cost
Multilingual model with top-tier quality across 100+ languages for localisation at scale.
Best for: Localisation, multilingual campaigns, transcreation
8.6
Quality
7.5
Speed
9
Cost Eff.
Connected
View model →
Seedream 4.0
ByteDance Seedream ·
seedream-4-0
Production
Image Generation
Fastest, Lowest Cost, Best for Social Content
Very fast image generation with strong aesthetics, tuned for social-scale content volume.
Best for: Social content, rapid iterations, thumbnails
8.6
Quality
9.5
Speed
9
Cost Eff.
Connected
View model →
Grok 4.1
xAI Grok ·
grok-4-1
Production
Research and Search
Fastest
Real-time knowledge model with live web and X data for current-events research and trend tracking.
Best for: Trend research, social listening, cultural insight
8.5
Quality
8.5
Speed
6.5
Cost Eff.
Connected
View model →
Kling 2.5 Turbo
Kling AI ·
kling-2-5-turbo
Production
Video Generation
Fastest, Best for Social Content
Fast, affordable video generation with reliable motion for social-first formats.
Best for: Social video, quick drafts, volume production
8.5
Quality
9
Speed
9
Cost Eff.
Connected
View model →
Sonic 3
Cartesia ·
sonic-3
Production
Voice and Text-to-Speech
Fastest
Ultra-low-latency streaming TTS built for real-time voice agents and live experiences.
Best for: Real-time voice, interactive experiences
8.5
Quality
10
Speed
8
Cost Eff.
Connected
View model →
Dream Machine
Luma AI ·
dream-machine-2.5
Production
Video Generation
Recommended, Fastest
Luma's Dream Machine platform API powered by Ray 3 — fast, cinematic text- and image-to-video with strong camera and physics control.
Best for: Cinematic clips, product shots, storyboard animatics
8.5
Quality
8.5
Speed
7.5
Cost Eff.
Connected
View model →
Firefly Image 5
Adobe ·
firefly-5
Production
Image Generation
Best for Enterprise, IP Indemnified
Commercially-safe image generation trained on licensed content, deeply integrated with Creative Cloud.
Best for: Brand-safe campaigns, enterprise content, print
8.5
Quality
7.5
Speed
6.5
Cost Eff.
Connected
View model →
Mistral Large 3
Mistral AI ·
mistral-large-3
Production
Text and Language
EU Data Residency
Europe's frontier language model with strong multilingual output, EU data residency and open deployment options.
Best for: EU-compliant copy, multilingual campaigns, RAG
8.5
Quality
8
Speed
7.5
Cost Eff.
Connected
View model →
Command R+ 2
Cohere ·
command-r-plus-2
Production
Embeddings and Knowledge Retrieval
Recommended, Best for Enterprise
Retrieval-optimised model with first-class embeddings and grounded RAG answers over brand knowledge.
Best for: Brand knowledge bases, retrieval, enterprise search
8.4
Quality
7.5
Speed
8
Cost Eff.
Connected
View model →
Seedance 1.5 Pro
ByteDance Seedance ·
seedance-1-5-pro
Production
Video Generation
Lowest Cost
Budget-friendly multi-shot video with solid coherence for high-volume social pipelines.
Best for: Social video at scale, variations, A/B creative
8.3
Quality
8.5
Speed
9.5
Cost Eff.
Connected
View model →
Stable Audio 2.5
Stability AI ·
stable-audio-2-5
Production
Sound Effects
Recommended, Lowest Cost
Fast sound-effect and ambience generation with precise duration control at low cost.
Best for: Sound design, UI sounds, ambient beds
8.2
Quality
8.5
Speed
9.5
Cost Eff.
Connected
View model →
Llama 4 Maverick
Meta Llama ·
llama-4-maverick
Production
Text and Language
Lowest Cost
Open-weight workhorse with strong quality-per-dollar, deployable on any inference provider.
Best for: Social copy, drafts, high-volume generation
8
Quality
8
Speed
9.5
Cost Eff.
Connected
View model →
Llama 4 on Groq
Groq ·
llama-4-scout-groq
Production
Text and Language
Fastest
Llama 4 served on Groq LPUs at extreme token throughput for real-time drafting and chat.
Best for: Real-time ideation, instant drafts, chat UX
7.5
Quality
10
Speed
9
Cost Eff.
Connected
View model →
