Models
Every model, one place.
The foundation models behind LetzAI — image, video, upscaling and text — picked by benchmark and available through one credit balance and one API.
- Sep 2026
GPT Image 2.5
OpenAIOpenAI's GPT Image 2.5. Stronger instruction-following than 2.0; different aesthetic. Precise mask-based editing up to 4K.
- Sep 2026
GPT Image 2.5 Upscaler
OpenAIAI upscaling via OpenAI GPT Image 2.5 edit with intelligent detail enhancement up to 4K resolution.
- Sep 2026
Gemini 3.8 Flash
GoogleGoogle's most intelligent workhorse — gains over 3.7 Flash on software engineering, agentic tasks, and multi-step reasoning. 1M context, native thinking (low/medium/high), Google Search grounding, and URL context. Via Vertex AI. Introductory pricing through Dec 31 2026.
- Sep 2026
Claude Fable 5.1
AnthropicAnthropic's highest-capability widely released model for long-running agents and the hardest work. Adaptive thinking; 1M context; same $10/$50 as Fable 5 with cheaper cache reads.
- Aug 2026
MiniMax H3 Max
MiniMaxExtremely fast MiniMax video — great for quick tests. Fal's post-trained H3 variant with stronger prompt adherence and aesthetics; text-to-video, image-to-video (optional last frame), and multimodal reference-to-video (up to 9 images, 3 videos, 3 audio; 12 files total), 5–15s at 480p or 768p.
- Aug 2026
Gemini Omni 1.1 Flash
GoogleGoogle's multimodal video model with native audio, now up to 4K. Handles text-to-video, image-to-video (first and last frame), reference-to-video, and prompt-driven video editing from a single endpoint.
- Aug 2026
Seedance 2.5 Video Upscale
ByteDanceUpscale a video to 1080p via Seedance 2.5 reference-to-video. Max 30s. 4K is not available on 2.5.
- Aug 2026
Seedance 2.5 Enterprise Video Upscale
ByteDanceOrg-scoped Seedance 2.5 Studio video upscale to 1080p. Max 30s. 4K is not available on 2.5.
- Aug 2026
BytePlus Video Enhancer
ByteDanceCinema-grade BytePlus MediaKit enhancement (Professional): 30+ algorithms, AI super-res, defect repair, and color. Keeps the original footage (does not regenerate).
- Aug 2026
Grok 4.6
xAIxAI's newest flagship for coding, agentic tool calling, and long-running visual work. 500k context; reasoning effort low/medium/high/xhigh (default high). Knowledge cutoff February 1, 2026.
- Aug 2026
Seedance 2.0 Video Upscale
ByteDanceUpscale a video to 1080p or 4K via Seedance 2.0 reference-to-video. Max 15s.
- Aug 2026
Seedance 2.0 Enterprise Video Upscale
ByteDanceOrg-scoped Seedance 2.0 Studio video upscale to 1080p or 4K. Max 15s.
- Aug 2026
Grok Imagine 2.0
xAIxAI's newest image model — precise instruction following, sharp typography, and iterative editing with up to 5 reference images. $0.04/image; 1K and 2K.
- Aug 2026
Seedance 2.5 Enterprise
ByteDanceByteDance's professional Seedance 2.5 Studio variant with private asset library support. Automatically uploads reference images to BytePlus virtual portrait library for enhanced character consistency. Up to 30s clips with up to 50 multimodal references (30 images / 10 videos / 10 audio) and synchronized audio. 480p, 720p, and 1080p.
- Aug 2026
Flux 3 Video
Black Forest LabsBlack Forest Labs' Flux 3 Video — generate and animate video up to 20s at HD or Full HD with synchronized audio. Supports text-to-video, image-to-video (start frame), and first+last frame keyframes. No Omni / multimodal reference stack.
- Jul 2026
MiniMax H3
MiniMaxMiniMax's frontier video model at 2K. Smart-routes text-to-video, image-to-video (optional last frame), and multimodal reference-to-video (reference images, one video, one audio) while keeping subjects consistent.
- Jul 2026
Claude Opus 5
AnthropicAnthropic's strongest Opus — step-change over 4.8 for agentic coding, long-horizon work, and professional knowledge tasks. Adaptive thinking; 1M context; near-Fable intelligence at Opus price ($5/$25).
- Jul 2026
GPT-5.6 Sol
OpenAIOpenAI's GPT-5.6 flagship for complex professional work — coding agents, long research, computer use, and multi-tool workflows. Native web search via the Responses API.
- Jul 2026
Seedream 5 Pro
ByteDanceByteDance's latest Seedream model with Chain-of-Thought reasoning, enhanced prompt following, reference consistency, and up to 4K resolution.
- Jul 2026
Seedance 2.5
ByteDanceByteDance's next-generation Seedance with up to 30-second native clips, up to 50 multimodal references (30 images / 10 videos / 10 audio), and synchronized audio. Supports text, image, first+last frame, and reference-to-video. 480p, 720p, and 1080p.
- Jul 2026
GPT Image 2 Upscaler
OpenAIAI upscaling via OpenAI GPT Image 2 edit with intelligent detail enhancement up to 4K resolution.
- Jun 2026
Claude Sonnet 5
AnthropicAnthropic's latest Sonnet — drop-in upgrade with adaptive thinking, stronger agentic coding, and 1M context. Introductory pricing through Aug 2026.
- Jun 2026
Nano Banana 2 Lite
GoogleGoogle's fastest and cheapest Gemini image model — optimized for high-volume generation where speed and cost matter most. Best for single-prompt text-to-image; not optimized for multiple reference inputs or multi-turn editing.
- Jun 2026
Nano Banana 2 Upscaler
GoogleFaster AI upscaling using Google Gemini 3.1 Flash with intelligent detail enhancement up to 4K resolution.
- May 2026
Pruna Upscaler
Pruna AIFast AI upscaling via Pruna P-Image-Upscale with optional detail and realism enhancement up to 128 megapixels.
- May 2026
Beeble SwitchX
BeebleBeeble's video-to-video model that swaps the background or scene in a clip while preserving the original subject, motion, framing and lighting. Not for character or outfit changes. Driven by an optional reference image plus prompt.
- May 2026
Seedance 2.0 Enterprise
ByteDanceByteDance's professional Seedance 2.0 Studio variant with private asset library support. Supports up to 4K (10-bit color), with text-to-video, image-to-video, and multimodal reference-to-video (up to 9 reference images, 3 reference videos, 3 reference audios) and synchronized audio.
- Apr 2026
GPT-5.5
OpenAIOpenAI's frontier model for complex reasoning, coding, and agentic work. Native web search via the Responses API.
- Apr 2026
GPT Image 2
OpenAIOpenAI's GPT Image 2 — the look people actually want. Strong typography, fine detail, and precise mask-based editing up to 4K.
- Apr 2026
Seedance 2.0
ByteDanceByteDance's most advanced video model with cinematic output, native audio, real-world physics, and director-level camera control. Supports up to 4K (10-bit color). Supports text, image, audio, and video reference inputs (up to 9 reference images, 3 reference videos, 3 reference audios) with first+last frame control.
- Apr 2026
WAN 2.7 Image Pro
AlibabaAlibaba's professional image generation model supporting text-to-image, image editing, and multi-reference generation with up to 4K high-definition output. Supports thinking mode for improved quality.
- Feb 2026
Nano Banana 2
GoogleGoogle's Gemini 3.1 Flash image generation model — fast, cost-effective generation and editing with up to 4K resolution, advanced text rendering, and Google Search grounding.
- Feb 2026
Nano Banana 2
GoogleGoogle Gemini 3.1 Flash via InferenceSH — fast, cost-effective image generation and editing with up to 4K resolution and Google Search grounding.
- Feb 2026
Kling O3 Pro Edit
KuaishouKuaishou's unified multimodal video editor with 7-in-1 capabilities including object removal, background swapping, style changes, and character consistency.
- Feb 2026
Kling V3
KuaishouKuaishou's flagship video model with native multilingual audio, up to 15 seconds, multi-shot storytelling, and reference image support.
- Nov 2025
Flux 2
Black Forest LabsBlack Forest Labs' 32B parameter model with multi-reference image support, 4MP editing, improved text rendering, and enhanced photorealism.
- Nov 2025
Nano Banana Pro
GoogleGoogle's state-of-the-art image generation model with improved text rendering, multi-turn editing, and professional-grade controls over lighting, camera, and composition.
- Nov 2025
Nano Banana Pro
GoogleGoogle Gemini 3 Pro via InferenceSH with improved retries and per-organization API key support.
- Nov 2025
Nano Banana Pro Upscaler
GoogleAI-powered upscaling using Google Gemini 3 Pro with intelligent detail enhancement up to 4K resolution.
- Oct 2025
Veo 3.1 Extend Video
GoogleGoogle's video extension model capable of extending existing videos up to 148 seconds with maintained consistency and style.
How we choose models
Tested by creative professionals.
We run every candidate through our own benchmark suite and keep only the models that hold up in real creative work.
See LetzBenchPricing
One credit balance for all of them.
Every model draws from the same credits. Compare per-generation costs and pick a plan.
See credit costs
Every model, one API.
Integrate LetzAI into your apps, workflows and agents with the same REST API the product runs on.
- 01
Get your API key
Subscribe to any paid plan, then find your key on the subscription page.
- 02
Generate an image
Send a POST request to create an image with any model.
POST https://api.letz.ai/images - 03
Retrieve your result
Poll the image endpoint with the returned ID until it's ready.
GET https://api.letz.ai/images/:id
Inference partners.
Model availability and pricing may change. Check the pricing page for current rates.