Our AI Models
From concept art to full cinematic video, Layer brings together the industry's leading AI models in one place.
Seedance 2
Professional-grade video model with cinematic quality, up to 4K output, and synchronized audio generation.
Seedance 2 Reference
Generate videos guided by reference images, videos, and audio with precise style and character control.
Seedance 2 Fast Reference
Fast reference-guided video generation with multi-modal inputs and synchronized audio.
Seedance 2 Fast
Fast, cost-effective variant of Seedance 2 with 720p output and synchronized audio generation.
GPT Image 2
OpenAI's latest image model with stronger text rendering, UI generation, and photorealism. Native output up to 4K with three quality tiers.
Grok Imagine Video 1.5
xAI's Grok Imagine 1.5 image-to-video model, animating images into 480p or 720p clips.
PixVerse v6
Latest PixVerse video model with improved quality and audio generation, supporting up to 1080p resolution.
Grok Imagine Video
xAI's video generation model capable of creating high-quality 720p video from text and images.
Grok Imagine Video Edit
xAI's video editing model for modifying existing videos using text prompts.
Happy Horse 1.1
Alibaba's text-to-video and image-to-video model generating 720p or 1080p clips with native audio and multilingual lip-sync.
Happy Horse 1.0 Edit
Alibaba's video editing model for modifying existing videos with text prompts and optional reference images.
Happy Horse 1.0
Alibaba's text-to-video and image-to-video model generating 720p or 1080p clips with native audio and multilingual lip-sync.
Happy Horse 1.0 Reference
Reference-guided video generation with up to 9 reference images for consistent characters and style.
Kling O1 Edit
Edit videos guided by a prompt or images.
Kling O1 Reference
Generate new videos guided by prompts, images or videos.
Kling O1
Generate new videos from first and last frame images.
Kling v2.5 Turbo Pro
High-speed video model. Produces rapid, professional-quality 1080p video for fast-paced content.
Kling V3 Pro
Generate high-quality videos with advanced control. Pro tier with negative prompts.
Kling V3 4K
Generate native 4K videos directly from prompts or images. Cinema-grade output in one step.
Kling O3 Pro Edit
Edit videos guided by prompts and reference elements. Pro tier.
Kling O3 Pro
Generate high-quality videos with prompts, images, or reference elements. Pro tier for premium quality.
Kling O3 4K
Generate native 4K videos with Kling's Omni 3 model. Cinema-grade output in one step.
Kling O3 4K Reference
Generate native 4K videos guided by reference images and elements. Cinema-grade output.
PixVerse v5.5 Fast
Speed-optimized PixVerse v5.5 variant for rapid video generation at up to 720p resolution.
PixVerse v5.5
Latest PixVerse video model with enhanced quality and audio generation, supporting up to 1080p resolution.
Veo 3.1 Fast
Speed-optimized Veo 3.1 model. Delivers high-quality video rapidly for dynamic workflows.
Minimax Hailuo-02 Standard
Robust video model producing crisp 768p video. Offers a solid balance of quality and reliable performance.
Veo 3.1
Flagship video update: refined control, enhanced visual fidelity, and improved subtle motion details.
Kling O3 Standard
Generate videos with prompts, images, or reference elements. Standard tier for balanced quality and speed.
Kling O3 Standard Edit
Edit videos guided by prompts and reference elements. Standard tier.
Kling V3 Standard
Generate videos with prompts and images. Standard tier with negative prompts.
PixVerse v5
Powerful, user-friendly video model producing high-quality, stylized 1080p video with consistent motion.
Runway Gen-4.5
Runway's high-fidelity image-to-video model with improved prompt adherence and motion quality.
Gemini 3.1 Flash-Lite Image Edit
Google's fastest, most cost-efficient Gemini image editing model.
Gemini 3.1 Flash-Lite Image
Google's fastest, most cost-efficient Gemini image generation model.
GPT Image 1.5
GPT Image 1.5 is the latest image generation model from OpenAI, with better instruction following and adherence to prompts.
Kling v2.6 Pro
Generate videos from images with native audio generation and fluid motion.
Gemini 3.1 Flash Image Edit
Google's fast, high-quality image editing model with multimodal reasoning.
Gemini 3.1 Flash Image
Google's fast, high-quality image generation model with multimodal reasoning.
Seedance 1.5 Pro
Next-generation Seedance video model with improved quality, 1080p output, and integrated audio generation.
Veo 3.1 Lite
Cost-effective Veo 3.1 variant. Balances quality and affordability for high-volume video creation.
Reve 2.0
Reve's flagship image model with best-in-class prompt adherence and text rendering, plus native editing.
Minimax Hailuo-2.3 Pro
Pinnacle of Hailuo T2V series. Cinematic quality with superior coherence, detail, and artistic control.
Minimax Hailuo-02 Pro
Premium video model engineered for professional-grade 1080p output, superior fidelity, and smoother motion.
Minimax Hailuo-2.3 Standard
Latest standard video model. Improved prompt understanding and visual consistency for daily creation.
Gemini 3 Pro Image Edit
Google's state-of-the-art image generation and editing model.
Gemini 3 Pro Image
Google's state-of-the-art image generation and editing model.
Wan 2.5
Powerful video model. Optimized for top-tier 1080p cinematic quality and consistency.
Seedance Pro
High-quality video model. Tuned for maximum visual fidelity and broadcast-quality 1080p output.
Grok Imagine Image Edit
xAI's image editing model for modifying existing images using text prompts.
Grok Imagine Image
xAI's image generation model capable of creating high-quality images from text prompts.
Veo 3
Next-gen video model with enhanced control over narrative, tone, and shot composition. Includes audio.
Luma UNI-1 Max
The highest-fidelity tier of Luma UNI-1, for hero-quality stills and premium edits.
Kling Image V3
Kling V3 image model. High-quality images with negative prompts, supports up to 2K resolution.
Gemini 3.1 Flash TTS
Google's most controllable TTS model with 200+ audio tags for vocal style and delivery.
Wan 2.6
State-of-the-art multimodal video generation model from Alibaba, with native audio support
Kling Image O3
Kling Omni 3 image model. High-quality images with text rendering capabilities up to 4K resolution.
Kling v2.1 Master
Premium 1080p video model. Maximum visual fidelity, capturing intricate details and lifelike expressions.
FLUX.2 [max] Edit
FLUX.2 [max]
Minimax Hailuo-2.3 Fast
High-speed T2V variant. Optimized for rapid creation, iteration, and social media content workflows.
Krea 2 Turbo
Speed-optimized open-source version of Krea 2 — high-fidelity images in seconds, with the full Krea aesthetic range.
LTX Video 2.0 Pro
High-fidelity video model designed for professional-quality results, offering superior detail and nuanced audio.
Seedream 4.0
Unified architecture for image generation and editing. Allows fluid movement from concept to refinement.
Seedream 4.5
A new-generation image creation model from ByteDance, for both generation and editing.
FLUX.2 [pro]
FLUX.2 [pro] Edit
Sora 2 Pro
Premium video version offering higher resolution (up to 1024p) and enhanced controls for pro projects.
Kling v2.1 Pro
Refined Kling model delivering professional-grade 1080p video with improved clarity and motion.
LTX Video 2.0 Fast
Versatile video model integrating video and audio creation in one seamless, speed-optimized workflow.
Veo 3 Fast
Speed-optimized Veo 3 variant for rapid video creation and iteration, ideal for short-form content.
Imagen 4 Ultra
Google's highest quality image generation model.
Gemini 2.5 Flash Image Edit
Google's image generation and editing model capable of multimodal reasoning.
Gemini 2.5 Flash Image
Google's image generation and editing model capable of multimodal reasoning.
FLUX.2 [flex] Edit
FLUX.2 [flex]
Sora 2
Next-gen video model generating long, high-fidelity 720p video with unparalleled narrative understanding.
ElevenLabs TTS V3
ElevenLabs' most expressive TTS model with inline audio tags for emotion and delivery.
Kling v2.0 Master
Master-grade video model. Enhanced realism and physics simulation for cinematic, high-impact clips.
Qwen Image 2 Pro Edit
Qwen Image 2 Pro image editing with higher quality prompt-guided transformations.
Qwen Image 2 Pro
Qwen Image 2 Pro tier with higher quality text-to-image generation and enhanced detail.
Seedream 5.0 Lite Edit
Intelligent image editing from ByteDance, with multi-reference support for creative advertising.
Seedream 5.0 Lite
Fast, high-quality image generation from ByteDance, optimized for creative advertising.
FLUX.2 [klein] 9B
FLUX.2 [klein] 9B Edit
Qwen-Image Edit 2511
Latest Qwen image editing model. Supports prompt-guided transformations with enhanced quality.
Imagen 4
Next-gen image model for professionals. Unparalleled prompt adherence and high-resolution output.
LTX Video 2.3 Fast
Speed-optimized LTX video model with extended duration support up to 20 seconds and portrait mode.
LTX Video 2.3
Latest LTX video model with sharper details, cleaner audio, and portrait support for professional content creation.
Qwen-Image 2512
Latest Qwen text-to-image model with enhanced quality, detail, and prompt adherence.
FLUX.2 [dev]
FLUX.2 [dev] Edit
Recraft V4.1 Pro
Professional design model with extended prompt support (10,000 characters) and advanced style customization.
Luma UNI-1
Luma's unified image model for high-fidelity generation and prompt-based editing.
Recraft V4.1
Professional design model with extended prompt support (10,000 characters) and advanced style customization.
Qwen-Image Edit 2509
Advanced image editing model. Enhanced performance for fine-grained manipulation.
Seedance Lite
Versatile and efficient video model. Optimized for speed and ideal for short clips and rapid prototypes.
GPT Image 1
Powerful, versatile OpenAI image model for creative and professional apps.
Recraft V4
Professional design model with extended prompt support (10,000 characters) and advanced style customization.
Kling v1.6 Pro
Powerful video model for high-fidelity, imaginative content with complex character motion.
Hunyuan Video 1.5
High quality open-source video model from Tencent Hunyuan.
Veo 2
Legacy video model generating high-definition, long-form cinematic video content.
xAI TTS V1
xAI's text-to-speech model with speech tags for expressive delivery in 20+ languages.
Qwen Image 2 Edit
Qwen Image 2 standard image editing with prompt-guided transformations and multi-image input.
Qwen Image 2
Qwen Image 2 standard text-to-image model with strong prompt adherence and diverse style support.
FLUX.1 Kontext [max]
Premium model for max editing performance, superior typography, and visual narrative consistency.
Wan 2.2
Advanced video model. Enhanced visual consistency and detail for high-resolution 1080p content.
FLUX.2 [klein] 4B
FLUX.2 [klein] 4B Edit
Imagen 3
Cutting-edge T2I model. Exceptional detail, photorealism, and accurate text rendering.
ElevenLabs Multilingual V2
Multilingual text-to-speech with natural voice selection and stability controls.
Z-Image Turbo
Ultra-fast 6B parameter image model from Tongyi-MAI, optimized for near real-time generation.
FLUX 1.1 [pro] Ultra
Delivers ultra-high-res (up to 4MP) images with superior photorealism, detail, and speed.
FLUX.1 Kontext [pro]
Pro-grade multimodal model for fast, iterative editing, style transfer, and consistency.
Qwen-Image Edit
Versatile, open-source image editor. Performs modifications via text instructions.
FLUX 1.1 [pro]
Next-gen FLUX model. 6x faster with enhanced prompt adherence and top-tier quality for production.
Runway Gen-4 Aleph
Runway's video-to-video model for restyling and transforming existing footage from a prompt.
Runway Gen-4 Turbo
Runway's speed-optimized image-to-video model for rapid, high-quality clip generation.
Runway Aleph 2
Runway's next-generation video-to-video model for restyling and transforming existing footage from a prompt.
Imagen 4 Fast
High-velocity Imagen 4 model, optimized for speed. Essential for quick iteration and interactive creative tools.
FLUX.1 SRPO [dev]
12B flow transformer fine-tuned with SRPO for exceptional photorealism and polished composition.
FLUX.1 [pro]
Flagship commercial T2I model offering superior prompt adherence, quality, and stylistic outputs.
Recraft V3
Versatile image model for graphic design. Generates legible, stylized text and scalable vector art (SVG).
Qwen-Image
Open-source T2I model. Excels at high-res images from complex text, notable for clear, stylized text.
FLUX.1 Krea [dev]
Open-weight model co-developed with Krea AI. Excels in photorealism and aesthetics.
FLUX.1 [dev]
Powerful, open-weight 12B image model. Excels in image quality, prompt adherence, and commercial use.
Minimax Video 01
Accessible video model. Generates engaging 720p clips from text with good visual consistency.
Stable Diffusion 3
Latest open-weight MM-DiT model. Major improvements in quality, prompt following, and text rendering.
Wan 2.1
State-of-the-art video model. Generates detailed, stylistically diverse clips with fluid motion.
FLUX.1 Kontext [dev]
Open-weight, multimodal model for context-aware image editing. Excels at iterative edits via text.
Hunyuan Video
Impressive open-source video model. Excels at stable, coherent sequences with high visual quality.
FLUX.1 [schnell]
Ultra-fast, open-source model. Generates high-quality images in 1-4 steps for rapid prototyping.
Ray 2 Flash
High-speed variant of Ray 2, optimized for rapid video creation. Perfect blend of speed and quality.
Ray 2
Large-scale, state-of-the-art video model for stunning realistic and coherent 1080p motion.
BiRefNet v2
High-quality image background removal.
SeedVR2 Image Upscaler
ByteDance's SeedVR2 model for high-quality image upscaling.
ElevenLabs Sound Effects
Generate custom sound effects from text descriptions with duration and prompt control.
Ideogram Remove Background
Fast, clean background removal from Ideogram.
Meshy V6
Meshy V6. High-fidelity 3D asset creation focusing on PBR, quad mesh, and face rigging.
Tripo v3.0
Latest 3D model for production-quality assets with superior textures and clean geometry.
Gemini Omni Flash Reference
Reference-guided video generation with up to 5 reference images for consistent characters and style.
Hunyuan 3D v3.1 Pro
Latest Hunyuan 3D model with enhanced quality, multi-view input, and PBR material support.
Video Subtitles
Topaz Enhance
Enhance image quality with advanced upscaling.
Gemini Omni Flash Edit
Google's video editing model for modifying existing videos with a simple text instruction.
ESRGAN Upscaler
Very fast upscaling with good quality.
Topaz Video Upscaler
Professional-grade video upscaling solution utilizing Topaz AI technology for high-quality enhancement.
Topaz Generative Enhance
Topaz Generative Enhance for enhanced details.
Bria Video Background Removal
Remove video backgrounds for advanced video editing.
Clarity Creative Upscaler
Upscale images with high fidelity or creativity.
Recraft Vectorize Image
Convert raster images to vector graphics using Recraft.
Gemini Omni Flash
Google's fast multimodal video model — 720p clips with synchronized native audio from text or a still image.
Bria Video Increase Resolution
Bria's advanced AI technology designed to increase the resolution of video content efficiently.
Recraft Crisp Upscale
Boost resolution while refining small details and faces.
Pixelcut Video Background Removal
Remove video backgrounds with Pixelcut for clean cutouts.
Qwen-Image Layered
Split an image into layers using Qwen-Image Layered.
Hunyuan Video Foley
Specialized model to automatically create and sync sound effects (foley) for video content.
Qwen-Image Edit 2511 Multiple Angles
Qwen Image Edit 2511 with the Multiple Angles LoRA. Re-renders the input image from a chosen camera angle.
ElevenLabs Music
Premium AI music generation from ElevenLabs with composition planning and section control.
ESRGAN Video Upscaler
ESRGAN model for video upscaling, enhancing resolution and detail.
Bria Video Background Removal v3
Bria's v3 video background removal for advanced video editing.
Pixelcut Background Removal
High-quality image background removal from Pixelcut.
Hunyuan 3D 3.0
Professional-grade 3D model optimized for high-quality, detailed assets with advanced features.
Rodin v2
Advanced 3D generation model creating high-quality, textured T-pose avatars from a single image.
Magi Distilled
Fast, efficient open-source I2V model. Animates still images using an autoregressive approach.
Hunyuan 3D v2 Mini
Lightweight, efficient 3D version optimized for less powerful hardware. Delivers good quality assets.
Reve 2.0 Remix
Reve's remix model - blends one to six reference images into a single new image, guided by a prompt.
Veo 3.1 Fast Reference to Video
Speed-optimized reference-guided Veo 3.1: character-consistent video from up to three reference images.
Recraft V4.1 Vector
Recraft's V4.1 text-to-vector model. Generates scalable vector art (SVG) directly from a prompt.
Qwen 3 TTS
Multilingual TTS with zero-shot voice cloning and prompt-based voice design.
Veo 3.1 Reference to Video
Reference-guided Veo 3.1: lock characters and objects across shots using up to three reference images.
Recraft V4.1 Pro Vector
Recraft's V4.1 Pro text-to-vector model. Generates scalable vector art (SVG) directly from a prompt.
Image to SVG
Convert raster images to clean, scalable SVG vector graphics with fine-grained control over detail.
Meshy Rigging
Auto-rigs humanoid 3D models and optionally applies a preset animation, returning rigged GLB/FBX.
Recraft Creative Upscale
Upscale for a sharper, cleaner, higher-resolution result.
Stable Audio
Open-source text-to-audio model for sound effects, field recordings, and instrument samples.
Minimax Music V2.6
Latest MiniMax music model with native audio-visual generation capabilities.
Minimax Music V2.5
Updated MiniMax music model with improved vocal quality and arrangement control.
Stable Audio 2.5
Enterprise-grade audio generation with multi-part compositions and audio inpainting.
CassetteAI Sound Effects
Real-time sound effects generation in under 1 second, up to 30 seconds long.
SAM Audio Separate
Foundation model for audio source separation using text, visual, or temporal prompts.
Inworld TTS 1.5 Max
Inworld's premium TTS model optimized for expressive game character voices.
ElevenLabs Text to Dialogue
Generate multi-speaker dialogue audio from structured text with per-speaker voice control.
Lyria 3
Google DeepMind's next-generation music model with enhanced composition and vocal quality.
Lyria 3 Pro
Premium tier of Google's Lyria 3 with highest fidelity and compositional complexity.
ElevenLabs Voice Changer
Transform the voice in an audio recording to a different target voice.
CassetteAI Music
Ultra-fast music generation producing 3-minute tracks in under 10 seconds at 44.1 kHz.
Lyria 2
Google DeepMind's high-fidelity 48 kHz music model with fine-grained creative control.
Minimax Music V2
AI music generator producing complete songs with vocals, lyrics, and full instrumentation.
DeepFilterNet3
Real-time speech enhancement and noise suppression at 48 kHz full-band audio.
Kling Video to Audio
Generate synchronized sound effects, dialogue, and ambient audio from video content.
ElevenLabs Audio Isolation
Isolate clean voice from noisy recordings by removing background noise and music.
Mirelo SFX 1.6
Sound effects generation and editing with text-to-audio and audio inpainting.
ACE-Step
Fast open-source music generation with lyrics alignment, remixing, and audio editing.
Gemini TTS
Google's text-to-speech model with multi-speaker support and natural expressiveness.
Hunyuan 3D v3.1 Fast
Fast Hunyuan 3D model for rapid 3D generation with PBR material support.
Trellis 2
Open-source, high-quality 3D model from Microsoft, leveraging a novel field-free sparse voxel structure.
Meshy V5 Remesh
Dedicated Meshy 3D tool for remeshing models to optimize topology and reduce polygon count.
Meshy V5 Retexture
Specialized Meshy 3D tool for quickly retexturing imported meshes using text prompts.
Bytedance Seed 3D
Powerful 3D model focusing on high-quality objects from a single image. Adept at geometry & texture.
Tripo Turbo v1.0
Speed-optimized 3D generation model designed for rapid prototyping and fast generation times.
SeedVR2 Video Upscaler
Powerful video upscaling model from ByteDance, optimized for high-quality resolution and fidelity improvement.
OmniHuman
Advanced video model bringing a still image of a person to life using audio, producing expressive videos.
AI Avatar
Specialized model for creating realistic, audio-driven talking avatars with accurate lip-sync and expressions.
Anything World Animate
Specialized mesh rigging model that automatically prepares 3D models with skeletons for animation.
Framepack
Highly efficient, open-source I2V model that generates video by predicting the next frame.
Tripo v2.5
Incremental 3D update. Refined performance, improved mesh topology, and texture fidelity.
Hunyuan 3D v2
Powerful, open-source 3D model producing high-res, textured 3D objects from text or image inputs.
Minimax Video 01 Live
Specialized video model. Optimized for a dynamic, live-action feel with naturalistic camera work.
Trellis
Open-source 3D model creating high-quality objects with realistic materials and geometry from text.
Imagen 3 Fast
Speed-optimized Imagen 3. Delivers high-quality images fast, ideal for real-time previews.
Reve 2.1
Reve's next-generation image model with strong prompt adherence and text rendering, plus native editing.
Reve 2.1 Remix
Reve's remix model - blends one to eight reference images into a single new image, guided by a prompt.
Seed Audio 1.0
Bytedance Seed Audio TTS with preset voices and zero-shot voice cloning.
No models match your search.