Skip to content

Changelog

What's new on Layer

Product updates, new models, improvements, and fixes — shipped continuously.

Rig a Photoshop file for Spine, tileable textures, and brand Reference Sets

Rig a Photoshop file for Spine, tileable textures, and brand Reference Sets

Hand the Layer agent a Photoshop file and get back a character rigged and animated for Spine. Also this week: textures that repeat without a seam, a brand you can carry in a Reference Set, and our regular refresh of the models behind your everyday tools, so you get the best results without having to pick.

What's new?

Rig a Photoshop file for Spine: Give the agent a layered Photoshop file and ask it to rig the character, creature or prop for Spine. It reads your layers as parts, builds a skeleton with bones and pivots, and animates an idle plus up to three clips you ask for, for Spine 4.1, 4.2 or 4.3. No layers? A flat image works too: Layer splits it into parts first. Every animation plays in the viewer before you export, so you can ask for changes in the same chat until the motion is right. New this week, a rig goes beyond the bones: a surface can breathe, squash or ripple; a boot or a glove declares its opening, so the limb shows inside the cuff instead of behind the painted rim; a character facing you gets a walk that steps in place; and each clip can carry its own face. A rig needs at least 200 Creative Units in the workspace to start.

The best models, without the homework: New models arrive every week, and we keep testing them so you don't have to. This week's refresh moves the defaults for a new image, an edit, upscaling, Create 3D, the 3D composer, retexturing and splitting a mesh into parts onto the models that do each job best. Open a tool and it is already set up; nothing to compare, nothing to change. GPT Image 2.5 also costs less than it did. You can still pick any model yourself, and the models page lists every one.

Tileable textures: Ask for a texture that repeats, and the seams are joined. On models that offer it, Tileable is in the advanced settings. You can also ask the agent, including on a model whose form has no switch. Pixel art and organic surfaces are joined without another generation. A structured texture, bricks or planks, still repaints the seam, and that step spends Creative Units. From a texture you already have, the agent can derive a normal map, a height map, ambient occlusion, roughness and emissive that tile with it. Those maps do not spend Creative Units. Pixel art is worked one cell at a time.

A brand, as a Reference Set: A Reference Set can carry a brand: a palette of up to 16 named colours, logo marks, a wordmark and fonts. Whenever you use the set, the palette is added to the prompt, except on audio models. Logos and fonts stay on the set. The agent places a logo when you ask for it, and it does not train on those files.

Insert a scene: On a video, Insert scene marks the stretch to replace. MiniMax H3 Max Insert writes a new scene into the clip, and you get the whole clip back. You spend Creative Units for the new scene. The clip you get back includes the original footage around it. A scene runs from 5 to 13 seconds, at 480p or 768p, and the stretch you replace has to begin more than a second and a half into the clip.

Clean up a recording: Denoise is available on an audio file, from the same place you denoise an image. VEED Clean Audio is the default: it keeps the speech, including quiet and distant words, at the same length as the source. It sits with the rest of audio cleanup.

Exact words, in your font: Ask the agent to set text in a font from your workspace. The words are rendered exactly, on a transparent background, ready to place on an image. That step does not spend Creative Units.

A workflow can drop weak results: A workflow can score every item in a list against a rubric you write, and pass on only what clears the bar. There is a judge for text, images, layered images, video and meshes. If nothing passes, the step stops and says so.

Bug fixes & improvements

Cancel several generations at once from the selection bar, when every tile you selected is still running. One confirmation covers the batch, and a run still waiting in the queue is not charged.

The app on a phone: Below a tablet width the sidebar is a drawer, a session is a gallery, and listings keep search. The panel sits under the work.

Weights you set stay put. Clearing a source image and dropping another keeps the similarity you last used. Double-click a weight to type it. Drop an image on the prompt and it is described into the prompt, without spending Creative Units. The LoRA picker opens on your trained sets.

Drag a video, a mesh or audio off the board into the field that accepts it. A video no longer fails as if it were an image.

Usage matches the balance. The breakdown includes today, and a grant's rollover line states that grant's own rollover.

Edit a published workflow's name, description and covers from its status, without starting the publish steps over. Generated covers are 4:3, the same shape as the tiles.

Inset reveal is a workflow cover style: the outputs fill the frame, and the inputs sit in a card with an arrow out of it. You can set the background colour.

A new workflow can start from one that exists. When you ask for a workflow, Layer looks at featured workflows first and offers a close match instead of rebuilding it.

In Blender, scene work has its own builder. Blockouts, lighting and cameras land in the open file, in small steps. It is offered on Blender turns, not on the web.

A playable's first build uses the art it already approved, and assets you attach on that first request go into the game. The art bible card states the game's rules. A playable exported for AppLovin reports its load and its challenge there. Other networks stay quiet.

Pin models and Reference Sets into a toolset. Each view also shows what you used recently.

Running a workflow app shows its cover, drop tiles for the media it needs, and examples from the board.

Text you generate, or a text file, reads as text in the viewer.

Dropping several items onto a group adds all of them.

A 3D animation takes one animation Reference Set.

An identity provider can turn a suspended member back on. The account comes back with the role it had. Access to a child workspace returns when that group is pushed again.

A chat turn no longer fails when the agent leaves the camera unset.

A per-second generation is charged the way it was quoted.

A voice clip longer than the provider accepts is refused before training starts.

Select a range of items with shift, and toggle one with the command key, in the session's items panel.

An upload from the library stays in the project you have open.

Workflow drafts show by default. A Published only switch hides them.

A colour you are typing is not reformatted until you leave the field.

Edit video no longer asks for a first or last frame. Those belong to a model that generates a clip, not one that changes one.

A character Reference Set opens on its portrait.

A template can prefill only the form it names.

Duration can stay on Auto, and Enhance stays inside the length the model accepts.

New models17
Grok Imagine Video 1.5 LiteVideo
xAI

xAI's lower-cost Grok Imagine Video 1.5 Lite, generating 480p, 720p, or 1080p clips from text or a still image.

FLUX.3 ImageImage
Black Forest Labs

Black Forest Labs' FLUX.3 image model for native 1K to 4K generation and multi-reference editing.

MiniMax H3 Max InsertVideo
MiniMax

Insert a new 5-13 second scene into a video with MiniMax H3 Max, then return to the original footage.

VEED Clean Audio 1.0Audio
VEED

Remove background noise from speech while keeping quiet and distant words intact.

Ideogram V4.5Image
Ideogram

Ideogram's text-to-image and editing model with accurate typography and three quality tiers.

P-ImageImage
Pruna AI

Pruna's sub-second text-to-image and multi-image editing model built for production volume.

P-VideoVideo
Pruna AI

Pruna's fast video model for text-to-video and image-to-video at 720p or 1080p.

P-Video 2Video
Pruna AI

Pruna's quality-focused successor to P-Video, at 720p or 1080p.

P-Video 2 ProVideo
Pruna AI

Pruna's highest-quality video model, with speed and quality modes at 480p or 768p.

P-Video AnimateVideo
Pruna AI

Pruna's motion transfer: a reference image performs the motion and audio of a source video.

P-Video AvatarVideo
Pruna AI

Pruna's fast, low-cost talking avatar: a photo lip-synced to an audio track.

P-Video Avatar TTSVideo
Pruna AI

Pruna's talking avatar that voices a written script with a built-in voice.

P-Video EditVideo
Pruna AI

Pruna's video editor: change a clip of up to 15 seconds with a prompt and optional reference images.

P-Video ReplaceVideo
Pruna AI

Pruna's character swap: replaces the person in a video with one from reference images.

P-Image UpscaleImage
Pruna AI

Pruna's sub-second upscaler, scaling images up to 8x and 128 megapixels.

ElevenLabs TTS V4Audio
ElevenLabs

ElevenLabs' expressive v4 speech model with audio tags, stability, and similarity control.

ElevenLabs TTS V4 TurboAudio
ElevenLabs

Lower-latency Eleven v4 Turbo speech with audio tags, stability, and similarity control.

Layer Plugins are Live

Layer Plugins are Live

Layer Plugins are live, in seven places: Photoshop, Premiere Pro, After Effects, Blender, Figma, Slack, and Chrome. Each one puts Layer where the work already happens: generate against the document, the project, the composition, the scene, the canvas, or the thread in front of you, and the result lands back in it. Every plugin is free with a Layer account, signs in with the account you use on the web, and generations spend Creative Units the same way they do in the browser. Everything you generate is saved to a Layer session, so your team finds it where it always looks.

Layer Plugins are Live

Photoshop: The Layer panel for Photoshop generates against the document you have open, on Photoshop 25.10 or later, on macOS and Windows. Your selection is the mask: marquee, lasso, or Select Subject, and only that area is regenerated, placed back as a new layer with a layer mask, feathered so it sits on the image. Inpaint, relight, change the camera angle, extend the canvas (a reframe of the whole document grows it, as its own undo step), split a flat image into layers, or remove a background, with the full model catalogue and your trained styles behind one prompt box. The panel is redesigned this week: it uses Photoshop's own type, can take a style-transfer prompt, suggests Reference Sets from the project you are in, says when a capture is read smaller than the document, and tells you when an update is out. Install it from the .ccx on the page with the Creative Cloud desktop app.

Premiere Pro: The Layer panel for Premiere Pro generates video and audio beside the edit you are cutting, on Premiere Pro 25.6 or later, on macOS and Windows. Describe a shot or a sound, and the result is imported into your project, ready to drag onto the timeline. It installs from a .ccx, the same way the Photoshop panel does.

After Effects: The Layer panel for After Effects puts generation next to the timeline, on After Effects 25.0 or later, on macOS and Windows. Describe a plate, an element, an end-card badge, or a sound, and the result arrives as footage at the top of the layer stack, sized to the region it was made for, in one undo step. Image tools work from the frame you are on. Footage the panel placed can be selected and changed with a prompt, reframed for a vertical placement, or upscaled. For any other clip, the panel asks you to generate one here or add it from your Layer session, then select that layer. It docks beside Effect Controls, and generation runs on Layer rather than on your machine, so a render never competes with it for the same cores. Install it from the .zxp on the page with a ZXP installer.

Blender: Layer for Blender is a native add-on for Blender 4.5 or later, on macOS, Windows, and Linux. Its tools and its agent sit in the 3D viewport and work on the scene you have open: meshes, textures, images, and video land in the file, and a change the agent makes is one undo. Restyle and motion to video can use the scene's animation. The add-on opens on the prompt bar, keeps the agent's history, updates itself, and says when a restart is due.

Slack: Layer for Slack makes a thread a session. A workspace admin connects it once. Mention Layer in a channel or a direct message, and that thread gets the same agent, your styles and Reference Sets, with the reply showing as it is written. Images, video, and audio are posted back into the thread, and a 3D model arrives as its rendered preview. A prompt sent from Layer into the thread names who sent it, and the thread links to the session, where the full files live.

Figma: The Layer plugin for Figma runs from any file. Generate backgrounds, icons, buttons, and characters in the panel and place them straight onto the canvas, with no export and re-import. Install it from the Figma Community.

Chrome: The Layer Reference Collector saves reference images from any website into your Layer workspace: one click, or a batch selection of many images at once, with optional auto-upload. It works in Chrome and Edge, and it is free with any Layer plan.

What's new?

Spine, from a PSD or a flat image: Ask Layer to rig a character, creature, or prop for Spine. It starts from a layered PSD, or from a flat image that it splits into parts, and it asks you to confirm the Spine version (4.1, 4.2, or 4.3), the clips, and the cost before it runs. A rig is a longer job: it needs at least 200 Creative Units in the workspace to start. The result is a skeleton with bones and pivots, an idle, and up to three clips you asked for, and you preview it right in Layer: the rig and every animation play in the viewer, whichever Spine version you chose, so you can check the motion, ask for changes in the same chat, and iterate before you export. Parts can bend as weighted meshes instead of hinging as rigid cutouts, and on Spine 4.2 and 4.3 hanging pieces such as hair, capes, and tails can carry physics.

350+ cinematic presets: A new library of Reference Sets names a technique so you do not have to describe it from scratch, each with a cover that shows it in action. Mention one in a prompt, or pick it from the video form, where they are grouped as Camera, Staging, Look, and Performance. There are twelve types:

  • Camera Movement: dolly, crane, orbit, whip pan, and more
  • Camera Angle
  • Camera Framing
  • Camera Lens
  • Scene Lighting
  • Scene Composition
  • Scene Atmosphere
  • Genre Look
  • Optical Effect
  • Visual Effect
  • Character Performance
  • Time and Motion: speed ramps, slow motion, and other changes of pace

Moves and changes of pace are for video. Every other type describes a single frame as well, so it works on image models too.

Extend and speech: Extend video defaults to MiniMax H3 Max Extend. You get the whole clip back, the source plus the new seconds. Gemini 3.8 Flash TTS is the default for speech.

Music, section by section: On audio models that take a composition plan, the advanced panel is a list of sections: what holds for the whole track, then each section's length, what enters, what stays out, and any sung lines. The same plan can be sent from the API and from MCP.

OBJ and FBX downloads: A 3D model can be downloaded as OBJ with its textures as PNG files, without spending Creative Units. Tripo models can also be downloaded as FBX, with the rig and the animation kept. That conversion spends Creative Units.

Several images, one Photoshop file: Exporting a selection of images as PSD can combine them into a single layered file: one layer per image, the first you selected on top. A PSD that already has layers keeps them as a group.

Upscale from the composer: Upscale an image or a video you upload, from the same composer you generate in. Image upscalers that support it expose creativity and resemblance there. Leave them unset and the model keeps its own defaults.

Save the session as a template: Save as template now keeps the session: its name, and the generation settings that were on the form. You can edit those settings later, instead of recapturing the session.

A rule in chat can become a skill: State a rule the work should follow, and the agent can save it as a Workspace Skill on the project you are in. It confirms what it will write, and what it will be called, before it saves, unless you already named the skill and the change. Later chats in that project follow it.

An animated mesh plays: A 3D model that carries an animation gets a short, silent, looping preview. It plays in the viewer and on the session tile, the same way a video does.

Bug fixes & improvements

Motion transfer: A character in a still can perform the motion of a reference clip.

Stop a workflow run from the header, or from a node that is running.

Share a workflow, or make it private. The list can filter to public or private. Export a published workflow as JSON from the canvas, and import a whole workflow through MCP.

Animation Reference Sets have their own slot on 3D models that play an animation, instead of sitting in the general tile.

Reference Set prompts: A set has a Prompts tab. Text marked always applied is added whenever that set is used. A prompt template is something you insert when you want it.

Mention a Reference Set and the suggestions follow the project and the model.

Command palette: ⌘K opens onto a board of the workspace: templates, projects, files, Reference Sets, workflows, models, and sessions. Typing narrows it to one list, with a preview of the row you have highlighted.

Workspace members: Select several people and change their role, suspend or activate them, or remove them. A members admin sees each person's groups by name. Last Activity stays hidden when it cannot be read, instead of reading Never for the whole workspace.

Photos upload the way they look. Phone pictures keep their orientation, and a HEIC file is stored as a JPEG.

Transparency from the model: When a model can return its own alpha, a transparent request uses that, instead of a second pass that cuts the background out.

Stack images without a generation: The agent can place images you already have into one file, a logo on key art or a character on a background. No model runs, and it does not spend Creative Units.

The agent can see the session board and pick from the assets on it.

A failed model is not a billing problem. When a provider fails, the agent no longer describes that as a problem with your Creative Units.

A prompt that is too long is refused when you submit it, and the agent no longer treats a generation as finished before it is.

Connectors are split into what you have connected and what you can add, and connecting happens in Layer's own dialog.

Playable controls: A playable can name how it is played in portrait, from a hold or a swipe to a joystick.

A session board stays responsive with hundreds of items on it.

Video follows its rotation in previews, sizes, and renditions. A reframe returns the whole clip you started from, and the reframe canvas fills the stage. While a canvas tool has the stage, you see the clip's still.

The training email names the style that finished.

Batch size can be set on Create Image and Prompt Generator nodes, on models that generate more than one result.

Search, sort, and filters share one layout on every listing.

Download as is available from a workflow lightbox.

A workflow cover can be generated when you publish, and a node that was missing one is filled in.

An unpublished workflow tells a colleague why it cannot be run yet.

The general style category is labeled Art Style.

Tripo meshes can come in low poly, and they face the front.

Music can follow an attached picture, on a model that accepts one, and a hard length or a request for no vocals is kept as your instruction.

Docs search works after you move to another page, including on the REST reference.

A creator can turn on a public page for their workspace. The address uses the workspace name.

Opening a project you are not a member of says so.

A space separates invite addresses, the same way a comma does.

Every slider uses the same handle, shows its value while you drag, and lets you type the number.

A chat stays in the project its session is filed under.

A model with no prompt says so, instead of offering a box it will ignore.

Skills can be searched, and they show up in the command palette.

Session templates that were named for an ad network are named for the job instead, such as voiceover variants or a playable ad storyboard.

Retexturing from the composer works again for Hi3D Texture and Meshy V5.

When a 3D animation preset ignores the prompt, the agent says so, and it names the preset.

A PSD in chat shows its rendered still.

A permission refusal comes back to the agent instead of ending the run.

A batch is offered only on models that can fill one.

A pixel-art job ends snapped, on the colours the art uses, without a separate ask.

Retexture can ask for PBR or an HD texture, and PBR stays off unless you turn it on. You can delete a version from its strip. A Meshy rig no longer comes back washed out.

New models19
Gemini 3.8 Flash Lite TTSAudio
Google

The lighter, lower-cost tier of Google's Gemini 3.8 text-to-speech model.

Gemini 3.8 Flash TTSAudio
Google

Google's expressive text-to-speech model with style direction and inline vocal events.

Lyria 3.5Audio
Google

Google DeepMind's latest music model, writing full-length songs with vocals from one prompt.

MiniMax H3 Max 3D to VideoVideo
MiniMax

Render a Blender proxy animation as realistic video with MiniMax H3 Max, keeping its camera and motion.

PixVerse VibeMV 1.0Video
PixVerse

PixVerse music video model that turns a song into a full video timed to the track.

Ray 3.2Video
Luma AI

Luma's cinematic video model: generate from text or frames, restyle footage, and reframe to any aspect ratio.

Seedream 5.0 FlashImage
ByteDance

Fast, budget-friendly image generation and editing from ByteDance's Seedream 5.0 line.

Seedream 5.0 Flash LayerizeImage
ByteDance

Split a finished image into editable layers, fast, with Seedream 5.0 Flash Layerize.

MiniMax H3 Max ExtendVideo
MiniMax

Extend a video clip with MiniMax H3 Max: 5-15 seconds of new footage at 480p, 768p, 1080p, or 2K.

Genjutsu Motion TransferVideo
Higgsfield

Higgsfield's video model that moves a referenced character through the motion of a source clip.

Genjutsu Object SwapVideo
Higgsfield

Higgsfield's video model that swaps an object in a source clip for one shown in reference images.

Soul 2Image
Higgsfield

Higgsfield's photographic text-to-image model for fashion, editorial and character imagery.

Soul CinemaImage
Higgsfield

Higgsfield's text-to-image model with a fixed cinematic, film-still look.

Kling V2.6 Pro Motion ControlVideo
Kling

Kling's motion control: the character in your image performs the motion of a reference video.

Kling V2.6 Standard Motion ControlVideo
Kling

Kling's motion control: the character in your image performs the motion of a reference video.

Kling V3 Pro Motion ControlVideo
Kling

Kling's motion control: the character in your image performs the motion of a reference video.

Kling V3 Standard Motion ControlVideo
Kling

Kling's motion control: the character in your image performs the motion of a reference video.

Meshy V7.13D
Meshy

Meshy V7.1. Text, image, and multi-image to 3D with PBR, quad mesh, and optional rigging.

Tripo P23D
Tripo

Tripo P2. Text or a single image to a textured 3D mesh with PBR, quad topology, and a 48 to 50,000 face budget.

Qwen Image 2.1, webhooks, and an onboarding conversation

Qwen Image 2.1, webhooks, and an onboarding conversation

This week brings Qwen Image 2.1 and new defaults for lip sync and video reframe, webhooks so an API call can hand back its result without polling, and an optional onboarding conversation that can build a first project with you.

What's new?

Qwen Image 2.1: Qwen Image 2.1 is the latest image model from the Qwen team at Alibaba, and one of the strongest open-weight models you can run today. It is built for prompt adherence, so a long brief lands closer to what you asked for, and it generates with control over size, guidance, seed, negative prompt and transparency. Qwen Image 2.1 Edit takes the same model to prompt-guided editing, combining or transforming up to 10 reference images, so a scene can be rebuilt from art you already have.

Video reframe and lip sync: LTX Video 2.3 Quality Outpaint is the default when you reframe a clip. MiniMax H3 Max Lip Sync is the default for lip sync, from a still plus a voice track, up to 2K.

Onboarding can start with the agent: A new user with an active workspace who has not answered onboarding can be offered an agent-led setup: it asks what you are working on, pulls your game's art in from its store page, and makes something with you. Taking it is optional, and that first conversation does not spend Creative Units. We are trying a few different welcomes, so not every new user meets the same one.

The agent leaves beta: The Agent tab no longer says Beta. Most agent actions still spend no Creative Units; the intensive ones do.

Webhooks: An API call can name a URL and receive the finished result there, signed, instead of polling. The response lists the deliveries to expect: the job's own outcome, plus a separate scoring event where the workspace scores automatically. A test call sends a signed sample to any URL, so you can check your verification before spending Creative Units. The webhooks guide has the payload, the signature headers, and what happens when your endpoint is down.

The project stays with you: The project you are in is one setting, carried across the app. Search, listings, and sessions stay inside it, and search finds session templates too.

A shared workflow runs: Opening a shared workflow adds it to your workspace as you run it. You no longer have to link it first. Publishing is three steps (Inputs, Outputs, Publish), with names and examples edited on the card.

Editable type in a Photoshop file: When a layered result carries its text, whether a layout the agent measured or a PSD that arrived with type layers, the copy comes back as live type in the face it was set in, rather than flattened into pixels.

Find a teammate: Workspace members can be searched by name or email, filtered by status, and exported. Built for studios with large teams.

Bug fixes & improvements

Filter video models by frame: The model list can show which video models take a first frame, a last frame, or both.

Workflow outputs, together: Select a run's outputs the way you select anything else, then download them, add them to the session, or make a Reference Set.

Sample Objects and Smart Extract Frame: Sample Objects draws a seeded pick from a list of assets, without duplicates. Smart Extract Frame chooses frames from a video.

Audio you can ask for: Instrumental, loop, and lyrics reach the models that support them, including from the agent.

A model page, three ways: Web, MCP, and API sit on one rail, and the API snippet is that model's own call.

Settings follow you into the next session, per workspace and per kind of asset. Only choices you actually changed are kept.

Connect MCP and API Docs sit in the sidebar footer.

An empty session offers your templates. Picking one fills the composer. Nothing is generated until you say so.

A favorite from the preview. The heart sits next to the filename while you are looking at the asset.

The thinking indicator changes with what the agent is actually doing, whether searching, estimating or generating, instead of a fixed spinner.

Featured Reference Sets open on general styles, rather than a wall of pose placeholders.

Clip lengths on a model page read as a range, such as 3–10s, instead of a row of chips.

CLI and skills have home pages at /cli and /agent-skills.

Playables have a landing page where you can play the result.

Top trends carry a short research summary, sources, and remix ideas.

Text from a video no longer fails when the clip is too large for the model. The video is reduced before a text model reads it. The same path covers Smart Extract Frame.

The workflow agent can pick a text model for a text node, from the same catalog the editor already listed.

Retexture a mesh offers the texture image those models require.

A copy of a view-only workflow lands on the session board, so you can get back to the copy you just made.

Continue where you left off is no longer hidden behind the workflow lightbox.

A parameter you select on the canvas reaches the agent.

The project you pick is the project you see. Switching project from settings opens that project's settings. Listing rails name the project they are scoped to. Opening a session no longer flashes the wrong project first, and a project change made inside a session sticks.

Reference Sets: a trained set says how it is being used, and retyping a prompt keeps the set attached.

Usage shows the day's total, and a parent workspace can list its children's balances. The previous model filters are back on the model list.

The 3D face limit can be typed, not only dragged.

A signup from the marketing site is not interrupted by the welcome dialog over a conversation they already started. New users are no longer met by a rewards popover and a feedback pitch.

Pixel art snaps to colours the artwork actually uses, including colour that lives in a small detail.

A request to edit your own playable is treated as work, not as a question about how the agent is built.

New models3

Layer CLI & SDK, Gemini masks, exports

Layer CLI & SDK, Gemini masks, exports

What's new?

Layer from the terminal: The Layer CLI and TypeScript SDK put generation, editing, and training where your scripts already are — one login, the same models and Reference Sets as the app. The agent skills teach a coding agent how to drive them.

Masked edits on Gemini: A masked edit is sent as an annotated init image, so Gemini sees the region you marked rather than inferring it from the prompt alone.

Command palette, with room to work: ⌘K opens a larger window with standard chrome, so searching projects, sessions, files, and models isn't cramped.

Arrow keys page the viewer in gallery order — next and previous without leaving the asset.

Exports name what they produced: A finished row is labeled with the container it wrote (zip, sequence, and so on). Exporting a video as a PNG sequence extracts the frames.

Usage, in more detail: Settings → Usage again shows the subscription and purchase detail. Chat agent spend is listed per agent, with readable token pricing.

Bug fixes & improvements

The session toolbar clears the docked chat sheet instead of the rail, so closing chat doesn't take the session chrome with it.

Projects join sessions, files, Reference Sets, and Workflows on the Search API — one listing for members and admins.

A finished export keeps its name when you come back to the transfers row.

GPT Image 2.5, FLUX.3 Edit & Meshy T2

GPT Image 2.5, FLUX.3 Edit & Meshy T2

What's new?

GPT Image 2.5 — Flare and Sunburst: OpenAI's GPT Image 2.5 is in the catalog for text-to-image and edit. Flare is now the top image-editing pick, ahead of GPT Image 2 and Gemini Flash Image.

FLUX.3 Edit: Prompt-based video editing on Black Forest Labs' FLUX.3 — describe the change, keep the shot.

MiniMax H3 Max Turbo: A faster H3 Max for text-to-video and image-to-video at 480p, 768p, and 1080p, when you want the same model family at draft speed.

Meshy T2: Meshy's flow-matching native mesh generator — text or a single image to 3D, with textures and PBR. On Meshy V7 and T2 you can now request ultra detail, remesh, and a texture image.

Camera moves on H3 Max: Author an orbit path with 2–12 keyframes (azimuth, elevation, distance over time) and H3 Max flies it. From the Camera move tool on an image, or from the generation form when the model can take a trajectory. A thumbnail of the path sits on the control once it's set.

Favorites across the board, gallery, and library: One heart on the tile. Filter a session to Favorites, or Group by favorites (⇧V) to collect them into a board group everyone in the session can see. The library sidebar has a Favorites view for what the workspace has picked out.

Video editor, from the clip: Open the editor from a video — full screen, with a cover on the tile — instead of assembling only from a blank timeline.

Several source images on edit models: Attach more than one image when the model can take them; generations fan out from the set.

Skills from the project: Create or import a skill in the project picker without leaving the page. Open a skill from the project sidebar. Project instructions convert into project-attached skills.

Asset type filter on the session board and gallery, replacing the items-panel type row.

Del and ⇧Del on the board: Del clears a failed generation; ⇧Del deletes the asset.

Timed votes on canvas objects, with stamps that outlive the round — the timer is only a timer.

QR code in Export: The lightbox Export section shows a QR code for the asset URL.

Session Creative Units in chat: The rail shows Creative Units used in the session.

Provenance on exports: C2PA claims are signed at file creation and on re-encoded exports, including video and audio.

A full-screen asset viewer again, from the session.

Retrain a Reference Set when a finetune predates the set's current files.

Edit Image Auto size matches the base image, the same Auto the composer already offered.

Downloads run over the websocket, so a file transfer no longer depends on a separate HTTP path.

Exports say what they're doing while they run, and they outlast a connection blip instead of failing silently.

Bug fixes & improvements

Seedream 5.0 Pro is public now that it left the provider's early-access program.

Ideogram is the default background-removal model.

The sessions listing can show every member's sessions, with an All tab.

The Asset Library uses the Search API, with the same facets you already use elsewhere.

Aspect ratio on the generation form follows an attached source image.

Audio duration is visible on every surface that lists an audio asset.

The sidebar wordmark goes home in the app, not to the marketing site.

Plan Builder team-size presets fit inside the card, and the next-payment date on a subscription is the period end.

Video editor closes the generation form when it opens, and dismisses from the scrim behind it.

Browser zoom capture is scoped to the canvas, so the rest of the page doesn't steal the gesture.

Failed training shows the provider's refusal instead of a blank server error.

A status-page incident can show a banner in the sidebar while it's ongoing.

New models5

MiniMax H3 Max, board comments & Recraft styles

What's new?

MiniMax H3 Max — top-ranked video, with the sound already in it: MiniMax's post-trained H3 is now in the catalog, and it is currently at the top of the blind-preference video leaderboards (Design Arena and Artificial Analysis). What it changes for a game team:

  • Audio is generated with the picture, not dubbed on after. Footsteps, impacts, ambience, and dialogue land on the frames they belong to, in native stereo — so a trailer beat or a cutscene is watchable straight out of the generation.
  • It hits the beats in the order you wrote them. Prompt adherence is the model's headline improvement over base H3: a multi-action shot list plays back as a shot list, not as a vibe.
  • Camera moves and character look hold across a shot. Directed pushes, tracks, and pans, with a character and an art direction that stay themselves from one cut to the next — the part that usually forces a re-roll.
  • Ship-ready range. Text-to-video and image-to-video, 5–15 seconds, 480p, 768p, 2K, or 4K, six aspect ratios from 21:9 down to 9:16, and optional first-to-last-frame control for transitions you actually specify. Draft at 480p, then run the keeper at 4K.

Comments on the session board: Press C or use the toolbar to pin a thread on a tile or on the board itself. Reply, @mention teammates, resolve, and filter resolved threads — pins follow their tile as you rearrange the canvas.

Live presence and cursors: See who's in the session from the avatar stack in the header, and follow teammates' cursors on the board in real time.

Recraft V4 custom styles: Train a Recraft style from your reference images and generate with it in both raster and vector, on Standard or Pro. One training run covers both output kinds of that tier.

Background exports: Queuing a download no longer holds the dialog open. Exports run in the transfers panel alongside uploads, keep going if you close the dialog, and can be downloaded again from the finished row.

Auto-reload Creative Units: Opt in per workspace to top up automatically when the balance drops below a threshold you set, so a long session doesn't stall mid-generation. Enterprise and child workspaces stay on their existing billing.

A bigger asset viewer, without leaving the session: Opening an asset expands it into the session canvas instead of covering the whole page — header and chat stay put. In the viewer, both rails resize, the pager sits on the stage, and you can zoom to pixels.

Command palette finds everything: ⌘K searches projects, sessions, files, reference sets, workflows, and models in one ranked list. Hits on files, sets, and workflows open as shareable filtered URLs.

Video editor, faster to assemble: A CapCut-style rail for Image, Video, Audio, and Transition; right-click the timeline to add a clip at that frame; pick assets from browse and they land on a new track at the playhead. After a render, the board focuses the output.

Workflows: Segment Mesh is a node (the last 3D editor op that was missing one). And when a draft is better than the published original, Publish as new workflow forks it instead of overwriting.

Reference sets: A strength meter on the set editor shows how ready the set is to train. Artists can now create and manage sets without the training permission. Project tiles for workflows and sets offer Edit in place.

Auto aspect ratio: Models that can take their output size from the source image or first frame now offer Auto in the size picker, so image-to-video follows the source instead of snapping to a preset.

Home composer follows the file you attach: Drop a source video, mesh, or audio onto the matching tab and the composer switches to a model that can actually read it — with a one-click undo if you wanted the original model.

Bug fixes & improvements

Board layouts stick. Grouping, moving, resizing, and hiding tiles now survive leaving and returning to a session — including text-only groups and ungroup.

Stop means stop. A wedged workflow run terminates when you press Stop, instead of sitting in the session.

Downloads are named for the asset, not a random export key, and GIF originals export instead of failing the metadata path.

3D canvases size from their layout box, so posed meshes aren't clipped by a transformed rect.

Opening the viewer stops a playing thumbnail underneath it.

The composer shows the picked source video's preview frame, not a placeholder glyph.

Hidden models are no longer blocked by an opt-in workspace allowlist.

Content-policy chips on the board update live as verdicts land, and a project policy page is reachable with reference images as attachments.

Earn Free Units can show a Rewards tab for program grants, and the achievement Claim link lands on that dialog.

Skill pills in the agent drawer show which user skills are applied, and changing skill-loading mode no longer dead-ends.

Right-click in workflow node text fields opens the browser menu again; tile context menus follow the current selection.

New models10

Lip sync, video editor & Interpolate

What's new?

Lip sync, video editor & Interpolate

Lip sync from the lightbox: Drive a talking clip from a video plus an audio track, or animate a still into speech, without leaving the asset viewer. Sync.so, LatentSync, and VEED models are all available from the same control.

Video editor on the session canvas: You can now drop a video editor onto the session board and assemble clips, titles, and audio there — the same timeline that used to live only in workflows, available wherever you already work.

Interpolate and slow motion: Frame interpolation is its own tool. Convert frame rate or slow a clip down by up to 8× with Topaz Apollo, Chronos, and Aion, from the lightbox, the editor, or a workflow node.

The full Topaz lineup: Dozens of Topaz models are now available across image and video. On the image side, that's the Gigapixel precision tiers (Standard V2, High Fidelity V3, Low Resolution V2, Standard MAX), the generative ones (Redefine, Recover, Bloom, Wonder), plus CGI, Text Refine, and Transparency Upscale for game-art edge cases — with Nyx, Denoise, and Dust-Scratch for cleanup. On the video side: Proteus, Iris, Dione, Artemis, Gaia, Rhea, Theia, Astra, the Starlight family, and SDR-to-HDR.

Score assets against studio standards: A workspace or project can describe what “on-standard” looks like — in words and with reference images — and finished assets are scored 0–100 with a rationale you can act on. Sort and filter the library by score, and re-score a library when the rules change.

Input guardrails: The same policy layer can screen a prompt and its input images before a generation runs, either blocking the request or warning and letting it through. Workspace rules always apply; a project can only add to them.

Layered image editor: Undo and redo in the layered viewer, plus lock, group/ungroup, and merge — with a Photoshop-style right-click menu on layer rows. Locks round-trip through PSD protected flags.

Miro integration: Pull board items into a session as references and push finished assets back to the board — no more exporting and re-uploading by hand. An admin authorizes the connector once for the workspace, then each project points at a board. See connecting Miro to Layer.

Custom fonts in video: Render timeline text and subtitles with a FONT file from your library. Opening a font now shows a type specimen — every face, a size waterfall, and a glyph grid — instead of a flat preview.

Bug fixes & improvements

Reference-set training now shows full-lifecycle progress and an ETA through export. Compare finetune versions side by side and pick the active one. Sets created in the editor get a thumbnail and a real name instead of staying blank and “Untitled.”

Workflows gain nodes for split-into-layers and explode-layers, plus rasterize, image slice/select/grid, stitch-from-images, lip sync, and image segment.

Clips no longer vanish on save in the video editor — duration-aware placement and a safer save path keep the timeline you built.

Workspace shell no longer crashes when the backend briefly blips.

Two-finger swipe on a trackpad no longer navigates away from the canvas.

Skinned and animated meshes fit the viewer from their posed skeleton instead of clipping.

Meshy honors the animation preset you picked when rigging.

ElevenLabs Sound Effects now respects the duration you set (0.5–22s).

Edit-LoRA training no longer wipes pair data on launch, and a training run’s status follows the published finetune.

Imagen 3 and Imagen 4 have been retired after the provider shut them down.

Deleted sessions no longer appear in detail views.

New models58
Hunyuan 3D UV Unwrap3D
Tencent

Rebuilds an existing mesh's UV layout so it is ready to texture. Geometry is kept as-is and the result is untextured.

FLUX.3 Video UpscaleVideo
Black Forest Labs

FLUX.3-powered source-faithful video upscaler, up to 4K.

FLUX.3 Video Upscale CreativeVideo
Black Forest Labs

FLUX.3-powered video upscaler that adds detail, up to 4K.

Hunyuan 3D v3.1 Part3D
Tencent

Splits a 3D FBX model into separate parts with Hunyuan 3D v3.1.

Topaz Themis 2Video
Topaz Labs

Topaz motion deblur, faithful to the source footage.

Topaz Denoise ExtremeImage
Topaz Labs

Topaz's most aggressive conventional denoiser.

Topaz Denoise MaxImage
Topaz Labs

Topaz's generative denoiser, rebuilding detail as it cleans.

Topaz Denoise NormalImage
Topaz Labs

Topaz general-purpose denoising at the source resolution.

Topaz Denoise StrongImage
Topaz Labs

Topaz denoising tuned for heavier noise.

Topaz NyxVideo
Topaz Labs

Topaz's high-quality video denoiser, preserving texture.

Topaz Nyx FastVideo
Topaz Labs

Topaz's half-price video denoiser for high-volume work.

Topaz Nyx HFVideo
Topaz Labs

Topaz's precision video denoiser for pipeline-ready output.

Topaz Nyx XLVideo
Topaz Labs

Topaz's video denoiser tuned for extreme noise.

Bria FIBO Edit Restore 1.0Image
Bria AI

Bria's one-shot photo restoration, trained on licensed data.

Topaz Dust-Scratch V2Image
Topaz Labs

Topaz film clean-up for dust and scratches.

Topaz Recover 3 RestoreImage
Topaz Labs

Topaz's generative restoration, rebuilding natural detail.

Topaz SDR to HDRVideo
Topaz Labs

Topaz SDR-to-HDR conversion for grading and mastering.

Hi3D v3.03D
Hitem3D

Hi3D v3.0 image-to-3D at 2048³ with PBR materials and face counts up to 5M.

Topaz AionVideo
Topaz Labs

Topaz's frame interpolator for extreme and complex motion.

Topaz ApolloVideo
Topaz Labs

Topaz's general-purpose frame interpolator, retiming footage up to 120fps.

Topaz ChronosVideo
Topaz Labs

Topaz's frame interpolator tuned for real-world motion.

Uthana Animate3D
Uthana

Retargets a Uthana motion — trained, or generated from a prompt or video — onto your character.

Uthana Auto-Rig Character3D
Uthana

Uthana's automatic character rigging model, preparing an uploaded mesh with a skeleton.

Topaz Bloom 2Image
Topaz Labs

Topaz's creative image upscaler, reinventing detail with an adjustable creativity dial.

Topaz Bloom RealismImage
Topaz Labs

Topaz's creative upscaler with its invented detail biased toward photorealism.

Topaz CGIImage
Topaz Labs

Topaz's precision upscaler for rendered art and CG rather than photographs.

Topaz High Fidelity V3Image
Topaz Labs

Topaz's detail-preserving precision upscaler for professional photography.

Topaz Low Resolution V2Image
Topaz Labs

Topaz's precision upscaler tuned for small, compressed sources.

Topaz Recover 3Image
Topaz Labs

Topaz's restorative upscaler, rebuilding natural detail in damaged images.

Topaz Recovery V2Image
Topaz Labs

Topaz's upscaler for extremely low-resolution sources.

Topaz RedefineImage
Topaz Labs

Topaz's prompt-guided generative upscaler.

Topaz Standard MAXImage
Topaz Labs

Topaz's precision-leaning generative upscaler.

Topaz Standard V2Image
Topaz Labs

Topaz Gigapixel precision upscaling, faithful to the original image.

Topaz Text RefineImage
Topaz Labs

Topaz's precision upscaler that keeps text and hard shapes crisp.

Topaz Transparency UpscaleImage
Topaz Labs

Topaz upscaling that preserves the alpha channel end to end.

Topaz Wonder 3.5Image
Topaz Labs

Topaz's generative image upscaler, rebuilding detail with fewer repeated patterns.

Topaz Artemis High QualityVideo
Topaz Labs

Topaz's precision upscaler that denoises and sharpens as it enlarges.

Topaz Astra 2Video
Topaz Labs

Topaz's creative video upscaler, reinventing detail for maximum visual impact.

Topaz Dione DVVideo
Topaz Labs

Topaz's deinterlacing precision upscaler for legacy tape sources.

Topaz Gaia 2Video
Topaz Labs

Topaz's animation-tuned precision upscaler, and its cheapest tier.

Topaz Gaia CGVideo
Topaz Labs

Topaz's precision upscaler for rendered and CG video.

Topaz Gaia HQVideo
Topaz Labs

Topaz's precision upscaler for refining already-clean footage.

Topaz IrisVideo
Topaz Labs

Topaz's precision upscaler specialised in recovering facial detail.

Topaz ProteusVideo
Topaz Labs

Topaz's precision video upscaler, enhancing detail without inventing it.

Topaz Proteus NaturalVideo
Topaz Labs

Topaz's softer precision upscaler, tuned to look untouched.

Topaz RheaVideo
Topaz Labs

Topaz's maximum-detail precision upscaler.

Topaz Starlight Fast 2Video
Topaz Labs

Topaz's fastest and cheapest generative video upscaler.

Topaz Starlight HQVideo
Topaz Labs

Topaz's highest-quality generative video upscaler.

Topaz Starlight MiniVideo
Topaz Labs

Topaz's generative video upscaler tuned for archival restoration.

Topaz Starlight Precise 2.6Video
Topaz Labs

Topaz's generative video upscaler, rebuilding detail that the source never had.

Topaz Starlight SharpVideo
Topaz Labs

Topaz's faster generative video upscaler with sharper output.

Topaz Theia Fine Tune DetailVideo
Topaz Labs

Topaz's manually tuned precision upscaler, biased toward detail.

LatentSyncVideo
ByteDance

Fast, affordable video-to-video lipsync that syncs mouth motion to any audio track.

OmniHuman 1.5Video
ByteDance

ByteDance's latest OmniHuman, bringing a still photo to life from audio with improved motion and expression.

sync.so Lipsync 2 ProVideo
Sync

High-quality realistic lipsync that preserves unique facial details from any new audio track.

sync.so Lipsync 3Video
Sync

sync.so's most powerful lipsync model, re-articulating a source video to any new audio track.

sync.so Avatar (Image to Video)Video
Sync

Turns a single still image into a talking character lip-synced to a voice track.

VEED Lipsync 2.0Video
VEED

Production-quality lipsync that re-articulates a source video to match any new audio track.

Decompose layers, Wan 3.0 & Meshy V7

What's new?

Decompose layers, Wan 3.0 & Meshy V7

Decompose complex imagery — multi-layer PNG and PSD export: You can now split a finished image into independently editable layers and export them as a multi-layer PNG or PSD. Buildings, props, and background come apart as labeled cutouts you can hide, refine, or composite without re-generating the scene.

Wan 3.0: Alibaba's Wan video line now produces clips of up to 30 seconds at 480p, 720p, or 1080p with optional native audio, plus first- and last-frame control so a single generation can carry a full shot rather than a silent few seconds.

Meshy V7: Image-to-3D — including multi-view from up to four angles — now produces PBR-ready meshes with game-ready topology, optional auto-rigging, and animation presets.

MiniMax Music 3: Generate complete songs up to five minutes from a music description and lyrics, and shape the arrangement with structure tags such as verse, chorus, bridge, and outro.

LTX Video 2.5: Lightricks' LTX 2.5 Pro and Fast are now available for text-to-video and image-to-video, including audio-to-video variants.

More image and video models: Grok Imagine Image 2.0, HiDream O1, Krea 2 Large, Ideogram V4, MAI Image 2.5, Wan 2.7 Pro, and Vidu Q3 Pro are now in the catalog.

New models18
Auto CaptionVideo
fal
VEED SubtitlesVideo
VEED
MiniMax Music 3Audio
MiniMax

High-performance MiniMax music model for complete songs up to five minutes with structure tags and seed control.

Wan 3.0Video
Alibaba

Alibaba's Wan 3.0 video model with up to 30s clips, 480p/720p/1080p output, and native audio.

Wan 3.0 ReferenceVideo
Alibaba

Reference-guided Wan 3.0 video with up to 10 images, 5 videos, and 5 audio files.

Grok Imagine Image 2.0Image
xAI

xAI's Grok Imagine Image 2.0 model for high-quality text-to-image generation and editing.

HiDream O1 Image 1.0Image
Hidream

HiDream's full O1 Image model — create, edit, and personalize images up to 2K in one native model.

LTX Video 2.5 FastVideo
Lightricks

Speed-optimized LTX 2.5 audio-video model with 720p–4K output and clips up to 20 seconds.

LTX Video 2.5 Fast Audio-to-VideoVideo
Lightricks

Speed-optimized LTX 2.5 mode that generates video timed to a supplied audio clip.

LTX Video 2.5 ProVideo
Lightricks

Quality-optimized LTX 2.5 audio-video model for high-fidelity 720p and 1080p output.

LTX Video 2.5 Pro Audio-to-VideoVideo
Lightricks

Quality-optimized LTX 2.5 mode for final visuals synchronized to music, dialogue, or soundtrack.

MAI Image 2.5Image
Microsoft

Microsoft's photorealistic image model with strong typography and native editing.

Meshy V73D
Meshy

Meshy V7. Text, image, and multi-image to 3D with PBR, quad mesh, and optional rigging.

Vidu Q3 ProVideo
Vidu

Vidu's Q3 Pro video model with up to 16s clips, 720p/1080p output, and native audio.

HiDream O1 Image DevImage
Hidream

HiDream's distilled O1 Image Dev model — create, edit, and personalize images up to 2K in one native model.

Ideogram V4Image
Ideogram

Ideogram's latest text-to-image model with crisp visuals, accurate text, and image-to-image support.

Krea 2 LargeImage
Krea

Krea's flagship text-to-image model for high-fidelity generations with distinctive aesthetic range.

Wan 2.7 ProImage
Alibaba

Alibaba WAN 2.7 Pro image model for high-quality text-to-image generation and multi-image editing.

Seedance 2.5

New models

Seedance 2.5 (1 of 2)

Seedance 2.5: ByteDance's Seedance video line takes a big step forward. You can now generate clips of up to 30 seconds with synchronized native audio — including lip-synced speech — so a single generation can carry a full spoken beat rather than a silent few seconds. Seedance 2.5 supports both text-to-video and image-to-video with first- and last-frame control for tighter shot composition, renders at 480p or 720p across cinematic aspect ratios (16:9, 9:16, 1:1, 4:3, and 21:9), and lets you dial in any duration from 4 to 30 seconds. It's built for advertising, social, and short-form production, where longer takes and on-screen speech are what make a clip usable.

FLUX.3 Video Extend and Grok Imagine Video Extend: You can now extend the duration of a generated video using FLUX.3 Extend directly from the video lightbox, with Grok Imagine Video Extend as a second extension option — giving you two distinct paths for lengthening clips without re-generating from scratch.

What else is new?

Seedance 2.5 (2 of 2)

Draw-video: You can now sketch motion arrows directly onto a still image to generate an on-device reference clip, which feeds into video generation slots as a motion reference — giving you direct control over how elements in a scene should move.

Hide assets instead of delete: Assets on your board can now be hidden rather than permanently deleted, so you can keep your workspace tidy without losing work you might want to revisit.

Thinking and web search on Gemini and Luma models: Thinking level controls and web search are now available when using Gemini and Luma models in the app, giving you more control over reasoning depth and the ability to ground generations in current information.

@Reference chips now support video and audio: When referencing assets in the composer using @mentions, video and audio files are now supported alongside images — so any asset type in your library can be brought into a generation as a reference.

Drag and drop into composer: You can now drag assets directly into the composer input rather than uploading or selecting from a picker, making it faster to bring existing assets into a new generation.

Home quick actions: Edit tools are now accessible directly under the home composer, so common actions are one step closer without navigating into a specific session or asset view.

Bug fixes & improvements

Gemini restored in Inpaint and Reframe: Gemini models are now available again as options inside Inpaint and Reframe workflows after being incorrectly excluded in a recent update.

Clearer inference error detail: Error messages returned from failed inference runs now include more specific detail about what went wrong, making it easier to diagnose and retry without guesswork.

Large workspace member search: Searching for members in workspaces with large team sizes now works correctly, resolving a performance issue that caused the search to return incomplete or no results.

Rodin Gen-2 and Gen-2.5 style fix: Resolved a 500 error caused by a style-path unpacking issue in Hyper3D Rodin Gen-2 and Gen-2.5 generation requests.

New models23
FLUX Pro VTOImage
Black Forest Labs

Virtual try-on from Black Forest Labs: dress a person photo with a garment reference.

Grok Imagine Video ExtendVideo
xAI

Extend an existing video with xAI's Grok Imagine, continuing motion from the source ending.

SAM 3.1 ImageImage
Meta

Meta's SAM 3.1 for fast multi-object image segmentation.

SAM 3.1 VideoVideo
Meta

Meta's SAM 3.1 for multi-object video segmentation and tracking.

SAM 3D Align 1.03D
Meta

Meta SAM 3D Align places body and object meshes into a spatially coherent scene from one image.

SAM 3D Body 1.03D
Meta

Meta SAM 3D Body reconstructs accurate human body shape and pose from a single image.

SAM 3D Objects 1.03D
Meta

Meta SAM 3D Objects reconstructs textured 3D geometry from a single real-world image.

Rodin Bang3D
Hyper3D

Segments a 3D mesh into parts with Hyper3D Bang!

Rodin Gen-23D
Hyper3D

Hyper3D Rodin Gen-2 delivers sharper geometry and cleaner textures from a single image or prompt.

Rodin Gen-2.53D
Hyper3D

Hyper3D Rodin Gen-2.5 pushes structural detail and surface quality further, with selectable quality.

Sonilo Music 1.1Audio
Sonilo

Licensed, commercial-use-safe music from a text prompt with exact duration control.

Sonilo Sound Effects 1.1Audio
Sonilo

High-quality, commercial-use-safe sound effects from text with exact duration control.

Sonilo Video to Music 1.1Audio
Sonilo

Frame-synced, licensed music scored from video pacing, mood, and timing.

Sonilo Video to Sound Effects 1.1Audio
Sonilo

Synchronized, royalty-free sound effects timed to actions visible in a video.

Sonilo Video Music 1.1Video
Sonilo

Mux frame-synced, licensed music onto any video; optionally keep original speech.

Sonilo Video Sound Effects 1.1Video
Sonilo

Add synchronized, royalty-free sound effects mixed into the finished video.

Hi3D3D
Hitem3D

Hi3D image-to-3D with multi-view support, PBR materials, and configurable face counts.

Hi3D Parts3D
Hitem3D

Splits a 3D model into parts for editing, printing, or modular asset workflows.

Hi3D Texture3D
Hitem3D

Textures an existing Hi3D-compatible geometry mesh from a reference image.

Seedance 2.5Video
ByteDance

ByteDance's Seedance 2.5 video model with up to 30s clips, 480p/720p/1080p output, and native audio.

Seedance 2.5 ReferenceVideo
ByteDance

Reference-guided Seedance 2.5 video with up to 30 images, 10 videos, and 10 audio files.

Seedream 5.0 Pro LayerizeImage
ByteDance

Split a finished image into independently editable layers with Seedream 5.0 Pro Layerize.

FLUX.3 ExtendVideo
Black Forest Labs

Extend an existing video clip with FLUX.3, continuing motion and native audio from the source ending.

MiniMax H3, Tripo 3D mesh ops & parallel sub-agents

What's new?

MiniMax H3, Tripo 3D mesh ops & parallel sub-agents (1 of 4)

New video model — MiniMax H3: MiniMax H3 is now available for text-to-video, image-to-video, and reference-to-video generation at stereo 2K resolution, giving you a high-quality video option with strong motion consistency across all three input modes.

AI controls apply down workspace hierarchy: AI settings configured at the workspace level now propagate down to sub-workspaces automatically, so teams managing multiple workspaces no longer need to configure controls in each one separately.

SVG viewer and Rasterize tool: SVG assets now open in a dedicated viewer with SVG-aware tools, and a new Rasterize option lets you convert any SVG to a flat image in one click when you need it as a raster format downstream.

MiniMax H3, Tripo 3D mesh ops & parallel sub-agents (2 of 4)

New 3D model suite — Tripo mesh operations: Six new Tripo operations are now available — Rig, Animate, Retarget, Segment, Remesh, and Retexture — alongside a multi-part mesh viewer. You can now take a generated 3D asset through a full production pipeline without leaving Layer.

Flexible groups: Groups now support configurable roles, usage limits, and member management, giving workspace admins finer control over how teams are organized and what they can access.

Session date filters: You can now filter your session board by date server-side, making it faster to navigate large session histories without loading everything first.

MiniMax H3, Tripo 3D mesh ops & parallel sub-agents (3 of 4)

Parallel sub-agent streams: The agent can now run multiple sub-tasks in parallel within a single chat turn, so complex multi-step requests complete significantly faster instead of running sequentially.

Object editor mesh operations: The 3D object editor now supports mesh operations directly, with multi-view INIT slots so you can work with complex assets from multiple angles in the same session.

Aspect ratio picker for edit-only image models: Edit-only image models now expose an aspect ratio picker, giving you output size control that was previously only available on generative models.

Workflow node descriptions: You can now add descriptions to individual nodes inside a workflow, making it easier to document intent and hand off complex graphs to other team members.

Edit workflow metadata outside publish: Workflow name, description, cover image, and node descriptions can now be updated without going through the publish flow, so keeping documentation current doesn't require a full republish.

Canvas tile placement: New tiles generated on the canvas now place beside your current selection rather than at a fixed position, keeping your workspace organised as it grows.

Bug fixes & improvements

MiniMax H3, Tripo 3D mesh ops & parallel sub-agents (4 of 4)

Workspace-disabled models hidden from pickers: Models that have been disabled at the workspace level no longer appear in model selection dropdowns, reducing confusion when certain models are intentionally restricted.

Video and mesh frame picks restored: Fixed an issue where selecting a specific frame from a video or mesh as a reference image input was not working correctly.

Hung blueprint runs fixed: Resolved a bug causing blueprint runs to hang indefinitely rather than completing or failing cleanly.

Chat stream recovery: The chat interface now recovers gracefully from mid-turn stream interruptions rather than leaving a generation in an unresolved state.

Library grid performance: Scrolling and loading performance in the asset library grid has been improved for workspaces with large asset counts.

New models11

Ready to scale your creative production?