Floyo
Floyo
Workflows
API
Pricing
Floyo
Floyo
Workflows
API
Pricing

floyoofficial

Bio under construction. Expect wild opinions & mistakes. Always learning, iterating. Here for good prompts, great lighting & snacks.

OG badge
OG badge

1282

Total Likes

567193

Total Views

527

My Workflows

AiVideo

API

image to video

Video

video generation

wan 2.5

Wan 2.5: Image to Video with Audio

27.1k

Z-Image Turbo · Text to Image

Image

Marketing

Photography

Production

Text2Image

z-image

Z-Image Turbo

Fast Image Generation in Seconds

Z-Image Turbo · Text to Image

Fast Image Generation in Seconds

Wan 2.2 14B: Image to Video + End Frame

image to video

lora

LoRAs

Video

Video Generation

wan 2.2

Generate high quality video from a start frame, as well as an optional end frame with this Wan2.2 14b Image to Video workflow!

Wan 2.2 14B: Image to Video + End Frame

Generate high quality video from a start frame, as well as an optional end frame with this Wan2.2 14b Image to Video workflow!

Qwen Image Edit 2509: Build a LoRA Dataset

character consistency

Dataset

Image

Image to Image

LoRA

LoRAs

Qwen Image Edit 2509

Create Character LoRA Dataset

Qwen Image Edit 2509: Build a LoRA Dataset

Create Character LoRA Dataset

 Nano Banana Pro: Generate & Edit Images

API

gemini 3 pro

Image

Image2Image

typography

Google just released Nano Banana Pro, and honestly, it's a pretty big step up from the original Nano Banana. The main thing? It can actually put legible text in images now. Like, real text that you can read, not the garbled nonsense most AI models spit out.

Nano Banana Pro: Generate & Edit Images

Google just released Nano Banana Pro, and honestly, it's a pretty big step up from the original Nano Banana. The main thing? It can actually put legible text in images now. Like, real text that you can read, not the garbled nonsense most AI models spit out.

Z-Image Turbo · Text or Image to Image

ai image generator

Image

Image to Image

Text to Image

z-image

Z-Image Turbo

Z-Image Turbo is Alibaba's open-source 6B model that turns a text prompt into photorealistic images in about 8 steps. Type a prompt, or rework an image.

Z-Image Turbo · Text or Image to Image

Z-Image Turbo is Alibaba's open-source 6B model that turns a text prompt into photorealistic images in about 8 steps. Type a prompt, or rework an image.

Wan 2.1 FusionX: Cinematic Image to Video

FusionX

Image to Video

Video

Video Generation

Wan

Created by @vrgamedevgirl on Civitai, please support the original creator!

Wan 2.1 FusionX: Cinematic Image to Video

Created by @vrgamedevgirl on Civitai, please support the original creator!

Wan 2.6 Reference to Video

VFX

Video

Video2Video

Video Production

Wan2.6

Wan 2.6 Reference to Video

Fast LoRA Training for Flux via Floyo API

API

Flux

LoRAs

LoRa Training

FLUX is great at generating images, but locking in a specific aesthetic or character is easier with a  LoRA. Here's how to create your own.

Fast LoRA Training for Flux via Floyo API

FLUX is great at generating images, but locking in a specific aesthetic or character is easier with a  LoRA. Here's how to create your own.

LTX 2.3 Pro Image to Video

API

Audio

Image to Video

LTX2.3

Video

LTX 2.3

LTX 2.3 Pro Image to Video

LTX 2.3

Grok Imagine: Image to Video with Audio

grok imagine

image to video

Video

video generation

Turn images into excellent video using the Grok Imagine

Grok Imagine: Image to Video with Audio

Turn images into excellent video using the Grok Imagine

SeedVR2 · Image Upscaler

API

Image

Image2Image

image restoration

SeedVR2

SeedVR Upscale

super resolution

Upscale to Extreme Clarity

SeedVR2 · Image Upscaler

Upscale to Extreme Clarity

Wan2.2 Animate Character

Animate

Animation

Filmmaking

LoRAs

Video

Video to Video

Wan2.2

Wan 2.2

Wan2.2 Animate Character

Wan 2.2

Flux Kontext Sketch to LineArt + Color Previz

Flux Kontext

Image

Lineart

Previz

Sketch to Image

Quickly convert rough sketches into polished lineart and colorized concepts. Ideal for early storyboards, character designs, scene planning, and other visual explorations.

Flux Kontext Sketch to LineArt + Color Previz

Quickly convert rough sketches into polished lineart and colorized concepts. Ideal for early storyboards, character designs, scene planning, and other visual explorations.

Flux Text to Character Sheet

Character Sheet

Controlnet

Flux

Image

Create a character and a range of consistent outputs suitable for establishing character consistency, training a model, and ensuring consistency throughout multiple scenes. Key Inputs Image reference: Use the included pose sheet to show range of positions Prompt: as descriptive a prompt as possible

Flux Text to Character Sheet

Create a character and a range of consistent outputs suitable for establishing character consistency, training a model, and ensuring consistency throughout multiple scenes. Key Inputs Image reference: Use the included pose sheet to show range of positions Prompt: as descriptive a prompt as possible

Qwen Image Edit 2509: Face Swap

Face Swap

Image

Image to Image

Lora

LoRAs

Portrait

qwen image edit 2509

Face Swap and Inpainting

Qwen Image Edit 2509: Face Swap

Face Swap and Inpainting

Image to Character Spin

360

Image2Video

Video

Wan2.1

See an image of a character spin 360 degrees. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Width & height: Default resolution settings are noted: Default image resize resolution works best for portrait images, if the image is landscape change from 480x832 to 832x480 Prompt: Follow example format: The video shows (describe the subject), performs a r0t4tion 360 degrees rotation. Denoise: The amount of variance in the new image. Higher has more variance. File Format: H.264 and more

Image to Character Spin

See an image of a character spin 360 degrees. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Width & height: Default resolution settings are noted: Default image resize resolution works best for portrait images, if the image is landscape change from 480x832 to 832x480 Prompt: Follow example format: The video shows (describe the subject), performs a r0t4tion 360 degrees rotation. Denoise: The amount of variance in the new image. Higher has more variance. File Format: H.264 and more

Z-Image Turbo: ControlNet Image to Image

Controlnet

Depth

Image

Image2Image

Photography

Portrait

Pose Control

Z-Image-Turbo

Image to Image

Z-Image Turbo: ControlNet Image to Image

Image to Image

Qwen 2509: Combine Multiple Images Into One Scene

2509

Image

Image2Image

multi-image

product photography

Qwen

Qwen 2509: Combine Multiple Images Into One Scene

Kling 2.6 Motion Control: Animate Images

Animation

Image to Video

Kling 2.6

motion transfer

Video

Create an excellent for movement for your characters using Kling 2.6 Standard Motion Control

Kling 2.6 Motion Control: Animate Images

Create an excellent for movement for your characters using Kling 2.6 Standard Motion Control

LTX-2 19B Fast: Text to Video + Audio

Audio

Filmmaking

LoRAs

LTX 2

LTX 2 Fast

Open Source

Text2Video

Video

Videography

A text video model using LTX 2

LTX-2 19B Fast: Text to Video + Audio

A text video model using LTX 2

Wan2.1 and VACE for Video to Video Outpainting

Outpainting

Video

Video to Video

Wan

Wan VACE video outpainting invites you to break free from the limits of the frame and explore endless creative possibilities.

Wan2.1 and VACE for Video to Video Outpainting

Wan VACE video outpainting invites you to break free from the limits of the frame and explore endless creative possibilities.

Qwen Image Edit - Edit Image Easily

Image

Image2Image

Qwen

Qwen Image Edit

Qwen Image Edit - Edit Image Easily

ComfyUI Flux LoRA Trainer

Flux

Image

LoRAs

LORA Training

Created by @Kijai on Github, please support the original creator!

ComfyUI Flux LoRA Trainer

Created by @Kijai on Github, please support the original creator!

LTX 2 Pro: Cinematic Image to Video

Animation

Audio

Filmmaking

Image2Video

LTX 2 Pro

Video

Video Editing

video with audio

Outdated model. Please go to LTX 2.3 Image to video workflow to use LTX 2.3

LTX 2 Pro: Cinematic Image to Video

Outdated model. Please go to LTX 2.3 Image to video workflow to use LTX 2.3

Nano Banana 2 · Image Generation & Editing

API

gemini flash image

Image

Image2Image

nano banana 2

Text2Image

typography

The top-ranked image model on Artificial Analysis and LM Arena. 4K output, text rendering, and subject consistency across 5 characters.

Nano Banana 2 · Image Generation & Editing

The top-ranked image model on Artificial Analysis and LM Arena. 4K output, text rendering, and subject consistency across 5 characters.

Multi-Image  Flux Ultra, Pro, Dev, Recraft+

API

Flux

Image

Text to Image

Start with a prompt, and get a different render from a range of unique models at the same time.

Multi-Image Flux Ultra, Pro, Dev, Recraft+

Start with a prompt, and get a different render from a range of unique models at the same time.

Flux Kontext - Sketch to Image

Flux

Image

Kontext

Sketch to Image

Bring your sketches to life in full color with Flux Kontext! Key Inputs Load Image – Upload the sketch you want to transform. Prompt – Describe the desired output style, such as: “Render this sketch as a realistic photo” or “Turn this sketch into a watercolor painting.”

Flux Kontext - Sketch to Image

Bring your sketches to life in full color with Flux Kontext! Key Inputs Load Image – Upload the sketch you want to transform. Prompt – Describe the desired output style, such as: “Render this sketch as a realistic photo” or “Turn this sketch into a watercolor painting.”

Wan 2.6: Multi-Shot Image to Video

Animation

Film

Image2Video

VFX

Video

Wan2.6

Turn a still into a multi-shot clip with audio using Wan 2.6 by Alibaba. Upload an image, describe the scene, hit run, and get a 720P video with synced sound.

Wan 2.6: Multi-Shot Image to Video

Turn a still into a multi-shot clip with audio using Wan 2.6 by Alibaba. Upload an image, describe the scene, hit run, and get a 720P video with synced sound.

Flux Character LoRA Test and Compare

Animation

Filmmaking

Flux

Game Development

Image

LoRA

LoRAs

Text to Image

Test and compare multiple epochs of a character LoRA side by side with preset prompts When training a LoRA, you'll usually have a few checkpoints throughout the process to test. This workflow lets you load up to 4 LoRAs to test side by side, making it easier to determine which one is right for you! Key Inputs: LoRA Loaders: Load each LoRA epoch for the same character in up to 4 groups. Groups Bypasser: Enable/disable groups as needed. If you only have 2 epochs to test, disable the back 2 groups! Triggerword: Simply add the trigger word for your LoRA and it will auto-fill in the default prompts. Leave blank if you're using your own custom prompts that include the trigger word. LoRA Testing Prompts: Default prompts work well to get an idea of how your character will look in different situations, but feel free to replace them with your own prompts (max 4).

Flux Character LoRA Test and Compare

Test and compare multiple epochs of a character LoRA side by side with preset prompts When training a LoRA, you'll usually have a few checkpoints throughout the process to test. This workflow lets you load up to 4 LoRAs to test side by side, making it easier to determine which one is right for you! Key Inputs: LoRA Loaders: Load each LoRA epoch for the same character in up to 4 groups. Groups Bypasser: Enable/disable groups as needed. If you only have 2 epochs to test, disable the back 2 groups! Triggerword: Simply add the trigger word for your LoRA and it will auto-fill in the default prompts. Leave blank if you're using your own custom prompts that include the trigger word. LoRA Testing Prompts: Default prompts work well to get an idea of how your character will look in different situations, but feel free to replace them with your own prompts (max 4).

Flux Dev: Text to Image + Image Input

Flux Dev

Image

image to image

photorealism

text to image

Flux Dev: Text to Image + Image Input

Flux Kontext and HD360 LoRA for 360 Degree View

Flux

Flux Kontext

Image

Image2Image

kontext

LoRAs

panorama

Flux Kontext 360° Workflow - Seamless Panorama Generation Input: Simply upload an image in the "Load Image from Outputs" node Output: A 360° Panoramic image

Flux Kontext and HD360 LoRA for 360 Degree View

Flux Kontext 360° Workflow - Seamless Panorama Generation Input: Simply upload an image in the "Load Image from Outputs" node Output: A 360° Panoramic image

Video to Video with Camera Control with Wan

LoRAs

[Video]

Video

Adjust the camera angle of an existing video, like magic.

Video to Video with Camera Control with Wan

Adjust the camera angle of an existing video, like magic.

FLUX.2 Klein 9B: Edit Images by Prompt

Flux

Flux.2 Klein

Image

Image2image

instruction editing

multi-image

Unified workflow: one model for text‑to‑image, image‑to‑image, and image editing

FLUX.2 Klein 9B: Edit Images by Prompt

Unified workflow: one model for text‑to‑image, image‑to‑image, and image editing

Wan2.1 Fun Control and Flux for V2V Restyle

Controlnet

Flux

Video

Video2Video

Wan2.1

Create a new video by restyling an existing video with a reference image.

Wan2.1 Fun Control and Flux for V2V Restyle

Create a new video by restyling an existing video with a reference image.

Image-to-Video with Reference Video (Prompt-Based Camera Rotation)

9:16

camera rotation

DWpose

image to video

pose control

reference video

Video

Wan2.2

Image-to-Video with Reference Video (Prompt-Based Camera Rotation)

Audio

hailuo 3.0

image to video

minimax h3

text to video

Video

video with audio

Generate 2K video with stereo sound from a start image and a prompt using MiniMax H3 (Hailuo 3.0), the open-weights model. Mute the image to go text-only.

MiniMax H3 Open Weights · Image & Text to Video

Generate 2K video with stereo sound from a start image and a prompt using MiniMax H3 (Hailuo 3.0), the open-weights model. Mute the image to go text-only.

360° Character Turnaround & Sheet Workflow

360 TurnAround

Image

NanoBanana

360° Character Turnaround & Sheet Workflow

Seedance 2.0 - Text to Video

seedance

seedance 2.0

text to video

Video

video generation

Generate up to 15-second videos with native audio from a text prompt using ByteDance's Seedance 2.0. Pick your aspect ratio, resolution, and duration.

Seedance 2.0 - Text to Video

Generate up to 15-second videos with native audio from a text prompt using ByteDance's Seedance 2.0. Pick your aspect ratio, resolution, and duration.

Flux Kontext - Quick & Easy

flux

flux kontext

Image

kontext

sebastian kamph

Load an image reference and use the smart Flux Kontext model to ask for anything. The model understands natural language and looks at your input image. Example: Put this man on a tropical island. This man is sleeping in a bed.

Flux Kontext - Quick & Easy

Load an image reference and use the smart Flux Kontext model to ask for anything. The model understands natural language and looks at your input image. Example: Put this man on a tropical island. This man is sleeping in a bed.

Seedance 1.5 Pro with Draft Mode

API

Floyo API

Image to Video

Seedance 1.5 Pro

Video

Draft mode lets you first experiment at a low cost by generating 480p draft videos

Seedance 1.5 Pro with Draft Mode

Draft mode lets you first experiment at a low cost by generating 480p draft videos

Z-Image Base: High-Detail Text to Image

concept art

Fine-tuning

Image

Text2Image

Z-Image

Z-image-base

Create sunning images using z-image base model (non distlled).

Z-Image Base: High-Detail Text to Image

Create sunning images using z-image base model (non distlled).

Wan2.1 and RecamMaster for V2V Camera Control

LoRAs

Recammaster

Video

Video to Video

Wan

Adjust the camera angle of an existing video, like magic.

Wan2.1 and RecamMaster for V2V Camera Control

Adjust the camera angle of an existing video, like magic.

FLUX.2 Klein 4B and LanPaint for Swap Clothes

Clothes Swap

Flux

Flux.2 Klein

Image

Image Editing

Image to Image

virtual try-on

Replace clothes using the Flux.2 Klein 4B

FLUX.2 Klein 4B and LanPaint for Swap Clothes

Replace clothes using the Flux.2 Klein 4B

Character + Outfit → High-End Editorial Shoot

character-to-photoshoot

Image

Nano banana

studio-photoshoot

Character + Outfit → High-End Editorial Shoot

VibeVoice: Single-Speaker Text to Speech

Audio

text to speech

TTS

VibeVoice

voice cloning

VibeVoice

VibeVoice: Single-Speaker Text to Speech

VibeVoice

Text to Image with Multi-LoRA

Flux

Image

LoRa

LoRAs

Text2Image

Create consistent images with multiple LoRA models.

Text to Image with Multi-LoRA

Create consistent images with multiple LoRA models.

Wan2.1 FusionX and MultiTalk - Image to Video

Animation

Audio

Filmmaking

Image to Video

Lipsync

Marketing

Multitalk

Video

Wan2.1

Turn any portrait - artwork, photos, or digital characters - into speaking, expressive videos that sync perfectly with audio input. MultiTalk handles lip movements, facial expressions, and body motion automatically.

Wan2.1 FusionX and MultiTalk - Image to Video

Turn any portrait - artwork, photos, or digital characters - into speaking, expressive videos that sync perfectly with audio input. MultiTalk handles lip movements, facial expressions, and body motion automatically.

FlashVSR Upscale Your Videos Instantly

FlashVSR

Upscale

Video

Video2Video

FlashVSR Upscale Your Videos Instantly

Flux Image Upscaler with UltimateSD

Flux

Image

UltimateSD

Upscale

A simple workflow to enlarge & add detail to an existing image. Key Inputs Image: Use any JPG or PNG Upscale by: The factor of magnification Denoise: The amount of variance in the new image. Higher has more variance.

Flux Image Upscaler with UltimateSD

A simple workflow to enlarge & add detail to an existing image. Key Inputs Image: Use any JPG or PNG Upscale by: The factor of magnification Denoise: The amount of variance in the new image. Higher has more variance.

Wan2.1 Start & End Frame Image to Video

Image2Video

Start and end frame

Video

Wan2.1

Used for image to video generation, defined by the first frame and end frame images.

Wan2.1 Start & End Frame Image to Video

Used for image to video generation, defined by the first frame and end frame images.

Flux Text to Image

Flux

Image

Text2Image

Create original images using only text prompts, which can be simple or elaborate. Key Inputs Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted

Flux Text to Image

Create original images using only text prompts, which can be simple or elaborate. Key Inputs Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted

Wan 2.7 Reference to Video with Motion Control

character design

consistency

film production

image to video

Video

video generation

wan

Wan 2.7 Reference to Video with Motion Control

Wan 2.7 Reference to Video with Motion Control

Wan 2.7 Reference to Video with Motion Control

Seedream 5.0 Lite · Generate, Edit and Fuse

Image

Image2Image

Image Editing

Seedream 5.0

Text2Image

Generate, edit, or blend images with Seedream 5.0 Lite, the ByteDance model that reasons through your instruction before drawing. Prompt it and hit run.

Seedream 5.0 Lite · Generate, Edit and Fuse

Generate, edit, or blend images with Seedream 5.0 Lite, the ByteDance model that reasons through your instruction before drawing. Prompt it and hit run.

Qwen Image Edit 2511: Composite a Photoshoot

composite

Image

Image to Image

portrait

product photography

qwen image edit 251

Reference Image

Drop a person into any background with the lighting you choose, using Qwen Image Edit 2511. Upload three images, hit run, and keep the identity intact.

Qwen Image Edit 2511: Composite a Photoshoot

Drop a person into any background with the lighting you choose, using Qwen Image Edit 2511. Upload three images, hit run, and keep the identity intact.

Image to Image with Flux ControlNet

Controlnet

Flux

Image

Transform your images into something completely new, yet retaining specific details and composition from your original using flexible controls. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Prompt: as descriptive a prompt as possible Denoise Strength: The amount of variance in the new image. Higher has more variance. Width & height: Try and match the aspect ratio of the original if possible.

Image to Image with Flux ControlNet

Transform your images into something completely new, yet retaining specific details and composition from your original using flexible controls. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Prompt: as descriptive a prompt as possible Denoise Strength: The amount of variance in the new image. Higher has more variance. Width & height: Try and match the aspect ratio of the original if possible.

Flux 2 Text-to-Image Generation

Flux2

Image

Marketing

Photography

Text2Image

Flux 2 Text-to-Image Generation

Nano Banana Pro for Multi Grid View of Product Ads

API

Ecommerce

Image

Image2Image

Nano Banana Pro

Product Ads

Create grids of different angles for your ecommerce products.

Nano Banana Pro for Multi Grid View of Product Ads

Create grids of different angles for your ecommerce products.

Flux Outfit Transfer

Ace+

Fashion

Flux

Image

Image to Image

Virtual Try-on

Virtual Outfit Try-On with Auto Segmentation Try virtual clothing on any subject using Flux Dev, Ace Plus, and Redux, with automatic segmentation. Great for concept previews, fashion mockups, or character styling. Key Inputs Outfit: Load the outfit image you want to apply. Make sure it's high quality — visible artifacts or distortions may carry over into the final result. Actor: Add the subject or character you want to dress. Ideally, use a clear, front-facing image. Human Parts Ultra: Choose which parts of the body the clothing should apply to. For example, for a long-sleeve shirt, select: torso, left arm, and right arm. This helps the model align the clothing properly during generation. Prompt: Default value works for most outfits, however you may try to adjust it to describe the desired outfit.

Flux Outfit Transfer

Virtual Outfit Try-On with Auto Segmentation Try virtual clothing on any subject using Flux Dev, Ace Plus, and Redux, with automatic segmentation. Great for concept previews, fashion mockups, or character styling. Key Inputs Outfit: Load the outfit image you want to apply. Make sure it's high quality — visible artifacts or distortions may carry over into the final result. Actor: Add the subject or character you want to dress. Ideally, use a clear, front-facing image. Human Parts Ultra: Choose which parts of the body the clothing should apply to. For example, for a long-sleeve shirt, select: torso, left arm, and right arm. This helps the model align the clothing properly during generation. Prompt: Default value works for most outfits, however you may try to adjust it to describe the desired outfit.

Z-Image Turbo + DyPE + SeedVR2 2.5 + TTP  16k reso

16k

8k

DyPE

SeedVR2

Text2Image

Upscale

Video

Z-Image Turbo

Z-Image Turbo + DyPE + SeedVR2 2.5 + TTP 16k reso

VEO3  Future of Video Creation

API

Audio

Floyo API

Text2Video

VEO3

Video

VEO3 Future of Video Creation

Video to Video with Control Image

AnimateDiff

Control Image

HotshotXL

LoRAs

SDXL

Video

Video2Video

Breathe life into a character from an image reference using motion reference from a video. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly and the style of your shot Load Video: Use any Mp4 that you would like to use for motion reference

Video to Video with Control Image

Breathe life into a character from an image reference using motion reference from a video. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly and the style of your shot Load Video: Use any Mp4 that you would like to use for motion reference

AI Influencer Ad Generator (Nano Banana + Wan 2.6)

Image

Influencer

Product

Video

Build your AI influencer, stage the product moment, and animate the full promo in one workflow.

AI Influencer Ad Generator (Nano Banana + Wan 2.6)

Build your AI influencer, stage the product moment, and animate the full promo in one workflow.

Image to 3D with Hunyuan3D

3D

3D View

Animation

Architecture

Filmmaking

Game Development

Hunyuan 3D

Hunyuan3D

Image to 3D

A simple workflow to create a detailed & textured 3D model from a reference image.

Image to 3D with Hunyuan3D

A simple workflow to create a detailed & textured 3D model from a reference image.

LoRA Training Video with Hunyuan

API

Hunyuan

LoRAs

LORA Training

Hunyuan is great at generating videos, but locking in a specific aesthetic or character is easier with a  LoRA.

LoRA Training Video with Hunyuan

Hunyuan is great at generating videos, but locking in a specific aesthetic or character is easier with a  LoRA.

Text to Character Sheet with a reference LoRA

Character Sheet

Controlnet

Flux

Image

LoRAs

Generate a character sheet using a prompt and a LoRA model of a particular person for more accurate renders. Key Inputs Load Image: Use any JPG or PNG of your pose sheet Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted at 1280px x 1280px Denoise: The amount of variance in the new image. Higher has more variance. ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.) Flux Guidance: How much influence the prompt has over the image. Higher has more guidance.

Text to Character Sheet with a reference LoRA

Generate a character sheet using a prompt and a LoRA model of a particular person for more accurate renders. Key Inputs Load Image: Use any JPG or PNG of your pose sheet Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted at 1280px x 1280px Denoise: The amount of variance in the new image. Higher has more variance. ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.) Flux Guidance: How much influence the prompt has over the image. Higher has more guidance.

FLUX.2 Klein 9B: Realistic Photo Enhancer

FLUX

FLUX.2 Klein

Image

Image2Image

LoRA

LoRAs

realism

Create realistic image but in an enhanced details using FLUX.2 Klein 9B and with LoRA

FLUX.2 Klein 9B: Realistic Photo Enhancer

Create realistic image but in an enhanced details using FLUX.2 Klein 9B and with LoRA

AniSora 3.2 and Wan2.2: Best Practices for Generating Smooth Character 3D Spin

3D

3D Spin

AniSora

Character Spin

Image2Video

Video

AniSora 3.2 and Wan2.2: Best Practices for Generating Smooth Character 3D Spin

SeedVR2 and TTP Toolset 8k Image Upscale

8k

Image2Image

SeedVR2

Upscale

Video

SeedVR2 and TTP Toolset 8k Image Upscale

Kling 3.0 Pro for Image to Video

Animation

Image2Video

Kling

Kling 3.0 Pro

Video

Turn images into a video using Kling 3.0 Pro

Kling 3.0 Pro for Image to Video

Turn images into a video using Kling 3.0 Pro

Image Inpainting

Flux

Image

Inpaint

Change specific details on just a portion of the image, sometimes known as inpainting or Erase & Replace. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Masking tools: Right-click to reveal the masking tool option, and create a mask of the desired area to inpaint Prompt: as descriptive a prompt as possible to help guide what you would like replaced in the masked area

Image Inpainting

Change specific details on just a portion of the image, sometimes known as inpainting or Erase & Replace. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Masking tools: Right-click to reveal the masking tool option, and create a mask of the desired area to inpaint Prompt: as descriptive a prompt as possible to help guide what you would like replaced in the masked area

Vertical Video Character Face & Actor Swap (Wan 2.2 Animate)

character replacement

character swap

image to video

LoRAs

masking

Points Editor

vertical video

Video

Wan2.2 Animate

WanAnimateToVideo

Vertical Video Character Face & Actor Swap (Wan 2.2 Animate)

Qwen Image Edit 2509: Change Camera Angle

Image

Image2Image

Multiple Angles

novel view

Qwen

Qwen Image Edit 2509

Qwen Image Edit 2509: Change Camera Angle

Text to Image + LoRA model

Flux

Image

LoRa

LoRAs

Text2Image

Create an image from a trained AI model of something specific ( a specific figure, outfit, art style, product etc) to ensure specific details within.

Text to Image + LoRA model

Create an image from a trained AI model of something specific ( a specific figure, outfit, art style, product etc) to ensure specific details within.

SeC Video Segmentation: Unleashing Adaptive, Semantic Object Tracking

SeC

Segmentation

Video

Video2Video

SeC Video Segmentation: Unleashing Adaptive, Semantic Object Tracking

FLUX.2 Klein 9B: Text to Image

Flux

FLUX2 Klein

Image

Photography

photorealism

Text2Image

Create a high quality image using 9B model of Flux 2 Klein

FLUX.2 Klein 9B: Text to Image

Create a high quality image using 9B model of Flux 2 Klein

Start/End Frame Multi-Video via Floyo API

API

Image to Video

Video

Compare between Luma Dream Machine and Kling Pro 1.6 via Fal API

Start/End Frame Multi-Video via Floyo API

Compare between Luma Dream Machine and Kling Pro 1.6 via Fal API

FLUX.2 Klein 9B: Image Inpainting

Flux

Flux.2 Klein

Image

Image2Image

Inpainting

LanPaint

Inpainting image using Flux.2 Klein and LanPaint

FLUX.2 Klein 9B: Image Inpainting

Inpainting image using Flux.2 Klein and LanPaint

Flux.2 Klein Image Expansion / Outpaint

Flux

Image

Image to Image

Klein

Outpaint

Video

Flux.2 Klein Image Expansion / Outpaint

Wan 2.7 Image to Video

animation

film production

image to video

Video

video generation

wan

Turn any still image into a short 6-second video clip with Alibaba's Wan 2.7 model. Upload your photo, describe the motion you want, and run. 1080P output.

Wan 2.7 Image to Video

Turn any still image into a short 6-second video clip with Alibaba's Wan 2.7 model. Upload your photo, describe the motion you want, and run. 1080P output.

Block-wise Image Upscaling with Qwen

Diffusion

Image

LoRA

LoRAs

Memory Efficient

Qwen Model

Upscale

Block-wise Image Upscaling with Qwen

Block-wise Image Upscaling with Qwen

Block-wise Image Upscaling with Qwen

Anima Preview 3 - Text to Image

Anima2

character design

concept art

fantasy

Image

Text2Image

Generate images with Anima 2, a model built for anime and fantasy art. Write a prompt, set your resolution, and get stylized results in one run. Free to try.

Anima Preview 3 - Text to Image

Generate images with Anima 2, a model built for anime and fantasy art. Write a prompt, set your resolution, and get stylized results in one run. Free to try.

Wan2.2 Fun Camera for Camera Control

Camera Control

Image2Video

Video

Wan2.2

Wan2.2 Fun Camera for Camera Control

Wan2.2 Fun and RealismBoost LoRA for V2V

Enhancer

LoRA

LoRAs

Video

Video2Video

Wan

Wan2.2 Fun and RealismBoost LoRA for V2V

 HunyuanVideo Foley: Create a Lifelike Sound

Audio

HunyuanVideo Foley

Video

Video2Video

HunyuanVideo Foley: Create a Lifelike Sound

Seedance I2V: Image to Video in Minutes

API

Floyo API

Image2Video

Seedance

Video

Seedance I2V: Image to Video in Minutes

Veo 3.1 Image to Video - First Frame and Optional Last Frame

API

Audio

Floyo API

Image2Video

Veo 3.1

Video

Veo 3.1 Image to Video - First Frame and Optional Last Frame

Wan Alpha Create Transparent Videos

Alpha

Text to Video

Transparent

VFX

Video

Video Editing

Wan

Wan Alpha Create Transparent Videos

Image to Video with Seedance Pro API

animation

film

Video

Image to Video with Seedance Pro API

Multiple Angle Lighting LoRA + 2511

Image

Image2Image

Image Editing

Lighting

LoRA

LoRAs

Multiple Angle Lighting

Qwen Image Edit 2509

Multiple Angle Lighting LoRA + 2511

DyPe and Z-Image Turbo for High Quality Text to Image

DyPE

Image

Photography

Portrait

Z-Image Turbo

DyPe and Z-Image Turbo for High Quality Text to Image

Image to Character Sheet

Character Sheet

Image

Image to Image

SDXL

Generate a character sheet with multiple angles from a single input image as reference. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly. If you're trying to create a full body output, a full body input must be provided.

Image to Character Sheet

Generate a character sheet with multiple angles from a single input image as reference. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly. If you're trying to create a full body output, a full body input must be provided.

MMAudio: Video to Synced Audio

Audio

MMaudio

Video

Video to Video

Generate synchronized audio with a given video input. It can be combined with video models to get videos with audio.

MMAudio: Video to Synced Audio

Generate synchronized audio with a given video input. It can be combined with video models to get videos with audio.

Image Inpainting with LoRA

Image

Inpaint

LoRa

LoRAs

Change specific details on just a portion of the image for inpainting or Erase & Replace, adding a LoRA for extra control.

Image Inpainting with LoRA

Change specific details on just a portion of the image for inpainting or Erase & Replace, adding a LoRA for extra control.

Scribble to Image

Controlnet

Image

SD1.5

Turn your scribbles into a beautiful image with only a drawing tool and a text prompt. Key Inputs Scribble: Create your scribble with the painting and design tools Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.)

Scribble to Image

Turn your scribbles into a beautiful image with only a drawing tool and a text prompt. Key Inputs Scribble: Create your scribble with the painting and design tools Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.)

Z-Image Turbo with Controlnet 2.1 and Qwen VLM

API

Controlnet

Floyo API

Image

Image2Image

LoRA

LoRAs

Z-Image Turbo

Creating Accurate Variety of Images

Z-Image Turbo with Controlnet 2.1 and Qwen VLM

Creating Accurate Variety of Images

Kling Omni One Video to Video Edit

API

Audio

film and animation

Floyo API

kling 2.5

Omni One

Video

video to video

Kling Omni One Video to Video Edit

Sketch to Image

Controlnet

Image

SD1.5

Turn your sketches into full blown scenes. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Prompt: as descriptive a prompt as possible Width & height: In pixels ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.)

Sketch to Image

Turn your sketches into full blown scenes. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Prompt: as descriptive a prompt as possible Width & height: In pixels ControlNet Strength: The amount of adherence to the original image. Higher has more adherence. Start Percent: The point in the generation process where the control starts exerting influence. (Have it start later, to let AI imagine first.) End Percent: The point in the generation process where the control stops exerting influence. (Have it end sooner, to let AI finish it off with some variation.)

Qwen Image Edit 2511 Restore Damage Old Photograph

Image

Image2Image

Qwen

Qwen Image Edit 2511

Restore Damage Old Photograph

Qwen Image Edit 2511 Restore Damage Old Photograph

Restore Damage Old Photograph

Qwen Image 2512 · Text to Image

Image

Photography

Qwen

Qwen Image 2512

Text2Image

Text to image

Qwen Image 2512 · Text to Image

Text to image

Text to Video and Wan with optional LoRA

LoRa

LoRAs

Text2Video

Video

Wan2.1

Generate a high-quality video from a text prompt and add in a LoRA for extra control over character or style consistency. Key Inputs Prompt: as descriptive a prompt as possible Load LoRA: Load your reference model here Width & height: Optimal resolution settings are noted File Format: H.264 and more

Text to Video and Wan with optional LoRA

Generate a high-quality video from a text prompt and add in a LoRA for extra control over character or style consistency. Key Inputs Prompt: as descriptive a prompt as possible Load LoRA: Load your reference model here Width & height: Optimal resolution settings are noted File Format: H.264 and more

Happy Horse 1.0 Reference to Video

character design

consistency

happy horse

image to video

reference to video

Video

video generation

Turn up to 9 reference images plus a prompt into a 5-second video with Happy Horse 1.0. Keep characters, products, and style consistent across the shot.

Happy Horse 1.0 Reference to Video

Turn up to 9 reference images plus a prompt into a 5-second video with Happy Horse 1.0. Keep characters, products, and style consistent across the shot.

Camera Angle Control with QwenMultiAngle

Image

Image2Image

Image Edit

LoRAs

Qwen Image Edit 2511

Create different angle of the image using Qwen Image Edit 2511 and with special node

Camera Angle Control with QwenMultiAngle

Create different angle of the image using Qwen Image Edit 2511 and with special node

LTX 2 19B Fast for Image to Video

Animation

Audio

Filmography

Image2Video

LoRAs

LTX 2

Open Source

Video

A workflow for ltx 2 image to video using distilled model

LTX 2 19B Fast for Image to Video

A workflow for ltx 2 image to video using distilled model

Light Restoration LoRA + Qwen Image Edit 2509 Image to Image

Image

Image2Image

Light Restoration

LoRAs

Qwen

Qwen Image Edit 2509

Light Restoration LoRA + Qwen Image Edit 2509 Image to Image

SAM3 for Video Masking using Text

SAM3

Video

Video2Video

Video Masking

Create a video masking using SAM3 and Text only.

SAM3 for Video Masking using Text

Create a video masking using SAM3 and Text only.

Qwen Multiangle Light · Image to Image For Anime

Image

image to image

multiangle

qwen image edit

relighting

Upload one image and this workflow generates five relighted versions simultaneously, each lit from a different angle and direction, with full control over azimuth, elevation, intensity, and color.

Qwen Multiangle Light · Image to Image For Anime

Upload one image and this workflow generates five relighted versions simultaneously, each lit from a different angle and direction, with full control over azimuth, elevation, intensity, and color.

Next-Level Motion from Images using MiniMax

API

Floyo API

Image2Video

Minimax

Video

Next-Level Motion from Images using MiniMax

Chatterbox Text to Speech

Audio

Chatterbox

TTS

Text to speech workflow using Chatterbox

Chatterbox Text to Speech

Text to speech workflow using Chatterbox

 Qwen Image Edit 2509 + Flux Krea for Creating Next Scene

Filmography

Flux

Flux Krea

Image

LoRAs

Photography

Qwen

Qwen Image Edit 2509

Qwen Image Edit 2509 + Flux Krea for Creating Next Scene

 Recraft V3 Image to Image - Style Transfer

digital illustration

Image

image to image

recraft

style transfer

text to image

Transform an existing image with Recraft V3. Upload a reference, write a prompt, set strength, and pick a style. Controls how much of the original survives.

Recraft V3 Image to Image - Style Transfer

Transform an existing image with Recraft V3. Upload a reference, write a prompt, set strength, and pick a style. Controls how much of the original survives.

api

Audio

hailuo 3.0

image to video

minimax h3

Video

video generation

Bring a photo to life with MiniMax H3 (Hailuo 3.0). Upload an image, describe the motion, add optional references to guide the look, and hit run for a 2K clip.

MiniMax H3 · Image to Video

Bring a photo to life with MiniMax H3 (Hailuo 3.0). Upload an image, describe the motion, add optional references to guide the look, and hit run for a 2K clip.

Qwen Image Edit – Multi-Angle Camera View

Image

Image to Image

LoRAs

Qwen

Qwen Image Edit – Multi-Angle Camera View

  Kling 2.6 Pro for Image to Video

Animation

Filmmaking

Image2Video

Kling 2.6 Pro

Video

Create stunning videos using Kling 2.6 Pro

Kling 2.6 Pro for Image to Video

Create stunning videos using Kling 2.6 Pro

Image Upscaler with LoRA

Flux

Image

LoRa

LoRAs

Upscale

Create a larger more detailed image along with an extra AI model for fine tuned guidance. Key Inputs Load Image: Use any JPG or PNG showing your subject clearly Load LoRA: Load your reference model here Prompt: as descriptive a prompt as possible Upscale by: The factor of magnification Denoise: The amount of variance in the new image. Higher has more variance.

Image Upscaler with LoRA

Create a larger more detailed image along with an extra AI model for fine tuned guidance. Key Inputs Load Image: Use any JPG or PNG showing your subject clearly Load LoRA: Load your reference model here Prompt: as descriptive a prompt as possible Upscale by: The factor of magnification Denoise: The amount of variance in the new image. Higher has more variance.

Image to Image Character Sheet Face Swap with Ace+

Character Sheet

Face Swap

Flux

Image

Take a character sheet and use a reference image to replace all the faces with that new person. Key Inputs Load Image: Use any JPG or PNG showing your pose sheet Load New Face: Use any JPG or PNG showing your subject clearly that you would like to swap into the pose sheet. Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted at 1024px x 1024px Keep Proportion: Enable keep_proportion if you want to keep the same size with input and output Denoise: The amount of variance in the new image. Higher has more variance.

Image to Image Character Sheet Face Swap with Ace+

Take a character sheet and use a reference image to replace all the faces with that new person. Key Inputs Load Image: Use any JPG or PNG showing your pose sheet Load New Face: Use any JPG or PNG showing your subject clearly that you would like to swap into the pose sheet. Prompt: as descriptive a prompt as possible Width & height: Optimal resolution settings are noted at 1024px x 1024px Keep Proportion: Enable keep_proportion if you want to keep the same size with input and output Denoise: The amount of variance in the new image. Higher has more variance.

Kling Omni One Image to Video

Audio

Image2Video

Kling

Omni One

Video

Kling Omni One Image to Video

Wan2.6 Text to Video

Animation

Film

Text2Video

Video

Wan2.6

Wan2.6 Text to Video

LTX 2.3 Image to Video with Two-Pass Upscaling

API

Audio

LTX

Video

LTX 2.3 Image to Video with Two-Pass Upscaling

FLUX.2 Klein 9B + SAM3 + GhostMannequin LoRA

FLUX

FLUX.2 Klein

Ghost Mannequin

Image

Image2Image

LoRAs

SAM3

Create a ghost mannequin clothes using flux.2 klein, SAM3 and Ghost mannequin LoRA

FLUX.2 Klein 9B + SAM3 + GhostMannequin LoRA

Create a ghost mannequin clothes using flux.2 klein, SAM3 and Ghost mannequin LoRA

Image Redux with Flux

Flux

Image

Redux

Create variations of a given image, or restyle them. It can be used to refine, explore, or transform ideas and concepts. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Width & height: In pixels Prompt: as descriptive a prompt as possible Strength (step 5: value): Strength of redux model, play around with the value to increase or decrease the amount of variation

Image Redux with Flux

Create variations of a given image, or restyle them. It can be used to refine, explore, or transform ideas and concepts. Key Inputs Image reference: Use any JPG or PNG showing your subject clearly Width & height: In pixels Prompt: as descriptive a prompt as possible Strength (step 5: value): Strength of redux model, play around with the value to increase or decrease the amount of variation

Qwen Image Edit 2511 Lightning - Multi-Image Edit

Image

Qwen

Text to Image

Qwen Image Edit 2511 Lightning - Multi-Image Edit

Happy Horse 1.1 · Image to Video

ai video

audio

happy horse 1.1

image to video

Video

Upload a starting image and describe the motion you want. Happy Horse 1.1 animates it into a cinematic video with synchronized audio, dialogue, and lip-sync at up to 1080p.

Happy Horse 1.1 · Image to Video

Upload a starting image and describe the motion you want. Happy Horse 1.1 animates it into a cinematic video with synchronized audio, dialogue, and lip-sync at up to 1080p.

GPT Image 2: Image Editing

e-commerce

gpt image 2

Image

image to image

inpainting

product photography

Edit images with OpenAI's GPT Image 2. Upload one or two images, write what you want changed, and the model rewrites the scene while keeping details intact.

GPT Image 2: Image Editing

Edit images with OpenAI's GPT Image 2. Upload one or two images, write what you want changed, and the model rewrites the scene while keeping details intact.

Wan2.1 + WanMOVE for Animating Movement using Trajectory Path

Animation

Image2Video

Video

Wan2.1

Wan Move

Wan2.1 + WanMOVE for Animating Movement using Trajectory Path

360 Degree Product Video Using Nano Banana Pro

Image

Image to Video

NanoBanana

Veo2

Video

360 Degree Product Video Using Nano Banana Pro

 Nano Banana Pro: Edit Any Image

gemini 3 pro

Image

Image2Image

Image Editing

Nano Banana Pro Edit

Nano Banana Pro: Edit Any Image

Clothing & Accessories Replacement

Ecommerce

Opensource

Outfit replacement

Video

Video to Video

Clothing & Accessories Replacement

Seedance 2.0 - Image to Video

animation

film production

image to video

vfx

Video

video generation

Turn any image into video with Seedance 2.0 by ByteDance. Built-in audio generation, start and end frame control, and clips up to 10 seconds

Seedance 2.0 - Image to Video

Turn any image into video with Seedance 2.0 by ByteDance. Built-in audio generation, start and end frame control, and clips up to 10 seconds

Seedance 2.0 Reference-to-Video

API

Seedance

Video

Seedance 2.0 Reference-to-Video

Image to Video with Multiframe Control

Image2Video

LTX

Video

Used for image to video generation, including first frame, end frame, or other multiple key frames. Key Inputs Load Image (Start Frame): Use any JPG or PNG showing your subject clearly to start your video Load Image (End Frame): Use any JPG or PNG showing your subject clearly to act as the last part of your video. Make sure it's the same resolution as the load image. Width & height: Optimal resolution settings are noted. LTX maximum resolution is 768x512 Prompt: as descriptive a prompt as possible

Image to Video with Multiframe Control

Used for image to video generation, including first frame, end frame, or other multiple key frames. Key Inputs Load Image (Start Frame): Use any JPG or PNG showing your subject clearly to start your video Load Image (End Frame): Use any JPG or PNG showing your subject clearly to act as the last part of your video. Make sure it's the same resolution as the load image. Width & height: Optimal resolution settings are noted. LTX maximum resolution is 768x512 Prompt: as descriptive a prompt as possible

Wan 2.1 Text2Image

Image

text2image

Wan2.1

Created by @yanokusnir on Reddit, please support the original creator! https://www.reddit.com/r/StableDiffusion/comments/1lu7nxx/wan_21_txt2img_is_amazing/ If this is your workflow, please contact us at team@floyo.ai to claim it! Original post from the creator: Hello. This may not be news to some of you, but Wan 2.1 can generate beautiful cinematic images. I was wondering how Wan would work if I generated only one frame, so to use it as a txt2img model. I am honestly shocked by the results. All the attached images were generated in fullHD (1920x1080px) and on my RTX 4080 graphics card (16GB VRAM) it took about 42s per image. I used the GGUF model Q5_K_S, but I also tried Q3_K_S and the quality was still great. The only postprocessing I did was adding film grain. It adds the right vibe to the images and it wouldn't be as good without it. Last thing: For the first 5 images I used sampler euler with beta scheluder - the images are beautiful with vibrant colors. For the last three I used ddim_uniform as the scheluder and as you can see they are different, but I like the look even though it is not as striking. :) Enjoy.

Wan 2.1 Text2Image

Created by @yanokusnir on Reddit, please support the original creator! https://www.reddit.com/r/StableDiffusion/comments/1lu7nxx/wan_21_txt2img_is_amazing/ If this is your workflow, please contact us at team@floyo.ai to claim it! Original post from the creator: Hello. This may not be news to some of you, but Wan 2.1 can generate beautiful cinematic images. I was wondering how Wan would work if I generated only one frame, so to use it as a txt2img model. I am honestly shocked by the results. All the attached images were generated in fullHD (1920x1080px) and on my RTX 4080 graphics card (16GB VRAM) it took about 42s per image. I used the GGUF model Q5_K_S, but I also tried Q3_K_S and the quality was still great. The only postprocessing I did was adding film grain. It adds the right vibe to the images and it wouldn't be as good without it. Last thing: For the first 5 images I used sampler euler with beta scheluder - the images are beautiful with vibrant colors. For the last three I used ddim_uniform as the scheluder and as you can see they are different, but I like the look even though it is not as striking. :) Enjoy.

Hyper3D Rodin V2 for Image to 3D

3D

3D Model

Hyper3D Rodin v2

Image to 3D

Rodin v2

Turn your images into 3D using Hyper3D Rodin v2

Hyper3D Rodin V2 for Image to 3D

Turn your images into 3D using Hyper3D Rodin v2

Video Masking with Sam2 Comparison

Masking

Segmentation

Video

Use a video clip and visual markers to segment/create masks of the subject or the inverse. Key Inputs Load Video: Use any Mp4 that you would like to segment or create a mask from Select subject: Use 3 green selectors to identify your subject and one red selector to identify the space outside your subject Modify markers: Shift+Click to add markers, Shift+Right Click to remove markers

Video Masking with Sam2 Comparison

Use a video clip and visual markers to segment/create masks of the subject or the inverse. Key Inputs Load Video: Use any Mp4 that you would like to segment or create a mask from Select subject: Use 3 green selectors to identify your subject and one red selector to identify the space outside your subject Modify markers: Shift+Click to add markers, Shift+Right Click to remove markers

 LTX 2.3 Face-Consistent Image to Video with VBVR

Audio

Image to Video

LoRAs

LTX2.3

Video

Turn a single portrait into vertical video with LTX 2.3. The VBVR LoRA holds face identity steady and gives motion the physical weight that I2V usually loses.

LTX 2.3 Face-Consistent Image to Video with VBVR

Turn a single portrait into vertical video with LTX 2.3. The VBVR LoRA holds face identity steady and gives motion the physical weight that I2V usually loses.

Tripo3D for Image to 3D

3D

3D Model

Image to 3D

Tripo3D

Tripo v2.5

Create 3D model using Tripo3D with v2.5

Tripo3D for Image to 3D

Create 3D model using Tripo3D with v2.5

Happy Horse 1.0 - Image to Video

consistency

film production

happy horse

image to video

product photography

Video

video generation

Animate a still image with Happy Horse 1.0. Upload a frame, describe the motion you want, get a 5-second clip with stable physics and consistent details.

Happy Horse 1.0 - Image to Video

Animate a still image with Happy Horse 1.0. Upload a frame, describe the motion you want, get a 5-second clip with stable physics and consistent details.

Vertical Video FX Inserter - Qwen + Wan 2.1 FunControl

fx-integration

Image

image-to-image

LoRAs

qwen

reference-image

upscaling

Video

video-conditioning

wan21-funcontrol

Vertical Video FX Inserter - Qwen + Wan 2.1 FunControl

Vertical Video FX Insterter / Element Pass with Seedream + Wan

Image

reference-image

seedream

upscaling

Video

video-conditioning

wan2.1funControl

Vertical Video FX Insterter / Element Pass with Seedream + Wan

Vertical Video Light & Mood Shift

Audio

Image

image-to-image

LoRAs

qwen

reference-image

Video

wan2.1 FunControl

Vertical Video Light & Mood Shift

Wan2.2 and Bullet Time LoRA: Transform Static Shots into Product Spins

Bullet Time

Ecommerce

Image2Video

LoRA

LoRAs

Product Demo

Video

Wan2.2 and Bullet Time LoRA: Transform Static Shots into Product Spins

FlatLogColor LoRA and Qwen Image Edit 2509

FlatLogColor

Image

LoRA

LoRAs

Photography

Qwen

Qwen Image Edit 2509

FlatLogColor LoRA and Qwen Image Edit 2509

Chroma 1 Radiance Text to Image

Chrome1 Radiance

Image

Macro Photography

Text2Image

Chroma 1

Chroma 1 Radiance Text to Image

Chroma 1

Wan2.1 InfiniteTalk Video to Video

Audio

InfiniteTalk

Video

Video to Video

Wan

Wan2.1 InfiniteTalk Video to Video

SAM3 Image Segmentation

Image

Image2Image

SAM3

Segmentation

SAM3 Image Segmentation

SAM3 for Video Masking using Points

SAM3

Video

Video2Video

Video Masking

Create a video masking using SAM3 and Points only.

SAM3 for Video Masking using Points

Create a video masking using SAM3 and Points only.

Z-Anime - Text to Image with SeedVR Upscale

anime

character design

concept art

Image

seedvr

text to image

upscaling

z-anime

Generate anime and illustration art from text with Z-Anime, then upscale to 1080p with SeedVR. Compare the base render and the upscaled version side by side.

Z-Anime - Text to Image with SeedVR Upscale

Generate anime and illustration art from text with Z-Anime, then upscale to 1080p with SeedVR. Compare the base render and the upscaled version side by side.

InfiniteTalk - Lip Sync Any Video to Any Audio

Audio

vid2vid

Video

wan

InfiniteTalk - Lip Sync Any Video to Any Audio

Anything2Real 2601A

Image

Image to Image

Anything2Real 2601A

Flux LoRA Trainer

API

Flux

LoRA

LoRAs

Trainer

Flux LoRA Trainer

FLUX.1 Kontext · 3D Print Style (Image to Image)

3D

3D Print

Flux Kontext

Image

image editing

Imageto3D

LoRAs

Mockup

FLUX.1 Kontext · 3D Print Style (Image to Image)

SVG Potracer + Qwen Image 2511 for Image to SVG

Image

Image to SVG

Qwen Image Edit 2511

SVG

SVG Potracer

Create SVG image using Qwen Image Edit and SVG Potracer node

SVG Potracer + Qwen Image 2511 for Image to SVG

Create SVG image using Qwen Image Edit and SVG Potracer node

Insert Products in Ecommerce Ads - NanoBanana Pro

Ecommerce

Image

Image to Image

NanoBanana

Reference Image

Insert Products in Ecommerce Ads - NanoBanana Pro

Seedance 2.0 for Jewelry Scene Animator

consistency

e-commerce

image to video

product photography

seedance 2.0

Video

video generation

Animate jewelry from product photos with Seedance 2.0. Upload a start frame and up to 6 reference angles, describe the camera move, and get a 5-second clip.

Seedance 2.0 for Jewelry Scene Animator

Animate jewelry from product photos with Seedance 2.0. Upload a start frame and up to 6 reference angles, describe the camera move, and get a 5-second clip.

Meshy v6 for Image to 3D Model

3D

3D Model

Image to 3D

Meshy v6

Create a 3D model from Image using Meshy v6

Meshy v6 for Image to 3D Model

Create a 3D model from Image using Meshy v6

🔥Create Stunning 10 Second 3D Spin Shots

3D

3D Spin

Floyo

Floyo API

Image2Video

Seedance

Spin

Video

🔥Create Stunning 10 Second 3D Spin Shots

Qwen Image Edit 2509 and Grayscale to Color LoRA

3D Render

Ecommerce

Image

LoRAs

Marketing

Product

Qwen Image Edit 2509 and Grayscale to Color LoRA

Kling Master 2.0 Create Engaging Video Content

API

Floyo API

Image2Video

Kling

Kling Master 2.0

Video

Kling Master 2.0 Create Engaging Video Content

GPT Image 1.5

GPT Image 1.5

Image

Image2Image

Image Editing

for Image Editing

GPT Image 1.5

for Image Editing

Wan 2.1 Vid2Vid Style Transfer with Ditto

animation

Ditto

lora

LoRAs

VACE

Video

Video2Video

Wan

Upload any video, describe a new style, and Wan 2.1 rewrites every frame. Ditto keeps motion and structure intact across anime, Pixar, clay, and dozens more.

Wan 2.1 Vid2Vid Style Transfer with Ditto

Upload any video, describe a new style, and Wan 2.1 rewrites every frame. Ditto keeps motion and structure intact across anime, Pixar, clay, and dozens more.

SRPO Next-Gen Text-to-Image

Image

SRPO

Text2Image

SRPO Next-Gen Text-to-Image

Ovi: Create a Talking Portrait

Audio

Image2Video

Lip Sync

Ovi

Video

Ovi: Create a Talking Portrait

Studio Relighting for Composited Products

Image

lightning lora

LoRAs

product lighting

relight composite

studio relighting

Studio Relighting for Composited Products

LTX 2.3 IC LoRA Union Control · V2V For Anime

Audio

ic-lora

LoRAs

ltx2.3

union control

Video

video to video

Upload a video and a style reference image. LTX 2.3 with IC LoRA Union Control restyles the full video to match the reference while preserving the original motion, body pose, and scene structure using blended depth, pose, and edge control.

LTX 2.3 IC LoRA Union Control · V2V For Anime

Upload a video and a style reference image. LTX 2.3 with IC LoRA Union Control restyles the full video to match the reference while preserving the original motion, body pose, and scene structure using blended depth, pose, and edge control.

LTX 2 Fast API for Image to Video

API

Audio

Filmography

Fimmaking

Floyo API

Image2Video

LTX 2 Fast

Video

Outdated model. Please go to LTX 2.3 Image to Video workflow to use LTX 2.3

LTX 2 Fast API for Image to Video

Outdated model. Please go to LTX 2.3 Image to Video workflow to use LTX 2.3

LTX 2.3 Two-Pass · Image to Video For Anime

audio

image to video

ltx 2.3

two-pass

Video

Upload an image and LTX Video 2.3 22B generates a 5-second video at 1080p with synchronized audio, using a two-pass pipeline that renders at low resolution first then upscales in latent space for sharp, detailed output.

LTX 2.3 Two-Pass · Image to Video For Anime

Upload an image and LTX Video 2.3 22B generates a 5-second video at 1080p with synchronized audio, using a two-pass pipeline that renders at low resolution first then upscales in latent space for sharp, detailed output.

 Seedance Text to Video: Create Stunning video

API

Floyo API

Seedance

Text2Video

Video

Seedance Text to Video: Create Stunning video

Wan2.1 and ATI for Control Video Motion: Draw Your Path, Get Your Video

ATI

Image2Video

Video

Wan

Wan2.1 and ATI for Control Video Motion: Draw Your Path, Get Your Video

Nano Banana · Edit (Image to Image)

API

Image

Image2Image

Nano Banana Edit

Nano Banana · Edit (Image to Image)

Seedance 2.0 Fast Reference-to-Video

API

Seedance

Video

Seedance 2.0 Fast Reference-to-Video

Z-Image Turbo · Text to Image For Anime

Image

text to image

z-image turbo

Write a prompt and Z-Image Turbo generates a photorealistic 1024x1024 image in 9 steps, using a 6-billion-parameter model distilled for speed with bilingual prompt support.

Z-Image Turbo · Text to Image For Anime

Write a prompt and Z-Image Turbo generates a photorealistic 1024x1024 image in 9 steps, using a 6-billion-parameter model distilled for speed with bilingual prompt support.

Qwen Image Edit 2509 + Multi-Angle LoRA for Camera

Camera Control

Image

Image2Image

LoRA

LoRAs

Qwen

Qwen Image Edit 2509

Re-render your subject from any camera angle with Qwen Image Edit 2509 and a Multi-Angle LoRA. Pan, tilt, rotate, wide-angle, or close-up. No trigger word.

Qwen Image Edit 2509 + Multi-Angle LoRA for Camera

Re-render your subject from any camera angle with Qwen Image Edit 2509 and a Multi-Angle LoRA. Pan, tilt, rotate, wide-angle, or close-up. No trigger word.

Kling 3.0 Pro Motion Control

animation

character design

image to video

kling

Video

video generation

Apply motion from a reference video to a still image with Kling 3.0 Pro.

Kling 3.0 Pro Motion Control

Apply motion from a reference video to a still image with Kling 3.0 Pro.

Image to 3D with Hunyuan3D w/ Texture Upscale

3D

Animation

Architecture

Flux

Game Development

Hunyuan 3D

Image to 3D

Upscaling

Create a 3D model from a reference image with Flux Dev texture upscaling.

Image to 3D with Hunyuan3D w/ Texture Upscale

Create a 3D model from a reference image with Flux Dev texture upscaling.

Grok Imagine for Text to Video

Filmogrpahy

Grok

Text2Video

Video

Create excellent videos using Grok Imagine for T2V

Grok Imagine for Text to Video

Create excellent videos using Grok Imagine for T2V

Grok Imagine: Edit Images with a Text Prompt

API

background removal

E-commerce

Grok Imagine

Image

Image Editing

Image to Image

Style Transfer

Edit images using Grok Imagine

Grok Imagine: Edit Images with a Text Prompt

Edit images using Grok Imagine

Static Watermark Remover

Video

Watermark Remover

Static Watermark Remover

Simple Self-Forcing Wan1.3B+Vace workflow

Vace

Video

Wan

Created by @davcha on Civitai, please support the original creator! https://civitai.com/models/1674121/simple-self-forcing-wan13bvace-workflow If this is your workflow, please contact us at team@floyo.ai to claim it! Original guide from creator: This is a very simple workflow to run Self-Forcing Wan 1.3B + Vace, it only uses a single custom node, which everyone making videos should have: Kosinkadink/ComfyUI-VideoHelperSuite. Everything else is pure comfy core. You'll need to download the model of your choice from here lym00/Wan2.1-T2V-1.3B-Self-Forcing-VACE · Hugging Face, and put it inside your /path/to/models/diffusion_models folder. This workflow can be used as a very good start for experimenting. You can refer to this [2503.07598] VACE: All-in-One Video Creation and Editing for how to use Vace. You don't need to read the paper of course, the information you are interested in is mostly at the top of page 7, which I reproduce in the following: Basically, in the WanVaceToVideo node, you have 3 optional inputs: control_video, control_masks, and reference_image. control_video and control_masks are a little bit misleading. You don't have to provide a full video. You can in fact provide a variety of things to obtain various effects. For example: if you provide a single image, it's basically more or less equivalent to image2video. if you provide a sequence of images separated by empty images: img1, black, black, black, img2, black, black, black, img3, etc... then it's equivalent to interpolating all these img, filling the blacks. A special case of this one to make it clear is if you have img1, black, black, ..., black, img2, then it's equivalent to start_img, end_img to video. control_masks control where Wan should paint. Basically if wherever the mask is 1, the original image will be kept. So you can for example pad and/or mask an input image, like this: and use that image and mask as control_video and control_mask, and you'll basically do a image2video inpaint and outpaint. If you input a video in control_video, then you can control where the changes should happen in the same way, using control_mask. You'll need to set one mask per frame in the video. if you input an image preprocessed with openpose or a depthmap, you can finely control the movement in the video output. reference_image node is basically an image that you feed to Wan+Vace that serves as a reference point. For example, if you put the image of someone's face here, there's a good chance you'll get a video with that person's face.

Simple Self-Forcing Wan1.3B+Vace workflow

Created by @davcha on Civitai, please support the original creator! https://civitai.com/models/1674121/simple-self-forcing-wan13bvace-workflow If this is your workflow, please contact us at team@floyo.ai to claim it! Original guide from creator: This is a very simple workflow to run Self-Forcing Wan 1.3B + Vace, it only uses a single custom node, which everyone making videos should have: Kosinkadink/ComfyUI-VideoHelperSuite. Everything else is pure comfy core. You'll need to download the model of your choice from here lym00/Wan2.1-T2V-1.3B-Self-Forcing-VACE · Hugging Face, and put it inside your /path/to/models/diffusion_models folder. This workflow can be used as a very good start for experimenting. You can refer to this [2503.07598] VACE: All-in-One Video Creation and Editing for how to use Vace. You don't need to read the paper of course, the information you are interested in is mostly at the top of page 7, which I reproduce in the following: Basically, in the WanVaceToVideo node, you have 3 optional inputs: control_video, control_masks, and reference_image. control_video and control_masks are a little bit misleading. You don't have to provide a full video. You can in fact provide a variety of things to obtain various effects. For example: if you provide a single image, it's basically more or less equivalent to image2video. if you provide a sequence of images separated by empty images: img1, black, black, black, img2, black, black, black, img3, etc... then it's equivalent to interpolating all these img, filling the blacks. A special case of this one to make it clear is if you have img1, black, black, ..., black, img2, then it's equivalent to start_img, end_img to video. control_masks control where Wan should paint. Basically if wherever the mask is 1, the original image will be kept. So you can for example pad and/or mask an input image, like this: and use that image and mask as control_video and control_mask, and you'll basically do a image2video inpaint and outpaint. If you input a video in control_video, then you can control where the changes should happen in the same way, using control_mask. You'll need to set one mask per frame in the video. if you input an image preprocessed with openpose or a depthmap, you can finely control the movement in the video output. reference_image node is basically an image that you feed to Wan+Vace that serves as a reference point. For example, if you put the image of someone's face here, there's a good chance you'll get a video with that person's face.

Z-Image Turbo + Chord Image to  PBR Material

Image

Text to Image

Z-turbo

Z-Image Turbo + Chord Image to PBR Material

Create Photorealistic Packaging from Dielines

Image

Nano banana

packaging-materials

product-packaging

Create Photorealistic Packaging from Dielines

Wan2.1 and FantasyTalking - Image2Video Lipsync

FantasyTalking

Image2Video

Lipsync

Video

Wan2.1

Create high quality lipsync video from image inputs with Wan2.1 FantasyTalking Key Inputs Load Image: Select an image of a person with their face in clear view Load Audio: Choose audio file Frames: How many frames generated

Wan2.1 and FantasyTalking - Image2Video Lipsync

Create high quality lipsync video from image inputs with Wan2.1 FantasyTalking Key Inputs Load Image: Select an image of a person with their face in clear view Load Audio: Choose audio file Frames: How many frames generated

Kling 3.0 Standard Motion Control

API

Floyo API

Kling

MotionControl

Video

Transfer movements from a reference video to any character image.

Kling 3.0 Standard Motion Control

Transfer movements from a reference video to any character image.

Kling V3 Pro Motion Control

animation

kling

Video

Kling V3 Pro Motion Control

LTX 2.3 + IC-LoRA Cameraman: Image to Video

Audio

Film Production

Image to Video

LoRAs

ltx2.3

Video

Animate a still image with LTX 2.3 22B while a Cameraman IC-LoRA copies the camera motion from a reference video. Audio is generated in the same pass.

LTX 2.3 + IC-LoRA Cameraman: Image to Video

Animate a still image with LTX 2.3 22B while a Cameraman IC-LoRA copies the camera motion from a reference video. Audio is generated in the same pass.

Flux Fill Dev Image Outpainting

Flux

Image

Outpaint

Extend your images out for a wider field of view or just to see more of your subject. Expand compositions, change aspect ratios, or add creative elements while maintaining consistency in style, lighting, and detail while seamlessly blending with the existing artwork.

Flux Fill Dev Image Outpainting

Extend your images out for a wider field of view or just to see more of your subject. Expand compositions, change aspect ratios, or add creative elements while maintaining consistency in style, lighting, and detail while seamlessly blending with the existing artwork.

Partial Modification Reference Image using Flux.2

Flux

Image

Image to Image

Inpainting

Paint a mask over the part of your product image you want to change, drop in a reference design, and Flux 2 Klein redraws only that region in four steps.

Partial Modification Reference Image using Flux.2

Paint a mask over the part of your product image you want to change, drop in a reference design, and Flux 2 Klein redraws only that region in four steps.

Qwen3 Thinking Prompt Enhancer

Open-source

Prompting

Qwen3 Thinking Prompt Enhancer

Chord for PBR Material Generation using Text to 3D

3D

Chord

Game Design

Image

PBR Material

Text to 3D

Ubisoft

Create a 3D Game Material Asset using Chord Model from Ubisoft

Chord for PBR Material Generation using Text to 3D

Create a 3D Game Material Asset using Chord Model from Ubisoft

Wan2.1 + SCAIL-2 for Character Motion Transfer

character animation

character replacement

LoRAs

motion transfer

scail 2

Video

video to video

wan 2.1

Transfer the motion from any video onto your own character with SCAIL 2, Z.ai's end-to-end character animation model built on Wan 2.1. Upload a video and a character photo, then hit run.

Wan2.1 + SCAIL-2 for Character Motion Transfer

Transfer the motion from any video onto your own character with SCAIL 2, Z.ai's end-to-end character animation model built on Wan 2.1. Upload a video and a character photo, then hit run.

Topaz Video Upscaler for Sharper Results

API

Floyo API

Topaz

Video

Video2Video

Video Upscale

Upload a video, pick your enhancement model and quality level, and Topaz Video AI sharpens, denoises, and upscales it. Audio is preserved. Output is H265 MP4.

Topaz Video Upscaler for Sharper Results

Upload a video, pick your enhancement model and quality level, and Topaz Video AI sharpens, denoises, and upscales it. Audio is preserved. Output is H265 MP4.

Hunyuan 3D Pro - Image to 3D Model

3D

Api

Hunyuan

Image

Image to 3D

Turn any photo into a production-ready 3D model with Hunyuan 3D Pro. Get a GLB file, a thumbnail render, and an interactive turntable preview in under 60 seconds.

Hunyuan 3D Pro - Image to 3D Model

Turn any photo into a production-ready 3D model with Hunyuan 3D Pro. Get a GLB file, a thumbnail render, and an interactive turntable preview in under 60 seconds.

Qwen Image Edit - Infinite Image Styles

Image

image to image

infinite style

qwen

Restyle any photo with text. Type a style, hit Run, get the same image as anime, watercolor, claymation, or any look in 4 seconds with Qwen Image Edit.

Qwen Image Edit - Infinite Image Styles

Restyle any photo with text. Type a style, hit Run, get the same image as anime, watercolor, claymation, or any look in 4 seconds with Qwen Image Edit.

Text to Video + Hunyuan LoRA

Hunyuan

LoRa

LoRAs

Text2Video

Video

Integrate a custom model with your text prompt to create a video with a consistent character, style or element. Key Inputs Prompt: as descriptive a prompt as possible. Make sure to include the trigger word from your LoRA below Load LoRA: Load your reference model here Width & height: resolution settings are noted in pixels Guidance strength (CFG): Higher numbers adhere more to the prompt Flow Shift: For temporal consistency, adjust to tweak video smoothness.

Text to Video + Hunyuan LoRA

Integrate a custom model with your text prompt to create a video with a consistent character, style or element. Key Inputs Prompt: as descriptive a prompt as possible. Make sure to include the trigger word from your LoRA below Load LoRA: Load your reference model here Width & height: resolution settings are noted in pixels Guidance strength (CFG): Higher numbers adhere more to the prompt Flow Shift: For temporal consistency, adjust to tweak video smoothness.

LTX 2.3 Pro Text to Video

API

Audio

Text to Video

Video

LTX 2.3 Pro Text to Video

VibeVoice Text to Speech Multi Speaker

Audio

Multi Speaker

TTS

VibeVoice

Speech Multi Speaker

VibeVoice Text to Speech Multi Speaker

Speech Multi Speaker

ElevenLabs Text to Speech

API

Audio

ElevenLabs

Floyo API

TTS

ElevenLabs Text to Speech

ElevenLabs Text to Speech

ElevenLabs Text to Speech

LTX 2.3 Audio to Video

API

Audio

Audio to Video

LTX

Video

LTX 2.3 Audio to Video

Vertical Video Background & Scene Rebuild

Image

Image to Image

LoRAs

Qwen

Reactor

Upscale

Video

Wan2.1 FunControl

Vertical Video Background & Scene Rebuild

 LTX 2.3 - Extend Video

Audio

image to video

ltx 2

text to video

Video

video generation

Add seconds to an existing video with LTX 2.3. Upload a clip, set the duration and mode

LTX 2.3 - Extend Video

Add seconds to an existing video with LTX 2.3. Upload a clip, set the duration and mode

Audio

hailuo 3.0

minimax h3

reference to video

Video

video editing

Swap a character or object into an existing clip with MiniMax H3 (Hailuo 3.0), the open-weights editor. Add a reference image, describe the change, and hit run.

MiniMax H3 Open Weights - Reference to Video

Swap a character or object into an existing clip with MiniMax H3 (Hailuo 3.0), the open-weights editor. Add a reference image, describe the change, and hit run.

Audio

first last frame

hailuo 3.0

image to video

minimax h3

Video

Animate between two images with MiniMax H3 (Hailuo 3.0). Upload a first frame and an optional last frame, describe the motion, and hit run for a short 2K clip.

MiniMax H3 · First & Last Frame to Video

Animate between two images with MiniMax H3 (Hailuo 3.0). Upload a first frame and an optional last frame, describe the motion, and hit run for a short 2K clip.

api

Audio

hailuo 3.0

minimax h3

text to video

Video

video generation

Type a prompt and get back a native 2K video clip with synchronized sound from MiniMax H3 (Hailuo 3.0). Describe the scene, pick a length, and hit run.

MiniMax H3 · Text to Video

Type a prompt and get back a native 2K video clip with synchronized sound from MiniMax H3 (Hailuo 3.0). Describe the scene, pick a length, and hit run.

LTX 2.3 Video Inpainting · Video to Video

Audio

LoRAs

ltx2.3

object replacement

Video

video inpainting

video to video

Replace or add objects in a video with a single generation pass using LTX-Video 2.3 and a dedicated inpainting LoRA. Upload a video, draw a mask, describe the change, and hit run. Faster than the multi-pass version.

LTX 2.3 Video Inpainting · Video to Video

Replace or add objects in a video with a single generation pass using LTX-Video 2.3 and a dedicated inpainting LoRA. Upload a video, draw a mask, describe the change, and hit run. Faster than the multi-pass version.

SeedVR2 · Video Upscaler For Anime

4x upscale

seedvr2

Video

video restoration

video upscaler

Upload a low-resolution video and SeedVR2 upscales it to 1080p in a single diffusion step, restoring fine detail, removing compression artifacts, and maintaining temporal consistency across every frame.

SeedVR2 · Video Upscaler For Anime

Upload a low-resolution video and SeedVR2 upscales it to 1080p in a single diffusion step, restoring fine detail, removing compression artifacts, and maintaining temporal consistency across every frame.

Qwen Edit 2511: Multi-Angle Camera For Anime

camera control

Image

Image2Image

LoRAs

qwen

A image-editing workflow that generates a new camera angle of an image you provide. You load a source image, set horizontal_angle, vertical_angle, and zoom, and the node auto-writes a camera-instruction prompt (or you write your own) that's fed into the Qwen Image Edit 2511.

Qwen Edit 2511: Multi-Angle Camera For Anime

A image-editing workflow that generates a new camera angle of an image you provide. You load a source image, set horizontal_angle, vertical_angle, and zoom, and the node auto-writes a camera-instruction prompt (or you write your own) that's fed into the Qwen Image Edit 2511.

Z-Image Turbo: ControlNet Image to Image For Anime

Controlnet

depth

Image

image to image

z-image

This workflow uses Z-Image Turbo with Fun ControlNet Union to transform a reference image into a new style while preserving its original structure. The input image is processed with a ControlNet preprocessor to retain pose, composition and spatial layout.

Z-Image Turbo: ControlNet Image to Image For Anime

This workflow uses Z-Image Turbo with Fun ControlNet Union to transform a reference image into a new style while preserving its original structure. The input image is processed with a ControlNet preprocessor to retain pose, composition and spatial layout.

LTX2.3 Lip Sync · Image + Audio to Video For Anime

Audio

audio driven

image to video

lip sync

LoRAs

ltx2.3

Video

Upload a portrait and an audio file. LTX 2.3 generates a 10-second 1080p video where the character speaks or sings in sync with your audio, using a three-pass upscaling pipeline with vocal separation, NAG, and static camera locking.

LTX2.3 Lip Sync · Image + Audio to Video For Anime

Upload a portrait and an audio file. LTX 2.3 generates a 10-second 1080p video where the character speaks or sings in sync with your audio, using a three-pass upscaling pipeline with vocal separation, NAG, and static camera locking.

Wan 2.2 14B · Image to Video + End Frame For Anime

end frame

image to video

interpolation

start frame

Video

wan2.2

Upload a start image and an end image, describe the motion, and Wan 2.2 generates smooth video between them using a dual-model pipeline with high-noise and low-noise passes for maximum quality in 6 steps.

Wan 2.2 14B · Image to Video + End Frame For Anime

Upload a start image and an end image, describe the motion, and Wan 2.2 generates smooth video between them using a dual-model pipeline with high-noise and low-noise passes for maximum quality in 6 steps.

 Adding Sparkles to the Jewelry with Nano Banana 2

e-commerce

Image

image editing

image to image

jewelry

nano banana 2

product photography

Upload a jewelry photo and Nano Banana 2 adds natural light reflections, specular highlights, and prismatic refractions to every gemstone and diamond. Hit run.

Adding Sparkles to the Jewelry with Nano Banana 2

Upload a jewelry photo and Nano Banana 2 adds natural light reflections, specular highlights, and prismatic refractions to every gemstone and diamond. Hit run.

Z-Image Base · Text to Image For Anime

high details

Image

text to image

z-image

Write a prompt and Z-Image Base generates a photorealistic image at 1024x1024 with 30 steps and the res_multistep sampler, using the undistilled 6B model for maximum detail and quality.

Z-Image Base · Text to Image For Anime

Write a prompt and Z-Image Base generates a photorealistic image at 1024x1024 with 30 steps and the res_multistep sampler, using the undistilled 6B model for maximum detail and quality.

Qwen 2511 · Composite a Photoshoot For Anime

composite

Image

image to image

multi-reference

photoshoot

qwen image edit

Upload three images, a subject, an environment, and a prop or lighting reference, and Qwen Image Edit 2511 composites them into a single photorealistic editorial scene with matched lighting, scale, and perspective.

Qwen 2511 · Composite a Photoshoot For Anime

Upload three images, a subject, an environment, and a prop or lighting reference, and Qwen Image Edit 2511 composites them into a single photorealistic editorial scene with matched lighting, scale, and perspective.

Anima Preview 3 · Text to Image For Anime

anima

anime

character art

illustration

Image

text to image

Write a prompt using natural language or Danbooru tags and Anima Preview 3 generates four anime-style illustrations at once, using a 2-billion-parameter model built specifically for anime, character art, and non-photorealistic styles.

Anima Preview 3 · Text to Image For Anime

Write a prompt using natural language or Danbooru tags and Anima Preview 3 generates four anime-style illustrations at once, using a 2-billion-parameter model built specifically for anime, character art, and non-photorealistic styles.

Gemini Omni Flash: Video Editing

API

Audio

Gemini

Prompt-Based Editing

Video

video to video

A prompt-based video editing workflow that lets you edit an existing video with a plain-language instruction using Google's Gemini Omni Flash model.

Gemini Omni Flash: Video Editing

A prompt-based video editing workflow that lets you edit an existing video with a plain-language instruction using Google's Gemini Omni Flash model.

Gemini Omni Flash · Image to Video

ai video

audio

gemini omni flash

image to video

Video

Upload a starting image and describe the scene. Gemini Omni Flash by Google animates it into a 6-second video with native audio, cinematic motion, and synchronized sound effects.

Gemini Omni Flash · Image to Video

Upload a starting image and describe the scene. Gemini Omni Flash by Google animates it into a 6-second video with native audio, cinematic motion, and synchronized sound effects.

Gemini Omni Flash · Text to Video

ai video

audio

gemini omni flash

google

text to video

Video

Write a prompt describing a scene and Gemini Omni Flash by Google generates a video with native audio at up to 8 seconds in 16:9, ready to download as MP4.

Gemini Omni Flash · Text to Video

Write a prompt describing a scene and Gemini Omni Flash by Google generates a video with native audio at up to 8 seconds in 16:9, ready to download as MP4.

Nano Banana 2 Lite · Image Editing

Image

image to image

multi-reference

Nano banana lite

Upload up to 14 reference images and describe the change you want. Nano Banana 2 Lite, Google's fastest Gemini image model, applies the edit in about 4 seconds and returns the result.

Nano Banana 2 Lite · Image Editing

Upload up to 14 reference images and describe the change you want. Nano Banana 2 Lite, Google's fastest Gemini image model, applies the edit in about 4 seconds and returns the result.

Nano Banana Lite: Text to Image

API

Fast Generation

Image

Text to Image

A single-node text-to-image workflow that generates images from a written prompt using Google's Nano Banana Lite model (served via fal.ai). You write a prompt, pick an aspect ratio, and hit run — a built-in system prompt automatically enriches simple descriptions with composition

Nano Banana Lite: Text to Image

A single-node text-to-image workflow that generates images from a written prompt using Google's Nano Banana Lite model (served via fal.ai). You write a prompt, pick an aspect ratio, and hit run — a built-in system prompt automatically enriches simple descriptions with composition

Nano Banana 2 Lite: Text to Image

API

Google

Image

NanoBanana

Text to image

This workflow generates images from a text prompt using Nano Banana 2 Lite, Google's fast image model served through fal.ai. A built-in system prompt automatically expands simple prompts into detailed, well-composed scenes while respecting explicit instructions on style, color, l

Nano Banana 2 Lite: Text to Image

This workflow generates images from a text prompt using Nano Banana 2 Lite, Google's fast image model served through fal.ai. A built-in system prompt automatically expands simple prompts into detailed, well-composed scenes while respecting explicit instructions on style, color, l

Microsoft Lens Turbo · Text to Image

Image

lens turbo

microsoft lens

photorealistic

text to image

Write a prompt and Microsoft Lens Turbo generates a photorealistic image in 8 steps at 1280x720, using a 3.8B parameter model that matches the quality of models twice its size.

Microsoft Lens Turbo · Text to Image

Write a prompt and Microsoft Lens Turbo generates a photorealistic image in 8 steps at 1280x720, using a 3.8B parameter model that matches the quality of models twice its size.

FLUX.2 Max Edit · Image to Image

flux 2 max

Image

image editing

image to image

multi-image

photorealistic

Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.

FLUX.2 Max Edit · Image to Image

Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.

FLUX.2 Max · Text to Image

flux 2 max

Image

photorealistic

text rendering

text to image

Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.

FLUX.2 Max · Text to Image

Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.

Seedance 2.0 ASMR Unboxing · Reference to Video

asmr

bytedance

product video

reference to video

seedance2.0

unboxing

Video

Upload a product photo and a hand/style reference, and Seedance 2.0 generates a 10-second vertical ASMR unboxing video with tapping sounds, paper folding, and slow satisfying reveals.

Seedance 2.0 ASMR Unboxing · Reference to Video

Upload a product photo and a hand/style reference, and Seedance 2.0 generates a 10-second vertical ASMR unboxing video with tapping sounds, paper folding, and slow satisfying reveals.

Stable Audio 3 for Text to Music

ai music generator

Audio

comfyui

instrumental music

sound effects

stability ai

stable audio 3

text to audio

Describe the track or sound you want and Stable Audio 3 Medium, Stability AI's open text-to-audio model, generates it. Write a prompt, set a length, and hit run.

Stable Audio 3 for Text to Music

Describe the track or sound you want and Stable Audio 3 Medium, Stability AI's open text-to-audio model, generates it. Write a prompt, set a length, and hit run.

Krea 2 for Text to Image

Image

image generation

krea 2

krea 2 turbo

krea ai

LoRAs

lora styles

open source

text to image

Turn a text prompt into a finished image with Krea 2 Turbo, Krea AI's fast open-source image model. Type what you want to see, hit run, and get a result in seconds.

Krea 2 for Text to Image

Turn a text prompt into a finished image with Krea 2 Turbo, Krea AI's fast open-source image model. Type what you want to see, hit run, and get a result in seconds.

Texture to PBR · Image to Material

albedo

Image

material

normal map

pbr

texture

Upload a seamless texture and this workflow extracts a full PBR material set including albedo, normal, metallic, roughness, ambient occlusion, and height maps using Marigold and Lotus models.

Texture to PBR · Image to Material

Upload a seamless texture and this workflow extracts a full PBR material set including albedo, normal, metallic, roughness, ambient occlusion, and height maps using Marigold and Lotus models.

Happy Horse 1.1 · Reference to Video

ai video

character consistency

happy horse 1.1

reference to video

Video

Upload up to nine reference images of characters, objects, or scenes, and Happy Horse 1.1 generates a cinematic video that preserves their identity, style, and detail with synchronized audio.

Happy Horse 1.1 · Reference to Video

Upload up to nine reference images of characters, objects, or scenes, and Happy Horse 1.1 generates a cinematic video that preserves their identity, style, and detail with synchronized audio.

Happy Horse 1.1 · Text to Video

alibaba

dialogue

happy horse 1.1

text to video

Video

Describe a scene in plain language and Happy Horse 1.1 generates a cinematic video with synchronized audio, dialogue, and lip-sync at up to 1080p.

Happy Horse 1.1 · Text to Video

Describe a scene in plain language and Happy Horse 1.1 generates a cinematic video with synchronized audio, dialogue, and lip-sync at up to 1080p.

Qwen Image Layered for Image Deconstruction

background removal

concept art

Image

image to image

qwen

Qwen Image Layered

vfx

Upload one image and Qwen Image Layered pulls it apart into a clean foreground layer and a background layer you can edit, swap, or composite on their own.

Qwen Image Layered for Image Deconstruction

Upload one image and Qwen Image Layered pulls it apart into a clean foreground layer and a background layer you can edit, swap, or composite on their own.

Jewelry Animator and  Sparkle with Seedance 2.0

e-commerce

image to video

nano banana

product photography

seedance

Video

video generation

Animate jewelry shots with Seedance 2.0 while Nano Banana 2 adds natural sparkle, so the camera orbits the piece and light catches every facet as it turns.

Jewelry Animator and Sparkle with Seedance 2.0

Animate jewelry shots with Seedance 2.0 while Nano Banana 2 adds natural sparkle, so the camera orbits the piece and light catches every facet as it turns.

Z-Image Turbo Text-to-Image + SDA Diversity LoRA

Image

LoRAs

Text To Image

Z-Image

Z-Image Turbo at 10 steps with the SDA LoRA on top, so different seeds give different poses, angles, and compositions instead of variants of one image.

Z-Image Turbo Text-to-Image + SDA Diversity LoRA

Z-Image Turbo at 10 steps with the SDA LoRA on top, so different seeds give different poses, angles, and compositions instead of variants of one image.

VOID Video Inpainting + SAM3 Text Masking

Inpainting

Video

Video to Video

Remove objects from video with VOID's two-pass model. Type what to erase, SAM3 builds the mask, then VOID fills the holes coherently across every frame.

VOID Video Inpainting + SAM3 Text Masking

Remove objects from video with VOID's two-pass model. Type what to erase, SAM3 builds the mask, then VOID fills the holes coherently across every frame.

Seamless PBR Texture Workflow

Image

Image to Image

PBR texture

Turn any reference image into a tileable wood texture with full PBR maps: basecolor, normal, roughness, metalness, and height. Built for Unreal and Blender.

Seamless PBR Texture Workflow

Turn any reference image into a tileable wood texture with full PBR maps: basecolor, normal, roughness, metalness, and height. Built for Unreal and Blender.

Hunyuan 3D Pro - Text to 3D Model

3D

API

Hunyuan

Image

Text to 3D

Generate a textured 3D model from a text prompt with Hunyuan 3D Pro. Describe an object, hit Run, and get a GLB file plus interactive 3D viewer in under 60 seconds.

Hunyuan 3D Pro - Text to 3D Model

Generate a textured 3D model from a text prompt with Hunyuan 3D Pro. Describe an object, hit Run, and get a GLB file plus interactive 3D viewer in under 60 seconds.

 Qwen Image - Text to 360° HDRI Panorama

Image

Qwen

Text to Image

Qwen Image - Text to 360° HDRI Panorama

Nano Banana Pro - Game Art Restyling

api

Image

image to image

nano banana

style transfer

Nano Banana Pro - Game Art Restyling

Jewelry Scene Compositor with Nano Banana 2

e-commerce

Image

image to image

nano banana 2

product photography

Drop a ring photo and a background into the same workflow. Nano Banana 2 composites the jewelry into the scene seven ways so you can cherry-pick the best take.

Jewelry Scene Compositor with Nano Banana 2

Drop a ring photo and a background into the same workflow. Nano Banana 2 composites the jewelry into the scene seven ways so you can cherry-pick the best take.

Jewelry Environment Creator with Nano Banana 2

concept art

e-commerce

Image

image to image

nano banana 2

product photography

style transfer

Upload three reference images and Nano Banana 2 generates a new jewelry environment that matches their color, lighting, and visual style. Mood board to scene.

Jewelry Environment Creator with Nano Banana 2

Upload three reference images and Nano Banana 2 generates a new jewelry environment that matches their color, lighting, and visual style. Mood board to scene.

Happy Horse 1.0 Video Editing

consistency

film production

happy horse

style transfer

vid2vid

Video

video generation

Edit any video with Happy Horse 1.0 by uploading up to 5 reference images. Swap backgrounds, change subjects, or shift style. Original motion stays intact.

Happy Horse 1.0 Video Editing

Edit any video with Happy Horse 1.0 by uploading up to 5 reference images. Swap backgrounds, change subjects, or shift style. Original motion stays intact.

Happy Horse 1.0 - Text to Video

animation

film production

happy horse

text to video

Video

video generation

Generate cinematic video with synchronized audio from a text prompt using Alibaba's Happy Horse 1.0. Pick resolution, aspect ratio, and clip length up to 15s.

Happy Horse 1.0 - Text to Video

Generate cinematic video with synchronized audio from a text prompt using Alibaba's Happy Horse 1.0. Pick resolution, aspect ratio, and clip length up to 15s.

Image to Talking Video - LTX 2.3 + ElevenLabs UGC

Api

Audio

Audio to Video

Ltx2.3

Video

Image to Talking Video - LTX 2.3 + ElevenLabs UGC

Graphic Design Recomposer - Reframe Ads

Api

Image

Image to Image

Nano banana

Outpainting

Graphic Design Recomposer - Reframe Ads

Qwen 3.5 9B for Open Source LLM and VLM

image to text

llm

open source

qwen

text generation

vlm

Run Qwen 3.5 9B in ComfyUI as a text-only LLM or as a vision language model. Attach an image or a video, write your prompt, and get text back.

Qwen 3.5 9B for Open Source LLM and VLM

Run Qwen 3.5 9B in ComfyUI as a text-only LLM or as a vision language model. Attach an image or a video, write your prompt, and get text back.

Flux 2 Klein 9B + KV Cache for Image Editing

flux

flux 2 klein

Image

image to image

style transfer

Edit images with Flux 2 Klein 9B in 4 steps. KV Cache speeds every run by reusing attention work across steps. Upload an image, describe the edit, hit Run.

Flux 2 Klein 9B + KV Cache for Image Editing

Edit images with Flux 2 Klein 9B in 4 steps. KV Cache speeds every run by reusing attention work across steps. Upload an image, describe the edit, hit Run.

PixVerse C1 - Image to Video

animation

concept art

film production

image to video

pixverse

pixverse c1

Video

Animate a reference image into cinematic video with PixVerse C1. Pick your duration up to 15 seconds, resolution up to 1080p, and optional native audio.

PixVerse C1 - Image to Video

Animate a reference image into cinematic video with PixVerse C1. Pick your duration up to 15 seconds, resolution up to 1080p, and optional native audio.

PixVerse C1 - Text to Video

film production

pixverse

pixverse c1

text to video

vfx

Video

video generation

Generate cinematic video from text with PixVerse C1. Up to 1080p, up to 15 seconds, with optional native audio synchronized in the same generation pass.

PixVerse C1 - Text to Video

Generate cinematic video from text with PixVerse C1. Up to 1080p, up to 15 seconds, with optional native audio synchronized in the same generation pass.

Flux 2 Klein 9B Panorama Inpainting

flux

flux 2 klein

Image

image to image

inpainting

outpainting

panorama

Edit 360 panoramas with Flux 2 Klein 9B. Select a region, describe the change, and the edit gets composited back into your full panoramic image. No warping.

Flux 2 Klein 9B Panorama Inpainting

Edit 360 panoramas with Flux 2 Klein 9B. Select a region, describe the change, and the edit gets composited back into your full panoramic image. No warping.

Flux 2 Klein 9B + 360 Panorama ERP LoRA

flux

flux 2 klein

Image

image to image

lora

LoRAs

outpainting

panorama

Turn any image into a full 360 equirectangular panorama with Klein 9B and a 360 ERP outpaint LoRA. Cut flat camera shots at any angle from the result.

Flux 2 Klein 9B + 360 Panorama ERP LoRA

Turn any image into a full 360 equirectangular panorama with Klein 9B and a 360 ERP outpaint LoRA. Cut flat camera shots at any angle from the result.

Seedance 2.0 Fast - Image to Video with Audio

image to video

seedance 2.0

Video

video generation

Animate any image into video with ByteDance's Seedance 2.0 Fast. Built-in audio generation, start and end frame control, and multiple aspect ratios. No setup needed.

Seedance 2.0 Fast - Image to Video with Audio

Animate any image into video with ByteDance's Seedance 2.0 Fast. Built-in audio generation, start and end frame control, and multiple aspect ratios. No setup needed.

Z-Image Turbo + SDA LoRA for Diverse Text to Image

concept art

Image

lora

LoRAs

portrait

SDA

text to image

z-image turbo

Generate images with Z-Image Turbo while the SDA diversity LoRA stops every seed from producing the same pose and composition. 8 steps, 2x upscale to 2048.

Z-Image Turbo + SDA LoRA for Diverse Text to Image

Generate images with Z-Image Turbo while the SDA diversity LoRA stops every seed from producing the same pose and composition. 8 steps, 2x upscale to 2048.

Corridor Key Green Screen Keying for Video

background removal

film production

vfx

Video

video generation

Upload green screen footage and get a clean alpha matte plus composite preview. Corridor Key's neural network handles hair, motion blur, and transparency.

Corridor Key Green Screen Keying for Video

Upload green screen footage and get a clean alpha matte plus composite preview. Corridor Key's neural network handles hair, motion blur, and transparency.

Wan 2.7 Pro Unified Image Editing

concept art

e-commerce

Image

image to image

portrait

style transfer

text to image

wan

Upload an image, describe what you want changed, and Wan 2.7 Pro rewrites it. Style transfers, scene edits, and generation with thinking mode built in.

Wan 2.7 Pro Unified Image Editing

Upload an image, describe what you want changed, and Wan 2.7 Pro rewrites it. Style transfers, scene edits, and generation with thinking mode built in.

Wan 2.7 - Text to Video

Alibaba

Audio

Text to Video

Video

Wan 2.7

Generate video from a text prompt using Alibaba's Wan 2.7 model. Set your resolution, aspect ratio, and duration, then hit Run. Audio input supported.

Wan 2.7 - Text to Video

Generate video from a text prompt using Alibaba's Wan 2.7 model. Set your resolution, aspect ratio, and duration, then hit Run. Audio input supported.

BitDance 14B - Text to Image

bitdance

Image

T2V

text to image

Generate photorealistic images from text prompts using BitDance 14B, a 14-billion parameter autoregressive model that predicts up to 64 visual tokens per step.

BitDance 14B - Text to Image

Generate photorealistic images from text prompts using BitDance 14B, a 14-billion parameter autoregressive model that predicts up to 64 visual tokens per step.

Qwen3-VL Image and Video Captioning

Captioning

LLM

Prompt Generator

Qwen3VL

VLM

Upload an image or video and get a detailed text description from Qwen3-VL. Choose your model size, pick a preset prompt, or write your own. Runs in your browser.

Qwen3-VL Image and Video Captioning

Upload an image or video and get a detailed text description from Qwen3-VL. Choose your model size, pick a preset prompt, or write your own. Runs in your browser.

Whisper Speech-to-Text and SRT Subtitle Generator

audio

speech to text

srt

STT

subtitles

transcription

whisper

Upload any audio file and Whisper transcribes it into text with word-level and segment-level SRT subtitle files. Auto language detection included.

Whisper Speech-to-Text and SRT Subtitle Generator

Upload any audio file and Whisper transcribes it into text with word-level and segment-level SRT subtitle files. Auto language detection included.

Auto Subtitles with Whisper - Video to Video

subtitling

vid2vid

Video

video generation

Upload a video and get it back with burned-in subtitles. Whisper transcribes the audio, then the text gets placed frame-by-frame with word-level timing.

Auto Subtitles with Whisper - Video to Video

Upload a video and get it back with burned-in subtitles. Whisper transcribes the audio, then the text gets placed frame-by-frame with word-level timing.

Z-Image Turbo Inpainting

controlnet

Image

inpainting

z-image-turbo

Z-Image Turbo Inpainting

Z-Image Turbo Inpainting

Z-Image Turbo Inpainting

Voice Changer using TTS Audio Suite (ChatterBox)

audio

Audio2Audio

Chatterbox

tts

TTS Audio Suite

voice conversion

Convert any voice to match a target speaker using ChatterBox TTS. Upload source and narrator audio, run it, get back a converted MP3. No voice training needed.

Voice Changer using TTS Audio Suite (ChatterBox)

Convert any voice to match a target speaker using ChatterBox TTS. Upload source and narrator audio, run it, get back a converted MP3. No voice training needed.

Recraft V3 Text to Image

API

FloyoAPI

Image

Recraft

Text2Image

Generate images with Recraft V3 from a text prompt. Choose a preset size or custom dimensions, pick a style, and run.

Recraft V3 Text to Image

Generate images with Recraft V3 from a text prompt. Choose a preset size or custom dimensions, pick a style, and run.

FireRed Image Edit - Makeup Transfer

Flux

Image

Image to image

Opensource

Style transfer

Transfer makeup from a reference photo onto a portrait using FireRed Image Edit 1.1. Upload a face and a makeup reference, and the model applies the look while keeping pose and facial features intact.

FireRed Image Edit - Makeup Transfer

Transfer makeup from a reference photo onto a portrait using FireRed Image Edit 1.1. Upload a face and a makeup reference, and the model applies the look while keeping pose and facial features intact.

Kling O3 Video to Video — Standard Reference

API

Video

Video to Video

Kling O3 Video to Video — Standard Reference

Kling Image to Video with Reference Control

API

image to video

kling

text to video

Video

Kling Image to Video with Reference Control

Capybara for Image Editing

Capybara

Image2Image

Image Editing

Video

Edit your cool images using Capybara

Capybara for Image Editing

Edit your cool images using Capybara

LongCat for Text to Image

Image

LongCat

Text2Image

Create cool images using the LongCat

LongCat for Text to Image

Create cool images using the LongCat

FLUX.2 Klein 9B · Image to Image

FLUX

Flux.2 Klein

Image

Image2Image

Image Editing

LoRA

LoRAs

Edit images with consistency of the subject or things using Flux.2 Klein 9B and a LoRA

FLUX.2 Klein 9B · Image to Image

Edit images with consistency of the subject or things using Flux.2 Klein 9B and a LoRA

FLUX.2 Klein 9B + Virtual Tryon LoRA

Flux

Flux.2 Klein

Image

LoRAs

Tryon

VTON

Try a clothes using Flux.2 Klein 9B and tryon LoRA from

FLUX.2 Klein 9B + Virtual Tryon LoRA

Try a clothes using Flux.2 Klein 9B and tryon LoRA from

Sopro for Text to Speech

Audio

Audio2Audio

SoproTTS

Text to Speech

TTS

Turn your text to excellent speech using SoproTTS

Sopro for Text to Speech

Turn your text to excellent speech using SoproTTS

SopranoTTS for Text to Speech

Audio

Soprano

Text to Speech

TTS

Turn speech using Soprano TTS

SopranoTTS for Text to Speech

Turn speech using Soprano TTS

LTX 2.0 – Prompting & Dynamic Camera Movement

Audio

LoRAs

Opensource

Text to Video

Video

LTX 2.0 – Prompting & Dynamic Camera Movement

Whisper STT

AILab

Audio to Text

Speech to Text

STT

Transcribe

Create a text from speech using Whisper STT

Whisper STT

Create a text from speech using Whisper STT

ACE-Step 1.5 for Music Generation

ACE-Step 1.5

Audio

Music Generation

Text to Audio

Create stunning music using ACE Step 1.5

ACE-Step 1.5 for Music Generation

Create stunning music using ACE Step 1.5

Qwen Image Edit 2511 and VNCCS Utils - Visual Pose

Image

Image Editing

Qwen

Qwen Image Edit 2511

VNCCS Utils

Create different position of person using VNCCS custom node and Qwen Image Edit 2511

Qwen Image Edit 2511 and VNCCS Utils - Visual Pose

Create different position of person using VNCCS custom node and Qwen Image Edit 2511

Z-Image Base - Text to Image w/ LoRA

Base

Image

LoRA

LoRAs

Text to Image

Z-image

Run Z-Image Base with a custom LoRA

Z-Image Base - Text to Image w/ LoRA

Run Z-Image Base with a custom LoRA

Vidu Q3 for Image to Video

Animation

Image2Video

Video

Vidu Q3

Turn to images to real life

Vidu Q3 for Image to Video

Turn to images to real life

Seedream 4.5 Unified for Image Generation

Image

Image2Image

Image Editing

Seedream 4.5

Text2Image

typography

An all purpose Seedream 4.5 for image generation

Seedream 4.5 Unified for Image Generation

An all purpose Seedream 4.5 for image generation

Minimax Speech 2.8 HD for Text to Speech

Audio

Minimax

Minimax Speech 2.8 HD

TTS

Create realistic speech using Minimax speech 2.8

Minimax Speech 2.8 HD for Text to Speech

Create realistic speech using Minimax speech 2.8

Qwen Multiangle Light with Qwen Image Edit 2511

Image

Image Edit

Qwen

Qwen Image Edit 2511

Relighting

Relighting images using Qwen multiangle light node

Qwen Multiangle Light with Qwen Image Edit 2511

Relighting images using Qwen multiangle light node

Grok Imagine: Fast Text to Image

Grok

Image

photorealism

Text2Image

Create cool images using Grok Imagine

Grok Imagine: Fast Text to Image

Create cool images using Grok Imagine

Qwen Image Max Edit for Editing Images

API

Image

Image2Image

Image Editing

Qwen Image Max Edit

Editing images using the flagship model of Qwen Image Max Edit

Qwen Image Max Edit for Editing Images

Editing images using the flagship model of Qwen Image Max Edit

Pixverse Swap for Image to Video Swap

Floyo API

Image2Video

PixVerse

Video

You can swap object,, character and background using PixVerse

Pixverse Swap for Image to Video Swap

You can swap object,, character and background using PixVerse

Kandinsky for Text to Video

Filmmaking

Kandinsky

Text2Video

Video

Videography

Creating excellent videos using Kandinsky

Kandinsky for Text to Video

Creating excellent videos using Kandinsky

3D Products with Logo - Wan2.6 Image to Video

Image

Image to Video

Text to Image

Video

Wan2.6

3D Products with Logo - Wan2.6 Image to Video

Create Product Demo from Concept to Video

GPT-Image 1.5

Image

Image2Image

Image2Video

Kling 2.6

Text2Image

Video

VLM

Create a high quality demo for your products using Kling 2.6 Image to Video

Create Product Demo from Concept to Video

Create a high quality demo for your products using Kling 2.6 Image to Video

Insert Product into Existing Ad

Ecommerce

Image

NanoBanana

Reference Image

Insert Product into Existing Ad

Character Reshoot using Qwen Edit 2511 + Kling O1

Audio

Image

Image2Video

Kling Omni One

LoRAs

Next Scene LoRA

Qwen Image Edit 2511

Reference2Video

Video

Creating a reshoot for a character

Character Reshoot using Qwen Edit 2511 + Kling O1

Creating a reshoot for a character

LTX 2 19B Pro for Text to Video

Audio

Flimography

LTX 2 Pro

Open Source

Text2Video

Video

Videography

An open source LTX 2 Pro for Text to Video

LTX 2 19B Pro for Text to Video

An open source LTX 2 Pro for Text to Video

Video Detailer using LTX 2 Vid2Vid

Audio

LoRAs

LTX 2

Vid2Vid

Video

Video Detailer

Video Editing

It can enhance the detail of the video

Video Detailer using LTX 2 Vid2Vid

It can enhance the detail of the video

Camera Angle Creation using Image2Vid

Camera Control

Image

Image2Vid

LoRAs

Qwen Image Edit 2511

Vid2Vid

Video

Wan2.6

Using witness cameras to recreate additional shots that were not captured by principal photography

Camera Angle Creation using Image2Vid

Using witness cameras to recreate additional shots that were not captured by principal photography

Multi Model for Voice Convesion and Text to Speech

Audio

ChatterBox

Higgs

Text to Speech

TTS

VibeVoice

A workflow of TTS Audio Suite which can to use different type of audio models.

Multi Model for Voice Convesion and Text to Speech

A workflow of TTS Audio Suite which can to use different type of audio models.

Create Cinematic Poster & Ad from Your Product

Image

poster-design

product-ad

Seedream

VLM

Create Cinematic Poster & Ad from Your Product

Wan2.1 + SCAIL for Animating Images for Movement

Image2Video

SCAIL

Video

Wan

Wan2.1 + SCAIL for Animating Images for Movement

Change Product Shots with NanoBanana Pro

Ecommerce

Image

Image to Image

NanoBanana

Reference image

Change Product Shots with NanoBanana Pro

joy caption fine controls

captions

#dataset

detailed

Image

Lora

LoRAs

tool

training

joy caption fine controls

Vertical Video Prop & Object Replacement Using Seedream + Wan 2.2

Image

Image to image

LoRAs

Reference Video

Seedream

Video

Wan2.2

Vertical Video Prop & Object Replacement Using Seedream + Wan 2.2

Veo 3.1 Image to Video

API

Audio

Floyo API

Image2Video

Veo 3.1

Video

Veo 3.1 Image to Video

Kling Omni 1 Reference to Video

API

Audio

Floyo

Kling Omni One

Reference2Vid

Video

Kling Omni 1 Reference to Video

Kling 2.5 Image to Video

Animation

API

Filmography

Floyo

Floyo API

Image2Video

Kling 2.5

Video

Kling 2.5 Image to Video

Vertical Video Scene Extension & Coverage Generator using Seedream +Wan

first-last frame

Image

reference image

Seedream

Video

wan2.2

Vertical Video Scene Extension & Coverage Generator using Seedream +Wan

Vertical Video Scene Extension & Coverage Generator

first-last frame

Image

LoRAs

qwen

reference-image

Video

wan2.2

Vertical Video Scene Extension & Coverage Generator

Realistic Product or Props Replacement

Animate

Image

LoRAs

Qwen_2509

Realistic

Video

wan2.2

Realistic Product or Props Replacement

MiniMax Text-to-Video will Bring Your Creative Concepts to Life with Realistic Motion

API

Floyo API

Minimax

Text2Video

Video

MiniMax Text-to-Video will Bring Your Creative Concepts to Life with Realistic Motion

Boost Your Creative Video: Comprehensive Solutions

API

Floyo API

Image2Video

Seedance

Video

Boost Your Creative Video: Comprehensive Solutions

Modify the Image using InstantX Union ControlNet

Controlnet

Image

Image2Image

InstantX Union Controlnet

Qwen

Modify the Image using InstantX Union ControlNet

Kling 3.0 for Video Generation

Image2Video

Kling 3.0

Text2Video

Video

Coming soon page for Kling 3.0

Kling 3.0 for Video Generation

Coming soon page for Kling 3.0

LTX 2 Retake Video for Video Editing

API

Audio

Floyo API

LTX 2 Retake

Video

Video2Video

Video Editing

Outdated model. Please go to LTX 2.3 Retake Video workflow to use LTX 2.3

LTX 2 Retake Video for Video Editing

Outdated model. Please go to LTX 2.3 Retake Video workflow to use LTX 2.3

LTX 2 Fast API for Text to Video

API

Audio

Filmmaking

Filmography

Floyo API

LTX 2 Fast

Video

Outdated model. Please go to LTX 2.3 Text to Video workflow to use LTX 2.3

LTX 2 Fast API for Text to Video

Outdated model. Please go to LTX 2.3 Text to Video workflow to use LTX 2.3

Amazon Bedrock - Text to Multi-Image with SDXL, Titan and Nova Canvas

API

Bedrock

Image

Nova Canvas

SDXL

Text to Image

Titan

Generate and compare images between 3 different models powered by Amazon Bedrock. Key Inputs Prompt: as descriptive a prompt as possible Models SDXL: Solid all-around performer with strong prompt adherence and wide style range Titan: Versatile model with built-in editing features and customization flexibility Nova Canvas: Quick iterations with creative flair, ideal for brainstorming and concept exploration

Amazon Bedrock - Text to Multi-Image with SDXL, Titan and Nova Canvas

Generate and compare images between 3 different models powered by Amazon Bedrock. Key Inputs Prompt: as descriptive a prompt as possible Models SDXL: Solid all-around performer with strong prompt adherence and wide style range Titan: Versatile model with built-in editing features and customization flexibility Nova Canvas: Quick iterations with creative flair, ideal for brainstorming and concept exploration

DyPE + Z-Turbo · Text to Image For Short Drama

dype

high resolution

Image

LoRAs

text to image

z-turbo

Generate sharp 2K images from a text prompt using Z-Image Turbo with DyPE resolution scaling and a DeJPEG cleanup LoRA. Type a prompt and hit run.

DyPE + Z-Turbo · Text to Image For Short Drama

Generate sharp 2K images from a text prompt using Z-Image Turbo with DyPE resolution scaling and a DeJPEG cleanup LoRA. Type a prompt and hit run.

Z-Img Turbo + SDA · Text to Image For Short Drama

diversity

Image

LoRAs

photorealistic

sda

text to image

Write a prompt and Z-Image Turbo generates a photorealistic image in 8 steps, with the SDA diversity LoRA ensuring each seed produces a distinct composition, pose, and camera angle. Output upscaled to 2048x2048.

Z-Img Turbo + SDA · Text to Image For Short Drama

Write a prompt and Z-Image Turbo generates a photorealistic image in 8 steps, with the SDA diversity LoRA ensuring each seed produces a distinct composition, pose, and camera angle. Output upscaled to 2048x2048.

IndexTTS2 Voice Cloning with Emotion Control

Audio

text to speech

voice cloning

Emotion Control

IndexTTS2 Voice Cloning with Emotion Control

Emotion Control

Morse Code to Speech

accessibility

amateur radio

Audio

audio processing

chatterbox tts

morse code

text to speech

transcription

Upload Morse code audio, decode it to readable text, and convert to natural speech using ChatterboxTTS. Perfect for amateur radio transcription and accessibility.

Morse Code to Speech

Upload Morse code audio, decode it to readable text, and convert to natural speech using ChatterboxTTS. Perfect for amateur radio transcription and accessibility.

Microsoft Lens Turbo · Text to Img For Short Drama

Image

lens turbo

photorealistic

text to image

Write a prompt and Microsoft Lens Turbo generates a photorealistic image in 8 steps at 1280x720, using a 3.8B parameter model that matches the quality of models twice its size.

Microsoft Lens Turbo · Text to Img For Short Drama

Write a prompt and Microsoft Lens Turbo generates a photorealistic image in 8 steps at 1280x720, using a 3.8B parameter model that matches the quality of models twice its size.

FLUX.2 Consistency LoRA · I2I For Short Drama

consistency lora

flux

Image

image to image

klein

LoRAs

Edit any image with a text instruction using FLUX.2 Klein 9B by Black Forest Labs, with a Consistency LoRA to keep results close to the original. Upload a photo, describe the change, and hit run.

FLUX.2 Consistency LoRA · I2I For Short Drama

Edit any image with a text instruction using FLUX.2 Klein 9B by Black Forest Labs, with a Consistency LoRA to keep results close to the original. Upload a photo, describe the change, and hit run.

Kling 3.0 Standard for Image to Video

Image2Video

Kling

Kling 3.0 Standard

Video

Animate the images using Kling 3.0 Standard

Kling 3.0 Standard for Image to Video

Animate the images using Kling 3.0 Standard

Kling 3.0 Standard for Text to Video

Filmography

Kling

Kling 3.0 Standard

Text2Video

Video

Create videos using Kling 3.0 Standard

Kling 3.0 Standard for Text to Video

Create videos using Kling 3.0 Standard

FLUX Luxury Furniture Generator with LoRA Enhancem

Flux

LoRAs

text to video

Video

Generate high-end furniture and interior design images with FLUX and multiple LoRAs. Perfect for product catalogs, showrooms, and design visualization.

FLUX Luxury Furniture Generator with LoRA Enhancem

Generate high-end furniture and interior design images with FLUX and multiple LoRAs. Perfect for product catalogs, showrooms, and design visualization.

Qwen Multiangle Light · Img to Img For Short Drama

Image

image to image

multiangle

qwen image edit

Upload one image and this workflow generates five relighted versions simultaneously, each lit from a different angle and direction, with full control over azimuth, elevation, intensity, and color.

Qwen Multiangle Light · Img to Img For Short Drama

Upload one image and this workflow generates five relighted versions simultaneously, each lit from a different angle and direction, with full control over azimuth, elevation, intensity, and color.

Change Emotion using IndexTTS2

Audio

Audio to Audio

voice clone

Clone any voice and change the emotional delivery. Upload audio, type new text, adjust emotions, and get speech with different feelings using the same voice.

Change Emotion using IndexTTS2

Clone any voice and change the emotional delivery. Upload audio, type new text, adjust emotions, and get speech with different feelings using the same voice.

Orion 4D QR Code Generator from Text or URL

e-commerce

Image

QR Code

text to image

Type a URL or any text and get three scannable codes at once: a standard QR code, an Aztec barcode, and a branded creative QR with your logo in the center.

Orion 4D QR Code Generator from Text or URL

Type a URL or any text and get three scannable codes at once: a standard QR code, an Aztec barcode, and a branded creative QR with your logo in the center.

Audio to Modem Encryption using Orion4D Secret

audio

audio encryption

comfyui utility

encryption

modem

orion4d

text to audio

Encode text into audio signals using three modem methods (AFSK, OFDM, Spectral), then decode and verify integrity. Built with Orion4D nodes in ComfyUI.

Audio to Modem Encryption using Orion4D Secret

Encode text into audio signals using three modem methods (AFSK, OFDM, Spectral), then decode and verify integrity. Built with Orion4D nodes in ComfyUI.

Meshy v6 Text to 3D Model

3D

3D Asset

Meshy v6

Text to 3D

Create a 3D using Meshy v6 text to model

Meshy v6 Text to 3D Model

Create a 3D using Meshy v6 text to model

Vidu Q3 for Text to Video

Text2Video

Video

Videography

Vidu Q3

Create good videos with Vidu Q3

Vidu Q3 for Text to Video

Create good videos with Vidu Q3

Qwen Edit 2511: Multi-Angle Camera For Short Drama

camera control

Image

image to image

LoRAs

qwen

A image-editing workflow that generates a new camera angle of an image you provide. You load a source image, set horizontal_angle, vertical_angle, and zoom, and the node auto-writes a camera-instruction prompt (or you write your own) that's fed into the Qwen Image Edit 2511.

Qwen Edit 2511: Multi-Angle Camera For Short Drama

A image-editing workflow that generates a new camera angle of an image you provide. You load a source image, set horizontal_angle, vertical_angle, and zoom, and the node auto-writes a camera-instruction prompt (or you write your own) that's fed into the Qwen Image Edit 2511.

FLUX.2 Klein Image Expansion · I2I For Short Drama

flux 2 klein

Image

image expansion

outpaint

Video

Expand any photo with new content on any side using FLUX.2 Klein 9B for outpainting and SeedVR2 7B Sharp for upscaling. Upload a photo, set the expansion width, and hit run.

FLUX.2 Klein Image Expansion · I2I For Short Drama

Expand any photo with new content on any side using FLUX.2 Klein 9B for outpainting and SeedVR2 7B Sharp for upscaling. Upload a photo, set the expansion width, and hit run.

Audio to Spectogram using Orion4D Secret

audio

audio encryption

Image

image to image

spectrogram

Turn any image into audio by encoding it as a spectrogram. Upload a picture, set your frequency range, and get an MP3 that shows your image when visualized.

Audio to Spectogram using Orion4D Secret

Turn any image into audio by encoding it as a spectrogram. Upload a picture, set your frequency range, and get an MP3 that shows your image when visualized.

Kling 3.0 Pro for Text to Video

Filmography

Kling 3.0 Pro

Text2Video

Video

Create videos using Kling 3.0

Kling 3.0 Pro for Text to Video

Create videos using Kling 3.0

MMAudio V2: Add Sound Effects to Any Video

Audio

audio generation

foley

mmaudio v2

sound design

Video

video to audio

video to video

Generate synchronized sound effects for any video using MMAudio V2, the open-source video-to-audio model from Sony AI and University of Illinois. Upload a video, describe the sounds, and hit run.

MMAudio V2: Add Sound Effects to Any Video

Generate synchronized sound effects for any video using MMAudio V2, the open-source video-to-audio model from Sony AI and University of Illinois. Upload a video, describe the sounds, and hit run.

Audio Watermarking using Orion 4D Secret

audio

audio watermarking

content tracking

provenance

steganography

Upload any audio file, write the text to hide, and get back two watermarked copies. Steganography packs more data. Frequency encoding survives MP3 compression.

Audio Watermarking using Orion 4D Secret

Upload any audio file, write the text to hide, and get back two watermarked copies. Steganography packs more data. Frequency encoding survives MP3 compression.

Text Manipulator

Cipher

Text

Encode plain text into cipher characters and decode it back. Paste a message, pick a cipher style, and get a disguised version only the matching decoder reads.

Text Manipulator

Encode plain text into cipher characters and decode it back. Paste a message, pick a cipher style, and get a disguised version only the matching decoder reads.

Segment Anything 2 for Creating Video Mask

SAM2

Segment Anything 2

Video

video2video

Video Mask

Create a video mark frame by frame using Segment Anything 2

Segment Anything 2 for Creating Video Mask

Create a video mark frame by frame using Segment Anything 2

Minimax Music 2.6 - Text to Music

Audio

instrumental

minimax music 2.6

music generation

song generation

soundtrack

text to music

Text-to-music with Minimax Music 2.6. Generate songs with vocals and backing from a style prompt and lyrics, or toggle instrumental mode for score only.

Minimax Music 2.6 - Text to Music

Text-to-music with Minimax Music 2.6. Generate songs with vocals and backing from a style prompt and lyrics, or toggle instrumental mode for score only.

Pixverse C1 Reference to Video

character design

consistency

film production

image to video

pixverse

pixverse c1

Video

video generation

Upload up to 7 reference images, tag each one, and generate a video that composes your characters, objects, and backgrounds into one scene with Pixverse C1.

Pixverse C1 Reference to Video

Upload up to 7 reference images, tag each one, and generate a video that composes your characters, objects, and backgrounds into one scene with Pixverse C1.

Pixverse C1 Transition

animation

film production

image to video

pixverse

transition

Video

video generation

Upload a start image and end image, describe the motion, and Pixverse C1 generates a video transition between them. Set duration, resolution, and audio.

Pixverse C1 Transition

Upload a start image and end image, describe the motion, and Pixverse C1 generates a video transition between them. Set duration, resolution, and audio.

MoGe-2 · Panorama to 3D Mesh

3D

3d mesh

image to 3d

moge

panorama to 3d

Turn a 360 panorama into a textured 3D mesh with MoGe-2, Microsoft's geometry model. Upload a panorama, hit run, and download a GLB you can open anywhere.

MoGe-2 · Panorama to 3D Mesh

Turn a 360 panorama into a textured 3D mesh with MoGe-2, Microsoft's geometry model. Upload a panorama, hit run, and download a GLB you can open anywhere.

TripoSplat · Image to 3D Turnaround

3D

3d asset

gaussian splatting

image to 3d

triposplat

Video

Turn one photo into a 3D Gaussian splat with TripoSplat by VAST-AI, plus colour, depth, normal, and clay turntable videos. Upload an image and hit run.

TripoSplat · Image to 3D Turnaround

Turn one photo into a 3D Gaussian splat with TripoSplat by VAST-AI, plus colour, depth, normal, and clay turntable videos. Upload an image and hit run.

Qwen Image Max for Text to Image

Image

Qwen Image Max

Text2Image

Create a high quality using the flagship model of Qwen Image

Qwen Image Max for Text to Image

Create a high quality using the flagship model of Qwen Image

Qwen Thinking Prompt Refiner

prompt-refinement

qwen-thinking

Qwen Thinking Prompt Refiner

LTX-2.3 · Motion Transfer (Pose)

Audio

LoRAs

ltx

motion transfer

Video

video to video

Transfer the movement from any reference video onto your portrait with LTX-2.3 by Lightricks, audio and all. Upload a photo and a clip, then hit run.

LTX-2.3 · Motion Transfer (Pose)

Transfer the movement from any reference video onto your portrait with LTX-2.3 by Lightricks, audio and all. Upload a photo and a clip, then hit run.

FLUX.2 Klein 4B for Text to Sprite Sheet

Flux

Flux.2 Klein 4B

Image

Image2Image

LoRAs

Create sprite sheet for game characters in using flux 2 Klein 4B

FLUX.2 Klein 4B for Text to Sprite Sheet

Create sprite sheet for game characters in using flux 2 Klein 4B

Moonvalley Marey - Text to Video

cinematic

film production

moonvalley marey

text to video

Video

video generation

Generate cinematic 5 or 10 second 1080p video clips from a text prompt with Moonvalley Marey, a commercially-safe model trained only on licensed footage.

Moonvalley Marey - Text to Video

Generate cinematic 5 or 10 second 1080p video clips from a text prompt with Moonvalley Marey, a commercially-safe model trained only on licensed footage.

FLUX.2 Klein 4B for Image Outpainting

Flux

Flux.2 Klein

Image

Image Outpainting

LoRAs

Outpaint image using Flux 2 Klein 4B using LanPaint and Outpaint LoRA

FLUX.2 Klein 4B for Image Outpainting

Outpaint image using Flux 2 Klein 4B using LanPaint and Outpaint LoRA

Moonvalley Marey Image to Video

film production

image to video

marey

moonvalley

vfx

Video

video generation

Turn an image into cinematic 1080p video with Marey, Moonvalley's video model trained on licensed footage. 5s or 10s clips at 24fps, safe for commercial work.

Moonvalley Marey Image to Video

Turn an image into cinematic 1080p video with Marey, Moonvalley's video model trained on licensed footage. 5s or 10s clips at 24fps, safe for commercial work.

Qwen Image Edit – Portrait Light Migration

Image

Image to Image

Lightning

Portrait

Qwen

Qwen Image Edit – Portrait Light Migration

Moonvalley Marey Motion Transfer - Video to Video

vid2vid

Video

video generation

Moonvalley Marey Motion Transfer - Video to Video

Isometric Miniatures from a Selfie

2*2 Grid

Image

NanoBanana

Isometric Miniatures from a Selfie

Moonvalley Marey Pose Transfer - Video to Video

vid2vid

Video

video generation

Moonvalley Marey Pose Transfer - Video to Video

LongCat-Image-Edit - Instruction Image Editing

concept art

consistency

Image

image to image

longcat-image-edit

portrait

style transfer

Upload one image, write an instruction, and LongCat-Image-Edit rewrites the parts you describe while keeping the rest identical. Bilingual prompts, 8 steps.

LongCat-Image-Edit - Instruction Image Editing

Upload one image, write an instruction, and LongCat-Image-Edit rewrites the parts you describe while keeping the rest identical. Bilingual prompts, 8 steps.

ACE-Step 1.5 XL - Text to Music

ace-step

ace-step 1.5 XL

Audio

audio generation

instrumental

lyrics

music generation

text to music

Generate full songs with ACE-Step 1.5 XL Base. Write a style prompt, add structured lyrics like [Intro] [Verse] [Chorus], pick BPM and key, get an MP3.

ACE-Step 1.5 XL - Text to Music

Generate full songs with ACE-Step 1.5 XL Base. Write a style prompt, add structured lyrics like [Intro] [Verse] [Chorus], pick BPM and key, get an MP3.

LTX-2.3 · Clean Plate (Video Object Removal)

Audio

clean plate

LoRAs

ltx-2

object removal

Video

video inpainting

Erase people and moving objects from a video and rebuild the scene behind them with LTX-2.3 by Lightricks. Upload a clip, describe the empty shot, hit run.

LTX-2.3 · Clean Plate (Video Object Removal)

Erase people and moving objects from a video and rebuild the scene behind them with LTX-2.3 by Lightricks. Upload a clip, describe the empty shot, hit run.

Nano Banana Pro Storyboard: 1 Image to 5 Shots

Image

image to image

nano banana

Nano Banana Pro Storyboard: 1 Image to 5 Shots

LTX-2.3 · Video Colorization

Audio

black and white

colorize video

LoRAs

ltx-2

Video

video colorization

video restoration

Colorize black-and-white or faded video with LTX-2.3 by Lightricks. A local Gemma model reads the scene and writes the colour prompt, then you hit run.

LTX-2.3 · Video Colorization

Colorize black-and-white or faded video with LTX-2.3 by Lightricks. A local Gemma model reads the scene and writes the colour prompt, then you hit run.

ERNIE Image - Text to Image

concept art

ernie image

Image

prompt enhancement

text to image

Generate images with Baidu's ERNIE Image model. Write a short prompt and let the built-in AI enhancer expand it into rich detail. Toggle the enhancer on or off.

ERNIE Image - Text to Image

Generate images with Baidu's ERNIE Image model. Write a short prompt and let the built-in AI enhancer expand it into rich detail. Toggle the enhancer on or off.

FLUX.2 Klein 9B · Multi-image Edit + Upscale

flux2

flux.2 klein

Image

image editing

image to image

Edit one photo using a second as reference, then refine and 4x upscale it with FLUX.2 Klein 9B by Black Forest Labs. Upload two images, describe the edit, run.

FLUX.2 Klein 9B · Multi-image Edit + Upscale

Edit one photo using a second as reference, then refine and 4x upscale it with FLUX.2 Klein 9B by Black Forest Labs. Upload two images, describe the edit, run.

Dissolve image lighting Kontext V2

Flux

Image

Image to Image

Dissolve image lighting Kontext V2

Qwen3 ASR: Transcribe Audio

asr

audio

qwen

speech to text

subtitles

transcription

Upload audio and Qwen3's ASR engine returns the transcript, word-level timing for SRT subtitles, and an optional translation to English. Language auto-detected.

Qwen3 ASR: Transcribe Audio

Upload audio and Qwen3's ASR engine returns the transcript, word-level timing for SRT subtitles, and an optional translation to English. Language auto-detected.

Trellis 2 Image to 3D

3D

3d generation

glb export

image to 3d

pbr materials

textured mesh

trellis 2

Upload an image and Trellis 2 builds a textured 3D mesh with PBR materials. Outputs GLB ready for Blender, Unity, or Unreal in about a minute on an H100.

Trellis 2 Image to 3D

Upload an image and Trellis 2 builds a textured 3D mesh with PBR materials. Outputs GLB ready for Blender, Unity, or Unreal in about a minute on an H100.

LongCat AudioDiT for Voice Clone

Audio

audio generation

film production

longcat

text to speech

voice cloning

voiceover

Clone any voice from a short audio sample with LongCat AudioDiT 3.5B. Upload a reference clip, type what you want it to say, and get speech in that voice.

LongCat AudioDiT for Voice Clone

Clone any voice from a short audio sample with LongCat AudioDiT 3.5B. Upload a reference clip, type what you want it to say, and get speech in that voice.

Krea 2 · Drawing to Image

drawing to image

Image

image editing

krea 2

krea 2 turbo

LoRAs

sketch to image

Turn a rough sketch into a photorealistic image that keeps your layout, using Krea 2 Turbo and an Identity Edit LoRA. Upload a drawing, describe it, hit run.

Krea 2 · Drawing to Image

Turn a rough sketch into a photorealistic image that keeps your layout, using Krea 2 Turbo and an Identity Edit LoRA. Upload a drawing, describe it, hit run.

Create Magazine Cover & Package Design

Ecommerce

Image

NanoBanana

Reference Image

Create Magazine Cover & Package Design

Mage Flow Edit· AI Image Editing

ai image editing

Image

image editing

mage flow edit

Edit, combine, or transform one or two images using natural language. Change backgrounds, merge scenes, swap outfits, add objects, or create entirely new compositions in just a few seconds.

Mage Flow Edit· AI Image Editing

Edit, combine, or transform one or two images using natural language. Change backgrounds, merge scenes, swap outfits, add objects, or create entirely new compositions in just a few seconds.

VoxCPM2 for Text to Speech

audio

multilingual

text to speech

tts

voice cloning

voice design

voxcpm2

Turn text into spoken audio with VoxCPM2. Describe the voice you want in plain language, type what it should say, and hit run. 30 languages, 48 kHz output.

VoxCPM2 for Text to Speech

Turn text into spoken audio with VoxCPM2. Describe the voice you want in plain language, type what it should say, and hit run. 30 languages, 48 kHz output.

VoxCPM2 for Voice Cloning

audio

multilingual

open source

text to speech

tts

voice clone

voxcpm2

Upload a short voice sample and type what you want it to say. VoxCPM2 clones the voice and generates new speech in 30 languages at 48 kHz studio quality.

VoxCPM2 for Voice Cloning

Upload a short voice sample and type what you want it to say. VoxCPM2 clones the voice and generates new speech in 30 languages at 48 kHz studio quality.

LongCat AudioDiT for TTS

Audio

audiodit

audio generation

longcat

text to speech

tts

Turn text into spoken audio with LongCat AudioDiT 3.5B, Meituan's open-source diffusion TTS model. Clean voice quality in English and Chinese, no setup.

LongCat AudioDiT for TTS

Turn text into spoken audio with LongCat AudioDiT 3.5B, Meituan's open-source diffusion TTS model. Clean voice quality in English and Chinese, no setup.

Upscaling Images to 4k using Qwen Image Edit 2511

4k Upscale

Image

Image2Image

Image Edit

Image Upscale

Qwen Image Edit 2511

Upscale to 2k to 4k

Upscaling Images to 4k using Qwen Image Edit 2511

Upscale to 2k to 4k

Product Placement in Video

Image

Image to Video

Kling

Nano banana

Video

Product Placement in Video

Change in the Character using Image2Vid

Audio

Image

Image2Image

Image2Video

Kling Omni One Video Edit

Qwen Image Edit 2511

Video

Editing the character in the video without losing quality using video to video workflow

Change in the Character using Image2Vid

Editing the character in the video without losing quality using video to video workflow

ai video generator

black forest labs

flux 3 video

text to video

Video

FLUX 3 is Black Forest Labs' multimodal video model that turns a text prompt into an HD clip up to 20 seconds, with multi-shot scenes. Type a prompt and run.

FLUX 3 · Text to Video

FLUX 3 is Black Forest Labs' multimodal video model that turns a text prompt into an HD clip up to 20 seconds, with multi-shot scenes. Type a prompt and run.

LongCat AudioDiT for Multi Speaker TTS

Audio

audiodit

dialogue

longcat

multi-speaker

text to speech

voice cloning

Clone two voices from short audio samples and generate dialogue between them with LongCat AudioDiT 3.5B. Upload your references, write your script, hit run.

LongCat AudioDiT for Multi Speaker TTS

Clone two voices from short audio samples and generate dialogue between them with LongCat AudioDiT 3.5B. Upload your references, write your script, hit run.

Generate Fashion Billboard Using Outfit Image

Ecommerce

Image

NanoBanana

Reference Image

Upload the outfit image to generate fashion billboard

Generate Fashion Billboard Using Outfit Image

Upload the outfit image to generate fashion billboard

ai video

flux 3

flux 3 video

image to video

FLUX 3 is Black Forest Labs' multimodal video model that animates a still image into an HD clip up to 20 seconds. Upload an image, describe the motion, and run.

FLUX 3 · Image to Video

FLUX 3 is Black Forest Labs' multimodal video model that animates a still image into an HD clip up to 20 seconds. Upload an image, describe the motion, and run.

LongCat AudioDiT 3.5B - TTS, Voice Clone, Multi-Sp

Audio

longcat

text to audio

tts

LongCat AudioDiT 3.5B - TTS, Voice Clone, Multi-Sp

LTX-2.3 IC-LoRA HDR for SDR to HDR Video

Audio

LoRAs

LTX2.3

Vid2Vid

Video

Convert SDR video to 16-bit linear HDR with LTX-2.3 and the HDR IC-LoRA. Get EXR frames ready for color grading pipelines, plus a tone-mapped SDR preview.

LTX-2.3 IC-LoRA HDR for SDR to HDR Video

Convert SDR video to 16-bit linear HDR with LTX-2.3 and the HDR IC-LoRA. Get EXR frames ready for color grading pipelines, plus a tone-mapped SDR preview.

first last frame

flux 3

flux 3 video

image to video

FLUX 3 is Black Forest Labs' video model that builds the motion between two keyframes. Upload a start and end image for an HD clip up to 20 seconds.

FLUX 3 · First & Last Frame to Video

FLUX 3 is Black Forest Labs' video model that builds the motion between two keyframes. Upload a start and end image for an HD clip up to 20 seconds.

Seedance 2.0 for Camera Angle and Motion Control

camera angle control

camera motion

film production

seedance 2.0

vfx

vid2vid

Video

video generation

Take any video and re-shoot it with a new camera move. Seedance 2.0 keeps your scene intact and pulls the camera motion from a reference clip you upload.

Seedance 2.0 for Camera Angle and Motion Control

Take any video and re-shoot it with a new camera move. Seedance 2.0 keeps your scene intact and pulls the camera motion from a reference clip you upload.

Nano Banana 2 for  Image to 360 Panorama

360

concept art

Image

image to image

nano banana

panorama

vfx

Turn any photo into a 360 equirectangular panorama with Nano Banana 2. Upload an image, hit run, and get a 2:1 wrap ready for VR scenes and skyboxes.

Nano Banana 2 for Image to 360 Panorama

Turn any photo into a 360 equirectangular panorama with Nano Banana 2. Upload an image, hit run, and get a 2:1 wrap ready for VR scenes and skyboxes.

Flux.2 Klein Enhancer for Identity Transfer

character design

concept art

consistency

flux 2 klein

Image

image to image

portrait

Edit a photo of a person with Klein 9B while an identity transfer enhancer keeps their face and likeness locked, so they still look like themselves after.

Flux.2 Klein Enhancer for Identity Transfer

Edit a photo of a person with Klein 9B while an identity transfer enhancer keeps their face and likeness locked, so they still look like themselves after.

HunyuanVideo 1.5 for Image to Video

Animation

Filmmaking

HunyuanVideo 1.5

Image2Video

Video

HunyuanVideo 1.5 for Image to Video

Qwen Image 2512 and NVIDIA PiD Text to 4k Image

4k

concept art

Image

portrait

qwen

text to image

upscaling

Generate with Qwen Image 2512 and let NVIDIA PiD decode straight to 4K. One prompt, one run, a 4096px image with no separate upscale pass needed.

Qwen Image 2512 and NVIDIA PiD Text to 4k Image

Generate with Qwen Image 2512 and let NVIDIA PiD decode straight to 4K. One prompt, one run, a 4096px image with no separate upscale pass needed.

NVIDIA PiD for 4k Image Upscale

4k

concept art

flux

Image

image to image

product photography

upscaling

Upload any image and NVIDIA PiD rebuilds it at 4K in four diffusion steps. No tiling, no separate upscale model. Add a caption for guidance, or leave it blank.

NVIDIA PiD for 4k Image Upscale

Upload any image and NVIDIA PiD rebuilds it at 4K in four diffusion steps. No tiling, no separate upscale model. Add a caption for guidance, or leave it blank.

GPT Image 1.5  Text to Image

ECommerce

GPT Image 1.5

Image

Text2Image

GPT Image 1.5 Text to Image

SAM3.1 for Image and Video Segmentation

Image to Image

SAM 3.1

Segmentation

Video

Video to Video

Segment and track any object in images and video with SAM 3.1, Meta's concept segmentation model. Type what you want, hit run, and get clean masks back.

SAM3.1 for Image and Video Segmentation

Segment and track any object in images and video with SAM 3.1, Meta's concept segmentation model. Type what you want, hit run, and get clean masks back.

Kling Omni One Image Edit

Audio

Image2Image

Image Editing

Kling Omni One

Video

Kling Omni One Image Edit

HunyuanImage 3.0 Text to Image

API

Floyo API

HunyuanImage 3.0

Image

Text2Image

HunyuanImage 3.0 Text to Image

Ovis Text to Image

Image

Ovis

Text2Image

Typography

Ovis Text to Image

 Create a Fashion Shoot  - NanoBanana + Kling

Audio

Ecommerce

Image

Image to Video

Multi Reference Image

Video

Create a Fashion Shoot - NanoBanana + Kling

Bernini-R - Generation & Edit for Image and Videos

Bernini-R

Image Editing

Image Generation

Multi-Model

R2V

Video

Video Editing

Video Generation

Edit or generate video with Bernini, ByteDance's open-source unified video model. Upload a clip, add reference images, describe the change, and run it.

Bernini-R - Generation & Edit for Image and Videos

Edit or generate video with Bernini, ByteDance's open-source unified video model. Upload a clip, add reference images, describe the change, and run it.

Anima Turbo: Restyle Any Photo Into Anime Art

anime

anime-turbo

illustration

Image

LoRAs

restyle

Turn a photo or sketch into an anime illustration with Anima Base v1.0 plus the Turbo LoRA. Upload an image, describe the look you want, and hit run.

Anima Turbo: Restyle Any Photo Into Anime Art

Turn a photo or sketch into an anime illustration with Anima Base v1.0 plus the Turbo LoRA. Upload an image, describe the look you want, and hit run.

flux 2 video

flux 3

keyframes to video

text to video

FLUX 3 is Black Forest Labs' video model that hits up to 10 keyframes on a timeline you set. Upload your frames, place them, and get an HD clip with sound.

FLUX 3 · Keyframes to Video

FLUX 3 is Black Forest Labs' video model that hits up to 10 keyframes on a timeline you set. Upload your frames, place them, and get an HD clip with sound.

Gemini 3.1 Flash TTS for Text to Speech:

audio

gemini

gemini 3.1 flash tts

google

multi-speaker

text to speech

tts

voiceover

Turn any script into natural spoken audio with Gemini 3.1 Flash TTS, Google's text-to-speech model. Type your text, pick a voice, describe the tone, and hit run.

Gemini 3.1 Flash TTS for Text to Speech:

Turn any script into natural spoken audio with Gemini 3.1 Flash TTS, Google's text-to-speech model. Type your text, pick a voice, describe the tone, and hit run.

DramaBox: Direct a Voice Performance

Audio

dramabox

expressive speech

opensource

resemble ai

text to speech

Turn a scene-style prompt into a performed audio clip with DramaBox, Resemble AI's expressive open-source TTS model. Describe your speaker and delivery, optionally clone a voice, and hit run.

DramaBox: Direct a Voice Performance

Turn a scene-style prompt into a performed audio clip with DramaBox, Resemble AI's expressive open-source TTS model. Describe your speaker and delivery, optionally clone a voice, and hit run.

ComfySketch for Creating Images

ComfySketch

Image

Image2Image

Sketch2Image

Draw cool images using comfysketch

ComfySketch for Creating Images

Draw cool images using comfysketch

Mage Flow for Text to Image

ai image generator

int8

mage flow

mage-flow

microsoft

open source image model

qwen3-vl

text to image

Generate images from a text prompt with Mage-Flow, Microsoft's open-weight 4B image model. Pick an aspect ratio, describe the picture you want, and hit run.

Mage Flow for Text to Image

Generate images from a text prompt with Mage-Flow, Microsoft's open-weight 4B image model. Pick an aspect ratio, describe the picture you want, and hit run.

Gemma 4 E4B: Ask Anything About an Image

gemma 4

google deepmind

image to text

multimodal llm

opensource

vision language

Upload an image, ask a question or give an instruction, and get a written answer from Gemma 4 E4B, Google DeepMind's open-weights multimodal model. Add audio to transcribe or analyze it alongside the image.

Gemma 4 E4B: Ask Anything About an Image

Upload an image, ask a question or give an instruction, and get a written answer from Gemma 4 E4B, Google DeepMind's open-weights multimodal model. Add audio to transcribe or analyze it alongside the image.

LTX 2.3 for Text to 360 VR  Video Panorama

360 video

Audio

audio video

equirectangular

immersive

panorama

text to video

Video

vr

Describe a scene and LTX-2.3 by Lightricks generates a 360 video with synchronized sound that you can look around inside. Type a prompt and hit run.

LTX 2.3 for Text to 360 VR Video Panorama

Describe a scene and LTX-2.3 by Lightricks generates a 360 video with synchronized sound that you can look around inside. Type a prompt and hit run.

BiRefNet for Microscope Auto Segmentation

birefnet

Image

microscope

object isolation

rmbg

Drop in a microscope photo and BiRefNet finds the main subject, cuts it out, places it on a clean background, then upscales and sharpens it. Upload an image and hit run.

BiRefNet for Microscope Auto Segmentation

Drop in a microscope photo and BiRefNet finds the main subject, cuts it out, places it on a clean background, then upscales and sharpens it. Upload an image and hit run.

ai video generator

comfyui video

image to video

lightricks

ltx 2.3

ltx director

text to video

video with audio

Build a clip on a timeline with LTX 2.3, the open-weight 22B model that writes picture and synced audio in one pass. Drop an image in, write the shot, run.

LTX 2.3 and LTXDirector for Video Generation

Build a clip on a timeline with LTX 2.3, the open-weight 22B model that writes picture and synced audio in one pass. Drop an image in, write the shot, run.

LTX 2.3 + Rogala for Prompt Relay and Overlay Txt

Audio

concept art

image to video

prompt relay

rogala

text overlay

vfx

Video

video generation

Turn a photo into a multi-scene video with synced audio and burnt-in text using LTX 2.3 and Prompt Relay. Upload an image, script your beats, hit run.

LTX 2.3 + Rogala for Prompt Relay and Overlay Txt

Turn a photo into a multi-scene video with synced audio and burnt-in text using LTX 2.3 and Prompt Relay. Upload an image, script your beats, hit run.

R2A for Klein · Image to Image For Anime

cel shading

flux 2 klein

Image

image to image

real to anime

style transfer

Upload any photo or render and FLUX.2 Klein converts it into an anime-style illustration with cel shading, clean lineart, and vibrant color, then upscales the result to high resolution.