Every model, one typed interface

89 curated nodes — image, video, and language models, media utilities — each hand-translated into blitflow's node API with typed ports, real defaults, and a pinned version. Open one to run it from the browser or copy the call for your code.

Read the run-a-node guide
Background Remover input
Background Remover output

Background Remover

v1.0.0

Remove the image background → transparent PNG.

BiRefNet (Cutout) input
BiRefNet (Cutout) output

BiRefNet (Cutout)

v1.0.0

High-resolution background removal with fine edge matting → transparent PNG.

Bria Eraser (Remove Object)

v1.0.0

Erase the masked region and fill it from its surroundings.

Bria Expand (Outpaint) input
Bria Expand (Outpaint) output

Bria Expand (Outpaint)

v1.0.0

Extend an image past its edges to fill a wider canvas.

Clarity Upscaler input
Clarity Upscaler output

Clarity Upscaler

v1.0.0

Detail-enhancing upscale for illustrated/artistic images. Up to 13000px.

Color Matcher input
Color Matcher output

Color Matcher

v1.0.0

Match one image's palette to a reference, or fix white balance.

ControlNet Pose

v1.0.0

Generate an image from a prompt while preserving a detected human pose.

ControlNet Scribble

v1.0.0

Generate a detailed image from a prompt while preserving a scribbled composition.

ControlNet Tile (Detail Upscale) input
ControlNet Tile (Detail Upscale) output

ControlNet Tile (Detail Upscale)

v1.0.0

Upscale tile by tile, adding detail while holding composition.

Depth Anything 3 Base

v1.0.0

Estimate depth from a single image → depth map (near = bright) with a confidence map and estimated camera intrinsics.

Depth Anything 3 Mono input
Depth Anything 3 Mono output

Depth Anything 3 Mono

v1.0.0

Best-quality single-image depth → depth map (near = bright) plus a sky mask.

Depth Anything 3 Small

v1.0.0

Estimate depth from a single image → depth map (near = bright) with a confidence map and estimated camera intrinsics.

Depth Anything V1 Base

v1.0.0

Estimate depth from a single image → a depth map (near = bright) plus a colored preview.

Depth Anything V1 Large

v1.0.0

Estimate depth from a single image → a depth map (near = bright) plus a colored preview.

Depth Anything V1 Small

v1.0.0

Estimate depth from a single image → a depth map (near = bright) plus a colored preview.

Depth Anything v2 (Depth Map) input
Depth Anything v2 (Depth Map) output

Depth Anything v2 (Depth Map)

v1.0.0

Read a depth map off any image, ready for depth-guided generation.

Depth Anything V2 Small

v1.0.0

Estimate depth from a single image → a depth map (near = bright) plus a colored preview.

Florence-2 (Caption & Detect) input

{'<DETAILED_CAPTION>': 'The image shows a golden retriever puppy sitting in a field of daisies, surrounded by lush green grass and a bright blue sky.'}

Florence-2 (Caption & Detect)

v1.0.0

Caption, tag or read text out of an image.

a lighthouse in a storm, spray lit by the beam, dramatic photographic realism

Flux 1.1 Pro output

Flux 1.1 Pro

v1.0.0

Top-tier FLUX text-to-image with strong prompt adherence

Flux 2 Klein

v1.0.0

Sub-second FLUX.2 generation for drafts and latency-critical steps

Flux 2 Pro

v1.0.0

FLUX.2 text-to-image up to 4MP, optionally guided by a reference image

Flux Canny input
Flux Canny output

Flux Canny

v1.0.0

Regenerate an image while preserving exact outlines and silhouettes from the reference.

Flux Depth input
Flux Depth output

Flux Depth

v1.0.0

Regenerate an image while preserving depth/perspective structure from the reference.

an astronaut resting in a hammock between two palm trees on an alien beach, golden hour, cinematic

Flux Dev output

Flux Dev

v1.0.0

High-quality text-to-image with FLUX.1 [dev]

Flux Fill (Inpaint) input
Flux Fill (Inpaint) output

Flux Fill (Inpaint)

v1.0.0

Fill masked areas of an image guided by a text prompt. White mask = fill, black = keep.

Flux Fill Pro

v1.0.0

Inpaint a masked region from a prompt at pro quality

Flux Kontext input
Flux Kontext output

Flux Kontext

v1.0.0

Edit an image with text instructions while preserving character identity and style consistency.

Flux Kontext Pro

v1.0.0

Instruction edit holding identity and style, at Kontext's pro quality tier

Flux Redux (Variations) input
Flux Redux (Variations) output

Flux Redux (Variations)

v1.0.0

Generate style-consistent variations of an image. No text prompt required.

a red panda barista pulling an espresso shot, warm morning light, shallow depth of field

Flux Schnell output

Flux Schnell

v2.0.0

Fast, cheap text-to-image with FLUX.1 [schnell]

GPT Image 1 Mini

v1.0.0

Text-to-image with GPT Image 1 Mini, optionally guided by a reference image

GPT Image 1.5

v1.0.0

Text-to-image with GPT Image 1.5, optionally guided by a reference image

GPT Image 2

v1.0.0

Text-to-image with GPT Image 2, optionally guided by a reference image

Grok Imagine 2

v1.0.0

Grok Imagine 2.0 text-to-image

Grounded SAM (Text to Mask) input
Grounded SAM (Text to Mask) output

Grounded SAM (Text to Mask)

v1.0.0

Turn a plain-text description of a subject into a cutout mask.

Grounding DINO (Detect) input
Grounding DINO (Detect) output

Grounding DINO (Detect)

v1.0.0

Locate objects named in plain text, as boxes plus an annotated image.

IC-Light (Relight) input
IC-Light (Relight) output

IC-Light (Relight)

v1.0.0

Relight a subject to match a described lighting setup.

a bold gig poster that reads "NIGHT MARKET" over neon street food stalls

Ideogram v3 Balanced output

Ideogram v3 Balanced

v1.0.0

Ideogram v3 at the middle speed/quality rung

an art-deco book cover that reads "THE GILDED HOUR", gold foil on deep teal

Ideogram v3 Quality output

Ideogram v3 Quality

v1.0.0

Highest-fidelity Ideogram v3 text-to-image

a vintage travel poster that reads "KYOTO" with cherry blossoms and a pagoda, bold clean typography

Ideogram v3 Turbo output

Ideogram v3 Turbo

v1.0.0

Fast text-to-image with strong, legible text rendering

a festival poster that reads "HARVEST MOON" over paper-cut hills

Ideogram v4 Balanced output

Ideogram v4 Balanced

v1.0.0

Ideogram 4.0 at the middle speed/quality rung

an enamel pin design that reads "TRAIL CREW", crisp lettering on forest green

Ideogram v4 Quality output

Ideogram v4 Quality

v1.0.0

Ideogram 4.0 at its highest-fidelity rung

an anime-style rooftop at dusk, city lights blooming behind a lone figure

Krea 2 Medium output

Krea 2 Medium

v1.0.0

Expressive illustration and anime styles.

Nano Banana 2

v1.0.0

Conversational image generation and editing with Nano Banana 2

Nano Banana Pro

v1.0.0

Conversational image generation and editing with Nano Banana Pro

Pruna Image Edit (Fast) input
Pruna Image Edit (Fast) output

Pruna Image Edit (Fast)

v1.0.0

Sub-second instruction editing, with task presets and a second reference image

Pruna Upscale input
Pruna Upscale output

Pruna Upscale

v1.0.0

Upscale to a target megapixel count, up to 128 MP.

Qwen Image Edit Plus input
Qwen Image Edit Plus output

Qwen Image Edit Plus

v1.0.0

Instruction editing with strong text rendering inside the image.

Real-ESRGAN (Upscale) input
Real-ESRGAN (Upscale) output

Real-ESRGAN (Upscale)

v1.0.0

Upscale and restore images 2–4×. Fast and cheap.

Recraft Crisp Upscale input
Recraft Crisp Upscale output

Recraft Crisp Upscale

v1.0.0

Sharpen and clean up an image without reimagining it

a minimalist flat illustration of a steaming coffee cup, soft pastel palette

Recraft v3 output

Recraft v3

v1.0.0

High-quality text-to-image with style control

a flat vector illustration of a mountain cable car, muted alpine palette

Recraft v4.1 output

Recraft v4.1

v1.0.0

Design-led text-to-image with precise prompt adherence.

a knight in silver armor — four-angle walk cycle

Retro Diffusion Animation (Pixel Art) output

Retro Diffusion Animation (Pixel Art)

v1.0.0

Pixel-art animation spritesheet generation (walk/idle cycles, 4-direction characters)

a brass treasure chest overflowing with gold coins

Retro Diffusion Fast (Pixel Art) output

Retro Diffusion Fast (Pixel Art)

v1.0.0

Fast pixel-art / retro game-asset generation

an ancient spellbook with glowing runes on the cover

Retro Diffusion Plus (Pixel Art) output

Retro Diffusion Plus (Pixel Art)

v1.0.0

High-quality pixel-art generation

a majestic snow leopard perched on a rocky cliff at dawn, ultra-detailed, photorealistic

SDXL output

SDXL

v1.0.0

Text-to-image with Stable Diffusion XL

a weathered brass diving helmet on a workshop bench, shafts of dusty light

Seedream 4.5 output

Seedream 4.5

v1.0.0

Generate or re-mix from references, up to 4K.

Seedream 5.0 Pro

v1.0.0

Seedream 5.0 text-to-image with reference re-mixing and strong typography

Segment Anything 2 (Masks) input
Segment Anything 2 (Masks) output

Segment Anything 2 (Masks)

v1.0.0

Segment every region of an image into one combined mask.

Claude Haiku 4.5

v1.0.0

Prompt Claude Haiku 4.5

Claude Sonnet 4.6

v1.0.0

Prompt Claude Sonnet 4.6

DeepSeek V4 Pro

v1.0.0

Prompt DeepSeek V4 Pro

Gemini 3 Flash

v1.0.0

Prompt Gemini 3 Flash

Gemini 3.1 Pro

v1.0.0

Prompt Gemini 3.1 Pro

GPT-5.4

v1.0.0

Prompt GPT-5.4

GPT-5.4 Pro

v1.0.0

Prompt GPT-5.4 Pro

Grok 4.6

v1.0.0

Prompt Grok 4.6

GPT-4o Transcribe

v1.0.0

Speech-to-text with GPT-4o, stronger on accents and noise

The gate is open. Follow me, and stay close.

Kokoro (Text to Speech)

v1.0.0

Text-to-speech in 46 voices across nine languages

a warm lo-fi hip hop loop with a dusty rhodes chord

MusicGen Looper (Seamless Loop)

v1.0.0

Generate a seamless music loop at a fixed tempo.

a heavy wooden door creaking open in a stone hall

Stable Audio 2.5 (SFX & Music)

v1.0.0

Generate a sound effect or a music bed from a prompt.

TTS-1 HD

v1.0.0

Text-to-speech with OpenAI's HD voices

How old is the Brooklyn Bridge?

Whisper

v1.0.0

Speech-to-text with Whisper

Convert Audio Format

v1.0.0

Transcode audio to mp3, wav, ogg, m4a, or flac (ffmpeg)

Convert Image Format input
Convert Image Format output

Convert Image Format

v1.0.0

Re-encode an image as PNG, JPEG, WebP, or AVIF (sharp)

Extract Frame output

Extract Frame

v1.0.0

Capture a still image from a video at a timestamp (ffmpeg)

Image Input input
Image Input output

Image Input

v1.0.0

Ingest an uploaded or inline image (data: URI, base64, or URL) and host it as a reusable image artifact

Resize Image input
Resize Image output

Resize Image

v1.0.0

Resize or rescale an image to target dimensions (sharp)

Video to GIF output

Video to GIF

v1.0.0

Convert a video (or a clip of it) into an animated GIF (ffmpeg)

Kling 3.0 Image-to-Video

v1.0.0

Animate an image with Kling 3.0, in std and pro tiers

Kling 3.0 Text-to-Video

v1.0.0

Text-to-video with Kling 3.0's motion quality, in std and pro tiers

Seedance 2.0 Fast

v1.0.0

Seedance's draft tier — fast text- or image-to-video up to 720p

a paper boat drifting down a rain-slicked cobblestone gutter, autumn leaves swirling past, macro shot

Seedance 2.5

v1.0.0

Text-to-video with Seedance 2.5, optionally guided by a first frame and reference images

TRELLIS (Image to 3D) input

TRELLIS (Image to 3D)

v1.0.0

Turn one image into a textured GLB mesh and a turntable render.

Veo 3.1

v1.0.0

Cinematic text- or image-to-video with native audio, dialogue and sound effects

Veo 3.1 Fast

v1.0.0

Veo 3.1's draft tier — same controls at a quarter of the rate

Wan 2.2 (Image to Video) input

Wan 2.2 (Image to Video)

v1.0.0

Animate a still image into a short clip

ink spreading through water, slow motion, high contrast

Wan 2.2 (Text to Video)

v1.0.0

Generate a short clip from a prompt alone.

Wan 2.6 Image-to-Video

v1.0.0

Animate an image with Wan 2.6, with native audio generation