Model provider · Bring your own keyReplicate in a visual AI workflow
Run Replicate's image, video, and audio models in a visual AI workflow — hundreds of community and first-party models, run with your own key at Replicate's list price.
Replicate hosts a large, fast-moving catalog of open and commercial models behind one API — image, video, and audio generators from FLUX and Stable Diffusion to Kling, Wan, and more. NodeTool surfaces each as a node you can drop onto the canvas.
Chain Replicate models into a pipeline — prompt to image, image to video, transcript to speech — and the whole graph is reusable and shareable. Because every model is a node, comparing two Replicate models on the same prompt is a matter of wiring both into the same input.
Replicate runs on your own key in NodeTool: set your `REPLICATE_API_TOKEN` and calls go straight to Replicate at their price. The model list below comes from the Replicate node manifest, so it tracks what the provider ships.
Requests go straight to Replicate with your key at their list price. NodeTool never sits in the middle and never adds a markup.
Supported Replicate models
Each model below is its own node in NodeTool — the id is what Replicate serves it under. This list is generated from Replicate's node manifest, so it tracks what the provider actually ships.
Image models
310 total851-labs/background-removerRemove backgrounds from images.
aaronaftab/mirage-ghibliGhiblify any image, 10x cheaper/faster than GPT 4o
adirik/flux-cinestillFlux lora, use "CNSTLL" to trigger
ai-forever/kandinsky-2.2multilingual text2image latent diffusion model
alexgenovese/upscalerGFPGAN aims at developing Practical Algorithms for Real-world Face and Object Restoration
anon987654321/ra2arielreplicate/deoldify_imageAdd colours to old images
arthuryeti/dwiss-qwen-2black-forest-labs/flux-1.1-pro-ultraFLUX1.1 [pro] in ultra and raw modes.
black-forest-labs/flux-1.1-pro-ultra-finetunedInference model for FLUX 1.1 [pro] Ultra using custom `finetune_id`.
black-forest-labs/flux-2-devQuality image generation and editing with support for reference images
black-forest-labs/flux-2-flexMax-quality image generation and editing with support for ten reference images
black-forest-labs/flux-2-klein-4bVery fast image generation and editing model.
black-forest-labs/flux-2-klein-4b-baseUn-distilled version of FLUX.2 [klein].
black-forest-labs/flux-2-klein-4b-base-loraA version of FLUX.2 [klein] 4B-base that supports fast fine-tuned lora inference
black-forest-labs/flux-2-klein-9b4 step distilled version of FLUX.2 [klein].
black-forest-labs/flux-2-klein-9b-baseUn-distilled version of FLUX.2 [klein].
black-forest-labs/flux-2-klein-9b-base-loraA version of FLUX.2 [klein] 9B-base that supports fast fine-tuned lora inference
black-forest-labs/flux-2-maxThe highest fidelity image model from Black Forest Labs
black-forest-labs/flux-2-proHigh-quality image generation and editing with support for eight reference images
black-forest-labs/flux-canny-devOpen-weight edge-guided image generation.
black-forest-labs/flux-canny-proProfessional edge-guided image generation.
black-forest-labs/flux-depth-devOpen-weight depth-aware image generation.
black-forest-labs/flux-depth-proProfessional depth-aware image generation.
black-forest-labs/flux-devA 12 billion parameter rectified flow transformer capable of generating images from text descriptions
black-forest-labs/flux-dev-loraA version of flux-dev, a text to image model, that supports fast fine-tuned lora inference
black-forest-labs/flux-fill-devOpen-weight inpainting model for editing and extending images.
black-forest-labs/flux-fill-proProfessional inpainting and outpainting model with state-of-the-art performance.
black-forest-labs/flux-kontext-devOpen-weight version of FLUX.1 Kontext
black-forest-labs/flux-kontext-dev-loraFLUX.1 Kontext[dev] image editing model for running lora finetunes
black-forest-labs/flux-kontext-maxA premium text-based image editing model that delivers maximum performance and improved typography generation for transforming images through natural languag…
black-forest-labs/flux-kontext-proA state-of-the-art text-based image editing model that delivers high-quality outputs with excellent prompt following and consistent results for transforming…
black-forest-labs/flux-krea-devAn opinionated text-to-image model from Black Forest Labs in collaboration with Krea that excels in photorealism.
black-forest-labs/flux-proState-of-the-art image generation with top of the line prompt following, visual quality, image detail and output diversity.
black-forest-labs/flux-pro-finetunedInference model for FLUX.1 [pro] using custom `finetune_id`
black-forest-labs/flux-redux-devOpen-weight image variation model.
black-forest-labs/flux-redux-schnellFast, efficient image variation model for rapid iteration and experimentation.
black-forest-labs/flux-schnellThe fastest image generation model tailored for local development and personal use
black-forest-labs/flux-schnell-loraThe fastest image generation model tailored for fine-tuned use
bria/eraserSOTA Object removal, enables precise removal of unwanted objects from images while maintaining high-quality outputs.
+ 270 more image models available in Replicate.
Video models
156 totalalibaba/happyhorse-1.0Alibaba's Happy Horse 1.0 generates videos from text prompts or animates a single image into video.
alibaba/happyhorse-1.1Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images.
andreasjansson/wan-1.3b-inpaintInpainting and video2video experiments with Wan 2.1
anotherjesse/zeroscope-v2-xlZeroscope V2 XL & 576w
arielreplicate/deoldify_videoAdd colours to old video footage.
arielreplicate/robust_video_mattingextract foreground of a video
bria/video-erase-objectA high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency
bria/video-increase-resolutionUpscale videos up to 8K output resolution.
bria/video-remove-backgroundAutomatically remove backgrounds from videos -perfect for creating clean, professional content without a green screen.
bytedance/dreamactor-m2.0Animate any character, humans, cartoons, animals, even non-humans, from a single image + driving video
bytedance/latentsyncLatentSync: generate high-quality lip sync animations
bytedance/omni-humanTurns your audio/video/images into professional-quality animated videos
bytedance/omni-human-1.5A film-grade digital human model that generates realistic video from a single image, audio clip, and optional text prompt.
bytedance/seedance-1-liteA video generation model that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 720p resolution
bytedance/seedance-1-proA pro version of Seedance that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 1080p resolution
bytedance/seedance-1-pro-fastA faster and cheaper version of Seedance 1 Pro
bytedance/seedance-1.5-proA joint audio-video model that accurately follows complex instructions.
bytedance/seedance-2.0ByteDance's multimodal video generation model with native audio, multimodal reference inputs, and intelligent duration control.
bytedance/seedance-2.0-fastA faster variant of Seedance 2.0 for quicker video generation with multimodal inputs and native audio.
bytedance/seedance-2.0-miniA lower-cost variant of Seedance 2.0 for high-volume video generation with multimodal inputs and native audio.
bytedance/video-upscalerUpscale and enhance video up to 4K at 60fps, with scene-aware presets for AI-generated content, short dramas, UGC, and film restoration.
character-ai/ovi-i2vOvi: generate videos with audio from image and text inputs
chenxwh/video-retalkingAudio-based Lip Synchronization for Talking Head Video
cjwbw/aniportrait-audio2vidAudio-Driven Synthesis of Photorealistic Portrait Animations
cjwbw/sadtalkerStylized Audio-Driven Single Image Talking Face Animation
decart/lucy-edit-2Edit and transform videos with text prompts and reference images.
easel/ai-avatarsUse one or two face images to create AI avatars
fictions-ai/autocaptionAutomatically add captions to a video
flux-kontext-apps/restyle-video-frameUse flux-kontext-pro to change the first or last frame of a video.
fofr/audio-to-waveformCreate a waveform video from audio
fofr/kontext-ps1FLUX Kontext fine-tune that let's you restyle any image as a PS1 or PS2 video game
google/veo-2State of the art video generation model.
google/veo-3Sound on: Google’s flagship Veo 3 text to video model, with audio
google/veo-3-fastA faster and cheaper version of Google’s Veo 3 video model, with audio
google/veo-3.1New and improved version of Veo 3, with higher-fidelity video, context-aware audio, reference image and last frame support
google/veo-3.1-fastNew and improved version of Veo 3 Fast, with higher-fidelity video, context-aware audio and last frame support
google/veo-3.1-liteGoogle's cost-efficient video generation model with native audio, optimized for high-volume applications
heygen/avatar-ivCreate realistic talking avatar videos from text with HeyGen's Avatar IV engine
heygen/avatar-vCreate realistic talking avatar videos from text with HeyGen's Avatar V engine — the newest, highest-quality avatar engine with cross-reference-driven animat…
heygen/lipsync-precisionHigh-accuracy lip-sync: replace or dub audio on any video with avatar-inference lip sync
+ 116 more video models available in Replicate.
Audio models
62 totaladirik/styletts2Generates speech from text
afiaka87/tortoise-ttsGenerate speech from text, clone voices from mp3 files.
andreasjansson/musicgen-looperGenerate fixed-bpm loops from text prompts
chenxwh/openvoiceUpdated to OpenVoice v2: Versatile Instant Voice Cloning
cjwbw/parler-ttslightweight text-to-speech (TTS) model, trained on 10.5K hours of audio data
cjwbw/voicecraftZero-Shot Speech Editing and Text-to-Speech in the Wild
elevenlabs/flash-v2.5ElevenLabs's fastest speech synthesis model
elevenlabs/musicCompose a song from a prompt or a composition plan
elevenlabs/turbo-v2.5High quality, low latency text to speech in 32 languages
elevenlabs/v2-multilingualGenerate multilingual text-to-speech audio in over 30 languages
elevenlabs/v3The most expressive Text to Speech model
fermatresearch/spanish-f5-ttsA F5-TTS fine-tuned for Spanish
google/gemini-3.1-flash-ttsGoogle's fast, expressive text-to-speech model with 30 voices and 70+ language support
google/lyria-2Lyria 2 is a music generation model that produces 48kHz stereo audio through text-based prompts
google/lyria-3Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model
google/lyria-3-proGenerate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model
inworld/realtime-tts-1.5-maxHighest-quality realtime text-to-speech with <200ms latency, emotion control, and 15-language support
inworld/realtime-tts-1.5-miniUltra-fast, cost-efficient realtime text-to-speech with ~120ms latency and 15-language support
inworld/realtime-tts-2Most expressive text-to-speech model from Inworld, with natural-language steering, real-time latency, and multilingual support across 100+ languages.
inworld/tts-1.5-maxHighest-quality text-to-speech with <200ms latency, emotion control, and 15-language support
inworld/tts-1.5-miniUltra-fast, cost-efficient text-to-speech with ~120ms latency and 15-language support
jaaari/kokoro-82mKokoro v1.0 - text-to-speech (82M params, based on StyleTTS2)
lucataco/ace-stepA Step Towards Music Generation Foundation Model text2music
lucataco/csm-1bCSM (Conversational Speech Model) is a speech generation model from Sesame that generates RVQ audio codes from text and audio inputs
+ 38 more audio models available in Replicate.
Text models
127 totalabiruyt/text-extract-ocrA simple OCR Model that can easily extract text from an image.
adirik/e5-mistral-7b-instructE5-mistral-7b-instruct language embedding model
andreasjansson/blip-2Answers questions about images
andreasjansson/clip-featuresReturn CLIP features for the clip-vit-large-patch14 model
andreasjansson/clip-featuresReturn CLIP features for the clip-vit-large-patch14 model
andreasjansson/llama-2-13b-embeddingsLlama2 13B with embedding output
anthropic/claude-3.5-haikuAnthropic's fastest, most cost-effective model, with a 200K token context window (claude-3-5-haiku-20241022)
anthropic/claude-3.7-sonnetThe most intelligent Claude model and the first hybrid reasoning model on the market (claude-3-7-sonnet-20250219)
+ 119 more text models available in Replicate.
Frequently asked questions
- How do I connect Replicate to NodeTool?
- Add your Replicate token as `REPLICATE_API_TOKEN` in NodeTool. Replicate nodes then call the API directly with your token.
- Which Replicate models are available?
- The image, video, and audio models in Replicate's node manifest — several hundred, each a separate node. See the catalog below.
- What does Replicate cost through NodeTool?
- Replicate's own list price. NodeTool is bring-your-own-key and adds no per-generation fee.
Other providers
Run Replicate your way.
Download NodeTool Studio and build across image, video, audio, and text with your own keys.