Skip to main content

Replicate Models

4,127 entities

barakplasma/jev-omni

model

Jev-Omni (#1 on Image JevBench v0.1): ask a typed yes/no or multiple-choice question about an image, get a calibrated probability per option from one forward pass. Pinned akhilaaa3/Jev-Omni revision, Apache-2.0.

black-forest-labs/flux-3-image

model

FLUX 3 is Black Forest Labs' text-to-image and image-editing model. Generate images from a prompt, or edit with up to 10 reference images.

ideogram-ai/ideogram-4-5

model

Generate and edit images with Ideogram 4.5, the latest for precise image generation.

ideogram-ai/ideogram-4-5-precise-edit

model

Precisely edit source images with Ideogram 4.5.

onesoltech/cadence-fast

model

Adds punctuation to text

aiviostudio/salmonn-2025

model

lucataco/interactiveomni-8b

model

A unified omni-modal model that can simultaneously receive inputs such as images, audio, text, and video and directly generate coherent text and speech

dvyio/sdxl-polaroid

model

Photos on Polaroid, including hands holding Polaroid photos

dvyio/sdxl-spectrogram

model

SDXL trained on a variety of spectrograms

dvyio/sdxl-victorian-britain

model

SDXL trained on photographs from Victorian-era Britain

dvyio/sdxl-film-noir

model

SDXL fine-tuned on film noir movie stills

dvyio/sdxl-soviet-propaganda

model

SDXL fine-tuned on Soviet propaganda posters

untapped/glance-qwen3-vl-4b

model

Typed answers about an image, read from unchanged Qwen3-VL-4B logits in one forward pass.

amberwhitehead/llama2-7b-chat-gptq

model

prakhar-bhartiya/meta-tribev2-social-media-content-signal

model

Predicts virality of short videos using a Meta TRIBE v2 brain-encoding model

jdreetz/faces

model

deepsbhat1984/music_voice_clone_pro_depricated

model

Sing any song in any voice. Give a song + a short voice clip and get the song re-sung in that voice, with the original music kept intact. Zero-shot, no training. Works in any language - English, Spanish, French, Italian, Hindi, and more.

littletry/video-nsfw-filter

model

Detect NSFW content in sampled video frames with adjustable sensitivity, per-frame scores, and optional screenshots.

hautechai/grounding-dino

model

Grounding DINO: zero-shot text-prompted object detection (SwinT-OGC). H100 build.

prunaai/p-video-2-pro

model

Pruna's highest-quality video model. Generate video from text or images with speed and quality modes, first- and last-frame control, and durations up to 15 seconds.

quiverai/arrow-1.1

model

Generate editable SVG graphics from text prompts with Quiver's Arrow 1.1 model. Great for icons, logos, and illustrations.

ultralytics/yolo26-seg

model

Ultralytics YOLO26 instance segmentation (COCO-Seg), selectable size n/s/m/l/x.

tuannha/f5-tts-vi

model

Vietnamese F5-TTS. Released by EraX-AI team

openai/gpt-image-2.5-flare

model

OpenAI's fastest model for high-quality, everyday image generation. Generate and edit images from text and image inputs with strong instruction following and sharp text rendering.

openai/gpt-image-2.5-sunburst

model

OpenAI's most capable image model, built for workflows where editing precision matters most. Generate and edit images from text and image inputs with strong instruction following, sharp text rendering, and detailed control.

Page 1 of 166