

beautyyuyanli / multilingual-e5-large
multilingual-e5-large: A multi-language text embedding model
78.4M runs


lucataco / moondream2
moondream2 is a small vision language model designed to run efficiently on edge devices
20.5M runs


prunaai / p-image
A sub 1 second text-to-image model built for production use cases.
21.2M runs


prunaai / z-image-turbo
Z-Image Turbo is a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.
56.7M runs

Google's latest fast image generation model with sharp text rendering, conversational editing, multi-image fusion, and up to 4K output
15.8K runs

prunaai/p-video-2-proPruna's highest-quality video model. Generate video from text or images with speed and quality modes, first- and last-frame control, and durations up to 15 seconds.
252.1K runs
Wan 3.0 generates video from a text prompt or a starting image, with cinematic motion and support for 480p, 720p, and 1080p output up to 30 seconds.
23.3K runs

bytedance/seedance-2.5ByteDance's flagship multimodal video model with native audio, native 30-second generation, and large multimodal reference sets.
216.2K runs

Foundation image model from Krea, tuned for expressive illustration, anime, and painterly styles. Fast and consistent across artistic directions.
36.7K runs

Most expressive text-to-speech model from Inworld, with natural-language steering, real-time latency, and multilingual support across 100+ languages.
85.9K runs

openai/gpt-image-2OpenAI's state-of-the-art image generation model. Create and edit images from text with strong instruction following, sharp text rendering, and detailed editing.
30.6M runs

Anthropic's most capable model with a step-change improvement in agentic coding, better vision, and stronger multi-step reasoning
538.7K runs

Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics
58.9K runs

bytedance/seedream-5-liteSeedream 5.0 lite: image generation with built-in reasoning, example-based editing, and deep domain knowledge
3.5M runs
Generate videos using xAI's Grok Imagine Video model
2M runs

The highest fidelity image model from Black Forest Labs
4.2M runs
Official models are always on, maintained, and have predictable pricing.

Google's latest fast image generation model with sharp text rendering, conversational editing, multi-image fusion, and up to 4K output

FLUX 3 is Black Forest Labs' text-to-image and image-editing model. Generate images from a prompt, or edit with up to 10 reference images.

Precisely edit source images with Ideogram 4.5.

Generate and edit images with Ideogram 4.5, the latest for precise image generation.

Pruna's highest-quality video model. Generate video from text or images with speed and quality modes, first- and last-frame control, and durations up to 15 seconds.

OpenAI's fastest model for high-quality, everyday image generation. Generate and edit images from text and image inputs with strong instruction following and sharp text rendering.

OpenAI's most capable image model, built for workflows where editing precision matters most. Generate and edit images from text and image inputs with strong instruction following, sharp text rendering, and detailed control.

Generate videos from text prompts using Alibaba's Wan 3.0 Prime model. Up to 1080p and 30 seconds, with 480p, 720p, and 1080p output.

Google's fast multimodal video generation and editing model with native audio, using the Interactions API

Recraft V4 Styles Pro SVG generates high-quality editable vector images that match a reusable style or attached reference images.
Wan 3.0 generates video from a text prompt or a starting image, with cinematic motion and support for 480p, 720p, and 1080p output up to 30 seconds.

Qwen-Image-3.0 generates and edits images with accurate text rendering, complex layouts, and photographic detail.

Qwen-Image-3.0-Pro generates and edits images with dense, accurate text rendering, complex multi-element layouts, and photographic detail.

Upscale videos to higher resolution with FLUX super-resolution. Precise mode sharpens and stays faithful to the source; creative mode restores and invents fine detail.
P-Video-2 is the quality-focused successor to P-Video from $0.025/s
Edit videos with a text prompt and optional reference images.

xAI's Grok Imagine Image 2.0 — text-to-image generation and editing with a quality control and output up to 2k

Fast video generation with text-to-video and image-to-video, portrait and landscape support, synchronized audio, and frame interpolation. Up to 20 seconds at 1080p, and 4K resolution.

Translate audio and video into 90+ languages while preserving each speaker's voice, emotion, and timing

ByteDance's flagship multimodal video model with native audio, native 30-second generation, and large multimodal reference sets.
Use AI to generate images & photos with an API
Use AI to understand, describe, and caption videos with an API
Use AI for text-to-speech or to clone your voice via API
Use AI to generate images from a face with an API
Use AI to generate videos with an API
Use AI to upscale and enhance images with an API
Use AI to generate music with an API
Use AI to edit any image via API
Use AI to transcribe speech to text with an API
Use AI For Optical Character Recognition (OCR) to extract text from images via API
Use AI to remove backgrounds from images and videos with an API
FLUX AI models by Black Forest Labs: image generation & editing via API
Use AI to restore images via API
Use AI to upscale, restore, extend, and enhance videos with an API
Detect NSFW content in images and text
Classify text by sentiment, topic, intent, or safety
Identify speakers from audio and video inputs
Replace faces across images with natural-looking results.
Transform rough sketches into polished visuals
Generate custom emojis from text or images
Create anime-style characters, scenes, and animations
Use AI to generate videos from images with an API
Chat with images — visual Q&A, analysis, and reasoning via API
Use AI to generate captions and descriptions from images with an API
Use AI to edit, restyle, extend, and remix videos with an API
WAN family of models: open-source video, image, and audio generation
Generate 3D objects, meshes, and textures from text or images with an API
Official models are always on, predictably priced, and have a stable API.
Explore Large Language Models (LLMs) for chat, generation & NLP tasks via API
Try AI Models for free: video generation, image generation, upscaling, and photo restoration
Use AI to generate lipsync videos with an API
Use AI to control image generation with an API
Embedding models for AI search and analysis
Use AI object detection and segmentation models to distinguish objects in images & videos
Flux fine-tunes: build and run custom AI image models via API
Kontext fine-tunes: Build custom AI image models with an API
Create songs with voice cloning models via API
AI media utilities: auto-caption, watermark, frame extraction & more via API
Browse the diverse range of qwen-image fine-tunes the community has custom-trained on Replicate.
sprited / kimodo
Kimodo text-to-motion: SOMA and G1 models, sequential prompts, motion constraints, multi-sample generation, GLB/BVH/NPZ and Mixamo FBX export, with animated previews. Unofficial NVIDIA Kimodo deployment.
2 runs

google / nano-banana-2.1
Google's latest fast image generation model with sharp text rendering, conversational editing, multi-image fusion, and up to 4K output
15.8K runs


andrewstaab / hornet
72 runs


hardtunesk / clon-tinder
11 runs


newwork062026-lab / stocksurgemodel
133 runs

black-forest-labs / flux-3-image
FLUX 3 is Black Forest Labs' text-to-image and image-editing model. Generate images from a prompt, or edit with up to 10 reference images.
13.7K runs


gregpriday / crisp-vector
Raster logos to clean, compact SVGs with the background removed. Batch-friendly: pass many images (or a zip) and they are vectorised in parallel.
120 runs


ideogram-ai / ideogram-4-5-precise-edit
Precisely edit source images with Ideogram 4.5.
2.1K runs


ideogram-ai / ideogram-4-5
Generate and edit images with Ideogram 4.5, the latest for precise image generation.
1.7K runs
sprited / anisora
Animate a still image with a motion prompt. Unofficial AniSora V3.2 deployment using FP8-scaled weights, with MP4, WebM, and animated WebP output.
13 runs


devgmstudios / nsfw-filter-adjustable
Adjustable Stable Diffusion NSFW safety checker. Higher threshold = less aggressive (fewer flags); 0 = most aggressive (stock). Fork of m1guelpf/nsfw-filter.
13.6K runs


karthiksync / blitz
Make structured choices from a supplied case record. Ask one question or score several questions against the same state; get a selected choice and probabilities for the supplied options.
7 runs