Explore
meta/meta-llama-guard-2-8b
A llama-3 based moderation and safeguarding language model
zsxkib/pulid
📖 PuLID: Pure and Lightning ID Customization via Contrastive Alignment
cjwbw/openvoice
Updated to OpenVoice v2: Versatile Instant Voice Cloning
fofr/video-morpher
Generate a video that morphs between subjects, with an optional style
snowflake/snowflake-arctic-instruct
An efficient, intelligent, and truly open-source language model
meta/meta-llama-3-70b-instruct
A 70 billion parameter language model from Meta, fine tuned for chat completions
I want to…
Generate images
Models that generate images from text prompts
Edit images
Tools for manipulating images.
Caption images
Models that generate text from images
Restore images
Models that improve or restore images by deblurring, colorization, and removing noise
Upscale images
Upscaling models that create high-quality images from low-quality images
Get embeddings
Models that generate embeddings from inputs
Use a language model
Models that can understand and generate text
Extract text from images
Optical character recognition (OCR) and text extraction
Train a language model
Language models that you can fine-tune using Replicate's training API.
Use a face to make images
Make realistic images of people instantly
Chat with images
Ask language models about images
Transcribe speech
Models that convert speech to text
Use handy tools
Toolbelt-type models for videos and images.
Generate music
Models to generate and modify music
Generate videos
Models that create and edit videos
Generate speech
Convert text to speech
Make 3D stuff
Models that generate 3D objects, scenes, radiance fields, textures and multi-views.
Get structured data
Language models that support grammar-based decoding as well as jsonschema constraints.
Popular models
SDXL-Lightning by ByteDance: a fast text-to-image model that makes high-quality images in 4 steps
A text-to-image generative AI model that creates beautiful images
Practical face restoration algorithm for *old photos* or *AI-generated faces*
Real-ESRGAN with optional face correction and adjustable upscale
Practical face restoration algorithm for *old photos* or *AI-generated faces*
Latest models
Train SDXL 1.0 with LoRA | mixed precision bf16 and save precision fp16
Just some good ole beautifulsoup scrapping URL magic. (some sites don't work as they block scrapping, but still useful)
High resolution image Upscaler and Enhancer. Use at ClarityAI.cc. A free Magnific alternative. Twitter/X: @philz1337x
Demucs is an audio source separator created by Facebook Research.
Projection module trained to add vision capabilties to Llama 3 using SigLIP
llava-phi-3-mini is a LLaVA model fine-tuned from microsoft/Phi-3-mini-4k-instruct
PyTorch implementation of AnimeGAN for fast photo animation
Function calling with llama-3 with prompting only.
Qwen1.5 is the beta version of Qwen2, a transformer-based decoder-only language model pretrained on a large amount of data
Qwen1.5 is the beta version of Qwen2, a transformer-based decoder-only language model pretrained on a large amount of data
Hyper-SD: Trajectory Segmented Consistency Model for Efficient Image Synthesis
AbsoluteReality V1.8.1 Model (Text2Img, Img2Img and Inpainting)
Phi-3-Mini-128K-Instruct is a 3.8 billion-parameter, lightweight, state-of-the-art open model trained using the Phi-3 datasets
Phi-3-Mini-4K-Instruct is a 3.8B parameters, lightweight, state-of-the-art open model trained with the Phi-3 datasets
This is wizard-vicuna-13b trained with a subset of the dataset - responses that contained alignment / moralizing were removed
Newest reranker model from BAAI (https://huggingface.co/BAAI/bge-reranker-v2-m3). FP16 inference enabled. Normalize param available
Best-in-class clothing virtual try on in the wild (non-commercial use only)
Generate a video that morphs between subjects, with an optional style
An efficient, intelligent, and truly open-source language model
Make stickers with AI. Generates graphics with transparent backgrounds.
yuan2.0-2b-mars是源2.0-2B模型的2024年3月版本,源2.0 是浪潮信息发布的新一代基础语言大模型。我们开源了全部的3个模型源2.0-102B,源2.0-51B和源2.0-2B。并且我们提供了预训练,微调,推理服务的相关脚本,以供研发人员做进一步的开发。源2.0是在源1.0的基础上,利用更多样的高质量预训练数据和指令微调数据集,令模型在语义、数学、推理、代码、知识等不同方面具备更强的理解能力。
Idefics2 is an open multimodal model that accepts arbitrary sequences of image and text inputs and produces text outputs
IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
FlashFace: Human Image Personalization with High-fidelity Identity Preservation