AI Tools.

Search

text to image models

6 models · ranked by HuggingFace downloads

stable-diffusion-xl-base-1.0

by stabilityai

SDXL Base 1.0 is Stability AI's flagship text-to-image diffusion model, operating at 1024x1024 native resolution with a dual-text-encoder architecture. It produces significantly higher-quality images than SD 1.5 and 2.x, especially for complex compositions.

1,677,646 ↓ · 8,106 ♡

dreamshaper-7

by Lykon

dreamshaper-7 is a community fine-tune of Stable Diffusion 1.5 by Lykon, optimized for photorealistic portraits, artistic illustrations, and anime-adjacent styles. Version 7 is a direct successor improving skin texture, lighting coherence, and reducing NSFW content leakage versus earlier releases. It uses the CreativeML OpenRAIL-M license which restricts some commercial uses.

869,123 ↓ · 64 ♡

stable-diffusion-3.5-medium

by stabilityai

Stable Diffusion 3.5 Medium is Stability AI's mid-tier SD3 variant using a Multimodal Diffusion Transformer (MMDiT) architecture. At a smaller parameter count than SD3.5-Large, it offers faster generation while maintaining SD3's improved text rendering and prompt adherence over SD2/SDXL. License restricts commercial use above certain revenue thresholds.

616,212 ↓ · 1,007 ♡

Realistic_Vision_V5.1_noVAE

by SG161222

Realistic Vision V5.1 is a photorealism-focused Stable Diffusion 1.5 fine-tune that has accumulated substantial community use for portrait and product photography generation. The 'noVAE' variant ships without the VAE weights, requiring users to supply a separate VAE (typically the SD 1.5 base VAE or the EMA840k variant), which reduces checkpoint file size. It is designed for integration into A1111, InvokeAI, and ComfyUI workflows.

423,018 ↓ · 262 ♡

HunyuanImage-3.0

by tencent

HunyuanImage-3.0 is Tencent's third-generation text-to-image diffusion model using a Mixture-of-Experts transformer backbone. It targets photorealistic and stylised image generation with improved prompt adherence over the previous HunyuanImage series. The MoE architecture selectively activates expert layers per token, balancing quality and compute.

342,485 ↓ · 1,094 ♡

stable-diffusion-xl-1.0-inpainting-0.1

by diffusers

This is an SDXL-based inpainting model from HuggingFace Diffusers, fine-tuned specifically for masked region infilling using Stable Diffusion XL's 1024px native resolution. Unlike SD 1.5 inpainting models, the SDXL base enables generating higher-resolution inpaints that blend more naturally with surrounding image context. The model uses the StableDiffusionXLInpaintPipeline.

307,203 ↓ · 374 ♡