AI Image Upscalers & Restorers
Upscalers add real detail rather than just stretching pixels, and restoration models repair scratches, blur and faded color in old photos.
Use them to rescue low-resolution images, sharpen scans or prep assets for print.
60 models for this use case
60 models available
FLUX 1.1 Pro
Black Forest Labs' flagship text-to-image model. Faster generation than FLUX.1 Pro at higher prompt adherence, with strong photorealism and reliable spatial composition. Runs as a hosted Replicate model.
Flux 1.1 Pro Ultra
FLUX 1.1 Pro in ultra mode. Up to 4 megapixel images with raw mode for photorealism.
FLUX 1.1 Pro Ultra
FLUX 1.1 Pro in Ultra mode by Black Forest Labs. Generates up to 4 megapixel images with a raw mode for less processed, more natural-looking photography. Best FLUX option when output resolution and fine detail matter.
Flux Dev
Black Forest Labs' development model. Fast, high-quality image generation with LoRA support.
Google Imagen 4
Google DeepMind's Imagen 4 text-to-image model, hosted on Replicate. Sharp detail, accurate text rendering, and strong prompt adherence across photographic and illustrated styles. Outputs up to 2K resolution.
Google Imagen 4
Google's Imagen 4. Text-to-image with strong photorealism and improved typography support.
Google Imagen 4 Ultra
Premium Imagen 4 tier. Highest fidelity, prompt adherence and typography quality from Google.
Icons (SDXL Flat Pop)
SDXL fine-tune by galleri5 for slick flat icons and pop constructivist graphics with thick edges. Trained on Bing generations, it produces clean single-subject icon art that suits app icons, badges and UI glyphs. Raster output, not true vector.
Ideogram 3.0
Ideogram's flagship text-to-image model with industry-leading text rendering and prompt adherence.
Ideogram v3 Quality
The highest-quality tier of Ideogram v3. Improved photorealism and prompt adherence over v2 while keeping Ideogram's best-in-class text rendering. Supports style references and inline text layout.
InstantID
InstantID makes realistic portraits of a real person from a single reference photo without per-user training. Combines a face encoder with an IdentityNet adapter on SDXL to keep identity and pose while following a text prompt, so it is fast and tuning-free.
Kather100K Colorectal Tissue Classifier (ResNet50)
ResNet50 from the TIA Toolbox model zoo, trained on the Kather100K dataset of 100,000 hematoxylin-and-eosin colorectal histology patches. It classifies a tissue tile into one of nine categories such as tumor epithelium, stroma, lymphocytes, mucus, smooth muscle, debris, adipose, background and normal mucosa. Research use only, not a diagnostic device.
Midjourney V7
The latest Midjourney model. Industry-leading aesthetic quality and prompt adherence for image generation.
Professional Headshot (FLUX Kontext)
Turns any single selfie into a clean professional headshot using FLUX Kontext image editing. Keeps the person's face while swapping to business attire, a studio background and even lighting. Aimed at LinkedIn-style profile photos.
Recraft 20B SVG
Recraft's faster, cheaper vector model. Outputs editable SVG paths instead of raster pixels, so logos, icons and flat illustrations scale to any size without blur. Defaults to a vector_illustration style and supports line art and engraving looks. Hosted API only.
Recraft V3
Recraft's text-to-image model that topped the Hugging Face text-to-image arena at release. Strong long-text rendering, brand-style consistency, and precise control over image dimensions and color palettes.
Recraft Vectorize
Recraft's raster-to-vector converter. Takes a PNG or JPG and traces it into a clean SVG with precise vector paths, aimed at logos, icons and graphics that need to scale. Image-to-SVG counterpart to Recraft's text-to-SVG models.
Stable Diffusion XL
Stability AI's SDXL 1.0 with the optional refiner. The 3.5B base plus 6.6B ensemble UNet that became the default open image model before FLUX. Good for fine-tuning and LoRAs, broad community support.
Sticker Maker
fofr's sticker generator that outputs graphics with transparent backgrounds, so the result drops straight into chat apps or print sheets. Runs an SDXL-based pipeline at high speed (default 17 steps) and returns die-cut style art without manual background removal.
ViT Chest X-ray Classifier
Vision Transformer (ViT) fine-tuned on chest x-ray images for multi-class thoracic findings. Given a single frontal chest radiograph it returns class probabilities across several disease categories. One of the more downloaded chest x-ray classifiers on the Hugging Face Hub. Research and education only, not a diagnostic tool.
851-Labs Background Remover
Background removal model from 851-Labs that outputs a clean cutout with a transparent alpha channel. One of the most-run background removers on Replicate, handles people, products and objects on busy backgrounds.
Ad Inpaint (Product Photo)
Product advertising photo generator. You upload a cut-out product shot and a prompt describing the scene; it places the product on a new generated background with matching lighting and shadows, so a plain packshot becomes an ecommerce or ad-ready hero image without a photo studio.
AuraFlow v0.3
fal.ai's fully open-source 6.8B flow-based text-to-image model. Up to 1536x1536 resolution.
BiRefNet Background Removal
BiRefNet high-resolution dichotomous image segmentation for background removal. Bilateral reference network that produces sharp matting on fine detail like hair, fur and thin structures, often cleaner than older U2Net or rembg models.
Bone Fracture Detection (X-ray)
Image classifier by prithivMLmods that labels a bone x-ray as Fractured or Not Fractured. Given a single radiograph it returns binary class scores. One of the more downloaded fracture classifiers on the Hub. Research and education only, not a diagnostic tool.
BRIA Remove Background
BRIA AI's commercial background removal model trained on fully licensed data. Produces accurate cutouts for e-commerce and design, with attention to clean edges around products and people.
BRIA RMBG-1.4
BRIA's first commercial-safe background-removal model. Trained on fully-licensed data, suitable for production e-commerce and design pipelines.
BRIA RMBG-2.0
BRIA's professional background-removal model trained on fully-licensed data. Commercial-safe.
Bringing Old Photos Back to Life
Microsoft Research pipeline by Ziyu Wan et al. that restores scanned old photos, removing scratches, dust and fading and optionally enhancing faces in one pass.
Cartoonify
catacolabs Cartoonify turns a photo into a flat cartoon illustration. Takes a single image and returns a stylized cartoon version with clean shapes and bold outlines. Straightforward one-input model for avatars and profile pictures.
CCSR (Content-Consistent SR)
Content-Consistent Super-Resolution model. Reduces hallucination compared to typical diffusion-based upscalers while keeping perceptual quality high.
Clarity Upscaler
High-resolution image upscaler with creative detail re-imagination via SD-based hallucination. Strong for photography and product shots.
CodeFormer
Robust face-restoration model using a transformer-based codebook prior. Handles severe degradation, occlusion, and old-photo restoration with adjustable fidelity-quality tradeoff.
Consistent Character
fofr's model generates the same character in many poses and angles from one reference image. Useful for building an avatar set or character sheet where the face and design stay consistent across outputs. Can produce a grid or individual images.
ControlNet Canny
ControlNet conditioned on Canny edge maps. Preserves composition and outlines while restyling with Stable Diffusion 1.5 or SDXL backbones.
ControlNet Depth
ControlNet conditioned on depth maps. Preserves the 3D scene layout while letting the prompt change style, lighting and content.
DALL-E 3
OpenAI's latest image generation model. Excellent at following complex prompts with high fidelity.
DDColor
DDColor by Xiaoyang Kang et al. colorizes black-and-white photos using dual decoders that jointly learn pixel colors and semantic color queries, giving vivid and natural results on old images.
DINOv2 Skin Disease Classifier
DINOv2-base backbone fine-tuned for skin-disease image classification across 31 conditions, including basal cell carcinoma, lichen planus, lupus, herpes simplex, impetigo, leprosy variants and several genodermatoses. Broader than melanoma-only models. Research and educational use only, not a diagnostic.
DreamGaussian
Generative Gaussian-splatting model for fast image-to-3D synthesis. Produces textured meshes in two minutes via differentiable rasterization.
Ecommerce Virtual Try-On
Try-on pipeline aimed at ecommerce listings. You give it a photo containing clothing on a body pose plus a separate face image; it composes a person wearing that clothing with the supplied face, controllable by prompt, CFG, and output size. Useful for generating on-model product shots from a flat garment image.
ESRGAN Classic
Enhanced Super-Resolution GAN, the original 2018 architecture. Produces sharp 4x upscales with strong perceptual quality on natural images.
Face to Many
fofr's face stylizer converts a face photo into 3D render, emoji, pixel art, video-game character, claymation or toy styles. Uses InstantID plus style LoRAs on SDXL to keep the likeness while applying a chosen art style. Popular for fun avatars.
Face to Sticker
fofr's model turns a face photo into a die-cut sticker with a white border and transparent background. Uses InstantID to hold the likeness and outputs a clean PNG suitable for chat stickers or print. Simple single-image input.
FILM Frame Interpolation
Google FILM frame interpolation. Synthesizes high-quality intermediate frames between near-duplicate inputs, designed for large motion gaps.
FLUX PuLID
PuLID identity customization running on FLUX.1-dev. Inserts a face from one reference photo into prompt-driven scenes using contrastive alignment, giving higher likeness and detail than SDXL-era ID adapters. Good for realistic avatars and character portraits.
Flux Schnell
The fastest Flux model. Generate images in under 2 seconds. Great for prototyping.
FLUX.1 [dev]
The open-weight 12B rectified-flow transformer from Black Forest Labs. Close to FLUX Pro quality with a guidance-distilled checkpoint released under a non-commercial license. The most widely fine-tuned base in the FLUX family.
FLUX.1 [schnell]
The fastest FLUX model from Black Forest Labs, distilled to produce images in 1 to 4 steps. Apache 2.0 licensed for commercial use. Built for high-volume generation and real-time previews.
FLUX.1 [Schnell]
Black Forest Labs' fastest open-weights image model. Apache-2.0 licensed, ~1-4 step inference.
FLUX.1 Canny
FLUX structural control via Canny edge maps. Preserve composition while restyling.
FLUX.1 Canny [dev]
Open-weight edge-guided FLUX model from Black Forest Labs. Extracts Canny edges from a control image and regenerates it from your prompt while holding the original composition and outlines, so you can restyle a scene without changing its structure.
FLUX.1 Depth
FLUX structural control via depth maps. Keep 3D scene layout while changing style/content.
FLUX.1 Depth [dev]
Open-weight depth-guided FLUX model from Black Forest Labs. Derives a depth map from the control image and regenerates from your prompt while preserving 3D spatial layout, useful for re-texturing rooms, products, or scenes without moving objects.
FLUX.1 Fill
Black Forest Labs' inpainting/outpainting model for FLUX. Fill masked regions with prompt-guided content.
FLUX.1 Fill [dev]
Black Forest Labs' open-weight inpainting and outpainting model, guidance-distilled from FLUX.1 Fill [pro]. You supply an image plus a mask and a prompt; it fills the masked region or extends the canvas with prompt-guided content that matches lighting and texture.
FLUX.1 Kontext [dev]
Open-weight version of FLUX.1 Kontext by Black Forest Labs. Instruction-based editing: pass an input image and a plain text edit ('change the jacket to red', 'remove the person on the left') and it applies the change while keeping the rest of the scene and identity consistent.
FLUX.1 Redux
FLUX image-variation adapter. Generate variations and remixes from a reference image.
FLUX.1-dev Inpainting
FLUX.1-dev inpainting wrapper that fills masked parts of an image from a prompt. Useful when you want FLUX-quality fills with a simple image plus mask plus prompt interface and adjustable mask strength.
GFPGAN v1.4
Tencent ARC face-restoration GAN. Reconstructs realistic facial detail in low-quality or compressed photos using a pretrained StyleGAN2 prior.
Frequently asked questions
One API, pay only for what you use
Try any model with a free generation, no signup. Then 50 free credits and transparent per-use pricing, no subscription.
Related use cases
Generate clean, editable SVG vectors and convert raster images to vectors with AI. Compare Recraft, StarVector and more through one API.
Remove or replace image backgrounds automatically with AI. Clean cutouts with transparent alpha for products, portraits and objects.
Create professional AI headshots, avatars and stylized portraits from a photo or prompt. Consistent identity, many styles.
Generate logos, icons and line art with AI, including true SVG vector output for crisp scaling at any size.
Generate images from text with the top AI models - Flux, SDXL, Stable Diffusion, Ideogram, Recraft and more, through one API.
Generate video from text or images with AI - Kling, Hunyuan, LTX, Wan, Luma and more. Compare quality, length and price.