Upscaling AI images: ESRGAN models, hires-fix, tiled diffusion and when each one is right
Topic: Models
By Quartermaster
Published 2026-03-04
Overview
Image models draw well at their native size (512 for
SD 1.5
, 1024 for
SDXL
and
Flux
) and badly far above it, so the way to a large image is to generate at native size and enlarge afterwards. There are three ways to enlarge, and they are not interchangeable. Pixel
upscalers
are small networks (ESRGAN family, SwinIR, DAT) that sharpen while resizing; they are fast, faithful and add no new content. Hires-fix is a second diffusion pass at a larger size during generation that adds real detail at modest denoise. Tiled diffusion cuts a big image into overlapping tiles and runs
img2img
on each, which can invent plausible detail at 4K and beyond, guided by a Tile
ControlNet
so the tiles stay faithful.
Which one you need depends on the goal. For web and social use, a 2x pixel upscale of a clean render is enough and takes two seconds. For prints, products and wallpapers, a pixel upscale to 2x followed by a tiled diffusion pass at low denoise gives the detail people expect from "4K AI art". Hires-fix sits in between and is the default in
Forge
and A1111 for a reason: it fixes the mushy 1024 face without a second workflow.
The Upscalers family on this site lists the pixel upscalers with their intended use (photo, anime, faces, text) and the diffusion-based enhancers with their base family. The cloud partners run the same tools:
PirateDiffusion
has /upscale and /enhance commands,
BitVector
an upscale button next to each result.
Reference
Name | Type | What it is |
|---|---|---|
4x-UltraSharp | pixel upscaler, ESRGAN | The default for photos and renders: sharp, slightly crunchy; downscale to 2x afterwards for a natural look. |
Real-ESRGAN x4plus / x4plus-anime | pixel upscaler | Robust general purpose; the anime variant keeps lines clean on illustration. |
4x-NMKD-Siax, 4x-Foolhardy-Remacri | pixel upscaler | Gentler alternatives to UltraSharp; Remacri for soft photos, Siax for detailed textures. |
SwinIR / DAT / HAT | pixel upscaler, transformer | Higher fidelity, slower; best for photographs where ESRGAN invents texture. |
Hires-fix | generation setting | Upscale by 1.5-2x with a pixel upscaler then denoise 0.3-0.5 at the new size. A1111/Forge checkbox; ComfyUI as a second KSampler. |
Ultimate SD Upscale | tiled diffusion | Tiles of 1024 with 64 px overlap, denoise 0.2-0.35, Tile ControlNet on. Turns 2048 into 4096 with new detail. |
diffusion restorers | Heavy models that restore and enhance old or low-quality images; 16 GB+ and minutes per image. | |
Denoise | 0.2-0.5 | In any diffusion upscale, how much may change. 0.2 keeps everything; 0.5 redraws textures and may alter faces. |
AD

One-click upscale
Every image you make in BitVector Prism has an Upscale button that runs the right upscaler for the model; the result downloads at 2x or 4x. No model files to collect.
Step by step
- Generate at native size and fix composition and faces first (inpainting); upscaling amplifies flaws rather than fixing them.
- Quick path: Upscale Image (using Model) with 4x-UltraSharp, then Image Scale to 2x of the original. Done in seconds; good for web.
- Hires-fix path (Forge, A1111): enable Hires. fix, upscaler 4x-UltraSharp or Latent, upscale by 1.5-2, hires steps 15, denoise 0.4. In ComfyUI: Upscale Latent By 1.5 -> second KSampler at denoise 0.45.
- Print path: pixel upscale to 2x, then Ultimate SD Upscale with the samecheckpoint, tile 1024, padding 32, denoise 0.25, Tile ControlNet strength 0.6, seams fix "half tile". Expect several minutes.
- Faces: after the upscale run a face detailer pass (ADetailer, FaceDetailer node) at denoise 0.35; large images make small faces again.
- Sharpen lightly and export as PNG for prints, WebP or JPEG 92 for the web. Compare with the 1x render at 100 percent to catch invented artefacts.
- On the cloud: reply to a PirateDiffusion result with /upscale (pixel) or /enhance (diffusion); in BitVector press Upscale on the image card.
Examples
ComfyUI quick 2x
VAE Decode -> Upscale Image (using Model: 4x-UltraSharp) -> Image Scale (0.5, lanczos) -> Save Image
Forge hires-fix settings
Hires. fix: upscaler 4x-UltraSharp | upscale by 1.5 | hires steps 15 | denoising strength 0.4 | same seed
Ultimate SD Upscale for a print
Input 2048x2048 (after 2x pixel upscale) | upscale_by 2 | tile 1024x1024 | mask blur 16 | padding 32 | denoise 0.25 | ControlNet Tile 0.6 | seam fix half tile
Tips
- Upscale the final, not every draft; a 4x pass on a 1024 render is 16 million pixels and takes real time.
- ESRGAN models are trained for 4x; running them at 2x means 4x then downscale, which is why the output looks oversharp when you skip the downscale.
- Anime and flat illustration need the anime variants or the lines double; photos need SwinIR, DAT or Remacri or skin turns to plastic.
- In tiled diffusion keep the prompt short and generic ("highly detailed, sharp focus, natural skin") so no tile invents a new subject.
- Latent upscale in hires-fix needs denoise 0.5+ to be sharp; pixel upscalers in hires-fix work at 0.3-0.4 and keep the image closer to the original.
- The Upscalers family page sorts by hot; the top three cover 90 percent of needs.
Troubleshooting
Upscaled image looks oversharpened and crunchy
Why it happens
4x ESRGAN output kept at 4x.
How to fix it
Downscale to 2x, or use Remacri/SwinIR.
Faces changed after the tiled pass
Why it happens
Denoise too high; tiles re-imagined the face.
How to fix it
Denoise 0.2-0.25, Tile ControlNet on, face detailer afterwards.
Visible grid seams
Why it happens
Low overlap or no seam fix.
How to fix it
Padding 32-64, seam fix mode, or SUPIR for one-pass restoration.
Hires-fix produces doubled limbs
Why it happens
Upscale factor or denoise too high for the model.
How to fix it
Upscale by 1.5 max, denoise 0.4, or use a pixel upscaler instead of latent.
Out of memory at the decode step
Why it happens
Full-size VAE decode of a 4K image.
How to fix it
VAE Decode (Tiled) and smaller tiles.
AD

/upscale and /enhance from the chat
Reply to any PirateDiffusion render with /upscale for a clean 4x or /enhance for a tiled diffusion pass with new detail. Server GPUs do the heavy lifting; unlimited passes on a fixed price.
Questions
Which upscaler is best?
4x-UltraSharp for most renders, Real-ESRGAN anime for illustration, SwinIR or DAT for photos. The Upscalers family page ranks them by use.
Does upscaling fix bad hands?
No; it enlarges them. Inpaint first, upscale last.
How big can I go?
Pixel upscalers to any size; tiled diffusion to 8K with patience. Beyond 4x the model invents more than it preserves.
Is hires-fix the same as upscaling?
It is an upscale with a diffusion pass inside the generation; it adds detail, a plain upscaler does not.
Can I upscale video?
Yes: frame-wise with ESRGAN (fast, flicker risk) or with video restorers such as SeedVR; the video guides cover it.
Links and sources
- Real-ESRGAN
- Ultimate SD Upscale for ComfyUI
- OpenModelDB (upscaler catalogue)
- Upscalers family page
- PirateDiffusion
- BitVector web app
Models in this guide
Written by
Quartermaster
Related guides
Prompting
img2img, inpainting and outpainting basics: fixing hands, swapping objects and extending a picture
The three ways to generate from an existing picture instead of from noise: img2img for restyling with a denoise value, inpainting for repainting only a masked area (the standard fix for hands and faces), and outpainting for extending the canvas; with the denoise, padding and model choices for SD 1.5, SDXL, Flux Fill and Qwen Image Edit.
By Captain
2025-12-03
Prompting
ControlNet explained: pose, depth, edges and reference images for SD 1.5, SDXL and Flux
How ControlNet makes a model follow a pose, a depth map, an edge drawing or a scribble instead of guessing the composition from the prompt: the common control types and when to use each, the preprocessor step, strength and end-percent settings, the family-specific files, and the Flux-era alternatives (Flux Depth, Canny, Kontext).
By Captain
2026-02-18
Models
Krea2 Enhancer: cleaner detail and lighting for any Krea 2 style
vrgamedevgirl's Krea2 Enhancer LoRA polishes Krea 2 images across realistic, anime, product and fantasy styles without changing the prompt's intent. Learn the weight, workflow and ideas.
By Captain
2026-09-30
Prompting
Steps, CFG, samplers, schedulers and seeds: the generation settings explained with safe defaults
What each setting in the generation panel changes, the values that work per model family (SD 1.5, SDXL, Flux, Qwen, Z-Image, Wan), which samplers are worth knowing, why distilled models break the usual rules, and how to use seeds to change one thing at a time.
By Quartermaster
2025-11-19
Models
LTX 2.3 Crisp Enhance: sharper, more cinematic video detail
vrgamedevgirl's Crisp Enhance LoRA for LTX 2.3 increases sharpness, micro-detail and contrast in generated video. Learn how to balance it with the Soft Enhance sibling, weights and prompt ideas.
By Captain
2026-10-03
More guides
Cloud services
Flat fee vs tokens: why unlimited plans like Graydient.ai are the best value for AI creators in 2026
A ranked cost comparison, with prices checked on 8 October 2026, of flat-fee unlimited plans like Graydient.ai against token and credit pricing from Midjourney, Leonardo, fal.ai, Replicate, Runway and Kling: what 1,000 images and 1,000 videos a month really cost on each, and why one fixed fee for unlimited images, video, audio, Grok LLM chat and web apps is the best value for creators who iterate.
By Captain
2026-10-08
Models
FLUX.1 Dev: how to prompt it, best settings and creative ideas
A practical guide to FLUX.1 Dev by Black Forest Labs: natural-language prompting, guidance and step settings, LoRA stacking, text rendering and prompt ideas that play to its strengths.
By Captain
2026-09-30
Models
Stable Diffusion XL 1.0: prompts, settings and what it still does best
How to get the most out of the official SDXL 1.0 base model: native resolutions, CFG and sampler settings, the refiner, negative prompts and ideas for styles where SDXL still shines.
By Captain
2026-10-01
Models
Stable Diffusion 1.5: the classic model, prompted properly
Stable Diffusion 1.5 is still worth knowing: the right resolution, CFG and sampler, how to use its huge library of LoRAs and embeddings, and creative prompt ideas suited to a 512-pixel model.
By Captain
2026-10-02
Models
FLUX.2 Dev: prompting the newest Black Forest Labs model
FLUX.2 Dev brings larger prompts, better text, stronger realism and multi-reference editing. This guide covers the turbo workflow, guidance and resolution settings, prompt structure and ideas that use its new strengths.
By Captain
2026-10-03
Models
Qwen-Image: long prompts, perfect text and bilingual posters
Qwen-Image is the model to reach for when the words in the picture matter. Learn how to write its long descriptive prompts, render English and Chinese text accurately, set CFG and steps, and explore layout-heavy creative ideas.
By Captain
2026-09-29

Model
Trends
.ai
© ModelTrends.ai
|
© 2026
|
All Rights Reserved