Getting started with AI image generation: your first picture in ten minutes
Topic: Cloud services
By Captain
Published 2025-08-14
A plain-language start for complete beginners: what a model is, the three ways to run one (web app, chat bot, your own PC), how to write a first prompt, what the settings mean and how to tell a good result from a lucky one.
Overview
AI image generation turns a sentence into a picture. You type a description (the prompt), a model that has learned from millions of captioned images turns random noise into something that matches the description, and a few seconds later you have an image. Everything else in this hobby,
LoRAs
,
samplers
,
ControlNet
,
upscalers
, is a refinement of that one loop: describe, generate, look, adjust.
The model is the part that matters most. Different models have different strengths:
SDXL
and its anime fine-tunes draw illustration and characters,
Flux
and
Z-Image
follow long prompts and render readable text,
Wan
and
LTX
make video. ModelTrends.ai exists to show which models people are actually using this week, so you do not have to guess. Each model page shows example images with the exact prompt and settings that made them, which is the fastest way to learn.
You do not need a gaming PC to begin. The quickest first picture is on a cloud service where the models are already installed: a web app such as
BitVector
, where you pick a model and type, or the
PirateDiffusion
chat bot on Telegram, where you send a command. Running models on your own graphics card comes later, once you know which models you like and whether you want to spend the money on hardware.
Reference
Name | Type | What it is |
|---|---|---|
Prompt | text | What you want to see. Subject first, then setting, then style and lighting. Plain English works on modern models; keyword lists work on older ones. |
Negative prompt | text | What you do not want (blurry, extra fingers, watermark). Matters on SD 1.5 and SDXL; mostly ignored by Flux and other newer models. |
Model / checkpoint | choice | The brain. Pick from the hottest list on this site; the family decides the look and the prompt style. |
Size | pixels | 512x512 for SD 1.5, 1024x1024 for SDXL and Flux. Other shapes with the same pixel count work (832x1216 portrait, 1216x832 landscape). |
Steps | number | How many refinement passes. 20 to 30 is the normal range; more is slower and rarely better. |
CFG / guidance | number | How strictly the model follows the prompt. 5 to 7 on SDXL, 3.5 on Flux, 1 on turbo or lightning models. |
Seed | number | The starting noise. The same seed with the same settings gives the same picture, so you can change one thing at a time. |
Sampler | choice | The maths used to remove the noise. Euler or DPM++ 2M Karras are safe defaults; the model page shows what the examples used. |
AD

Make your first image without installing anything
BitVector Prism is a web app: pick one of the models from this site, type the prompt from the example above and generate. Nothing to download, nothing to configure, works on a phone.
Step by step
- Open the models browser on this site and sort by hot. Pick a base model (acheckpoint, not a LoRA) from the SDXL or Flux family; SDXL fine-tunes such as DreamShaper XL or Juggernaut XL are forgiving for a first try.
- Read the model page: look at three or four example images and their prompts. Notice the pattern: subject, scene, lighting, camera, style words.
- Press the Run button on the model page. BitVector opens a web app with the model preselected; PirateDiffusion opens a Telegram chat where you type /render followed by your prompt.
- Write a first prompt by copying an example and changing only the subject. Keep the size, steps and sampler the example used.
- Generate four images, not one. Random noise means some seeds are better than others; judging a model on one picture is a mistake.
- Change one thing: add a lighting word (golden hour, soft studio light), or a camera word (35mm, close-up), and generate again. This is how prompting is learned.
- Star the models you like with the heart on the card; the Favs page keeps them in your browser so you can come back without an account.
Examples
A first prompt for an SDXL photo model
a fisherman mending a net on a wooden pier at dawn, soft fog, warm light, 35mm photo, shallow depth of field
Negative: blurry, deformed hands, watermark, text | 1216x832 | DPM++ 2M Karras | 28 steps | CFG 6
The same idea on PirateDiffusion (Telegram)
/render <dreamshaper-xl> a fisherman mending a net on a wooden pier at dawn, soft fog, warm light, 35mm photo [[blurry, watermark]] /size:1216x832 /steps:28
A Flux prompt (no negative, lower guidance)
A candid photo of a fisherman mending a fishing net on a weathered pier at dawn. Soft fog over the water, warm low sun from the left, shallow depth of field, 35mm film look.
1216x832 | Euler | 24 steps | guidance 3.5
Tips
- The model is the biggest lever, the prompt the second, the settings the third. Beginners spend their time in the wrong order.
- Short, concrete prompts beat long vague ones. "A red bicycle leaning on a white wall, morning light, 50mm photo" is better than "beautiful stunning amazing 8k masterpiece bicycle".
- Portrait shapes for people, landscape shapes for scenes. Square is rarely the best shape for anything.
- Save the prompts you like in a text file. Every model page here shows the full prompt under its examples for the same reason.
- Anime and illustration models (Illustrious, Pony, NoobAI) use tag-style prompts; photo models (Flux, Z-Image, SDXL photo fine-tunes) use sentences. The family page explains which.
- Do not download anything until you have used a model on the cloud and liked it. Checkpoints are 2 to 24 GB each.
Troubleshooting
Every image looks the same
Why it happens
The seed is fixed, or the prompt is a list of generic quality words.
How to fix it
Set the seed to random, remove "masterpiece, 8k, best quality" from photo prompts, describe the scene instead.
Hands and faces are wrong
Why it happens
Small faces and hands at low resolution, or an old SD 1.5 model.
How to fix it
Use SDXL or Flux at 1024 pixels, frame the subject larger, or inpaint the face later (see the
inpainting
guide).
The picture ignores half the prompt
Why it happens
Too many ideas in one prompt, or CFG too low.
How to fix it
One subject, one setting, one style per prompt. Raise CFG by one on SDXL. Flux and
Qwen
follow long prompts far better than SD 1.5.
Text in the image is gibberish
Why it happens
SD 1.5 and SDXL cannot spell.
How to fix it
It takes minutes per image on my laptop
Why it happens
No dedicated GPU, or a card with less than 6 GB of memory.
How to fix it
Generate on the cloud (BitVector, PirateDiffusion) and keep the laptop for browsing results.
AD

Thousands of models in a chat window
PirateDiffusion runs on Telegram: send /render and a prompt and the image comes back in the chat. Every model and LoRA on this site with a Run button is already installed there, on a fixed monthly price with unlimited generation.
Questions
Do I need to pay to try this?
Both cloud partners linked from the model pages let you start without installing anything; local tools (ComfyUI,
Forge
) are free but need a graphics card with at least 8 GB of memory for current models.
Which model should a beginner start with?
An SDXL fine-tune for illustration or a
Flux dev
model for photos. The "which models to try first" guide on this site ranks them with reasons.
Is a LoRA a model?
A LoRA is an add-on that changes a base model. You need a base model (checkpoint) first; the LoRA goes on top. The LoRA guide explains it.
Can I make video too?
Yes, with Wan,
LTX-2
and
MiniMax H3
among others, but video needs far more memory locally. Start on the cloud; the video getting-started guide covers it.
Are the images mine?
Licences differ per model; the model page links to the licence. Most open models allow personal and commercial use, some research-only licences (early Flux dev builds) do not.
Why do my results look worse than the examples?
Usually a different size, sampler or step count. Copy the example settings exactly first, then change one thing at a time.
Links and sources
- Models browser (sort by hot)
- Weekly ranking
- BitVector web app
- PirateDiffusion on Telegram
- SDXL family page
- Flux family page
Models in this guide
Written by
Captain
Related guides
Models
Which AI image models to try first, and why: a ranked starting list
Seven models for a first week, chosen for forgiveness rather than hype: one SDXL photo fine-tune, one anime fine-tune, Flux dev, a fast turbo model, an inpainting model, a video model and an upscaler, each with the reason, the memory it needs and the prompt style it expects.
By Captain
2025-09-10
Prompting
Prompt writing basics for AI images: structure, order, weights and the two prompt dialects
How to write a prompt that works on the first try: the subject-scene-style order, why word position matters, the tag dialect of SDXL anime models against the sentence dialect of Flux and Qwen, attention weights, prompt length limits and a checklist for fixing a prompt that is being ignored.
By Captain
2025-10-22
Models
How generative AI image models work: diffusion, latents and text encoders explained simply
The mechanics behind Stable Diffusion, SDXL, Flux and the video models without the maths: what noise has to do with it, why models work in a compressed latent space, what the text encoder and VAE do, what steps and guidance really change, and why training data decides what a model can draw.
By Quartermaster
2025-08-28
Models
What is a LoRA and how to use one in ComfyUI, Forge, Automatic1111 and on the cloud
LoRAs explained for people who want results, not maths: what a LoRA file changes, why it must match the base model's family, trigger words, weights, stacking several LoRAs, the folder and syntax for each UI, and how to try a LoRA on the cloud before downloading anything.
By Captain
2026-05-20
Cloud services
PirateDiffusion: run any model from Telegram, no GPU needed
The commands that matter in the PirateDiffusion Telegram bot: /render with model trigger words, (( )) and [[ ]] weighting, #recipes, /adetailer, /highdef and /facelift upscaling, /remix, /inpaint, ComfyUI workflows with /wf /run:, and the new // skills that pick the model for you.
By Captain
2026-10-03
More guides
Models
FLUX.1 Dev: how to prompt it, best settings and creative ideas
A practical guide to FLUX.1 Dev by Black Forest Labs: natural-language prompting, guidance and step settings, LoRA stacking, text rendering and prompt ideas that play to its strengths.
By Captain
2026-09-30
Models
Stable Diffusion XL 1.0: prompts, settings and what it still does best
How to get the most out of the official SDXL 1.0 base model: native resolutions, CFG and sampler settings, the refiner, negative prompts and ideas for styles where SDXL still shines.
By Captain
2026-10-01
Models
Stable Diffusion 1.5: the classic model, prompted properly
Stable Diffusion 1.5 is still worth knowing: the right resolution, CFG and sampler, how to use its huge library of LoRAs and embeddings, and creative prompt ideas suited to a 512-pixel model.
By Captain
2026-10-02
Models
FLUX.2 Dev: prompting the newest Black Forest Labs model
FLUX.2 Dev brings larger prompts, better text, stronger realism and multi-reference editing. This guide covers the turbo workflow, guidance and resolution settings, prompt structure and ideas that use its new strengths.
By Captain
2026-10-03
Models
Qwen-Image: long prompts, perfect text and bilingual posters
Qwen-Image is the model to reach for when the words in the picture matter. Learn how to write its long descriptive prompts, render English and Chinese text accurately, set CFG and steps, and explore layout-heavy creative ideas.
By Captain
2026-09-29
Models
SDXL-Lightning: four-step generation on any SDXL checkpoint
ByteDance's SDXL-Lightning LoRA turns a 30-step SDXL render into a 4-step one. Learn the step counts, the CFG you must use, the sampler, and how to combine it with your favorite checkpoints and style LoRAs.
By Captain
2026-09-30

Model
Trends
.ai
© ModelTrends.ai
|
Made in Japan
|
© 2026