Originally posted to: https://huggingface.co/Qwen/Qwen-Image-2512
GGUF available here: https://civarchive.com/models/2267491
We are excited to introduce Qwen-Image-2512, the December update of Qwen-Image’s text-to-image foundational model. You are welcome to try the latest model at Qwen Chat. Compared to the base Qwen-Image model released in August, Qwen-Image-2512 features the following key improvements:
Enhanced Human Realism Qwen-Image-2512 significantly reduces the “AI-generated” look and substantially enhances overall image realism, especially for human subjects.
Finer Natural Detail Qwen-Image-2512 delivers notably more detailed rendering of landscapes, animal fur, and other natural elements.
Improved Text Rendering Qwen-Image-2512 improves the accuracy and quality of textual elements, achieving better layout and more faithful multimodal (text + image) composition.
We conducted over 10,000 rounds of blind model evaluations on AI Arena, and the results show that Qwen-Image-2512 is currently the strongest open-source model—while remaining highly competitive even among closed-source models.
Description
FAQ
Comments (28)
Is it normal that generating a single 9:16 (928×1664) image takes 3 minutes and 26 seconds on an RTX 3090 with 24 GB VRAM?
Or did I misconfigure something in my workflow?
Qwen Image is slow
Steps reduced to 10; now 4X+ faster.
2 minutes and 42 seconds on an RTX 5080 with 16 GB VRAM 64GB RAM?
@zzkszzks603 ah, use Lightning then
use lightning lora 8 steps
Only 3 minutes?
Mine take 4.5-6 minutes (30 steps) on RTX 5070.
considering the model is 39 GB I'd say that's pretty good lol
This one functions rather weirdly. On the site generator, I choose checkpoint 2511 in order to edit, but it shows 2512 in the generator. 2511 should be good with text, but with 30 steps it messes it up, and prompting should be more in the spirit of "inspired by," not like in the spirit of image editing. Yeah, I also, as far as I remember, img2img on the site has like 50% degradation. Overall... Confusing...
Status - "Dormant"....
Talk about an absolute sleeper hit; this model is amazing. The prompt adherence/natural language understanding, & image edit capabilities are outstanding. I also tend to prefer the base style outputs over FLUX (even the newer ones). It has nicer human faces; less wax/FLUX double chin nonsense going on. It's also less censored.
It's crazy it's not way more popular.... There's low VRAM GGUF variants anyone should be able to use, but those pages aren't popular either....
facts bro, it stupid, i tried this on god damn int4 and the quality and prompt adherence was still better than flux 2 max, chatgpt image or nano banana 2. and it looks better than all of them, especially the hands
This checkpoint is literally so damn good, it actually feels like cheating using it....lol
It turns even low effort, garbage prompts into gold lmao.
I have the thing running right now in the background, & it's spamming out 1584x1072 images one after the other in like 15-20seconds like a golden goose. It's becoming redundant favouriting the good images, because virtually all of them are.
You almost have to actively, & intentionally go out of your way to try & generate bad anatomy on purpose before it'll do it.
PS - I went away, & then came back after posting that last comment...... & in that brief time, it generated another 117 high resolution images; of which, only literally 3 images had minor anatomy defects! All the rest had perfect fingers, hands, etc.
PS, PS - I ended up generating over 2,000 images tonight.....lol. But I've never had so much fun using a checkpoint before. My ribs are literally aching from pain I've been laughing so hard generating so many insane things ^^
@ShadowCell Could you point me towards a decent workflow? Are you using it with the lightning lora?
@Themacsback https://huggingface.co/Wuli-art/Qwen-Image-2512-Turbo-LoRA/tree/main Wuli-Qwen-Image-2512-Turbo-LoRA-4steps-V3.0-bf16.safetensors
Try V3 - It seems to be quite a step up over others.
(As for workflows) Here's something I posted elsewhere........
"I use a lot of things, but I'm moving more & more over to SwarmUI. It's biggest advantage is that it has a fully working ComyfiUI backend which you can switch over to in 5 seconds (or without even firing Swarm up). & SwarmUI/Comfyi gets the latest/most updates to support new things/optimisations.
Personally, I'm not a fan of all this workflow/node bullshit lol. It's just too time consuming, & clunky most of the time (especially if you need to edit someone else's spaghetti workflow). With SwarmUI, you have the added control of that, or if you don't need it, & just want to change settings quickly, then you have the Forge like UI frontend of SwarmUI you can use instead.
https://github.com/mcmonkeyprojects/SwarmUI If anyone is curious; scroll down & "Download The Install-Windows.bat file". You basically just double click that, & it auto-downloads everything you need in 10mins or less. It's the easiest setup ever. For n00bs; It'll also automatically auto-download any clips or VAE, etc, you need for checkpoints if you have the wrong one, or none at all."
looks like 4 step Turbo Lora here is V1.0 Wuli (d71387110de235aa5405e6c03a464b0d). See compare to Wuli v2 and v3 and also lighting 4/8 steps in article
When I try to run this with the Turbo Lora in Forge Neo with 16gb VRAM/32 gb RAM I seem to get a crash relating to lack of memory. How much is actually needed, or did I misconfigure something?
if the model is 19 gb big, you need a minimum of 19 gb of memory. In practice, your operating system needs to run too, so you probably want a 24 gb vram gpu.
Model realism leaking
Is this a Checkpoint model or a diffusion model? Which workflow should I use for it?
nide
https://huggingface.co/Wuli-art/Qwen-Image-2512-Turbo-LoRA/tree/main V3 An amazing 4-8 step Turbo LoRA (You def want it)
1:1 1328x1328
16:9 1664x928
9:16 928x1664
4:3 1472x1104
3:4 1104x1472
3:2 1584x1056
2:3 1056x1584
These are the officially supported aspect ratios/resolutions. Qwen works much better if you stick to them (more so than some other checkpoints I've personally found).
Спасибо бро, очень помогло!
Рад помочь, товарищ
Half of the publications are missing
I can't publish a post
Looks like you're posting again (I can see your images in feed in the last 6 hours). Like you, I kind of thought I was shadow banned or something for some unbeknownst reason. Looks like it's civitai servers just being absolute potato shit again though lol.
My images were constantly going missing, or showing as 'not posted', etc.
Details
Files
qwenImage2512_imageEdit2511.safetensors
Mirrors
Qwen_Image_Edit_2511_BF16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
Qwen Image Edit 2511 [BF16].safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors
qwen_image_edit_2511_bf16.safetensors




