[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


Discussion and Development of Local Image, Video, and Music Models

Previous: >>109861523

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP
Neural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/neo_collage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: Qwen_image_2.1_00032.png (2.34 MB, 1760x1312)
2.34 MB PNG
I think this might be due to the sampler but I need to adjust some values, not sure on the best combo yet
>>
Blessed thread of frenship
>>
File: Qwen_image_2.1_1_2.jpg (506 KB, 1184x1776)
506 KB JPG
>>
Can someone do "closed eyes + moaning = pooping"?
>>
>>109864888
You can run the model right?
>>
>>109864856
thanks for another schizo bake ran
>>
>inb4 ffaze is lonely again
>>
File: Qwen_image_2.1_1_4.jpg (578 KB, 1184x1776)
578 KB JPG
>>
File: Qwen_image_2.1_00034.png (2.28 MB, 1760x1312)
2.28 MB PNG
I used eyes instead of eye
>>
>>109864926
Is it ran baking threads with schizo links? It all makes sense now.
>>
>julien
>>
>>109864954
can you neg "ugly"?
>>
File: flux_sucks.png (2.76 MB, 1072x1680)
2.76 MB PNG
>>109864939
I've been thoroughly unimpressed with the new Qwen. But I guess it's better than Flux2Klein
>>
I have an image I'm trying to edit using natural language. But ChatGPT is basically telling me it won't process sensitive content when I upload the image there. I found that odd since there is no nudity or public figure, but whatever.
If I want to do this locally do I need a diffusion model or do I need a multi-modal reasoning model?
The best machine I have access to at the moment is a gaming laptop with 16 GB of VRAM and it's like an RTX 2090 or something from like 4 years ago. My understanding is its weak for AI.
I guess I could rent a private cloud GPU somewhere but that sounds complicated.
>>
>>109864974
He would cease to exist if he did that
>>
>>109865015
You can run qwen image 2.1, unlike the old one this works in fp16 without nan's so it won't be too slow.
>>
File: Qwen_image_2.1_1_9.jpg (777 KB, 1184x1776)
777 KB JPG
>>109864992
flux can be great but it requires a large lora collection sadly
>>
i just updated comfy to try the new qwen model and it's taking so much longer to start. what tehf uck
>>
>>109865045
you new? comfy tends to fail catastrophically with every other update.
>>
This is the guy seething in the thread btw
>>
>>109865068
i think it was just the first run, seems fine now. i kept a backup of my old one just in case though
>>
File: Qwen_image_2.1_1_11.jpg (457 KB, 1184x1776)
457 KB JPG
>>
File: asdfafs.png (178 KB, 1008x853)
178 KB PNG
>>109864856
Anyone tried the Fast H3 for Minimax yet? Regular H3 is painfully slow and this seems like a good way to rough draft and refine prompts more quickly.

Also

>what loras do you guys use for H3?
>are there any good image gens out there? currently using z-image for text and quen for image to image and just wondering if there's anything better
>>
>>109865072
the only trans seething here is you ran
let us discuss local tech in peace
>>
>>109865030
Yeah and right now out-the-box stuff is where it's at. It's not like a video game where we'll come back to Flux2Klein in 2030 like I go back to modded Skyrim in 2026
>>
>>109865092
It's shit quality
>>
>>109865099
Are you drunk again?
What is
>>109864926
>>109864958
>>109865016
>>109865068
>>
>>109865123
have you maybe considered that everyone here fucking hates catjak and her stupid spam?
>>
>>109865119
Nobody goes back to give support to these older models
>>
File: 1740201574106435.jpg (2.49 MB, 1664x2432)
2.49 MB JPG
>>109864881
>>
File: 0_00013.png (2.25 MB, 1280x1856)
2.25 MB PNG
>>
>>109865129
I agree with you but there is a report button that gets the job done instead of complaining about it and shitting up the thread more than it already is by the mentally ill troon
>>
>>109865092
the power of H3 is that you don't really need LoRAS. It's good out of the box. Not perfect, but still very good.

The issue with Fast H3 and everything that speeds it up is that real H3 gens will probably look nothing like the Fast H3 gens. Drafting is done better at like 0.1MP on a regular H3 gen rather than piling up speed up stuff on it.

As for the image gen question, yeah. Flux2Klein. Krea2.
>>
>>109865129
what spam?
explain succinctly and i may see things your way
>>
Literal ban evading cock sucking schizo projecting
Please look at all his /d/ post if you can stomach it.
>>
which schizos are real and which schizos just exist in the minds of the real schizos
>>
>>109865172
Better question is, where the fuck are the jannies and mods and why do they allow this shit to go on?
>>
Let's settle this
https://strawpoll.com/xVg719K1Ryr
>>
File: the-heist.jpg (861 KB, 2304x1792)
861 KB JPG
>>
>>109865188
No poll is safe from your proxyfagging, ban evader
>>
>>109865179
Same reason why the mods let us keep the rentry and delete your post. Every diffusion thread wants you gone just like how your last ban begging post got deleted.
Oh I forgot you only have a wrapper and can't add the new model until the sdcpp devs do the actual work.
>>
>>109865193
If you keep going I am considering these as spam
>>
File: file.png (54 KB, 841x606)
54 KB PNG
>>109865200
;)
>>
File: 0_00015.jpg (1.32 MB, 2928x2304)
1.32 MB JPG
>>
File: file.png (22 KB, 964x188)
22 KB PNG
Wow so many people from sweden in //ldg/ and voting in the poll all of the sudden
>>
>>109865146
Notice how doesn't reply to your post
>>109865188
GOT EEEEEM
>>
File: the-massage.jpg (2.29 MB, 4096x3072)
2.29 MB JPG
>>109865209
considering this is my first post in this thread, I'm not sure what you're on about.
>>
>>109865129
>her
>>
File: Qwen_image_2.1_1_16.jpg (360 KB, 1184x1776)
360 KB JPG
>>109865131
>>
>>109865241
Thought you were the SpongeBob fried qwenslop poster. Please excuse me Biden + toph fried krea2 spammer
>>
>>109865241
the failed developer of tranustudio probably saw a jewish rat and thought it was a personal attack towards xir
>>
File: cool-toph.jpg (650 KB, 2304x1792)
650 KB JPG
>>109865253
>>109865262
S'all good frens, have a cool Toph on me :)
>>
>>109865272
Keep em to yourself. Treasure them privately
>>
why does the schizo aggro at any criticism of cumstainui? do we unironically have paid shills here now?
>>
>>109865288
He does it for free
>>
New qwen 2.5 can do cock and balls fine but can barely do pussies the fuck? What is this double standard?
>>
can i be runnings the qwen 2.1 Image Model with 4gb of VRAM?
>>
File: toph-skating.jpg (985 KB, 2944x2176)
985 KB JPG
>>109865301
to be fair, a pee-pee is a lot easier to draw than a puh-puh
>>
i downloaded the latest comfyui (0.36.0) and am trying to get qwen 2.1 image edit to run, but it's telling me i'm missing QwenImage21Cache and TextEncodeQwenImage21
>>
>>109865301
Lead Google engineer bias
>>
>>109865315
>i downloaded the latest comfyui
oof...
>>
>>109865315
Pull, don't download releases like a pleb.
>>
>>109865324
>>109865323
oh, fucking comfy niggers
>>
>>109865310
I gave it a reference image of what the vajeyjay should look like and it still DOESN'T wanna render it.
>>
>>109865315
Update again, also you don't need the cache node for it to work.
>>
>>109865339
>blaming your technological ineptitude on devs
I am 100% sure you use windows.
>>
>cumfart shat itself again
to nobody's surprise
>>
>>109865353
Running fine for me, are you one of those retards that can't update the requirements.txt?
>>
File: the-goon-coin.jpg (1.15 MB, 2944x2176)
1.15 MB JPG
>>109865343
I'm sorry to hear, fren. Hopefully it's resolved by someone who likes puh-puh's eventually
>>
>>109865353
Are you retarded?
>>
>>109865288
>>109865290
>>109865323
>>109865353
So where you from, Jules
Sweden, Germany, Hungary or the US?
>>
File: Qwen_image_2.1_1_21.jpg (226 KB, 832x1248)
226 KB JPG
git pulled today after not using comfy for months. worked instantly.
>>
>unironic comfyui defence force
Fucking grim.
>>
>>109865363
Why should "professional software" require anything other than an update button that just works?
>>
>>109865379
no one here has fantasized about cumfart's cock, let alone sucked it except for you, you worthless retard
>>
>>109865378
Stop generating spongeslop I want anime books on my racism fourm
>>
>>109865384
So you are slow and can't read basic instructions. Stick to the release binaries for windows users. Just wait your turn until the next update
>>
>>109865365
It's really no fair. I spent two hours trying to get a crumb for nothing but the moment I input penis in the prompt I get a 10 inch column of veiny bitchbreaking schlong.
>>
>update comfy
>it works like it does every time
I have no idea why updating comfy seems to be this unsolvable riddle to you faggots
Like what are even doing? Updating via git worked and on my laptop comfy desktop updates also work
Are you guys retarded?
>>
No thanks. Im sticking with Illustrious + Anima + Photoshop
>>
>>109865315
You can use the latest master version instead of the latest stable, it should be 0.37.0
>>
>>109865411
I mean, it's probably the failed dev, so I am not really surprised at its level of tech illiteracy
>>
File: Screenshot_3617.png (13 KB, 1468x107)
13 KB PNG
oof
>>
>>109865422
What gpu and res?
>>
File: the-coolest-skater.jpg (1.24 MB, 2944x2176)
1.24 MB JPG
>>109865406
we can only hope it'll work one day!
>>
>>109865411
Maybe you just have a basic bitch install and workflow.
>>
>>109865432
3050ti, 1mpb so im guessing 1024x1024
>>
>>109865436
Post log then, i have a lot of stuff installed + vibe coded nodes
So show your log + cli output if any, we will help you with your serious problem anon
>>
>>109865441
I m getting half of that on an rtx 2060. Do you have any attention nodes enabled?
>>
File: zula.jpg (1.05 MB, 2944x2176)
1.05 MB JPG
>>109865435
>>
>>109865448
not that im aware of, using the default qwen edit template
>>
File: Qwen_image_2.1_00040.jpg (1.03 MB, 1248x1856)
1.03 MB JPG
I don't think this works too complex. Also it's funny I can sit back and just watch the fireworks.
>>
File: katara-goon-coin.jpg (1.3 MB, 2944x2176)
1.3 MB JPG
>>109865463
the trilogy is complete
>>
>>109865465
Do you perhaps not have enough ram? that can be a cause of slowdown.
>>
>>109865446
Shockingly, my install hasn't nuked itself in about six months. I gotta pull today though so we'll see if my luck (and it IS luck) holds.
>>
File: Qwen_image_2.1_1_27.jpg (470 KB, 1184x1776)
470 KB JPG
>>109865401
>>
>>109865503
So you're a retarded drama faggot
Got it
>>
>>109865411
retards update python_deps which basically always borks your installation. you almost never need to run it
>>
File: goon-ui.jpg (1.16 MB, 2944x2176)
1.16 MB JPG
>>109865514
>>
>>109864856
>mfw Resource news

09/19/2026

>HyperFlow: Video Rebirth's 8-step LoRA for MiniMax-H3
https://github.com/Saganaki22/ComfyUI-Hyperflow

>HyperFlow 8-Step LoRAs ComfyUI Conversions
https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI#update--hyperflow-8-step-loras

>Day-0 Sol Engine + Sol-Attn for HyperFlow
https://github.com/NVlabs/Sana/tree/sol-engine/models/minimax_h3/HyperFlow

>Local Dream v3.0.0-alpha.1
https://github.com/xororz/local-dream/releases/tag/v3.0.0-alpha.1

>ComfyUI-NodeSnapshots: 2-3x your frontend FPS with caching
https://github.com/SparknightLLC/ComfyUI-NodeSnapshots

>Almost 86% of Japanese game developers are using AI in their workflows, up from 51% in 2025
https://www.techspot.com/news/113892-nearly-86-japanese-game-developers-using-generative-ai.html

>MiniMax-H3-Singularity GGUF
https://huggingface.co/Abiray/MiniMax-H3-Singularity-GGUF

>ComfyUI-DAAM-Pack: See which prompt tag shaped which part of the image
https://github.com/alchemine/comfyui-daam-pack

09/18/2026

>Perceptual Refinement of an End-to-End Video Streaming Pipeline via Generative AI Layers
https://github.com/emanuele-artioli/presley

>PACE: Precise AI Cinematic Expression: A Typed Specification for Script-Grounded Previsualization and Geometric Conformance
https://github.com/StudioPiLabs/pace-core

>LingBot-World 2.0 realtime
https://github.com/kaarelkaarelson/lingbot-world-v2-realtime

09/17/2026

>Minimax H3 Latent Upscaler (3D) — BF16 Conservative v5
https://huggingface.co/Asirus/Minimax-H3-Latent-Upscaler-BF16-MAXQUALITY

>Kijai MiniMaxH3 INT8 VAE
https://huggingface.co/Comfy-Org/MiniMax-H3/commits/main/vae/minimax_h3_video_vae_int8_convrot.safetensors

>Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control
https://zing.loopit.me

>comfyui-SelfLift: Progressive-resolution sampling for ComfyUI
https://github.com/facok/comfyui-SelfLift

>LTX-2.5-uncensored-v1.1-FP8
https://huggingface.co/frtertaer/LTX-2.5-uncensored-v1.1-FP8
>>
>>109865556
Fuck off unemployed loser
>>
>>109865556
Thanks bro.
>>
File: just-got-this-snack.jpg (1.21 MB, 2944x2176)
1.21 MB JPG
just bought this, should I eat it?
>>
>>109865575
No he can still be saved!
Bring him to the nearest coin fountain
>>
File: debo_hd_k2_00003_.png (2.52 MB, 1664x1069)
2.52 MB PNG
>>109865572
np :)
sorry for nothing fresh today, but I missed getting yesterday's stuff up in the last thread
>>
File: Qwen_image_2.1_00004.png (1.28 MB, 1920x1088)
1.28 MB PNG
is the default cumfartui workflow broken, as usual?
i set the cfg to 6 and its still coming out ghosted.
>>
Ali Baba And The 40 Billion GPT Outputs
>>
File: restored-rat.jpg (1.57 MB, 2944x2176)
1.57 MB JPG
>>109865582
You were right, anon!

He's getting a bit too powerful though...
>>
>>109865586
I was being sarcastic.
>>
File: the-rat-reborn.jpg (3.11 MB, 4096x3072)
3.11 MB JPG
>>109865637
must be fun being unemployed and just attempting to ragebait like this - a true dream!

你不太聪明。
>>
files.catbox.moe/5o9myc.mp4
>>
File: debo_hd_k2_00006_.png (2.14 MB, 1664x1069)
2.14 MB PNG
>>109865664
you'd think he's come up with new material with all the free time he has
>>
>>109865637
No I wasn't, Ran.
>>
File: the-friendly-rat.jpg (2.26 MB, 4096x3072)
2.26 MB JPG
I met another rat when visiting 北京 wutongs.

This one seems less Jewish and more friendly!
>>
Crazy how both rentry schizos are just making shit up and seething instead of posting elsewhere
>>
File: Qwen_image_2.1_00006.png (3.26 MB, 2368x1344)
3.26 MB PNG
>>109865598
>this is the same prompt at 50 steps
wha
>>
>>109865092
it fucking trash. Fries image and barely faster than sparse attention
>>
File: Krea2_turbo_03988_.jpg (1.49 MB, 1776x2368)
1.49 MB JPG
>>109865725
The defaults are fucking garbage, they didn't test it before pushing that shit workflow out.
>>109865710
Because they all want to be here. I don't need to hide to make my feelings known either which makes it more pathetic.
>>
>>109865556
Thanks for the news, friend.
>>
Model looks awful and slopped. You can tell they're using a dogshit VAE.
>>
File: peace-out.jpg (3.37 MB, 4736x2688)
3.37 MB JPG
Yet another person I met on my 北京 adventures was former president Joe Biden!

He was quite friendly, and agreed to let me photograph him <3
>>
>>109865735
I would like to try this style, do you use lora or is it just a krea2 prompterinos?
>>
>>109865556
thanks!
>>
File: Qwen_image_2.1_1_36.jpg (855 KB, 1184x1776)
855 KB JPG
>>
>>109865556
>>HyperFlow: Video Rebirth's 8-step LoRA for MiniMax-H3
>https://github.com/Saganaki22/ComfyUI-Hyperflow
I'm interested to know if anyone has reviews on this yet
>>
File: Qwen_image_2.1_1_37.jpg (517 KB, 1184x1776)
517 KB JPG
>>
File: Qwen_image_2.1_1_38.jpg (650 KB, 1184x1776)
650 KB JPG
>>
>>109865807
nta but there has never been a pro-ani post sharing a gen
you clearly don't care about this tech you're just here to make a quick buck shilling your dead project
grim
>>
is there an offcial qwen 2.1 image/edit prompt guide?
>>
File: 1789919185950105.jpg (43 KB, 850x319)
43 KB JPG
>>109865842
https://github.com/QwenLM/Qwen-Image-2.1/blob/main/prompt_rewrite/prompts/system_prompt_edit.txt
Try this test too,
>>109865841
If you look at the idiot get baited by the poll you'll understand how low IQ they are.
>>
File: black_hole.jpg (609 KB, 1615x1055)
609 KB JPG
>>
File: Qwen_image_2.1_1_40.jpg (314 KB, 1184x1776)
314 KB JPG
>>
>>109865860
Should be self inflicted tbqh
>>
>>109865860
and this is why anistudio is unsafe to run
no thanks no one in /ldg/ is going to run mystery binary blobs from a murderous mentally unstable fuck who is on record saying he will "get rid of" his love rivals irl
>>
>>109865556
thanks bro
>>
>>109865860
If you get this triggered please stop posting here
>>
>>109865860
>>109865867
>post times
Do you normally struggle with keeping your suicidal thoughts internally?
>>
File: Qwen_image_2.1_1_41.jpg (264 KB, 1184x1776)
264 KB JPG
>>
File: 1754318122369966.jpg (309 KB, 1125x843)
309 KB JPG
what does the nsfw part look like?
>>
File: joe-snowden.jpg (740 KB, 2048x1536)
740 KB JPG
>>109865830
I like the colors on this one, anon!

A bit fried, but then again, so are all of my gens.
>>
>>109865886
vageens are cthulhu tier
benis and balls are top notch
so basically a model for faggots
>>
File: debo_hd_k2_00007_.png (2.56 MB, 1664x1069)
2.56 MB PNG
>>109865841
>inb4 debo doesnt count
I'm pro ani, insofar as I hope he can achieve more success with his projects. he's smart and ambitious and I hope he continues to be encouraged to pursue his ideas. I've been inspired by him in the past, so its a shame he'd been scared off from posting here
>>
>>109865789
A lot more slopped than what people say krea2 is
>>
>>109865910
damn, debo is even more based than I thought
>>
>namefag1: i like namefag2!
why can't these faggots not just fuck off
>>
>>109865910
I think he doesn't want the mentally unhinged retard throwing a giga tantrum every time he wants to give an update. If you are really pro ani, you'd also stop advertising comfy slop
>>
File: Qwen_image_2.1_1_43.jpg (638 KB, 1184x1776)
638 KB JPG
>>
>>109865910
no surprise debo, you both have rentries for shitting up threads because both of you have no irl social life
>>
ani really should post updates here about anistudio
>>
File: Qwen_image_2.1_1_45.jpg (444 KB, 1184x1776)
444 KB JPG
>>
They released Qwen because it was a training accident.
What a disappointment.
>>
kill catjak and piss on his corpse
>>
this thread will be the one I finally take down pissbuttfag for good
>>
>>109865989
Looks like a distilled gpt image
>>
i dont understand why the qwen chinks even release the trash they produce
every single one of their models underperforms. every single one. like why even bother at this point
even their mascot is shitty
>>
File: Qwen_image_2.1_00012.png (3.86 MB, 2368x1344)
3.86 MB PNG
This is retarded, so can you NOT give this thing a specific aspect ratio different from the input images? I try to do that with the default workflow and it just generates nonsense like picrel.
>>
>>109866023
BUT MUH BENCHMARKS!!!
>>
>>109866026
Should be possible. Post your graph.
>>
not a single one of these qwen gens are good. like im not remotely curious in downloading this shit. back to h3/krea2/klein9b
>>
>>109866031
https://files.catbox.moe/llny97.png
>>
>>109866026
Did you switch the toggle to do it correctly on the node?
>>
>>109864954
The eyebrows look better in the other one.
>>
I want to stab my eyes everytime I see a gpt image gen with that fucking weird noise and dark / over saturated image and qwen image 2.1 is only going to make it worse.
>>
>>109865968
>>109865976
>>109866005
Least conspicuous retard ever lmao
>>
>>109866057
yes, and that made it worse kek
>>
File: res.png (91 KB, 1123x596)
91 KB PNG
>>109866040
Unless there is some unknown faggotry, you could do it like this. I don't even understand the meaning of this resolution selector node because it just obfuscates the meaning of a resolution anyway.
>>
did a VAE round trip comparison (load -> encode -> decode -> save) on an actual stock photo

Qwen 2.1 VAE is better than the OG Qwen VAE but still worse than both the Flux.1 and Flux.2 VAE I'd say

4Chan says "image resolution too large" so boxed it:
https://files.catbox.moe/pyangf.png
>>
File: ComfyUI_1254.png (1.38 MB, 1152x896)
1.38 MB PNG
balls
>>
File: qwen fox.png (1.89 MB, 2312x1057)
1.89 MB PNG
>>109866080
pretty sure we're accomplishing the same end goal, i'm just doing it automatically by MP.
>>
File: Qwen_image_2.1_1_53.jpg (145 KB, 1376x768)
145 KB JPG
>>
>>109866093
This wasn't the original question. You asked about the resolution.
Seems like you don't understand the concept of resolution.
>>
>>109866112
You don't understand the concept of love.
>>
File: bleeeh!.jpg (316 KB, 1152x896)
316 KB JPG
>>
>>109866115
Sure enough, I do understand why this thread is full of shitters and faggots and why no one will ever post any images in here either.
>>
File: Qwen_image_2.1_00018.png (2.08 MB, 1714x1227)
2.08 MB PNG
Why cumfy set cfg at 1.0 ( is it a distilled model).
The proportions are not realistic at all. More like just a filter
Turn this stylized 2D character into a realistic human photo. Replace the exaggerated cartoon body proportions with natural adult human anatomy and realistic proportions. Make the skin, face, hair, clothing, and materials fully photorealistic.
>>
File: Qwen_image_2.1_00024.png (2.04 MB, 1714x1227)
2.04 MB PNG
>>109866126
cfg 2.5, for example, does improve prompt adherence, but it shows signs of error
>>
>>109866126
>club foot
>>
File: Qwen_image_2.1_00052.png (2.74 MB, 1136x2016)
2.74 MB PNG
>>
File: Qwen_image_2.1_1_57.jpg (453 KB, 1376x768)
453 KB JPG
>>
File: Qwen_image_2.1_00026.png (2.08 MB, 1714x1227)
2.08 MB PNG
they don't want you to use a higher cfg because they want to hide their gpt image's footprint
>>
>>109866179
I think that's the case kek, it's still a better edit model and I hope someone finetunes it for that alone
>>
File: garaxy.jpg (731 KB, 1611x1111)
731 KB JPG
It's getting pretty limited.
>>
any finetunes happening or is it over for coomers
>>
File: galaxy.jpg (950 KB, 1900x1333)
950 KB JPG
One more.
>>
>>109866204
new models are releasing too rapidly for anyone to dump serious money and time into a finetune.
>>
File: Qwen_image_2.1_1_61.jpg (1013 KB, 1184x1776)
1013 KB JPG
>>
>>109866212
H3 is likely to be SOTA for a while. I can see an ecosystem around it like LTX 2.1 given how long it was in the lead. I suspect no one is going to try and release a competitor to H3 for a while. For image diffusion, I would say so but for most people, it's been diminishing returns since the original FLUX model came out. The only thing worthwhile is anime and cartoons which still aren't served well by the models coming out that well.
>>
File: Qwen_image_2.1_00055.png (3.55 MB, 1328x2368)
3.55 MB PNG
>>
>>109866212
only the inept pedophiles are lurking for finetunes lmao
morons cant even go on the deepweb to download their bullshit
>>
>>109866247
>totallyaloliconmodel.safetensors.exe
>>
I'm trying to use ADetailer in ForgeNeo and I can't get the MediaPipe Face Mesh to work. The face_yolo stuff works fine but the mediapipe just doesn't detect any faces/eyes ever. Is this a compatability issue? Would changinge models/sampling work at all? I'm going for realistic rather than anime or cartoon models.
>>
>>109866212
>new models are being released too rapidly
More like bad models are released, and nobody gives a fuck since training costs have increased exponentially compared to 2 years ago.
LTX 2.5 gets a fraction of H3's attention because it's a bad model. Even H3 is a cuck model itself since it's so hard to finetune.
>>
File: Qwen_image_2.1_1_65.jpg (798 KB, 1184x1776)
798 KB JPG
>>
oops
>>109864570
well to be fair it's not intuitive

>>109863995
thanks ill give it a try, didnt want to force my pc after that so I didnt even try again, I did deleted the models since I assumed it would be impossible with amd gpu and only 32gb, ill give it a try another day, ill look the options in the templates though, in the meantime
>>
File: Qwen_image_2.1_00055.jpg (296 KB, 1216x832)
296 KB JPG
>>
File: 00066-1492369493.png (1.25 MB, 2048x512)
1.25 MB PNG
deleted krea2. back to sdxl
>>
>>109866242
So you reposted it again, this time with a new model.
>>
File: file.png (1.13 MB, 1343x6709)
1.13 MB PNG
Why iss the mysterious trani defender trying to pass an old image as a qwen gen?
Can xe not make gens?
>>
>>109866322
That's probably better for pixel art because SDXL is more flexible in terms of artists and art styles.
>>
>>109866329
It's not a gen, it's a photo of catjack
>>
>>109866340
i wasn't talking to you, raped retard
no one uses your useless dogshit wrapper btw
>>
>trying to do something simple with krea
>just a women kneeling between a shota's legs and licking his boots
>krea2 cant even do it

Ok, i am hoping its just my shitty prompt. Whats so bad about it?

**Anime style digital illustration** — a vibrant, photorealistic anime-style digital illustration featuring a guild hall teeming with adventurers. The scene is alive with intricate details and dynamic lighting that highlights every textured surface: polished wooden floors reflecting ambient glow, ornate metal fixtures gleaming under flickering torchlight, and hanging banners fluttering in the breeze. A woman kneels between a boy’s legs, her red pixie-cut hair framing her face as she licks his boots with deliberate intensity; her brown skin glows under warm light, her muscular abdomen tensed beneath skimpy bikini armor that clings to every curve. Her eyes are red with slit pupils, fixed on the boy, while her breasts, each the size of her head, rest naturally against her torso. Her hips are wide, thighs thick and toned, emphasizing her powerful physique. The boy lounges on a high-backed chair, smugly smirking, one hand resting casually on his thigh, the other holding a cup of steaming tea. The room pulses with activity: adventurers in varied armor move behind them, weapons drawn or idle, all caught mid-action — some arguing, others laughing — their expressions animated, their movements fluid. The composition centers on the pair, with shallow depth of field blurring the background into a riot of motion and color. Lighting cascades from above and below, casting sharp shadows and highlighting skin textures, armor plating, and fabric folds with hyperreal precision. The mood is intense, sensual, and charged with energy — a moment frozen in the heat of camaraderie and dominance.
>>
File: the photo.jpg (493 KB, 1536x1152)
493 KB JPG
>>109866306
>>
>>109866348
this is too ai-slopped writing. try to separate your character descriptions into separate paragraphs, then describe exactly which character is interacting with who and what.

the LLM-style writing is too verbose, so it ends up with too much and too little to focus on.
>>
>>109866334
yep. krea2 gives me tumblr pixel art.
>>
File: space.jpg (956 KB, 1920x1320)
956 KB JPG
>>
>>109866328
Krea2 censors certain words I would like a non text refusal lora for that, also I never did a soft shelled helmet before.
Qwen is way less safety slopped and smaller I do think it would be the best bet for a finetune even with it's flaws and comparing it to the non turbo krea2 model it has it beat imo, unless krea actually releases the real base model used for turbo
>>
>>109866357
It's that and even if SDXL might melt up details, it doesn't matter when you quantize them.
>>
>>109866359
looks like his shit hose broke loose
>>
>>109866368
Adults are posting images now, kid.
>>
File: beaner.jpg (408 KB, 1536x1152)
408 KB JPG
>>
>>109866375
hey brother, if you want to post pics of India's astronauts, have fun. I'm not here to stop you.
>>
why does he keep pretending to be different people when every janny sweep just proves its one person arguing with himself?
>>
File: Qwen_image_2.1_00031.png (1.97 MB, 1714x1227)
1.97 MB PNG
>>109866179
>Added hag, Indian to negative prompt. 5.0 cfg
Improved. Nice
A fine-tune to remove GPT image texture will probably save it
>>
File: shit-hose.jpg (435 KB, 1536x1152)
435 KB JPG
>>109866368
>>
>>109866392
I'm not from India, I am actually from Pakistan.
>>
>>109866023
It's just gatekeeping their best models for API. Qwen Image 3 wasn't bad, API only. Same with every other major breakthrough they made. They're now a meme company.
>>
File: 1777212914776044.png (1.27 MB, 1071x1061)
1.27 MB PNG
I don't know how to describe a trailcam image to krea to make it stop looking so high quality
>>
>fine tunes will fix it!!!
AHAHAHAHAHAHAHAHAHA
>>
File: 1770908212196318.jpg (700 KB, 2416x1697)
700 KB JPG
so how does qwen 2.1 editing compares to Klein?
>>
>>109866427
Have you tried feeding a real trailcam pic to an llm and have it describe?
>>
>>109866427
That's a cool image.
>>
>>109866427
That's what trailcams look like, though...
>>
>>109864856
newfag here, sorry for the newfag questions
can i create porn with an old 8 GB VRAM card?
ideally in a photorealistic style
adding reference images, editing images, etc., is it possible or would i need to spend $10k on a new card?
>>
4090 for $950 .. good deal?
>>
File: 1768695556833621.jpg (20 KB, 680x459)
20 KB JPG
>>109866405
hey man. have you been abducted, and are you trying to convey something to us, using the same theme?
>>
>>109866456
Even if you had to give up your mouth and bussy I would still say it's too good to be true
>>
>>109866427
>to make it stop looking so high quality
it can't make "low quality" images because of how its trained...
>>
File: Return_00405_.jpg (1.18 MB, 1776x2368)
1.18 MB JPG
Ok 'm looking at this qwen default WF and I have to ask
Why is the seed locked and obscured from the user?
What the fuck is comfy doing?
Why would you do that to a non turbo model?
Testing higher text encode resolutions now too
>>
>>109866454
oh and i should add that i'm mostly interested in generating images, not videos
do the models come pre-censored or?
>>
>>109866493
Thank you for this avatarfagging.
>>
>>109866493
To add it's fixed to seed 0 and doesn't change
>>109866497
Get your terms correct wino
>>
>>109866493
Way to fud any alternatives then complain you are stuck with pootorch garbage wrappers
>>
File: Qwen_image_2.1_00009.jpg (304 KB, 1376x1376)
304 KB JPG
>>
how do qwen draw bagene and dock, good sirs?
>>
>>
>>109866462
I think I may need help
>>
>>109866501
Thanks for replying.
>>
>>109866348
I mean yeah, Krea 2 literally cannot even do facial expressions without one of the censor bypass things, you need to use on of those
>>
>>109866493
I recall the comfywiki guy being the one who creates at least some of the default WFs. He's a literal retard who does not know how to use comfy as you can see. :/
>>
File: qwen21.png (2.62 MB, 1120x1680)
2.62 MB PNG
Doesn't know overwatch? lol
>>
>>109866356
frankly this was my last ditch effort, the help of a llm. I just cant make it happen. I am thinking its censorship or something.
>>
File: 1770089942787214.png (2.53 MB, 1120x1680)
2.53 MB PNG
anyway, maybe next year
>>
File: Drive.png (2.75 MB, 1800x1200)
2.75 MB PNG
>>
>>109866557
see:
>>109866545
>>
>>109866569
yeah just saw this comment. Guess i will have a look on those filter bypass things. I thought abliterated would be enough.
>>
>>109866577
you actively SHOULDN'T use anything other than the stock text encoder, it will never under any circumstances fix anything and will likely give you worse outputs, the censorship has nothing to do with that, that's not how it works
>>
>>109866577
>abliterated
literally retard traps dont use them
>>
>>109866593
hear it from the guy who made the Heretic abliteration method himself if you don't believe me BTW:
https://www.reddit.com/r/StableDiffusion/comments/1vmdxzk/psa_im_the_creator_of_heretic_and_i_advise_you_to/
>>
>>109866577
>I thought abliterated would be enough.
the redditards got another one
>>
qwen image 2.1
>hyperdistilled overbaked version of gpt image 1.0
>shit realism
>slopped colors
>unstable anatomy
>slopped plastic skin
>shit speed
>no unique things about the model beyond (shit) transparency genning/editing
How many of these slopped models are Qwen gonna release that all look the same before they realize maybe training on primarily dogshit synthetic slop data is not the way forward?
>>
File: file.png (167 KB, 1180x206)
167 KB PNG
>>109866624
If only they read the Z paper...
>>
>>109866637
>>109866624
Have any of you tried actually testing it, I think the default comfy wf is the problem, it's better than krea2 raw by a wide margin.
>>
>>109866493
saved...will chatbot it later with some of the others
>>109866497
"avatar"
>>
>>109866650
What I wrote I wrote after testing it, it's shit. It's literally better to just train a gpt 1.0 lora on top of krea2 or something than to use this model.
>>
File: rocking-it.jpg (420 KB, 1536x1152)
420 KB JPG
>>
>I know a lot of people are not super concerned about the licenses, and I get that. Most people have never been in a deposition being grilled for hours by a team of lawyers going over what the word "is" means in a document. I have. Most people have never been sued in a contract dispute. I have. Most people have not had to pay lawyers hundreds of thousands of dollars to defend them in court on a frivolous lawsuit that you will win, but had to fight anyway, and the legal fees alone bankrupted your company. I have. More than once. I have never lost a court case, but I have been invited to them many times. And when you are invited to one, you want to be 100% sure and confident that you are in the right. Trust me, the licenses DO matter.
Ostris what do you mean ?
https://twitter.com/ostrisai/status/2101714386356511174
>>
>>109866680
I'm no longer concerned for the license when I saw how shit the model is
>>
File: peace-plaza.jpg (827 KB, 2048x1536)
827 KB JPG
>>
File: Qwen_image_2.1_1_74.jpg (723 KB, 1184x1776)
723 KB JPG
>>
>>109866656
Samefag homo.
>>
>>109866563
my beautiful bride to be
>>
File: 1768575216584969.png (1.28 MB, 1382x1058)
1.28 MB PNG
>>109866452
Idk it does this weird thing where some sections are more high quality than others and it just looks like a photoshop.
>>
File: paris.jpg (813 KB, 1800x1200)
813 KB JPG
>>
>>109866650
why would anyone care about it being better than Raw though, they explicitly say Raw is only for full finetuning
>>
>>109866517
im aroused
>>
File: AnimaTurbo_Output_12515.png (2.5 MB, 1248x1872)
2.5 MB PNG
>>
is ideogram4 kill
>>
>still falling for the "raw is only for training" warning
Please do not come here again retard
>>
File: joe-smurfden.jpg (1.08 MB, 2048x2048)
1.08 MB JPG
>>
>>109866804
it's exactly as ass as they say it is for inference tho
>>
>>109866624
it's a light-weight ('compact and efficient') edit model mainly for graphic design and typography assets, with an emphasis on instruction following for multiple image sources. it looks like it's supposed to be a more functional all-in-one for stuff like background removal. basically just a speedy webslop tool.
>>
>>109866799
it's not super popular at least
>>
>>109866799
the whole ideogram line is dead, it's successor only stands to be more pozzed than 4
>>
>expected input with shape [*4096], but got input of size[1, 332, 3584]
Cool default workflow, Comfy...
>>
>>109866817
I can spot your skill issue from three boards away
>>
File: the rower jower.jpg (629 KB, 1664x1280)
629 KB JPG
>>
>>109866799
who
>>
File: sped-up-joe.jpg (542 KB, 1664x1280)
542 KB JPG
SLOW DOWN, JACK!
>>
>>109866760
one of the specific things i HATE about AI, in fact probably the only thing really, is that legit /x/ content is forever fucked
>>
File: AnimaTurbo_Output_353626.png (2.54 MB, 1248x1872)
2.54 MB PNG
>>
>>109866729
uh oh....someone is posting their coocoo and exposing they have boogeymen living rent free in their head. lol, lmao sit down kid and take your meds you schitzo

>>109866493
not me
>>109866656
me
>>
File: browsin.jpg (487 KB, 1664x1280)
487 KB JPG
>>
>>109866851
the model like five people used because it speaks json and requires manual effort. results were good, but it's a lot of frontload for something that's a lot less useful than photoshop + indesign.
>>
>>109866868
krea2 could never
>>
File: ComfyUI_1299.png (2.71 MB, 1248x1856)
2.71 MB PNG
>>109866868
>>
>>109866462
>>
File: QwenImageEdit2dot1_0013.png (3.5 MB, 1440x1376)
3.5 MB PNG
>>
>>
>>109866843
Please enlighten us then retard kun
I have a feeling you're some bitter retard from /sdg/
>>
>>
File: QwenImageEdit2dot1_0029.png (1.28 MB, 704x1120)
1.28 MB PNG
>>
File: 1773606241428303.png (560 KB, 794x1001)
560 KB PNG
>>109866863
I actually don't think I have seen any ai cryptid pics. Even on /x/. AI stuff in the wild is the overbaked slop youtube thumbnails for gameplay videos with 200 views, the lady yelling at wacky cat doorbell camera videos, or 1girl catfishes on twitter.

I'd like to see more cool /x/ stuff with AI honestly. Open this catbox btw it's really cool:
https://files.catbox.moe/wl7o61.png
>>
>>109866913
kek
>>
>>109866975
i want to eat this pepe for some reason
>>
sulfur 3 status finally?
>>
>>109865411
that's because you're a newfag
>>
File: AnimaTurbo_Output_1241151.png (2.78 MB, 1872x1248)
2.78 MB PNG
>>
File: 00003-2735890242.jpg (836 KB, 3072x1920)
836 KB JPG
>>109866023
they only gatekeep this goodies for the api revenue and give local ai cucks the trash experiments and failed bakes but rebrand it as "open source" to save face to their investors.
>>
>>
File: AnimaTurbo_Output_121515.png (1.87 MB, 1872x1248)
1.87 MB PNG
>>
>>109867101
nice, pretty crisp 3DPD style
>>
>>109866073
the only personal use case for gpt image 2 for me was making synthetic images for certain characters for lora training on krea2 who have very little original media appearances. Annoying part was that i had to spend a lot of credits by to crank the settings to high thinking and very high resolution to get a decent, useful and high quality synthetic image from gpt image 2. The qwen team most likely used shitty low res 1mp gpt image 2 and nanobanana 2 lite images for the majority of their dataset.
>>
It's funny that the Chinese love releasing shit on Mondays. Why do they do that?
>>
File: anons-edit.jpg (1.18 MB, 3072x1920)
1.18 MB JPG
>>109867101

>>109867045
It's made out of Marzipan
>>
is it me or gwen 2.1 edit creates a weird grid pattern on the texture?
zoom in
>>
>>109867151
they want to do it when the market is open hopefully to crash US tech shares
>>
File: POG2.jpg (52 KB, 384x416)
52 KB JPG
>>
File: QwenImageEdit2dot1_0051.png (2.47 MB, 1600x1216)
2.47 MB PNG
>>
File: 00005-659833427.jpg (556 KB, 2880x1536)
556 KB JPG
>>109867122
its soo fucking at 3dcg styles and learning very well from lora training in replicate rendering styles.
>>
File: QwenImageEdit2dot1_0052.png (2.14 MB, 1024x1376)
2.14 MB PNG
>>
>>109867181
it do yes, VAE still poop
>>
>>109867199
kek
>>
>>109867199
asian women are consistently nicer to me for appreciating being smart than white women
>>
>>109867242
oh, I totally agree, POC women tend to have less of a distorted view of reality - issa funny meme tho!
>>
>>109867242
that's more of a cultural thing. they are taught to always be respectful. being passive aggressive is more their style
>>
File: cursed.jpg (1.06 MB, 2880x1536)
1.06 MB JPG
>>109867202
>>
>>109867251
I was being less crude but to be more direct all my girlfriends have been asian that I have had sex with despite being just as open to white women on apps
>>
File: QwenImageEdit2dot1_0065.png (1.89 MB, 1472x1472)
1.89 MB PNG
>>
File: debo_hd_k2_00013_.png (2.35 MB, 1664x1069)
2.35 MB PNG
>>109867205
I'm quite enjoying this series
>>
>>109867311
hands look burned off
>>
File: QwenImageEdit2dot1_0066.png (1.7 MB, 1472x1472)
1.7 MB PNG
>>109867311
ty fren
>>
>>109867254
this is what demons look like
>>
>>109867311
>>
>>109867255
foreign asian women? because they are obsessed with white men. american asian women are no different than regular white women
>>
File: QwenImageEdit2dot1_0086.png (885 KB, 1024x544)
885 KB PNG
>>
>>109867254
this guy has been gen'ing for probably several years now and still makes the most dogshit gens possible.
>>
>>
File: Qwen_image_2.1_00060.jpg (1.34 MB, 1536x2048)
1.34 MB JPG
I'm going to see how much it improves with a second high res pass but I see potential
>>
>>109867351
fun fact, I'm not that guy, I ran his gen through Qwen 2.1 edit :)
>>
>>
>>109867211
Can you change to another VAE?
>>
>>
>>
File: QwenImageEdit2dot1_0092.png (1.01 MB, 896x1184)
1.01 MB PNG
>>
>>
File: Qwen_image_2.1_00125.png (2.07 MB, 928x1376)
2.07 MB PNG
>>
File: QwenT2I_0013.png (2.18 MB, 1248x1248)
2.18 MB PNG
>>
File: bite-of-joe.jpg (463 KB, 1440x1440)
463 KB JPG
>>
>>109867432
>we have gpt lite at home
>>
>>
qwen 2.1 is very crusty, any anons have recommendations on how to resolve this?
>>
>>109867471
beautiful work anon, can we get this with sound (as a catbox)?
>>
File: 00416-2237860494re.png (2.38 MB, 1152x1472)
2.38 MB PNG
>another low-param image model
What is the use case?
>>
>>109867510
being actually trainable
>>
File: ComfyUI_08648.jpg (1.86 MB, 1696x2528)
1.86 MB JPG
>>109867472
Linear Quadratic cleans it up some and doesn't make it brutally slower like other Schedulers. This 2mp image took 77s at bf16 on a 4090.
>>
>>109867373
not really
>>
>>109867515
actually trainable as oppose to what? I feel like people who claim XYZ recent model were all using AI Toolkit which is consistently the worst software perhaps ever written
>>
>>109867541
>>109867541
>>109867541
>>
>>109867478
i dont do sound because 4chon doesn't allow it so i just disable those nodes
>>
>>109867515
So you're saying it's DIY beta software and the user is expected to finish it himself?
>>
>>109867535
no one's gonna do anything with some quadrillion parameter model, once the novelty wears off it's dead, small models actually have the chance of getting community development and not being forgotten after a month
>>
>>109867472
When in doubt
Add Indian, old hags to the negative prompt
Works every time for me.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.