[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Discussion and Development of Local Image, Video, and Music Models

Previous: >>109850857

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP
Neural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/neo_collage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
blessed thread of frenship
>>
thread is maintained at low quality, always. I really hate how anons can't evict one schizo that shits this place up all the time and shits on anistudio dev and news poster. 4chan is cuckshit nowadays
>>
no one is forcing you to stay sweatie
>>
I Love LDG
>>
File: Donuts.mp4 (2.49 MB, 1376x768)
2.49 MB
2.49 MB MP4
>>109861543
How would you even evict someone on an anonymous image board?
>>
>>109861610
Also you need to figure out the notion on these because the 2.5d is too uncanny. I like the aesthetic
>>
https://x.com/realrebelai/status/2101515372490854573

was qwen image 2.1 trained with those garbage chatgpt outputs? it's garbage
>>
>>109861621
I'm fairly sure he already had his fair share of temp bans. Doesn't seem to work.

>>109861623
Fighting 2.5D first in image generation, then in video generation is such an uphill battle, that I chose to just live with it.
>>
i'll just run sdxl
>>
File: ComfyUI_08627.png (3.55 MB, 1200x1800)
3.55 MB PNG
>>109861659
It's weird the Krea2 missed the wet fur and rain, but did better on the structure over the pumps than the other two.
>>
>>109861523
Paid models in CivitAI deserve a special place in hell. CivitAI pirate when ???
>>
do you guys rename your loras so they're easier to find or is there some kind of comfyui plugin that makes it easier to sort them when you have tons for a bunch of different models?
>>
why are all tran bakes like this?
>>
Is there any way to run wan2.1 or minimax h3 on a 9070xt and 32gb?
tried it and crashed, so either I just dont have enough ram gg or I need a different model etc
>>
File: 1778862477830495.jpg (830 KB, 1248x1824)
830 KB JPG
>>
File: ComfyUI_07246_.png (1.37 MB, 1920x1080)
1.37 MB PNG
>>
>>109862193
> 9070xt
It’s over before it even began
>>
>>109862193
i was able to run wan on 12/32
>>
>>109861999
lora manager
>>
>>109862327
my attempt just crashed the other day, so dunno
>>
>>109862362
it was year ago cumfart could broke something
and i was using reserve vram arg still do
>>
File: 905007855472132.png (2.89 MB, 1280x1856)
2.89 MB PNG
>>109862259
kawaii
>>
>>109862384
why do people put up with this fucking bullshit? comfy has been truly awful for last 2 years
>>
>>109862749
what can we do
>>
>>109862193
just press run on the default workflow its not that hard
>>
File: 1066024978723373.png (3.31 MB, 1280x1856)
3.31 MB PNG
>>
>>109862762
use one of the dozen proper uis that are not cumfart? preferably not written in poothon too
>>
>>109862825
do these uis have nodes
what are the backends for these uis
>>
Anybody else getting constant OOMs after the latest Comfy update? What did this retard fuck up this time?
>>
>>109862837
just roll back, gotta get used to this
>>
>>109862857
Doesn't this piece of shit have corporate backing? How is this allowed to happen?
>>
File: latinayoga.png (775 KB, 864x720)
775 KB PNG
>making your own coombait
Based based based based based based
>>
>>109861543
>>109861621
Shut the fuck up and stay out, we all hate you
>>
File: 1617234637622.jpg (93 KB, 900x675)
93 KB JPG
>>109862749
i disagree. comfy is better now, and the int8 format is kino. better than gguf
>>
>>109862916
>This is the guy seething everyday because he can't shill his zero user software.
We all saw you get banned for samefaging after you shat yourself over a single gen
>>
>tested out faceswapping with h3 for the first time

Lord have mercy, for I am about to lose my penis..
>>
ran meltie time
>>
>>109862909
new?
>>
>>109862948
?
>>
File: 90934210741246.png (1.98 MB, 1280x1856)
1.98 MB PNG
>>
>>109862825
>>109862946
you are so raped, Julien
>>
>>109862946
How could you run one?
>>
File: 1773038135341468.png (15 KB, 868x364)
15 KB PNG
Suggestion, we should add all of these UIs to the OP, since it looks like there's someone asking for one not written python almost every thread for some reason
>>
>>109863003
personally couldnt care less about corposhit bloat
>>
Qwen 2.1 is about to release and no post about it, /ldg/ is truly on the gutter
>>
>>109863029
well it hasnt released yet
>>
https://qwen.ai/blog?id=qwen-image-2.1
https://qwen.ai/blog?id=qwen-image-2.1
https://qwen.ai/blog?id=qwen-image-2.1
>>
You have to have something seriously wrong with you to post those things after a self dox. You have to be really mentally ill to come to a thread daily that hates you because of the laundry list of actions that include harassment of other UI devs followed by constant advertisement and shit yourself when anons post things you decided to post yourself with your avatar.
Then you have the nerve to call for bans when you have been caught ban evading for years and even gloat about catching bans even in the images posted.
>>
>>109863037
https://huggingface.co/Qwen/Qwen-Image-2.1
what is this then?
>>
it's up on comfy HF

is the vae different from previous qwen vae?
>>
>>109863066
damn i willed it into existence lol, shouldhave wished for something else
>>
File: 1788667391568859.png (91 KB, 320x180)
91 KB PNG
>please lord let qwen 2.1 run well on my 16gb 4080 and let me recreate characters perfectly without the need of loras
>>
>>109863029
Does not look like a step change in quality. Maybe it's better in instruction following than previous models.
>>
File: 1778666434638844.png (1.2 MB, 781x968)
1.2 MB PNG
>>109863083
my god i'm not sure my dick is ready for this
>>
>>109863116
That show was fucking stupid. Not funny at all.
>>
>>109863003
Some of them are probably useful, but I get the impression people don't really understand what parts of the most prominent UIs are actually python and which ones are not.

>>109863029
There's a post in this very thread calling it garbage
>>109861659
>>
>>109863066
>Image Editing
bros? is this finally it?
>>
>>109863120
but that intro though...
>>
File: 1761488902325062.png (119 KB, 829x958)
119 KB PNG
>>109863116
>ted dansen
it's over. this is sd3 all over again
>>
>>109862969
God I love big titties
>>
>>109863122
Maybe I'm biased, but I have trouble believing such a fairly small model and text encoder will perform all that well.
>>
>you need 25gb of VRAM to run this model without degradation
I expect a ton of tears ITT
>>
https://huggingface.co/Comfy-Org/Qwen-Image-2.1
>>
>write "1girl, solo" in Qwen Image 2.1
>it actually outputs anime girl
>>
I had fun with illustrious models for a month or so but now I'm starting to recognize a pattern and am unable to break it. It was a source of dopamine and now that's gone, I feel unmotivated to continue genning, it's all the same slop and the images even have small defects most of the time.
Is there any photorealistic model that can actually do photorealistic stuff and not this uncanny repetitive shit?
>>
>>109861818
When there is something remotely worth pirating
>>
File: 1767494314444169.png (24 KB, 615x238)
24 KB PNG
>>
>>109863147
Are you face-blind or trolling ? The reference image has Woody Harrelson in it, not Ted Danson
>>
File: image.jpg (2.22 MB, 2048x2048)
2.22 MB JPG
>>
File: 1773618210574918.png (38 KB, 861x214)
38 KB PNG
>>109863155
the transformer is very small
>>
>>109863161
Oh and not to mention that the faces are all copies of each other. Very difficult to make actually different looking people after you've iterated through different eye colors, hairstyles, hair colors and ethnicities.
>>
File: 1769837077518723.png (151 KB, 568x388)
151 KB PNG
>lora training is back on the menu
>>
>>109863155
It's smaller than Krea 2 you retard
>>
>>109861999
lora manager. never modify or fuck with the original files. archiving 101
>>
>>109862837
No. I haven't gotten an OOM in like a year and I update daily.

>>109862871
Retard.
>>
Remember if you have under 24gb of VRAM your opinion on this model doesn't matter, you're not actually running the real model
>>
File: 1767389238599786.png (60 KB, 262x192)
60 KB PNG
>i have more than 16gb vram huhuhu
you're being sent straight to the gulags after the gpu uprising faggot
>>
you just know this is one of those models that fries itself to ash if you try to combine more than one undertrained lora
>>
1. what sampler/scheduler for qwen 2.1?
2. is the cache node necessary? i've seen kijai add it in the PR
>>
>>109862909
I wish she was real
>>
>>109863230
nearly everyone that has 24gb vram cards bought them before the ai boom. you've got to be insane to buy them now at 5x the msrp
>>
File: 88.jpg (338 KB, 1331x1026)
338 KB JPG
>>
File: 1763698826513711.png (564 KB, 684x448)
564 KB PNG
>new sota model
>edit
>lightweight
me when i become one with the latent noise tonight
>>
qrd on qwen now that the dust has settled?
>>
>>109862837
I only ever get OOM if I remove '--disable-dynamic-vram' from my startup parameters

It's been like 4-5 months and dynamic-vram still is a fucking mess, they keep threatening removing the --disable-dynamic-vram option, but they clearly can't because it's needed.

Rattus is a great dev but this is truly a failure on his part
>>
>>109863254
Why would anyone pay even half of that? I bought my RX 9070 XT for 1000€.
>>
anyone getting a black image on their second gen? first gen works fine
>>
so is qwen image worth taking look at? wtf are you retards doing
>>
>>109863260
I'm only interested in the Edit capability, and it certainly looks way better than Flux Edit, so here's hoping it's a winner
>>
>>109863191
I hope khoya takes his time because I really can't be bothered to do them all again
>>
>>109863265
Someone must be buying them if the price keeps going up.

>RX 9070 XT
Yeah but AMD are shit for AI.
>>
>>109863276
>Yeah but AMD are shit for AI.
How so?
>Ubuntu 24.04
>Latest ROCM
I can gen stuff just fine.
>>
>>109863265
Because they want to do AI, AMD is fucking retarded, they're still not even trying to catch up

Oh wait, probably should change retarded to corrupt
>>
workflows
t2i: https://github.com/Comfy-Org/workflow_templates/blob/main/templates/image_qwen_image_2_1_t2i.json
edit: https://github.com/Comfy-Org/workflow_templates/blob/main/templates/image_qwen_image_2_1_image_edit.json
>>
>>109863284
Please elaborate how they are corrupt? I haven't followed, I just installed ROCM and things work fine.
>>
File: 1776514059331471.png (306 KB, 407x483)
306 KB PNG
>clicked update all instead of update comfy
>>
>>109863279
They will always be second class in terms of support, compatibility and performance tuning. Everything is built around NVIDIA/CUDA. If AMD gpu's were remotely comparable then everyone would be buying them just as much as nvidia gpus.
>>
File: file.png (1.38 MB, 817x1094)
1.38 MB PNG
holy gpt 2 diltill
>>
>>109863294
might as well roll back. updating python deps borks your comfy. i actually deleted that script just so i never accidentally click it
>>
>>109863295
But my GPU cost 1000€ a year ago and it has 16GB of VRAM. I don't see myself paying 10k€ for an NVIDIA GPU.
>>
>>109863297
what is this, 2022? I havent seen realism gens this bad since SDXL days.
>>
>>109863292
What things ? And at what speeds ? Practically no AI trainers offer AMD support because it's so broken and slow.
>>
>>109863147
reminds me of a less horrifying version of klaus kinski
>>
unless the nodes or officials workflows are fucked up this is unusable for anime/illustration, the patchworky "details" and that fucking grime is everywhere

realistic edit looks pretty good though, and it's fast
>>
>>109863312
I haven't tried training but comfy, automatic, sdnext just work for me. Only thing I have to make sure is that I install the AMD ROCM versions of pytorch deps and those are available as wheels.
>>
File: file.png (3.27 MB, 1248x1664)
3.27 MB PNG
is it just me or prompting anything nsfw just fucks the image up?
>>
File: 1762975452459861.png (42 KB, 489x545)
42 KB PNG
>update comfy
>see this
what?
>>
File: 1787297561379881.jpg (94 KB, 361x345)
94 KB JPG
>>109863265
>1000 euro for AMDumb
serious nigga? thank fuck I got my 5090 last year when it was 'normal' price.
>>
>qwen image 2.1
>smaller model than krea and overall worse in terms of both knowledge and prompt understanding
>gpt slopped to the max
>guidance distilled model like minimax h3
>no base model released
>lora training will be just like minimax: bunch of cope settings and modified training objectives to try to preserve the distillation but it doesn't really work
>non commercial license with no minimum revenue carve-out like krea
what is the point
>>
File: 844927411144811.png (1.8 MB, 1280x1856)
1.8 MB PNG
>>109863059
Looks really good. Should a very nice upgrade from Klein 9B.
>>
>>109863331
It does alpha channel!
>>
>>109863330
What was the normal price even? Can't remember.
>>
>>109863301
Most normal people aren't paying that either. We are weathering the storm and hoping something happens because this isn't sustainable. I still wouldn't ever waste money on an AMD gpu though.
>>
>>109863327
SD3 bros we are so back.
>>
>>109863322
Again, what speeds ? You can run AI diffusion on macs as well, but it's pointless since it's so insanely slow

I'm saying this as someone who would LOVE competition in this space, but AMD is still so far behind in performance and support, macs at least can use large LLM's somewhat performatively due to the unified memory (although anything image / video related crawls), but AMD has no such feature, it's just cheaper and MUCH slower with very little and very buggy support

If you want to use local AI efficiently right NOW, Nvidia is sadly the only viable option
>>
File: offbyone.jpg (221 KB, 1035x785)
221 KB JPG
>>109863066
pretty decent but still meh, doubt klein would do much better
>>
>>109863327
surely this is some unfinished lodestone finetune
>>
>>109863372
>doubt klein would do much better
klein is awful at characters. but it excels when you combine lora+reference.
>>
File: Krea2_turbo_hr_fix_00210_.jpg (2.81 MB, 2512x3344)
2.81 MB JPG
>>109863254
They called me crazy but look at me now
So is this model better or worse than krea?
I'm pretty sure the only thing worth it will be the edit model which imo might be a saving grace for models like Anima that fail at text
>>
>>109863331
>what is the point
That it is a better edit model than Flux Klein, are you retarded ?

That said, Krea devs said they are cooking a edit model, so there's a GOOD chance it will wipe everything else up until that point

Meanwhile I want the Minimax image model they hinted at, it's insane how much the video model knows
>>
>>109861659
As long as there's something to scratch my i2i itch. Hope it's better than klein and it better be much better.
>>
Kek, made an accidental troon moment.
>>
>>109863381
>but it excels when you combine lora+reference.
Every model does that
>>
>>109863386
I'm a huge fan of 2512 because it was very easy to train, in fact it's kinda garbage on its own, but if 2.1 hasn't got that then it's doa to me. I stuck with qwen until krea2 dethroned it
>>
>>109863404
Why did she turn into a faggy jew ? Delete this
>>
been using krea2. It isnt bad... but am i correct in assuming it isnt very good at porn? Specially, stylized porn?
>>
>>109863245
A year ago I bought a used RTX 3090 for $600. One of the ports was a bit rusty, but it works like a charm
>>
>>109863368
5 it/s for image gen on illustrious models 832x1216 resolution
>>
wtf, qwen 2.1 is terrible
>>
>>109863419
Kroma is your friend: https://huggingface.co/lodestones/Kroma

I would suggest using the extracted loras instead of the full checkpoints: https://huggingface.co/silveroxides/Kroma-LoRA/tree/main
>>
>>109863419
yes. if you want that, use anima (if it's anime). if you want 3dpd porn, i don't know what you can use desu
>>
>>109863422
As in SDXL based Illustrious ? Haven't used SDXL stuff in years.
>>
>>109863478
I have no idea. Latest Cyberrealistic illustrious v12 for example.
>>
why would you ever need more than SDXL juggernaut?
>>
the hype died down fast
>>
>>109863442
chicken korma
>>
This model seems dumber than krea2
>>
>>109863507
there was never any hype for qwen 2.1 it was clear from the moment people started posting images that it was slopmaxxed just like previous qwen models
>>
File: example-01.png (290 KB, 3333x1052)
290 KB PNG
>>109863514
But the benchmarks say it's better than Nano Banana??
>>
File: Qwen_image_2.1_00005.jpg (1.33 MB, 1776x2368)
1.33 MB JPG
>>109863507
Well it can do nudity out the box
>>
>thread got flooded with qwen shills again
See ya ldg.
>>
>>109863564
I don't see any so please never come back again
>>
>>109863276
>Someone must be buying them if the price keeps going up.
I think it's more like every time the product changes hands, the new owner is hoping they can find a bigger idiot bagholder and turn a profit. The enabler here is newegg and ebay provide no disincentive towards listing that never sell. Ebay used to punish greed this way before BIN - if you asked for delusional money, your listing expired with zero bids and you paid to list it again.
>>
File: 1773458396090070.png (273 KB, 589x422)
273 KB PNG
it's not good
>>
While not perfect this model is less safety slopped than Krea by a country mile,
I included the actual catbox to give a example for once
https://files.catbox.moe/5zeq9n.png
>>
How's the new qwen for editing compared to flux klein?
>>
>>109863611
nice cock bro, very lovecraftian
>>
File: 1789914263409748.jpg (479 KB, 1858x1066)
479 KB JPG
look at what they need to mimic a fraction of our power
>>
i think yue2 is my favorite release of lately, can other anons share some creations?
https://vocaroo.com/1m3kJE7AmPbY
>>
>>109863621
I'm sure if you describe it enough it will look better, this is base exploration and to see if the model is worth training. I think you should start small then go big. The default anime style looks fucking horrid imo
>>
Saying this model is not good is a dramatic understatement. wew. it's bad.
>>
ani said he doesnt see any potential in qwen so its probably garbage. he's probably too polite to call it what it is
>>
SOAD yue2 lora, trained it on every single one of their studio songs
kinda ok it gets Serj's voice but the riffs and guitar doesn't sound like Daron's style
https://vocaroo.com/1cOWmUYtSCIm
https://files.catbox.moe/vxg4if.safetensors
>>
Yeah, it's as slopped as Flux Klein Edit, not worth your time.

I guess it's back waiting for the Krea team to release their upcoming Edit model.

There's been a few 'also-ran' models recently, this and LTX 2.5, BFL at least avoided shame by not releasing their video model against Minimax
>>
>>109863665
Yeah nobody's even posting good gens. It's obviously trash.
>>
>my butt buddy said
>>
4mp at 100 steps takes 3 minutes on king hardware
>>
File: file.png (2.51 MB, 1500x1500)
2.51 MB PNG
Juice WRLD lora
https://vocaroo.com/11UjN9obQ5pS
https://files.catbox.moe/b5k9od.safetensors
>>
>>109863708
who are you quoting ran? can we discuss new models without you sperging out?
>>
>>109863665
hang yourself Julien
>>
is there any place you can pirate civitai stuff? Cause holy shit, there is a dude asking 16k buzz for a lora
>>
File: Krea2_turbo_04242_.jpg (1.9 MB, 1776x2368)
1.9 MB JPG
>Comfy once again giving shit tier defaults
Please crank you cfg past 1 anons
>>
>>109863722
you can also stop talking about yourself in third person, you cocksucking leaf retard
>>
>>109863729
What would I do that on a distilled model?
>>
>>109863729
why are all your gens so fucking ugly?
>>
>>109863549
is there a name for awful detailing like this in newer models? it's either weird membrane like in your pic or layer of dirt so that there's never a solid patch on on color

i don't remember retarded sdxl doing that
>>
>>109863743
yeah that rat is disgustingly repulsive
>>
>>109863743
Why are you so fucking ugly ?

That said, the rat is blocking the view of the boobs, so you are technically correct, this is a bad image.
>>
>>109863743
catjak is braindead welfare leech, please understand
genning ugly slop is the most intellectually demanding task she encountered in years
>>
>>109863738
>>109863750
https://files.catbox.moe/dhoxgr.png
I see a jump in quality
Just saying you can compare the two catboxes I'm even on the extreme end of cfg
>>
Is video gen even possible with 16GB VRAM and Rocm or is just a pipe dream right now?
>>
>>109863729
what is she looking at? can you at least post something that isnt completely slopped? or even better just fuck off already
>>
>>109863767
>catjak is braindead welfare leech
wow can't believe the developer of tranustudio is an ESL retard despite being a leaf
>>
>>109863729
Nice gen.
>>
File: Krea2_turbo_04297_.jpg (1.57 MB, 1776x2368)
1.57 MB JPG
Listen man you can seethe all you want, nothing is going to change the shit you posted that anon pointed out, now fuck off and let us test this new model while you sit in the corner.
>>
File: 1776665500018468.png (597 KB, 1144x1047)
597 KB PNG
>put the person in picture 1 on a chair in front of a white background, wearing the outfit in picture 2
the edit function barely even works.
>>
>>109863791
Dude's right though, your gens are repulsive and repetitive.
>>
>>109863799
crank you cfg up first anon 6-8
>>
>>109863799
Why would you do that
>>
>>109863807
>more talking in third person
lmao what a raped retard
Julien the ESL raped retard
>>
>>109863799
In this example that is a GOOD thing

It's never too late to give up anon
>>
Ignoring the seething fail dev I'm going to do a real test going to see how the qwen 3.8 handles the prompt writing, the model is uncensored but I don't really care about nsfw and want to make something funny
>>
>>109863818
I have no clue who or what are you talking about. You are deranged, dude.
>>
>>109863799
But it did what you asked it to do???
>>
>>109863774
Okay. I will admit that penis looks a lot more like a penis.
>>
File: image_02058.png (2.65 MB, 1920x1080)
2.65 MB PNG
I'm going to take another shot at doing a full-show LoRA of Oregairu for Krea2. I had my hermes-agent balance the character frequency distribution by purning and duplicating where needed. I want to see if this avoids having the LoRA pull gens towards characters who appear frequently in the show. There's about 4K images in the dataset, but due to repeats on lesser characters, it's more like an effective 1K.
I'll be watching my test outputs to see how close ir gets to being able to do picrel by character name only, all those characters are properly named in the captions.
>>
File: Qwen_image_2.1_00021.png (3.31 MB, 1440x1440)
3.31 MB PNG
bf16
>>
File: Qwen_image_2.1_00020.png (3.32 MB, 1440x1440)
3.32 MB PNG
>>109863847
int8_convrot
>>
>>109863847
Slopped to high heavens, actually reminds me gpt output.
>>
>>109863840
i think it's brave of you to admit that
>>
>>109863847
>>109863853
Top of the left earphone/hairband thing is slightly different.
>>
>>109863818
local diffusion?
>>
>>109863830
>I have no clue who or what are you talking about.
Yeah, that's the average person's reaction to the raped ESL retard and its delusions of fame and fortune
>>
>>109863868
It's so pathetic you try to dog whistle the mod to justify your self inflicted verbal ass beatings. Go finish your drink so you can black out already you cum guzzling leaf.
We're trying to enjoy a new model
>>
File: file.png (3.95 MB, 1440x1440)
3.95 MB PNG
>>
>>109863868
Yeah, seethe more trani
>>
>>109863404
Ranma 1/2
>>
>>109863799
where did those shoes come from?
>>
>>109863883
>>109863885
>>109863892
what are you even talking about dude
can you please stop flooding ldg with your off-topic bullshit?
>>
File: file.png (2.36 MB, 768x1376)
2.36 MB PNG
a woman holding an apple
>>
>>109863903
>ani and thus anistudio is offtopic bullshit
you said it, not me
>>
File: file.png (2.21 MB, 768x1376)
2.21 MB PNG
>>109863906
>>
>feed it 4 reference images of the same person
>every photo has 4 copies of the same person in it
is there an official prompting guide somewhere?
>>
Stop replying to him he's trying to get his dick sucking post deleted. Just link those images to newfags and they will understand.
He's already spiraling because multiple anons are shiting on him
He's going to spam reports next which is why he tries to frame his reporting reason
>>
>>109863913
NTA but AniStudio is a UI for diffusion and image genning. Why would it be off-topic?
>>
>>109863917
When I take a shit, my prompting guide is the shit smeared toilet paper I throw into the bowl after it. I assume this is the same.
>>
File: Krea2_turbo_04219_.jpg (1.77 MB, 1776x2368)
1.77 MB JPG
>>109863917
https://github.com/QwenLM/Qwen-Image-2.1/tree/main/prompt_rewrite/prompts
>>
>>109863003
No it also can't be JavaScript or typescript as well
>>
>>109863938
can you explain what is the "36 stars" meme is about? you seem to be completely obsessed with it
>>
>>109863003
Look at the OP.
>>
>>109861621
>Swept up by janny and deposited into the nearest bin as many times as needed for him to decide to fuck off forever. Jannies are blind and need a little hint from anons but it doesn't require posting off topic tripe to do so
>[Deleted]
Thanks jannies for showing us who really needed to get put in a bin :]
Seethe more Trani :]
>>
>>109863925
>NTA
raped lmao
>>
File: 451725704390795.jpg (1.82 MB, 2560x2227)
1.82 MB JPG
>>109862969
Looks like if the image is anime style and the reference is a photograph it doesn't manage to fit the reference to the anime style very well.

Left is >Change the hat in <image1> to the hat in <image2>. Keep the exact same style as in <image1>.

Right which is a bit better is >First convert <image2> to the same anime art style as <image1> then change the hat in <image1> to the hat in <image2>.

Maybe there is some other way to prompt it that would work better.
>>
>>109862193
>>109862362
That should be enough to run Minimax (int8 convrot), I think; it's bigger/stronger than my machine. Try launching with --disable-smart-memory so it doesn't try keeping more than one model in VRAM at a time.

If the VRAM is running out toward the end, you can also try setting an environment variable for COMFYUI_ENABLE_MIOPEN=1, which should enable cudnn on AMD and greatly reduce peak VRAM during VAE decode.

Other than that, you can launch with --use-ck-attention to speed up genning by up to a third.
>>
>>109863750
It's what you get when you train on GPT images.
>>
>>109863990
First of all, there were no high ranking officials in the Schutzstaffel, and secondly, i highly doubt Hitler would ever approve of such an outfit
>>
>>109861523
>Neural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel
>GTK
Rape All Gnome Fags
>>
>>109863727
what lora
usually not even checkpoints are that expensive. most i've seen was 10k from pornmaster shit
>>
do i have 90k buzz on civitai? yes. will i ever spend it? no
>>
>>109864024
frankly i dont ever remembere.it was so egregious i noped out.
>>
I guess the biggest issue is that the model's outputs are a bit small at 2.2mp
I'm worried how well it would do at a higher value for now
>>
File: Qwen_image_2.1_00015.jpg (920 KB, 1856x1248)
920 KB JPG
I'm feeling bullish on this model
>>
M-Minimax Image will sa-save us, t-trust the plan!
>>
File: 1762942665737106.png (3.64 MB, 1024x1536)
3.64 MB PNG
Yeah for imagegen the new qwen is fucking awful unless i'm missing something about its generation parameters, no matter what i feed to it, the result is some sloppified GPT tier garbage, not particularly creative either, we really got spoiled by krea. Some outputs feel like 1.5 quality wise.
This used the same prompt as this krea pic >>109862259
Gonna try editing, but i never had fun time with qwen edit to begin with.
>>
File: chat-is-this-real.mp4 (1.32 MB, 1024x1024)
1.32 MB
1.32 MB MP4
now that the dust settled down, what's the verdict?
given it's qwen, well...
>>
>>109864074
why is it so ugly and artifacted
>>
>>109864085
Did you follow the prompting guide?
>>
>>109863975
There needs to be a new class of moderation called "schizo babysitter" so they can snipe you immediately. Hell, make an agent that is trained on your coded language that bans you immediately for using it
>>
>>109864095
It might be a good swap model with something like krea or anima imo
>>
uh, krea2 is slopped to hell and back. Is there any finetune that doesnt make it dumb?

btw, is int8 too dumb?
>>
>>109863975
I am that anon and I am not banned. You just used proxies to autojanny again
>>
>>109863975
you lost hlky
>>
was qwen ever good?
>>
File: file.png (1.25 MB, 2389x1201)
1.25 MB PNG
wrong emotions, I don't think any edit model can do this though
>>
>>109864097
>>109864111
>>109864112
Having a melty I see :]
>>
>>109864126
Can you make the cfg 6
>>
>>109864117
no, i can only assume that there's a billion chinks shilling their trash
large llm are shit, small llm are mediocre, all their image gen is trash
>>
>>109864117
When there wasn't any edit models way back and it was pretty much the only option yeah but Klein came out and it's been downhill ever since
>>
>>109864110
>uh, krea2 is slopped to hell and back.
Compared to what ?
>>
File: Qwen_image_2.1_00030.png (2.27 MB, 2368x896)
2.27 MB PNG
>>109864134
>>
>>109864001
How is this further behind than anistudio despite having contributors?
>>
>>109864154
i mean the nsfw
>>
>>109863914
>No thumb

Which problem will science solve first: FTL travel or the ability to generate an image of a human hand without risk of an eldritch atrocity?
>>
>>109864158
>>109864001
>trani has another melty because he's jealous of another dev
ROFL
>>
File: Qwen_image_2.1_00018.jpg (1.45 MB, 2432x1824)
1.45 MB JPG
Personally I think this is pretty good for a base model/pass because it's not as safety slopped, it might be a good foundation for a anime finetune if people can't figure out krea2.
>>109864157
Damn, let me play around with it I guess going to use the official llm system prompt.
>>
>qwen
>total dogshit
the usual, then
>>
>>109864172
I also think using imgui is retarded but I already have my own vibecoded front end so it's whateves
>>
File: 1689894020672879.gif (504 KB, 200x241)
504 KB GIF
3 minutes for a 720p image in krea 2 turbo
>>
>>109863399
i don't know what it's better than or worse than but it's shit
>>
>Alibaba fires original Qwen team
>Brings in Google jew to take over
>Worst qwen models ever
What did they mean by this ?
>>
File: HSqql0lXcAApYP5.jpg (406 KB, 3600x1440)
406 KB JPG
chat is this real
>>
>>109864172
Why didn't you put anistudio in the OP?
>>
>>109864194
and then the vae ruins it all
>>
File: cruelty.jpg (43 KB, 850x319)
43 KB JPG
>>109864174
prompt:

Use image 1 as the base image.
For each emotion label written in image 1, replace the illustrated emotion/face with the face from image 2, modifying the facial expression to accurately match the written emotion text (e.g., happy, sad, angry, surprised, etc.).

Keep the same person’s identity from image 2 for all replacements.

Ensure:

The facial expressions clearly correspond to the emotion text in image 1

Face orientation, lighting, skin tone, and proportions look natural and consistent

The original layout, text, background, and positioning from image 1 remain unchanged

Only the emotion visuals are replaced, not the text itself

The final image should look cohesive and realistic, as if the same person is expressing all listed emotions.
>>
>>109864207
klein 9b is better than this
>>
>>109864207
Benchmarks are just grifter head canon
>>
how is qwen 2.1 for image editing?
>>
>>109864174
nobody asked the opinion of a nocode faggot who contributed nothing to open source and ldg except for ugly dogshit gens
>>
>>109864207
>longcat
>boogu
>kling
ahhh these take me way back...
>>
>>109864228
's alright
>>
File: Qwen_image_2.1_00028.png (1.16 MB, 1216x832)
1.16 MB PNG
Eh, it's ok as a draft image editor

Dimly lit, atmospheric interior shot centered on a man in <picture 1> seated at a bar counter made of dark wood. His look is depressive/defeated as he brings a lit cigarette to his lips with his right hand, from which a plume of grey smoke drifts up into the upper part of the frame. A ring is visible on one finger of that hand.

He wears a deep red zip-up racing-style jacket covered in sponsor markings. His left hand rests flat on the wooden surface beside a short glass tumbler holding whiskey. To his right on the table stands a dark brown beer bottle with its neck visible, next to an ashtray at the far edge.

The background is softly out of focus and shadowy, suggesting a bar: shelves lined with bottles glow faintly on the left, and several other patrons are seated as blurred silhouettes (one in dark clothing facing away on the left, others toward the right). Warm overhead lights produce soft circular bokeh highlights across the scene. The overall lighting is low-key and cinematic, casting much of the image in shadow while illuminating the man's face, jacket, hand, and smoke, creating a moody, introspective tone.
>>
>>109864207
I'm sick of scores it's clear at this point it's not at all an indication of actual model quality.
>>
File: Qwen_image_2.1_00032.png (1.74 MB, 800x1216)
1.74 MB PNG
>>109864217
>>
>>109864261
Impressive, but let's see you change Steven Seagal's expression
>>
>>109864261
>>109864217
I'm testing a style transfer then I'll look into this
>>
>>109863846
The zero-lora krea2 preview. Man... base krea2... "you tried" award.
>>
>>109864288
still better than anima slop
>>
>give it a klein 9 gen as a reference
>spits out an image that looks worse
i give up. what a pointless model.
>>
>>109864230
nice self-description
>>
>>109864299
Anything's better than anima.
>>
>prompt h3 for loitering leering creeps
>exclusively generates indian men
i think we have to give it some credit
>>
File: 0_00010.png (2.8 MB, 1280x1856)
2.8 MB PNG
It works, and I'd say it's an upgrade over Klein, so that's good, but it's not amazing

>Change her color scheme to black and yellow. Change her hair color to pink gold.
>>
File: 812115210686280.png (1.9 MB, 1280x1856)
1.9 MB PNG
>>109864327
>>
>>109864327
But Krea identity edit could already do that...
>>
File: Qwen_image_2.1_00041.png (444 KB, 736x608)
444 KB PNG
remove.bg has been achieved internally
>>
>>109864341
>>109864327
its a very literal edit model. you use it to edit images. but it's very bad at creating new images using the subjects in the references, which is what i think is the most important part.
>>
>>109863337
$3k. Started around $2.2k
>>
>>109863331
>what is the point
Making consistent character datasets for Krea 2 loras it looks like.
>>
>>109864425
I remember anons trying to shit on anons who got the AiB models that were close to 3k.
Who's laughing now when the door has been closed?
WHO'S FUCKING LAUGHING NOW
>>
File: f2k.jpg (434 KB, 1918x1218)
434 KB JPG
>>109864217
this was the best of about a dozen rng attempts with f2k9b, really hit and miss and the likeness isn't there imo, great test idea
>>
>>109863419
Does porn just fine for me, even cunny. You have to use a nsfw lora and the filter bypass. I guess it doesn't do any super crazy positions you'd see on danbooru but it's most likely because I didn't prompt it correctly.
>>
>>109864464
>filter bypass
got the link?
>>
>>109863729
Oh okay qwen isn't so ba—
>filename
Oh never mind. Guess I should go back to lurking.
>>
>>109864510
get your eyes checked while at it
>>
>>109864310
Shit in, shit out

What did you expect ?
>>
>>109864457
turned her into Kathy Griffin lmao
>>
>>109864530
>What did you expect ?
shit in, good out
>>
>>109863846
Do you really need that many pictures for Krea 2? In my experience I rarely go over 30 to learn the style, it just has to train for 2000+ steps for anime. As long as everything is tagged it doesn't overfit but yes you still need an even distribution.
>>
>>109863419
Try a small decensor lora/node like Fedor Bypass, if you haven't yet. That helps with porn at the cost of some prompt bleed. Not sure about stylized porn though, since it isn't trained on artist tags.

>>109863774
Well that makes a huge difference.

>>109863778
I have it working on my iGPU set to 16GB, albeit slowly and small. See >>109863995
>>
File: Qwen_image_2.1_00006.png (1.98 MB, 1024x1024)
1.98 MB PNG
>>
>>109864448
AIB models are still overpriced (and inferior).
FE will never fail and you will never need to repaste it. And if it does, Nvidia is guaranteed to handle warranty, unlike the Chinese scam customer support from MSI and aSUS.
>>
File: Qwen_image_2.1_00008.png (1.75 MB, 1024x1024)
1.75 MB PNG
holy shit its amazing
>>
>>109864543
Many anons can't take the 5 minutes to adjust values and always take base settings as a source of truth, it's kind of sad desu
>>
File: qwen-image-2.1-00035-.png (2.17 MB, 1664x2496)
2.17 MB PNG
default settings must be shit
>>
>>109864580
What are your settings
>>
>>109864587
I'm guessing he upped the cfg similar to what was discussed in the blowjob catboxes
>>
File: Qwen_image_2.1_00012.png (1.96 MB, 1024x1024)
1.96 MB PNG
>>
>>109864587
25 steps, cfg 3, euler/simple
>>
>>109864472
https://files.catbox.moe/3ajoxt.safetensors
It's the one from civit. It's not strong enough on its own and you still need a nsfw lora, even for sfw gens but together you get the least censorship. If you raise the weight above 1 it works on its own but at the cost of quality hence the need for another lora.
>>
How long will a RTX 5090 last?
How long will a RTX Pro 6000 last?
How long until you need even more VRAM to run better models?
Is the only way to better models more parameters?
>>
>>109864606
nsfw lora uh? Is that the top downloaded one? snops or something
>>
>>109864621
>5090 stops getting updates in September 2027
>6000 stops getting updates in January 2028
>48GB floor will hit in Early 2027
>No shit
>>
>>109864638
Where do you come up with retarded shit like this?
>>
File: 1789923102511.jpg (161 KB, 1544x796)
161 KB JPG
Yeah as expected, Qwen is CCP cucked, it does some basic nsfw (Artistic nudity, haven't tried more) and surprisingly, it accepted Taiwan, but some stuff is a bit too much for it it seems
Gonna wait for the censorship bypass to continue
Local btw, checking if the model or the TE is cucked
>>
I can't comprehend how it is this shit. Where did the ZTurbo team inside Alibaba go? How did they allow this to happen?
>>
>>109864457
>>109864261
>>109864217
Worked out pretty well for me, I would just need to refine the prompt and negs did several gens and it does what I'm looking for, I could keep it more consistent with negs
>>
>Huihui-Qwen3-VL-4B-Instruct-abliterated

just saw this on a workflow i got. Is this snake oil or does it help uncensor krea2?
>>
>>109864710
that cruelty part was trained on images of Jinx
>>
>>109864714
yes, the default encoder doesn't allow gaping
>>
>>109864628
The one I use is this one. I can never find it on civit anymore.
https://huggingface.co/Kutches/Kr3a/blob/main/Krea2_HMNSFW_AIO.safetensors
I should probably make my own because it's a little too trigger happy with pov shots but it works really well.
>>
>>109864710
Still looks like shit like everything else you post
>>
File: pg006.jpg (212 KB, 800x1200)
212 KB JPG
>>109864710
now this one
>>
>>109864621
Models have plateaued and ngrams are here. The world is plunging into an energy crisis so they will need more efficiency for the rest of this decade. Luckily China is leading the charge on that and Trump told Sama and Dario to get back to work. Things will get better on all fronts.
>>
>>109864742
I want to try one more thing and I'll try that after
>>
>>109864714
I have never seen it do better on an image quality basis for things in general over the default TE which makes sense. You are able to generate some things the base TE can't output where Krea 2 is the worst at but most sane diffusion models use the base model which makes that not an issue. You'll need probably some sort of inpainting or something to get it up to par.
>>
>>109864710
Why do you evade bans?
>>
>>109864157
>anger
Kek, accurate.

>>109864261
Cool stress-test, although it's fumbling hard.

>>109864194
I get about 2 minutes on a laptop iGPU. My condolences.
>>
File: Qwen_image_2.1_1.jpg (398 KB, 1440x1440)
398 KB JPG
>>
File: demo.jpg (340 KB, 1243x1452)
340 KB JPG
final got the demo to gen, enhance prompt turned on
https://huggingface.co/spaces/Qwen/Qwen-Image-2.1
>>
File: Qwen_image_2.1_1_1.jpg (317 KB, 1184x1776)
317 KB JPG
>>
What was the deal with the schizo calling the model shit?
Anons are already making the model work and it's only been a few hours.
>>
>>109864856
>>109864856
>>109864856
calm and reasonable new /ldg/
>>
File: the rat.jpg (801 KB, 2304x1792)
801 KB JPG
>>
Step 3500. Much better than what I was getting out of my first attempt. Flattening the character distribution and manually correcting captions made a big difference.
>>
>>109865961
no, dont eat the ojousama drills!
>>
>>109864570
well to be fair it's not intuitive

>>109863995
thanks ill give it a try, didnt want to force my pc after that so I didnt even try again, I did deleted the models since I assumed it would be impossible with amd gpu and only 32gb, ill give it a try another day, ill look the options in the templates though, in the meantime
>>
File: 00002-4122821924.jpg (607 KB, 3072x1856)
607 KB JPG
>>109864110
>krea 2 slopped
vramlet seething or shitposting?



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.