[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Discussion and Development of Local Image, Video, and Music Models

Previous: >>109318601

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Z
https://huggingface.co/Tongyi-MAI/Z-Image

>Qwen
https://huggingface.co/collections/Qwen/qwen-image

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>LTX-2.3
https://huggingface.co/collections/Lightricks/ltx-23

>Wan
https://github.com/Wan-Video/Wan2.2

>Chroma
https://huggingface.co/lodestones/Chroma1-Base
https://rentry.org/mvu52t46

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
>mfw Resource news

07/19/2026

>Kura: Experiment workspace for LoRA training AI agents and image-comparison runs
https://github.com/nomadoor/Kura

>China bans AI “boyfriends” and “girlfriends” over addiction and birth rate concerns
https://www.dexerto.com/entertainment/china-bans-ai-boyfriends-and-girlfriends-over-addiction-and-birth-rate-concerns-3388737

>Local AI Toolkit: Curated Collection of Locally-Runnable AI Models
https://huggingface.co/m15dg/local-ai-toolkit

>Krea 2 ReID Reference
https://huggingface.co/yijunwang2/krea2-reid

07/17/2026

>Rare Concept Generation via Counterfactual Inference in Diffusion Models
https://github.com/200204jzy/CI-Diff

>Uni-AdaVD: Universal Concept Erasure for Visual Generation via Orthogonal Value Decomposition
https://github.com/QifanZhou/Uni-AdaVD

>MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators
https://harahan.github.io/meanflownft-project-page

>TanGO: Training-Free 3D Editing via Tangent-Space Guidance and Optimization
https://github.com/siw00-lim/TanGO

>GlobalForge: Towards Robust AI-Generated Image Detection
https://anonymous.4open.science/r/GlobalForge-BE0F

>Knowing You at First Glance: Inferring Apparent Personality from Faces
https://github.com/MrHuan3/GlanceFace

>GNM: Generative aNthropometric Model and Ecosystem
https://github.com/google/GNM

07/16/2026

>Reflecting Process Expertise in Procedural Material Generation
https://materialapprentice.github.io

>PromptForge LD: Shot-writer for LTX Video in ComfyUI using llama + lm studio + ollama
https://github.com/Brojakhoeman/Prompt-Forge-LD

>Qwen3-VL-4B-Instruct Heretic (ComfyUI)
https://huggingface.co/DreamFast/Qwen3-VL-4b-Heretic-ComfyUI

>The AI Backlash Has Tech Executives Fearing for Their Lives
https://www.wsj.com/us-news/the-ai-backlash-has-tech-executives-fearing-for-their-lives-30c43972
>>
Anyone have experience making videos from images?
Asked this last thread, but by the time it makes it to the sampler, is the GPU supposed to be basically idle before the preview appears?
Any way to speed up that idle to active process?
>>
File: 1768171487596177.png (3 MB, 1760x1176)
3 MB PNG
hmmm... actually not sure who'd win this one.

I mean obviously Kiara would win - the first time.
>>
anons you have no idea how it feels to succeed at life when 80% of you're gens are in the collage
>>
File: ComfyUI_01896_.png (3.89 MB, 1920x1200)
3.89 MB PNG
>>
>>109320475
give your hardware a pat on the back, and then kys
>>
>>109320475
it's 1 person, the same person, picking the images. you act as if /ldg/ as a community voted and choose your images. I can promise you, every single image from that would get 0 interaction on twitter or even civitai.
>>
its clearly more than a single anon doing the faggollages desu
>>
>>109320475
I'm not the usual baker btw, he's sometimes asleep at this time and "someone" flamed the chubby anime girl genner so I thought I should bake.
>>109320492
correct
>>
>>109320463
good to poop
vram>system ram>poop(page) file
you want your diffusion model to live in your vram. it should load from an SSD into your vram, the model initializes and start iterating, all is well.
if you don't have enough vram to hold the entire model in vram it gets bumped to system ram, then things get slow, if you get pushed into the page file, you are fucked.
you can unload your text encoder so it doesn't shit up your vram, and you can run gguf models to lower their size.
worst case, the model you use shouldn't be more than 20-30% larger than you vram.
aside from that, initialization/the first iteration can take awhile, but it shouldn't be crazy long, like several minutes.
>>
File: corearchitecture.jpg (2.18 MB, 3144x1868)
2.18 MB JPG
>>
no local video model is good for anime. they all generate rotoscoped linear overly smooth shit. good animation relies on keyframes to make the animation pop. all ai anime animation feels like it was made in flash using the tween tool. its so bad
>>
>>109320499
Alright some great advice, first time I’ve seen anything about the pagefile.
Got a question though, I’ve been testing this out for hours, but for some reason the overall process is 3x slower than it was yesterday, even after turning my unit off and on again.
Might be related, but even my games felt slower than usual, but that was fixed after a restart, so I hope my equipment isn’t dying this fast.
What’s the closest thing to a “restart” or “clean state” when it comes to the wanvideo model (I used Kijai’s).
Preferably, I don’t wanna reinstall comfyui.
>>
File: 1759422911614413.jpg (2.53 MB, 2256x1504)
2.53 MB JPG
Oh, and by the way... Kiara would TEAM UP.
>>
>>109320406
Thanks anon, my Anima-Krea 2 refinements have some life to them thanks to my color transfer node.
>>
File: 280375803579250.png (78 KB, 830x688)
78 KB PNG
>>109320560
if you are on windows you can download rammap, it's a microsoft app that shows you everything that is using your ram, it helps troubleshoot.
you can set your text encoder to use system ram in comfy, so do that.
best bet is download a gguf, just google wan ggufs and test small to big until you find a good balance of performance and quality.
and loras hangout in your ram too, so if you are stacking a bunch of them they can push you over the edge.
another thing, you can add your models into a subgraph to retain the connections. it lets you switch around nodes without having to rewire your workflow. makes fucking around and testing things way less annoying.
>>
Can Krea 2 do style mixing like you could do on SDXL?
>>
>gemini doxxed me in my prompt for no reason

t-thanks
>>
>>109320624
Naming different artists? I don't see why not. The model knows more styles than SDXL out of the box.
>>
File: 1783939318200263.jpg (436 KB, 1200x864)
436 KB JPG
>>109320606
last time i wasted my time with image generation it was using ComfyUI, what is this UI ive never seen this before
>>
>>109320779
...you mean anon's screenshot? That's clearly ComfyUI. What? Also, nice gen.
>>
>>109320770
Yeah mixing artists and LoRAs. Good to know.

Might actually nuke all my IL LoRAs and start over fresh with just Krea 2.
>>
>>109320900
I nuked all IL and Chroma models / loras
>>
File: ComfyUI_05509_.png (1.98 MB, 1920x1080)
1.98 MB PNG
>>
make sure to strip the metadata from your gens before you post them. some loser figured out how to extract personal info from them using 4chan server fuckery
>>
>>109320471
>>109320571
Hello Catjack
>>
File: ComfyUI_05278_.png (1.65 MB, 1024x1024)
1.65 MB PNG
>>
File: connelly_00003_.jpg (741 KB, 2048x1536)
741 KB JPG
>>
File: connelly_00024_.jpg (650 KB, 1536x2048)
650 KB JPG
>>
File: 0 users.jpg (1.73 MB, 979x2558)
1.73 MB JPG
>>109321224
>>109321239
Hello Ani! :)
>>
File: ComfyUI_43325398_.png (3.59 MB, 1920x1080)
3.59 MB PNG
>>
>>109321343
>>109321408
Where are you getting these loras from?
>>
>>109321475
i just hacked his jpg vram, looks like he trained them locally.
>>
File: connelly_00054_.jpg (905 KB, 1536x2048)
905 KB JPG
>>109321475
I trained it. Coping and seething because the model doesn't properly know Walkman

>>109321485
well done hackerman
>>
Ahh, another day of genning 5000 futanari images.
>>
File: connelly_00056_.jpg (650 KB, 1536x2048)
650 KB JPG
>>
what is this https://huggingface.co/Comfy-Org/Krea-2/blob/main/loras/krea2_style_reference.safetensors
uploaded 2 hours ago
>>
>>109321624
Another snake oil release to keep you hooked.
>>
File: connelly_00071_.jpg (655 KB, 2048x1536)
655 KB JPG
>>
>>109321624
its just the ostris lora, it has the exact same sha-256
>>
a painful blackpill to swallow is that majority of the based character lora makers of the sd 1.5 and sdxl days won't be coming back to making krea2 loras. 4 weeks later and all that's being posted on civitai is fake ai avatar females, artstyle loras and nsfw hardcore lewd concepts.
>>
>>109321694
Character loras are one of the easiest things to train. Are you incapable of doing it yourself?
>>
>>109321705
This. Also, most of the loras from the sd 1.5 era were terrible. Quantity does not equal quality.
The real bottleneck for advanced genners is prompting and finding ways to automate it, which is far more technical and less flashy. The idea of shitting out loras is no longer interesting
>>
>>109321694
I thought civitai banned real person loras?
>>
>>109321694
I can't buy crypto so I can't buy buzz, otherwise I would've trained a bunch of style LoRAs by now.
>>
>>109320441
sfw vageen department here, we're impressed with the pink bunny from last thread and OP collage, keep up the good work
>>
>>109321676
>sha-256
are we still living in 2023, gramps?
>>
File: ComfyUI_ID4_00187_.png (2.1 MB, 1936x1088)
2.1 MB PNG
>>109321694
I'd like to do a proper anime finetune, but it's way beyond what's possible with a single GPU, even a powerful one. Best I can do is show style loras: https://huggingface.co/quarterturn/id4_nichijou
That's 1.34K images and took about 24 hours of GPU time.
>>
>>109321768
huh?
>>
>>109321587
>>109321505
Nice touch with the era-appropriate clothing. The likeness is also very close, if not perfect.
>>
>>109321768
Why not? It's enough to show that it's ostris lora and hugging face shows it in the ui. Why wasting time checking autov3
>>
It probably knows her though.
>>
why is z image so good with paragraph slop prompts instead of booru tags? i have to use an llm to shit out verbose garbage instead of using human readable tags
>>
>>109321813
if verbose natural language is harder to comprehend than tags its an iq problem on the end user.
>>
>>109321817
purple prose is read for pleasure not for sharing information.
>>
>>109321813
prompt with tags -> Z-engineer -> natural language.
>>
File: connelly_00095_.jpg (756 KB, 1536x2048)
756 KB JPG
>>109321784
Career opportunities era look is surprisingly hard look to nail since she still had baby fat left. In Hot spot she looks completely different
>>
File: caption.jpg (509 KB, 1938x1250)
509 KB JPG
>>109321782
How did you do your captions?
I'm currently .json captioning a 500 image dataset for an ID4 lora and it's taking FOREVER, no (local) llm can get the bounding boxes right and I'm having to manually adjust them all using this
https://github.com/Auryg/Ideogram-Json-Captioner
>>
>>109321832
Well you nailed it if that's your lora.
>>
File: Ideogram__01423_.jpg (1.09 MB, 2048x2048)
1.09 MB JPG
>>109321835
That said I do a quick test train the dataset raw with just the art_style captioned and it came out okay. But it can't replicate the characters very well
>>
>>109321835
I used a slopcoded python script which feeds the images to qwen3.6-35b-a3b. It's slow. Full HD images take two minutes per image. It might be overkill to give it that much info. It's overall very accurate but does misidentify characters if it can't see the expected details.
The dataset is here: https://huggingface.co/datasets/quarterturn/nichijou-hd-json-captioned
>>
>>109321768
> new is better
>>
File: 00062-999936776.jpg (728 KB, 2880x2112)
728 KB JPG
>>109321705
hardest thing is image hunting, dataset sorting out, captioning progress (longest and painful aspect). you can't cheat the process unless you don't care for having a accurate enough looking and functioning lora.
>>
File: connelly_00101_.jpg (630 KB, 1536x2048)
630 KB JPG
>>109321840
Thanks. I'm not happy but dataset can always be improved. I'll upload it
>>
>>109321885
You did not answer the question about a lora.
>>
>>109321866
Are these not just plain text captions?
Picrel is what one of my captions looks like. Basically the same format you use to prompt the model
>>
>>109321827
>is read for pleasure
only if you are gay or a woman or a gay woman
>>
>>109321885
Not really you can throw together a character lora dataset in a couple hours minus training time
>>
>zittards don't read
who'd have thunk?
>>
>>109321900
I tested his previous Stellar Blade LoRAs and they're overcooked. Tested the character in different outfits, and all outfits give her some NSFW aspects, like being skintight or bottomless, or an umprompted camel toe. He put effort into them though, but his LoRAs aren't good, he's failing at the captioning process, like some anons said in prior threads
>>
>>109321917
i guess that who rated the prompting for llms and nu diffusion models
>>
>>109321909
If you look at the files, the json captions are there. The dataset just shows the plaintext description. I wanted to incorporate the json data into the dataset but kept running into issues where the dataset browser in HF didn't like it, so I gave up and just included the text caption.
>>
Here's the captioner I used for my Nichijou dataset: https://github.com/quarterturn/ollama-captioner
I run qwen3.6-35b-a3b at full context so it can look at high-res images, so despite it being only 3b active it takes almost all of my 48GB VRAM. With some adjustments it can probably work in 32GB, or maybe less, if you go below q8, but I don't want to because I think it will hurt accuracy.
>>
>>109321984
>despite it being only 3b active
this doesn't mean it will only take 3b worth of vram
>>
>The Trump administration is showing signs it could ban cutting-edge Chinese AI models
https://x.com/kimmonismus/status/2079167072571978033

This is funny, the ban's not going to even work, all you need is a phone connecting to a random website outside of the country that hosts Kimi somewhere. Even employees can do that. And even if it does work, your country is forcing itself to use the more expensive and (eventually dumber) product. That's hilarious.
>>
>>109321959
>>109321984
Oh cool. I'm a 24GB vram pleb though. I might try Gemma 26B-A4B I doubt it'll work so it'll be back to manually bboxing for me most likely
>>
>>109322009
It'll work on 24GB you just need to adjust the context size on your back end and feed it smaller images. Maybe 480p size is enough if you don't care about it reading small text.
>>
>>109322002
I swear someone important already said this but the result of this would be COMPANIES not using chinese models.
They want idk the big 4 consulting companies to keep slurping openai not getting dangerous thoughts about buying their own equipment to run deepseek
>>
https://huggingface.co/xixxix-HF/JenniferConnellyEarly90s_krea2
>>
>>109321946
how do i improve of the captioning progress, is there better training settings. i thought the lunafrya, reema rochelle and artoria pendragon (lancer alter) loras were pretty good. The lily lora was fairly functionable but this raven lora is not looking great. Are the better training setting to use? pic related is the settings I've been using for the longest time with local 5090/128gb ram setup. change the learning rate, change the optimzer, reduce the steps, try lokr instead of lora?
>>
>>109322002
>api story
no one care bruh
>>
>>109322039
You want to use sigmoid for character/likeness loras
>>
>>109322033
Yeah I suppose it doesn't really affect anything for hobbyists, what goes on inside companies is already a shrouded behind closed doors so whatever they're doing is completely irrelevant to the general public anyway.
>>
>>109322039
you must be trolling but i genuinely can't tell
>>
File: 1756760279035755.png (23 KB, 997x669)
23 KB PNG
I checked civit and couldnt find anything good I'm modding neverwinter nights and would like my character portraits to be like doom guy / look more damaged/fucked up depending on their current health. I assumed i would just find a gore lora and crank up the strength but I seriously cant find anything
>>
>>109322037
Thanks I will test it later.

>>109322063
I can confirm sigmoid works but the others may work as well.
>>
>>109322063
how about the learning rate, optimizer, steps? i had the caption dropout rate at 0 instead of the default 0.05. Do i increase the batch size to 2? i heard people hyping lokr, is it really better than loras?
>>
File: 362787484401138.png (2.71 MB, 1344x1728)
2.71 MB PNG
>>
File: ComfyUI_ID4_00193_.png (2.34 MB, 1456x1456)
2.34 MB PNG
Running the Nichijoui ID4 lora at full strength in a non-standard scene works reasonably well. I have no idea if it's the best possible lora but it seems to work OK.
>>
>>109322037
Thanks
>>
>>109322039
I don't know, man. I use Danbooru tags and train on Danbooru style so I can shuffle tags around without overcooking my characters.
As for testing your LoRA, same thing happened to me with Danbooru tagging when I trained an already existing character wearing an already existing concept of that specific outfit, then trained that same character again on top of it, same outfit.
If Krea already knows the character, the franchise, the series, and the outfit, I think your captioning strategy should be different, since you're not teaching it something it's unfamiliar with.
Also, just to be clear, I'm not the anon who criticized your tagging. I just wanted to remind the other anon that some anon said your tagging was bad, Krea isn't even my main model anyway.
>>
how do i use automagic v3 for ai toolkit? its not selectable.
>>
>>109322221
update
>>
Or just select automagic2 and replace it with "automagic3" in the config.
>>
>>109322111
Why are you quantizing the transformer with a 5090? Cache the text embeddings if you need to free up some vram and get rid of the trigger word
>>
>>109322221
>>109322247
is automagic3 good? ostris keeps rewriting it lol
>>
>>109322263
Seems to work pretty well.
>>
>>109321984
> qwen3.6-35b-a3b
> all of my 48GB VRAM
why? moe is for scenarios with low vram, you could take dense 27b and it would be much smarter and accurate
>>
Why do I see people saying changing the text_encoder also changes the image quality?
Isn't that the job of the VAE?

Maybe they just assume things.
>>
>>109322039
did you delete the blonde child girl lora from civitai?
>>
>>109322288
I tried it. It was much slower.
>>
>>109322247
i updated the ai toolkit, closed the nodejs tasks and restarted ai toolkit. it works.
>>109322263
redditors and people in the official discord hype it up.
>>109322258
adding the trigger word to that ui space is an overkill and duplicates it in all the txt files? would cache the text embedding hinder the performance? is it really a significant difference in training the krea2 transformer model at full fp16 vs fp8/q8.
>>
>>109322370
I don't really want to continue giving you advice, since you ignored previous anons about your atrocious captioning process and trained the lora anyway
>>
>>109322120
Kino
>>
>>109322311
it's still there including the z image base version. you have to disable the browsing level settings.
https://civitai.red/models/2761525/lunafreya-nox-fleuret-final-fantasy-xv-krea-2
https://civitai.red/models/2727301/lunafreya-nox-fleuret-final-fantasy-xv
>>
Wtf lol
>>
File: autumn mountains.png (3.78 MB, 1920x1200)
3.78 MB PNG
>>
>>109322120
> the phtanom of the opera
https://www.youtube.com/watch?v=tXvFL9RMjGU&t=638s
>>
>>109322039
fyi, you do not need a 64+ rank lora unless you're making concept/style loras with a big dataset. the higher the rank, the more the lora overfits and less flexible it is. it is not a representation of lora quality. this is why you will almost never see high rank character loras.
>>
>>109322297
changing the text encoder to a less quantized one will result in better quality, but i dunno about using a different text encoder the problem wasn't trained on. i cant imagine it offers any benefit because the model wont understand whatever new knowledge that text encoder has, assuming its even compatible.

this is why using abliterated/uncensored text encoders does absolutely nothing for image gen.
>>
Just got out of a coma
What's the met these days?
>>
>>109322415
> you have to disable the browsing level settings
but I can't see the rest then wtf
>>
>>109322572
going to reduce the rank to 32 but keep the res 1024 with the auomagic v3 and sigmoid. the other anon recommended not quantizing the transformer model but training installed and did move an it at all. Going back to the qfloat8 setting. cache text embeddings a good idea or a risky move?
>>
>>109322599
uncovering secrets embedded by python in gens
>>
File: screenshot.1784557266.jpg (37 KB, 750x114)
37 KB JPG
>>109322631
>the other anon recommended not quantizing the transformer
you should be using qint8(only can put it via advanced option since UI doesnt list it for some dumb reason)
pic-related

>cache text embeddings a good idea or a risky move?
do not cache text embeddings.
select unload TE and select cache latents.
uncheck low vram, you have a 5090. its not needed
make sure lora is bf16 of course
>>
>>109322631
> the other anon
other anons don't train loras or theirs are inferior to yours
>>
>>109322678
so type qint8 for both the transformer and text encoder? so will it redownload another model krea2 raw mode of quantize the full bf16 model to the qint8 format?
>>
>>109322713
imo, someone should make a benchtest for loras. a simple test that shows how flexible it is when used with style loras, concept loras and other metrics.

if models can have bench tests, why not lora? we really need a somewhat objective metric for testing quality loras
>>
File: ComfyUI_temp_cbjot_00028_.jpg (1.71 MB, 1056x2048)
1.71 MB JPG
>>
File: screenshot.1784558081.jpg (169 KB, 721x352)
169 KB JPG
>>109322764
>so type qint8 for both the transformer and text encoder
yes

>will it redownload another model krea2 raw mode of quantize the full bf16 model to the qint8 format?
no
source:claude
>>
any interesting wan lora released recently?
>>
>>109322788
final quest before i click update job. leave the trigger word blank? I'm wondering if it duplicates and doubles the same "stellar blade raven, raven(stellar blade)," tag that on every txt file already.
>>
>>109322825
if you have the same caption in every prompt, it doesn't matter if you leave the trigger word blank, it will still act like a trigger word
>>
>>109322825
you can leave it blank since it's already in the caption files.

>>109322823
yeah iGoon released another banger today
https://litter.catbox.moe/mjkayoeb79oyytjs.gif
>>
>>109322858
its not a release if its gated behind fanvue
>>
>>109322823
Minimum ram needed to Gen this kind of video?
>>
>>109322858
multi scene lora?
>>
>>109322918
its a scene change/blink lora. you wont find any of his loras on civitai because its banned so you have to join his discord

>>109322889
anon never said it had to be public ;)
>>
>>109322945
>you have to join his discord
i dont know why you would lie to him when theres nothing on his discord. its just a gateway to his fanvue
>>
thanks for the help and suggestions anons. i hope this one is baked a lot better than the previous one. if it fails, i might bump of the res to 1280 since i have enough vram for it judging by pic related showing it consume only 21gb of vram or perhaps changing the batch size from 1 to 2 will make a difference. might have to diverse the dataset a lot better with raven in different outfits and and add fanart into the mix.
>>
slap lora. it's sad that you need lora for every specific action
>>
>>109322962
you dont have to but you can see what lora he will be making next and vote on it.

>>109322976
if it comes out bad its 90% of the time due to the dataset. use klein 9b add more poses and stuff
>>
>>109322991
i can vote on loras i won't even have access to? wow, sign me up
>>
Impregnation with teenage Jennifer Connolly
>>
>>109323013
have you tried just pirating them?
>>
>>109323025
im unaware of any lora pirating sites.
it's always been frowned upon to paywall loras especially when they're trained on a free open source model
>>
>>109323047
You are really ignorant if you can't find civitai copy site.
>>
>>109323047
why? training video loras is time consuming and expensive. $5 is nothing by comparison. you probably spend more buying starbucks everyday
>>
It's weird. Krea 2 Identity lora is better at changing outfits and retaining the image than Klein, but Klein can do more complex editing, so now I'm in this weird place where I have to try both models to see which gives me the best results in certain edge cases. Very annoying.

Can't wait until Krea releases their edit model.
>>
>>109323099
Thank you for letting us know.
>>
>>109323105
Your welcome. I advise anyone to try both Krea2 & Klein for editing since each has their own strengths and weaknesses. Neither is better than the other yet.
>>
>>109323057
feel free to share any
>>109323059
im not funding a jeet who thinks it's okay to make money off an open source model. thats just not happening
>>
It's a good gen if I want to have sex with her.
>>
File: 75838.jpg (244 KB, 1280x1600)
244 KB JPG
>>109322037
danke
>>
>>109322507
mexican detected
>>
File: gob girl.png (3.3 MB, 1440x1440)
3.3 MB PNG
>>
>>109323138
I don't spoonfeed retards who don't know how to use a search engine.
>>
File: 154219CUI_00001_.png (950 KB, 768x1344)
950 KB PNG
>>
File: ComfyUI_02827_.jpg (979 KB, 2024x2696)
979 KB JPG
>>109322037
Cool
>>
Impregnation with Jennifer Connolly while her legs are wrapped around me and I am deeply kissing her
>>
>Z-Image
No ControlNet
>Anima
No ControlNet
>Krea2
No ControlNet
>Ideogram
No ControlNet
>Every other new model
No ControlNet

Why?
>>
>>109323360
https://github.com/facok/comfyui-krea2-controlnet
What's this then?
>>
>>109323360
zimg, zanama and krea have controlnets
>>
>>109323366
its depth control and not canny so its useless as fuck
>>
>>109323355
gm sar!
>>
>>109323360
Z-Image was supposed to release an edit model, making control nets pointless.

Krea2 is supposed to release an edit model, making control nets pointless.

Anima has controlnets.

Nobody gives a fuck about Ideogram.
>>
File: connelly_00185_.jpg (546 KB, 1536x2048)
546 KB JPG
>>109323343
>>109323160
>>109322137
I uploaded factor 4 lokr there too, sidegrade? upgrade? hard to say
>>
>>109323360
>>109323481
>>109323370
Is there a "rough sketch input to generated image" controlnet for ZIT and how can I train one if there is none?
>>
>>109323360
don't need 'em anymore
>>
next faggollage should only be teenage jen con
>>
>>109323757
why?
>>
>generate incredible femboys
>dont want everyone to know im gay
>>
>>
>>109323834
the tribulations of a based gensman
>>
After seeing all the variations of female faces in the modern slopped models in the last 200,000 ganned faces I've looked at, I'm going back to Chroma 48.
>>
>>109323942
I mean the absolute best Chroma can do is basically this person's gen except with a background that isn't quite as coherent
>>109323910
>>
>>109323834
>generate incredible femboys
I sooo believe you, nogen anon.
>>
>>109323990
real: this comment exists
gay: anon wants to see other anon's amazing femboys
>>
>>109323942
*feces
Sorry for the typo
>>
File: debo_si_k2_00005_.jpg (798 KB, 2047x1366)
798 KB JPG
>>
krea 2 doesnt seem to mix loras well
>>
>>109321694
Styles are solved though. Just use the style transfer method, you get way more styles in the domain of styles that were only possible with cloud models like Dalle 3
>>
>>109323965
>I mean the absolute best Chroma can do is basically this person's gen except with a background that isn't quite as coherent
No, for 1girl realism, Chroma with 16 generated images will give you 11-13 fucked worthless images, 2-4 good unique images, and 1 kinosovl unique image no other model can. Shit backgrounds, slight blurriness, grid artifacting with huge prompts and some body horror are easily fixable after if you care and not important.

If you care about uniqueness in your 1girls Chroma is still unbeaten. The closest is ZIT when using a LoRA and generating at 2000 by 2000 which makes it break free into generating more unique faces than usual but not by a huge amount. Every other model is either too slopped, or unstable for realism.
>>
>>109321694
the ai avatar girls annoys me so much. who the fuck is using that shit? feels like a bot farm is just churning them out for more bots to use on instagram
>>
File: ComfyUI_02877_.jpg (776 KB, 3120x1752)
776 KB JPG
>>109323624
The lora performs better, but it isn't a fair comparison because the rank 64 lokr has roughly half the trainable parameter capacity.
A lokr rank of at least 128 with the same 4x4 factor would be needed for a closer comparison with a rank 32 lora
>>
>1 good image out of 16 isn't unstable
>>
>he doesn't gen with a batch of 32
>he doesn't have a 300B parameter VLM filtering the duds
ngmi
>>
>>109324162
By other models being unstable for realism I meant to do it at all, they can't give you 1 unique unslopped image in 1000, because all modern models have coalesced into low seed and low creativity variance because of their novel-like captioning during training in order to better follow user prompts.
>>
You should only be able to post itt if your GPU is actively genning/training and it detects you have 24GB VRAM 64+GB RAM btw.
>>
File: 20-48-2026.jpg (533 KB, 768x1152)
533 KB JPG
Gotta say, I was expecting a troll, but it’s good. Here's a fast slop gen.
>>
>>109324197
post some goated chroma gens.
some of you folks need to take off the rose-colored glasses.
>>
File: Krea2_turbo_05805_.jpg (730 KB, 1472x2176)
730 KB JPG
it's okay
>>
>>109324248
Nice. Cool spitfire too.
>>
>>109324224
>rose-colored glasses
How can that be the case when anyone can click a button to queue a workflow with one model and see it's output compared to the other?
>>
>>109324311
lets see your outputs then
>>
>>109324153
>A lokr rank of at least 128 with the same 4x4 factor would be needed for a closer comparison with a rank 32 lora
Interesting, thanks! I had wrong information.
>>
the only Chroma user I've seen was some weirdo that kept putting some asian woman in degrading scenarios
>>
>>109321984
This is late but you're doing it wrong. I don't use ollama but it must have some sort of equivalent to the llama.cpp --cpu-moe flag. Ask Claude for help or read the docs.
>>
File: ComfyUI_04257.png (2.62 MB, 1920x1080)
2.62 MB PNG
>>109324213
Well, I still get to play... cool I guess.
>>
>>109324347
i thought irl jenny finally caught up to you
>>
>>109324347
needs more noise
>>
File: ComfyUI_02884_.jpg (722 KB, 2024x2696)
722 KB JPG
>>109324248
Nice idea
>>109324321
No problem
>>
File: debo_si_k2_00006_.jpg (1.04 MB, 2047x1366)
1.04 MB JPG
>>109324213
shouldn't you post a gen with this?
>>
>>109323118
>Neither is better than the other yet.
This was always the story for this local branch of this hobby
>>
>>109324135
didn't krea developers decide not to release the style transfer method?
>>
File: which.png (2 KB, 261x44)
2 KB PNG
which one do I delete FOR EVER?
>>
>>109324463
yourself
>>
z-image is useless unless you have porn addiction
>>
File: ComfyUI_03123.png (2.49 MB, 1920x1080)
2.49 MB PNG
>>109324356
I wish! I've just been busy with other bullshit though.

>>109324357
Honestly, I should've pulled noise from the movie I was aiming for itself (Virtuosity).
>>
>>109324463
who is Ever and why do they want one of the models deleted?
>>
>implying all anons here don't have a porn addiction to some degree
>>
>>109324474
english as third language in toher words, ETL
>>
>>109324468
in other words it's extremely useful
>>
>>109324214
>>109324248
>>109324375
always nice reminders she was so hot in the 80s in movies. looked her up again. might make some ai of her now.
>>
>>109324468
but a big selling point of krea2 is the out of the box nsfw?
>>
>drooling over 80s actress
is anon in their 60s?
>>
we're missing out on so many early 1900s beauts
>>
>>109324553
definitely some boomer.
>>
File: connelly_00328_.jpg (596 KB, 1512x2048)
596 KB JPG
>>109324553
Nope, it's my pon farr
>>
File: 4512245412.jpg (266 KB, 2514x1585)
266 KB JPG
>>109324450
They didn't release it but there's since been two separate methods for it
>>
>>109324553
wait...I am not attacking you but genuinely curious...how are you not attracted to females of any era? I mean we could have a chat about some beautiful women in time. that goofy 1960s batman show had a ton of beautiful females on it playing the roll of molls as were some of the villainess. the 1940s has some very lovely ladies and so on. I am a fan of the victorian era and 1920s and 1930s personally. all era's had beautiful ladies in various forms, attires and hairstyles and more. I look at that 1980s actress and go damn she is very lovely. same with many others. The female in the movie howard the duck in her panties is still one of those oh wow she's hot things. even that goofy jeane in ghostbusters is hot in many ways and I mean the old 1980s ghostbusters.
>>
File: Krea2_turbo_00588__75.png (2.36 MB, 1776x1332)
2.36 MB PNG
hehehe :]
>>
>>109324553
Boomer general, nothing wrong with browsing civitai's good quality loras while grilling steaks in the back garden God bless AI
>>
>Thread turns into generic deepfake slop
This is why you should never share your loras with indians
>>
>>109324613
is mine generic too?
you know the one
the good one
>>
File: images.jpg (21 KB, 400x224)
21 KB JPG
>>109324598
It's not forbidden by the laws of the universe that anon and his contemporary crush could end up together.
>>
>>109324600
Krea is really bad with posing by default, you can describe the sexiest scene but if you don't also describe a sexy pose they just kinda stand there
>>
>>109324463
all mixeds should be eradicated
>>
>>109324613
friday to sunday: zoomers curious about AI
monday onward: retired boomers take over /ldg/
>>
>>109324579
which do you recommend?
>>
>cries for weeks about there not being enough loras
>cries when people start sharing loras
just shut up, idiot
>>
I just want to faceswap and remove clothes, I'm a simple person.

How do I do that?

For clothing removal there was a website called deepsukebe some years ago, but I want to do it offline now.

I just want images, my hardware is not good enough for video anyway.

that's why I mentioned faceswap (can pick an image from ant porn and add the face I want) and clothing removal (get a photo of the person I chose and have all the clothes removed, like deepsukebe used to
>>
>>109324643
yes saar post more jenifer conoly 1girl bikini mirror selfie saar
>>
>>109324645
This is for pictures of yourself, right? It's much easier to buy a camera stand, take off your clothes and take some pictures.
>>
>>109324468
in other words it's still the best model.
>>
File: Krea2_turbo_00591__75.png (2.72 MB, 1254x1884)
2.72 MB PNG
>>
>>109322037
>single brow in every shot
throwing up in the mouth
>>
>>109324642
I've only used the first I came across which was good enough
https://github.com/BigStationW/ComfyUi-Scale-Image-to-Total-Pixels-Advanced

Ostris trained a LoRA so it might be worth looking into as well
>>
File: ComfyUI_02895_.jpg (706 KB, 3120x1752)
706 KB JPG
>>
i asked claude to build me a decompiler for a hentai game to extract the resources so i can then use those raw resources to make a lora with. absolutely GOD like ai
>>
>>109324688
>https://github.com/BigStationW/ComfyUi-Scale-Image-to-Total-Pixels-Advanced

Meant to link
https://github.com/BigStationW/ComfyUi-Untwisting-RoPE
>>
>>109324691
what
>>
>>109324680
Slop
>>109324691
Kino
>>
>>109324657
>This is for pictures of yourself, right?
No
>It's much easier to buy a camera stand, take off your clothes and take some pictures

She died in 2015.
>>
>>109324717
>She died in 2015.
let her rest in peace then
>>
>>109324717
>No
No one here makes gens of other people officer
>>
photograph of Jennifer Connelly (actress) with clean two eyebrows holding a Gilette razor
>>
File: Krea2_turbo_00599__75.png (2.27 MB, 1254x1884)
2.27 MB PNG
>>
>>109324695
>Untwisting-RoPE
It was shit. I tried it when it was shilled here for ZIT.
Post your gens with Krea that showcase it not being shit.
>>
File: file.png (948 KB, 1920x1200)
948 KB PNG
>>
>>109324375
Excellent! Maybe should be holding a Polaroid camera to be more era-appropriate.
>>
File: Krea2_turbo_00601__75.png (2.2 MB, 1254x1884)
2.2 MB PNG
>>
>>109324153
>>109321832
Beautiful, but a few of these look like they need a bit more eyebrow tweezing.
>>
>>109324138
>>109323965
>Muh background issues

My Chroma-Krea wf oneshots fixing the background and most of the limb issues anon. It's by far the most advanced way to fix Chroma gens. You think pic rel is easy to oneshot just with raw Chroma? Lol

https://files.catbox.moe/y7p7mk.png

It's fully compatible with NSFW (that's the power of it). Though, generally, I agree, anons have no idea what Chroma was capable of if they claim its backgrounds or gens weren't coherent (it was possible to get good ones with good seeds)
>>
File: file.png (3.19 MB, 1920x1200)
3.19 MB PNG
>>
File: ComfyUI_02900_.jpg (793 KB, 3120x1752)
793 KB JPG
>>
File: 00174-2391719349.png (2.38 MB, 1152x2048)
2.38 MB PNG
Computer, make them kiss.
>>
>>109324835
>model doesnt understand aspect ratio and refuses to crop characters so it makes them as tall as the height of the image and now they're 7ft tall reaching the ceiling
>>
>>109324824
???
what even is the purpose of Chroma in that workflow
>>
>>109324845
so he can generate pizza w no filter.
>>
>>109324843
I don't think so
>>
>>109324845
i havent looked at it but i but its the using chroma as a first pass at low steps then krea as the 2nd pass. not a bad idea considering chroma has better nsfw capabilities

>>109324856
you could already do that by just doing i2i with any chroma gen.
>>
>>109324871
>not a bad idea considering chroma has better nsfw capabilities
what's something Chroma can do that Krea with a NSFW lora can't?
but anyway, I'm asking because of "You think pic rel is easy to oneshot just with raw Chroma?" so NSFW capability isn't relevant
>>
>>109324856
lol wat? why would pizza trigger the filter
>>
>>109324884
>what's something Chroma can do that Krea with a NSFW lora can't?
natively do porn without needing a specific lora for every little thing. that is the power of an uncensored finetune.
>>
>>109324824
give wf with only chroma v48 (best version)
>>
>>109324845
>what even is the purpose of Chroma in that workflow

Try to generate this image >>109324824 with just Krea and no LoRAs... Not a single model has caught up to Chroma's prompt understanding anon. That is Chroma's purpose. You can't defeat a base model that simply is SOTA at NSFW.
>>
>>109324893
>a specific lora for every little thing
I said "a NSFW lora", not "a NSFW pose/concept lora"
you might be thinking of SDXL models
>>
>>109324824
>based chroma asianfootfag became a kreatard
owari
>>
>>109324803
>>109324753
can't resist that ass.
>>
>>109324901
Those jumbo general NSFW loras like MysticXXX and SNOFS tend be shit and can only do the basics. With Krea 2 in particular, they also tend to change the identity and don't mix well.

Using no loras will always be better than using loras.
>>
File: 703286982.jpg (529 KB, 1920x2560)
529 KB JPG
>>109323624
beats me. loras in picrel:
> <lora:krea2filterbypass:1> <lora:fedor_bypass:3> <lora:AmateurSlider-KREA2_v1:-2> <lora:Krea 2 - Naturally Sagging Breasts:1>
>>
>>109324899
>and no LoRAs
there is no downside to the TextFusion lora, why would that disqualify Krea?
>>
>>109322037
Gratias, amico. Mind sharing what your captioning strategy is? How much of the output fidelity of a lora dependent on captioning vs. general data quality/size?
>>
>>109324923
Okay cool but the question is what's something Chroma can do that Krea with a NSFW lora can't?
>>
>>109324628
ah so his mental block is because he can't take that 1930s beauty and impregnate her and have her in the kitchen. figures.
>>
>>109324938
I just told you the flaws of Krea + NSFW lora usage which is not a problem with Chroma. See >>109324899
>>
File: 1754444711605995.jpg (99 KB, 971x648)
99 KB JPG
>>
guys, i wanna start one of those tiktok/instagram influencer profiles with a fake AI lady, what would be the general workflow to prompting such? i've been making a lot of art, but i'm clueless when it comes to "real things"

if anyone can offer a hand i'd be grateful
>>
>>109324958
Go away
>>
jeet means victory
>>
>>109324955
>I just told you the flaws of Krea + NSFW lora usage
OKAY COOL BUT WHAT'S SOMETHING CHROMA CAN DO THAT KREA WITH A NSFW LORA CAN'T?
SHOW IT
SHOW
IT
FUCK
>>
>>109324958
oversaturated market. come back 2 years ago
>>
melty status?
>>
>>109324981
also, why can't retards follow simple instructions?
>>
File: connelly_00387_.jpg (701 KB, 2048x1640)
701 KB JPG
>>109324929
Sidegrades in that example I'd guess

>>109324934
example
>20-years-old Jennifer Connelly (young actress). Close-up photograph of a young woman with fair skin, dark brown hair, and green eyes. She has red lipstick, slightly parted lips, and wears a single, dangling, diamond and pearl earring on her right ear. The background is blurred with green foliage. Her expression is serious and focused.

>How much of the output fidelity of a lora dependent on captioning vs. general data quality/size?
Dataset quality is the most important thing, but you can ruin lora with wrong captioning.
>>
>>109324957
she had a unibrow when she was that age
>>
File: 1738542058664384.jpg (155 KB, 1072x1256)
155 KB JPG
>>109324691
>the thing on the right
i quite like female aliens, but i suppose i do have limits after all...
>>
File: 1753260253299600.png (366 KB, 686x386)
366 KB PNG
I swear 1girls are starting to become such unoriginal slop that I swear I've seen them before.

1girl slop creators are the "An animator was messing around" of AI
>>
File: flowerglass.jpg (1.48 MB, 1920x1200)
1.48 MB JPG
>>
File: 00071-1016052298.jpg (624 KB, 1856x2880)
624 KB JPG
the second raven kre2 lora is a lot more stable and the body looks more accurate than before.
https://gofile.io/d/P2wXQG
>>
>>109325038
based
>>
>>109324991
>example
>20-years-old Jennifer Connelly (young actress). Close-up photograph of a young woman with fair skin, dark brown hair, and green eyes. She has red lipstick, slightly parted lips, and wears a single, dangling, diamond and pearl earring on her right ear. The background is blurred with green foliage. Her expression is serious and focused.

That's it? Here I was captioning data with huge captions. Thanks again.
>>
>>109325038
did you again describe every facial feature in the captions with no dropout and therefore it is required to copy a whole sentence to get the likeness?
>>
>>109325038
damn that turned out pretty damn good. congrats
>>
>>109325057
you're retarded
>>
>>109325038
I don't have porn addiction | who is this?
>>
>>109325049
It's better to go for shorter, but correct captions. Increase prose for usual poses and expressions, test the base model what kind of expression and pose it makes before using it with lora.
>>
>>109324958
you know do kneedful?
>>
>>109325057
fyi, you only need to describe pose/scenery changes in captions. hell, you can probably just use a trigger word with no captions and get good results too.
>>
>>109325057
That's retarded. You caption things you don't want to appear as default alongside your trigger word.
>(Character/trigger) wearing (clothes) doing (pose/action) in a (room/setting/background)
>>
>>109324896
>chroma v48 (best version)

I tried an int8 version of it but too much limb horror and it was far too slow, and I have a skill issue with getting svdq (nunchaku) to work properly with 1 HD
https://huggingface.co/tonera/Chroma1-HD-SVDQ

Once I get that I'll have it working with 1 HD at least.
>>
>2026
>people still don't understand how to caption a lora
How difficult is "caption as if the model already knows the concept" to understand?
>>
>2026
>i added a unibrow to my deepfake lora through sheer hallucination and pretend im above all
>>
>>109325106
Like this is bad caption >>109325049
There's no need to caption her fair skin and green eyes, you want your lora to associate "Jennifer Connelly" with those attributes by default so you leave them out of your caption
>>
>>109325126
I tried overcaptioning, I tried undercaptioning, in the end nothing really matters.
>>
>>109325063
raven from stellar blade
>>
File: ComfyUI_01910_.png (3.44 MB, 1920x1200)
3.44 MB PNG
>>
>>109325149
you can have them in the caption to make it more versatile but then you should implement tag dropout
>>
File: 00076-2573571019.jpg (361 KB, 2880x1664)
361 KB JPG
>>109325051
not really, seems to be a lot more stable but gray cybernetic suit is too complex in its design, the basic triggers won't be enough to accurate enough. Not sure how to further improve the accuracy of gray cybernetic cybersuit to 95% believability as the og presentation of it. tired of putting this much time into this lora. i will to give you dataset if your interested in trying it for yourself. i want to move on to different females from other franchises.
>>
>>109325152
yeah, people place way too much importance on captioning. this isnt sd 1.5 days. unless you're training a complex concept the model has 0 knowledge of, you really dont need to go crazy with it
>>
File: 1753347763036915.jpg (519 KB, 1424x2216)
519 KB JPG
>>109325038
first attempt i just used `stellar blade raven, pool table` and let the enhancer do the rest
dope lora desu
>>
>>109324991
>>109324470

Forget about i2v t2v, what I want is Lora2Joi
>>
>>109325173
brainlet.
>>
>>109324931
Not counting it out, but Krea gives too much slop by default, and LoRA coping ain't fixing only what realignment from a full finetune can fix.
>>
File: 1768525842787036.jpg (2.92 MB, 2352x1568)
2.92 MB JPG
This is lowk why I don't share LoRAs because everyone (except me) is either too porn addicted to trust or too moronic to understand the artistic value.
>>
man you guys are real saints
>>
Krea2 is king.

https://files.catbox.moe/cdrjx9.png
>>
^okay you deepfaked someone naked. Now what?
>>
File: Krea2_turbo_00606__75.png (2.19 MB, 1254x1884)
2.19 MB PNG
Slop
>>
>>109325113
I think int8 in general probably fucks up gens more, same with HD Flash's int8 version, but if you can get it working with fp16 and don't mind the wait time just replace the int8 HD Flash/t5 loader with Chroma's standard one from the wf.
>>
the most annoying thing about captioning is explaining to retards how to caption properly
>>
File: 929771076678734.png (1.78 MB, 1152x1472)
1.78 MB PNG
>>
>>109325262
if your dataset is good you literally don't need to caption. just leave it blank or add a trigger word.
>>
>>109325271
this.
anons being autistic about captioning is the worst thing about captioning. almost like that one anon who kept insisting you need to use regularization datasets for good loras when that hasn't been relevant in ages
>>
White people don't caption. Typing out your gay little prose is jeet coded.
>>
>>109325086
>>109325149
>>109325161
Thanks anons. I'm going to try and train a new lora tonight. Another thing I struggle with with character loras and people is that I tend to lose consistency depending on the distance/zoom of the generated image. Any tips for that?
>>
File: 1769588483000773.jpg (204 KB, 1424x873)
204 KB JPG
>>
>>109325271
this is retarded advice, you're teaching the model to lose the conditioning on text
also you should use regularization datasets, they will ALWAYS be relevant until an entirely different way to train models is created
>>
>>109325295
you need to give me money if you want more help
>>
>>109325309
Furkan?
>>
File: 340755983739186.png (327 KB, 640x480)
327 KB PNG
>>
>>109325313
now that's a name I hadn't seen in a long time
checked his reddit account and it seems the grift is expanding into new content
>>
it needs to be sticked that you caption things YOU DO NOT WANT BAKED into the character. Yes, it seems confusing because you think of captioning as trying to describe the image, but that's not the case.

If you describe every single image, you will need to retype that out every single time you want to use the lora.
>>
>>109322766
This will never happen. People that post images of overcooked loras they've created are usually only farming for praise.
>>
File: 00016-2835371192.jpg (813 KB, 1856x2880)
813 KB JPG
>>109325262
I'm tired of arguing with shit you. The more complex the character design is, the harder and more precise the captions have to be. someone like as simple as lunafreya is a piece of cake but someone like eve, lily and raven are going to be an absolute pain in the ass to caption properly because of how unusual and overly complex their character designs are. remember krea2 is 12B model not some 40B+ sorta model. You can't bullshit your way through the captioning process.
>>
>>109325353
the problem is nobody criticizes them really. they just dont use the lora and the creator has no idea, so they just keep pumping out more slop. with no global standards in place, this will keep happening.
>>
File: 442865147995474.png (381 KB, 480x640)
381 KB PNG
>>
>>109325344
this is stupid. thats not how lora works.
>>
>>109325361
You clearly have no idea what you're doing. People have already explained why it's bad to be overly verbose. You just describe her outfit in a consistent way across your captions and it'll be perfectly replicable
>>109325388
Yes it is
>>
File: 99.jpg (378 KB, 772x767)
378 KB JPG
>>109325388
>>
>clanker
>>
>>109325388
that is how loras work.
you have a female character that always has blond hair, a pink streak and a cowboy hat. you never caption her hair, the streak or the cowboy hat.
>>
File: 501035864721779.png (286 KB, 640x480)
286 KB PNG
>>
>>109325405
Can you ask it how to caption for a style LoRA please?
>>
>>109325423
Why can't you do it?? I'm tired of wasting my precious tokens on /ldg/ misinformation.
>>
>discussions about the basics of captioning like its 2023
did i step through a time portal or what?
>>
>>109325411
You should 100% caption the hat. Hair color can easily be changed when using the lora but hard baking the hat into the trigger can be a bigger issue
>>
>>109325423
trigger word then everything in the image that isn't a part of the style.
drippy watercolor, don't mention drippy watercolor.
>drippy watercolor painting of a woman in a boat
style lora would be
>dr1p, a woman in a boat
>>
>>109325437
>2023
newfag
>>
>>109325388
it's at least partly how it works IF the base model/lora understood properly.

if you caption a fedora hat that is something a person or character is wearing it, if the training "understood" it will be likely to omit it during inference if you don't mention it.

if it is not captioned it is indeed more likely that it will generally show up unprompted too

but it still depends on the model and training
>>
File: image.png (60 KB, 845x907)
60 KB PNG
>>109325432
Ok? it's literally free
>>
>>109325442
how?
>>
File: 1776139377580082.png (522 KB, 541x1440)
522 KB PNG
>>109325405
Claude is usually wrong about things a majority of the time - and he is wrong here.

AI in 2026 is actually pretty fucking bad; so don't use them to try and "gotcha" people, otherwise you'll look stupid as fuck.
>>
>grok (dumbest saas LM)
>>
>Retard acts like an authority on loras after training 2 bad barely functional loras
>Gets BTFO
>>
>>109325463
Here's your (You)
>>
>things that didn't happen
>>
>>109325271
ok malcom
>>
>>109325458
what point do you think this is making?
>>
File: file.png (3.02 MB, 1920x1200)
3.02 MB PNG
>>
excellent thread. this is what i like to see. actual discussion relating to using local models, not 1girl spam and ani drama

good job brothers
>>
>>109325474
>3 focal points
slop
>>
>>109325462
https://huggingface.co/CompVis/stable-diffusion-v-1-4-original/commit/6647636e3f542cc0b06fb71d1aff7357ac45d303
>commited on Aug 20, 2022
>>
>>109325489
>>109325489
>>
>>109325361
if you want to be able to individually prompt every accessory it's an arsepain

else it's not particularly harder. we had complex outfits (but not promptable to every individual accessories) even on SDXL tunes
>>
File: 1781173367110357.png (67 KB, 852x929)
67 KB PNG
>>109325472
You don't know how to read?
It's telling us to not describe the style when making a style LoRA. Similarly and contrarily we shouldn't caption the character's traits when doing a character LoRA.
>>
>>109325499
* basically you get eve, [outfit1|outfit2|outfit3] and it'll give you eve and the outfit.

depending on the model and training it's more semantically close or exactly looking like trained
>>
>>109325521
anon you need to work on your reading comprehension
>>
File: 00007-1919225412.jpg (351 KB, 2816x1664)
351 KB JPG
>>109325458
this is bullshit advice. you caption what you see in the image, how is its represented and whether you want it to be replicated. the purpose of the lora is to replicate the character or concept you desire with accurate details that match the detail of the og representation of the character or concept to various degrees that fits your needs.
>>
>>109325545
That's wrong and your bad loras are proof of that. The dataset being trained is the same but you're simply making it harder to replicate your character with your overly verbose captions and thus a worse lora
>>
File: 00065-45.png (3 MB, 1728x1344)
3 MB PNG
Any music anon around? What is the best voice clone for song covers right now? I am still using UVC + Mangio Crepe circa 2023.

>>>/wsg/6199339
>>
>>109322120
is that Resident Evil
>>
>>109325925
yes
>>
>>109321694
>majority of the based character lora makers of the sd 1.5 and sdxl days won't be coming back to making krea2 loras

Because they're busy updating LoRAs for Anima



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.