[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: collage_1784266112_1.jpg (3.4 MB, 6287x4689)
3.4 MB JPG
Discussion and Development of Local Image, Video, and Music Models

Previous: >>109291542

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Z
https://huggingface.co/Tongyi-MAI/Z-Image

>Qwen
https://huggingface.co/collections/Qwen/qwen-image

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>LTX-2.3
https://huggingface.co/collections/Lightricks/ltx-23

>Wan
https://github.com/Wan-Video/Wan2.2

>Chroma
https://huggingface.co/lodestones/Chroma1-Base
https://rentry.org/mvu52t46

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Blessed thread of friendship
>>
>mfw Resource news

07/16/2026

>Reflecting Process Expertise in Procedural Material Generation
https://materialapprentice.github.io

>PromptForge LD: Shot-writer for LTX Video in ComfyUI using llama + lm studio + ollama
https://github.com/Brojakhoeman/Prompt-Forge-LD

>Qwen3-VL-4B-Instruct Heretic (ComfyUI)
https://huggingface.co/DreamFast/Qwen3-VL-4b-Heretic-ComfyUI

>The AI Backlash Has Tech Executives Fearing for Their Lives
https://www.wsj.com/us-news/the-ai-backlash-has-tech-executives-fearing-for-their-lives-30c43972

>George Lucas likens AI sceptics to luddites clinging to horses and carts
https://www.theguardian.com/film/2026/jul/15/george-lucas-likens-ai-sceptics-to-luddites-clinging-to-horses-and-carts

07/15/2026

>PiD v1.5 Checkpoint
https://research.nvidia.com/labs/sil/projects/pid/comparison.html#qwenimage

>RFMSR: Residual Flow Matching for Image Super-Resolution
https://github.com/Faze-Hsw/RFMSR

>Contrastive-Augmented Flow Matching for Style-Content Disentanglement
https://github.com/CompVis/SCFlow/tree/main#-catfm-follow-up

>Let RGB Be the Language of Vision
https://github.com/yangtiming/RINO

>SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning
https://spherelab.ai/symbomni

>Hack Reveals Suno AI Music Generator Scraped YouTube, Deezer, and Genius
https://www.404media.co/hack-reveals-suno-ai-music-generator-scraped-youtube-deezer-and-genius

>Wan Dancer GGUFs
https://huggingface.co/realrebelai/Wan_Dancer_GGUFs/tree/main

>Krea2 Trainer
https://github.com/hinablue/krea2-trainer

>Krea 2 Identity Edit
https://huggingface.co/conradlocke/krea2-identity-edit

07/14/2026

>Bonsai 27B: The First 27B-Class Model to Run on a Phone
https://prismml.com/news/bonsai-27b

>DynEval: Holistic Evaluations of T2I Generative Models in the Wild
https://vcl-iisc.github.io/dyneval

>Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation
https://huangrh99.github.io/SpectraReward
>>
>mfw Research news

07/16/2026

>Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation
https://arxiv.org/abs/2607.13125

>VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
https://zhxie0117.github.io/VideoRAE

>VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation
https://arxiv.org/abs/2607.13527

>From Pixels to States: Rethinking Interactive World Models as Game Engines
https://arxiv.org/abs/2607.14076

>Nexus: Native Mesh Generation with Diffusion
https://arxiv.org/abs/2607.13563

>MultiAnimate: A Unified Framework for Controllable Multi-Character Animation
https://arxiv.org/abs/2607.13415

>DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation
https://arxiv.org/abs/2607.13365

>Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment
https://arxiv.org/abs/2607.13941

>Fine-grained CLIP fine-tuning with self-annotated region alignment
https://arxiv.org/abs/2607.13661

>Bring Music The Horizon: Music-Driven 360$^ $ Video Generation
https://etoile-et-toi-mp3.github.io/BMTH_Project_Page

>AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization
https://arxiv.org/abs/2607.13805

>DNA: Dual-stage Native Attribution for Generated Image Source Tracing
https://arxiv.org/abs/2607.13685

>Attention-Free and Lightweight Token Reduction for Efficient Vision-Language Models
https://arxiv.org/abs/2607.13500

>Music-to-Dance Generation via Atomic Movements
https://arxiv.org/abs/2607.13978

>LPM: Industrial-Scale Generative Video Restoration
https://arxiv.org/abs/2607.13460

>Fre-Res: Frequency-Residual Video Token Compression for Efficient Video MLLMs
https://arxiv.org/abs/2605.16366
>>
>>109294549
Which Krea2 model can I use on a 12GB 3060?
Anyone got a link?
>>
gguf?
>>
should i use fp16 accumulation? it gives 20% performance uplift on my hardware
>>
>>109294624
What hardware? iirc with krea when I ran some benchmarks it ended up being slightly faster to have it off when using sageattention+int8
>>
>>109294617
turbo int8 convrot, probably

https://huggingface.co/Comfy-Org/Krea-2/tree/main/diffusion_models
>>
cozy breas
>>
So were anons really arguing with an LLM in the last thread
>>
File: 713286211551052.jpg (679 KB, 1664x2432)
679 KB JPG
>>109294675
Are LLMs not superhuman at arguing?
>>
I don't think it's an LLM anon...
>>
File: ComfyUI_00750_.png (1.55 MB, 1024x1024)
1.55 MB PNG
>>
>>109294675
>So were anons really arguing with an LLM in the last thread
at least it killed the schizobake quicker lol
>>
>>109294659
Sankyu, I seem to be getting an OOM error.
>>
File: vn.png (2.6 MB, 1920x1080)
2.6 MB PNG
>>
>>109294817
let's see.. turbo int8 is already sitting at 13gb and you need the text encoder too, so add another 5. there is probably a way
https://huggingface.co/vantagewithai/Krea-2-Turbo-GGUF/tree/main
maaybe this?
https://huggingface.co/InsecureErasure/Krea2-Turbo-mixed-NVFP4/tree/main
then a quant of the text encoder..
https://huggingface.co/models?other=base_model:quantized:Qwen/Qwen3-VL-4B-Instruct
you either switch back and forth between the model and the text encoder (painful when you change prompt) or you choose a combo of small quants for both the model and text encoder
>>
>>109294817
>>109294911
Is your comfyui on the latest version? I was under the impression the dynamic memory management should handle this so it doesn't OOM. Also depends on how much RAM you have though, I assumed at least 32GB.
>>
File: 1761917742310702.jpg (2.84 MB, 1832x2448)
2.84 MB JPG
>>109294549
>that non AI image of my Kiara in the OP

My Kiara LoRA is still in the oven, why would you do this
>>
File: 00000-1918574083.jpg (2.28 MB, 1344x2240)
2.28 MB JPG
>>109294911
Thanks, I managed to gen something with this model:
https://civitai.red/models/2762717

It only took 8 minutes...
>>
File: 1134703996.jpg (34 KB, 474x461)
34 KB JPG
>she slowly floats in close to the camera. she raises a bright ball of plasma in her hand. the plasma illuminates her
worst mistake of my life
>>
You should rename these threads as Local Unemployment General or Wooden Eye General.
>>
>>109294928
I (person who replied with the links) am on the actual latest comfy build and had to disable that feature via --disable-dynamic-vram after repeatedly getting stuck in an endless "initializing model.." loop (24/32).
but yeah he should be able to get the ball rolling if at least the checkpoint fits into his 12gb of precious vram with enough headroom.
>>109294952
8 minutes can't be right.
you can try this
https://huggingface.co/comfyanonymous/int4_tests/tree/main/split_files/diffusion_models
together with this
https://github.com/viralvfx/ComfyUI-INT4-Fast

>>109294971
lol ?
>>
>>109286223
Anyone know which models and LoRAs were used for Maria in this style?
>>
>>109294997
kek
>>
File: ComfyUI_05106_.png (1.03 MB, 1280x1280)
1.03 MB PNG
>>
File: 616903730220182.jpg (449 KB, 2688x1536)
449 KB JPG
>>109295050
https://civitai.red/models/2761579/mgwr-m87?modelVersionId=3107943
https://civitai.red/models/1888846/big-asships-narrow-waist-or-krea-2-klein-qwen-image?modelVersionId=3071384
https://civitai.red/models/1972981/sex-nudes-other-fun-stuff-snofs?modelVersionId=3072664

Mostly that first one gives it the look.
>>
>>109295097
Anon delivers, thanks.
>>
>>109294997
Tbh local genning is now a bit too expensive for the truly unemployed.
>>
>>109295140
Then you are the blind, wooden eye.
>>
>>109295147
No, anon does have a point tho
>>
>>109295153
Yes, you are also a retard.
>>
>>109295194
That's very rude
>>
>>109295203
Rude but accurate.
>>
holy meltie
>>
>>109295253
When was the last time you posted a gen?
>>
>>109295140
kind of the perfect hobby for a fulltime wagie because you can gen and train while you are at work.
>>
>>109295323
Which of you never do because you're squatting here 24/7 just because of the attention.
>>
>>109295329
everyone here could be an 80h a week wagie that queues before work and shitposts on their pall day, or they are perma neets who gen and post 24/7.
what difference does it make?
>>
>>109295384
*their phone all day
autocorrect kind of fucked me on that one
>>
>>109295384
what do you mean?
>>
>>109295413
didn't mean nothin, i just post for attention like you.
>>
File: krea2lora2character.png (2.34 MB, 1920x1080)
2.34 MB PNG
trained 2 different characters in 1 krea2 lora
>>
File: gah42qedhabjhda.png (167 KB, 606x552)
167 KB PNG
>>
>>
>>109295471
Netflix et al are cancer which have killed real cinema.
>>
File: 1770040599152020.jpg (2.12 MB, 2352x1568)
2.12 MB JPG
>tfw you fucked up the lora
feels bad

In my defense, my AI agent lied several times and lowk was sabotaging me the whole time. Plus literally how do you even describe this shit?
>>
File: meh_00009_.jpg (894 KB, 1536x2048)
894 KB JPG
>>109295529
what happened? that sucks
>>
>>109295529
enamel bull horns, gold tassels, arm sleeves, vinyl thigh high heels, shirt is basically a drop-armhole tanktop with straps.
>>
File: 1779752419008100.jpg (2.65 MB, 2352x1568)
2.65 MB JPG
>>109295565
I usually have minimax first describe a character with very unique outfits, so I can get a clue about what the hell I am actually looking at - with Kiara you can't just put "shorts, crop top" etc.
But my AI agent SWITCHED my openrouter model over to some 8b model and basically described everything like a moron... No big deal right? we'll just switch it over and redo it.

So I do that and everything is going okay - I caption everything with minimax-m3, I go through and I manually fix the shit that sucks, whatever.

So I got the dataset completed, and I have my AI agent just do a once over for me checking for spelling errors and other weirdness - this motherfucker while I was in the bathroom changed ALL the terminology to the old stuff the 8b model described it with.

The only reason I didn't even catch it is because I had been working on this for too long and I was getting annoyed so I just tossed it into runpod only to find out it had fucked up captions too late...

anyway... RIP
>>
>>109295644
Sounds like a configuration issue.
>>
File: 1775164753899600.jpg (1.73 MB, 1248x1824)
1.73 MB JPG
Kreia's big issue is that it's emotionless as fuck. No matter how much do i try tard wrangling the model into showcasing emotions it just seems to be incapable of doing them.
>>
>>109295661
almost like AI is... soulless
>>
File: 1773471023588920.jpg (981 KB, 1248x1824)
981 KB JPG
>>
>>109295097
box?
>>
File: OIP-3230204450.jpg (24 KB, 474x457)
24 KB JPG
>>109295661
>emotionless as fuck
Is this just anime illlustration being bad, or does base have this too? I'm so fucking tired of this shit, another fail base model, there's always something wrong with it, always some quirk that keeps the model from ever being complete. Time passes, new models drop, and it's the exact same story every single time.
HOW THE FUCK CAN A 12B PARAMETER MODEL NOT DO EMOTIONS!?!?! FOR FUCK'S SAKE
LOCAL MODELS ARE CURSED
>>
File: meh_00043_.jpg (825 KB, 1536x2048)
825 KB JPG
>>
Every base model is the same betrayal wearing a different filename.
>>
File: 7777.jpg (229 KB, 2048x1024)
229 KB JPG
>>109295661
You should always run NSFW loras and jailbreaks at some strength even if you think you're doing something innocuous. It's incredible how restricted the model is by default.

Just as an example
>Photograph, nikon d7500. Closeup of a man crying in despair, extreme pain and anguish from loss.
With and without NSFW loras.
>>
>>109295694
You are describing an image to a chinese llm model (qwen) which then gets translated to vector space.
If Gemma 4 was in place of qwen, results would be way more interesting.
I don't understand why everyone is defaulting to a slop trained llm in the first place.
>>
File: 1783724635810161.jpg (3.02 MB, 1656x2216)
3.02 MB JPG
>>109295668
>tfw no matter how hard I try my AI art will never have sovel

Feels bad
>>
I see that Seedance allows you to just load in reference images of the character you are trying to generate without actually going through training. Is it possible with any local model?
>>
Do you need to an high-res/second pass with krea2?
>>
>>109295746
if it couldn't then /r/ would still exist.
>>
>>109295752
It's not that great in my opinion. Gen a large enough image first then use rtx upscale to blow it up
My knowledge may not be current but that's what I do
>>
>>109295742
>the universe
>coom-coded nun
checks out
>>
>>109295731
Eyes closed, mouth open, the gemoetry of pain. No tears or flushed skin or trembling or strained veins on the forehead.
12b model
2026
Local
>>
File: meh_00065_.jpg (950 KB, 1536x2048)
950 KB JPG
>>
File: 1784276840.png (1.16 MB, 888x1184)
1.16 MB PNG
>>
>>109295694
But we're the ones who shilled it, trained LoRAs on it, and said "it has a lot of potential and it's a shame nobody's training on it." We are the first misinformation spreaders, we did this to ourselves.
>>
>>109295742
Slop
>>
File: meh_00069_.jpg (1.07 MB, 1536x2048)
1.07 MB JPG
>>109295788
That's great
>>
>wan
>high noise preview shows exactly what I want
>low noise preview and finished video doesn't have what i want
how do i fix this
>>
>>109295795
Sub 90 IQ poster right there.
>>
File: 1784281815084833.png (1.32 MB, 888x1184)
1.32 MB PNG
>>109295788
>:3
>>
>>109295795
this nigger thinks he's downvoting in plebbit, rofl
>>
File: 00006-1539633508.jpg (1.19 MB, 2048x1216)
1.19 MB JPG
>>109294952
>>109295013
I managed to get it down to 4 mins with a 2-pass, I'm probably still doing something wrong with my settings but I'm happy it shits out an image now.
>>
>>109295806
Low noise is the only one that matters. Wan is even usable with just a low noise pass
>>
File: 1779628730549038.jpg (3.19 MB, 1656x2216)
3.19 MB JPG
>slop
It's okay she saved some for you.
>>
>>109295803
ty. yours look awesome too.
>>
I always preferred gruel
>>
Not to derail, I know this general is for ComfyUI troubleshooting general more than gen quality, but those calves look unrealistically massive.
>>
>>109295879
thanks bro
i've been working out
>>
>>109295841
Is that Anima? Looks good.
>>
File: emotionlesskrea2.jpg (904 KB, 2560x3200)
904 KB JPG
>>109295661
<lora:fedor_bypass:5>
>>
>Sagging breasts is the 4th most popular Anima lora
>>
>>109295908
that's not what he meant
>>
hi sorry for noob question, pretty new to this:

let's say i have an album of 10 pictures of the same subject, but they all vary *slightly* in style, is there an easy-ish way to have comfyui iterate a reference image over all of them in one workflow?
>>
>>109295908
A pointless comparison if you don't say what the expressions are supposed to be
>>
>>109295914
what did he meant?
>>
>>109295919
Image edit workflow allows you to pipe in multiple references. I don't know about Krea2 but Klein9B is great.
Usually just one image is enough to reference a single subject but whatever.
>>
>>109295919
https://github.com/remingtonspaz/ComfyUI-ReferenceChain
>>
>>109295935
>>109295938
thank you anons
>>
>>109295956
aiieee this is a sfw board you can't post violence anon you'll get banned!
>>
>>109295908
A pointless comparison because the person in each example is slightly different.
>>109295924
Probably destroys prompt adherence
>>
>>109295964
That's borderline guro. Perhaps read the rules. Oh wait, you are illiterate. Sorry.
>>
File: 1754954916036610.jpg (755 KB, 768x1280)
755 KB JPG
>>109295908
Can you do this in Comfy with Krea?
Their site lets you adjust a bunch of parameters and add mood boards.
>>
File: 1770548821355558.png (25 KB, 961x481)
25 KB PNG
>>109295979
Wrong image.
>>
/gdg/ - Guro Diffusion General when?
>>
>>109295979
Slop
>>
File: 1768342437123066.jpg (1.55 MB, 2560x1440)
1.55 MB JPG
>>109295905
I don't think I'll ever use anima again
>>
Is there a list of AMD GPU that can or can't do AI?
9070 xt btw, or if one can do it every 9070 xt can?
dunno how it works, im new to local, and im choosing gou for my rig
>>
>>109295998
better than the stuff you post
>>
>>109296000
>im choosing gpu
if you hasn't bought an amd gpu yet, just buy a nvidia gpu
>>
>>109295999
Why? :(
Tdrusell desu
>>
>>109296000
Save yourself a headache and buy Nvidia
>>
File: 122170228971507.png (1.91 MB, 1152x1472)
1.91 MB PNG
>>
>>109296007
>>109296012
have you seen the prices?
>>
File: amazon.png (892 KB, 1024x1024)
892 KB PNG
>>109295970
strange I've posted multiple throat slash/mutilated corpse gens before and never got a warning oh well
>>
>>109296022
well buy an amd gpu then
>>
>>109296022
Yes but if you're seriously interested in local AI you should pay the extra couple hundred dollars more.
>>
>his model use anything other than euler and simple
>his model needs loras
>his model needs especially made custom nodes to diffuse properly
lol
lmao
>>
>>109296022
>>109296038
Don't listen to this guy, he owns NVIDIA stock. (he's right though)
>>
>>109296022
Please buy GGUF coins sirs!
>>
File: motorcycle.png (1.66 MB, 1024x1024)
1.66 MB PNG
>>
>>109296068
sks- stiff krea slop
>>
>>109296086
If you look, the motorcycle is parked on its kickstand, it's obvious it's going to be stiff.
>tires in motion
Oh no, the anon is a low effort slopper.
>>
i just can't take slopanon seriously after seeing his gen.
>>
>>109294624
>should i use fp16 accumulation?
Yes. Fast fp16 accumulation just use fp16 for the accumulation step instead of fp32. Fp32 precision in these calculations does not matter. This is not theoretical. There are a couple of links in the ldg lazy getting started guide called "Encoder Quant Rundown" or something that gets into this more
>>
>>109296086
Samozaryadny karabin Simonova*
>>
File: 1763271012200017.jpg (2.15 MB, 2560x1440)
2.15 MB JPG
>>109296011
because it can't do stuff like this.
>>
File: ComfyUI_00333_.png (960 KB, 1024x1024)
960 KB PNG
>>
>>109296000
>Is there a list of AMD GPU that can or can't do AI?
You want to ask an AI for which AMD GPUs support the latest ROCm. I wouldn't go older than the 7900 XTX

Your only three options are 9070 XT, 7900 XTX and Radeon Pro 32gb whatever the model is (navi48)

I would not recommend buying a 16GB card if you want to do LTX video. I would not recommend an AMD card at all to do video generation. If you are fine with doing everything other than LTX video, 16GB is enough

You will not be able to use int8 convrot models because int8 is slower than fp16 on RDNA. You will be a GGUF and fp16 and even fp32 merchant.

What do you want to make? Like actually?
>>
File: ComfyUI_00334_.png (1.83 MB, 1024x1024)
1.83 MB PNG
>>
File: 533370187938312.png (1.99 MB, 1472x1152)
1.99 MB PNG
>>
>>109296197
>I would not recommend buying a 16GB card if you want to do LTX video
Is this just a AMD thing? LTX distilled fp8 worked fine for me on 16GB vram 4070ti super
>>
>>109296184
Krea?
>>
>>109296224
Anima with FF lora
>>
File: krea2lorafreefuse.png (805 KB, 1024x1024)
805 KB PNG
used this https://github.com/yaoliliu/FreeFuse with krea 2 and 2 loras
>>
File: 1772523014914926.jpg (2.72 MB, 2560x1440)
2.72 MB JPG
>>109296224
sdxl actually
>>
>>109296215
You end up needing more VRAM on AMD all the time in practice for most models. When I was using the 32gb Radeon pro it felt like it was actually a 26gb card


Oh did I mention the crashing? Even if you do everything right sometimes the driver or hardware will just shit itself. I would never train a model using AMD consumer hardware (their datacenter stuff is fine)
>>
File: freefusekrea2.png (1.09 MB, 1024x1024)
1.09 MB PNG
>>
File: 378192830834006.png (2.08 MB, 1472x1152)
2.08 MB PNG
>>
Well I pretty much only gen futanari now so I can't post anything.
>>
>>109296307
Bulges exists
>>
Where is Anima int8?
>>
>>109295979
Crap
>>
>>109296307
same, muscle futas with chastity cages is all I gen anymore
>>
>>109296325
>>109296313
And yogapants or why not corn syrsup leggings? Futa feet are erotic
>>
>>109295908
ConditioningKrea2Rebalance node provides more granular adjustments.
>>
>>109295908
1girl and bypass lora is an oxymoron.
>>
the AIbooru has to be the biggest slop factory on the net
>>
File: 981581833155865.png (2.77 MB, 1152x1472)
2.77 MB PNG
>>
>>109296413
Better than danbooru desu
>>
>>109296507
Unironically this. The sheer amount of amateur drawfags uploading to that site every day is insane. At least AI sloppers put more effort into uploading to the AI booru than the pencil sloppers who think they're Akira Toriyama after drawing for a month
>>
DanBooru is the CivitAI of drawfags
>>
>>109296215
>LTX distilled fp8 worked fine for me on 16GB vram 4070ti super
depends on system RAM.
how much system RAM do you have?
>>
>>109296533
64
>>
File: hdusia23hgs3.png (360 KB, 2016x1367)
360 KB PNG
>>109296523
>>109296530
>>
/r/ing' sfw coomer gens
>>
>>109296571
That's the point. It's not only depended on your card.
I have the same card but only 16GB system RAM. So it doesn't work on my computer.
>>
I just use my iphone 2 gen idk why yall need nerdstations for cooming
>>
>>109296603
>Pick the pencil up...now!
>entire image was made in 5 seconds in mspaint with 0 pencils involved
>>
>>109296656
>16GB system RAM
That's rough
>>
>>109296672
Gotta use it for something.
>>
File: 600587352871406.png (2.51 MB, 1024x1600)
2.51 MB PNG
>>
>not generating thoughts2vid in your own mind through an hypnagogic model
Ngmi
>>
i cant get z-image to do femboys. as soon as i start getting it to look feminine it adds breasts
>>
>>109296937
2 more weeks
>>
>>109296779
yup. I paid $360 for it too. Jews man!
>>
>>109296941
Another W for ZITGODS dunking on faggots.
>>
File: mangaloratest.png (1.84 MB, 1280x1280)
1.84 MB PNG
>>
File: 51632512561232563.jpg (1.45 MB, 1238x2200)
1.45 MB JPG
>>
File: 1688740382.jpg (11 KB, 155x323)
11 KB JPG
>>109294549
can I make porn of this character
it must be on model
I'm serious
>>
>>109297208
that is an animal
>>
File: Ideogram__01210_.jpg (1017 KB, 1536x2048)
1017 KB JPG
>>
File: it123_V-0.93hires2_00001_.png (2.93 MB, 1408x1920)
2.93 MB PNG
Is there a node which will save images with a higher level of png compression? Comfy keeps saving these >4MB images. A quick resave in Krita brings them below 4MB. Krita is doing something to get the PNG compression more efficient, and I'd like comfy to do this as well.
>>
>>109297265
that is a mythological creature
>>
File: it16_V-0.93hires2_00001_.png (2.75 MB, 1536x1920)
2.75 MB PNG
>>109297282
He's my wow character, a dracthyr mage.
>>
File: Capture.png (61 KB, 994x585)
61 KB PNG
>>109297265
"ComfyUI natively saves PNG images with a fixed compression level of 4, as higher levels (5-9) provide negligible file size benefits while drastically increasing save times. To achieve different PNG compression levels, you must install a custom node or modify the core source code.

Custom Node Solutions
ComfyUI-Image-Compressor: A custom node that allows you to set the compression_level parameter for PNGs, ranging from 0 (no compression) to 9 (maximum compression). This node also supports batch processing and format conversion.
Save Image Extended: While primarily focused on formats like JPEG and WebP, this node offers advanced metadata and folder organization but does not explicitly override the native PNG compression level in the same granular way as the compressor nodes.
Core Code Modification
If you prefer not to install custom nodes, you can change the PNG compression level by editing ComfyUI's server.py file. Locate the line:

original_pil.save(filepath, compress_level=4, pnginfo=metadata)

Replace 4 with your desired level (0-9). For example, changing it to 9 will maximize compression but significantly slow down the save process. Alternatively, use the fpnge library for faster PNG compression."
>>
>>109297291
that is a dragon from european folklore
>>
File: ComfyUI_temp_drxna_00013_.png (1.94 MB, 1024x1024)
1.94 MB PNG
>>
>>109297301
Wow, thanks!
>>
>>109297301
Comfyui can't save images async? What the actual fuck
>>
>>
>>109295507
They didn't do anything though? They just followed the market demands.
>>
File: ComfyUI_temp_drxna_00030_.png (2.24 MB, 1024x1024)
2.24 MB PNG
>>
File: ComfyUI_temp_drxna_00033_.png (2.26 MB, 1024x1024)
2.26 MB PNG
>>
Losslessly compress your png images to save ~10% valuable space so you can gen more kino

https://github.com/oxipng/oxipng
>>
>>109297776
just use webp
>>
>>109297796
just save the workflow and not the image themselves
>>
File: image.png (127 KB, 1175x254)
127 KB PNG
>>109296000
>>109296022
saar
>>
>>109297776
lossless jxl would save ~40% btw
>>
File: 1753932994880551.jpg (897 KB, 768x1280)
897 KB JPG
>>109297776
I use webpee.
>>
>>109297265
consider saving in webp
>>
gen my cock bigger pls
>>
>>109297674
You don't understand at all. I didn't expect you to do so anyway. Proves the fact that most normies are blind and stupid.
>>
>>109297835
And preserve embedded wf?
>>
File: 2544536106749.png (2.66 MB, 1024x1600)
2.66 MB PNG
>>
>>109297939
>Proves the fact that most normies are blind and stupid.
doesn't that just prove his point about them following the market?
>>
>>109297939
>companies follow the market
>???
>yeah actually you don't understand it's actually *schizo babble*
>>
>>109297944
yes

there's the caveat that comfyui doesn't parse it yet so you need to copy/paste it as text
>>
krea 2 DESPERATELY needs a finetune to teach it how to generate skin
>>
>>109297939
muh free markut is an endemic belief. everyone is trained to think any economic event is purely rational and good. it's the foundation of our system. never mind that big business never uses free market systems within the business
>>
>>109297980
vae makes it impossible
>>
>>109297843
I would, but my goal is to post on 4chan with minimal friction.
>>
>>109297989
cant you force it to learn another vae?
>>
Why does this dude always try to derail the thread with off-topic
>>
>>109297980
what is wrong with the skin it generates?
>>
I love smooth skin haha I like running my fingers over beautiful smooth pale skin i want to wear smooth pale skin
>>
>>109297994
please explain
>>
>>109298019
could you post an example of a "good skin" gen?
this is an imageboard, you can post pictures instead of mindless fuddy vagueposting.
>>
>>109297980
skill issue
>>
>>109298039
look in the mirror
>>
>>109298044
if you require 3 differeent loras, 4 custom nodes, a weird jeeted workflow, a very specific prompt, and roll 10+ seeds to find a gen with good skin, you desperately need a finetune.
>>
>>109298052
done, i'm looking quite handsome today i must say. summer tan does wonders to my aryan brahmin skin.
now could you post something you think is a "good skin" gen.
>>
>>109298069
Good water is easier to gen. Great water isn't.
>>
File: 1772849711699020.png (2.48 MB, 1152x1600)
2.48 MB PNG
>>109298039
could you do this with krea?
>>
>>109298082
profound. oh wait no I mean retarded.
>>
>>109298086
why couldn't you?
>>
tfw anon is unable to objectively measure the "realness" of skin texture
>>
>>109296184
extremely based gen and style.
>>
File: 1757163165732247.png (2.66 MB, 1152x1600)
2.66 MB PNG
>>109298086
best attempt after 10 gens on krea 2
>>
>>109298086
A gen I already had. the noise can be reduced with better sampler settings.
>>
>109298171
BWAAAAAAAAHAHAHA
>>
File: 24.png (1.66 MB, 1456x1120)
1.66 MB PNG
schizo thread
>>
>tfw there is still no Ken Sugimori LoRA that can do his old style justice
I needs to have the grain, the washed out colors, the watercolor fade.

>>109297265
WAS Node Suite > Image save node
>>
>>109298086
my best attempt
>>
very straight thread, /ldg/
>>
File: ComfyUI_temp_drxna_00062_.png (1.85 MB, 1024x1024)
1.85 MB PNG
>>
generic eastern european teen egirl something to straighten up the threat
unrelated:
https://unsplash.com/s/photos/portrait-woman
half decent ressource, also nice for a quick reality check
>>
>>109297039
Tse would never do that
>>
please stop posting homosexual tummy i am at work
>>
File: Ideogram__01230_.png (1.57 MB, 1152x1600)
1.57 MB PNG
>>109298086
'deo'ram4
>>
>>109298250
All you need is 50 images my guy and you've got yourself a style lora.
>>
>>109298363
I don't have buzz.
>>
>>109298366
boobies!
>>
>>109298366
just lost my job in the microslop support, thanks sir.
>>
meant to post the catbox:
https://files.catbox.moe/3rl0ux.jpg
>>
>male bodies stay
>female stuff deleted
HMMMMM
>>
>>109298373
vagina too.
>>
>pull
>brick comfy
>can't participate in tummy gen
>miss accidental booba post while alt-tabbed
fml
>>
>>109298383
how many times do I have to tell you jannies are trannies and pedophiles?
>>109281620
>>
>>109298383
>muscle mommies are gay
ok.
>>
why don't you just use catbox
>>
im so lonely bwos..... :(
>>
>>109297956
jfc
do you think this is a good image?
>>
can i have some cute anime muscle mommies?
>>
>>109298393
>>
>>109298431
I like it
>>
File: anistudio will save local.jpg (254 KB, 1868x1020)
254 KB JPG
>>
>>109298435
it is absolute shit
>>
>>109298431
Let's see your good gens anon.
>>
>>109298436
Instead of python, they should use Scala!
>>
>>109298436
we get it, catjack
>>
>>109298436
Please god no.... don't tell me ani lives quebec..... oh my god....
>>
>>109298446
my fart looks better than that 70 nodes gen
>>
>>109298436
>he fell for my troll
hahaha what a lolcow
hope he had high hopes lmao
>>
>he still tries to shit up the blessed bread
no one is coming back to /s*g/ thread schizo
>>
>>109298459
post one of your gens
>>
File: 246887977807951.jpg (822 KB, 1792x2304)
822 KB JPG
>>109298431
I think so.
>>
I'm trying to use Klein 9b and Lustify for inpaint in Comfyui but getting really shit results. Not sure what I'm doing wrong, I'm a retard admittedly. What are best models currently for inpainting on realistic images and anyone have sample workflows they could direct me to?
>>
>>109298436
tag explorer anon what is your next project?
>>
>Posted 1 day ago
>55/36
rough
>>
>>109298495
meant for >>109298461
what is your next project tagexplorer-san
>>
I'm new to this. Got a dumb question. What's the difference between the official repos and Comfy's repos? For example,
Official: https://huggingface.co/ideogram-ai/ideogram-4-fp8/tree/main/transformer
Comfy: https://huggingface.co/Comfy-Org/Ideogram-4/tree/main/diffusion_models
The official file is fp8 so it should be identical to Comfy's fp8 but they're different file sizes so something's changed? Is official or Comfy's version generally preferred?
>>
>>109298513
the comfy repos usually try to put all the files you need in one place. I usually prioritize grabbing from the comfy repos.
>>
>>109298513
>or Comfy's version generally preferred?
Preferred in that I'm not gunna potentially spend time getting the official version to work when Comfy loves to force the users hand into downloading their repackaged versions.
>>
>>109298513
they basically apply a small lora to all models to make sure it can generate the comfyui logo on inference.
>>
>>109298504
FWIW you're replying to schizoanon (I believe at least) once again impersonating me for some reason, anyway I don't have much planned, I'm working on a large AI character card project but that's not really of interest to this general.
>>
>>109298539
didn't ask
>>
>>109298526
>>109298526
>>109298534
So use Comfy repo for least friction, got it, thanks. But do you know if Comfy makes any changes to the models they repackage? I couldn't find anything on their methodology; is it normal that the file's been changed.
With Ideogram in particular, the source is only fp8 so I'm worried Comfy upcast it to bf16 to make all their quants, then uploaded their own requanted fp8 with additional rounding errors instead of the original.
>>
>>109298391
>vagina
who gives a shit, are u gay?
>>
yall got any selfsuck gens
>>
>>109298587
I wouldn't trust lab provided quants over comfy just from the fact that comfy needs to quant every new model that comes out where the lab only ever has to quant their own stuff which they often don't even bother to do.
>>
File: 1-2.png (606 KB, 2301x1047)
606 KB PNG
>>109298489
you can use any model for inpainting. I use lustify to inpaint lustify gen faces / tits / etc and flux klein9b for hands and feet and clothing and other details, like necklaces and so on. can you rebuild these? left side is a generic flux2 klein inpainting workflow. a load image node on the left, manual masking, adjust denoise to taste.
right side is a lustify face inpaint workflow with the detection part of the facedetailer node outsourced and an added controlnet.
both use a similar pipeline: inpaint crop and stitch (there is a tutorial video on youtube), differential diffusion, inpaint model conditioning
>>
>>109298616
nah
>>
Now that the dust has settled how do we feel about SDXL?
>>
>>109298625
How do we know what else he shoved into it? I don't actually trust safetensors preventing pickled binaries
>>
>>109298635
the downfall of optimism.
>>
>>109298635
SD 1.5 had soul
>>
>>109298436
based ani
>>
>>109298635
sdxl at the very least did not suffer from catastrophic forgetting like anima
>>
File: ComfyUI_temp_euzig_00070_.png (1.84 MB, 1024x1024)
1.84 MB PNG
>>
>>109298637
>he
>I don't actually trust safetensors preventing pickled binaries
Take that tinfoil hat off. comfy isn't a single person anymore.
>>
>>109298628
>local
>>
>>109298660
go compare artist knowledge on WAI v17 vs anima aesthetic 1.1, and tell me which one catastrophically forgot more
>>
>>109298628
Damn ok that image made me realise I'm way out of my depth here, but thank you anon. I'll save that and try study up on it. Appreciate the detailed explanation.
>>
>>109298678
and how did that work out? an open source project shouldn't be entirely dependent on cloud providers and API wrapper cancer
>>
>>109297975
which program ur using to convert the png to jxl while keeping all metadata?
>>
>>109298689
>entirely dependent on cloud providers and API wrapper cancer
what are you on about?
>>
>>109298689
>he's too far gone.
>>
>>109298380
>metadatalet
Ugh, why even catbox.
>>
>>109298685
hes b8ing you that post was b8
>>
>>109298400
saved.
>>
baiting anon into a sexual relationship
>>
>>109298715
too much compromising data there.
>>
>>109298687
just use the 'inpaint model conditioning' node (allows you to use a denoise value lower than 1) and copy the settings for the sampler/scheduler, etc, that should give you some sort of baseline
>>109298681
very helpful
>>
People have been a lot more reluctant to catbox every since the *incident*
>>
>>109298743
I don't catbox in the first place
I raw post and take the ban on right on my balls
>>
>>109298743
what are you babbling about? not everyone is an autism who knows the internets everything. info drop and dont be vague
>>
>>109298761
based.
>>
>>109298726
What data could be compromising? Lora names?
>>
What's the best way to ensure consistent clothing colors? If I'm prompting for white, It'd be a real shame if I got blue accents constantly for some reason...
>>
>>109298787
I'm not going to tell you.
>>
>>109298789
model? I rarely need precise colors but one time I did when using klein and I found using hex color codes gave pretty good results. I imagine it would work with krea, likely better than it did with klein.
>>
>>109298743
only the weakwilled do not catbox
>>
heh
>>
>>109298784
Julien the lolcow doxxed himself for updoots
>>
>>109298836
You turned her into a shiny flux face buttchin girl
>>
>>109298836
>slop
>SLOP
>S L O P M A X X E D
>>
Fresh when ready

>>109298857
>>109298857
>>109298857
>>
File: images.jpg (20 KB, 447x447)
20 KB JPG
>>109298836
cum right? because is white, yes, is cum she is crying coom,
>>
>>109298868
the prompt is there.
>>
>>109298836
stop using the badly captioned realism engine lora, shit is cooked
>>
>>109298883
nta but I prefer Realism Engine's bodies to other loras
how do you know it wasn't captioned right?
>>
>>109298883
idk I found v3 can be good. but it does turn tears into cum.
>>
>>109298883
no loras were used.
>>
>>109298897
he doesn't know
>>
>>109298483
now this a good one
>>
>>109298691
save-image-extended-comfyui does it

i saw there's also a fork that can load images https://github.com/koloved/save-image-extended-comfyui but the brotli compressed metadata won't be supported by too many other tools (including your image viewer)
>>
>>109294659
update comfyui and disable any memory startup flags such as --lowvram as that messes with the dynamic vram, and it absolutely will work.
>>
File: janet.jpg (95 KB, 1024x1280)
95 KB JPG
>>109295965
haven't noticed an issue with prompt adherance yet but sometimes get nasty artifacts:
https://files.catbox.moe/oqn8z0.png
>>109296372
yeah, it's pretty cool, i just hate comfyui



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.