Weekend EditionDiscussion and Development of Local Image, Video, and Music ModelsPrevious: >>109730780https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
>>109736456
as always, nothing happens.a week just like any other.
how do you feel about being in pic related, anon?
>>109736499hell yeah>-AR 1:2GET THAT CLOUD SHIT OUT OF HERE
>>109736499this is just like me frfr!
>>109736499Very funny. Raffed hard.
What should I prompt?
>>109736499>>109736554kinos
>>109736559You should make kinos instead of being the corner cuck. I need you to lock in anon
>>109736565but I'm currently working on figuring out the best possible compromise between speed and videoquality anon, my gpu is busy
>>109736579No excuses, one does not have the mandate to call for kino without making kino.
>>109736588no yuo kino!
>>109736585Why do you post this when anons regularly post local gens that clear you?
>>109736601Grandiose delusional thoughts.
>>109736601>a strong workflow more than makes up for his shortcomings!
>>109736601He does it to piss you off
>>109736449>mfw Resource news09/05/2026>Intern Lumina U2: Multi-Codebook Diffusion Large Language Model for Omni-Visual Understanding and Image Generationhttps://internlm.github.io/InternLumina-U2>Musk’s xAI loses court bid to block Minnesota's AI ‘nudification’ banhttps://www.reuters.com/legal/litigation/musks-xai-loses-court-bid-block-minnesotas-ai-nudification-ban-2026-09-04>Add Sparse Attention node- #16072https://github.com/Comfy-Org/ComfyUI/pull/16072>MiniMax-H3 FL2V 8-Step Motion Enhancerhttps://huggingface.co/rzgar/minimax-h3_fl2v_8Step_motion_enhancer>Qwen3.8-Flash-Next-NVFP4 https://huggingface.co/nvidia/Qwen3.8-Flash-Next-NVFP409/04/2026>lightx2v/Minimax-h3-Turbo · FL2V Turbo 4-step v1.2 (768p)https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/52#6a9a890895a616c64799324f>ComfyUI NVIDIA DLSS 5 Visual Enhancerhttps://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-Interpolation>Viggle-Animate: Character Replacement in Video from a Single Repainted Frame https://huggingface.co/Viggle/Viggle-Animate>DSAQuant: Denoising-Stage-Aligned Quantization-Aware Training for Video Generationhttps://robbyant-research.github.io/DSAQuant>Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuationhttps://github.com/AMAP-ML/StateAgent>FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlowhttps://byeongjun-park.github.io/FlashRender>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipeshttps://huggingface.co/inclusionAI/LLaDA-Image>ComfyUI-VDN-H3: v1.4.0 — Faster streaming, VRAM-aware buffer retention Latesthttps://github.com/Saganaki22/ComfyUI-VDN-H3/releases/tag/v1.4.0>AetherScale for ComfyUI: GPU-native NVIDIA video enhancementhttps://github.com/vizart-vj/ComfyUI-AetherScale
>mfw Research news09/05/2026>Generalization over Memorization: Generalization-Aware Diffusion Adaptation for Single-Image Multi-View Synthesishttps://arxiv.org/abs/2608.29233>Test-Time Scaling for Video Diffusion Models via Diagnosis-Guided Candidate Recyclinghttps://arxiv.org/abs/2608.29322>Streaming4D: Accelerate 4D World Models via Block-wise Video Generation and Incremental Reconstructionhttps://arxiv.org/abs/2609.00610>Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensicshttps://arxiv.org/abs/2609.02268>Visual Framing for News Stance Detection via Image Generationhttps://arxiv.org/abs/2609.00685>Controllable Image Captioning with Prompt-Conditioned Scene Rewardshttps://focus-emnlp2026.github.io>TimeSteer: Inference-Time Speech Scheduling in Joint Audio-Visual Diffusion Modelshttps://arxiv.org/abs/2609.01277>Diffusion Based Unpaired Data Learning for Inverse Problemshttps://arxiv.org/abs/2609.01370>ExpArt-KG: Artwork Image Description Generation through Iterative Exploration of Knowledge Graphshttps://arxiv.org/abs/2609.00629>ViTAL-X: Video-Text Alignment with Cross-Modal Temporal Editshttps://arxiv.org/abs/2609.00505>ASSERT: Adaptive Stochastic Sampling for Robust Diffusion Models on Analog Compute-in-Memory Hardwarehttps://arxiv.org/abs/2609.00955>From Detection to Localization: A Unified Forensics Framework for Fully Synthetic and Tampered Imageshttps://arxiv.org/abs/2609.02640>Sketch2Inspire: Structure-Sensitive Evaluation for Product Retrievalhttps://arxiv.org/abs/2608.29364>A Lagrangian View of Flow Matchinghttps://arxiv.org/abs/2609.00198>Beyond Language Priors: Diagnosing and Fixing Visual-Origin Hallucinations in Multimodal LLMhttps://arxiv.org/abs/2609.00231>SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Modelshttps://arxiv.org/abs/2608.29974
>>109736647>>109736651thank you for the news!
>>109736647>>109736651Fuck off
>>109736670np :)
is the latest 4 step lightx2v minimax or the older 8 step turbo lora better?
>>109736795Is 20 steps raw or 8 step lora better?
>>109736499absolute kino
consume more energy that it would take to sustain all of humanity for decades to run his >datacenter just to analyze every mathematically possible solution to some obscure math riddle from 200 years ago that nobody cares about just to get more of your tax money from the governmentthat's alignment all right
>>109736785looks kinda like my cat, picrel (but mine is chonkier)
>>109736853very high quality loafcats should be allowed in space
>>109736853Talk about the wrong stuff...
>>109736850my computer doesn't use that much. this is the local ai thread and we don't pay anything other than electric bill
>>109736868>cats should be allowed in spacei can guarantee once space travel and colonization is a thing, the first thing people will bring to other planets is their cats, they've won the evolutionary race by befriending the apex predator>>109736886huh?
>>109736647>>109736651>>109736670>>109736785>>109736853>>109736899Real dented zone post rn
If it bleeds..https://streamable.com/uv0z1a
>>109736637SD thread only exist as a coping mechanism and pure spite at this point
>>109736913You do know everyone thinks you are mentally ill because you can't stop acting like a child right?
>>109736943>he's upset
>>109736899It's a reference to the movie Armageddon. It's generally said that astronauts should be made of/have the right stuff.
>>109736954oh sorry lol my dumbass brain didn't get the reference even though i've seen the movie>>109736947honestly i'd rather have a beer with debo even though he's an avatarfag than with you, and that says a lot
I find it odd whenever wheelchair gens get posted he goes right to physical violence and doxxing
>>109736899Local diffusion?
>>109736978I would post some gens but I'm not at my atelier currently.
>>109736995yes. cats are pretty much diffusers, they can diffuse through very small spaces as if by mere entropy
>>109737104How odd you bring the dead discussion you have with yourself in /sdg/ here also if you're not him go talk about that shit in the containment thread seems that they all magically decided to post cat gens and not really talk with each other in /sdg/
>>109737053What does this have to do with catjack having another meltdown?
>>109737204Can you define meltdown because>>109736961Is a actual meltdown
>>109737203i literally don't understand what you are talking about, i'm a fucking tourist who visits these threads once every 3 weeks
cozy breas
t
>>109737261I'd recommend simply ignoring that anon. he's a bad faith schizophrenic drama-baiter
>>109737301Wow they really amped the graphics for Kitten Space Agency.
He's having a rough time today isn't he. I know the goal is to be annoying but imagine getting angry and seething to the point you make post like>>109736961
>>109737320DLSS: on
>>109737323That anon is correct. If we actually knew how pathetic you are irl you'd never come back
>>109737497Doesn't stop those other guys from doing daily antics why do you guys always complain/talk offtopic stuff and never post any gens at least?There's a general for that called /sdg/ is it because nobody actually post there or that even you guys find it dull and unfulfilling?
jealous locusts...
>>109737686People don't want to post gens when you are around
WHEN IS GEMMAPROMPT GETTING UPDATED?
>>109737916It's dead just like coomkit and all projects that got a dozen commits and suddenly stopped. MiniConstruct has a couple more commits into it before suddenly disappears too.
>>109737916because you touch yourself at night
>>109737686/sdg/ got taken over by some meth head a long time ago. he does nothing but post SDXL abominations and butterflies. he has singlehandedly shitted up /sdg/ permanently beyond recognition
>>109737963All these project creators could learn a thing or two from Ani. He has serious dedication and doesn't just abandon his passion project.
>>109737916JUST MAKE YOUR OWN!
>>109737916https://github.com/whp199/GemmaPrompt/tree/master/skills isn't that the essence of it?I think you could run ~that with most ways to run LLM.
>>109737981why are you being such a lazy fuckwit, anon?
>>109737985I made my own I don't rely on gemmaprompt or miniconstruct or whatever else is out there.
>>109737970I think it's the wheelchair dent, he uses 3 personas then a few other anons that are clearly mentally ill reply to him sparsely throughout the day.>>109737908This post is by him too, he keeps dog whistling old post said about him and wonders why it falls on deaf ears. He typically gets drowned out when actual happenings take place.
Fresh video.https://litter.catbox.moe/pw0ltxiibdu98fv1.mp4
>>109738395What the fuck was her problem?
>>109738395kekd and cute butto
>>109738395While I don't exactly get the correlation between mentos and crime, I did enjoy the video and it's definitely high quality. So, very nice gen anon
>ComfyUI-SeedVR2>Unable to allocate 23.7 MiB for an array with shape (1080, 1920, 3) and data type float32Anything I can do to fix this error? 16g VRAM & 64G RAM. Same error if video is 1080p or 720p.
>>109738436>>109738446also very nice gens anon, is that fl2va or ref2va?
>>109738450If you have something stolen from you you can counteract the negative mental effects by popping a mentos and thus the overall happiness in the world increases. However, if you kill your assailant with a rocket launcher it goes down so don't do that.
>>109738477ref2va
>>109738395Best thing I’ve seen on here lately lol
>>109738349sexy dom debos are so in right now
Is there any way to apply a lora to only one character in the picture while other characters are unaffected? im using comfyui with krea 2
>>109738466face aside that's kinda hot
Can video/image gen benefit from multi GPUs apart from generating from both cards at the same time? Like actually accelerating generation of a run?
>>109738395uhhh... kino has arrived ?
>>109738571I was about to answer this question unironically, but then I saw this >>109738590You're getting better attentionseeking boy
>>109738430https://litter.catbox.moe/s07q6y05wk3fhd42.mp4I tried audio ref but adding too many refs causes the prompt to shit itself
>>109738395Heartwarming! Pure kino.
>>109738629from my experience so far H3 has a hard time with audio references unless the audio is being used 1:1 over the entire video length.It does work as reference too, but then it needs to be a good 5sec reference with enough spoken words, but not with too many otherwise it hallucinates spoken words even though you didn't prompt it. Also very sensitive to resolution changes
>>109738395she cuteit would be cooler if the guy exploded in a ball of fire
>>109738682I agree, in an explosion of some kind. But I think those three large gibs saved it. Those were good gibs.
Next MiniConstruct update will be a really fat one.
>>109738395quality entertainment. made me think of those 80s movies that had fun shit like this happen in them
>>109738395how make 30 second video, pls tell sir
>>109738598Take your meds schizo.
>>109738792Then how do you explain this? Also newfags who don't even know how to do something as simple as the question that was asked don't know about schizo lore you schizo
>>109738395>dead link:/
>>109738817>Then how do you explain this?You're fucking crazy that's what
>>109737916this is exactly why I roll my own shit now. if its open source its not a big deal since it's easy for AI to extend it, but if its closed, you're shit out of luck if you want new features or bug fixes
>>109738395re-upload. i want to see why this got so many (You)'s
>>109738395She looks like Brooke Shields when she was young.
I read the guide but maybe I'm just retarded. Is there a containerized/service based front-end / back-end split? On the general LLM side, I use Open WebUI and llama-server. I have the WebUI in a quadlet and the llama-server as a bare metal service. Does SD have that equivalent?
>>109738395pretty nice
>>109738528you really need a job debo
>>109738830Also, you have the (you) on the schizo post that has been posted a 100 times not on the "potential" new guy asking how to do the lora thing, lmao, fuck off schizo faggot
>>109738839>Brooke Shieldstime is so cruel
>>109738857Indeed.
>>109738856>fuck off schizo faggotI called you out on being schizo which you are. It's often the crazy people who deflect back whatever they're being called for.
>>109738725
>>109738840not typically, it's usually just a web service thing without a split into backend1, backend2, middleware1-4, front end or whatever enterprise thing you could do.well yes the front end-client is maybe your browser if you want to see it that way
>>109738901I respect the commitment to the bit schizo, but you lost when you posted an image of yourself claiming the schizo post that has been posted a trillion times already, instead of the believable post, please try something original... maybe a wheelchair gen?
>>109738880The H3 video from the other Anon must have used references from when she was 14 or somewhere around that age.
>>109738880pretty woman indeed, I love her eyes
>>109738743I wouldn't do something like that in one go; that's the sort of video you want to edit in a non-linear video editor from multiple clips and external audio for the music.
Oh no! My beer!
Jack-o-Shoes
Not looking good for Qwen in the H3 prompt writing tests
>>109739304are they really "hard expectations" if they can all be cleared?
>>109739304Are you giving them any prompt writing skills or are you letting them wing it?
>>109738901Really can be seen with both rentry schizos
>>109739336>>109739338I'm implementing a new story planner for creating multi-generation prompt sequences.>Qwen has a systematic Story Build problem>Across the Build cases, Qwen repeatedly refuses to populate blank fields on existing Generations.>In story-basic-four-beat, Generation 1 came back completely blank:>creativeRequest = "">incomingState = "">intendedOutgoingState = "">cameraIntent = "">while Generations 2–4 were quite good.>The JSON is structurally legal, but applying that Story Plan would leave incomplete planning fields.>This is a real Qwen/Story-Planner compatibility problem, not merely a harsh benchmark. The same behavior occurs in one-Generation cases whose snapshot inputs actually were deterministic.
>>109739304I have been saying this for days now
What should I generate?
anyone got a download link for this?
>>109739576>lora trained on 1000 or so porn images>28 * $15 = $420>each version probably needs to be purchased separate so he's made many times thatIs it really that easy? Let's say I have figured out how to train Minimax H3, way better than everyone else's shitty loras. Why shouldn't I just pull a grift like this?
Has local video models been tuned down enough yet to work on a 12GB card or should I still cope?
>>109739576it's been up on the site
>>109739702>Why shouldn't I just pull a grift like this?because you most likely won't make something better than everyone else. if you could, you'd have already done it.
>>109739702>Why shouldn't I just pull a grift like this?Because you have integrity.
>>10973971312GB is just under the very smallest model that is 12.9GB large. Get a 5060ti with 16gigs and you can run nvfp4 comfortably, or get a used 3090 and you can barely run the int8 models, I assume you'll suffer though if you try to gen large or long videos.But before you buy any hardware there are like a billion 6GB vram zomg workflows on civitai, just try it out. But you need at least 32gb ram, better would be 64gb otherwise it's going to rape your ssd, if you have less than 32gb ram and it will be really slow
>Krea 2Is that style in base or is there a lora I need to get?>H3Have there been any major speed ups since release? I got a turbo lora a few days after but haven't touched it since
https://n.uguu.se/slFHwXFo.mp4
>>10973981050k likes on goontoob
>>109739810I don't get
>>109739826filtered
>>109739800Krea 2, plus the 'Text Fusion Refusal Reduction' LORA.For H3, just use Turbo LORA with decent MP size (0.6+MP).
>>109740120*correction, Krea 2 Turbo, base is ass for good images, sadly.
>>109740120Thank you.Is there a best lora for h3 or any will do?
catbox is dog shit what the fuck
>>109740152Is there anything better?
>>109736499masterpeicethis is one one of the best everand checked
>>109740134Personally, I have not used any LORAs, sorry! Might be best to ask other anons in this thread, I find that H3 works best with just the turbo lora, or without any loras :)
>>109739718what site?
>>109740164I meant the best turbo lora since I found a few
>>109739810would've been funny without the ebonics
>>109740174either one of these
deleted 600gb of wan loras. h3 is so much better im never going back.
>>109740243retard
>>109736449What model was used for bruce willis?
>>109740315Probably Krea
>>109740315>>109740315Krea indeed.
>>109738658How do you promt the audio to be used 1:1? I have trouble expressing myself to Krea.
>>109740321>>109740363Ah, thanks. I was hoping it was a model I could use with vapourkit but sadly not.
>>109740376*minimax h3
>>109740380What is vapourkit?
>>109740388It's a program that can be used to apply filters over preexisting videos. I've been messing about with DLSS5 in it.https://github.com/Kim2091/vapourkit
>comfy running super slow>dont know why>ask my little ai buddy>afterburner was somehow power limited to 27%I don't know how the fuck that happened. I never touched it?
>>109740431Leave some power for other users
>>109740447Think of the data centers!
>>109740407Does it give good quality upscales? Does it also fill in detail?
>>109740376Use the ref2va model, put audio into audio ref_audio_0then under subject definition:<Audio 1> is the synchronized audio track of <Video 1> and is reused in the target video.In summary add:["word"+ audio reuse] retention anal:<Audio 1>: fully_copy - <Audio 1> is reused 1:1 as the target video's complete final audio track.This should work, if there are people talking you still need to prompt when they talk with an accurate timestamp at the prompt, though you can also try without.
Some absolute garbage in this threadDo better idiots
>>109740459I can't really say as I only just started trying it out.
>>109740517what do you have to offer
>>109740543Wow... What a beautiful gen. Thank you.
>>109737916GemmaPrompt dev here, what update would you like to see? It already works perfectly for my use case.
>>109740490Thanks this is very helpful. I hate intricate prompt crafting and h3 is the worst.
Why do latents have to be so big? Can't somebody come up with a compression algorithm for them?
>got training for anima working on AMD>its still painfulim glad it works but just barely. honest to god i should just start forcing my way into some of these projects, forking it and adding amd support but FUCK maintaining any of that code
Is comfyui still the best?
>>109740832The best at what?
I wish there was /leg/ for 2D artwork. I'm desperately in need of advanced models and techniques but all I get is illustrious sloppaAnyway, are there any decent edit models for 2D art?
>>109740911/ldg/ duhGiving away early morning phone posting like a retard
>>109740431shortcut?
>>109740852being a ui
I am a retarded touristdo any of the models that come natively supported in wanGP support NSFW gens or do I need to get comfyui
>>109738395dude please re-upload and upload it on /wsg/. This treasure needs to be shared and I want to add it to my collection of awesome AI videos
>>109738835Here you go.https://mega.nz/file/3xRg1T5a#9FHCr7-KH25hDw6GAXJBVAE7t7ykFPAny4pQWRJenTI>>109738743I did this exactly: >>109739045.The continuity plugins I've tried so far destroy the quality of the later clips. At least with references and a turbo lora.
>>109741042H3 is (Kinda) uncensored, sometimes it will force clothing but 9/10 times it will just work out of the box (From personal experience, mileage may vary)Krea2 to create the first frame/image, on the templates choose the refusal reduction
hellohttps://litter.catbox.moe/ctrcsw.webmbye
>>109741090awesome, thanks.cool idea, and well executed. looks just like a real ad. I love the cheeky way she pops that last mentos. It's little nuances like these that when you find them in AI videos really get you hooked.
>>109741025it's the worst at that and it spies on you. the only thing it's best at is being a backend but even then it's unstable as fuck
>>109740490With your help this is the closest I've been able to get.https://files.catbox.moe/63khwb.mp4I just can't get it to replace the original voice with the second <Audio> clip I provided despite multiple prompt revisions. I guess I'm just expecting too much here. Maybe it's just something small and stupid that I am missing.
>>109741382H3 is a vastly superior model but this is one area where LTX actually shines - not too much movement, good lip syncing, super quick renders and you can really push the boat out on clip length. I made this a few months ago when testing LTX2.3https://litter.catbox.moe/jkizq1ndm1ss1y2u.mp4
>>109741440hmm. maybe I'll look into that. So you can do audio only stuff with LTX?
Can someone recommend a nice comfyui workflow with all I could need for SDXL models?
>>109741440If I were to create some yapping only video, I'd always choose LTX. I can just provide the audio I want instead of an empty audio latent, write what is said in the prompt for better adherence, and have a perfectly adequate video in the end. 160s for 1080p and 30s video for me.
If i generated a test video in shit quality, but it turned out to be great, can i throw the same seed and make it exactly the same in high resolution?Asking before i waste 50 minutes waiting for gen to finish.
>>109738528
>>109741675I would but I don't talk to brown faggots.
>>109741797japs are yellow tho
>>109741717That's unfortunately not how this works, but you could get a feeling for whether or not the higher res video goes broadly in the right direction even before completely generating it. Just use "Model Preview Override" node with taeh3 and you'll pretty much know if it does after like 6 steps. The preview is usually clear enough that you can judge whether or not something has gone horribly wrong at this point, at least visually.
>>109741717i assume you can probably upscale your shit quality vid and v2v it to keep the same composition
>>109741717No, the only option would be to upscale the result
I wish my system didn't have a melty when and only when I use local models.I'd be generating so much degenerate filth you wouldn't believe it.
>>109741931>>109741949From what i know:>crap in>crap outI'll try some sort of upscaling just to see if it's worth anything. Any recommendations?
>>109736499>Ciaran MalikLiterally who? That better not be you, anon.Funny gen though.
>>109742036any dumb upscale and v2v half steps / .5 denoise
>>109734278Nice. Which old school mangaka were you going for with this?>>109733118Implessive. Could be straight out of the animoo.>>109731449You ... you good, anon? Yiff twice if you're not.
>>109741382Well you asked for a 1:1 audio replacement, you never said anything of dual audio clip changes.I already said it's quite difficult and you also need to change things.Here is what worked for me a few days ago.Important is, if you're using music then the audio clip of the music cannot under any circumstance be louder then the speaking volume reference clip. Simply castrate the audio clip in sneedacity or audacity and try with this:<Audio 1> is the dialoge and pacing reference for <Subject 1>, but she retains her own feminine and soft voice.<Audio 2> is the voice reference for <Subject 1>.[video editing + audio reference] The target video is an edited version of <Video 1>. <Subject 1> replaces <Subject 2> in <Video 1>. <Audio 1>: weak_reference - the target speaker follows <Audio 1>'s delivery without copying the original signal.<Audio 2>: weak_reference - <subject 1>'s voice tone and pitch and general soundyou guys really need to be more specific when asking for shit
quality of life features added:-tag highlighting. mouse over on a tag such as <subject n> highlights every instance of that tag. -tag thumbnails. a small thumbnail display what the tag represents is also displayed.-send to ComfyUI. all media content + prompt are can now be sent directly to comfyui workflows. no need to go back and forth-nsfw writing guide. the guide instructs the prompt to target nsfw content. less guesswork for the llm
If heaven isn't all of your generated 1girls raping you forever, I'm not going.
>>109741717NoIf anything. Higher resolutions will fuck up your prompt
I decided to give comfyUI a second chance, this time staying away from video generation to see if it'll give me just as much trouble. Going with something simpler, which model (Local, not cloud) offers default image to image conversion.You know what I'm doing.You know.You know what I wanna do.(Degeneracy)
>>109742312Not sure why you had so much grief with video generation. It's usually just a matter of using some inbuilt template and downloading the models. Image to image can sort of be done with any image model by feeding the image in as a latent with lower denoise. I assume you're looking for image editing, which is quite okay with flux 2 klein 9b or flux 2 dev (if you can run that).
>>109742414No, it's not on comfy, it's something to do with my GPU. There's something that generative UI does that accesses my GPU in a way no videogame or stress test does which makes my GPU fuck up the drivers and I have to re-install them from scratch to be able to use steam. It's bizzare but I don't have the kind of money where I'd be willing to be replacing hardware in this economy so I just have to deal.Honestly I'm hoping it's video generation itself and not gen in its entirety.Actually I was still looking shit up after asking that and I did end up on flux clein as well. I'm downloading 4b right now (9b needs an account if I'm reading this error message right.) but I digress. An anon previously told me but I forgot, what's the keyboard button I press after putting the model files into their respective folders to have it update?
>>109742042i'm not that retarded. here's a hint: check the filename.
>>109742468Welp, got it set up and running.>GPU Utilization 100%>GPU board power 230-260w>GPU temp 65C>GPU memory utilization 19896MB>GPU Memory clock 2487 MHz>GPU Memory temp 82C>CPU Utilization 10%Running stable so far and at this rate it seems like my test gen will take about 10 minutes.Wish me luck lads.
>>109742468It's just "R".
>>109742680Thank you. I'll make sure to write it down this time, so I don't have to exit/run every time I do something new with the model.I ran a test run, seems to have no issues so far. I'm still threading lightly, but after it finished I closed out ComfyUI and tried running steam and this time nothing got messed up, so it might have been MinMax H3 that my card specifically just doesn't agree with.I might start experimenting now. Klein 4b ran an exact 50/50 blend of the two faces and ONLY the faces, and I guess it's up to me now to play with prompts until I can make it stop doing that, assuming my shit doesn't break again, that's gonna have me on edge for a while until.
>>109742468>>109742734iirc klein 4b is dogshit and there's no reason to use it you can almost certainly run 9b. just go download the file rather than expecting it to autodownload via comfy
>>109742816>iirc klein 4b is dogshitI'm starting to see that, it keeps putting clothes on my nekkid womyns. >just go download the file rather than expecting it to autodownload via comfySorry man I know it's stubborn but I'm not getting another count with anything for shit.
>>109742816Is klein 9b censored?
>>109742844thats okay, maybe you can get it from civit or get some quant or sloptune or something of it from huggingface that doesnt need the agreement clicked. if you cant find it, just give up and don't do anything, don't waste your time with the 4b>>109742877IIRC similar in kind to H3 but worse, i.e. it's quite stupid/bad at genitals and outright sex but generally doesn't really *refuse* in the way previous BFL models did (by being lobotomized to hell if the prompt had anything to do with human anatomy), and by its nature as an image edit model much like H3's reference model there's not much that filtering can really do against "here is a sexy pose, here is hatsune miku, put her in the sexy pose", swapping characters into a pre-existing image. Might take a couple of rerolls but it's a fast enough model to run. With that said it was a mediocre model on release, worse quality than the competing ZIT and only relevant because it was an image edit model (plus ZIT seed variety made it unusable for most use cases). so it will probably seem even worse today with Krea2 to compare to for low-filtered raw gen and H3 to compare to for reference consumption (albeit into videos instead of images). honestly id consider just using h3 for image edits with a 3 second gen at 1MP but ive not had any image edit tasks i wanted to do to test this out
>>109742931so you are saying coomers still haven't uncucked it yet?. It's been a year; there must be some image edit that's usable
>>109740587it needs to be able to analyze videos for video references, and remove the limit on shot count
>>109742877>>109743006just use a nsfw lora you retard
>>109743006https://huggingface.co/darknight9121/FLUX.2-klein-base-9B-bucket-uncensoredIs this not it?
>>109743053this is so washed out it may as well be a SD1.5 gen
>>109743006slowly realizing that local is dead
>>109743071>local just got the best video model, comparable to sota api models in quality>durrrrr local rrr ded, lul
>>109743071I mean for image generation local is still king. And for uncensored video generation as well.Only LLM applications are better over api due to the massive context size and tensor parallelism that datacentres allow
Is there anything that helps you track trigger words for loras?don't you tell me to write it down
does anyone use any magic one-liners for h3 i2v for 2d animation? i want to automate some of the reaction and movement effects by genericizing them into high-level instructions so it picks up more style queues from the actions and environment.ultimately, i'd just like to spend less time prompt writing than i spend genning. i've got a dedicated 3090 for prompt writing, but the output needs enough tweaking and cleanup that it's usually simpler for me to artisanally handcraft each scene.>>109742931h3 seems resistant to terminology that directly implies sexualization. terms like seductive, pleasurable, etc don't really do anything, but it knows what to do if you prompt longhand for an expression or reaction. i'll describe a scene as 'racy' or use associated terms that imply it shouldn't steer away from mature details of a certain theme, but i haven't found a prompt that universally keeps it from that kind of self-censorship.
>>109743141No the real trick is to not download garbage loras that rely on triggerwords unless they are literally character loras in which case the trigger word is self explanatory.If you download jeffs lora and he uses the triggerword @jeffAnaL32Rp don't expect any sympathy from us.Though most loras work without trigger words regardless
>>109743141comfyui lora manager
>>109743053Why on earth would you use Klein for NSFW ? Are you literally retarded ?
>>109743071I can see your nose from here
Sometimes I gen something in chatgpt and post it here to keep you guys on your toes
>>109743182NTA but not retarded, just new.I'm sure when I get to the point where I've learned enough about workflows I can assemble my own. But as a starter,I kind of have to go with what's given to me based on my extremely limited knowledge, and that knowledge right now is "Find model, download model, put files in the folder. select A and B, input prompt, click run."There's not a lot of models these templates I've found that actually start off by letting you input two files and prompt what to do with them which is the stage I'm at right now. Just playing around as early experimentation with the tech.
>>109743220>not using local like the rest of us
>>109743237See, I'm so fucking new I'm still mixing up terminology. I meant to say>>109743237>and that knowledge right now is "Find template, download model, put files in the folder. select A and B, input prompt, click run."
First Klein 9B gen and I want to vomit alreadyIs this the best local can do?
>>109743424Indeed this is the best local can do, we don't have any better models than klein 9B the model that nobody uses, no sir. Now would you please remove yourself from this general?
>>109743424no. Krea2 is better and has a lot of knowledge if fighting games. Why not ask what model is sota?
>>109743436>>109743439But I asked which is the best image edit that everyone uses, and anon said Klein
>>109743447anon is just our appointed representative to deal with newfags. you should have asked anon, instead.
>>109743461this so much
>>109743424Obviously not. Flux 2 klein is a smaller variant of the larger 32B model as the name implies. That will most likely handle stuff better, but it's not that great either.Only in terms of NSFW image editing, local is still ahead, just because cloud only models tend to completely refuse that.
>>109743447because that anon "assumed" you knew what an edit model is when your query is for img2img. It's not my fault anons are retarded
>>109731169Nobana (kinda), my beloved <3>>109738454>ratemybandNeed to all be hung by the neck until dead/10.
>>109743478the prompt is only turn this image into a realistic photo. it's not img2img
>>109743504That's literally img2img, are we being retarded rn on purpose?
>>109743504>turn this <image> into <image>well that's totally different. i have a workflow for that if you're interested, and it's only got 71 custom nodes and four different pytorch version dependencies.
Can I just quickly say that there are currently two anons that are new and talking about img2img. The reference anon you're talking to is not the same as me, the anon that wants to make coomer shit by replacing women in gooner shots with women that don't originally have gooner shots.
I love coming back from a long break of genning and discovering that there are new great models. Krea2 seems way better with way less effort than Z-image.https://files.catbox.moe/bz6v7a.jpg
>>109743521 It's not the same thoughimg2img add noise then denoise on image image edit is just a branch of InstructPix2Pixtechnically, it's denoise strength vs guidance scale.I maybe stupid but you are not fooling me
>>109743590stfu you retard, it's not even funny anymore we all used up our giggles already see >>109743549who is now even scared of being mistaken for you
>>109743504Not the best example, but stylised to realistic can be fairly easily be done by feeding the original image as a latent into a model with a denoising strength lower than 1, and prompting for a realistic gen.I'm pretty shit at prompting for realism in Krea2, as I almost never do it, but I think that gen kinda shows that it works.
>>109743580>https://files.catbox.moe/bz6v7a.jpgShow workflow (and prompt)
hanging out here is making me smell like curry
>>109743702We have browns here, but they are not indian
>>109743424yikes this is bad
>>109743424AHAHAHAHAHA LOCAL IS SUCH A JOKE
all me (you) btw
>>109743654https://files.catbox.moe/8asevw.pngGood luck with that. I'm tweaking the workflow constantly and not using everything in it right now so it's messy. I'm still adjusting from zimage to krea.Maybe someone can tell me if I'm doing anything wrong with the krea generation in the sampler or anything.
>>109743713>picrelAHHH! You can't just bust out spoopy stories like that outta nowhere, anon!
>I've been unknowingly using ref2va the entire time for i2v fl2va promptssigh...
GPT IMAGE 2.5 IS INSANEEEEEE
>>109743424You need a really high step count like 200 for the best resultsDon't forget to crank cfg to at least 17 to counter the step count and do warmup gens (gen in batches of 16, after the first 14 the results will be magnificent)
>>109744057>crank cfg to 17The sigma of the scheduler only works in even numbers, all images will come out looking like shit. I'd personally set cfg to either 16 or 18
>>109744069>The sigma of the scheduler only works in even numbers, all images will come out looking like shitthat's because you're not using an abliterated text encoder, odd number chads stay winning
>>109743580Coomer anon here.I installed Krea2, but unlike with klein which was heavily censored but at least actually used reference material to swap out elements, Krea2 just generates random profile shots vaguely inspired by reference shots. Anything I can change in picrelated to make it more accurate? Or do I just need to to start learning how to properly write prompts?(Or is Krea2 even more censored than Klein and I'm wasting my time?)
>>109744105>pic related>shows absolutely nothing worth of valuekek
>>109744121That's the whole thing. That's how krea2 template opened. Brother I did say I am brand new and don't know shit in like seven different posts so far.
>>109744135didn't the other anon upload his workflow for you?
>>109744105Dunno, haven't messed with klein, or i2i or reference material at all, I just do t2i. I'm no expert. Krea with nsfw loras is the best at nsfw I've seen so far. I've gotten basically zero body horror or anything like that, but I had the filter bypass and nsfw loras on from my very first gen with it.
>>109744105you wanted img2img so you have to denoise like everything fucking else. What the fuck do you even want? Edit or img2img because they aren't the same
>>109744012it's rotating the pixels
>>109744142Oh god, thanks for that, I forgot that was a thing. Holy shit that's a lot of boxes, I have a LOT to learn. >>109744156>you wanted img2img so you have to denoise like everything fucking else. What the fuck do you even want? Edit or img2img because they aren't the sameimg2image I am assuming (Replace person in image 2 with person in image 1)But I'm not usure of the distinction because that feels like editing too? What's the official difference between the two terms?
>>109744105>Anything I can change in picrelated to make it more accurate?Not really, the default workflow is good as is.>Or do I just need to to start learning how to properly write prompts?Prompting is overrated for the most part. You really only need to keep iterating until you get what you want.Remember to flush your GPU cache every now and then to clear out any traces of previous prompts that might show up in future generations.
>>109744192by the way before we waste any more time trying to help you, do you have comfyui manager installed? Cause if yes the thing you need/want is extremely easily accessible, but not without it
>>109738384Nice, eerie feel. What's the subject though? A fallen star in the mountains? Also: >>109742078>>109738395>link d00dAnyone DL'd it? Reup?
>>109743424After investigating a bit. Black Forest Lab recommends cfg 4.0, but the default workflow in comfy uses 5.0 for whatever reason. Also, it didn't warn me that I shouldn't leave negative prompt empty. Simply adding "Indian" to negative fixed most of it. Default workflows were probably written by AI or worse, jeets Apparently, BFL also shilled their distilled model more than the base one for whatever reason. But I don't care, I'm going to use that to avoid these brown-coded nodes
>>109744245>do you have comfyui manager installed?I have whatever is in the ComfyUI_windows_portable_AMD
>>109744251>linksorry bud, only for 4chan premium users
>>109744260Then check out my retard proof guide https://rentry.org/tkdupekk and install comfyui manager as well. simply skip to step 3 and install it, without it you wont get far after that we can talk
>>109742493Can't be arsed to right now but thanks for the hint, anon.>>109744261qq
What's the current best workflow for h3 reference to video? Still the default one with the default model and Spectrum+Sage Attention?
>>109744309switch sage for cka and yes, that is the current best workflow in terms of lossless quality with maximum speedups
>>109744326And Spectrum still uses the same default values as always?
>>109744349yes the stock values from when you put in the node are the best
>>109743447image edit is pretty much a dead endH3 same thing, the hype is just a few anons lying to themselves to try to justify spending on GPUs which they couldn't afford. it's completely fucked, barely usablet2i with krea 2 I haven't explored that much, there might be something there, but don't get your hopes up either
i had another dream about generating the same prompt over and over again
>>109743424>>109743071
How does Spectrum in it's current state compare to using the lightx2v turbo lora at 8 steps?I remember Spectrum would often have output issues like if you wanted camera shake.
>>109744408just fyi, catjack literally lies all the time about ani and confirmed cannot read english
>>109744038>local
>>109744462We figured that out years ago. If only mods would figure that out
>>109744440Spectrum takes twice as long at 32 steps compared to the turbo lora at 8 steps (ref2va). Spectrum having any output issues is 100% on your end since it should neither be visible or noticeable due to the way it works.
>>109744408what did julien do to yoland? kek
>tfw haven't genned anything for over a week because i've been vibing a prompt writer the entire time
https://litter.catbox.moe/2qmspj.mp4
>>109744478Not true. Clanker even told me Spectrum is known to cause issues with the way it handles model data.There is absolutely a problem with how it handles camera shake. "Slight camera shake" in the prompt often becomes a vibrating/rumbling camera that looks horrible.
>>109743182you can use klein to nudify, you know, like an edit model, you dumb faggot>>109743070sybau
>>109744485nothing according to that image and video. so why fixate on nothingburgers?
>>109744492I'm getting pretty close to abandoning my script writer app at this point. the token burn is crazy, the process runs for hours, and it still can't produce usable h3 prompts at the end of it all.
>>109744506Skill issue, don't care. Either you're on the spectrum or you're not. You're using turbo loras and argue with me about video quality as if I couldn't piss in your face with an ltx gen and you couldn't tell the difference
>>109744525but comfy said "you can go near him, i just don't think it's a good idea when you shit on him that [...]"is reading hard for you anon?
>>109744462>>109744485>>109744539>schizo is still reliving the past by looking at his discord friends post from january '25Let go schizo, we really don't care
>>109744539Is being banned being able to do the thing he was apparently banned for?
>>109744551He dedicated his life to this mission. Not sure what the point of it all is since he just became a concern troll lolcow
how do i join the ldg discord server?
>>109744608There is a schizo who posts images with a cat that has a bag over his head. You can ask him for an invite. Now I would warn you of the consequences but since you're asking for a discord invite it is clear you don't belong here
>>109744567i think a hard requirement for being a lolcow is to doxx yourself
How to make img2img end result very realistic? I tried throwing in the generated picture into flux klein and z-image-turbo but it looks like ai slop with dirt on top.
>>109744673Maybe if you pass it through sdxl then through sd1.5 and any other of the extremely outdated models it'll get better?
>>109744673anon just stop. for your own sake. it's not gonna work anyway. people will just bait you and waste your time
>>109743009gotcha, I'll have that over to you by end of day wednesday 9/9
>>109744673if the generation itself looks too sloppa then you probably cant fix it. if it looks pretty close to realistic then you can do a bunch of editing to make it look like it was taken with a cheap digital camera to maybe hide the vae artifacts. this involves compressing the shading, sharpening the image, simulating the bayer filter, and adding some noise. you need to learn the post-processing that phones go through. take a picture with your own phone and zoom in to see what the noise pattern looks like
Any good workflows for generating coherent long videos with Minimax H3? I've looked at some and they seem clunky.
>>109744733no, its all snake oil. the only option is to make multiple 15s clips, manually review/regenerate as needed, and stitch them together
>>109744705anon he's clearly taking the piss. Nobody who is new to image gen would even manage to find obscure and outdated models like flux klein or z-image.You can see for yourself, go to civitai, check out models sorted by newly released and you won't find any of the models mentioned.>>109744733I've considered making one but I never saw the point cause unless you have a shot/cut/scene that is longer than 15 seconds there is no need for it.Though I might make an infinite continuation workflow someday just to make a "walking infinitely through the city" vlog type video to see how much the ai can hallucinate
>>109744524lol tranny big mad
>>109744534You're not very intelligent, anon.
>>109742160If your a tranny it's uno reverse kek
>>109744773I'm in the 98th percentile when it comes to intelligence, have you considered it might be you?
>>109744626I thought it was denying reality and advertising your mental illness
>>109744531I'm making something like this, but it's quick and modular. Give a story idea and some number of generations, and the writer will set up the framework for each generation. Information like incoming state, outgoing state, camera state and sequence context. Benchmarks are looking pretty good so far.
it's pretty sad that this entire thread got reduced to some pathetic jeetoids trying to sabotage othersnobody is going to buy your shitty gens, sorry
>>109744800Nah you're clearly a moron, and don't know what you're talking about.
>>109744829Ok then, mind posting a single good video gen of yours to show us you know what you're talking about?
>>109744673Dunno what you have in mind, but I would assume going with Krea2 and .6 denoise with a good prompt describing every detail of the character and also describing it as cosplay should get you fairly far.Again, I'm shit at prompting for realistic stuff, so the example could probably look a lot better, if you let some LLM do that or something.
The H3 reference model works maybe 1 in every 10 tries, and each attempt takes like 15 minutes. Useless model.Whoever says it's good must only be genning talking head videos
>>109744867I'm trying to give it some references for a character and then another reference for camera view, etc. I have it set up exactly as their prompt guide says but it still sometimes ignores my character and gens something straight from the camera ref
>>109744816I hope you have more success than I've ended up with. when the app was simpler, the stuff it would gen was mostly incoherent. I added in my layers to track trajectories of story beats and characters, plus review/repair phases to improve the visual storytelling. now its just permanently trapped in trying to pass the continuity checks and never succeeding
>>109744867been my experience as well and 1 in 10 is pretty generousif you wanna gen zero movement and 1girl and the most obvious position like standing straight yeah it might work but that's about it
>>109738430>>109738436>>109738446Adorbs.
>>109739137>>109739236LMAO, WTF are you even doing at this point, anon? You got Joefever or something?
>>109744531Don't treat this like a job. If it just becomes a source of frustration, you should probably cut back the scope or abandon it.>>109744867I'm assuming you still did something wrong, because it worked out pretty flawlessly for me so far.
>>109744923more like cut back the cope
>>109744947>>109744947
>>109736449ok what models and stuff do i need for editing?
>>109744923>you should probably cut back the scope or abandon it.a bit of a sunken cost fallacy at this point. "just one more improvement and maybe it'll work", he said 50 times now.maybe if I sic astra on an architectural review, I can find a new path forward. or maybe I should call it a failure and drop the whole idea
>>109744966I understand. Astra is supposed to be pretty good at this sorta stuff, but I would assume it will also hit a wall, because I don't think longer form gens like full episodes consisting of mutiple linked gens are something anyone has really figured out yet, so there's a lot of snake oil it will find when researchning that topic.
>>109744848That guy is not me wtf
>>109745072Sorry anon, I mistook you for anon
Give me some prompts to try I'm not creative right now.
>>109745335>>109744947
>>109741717You have to use original video as reference.