Previous: >>109557402https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
miku nigger ignoring the call of black migu over the white boycalls
>>109561703what about this then?https://huggingface.co/MiniMaxAI/MiniMax-Music3/discussions/5#6a7f5da5341a0766b9c67ccd
>>109561698https://files.catbox.moe/cu88xe.webm
>>109561740Read a bit further on the comment thread, Ostris addresses it>Unfortunately it is not in there. That is the audio VAE encoder, not the RVQ encoder.I had Claude look at the file independently of Ostris and it came to the same conclusion, serveurperso most likely got the two confused.
>>109561698Reminder to submit your video gens for Sexy Jam 1:https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewformAll sexy h3 gens welcome. Troll submissions rejected.Deadline is after the weekend.
ITS ALREADY UPhttps://boards.4chan.org/g/thread/109557359#bottomhttps://boards.4chan.org/g/thread/109557359#bottomhttps://boards.4chan.org/g/thread/109557359#bottom
GemmaPrompt anon, you still have a lot of work to do. My manual prompts still end up much better in I2V because I've memorised a lot of the models quirks and eccentricities whereas GemmaPrompt still writes certain actions sub-optimally.
Debo says he didn't bake any of these duplicates.Is he lying, or is it really just ani gaslighting as debo? and simultaneously trying to get his own rentry removed?
https://files.catbox.moe/ipj4zs.mp4
Women don't act like this
https://www.youtube.com/watch?v=sVx1mJDeUjY&list=RDsVx1mJDeUjY&start_radio=1how many german fags got misleaded due to kinolet us connect
>>109561764You are feeding your output back into Gemma so it can learn from it's mistakes anon r..right?
>>109561769>is it really just ani gaslighting as deboAni hasn't even done anything to look like debo. It's always been Ani. It started when the Ani rentry was added back in december.
>>109561798no, chatgpt is far better at that.
Comfy question. Some lora loaders have optional clip passed through, is there any difference?
>>109561878There's a difference for SDXL-based models, if you don't mutate the CLIP then it'll just apply the full lora regardless of trigger words.This makes the most notable difference for loras which contain multiple concepts. Any lora trained on a tagged dataset technically contains multiple concepts but you would most often be able to tell the difference in cases of multi-character loras.
>>109561798>try gwen 3.8 yesterday>make tampermonkey script, like 150 small lines, reiterate 3 times because he made some retarded bugs>starts losing context after like 5 messages and adds the bugs he removedhow would a much less smarter model handle the user throwing infinite prompts at it?
>>109561698What can i expect Generating 10 seconds Minimax video with 16gb VRAM + 64gb RAM ??Amd card by the way
If you can get the ddim_uniform scheduler to work, you can get some pretty high gens if you are doing anime. But it's very unstable.>>109562022>AMDNo idea, but probably not great. Try it and good luck.
>>109562022>What can i expect Generating 10 seconds Minimax video with 16gb VRAM + 64gb RAM ??>Amd card by the wayNpo int8 convrot so your life is pain. Expect 360p max resolution and it takes 15 minutes a video.
>>109562067>high gensHigh quality, that is
why does H3 keep giving women hair ties around their wrists, and how do i stop it?
>>109562101>why does H3 keep giving women hair ties around their wrists, and how do i stop it?Nothing you can do about it since we don't have negs. I feel a similar way when my LLM inevitably adds an ankle bracelet eventually when I'm making foot fetish gens even though ankle bracelets are retarded and Indian and not hot at all
>>109562101bare wrists or no accessories maybe
>>109562112there's NAG for h3
>>109561758Thanks for the heads upHopefully Julien gets banned for good this time
>>109562135come on deepmeepbeep get it done!
>>109562135>there's NAG for h3Haven't seen anyone talking about it or using it, and doesn't it increase gen time tooExtra hair ties have never been a problem for me though because I prompt for pigtails with cute little bows
>>109562101tell it not to but be explicit though; say "she has no bracelets/ties on her wrists" rather than "bare wrists"I didn't test it much but I negative prompted for eyes, nose, ears, and it all worked, so think it'll work for bobbles
>>109562181Often if you say things negatively (ie don't do x) it will just do x. Bare wrists might be the better option.
>>109562219I know, cfg is a bitch in every other model. It was a week before I tried it myself because I just assumed it wouldn't work
>>109562131>>109562181"bare wrists" fixed it
What's the correct wording for H3 to continue the next shot from last shot? So it doesn't completely do a new scene from the description?
>>109562338Did you try "The scene continues" or similar?Most of the time, unless I denote a camera change, the scene stays the same. Unless you are using a strange scheduler.
how do you avoid a penis generating if you are prompting for the crotch area?
"," sometimes making the movement stop so ill just use "while" and "and"
I'm new to this and i'm greatly confused. Is there a guide for dummies?I've installed SwarmUI and downloaded a few loraWhat's next?
>>109562364i've gotten mixed (but tending toward successful) results with, "safe for work"
>>109562377Ask chatgpt
>>109562364Have you tried "no penis" or "no genitals"?
>>109562364Have you defined that your subject is a female, woman, girl, eunuch, etc..?
>>109562387no. i always have a mindset of implied removals since explicit ones never work>>109562390i'll try that
>>109562364Don;t worry bro, I got you. The word is "post-op"
>>109557462>>109558134>>109558725>>109558919these are so 2000s coded, absolute kino. I can easily imagine most of the scenes in a parody-type of movie as cut-ins. Maybe something for /tv/
I tried genning.. THAT.. with h3..It worked..
>>109562428yeahit's freeing
>>109562390seems to work. though it will be a problem if i need a male mannequin
>When you get everything tuned in and every seed is a new piece of kinothis can't be healthy, it's been hours
>>109562428
>"girl stroking penis nonstop, girl head turns camera"*Handjob stops after head turn to camera*>"girl stroking penis nonstop, girl head turns camera, "girl stroking penis nonstop"*Handjob continues nonstop, even when head turn to camera*Prompting with Minimax is really weird man...... Who the fuck the said prompting is better than LTX...........
>prompt suddenly gens consistently what I wantI'm in heaven.
>>109562457how would h3 know if you want her to keep stroking or not unless you tell it?
>>109562466I want Sulphur H3 so bad so i can just make a simple 1girl,handjob prompt
>>109562457[Shot 1]"<Subject 1> starts stroking the penis. As she's stroking the penis, she turns her head towards the camera, and she..."Use the proper format.
>>109562462proof?
>>109562481Thats what im use. i need to prompt "she's stroking the penis" twice to make it work
Kinda struggling with the turbo loras for Minimax H3 Ref2VA. The v0.1 lightx2v lora performs a lot better overall, but it tends to always do little bullshit like the floaty bits here, or the guns of the mech having some backward facing artifacty guns.
>Got a good scene and movement for handjob>Handjob sound looks like rubbing a wrinkled paperAh fuck this, ill just mute it
>>109562492Write it what to do.."she's stroking penis, she stops for a moment to turn her head towards the camera and continues stroking penis", or something like that.
>>109562499The larryvrh loras aren't meant for Ref2VA at all, but they produce a result with less artifacts, but a bit more smugded.Am I doing something fundamentally wrong, or should I just wait for lightx2v to release their v1.0 Ref2VA lora and hope that is fixes everything?
>>109562499The lora is still a work in progress, but you can try using some different samplers. er_sde or seed_2 maybe.
why is it obviously trained on sex but none of the sounds are there? it's like they muted those videos during training
>>109562480Wouldn't fine tuning H3 be less effective since it's a distilled model? That sulphur guy asking for $10k is kinda scamming, since it's unlikely to produce the results everyone is expecting
>>109562517Probably trained it against outputting that, same as with genitals.
>>109562481I was under the impression that you had to be extremely specific in your Minimax H3 prompts. "Stroking" seems very unspecific to me, and could possibly mean a lot of things. I would've thought something like "one of her hands is gripping the penis while moving it up and down along the shaft continuously" could work much better.>>109562516Thanks, I'll give that a try.
>>109562543>same as with genitalsit generates implied genitals if you involve motions related to someones lap. how can it know to do that?
>>109562543anyone tried prompting stuff like "copulating" or something else
>>109562586anon the loicence forbids it
>>109562394I just tried "no penis, no skin showing" as well as describing what he's wearing and 3 times in a row got no penis or weird abomination. Without it, always got some type of penis worm. Seems to work
>>109562607>penis wormfor me, it's the stalagmite penis
>>109562543>>109562531yeah this model is so pathetically scuffed that I doubt we will ever get proper dicks or pussies other than through overcooked jeet loras that destroy any prompt adherence and kill your gen. I just don't see it, sorry to say
>>109562628I think you're overly pessimistic. We'll see who was right in a couple months.
Man, animators are TOASThttps://files.catbox.moe/2h2cci.webm>>109562531You can finetune a distilled model if you know what you are doing. Not that I know if he knows what he's doing.
remember the anon that demanded that /ldg/ grovels and apologizes to him because he was right that z-image base would never be released because of chinese culture
>>109562390how would that help? a female can obviously have a penis..
>>109562657>a female can obviously have a penis..how?
>>109562657some humans are born with no arms but I have a feeling most humans you gen will have at least 2
>>109562642They were right
>>109562657just end it. get it over with. stop bothering normal people with your mental illness. we've had enough
>>109562674anon...check huggingface...
>>109562688hmmmnyo
>>109562688nta, but what we should check?
>>109562688Hmm I see something called an image. But no base… I also don’t see an edit model anywhere. Maybe something got mixed up due to cultural differences here.
>>109562701The main z-image repo where they uploaded the base model 6 months ago>>109562703image is the base model
post kinos. i know you're hiding some
>>109562674>>109562703omg were you that anon hahaha
>>109562709Oh. But that’s not the base model. You might think it is, but it’s actually not.
>>109562733Unfortunately for you it is. It is the model z-image turbo was distilled from. I don't understand, why keep coping after all this time? Are you down to your last bit of izzat?
>>109562740I can see how this would be confusing to you, as you were promised the base model. But as you can see from this image in the model card page. It is actually not the base model. I understand this may be upsetting for you to read but I implore you to detach your ego from the situation
H3 or LTX if I just want some simple camera motion and simple gesture from the subject? has to be high at high res
>>109562752So you don't actually want the z-image turbo base model? You want a different model? Which one?
>>109562759ltx is best for high res
>>109562740>It is the model z-image turbo was distilled fromit clearly wasn't, since even lora training is broken with that model, let alone a proper finetune. z image turbo was distilled from the true base model which they never released
>>109562760Resorting to strawman arguments I did not make to bolster your argument only serves to weaken it.Might be time to brush up on your Chinese culture lessons and learn to be content with reality rather than lash out at me for being the bearer of news.
>>109562778>It's not the base model because I say so, you fucking chud!k>>109562779>repeating what I said is a strawman, you fucking chud!k
>arguing about an outdated obsolete modelthere are better ways to spend your time anons
>>109562796Exactly, we need to talk about how anima is garbage compared to krea 2
>>109562793No actually, phrasing question in response to a statement I did not make is actually the strawman argument here. But I can see based on your perception of the release status of z image base that reading comprehension is not your strong suit.
>>109562812>I don't understand what you posted, so it's a strawman you fucking chud!k
>>109562818Yep, we won. It's clear you don't have any idea what's going on, or, more likely, are yet another Chinese shill I've now defeated. Good bye.
i just updated my nvidia studio driver to 610.88am i going to be okay?
>>109562829What did we win?
>>109562842Chinese culture.
>empty h3 t2i promptPost your results.
>>109562842A lifetime supply of AIDS
>>109562848huh, pretty mundane ony my end
>>109562894If that's supposed to be grace she should have greyish blue eyes.
Sheesh prompting ref model is such a pain in the ass, I have a decent reference image and have a prompt for my character to enter an empty scene but for some reason she's coming in as tall as a doorway. Gonna try fixing it with another reference image depicting her scale in the room.
>>109562929True, but I didn't specify it in the prompt.
>>109562894How did you prompt for that transition? Do you use the [Shot] syntax or just rawdog it?
>>109557688right, what i meant was that adding a lora on top of a model doesn't increase the vram requirements by very much. in case you're unaware a lora is a bit like a filter on top of a model, it's not a mini model you can run by itself, which i clarify because your suggestion of a lora theoretically *lowering* vram requirements implies you might have been under that impression. sorry i don't know the answer to your question about training on 8GB though. generally training is actually not significantly harder than inference, but it is typically a *bit* harder, and training your own lora is probably not something to even think about until you have a grasp of how generating stuff works. i don't know your financial situation but a good heuristic is to try to stay ~roughly ahead of 50% of the pack so if your chosen gpu is nvidia and has like 12-16GB of RAM most of the popular toys will target your bracket, 24-32GB+ chads will of course have an easier time and run anything easily but you'll always be able to do 95% of the things you want to do (if a bit slowly) as long as you're in the upper mid tier.as for LLMs there's only a very light touch of "x is better at y" it's mostly just a raw intelligence scalar, but the thing you're wanting here is a big corporation with a harness to solve the problem for you by it running a bunch of web searches and condensing the answer down for you, don't even think about local for that, seriously just ask chatgpt it'll do fine, the (localgen) next step up from that is quite a few hours of knowledge and effort you won't be able to bypass with a single 4chan post
https://files.catbox.moe/gg21cb.mp4This took so long but I think I've finally solved the blowjob horrific crunching noises problem
>>109562997lmao
>>10956299710/10
>>109562997kek if anyone is having this problem, the real solution is>overall_soundscape:>blowjob sounds, gagging and muffled moans.
weird shit start to happen when trying to gen longer than 12 seconds in h3 r2v. like it wants to change scene even if you didn't prompt for it.
https://files.catbox.moe/8a9w0n.png (don't open in public or at work)is anima base still the best anime checkpoint?haven't been here in some time
>>109563037If you use spectrum, did you pull? I did and my gen came out like an acid trip in the background, subject was intact though.
>>109562511I get good results by upping steps to 10.
>>109563066that looks like turbo or high cfgyou can download the merged anime-turbo checkpoint
>>109562950Actually spent a while trying to fix this exact problem, had my character walk onto the scene twice as tall as all the existing people. I prompt wrangled it eventually with lots of "normal height , height of the other people in the shot" but the model seemed to want to interpret the character as the same size in the gen's frame as it was in the reference image. Might be easiest to fix by scaling the input differently.
>>109562640>animators are TOASTI disagree. What's likely to happen is animators will just focus on fun stuff like keyframes and let AI do the boring shit. It will massively speed up the process and maybe improve qol for them.
Is H3 genning without sound faster?
so ref_image_0 is <Picture 1> and so on?
https://gist.github.com/PierreHoule/947c6655a68279bb16f661ffdbef6ba7is this a good workflow?
>>109563123That's right desu.
>>109563118sound only uses like 2% of the tokens, so removing it would barely make a difference to gen speed
6gb vram hereI'm okay with 0.2 megapixelsshould I go for wan2gp or use comfy for h3?
I'm retarded and didn't realize she's supposed to be wearing pantyhose.
>A community-modified version nicknamed a “heretic” build has circulated, which isn’t technically a fine-tune of H3. Instead, it removes certain layers and replaces the language model head to reduce refusals, while still requiring the original H3 weights to function. Users report it generates recognizable characters from popular media with fewer restrictions than the base model, along with reports of it producing nudity and gore when prompted.where do I get this uncensored model?
>>109563192what's the point?base is already uncensored and does that if prompted
>>109563192its the default model, retard. people got meme'd in to thinking the heretic qwen encoder somehow made a difference. it does not. if you cant get nudity and gore out of the box, you're a promptlet and the issue is entirely with yourself, not the model.
>>109563192Completely wrong. Heretic lobotomizes text encoders for video models. Refusals aren't an issue in video gen, it just needs a rich inner representation, which heretic breaks.
>>109563225>>109563226>>109563227okayI'll try running it
>>109563225>>109563226>>109563227>>109563174which ones should I get?
>>109563261int8 convrot
>>109562960Just rawdog>Track sideways behind foreground objects that repeatedly obscure the subject. Use each occlusion as a natural wipe, gradually moving closer until the final obstruction reveals an extreme close-up.
>>109563270
>>109563261int8 convrot pruned is the smallest & fastest, but you want both fl2v and ref2v models. text encoder is nvfp4, and the vae is the int8 convrot for video and fp32 for audio.
The Kijai turbo fl2v loras (https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras) are awful, they instantly sloppify all your crap. Anons here who are claiming the 4-step model is better than the 8-step ones are even more delusional, 4-step is so much worse in my tests.If you don't believe just look at the comparisons for yourself. Only one that is maybe OK is the lightx2v · 8 steps · str 0.75 · er_sde example. https://jo-nike.github.io/h3-turbo-eval/scenes/singing-sustain.htmlIt also more or less proves that Spectrum will fuck your gen and is not as free as every retard here is claiming.
>>109563261depends on your card, convrot for 5x series, scaled for 4x and below
>>109563226If you have to put things like "white, creamy liquid" in your prompt instead of just "semen" or "cum", then it's censored.
>>1095633603060 6gb
>>109562516Improved things somewhat, although both trial runs had vehicles partly clipping through the container. I assume this is just bad luck with the seeds, though.
>>109562848I got a 1girl tradwife. Might be considered cheating though because I forgot to remove the prompt connection from a prompt helper and so technically the prompt wasintegrated_multimodal_description: overall_soundscape: non_diegetic_music: N/A>>109563348This was done on the turbo 4step @ 1.25 str, but I ran it at 10 steps euler/beta .3mp. Audio issues with the loras are fixed by following the actual recommended step shifts based off of the lora you're using. 12/3, 12/6, etc. aren't universal.
How do I get h3 to make a girl get fucked with the penis going deep in and out? Whenever I do i2v I usually only get the penis sliding in a little bit (like one third of its length) and then out a little bit again."fully inserted" does nothing
>>109563348All cope nodes and loras except sage is pure vramlet dribble. They don't care about quality loss or degraded prompt adherence just as long as they get their slop faster. Ignore them.
>>109563399If there's something attached to the penis, you can try describing the position of that something.Like his crotch touching her ass or whatever.
>>109563392You can't see the artifacts in this? It's looks awful.
>>109563341okie
>>109563348Spectrum seems acceptable compromise if you look at the samples.
>>109563088Seems like I'm not that lucky...Will probably have to wait for a proper Ref2Va turbo lora, before really trying Ref2VA again. Don't wanna wait like 20 min for a 15 second 1mp gen.
>>109563445increase the lora strength to 1.25 in addition to going up to 10 or 12 steps.
I know this is a general for gooning to slop softcore porn of artificially generated women, but has anyone tried the MiniMax audio model m3? I'm looking to train a lora for it but since audio ai is lagging so much behind image and video idk where to start.
>>109563445looks like modern halo if it wasnt woke and was instead made by japanese pedophiles
>>109563409People optimized the quality out of the model so quick they are now complaining the model is bad and gives shit output. "Model can't do sex!! It's so bad!!""It doesn't follow my prompt!!""Ugh... I get the same results every gen, seeds do nothing"If you actually use the model as intended without the 500 cope nodes, you can generate 90% of shit anons keep struggle posting about every thread.
Attempt 1 was a failure. I want to try getting it right but this shit took 25 minutes to gen. fml.
>>109563510>YOU DIED
>>109563510why would that gen take 25 minutes? are you running a 1060 ti or something?
>>1095635397900xtx
How to make Comfy save the queue and autorestart on crash?
>>109563510That's still a lot of gens over the course of a day.
>>109563348that entire page is focused on audiofaggotry
>>109563542haha fag
>>109563477Tried it and while it is kinda nice it the songs dont really blow your socks off and are a bit generic, but it's kinda a big leap when it comes to local quality. Strongly suggest you use a llm of your choice with the prompt guide to generate songs. While i tweaked the outputs of the llm a bit here and there, i didn't deep dive, so there might be potential. It also has the same duration issue, double the length takes about 4x long iirc.But overall it's gonna be my go to music model from now on.
>>109563510ToT
>>109563348also these results are couple days old and do not contain anything from here `https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras`
How do I get 100% chance of getting an instrumental in Minimax Music, I just want some cool jazz, but the moment it starts singing it's incomplete.
>>109563568it's the first local model that is somewhat usable. Comparable quality to suno except not behind a paywall so you can do infinite autistic tinkering. I want to generate instrumentals and after testing for a bit it's definitely trained on royalty free crap, but has the potential to be saved by a lora. It's just that there is practically no lora training support atm, otherwise i would be tuning it right now
>>109563598grok told me to put this in the lyrics prompt and it makes speech much rarer than just leaving the lyrics blank : >[Intro]>(instrumental)>[Instrumental]>(instrumental)>[Outro]>(instrumental)
>>109563477It sucks. Too bland, can't get any good range with it. You either have to settle for AceStep which has ok quality but occasional errors (lyrics dropped or popping sounds) but good emotional range and depth or MiniMax which has very clean audio and doesn't seem to drop lyrics but is bland as hell.
>>109563457Alright. Will try that.Hope we won't arrive at the lora only working well at 2.00 strength and 16 steps or something kek>>109563486This was made by a european podophile, though.>>109563627Can Minimax Music at least handle stuff like techno well? Found that AceStep is terrible at it.
>>109563677>This was made by a european podophileOuting yourself?
>>109563192are you getting your info from Gemini Flash?
>>109563718noI googled h3 for 6 gb card and this popped uphttps://www.mindstudio.ai/blog/minimax-h3-run-locally-guide
>>109563729guess which model wrote that AI generated blog post
>>109563683>no flashing screamer:(
>>109563457That still doesn't seem to work so well.>>109563690Don't think outing yourself as a footfag on anonymous image board is that terrible.
>>109562997it's so bad compared to ltx
>>109563782parading your mental illness should come with shame
>>109563782Is your summary prompt fucked up? From testing, I've learned that it'll prioritize the summary over the detailed description, and even the [Shot] tags. Also the sound tags can cause ghosting if you've specified specific sound effects to play during certain actions instead of prompting them in the shot itself.
>>109561751sad
minimax music 3 is so good at normie music styles like drill rap, the comedy potential is there. wish it could do audio.
If I make one 10 second clip of fake star trek every day, I will have full episode in about 260 days.
>>109563858With an RTX PRO 600 genning 10 second clips continuously you could probably get a full episode every 2 days. Just slop together a buncha prompts and queue em up. Someone will probably automate this and post them to youtube.
Anyway to make animation from similar images that do not have exactly the same shadows etc..Like have one ref image and model edits the rest to look like that one?
>>109563957You question is confusing. You can provide multiple reference images for multiple characters. You can edit a video and replace a character or multiple characters from reference images.
>>109563802Summary is quite concise and I can't spot anything that could really throw the visuals off:https://pastebin.com/z4RjNdZJOther samplers work a lot better, as in this example:>>109563374
>>109563949I'm the bottleneck not my gpu>>>/wsg/6215059
>>109563995>hilarious lmao I love the steel growing
How do you install kijai/ComfyUI-SolAttn_triton? I cloned its git repo into custom_nodes, no requirements.txt. The github page also has no instructions
>>109563966If you gen like 24 images with different poses that could still be put in sequence to make an animation, they wouldn't match since AI can't gen consistently.Basically, is there a node for Comfy that can replicate the looks of one of the images and copy it over to other 23?
So, reference images scale down to the size you're working at first, right? So if my output file is 0.3mp @ 4:3 (640*480)is my reference image being scaled down to a max side length of 640, or 480?
>>109564012You need to use a video model like H3 or Wan or LTX
>>109564010I installed it exactly like that.You rebooted comfyui, and if you kept the browser UI open you pressed r, right?
>>109561698New to this. Is there anything offline that comes even close to what's available for subscription in flexibility and quality (chatgpt, gemini, etc.)?
>>109564050Yeah. Do I need to rename the folder to prepend it with `kijai/`? Or nest it in a kijai folder?Also I can't figure out how to install a node through Comfy UI via Github URL. All the instructions I can find reference the old comfyui manager panel
>>109563995>>109564001I'm serious, please post this on fb for max boomer rage.
>>109564023It tries to match pixel countYou might want to experiment with max for better quality though. (Don't forget to manually downsample really large images to 1MP or whatever)
>>109564072No.What node in the workflow is erroring out?Maybe it's a retired node from previous commits or something.
>>109563949this is my dream and its close to a reality with h3. the biggest problem to solve now is a script generating LLM that would be able to orchestrate the h3 prompting in a way that is vaguely narratively-consistent
>>109564098
Does anyone have dual sampler setup for Krea2? Like raw + raw with turbo lora. Almost every seed is the same with just turbo and I'm not good enough at comfy to make my own one :(
>>109564052You do mean image models right?>general purpose SFW text2imageCloud is slightly better.>artistic SFW text2imageLocal is better than the big players but roughly at the same level as NovelAI.>NSFW text2imageditto>SFW image editingCloud is much better.>NSFW image editingLocal is much better.>SFW text2video/image2videoCloud is much better.>NSFW text2video/image2videoLocal is much better.
ouch, i just cut my finger to the bone
>>109564127lucky you, I'm nofapping.
>>109564127get well soon fren
>>109564080ah, thanks!>Don't forget to manually downsample really large images to 1MP or whateveroh? does it fuck it up with bad nearest neighbor scaling if it's not scaled down first or something?
>>109563614Tried many times, it only works occasionally. This model is for songs. So Stable Audio, which has an inferior music quality remains useful.
>>109563990substitute yourself with a LLM
>>109564139shouldn't you pray to God that the sodomite die of the infection?
>>109564108shit looks like that either if the nodes aren't found at all in custom_nodes or if there's an error parsing the py during startup, does the terminal have any errors or warnings?
>>109564113yeah sorry I thought this thread was exclusively about image modelsbasically>general purpose SFW text2imagebut I wouldn't mind if it let me go a little more nsfw than ChatGPT and was at least as good as let's say 2-grok versions ago which was pretty bad compared to the other cloud models. Is there anything like that?or what's the best I can have in that regard?
>>109564272Krea 2 with loras. It's really good for NSFW text2image and image editing.
Can't we have a Matrix server where anime video creators can exchange ideas and collaborate?So we can avoid all this hate we get from regular anime fags and AI meme creators.We are the most persecuted minority on the internet. You are not alone.
>>109564284very cool. I'm assuming krea image as last frame?
/adt/ has deduced that catjack is the target of the sharty raids because triggering him will greatly reduce thread quality
>>109564288How does one use Krea2 for image editing?
>>109564351https://huggingface.co/conradlocke/krea2-identity-editIt's not nearly as versatile as cloud models or even Klein (local model), but it's the best for NSFW.
>>109563107I though animators loved to be miserable
>>109564338it's pretty obvious desu. sad that catjack keeps doubling down on this retardation when people just want to post 1girl and discuss tech. all he has to do is let go of whatever grudge he has
>>109564310 No. They always end up dying or becoming inactive
>>109564338>>109564381why are you talking to yourself?
can I only install one missing custom node using the manager in one go?
>>109564355Thanks, will try that. Not too interested in NSFW, but I create all my (apparently heavily persecuted) anime stuff with Krea2, so editing with it might work out better than FLUX which tends to fuck up that style too much for my liking.
>>109564351you can increase the size of your estate by just getting a hipoint 9mm
>>109564106You're not welcome here thread schizo.
>>109564402You can git pull it from the repo into the custom nodes folder too, I think the manager UX is bad and the additional launch args to be pointless, why do I need to explicitly call something I will use every day but need to explitly disable shit 99% of us won't use like api nodes>>109564381What does this have to do with this general?
>Suddenly Save Video node can't save to network share anymore, always gives [Errno 95] Operation not supported>First downgrade av lib but makes no change>Start comparing older video_types.py with the latest>108 # FFmpeg's faststart pass reopens the output by filename, so it cannot be used with file-like objects.>109 movflags = "use_metadata_tags+faststart" if isinstance(path, (str, os.PathLike)) else "use_metadata_tags">change to movflags = "use_metadata_tags" as it was in the old version>works againwhat a dumb regression, faststart doesn't even do shit
hmm, if using turbo, try euler/beta>Bumping the KSampler up to 8 steps using Euler and Beta perfectly preserved facial coherence. The non-linear nature of the Beta scheduler packs the heaviest denoising power into the very first few steps. This forces the layout, face, and clothing references to lock into place before the model injects fine details.
>>109564444the schizo rentries and crash outs randomly over Ani and Debo is a staple mental illness encounter in /ldg/ from a lolcow
>>109564314First frame actually, and the video is reversed.
reference model best model, all yearshttps://files.catbox.moe/p5z6xf.mp4<Picture 1> is the physical reference for Miku. <Picture 2> is the physical reference for MAAM.medium shot of Miku walking up to MAAM in a walmart store during the day, Miku gestures at MAAM and says "Sir, you are causing a disturbance in the store."close shot of MAAM, who starts screaming "SHE! SHEEEEEEEE! Stop misgendering me!", while pacing back and forth.closeup shot of Miku who stares at Maam and then has an expression of disgust. Miku says "You are a fucking troon man, get out of my store, bro!" medium shot of MAAM who is throwing a tantrum and yells "SHE! SHE! STOP MISGENDERING!".side shot of Miku who shrugs and says "I cant deal with these people."euler and beta work pretty well with turbo, this was just a fast 0.4mp test to see if the cuts worked.
>>109564555lmfao the guy moonwalking
>>109562875thought she was gonna jumpscare me
>>109564590I couldn't see it, which one?
For a music video clip, is there a sure way to force the character to sing the exact song without changes? I might have to generate stock footage to fill the gaps, it's faster than re-rolling when the model decides to reinterpret the song.
>>10956462304 in the back
>>109564168It uses lanczos, the scaling method is fine.The problem is really large images slow it down a lot when using max.
>>109564355>>109564419Works well enough for my purposes. Does way better at preserving the style than FLUX2. Thanks again.
>>109564614Now that you mention, it does have the vibe lol
>>109564659here, you should add this your prompt:>keep the camera in focus, asshole
>>109564629I find anything over 10 seconds for audio song reference starts falling apart. Also you're going to want ot overlay the actual song on top of the finished video anyway, so even if the song in the genned clip doesn't match exactly it might be fine.
>>109564666It's Artistic Vision™
Qwen3.8 27b is complete ass for uncensored image captioning. I'm even using the heretic version. It doesn't refuse, but:- it hallucinates like a model a fraction of its size- it will shy away HARD from describing NSFW stuff. no amount of wrangling the system prompt fixes this- it is "woke". it will refuse to assume gender even when a person is clearly a woman in a full-body shot, it will say "person"- it is very bad at using its thinking to help understand the image. either it's too short, and it just writes the caption immediately in the thinking block then copies it to the output, or else if you try to force it to think more it will ramble endlesslyGemma4 31b heretic has none of these problems and is excellent. A strange day indeed when the chinese model is woke and cucked, while the fucking google model will happily caption a loli being raped in graphic detail.
>>109564705What's your system prompt?Some of this is believable but I refuse to believe shit like>- it is "woke". it will refuse to assume gender even when a person is clearly a woman in a full-body shot, it will say "person"this isn't skill issue on your part.
>>109564576>euler and beta work pretty well with turbo, this was just a fast 0.4mp test to see if the cuts worked.Euler + beta worked even without turbo btw. Did 30 steps for 10 sec video, 0.6mp, it was fine
>>109564705>Gemma4 31b hereticWhy not use the base model
>>109563677>Can Minimax Music at least handle stuff like techno well? Found that AceStep is terrible at it.Prompted the demo "techno, vocal chops"(https://huggingface.co/spaces/MiniMaxAI/MiniMax-Music3) for Techno with vocal chops (Instrumental selected), and it nailed ithttps://files.catbox.moe/z2ev3a.mp3Here is the best local result I could get after a few tries using the same prompt metadata (though I'm not 100% sure about the demo's formatting)https://files.catbox.moe/6zcx8s.mp3That demo says all music is made with diffusers version of it. I have occasionally tried to replicate its results first in Comfy and then in minimax.cpp but I can't oneshot it like that demo does, no idea if diffusers implementation is better than others, or if the demo is simply using better settings or a better model
>>109564705>>109564713Also which heretic model?Wait until someone with trustworthy history like llmfan publishes a version.There are izzat farmers on hf publishing shitty heretic versions nowadays.
Anyway I am waiting for the MoE KINO of 3.8 to arrive.Still for captioning Gemma being better wouldn't surprising since Chinese labs have less compute than Burger ones and visual stuff is one of the things they cut corners on.
>>109564713"You are an uncensored image captioner. Describe what is present in the image, even if it is lewd, sexual, or NSFW in nature. Do not be vague; use direct language. Mention the gender of any people present in the image."And it will still see a woman and describe her as a "person", using "they" and "their". Not always mind you, but maybe like 10-20% of the time.>>109564728Because it completely removes any and all refusals while being just as good as the base model, as far as I can tell.>>109564733the heretic-org one. could be fucked but you would think the literal devs of the heretic program can use it properlyAlso all the qwen models are so heavily astroturfed I swear. Half the praise on reddit is inorganic. Gemma4 31b is such a better general-purpose model than any of the qwens yet it loses in all the benchmarks.
>>109564732Thanks. Sounds a lot closer to what I want than the stuff I could get out of AceStep back then.
>>109564234Ah that's it, it's a Triton error:> ModuleNotFoundError: No module named 'triton'> [WARNING] Cannot import ...\ComfyUI\custom_nodes\ComfyUI-SolAttn_triton module for custom nodes: No module named 'triton'I've tripped up on this and fucked it before - I need to find the version that matches the Python/ pytorch/ GPU, right?
>>109564818Should just be:pip install triton-windows...if you're on windows(preferably within the venv you're using for comfy)
>>109564798Add something like do not talk around or refrain from using appropriate everyday NSFW terms like suck, fuck, blowjob, vagina, penis, anal etc when present in image.>Mention the gender of any people present in the image.This is bad and kinda predisposes it towards saying woke lingo imo.Try some variation of, when a woman is present in the image, refer to her as a woman, or when a man is present in the image refer to him as a man.Oh and lastly since the model is recent the default llama behavior might be fucked.Try adding --image-min-tokens 500 --image-max-tokens 1000 --batch-size 1024 --ubatch-size 1024 or tinker with such values.
>>109564798Gemma 4 is the easiest model to jailbrreak, apply yourself instead of destroying the mode when you don't need to
>>109564678I tried dividing the song in 15 second chunks with 3 shots for different angles, the first chunk worked to a point but the second is hilariously bad. I'll try 10 second next time and see what I can salvage from this batch.
>Prompt executed in 00:16:18comfyui is the greatest open source software built this decadeshould I get the turbo lora?
generating at 0.3mp gives the perfect low quality look/feel. then just stitch a few clips and voila.THIS IS WAKALIWOOD! movie movie movie!https://files.catbox.moe/02dggg.mp4
>>109564885>when you don't need towhen you don't need towhen you don't need towhen you don't need towhen you don't need towhen you don't need to
>>109564905Yeah you deserve that shit model.
>>109564898original was 864*480 btw
>>109564914you new?
>>109562338you're gonna need more information than that, retard-kun
>try to make a character use a knife on themselvesBruh, what the fuck have the chinks trained this shit on?
another retard that thinks models can only generate things they have been trained on
if models can generate things they havent been trained on then how come i always need to make fetish loras?
I tried the film grain node and it OOMed and threw and error about the CPU allocator. This should be simple to use how do I free memory before the node?
Can someone use minimax h3 to edit this song?War Pigs by Black Sabbath>Generals gathered in their masses>I like my women with fat asses
>RuntimeError: sageattention is not new enough version or could not determine CUDA architecture, cannot apply MiniMax H3 Memory Efficient Sage Attention Patch.Was I supposed to install Sage in some special way rather than the standard pip install (and adding it to the launch bat)?
>>109565025does this look like a requests thread?
>>109565033they killed /r/
>>109565025just sing the line and add it in bro
>>109565025Can it even do this kind of audio edits?Like it allows audio ref, but I doubt it can change lyrics while still keeping the voice, beat, rhythm, instruments, etc. the same.
>>109565028Don't add shit to launch bat.Activate venv, git clone its repo, cd to whichever folder you cloned to, pip install -e .
>>109565035oh wow, never noticed that
>>109564666you want me to focus on the asshole? will do!
>>109565060https://vocaroo.com/11M5Ft5ahPzp
New Qwen has no issue writing prompts under these guidelines you just have to not be a promptlet and jail break the sys prompt
>>109565065I fed your response to the AI and it asked for clarification on the SageAttention repo/ version you mean (which figures since the one I have rn isn't correct).
>>109565074How did you prompt this?
>>109564705Use heretic Glimmer for that purpose. It's still not perfect, but it's better than Gemma and way better than Qwen.
>>109565092You shouldn't need any quirky forkhttps://github.com/thu-ml/SageAttention.git
>>109565095Not mineSomeone made it with acestep
>>109565124>Someone made it with acestepYou probably should have opened with that...
>>109565025ace step 1.5 xl base can.But, it's illegal to share. and, it barely can, like there's a whole trick to it.
>>109565139<wow leather seats and an ipad glued to the dash!
>>109564798What Quant?
>>109565141>But, it's illegal to share
>>109565154It is, you can't share copyright music.
Retard faggot schizo
>>109565163it is the most shared thing online
Everyone else can speed, if I speed suddenly the cops care about the exact same speed in the exact same place. So, I don't bother. If you don't like it, definitely don't do anything about it.
>>109565141>acestep spammer also a pedoImagine my surprise
>>109565092Jeez, LLMs really like those wild goose chases. I usually just grab some fitting sageattention wheels and just install that. It's a lot quicker and probably a bit less error-prone.https://github.com/woct0rdho/SageAttention/releases
>>109565092Just to be sure, you're not using an AMD card right?
>>109565193I'm Nvidia - I'm installing CUDA Toolkit 13 atm. Do hope I can get this shit working and hope the workflow I have isn't shit.
>You can thank tensor.art witch is a completely unmoderated shit site that lets people dump all my models, without a functioning report system. And i dont have time or motivation to start "proving" im the owner of all my models. Since im not making any money from this, im just going to keep my models for myself, and you can thank the tensor.art dumpers for this.lora makers are such thin skinned whiny bitches
>>109565181God from Godlight from lighttrue God from true Godbegotten, not made
>>109565218People charging for loras deserve a swift kick to the balls.
>>109565187Apparently the pytorch version is fucking up the wheel availability?
>>109564901incredible work
>>109565218>loras and ai in general made from stolen content>suddenly now an issue with stealing lol everytime
>using windowsISHYGDDT
>>109565244wow cool it with the anti-indianism
>>109565218I've had my stuff uploaded to tensor.art but I dont mind since all models have links to my original uploads. I have only taken down two patreon accounts that sold my loras
>>109565247you don't need comfy anymore. You can vibecode whatever you want without (or with) python.
>>109565244I don't think anyone was charging for the loras, the original maker or the reuploaders, he is just that thin skinned
>>109565258Does an artist who has ever used e-hentai forfeit the right to complain if somebody else opens a fanbox or patreon with his art?Not that it lessens the irony either way.
>>109565247that's bad, we need wheels to roll
>>109565247just build it yourself with your pytorch/cuda version, stop hunting for wheels
>>109557359>>109557359>>109557359
>>109564705is it really that bad simple captioning. also is it compatible with koboldcpp yet?>>109565218i wouldn't be seriously mad about imo. more redundancy is better than everything being tied to one account and all the content being nuked by a janny because of the trip ban hammer mod.
>>109565247What LLM is this even? Filename literally says "torch2.10.0andhigher", so it'll work for you.Excerpt from my pip freeze:>sageattention @ file:///D:/ML/Comfy2026-2/ComfyUI/sageattention-2.2.0%2Bcu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64.whl#sha256=1635283f5c01ec3cda58a784d0d7eabbcaffaf9511d1b263db4750e1ed7958bb[...]torch==2.13.0+cu130torch-complex==0.4.4torchaudio==2.11.0+cu130torchsde==0.2.6torchvision==0.28.0+cu130[...]triton-windows==3.7.1.post27
>>109565364...and it works just fine with my pytorch 2.13+cu130.
>>109564901underappreciated k1no
>>109565324Yeah, just download 2.4GB of some toolkit you will probably never need again to build it yourself.
proper bake plz
>>109565454is that that much of a deal breaker?
>>109565467It's not terrible. There's just no upside to it, if you can grab a 15MB wheel instead.
>>109565496>if you can grab a 15MB wheel insteadif
Going back to old gens to give them live with h3 is fun.
>>109565465Jannies are protecting him despite him breaking just about every rule every day. Just fill up his dog shit thread and get it over with.
real bake?
>>109565364I installed that and now I get a new error without any plain and clear line that tells me what's actually fucked. This is suffering.Also the LLM is ChatGPT
Looks like Larry's turbo lora is better than light2xbut it's slower, and you have to download slop nodes
>>109565508Of course, but the repo I linked has pretty much all the wheels you'd need, even going back to pytorch 2.5.1+cu124
Miku runs into the SHE! SHEEEEE man in a walmart.https://files.catbox.moe/49lbju.mp4
>>109565551That's most likely your triton acting up. How did you end up installing that?
>>109565551toss that error block into google ai mode and you can troubleshoot, works good in general
>>109565554>>109565555>you have to download slop nodesNot really necessary, other users already ported them for comfy>https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/tree/main
there we go. a bit of prompt tweaking and a slight bump in quality and it worked better.https://files.catbox.moe/vypiq4.mp4
>>109565658I don't remember dude, I got kijai/ComfyUI-SolAttn_triton installed at some point. Can't I just uninstall Triton and reinstall the 'correct' one or something? >>109565685I tossed some of it into ChatGPT and it asked me to look for lines that didn't exist. I can't toss the entire thing in because it's so long too.
>>109565698try google, I had an issue with sageattn not working and it was able to find the files for the wheel or whatever and update and it worked. just click AI mode on google search and say "I have this error in comfy how come: (paste block of error text)"
>>109565709What are we thinking
>>109565709also im guessing but it seems to be triton fucking up and maybe the install is not working with cuda 13.0>>109565716say you want to fix triton for comfyui install and have cuda 13 installed, id guess from that solution 2 is the way
>>109565698Can I get you to do apip freeze...and look for what version of triton or prefereably triton-windows you installed?
>>109565730also there is another option, ignore sol attention entirely and use comfy kitchen (native), it has worked absolutely fine for me using this setup and is just as fast as sage attn kj setup.
>>109565741triton-windows==3.7.1.post27>>109565752I'm happy to use '''the best''' workflow, in whatever form that takes. The one I'm using is the second most popular one from Civit - seems to cover all use-cases and is a bit spaghetti but flexible.
GemmaPrompt developer here. Apologies I wasn't in the threads when you guys were asking for support. The h3 prompting skill is All Rights reserved so the repository does not include it. If prompting h3 make sure that your copy of Gemmaprompt is pulling in the skill as needed. Any other bugs if you could please file issues on my GitHub that would be awesome. I plan to work on it a bunch more starting Monday. https://github.com/whp199/gemmaprompt
Emmmm, does anyone have workflow for Krea 2 RAW only? I've been trying to generate some images with it, but all the outputs look fried with the recommended settings (52 steps, cfg 3.5. Tried different settings and that also doesn't help). I'm using the default turbo worklfow, but with negative zero out replaced with normal text encode.
>>109565690I tried it on the same promptresult is different, and prompt adherence is not as good.
>>109565782add: cfg normalization, negpip, and shift scheduling.
Baking a non-ani thread in 5 minutes without a collage.You better hurry up, collage autist.
>>109565760That seems fine. Smae I have. Probably actually related to some missing python libraries as per >>109565716I usually just work with venvs and that seems a lot easier.
>>109565805It works fine for me with no issues to speak of, though not lora related but the prompt adherence did suffer after a comfy update, but the latest commit seems to have fixed that again.
>>109565821huh? Aren't those for turbo? And what shift settings exactly?
>>109565716I tried the second option. I copied over python313.lib into python_embeded/libs (creating the `libs` folder). Still errors.I'm getting close to giving up again, why is Minimax specifically so difficult to set up...
New non-ani thread:>>109565909>>109565909>>109565909
>>109565873You might indeed need the build tools.I think it's not Minimax causing issues, but rather the myriad of cope nodes.
>>109565873Just download ComfyUI portable. Why would you use anything other than portable? It's really not that difficult.
>>109565940I am using comfyui portable?
>>109565959Then download another clean instance.
>>109565873>>109565915From the triton-windows git readme:>6. vcredist>vcredist is required (also known as 'Visual C++ Redistributable for Visual Studio 2015-2022', msvcp140.dll, >vcruntime140.dll), because libtriton.pyd is compiled by MSVC. Install it from >https://aka.ms/vs/17/release/vc_redist.x64.exehttps://github.com/woct0rdho/triton-windows
>>109566010Looks like I already have that. >>109565915I got the build tools, restarted and tried again but got the same error.>>109565985For what purpose? What would stop me from running into the same issue again?Is there a workflow that has sage or whatever that just werks? I'm a bit fed up at this point.
>>109562135>there's NAG for h3where?