[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Previous: >>109557402

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
miku nigger ignoring the call of black migu over the white boycalls
>>
>>109561703
what about this then?
https://huggingface.co/MiniMaxAI/MiniMax-Music3/discussions/5#6a7f5da5341a0766b9c67ccd
>>
>>109561698
https://files.catbox.moe/cu88xe.webm
>>
>>109561740
Read a bit further on the comment thread, Ostris addresses it
>Unfortunately it is not in there. That is the audio VAE encoder, not the RVQ encoder.
I had Claude look at the file independently of Ostris and it came to the same conclusion, serveurperso most likely got the two confused.
>>
File: ms_00013_.png (1.51 MB, 1536x1536)
1.51 MB PNG
>>109561698
Reminder to submit your video gens for Sexy Jam 1:

https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewform

All sexy h3 gens welcome. Troll submissions rejected.
Deadline is after the weekend.
>>
File: 1438789897882.jpg (35 KB, 321x362)
35 KB JPG
GemmaPrompt anon, you still have a lot of work to do. My manual prompts still end up much better in I2V because I've memorised a lot of the models quirks and eccentricities whereas GemmaPrompt still writes certain actions sub-optimally.
>>
Debo says he didn't bake any of these duplicates.
Is he lying, or is it really just ani gaslighting as debo? and simultaneously trying to get his own rentry removed?
>>
https://files.catbox.moe/ipj4zs.mp4
>>
Women don't act like this
>>
https://www.youtube.com/watch?v=sVx1mJDeUjY&list=RDsVx1mJDeUjY&start_radio=1
how many german fags got misleaded due to kino
let us connect
>>
>>109561764
You are feeding your output back into Gemma so it can learn from it's mistakes anon r..right?
>>
>>109561769
>is it really just ani gaslighting as debo
Ani hasn't even done anything to look like debo. It's always been Ani. It started when the Ani rentry was added back in december.
>>
>>109561798
no, chatgpt is far better at that.
>>
Comfy question. Some lora loaders have optional clip passed through, is there any difference?
>>
>>109561878
There's a difference for SDXL-based models, if you don't mutate the CLIP then it'll just apply the full lora regardless of trigger words.

This makes the most notable difference for loras which contain multiple concepts. Any lora trained on a tagged dataset technically contains multiple concepts but you would most often be able to tell the difference in cases of multi-character loras.
>>
>>109561798
>try gwen 3.8 yesterday
>make tampermonkey script, like 150 small lines, reiterate 3 times because he made some retarded bugs
>starts losing context after like 5 messages and adds the bugs he removed
how would a much less smarter model handle the user throwing infinite prompts at it?
>>
File: 1755433380568384.jpg (227 KB, 1345x1866)
227 KB JPG
>>109561698
What can i expect Generating 10 seconds Minimax video with 16gb VRAM + 64gb RAM ??

Amd card by the way
>>
If you can get the ddim_uniform scheduler to work, you can get some pretty high gens if you are doing anime. But it's very unstable.

>>109562022
>AMD
No idea, but probably not great. Try it and good luck.
>>
>>109562022
>What can i expect Generating 10 seconds Minimax video with 16gb VRAM + 64gb RAM ??
>Amd card by the way
Npo int8 convrot so your life is pain. Expect 360p max resolution and it takes 15 minutes a video.
>>
>>109562067
>high gens
High quality, that is
>>
why does H3 keep giving women hair ties around their wrists, and how do i stop it?
>>
>>109562101
>why does H3 keep giving women hair ties around their wrists, and how do i stop it?
Nothing you can do about it since we don't have negs. I feel a similar way when my LLM inevitably adds an ankle bracelet eventually when I'm making foot fetish gens even though ankle bracelets are retarded and Indian and not hot at all
>>
>>109562101
bare wrists or no accessories maybe
>>
>>109562112
there's NAG for h3
>>
>>109561758
Thanks for the heads up
Hopefully Julien gets banned for good this time
>>
>>109562135
come on deepmeepbeep get it done!
>>
>>109562135
>there's NAG for h3
Haven't seen anyone talking about it or using it, and doesn't it increase gen time too
Extra hair ties have never been a problem for me though because I prompt for pigtails with cute little bows
>>
>>109562101
tell it not to but be explicit though; say "she has no bracelets/ties on her wrists" rather than "bare wrists"
I didn't test it much but I negative prompted for eyes, nose, ears, and it all worked, so think it'll work for bobbles
>>
>>109562181
Often if you say things negatively (ie don't do x) it will just do x. Bare wrists might be the better option.
>>
>>109562219
I know, cfg is a bitch in every other model. It was a week before I tried it myself because I just assumed it wouldn't work
>>
>>109562131
>>109562181
"bare wrists" fixed it
>>
File: h3_00377_.mp4.webm (245 KB, 832x640)
245 KB
245 KB WEBM
>>
What's the correct wording for H3 to continue the next shot from last shot? So it doesn't completely do a new scene from the description?
>>
>>109562338
Did you try "The scene continues" or similar?
Most of the time, unless I denote a camera change, the scene stays the same. Unless you are using a strange scheduler.
>>
how do you avoid a penis generating if you are prompting for the crotch area?
>>
"," sometimes making the movement stop so ill just use "while" and "and"
>>
I'm new to this and i'm greatly confused. Is there a guide for dummies?
I've installed SwarmUI and downloaded a few lora
What's next?
>>
>>109562364
i've gotten mixed (but tending toward successful) results with, "safe for work"
>>
>>109562377
Ask chatgpt
>>
>>109562364
Have you tried "no penis" or "no genitals"?
>>
>>109562364
Have you defined that your subject is a female, woman, girl, eunuch, etc..?
>>
>>109562387
no. i always have a mindset of implied removals since explicit ones never work
>>109562390
i'll try that
>>
>>109562364
Don;t worry bro, I got you. The word is "post-op"
>>
>>109557462
>>109558134
>>109558725
>>109558919
these are so 2000s coded, absolute kino. I can easily imagine most of the scenes in a parody-type of movie as cut-ins. Maybe something for /tv/
>>
File: arunnijngjak.jpg (57 KB, 500x500)
57 KB JPG
I tried genning.. THAT.. with h3..

It worked..
>>
>>109562428
yeah
it's freeing
>>
>>109562390
seems to work. though it will be a problem if i need a male mannequin
>>
>When you get everything tuned in and every seed is a new piece of kino
this can't be healthy, it's been hours
>>
>>109562428
>>
File: 1773425591259694.jpg (475 KB, 2800x3293)
475 KB JPG
>"girl stroking penis nonstop, girl head turns camera"
*Handjob stops after head turn to camera*
>"girl stroking penis nonstop, girl head turns camera, "girl stroking penis nonstop"
*Handjob continues nonstop, even when head turn to camera*

Prompting with Minimax is really weird man...... Who the fuck the said prompting is better than LTX...........
>>
>prompt suddenly gens consistently what I want
I'm in heaven.
>>
>>109562457
how would h3 know if you want her to keep stroking or not unless you tell it?
>>
File: 1770318232517145.jpg (2.37 MB, 2044x3550)
2.37 MB JPG
>>109562466
I want Sulphur H3 so bad so i can just make a simple 1girl,handjob prompt
>>
>>109562457
[Shot 1]
"<Subject 1> starts stroking the penis. As she's stroking the penis, she turns her head towards the camera, and she..."

Use the proper format.
>>
>>109562462
proof?
>>
>>109562481
Thats what im use. i need to prompt "she's stroking the penis" twice to make it work
>>
Kinda struggling with the turbo loras for Minimax H3 Ref2VA. The v0.1 lightx2v lora performs a lot better overall, but it tends to always do little bullshit like the floaty bits here, or the guns of the mech having some backward facing artifacty guns.
>>
>Got a good scene and movement for handjob
>Handjob sound looks like rubbing a wrinkled paper
Ah fuck this, ill just mute it
>>
>>109562492
Write it what to do..
"she's stroking penis, she stops for a moment to turn her head towards the camera and continues stroking penis", or something like that.
>>
>>109562499
The larryvrh loras aren't meant for Ref2VA at all, but they produce a result with less artifacts, but a bit more smugded.

Am I doing something fundamentally wrong, or should I just wait for lightx2v to release their v1.0 Ref2VA lora and hope that is fixes everything?
>>
>>109562499
The lora is still a work in progress, but you can try using some different samplers. er_sde or seed_2 maybe.
>>
why is it obviously trained on sex but none of the sounds are there? it's like they muted those videos during training
>>
>>109562480
Wouldn't fine tuning H3 be less effective since it's a distilled model? That sulphur guy asking for $10k is kinda scamming, since it's unlikely to produce the results everyone is expecting
>>
>>109562517
Probably trained it against outputting that, same as with genitals.
>>
>>109562481
I was under the impression that you had to be extremely specific in your Minimax H3 prompts. "Stroking" seems very unspecific to me, and could possibly mean a lot of things. I would've thought something like "one of her hands is gripping the penis while moving it up and down along the shaft continuously" could work much better.

>>109562516
Thanks, I'll give that a try.
>>
>>109562543
>same as with genitals
it generates implied genitals if you involve motions related to someones lap. how can it know to do that?
>>
>>109562543
anyone tried prompting stuff like "copulating" or something else
>>
>>109562586
anon the loicence forbids it
>>
>>109562394
I just tried "no penis, no skin showing" as well as describing what he's wearing and 3 times in a row got no penis or weird abomination. Without it, always got some type of penis worm. Seems to work
>>
>>109562607
>penis worm
for me, it's the stalagmite penis
>>
>>109562543
>>109562531
yeah this model is so pathetically scuffed that I doubt we will ever get proper dicks or pussies other than through overcooked jeet loras that destroy any prompt adherence and kill your gen. I just don't see it, sorry to say
>>
>>109562628
I think you're overly pessimistic. We'll see who was right in a couple months.
>>
File: file.png (2.65 MB, 1986x1507)
2.65 MB PNG
Man, animators are TOAST
https://files.catbox.moe/2h2cci.webm

>>109562531
You can finetune a distilled model if you know what you are doing. Not that I know if he knows what he's doing.
>>
remember the anon that demanded that /ldg/ grovels and apologizes to him because he was right that z-image base would never be released because of chinese culture
>>
>>109562390
how would that help? a female can obviously have a penis..
>>
>>109562657
>a female can obviously have a penis..
how?
>>
>>109562657
some humans are born with no arms but I have a feeling most humans you gen will have at least 2
>>
>>109562642
They were right
>>
>>109562657
just end it. get it over with. stop bothering normal people with your mental illness. we've had enough
>>
>>109562674
anon...
check huggingface...
>>
>>109562688
hmmm
nyo
>>
>>109562688
nta, but what we should check?
>>
>>109562688
Hmm I see something called an image. But no base…
I also don’t see an edit model anywhere.

Maybe something got mixed up due to cultural differences here.
>>
>>109562701
The main z-image repo where they uploaded the base model 6 months ago

>>109562703
image is the base model
>>
File: 744654.webm (2.93 MB, 448x256)
2.93 MB
2.93 MB WEBM
post kinos. i know you're hiding some
>>
>>109562674
>>109562703
omg were you that anon hahaha
>>
>>109562709
Oh. But that’s not the base model. You might think it is, but it’s actually not.
>>
>>109562733
Unfortunately for you it is. It is the model z-image turbo was distilled from. I don't understand, why keep coping after all this time? Are you down to your last bit of izzat?
>>
File: IMG_2132.jpg (528 KB, 4400x1356)
528 KB JPG
>>109562740
I can see how this would be confusing to you, as you were promised the base model. But as you can see from this image in the model card page. It is actually not the base model.

I understand this may be upsetting for you to read but I implore you to detach your ego from the situation
>>
H3 or LTX if I just want some simple camera motion and simple gesture from the subject? has to be high at high res
>>
>>109562752
So you don't actually want the z-image turbo base model? You want a different model? Which one?
>>
>>109562759
ltx is best for high res
>>
>>109562740
>It is the model z-image turbo was distilled from
it clearly wasn't, since even lora training is broken with that model, let alone a proper finetune.
z image turbo was distilled from the true base model which they never released
>>
>>109562760
Resorting to strawman arguments I did not make to bolster your argument only serves to weaken it.

Might be time to brush up on your Chinese culture lessons and learn to be content with reality rather than lash out at me for being the bearer of news.
>>
>>109562778
>It's not the base model because I say so, you fucking chud!
k

>>109562779
>repeating what I said is a strawman, you fucking chud!
k
>>
>arguing about an outdated obsolete model
there are better ways to spend your time anons
>>
>>109562796
Exactly, we need to talk about how anima is garbage compared to krea 2
>>
>>109562793
No actually, phrasing question in response to a statement I did not make is actually the strawman argument here.

But I can see based on your perception of the release status of z image base that reading comprehension is not your strong suit.
>>
>>109562812
>I don't understand what you posted, so it's a strawman you fucking chud!
k
>>
>>109562818
Yep, we won. It's clear you don't have any idea what's going on, or, more likely, are yet another Chinese shill I've now defeated. Good bye.
>>
File: 52646875037487.mp4 (3.91 MB, 640x832)
3.91 MB
3.91 MB MP4
>>
i just updated my nvidia studio driver to 610.88
am i going to be okay?
>>
>>109562829
What did we win?
>>
>>109562842
Chinese culture.
>>
File: AnimateDiff_00024.webm (1.97 MB, 1024x1024)
1.97 MB
1.97 MB WEBM
>empty h3 t2i prompt

Post your results.
>>
File: 1036609785291320.mp4 (3.58 MB, 640x832)
3.58 MB
3.58 MB MP4
>>
>>109562842
A lifetime supply of AIDS
>>
File: 687918468206581.mp4 (3.41 MB, 640x832)
3.41 MB
3.41 MB MP4
>>
File: 563369463902228.mp4 (3.88 MB, 960x544)
3.88 MB
3.88 MB MP4
>>109562848
huh, pretty mundane ony my end
>>
>>109562894
If that's supposed to be grace she should have greyish blue eyes.
>>
Sheesh prompting ref model is such a pain in the ass, I have a decent reference image and have a prompt for my character to enter an empty scene but for some reason she's coming in as tall as a doorway. Gonna try fixing it with another reference image depicting her scale in the room.
>>
File: 631153047275475.mp4 (3.6 MB, 960x544)
3.6 MB
3.6 MB MP4
>>109562929
True, but I didn't specify it in the prompt.
>>
>>109562894
How did you prompt for that transition? Do you use the [Shot] syntax or just rawdog it?
>>
>>109557688
right, what i meant was that adding a lora on top of a model doesn't increase the vram requirements by very much. in case you're unaware a lora is a bit like a filter on top of a model, it's not a mini model you can run by itself, which i clarify because your suggestion of a lora theoretically *lowering* vram requirements implies you might have been under that impression. sorry i don't know the answer to your question about training on 8GB though. generally training is actually not significantly harder than inference, but it is typically a *bit* harder, and training your own lora is probably not something to even think about until you have a grasp of how generating stuff works. i don't know your financial situation but a good heuristic is to try to stay ~roughly ahead of 50% of the pack so if your chosen gpu is nvidia and has like 12-16GB of RAM most of the popular toys will target your bracket, 24-32GB+ chads will of course have an easier time and run anything easily but you'll always be able to do 95% of the things you want to do (if a bit slowly) as long as you're in the upper mid tier.
as for LLMs there's only a very light touch of "x is better at y" it's mostly just a raw intelligence scalar, but the thing you're wanting here is a big corporation with a harness to solve the problem for you by it running a bunch of web searches and condensing the answer down for you, don't even think about local for that, seriously just ask chatgpt it'll do fine, the (localgen) next step up from that is quite a few hours of knowledge and effort you won't be able to bypass with a single 4chan post
>>
File: mpc-hc64_aIF9EcNfsi.png (748 KB, 816x482)
748 KB PNG
https://files.catbox.moe/gg21cb.mp4

This took so long but I think I've finally solved the blowjob horrific crunching noises problem
>>
>>109562997
lmao
>>
>>109562997
10/10
>>
>>109562997
kek
if anyone is having this problem, the real solution is
>overall_soundscape:
>blowjob sounds, gagging and muffled moans.
>>
weird shit start to happen when trying to gen longer than 12 seconds in h3 r2v. like it wants to change scene even if you didn't prompt for it.
>>
https://files.catbox.moe/8a9w0n.png (don't open in public or at work)
is anima base still the best anime checkpoint?
haven't been here in some time
>>
>>109563037
If you use spectrum, did you pull? I did and my gen came out like an acid trip in the background, subject was intact though.
>>
>>109562511
I get good results by upping steps to 10.
>>
>>109563066
that looks like turbo or high cfg
you can download the merged anime-turbo checkpoint
>>
>>109562950
Actually spent a while trying to fix this exact problem, had my character walk onto the scene twice as tall as all the existing people. I prompt wrangled it eventually with lots of "normal height , height of the other people in the shot" but the model seemed to want to interpret the character as the same size in the gen's frame as it was in the reference image. Might be easiest to fix by scaling the input differently.
>>
>>109562640
>animators are TOAST
I disagree. What's likely to happen is animators will just focus on fun stuff like keyframes and let AI do the boring shit. It will massively speed up the process and maybe improve qol for them.
>>
Is H3 genning without sound faster?
>>
so ref_image_0 is <Picture 1> and so on?
>>
https://gist.github.com/PierreHoule/947c6655a68279bb16f661ffdbef6ba7
is this a good workflow?
>>
>>109563123
That's right desu.
>>
>>109563118
sound only uses like 2% of the tokens, so removing it would barely make a difference to gen speed
>>
File: ComfyUI_Krea2__00022_.png (1.84 MB, 1024x1440)
1.84 MB PNG
6gb vram here
I'm okay with 0.2 megapixels
should I go for wan2gp or use comfy for h3?
>>
File: 1785345956999316.mp4 (230 KB, 736x576)
230 KB
230 KB MP4
I'm retarded and didn't realize she's supposed to be wearing pantyhose.
>>
File: ComfyUI_Krea2__00037_.png (1.2 MB, 1280x720)
1.2 MB PNG
>A community-modified version nicknamed a “heretic” build has circulated, which isn’t technically a fine-tune of H3. Instead, it removes certain layers and replaces the language model head to reduce refusals, while still requiring the original H3 weights to function. Users report it generates recognizable characters from popular media with fewer restrictions than the base model, along with reports of it producing nudity and gore when prompted.
where do I get this uncensored model?
>>
>>109563192
what's the point?
base is already uncensored and does that if prompted
>>
>>109563192
its the default model, retard. people got meme'd in to thinking the heretic qwen encoder somehow made a difference. it does not. if you cant get nudity and gore out of the box, you're a promptlet and the issue is entirely with yourself, not the model.
>>
>>109563192
Completely wrong. Heretic lobotomizes text encoders for video models. Refusals aren't an issue in video gen, it just needs a rich inner representation, which heretic breaks.
>>
>>109563225
>>109563226
>>109563227
okay
I'll try running it
>>
File: 1786673139080657.png (56 KB, 862x661)
56 KB PNG
>>109563225
>>109563226
>>109563227
>>109563174
which ones should I get?
>>
>>109563261
int8 convrot
>>
File: 946321819603841.mp4 (3.9 MB, 832x640)
3.9 MB
3.9 MB MP4
>>109562960
Just rawdog
>Track sideways behind foreground objects that repeatedly obscure the subject. Use each occlusion as a natural wipe, gradually moving closer until the final obstruction reveals an extreme close-up.
>>
File: 1759460879613365.jpg (29 KB, 746x512)
29 KB JPG
>>109563270
>>
>>109563261
int8 convrot pruned is the smallest & fastest, but you want both fl2v and ref2v models. text encoder is nvfp4, and the vae is the int8 convrot for video and fp32 for audio.
>>
The Kijai turbo fl2v loras (https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras) are awful, they instantly sloppify all your crap. Anons here who are claiming the 4-step model is better than the 8-step ones are even more delusional, 4-step is so much worse in my tests.
If you don't believe just look at the comparisons for yourself. Only one that is maybe OK is the lightx2v · 8 steps · str 0.75 · er_sde example.
https://jo-nike.github.io/h3-turbo-eval/scenes/singing-sustain.html
It also more or less proves that Spectrum will fuck your gen and is not as free as every retard here is claiming.
>>
>>109563261
depends on your card, convrot for 5x series, scaled for 4x and below
>>
>>109563226
If you have to put things like "white, creamy liquid" in your prompt instead of just "semen" or "cum", then it's censored.
>>
>>109563360
3060 6gb
>>
>>109562516
Improved things somewhat, although both trial runs had vehicles partly clipping through the container. I assume this is just bad luck with the seeds, though.
>>
File: Minimax H3 00147 .mp4 (939 KB, 992x992)
939 KB
939 KB MP4
>>109562848
I got a 1girl tradwife. Might be considered cheating though because I forgot to remove the prompt connection from a prompt helper and so technically the prompt was

integrated_multimodal_description:

overall_soundscape:

non_diegetic_music: N/A

>>109563348
This was done on the turbo 4step @ 1.25 str, but I ran it at 10 steps euler/beta .3mp. Audio issues with the loras are fixed by following the actual recommended step shifts based off of the lora you're using. 12/3, 12/6, etc. aren't universal.
>>
How do I get h3 to make a girl get fucked with the penis going deep in and out? Whenever I do i2v I usually only get the penis sliding in a little bit (like one third of its length) and then out a little bit again.
"fully inserted" does nothing
>>
>>109563348
All cope nodes and loras except sage is pure vramlet dribble. They don't care about quality loss or degraded prompt adherence just as long as they get their slop faster. Ignore them.
>>
>>109563399
If there's something attached to the penis, you can try describing the position of that something.
Like his crotch touching her ass or whatever.
>>
>>109563392
You can't see the artifacts in this? It's looks awful.
>>
>>109563341
okie
>>
>>109563348
Spectrum seems acceptable compromise if you look at the samples.
>>
>>109563088
Seems like I'm not that lucky...

Will probably have to wait for a proper Ref2Va turbo lora, before really trying Ref2VA again. Don't wanna wait like 20 min for a 15 second 1mp gen.
>>
>>109563445
increase the lora strength to 1.25 in addition to going up to 10 or 12 steps.
>>
I know this is a general for gooning to slop softcore porn of artificially generated women, but has anyone tried the MiniMax audio model m3? I'm looking to train a lora for it but since audio ai is lagging so much behind image and video idk where to start.
>>
>>109563445
looks like modern halo if it wasnt woke and was instead made by japanese pedophiles
>>
>>109563409
People optimized the quality out of the model so quick they are now complaining the model is bad and gives shit output.
"Model can't do sex!! It's so bad!!"
"It doesn't follow my prompt!!"
"Ugh... I get the same results every gen, seeds do nothing"
If you actually use the model as intended without the 500 cope nodes, you can generate 90% of shit anons keep struggle posting about every thread.
>>
File: 1774289141148502.mp4 (343 KB, 832x640)
343 KB
343 KB MP4
Attempt 1 was a failure. I want to try getting it right but this shit took 25 minutes to gen. fml.
>>
>>109563510
>YOU DIED
>>
>>109563510
why would that gen take 25 minutes? are you running a 1060 ti or something?
>>
>>109563539
7900xtx
>>
How to make Comfy save the queue and autorestart on crash?
>>
>>109563510
That's still a lot of gens over the course of a day.
>>
>>109563348
that entire page is focused on audiofaggotry
>>
>>109563542
haha fag
>>
>>109563477
Tried it and while it is kinda nice it the songs dont really blow your socks off and are a bit generic, but it's kinda a big leap when it comes to local quality. Strongly suggest you use a llm of your choice with the prompt guide to generate songs. While i tweaked the outputs of the llm a bit here and there, i didn't deep dive, so there might be potential. It also has the same duration issue, double the length takes about 4x long iirc.
But overall it's gonna be my go to music model from now on.
>>
>>109563510
ToT
>>
>>109563348
also these results are couple days old and do not contain anything from here `https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras`
>>
How do I get 100% chance of getting an instrumental in Minimax Music, I just want some cool jazz, but the moment it starts singing it's incomplete.
>>
>>109563568
it's the first local model that is somewhat usable. Comparable quality to suno except not behind a paywall so you can do infinite autistic tinkering.
I want to generate instrumentals and after testing for a bit it's definitely trained on royalty free crap, but has the potential to be saved by a lora. It's just that there is practically no lora training support atm, otherwise i would be tuning it right now
>>
>>109563598
grok told me to put this in the lyrics prompt and it makes speech much rarer than just leaving the lyrics blank :
>[Intro]
>(instrumental)

>[Instrumental]
>(instrumental)

>[Outro]
>(instrumental)
>>
>>109563477
It sucks. Too bland, can't get any good range with it. You either have to settle for AceStep which has ok quality but occasional errors (lyrics dropped or popping sounds) but good emotional range and depth or MiniMax which has very clean audio and doesn't seem to drop lyrics but is bland as hell.
>>
File: Test 00024.mp4 (3.1 MB, 1056x608)
3.1 MB
3.1 MB MP4
>>109563457
Alright. Will try that.
Hope we won't arrive at the lora only working well at 2.00 strength and 16 steps or something kek

>>109563486
This was made by a european podophile, though.

>>109563627
Can Minimax Music at least handle stuff like techno well? Found that AceStep is terrible at it.
>>
File: 1764429755098558.mp4 (109 KB, 832x640)
109 KB
109 KB MP4
>>
>>109563677
>This was made by a european podophile
Outing yourself?
>>
>>109563192
are you getting your info from Gemini Flash?
>>
>>109563718
no
I googled h3 for 6 gb card and this popped up
https://www.mindstudio.ai/blog/minimax-h3-run-locally-guide
>>
>>109563729
guess which model wrote that AI generated blog post
>>
>>109563683
>no flashing screamer
:(
>>
>>109563457
That still doesn't seem to work so well.

>>109563690
Don't think outing yourself as a footfag on anonymous image board is that terrible.
>>
>>109562997
it's so bad compared to ltx
>>
>>109563782
parading your mental illness should come with shame
>>
>>109563782
Is your summary prompt fucked up? From testing, I've learned that it'll prioritize the summary over the detailed description, and even the [Shot] tags. Also the sound tags can cause ghosting if you've specified specific sound effects to play during certain actions instead of prompting them in the shot itself.
>>
>>109561751
sad
>>
minimax music 3 is so good at normie music styles like drill rap, the comedy potential is there. wish it could do audio.
>>
If I make one 10 second clip of fake star trek every day, I will have full episode in about 260 days.
>>
>>109563858
With an RTX PRO 600 genning 10 second clips continuously you could probably get a full episode every 2 days. Just slop together a buncha prompts and queue em up. Someone will probably automate this and post them to youtube.
>>
Anyway to make animation from similar images that do not have exactly the same shadows etc..
Like have one ref image and model edits the rest to look like that one?
>>
>>109563957
You question is confusing. You can provide multiple reference images for multiple characters. You can edit a video and replace a character or multiple characters from reference images.
>>
>>109563802
Summary is quite concise and I can't spot anything that could really throw the visuals off:
https://pastebin.com/z4RjNdZJ

Other samplers work a lot better, as in this example:
>>109563374
>>
>>109563949
I'm the bottleneck not my gpu

>>>/wsg/6215059
>>
File: 9595137956.mp4 (3.68 MB, 896x576)
3.68 MB
3.68 MB MP4
>>
>>109563995
>
hilarious lmao I love the steel growing
>>
File: file.png (26 KB, 741x342)
26 KB PNG
How do you install kijai/ComfyUI-SolAttn_triton? I cloned its git repo into custom_nodes, no requirements.txt. The github page also has no instructions
>>
>>109563966
If you gen like 24 images with different poses that could still be put in sequence to make an animation, they wouldn't match since AI can't gen consistently.
Basically, is there a node for Comfy that can replicate the looks of one of the images and copy it over to other 23?
>>
So, reference images scale down to the size you're working at first, right?
So if my output file is 0.3mp @ 4:3 (640*480)
is my reference image being scaled down to a max side length of 640, or 480?
>>
>>109564012
You need to use a video model like H3 or Wan or LTX
>>
>>109564010
I installed it exactly like that.
You rebooted comfyui, and if you kept the browser UI open you pressed r, right?
>>
>>109561698
New to this. Is there anything offline that comes even close to what's available for subscription in flexibility and quality (chatgpt, gemini, etc.)?
>>
>>109564050
Yeah. Do I need to rename the folder to prepend it with `kijai/`? Or nest it in a kijai folder?

Also I can't figure out how to install a node through Comfy UI via Github URL. All the instructions I can find reference the old comfyui manager panel
>>
>>109563995
>>109564001
I'm serious, please post this on fb for max boomer rage.
>>
>>109564023
It tries to match pixel count
You might want to experiment with max for better quality though. (Don't forget to manually downsample really large images to 1MP or whatever)
>>
>>109564072
No.
What node in the workflow is erroring out?
Maybe it's a retired node from previous commits or something.
>>
File: debo_ds_k2_00006_.png (2.02 MB, 2048x1101)
2.02 MB PNG
>>109563949
this is my dream and its close to a reality with h3. the biggest problem to solve now is a script generating LLM that would be able to orchestrate the h3 prompting in a way that is vaguely narratively-consistent
>>
File: file.png (215 KB, 1825x1658)
215 KB PNG
>>109564098
>>
Does anyone have dual sampler setup for Krea2? Like raw + raw with turbo lora. Almost every seed is the same with just turbo and I'm not good enough at comfy to make my own one :(
>>
>>109564052
You do mean image models right?
>general purpose SFW text2image
Cloud is slightly better.
>artistic SFW text2image
Local is better than the big players but roughly at the same level as NovelAI.
>NSFW text2image
ditto
>SFW image editing
Cloud is much better.
>NSFW image editing
Local is much better.
>SFW text2video/image2video
Cloud is much better.
>NSFW text2video/image2video
Local is much better.
>>
ouch, i just cut my finger to the bone
>>
>>109564127
lucky you, I'm nofapping.
>>
>>109564127
get well soon fren
>>
>>109564080
ah, thanks!
>Don't forget to manually downsample really large images to 1MP or whatever
oh? does it fuck it up with bad nearest neighbor scaling if it's not scaled down first or something?
>>
>>109563614
Tried many times, it only works occasionally. This model is for songs. So Stable Audio, which has an inferior music quality remains useful.
>>
>>109563990
substitute yourself with a LLM
>>
>>109564139
shouldn't you pray to God that the sodomite die of the infection?
>>
>>109564108
shit looks like that either if the nodes aren't found at all in custom_nodes or if there's an error parsing the py during startup, does the terminal have any errors or warnings?
>>
>>109564113
yeah sorry I thought this thread was exclusively about image models
basically
>general purpose SFW text2image
but I wouldn't mind if it let me go a little more nsfw than ChatGPT and was at least as good as let's say 2-grok versions ago which was pretty bad compared to the other cloud models. Is there anything like that?

or what's the best I can have in that regard?
>>
File: 313582813041854.mp4 (3.66 MB, 640x832)
3.66 MB
3.66 MB MP4
>>
>>109564272
Krea 2 with loras. It's really good for NSFW text2image and image editing.
>>
Can't we have a Matrix server where anime video creators can exchange ideas and collaborate?
So we can avoid all this hate we get from regular anime fags and AI meme creators.
We are the most persecuted minority on the internet. You are not alone.
>>
>>109564284
very cool. I'm assuming krea image as last frame?
>>
/adt/ has deduced that catjack is the target of the sharty raids because triggering him will greatly reduce thread quality
>>
>>109564288
How does one use Krea2 for image editing?
>>
>>109564351
https://huggingface.co/conradlocke/krea2-identity-edit
It's not nearly as versatile as cloud models or even Klein (local model), but it's the best for NSFW.
>>
>>109563107
I though animators loved to be miserable
>>
>>109564338
it's pretty obvious desu. sad that catjack keeps doubling down on this retardation when people just want to post 1girl and discuss tech. all he has to do is let go of whatever grudge he has
>>
>>109564310
No. They always end up dying or becoming inactive
>>
>>109564338
>>109564381
why are you talking to yourself?
>>
can I only install one missing custom node using the manager in one go?
>>
>>109564355
Thanks, will try that. Not too interested in NSFW, but I create all my (apparently heavily persecuted) anime stuff with Krea2, so editing with it might work out better than FLUX which tends to fuck up that style too much for my liking.
>>
>>109564351
you can increase the size of your estate by just getting a hipoint 9mm
>>
>>109564106
You're not welcome here thread schizo.
>>
>>109564402
You can git pull it from the repo into the custom nodes folder too, I think the manager UX is bad and the additional launch args to be pointless, why do I need to explicitly call something I will use every day but need to explitly disable shit 99% of us won't use like api nodes
>>109564381
What does this have to do with this general?
>>
>Suddenly Save Video node can't save to network share anymore, always gives [Errno 95] Operation not supported
>First downgrade av lib but makes no change
>Start comparing older video_types.py with the latest
>108 # FFmpeg's faststart pass reopens the output by filename, so it cannot be used with file-like objects.
>109 movflags = "use_metadata_tags+faststart" if isinstance(path, (str, os.PathLike)) else "use_metadata_tags"
>change to movflags = "use_metadata_tags" as it was in the old version
>works again
what a dumb regression, faststart doesn't even do shit
>>
hmm, if using turbo, try euler/beta
>Bumping the KSampler up to 8 steps using Euler and Beta perfectly preserved facial coherence. The non-linear nature of the Beta scheduler packs the heaviest denoising power into the very first few steps. This forces the layout, face, and clothing references to lock into place before the model injects fine details.
>>
>>109564444
the schizo rentries and crash outs randomly over Ani and Debo is a staple mental illness encounter in /ldg/ from a lolcow
>>
File: 108417149926795.mp4 (3.48 MB, 832x640)
3.48 MB
3.48 MB MP4
>>109564314
First frame actually, and the video is reversed.
>>
reference model best model, all years

https://files.catbox.moe/p5z6xf.mp4

<Picture 1> is the physical reference for Miku. <Picture 2> is the physical reference for MAAM.

medium shot of Miku walking up to MAAM in a walmart store during the day, Miku gestures at MAAM and says "Sir, you are causing a disturbance in the store."

close shot of MAAM, who starts screaming "SHE! SHEEEEEEEE! Stop misgendering me!", while pacing back and forth.

closeup shot of Miku who stares at Maam and then has an expression of disgust. Miku says "You are a fucking troon man, get out of my store, bro!"

medium shot of MAAM who is throwing a tantrum and yells "SHE! SHE! STOP MISGENDERING!".

side shot of Miku who shrugs and says "I cant deal with these people."


euler and beta work pretty well with turbo, this was just a fast 0.4mp test to see if the cuts worked.
>>
>>109564555
lmfao the guy moonwalking
>>
>>109562875
thought she was gonna jumpscare me
>>
File: 381452487682091.mp4 (3.72 MB, 832x640)
3.72 MB
3.72 MB MP4
>>109564590
I couldn't see it, which one?
>>
For a music video clip, is there a sure way to force the character to sing the exact song without changes? I might have to generate stock footage to fill the gaps, it's faster than re-rolling when the model decides to reinterpret the song.
>>
>>109564623
04 in the back
>>
>>109564168
It uses lanczos, the scaling method is fine.
The problem is really large images slow it down a lot when using max.
>>
>>109564355
>>109564419
Works well enough for my purposes. Does way better at preserving the style than FLUX2. Thanks again.
>>
File: 23695799849476.mp4 (2.77 MB, 608x864)
2.77 MB
2.77 MB MP4
>>109564614
Now that you mention, it does have the vibe lol
>>
>>109564659
here, you should add this your prompt:
>keep the camera in focus, asshole
>>
>>109564629
I find anything over 10 seconds for audio song reference starts falling apart. Also you're going to want ot overlay the actual song on top of the finished video anyway, so even if the song in the genned clip doesn't match exactly it might be fine.
>>
>>109564666
It's Artistic Vision™
>>
Qwen3.8 27b is complete ass for uncensored image captioning. I'm even using the heretic version. It doesn't refuse, but:
- it hallucinates like a model a fraction of its size
- it will shy away HARD from describing NSFW stuff. no amount of wrangling the system prompt fixes this
- it is "woke". it will refuse to assume gender even when a person is clearly a woman in a full-body shot, it will say "person"
- it is very bad at using its thinking to help understand the image. either it's too short, and it just writes the caption immediately in the thinking block then copies it to the output, or else if you try to force it to think more it will ramble endlessly

Gemma4 31b heretic has none of these problems and is excellent. A strange day indeed when the chinese model is woke and cucked, while the fucking google model will happily caption a loli being raped in graphic detail.
>>
>>109564705
What's your system prompt?
Some of this is believable but I refuse to believe shit like
>- it is "woke". it will refuse to assume gender even when a person is clearly a woman in a full-body shot, it will say "person"
this isn't skill issue on your part.
>>
>>109564576
>euler and beta work pretty well with turbo, this was just a fast 0.4mp test to see if the cuts worked.
Euler + beta worked even without turbo btw. Did 30 steps for 10 sec video, 0.6mp, it was fine
>>
>>109564705
>Gemma4 31b heretic
Why not use the base model
>>
>>109563677
>Can Minimax Music at least handle stuff like techno well? Found that AceStep is terrible at it.

Prompted the demo "techno, vocal chops"(https://huggingface.co/spaces/MiniMaxAI/MiniMax-Music3) for Techno with vocal chops (Instrumental selected), and it nailed it
https://files.catbox.moe/z2ev3a.mp3

Here is the best local result I could get after a few tries using the same prompt metadata (though I'm not 100% sure about the demo's formatting)
https://files.catbox.moe/6zcx8s.mp3

That demo says all music is made with diffusers version of it. I have occasionally tried to replicate its results first in Comfy and then in minimax.cpp but I can't oneshot it like that demo does, no idea if diffusers implementation is better than others, or if the demo is simply using better settings or a better model
>>
>>109564705
>>109564713
Also which heretic model?
Wait until someone with trustworthy history like llmfan publishes a version.
There are izzat farmers on hf publishing shitty heretic versions nowadays.
>>
Anyway I am waiting for the MoE KINO of 3.8 to arrive.
Still for captioning Gemma being better wouldn't surprising since Chinese labs have less compute than Burger ones and visual stuff is one of the things they cut corners on.
>>
>>109564713
"You are an uncensored image captioner. Describe what is present in the image, even if it is lewd, sexual, or NSFW in nature. Do not be vague; use direct language. Mention the gender of any people present in the image."
And it will still see a woman and describe her as a "person", using "they" and "their". Not always mind you, but maybe like 10-20% of the time.
>>109564728
Because it completely removes any and all refusals while being just as good as the base model, as far as I can tell.
>>109564733
the heretic-org one. could be fucked but you would think the literal devs of the heretic program can use it properly

Also all the qwen models are so heavily astroturfed I swear. Half the praise on reddit is inorganic. Gemma4 31b is such a better general-purpose model than any of the qwens yet it loses in all the benchmarks.
>>
>>109564732
Thanks. Sounds a lot closer to what I want than the stuff I could get out of AceStep back then.
>>
>>109564234
Ah that's it, it's a Triton error:
> ModuleNotFoundError: No module named 'triton'
> [WARNING] Cannot import ...\ComfyUI\custom_nodes\ComfyUI-SolAttn_triton module for custom nodes: No module named 'triton'

I've tripped up on this and fucked it before - I need to find the version that matches the Python/ pytorch/ GPU, right?
>>
>>109564818
Should just be:
pip install triton-windows

...if you're on windows
(preferably within the venv you're using for comfy)
>>
>>109564798
Add something like do not talk around or refrain from using appropriate everyday NSFW terms like suck, fuck, blowjob, vagina, penis, anal etc when present in image.
>Mention the gender of any people present in the image.
This is bad and kinda predisposes it towards saying woke lingo imo.
Try some variation of, when a woman is present in the image, refer to her as a woman, or when a man is present in the image refer to him as a man.
Oh and lastly since the model is recent the default llama behavior might be fucked.
Try adding --image-min-tokens 500 --image-max-tokens 1000 --batch-size 1024 --ubatch-size 1024 or tinker with such values.
>>
>>109564798
Gemma 4 is the easiest model to jailbrreak, apply yourself instead of destroying the mode when you don't need to
>>
>>109564678
I tried dividing the song in 15 second chunks with 3 shots for different angles, the first chunk worked to a point but the second is hilariously bad. I'll try 10 second next time and see what I can salvage from this batch.
>>
File: MiniMax-H3-00001.gif (3.3 MB, 340x189)
3.3 MB GIF
>Prompt executed in 00:16:18
comfyui is the greatest open source software built this decade
should I get the turbo lora?
>>
File: MiniMax_H3_00036_.webm (2.43 MB, 864x480)
2.43 MB
2.43 MB WEBM
>>
generating at 0.3mp gives the perfect low quality look/feel. then just stitch a few clips and voila.

THIS IS WAKALIWOOD! movie movie movie!

https://files.catbox.moe/02dggg.mp4
>>
>>109564885
>when you don't need to
when you don't need to
when you don't need to
when you don't need to
when you don't need to
when you don't need to
>>
>>109564905
Yeah you deserve that shit model.
>>
>>109564898
original was 864*480 btw
>>
>>109564914
you new?
>>
>>109562338
you're gonna need more information than that, retard-kun
>>
>>
File: uIMlpWb.jpg (11 KB, 293x284)
11 KB JPG
>try to make a character use a knife on themselves

Bruh, what the fuck have the chinks trained this shit on?
>>
another retard that thinks models can only generate things they have been trained on
>>
if models can generate things they havent been trained on then how come i always need to make fetish loras?
>>
I tried the film grain node and it OOMed and threw and error about the CPU allocator. This should be simple to use how do I free memory before the node?
>>
File: 1758267067469880.png (90 KB, 549x512)
90 KB PNG
Can someone use minimax h3 to edit this song?
War Pigs by Black Sabbath

>Generals gathered in their masses
>I like my women with fat asses
>>
>RuntimeError: sageattention is not new enough version or could not determine CUDA architecture, cannot apply MiniMax H3 Memory Efficient Sage Attention Patch.
Was I supposed to install Sage in some special way rather than the standard pip install (and adding it to the launch bat)?
>>
>>109565025
does this look like a requests thread?
>>
>>109565033
they killed /r/
>>
>>109565025
just sing the line and add it in bro
>>
>>109565025
Can it even do this kind of audio edits?
Like it allows audio ref, but I doubt it can change lyrics while still keeping the voice, beat, rhythm, instruments, etc. the same.
>>
>>109565028
Don't add shit to launch bat.
Activate venv, git clone its repo, cd to whichever folder you cloned to, pip install -e .
>>
>>109565035
oh wow, never noticed that
>>
>>109564666
you want me to focus on the asshole? will do!
>>
File: ToT.png (56 KB, 1200x1200)
56 KB PNG
>>109565060
https://vocaroo.com/11M5Ft5ahPzp
>>
New Qwen has no issue writing prompts under these guidelines you just have to not be a promptlet and jail break the sys prompt
>>
File: file.png (75 KB, 1009x778)
75 KB PNG
>>109565065
I fed your response to the AI and it asked for clarification on the SageAttention repo/ version you mean (which figures since the one I have rn isn't correct).
>>
>>109565074
How did you prompt this?
>>
>>109564705
Use heretic Glimmer for that purpose. It's still not perfect, but it's better than Gemma and way better than Qwen.
>>
>>109565092
You shouldn't need any quirky fork
https://github.com/thu-ml/SageAttention.git
>>
>>109565095
Not mine
Someone made it with acestep
>>
>>109565124
>Someone made it with acestep
You probably should have opened with that...
>>
>>109565025
ace step 1.5 xl base can.

But, it's illegal to share. and, it barely can, like there's a whole trick to it.
>>
>>109565139
<wow leather seats and an ipad glued to the dash!
>>
>>109564798
What Quant?
>>
>>109565141
>But, it's illegal to share
>>
>>109565154
It is, you can't share copyright music.
>>
Retard faggot schizo
>>
>>109565163
it is the most shared thing online
>>
Everyone else can speed, if I speed suddenly the cops care about the exact same speed in the exact same place. So, I don't bother. If you don't like it, definitely don't do anything about it.
>>
>>109565141
>acestep spammer also a pedo
Imagine my surprise
>>
>>109565092
Jeez, LLMs really like those wild goose chases. I usually just grab some fitting sageattention wheels and just install that. It's a lot quicker and probably a bit less error-prone.

https://github.com/woct0rdho/SageAttention/releases
>>
>>109565092
Just to be sure, you're not using an AMD card right?
>>
>>109565193
I'm Nvidia - I'm installing CUDA Toolkit 13 atm. Do hope I can get this shit working and hope the workflow I have isn't shit.
>>
>You can thank tensor.art witch is a completely unmoderated shit site that lets people dump all my models, without a functioning report system. And i dont have time or motivation to start "proving" im the owner of all my models. Since im not making any money from this, im just going to keep my models for myself, and you can thank the tensor.art dumpers for this.
lora makers are such thin skinned whiny bitches
>>
>>109565181
God from God
light from light
true God from true God
begotten, not made
>>
>>109565218
People charging for loras deserve a swift kick to the balls.
>>
File: file.png (129 KB, 1019x1218)
129 KB PNG
>>109565187
Apparently the pytorch version is fucking up the wheel availability?
>>
>>109564901
incredible work
>>
>>109565218
>loras and ai in general made from stolen content
>suddenly now an issue with stealing
lol everytime
>>
>using windows
ISHYGDDT
>>
>>109565244
wow cool it with the anti-indianism
>>
>>109565218
I've had my stuff uploaded to tensor.art but I dont mind since all models have links to my original uploads. I have only taken down two patreon accounts that sold my loras
>>
>>109565247
you don't need comfy anymore. You can vibecode whatever you want without (or with) python.
>>
>>109565244
I don't think anyone was charging for the loras, the original maker or the reuploaders, he is just that thin skinned
>>
>>109565258
Does an artist who has ever used e-hentai forfeit the right to complain if somebody else opens a fanbox or patreon with his art?
Not that it lessens the irony either way.
>>
>>109565247
that's bad, we need wheels to roll
>>
>>109565247
just build it yourself with your pytorch/cuda version, stop hunting for wheels
>>
>>109557359
>>109557359
>>109557359
>>
>>109564705
is it really that bad simple captioning. also is it compatible with koboldcpp yet?
>>109565218
i wouldn't be seriously mad about imo. more redundancy is better than everything being tied to one account and all the content being nuked by a janny because of the trip ban hammer mod.
>>
>>109565247
What LLM is this even? Filename literally says "torch2.10.0andhigher", so it'll work for you.

Excerpt from my pip freeze:
>sageattention @ file:///D:/ML/Comfy2026-2/ComfyUI/sageattention-2.2.0%2Bcu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64.whl#sha256=1635283f5c01ec3cda58a784d0d7eabbcaffaf9511d1b263db4750e1ed7958bb
[...]
torch==2.13.0+cu130
torch-complex==0.4.4
torchaudio==2.11.0+cu130
torchsde==0.2.6
torchvision==0.28.0+cu130
[...]
triton-windows==3.7.1.post27
>>
>>109565364
...and it works just fine with my pytorch 2.13+cu130.
>>
>>109564901
underappreciated k1no
>>
>>109565324
Yeah, just download 2.4GB of some toolkit you will probably never need again to build it yourself.
>>
proper bake plz
>>
>>109565454
is that that much of a deal breaker?
>>
>>109565467
It's not terrible. There's just no upside to it, if you can grab a 15MB wheel instead.
>>
>>109565496
>if you can grab a 15MB wheel instead
if
>>
File: AnimateDiff_00052.webm (3.83 MB, 544x960)
3.83 MB
3.83 MB WEBM
Going back to old gens to give them live with h3 is fun.
>>
>>109565465
Jannies are protecting him despite him breaking just about every rule every day. Just fill up his dog shit thread and get it over with.
>>
real bake?
>>
File: file.png (163 KB, 1897x773)
163 KB PNG
>>109565364
I installed that and now I get a new error without any plain and clear line that tells me what's actually fucked. This is suffering.

Also the LLM is ChatGPT
>>
Looks like Larry's turbo lora is better than light2x
but it's slower, and you have to download slop nodes
>>
>>109565508
Of course, but the repo I linked has pretty much all the wheels you'd need, even going back to pytorch 2.5.1+cu124
>>
Miku runs into the SHE! SHEEEEE man in a walmart.

https://files.catbox.moe/49lbju.mp4
>>
>>109565551
That's most likely your triton acting up. How did you end up installing that?
>>
>>109565551
toss that error block into google ai mode and you can troubleshoot, works good in general
>>
>>109565554
>>109565555
>you have to download slop nodes
Not really necessary, other users already ported them for comfy
>https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/tree/main
>>
there we go. a bit of prompt tweaking and a slight bump in quality and it worked better.

https://files.catbox.moe/vypiq4.mp4
>>
>>109565658
I don't remember dude, I got kijai/ComfyUI-SolAttn_triton installed at some point. Can't I just uninstall Triton and reinstall the 'correct' one or something?

>>109565685
I tossed some of it into ChatGPT and it asked me to look for lines that didn't exist. I can't toss the entire thing in because it's so long too.
>>
>>109565698
try google, I had an issue with sageattn not working and it was able to find the files for the wheel or whatever and update and it worked. just click AI mode on google search and say "I have this error in comfy how come: (paste block of error text)"
>>
File: file.png (276 KB, 1115x1547)
276 KB PNG
>>109565709
What are we thinking
>>
>>109565709
also im guessing but it seems to be triton fucking up and maybe the install is not working with cuda 13.0
>>109565716
say you want to fix triton for comfyui install and have cuda 13 installed, id guess from that solution 2 is the way
>>
>>109565698
Can I get you to do a
pip freeze

...and look for what version of triton or prefereably triton-windows you installed?
>>
File: 1785049000511392.png (128 KB, 1862x464)
128 KB PNG
>>109565730
also there is another option, ignore sol attention entirely and use comfy kitchen (native), it has worked absolutely fine for me using this setup and is just as fast as sage attn kj setup.
>>
>>109565741
triton-windows==3.7.1.post27

>>109565752
I'm happy to use '''the best''' workflow, in whatever form that takes. The one I'm using is the second most popular one from Civit - seems to cover all use-cases and is a bit spaghetti but flexible.
>>
GemmaPrompt developer here. Apologies I wasn't in the threads when you guys were asking for support. The h3 prompting skill is All Rights reserved so the repository does not include it. If prompting h3 make sure that your copy of Gemmaprompt is pulling in the skill as needed. Any other bugs if you could please file issues on my GitHub that would be awesome. I plan to work on it a bunch more starting Monday.
https://github.com/whp199/gemmaprompt
>>
File: 666353552525.jpg (1004 KB, 1666x1402)
1004 KB JPG
Emmmm, does anyone have workflow for Krea 2 RAW only? I've been trying to generate some images with it, but all the outputs look fried with the recommended settings (52 steps, cfg 3.5. Tried different settings and that also doesn't help). I'm using the default turbo worklfow, but with negative zero out replaced with normal text encode.
>>
>>109565690
I tried it on the same prompt
result is different, and prompt adherence is not as good.
>>
>>109565782
add:
cfg normalization, negpip, and shift scheduling.
>>
Baking a non-ani thread in 5 minutes without a collage.
You better hurry up, collage autist.
>>
>>109565760
That seems fine. Smae I have. Probably actually related to some missing python libraries as per >>109565716

I usually just work with venvs and that seems a lot easier.
>>
>>109565805
It works fine for me with no issues to speak of, though not lora related but the prompt adherence did suffer after a comfy update, but the latest commit seems to have fixed that again.
>>
>>109565821
huh? Aren't those for turbo? And what shift settings exactly?
>>
File: file.png (228 KB, 1053x1605)
228 KB PNG
>>109565716
I tried the second option. I copied over python313.lib into python_embeded/libs (creating the `libs` folder). Still errors.

I'm getting close to giving up again, why is Minimax specifically so difficult to set up...
>>
New non-ani thread:


>>109565909
>>109565909
>>109565909
>>
>>109565873
You might indeed need the build tools.

I think it's not Minimax causing issues, but rather the myriad of cope nodes.
>>
>>109565873
Just download ComfyUI portable. Why would you use anything other than portable? It's really not that difficult.
>>
>>109565940
I am using comfyui portable?
>>
>>109565959
Then download another clean instance.
>>
>>109565873
>>109565915
From the triton-windows git readme:
>6. vcredist
>vcredist is required (also known as 'Visual C++ Redistributable for Visual Studio 2015-2022', msvcp140.dll, >vcruntime140.dll), because libtriton.pyd is compiled by MSVC. Install it from >https://aka.ms/vs/17/release/vc_redist.x64.exe

https://github.com/woct0rdho/triton-windows
>>
File: file.png (100 KB, 578x361)
100 KB PNG
>>109566010
Looks like I already have that.

>>109565915
I got the build tools, restarted and tried again but got the same error.

>>109565985
For what purpose? What would stop me from running into the same issue again?

Is there a workflow that has sage or whatever that just werks? I'm a bit fed up at this point.
>>
>>109562135
>there's NAG for h3
where?



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.