[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: Krea2_turbo_hr_fix_00043_.jpg (3.4 MB, 2368x3544)
3.4 MB JPG
You know what it is (stop making me do OP)
Previous: >>109547997

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
clussy.
>>
niguh
>>
Blessed thread of frenship
>>
>>109550423
getting pied and fucked by hot clussy
>>
I can't believe ai music will ever replace such inspiring artists as nepo baby billionaires writing about dumping their boyfriends 15 times over with like 2 unique lines a song
>>
uh oh melty
>>
>>109550446
Blame the audience.
>>
how can bollywood compete?

https://files.catbox.moe/q045me.mp4
>>
>>109550412
we already went over this last thread faggot. we didn't need a bake yet.
>>
>>109550452
>blame the audience
this is a sfw board tho
>>
>>109550453
https://files.catbox.moe/pbfzhr.mp4
>>
>>109550446
I just realized it's the same thing for the visual art world too. Wtf is wrong with this earth. :(
>>
>>109550459
But you don't post gens are contribute to meaningful discussion
>>
h3 is way too slow and soulless. i can generate 4 or 5 ltx videos at higher quality by the time h3 can finish one
>>
>>109550412
Anon this is the final warning. Dont make an extremely baity pic on the first frame. Also attempt to do a collage
>>
>>109550498
>But you don't post gens
says who?
>contribute to meaningful discussion
You're the one early baking and insisting on not making a collage.
>>
>>109550515
Says me because you won't do anything other than complain
Miku images will continue until morale improves
>>
>>109550527
I'm not the only one complaining.
>>
I wonder where the original collage anon is and if he is okay. Also the anon who inspired him to start doing collages. Also the anon who vibe coded the original collage script.

You may not believe me but those are three separate anons.
>>
>>109550511
Higher resolution =/= higher quality.
LTX can only do static video shot. Any attempt of moving the character will result of massive style drift.

Prove me wrong
>>
File: 000.jpg (233 KB, 963x707)
233 KB JPG
added the guides from h3. these work on top of the existing prompt guides. they are more like guide templates. also exposed some settings that i doubt ever need to be changed

loving MY prompt enhancer
>>
>>109550511
Well you can generate certain videos faster, it's not that much faster. LTX is pretty much dead, Wan is dead as well. If it took an hour even on a 5090 to gen a 10 sec video on H3 then it would be valid still, but it doesnt so LTX and Wan are dead.
>>
File: 3453435345.gif (21 KB, 324x316)
21 KB GIF
can you llm do THIS
>>
>>109550540
Looks nice, but I won’t be running any large LLMs on the same PC I’m genning with.
>>
>>109550544
Not him but LTX still had its uses. At least for porn. WAN is the one who 100% dead
>>
Retard question and perhaps best for a different thread but is there a difference between an actual tool created from the guide that agents call to vs simply pointing your agent at the markdown and saying "use this"?
>>
>>109550535
>Any attempt of moving the character will result of massive style drift.
if you prompt for a realistic video with a realistic character, then the video will be realistic
>>
>>109550555
Why not? It's not like you need to generate prompts and videos at the same time. Staging is important, and arguably the most important aspect of H3 is the prompt. Once you get a solid prompt, you can unload the model. Works perfectly fine for me.
>>
>>109550412
I wanna get a titjob from this incredibly hot miku
>>
https://litter.catbox.moe/c7d1kh.mp4
genning all day and dying. headache and sweating, cant get enough of this shit.
>>
>>109550559
>but LTX still had its uses. At least for porn
only with Loras and it's pretty ropey for porn at beat. Once the Loras flood in for H3 and get better it's gam over. LTX 2.5 wont even be used. I say this having use LTX 2.3 and pretty please with it as it was far faster than Wan2.2, had audio and did longer gens (20 secs) for even below the gen time of Gen 2,2 doing maybe 6 secs. People really are not going to use an older model when H3 is sitting there and isn't that much slower.
>>
>>109550567
I tried using them and it was always a hassle getting them running every time I started genning. I’d much rather just use Grok, Venice, or ChatGPT.
>>
>>109550563
Ok now try moving your realistic character to the side and see what happens
>>
there is no difference. a tool would just be more streamlined
>>
File: 1767380873519558.webm (3.48 MB, 1280x712)
3.48 MB
3.48 MB WEBM
https://files.catbox.moe/25gwci.mp4
my first working test with h3, finally had some time
>>
I shiggy diggy, 2 img reference

https://files.catbox.moe/2kzwu0.mp4
>>
File: 8934589873.mp4 (3.11 MB, 1216x672)
3.11 MB
3.11 MB MP4
>>109550527
she really is everywhere
>>
has anyone done any crazy fight animations or is it all porn and memes
>>
>>109550599
slight variation

https://files.catbox.moe/g9fjjn.mp4
>>
>>109550594
ok it worked. now what?
>>
>>109550583
For simple humping shots LTX 2.3 still good. You dont need to waste resource on H3 for it
>>
>>109550615
Webm it duh
>>
george is getting upset!

https://files.catbox.moe/13njj9.mp4
>>
>>109550626
later, i'm kind of busy
>>
With music generation, do people always just generate exclusively from a text prompt?

What about using existing libraries of samples and VSTs? Can local AI be leveraged to use Kontakt or something? Or used to write midi sequences?
>>
I know someone posted earlier how to make a 3D reference in to a 2D animation, but I can't find the fucking post now. Can someone send help?
>>
>>109550667
>Can local AI be leveraged to use Kontakt or something?
This is the one interesting thing I've seen nobody do yet. Why hasnt an AI-based instrument come out? It'd be way better than sample libraries.
>>
>>109550617
>You don't need to waste resource on H3 for it
how is "wasting resources" simply switching to another model? This is the real problem with most genners (again I include myself). The lack of creativity, genning the same shit over and over, LTX relies on Loras which do the same shitty actions. H3 offers a multitude of creative outlets with the ref model, not to mention what even the t2v or i2v can do because the model is far far superior.
>>
>>109550641
higher quality george (0.6mp vs 0.3 to test)

https://files.catbox.moe/aqrsp6.mp4
>>
>>109550679
>Why hasnt an AI-based instrument come out? It'd be way better than sample libraries.
there are physically modeled sound engines that got configured through deep learning
>>
>>109550676
like realistic video turned into anime?
>>
>>109550699
Yeah. Like specifically I was trying to feed it picrel and turn it in to a "1970's 2D-Animated Hanna-Barbera cartoon" character prompt for a vid, since it doesn't know Ron for whatever reason.
>>
File: NothingTastesBetter-2.jpg (3.06 MB, 2496x1920)
3.06 MB JPG
https://files.catbox.moe/u0kn5k.png
>>
Does MP affect audio quality?
>>
>>109550684
is that floydsbane?
>>
>>109550725
no
>>
>>109550683
Im horny and i want to make my anime girl humping a guy while handjobbing a guy on the right.

LTX 2.3 is great for quick shit like that
>>
File: 02368-1644568069.png (2.58 MB, 1280x1600)
2.58 MB PNG
Genning Na'vi girls has been my thing for the past couple of days.
This thing is magical.
Last time I tried video generation was with LTX 2.3 and I wasn't all that impressed by it.

https://files.catbox.moe/bth1zx.mp4
>>
>>109550725
No, it's more of a sampler/scheduler thing, probably why the turbo gens always sound worse than a full 20-25-50 steps gen.
>>
>>109550412
I love you, anon.
>>
>>109550707
You can be very specific in your H3 prompt and it will probably work, but it might be easier to just edit your reference image with something like Flux Klein and turn it into a cartoon style.
>>
I hope minimax releases a ref2music and a ref2image model, on top of their fabled spatial upscale model.
So far everything they've released blew my expectations.
>>
is there any model better for editing pictures other than mageflow or qwenimageedit? they're both pretty good but they usually fail at retaining artist style
>>
File: amd bros.png (132 KB, 395x255)
132 KB PNG
>RTX6000 now more than a yuopoorean makes in a year
>MAC ultra M3 like $30k for 512GB of ram
JUST
>>
>>109550785
I'm hating myself from not getting it when it just got released, now I just can't justify the cost.
>>
>>109550785
>europoors live on less than 16k annually
holy fuck i didnt know things were that dire
>>
>>109550781
Klein edit, which is also far from perfect.
And that's it, we don't have best in class for that stuff locally yet.
>>
>>109550747
BRAAAAAAAAAAAAP
>>
>>109550803
i'm european and after taxes make double that per year
>>
File: ms_00013_.png (1.51 MB, 1536x1536)
1.51 MB PNG
Alright, Sexy Jam 1 starts now.
Generate your sexy video using Minimax H3 (or LTX if you're a hipster), upload to catbox/moepantsu, then post the link here: https://forms.gle/TmxQWnNH1uS1M3tY7
Can be any type of video so long as it fits the sexy theme.

Rather than one big fat collage containing every video, the final result will instead me a montage (each video playing in sequence), uploaded to youtube (and/or catbox).
You have 48 hours. GO.
>>
>>109550789
I have two the macs. I can't believe they're worth so much more.
>>
>>109550821
The 512GB versions? How much did you pay for them?
I don't even think they sell that version.
>>
>mom found the incest deepfake vids i accidentally left on the family computer
>>
Can someone please just answer whether or not I should be using the "Mem Efficient Sage Attention" patch? I keep hearing conflicting information about this. Who is it for? If it trims vram usage, doesn't that slow down the gen time? Why would I reduce vram usage?
>>
my god. you can make tiktok slop even. it knows literally every piece of garbage.

im dying. (crying emoji)

https://files.catbox.moe/yidj3x.mp4
>>
>>109550814
Based
>48 hours
Yessir
>>
File: MiniMax_H3__00073.mp4 (3.21 MB, 800x800)
3.21 MB
3.21 MB MP4
>>
>>109550836
<Picture 2> is the physical reference for George with the same clothing.

[Scene Start]
00:00.000 - George is standing in front of a brick wall.

00:02.000 - George drops a pencil on the floor. [Sound cue: Massive, ear-shattering, bass-boosted VINE BOOM sound effect]. The camera aggressively zooms directly into the creator's eyes, freeze-framing for a millisecond.

00:04.000 - Camera cuts back out. George blinks. [Sound cue: Massive, ear-shattering, bass-boosted VINE BOOM sound effect].

00:06.000 - George opens their mouth to speak. George says "I am cryin bruh" [Sound cue: High-pitched nasal TikTok Trickster AI voice filter narration].

00:08.000 - George stares blankly. [Sound cue: Sharp Windows XP Error chime]. [Visual overlay: 3 yellow digital crying meme emojis with large streams of blue cartoon tears shooting out of its eyes appears floating over the center of the screen.]
>>
>>109550834
sage, easycache and cache do both
>>
>>109550833
catastrophe scenario
>>
The problem with stating overpriced shit like Blackwell 6000's is they really are not designed for a model that even runs on 8gb cards from 2018. Not that the model isn't really suited to higher spec cards but the point stands, all that is gained is higher resolution and maybe longer gens over a 8gn version, a 16gb card will do 20 sec gens easily. It's like gayming, the higher cards won't make the game actually better apart from the graphics and frame rate, you can't improve the gayme or model's output content by spending thousands on a card.
>>
>>109550833
Hopefully you saved it onto an external so you can remember your family when they kick you out and disown you.
>>
File: 1762784875542975.jpg (1.01 MB, 4078x3682)
1.01 MB JPG
refer to an anon's image, you can see it doesn't really matter (for speed and vram) as long as you don't stack them
>>
>>109550836
>even the AI has brainrot
Now imagine the Terminator AI apocalypse with T100s playing Vine boom as they shoot humans.
>>
>>109550854
Yep, and next generation is 2 years away, so we're stuck with this.
>>
>>109550836
we have done it, peak brainrot.

https://files.catbox.moe/9ued75.mp4
>>
I can't stand floyd gens
>>
>>109550859
this is how we stop the tiktok shit, mass AI slop to overwhelm the algorithm.
>>
>>109550859
i misread shoot humans as short humerus
i fucking hate this place
>>
>>109550857
did you invert the colors of anons test results kek idk why i find that funny
>>
>>109550392
>>109550383

Alright, so for Minimax Music maybe Comfy fucked something up in his implementation. I was only able to generate a few songs up to 60 seconds using "Simple", but take a look at results

"A 1980s rock song"
https://files.catbox.moe/w5tatd.wav

"Melodic house, female vocals"
https://files.catbox.moe/lo5vfd.wav

"Japanese shoegaze, female vocals"
https://files.catbox.moe/hnscm9.wav

These are way better than results I was getting on Comfy, and basically blow what ACEStep XL can do out of the water for those prompts. I'll have to try the diffusers implementation then.
>>
>>109550901
Forgot link, these were generated with official HF Demo
https://huggingface.co/spaces/MiniMaxAI/MiniMax-Music3
>>
>>109550901
Seeing how the fucking models wouldn't work on blackwell for the first couple of hours I'm pretty sure that's the case
>>
>>109550901
The japanese shoegaze one is really good. If the riff melody was extended a bit, it'd be a total earworm.
>>
File: ichud.jpg (1.61 MB, 4096x3072)
1.61 MB JPG
>>109550827
You can't buy them any more, I think the biggest one they sell is 128GB of ram. They were like 12k each?
>>
>>109550901
>"A 1980s rock song"
failure. always the best benchmark to hear if a model is slopped or not
>>
what should i generate?
>>
how are llms so good at generating brainrot slop prompts, how is minimax good at generating it. Oh right, China owns Tiktok, my bad.

https://files.catbox.moe/mzajnh.mp4
>>
File: collage-spaced-40cf61.mp4 (3.8 MB, 1440x1436)
3.8 MB
3.8 MB MP4
>>109550412
one more week until I get my 5090
>>
>>109550934
kinos
>>
>>109550936
>zesty Hitler in the collage
it's over
>>
File: 7b1.png (524 KB, 600x568)
524 KB PNG
>>109550864
>>109550935
is this what Tiktok is actually like?
>>
>>109550936
that Hitler gen
>>
>>109550936
what the heck is going on with the saturation and colors here
>>
>>109550955
sadly yes, pretend you did enough hard drugs to destroy your brain and then you headbutted a soundboard.
>>
>>109550814
desu im nervous about my 1girl not looking sexy enough

what if anons 1girl is sexier?
>>
>>109550944
right away sir
>>
File: 1763127782092242.mp4 (1.03 MB, 544x800)
1.03 MB
1.03 MB MP4
>>
this is how we save tiktok. we flood all the people off with infinite floyd brainrot.

https://files.catbox.moe/s79nqv.mp4
>>
its up, get it before it's deleted
https://civarchive.com/models/2856467?modelVersionId=3226233
>>
>>109550927
That sounds like it to me though
>>
>>109550890
not me I just was too lazy to find the original post
>>
>>109550976
if you're lucky you will be the only one to make a submission
>>
>>109550926
It's crazy it's an investment now.
What are you even running on them, giant LLMs?
>>
>>109550936
How much did you pay for it?
>>
cyka blyat?
https://litter.catbox.moe/uaffdhbbs2s7zui0.png
>>
>>109550814
Does it have to be SFW?
>>
>>109551004
4bit deepseek but right now they're both doing mimax h3. Speeds are okay.
>>
>>109550993
Why would it be deleted?
>>
>>109551014
Because it's garbage.
>>
File: 438229116895470.mp4 (3.8 MB, 544x960)
3.8 MB
3.8 MB MP4
>>
>>109551012
I guess you can run multiple ones on parallel on the same box but not sure compute can keep up on that.
>>
>>109550814
I'm sadly extremely busy in the next 48h hours so will not be able to participate but excited to see everyones gens
>>
anyways. the new ref2v lora is working well, im using 8 steps. the 1.0 fl2v is already very good at 8 steps. there will be a 1.0 for the reference one but despite that it's already working really well (for me) over the spectrum setup. just swap that with the lora after model loader.
>>
>>109551019
needs a bit more expressive face, otherwise it's cute
>>
>>109551011
No.
If I can't upload it to youtube then the montage will be on catbox.
>>
>>109550552
no but 5 minutes in after effects can
see, I'm skilled so AI doesn't threaten me, its adds to what I already can do.
>>
>>109550995
it sounds like rock with a modern style sound
>>
>>109551012
>Speeds are okay
What are your speeds?
>>
>>109550814
my gen would get it taken down, no thanks.
>>
>>109551022
You're not too busy to post in this thread but you're too busy to spend 100-300 seconds generating a video?
>>
>>109551041
Anon you're allowed to submit NSFW.
>>
File: output.mp4 (778 KB, 1184x768)
778 KB
778 KB MP4
That was a pain in the ass, wish the reference model would just read my mind.
>>
>>109550814
>gen only has to be sexy
>doesnt have to be a girl
I am in.
>>
>>109550926
speeds are probably doggy but great LLM machines
>>
>>109551054
Correct.
>>
>>109551053
looks good but it's a headswap. would be interesting to see if the body could be changed while retaining the action.
>>
>>109550908
"Japanese 1980s city pop, female"
https://files.catbox.moe/lnme2r.wav
>>
>>109551069
Any mention of breasts seems to create nips. Idk what else you would expect to be changed.
>>
>>109551039
.4MP 5 seconds is like 66 second/IT for reference. I'll start a few 2MP for 20 seconds tonight and report the total time, but I'm thinking hours.

still, 2 more boxes to gen with
>>
>>109551079
try chest, or body type
>>
>>109551053
Why am I suddenly remembering Golden Boy?
>>
>>109551081
>66 second/IT for reference.
that sounds like hell no cap
>>
>>109550901
Pretty impressive honestly
>>
File: 1783566168183450.mp4 (742 KB, 512x800)
742 KB
742 KB MP4
>>
>>109551072
Please make a ticket or alert comfy, I'm so fucking tired of these rush half assed releases
>>
>>109551110
kino
>>
>>109551100
anon I have a 5090 and RTX 6000, it's free gens
>>
File: file.png (165 KB, 253x398)
165 KB PNG
How we doing genners?
>>
>>109550901
ComfyOrg fucked up their implementation of ACEStep XL too. At least initially idk about now.
>>
anyone have a suggestion for the LLM model to use for prompting H3? I used Opus 5 for this: https://odysee.com/@MLP:2/Posey-Mogs-Misty:b

but it's so wordy I figure there's probably a better model for this sort of thing
>>
>>109551134
What website is this saar?
>>
jesus christ, why is the venti faggot here
what's next, ryanposting?
>>
>>109551136
i use qwen 3.6-27B abliterated locally and get very good results with it
>>
>>109551148
>>109550081
>>
>>109551136
I use Gemma4 26B, some of the local models do pretty good.
>>
can I get good gens on a 12gb gpu or do I have to load up on cope nodes just to get smeared shit?
>>
you asked that question already
>>
>>109551069
I got the meiya one working last thread
>>109548056
>>
>>109551157
How much RAM do you have? Comfy is pretty good about offloading to RAM now. It slows it down but really not as bad as it used to be.
>>
long dick general
>>
File: ComfyUI_00151_.png (1.88 MB, 1024x1024)
1.88 MB PNG
>>109551149
>>109551151
I avoided local so I wouldn't have to load/unload models in my 24gb of VRAM :( I guess I could just openrouter something. Actually I haven't tried Luna or zai models yet
Gemma I've used for short form JP writing and it does seem like it might be a good model for more creative concise output tbqh
>>
>>109551053
god damnit
https://files.catbox.moe/l0cot4.mp4
this is what I meant
>>
/vp/ has a video request for you, /ldg/:

>>>/vp/59505200
>>
File: 374541427336853.mp4 (3.75 MB, 544x960)
3.75 MB
3.75 MB MP4
>>109551027
Yeah, I didn't prompt for any expression so it's a bit bland.
>>
>>109551114
Seems like all their focus is on nodes 2.0, which is why everything else is broken to oblivion like subgraphs and such if you aren't on 2.0
It's also one of the reasons why they haven't bothered fixing the reported bugs for the old legacy nodes, since they it's a consequence for changing the code to fit 2.0
>tl;dr
They won't give a fuck unless the bug is nodes 2.0 related.
>>
>>109551177
I used Gemma before too, but in my experience while it was better for the actual prose itself, it wasn't as robust with the template as I'd like it to be. they both have their strengths.
>>
>>109551135
It seems that way, I do not have confirmation yet. I'm using "Simple" prompts on the HF frontend, that seems to run it through their agents and do some magic. I will attempt to recreate the results I got with the Agent Skills and lyrics on HF diffusers implementation, then compare it to results I'm getting locally. So far either their prompt rewriter simply is superior, or Comfy's implementation is duped.
>>
File: file.png (146 KB, 262x393)
146 KB PNG
Sorry, I forgot there are certified faggots on this general.

How we doing genners?
>>
>>109551191
That's a man.
>>
>>109551186
but 2.0 is bad and breaks a ton of community nodes, why dont they just do one thing at a time
>>
>>109551171
hell yeah brother
https://litter.catbox.moe/o06ryfx2gcv3p3kz.png
>>
>>109551186
customized and personalized interfaces are becoming easier and easier to create. they are so fucking retarded.
>>
Another /vp/ request:

>>>/vp/59505222
>Could you gen something with Lorelei undressing or bathing, i think she needs more love around here.
>>
>>109551195
you dont know what genitals they were generated with
>>
File: Krea2_turbo_00875_.jpg (2.24 MB, 1672x2512)
2.24 MB JPG
>>
>>109551196
>bad decisions and lack of management skills
>>109551205
>>109551205
>they are so fucking retarded
You said it, except for kj
>>
>>109551136
>2026
>mlp
>>
can someone make hunter schafer a woman so I can jerk off to him without it being gay?
>>
>>109551203
Are they okay?
>>
>>109551119
never mind carry on
>>
So is anyone here actually going to do this sexy jam thing?
>>
>>109551203
curious, why arent you on some ai porn subreddit. why come here. i just dont get it
>>
>>109551232
Sure.
>>
File: Flux2-Klein_00046_.png (1.39 MB, 832x1216)
1.39 MB PNG
>>
File: ComfyUI_00002_.mp4 (1.66 MB, 672x960)
1.66 MB
1.66 MB MP4
Amazing model!
>>
>>109551019
needs more pantsu from the back
>>
holy mother of fuck the ref2v turbo lora is SO MUCH FUCKING SLOWER than fl2v turbo
>>
reminder to use the uncensored text encoder

https://huggingface.co/sakamakismile/Qwen3-VL-32B-Heretic-MiniMax-H3-NVFP4
>>
>>109551263
Retard.
>>
>>109551263
Genius.
>>
>>109551263
model not needed. stop vibeslopping
>>
>>109551263
Any comparisons?
>>
>>109551263
there is no need to use a heretic text encoder, but you should be using a quantized vae. saves a lot of time on vae decode.
>>
>>109551278
VAE decode is the fastest part of the process why would I care about saving a few seconds?
>>
>>109551136
opus is the worst model as it loves to moralfag, use gemma31B locally
>>
>>109551284
he could be a yuropoor that pays out the ass for power.
>>
>>109551191
you're weird
>>
>>109551263
The heretic creator himself said this wont do shit for h3
>>
>>109551284
>a few seconds?
kek
>>
>>109551242
her face got 10 years older by the end of the video
>>
>>109551284
it feels like my decode takes forever
>>
is there no 8step ref2va turbo lora? all the ref loras seem to be 4step
>>
taffy vaginas
>>
>>109551308
i think the 8 step one comes later? i dunno
>>
>>109551308
Are you stupid? The LoRAs work on both fl2v and ref2v.

https://huggingface.co/Kijai/MiniMax-H3_comfy/blob/main/loras/minimax_h3_fl2v_lightx2v_turbo_8step_v1.0_resized_avg_rank_24_bf16.safetensors
>>
>minimax cant make ar-
2 references. one for the setting, one for the retard.

https://files.catbox.moe/4gqyxj.mp4
>>
Is the widowmaker anon who worked on that exquisite image replacement prompt here? How do you tell the model what the "replacement object is"?
>>
So what uncensored local model should I use with this?

https://github.com/whp199/GemmaPrompt

Will an MoE model work well?
>>
>>109551255
try this, im using 8 steps, seems fine to me

https://huggingface.co/Kijai/MiniMax-H3_comfy/blob/main/loras/minimax_h3_ref2v_lightx2v_turbo_4step_v0.1_resized_avg_rank_20_bf16.safetensors

note if you use video reference inputs it's slower. image or sound, normal.
>>
File: 1777187272700010.png (265 KB, 357x422)
265 KB PNG
>>109551319
>the lora works on the model it wasnt trained on
>>
quick gimme your BEST Minimax H3 workflows, there's too much trash on civitai
>>
>max_position_embeddings = 262144
Remember when you used to have separation for the text every few hundreds tokens because it had to be clipped?
>>
>>109551319
>works on both models
>yet the trained two different ones
how does this make sense to you
>>
>>109551292
>he thinks he can shame me out of this general
>>
>>109551319
What is the difference between this

https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras

and this

https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

Why is the second one way bigger?
>>
>>109551319
show us a gen anon, show us what you can do
>>
>>109551336
honestly the default comfy workflows may still be the best with just slight modifications
>>
>>109551338
i member
>>
File: rh_00009_.png (1.88 MB, 1536x1536)
1.88 MB PNG
so what exactly makes a sexy video?
>>
File: 1785217666440908.mp4 (905 KB, 544x800)
905 KB
905 KB MP4
third seed's the charm
>>
>>109551350
https://www.youtube.com/watch?v=_VaZGg9dXUE
>>
File: Krea2_turbo_00888_.jpg (1.93 MB, 1672x2512)
1.93 MB JPG
>>
>>109551336
I like the one with the KJNode for Model Preview Override because you can see a low res preview and stop it if it's fucking up the current gen.
>>
>>109551350
>so what exactly makes a sexy video?
for me, its a naked girl
>>
female topless

I'm very creative prompter.
>>
>>109550814
this is a bait. I'll waste my time making a video then will get excluded because I'm not in the discord clique. ldg boys club cunts
>>
>>109551380
i just prompt sexy
>>
>>109551381
no you'll get excluded if your video doesn't have a sexy undertone
>>
>>109551374
for me it's a fwc (fat white cock) and sexy girls licking it but it can also be a black cock if the girls are jewish
>>
File: MiniMax_H3__00076.mp4 (1.47 MB, 1056x608)
1.47 MB
1.47 MB MP4
>>
>>109551381
>I'll waste my time making a video
but you already waste your time making sexy videos, don't you?
>>
>>109551350
loli panty shots
what else?
>>
Are we at a point where I can create references of my favorite loli doujins and anime these short stories?
>>
>>109551374
kino
>>
File: MiniMax_H3_00019_.webm (2.16 MB, 496x896)
2.16 MB
2.16 MB WEBM
I tried this meme 4 times but cant get it to work. here's the best failed attempt of the lot
>>
>>
>>109551447
Always was
>>
The Spectrum node got a lot of commits today. Anyone test it yet with today's update?
>>
>>109550958
I'm testing the new collage maker I vibe coded so the colours might be messed up. I think there's an export-parity bug with the source files
>>
>>109551412
Learning to become proficient at a new technology is never a waste of time
>>
>>109551465
It was funny and cute until the CURRENT THING part
>>
>>109551493
he's always loved raw milk though
>>
File: MiniMax_H3__00077.mp4 (1.34 MB, 1056x608)
1.34 MB
1.34 MB MP4
>>
>>109551489
is there a single portable skill we can gain from telling a toaster to make better big titty girls
>>
I know everyone is on videos now but I'm still doing images. What is like the general workflow?
I haven't touched anything since A111 days and just been using a basic workflow in ComfyUi. I used to just prompt multiple small images, adjust the prompt, then find one I like maybe rerun the seed with higher steps to see if it changes anything and then upscale it. Is that still the go?
Also does Comfyui have X/Y graph for finding settings or either/or statements in prompts? ie "A [red|blue} dog" and then it would choose either "A red dog" or "A blue dog", so then with multiple either statements you could get variety each time
>>
>>109551511
>>109551408
>>109551053
why do you have all these videos of my wife i have never seen before?
>>
>>109551447
maybe in another year
we're getting close though
>>
>>109551520
the only thing in your post than an llm wont answer for you is that the meta for image gen is cfg normalization, negpip, and shift scheduling. itll tell you what those are thoughever.
>>
>>109551512
inference and training and python seem like good skills, it got me from fucking around with boring ass lists and dicts and gay nerd shit like that to actually having a renewed interest in some coding projects, and now I have something to make coding projects for
its like crack
the fact I created my dream girl and she'll do anything I say and loves me till my computer dies is very addicting
>>109551520
krea 2 one gen to an acceptable resolution, then a denoised upscale pass with another sampler at 0.25-0.3 and a latent upscaler
picrel is krea2 using that process
https://files.catbox.moe/rx44ev.png
>>
File: StyleTest.webm (697 KB, 1832x1056)
697 KB
697 KB WEBM
T2VA.
I hope someone makes a list of all the thing this is trained on.
>>
why did jannies wordfilter voldy in the first place?
>>
reminder, smaller fl2v/ref loras are up (use 8 steps)

https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras
>>
does the guy who made this claudeslop not realize anima DOES support natural language? he seems to be confusing it with illustrious
>>
File: z-image_09931_.png (3.42 MB, 1984x1536)
3.42 MB PNG
>>109551520
>I know everyone is on videos now
imggen is still cool and you need something to feed into i2v
>>
>>109551555
He had a captcha bypass in his repo.
>>
>>109551574
krea 2, klein edit 9b, and illustrious (still great for anime) are all tools that you can use with i2v or reference. image gen will always be important.
>>
>>109551585
oooooohhh right, i forgot he was the original developer of that before druno forked it
>>
>>109551521
she gets around if you know what i mean
>>
File: output.webm (3.88 MB, 1184x768)
3.88 MB
3.88 MB WEBM
Not really what I wanted his face to look like but I'm not sitting around for another gen.
>>
File: anima_00059_.jpg (454 KB, 1424x1784)
454 KB JPG
good morning sirs. I spent a while today trying to recreate indian miku from the other thread. harder than it looks. anyway this is an anima finetune
>>
Ok, and what the fuck about if you want to use a specific image as both the first frame as a reference for <Subject 1>?
>>
>>109551607
reddit ahh UI
>>
>>109551570
>>109551607
I'm sure anon would appreciate you opening an issue about it...
>>
>>109551607
Lol not a problem in my custom vibe coded interface.
>>
>>109551607
wipe it saar
>>
>>109551607
also where the fuck are the audio references? did the author forget minimax is an audio model?
>>
>>109551465
You'll never have real udders.
>>
>>109551353
the feet are wrong. :^)

>>109551242
lxt wraithface
>>
>>109551600
>wakes up smiling
extremely unrealistic
>>
>>109551624
you're not my real mother
>>
After genning myself having countless sexual encounters with attractive women, I've noticed a profound shift in my confidence and self image.
I feel better about myself and carry myself with a genuine sense of abundance. It's almost as if my subconscious genuinely believes I'm at the top of the social hierarchy. People respond to that confidence which only reinforces it further, creating a positive feedback loop
>>
>>109551234
>reddit
reddit sucks. the stuff on civitai is better than that shit.
>>
we need ltx 3.0
>>
File: 1757629642430506.mp4 (1.22 MB, 736x544)
1.22 MB
1.22 MB MP4
>>109550814
>>
>>109551447
Technically you can make a hentai animations with better animation quality than queen bee right NOW..... But, like in all things. you need some effort.
>>
File: RH_00016_.png (1.21 MB, 1536x1024)
1.21 MB PNG
>>
File: new.jpg (29 KB, 256x256)
29 KB JPG
so this is the power of local models?
>>
Is sol-attn old and busted or are we still using it, and does it play nice with loras
>>
>>109551678
we all using ck now
>>
File: 1783520671828666.png (77 KB, 1631x294)
77 KB PNG
>>109551678
this at 8 steps (default simple/res multistep) working good for me, no more spectrum cause it doesnt work with the lora (I think)
>>
>>109551684
what does it do?
>>
File: 1409862727217.jpg (47 KB, 499x499)
47 KB JPG
>>109551690
>comfy kitchen node instead of sage
>>
>>109551696
yes, the reason being comfy works natively with comfyui, and there is no int8 compression or whatever, so it's better quality or faster
>Yes, Sage Attention (via KJ's patch or command flags) uses low-precision quantization compression, whereas Comfy Kitchen Attention does not rely on that specific attention compression.
>>
File: 7657389987348.mp4 (3.93 MB, 1216x672)
3.93 MB
3.93 MB MP4
>>
>>109551690
where does the modelattentionbackend node come from. I dont have it
>>
>>109551696
sage can be a compatibility headache on 50 series GPUs because SM120 often requires specific builds and matching CUDA/PyTorch versions. for most 50 series users, comfy kitchen is the simpler and better option
>>
>>109551725
Update u're comfyui sweaty
>>
>>109551725
natively in v0.32.0 and newer, if you have comfy updated you will have it (update comfyui in update folder)
>>
here I go, updatin again
>>
shocking.

https://files.catbox.moe/w1zc78.mp4
>>
>>109551684
Nah ck bites
But because sol-attn directs the model's attention it would override any other attention setting node, yes?
>>
>>109550412
> stop making me do OP
just remove the tranny's links, it will sound the alarm and wake him up
>>
>>109550803
Actually I make half that much
>>
Is Wan old news? Nobody is mentioning it?
>>
>>109551791
It crawled so MiniMax H3 could run.
>>
>>109551802
>>109551802
>>109551802
>>
Warning, this is a debo thread: >>109551803

Do not post in it. We're not even at the bump limit.
>>
waiting for the non-debo thread
>>
baking
>>
File: Untitled.png (123 KB, 1648x984)
123 KB PNG
lower strength spread out across multiple sex loras makes it so your anime doesn't turn realistic/blurry
>>
Trying out the minimax MysticXXX lora. This fucking retard didn't train it with guidance loss, so it removed the distillation, meaning you have use it with CFG, meaning spectrum doesn't work, and also it won't stack right with other loras, and also even combining it with the distill loras is wonky since the distills don't work properly with CFG (which as previously mentioned, the lora requires)

No base model completely ruins minimax's potential. 99% of people are too stupid to train the guidance-distilled one properly.
>>
>hmmpf why do i always have to make the thread
>NOOOOOOOOO WHAT THE HECK WHY DID YOU MAKE A THREAD?
>>
Poor debo. This moron is so desperate to get his rentry removed from the OP but he just keeps failing. Everyone can recognize him.
>>
>>109551773
ck is 100% better for me over patch sage kj. it is slightly faster or in any case, comparable with better quality (not a huge diff, but a difference)

ck and turbo lora 8 steps works nice for me.
>>
So is this comfyui kitchen cope better than KJ's sage attention or what? Cause KJ has always produced better nodes than Comfycore.
>>
turbo and ck are cope, your gens are ugly
>>
>>109551825
>power lora loader
>2026
>>
File: 1765569637844639.png (267 KB, 609x781)
267 KB PNG
>>109551723
ToT

Also on side note.
Turbo Lora is great but it killed the prompt adherance for me. I wonder if its going better in V2 or V3 did Lightx2v makes a new version in WAN ?
>>
>>109551856 >>109551856
>>
When ready
>>109551858
>>109551858
>>109551858
>>109551858
>>
>>109551865

>>109551863
>>
>>109551872
Deleted it
>>
>>109551607
Read the reference labels, dipshit. img 1 = subject 1, img 1 = first frame
>>
>one thread with a different OP paste and old collage from a week ago

>one thread with the same OP paste and a collage with gens from the previous thread

hm....
>>
>>109551901
I choose the one that came first.
>>
>>109550901
slop.
horrid utter slop.
>>109551135
no. his original code was very good.
fuck up occurred after the original code. and it is continuous.
i claim this:
>>109551879
>>
>>109550582
xDDDDDD
and the end frame



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.