[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File deleted.
REMINDER TO SUBMIT YOUR VIDEO FOR SEXY JAM:
https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewform

Previous: >>109565909

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
>at the gym and all i can think of is exploring the h3 tech i found out last night by experimenting
>>
>>109569562
What tech? The 1s-chaining?
>>
HOW TO STOP MALES MAKING GRUNTING SOUNDS AAAAAAAAAAAAAA
FUCK MINIMAX FUCK YOU
>>
>>109569567
nta, but the attention is not linear in H3, that is why it takes so fucking long for 10 second clips. But when you chain them, the 1+1+1... is so much fucking faster. People have to do this
>>
>>109569567
What's one chaining? Ref model, generate a 1 second clip, add it as ref, generate the next 1 second clip, etc? Is there a work flow?
>>
AUDIO ON MINIMAX IS SO FUCKING FUCKING BAD
FUCK MINIMAX
NO WONDER THEIR MUSIC MODEL IS SHIT
>>
>>109569586
qrd on this, im over here genning 23 second clips all at once
>>
>>109569590
Well I don't know how it works but anon mentioned it in >>109569089
might be >>109569586
would love to see an example gen with workflow
>>
but chaining 1s clips would make there be no temporal consistency
>>
>>109569613
1 second in the context (as a ref) might be enough. Or, maybe combining them with ffmpeg automatically and feeding in a rolling three second clip or something. I bet claude could cook up a node...
>>
>>109569639
i was invited in
>>
>>109569586
Can someone make an example gen of the output when doing this?
>>
>>109569595
Use the full model you vramlet
>>
>>109569504
>https://rentry.org/debo
>https://rentry.org/animanon
Can anyone explain how does this shit help the threads and how is it relevant at all? Seems like some troll bullshit to me
>>
>>109569650
based OP inviting the good posters
>>
>>
>>109569504
>/v/irgins general
>>
File deleted.
>>
File deleted.
>>
>>109569693
hmmm
Yep, safe. I checked carefully.
>>
File deleted.
>>109569702
you sure?
>>
File: kekekekkeeeeeeeeek.png (1.17 MB, 864x1184)
1.17 MB PNG
nah its over JSID already bruh wth is this keeeeeeeeeeeeeek
>>
>>109569693
Catbox?
>>
File deleted.
>>109569722
spam
@EightGo is the artist
>>
File deleted.
>>
my god, what a miserable shithole.
/cancer diffusion general/
>>
Ah.
Misunderstood what was going on.
my bad fellas. your schizo is annoying
https://litter.catbox.moe/qdqzsk.webp would have been next
>>
>>109569742
save us
>>
>>109569752
Oh, you're from another general, and you came here after the schizo linked to this thread in your general.
Yeah, let me explain. Our schizo (animanon) doesn't like that there's a rentry exposing him in our OP, so he spergs out year long about it in multiple ways. This time, he makes threads without the rentry (that no one wants to use), and spams crosslinks to the thread with the rentry to make it look bad/to get it deleted.
>>
>>109569752
let me explain. will smith kino poster makes trannies seethe so much, that they spam all of /g/ whenever he makes a thread
>>
That retard trying to falseflag
>>
>>
>>109569767
What I don't get is how are those rentries related to local diffusion. Why waste space on off-topic bullshit?
>>
>>109569662
No one able to run FP32 bro
>>
>>109569807
well anime anon is a profession c++ programmer who is on the cutting edge of stable diffusion development, so that is relevant.
do you complain when people talke about sam altman in a saas generals?
>>
>>109569882
They don't put Sam Altman slander in their OPs.
>>
File: comf3.png (302 KB, 3130x1305)
302 KB PNG
I thought comfy was isolated to local and manager used python backend to fetch updates which I can control with firewall.
So jeets have been stealing my prompts?
>>
>stable diffusion development
>>
>>109569890
>corpo UI is a botnet
wow who could have thought
>>
File: comf4.jpg (14 KB, 369x234)
14 KB JPG
>he uses cumshart manager™
its time to start fresh
>>
File: file.png (10 KB, 366x207)
10 KB PNG
>>109569890
>>109569913
I use comfyui manager, the custom node tho, not the builtin one
>>
File: 46757567857665.gif (260 KB, 220x176)
260 KB GIF
>>109569890
i've got your IP address now, it is going to get much worse for you.
>>
>>109569504
You might need to update the ani rentry with his Discord messages.
(Death threats, not being able to be near yoland 'n shit like that.)
>>
>>109569879
>laughs in 5090
>>
H3 isnt as great as i thought.
Sound are TERRIBLE
Gibberish voice will happen no matter what. even when you prompt properly.
>>
>>109569945
use 20 steps no cope nodes except kitchen
>>
>>109569890
How could it be isolated ? Every custom node you install can go online for whatever purpose it wants.
>>
>>109569945
Use fp16 model
Use more steps
If you don't want any talking use dialogue:N/A
>>
File: file.png (24 KB, 321x216)
24 KB PNG
>>109569879
??? There are other models besides fp32 for audio ???
>>
>>109569945
>even when you prompt properly.
doubt you're doing that thobeit
>>
>>109569950
>>109569954
>>109569959

Already did. Tried with gangbang scene. And grunting male voice overwhelm the girl sound no matter what.

See my prompt on >>>/gif/31048020
>>
>>109569963
>grunting male voice overwhelm the girl sound no matter what
based loud porno actor
>>
>>109569963
Your shit is too complicated and says nothing about voice levels so it just makes shit up
>>
>>109569968
>>109569954
>>109569959
Hear the horror yourself https://litter.catbox.moe/c4pfhp7vlyrsuzjv.mp3
Fuck this shit
>>
File: worksinmycountry.png (6 KB, 576x649)
6 KB PNG
>>109569890
I just cloned the repo and set up a venv, and I still use manager
>>
>>109569973
>Says nothing about voice levels so it just makes shit up

How ?? I already follow the R2V and I2V guide properly dude.
>>
>>109569974
he sounds just like me for real
>>
>>109569963
try
><Subject 2> remains silent
><Subject 3> remains silent
etc.
>>
>>109569985
Thats what i did.
Still "ngggggggrrrhhhh" and "grrrrrrrrrhhhh"
I even want to cut tha audio vae to cut out the sound completely, but it seems pointless so ill just to my porn on LTX instead
>>
retarded french pedo
>>
>>
yikes is that hentai in my local diffusion general
>>
File: Krea2_turbo_00888_.jpg (1.93 MB, 1672x2512)
1.93 MB JPG
>>109569742
Then go back to your dead schizo thread if you're going to complain
Nobody wants to spam GM and share shitty music that nobody engages with. There's a reason why everyone post here and not in /sdg/
>>
unc is crashing out
>>
>>109569978
>how?
By using words, moron
>muh guide
Try to write actual sentences describing what you want
>>
>>109570062
I been doing this since 8 hours ago. I know what im doing. Minimax cannot do more than 4 subjects
>>
>>109570066
>I know what im doing.
Clearly not since you're copying the chatgpt ahh guide instead of writing descriptive, verbose text.
>>
>>109570076
You cant even write reply properly.
>>
>>109570076
God i hate zoomerspeak like you so much
>>
>>109570082
>>109570085
>not addressing the point made
>>
File: ed5.png (131 KB, 680x1112)
131 KB PNG
>God i hate zoomerspeak like you so much
>>
This reminded me
Has there been any tests on how well H3 and qwen TE responds to English vs chinese prompts
>>
>>109570093
You can prove me wrong by making I2V from a gangbang scene. Do it now
>>
>>109570099
prove that you're unable to stay on topic by generating a gangbang scene? that doesn't make sense anon
>>
File: MiniMax_H3_00077_.mp4 (2.41 MB, 928x672)
2.41 MB
2.41 MB MP4
Minimax Music hard trance banger
https://files.catbox.moe/ozrq4c.mp3
>>
>>109570106
how long does each minute of music take to generate? i want to try it out
>>
>>109570104
>>>/gif/vdg
Go
>>
>>109570109
I think you're very confused
>>
>>109570116
Then you're are wrong.
>>
>>109570119
no I don't think so
>>
>>109570104
Can't, too busy creating child porn
>>
>>109570126
what a sad lonely man,
>>
>>109570133
Meant for >>109570099
>>
File: jeeeeeeeeeeej.png (269 KB, 432x621)
269 KB PNG
>>
>>109570134
kiss me already
>>
>When the character says ayy-non instead of ah-non
>>
>>109570158
Just write it how you pronounce it
>uhnon
>>
>>109570158
>>109570166
I pronounce it a-non
>>
>>109570172
we know you do, redditor
>>
>>109570037
make me
>>
>>109570108
>how long does each minute of music take to generate

I'm on minimaxmusic.cpp. It depends, on a 3090 for 60s it's 1.8 mins per gen, and for 4ish mins for 2 mins, and so on (note longer gens take a while longer, it's not just x2). You may be able to get faster gens on Comfy since that uses a pruned INT8 model, but since I use gguf and I want highest quality possible those are the times I get.
>>
File: output_smal.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
>>109570176
>>109570181
Nobody cares go back to your containment, you have nothing to offer other than garbage and complaining. You make nothing of note and you only seethe at this thread while crying with the main schizo. You're not even worth a funny gen because it will be just you alone posting in a dead thread.


wait.....you can't run H3 so that's why you're seething
>>
>>109570205
you think GGUF Q8 beats int8 convrot?
>>
>>109570205
not bad
>>
>>109570215
Dunno if it does for Minimax, but this model is very hard to prompt and for some reason Comfy only provided the pruned int8 convrot text encoder. But the text encoder is doing most of the work on this model, so that was my rationale for not using it and sticking to .cpp right away. It'd be trivial to convert the bf16 text encoder to int8 convrot but takes a while. At first, I thought something may have been wrong with Comfy since the model needs to be prompted very specifically unlike other music models I've used like ACEStep XL, but it may be fine after all as I learned to prompt it using the agent skills and https://huggingface.co/spaces/multimodalart/minimax-music3-prompting-guide

Also the official HF demo which I can borrow prompts from. It'll be a lot easier to steer this model once they release the missing encoder (or someone distills it). The model is able to mimic musical structure of any existing song very well if you describe it accurately, so no LoRAs may be needed at all for most things, but it may take surgical precision.
>>
>>109570215
gguf is much slower from my experience
>>
>>109570302
>The model is able to mimic musical structure of any existing song

This last song I prompted was intentionally in the style of https://www.youtube.com/watch?v=z5LW07FTJbI
Claude oneshotted the prompt, and the model pretty much nailed the square waves and vibe of the song right away. I've found that once the musical elements of a song are known and described well, Minimax excels every single time at creating the music.
>>
the hype for Flux 3 is so dead lmao
>>
>>109570332
I'm still looking forward to it. Not for video, as it's pretty unlikely it'll beat Minimax in terms of quality or LTX in terms of speed, but for image editing.

Think FLUX2 was in a similar situation. Z-Image came out pretty close to its release, and it could neither beat that in terms of quality or speed, but rather its ability to edit images.
>>
Getting in-and-out movement with tentacles in H3 in ref2v is pretty difficult. It really wants to do a continuous inwards movement even when I add a timestamp for every individual in/out movement.
>>
fuck you LTX for mangling nipples, can't even let that one alone?
>>
>>109570413
Take the all-the-way-through pill
>>
>>109570449
Can it actually do that?
>>
>>109570449
I'm definitely going to try that. Time to stick it in the ear.
>>
File: 1769477667677184.png (3.78 MB, 3609x7679)
3.78 MB PNG
>state of cloud
>>
>>109570214
Catjak confirms his schizo status once again and proves that he was the sdg spammer. Have fun in your rotten shithole without community. You're just a bunch of losers posting slop into the void.
>>
File: PepeTheGrok.png (135 KB, 792x695)
135 KB PNG
>>109570595
Doesn't have to be that way. It's quite an awkward Pepe, though...
>>
>>109570595
local has surpassed cloud for me, thanks to H3.
the biggest advantage cloud had was image editing, but almost every edit i wanted to make got refused. that's just not a problem with H3
>>
>>109569504
Huh?
Wasn't this thread deleted?
>>109570640
>>109570673
Is this really how you should be spending your time, Julien?
From what I've glanced looking at past threads, you seem to have more urgent matters to attend to
>>
>>109570640
sdg had so much more soul its unreal
schizos like catjak ruined a good thing and still act smug about it years later
>>
>>109570697
yeah i can honestly say that i finally have grok imagine at home with H3
>>
File: lucifel99.jpg (539 KB, 1790x1658)
539 KB JPG
>>
>>109570771
eh?!
>>
File: grafik.png (59 KB, 576x510)
59 KB PNG
Is he wrong?
t.other Anon
>>
>>109570771
can you change references, as you go?
>>
>1 hour for a 15 second video
thank you custom nodes
>>
>>109570781
theoretically yes. if you aren't using jeet code
>>
Why is China releasing so many good local models?
The answer is simple. They're willing to take a short term financial loss to become the face of local AI.
The endgame is hardware. China is building its own AI accelerator ecosystem to eliminate its dependence on Nvidia.
It's essentially the same strategy Nvidia used with gamers, but on a much larger scale. As local AI becomes integrated into everyday computing, almost everyone becomes an "AI enthusiast." Offices, hospitals, schools and businesses will either run these models locally on their own hardware or use cloud services running on the same hardware.
That's why going API only from the start makes no sense. Open models build the ecosystem first. Millions of users test them, find their limits, optimize them, build tools around them and integrate them into real world workflows. They're effectively a massive decentralized beta testing and R&D network.
Then, when Chinese hardware is ready to compete with Nvidia at scale, the models, tooling and users are already there.
>>
>>109570833
slop writing still comes trough
>>
>>109570413 (me)
The solution is to specify the actual direction in addition to writing in / out. Pull back towards the right, push in towards the left, and so on.
>>
>>109570448
the funny thing is, it was literally jews who created porn
>>
File: MiniMax-H3-00015.mp4 (3.47 MB, 736x576)
3.47 MB
3.47 MB MP4
>>
>>109570106
I got music from my 5 second gens stuck in my head one or two times.
>>
>>109570931
same, i have lots of little kinos that i wish could be extended
>>
>>109570697
How do you edit images with H3?
>>
File: output_8.mp4 (3.82 MB, 2048x1130)
3.82 MB
3.82 MB MP4
>>109570699
I'm confused to why he always replies to himself desu
I'm going to make a new H3 gen of him crying in a padded room through a cctv feed, need to think of what I want to add. I don't think he can run H3 either which is hilarious and sad
>>
Are there any particularly good scheduler + sampler combos for H3? Without any lora applied.
>>
>>109570976
kek
>>
File: erecting my autocannon.webm (1.71 MB, 1280x896)
1.71 MB
1.71 MB WEBM
>>
>>109569562
I think AR glasses and big foldables are going to take off because it sucks not to be home and genning. I don't think this will always be a minor hobby for people in the loop like it is now. I think genning will be one of the most popular activities in man hours period. And the best time to gen is in the background when you aren't waiting for it.
>>
>every ass shake prompt i try ends up way too fast or too clunky
Prompt Gods i need your help
>>
>>109571025
prompt it as a dance and ask an LLM which tiktok dances have ass shaking, or just prompt hip sway to the fast rhythm or some shit
>>
>[File deleted]
>>
File: Test 00009(1).mp4 (3.7 MB, 1376x768)
3.7 MB
3.7 MB MP4
Waited 16 min for this shit. Thought it'd turn out better if I tried going with the full 20 steps.

Should've honestly just done three different scenes instead of trying to cram all that into a single gen.
>>
File: i_00098_.png (941 KB, 768x1376)
941 KB PNG
>>109571099
>Waited 16 min for this shit.
actually looks pretty good.
>>
File: 1774352263376224.jpg (440 KB, 2551x2480)
440 KB JPG
>>109569504

>Using FLF2V model with R2V nodes
>Good motion.....
>.......Worse prompt adherance
>Using R2V model with R2V nodes
>Bad motion.....
>.......Good prompt adherance

Really got monkey paw'd here
>>
what to use to enhance my prompts for h3 between gemma 31b, qwen 3.8 27b and glimmer?
they all support image input I think
>>
so this eros guy who created sulphur just scammed 10k $ from his followers? kek
>>
What's the speed difference between a 3090 and 5090? Currently on AMD and want to switch to Nvidia.
>>
>>109571294
?? qrd ?
>>
>>109571299
5090 is the fastest obviously. But its used car price so go for 5090 if you have the cash
>>
>>109571300
google it, french pedotroon
>>
>>109571299
like, 2x
>>
>>109571309
Significantly faster?
>>
>>109571311
Dumb debo
>>
>use "extremely" and get the speed i want in a thrust but gives me an anime shockwave effect
>use "very fast" and it's like 1/5th of the speed with no effect
>>
>>109571324
>he is fucking her at 4Hz
>>
>>109571312
Seriously? I thought 5090 would be like 4-5 times faster than 3090.
>>
>R2V
>tell it that camera stays fixed
>still changes the scene half way the shot occasionally
Annoyed.
>>
>>109571347
that's not how generational performance improvement goes
>>
>>109571347
theyre both fatass cards that fit the model, only so much speedup you can get while still being on effectively the same amount of VRAM and only a couple of generations apart
>>
>>109571363
sucks to suck
>>
>>109571336
Now that I think about it how is this not a thing, like guys keeping track at what frequency they can fuck at and who can hit the fastest. Some baby dick could get a reputation for fucking like a sewing machine
>>
>>109571363
>"camera stays fixed"
This wording exactly? I know the R2V prompt guide doesn't mention it explicitly but you might as well use the wording supplied in the base model prompt writing guide.
>>
>>109571392
>Some baby dick could get a reputation for fucking like a sewing machine
stop humble bragging
>>
>>109571392
gonna need water jets to keep the thing cool at some point
>>
>>109571347
Yeah that's pretty close to moore's law, which is dead
>>
>>109571347

It can be a lot faster, depends on the task.
While it's not the 3090, but I went from a 3080 to 5090 and it's like 4x faster in image generation and can generate a shitton more images at the same time, so the speedup can be even greater in that way.
It's hard to compare them like apples to apples, because that 32gb of memory is such a big deal even when comparing to 24gb.
I'd recommend not getting a card with 24gb as it's the worst cope memory class. It'll just piss you off. It almost gets you to experience the really good stuff, but you'll still be a bit too short on everything and have to go with the smaller options.
You don't have to worry about not being able to load models into 32gb, but you have to give a damn at 24gb.
If you have the money then get the 5090. If you can't afford it then get the 3090.
>>
I wouldn't buy at all because when prices drop which they will in a couple of years and the next big thing cost less, you will be seething be it 2 or 5 years from now
>>
I'll be seething about something regardless so I'm buying now
>>
>>109571436
If anon is loaded with cash he should buy now so he can have fun for those 2-5 years at least.
If he's one of the retards who'd buy this shit in credit then yeah please don't do this, get another hobby for now.
>>
>>109571363
Use R2V model. I2V model tends prompt adherance is bad for complex movement
>>
>>109571436
>which they will in a couple of years
keep dreaming
>>
>>109571436
Can you wait until 2040 ?
>>
File: MiniMax-H3-00019.mp4 (3.41 MB, 864x480)
3.41 MB
3.41 MB MP4
put me in the collage
>>
>>109571436

If you're employed even as a burger flipper then the price of this GPU shouldn't be any kind of an issue to deal with, even if you have to use credit.
Almost every average person spends more than the price of 5090 on pure frivolities every single year without even noticing they do it and that's not even taking actual hobbies into account.
The advice to just wait for years is retarded, unless you live in a homeless shelter or something.

Besides there's no guarantee of price drops or availability. AI demand isn't going anywhere even if the companies got fucked tomorrow.
Let's say that 6090 comes out in 2028 or even 2029, which is a realistic expectation. It'll be out of stock at the beginning for god knows how long, as we are now all used to the halo product going up in price for 3 generations, so might be +3 grand in real terms.
Then you order one and have to possibly wait for +6 months to get it.
So now it's almost 2030 and you finally have your GPU and you "saved" like 2k.
And let's be real you won't have that 2 grand available anywhere at the time, because you didn't save it, you spent it on some bullshit over the years.
>>
>>109571436
>which they will in a couple of years
lol
lmao even
>>
>>109571466
I don't need to because I read the market
>>109571460
It will happen it's all fake inflation and lawsuits are piling up
>>109571515
I agree but based on the average poster you think anons here have that sort of discipline
>>109571526
Doomer with a shit rig
>>
just made my childhood sweetheart kiss me on the lips (using adult pics of us)
I feel giddy
>>
>>109571515
Kind of afraid the 6090 will just have the same 32GB of VRAM as the 5090.
>>
>>109571558
lol
>>
>>109571541
So since you read the market, when 5090 or RAM prices going to get down ? two weeks ?
>>
>>109571126
>>
>>109571458
I use R2V, lmao, I2V works much better, but I want to add certain details from multiple ref images.
I think model is confusing something in the prompt. I'll make some changes and queue more and see what works.
>>
>>109571363
The phrase you're looking for is static view
>>
Is 5060 TI 16GB that much faster than 4060 TI 16GB?
>>
>>109571606
nvfp4 is useless so no
>>
>>109571558
this. it's been genuinely therapeutic seeing myself have sex with every hot girl from my college who didn't even know i existed (using adult pics).
i feel this strange sense of confidence in myself that i haven't felt in years. it's like... everything's going to be okay
>>
>>109571558
>>109571628
This is the kind of people that post here? Jesus, man. No wonder it's always a mess.
>>
>>109571515
>If you're employed even as a burger flipper then the price of this GPU shouldn't be any kind of an issue to deal with
Most people aren't willing to drop 5k on a computer part
>6090
>costing 3k
Lol, it'll be 7k minimum in the current economy, not that it will even be released since gaming is dead and all the big names are only focusing on AI workstations
>>
fp32 or fp16 vae? what's different?
>>
File: file.png (572 KB, 1504x773)
572 KB PNG
I'm about to join the big leagues!
>>
>>109571633
JSID already
>>
>>109571515
>Almost every average person spends more than the price of 5090 on pure frivolities every single year without even noticing
Just because the average person is retarded doesn't mean you should be.
>>
>>109571633
stop projecting. we all know the kind of stuff you gen
>>
>>109571436
We haven't even seen the profitability shot from AI though. They aren't doing all this to replace google and wikipedia and make brainrot scrolling videos. Soon they can start deleting wages and capital. There won't be an "AI is good enough" phase where they stop buying all hardware. They won't stop until it can perform every information job.
>>
>>109571597
Weird, i got the opposite.
I guess this is the specific image issue again. Some pic is better in i2v model, but some pic is better in r2v model.
>>
>>109571638
oh never mind, other is for audio.
is it possible to disable audio to save resources though if I don't need sound?
>>
This model is so fucking versatile
https://reddit.com/r/StableDiffusion/comments/1vp9nvj/pushing_minimax_h3_v2v_to_the_absolute_limit/
https://reddit.com/r/StableDiffusion/comments/1vpt4pz/minimax_h3_day_of_the_tentacle_remaster/
>>
if i wanted to be on reddit, then i would be on reddit. stupid ahh reddit nga
>>
>>109571684
/ldg/ should start making cool shit then
>>
>>109571558
In a few years you will have a sexbot that you can overlay photoreal guassian splat graphics over with an XR headset. You will bathe with her every day and cum in her every night with no consequences and no hassle, and she will be your personal assistant that does shopping and housekeeping. She will perform better that any human possibly could. She will never present a challenge and will always be able and willing.

Then you will feel emptiness, not because she isn't real, but because you know if you scored the real thing it would not turn out this well. And this is all there is. You live on an island of cumming in an ocean of nihilism until you die.
>>
>>109571684
>ahh nga
If I wanted to be on TikTok, ...
>>
File: yay.gif (511 KB, 840x488)
511 KB GIF
>>109571706
can't wait bros
>>
>>109571706
>In a few years
10 minimum. More likely closer to 15 or 20.
>>
>>109571708
you are on tiktok
>>
>>109571593
Jackass reread my post, I bought early
>>
>>109571558
I'd make my HS crush kiss me, but I don't keep any pics of her :(
Shame, I never confessed. I think she liked me too. She offered drink a few times, but I was too much of a beta.
>>
>>109571684
get the fuck out of here you dumb nigger
>>
>>109571744
>>109571558
i can only assume you guys are false flagging journos trying to start shit for content, maybe keep the illegal posts in /b/ or anywhere but here
>>
>>109571681
no one cares about making gay ass dinosaur shit bro. i just want to gen videos of my winkie inside prime megan fox (transformers (2007)
>>
>>109571719
Remember that sitting is 2nd hand smoke, you are more likely to survive in a big vehicle with side airbags, food should be well done, stay away from animals, etc
>>
>>109571756
>reddit ahh nga telling me to leave
>>
>>109571763
>wanting to fuck that whore instead of the dinosaur
Faggot
>>
>>109571758
>kissing is illegal
fuck off, fag
>>
>https://huggingface.co/TenStrip/10Eros-Max
Some porn weight merge. No idea about quality
But this dude made some model before. I didn't
>>
>>109571765
>sitting is 2nd hand smoke
What?
>>
I'm too successful with T2V to try R2V right now

I just want to know, can a few pictures of a subject compete with a lora?
>>
i wouldn't even consider buying an nvidia gpu if the amd fucktards would make their software not suck. there's no reason my 7900 xtx shouldn't be getting faster speeds
>>
>>109571797
Its too late anon, the blood clot in your leg has already formed
>>
>>109571801
character loras are completely pointless
>>
>>109571819
lmfao brainlet take
>>
>>109571811
Someone said Vulkan on Linux is faster than any Nvidia with Cuda.
But I'm just an end user with plug and play stuff, so no idea.
>>
>>109571846
Does Comfy even support Vulkan? Vulkan has been pretty nice for LLMs.
>>
its up
https://www.youtube.com/watch?v=LRaaflMtpAc
>>
>>109571843
goofy ahh nga
>>
>>109571796
>massive wall of claudeslop explaining how he supposedly managed to merge entirely different models into the minimax h3 weights
Yeah I'm gonna go with "this is literally impossible and you are suffering from AI psychosis". Haven't tried the model but I guarantee all that nonsense is essentially equivalent to a small random perturbation of the weights.
>>
>>109571864
>custom node
Ew
>>
seems like the rebalancer node is still basically better than all the loras for krea2 uncensoring/uncucking of its ability to follow body-related instructions
>>
>>109571856
sd.cpp supports Vulkan
>>
>>109571856
probably means some specialized cpp other than comfy.
>>
>>109571873
Care to show examples?
>>
>>109571870
>>
>>109571870
You think people download that model to prompt modern arts?
If it can draw bagina, it does what it supposed to do. 1 job
>>
sd.cpp where's the nodes?
>>
>>
>>109571595
based
>>
File: file.png (6 KB, 297x108)
6 KB PNG
if you want quality you gotta go high res and many steps and wait for it
>>
File: 1766979941147478.jpg (12 KB, 417x201)
12 KB JPG
what if I use the turbo lora but don't reduce the steps from 20?
>>
>>109572154
just lower the lora strength so it doesn't cook the image
>>
>>109572154
hmmm
decrease lora strength
or decrease step count & increase MP
>>
Is there a trick to getting h3 to not add unprompted motion? I'm trying to do a static shot that has very little movement, but it keeps adding motion or moving the camera.
>>
static camera, steady camera or similar
>>
>>109572133
I've seen the original image of this.
>>
>>109572154
just ask them to train 20steps lora that mimic 50 steps
>>
>https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md
>The camera holds a static shot
>>
>>109572213
>>109572253
>>
>>109572253
Using R2V, I have static shot mentioned multiple times and first frame ref set to <Picture 1>, still it starts the SHOT 1 from completely hallucinated scene..
>>
>>109572213
in ref or t2v? never had camera issues with t2v myself but ref has sometimes tried to force the pictures into the output, I still don't have a firm grasp on the minutae of that model
>>
>>109572276
you need to use timestamps to mark the beginning of your actions. any scene descriptions go before that
>>
>>109572227
>>109572264
I have static shot in the prompt. Really it's the subjects moving around, idle animations, etc that is the problem. I'm trying to make them motionless with some animation in the background.
>>
>>109572284
then use
>The camera holds a static shot as action # goes on.
>>
>>109572213
>>109572253
It's also a good idea not to just blindly give the guide to an LLM, but to actually dissect how the video model itself interprets prompts.
For example:
"The camera holds a static shot."
Subject (The camera) + Action (holds) + Object (a static shot)
Breaking prompts down like this helps you understand how the video model interprets different parts of a prompt, rather than just relying on an LLM to follow the guide
>>
>>109571558
>just made my childhood sweetheart kiss me on the lips (using adult pics of us)

just made my sister touch my butthole using an adult picture of her, and a baby picture of me in the bath.

CIVILIZATION IS OVER.

You can either embrace the complete demolition of consent by the machine, or you can walk into the woods and eat grubs forever.
>>
Anyone having issues with the generated videos, even at 1mp, are very low quality overall? I'm seeing a ton of artifacting.
>>
Defining just one shot with multiple time stamps after it gave better results than multiple shots. But that was just one gen, gotta try more.
>>
are you guys using sage attention?
>>
>>109572359
no, we are all using ck with spectrum
>>
>>109572345
you mean color banding? probably from video compression
>>
>>109572359
yes, you dont?
>>
>>109572359
what does that even mean? I use the wf that came with comfyui H3 tutorial
>>
I have 64GB of RAM, with two video cards, one's a 5060, they have 26GB of VRAM between them.

In both ComfyUI and Wan2GP, I got it to work once and then it OOMs every time. even after I switch to the pruned one, and rebooted.

got Comfy running on cu130 just for this.

what do.
>>
>>109571294
arent sulphur and eros creators different people?
>>
>>109572407
lower resolution, use smaller models, do less stuff with wf at once, etc..
>>
>>109572407
downgrade a little. you'd still be on cu130 so you still get the h3 speed gainz. bleeding edge torch is trash. these are the comfy ~v0.28 defaults:
python_embeded\python.exe -m pip uninstall -y torch torchvision torchaudio && python.exe -m pip uninstall sageattention -y && python.exe -m pip cache purge
python_embeded\python.exe -m pip install torch==2.9.1 torchvision==0.24.1 torchaudio==2.9.1 --index-url https://download.pytorch.org/whl/cu130
python_embeded\python.exe -m pip install https://github.com/woct0rdho/SageAttention/releases/download/v2.2.0-windows.post5/sageattention-2.2.0+cu130torch2.9.1.post5-cp310-abi3-win_amd64.whl
python_embeded\python.exe -m pip install -U "triton-windows==3.5.1.post23"
>>
>>109572372
It's not the video compression. I've tested with png and proress, 8bit, 10bit. It's even in a fucking bf16 model of h3.
>>
You know what we really need for minimax is more vibe coded custom nodes.
>>
>>109572478
they're different people, he's just trolling.
the part about the sulpher guy scamming reddit out of $10k is true though. you can't realistically finetune a distilled model and expect good results, and doing it properly would cost way more than $10k
>>
>>109572497
i don't have that problem
>>
>>109572080
kek what the fuck
>>
>>109572359
i use CK+sol+turbo
I don't care about quality anymore. This model is so slow
>>
>>109572345
wtf am i looking at, just catbox
nobody mod this shit. unless this is CP or gay pony then don't
>>
>>109572407
Use the chunk feedforward and lowvram nodes, start comfyui with --lowvram flag.
>>
why can't models offload to another GPU instead of system ram to make things faster?
>>
File: MiniMax_H3_00039_.webm (509 KB, 960x544)
509 KB
509 KB WEBM
>>109571801
>a few pictures
I gave it exactly 1 (one) picture and it just werked even though it was a non-standard, complex design
first try it fucked up her mouth by trying to make it open normally, but after adding a specific description of the way her mouth is shaped to the prompt, it replicated it perfectly
>>
File: Fuck Comfy.jpg (54 KB, 823x476)
54 KB JPG
>>109569890
>>109569913
>>109569920
Fuck, that was a total blind spot for me too, although mine seems to be clean, but still, the fact the browser isn't within the firewall rules and can connect to anything is a security flaw. Now I really have to run it in a separate browser process with separate local rules.
>>
something's telling me that minimax will release their image model. i just have a good feeling about it
>>
>>109571642
lmfao.. yeah... right.
>>
File: MiniMax_H3__00103.mp4 (1.53 MB, 800x1056)
1.53 MB
1.53 MB MP4
>>
>>109570990
this is pretty good
>>
>>109572710
Didn't they already say they would?
>>
>>109572719
I'm reporting this bitch to the state board
>>
>>109569890
Everyone with a functional brain knows what's the reason for that.
>>
Is there a mistake here? In the instructions for I2VA and L2VA say to use <Picture 1> but the FL2VA just says Picture 1 without the <>. Is that actually correct?
>>
>>109572653
Gpus talk through pcie port, dood. You'll get bottlenecked anyway
You can use two gpus to gen 1girl x2 though.
>>
WHERE ARE THE FINETUNES
>>
>>109572777
>finetroon
>>
>>109572710
I want the image model just so you can test what terms it understands faster

My current beef is "white balance" or "no color grading" or "no color filters" don't seem to do shit even though that's a really simple request. As you raise the resolution, it starts drawing from modern film making and you literally see perfect white balance become teal filtered. My only fix is to prompt blue sky or blue water, but if the scene didn't have those I don't know what I'd do. Prompt something vibrantly orange? Teal filters would dull orange so they shouldn't be able to co-exist.
>>
can minimax make videos of girls singing this song?
https://www.tiktok.com/discover/i-like-the-bbc-song
>>
>>109572228
what do you think the original image is
>>
>>109572709
You can run it in a sandbox coupled with a proxy so that the only thing it can do is communicate via its socket through the proxy which in turn listens on 127.0.0.1:8188
>>
>>109572758
Prompting with minimax is so retarded.
Bring natural language back ffs
>>
>>109572846
it is retarded but it's also easy to ask chatgpt to make the prompt after giving it the docs
>>
>>109572832
I know, I am just shocked that such a thing slipped through my mind for years, even though I am aware about all of it.
>>
>>109572856
"Sorry I cant do that" - CuckGPT
>>
>>109572868
use gemma for nsfw prompts it works just as well
here, example command you can use for 24gb vram

  llama-server \
-hf unsloth/gemma-4-31B-it-GGUF:UD-Q4_K_XL \
-ngl all \
-c 128000 \
-np 1 \
-fa on \
-ctk q4_0 \
-ctv q4_0 \
--no-mmproj-offload \
--host 127.0.0.1 \
--port 8080
>>
>>109572825
Do you have it? Probably from Warcraft.
>>
>>109570976
The api guy should be schizo anon
Like go over to julien, kick him hard and go on a date with comfy
>>
>>109572903
schizoanon why would you stoop to referring to yourself in third person :^)
>>
>>109572913
Schizo anon saved the whole ecosystem from that pedo and his vibe coded garbage
He should get a medal or something
>>
>>109569569
what a gay
>>
>>109572926
I agree, even though some may have worked as much as you :^)
>>
File: Test 00027.mp4 (3.99 MB, 1076x600)
3.99 MB
3.99 MB MP4
Think I'm over my honeymoon phase with Minimax. Thought it could handle pretty much anything at first, because it could do so much more than other local models, with its ability to generate multiple shots of a simple input image being the most impressive to me.

But now I know I shouldn't get overambitious, especially with gunfights and more complex scenes.

Will probably also move towards more one-off stuff, instead of trying to chain some generations together to tell a sort of longer story.
>>
>>109571473
The turbo lora is shit
>>
>>109572936
Have you tried keyframes?
>>
>>109572987
>>109572987
>>109572987
>>
>>109572960
No, because I don't know how I would even create them in a consistent enough style with consistent enough characters.
>>
>>109572709
Disable api nodes, dent
>>
>>109572936
I'm sure you're prompting all sorts of shit that's not coming through but maybe you need to timestamp the seconds when the shot begins and when the kills are supposed to happen so it's not 9mm vs 9mm shootout

Other than that it's not bad. The worst problem is random bullet hole placement. Also no reciprocating ejection port and spent brass looks weird.

Also lingering dust from the bullet holes and reciprocating ejection port with spent brass would
>>
>>109573046
Yeah, I would probably have to autistically script the whole scene out with very precise details. In this case I only said that each shot should create a bullet hole if it hits a wall, or blood and sparks if a person is hit. That might've been way too general.
>>
>>109573090
Yeah I don't think conditional statements like "if this do that" work in prompts. You should write a prompt like the image or video already exists and you are describing what you see and don't see.
>>
>>109571473
you using wan 1.0?
>>
>>109573115
Sometimes it does work, especially with models that have a good world understanding. Does work okay with water splashes for example. Also worked completely fine with blood sprays for me.

Evidently doesn't work with bullet holes.
>>
How do you proompt anime characters in a photorealistic scene interacting with real people?
>>
>>You can totally run comfyUI even without an nvidia guise (not clickbait!)
>32g ram minimum
I'm going to kill someone and it's probably me
>>
>>109573264
detailed_description:
Roger Rabbit-style, in the style of DeviantArt 'me and my waifu' images.
[Shot 1] whatever goes here
>>
>>109571473
why is your output quality destroyed
>>
>>109573189
Yeah it works in the sense that the models have cause and effect. So for example, if I type a fat woman jumps off a roof towards a pool, if she lands in the pool then splash water. I can remove "if" and the same thing will happen, the model will probably just ignore "if". But you could be confusing the model with sentence structure. Think about the process backwards. Someone tagged photos of people jumping in pools and they would have no reason to describe any of the content with "if". It's probably a meaningless term right now.
>>
>>109573309
Didn't work. It made the whole thing anime.
>>
>>109573504
Yeah, splashes might be a special case as the model will do them anyway, but blood splatter is not something the model will usually do by itself. An "if" definetly works for those, as I've found out multiple times.
>>
>>109571363
>[Shot 1] At 00:00:00, The camera holds for a static shot while ...

READ THE FUCKING OFFICIAL PROMPT GUIDE YOU FUCKING PIECE OF SHIT.
>>
>>
>>109574123
How are you prompting that perspective?
>>
https://huggingface.co/ethanfel/Qwen3-VL-32B-Ultra-Heretic-H3-ComfyUI-INT8-ConvRot
Anyone tested this? Does it help to generate nsfw videos?



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.