[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: MiniMax_H3_00562_.webm (858 KB, 768x960)
858 KB
858 KB WEBM
Previous: >>109482228
https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Blessed thread of frenship
>>
stop having a melty just because you didn't start the thread. you and will smith schizo are both fucking losers.
>>
next bread will be collage
>>
its been 3 days and i cant stop gennning
>>
>>109483997
Purposefully not including the full OP is trolling
>>
>>109484001
you telling me i posted kinos in the last thread for nothing?
>>
>>109483964
How do you capture likeness of real people in minmax? Refs alone won't work for me, its always a bit off.
Is a lora my only hope?
>>
File: file.png (22 KB, 233x175)
22 KB PNG
>>109484014
>it's not a real OP if it doesn't include my manifesto
>also he's the schizo not me
>>
>inb4 n*gbo malware spam
>>
>>109484025
prompt issue
>>
>>109484019
the collage userscript lets you do custom collages with any images
pray baker sees your previous kino
>>
>>109484014
You've done him
>>
How is debo so autistic? How can he care this deeply about trying to be the OP?
Why can't autistic people handle rejection?
>>
>>109484046
>Does ComfyUI have the ability to offload to system RAM on Linux systems if the GPU runs out of VRAM?
Yes. Anon will lament Comfys memory handling but if you're using a normal install it works great.
>>
File: 1756687653626751.png (677 KB, 800x487)
677 KB PNG
>>109484058
>Why can't autistic people handle rejection?
Idk but I really wish orange man could have found a cure for autism, would solve a lot of problems on the internet lol
>>
Why are subgraphs still so broken in Comfy?
>>
>>109484013
i know its wrong for me to generate videos of grills i once had a crush on (not morally, but mentally, for me) but i cant stop
h3 is the most dangerous model ever to be released unironically
>>
File: MiniMax_H3_00576_.webm (1.67 MB, 672x960)
1.67 MB
1.67 MB WEBM
>>
LEAVE THE PEDOTROON THREAD. COME HERE
>>109483968
>>109483968
>>109483968
>>
File: 334644052356012_00001_.mp4 (3.79 MB, 1184x800)
3.79 MB
3.79 MB MP4
Pretty good at muppet motions, got some Jim Henson in there.
Sound: https://files.catbox.moe/fxfsns.mp4
>>
>>109484086
Im that OP. We stick her instead
>>
>>109484086
It's not working debo. You've realized nobody's coming to your thread so you're trying to gaslight as the other OP in hopes of starting up drama.
The problem is you're just too stupid, and don't know how to be discreet.
>>
File: 1783686757903883.mp4 (835 KB, 512x800)
835 KB
835 KB MP4
>>
>>109484086
>PEDOTROON
why?
>>
>>109484086
we all know you're the troll baker trying to split the thread lmao. nice try
>>
>>109484093
>It's not working debo.
not sure that's debo, that OP has the rentries in there
>>
>>109484094
So you can pretend doing something in blender? Genius move
>>
nasty ahh video. im going to the other one bye
>>
>>109484099
>illiterate retard fails to understand a simple post
>>
>>109484086
You're on the wrong website. Go back.
>>
Turbo Lora works really well for me now. At least dont make too much movement for it. Genning time is as fast as LTX too
>>
>>109484013
>>109484079
Someone needs to start a rehabilitation clinic of terminal genners
>>
What are the cool kids running now? Last time I tried any local diffusion stuff. Wan 2 was the big thing but AMD cards still got a ways to go. I imagine it's gotten a bit better. I'm using a 9070 XT 16GB.
>>
>>109484112
>Turbo Lora works really well for me now.
I still have the broken sound, even with the latest commit...
>>
>>109484089
more anya!
>>
>>109484119
Use the non Pruned Int8 model and download the latest version of Turbo Lora
>>
>>109484127
>Use the non Pruned Int8 model
NUH UH, FUCK THAT
>>
>>109484112
>Genning time is as fast as LTX too
not for me. decode stage takes too long
>>
>>109484013
but wait, I was gennig all the time wit hWan2.1, 2.2 and LTX. Everything changes but at the same time nothing changes.
>>
>>109484134
>decode stage takes too long
go for the int8 vae, it's 2x faster
https://huggingface.co/Kijai/MiniMax-H3-experimental/blob/main/minimax_h3_video_vae_int8_convrot.safetensors
>>
>>109484110
Don't bother, it's just debo having an autistic fit.
>>
>>109484025
enhanced prompt?
>>
>>109484127
>non Pruned Int8
based, getting better gens with it and at the same speed
>>
>>109484141
And just when he left /adt/ in peace, he just had to come back here of all places.
>>
>>109484132
Its only 5 seconds slower compared to pruned model now with Turbo Lora enabled
>>
fastest decent quality 5 step WF. Seems like the trick was you needed a special node to load the turbo lora with. https://github.com/Larryvrh/ComfyUI-MiniMax-H3-Turbo
The comfy converted ones sucked
https://files.catbox.moe/z5zu2y.json
https://files.catbox.moe/3zfjtb.mp4
>>
>>109484167
I don't care I'm not downloading an additional 33b model just to make a turbo lora work, I'll wait for a lora that works on the pruned version, period
>>
Check out /vp/, they're actually talking about Anima costing a lot of buzz on civitai:

>>>/vp/59474109
>Looks like it costs considerably more buzz to gen with Anima on Civit. With Illustrious, making a Blue gen would cost 5 buzz; on Anima, it's 14 buzz. I'll still experiment with it for Blue at least but even if I like the results, I doubt I'll abandon Illustrious yet for general genning.
>>
>>109484086
/r/ levels of mental illness
>>
>>109484112
it fucks up the sound too much and makes everything live action look overly detailed and thus too fake for low res
>>
>>109484140
i don't think wan2gp supports that yet sadly
>>
>>109484177
>poorfags crying
This interests me not
>>
imagine needing buzz to gen
just imagine it
>>
>>109484184
this lol
>>
i seriously lack the self control required to have a healthy relationship with ai
>>
Sulphur is at $7400 out of 10k now.
>>
this baker war situation is getting ridiculous.

can you just make a proper bake with a collage? it's not that hard.
>>
>>109484193
I'm in the wrong grifting business
>>
anon I'm from wan 2.2 era. is h3 uncensored? how does it compare to wan
>>
>>109484189
We all do.
>>
>>109484195
>can you just make a proper bake with a collage? it's not that hard.
if it's not that hard then why won't you do it?
>>
the way of the future
>>
>>109484202
>is h3 uncensored?
yes, it generates unwanted penises if you have a certain POV prompt
>>
>>109484202
>is h3 uncensored?
still no genitals but you can just bring in video and image references, so the answer to the spirit of your question is yes

>>109484202
>how does it compare to wan
superior in every way except generation time. i will never be using wan again. 5 seconds is just too little. it was always a cope.
>>
https://files.catbox.moe/cl55b3.mp4
>>>/wsg/6209484
>>
File: AnimateDiff_00154.mp4 (2.68 MB, 672x1216)
2.68 MB
2.68 MB MP4
>>
>>109484168
thanks for the heads up anon, I will try that out
>>
I'm glad I can prompt AI to do my job now, so I can spend all day figuring out how to prompt things instead.
>>
>>109483964
OP image workflow?
>>
>>109484218
oh how long does it support then?
>>
>>109484247
prompt please?? good morning
>>
>>109484248
>oh how long does it support then?
15 seconds / 362 frames at 24fps (no more 16fps shit either i just remembered that. the year of wan was truly a cope year but it was still pretty awesome honestly)
>>
>>109484247
https://h.uguu.se/RfsjLdtl.mp4
>>
>>109484218
>5 seconds is just too little. it was always a cope.
not only that but wan has no sound and gives you slow mo shit, H3 is like another 10 levels above
>>
>>109484179
Just tried it. It has problem with realism.
I turned the lora str down to 0.8 and it fixed it
>>
>>109484025
you're doing it wrong. Provide a small 3 second clip of the person and a few images and follow the prompt guide. It can replicate and maintain anyone perfectly. Much better than even good loras for wan.
>>
File: qumd'umpe.mp4 (1.07 MB, 1056x608)
1.07 MB
1.07 MB MP4
>>109484116
minimax h3 for video
krea (some also ideogram or z-image-turbo but krea probably suits more people best) for images

you'll definitely want to update everything for the earlier it's less than a handful days old
>>
>>109484264
how do people extend beyond 15 secs? still using with freelong/svi?
>>
kimi is THE least censored model with the tinyest of prefills. the fuck are people on about
>>
>>109484278
the model can go for really long videos if you have enough vram and patience
>>
>>109484281
>kimi
>>>/g/lmg
>>
The ref model's understanding of physics seems to be worse than the fl model.
>>
>>109483964
What model did you use to generate the anime girl?
>>
>>109484292
yeah, the ref model doesn't seem to have been trained as well as the default model, they should have gone for something unified desu
>>
does h3 have better geometry understanding than wan 2.2? I mean does it distort the background geometry when changing viewpoint
>>
>>109484298
WAI. There's an Ace Trainer GSC lora on Civit.
>>
File: 1768594663869925.mp4 (1.42 MB, 608x704)
1.42 MB
1.42 MB MP4
>>
Why is OP asking himself how he made his own gen? Is that considered catastrophic forgetting?
>>
>>109484306
>does h3 have better geometry understanding than wan 2.2? I mean does it distort the background geometry when changing viewpoint
if you still use wan 2.2 we're gonna laugh at you dude, stop asking retarded questions and set up the model
>>
MMH3 is coming to 4 gb eventually right fellow vramletbros? R-right?
>>
File: qumd'umpe_variant.mp4 (1.53 MB, 1056x608)
1.53 MB
1.53 MB MP4
>>109484278
it can work for more than 15s. just like that.
OR people gen scenes with subjects references, after all if subjects stay the same scene cuts are often not even that important
OR firstframe/lastframe

before that SCAIL2 video references for WAN were also a nice easy method but of course pretty much everyone prefers to gen mm-h3 now
>>
File: test.webm (1.45 MB, 1550x2048)
1.45 MB
1.45 MB WEBM
First time using reference, audio seems good but I need to refine the prompt to keep the elements like text on the clothes consistent will try that after this
>>
>>109484202
wan is completely dead
>>
>>109484306
>I mean does it distort the background geometry when changing viewpoint
not even close, it's really clean >>>/wsg/6208323
>>
>>109484116
>I'm using a 9070 XT 16GB.
you're fucked for H3
I tried setting up a 32GB Radeon Pro R9700 to do H3 and ROCm can't int8 convrot so you're slower, and there's bugs in comfy for rocm right now so you're even slower, and you have to use fp8 scaled so you're shittier quality than int 8 convrot.

the best i could do for 20 steps on a 32gb radeon card in 10 minutes is 360p. good luck to you and your 16gb of the same generation
>>
>>109484309
Thanks for replying
>>109484317
Meds
>>
https://n.uguu.se/azRFebCN.mp4
>>
>>109484306
>>109484324
that one is the most impressive to me >>>/wsg/6208793
>>
>>109484322
holy cum
>>
Video generation has made me realize how uncreative I truly am
>>
>>109484341
I half expected a jump scare...
>>
>>109484341
Die faggot
>>
>>109484338
>he fell for the amd meme
we warned you
>>
>>109484341
live faggot

(and maybe make one featuring cynthia)
>>
>>109484351
>Gen a creative prompt
>Vid Gen Gen'd prompt
I truly got nothing... besides person undresses or something along those lines.
>>
does h3 support keyframe reference like wan vace?
>>
Uh.. should I practice this in a VM? Am I wasting my time if I use integrated graphics and 32GB ddr4? I just want to maybe upscale some images and make them less blurry, or perhaps uncensor some photos/videos. Maybe take some DVDrips I made with makemkv and upscale them to UHD
>>
Is debo beginning his transition into barneyfag mkII?: >>109484357
>>
>>109484375
>integrated graphics and 32GB ddr4
Anon I-
>>
So what's the current node/lora setup for minimax? This shit evolves so fast I can't keep up.
>>
>>109484341
https://n.uguu.se/kpqHeHGM.webm
>>
>>109484351
turns out AI won't make you an artist after all
>>
File: MiniMax_H3_00258_.mp4 (3.39 MB, 768x960)
3.39 MB
3.39 MB MP4
>>
>>109484371
https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md
>Use a standalone <Picture N> when the reference image itself serves as a shot's first frame, keyframe, last frame, edited keyframe, or composition anchor
>>
>>109484371
Think so. I saw something about that in this node:
https://github.com/jlucasmcrell/ComfyUI-H3-Multishot
>>
File: 1770700877002404.png (78 KB, 1798x449)
78 KB PNG
>>109484221
ok I can confirm that his PR works, the audio is fixed with the turbo lora
https://github.com/Comfy-Org/ComfyUI/pull/15243
you can go for that lora
https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/blob/main/minimax_h3_turbo_4step_ema_ckpt850_pruned_comfyui.safetensors

default [20 steps] + spectrum (10.12 mn)
https://files.catbox.moe/cl55b3.mp4

turbo lora, strength = 1 [8 steps] (6.23 mn)
https://files.catbox.moe/s0rs8k.mp4
>>
gotta see if I can run H3 on my poverty 8gb
>>
>>109484362
i have it from my job and use it as an fp16 merchant for normal pytorch work in which it does fine
i would never recommend anyone purchase it for any reason outside a corporate envirobnment for specific use cases
>>
>>109484397
the thing that I notice is that the turbo lora slopifies the gen though, the original kept the low res image style while the turbo lora one makes it HD
>>
>>109484351
skill issue
>>
File: 1770580683182491.mp4 (1.16 MB, 544x768)
1.16 MB
1.16 MB MP4
sd 1.5 gen from mid 2023 into h3
>>
>>109484397
extremely cute migu voice here
Amazing how much keeps, though it's unfortunate it couldn't keep track of the train environment
>>
File: 1780549433887398.png (74 KB, 1653x204)
74 KB PNG
>>109484384
>So what's the current node/lora setup for minimax? This shit evolves so fast I can't keep up.
here's what i'm rocking
only arguments right now between people are whether spectrum or easycache or another cache node is better. i like spectrum the best and will be keeping it for now.
low vram and feed forward nodes do not seem to impact quality at all from what i can tell
>>
File: AnimateDiff_00160.mp4 (3.29 MB, 672x1216)
3.29 MB
3.29 MB MP4
>>
goddamnit, just because I ask for a broken bone doesn't mean I literally want to see the bone...

https://d.uguu.se/WbKeHkJX.webm
>>
>>109484425
>Amazing how much keeps, though it's unfortunate it couldn't keep track of the train environment
it's gonna be better the turbo lora isn't even finished yet, godspeed
>>
File: Chunk Feed and Sneed.png (6 KB, 222x111)
6 KB PNG
>>109484427
>>
Can h3 use a random voice clip as a reference? "Make <Subject 1> sound like <Audio 1> while reading <d>"
>>
ok h3 seems cool
>maybe uncensored
>15 secs+
>more geometry understanding
>support keyframe

gotta try it on my dgx spark. last time if I want to overcome the issues above I had to
>lora for uncensored gens
>freelong for long video
>learn nvidia lyra-2 but too much effort to bother with
>vace for keyframe

maybe h3 struggles with mirror/reflection/object behind glass too. we'll see
>>
>>109484432
>BRUTALITY
>>
>>109484381
I'm a poor old fool, lad
>>
>>109484453
A Sneed & Feed layer is a feedforward neural network component used in text-to-video diffusion models to process and transform feature representations. It takes input features, passes them through a larger hidden layer (typically using ReLU or GELU activation functions) to capture complex non-linear relationships, then projects back to the original dimension. This two-stage expansion-contraction allows the model to learn richer representations of visual content while keeping computational costs reasonable, helping it generate coherent video frames that align with text prompts across multiple temporal steps.
>>
File: 1766981406422032.mp4 (910 KB, 704x1056)
910 KB
910 KB MP4
https://files.catbox.moe/ivxp7j.mp4
>overall_soundscape: clattering bones
>>
>>109484458
nsfw movements (like gagging on a cock) can be prompted and will look good

feed it an anime sex gen without a prompt and itll have her orgasm or the guy thrust it clearly knows
>>
>>109484453
>>109484475
Fuck i was eating while reading this
almost got seeded and feeded into the light
>>
>>109484463
This is one of those hobbies that filters out those with less-than-ideal hardware, sadly :(
>>
Could I ask for a workflow for Krea2 with upscaling? I have 24gb vram and 64 ram
>>
File: almost unpleasant.png (159 KB, 384x390)
159 KB PNG
>>109484476
>haha what clattering bones that's silly
>mfw this sounds better than when we intentionally tried to prompt for realistic bj/popsicle succ sounds

what the fuck H3
>>
>>109484427
>whether spectrum or easycache
Yeah, I'm on the cache camp.
>>
>>109484497
it was fed on spooky bones and souls.
>>
File: MiniMax_H3_00048.mp4 (2.97 MB, 832x1248)
2.97 MB
2.97 MB MP4
>>
Friendly tip that if you are upscaling images for ref pictures, topaz gigapixel absolutely shits all over sneedvr2

https://slow.pics/c/CxiGaUhz
>>
https://files.catbox.moe/7ssawi.mp4

With audio :3
>>
>>109484476
does she have liver disease?
>>
File: 486390448426.mp4 (3.82 MB, 544x768)
3.82 MB
3.82 MB MP4
>>109484427
What about the turbo lora? Is that not better than cache?
>>
>>109484511
don't care + I'm deaf. anyway let me get an order of SONLEX
>>
hey guys oldfag here but new to ai degenning, got my pc set up today 4090, 64gigs ram, fat ssd, downloading the pruned minimax model, will the int8 work or no? im retarded also
>>
File: Megamind-meme.png (129 KB, 680x447)
129 KB PNG
>>109484511
>no lewding?
>>
>tifa tit reduction
>wearing the cringe censored tank top
GET THE FUCK OUT
RIGHT NOW
>>
>>109484515
bery bery cute
>>
>>109484526
I have made a bunch of porn with these two

Unsure if I should share them desu
>>
>>109484525
Welcome fellow oldfag retard. Yes, the int8 convrot will work on a 4090.
>>
>>109484491
Better than my 2nd generation i5 and ddr2 that I was previous running; I keep that desktop next to my new one. I should probably take the drive out of it, but I don't have a 3.5 inch bay in the new case.
>>
>>109484518
>What about the turbo lora? Is that not better than cache?
turbo lora fucks up audio to the point where it's a non-starter for me. i can wait 10 minutes for 15 seconds, i used to wait 5 minutes for 5 seconds of WAN.
>>
>>109484537
please?
>>
>>109484537
>I have made a bunch of porn with these two
its funny how the initial wave of sfw video posts has slowed down because everyone has learned that H3 is very, very good for porn and too slow to waste gen time on shitposts
>>
>>109484525
I'm running it on a 5060ti 16gb on 48gb ram running the int8 model (you need a custom node to load it). 10sec i2v gens take about 6-10mins with all the optimisations (20 step easycache + sageattn) on some workflow I downloaded. Almost exhausts my ram but seems to not be using all my vram right now (around 10gb if I recall). Audio sometimes seems messed up though with voice lines but not sure if it's the optimisations or my prompting.
>>
>>109484351
i'm sure it comes in phases for most.

also i get the feeling a good bunch of the "creative" people prompted llm, you could do that too
>>
damn, i remember when SVD was blowing my mind. shit got good FAST
>>
>>109484456
I was trying that out earlier. It somewhat worked but couldn't quite get the generated audio to match the reference audio to an acceptable level. I imagine someone will figure out how to it better
>>
>>109484556
what res? less than 0.5mp? thats pretty fast for 20 steps and 10 seconds.
>>
File: MiniMax_H3_00587_.webm (828 KB, 896x704)
828 KB
828 KB WEBM
>>
>>109484476
Why does she have a scrape on the bridge of her nose?
>>
AI genning has taught me how to describe my fetish in excruciatingly literal detail.
>>
>>109484567
This migu is a refurbish model with slight cosmetic wear.
>>
File: 1770478835565352.mp4 (1.47 MB, 640x640)
1.47 MB
1.47 MB MP4
>>
>>109484525
literally just plug everything into the template oldfag
after everything works worry about optimizations
>>
>>109484518
turbo lora is dogshit and the creators should be shot for wasting everyone's time. It's barely even faster than optimised 20 steps and messes up motion constantly.
>>
>>109484565
iirc it's 0.5mp with rtx upscaler in the workflow. I did bump it up to about 0.6-7 and that took about 10 mins for gens
>>
>>109484588
it's not ready yet...
>>
>>109484528
>wearing the cringe censored tank top
this post is how i found out tifa is 15 because i never grew up with final fantasy and thought it always looked boring. one of my first hentai was "tifa on the machine" or something like that i found on pornhub
i always thought she was like lara croft age, she has lara croft tits so why wouldnt she be lara croft age her face just looks young because she's asian
>>
>>109484563
Gonna lurk more then
>>
>>109484566
this is a very clean style. high quality looking gen, very nice.
>>
Is topaz worth it for upscaling?
>>
>>109484597
>tifa is 15
Nah... in the main game they're all in their 20s. Maybe you're thinking of Zack's flashbacks.
>>
File: MiniMax_H3_00696_.mp4 (3.56 MB, 1248x832)
3.56 MB
3.56 MB MP4
https://files.catbox.moe/7qfdm8.mp4
>>
>>109484556
>Audio sometimes seems messed up though with voice lines but not sure if it's the optimisations or my prompting.
if you're using the turbo lora it's that
>>
Why can't flux2 make a good character sheet?
>>
File: 779797171146625.mp4 (3.44 MB, 576x736)
3.44 MB
3.44 MB MP4
>>109484545
>>109484588
It's been pretty decent in my limited testing.

https://files.catbox.moe/g7o7h8.mp4
>>
>>109484611
oh ok
but i skimmed a bunch of news posts and it said shes 15 in a cutscene where they added a shirt so you couldn't see her cleavage
>>
Why does last frame ref image for H3 gen stretch vertically? it's obviously not exactly as ref as prompted.
>>
>>109484611
out of ten!
>>
>>109484614
Say the line Bart
>>
>>109484614
I can tell it's some gay shit before even checking
>>
>>109484547
Probably can't...
>>
>>109484601
previously with LTX I was using voice cloning such as OmniVoice and then using that generated audio in my video gens. They may still be a useful option with H3
>>
>>109484605
Made with WAI too.
Ppl here write WAI off way too easily. Sure anima is pretty cool, but WAI still has plenty of capability that anima can't yet produce. Much of this is thanks to lora support, where WAI is still the champion.
>>
>>109484615
Haven't tried the turbo lora yet, have little complaints with the current setup. It's more that the characters would sometimes say gibberish, slightly lower prompt adherence, the audio quality is there though. Haven't bothered yet with prompt writing via llm or whatever, just wrote simple prompts.
>>
>>109484640
Clever
>>
File: 87456132115.jpg (46 KB, 1280x720)
46 KB JPG
>>109484639
>>
>>109484338
There is a Wan2GP version for AMD cards, wonder if that would work. I don't have a AMD card tho, NVIDIA. As always people just have to hope something is fixed
>>
i'm a colossal faggot, what repo is the low mem attention and chunk forward update on. I pulled the main branch and I don't seem to have it
>>
Flux 3 is gonna curbstomp Minimax if the dev version outputs are even close to the API version
>>
>>109484597
Girls stop growing by the time they're 16.
>>
How i clone voice with minimax for a character?
>>
>>109484626
Maybe crop the image to the same aspect ratio as the video?
>>
>>109484514
official seedvr2 implementation is weird as fuck
kinda feels like it cant detect the faces or tiny objects properly
but the original seedvr2 that requires custom nodes is fucked up because it leaves visible upscaling artifacts so you gotta gen 5 diff seeds then merge them together like a long exposure shot
>>
https://files.catbox.moe/rkz0k9.mp4
>>
>>109484660
if it can't run on my hardware out of the box then it's as good as dead to me
>>
>>109484641
there's a WAI anima and a WAI Illustrious, so no one knows what you're talking about buddy
>>
>>109484673
it's pretty obvious by the fact he said anima twice that he means WAI ANIMA. pedantic retard.
>>
>>109484666
Yeah. I tried the setting up the original with old nodes recently but gave up because it wasn't worth it. gigapixel does it in literally 10% of the time with 1/3 of the file size
>>
is it time I try to ref model?
>>
>>109484660
no chance it's gonna dethrone minimax, or else it'll be too big, or else it'll be too cucked, or else dev will be really inferior to max, or else a combinaison of the 3
>>
File: tifa.png (1.66 MB, 924x756)
1.66 MB PNG
>>109484624
Still stacked back then, but you know... japan... What you read is probably about the new one.
>>
>>109484661
>Girls stop growing by the time they're 16.
maybe on average but not all of them for sure. maturation in the face especially can happen until 20 or so, but i'll respond with an actual statistic that the most attractive age for sexually normal heterosexual men is 14-19 normally distributed with around 16-17 being the peak
>>
>>109484614
decent breasts boobiling, almost expect to hear balloon sfx
>>
>>109484660
I've seen comparisons and minimax had better output imo.
>>
>>109484684
You should know it's much slower than FL
>>
>>109484682
no that would be even weirder, because WAI-Anima is still anima.
>>
>>109484673
Didn't even know there was a "WAI anima" model.
Whatever, Illustrious then.
>>
>>109484597
>>109484611
>>109484689
>arguing about a fictional character's age like some redditards
>>
I had Claude make a version of nag for h3, where should I post it?
>>
https://d.uguu.se/NdMzBsXT.webm
>>
>>109484707
As I thought!

I'll have your apology for calling me a pedantic retard now >>109484682
>>
>>109484690
I simply mean that their boobs won't get any larger with age beyond that (besides weight changes or pregnancy), if anything it's all downhill from there.
>>
>>109484713
>I had Claude make a version of nag for h3, where should I post it?
you should post a side by side and prove that it works first
>>
>>109484641
>older model has more community loras than newer model
STOP THE PRESSES
>>
>>109484734
There is almost no support for the cosmos series. So I doubt anyone will go over the trouble.
>>
>>109484727
they can lock in a 10% bonus to breast size if they take estradiol birth control around 16-18 for a few years
>>
>>109484741
>There is almost no support for the cosmos series.
?
>>
Any good character sheet lora for flux?
Found one on civit but it wasn't better than without the lora
>>
>>109484641
After Krea2 got the chroma lora there's no reason for anima to exist desu. Once anima is given more knowledge nobody other than gpu bag holders will bother with that model. It's a slight upgrade that can never beat krea 2, it can't match it in quality prompt adherence or even size. What's the fucking point when I can't even get past basic resolutions without the model losing quality when I can go up to 12mp+ with krea2
>>
>>109484719
I'm sorry you're a pedantic retard :^)
>>
what's this?
https://github.com/duckyshell/ComfyUI-MiniMaxH3-FirstBlockCache
https://www.reddit.com/r/StableDiffusion/comments/1vhlfmw/minimax_h3_firstblockcache_for_comfyui_3033_lower/

30% faster gen time is claimed
>>
>>109484716
why do you keep reposting the same gen?
>>
File: .mp4 (3.36 MB, 1056x608)
3.36 MB
3.36 MB MP4
>>109484559
yup. very nice that it progressed so fast.

>>109484662
it's in the prompting guide complete example
https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md#7-complete-example

ymmv based on what you need, perhaps it's better to still do the voice cloning stuff with like TTS-Audio-Suite and then supply the already cloned audio
>>
>>109484761
another dubious fully vibe coded implementation of sol. You can be for sure that I won't be the one to test it out.
>>
>>109484514
seedvr looks better though?
>>
>>109484759
lol
>>
>>109484774
GPT-5.6-Sol?
>>
How do you stop comfyui's ui from getting slow as you continue doing gens?
i clear vram and unload models and stuff, but the UI itself just goes to a complete crawl if it's open too long. Makes it take foever just to edit text.
>>
>>109484761
>cache
meh, like wan and ltx they're not useful, turbo loras is what people will be using, they work >>109484397
>>
>>109484800
well why not use both
>>
>>109484749
the base mode arch for Anima, cosmos wasn't built for image generation tasks, so people didn't bother writing code and tools for it, which mean you have to make tools from the ground up unlike illustrious, where you had to port existing code and tools.
>>
>>109484804
how do you cache on 4 steps?
>>
>>109484810
I see
>>
>>109484734
You have autism.
>>
>>109484784
sol-attn
>>
>>109484807
Name one tool that is commonly used with Illustrious that does not work with Anima

>>109484817
Yes
>>
>>109484759
You are an idiot.
>>
>>109484800
>they work
>links to post that shows it is dogshit the ruins everything
the delusion and cope of turbofags must be studied.
>>
Pruned vs nonpruned BF16 is it really lossless? I find it impossible to believe you can just remove 26B weights without any degradation at all
>>
File: 1764886870624417.jpg (76 KB, 340x1801)
76 KB JPG
>>109484683
oh its the topaz stuff. guess they rebranded again. ive been using their video upscaler for years now. i do x3 fps and 2x res and they turn out fine as long as source footage doesnt have hallucinations
https://files.catbox.moe/6ni19y.webm
https://desu-usergeneratedcontent.xyz/g/image/1785/99/1785996278746.webm
https://arch-img.b4k.dev/vg/1785974605143.webm
>>
>>109484820
>Name one tool that is commonly used with Illustrious that does not work with Anima
good luck using an SDXL contronet model in Anima.
also touch some code.
>>
File: test3.webm (2.16 MB, 1550x2048)
2.16 MB
2.16 MB WEBM
I can improve this I think, but perhaps 10 seconds makes sense for stuff like this,
>>
>>109484829
>autist doesn't understand that the lora is still undertrained
many such cases
>>
>>109484832
woah these are good as shit
so is topaz free? i've heard of it for years but never tried it.
>>
This turbo shit is just about faster gens right? Nice and all but when do we get the important stuff? Actual anatomy loras
>>
>>109484831
he removed 13b not 26b, no it's not lossless but the quality is the same, it's equivalent to get another seed
>>
>>109484834
>good luck using an SD1.5 controlnet model in SDXL
Kek
>also touch some code.
https://huggingface.co/kohya-ss/Anima-LLLite
>>
Why is /vp/ like this? >>>/vp/59487585
Are they stuck in 2023? They're all using paid generators.
>>
https://h.uguu.se/maTwdere.webm
>>
File: 1786057569244818.mp4 (389 KB, 512x800)
389 KB
389 KB MP4
>>
>>109484858
You mad because they get the (you)s there instead of you. So take one from me instead
>>
Still getting hypervisor errors (assuming OOM related) even after the chunk + low vram attention nodes. 16gb vram, feel like this shouldn't be happening - any fix suggestions?
>>
>>109484868
You have autism, anon.
>>
>>109484845
while anatomy is important to fix, it can be sidestepped if you provide the anatomy to begin with (such as with I2V have the dick already in the image)
>>
File: 1758033387092226.jpg (7 KB, 114x124)
7 KB JPG
This is not good /ldg/bros
My body cant keep up with this. I need to rest. I cant stop gooning and genning for 3 days straight. Low semen count. Give me a fucking break please someone please stop updating.
>>
>>109484882
Drink some alcohol or soemthing and go to sleep.
>>
>>109484841
nah you need to crack it. pretty sure they dont even offer a one time purchase option anymore. i got my copy from rutracker but maybe there are better sources somewhere on fmhy.
>>
>>109484835
>my kind of gen.
>>
>>109484880
Meant to quote >>109484858 btw
You've been there fishing for yous for several days now. Give it a rest.
>>
>>109483964
can a radeon 7600 8 gb get me by?
>>
File: yar.mp4 (2.99 MB, 1376x768)
2.99 MB
2.99 MB MP4
>>109484835
yes and either 10s or you just prompt a little more motion to happen so it's not just essentially a breathing animation.
>>
>>109484857
>only for inpainting
my point still stands
>>
>>109484903
Depends on what you want to gen.
I did images on my 6000 series.
>>
>>109484887
brutal. guess i'll have a look around when i have the time.
>>
>>109484915
endless midna nudes and lewds just as the hello world of hot
>>
>>109484923
Should be doable.
I did it on Linux though so can't say about Microslop Winblows
>>
Updated comfy and now every reference gen with an audio clip tries to make the character say whatever is in the audio, literal same seeds I've used before with the same clip as well.
Something it's fucked up.
>>
>>109484893
Anon, you have autism. Try to get out a little more. Being terminally online isn't good for you.
>>
>>109484912
https://civitai.com/models/2718356/anima-control-pose?modelVersionId=3070469
https://civitai.com/models/2443202/anima-canny-control-lora-controlnet-like?modelVersionId=2748244
I stopped needing cnets for upscaling after Anima anyway
>>
>>109484923
>endless midna nudes
shit that reminds me to do some goblin/shortstack videos soon. thatll have to be first thing in the morning im fucking burned.
>>
>>109484658
https://github.com/kijai/ComfyUI-KJNodes
>>
>>109484214
very cool
>>
What is the longest gen you've made so far Anon? Single gen or combined shots. Feel free to share I'm curious about how you're doing.
>>
>>109484963
close to 2 minutes. i don't have that video anymore
>>
If we get to h4 I can only pray that someone remakes OPM Season 3
>>
>>109484982
be the change you want to see
>>
>>109484397
>he tested in on anime
OF COURSE IT LOOKS FINE FOR CARTOONS RETARD
>>
Ref subject transfer seems kind of spotty, even with a properly formatted prompt for it based on the prompt guide. Sometimes it works, sometimes it doesn't. Sometimes it doesn't transfer every aspect.
I don't like it.
>>
>>109484903
for images it's not too terrible

for video ... you can even try the good models if you have a bunch of system RAM to swap to, but it may be slow enough in most cases that it's just "try" or "queue up overnight"
>>
>>109485004
why are you screaming schizo? forgot to take your meds?
>>
>>109485008
>that gen
aww :(
>>
File: output.webm (2.31 MB, 720x224)
2.31 MB
2.31 MB WEBM
>>109484728
Here it is going from strength 1 (does nothing) to strength 5. Prompt was

"The 3D creature spins in place, performing a 360."

Because the initial frame has a glimpse of gold between her legs where her tail is, the model wants to add a tail to her. The negative was

"The 3d cg creature has a tail."

Ignore the weird artifacting around her border the original image had alpha around her for a transparent background.

It's doubles sampling time though, requires sampling twice at each step.
>>
>>109485012
because you're pushing shitty half-baked solutions that were rushed out in 3 days that only work on cartoons instead of normal shit.
>>
larryvrh 4 step lora is working too well i dont know if i should keep the updated version or not
>>
>>109485027
You can post it then, just make a rentry
>>
>>109485037
So this one is no longer the meta one? https://huggingface.co/QrusherZA/H3_Turbo_ComfyUI/tree/main
>>
>>109485046
No i use the one from larryvrh himself.
But be warned : His latest updated nodes (1.2.2) causes video to oversharpened and oversaturated like crazy
>>
>>109485056
>His latest updated nodes (1.2.2) causes video to oversharpened and oversaturated like crazy
I don't think you need his nodes anymore, I can load with the regular lora node now and it works fine
>>
Just wait for a proper turbo lora. All these are cope
>>
>>109485061
I did that and it causes audio issues
>>
im still using the original lora. i think the training went off course on the later training steps
>>
>>109485071
did you update comfyui? for me the sound is fine since kijai's last pr
>>
File: MiniMax_H3_00267_.mp4 (2.76 MB, 736x576)
2.76 MB
2.76 MB MP4
>>
>>109484942
Anyone else?
>>
>>109485072
>im still using the original lora. i think the training went off course on the later training steps
yeah, looks like 850 is overcooked
https://files.catbox.moe/cdzo4k.mp4
>>
>>109485090
if you got stuff in subgraphs it might be the time to unpack because they're all broken as fuck
>>
>>109485094
you tried using it at a lower strength?
>>
>>109485081
Yes, the stable version
>>
>>109485095
No, is a plain default template and I have no custom nodes at all installed other than KJnodes.
>>
>>109484779
you are tranime-brained I'm sorry
>>
move when ready:

>>109483968
>>109483968
>>109483968
>>
>>109485094
overcookied
>>
>>109485111
>stable
I think the pr is on nightly though, not sure
>>
>>109485128
>>109485128
>>
FRESH
>>109480220
>>109480220
>>109480220
>>109480220
>>109480220
>>
>>109485135
You're hopeless, debo
>>
File: c.mp4 (1011 KB, 864x480)
1011 KB
1011 KB MP4
>>109485015
yea. :o
>>
>>109485147
>t. ranni
>>
>>109483964
good gen
>>
>>109484276
you don't know what a good lora is. there is no comparison.
the ref model is fine but for faces it's still way behind
>>
Anyone using this wf?
https://civitai.com/models/2834514/minimax-h3-t2v-i2v-ref2v-advanced-filmmaking-workflow-or-all-speedups-qol-features

getting a 5sec gen done at 437 secs
3090
6 steps
1.4mp
using turbo lora at 1.40 and another lora

UPDATE
enabled sol-attn and the gen was done at 384.93

noticed that the sound is very bad.. I added a nsfw lora after the turbo lora, unsure if related



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.