[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: output_smal.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
Previous: >>109542134

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Catjack is really getting unhinged thinking about Ani and Debo all the time
>>
>>109544621
did you forget how real lolcows act or something?
>>
Reminder R2V Turbo Lora is coming soon so you dont need to use I2V lora anymore
https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/29
>>
>>109544675
Oh shit, was not expecting a dedicated model for that, ref still needs some improvement, I use the other model over the ref in many cases
>>
File: 1772144910739463.webm (3.05 MB, 1280x800)
3.05 MB
3.05 MB WEBM
>>109544588
This took too much effort.
3 Videos combined. Shot 2 took a while to get it right.
Shot 1 and 2 use Sage. Shot 3 use Kitchen.
Want to make more shots but i dont have autism to do this

https://litter.catbox.moe/1koukn.mp4
>>
>>109544746
go post here instead >>>g/adt/
>>
update comfy (folder), new comfykitchen seems a bit faster

https://files.catbox.moe/98xd5r.mp4
>>
>>109544759
why
>>
>>109544746
Based
>>
>>109544767
they post more child coded gens there so they would probably enjoy your pedo vids more than here
>>
File: kekekekkeeeeeeeeek.png (1.17 MB, 864x1184)
1.17 MB PNG
>>109544759
>kekypow doesnt know how to link
KEEEEEEEEEEEEKEKEKEK nah JSID already wth unc doin bruh
>>
>>109544774
If you see >>>/a/ you get a stroke then
>>
>>109544774
why instead then
we love it here too
>>
>>109544746
Hot
>>
File: DARKNESS for breakfast.png (3.48 MB, 1536x1920)
3.48 MB PNG
>>109544774
what are you on about, faggot? we all loved that skaterkino video the pedoanon shared.
Doesn't mean we all avow the subject matter, but y'know.
Point is i need you to kill yourself PRONTO.
>>
>>109544782
kek
>>
File: 1780433324849908.png (1.2 MB, 832x1216)
1.2 MB PNG
>>109544746
I love megumin body so much
>>
>>109544804
Subject matter aside is something that high quality possible with H3? Never used cloudshit before.
>>
Shaking my head in disapproval so the agent that is watching me knows that I don't condone that type of material.
>>
>>109544814
>megumin body
Just say hebe
>>
>>109544825
I think it could be once we get the 'scaler. I don't know about the voice cloning though we may have already seen its ceiling. Bright side is the length is very easy to do about as well because you can clearly see the scene cuts every 10-15 seconds or so in his vid.

>>109544831
I don't have a webcam to show my agent that i'm clearly shaking my head in disapproval, but believe me that i'm shaking my head as i type this.
>>
https://files.catbox.moe/2tmid3.mp4
>>
>>109544835
Can probably get better voice cloning using a dedicated model, right?
>>
>>109544746
kino
>>
Keeping everything consistent between scenes is a bigger problem than scene length desu
>>
Now do Saikawa and Kanna.
>>
>>109544746
Dunno if it's because of kitchen but I like the first 2 more. She looks kinda 3D in the 3rd one.
>>
>>109544871
seems like an easy problem to solve. you generate a very fast panned video of the scene beforehand and then keep using that as a reference for all the shots that will use that scene
>>
>trying to hand it a pussy ref
>refuses to use it
>>
>Catjack wants to be a pedo now
>>
>>109544814
>picrel
I will do things with this picture, you know that? She'll be a mega slut.
>>
>>109544857
Depends, i haven't kept up with voice cloning models in a while. Funny thing about those, they always release a 1.0 model then no further iterations after that, USUALLY better models came and dethroned them but i don't know if that's the case.
>>
>>109544857
reference model is 100% fine. it can copy gilbert gottfried.

https://files.catbox.moe/32pvfv.mp4
>>
>>109545013
oh yeah? do picrel.
>>
File: MiniMax_H3_00528_R.mp4 (3.91 MB, 832x1504)
3.91 MB
3.91 MB MP4
>>
new comfykitchen is slightly faster, 206 vs 220s yesterday at the same 0.4mp size

pokemon ep about ltx 2.5:

https://files.catbox.moe/91vhjq.mp4
>>
Has anyone managed to get rain working? I'm assuming I'm getting fucked by my resolution limitations atm (0.3MP) :(
>>
File: h3_00070_.mp4.webm (1.59 MB, 896x1184)
1.59 MB
1.59 MB WEBM
>>109545060
maybe? I dunno, you decide
>>
>>109544952
catbox them when you do, please
>>
>>109545045
>new comfykitchen
Is this just the new ComfyUI nightly
>>
>>109545111
I did update in the folder and it pulled a ck update
>>
i hate the poopschizo vs pissbuttfag shit but that OP video is pretty good
>>
kek, if you amplify "cute" in ltx, it makes the girls have cat ears
>>
File: 1761981437526228.mp4 (3.84 MB, 2048x590)
3.84 MB
3.84 MB MP4
>>109545060
kitchen and spectrum on Left (my 1st warm run) and sol and sage on Right, kitchen was faster and apart from the random Sadako Yamamura hair with kitchen the rain is fine in both???
>>
there we go, i2v:

the red hair anime girl, Misty from the anime Pokemon walks back and forth in frustration as the yellow pokemon Psyduck stands in place. Misty says "oh man this LTX model is fucking awful, I hate it!".

suddenly, the yellow psyduck pokemon with the same expression walks to a computer with a white CRT monitor, with the text "Minimax H3" on the screen on a comfyui interface, and starts typing rapidly, and an anime girl with very large breasts wearing a black bra and panties shows up on the screen.

https://files.catbox.moe/8c3jvu.mp4
>>
>>109545146
where's the rain?
>>
>>109545146
Buy some glasses, left is obviously worse
>>
>>109545177
explain how the rain is worse in the left one.
>>
File: MiniMax_H3_00530.mp4 (2.06 MB, 1280x960)
2.06 MB
2.06 MB MP4
>>
>>109545181
I RECOGNIZE THOSE EYES AND PAIR OF TITS!
>>
>>109545159
The same reason why they used to pour simulated drenching rain in movie shots. Light rain in low res doesn't show up.
>>
>>109545181
prompt for this style? muh nostalgia
>>
File: coomgenner psyduck.png (510 KB, 640x640)
510 KB PNG
>>109545147
god he's literally me
>>
>>109545201
it's i2v
>>
Have any of you guys been able to get H3 to do anime to realistic? I've been trying with the fl2va model and the anime image as a ref but the results are inconsistent.
>>
>>109545280
it can do like cosplay but the hair always looks like a singlepiece plastic wig
>>
>>109545280
>inconsistent
Yeah that's the story with trying to radically alter a video based on a ref, sometimes it'll work really well and other times it will do nothing or behave in really unexpected ways. I haven't gotten a prompt that consistently performs the translation correctly
>>
can someone bring catjack back to the asylum? He's schizoing out in /adt/
>>
File: vid_00189_.jpg (758 KB, 1384x1816)
758 KB JPG
>>109545181
That image is atleast 70 years old
>>
>>109545332
out of 10!
>>
>>109545332
it's quite literally older than this board innit
>>
>update comfy
>cache extension starts throwing errors during generation
>it's not even in the workflow
>>
>>109545280
why don't u use image edit, then take the output to make the video
>>
>>109545179
Look at the face autist, not the fucking rain
>>
Anyone tried Comfy Kitchen with 4 or 8 step turbo lora ?
>>
>comfy kitchen is another layer of shitty memory management along with dynamic vcuck
>>
https://old.reddit.com/r/StableDiffusion/comments/1vn9duw/looks_like_we_might_be_getting_minimax_music_3/

minimax music soon
>>
how to gen h3 blowjobs where she isnt chewing on the damn thing?
>>
>>109545388
even with the loras it sounds like it.
>>
>>109545388
I can do inflating, would you like inflating
>>
>>109544926
I've had some successes and some failures, weird lips because the photo I used had those covered perhaps. The solution is more references
>>
>>109545388
Dont use explicit terms. I learned that from the cucked grok days
>>
File: MiniMax_H3_00532_R.mp4 (3.96 MB, 1024x1214)
3.96 MB
3.96 MB MP4
>>
>>109545376
We're talking about the rain you fucking spastic, you walked into a conversation about rain then spazzed out about something completely irrelevent, fuck off autismo.
>>
File: 4557845437894.png (18 KB, 900x806)
18 KB PNG
where is kino?
>>
>>109545388
I like to use "bobbing" for a lot of my nsfw prompts, whether it be head bobbing or body bobbing. No idea what the prompt for the sound though.
>>
>>109545366
Which one is best? I tried klein 9b before but it's kinda ass (and slow).
>>
>>109545419
>random Sadako Yamamura hair
Nah retard, you weren't talking about the rain
>>
>>109545438
>No idea what the prompt for the sound though
>overall soundscape: clattering bones
>>
File: orodSh_00201_.jpg (1.35 MB, 1776x2560)
1.35 MB JPG
>>109545343
>>109545346
Keyra Agustina genetic memory unlocked
>>
File: 1760644797795336.png (82 KB, 1228x502)
82 KB PNG
You boughted, right?
>>
>>109545455
>Only one topic can be talked about, and it's interchangable to match what autism level i'm currently at, I must win even when im wrong aaaaahh!
The topic was rain, i posted rain and a person, you sperged on the person ignoring the rain side of it so you could slide into an autistic fit about anything.
>>
>>109545489
hory shiet why did you put my random post on your gen hahaha
>>
>>109545492
why would i need that?
>>
>>109545510
Probably fed to llm, can't remember
>>
>>109545513
Because the more you buy, the more you save!
>>
>>109545405
prompt?
>>
>>109545495
Take your L schizo
>>
File: _122247422_pills.jpg (58 KB, 976x549)
58 KB JPG
>>109545492
Model progress plateaus leading to less demand on compute.
Capabilities continue to increase with compute, unlocking new uses, pushing prices even higher in the short term.
>>
>>109545588
"You will own nothing and be happy"
>>
File: VOLCANO.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
Wish I could post in higher res or catbox was working
>>
Ideogram 4.0 completely flew under the radar for me, why's no one using it? Noticed it's completely dead on civitai too.
>>
>debo making vids about himself
>>
File: 785567.webm (2.93 MB, 420x224)
2.93 MB
2.93 MB WEBM
kino alert
>>
>>109545664
Your resolution bro?
>>
>>109545588
I'll take the yellow red pill. China flooding the market with cheap GPU and RAM.
>>
>>109545664
ants are loving ltx
>>
File: 634643.webm (3.35 MB, 420x224)
3.35 MB
3.35 MB WEBM
>>109545684
based kinoisseur ants enjoying the full bitrate
>>
>this facerefiner
If it works this fixes my most significant problem with minimax, but I'm a moron idiot and can't tell if it's actually safe to run since anon was shitposting. any non meme takes? Is it just impact pack?
>>
>>109545692
the colony is pleased
>>
>>109545664
>>109545692
I was checking out your vids on the shitter last night on my phone and realized that's probably another good reason to make sure models can do low resolutions really well. It fit perfectly on that screen and looked surprisingly watchable.
Y'know till i went into landscape mode kek.
>>
>>109545386
https://minimax-ai.github.io/music3-demo/
>>
>>109545732
sweet can't wait to remake Oneyplays' hit Russian folk song "I STINK" in the style of 80's new wave.
>>
File: Krea2_turbo_hr_fix_00043_.jpg (3.4 MB, 2368x3544)
3.4 MB JPG
>>109545732
cool
>>
>>109545638
I for one am enjoying these. I have also nooticed no complainers ever gen anything. If we need a thread mandated schizo this guy is better than the others.
>>
>>109545762
I find it odd that they complain and can't produce anything not even a witty reply using a gen. Bantz is a lost art
>>
>>109545732
Finally the acestep retard with no ears can shut up about ace step.
>>
really frustrating when the model keep adding animation or dialog when you don't want it to
>>
>>109545785
It might be too large for his rig so he'll seethe about this model too
>>
https://files.catbox.moe/ye69xn.mp4
>>
how obsessed and mentally ill do you have to be to make the same avatarfag videos. it's just a waste of vram just to samefag yourself
>>
>>109545829
This is art and diffusion, how about you be an example and make something that gets anons talking.
I'm waiting
>>
File: MiniMax_H3_00536_R.mp4 (3.57 MB, 832x1504)
3.57 MB
3.57 MB MP4
>>
>>109545829
I am also waiting for you to post something based.
>>
>>109545850
I did. I posted 1girl
>>
File: Ryusei No-Audio.mp4 (3.83 MB, 752x420)
3.83 MB
3.83 MB MP4
Any recommendations for an image editing model that can handle anime style gens well?

I want to create a starting frame for a follow up Minimax H3 generation with those same two guys. Tried FLUX2 and that's not working all too well. Either I'm doing something completely wrong or that's not something the model handles too well. Don't really want to gamble on the Ref2VA model getting the initial scene right.
>>
>>109545708
>I was checking out your vids on the shitter last night on my phone
careful, you might end up staying too long and getting a hemorrhoid
>>
After some light testing I have decided that comfy kitchen is faster than sage but might be more braindead.

With the same seed:

>restrained character's ankles are freed for no reason
>lines of dialogue repeated twice
>closeups on entirely incorrect charaacters
>>
>>109545846
acceptable honkers, I like this pov camera. any weird prompting or is this a ref
>>
>>109545781
Putting effort into anything is a big no no in zoomer dogma
>>
File: Screenshot 21.png (69 KB, 3029x232)
69 KB PNG
rev2v TURBO RELEASED
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
>>
>109545829
>LEAVE THE AVATARFAGS ALONE!!
lol
>>
>>109545931
Aaalright time to see if this fucking sucks.
>>
File: MiniMax_H3_00032__.mp4 (1.6 MB, 968x556)
1.6 MB
1.6 MB MP4
>>109545732
my twink ass preparing for image to video and 2k upscale
>>
https://files.catbox.moe/4mnzyz.mp4
man I wish i could gen in higher res, but i hate waiting
>>
science!

the breasts of the asian girl go from large to a flat chest, shrinking slowly over time, as she sits and says nothing.

https://files.catbox.moe/11xlnq.mp4
>>
>>109545994
and science, but the other way:

the breasts of the asian girl go from large to extremely large, growing slowly over time, as she sits and says nothing.

need cartoon sound effects

https://files.catbox.moe/m9xz3y.mp4
>>
>>109545994
This is sick
>>
>>109546010
This is healthy
>>
>>109545759
What lora are you using?
>>
>>109545977
ZAMN
>>
>>109545146
left is insanely bad wtf are you smoking?
>>
File: 1762013983539800.jpg (3.77 MB, 3552x3169)
3.77 MB JPG
are you postprocessing son
>>
>>109546010
This is the kind of audio that makes me exclusively want to use this model for vaguely body horror or sci fi prompts. These sound effects are so ludicrous in a realistic scene (well breast expansion isn't realistic but you know what I mean)
>>
File: MiniMax_H3_00538_R.mp4 (3.48 MB, 1024x1214)
3.48 MB
3.48 MB MP4
>>
>>109545664
>>109545692
I love your kinos. can H3 really not make those gens?
>>
>>109546055
model does it for me
>>
>>109546010
one more variation!

the breasts of the asian girl go from large to extremely large, growing slowly over time making her red shirt to burst open from the stretching, revealing a white bra over her breasts, as she sits and says nothing.

not quite, I think 5s isnt enough time may need 10, but still a good output.

https://files.catbox.moe/d9k5o8.mp4
>>
>>109546086
*I also just realized I fucked up with bad engrish (making her shirt to bust open)

anyways. more gens now
>>
How would you describe AI image/video generation to a medieval peasant? Challenge: don't make it sound like witchcraft
>>
>>109546086
these are gens that Wan or LTX can do, static one girl with not much happening. I know that the point is not to simply dictate gens knowing that Wan or LTX could do them, but it's not exactly pushing the abilities of H3 with gens such as these
>>
>>109546032
the kroma one, I think it's better than whatever the fuck that finetune he's making is.
I don't understand that guy he does random shit creates garbage and his fans eat it up. He always almost does something good and it annoys the fuck out of me
>>
>>109546065
Holy fucking titties
>>
>>109545732
First music model I find sounds good. I suspect they actually had a proper human write the lyrics for the examples.
>>
>>109546106
like looking at a reflective pool and seeing whatever you want
>>
>>109545923
I have turbo fatigue
>>
>>109546106
They'd have trouble understanding portable cold light from a phone screen.
>>
>>109546091
dont waste your beer:

https://files.catbox.moe/vn6bej.mp4
>>
>>109545937
I just tried with 4 steps, simple+res_multi, and it broke the gen almost completely. Did you have any luck
>>
>>109546055
>uprez = waste of resources
What? what's the point then?
>>
>>109546055
you could just sharpen and add film grain to get the same result. no need to do all the 1x "upscaling" snake oil crap
>>
File: 1758054208580609.png (336 KB, 2560x2204)
336 KB PNG
>>109546135
it cleans up the image nicely
>>
>>109545977
holy shit
>>
>>109545923
why would rev2v need its own turbo
>>
I have a little scheme where I want to take a reference and use Krea to make a very high resolution image, like 20k by 10k. Is that even possible?
>>
>>109546176
helsie my wife
>>
>>109546176
I guess I see what you mean. but why wouldn't you upscale? as far as I can tell 2x upscale is virtually identical to the base res just bigger. I don't really see what resources you would be "wasting"
>>
File: 74433.webm (1.82 MB, 960x512)
1.82 MB
1.82 MB WEBM
>>109546073
>can H3 really not make those gens?
i'm not sure, i can try it out later since i will go to sleep now. h3 does have a bias towards modern style videos so it might work pretty good
>>
>>109546190
I think it's just cope from people using turbo thinking it was bad because they were using ref.
(It's bad because it's turbo)
>>
If the minimax music model's ability to use refs is as good as H3's we're eating really fucking good.
>>
Debo won.
>>
>>109546132
So, i first tested it using my last gen settings, and it looked about what i expected for a lightning lora
Second test i boosted steps to 8 and res to 1.0mp. It's.. actually passable. I figure it's mostly the extra step count that helped.
Pretty sure you're meant to use a minimum of 8 steps with res_multistep, i just use Euler simple. Gonna test with a better prompt and ref combination and post that.
>>
>>109546224
>prompt: make a song about shooting blackrock execs in the style of free bird
>>
>>109546224
They're doing I2A, where you input an image and it turns it into a song
>>
>>109546236
Yeah I'm just a huge faggot and wasn't using euler+simple. It does at least work but the quality kinda sucks, not sure if that's just the reference model being worse though
>>
File: ocee autism.png (565 KB, 693x462)
565 KB PNG
>>109546271
probably a combination of both, i'm usually very critical of these early lightning loras but from that one test it actually performed better than expected. Bog test is still cooking but i was looking at frames from the last test, i'm baffled H3 is so good it can differentiate two outfits described the same but look different in their reference images, even with a lightning lora raping it. It definitely doesn't kill coherency/smartitude that hard.
>>
Trying something.
>>
File: MiniMaxH3_00055.png (504 KB, 864x864)
504 KB PNG
Okay yeah this lightning lora fucks.

https://n.uguu.se/lmLVOwse.mp4
>>
>>109546055
rtx x2, add sharp + grain, scale 0.5
>>
>>109546318
FUCK i forgot to connect the bog audio reference, my bad.
>>
File: Krea2_turbo_hr_fix_00058_.jpg (3.25 MB, 2368x3544)
3.25 MB JPG
>no clapback gen
I rest my case
>>
>>109546318
which one though? light or the other one
>>
>>109545923
Nice. The peasants should be able to enjoy the wonders of H3 too, even if it'll just be slow motion crappy gens. Enjoy your turbo shit lora.
>>
Idea:
using minimax like an LLM.
Input a question and the video is the answer.
>>
File: 0000267716.jpg (113 KB, 1094x724)
113 KB JPG
>>109546055
Do you even post process.
>>
oh shit

https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

theres a ref2v lora now.
>>
>>109546351
holy raped.
>>
>>109546351
hollywood had ERNIE in 2010?
>>
Is lightx2v supposed to be better than
>minimax_h3_turbo_v4_step600_ema.safetensors
Because this shit has been amazing for upscaling.
>>
>>109546318
kek. this is why I didn't bother updating comfy kitchen.
>>
File: 1771627213261209.mp4 (1.34 MB, 864x480)
1.34 MB
1.34 MB MP4
>>109545019
NTA
>>>/wsg/6213587
>>
>>109545732
I hope it can do instrumental stuff without lyrics
>>
>>109546318
congratulations anon you bogged yourself
>>
>>109546386
RIP
>>
>>109546386
that's fuckin nice, good job.

>>109546388
you cursed me motherfucker, i threw in the reference audio and now he gibberishes for most of the video.
so there IS a cap on how long the ref audio should be, my bad. my retardation, i'll DO OVER.
>>
>>109546249
so you can't use audio references? that's gay
>>
>>109546249
>I2A
lol, what’s the use case?
>>
>>109546355
however, the latest 8 step turbo may work better so try both. you can use comfykitchen node, but bypass spectrum if you use it.

https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
>>
>doesn’t read the thread
>>
so is the music model actually out? theres model files on the huggingface page the demo page links to
>>
what?
>>
Huh?
>>
>>109545743
>>109545759
>>109545785
>>109545976
>>109546114
>>109546387
it's up
https://huggingface.co/MiniMaxAI/MiniMax-Music3
>>
elegg my beloved:

https://files.catbox.moe/h3bpxw.mp4
>>
>>109545799
> The full precision fits under 24GB of VRAM. With automatic CPU offloading, generation takes in ~22 GB; additionally streaming the language model layer by layer makes it fit even 8 GB video cards:
>>
>>109546386
lel
>>
>>109544588
Based OP
>>
Can someone post their kjnode settings for minimax when doing the two pass gen?
>>
>>109546519
No.
>>
File: MiniMaxH3_00057.png (749 KB, 864x864)
749 KB PNG
>>109546475
>>109546497
>even on 8gb video cards
damn these chinks are fucking KILLIN' it. Vramlets are gonna have a field day.
anyway speaking of vramlets here's bog with proper length source audio.
The lora seems legit enough, but i don't think i'd use this on high motion/serious productions, it probably kills motion pretty hard either way.
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

https://d.uguu.se/iBuPYjke.mp4
>>
anyone else have trouble downloading large models off of huggingface?
>>
>>109546534
no with aryan2c
>>
File: file.png (5 KB, 193x107)
5 KB PNG
Should I update to is everything gonna break?
>>
well damn, the 1.0 8 step lightx2v lora actually works well. and this was only 0.4mp and 8 steps.

https://files.catbox.moe/bcmzfg.mp4
>>
>>109546550
the cloud picture looks threatening
>>
>>109546550
Even with Stability Matrix I make codex do it. It updates the nodes and dependencies when needed.
>>
Ok we're back to throwing everything at the kitchen sink to replace a woman in a video. Anyone find a prompt that works? I'm combing the damn guide but can't get it to do what I want - just replace the woman, that's it. It leaves the source video basically unchanged.
>>
>>109546560
no fucked audio sounds appealing to me.
>>
sounds illegal im not going to help you
>>
>>109546055
topaz is better and faster than all this bullshit
>>
>>109546560
even the noise is much better, I should do an a/b comparison vid but the quality is better than before, with 20 steps + spectrum. now it's just model -> lora -> comfy kitchen -> sigma shift node

https://files.catbox.moe/3jumdl.mp4
>>
File: t.png (123 KB, 1469x612)
123 KB PNG
>>
>>109546607
oops, shift audio is at 3 and apparently it should be 6 for turbo? ill test more.
>>
File: OUTPUT_00024.mp4 (1.64 MB, 768x1376)
1.64 MB
1.64 MB MP4
Prompt:
>What is the meaning of life?
>>
>>109546604
>0.099s
>0.048s
>0.046s
>>
>>109546627
>the meaning of life is1girl
>>
>>109546303
Did this create mustard gas?
>>
>>109546604
doesn't topaz deepfry your shit?
>>
>>109546641
>>>/wsg/6213591
It did create something yellow.
>>
>>109546642
depends on the model. the starlight models take forever and deep fry already high res footage because they're meant for old low res video.
>>
>>109546618
success, sigma 6 is clean with turbo.

https://files.catbox.moe/gmhyw2.mp4
>>
did another test with an image that wasn't 400x400 pixels, same settings cept 0.7mp. Looked about as good as my non lightning gens, besides maybe a tiny bit more noise around moving parts of the video. It's very promising.
>>
https://files.catbox.moe/bkt3mo.mp4
You can really just use whatever character you want now, even with just one image. It's insane.
>>
>>109546604
thanks for popping into /ldg/ Mr Shecklesberg
>>
>Minimax actually released their SOTA music model
https://huggingface.co/Comfy-Org/MiniMax-Music-3
https://huggingface.co/MiniMaxAI/MiniMax-Music3

Base Chinks, I kneel. I listened to their demos, it's a pretty strong base and step up over ACEStep XL. I'll try it on Comfy first, but from my experience with ACEStep Comfy is just shit for audio (guess I'll have to wait for .cpp implementation of Minimax). Now to generate kinos
>>
>>109546681
thanks for chiming in poorfag. what, you can't do $60 a month?
>>
ALL HAIL MINIMAX
PRAISE CHINA
>>
>>109546690
>paying to upscale
lol
lmao even
>>
>>109546674
Why didn't you use the same prompt as the last ones? sigma 6 could be raping the fine audio details and now we can't tell.
>>
>>109546627
>>109546640
Based and monogamypilled
>>
>>109546705
im testing a bunch of stuff now, guitar by itself is clear, vocals and no guitar is clear, apparently 12/6 is the turbo default and 12/3 for no turbo. now im gonna try more elaborate prompts.
>>
>>109546674
TY for testing and sharing workflow too
>>
>>109546682
time to generate free bird, but except free bird it is "kneel floyd" and a song about stopping fentanyl abusers.
>>
>>109546718
From losing my mind last night on a prompt. maybe you can corroborate my findings but when you add spoken language to your prompt it seems to rape all other audio in favor of the voice.
>>
>>109546530
>https://d.uguu.se/iBuPYjke.mp4
kek, good one anon
>>
File: Begone vramlets.png (97 KB, 1139x484)
97 KB PNG
>>109546690
Go be poor somewhere else
>>
my biggest gripe of music models has been many have intentionally ruined gens if you try to remix or recolor a song (copy the rhythm).

I wonder if the minimax model can remix a song or make it in the same style.
>>
>>109546751
minimax doesn't seem to give a fuck about copyright so I believe.
>>
File: 1767578373770424.png (415 KB, 800x582)
415 KB PNG
>>109546758
once again China wins
>>
>>109546749
absolutely BODIED that KIKE
>>
>>109546749
You will always be poor to me if you run windows.
>>
>>109546749
While I admire the gear I feel like I would be eternally seething with the overall vram. I cannot understand why nvidia cheaped out and cut it at that size.
I had my finger on buy and missed my chance but I feel like I would be seething even more than my 32gb of vram.

It keeps me up at night desu
>>
>>109546674
the turbo lora uses 12/3 or 6/3 shift depending on what turbo lora you use. In your case, 12/3 at 4-8 steps.

https://github.com/ModelTC/Minimax-H3-Turbo/blob/main/README.md
>>
>>109546749
>haha, poor!
>96gm ddr5
pfffffffff
>>
these sigma settings are just suggestions btw.
>>
I figure the model already sets its own sigma(balls) based on the step count and sampler right? This sounds like more tinkerfagging that leads to placebo.
>>
>>109546749
>2 TB NVMes
>not 4 TB
>>
>>109546749
Honest question, what could this run that a rammaxxing poorfag couldn't?
>>
>>109546803
it doesn't. it just defaults to 12/3 if you don't specify
>>
>>109546809
He's basically the tallest midget because Nvidia kneecapped him from the magical 128gb
>>
still turbo, it worked:

integrated_multimodal_description:
[0s-10s] Medium shot of the anime girl sitting on a stool, passionately playing a mahogany acoustic guitar. Her fingers cleanly change chords on the fretboard as she strums. She looks directly into the microphone and sings clearly.
<d> "This is the story of my life, echoing through the night." </d>

overall_soundscape:
The clean, bright acoustic guitar strums loudly in a rhythmic pattern, perfectly timed with the finger movements. The female vocals are crisp, dry, and centered, with no background echo.

non_diegetic_music: N/A

https://files.catbox.moe/on95ew.mp4
>>
>>109546682
I'm downloading the model right now, and right off the bat it's obviously superior music quality to ACEStep, but looking at their demo
https://www.minimax.io/blog/minimax-music-3-0-next-generation-open-weights-production-ready-versatile-music-model

>Progressive house / EDM, 126 BPM, B-flat major.
>side-chained synths

But what the model outputted is basically slop that sounds like generic pop, which is concerning.
>>
>>109546782
hmm ill retest with the diff settings, was using 12/6.
>>109546841
this time with 12/3 shift: seems better, also holy shit what an improvement from the 0.1 lora version.

https://files.catbox.moe/5izt01.mp4
>>
>>109546682
>.cpp
This? https://github.com/0xShug0/audio.cpp
>>
File: output_4mb_no_audio.mp4 (3.96 MB, 1366x2048)
3.96 MB
3.96 MB MP4
So the ref model and ref lora at 6 steps
>>
>>109546809

Video generation and anything else that requires everything on a single card.
You can technically poormaxx just fine by buying a bunch of 5060 Tis to reach a ton of VRAM and it'll work perfectly well on LLMs, but it won't cut it without proper parallelism support in video or image generation.
I'm sure that a functional multigpu support all across AI products is only a matter of time though, as running multiple cards is basically a standard by now.
>>
File: ironmang1.mp4 (2.08 MB, 960x640)
2.08 MB
2.08 MB MP4
>>
>>109546868
Nah,
https://github.com/ServeurpersoCom/acestep.cpp

It's superior to Comfy's ACEStep implementation, more lightweight and faster.
>>
File: output_4mb_no_audio2.mp4 (3.91 MB, 1366x2048)
3.91 MB
3.91 MB MP4
>>109546894
10 steps with the turbo 8 lora not for this model
>>
>>109546894
Prompt? Couldn't never get drool from H3
>>
kek, i2v surprisingly good at music, still turbo 8 steps at 12/3 shift:

https://files.catbox.moe/681p1y.mp4
>>
>>109546922
So much worse lmao
>>
>h3 music
i dont care. just give us the image model
>>
>>109546940
I agree
>>109546935
Their tongues slide and intertwine with fluid, organic motion, and glistening strands of saliva stretch and snap between their lips.
>>
>>109546682
>>109546854

It's garbage. I prompted for 1980s synthpop and it output modern hip-hop garbage instead.
Using the official Comfyui workflow.
>>
>>109546954
stop being a promptlet
>>
>>109546954
Have you tried the prompting guide and using a llm?
>>
>>109546965
I prompted using both "official" method with an LLM and a simple phrase, both output garbage
>>
>>109546975
I don't believe you because you're not showing work
>>
>>109546936
I like this gen, same prompt. catchy bocchi

https://files.catbox.moe/vk0u6a.mp4
>>
>>109546954
I wonder if they're giving us a distillation rather than what's hosted on their API https://www.minimax.io/audio

Or maybe their API isn't that good. Unfortunate, ACEStep XL despite its worse audio quality is very good with genres due to its 4B DiT, and it's also very creative due to its arch.
>>
>>109546985
anon have you been genning the same video the entire week?
>>
>>109547006
...or maybe the comfy implementation fucked up somewhere
>>
>>109547011
no sir, I am currently testing the new turbo lora 8 steps with assorted images. otherwise id make a bocchi concert.
>>
>>109547013
I think they worked closely with comfy for the release
>>
>>109545676
Not going to happen for a few years if not a decade.
They have enormous domestic demand.
>>
>>109546985
last bocchi, tried 0.6mp from 0.4. getting the speed of spectrum gens at 0.4, pretty neat

https://files.catbox.moe/l23l52.mp4
>>
man I would be genning right now but I'm busy. I wanna fuck with that minimax music model
>>
>>109547036
You are measuring China by applying the same metrics as western companies, anon. Rookie mistake.
>>
File: Krea2_turbo_00045_.png (1.36 MB, 1368x768)
1.36 MB PNG
>>109545879
Krea 2 turbo is you best bet.
>>
https://files.catbox.moe/8qeg3g.png
https://files.catbox.moe/eajquo.png
>>
>>109547099
that dick looks retarded
>>
>>109547075
They are just very good engineers, not wizards.
Like everyone else they're constrained by lithography tools and it just takes time to make those.
Plus their datacenter buildout has to accelerate to catch up to the US.
>>
>>109547105
>m-my 4 inch peckerwrecker is more realistic!
cope
>>
>>109547099
this looks just like real life tbhdesu
>>
>>109547099
insanely slopped
>>
>>109547115
nothing to do with size, it looks like someone stretched it in photoshop then diffused it
>>
>>109547099
AI made me appreciate my average sized peepee
>>
>>109546969
Link to the prompting guide?
>>
>>109547119
https://files.catbox.moe/2qsvla.png
I did super low cfg so it gave me some weird stuff like TV on the couch
>>109547116
putting instagram filters on your gens makes it look more real
>>
>>109547137
>filters on your gens makes it look more real
yeah like real shit
>>
>>109547134
I'm still working so I can only shit out miku kissing gens but it should be in the repo like the video models
>>
>>109547137
Bro, what are you doing?
>>
https://docs.comfy.org/tutorials/audio/minimax/minimax-music-3
>>
lowk not realistic enough needs about 15.2 on the realism slider
>>
Trying to better understand the system's logic behind doing video editing with references. It really does seem to be a ton of prompting.

I'm trying to swap animated character into a live scene, and just asking the prompt to "replace this person with Subject 1" didn't work. Asking the prompt to "replace this person with a version of Subject 1 rendered as a photorealistic character" did.

Starting to think what's possible with ref2v is going to be incredibly gated by how you figure out prompting
>>
>>109547211
my detail slider lora only goes up to 10
>>
>>109547216
add another one dude
this is why comfy ui is the best software
>>
Big dogs aka 32gb chads
What model stack should we target with this music model?
>>
>>109547159
posting gens and triggering the libs (you)
>>
>>109547207
Seriously, I have tried everything and it sounds like this shit was only trained on royalty-free slop (which in turn is mostly rap, hip-hop and dubstep garbage)

AceStep XL is still superior by a lot.
>>
>>109546975
Crazy how you don't reply
WHAT ARE YOUR SPECS, WHAT ARE YOU DOING?
IF YOU DO NOT ANSWER YOU ARE A FRAUD
>>
>>109547242
You can't even show a sample or provide anything for us to work with.
I think /sdg/ is more your speed
>>
>>109547215
it isnt so bad, just add a picture (first input is picture 1) then set it up:

Use <Picture 1> for the physical identity of Minimaxguy (use any name) with the voice of <Audio 1>

then just prompt as usual, Minimaxguy is a character you can prompt like any other character with the appearance of that image. you dont even need the audio 1 part if no voice clone.
>>
>>109547270
*my bad you are doing video edits

yeah, video can be tricky, it's pretty specific: llms/google ai mode/etc can help for prompting, cause it can be picky
>>
>>109547246
3090
official comfy workflow
used the "official" prompting method, tried prompting both with an LLM and simple sentences.
Try it yourself.
Or just accept it, it's garbage. AceStep XL is still the local king.
>>
>>109547285
So you're a little dog and can't even provide basic detail such as the LLM model
Thanks for playing
>>
>>109547285
kys
>>
>>109547282
Yeah it's kind of a fun problem to solve though. I just want my DVa, I'm a very simple man.
>>
>>109547099
>>109547137
Ultraslop.
Just looks off.
>>
AVGN test, actually functional at 4 steps/15s/0.4mp, this is with the new turbo ref lora:

https://files.catbox.moe/za1aup.mp4

at 8 steps:

https://files.catbox.moe/f5hveo.mp4
>>
My next Debo gen is almost ready for next OP. I'll bake as soon as it's finished.
>>
>>109547340
you really gotta prompt your shit better, it's uncanny hearing AVGN so low energy and mushmouthed trying to push out so many lines in that timeframe with no energy.
H3 lets you section the prompt by timestamps, 0s-3s,3s-5s, etc. Try that and add emotion to his lines, that should help.
>>
>>109547355
You never posted one because I'm OP.
Don't get bitter because people are enjoying H3
>>
>>109547361
I think we should just take James soul and put it into the model itself.
>>
>>109547340
is it me or 4 steps looks better?
>>
>>109547374
agreed
>>
>>109547340
also what sampler/scheduler?
>>
>>109547085
Thanks, but that's already the model I was using for the images. Super hard to get the same looking vehicle and characters with just a lucky seed.
>>
File: 1779312112189679.jpg (104 KB, 480x360)
104 KB JPG
Bless minimax for giving us h3 but they really picked the worst time to do it. I'm dying, bros.
>>
I guess us 32gb chads can't actually test this model because we're all busy working
>>
>>109547013
I hope
Progressive EDM using Comfy's prompt format
https://files.catbox.moe/ndrgrk.mp3

UK Dubstep (just gave me hip hop)
https://files.catbox.moe/xhl7mm.mp3

... A shit DiT could theoretically get genres confused, but the LM is 8B so I expect much better results.

This 1980s rock result is much better than the other two
https://files.catbox.moe/3fc6fq.mp3
>>
>>109547291
GPT5.6 on chatgpt. Why does it matter?
>>109547293
Test it yourself and report back.

https://voca.ro/1oSe3ppqrswK
https://voca.ro/1RsTqdJEoR8c

Even when you manage to evade the hip-hop shit, it's meh
>>
>>109547366
Stop pretending to be me, Debo. Get a job.
>>
>>109547430
>I'm dying, bros.
in what sense
>>
>>109547398
default, simple/res multistep
>>
>>109547430
no air conditioning?
>>
>>109547285
Yea, so far based on my results ACEStep is significantly better. These results are barely ACEStep 1.5 tier in genre recognition at best. ACEStep 1.5 is at XL now, significantly closer to Udio/Suno.
>>
File: 1771686725746295.png (66 KB, 890x330)
66 KB PNG
What could Ostris' new Blackwell RTX PRO 6000, gifted to him by comfy himself, possibly be so busy with that it takes priority over a just released SOTA video model?
>>
>>109547509
>>109547547
Yeah. AceStep XL Turbo is still the local sota. You can effortlessly get some bangers out of it, unfortunately it has the lyric-skipping issues. The "quality issues" people report is not a big deal for me
>>
Eagerly awaiting anon's Namine lora
>>
File: ComfyUI_03131_.png (1.08 MB, 1024x1024)
1.08 MB PNG
finally, the Bogsylvania kino i couldn't produce on ltx 2.3.
(original image not mine, credit to an anon)
https://n.uguu.se/FYtlXJAj.mp4
>>
>>109547565
you accidentally created Fred Armisen
>>
>>109547513
1980s Jap city pop, more generic pop blend slop
https://files.catbox.moe/udubru.mp3
>>
>>109547430
this ain't the season for a 350w space heater running on max for 20 minute stretches
I haven't genned all day but I've drafted some prompts, should be cooler tomorrow
>>
okay the ref turbo lora is good, if you find it blurry use 8 steps.

https://files.catbox.moe/kan7cm.mp4
>>
>>109547590
huge improvement over your last runs, nice.
>>
>>109547576
Like I said, it looks like they trained on "royalty free" or public domain tracks likely out of fear of getting sued by the labels like Disney did with them for the video models.
They should have at least trained on Suno syntheticslop since that at least would provide some better prompt adherence for styles.
>>
>>109547576
>1980s Jap city pop, more generic pop blend slop
that's basically what that "music" is tho
>>
>>109547547
>>109547554
>>109547601
Any of you retards willing to share what version of the models you are using or are you faggots just spreading fud again
There are multiple sizes for both the text encoder and the model itself
>>
>>109547602
Negative, japanese 80s citypop has 70s funk + disco + soul influences mixed with 80s synth

>>109547618
Can you just shut the fuck up and test it yourself with the official comfy workflow? Because I am using exactly that and the same models listed there.
>>
time for a bit of ref2v silliness.

[Identity Mapping]
Forsen is represented by <Picture 1>

the setting is the forest in <Picture 2>. Include the cart in <Picture 2>. Forsen is sitting in the cart to the right of the blonde nordic man in <Picture 2>. Forsen says "so, where are we headed mister Nord?". The blonde nordic man says "oh shit!" and the cart rolls off a mountain cliff, down into a large pit, and explodes.

https://files.catbox.moe/3j2ulm.mp4
>>
>>109547554
ACEStep 1.5 will skip lyrics too though (unless you're using SFT, but SFT itself is not a musically impressive model). This has significantly better audio quality which captured way more details during its training unfortunately (due to 32 kHz training), so it's sad that it's so slopped. Maybe it can be better with a LoRA.

ACEStep XL when using the merged model with a LoRA is currently SOTA (audio quality issues won't be as bad as just using raw model).
>>
>>109547631
Nah, I think you're a promptlet and not bright, the two music fags in chain of generals are both mentally slow jackasses that have no actual skill.
You failed the basic assessment of testing and you're not giving me confidence.
>>
>>109547618
I'm using the fp16 DiT, plus int8 convrot text encoder (which should be a negligible hit to its quality if it's one at all).
>>
>>109547631
so let me get this straight, you prompted for a specific niche genre instead of detailing the actual styles? I have no doubt people will create Loras for these trained on whatever they like or what people might want, but to think models should include every single niche genre is laughable.
>>
>>109547657
Thank you that's all the info I wanted
>>
>>109547642
>ACEStep XL when using the merged model with a LoRA is currently SOTA (audio quality issues won't be as bad as just using raw model).
May I ask you, which merge is the best one?
>>
>>109547641
prompt change:
>The blonde nordic man says "to the moon!" and the cart flies high into the sky, entering orbit above the earth.

if you leave too much empty air without timestamps they do jibberish. but, to space!

https://files.catbox.moe/cy0ejw.mp4
>>
>>109547659
See
>>109547646
We have 2 "hardcore" music heads and they are based in /sdg/ but love to post here and shit it up
>>
>>109545732
could you post a prompt example?
>>
File: ComfyUI_00008_.mp4 (299 KB, 480x832)
299 KB
299 KB MP4
I tried generating a video and I have no idea what I'm doing
>>
>>109547601
Minimax H3 is significantly better at recognition from short music loops

https://files.catbox.moe/ug2d9x.mp4
https://files.catbox.moe/ddjlxi.mp4

So this is specially sad
>>
Wouldn't you test with common genres first?
>>
oh good, ref turbo lora (lightx2v, new one today) at 8 steps works fine with my simpsons meme test.

https://files.catbox.moe/ont77u.mp4
>>
>>109547712
based ltx enjoyer
>>
>>109547509
>Progressive EDM
I'm of the opinion that EDM should not have any lyrics besides small little one-shots
>This 1980s rock
This is absolutely not 80s rock lol.
Sounds like nickelback
also
>Neon Light
Every. god. damn. time.
>>
It's very very strange seeing lightx2v release a lightning lora that isn't complete trash.
Starting to think maybe.. It was never his fault, it was just the trash models we had before H3..

Nah lol his loras were always awful, this is the first time i think he's actually managed to accomplish what he set out to do. Not sure if toying with the strength of the lora does anything, i'm trying to test that now. But shaving off a whole 12 out of 20 minutes of my usual gen times is really nice.
>>
File: 1771753551593775.png (87 KB, 1029x1058)
87 KB PNG
>>109546787
I kneel A100 chad
>>
>>109547721
could be helpful to compare the lora performances on same prompt/same seed
>>
>>109547713
wait, these are sick lol.
>>
>>109547731
"EDM" doesn't mean anything. it has as much meaning as people saying "AI"
>>
>>109547713
They should just have trained the H3 DiT base to output only long music.
But from what we got, they seemed afraid to train on "real" songs and get sued.
Might be the reason Qwen Music will not see a local release too (besides Alibaba deciding to cuck everyone putting everything behind API)
>>
>>109547678
https://huggingface.co/Astral01/Ace-Step_v1.5_XL_Base_Turbo_0.3_Merge

Tried and tested with LoRAs
ZUTOMAYO
https://files.catbox.moe/kae6or.mp3
https://files.catbox.moe/38f38d.mp3

Fate Gear (Power metal band)
https://files.catbox.moe/7znul3.mp3
https://files.catbox.moe/fn2ioj.mp3

Alternatively, 0.5 merge can also be good, but it's more slopped than this merge. These are DiT-only results, LM disabled, CFG 12-20 (mostly 20), 50 steps, Scrag's VAE.
>>
The lora goes a long way to unlocking the power of the ref model for me because it makes iteration so much more possible. Being able to see "wait does this work" or "did I prompt this right at all" in a short period of time means it's easy to try different combinations.
>>
>>109547713
they sound very AI tho especially that rock one, very generic and very artifical.
>>
>>109547758
Assuming the lora doesn't kneecap the model's prompt adherance.
>>
>sued
wtf I thought the chinks were based and immune to the legal jew
>>
>wait until they realize using only 8 steps fucks with prompt adherence and coherence.
>>
>>109547775
thats why I always do 9 steps
>>
>>109547756
Great, ty anon, I will give it a shot. Btw, did you train the lora on AceStep XL Base or Turbo? Which trainer did you use?
>>
>>109547721
dare I say it, 8 step turbo > 20 step spectrum for coherence (that also skips steps in a diff way)

https://files.catbox.moe/qaf9h0.mp4
>>
>>109547773
I've tried it with gens that worked with fl2v before and prompt adherence is just as good.
>>
>>109547775
who are you quoting?
>>
>>109545388
https://civitai.red/models/2845331/blowjob?modelVersionId=3212494
works great, just simply prompting "sucking".
>>
>>109547774
Well, Minimax is facing a lawsuit from Disney right now, and that's why you are required to sign to some bs to get commercial license for the model
>>
>>109547791
now do 10/20 steps -> upscale -> 4/8 turbo
>>
Odds and minimax releases the image model this month, evens next month. Dubs and they release it in 20 minutes.
>>
>>109547713
>beyond the <x> where the <y> <z>
closed the tab immediately
>>
Error log


# ComfyUI Error Report
## Error Details
- **Node ID:** 37:13
- **Node Type:** MiniMaxMusic3TextEncode
- **Exception Type:** AttributeError
- **Exception Message:** AttributeError: 'RVQDepthDecoder' object has no attribute '_v_block'

Really comfy?
>>
>>109547795
>1girl dancing to numa numa
The prompt.
>>
>>109547785
I used Base, training directly on Turbo will yield poor results.

Check out this rentry I made here
https://rentry.co/s8fg8ber
>>
File: banner_v3.png (711 KB, 1536x512)
711 KB PNG
Anons please enjoy my new prompt creation/enhancement tool which runs fully locally using any openai comaptible backend (I use Gemma4 of course). Gemma-chan now knows H3, klein, krea2, anima, etc!
https://github.com/whp199/GemmaPrompt
>>
>>109547808
bitch it's like got like three actors with independent behaviors and emoting and dialogue and a bunch of mechanical shit and all kinds of cuts fuck outta here
>>
>>109547848
I like gemma with gem hair ornaments better.
>>
>>109547848
smug brat
>>
>>109547806
great they release something that doesn't work on blackwell
>>
>>109547848
wtf anon? this ain't gemma!!!
why you gotta NTR like this?
>>
So far the ref turbo works very well regarding quality, but there's one annoyance with it. It really wants to zoom in.
Turbo wants to pan the camera in with every prompt and I need to wrangle it to submission.
>>
is he right?

https://files.catbox.moe/txpp3x.mp4
>>
>>109547873
There's a new turbo?
>>
File: fluster_klein.png (993 KB, 1024x1024)
993 KB PNG
>>109547866
Gemma-chan isn't smart enough for coding but she is smart enough to prompt H3 well
>>
>>109547866
Gemma is a qt but she's a bit of a retard when it comes to coding.
>>
>>109547878
yes, 4 step but use it at 6-8 for quality, I used 8 for my simpsons test ref gens
>>
>>109547885
>>109547887
She's a good at executing code if you already know what you want tho.
>>
>>109547885
cut your nails freak
can't do any decent yuri with them like that
>>
>>109547756
oh my god I hate weeb-shit so much.
>>109547848
where does the compute come from?
>>
>>109547919
llama.cpp is a popular one, but that's an inference engine?
I've never bothered to learn the correct terminology, it loads the model
>>
>>109547919
Your computer?
>>
Are you faggots going to bake or let this thread die?
>>
>>109547944
Should I do it again?
>>
>>109547948
Do what again pussy?
>>
>>109547930
>>109547934
Okay so I guess that means I have to download the model, right?
>>
When using the turbo lora, are you generally supposed to disable spectrum?
>>
>>109547944
>page 5
chill my guy. we've got easily 5 hours if not more before this gets archived.
>>
File: ldg baking guide.mp4 (3.59 MB, 768x1376)
3.59 MB
3.59 MB MP4
Reminder if you're going to bake.
>>
File: MiniMax_H3__00061.mp4 (3.65 MB, 480x640)
3.65 MB
3.65 MB MP4
>>
Blessed thread of frenship
>>
>>109547960
embarrassed newfag here: is there a post limit for threads on /g/? like if it gets to 600 posts or something does it stop anon from making more replies?
>>
>>109547997
>>109547997
>>109547997
>>
>>109548000
There is no post limit.
>>
>>109547848
>https://github.com/whp199/GemmaPrompt
Can I feed it a pages of a script and it will generate the shots for the scene, and arramge tje shots in the minimum amount of clips to be generated.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.