[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: collage_1785978860.webm (3.22 MB, 2048x1751)
3.22 MB
3.22 MB WEBM
Discussion and Development of Local Image, Video, and Music Models

Previous: >>109472936

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
b t o' f
>>
>>109474099
haha forgot how to 4chan >>>/wsg/6208208
>>
File: 1777609953851742.webm (3.86 MB, 1920x912)
3.86 MB
3.86 MB WEBM
https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora/discussions/1#6a73cf519b0aee71a9a71bcd
How do you make that comfyui compatible?
>>
>>109474136
ask claude
>>
>>109474136
>How do you make that comfyui compatible?
Claude can.
>>
Anyone have any experience with using those sodimm to dimm adapters with a ryzen system?
I have an opportunity to get 128GB for $1000 CAD after tax and I'm very tempted to do it.
>>
>>109472473
>MiniMax H3 Cache node
Here's your premium clanker security audit for this repo. That'll be 27 mikubucks, anon: https://rentry.org/h3-cache-security-audit
>>
>>109474160
>PyTorch 2.5 or earlier
what is this, 2022? who uses pytorch 2.5 nowdays?
>>
File: MiniMax_H3_00347_.webm (1.42 MB, 832x640)
1.42 MB
1.42 MB WEBM
>>
>>109474136
claude can brute force it for you, but i wouldn't waste tokens on doing so since its not ready yet
>>
>>109474160
>exposes a pentagon vram backdoor
its over
>>
>>109474172
Am I deluding myself thinking wan could have done this ableit at 5 second cap, 16fps doesn't seem to matter for double frame animation
>>
it amazes me how after gaining all this power all some of you guys do is cartoon slop that could be animated better by a japanese incel working overnight
>>
>>109474186
what are you gonna do with this power?
>>
>>109474180
>Am I deluding myself thinking wan could have done this ableit at 5 second cap
yes, there's a reason why everyone stopped using wan. you're forgetting how janky it was
>>
>>109474186
Sorry should we be posting unfunny Seinfeld skits instead?
>>
>>
>>109474186
Why don't you show anons what you have been cooking?
>>
>>109474196
>unfunny
*funny, sorry for the typo
>>
Why is there an anime thread when the anime posted here is a million times better? And this thread is actually active?
>>
>>109474196
are you the one who posted that? kek
https://www.reddit.com/r/StableDiffusion/comments/1vgl66e/minimax_h3_can_do_seinfeld_clips_we_get_it_already/
>>
>>109474206
NAI/API keks shitter split off.
>>
>>109474206
>a million times better
https://www.youtube.com/watch?v=j95kNwZw8YY
>>
>>109474211
gb2r
>>
File: MiniMax_H3_00279_.mp4 (980 KB, 576x896)
980 KB
980 KB MP4
I believe this was requested.
>>
>>109474211
no but good job outing yourself
>>
>>109474224
>no
sure...
>>
>>109474196
no we should post the 999999th george floyd video
>>
>>109474196
>>>/gif/31000619
>>
>>109474228
lmaoo, now that's a funny Seinfeld skit
>>
The secret to good prompting is to write a prompt that generates a prompt.
>>
>>109474228
ok you got me that one was good
>>
>>109474196
Have some more Seinfeld slop.
https://files.catbox.moe/pwxx8b.mp4
>>
>lmaoo
>>
4 step lora converted to comfy
https://litter.catbox.moe/r5jvn74gi6cq6ssv.safetensors
its not good enough for 4 steps but it works well at 8 steps
preview video:
https://litter.catbox.moe/6h8jvuewz0p0fu8r.mp4
>>
>>109474195
except the motion on wan was far worse
>>
why does sageattention seem to have different install steps fucking everywhere? how much does it improve speed by on minimax? never bothered with it, I just really don't want to fuck up my environment rn
>>
>>109474241
mustard gas just came out of my computer
>>
idk man. I just can't be bothered to gen video.
>>
>>109474241
>preview from the comfy link and now your own
>r5jvn74gi6cq6ssv.safetensors
sus
>>
>>109474251
i'm disappointed by how stupid and retarded /g/ is in 2026. it's literally as easy as making a venv, cloning the sageattention repo, and then running the install bash script
>>
>>109474256
shartbox randomizes filenames
>>
>>109474256
you know litterbox renames the files random names, right?
>>
>>109474257
too many clicks. too much keyboard
>>
>>109474256
.unsafetensors
>>
>>109474257
it's literally just clicking the hamburger menu of the package in stabilitymatrix
>>
>>109474256
>safetensors aren't safe
holy retard
>>
sorry, not running your 0-day safetensor exploit. i just wont. hahahaha. sorry!
>>
>>109474241
>4 step lora converted to comfy
which one is it, there's 2 of them on the huggingface repo, the ema or the non-ema?
>>
be careful downloading safetensors from unverified sources, they may contain malware!
>>
File: 1759869850229563.mp4 (3.72 MB, 1920x1088)
3.72 MB
3.72 MB MP4
>>
>>109474251
Just have your ai install it for you anon
>>
>>109474238
>understatement
It's so great to finally have a model that isn't afraid of blood and violence
>>
>>109474274
non. And its not done yet but works good enough at 8 steps it seems
>>
>>109474273
>can't ask a LLM to verify if the safetensor is safe
insane amount of skill issue
>>
>>109474266
my sides
>>
>>109474278
kino alert. i am planning on doing a POV medieval battle after i finish the tank kino prompt
>>
>>109474257
>random dependency fuckery NEVER happens
>>
>>109474241
1.2 strength seems to work best
>>
File: 1756777124066326.mp4 (3.5 MB, 1920x1088)
3.5 MB
3.5 MB MP4
>>
>>109474300
how to get this perspective? pov? or saying first person perspective?
>>
>>109474300
We wuz kangs
>>
>>109474300
was expecting a sisyphus style ending where the rope breaks and the block rolls back downhill
>>
>>109474300
Cool.
>>
am i fucked with only 12gb vram?
>>
File: 1774778595188676.jpg (24 KB, 554x554)
24 KB JPG
>>109474300
The blocks didn't look like that brand new. or is this set in the future?
>>
>>109474314
no, ive got 4
>>
>>109474278
>no sharks
ngmi
>>
>>109474281
It's not bad. A gore LoRA would go a long way.
https://files.catbox.moe/rm6xoi.mp4
>>
>>109474314
no, comfy automatically offloads your gpu if you don't have enough memory, and if you still OOM use those flags
--vram-headroom 1 --disable-pinned-memory
>>
Anon just gave me a virus right?
>>
Is R2V really that good? How does output compare to i2V?
>>
how do I get so deep into this that I am installing custom repos and learning about diffusion model cache but still retain some sense of my humanity?
how do I get that piece of me back that cared about other things?
I used to care about so much stuff.
I researched so many different topics across so many different fields and genres.
why is it that the only reason I care about getting back to that place because I believe it will help me to create more interesting gens?
who am I now?
my brain is not what it used to be.
it's been 5 years into this deep dive, anon.
I'm tired.
>>
Working on a custom turbo lora loader for H3.
Almost there.

>>109474314
I have 12gb vram and it works great.
>>
Have you guys figured out what the best sampler for H3 is yet?
>>
>>109474340
just ask chatgpt to help you make more interesting gens
>>
>>109474337
>sees format incompatibility as viruses
damn, that's one funny techlet
>>
>>109474241
>[ERROR] ERROR lora diffusion_model.blocks.2.adaln_proj.linear.weight shape '[96768, 8]' is invalid for input of size 260112384
>>
https://huggingface.co/QrusherZA/H3_Turbo_ComfyUI/tree/main
here, on huggingface instead since you retards are scared of litterbox for some reason
>>
>>109474345
Res Multistep with Beta Schedule for max clarity.
>>
>>109474337
You've been pwned
Replace all your credit cards now
>>
>>109474346
the one thing that an LLM can not do is draw from divine inspiration.
I used to be able to.
It's gone now and it feels like it's never coming back.
>>
>>109474353
>620mb
anons file was 740mb
definitely virus.
>>
File: uiiuid.png (272 KB, 837x814)
272 KB PNG
>>109474353
>>
>over 80 replies in half an hour
>>
>>109474353
it's the same one as this one? >>109474241
because if it is, I still have errors on my console
>[ERROR] ERROR lora diffusion_model.blocks.45.adaln_proj.linear.weight shape '[96768, 8]' is invalid for input of size 260112384
>>
is H3 good at genning POOPING
>>
>>109474353
>huggingface
AAAAAIIIIIEEEEEEEEEEE
>>
File: 1755143009501545.mp4 (3.65 MB, 1920x1088)
3.65 MB
3.65 MB MP4
>>109474303
>how to get this perspective? pov? or saying first person perspective?
integrated_multimodal_description: [Shot 1] Live-action, cinematic, first-person POV, GoPro-style, starting on a massive, dust-caked earthen ramp spiraling up the side of an unfinished pyramid under a blistering white sun. The camera holds a POV tracking shot and shakes slightly as calloused, dust-covered hands grip thick hemp ropes tied around a colossal limestone brick on a wooden sled, the rope fibers biting into palms. The runner leans back and heaves, feet digging into loose sand and gravel with a gritty scrape, the sled groaning and lurching forward inch by inch with a deep wooden creak. The camera pedestals up with large amplitude at slow speed as the ascent continues, revealing sweat dripping onto the lens and, in the shimmering heat haze in the distance, another huge, completed pyramid towering perfectly against the blue sky. The camera pushes in with small amplitude at slow speed as the block is finally dragged onto the top platform, hands releasing the ropes and slapping dust from thighs while labored breathing echoes.\noverall_soundscape: Hemp ropes creak under extreme tension, the wooden sled groans and scrapes loudly against sand and stone, feet scrabble and slip on gravel. Heavy, exhausted panting and grunts dominate, with distant shouts of other workers, whips cracking faintly, and hot desert wind whistling past.\nnon_diegetic_music: Low, percussive tribal drums at a slow, heavy tempo with deep resonant hits that match each heave, joined by a sustained, dusty horn drone that swells slightly as the distant pyramid is revealed.
>>109474311
>was expecting a sisyphus style ending where the rope breaks and the block rolls back downhill
the possibilities are endless
>>109474317
>The blocks didn't look like that brand new
another test of world model knowledge. it uses modern ruined coliseum for coliseum gens too
>>
Is r2v meant to be way slower than i2v?
>>
>>109474370
try it anon
>>
>>109474370
Inquiring minds want to know. Surely somewhere in the dataset are animals shitting. It should therefore be able to approximate a human shitting... right?
>>
>>109474373
Yes
>>
File: HO_MdG6bwAAcD6r.jpg (106 KB, 1200x747)
106 KB JPG
https://x.com/DesignArena/status/2085109955590594995
H3 beating seedance is most tests
>>
>>109474373
I wouldn't phrase it like that but it's expected
>>
>>109474136
based on this and the wf i'm working on, extraordinary things are coming our way. I don't even think we are ready for this shit.
>>
>>109474324
Is that r2v?
>>
>>109474372
thx
>>
>>109474381
>mememarks
seedance is still on a league of its own, especially for high paced action, only this model doesn't have weird shit when it goes fast
>>
>>109474383
What are you working on anon?
>>
>>109474389
it only has weird shit because everyone is genning at cope resolutions.
>>
>>109474353
this caused mustard gas to leak out of my GPU
>>
>>109474376
>implying there isnt plenty of human videos
>>
>>109474391
nah, even if you try the API version of Minimax it's still not close to seedance
>>
File: 1758163259221666.jpg (175 KB, 1408x959)
175 KB JPG
>>109474113
Thread moves too fast...
....Who am i kidding. I waste two days for this and i still have deadline to finish LOL
>>
>claude please update comfy and ensure nothing gets borked
feelsgoodman
>>
>>109474393
all you need is a live google maps view of india
>>
>>109474387
Nah, straight t2v.
>>
>>109474395
ok but seedance can't do porn or funny copyright edits so who cares
>>
>>109474373
with video yes, it has to basically run the video as context meaning its your generated video + the ref's length
>>
>>109474381
you just angered the sneedance shill
>>
>>109474406
trvth nvke
>>
>>109474389
The cool thing is that H3 is open weights.
Researchers will continuously improve it. I mean hell, some random dude is training a distillation on like 8 H200s right now on day 4.
Look at how much researchers improved Wan and LTX over time and those models aren't nearly as capable.
>>
>APIfag suddenly calling jeetmarks a meme
total local victory
>>
>>109474353
All of my crypto is fucking gone guys.
>>
>>109474417
>jeetmarks
see, you call them jeetmarks, and you take them seriously?
>>
audio is definitely a bit wonky with turbo lora but results remain impressive.

will test with 12 steps.
>>
Show me your most powerful H3 gen right now.
>>
>>109474406
NSFW DOESN'T EVEN MATTER!
now i'll post 6 hilariously bad sneedance sex scenes from 7 months ago.
>>
File: 1758199016223931.mp4 (3.34 MB, 1920x1088)
3.34 MB
3.34 MB MP4
Seedance isn't even as good as H3 in some 1:1 comparisons. It also suffers from serious sameface and of course censorship and is expensive as fuck

Seedance was the best. Now you can reasonably make anything you want weith H3 if you're determined enough
>>
>>109474428
mine only have one metric; nuts busted and I can't post those
>>
>>109474427
theorically we're supposed to get better results with the turbo lora right? because the turbo lora is supposed to reproduce the 50 steps process, and we always go with 20 steps
>>
>>109474430
>Seedance isn't even as good as H3 in some 1:1 comparisons. It also suffers from serious sameface
Seedance 2.5 fixed those issues though, I too was fucking tired of the same face, but they made it almost 2x as expensive as SD2.0 that's ridiculous lol
>>
man youd think /e/ would be all over this shit since it can do tits at least. not much h3 activity in the /vp/ thread either
>>
File: 74453455.webm (1.16 MB, 448x256)
1.16 MB
1.16 MB WEBM
>>109474428
>>
So how much longer until some api corpo trains H3 on their "advanced" dataset and offers it as their own thing
>>
think of all the improvements that came out on wan / ltx. And this model is actually frontier level. Gonna be like 1000 papers / experiments improving it.
>>
>>>/wsg/6208904
>tubo lora 1.2, euler/simple 8 steps
>euler/simple 8 steps
>0.7mp

also wtf picrel
>>
>>109474373
reduce the reference video's length and fps and resolution until you get a s/it you can live with.
>>
>>109474440
not surprising considering you have to be intelligent to be on the cutting edge
>>
ok. 1.3 strength, 10 steps is best with turbo lora
https://litter.catbox.moe/6rvm2fnivwdbtvz3.mp4
>>
Thoughts? Any of you tried it with H3?

https://huggingface.co/sakamakismile/Qwen3-VL-32B-Heretic-MiniMax-H3-NVFP4
>>
>>109474453
try strength = 2, seems like it's the best according to that comparison video >>109474136
>>
>>109474463
Now we're talking, gen time?
>>
Does having 2 Nvidia cards work for this stuff?
I got a 4060 16 right now and am thinking of upgrading in general
>>
>>109474463
damn it looks even better than 40 steps.
>>
>>109474463
wtf, the turbo lora has almost the same movements than the no lora + 40 steps, impressive as fuck
>>
>>109474466
stop posting that snake oil clueless twitter influcer BS. Diverging from what the model was trained with = worse results. TE's are not censored
>>
Turbo lora absolutely cooks the audio. Least it's not slo-mo, I guess. But I'll wait until it's done baking or those light guys release theirs.
>>
turbo loras has been trained for about a day on 8x H100s. He expects about 1-2 more days
>>
>>109474453
Turbo Lora is out ???
>>
>>109474463
Workflow please?
>>109474487
Maybe idk
>>
>>109474353
https://files.catbox.moe/gvfnxj.mp4
>8 steps strength 1.2
too bad it destroys the audio, but that looks really promising
>>
File: 366462.webm (2.5 MB, 448x256)
2.5 MB
2.5 MB WEBM
am i supposed to use a certain shift scale value with the turbo lora?
>>
>>109474381
Who knows what simple tests they are doing though
>>
File: compressed_under_4mb.mp4 (3.81 MB, 1920x576)
3.81 MB
3.81 MB MP4
ok, here's your charity, anons. Cache node is better than spectrum for faster gens. Quality difference is mostly negligible for all of them. Use it to make fast test runs and then gen at full quality if you like it.

>everything done at 0.4 megapixel
>3090 Ti
>I2VA

uncompressed video here: https://files.catbox.moe/p00k43.mp4
>>
>>109474490
you just add this lora at strength 1.3 and you go for 10 steps >>109474353
>>
>>109474466
for the last time, uncensored text encoders will not magically force the model to make porn for you, otherwise everyone would be using uncensored text encoders for wan/ltx/krea/etc
>>
>>109474503
>Cache node
on which github? there's a lot of cache node that exist for minimax already? and what parameters?
>>
>>109474509
https://github.com/silveroxides/ComfyUI-UtilsCollection
>>
>>109474506
this has to be some kind of jeet psyop. why do people keep bringing it up in this thread?
>>
>>109474337
lol debo'd
>>
>>109474353
Is this legit ??
>>
>>109474496
ran it at 12 and that seemed to fix it
>>
>>109474503
>>109474509
https://github.com/silveroxides/ComfyUI-UtilsCollection
I meant to mention that. It's been a long day of genning, if you can't tell by the times on the bottoms of those videos.
>>
>>109474337
his loras were acounting for the additional 13b layers that we don't have because we're using the pruned model (20b, not the 33b)
>>
>>109474503
doing gods work anon. cheers.
>>
File: H3_Combine2.jpg (211 KB, 1213x1083)
211 KB JPG
>>109474428

https://streamable.com/6vwywc
>>
dear china
we will nuke you
unless you release a new image model
-regards
>>
>>109474517
it is, but the audio is bad though >>109474494

below is another try
>8 steps strength 2.0
https://files.catbox.moe/gvfnxj.mp4
>>
>>109474534
are you simply using the last frame of the previously generated video as your first frame for the next video?
>>
>>109474453
>same seed
>turbo 1.3
>res_multi / simple, 10 steps
>0.5mp
>8 / 6 sigmas
https://files.catbox.moe/3zahaf.mp4

Audio definitely sounds OK to me.
>>
File: 1782268437272481.png (118 KB, 1318x1012)
118 KB PNG
>>109474503
what parameters you went for?
>>
>>109474546
these are the best.
>>
>>109474540
>regards
kek
>>
File: 464536.png (274 KB, 1414x394)
274 KB PNG
i think the turbo lora ruined seed variance cause my gens look very similar now
>>
Can you niggas shut the fuck up and not generate loli feet for a second? This thread is always on top. Just shut the fuck up and goon already. It's been 3 hours.
>>
>>109474546
the ones in your image. They are the defaults for a reason.
>>
File: 1765581495657772.png (58 KB, 939x464)
58 KB PNG
>>109474545
oh yeah, that anon has a point, you have to increase the sigma of audio if you go for lower steps, that can explain things
>>
>>109474543
Yes and no. Some are hard cuts.
>>
>>109474503
It's so good, I now get a 10 second video in the same time it would me to generate 3 images in other models.
definitely worth it
>>
>>109474552
the main drawback of turbo is indeed losing seed variance.
>>
>>109474568
cap
>>
>>109474545
Lora solved
>>
>>109474573
i actually meant to quote the lora post. but I am using the cache node as well.
>>
>>109474566
So, are you manually going through an regenning everything, or are you using something like the director node?
>>
>>109474545
>>8 / 6 sigmas
8 for shift_video and 6 for shift_audio right?
>>
>>109474572
Can't that be fixed by injecting noise or whatever those krea nodes do
>>
File: AniStudio-01345.png (1.46 MB, 768x1408)
1.46 MB PNG
>>
>>109474579
yes. video is something you can play with tho. I like it lower.
>>
>>109474552
that is the trade off. It distills the model by basically baking in some of the steps meaning those steps will always been the same. You can balance it by using more steps vs lower weight.
>>
>>109474578
Manual prompting each gen. Cause the director is too restrictive.
>>
>>109474588
Also a wan workaround was to have the lora off / at lower weight for the first few steps, then on at higher weight for the last ones
>>
>halfway through genning
>want to change the prompt
goddamn i hate not being able to visualize what the video should look like before i send the prompt
>>
File: 1771997198399194.jpg (34 KB, 566x411)
34 KB JPG
>>109474546
da fuck isnt this a built in node now? or its just slopped based on the original node?
>>
>>109474407
I wonder if making the source video low res and low frame rate makes the gen faster
>>
FUCK SLEEP
AI GOONNA KILL ME
>>
>>109474634
it just fits it to whatever res your generated video is gonna be
>>
Gonna stick with the cache node and 20 steps until the turbo lora is done cooking. It's just slightly slower for far better quality.

Very promising tho.
>>
>>109474644
you got that right.
>>
When will a good one drop on civitai?

https://d.uguu.se/KUMfyeqB.mp4
>>
>>109474136
>>109474353
Which one is better Turbo ???
>>
>>109474659
sulphur dev plans to start training in a about 10 days, is waiting to fund it. And he has a much better dataset now that also includes dan / e621 on top of real stuff
https://www.reddit.com/r/StableDiffusion/comments/1vgdqei/sulphur_3_is_looking_for_funding/
>>
>>109474661
they're the same, one is not compatible with comfyui format, the other is compatible
>>
Why are people shilling Cache now? I thought Spectrum was best?
>>
>>109474666
The Qrusher one ??
>>
>>109474670
yeah?? duh? that one has comfyui name in it
>>
fapped twice today, blasted a big load out both times.

H3 is just too powerful. Honestly in disbelief over how capable it is. This should honestly kill off the porn industry, but I'm guessing the luddite movement is still big enough to reject it.
>>
>>109474584
best gen itt. catjack could never
>>
>>109474691
since when the luddites have even won a single war? technology has always advanced they can't do anything about it
>>
>>109474702
nuclear
>>
>>109474669
>anon learns that people are experimenting a 2 days old model
it'll take some time before the right settings will be found
>>
>>109474669
Poorfag hour that wants speed over quality.
>>
>>109474691
the output length is too short and waiting time too long to kill traditional porn
for gooners its getting good but for normal porn consumers, no
>>
Created a loader node for H3 Turbo using the original .safetensors lora file.

Supports gguf models and the sageattention node from kjnodes

https://files.catbox.moe/pq4dcw.7z

Node order.
Model loader - turbo lora loader - sageattention patch - ksampler/H3 director
>>
File: 17862.png (382 KB, 560x821)
382 KB PNG
>>109474707
not in france
>>
>>109474691
>fapped twice today, blasted a big load out both times.
physically impossible for me to climax to my own gens as i know what theyll look like as soon as i hit "go"
>>
>>109474344
What does your setup look like as far as resolution/step count/scheduler/sampler etc. I have 12gb vram and 32gb ram and my outputs are a bit too blurry but I wanna know the lowest settings I can get away with, while maintaining decent quality
>>
>>109474561
oldfag here. how do you connect this node? kek
>>
File: 1570187538152.png (180 KB, 519x533)
180 KB PNG
How good is R2V at mimicking Japanese voices? If I give it an audio sample of Japanese dialog and wrote some text in kanji or romaji would it be able to clone the voice accurately?
If I gave it blowjob asmr audio could it do that too?
>>
>>109474561
How to connect sigma shift ?
>>
>>109474588
so i think the plan is to use the lora for rapid prompt iteration and then turn it off when i want to start collecting interesting kinos
>>
>>109474728
>>109474733
are you serious? when you see a model -> model connection you put that between the two
>>
>>109474715
>the output length is too short and waiting time too long to kill traditional porn
What are you talking about? R2V can stitch videos together natively. The sky is the limit.
Also 10 second clips from H3 are already 100 times more arousing than any shit on pornhub. You can tailor it exactly to what works for you. I don't need a 1 hour goonerslop video to get horny.
>>
>>109474724
>he didn't write a big wildcard prompt
>>
>>109474734
you say that now but once you've seen the mock up you'll move onto the next thing
>>
>>109474738
but there's the
>lora stack
>Patch Sage Attention KJ
>MiniMax H3 Mem Eff Sage Attention Patch
>EasyCache or Spectrum Apply MiniMax H3
all connected to the model
where would you place the sigma?
>>
>>109474503
>sage + cache + turbo lora - 4m01s
>these settings (except the same 0.4mp I was using for the others) for turbo lora >>109474545
>clanker sloppa prompt: https://pastebin.com/7WM03VH1

The turbo LoRA works, but the quality takes a huge hit. Also, you will not be able to reliably make almost the same gen with this and then remove all optimizations to make a higher quality version of the same video. I think I'll stick with just the cache node for now and then when I like one I can make it look even better.
>>
Can you guys imagine using wan2.1 now? Any of you care to try a prompt comparison it to see just how bad it is?

I remember being amazed by wan2.1 last year, but looking back, goddamn it's so much worse
>>
>>109474745
>once you've seen the mock up you'll move onto the next thing
what next thing? hard to be faster than a turbo lora
>>
>>109474745
>you'll move onto the next thing
i had like 1000 videos of my fighter jet videos, i don't get bored that quickly
>>
>>109474725
See >>109473929

The Turbo lora needs some more cooking still.

>>109474748
Right before the ksampler node
>>
>>109474734
just know what you're getting anon. It won't be the same >>109474755
All of those used the same prompt and seed. Maybe a better turbo will come out soon.
>>
>>109474762
thanks for the spoonfeed :^)
>>
File: debo_sc_k2_00017_.png (2.87 MB, 1872x1007)
2.87 MB PNG
>>
>>109474755
lil sis thought she was the main character
>>
I'm minimaxxing.
>>
>>109474774
dishonored vibes
>>
File deleted.
https://n.uguu.se/SjpZwpqF.webm
>>
>>109474780
uhhhhhhhhhh
>>
>>109473929
anyone who recommends ggufs are either trolling or retarded. they dont "save vram". Int8 streams the weights, you dont have to fit it all in vram. GGUFS are forced to fit all in vram and are about 2.25x slower than int8
>>
2 days and basically all my gens have been porn and most of the rest is cringe waifu shit. Am I cooked chat?
>>
>>109474780
whoops didn't mean to upload that here
>>
>>109474669
see >>109474503
and judge for yourself.
>>
File: 1783666330901717.webm (3.86 MB, 656x816)
3.86 MB
3.86 MB WEBM
>>109474756
>I remember being amazed by wan2.1 last year, but looking back, goddamn it's so much worse
I still have my wan 2.1 gen, goddam they're bad and slow, but I was impressed too at the time, it was my first time making images move, it was like discovering fire lool
>>
File: wan22.webm (3.61 MB, 1600x608)
3.61 MB
3.61 MB WEBM
>>109474756
>>
>>109474780
hot
>>
>>109474503
how do you do these multiple outputs? Or are you stitching these together manually?
>>
File: 1782184302876629.png (44 KB, 738x442)
44 KB PNG
>>
File: kay.mp4 (1.34 MB, 480x832)
1.34 MB
1.34 MB MP4
>>109474756
ahh... wan2.1...
yep, those were the days
>>
File: 200w.gif (557 KB, 200x150)
557 KB GIF
>>109474815
>>109474825
>>
File: download.jpg (46 KB, 600x603)
46 KB JPG
are you behaving yourselves?
>>
File: Mik.webm (1.33 MB, 480x500)
1.33 MB
1.33 MB WEBM
>>109474835
>>
File: bones.png (1.27 MB, 2880x1920)
1.27 MB PNG
>>109470435
Can it render and keep rig bones?
>>
dunno why it made the viewer so fat here all i mentioned was a white shirt
>>
>>109474801
Roger that, downloading the Int8 version now to test. I had started off by using the Int4 version but the outputs were garbage, I erroneously assumed the Int8 version would have the similar issues.

Will report back.
>>
>>109474854
you know why
>>
>>109474850
>that endless yapping
kek, that's definitely a Wan render
>>
>>109474831
I used to manually make ffmpeg commands. Now I just get a clanker to write the command based on the filenames. There's probably some workflow that can do it for me, but I don't do this sort of thing enough to justify learning how to set that up.
>>
https://github.com/Comfy-Org/ComfyUI/pull/15334
kijai made the vae process 2x faster with int8 convrot now
>>
File: 145130.jpg (41 KB, 622x402)
41 KB JPG
>>109474835
>>109474850
>>
File: 1772027018504277.png (3.11 MB, 1920x1088)
3.11 MB PNG
this is the best model ever
>>
File: MiniMax_H3_00327_.mp4 (2.01 MB, 1056x704)
2.01 MB
2.01 MB MP4
>>
>>109474865
same. i don't code with ai but i still ask for commands
>>
>>109474860
int8 convrot is Q8 tier so you can have fun with it anon
https://github.com/BobJohnson24/ComfyUI-INT8-Fast/blob/main/Metrics.md
>>
>>109474872
this thread is meant for AI generated outputs
>>
>USE EULER + BETA instead of res_Multistep because res_multistep give disco lights.

Wait what ??
>>
File: 1681179301328839.png (679 KB, 791x658)
679 KB PNG
I didn't know seedance could do this until now with the plus ultra min version but

>take porn clip
>gen a goth girl or girl of your preference with grok or your choice local model
>swap them out in prompt
>you can even edit the existing video and add camera cuts
>you can even add extra characters watching and reacting

what a fucking time to be alive man. it's like some cyberpunk bullshit

https://files.catbox.moe/dwg40e.mp4
>>
>>109474881
for 4 steps, I don't have that problem with 8
>>
File: file.png (76 KB, 249x327)
76 KB PNG
i fucked up somewhere
>>
>>109474885
>inb4 transphobes start seething
>>
>>109474891
No that looks right
>>
>>109474885
>seedance
local models?
>>
>>109474904
>http://127.0.0.1:8188/

>Unable to connect

>Firefox can’t connect to the server at 127.0.0.1:8188
it doesnt work
>>
>>109474885
>>swap them out in prompt
what
>>
>>109474904
you leaked your ip retard.
>>
>>109474911
it actually will let you swap out the entire body of the person in the clip, including clothes.
>>
frick
>>
>>109474904
holy fuck, i know /g/ is tech illiterate but this is insane, just leaking your ip like that
>>
>>109474874
setting up hermes on a dedicated box was a game changer for finally getting around to all of those backburner projects. Just let the model do it for you.
>>
>>109474865
ah that makes sense, thanks for the tip
>>
>>109474900
i should have kept it going but i also noticed a typo
>>
File: loic36776912.png (164 KB, 850x474)
164 KB PNG
im about to pwn that noob
>>
>>109474919
oh, it does v2v with i reference?
>>
>>109474935
pls dont how will i keep generating ludos
>>
https://d.uguu.se/ZaJYlUGe.mp4
>>
>>109474938
part of the onmi reference with seedance. you basically just say

replace the female character in @video1 with the character @image1
>>
this turbo lora is excellent. one step closer to replacing ltx
>>
Am I supposed to get a shitload of errors loading that 4 step lora? What loader node am I supposed to use?
>>
>>109474943
are you using video references, this looks familiar.
>>
>>109474506
>text en
fuck off retard. it definitely makes a difference. You cannot simply say things like "fucked from behind" at get it to actually interpret it properly with the standard encoder. no one wants to describe their porn scene like it's a fucking clinical examination
>>
>>109474950
>Am I supposed to get a shitload of errors loading that 4 step lora?
no, you have to download this one >>109474353
>>
File: 1489951762876.jpg (8 KB, 251x201)
8 KB JPG
remember vace in wan?
>>
i need to learn to prompt
https://d.uguu.se/vEzxQGgq.mp4
>>
>>109474956
my fully erect penis just slid in and out of your moms mouth multiple times
>>
do turbo loras degrade the knowledge of the model? or is it just the seed variance that gets affected
>>
>>109474953
no, it's just an i2v prompt. You've probably watched a lot of porn an recognize what it was trained on.
>>
>>109474971
>>109474598
>>
>>109474967
I think it's pacing, I dunno what anybody else is doing but I'm rehearsing the scene in my head to figure out the timing
>>
>>109474969
>just saying bullshit because he knows I'm right
thanks you for your submission cuck
>>
>>109474974
that isn't what i asked
>>
>>109474979
>>109474588
>>
>>109474981
that also isn't what i asked
>>
>>109474833
Spectrum has a better quality but it's also slower, I will try the Provisional aggressive preset to get equivalent time and see if the quality is still better
https://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3#conservative-preset
https://files.catbox.moe/yvrs8h.mp4
>>
>>109474875
Genning with it now, but it's thrashing my ssd by using the pagefile instead of using system memory during inference (61% of 24gb)
>>
Is the turbo lora better than the optimization node schizo stack?
>>
>>109474996
>>109474463
>>
this vae decoder is too slow. someone needs to distill that as well
>>
File: 1779387398011737.png (450 KB, 600x600)
450 KB PNG
>>109474996
the audio is definitely impacted by that lora, we have to let him cook
>>
>>109475005
>>109474868
lol. All these anons asking shit already answered above
>>
>>109475019
>>109475019
>>
>>109475014
LOL tricked you into spoonfeeding me
>>
>>109474885
what provider are you using for seedance?
>>
is /e/ right?

>>>/e/3122899
>You can get around that with Wan by genning 5-second clips then using the fun-VACE checkpoints to generate transition frames between then. I think the main advantage of this Minmax is that is seems to understand 2D animation. Wan only knows 3D and realism so it gives a heavy render-like bias to everything, and doesn't understand anime faces. LTX is the same but much worse.

Still trying to make a case for wan over h3.
>>
>>109475005
true
saw this earlier https://github.com/Comfy-Org/ComfyUI/pull/15334
>>
>>109474534
kys you creepy fuck
>>
>>109474868
I put swapped the current VAE for this one, but my vids just returned as black
>>
>>109475028
Anon might have severe brain damage. H3 ref can do consistent characters WAY better than Wan + character lora. Just that is worth dumping wan for all time.
>>
https://files.catbox.moe/34eqhb.mp4
Better than the last version, I reckon.
>>
>>109474868
>https://github.com/Comfy-Org/ComfyUI/pull/15334

I bathe thee gratitude and love Kijagod.
>>
>>109475098
kek
>>
>>109475098
lmao that was pretty realistic
>>
>>109475092
you prob don't have latest comfy + comfy dependences + CU130+
>>
>>109474466
FFS ANON...

I might be wrong on certain understanding but a text encoder is just that, it has no gaurdrails, it has no samplers, it does not predict it just converts you text to numeric values called tokens that get sent into the image or video model. if it had any of those things it would be even excruciatingly slower. The guardrails are in the image or video model not the fucking text encoder. just fuck off, tired of your brain damaged shit since fucking 1 year ago.
>>
>>109475027
artcraft
>>
>>109475127
that is correct. But its retarded "ai influencers" constantly pushing this useless shit for updoots. The TE only acts as a translator and changing the TE that the model was trained on at all simply makes it diverge from what it was trained for = worse gens.
>>
>>109474466
and if you use an llm (big difference) to enhance a prompt then yes that will censor you and that is when you might want to use Heretic or abliterated versions of those llms to get past its safety filter for nsfw gens. But the text encoder versions will do fuck all and are a waste of compute, bandwidth and other peoples time.
>>
>>109475110
Is CU130+ basically a requirement for reference video generation?
The last time I tried updating, most of my workflows broke
>>
>>109475194
yes. Otherwise its far slower and takes far more vram
>>
>>109475201
Okay, making a backup of everything and making the attempt I guess.
I remember updating cuda means you also update sageattention, does someone have the link to the versions of each that match?
>>
File: file.png (31 KB, 720x317)
31 KB PNG
>>109475201
Is installing it really this simple?
>>
>>109475219
if new comfy yea, otherwise you should uninstall the old torch / just remove the venv first
Then do those, then the comfyui requirements, then triton then sageattention2.2.0+
There is a one click .bat somewhere out there that does it all
>>
>>109475248
Do I have to use “—disable-pinned-memory” in the startup when using cuda130? Showering across different guides rn.
>>
>>109475262
no. Maybe if you only have a tiny bit of ram
>>
>>109475248
If this is what my python package list looks like in stabilitymatrix, does this mean I already am using cuda130?

I saw on reddit that the comfyui console usually shows a warning if you aren't using cuda130 and mine isn't showing that either.
>>
>>109475218
>>>chatgpt
>>
File: file.png (7 KB, 275x152)
7 KB PNG
>>109475337
forgot pic oops
>>
>>109475201
well that sucks, i have a 3090, will cuda 13.0 make a difference? or will it only change for 40s and 50s?
>>
>>109475408
especially for 3090 as int8 is 2x faster instead of just 40% faster
>>
>>109474381
Why did they give this to us for free tho
>>
>>109475427
Turns out that hosting API services is crazy expensive when you don't have that much customers



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.