Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109472936https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
b t o' f
>>109474099haha forgot how to 4chan >>>/wsg/6208208
https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora/discussions/1#6a73cf519b0aee71a9a71bcdHow do you make that comfyui compatible?
>>109474136ask claude
>>109474136>How do you make that comfyui compatible?Claude can.
Anyone have any experience with using those sodimm to dimm adapters with a ryzen system?I have an opportunity to get 128GB for $1000 CAD after tax and I'm very tempted to do it.
>>109472473>MiniMax H3 Cache nodeHere's your premium clanker security audit for this repo. That'll be 27 mikubucks, anon: https://rentry.org/h3-cache-security-audit
>>109474160>PyTorch 2.5 or earlierwhat is this, 2022? who uses pytorch 2.5 nowdays?
>>109474136claude can brute force it for you, but i wouldn't waste tokens on doing so since its not ready yet
>>109474160>exposes a pentagon vram backdoorits over
>>109474172Am I deluding myself thinking wan could have done this ableit at 5 second cap, 16fps doesn't seem to matter for double frame animation
it amazes me how after gaining all this power all some of you guys do is cartoon slop that could be animated better by a japanese incel working overnight
>>109474186what are you gonna do with this power?
>>109474180>Am I deluding myself thinking wan could have done this ableit at 5 second capyes, there's a reason why everyone stopped using wan. you're forgetting how janky it was
>>109474186Sorry should we be posting unfunny Seinfeld skits instead?
>>109474186Why don't you show anons what you have been cooking?
>>109474196>unfunny*funny, sorry for the typo
Why is there an anime thread when the anime posted here is a million times better? And this thread is actually active?
>>109474196are you the one who posted that? kekhttps://www.reddit.com/r/StableDiffusion/comments/1vgl66e/minimax_h3_can_do_seinfeld_clips_we_get_it_already/
>>109474206NAI/API keks shitter split off.
>>109474206>a million times betterhttps://www.youtube.com/watch?v=j95kNwZw8YY
>>109474211gb2r
I believe this was requested.
>>109474211no but good job outing yourself
>>109474224>nosure...
>>109474196no we should post the 999999th george floyd video
>>109474196>>>/gif/31000619
>>109474228lmaoo, now that's a funny Seinfeld skit
The secret to good prompting is to write a prompt that generates a prompt.
>>109474228ok you got me that one was good
>>109474196Have some more Seinfeld slop.https://files.catbox.moe/pwxx8b.mp4
>lmaoo
4 step lora converted to comfyhttps://litter.catbox.moe/r5jvn74gi6cq6ssv.safetensorsits not good enough for 4 steps but it works well at 8 stepspreview video:https://litter.catbox.moe/6h8jvuewz0p0fu8r.mp4
>>109474195except the motion on wan was far worse
why does sageattention seem to have different install steps fucking everywhere? how much does it improve speed by on minimax? never bothered with it, I just really don't want to fuck up my environment rn
>>109474241mustard gas just came out of my computer
idk man. I just can't be bothered to gen video.
>>109474241>preview from the comfy link and now your own>r5jvn74gi6cq6ssv.safetensorssus
>>109474251i'm disappointed by how stupid and retarded /g/ is in 2026. it's literally as easy as making a venv, cloning the sageattention repo, and then running the install bash script
>>109474256shartbox randomizes filenames
>>109474256you know litterbox renames the files random names, right?
>>109474257too many clicks. too much keyboard
>>109474256.unsafetensors
>>109474257it's literally just clicking the hamburger menu of the package in stabilitymatrix
>>109474256>safetensors aren't safeholy retard
sorry, not running your 0-day safetensor exploit. i just wont. hahahaha. sorry!
>>109474241>4 step lora converted to comfywhich one is it, there's 2 of them on the huggingface repo, the ema or the non-ema?
be careful downloading safetensors from unverified sources, they may contain malware!
>>109474251Just have your ai install it for you anon
>>109474238>understatementIt's so great to finally have a model that isn't afraid of blood and violence
>>109474274non. And its not done yet but works good enough at 8 steps it seems
>>109474273>can't ask a LLM to verify if the safetensor is safeinsane amount of skill issue
>>109474266my sides
>>109474278kino alert. i am planning on doing a POV medieval battle after i finish the tank kino prompt
>>109474257>random dependency fuckery NEVER happens
>>1094742411.2 strength seems to work best
>>109474300how to get this perspective? pov? or saying first person perspective?
>>109474300We wuz kangs
>>109474300was expecting a sisyphus style ending where the rope breaks and the block rolls back downhill
>>109474300Cool.
am i fucked with only 12gb vram?
>>109474300The blocks didn't look like that brand new. or is this set in the future?
>>109474314no, ive got 4
>>109474278>no sharksngmi
>>109474281It's not bad. A gore LoRA would go a long way.https://files.catbox.moe/rm6xoi.mp4
>>109474314no, comfy automatically offloads your gpu if you don't have enough memory, and if you still OOM use those flags--vram-headroom 1 --disable-pinned-memory
Anon just gave me a virus right?
Is R2V really that good? How does output compare to i2V?
how do I get so deep into this that I am installing custom repos and learning about diffusion model cache but still retain some sense of my humanity?how do I get that piece of me back that cared about other things?I used to care about so much stuff.I researched so many different topics across so many different fields and genres.why is it that the only reason I care about getting back to that place because I believe it will help me to create more interesting gens?who am I now?my brain is not what it used to be.it's been 5 years into this deep dive, anon.I'm tired.
Working on a custom turbo lora loader for H3.Almost there.>>109474314I have 12gb vram and it works great.
Have you guys figured out what the best sampler for H3 is yet?
>>109474340just ask chatgpt to help you make more interesting gens
>>109474337>sees format incompatibility as virusesdamn, that's one funny techlet
>>109474241>[ERROR] ERROR lora diffusion_model.blocks.2.adaln_proj.linear.weight shape '[96768, 8]' is invalid for input of size 260112384
https://huggingface.co/QrusherZA/H3_Turbo_ComfyUI/tree/mainhere, on huggingface instead since you retards are scared of litterbox for some reason
>>109474345Res Multistep with Beta Schedule for max clarity.
>>109474337You've been pwnedReplace all your credit cards now
>>109474346the one thing that an LLM can not do is draw from divine inspiration.I used to be able to.It's gone now and it feels like it's never coming back.
>>109474353>620mbanons file was 740mbdefinitely virus.
>>109474353
>over 80 replies in half an hour
>>109474353it's the same one as this one? >>109474241because if it is, I still have errors on my console >[ERROR] ERROR lora diffusion_model.blocks.45.adaln_proj.linear.weight shape '[96768, 8]' is invalid for input of size 260112384
is H3 good at genning POOPING
>>109474353>huggingfaceAAAAAIIIIIEEEEEEEEEEE
>>109474303 >how to get this perspective? pov? or saying first person perspective?integrated_multimodal_description: [Shot 1] Live-action, cinematic, first-person POV, GoPro-style, starting on a massive, dust-caked earthen ramp spiraling up the side of an unfinished pyramid under a blistering white sun. The camera holds a POV tracking shot and shakes slightly as calloused, dust-covered hands grip thick hemp ropes tied around a colossal limestone brick on a wooden sled, the rope fibers biting into palms. The runner leans back and heaves, feet digging into loose sand and gravel with a gritty scrape, the sled groaning and lurching forward inch by inch with a deep wooden creak. The camera pedestals up with large amplitude at slow speed as the ascent continues, revealing sweat dripping onto the lens and, in the shimmering heat haze in the distance, another huge, completed pyramid towering perfectly against the blue sky. The camera pushes in with small amplitude at slow speed as the block is finally dragged onto the top platform, hands releasing the ropes and slapping dust from thighs while labored breathing echoes.\noverall_soundscape: Hemp ropes creak under extreme tension, the wooden sled groans and scrapes loudly against sand and stone, feet scrabble and slip on gravel. Heavy, exhausted panting and grunts dominate, with distant shouts of other workers, whips cracking faintly, and hot desert wind whistling past.\nnon_diegetic_music: Low, percussive tribal drums at a slow, heavy tempo with deep resonant hits that match each heave, joined by a sustained, dusty horn drone that swells slightly as the distant pyramid is revealed. >>109474311 >was expecting a sisyphus style ending where the rope breaks and the block rolls back downhill the possibilities are endless >>109474317 >The blocks didn't look like that brand newanother test of world model knowledge. it uses modern ruined coliseum for coliseum gens too
Is r2v meant to be way slower than i2v?
>>109474370try it anon
>>109474370Inquiring minds want to know. Surely somewhere in the dataset are animals shitting. It should therefore be able to approximate a human shitting... right?
>>109474373Yes
https://x.com/DesignArena/status/2085109955590594995H3 beating seedance is most tests
>>109474373I wouldn't phrase it like that but it's expected
>>109474136based on this and the wf i'm working on, extraordinary things are coming our way. I don't even think we are ready for this shit.
>>109474324Is that r2v?
>>109474372thx
>>109474381>mememarksseedance is still on a league of its own, especially for high paced action, only this model doesn't have weird shit when it goes fast
>>109474383What are you working on anon?
>>109474389it only has weird shit because everyone is genning at cope resolutions.
>>109474353this caused mustard gas to leak out of my GPU
>>109474376>implying there isnt plenty of human videos
>>109474391nah, even if you try the API version of Minimax it's still not close to seedance
>>109474113Thread moves too fast.......Who am i kidding. I waste two days for this and i still have deadline to finish LOL
>claude please update comfy and ensure nothing gets borked feelsgoodman
>>109474393all you need is a live google maps view of india
>>109474387Nah, straight t2v.
>>109474395ok but seedance can't do porn or funny copyright edits so who cares
>>109474373with video yes, it has to basically run the video as context meaning its your generated video + the ref's length
>>109474381you just angered the sneedance shill
>>109474406trvth nvke
>>109474389The cool thing is that H3 is open weights.Researchers will continuously improve it. I mean hell, some random dude is training a distillation on like 8 H200s right now on day 4.Look at how much researchers improved Wan and LTX over time and those models aren't nearly as capable.
>APIfag suddenly calling jeetmarks a memetotal local victory
>>109474353All of my crypto is fucking gone guys.
>>109474417>jeetmarkssee, you call them jeetmarks, and you take them seriously?
audio is definitely a bit wonky with turbo lora but results remain impressive.will test with 12 steps.
Show me your most powerful H3 gen right now.
>>109474406NSFW DOESN'T EVEN MATTER!now i'll post 6 hilariously bad sneedance sex scenes from 7 months ago.
Seedance isn't even as good as H3 in some 1:1 comparisons. It also suffers from serious sameface and of course censorship and is expensive as fuckSeedance was the best. Now you can reasonably make anything you want weith H3 if you're determined enough
>>109474428mine only have one metric; nuts busted and I can't post those
>>109474427theorically we're supposed to get better results with the turbo lora right? because the turbo lora is supposed to reproduce the 50 steps process, and we always go with 20 steps
>>109474430>Seedance isn't even as good as H3 in some 1:1 comparisons. It also suffers from serious samefaceSeedance 2.5 fixed those issues though, I too was fucking tired of the same face, but they made it almost 2x as expensive as SD2.0 that's ridiculous lol
man youd think /e/ would be all over this shit since it can do tits at least. not much h3 activity in the /vp/ thread either
>>109474428
So how much longer until some api corpo trains H3 on their "advanced" dataset and offers it as their own thing
think of all the improvements that came out on wan / ltx. And this model is actually frontier level. Gonna be like 1000 papers / experiments improving it.
>>>/wsg/6208904>tubo lora 1.2, euler/simple 8 steps>euler/simple 8 steps>0.7mpalso wtf picrel
>>109474373reduce the reference video's length and fps and resolution until you get a s/it you can live with.
>>109474440not surprising considering you have to be intelligent to be on the cutting edge
ok. 1.3 strength, 10 steps is best with turbo lorahttps://litter.catbox.moe/6rvm2fnivwdbtvz3.mp4
Thoughts? Any of you tried it with H3?https://huggingface.co/sakamakismile/Qwen3-VL-32B-Heretic-MiniMax-H3-NVFP4
>>109474453try strength = 2, seems like it's the best according to that comparison video >>109474136
>>109474463Now we're talking, gen time?
Does having 2 Nvidia cards work for this stuff?I got a 4060 16 right now and am thinking of upgrading in general
>>109474463damn it looks even better than 40 steps.
>>109474463wtf, the turbo lora has almost the same movements than the no lora + 40 steps, impressive as fuck
>>109474466stop posting that snake oil clueless twitter influcer BS. Diverging from what the model was trained with = worse results. TE's are not censored
Turbo lora absolutely cooks the audio. Least it's not slo-mo, I guess. But I'll wait until it's done baking or those light guys release theirs.
turbo loras has been trained for about a day on 8x H100s. He expects about 1-2 more days
>>109474453Turbo Lora is out ???
>>109474463Workflow please?>>109474487Maybe idk
>>109474353https://files.catbox.moe/gvfnxj.mp4>8 steps strength 1.2too bad it destroys the audio, but that looks really promising
am i supposed to use a certain shift scale value with the turbo lora?
>>109474381Who knows what simple tests they are doing though
ok, here's your charity, anons. Cache node is better than spectrum for faster gens. Quality difference is mostly negligible for all of them. Use it to make fast test runs and then gen at full quality if you like it.>everything done at 0.4 megapixel>3090 Ti>I2VAuncompressed video here: https://files.catbox.moe/p00k43.mp4
>>109474490you just add this lora at strength 1.3 and you go for 10 steps >>109474353
>>109474466for the last time, uncensored text encoders will not magically force the model to make porn for you, otherwise everyone would be using uncensored text encoders for wan/ltx/krea/etc
>>109474503>Cache nodeon which github? there's a lot of cache node that exist for minimax already? and what parameters?
>>109474509https://github.com/silveroxides/ComfyUI-UtilsCollection
>>109474506this has to be some kind of jeet psyop. why do people keep bringing it up in this thread?
>>109474337lol debo'd
>>109474353Is this legit ??
>>109474496ran it at 12 and that seemed to fix it
>>109474503>>109474509https://github.com/silveroxides/ComfyUI-UtilsCollectionI meant to mention that. It's been a long day of genning, if you can't tell by the times on the bottoms of those videos.
>>109474337his loras were acounting for the additional 13b layers that we don't have because we're using the pruned model (20b, not the 33b)
>>109474503doing gods work anon. cheers.
>>109474428https://streamable.com/6vwywc
dear chinawe will nuke youunless you release a new image model-regards
>>109474517it is, but the audio is bad though >>109474494below is another try>8 steps strength 2.0https://files.catbox.moe/gvfnxj.mp4
>>109474534are you simply using the last frame of the previously generated video as your first frame for the next video?
>>109474453>same seed>turbo 1.3>res_multi / simple, 10 steps>0.5mp>8 / 6 sigmashttps://files.catbox.moe/3zahaf.mp4Audio definitely sounds OK to me.
>>109474503what parameters you went for?
>>109474546these are the best.
>>109474540>regardskek
i think the turbo lora ruined seed variance cause my gens look very similar now
Can you niggas shut the fuck up and not generate loli feet for a second? This thread is always on top. Just shut the fuck up and goon already. It's been 3 hours.
>>109474546the ones in your image. They are the defaults for a reason.
>>109474545oh yeah, that anon has a point, you have to increase the sigma of audio if you go for lower steps, that can explain things
>>109474543Yes and no. Some are hard cuts.
>>109474503It's so good, I now get a 10 second video in the same time it would me to generate 3 images in other models.definitely worth it
>>109474552the main drawback of turbo is indeed losing seed variance.
>>109474568cap
>>109474545Lora solved
>>109474573i actually meant to quote the lora post. but I am using the cache node as well.
>>109474566So, are you manually going through an regenning everything, or are you using something like the director node?
>>109474545>>8 / 6 sigmas8 for shift_video and 6 for shift_audio right?
>>109474572Can't that be fixed by injecting noise or whatever those krea nodes do
>>109474579yes. video is something you can play with tho. I like it lower.
>>109474552that is the trade off. It distills the model by basically baking in some of the steps meaning those steps will always been the same. You can balance it by using more steps vs lower weight.
>>109474578Manual prompting each gen. Cause the director is too restrictive.
>>109474588Also a wan workaround was to have the lora off / at lower weight for the first few steps, then on at higher weight for the last ones
>halfway through genning>want to change the promptgoddamn i hate not being able to visualize what the video should look like before i send the prompt
>>109474546da fuck isnt this a built in node now? or its just slopped based on the original node?
>>109474407I wonder if making the source video low res and low frame rate makes the gen faster
FUCK SLEEP AI GOONNA KILL ME
>>109474634it just fits it to whatever res your generated video is gonna be
Gonna stick with the cache node and 20 steps until the turbo lora is done cooking. It's just slightly slower for far better quality.Very promising tho.
>>109474644you got that right.
When will a good one drop on civitai?https://d.uguu.se/KUMfyeqB.mp4
>>109474136>>109474353Which one is better Turbo ???
>>109474659sulphur dev plans to start training in a about 10 days, is waiting to fund it. And he has a much better dataset now that also includes dan / e621 on top of real stuffhttps://www.reddit.com/r/StableDiffusion/comments/1vgdqei/sulphur_3_is_looking_for_funding/
>>109474661they're the same, one is not compatible with comfyui format, the other is compatible
Why are people shilling Cache now? I thought Spectrum was best?
>>109474666The Qrusher one ??
>>109474670yeah?? duh? that one has comfyui name in it
fapped twice today, blasted a big load out both times.H3 is just too powerful. Honestly in disbelief over how capable it is. This should honestly kill off the porn industry, but I'm guessing the luddite movement is still big enough to reject it.
>>109474584best gen itt. catjack could never
>>109474691since when the luddites have even won a single war? technology has always advanced they can't do anything about it
>>109474702nuclear
>>109474669>anon learns that people are experimenting a 2 days old modelit'll take some time before the right settings will be found
>>109474669Poorfag hour that wants speed over quality.
>>109474691the output length is too short and waiting time too long to kill traditional pornfor gooners its getting good but for normal porn consumers, no
Created a loader node for H3 Turbo using the original .safetensors lora file.Supports gguf models and the sageattention node from kjnodeshttps://files.catbox.moe/pq4dcw.7zNode order.Model loader - turbo lora loader - sageattention patch - ksampler/H3 director
>>109474707not in france
>>109474691>fapped twice today, blasted a big load out both times.physically impossible for me to climax to my own gens as i know what theyll look like as soon as i hit "go"
>>109474344What does your setup look like as far as resolution/step count/scheduler/sampler etc. I have 12gb vram and 32gb ram and my outputs are a bit too blurry but I wanna know the lowest settings I can get away with, while maintaining decent quality
>>109474561oldfag here. how do you connect this node? kek
How good is R2V at mimicking Japanese voices? If I give it an audio sample of Japanese dialog and wrote some text in kanji or romaji would it be able to clone the voice accurately? If I gave it blowjob asmr audio could it do that too?
>>109474561How to connect sigma shift ?
>>109474588so i think the plan is to use the lora for rapid prompt iteration and then turn it off when i want to start collecting interesting kinos
>>109474728>>109474733are you serious? when you see a model -> model connection you put that between the two
>>109474715>the output length is too short and waiting time too long to kill traditional pornWhat are you talking about? R2V can stitch videos together natively. The sky is the limit.Also 10 second clips from H3 are already 100 times more arousing than any shit on pornhub. You can tailor it exactly to what works for you. I don't need a 1 hour goonerslop video to get horny.
>>109474724>he didn't write a big wildcard prompt
>>109474734you say that now but once you've seen the mock up you'll move onto the next thing
>>109474738but there's the>lora stack>Patch Sage Attention KJ>MiniMax H3 Mem Eff Sage Attention Patch>EasyCache or Spectrum Apply MiniMax H3 all connected to the model where would you place the sigma?
>>109474503>sage + cache + turbo lora - 4m01s>these settings (except the same 0.4mp I was using for the others) for turbo lora >>109474545>clanker sloppa prompt: https://pastebin.com/7WM03VH1The turbo LoRA works, but the quality takes a huge hit. Also, you will not be able to reliably make almost the same gen with this and then remove all optimizations to make a higher quality version of the same video. I think I'll stick with just the cache node for now and then when I like one I can make it look even better.
Can you guys imagine using wan2.1 now? Any of you care to try a prompt comparison it to see just how bad it is?I remember being amazed by wan2.1 last year, but looking back, goddamn it's so much worse
>>109474745>once you've seen the mock up you'll move onto the next thingwhat next thing? hard to be faster than a turbo lora
>>109474745>you'll move onto the next thingi had like 1000 videos of my fighter jet videos, i don't get bored that quickly
>>109474725See >>109473929The Turbo lora needs some more cooking still.>>109474748Right before the ksampler node
>>109474734just know what you're getting anon. It won't be the same >>109474755All of those used the same prompt and seed. Maybe a better turbo will come out soon.
>>109474762thanks for the spoonfeed :^)
>>109474755lil sis thought she was the main character
I'm minimaxxing.
>>109474774dishonored vibes
https://n.uguu.se/SjpZwpqF.webm
>>109474780uhhhhhhhhhh
>>109473929anyone who recommends ggufs are either trolling or retarded. they dont "save vram". Int8 streams the weights, you dont have to fit it all in vram. GGUFS are forced to fit all in vram and are about 2.25x slower than int8
2 days and basically all my gens have been porn and most of the rest is cringe waifu shit. Am I cooked chat?
>>109474780whoops didn't mean to upload that here
>>109474669see >>109474503and judge for yourself.
>>109474756>I remember being amazed by wan2.1 last year, but looking back, goddamn it's so much worseI still have my wan 2.1 gen, goddam they're bad and slow, but I was impressed too at the time, it was my first time making images move, it was like discovering fire lool
>>109474756
>>109474780hot
>>109474503how do you do these multiple outputs? Or are you stitching these together manually?
>>109474756ahh... wan2.1...yep, those were the days
>>109474815>>109474825
are you behaving yourselves?
>>109474835
>>109470435Can it render and keep rig bones?
dunno why it made the viewer so fat here all i mentioned was a white shirt
>>109474801Roger that, downloading the Int8 version now to test. I had started off by using the Int4 version but the outputs were garbage, I erroneously assumed the Int8 version would have the similar issues.Will report back.
>>109474854you know why
>>109474850>that endless yappingkek, that's definitely a Wan render
>>109474831I used to manually make ffmpeg commands. Now I just get a clanker to write the command based on the filenames. There's probably some workflow that can do it for me, but I don't do this sort of thing enough to justify learning how to set that up.
https://github.com/Comfy-Org/ComfyUI/pull/15334kijai made the vae process 2x faster with int8 convrot now
>>109474835>>109474850
this is the best model ever
>>109474865same. i don't code with ai but i still ask for commands
>>109474860int8 convrot is Q8 tier so you can have fun with it anonhttps://github.com/BobJohnson24/ComfyUI-INT8-Fast/blob/main/Metrics.md
>>109474872this thread is meant for AI generated outputs
>USE EULER + BETA instead of res_Multistep because res_multistep give disco lights. Wait what ??
I didn't know seedance could do this until now with the plus ultra min version but>take porn clip>gen a goth girl or girl of your preference with grok or your choice local model>swap them out in prompt>you can even edit the existing video and add camera cuts>you can even add extra characters watching and reactingwhat a fucking time to be alive man. it's like some cyberpunk bullshit https://files.catbox.moe/dwg40e.mp4
>>109474881for 4 steps, I don't have that problem with 8
i fucked up somewhere
>>109474885>inb4 transphobes start seething
>>109474891No that looks right
>>109474885>seedancelocal models?
>>109474904>http://127.0.0.1:8188/>Unable to connect>Firefox can’t connect to the server at 127.0.0.1:8188it doesnt work
>>109474885>>swap them out in promptwhat
>>109474904you leaked your ip retard.
>>109474911it actually will let you swap out the entire body of the person in the clip, including clothes.
frick
>>109474904holy fuck, i know /g/ is tech illiterate but this is insane, just leaking your ip like that
>>109474874setting up hermes on a dedicated box was a game changer for finally getting around to all of those backburner projects. Just let the model do it for you.
>>109474865ah that makes sense, thanks for the tip
>>109474900i should have kept it going but i also noticed a typo
im about to pwn that noob
>>109474919oh, it does v2v with i reference?
>>109474935pls dont how will i keep generating ludos
https://d.uguu.se/ZaJYlUGe.mp4
>>109474938part of the onmi reference with seedance. you basically just sayreplace the female character in @video1 with the character @image1
this turbo lora is excellent. one step closer to replacing ltx
Am I supposed to get a shitload of errors loading that 4 step lora? What loader node am I supposed to use?
>>109474943are you using video references, this looks familiar.
>>109474506>text enfuck off retard. it definitely makes a difference. You cannot simply say things like "fucked from behind" at get it to actually interpret it properly with the standard encoder. no one wants to describe their porn scene like it's a fucking clinical examination
>>109474950>Am I supposed to get a shitload of errors loading that 4 step lora?no, you have to download this one >>109474353
remember vace in wan?
i need to learn to prompthttps://d.uguu.se/vEzxQGgq.mp4
>>109474956my fully erect penis just slid in and out of your moms mouth multiple times
do turbo loras degrade the knowledge of the model? or is it just the seed variance that gets affected
>>109474953no, it's just an i2v prompt. You've probably watched a lot of porn an recognize what it was trained on.
>>109474971>>109474598
>>109474967I think it's pacing, I dunno what anybody else is doing but I'm rehearsing the scene in my head to figure out the timing
>>109474969>just saying bullshit because he knows I'm rightthanks you for your submission cuck
>>109474974that isn't what i asked
>>109474979>>109474588
>>109474981that also isn't what i asked
>>109474833Spectrum has a better quality but it's also slower, I will try the Provisional aggressive preset to get equivalent time and see if the quality is still betterhttps://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3#conservative-presethttps://files.catbox.moe/yvrs8h.mp4
>>109474875Genning with it now, but it's thrashing my ssd by using the pagefile instead of using system memory during inference (61% of 24gb)
Is the turbo lora better than the optimization node schizo stack?
>>109474996>>109474463
this vae decoder is too slow. someone needs to distill that as well
>>109474996the audio is definitely impacted by that lora, we have to let him cook
>>109475005>>109474868lol. All these anons asking shit already answered above
>>109475019>>109475019
>>109475014LOL tricked you into spoonfeeding me
>>109474885what provider are you using for seedance?
is /e/ right?>>>/e/3122899>You can get around that with Wan by genning 5-second clips then using the fun-VACE checkpoints to generate transition frames between then. I think the main advantage of this Minmax is that is seems to understand 2D animation. Wan only knows 3D and realism so it gives a heavy render-like bias to everything, and doesn't understand anime faces. LTX is the same but much worse.Still trying to make a case for wan over h3.
>>109475005truesaw this earlier https://github.com/Comfy-Org/ComfyUI/pull/15334
>>109474534kys you creepy fuck
>>109474868I put swapped the current VAE for this one, but my vids just returned as black
>>109475028Anon might have severe brain damage. H3 ref can do consistent characters WAY better than Wan + character lora. Just that is worth dumping wan for all time.
https://files.catbox.moe/34eqhb.mp4Better than the last version, I reckon.
>>109474868>https://github.com/Comfy-Org/ComfyUI/pull/15334I bathe thee gratitude and love Kijagod.
>>109475098kek
>>109475098lmao that was pretty realistic
>>109475092you prob don't have latest comfy + comfy dependences + CU130+
>>109474466FFS ANON... I might be wrong on certain understanding but a text encoder is just that, it has no gaurdrails, it has no samplers, it does not predict it just converts you text to numeric values called tokens that get sent into the image or video model. if it had any of those things it would be even excruciatingly slower. The guardrails are in the image or video model not the fucking text encoder. just fuck off, tired of your brain damaged shit since fucking 1 year ago.
>>109475027artcraft
>>109475127that is correct. But its retarded "ai influencers" constantly pushing this useless shit for updoots. The TE only acts as a translator and changing the TE that the model was trained on at all simply makes it diverge from what it was trained for = worse gens.
>>109474466and if you use an llm (big difference) to enhance a prompt then yes that will censor you and that is when you might want to use Heretic or abliterated versions of those llms to get past its safety filter for nsfw gens. But the text encoder versions will do fuck all and are a waste of compute, bandwidth and other peoples time.
>>109475110Is CU130+ basically a requirement for reference video generation?The last time I tried updating, most of my workflows broke
>>109475194yes. Otherwise its far slower and takes far more vram
>>109475201Okay, making a backup of everything and making the attempt I guess.I remember updating cuda means you also update sageattention, does someone have the link to the versions of each that match?
>>109475201Is installing it really this simple?
>>109475219if new comfy yea, otherwise you should uninstall the old torch / just remove the venv firstThen do those, then the comfyui requirements, then triton then sageattention2.2.0+There is a one click .bat somewhere out there that does it all
>>109475248Do I have to use “—disable-pinned-memory” in the startup when using cuda130? Showering across different guides rn.
>>109475262no. Maybe if you only have a tiny bit of ram
>>109475248If this is what my python package list looks like in stabilitymatrix, does this mean I already am using cuda130? I saw on reddit that the comfyui console usually shows a warning if you aren't using cuda130 and mine isn't showing that either.
>>109475218>>>chatgpt
>>109475337forgot pic oops
>>109475201well that sucks, i have a 3090, will cuda 13.0 make a difference? or will it only change for 40s and 50s?
>>109475408especially for 3090 as int8 is 2x faster instead of just 40% faster
>>109474381Why did they give this to us for free tho
>>109475427Turns out that hosting API services is crazy expensive when you don't have that much customers