Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109744947https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
Thank you for baking and respecting the customs of purity OP
>>109752199Thank you for thanking yourself like the petty little schizo you are
>>109752242Why do you think I'm OP?
>>109752245None of your (((concern)))
Blessed thread of frenship
She can't talk to us like thathttps://files.catbox.moe/shg476.mp4
>>109752323You still going on with that? You have been doing the same act for weeks now, it was cute at first but its stale now
>>109752323who is this fat asian girl with the silly hat i keep seeing
>>109752391Get new material
>>109752415Not when you bite every time mysterious ani defender.
>>109752420>>109752415
>>109752420>posting from Indiaevery time
Make em mad make em mad
the absolute state of local, sad
>>109752444Using the this general as a metric is foolish. I'm having the time of my fucking life.
>>10975239124/7 on 4chan, making sure you always bake like it's your job, proxying to ban evade and ofc an overweight slob, perfect self portrait schizo-kun.Also I always wanted to ask, what's with the the outdated model upscaling? Do you think if you upscale your slop from 2MP to 4MP that it looks better? Cause it still looks like shit, just in a bigger resolution
>>109752455The schizo is black. He doxxed that info years ago
>>109752469That would explain an insane amount to be desu
isnt debo black too. they are the only one that makes black tier gens.
>>109746493>RIP MiniConstruct 2026 - 2026I'm releasing a big update soon. Just finishing off Comfy API pipeline integration.
AGI IS HERE!https://www.businessinsider.com/nvidia-jensen-huang-agi-openai-astra-ai-2026-9>The chipmaking mogul took to X to congratulate OpenAI's team on the release of the company's newest and most powerful model, Astra, and made the notable statement that "AGI has arrived," a nebulous term that typically describes AI models that match or surpass human intelligence.
>>109752546meanwhile diffusion slop is still hallucinating blobs. we're still training loras like it's 2022 while API models can create an entire 3d model of a character from a single reference image.
>>109752546most likely a nothingburger to raise stock prices
>>109752577H3 is pretty good at using reference images. You basically don't need to use character loras at all anymore. Of course they'll always help, but it's not needed in most cases.
>>109752577llms + agents can draw better than diffusion which is funny
https://www.youtube.com/watch?v=akkwj9d943Ycan't stop watching this, it's strangely hypnotic
>>109752546If and when actual AGI is achieved, I highly doubt it'll come from current gen models. I think it will require an entirely new kind of LLM. Either way, some rich guy that has no hands on experience working with AI models is definitely not going to be the one to announce AGI. In reality it won't even be announced. It'll just happen.
>>109752483>>109752469>>109752455Doing this is why your rentry is in OP and the only defense you have is a rentry you wrote while seething that everyone laughs at btw. Keep being our drunken dancing monkey, the rentry on you is still being actively updated.
>>109752323ignore the detractors, this bit never gets old. mostly because i wanna fuck her.mind sharing the image reference you use for her, or is this somehow purely text?
>>109752577it's nvidia faultthey deliberately fuck up hardware market so nobody can do shit
>>109752666that's not nvidia's fault. any company that has complete dominance over the market would do the same. this is why monopolies and late stage capitalism are bad. nvidia is just playing the game, and winning, hard.
>>109752660schizo no schizing!
Can someone please tell me why anons are trying to remove the rentry links when everything within them are damning?
>>109752602it's clear diffusion is a dead-end. true AGI will be omnimodel, and a model with as much understanding of the world at a 5T LLM will of course make better images than a shitty 4b Qwen text encoder
>>109752796Will it be better than gemma prompt? Honestly I doubt it
>>109752796The truth is those models are going to filter anyone not over 24gb of vram but nobody wants to have that conversation.Honestly not to be a dick, most of the most annoying faggots in this hobby would vanish overnight if someone was strong enough to just admit it and let it happen.
What kind of hardware do you need to generate images and videos locally? I was thinking of bying a P40 24GB from Ebay, or an RTX 3060 12GBAlso, it seems like people are generating 512x512 images. Is it that these are initial generations which are then upscaled?Sorry for being a newb but i could not find the relevant info in the rentry links.
>>109752842You can use any internet search anon, I suggest you use chatgpt because both cards are going to be things 90% of anons don't use now.
>>10975284240 series minimum, 3080+, don't even bother with the P40's or anything even remotely that old, mikuboxes went out of vogue years ago.
>>109752857They were never good in the first place
>>109752855Thanks, I didn't want to use chatgpt but i'm reading an article about it (but it seems AI generated too)
>>109752871So much time has passed and the tech is hard capped at the 5000 series right now. Most anons upgraded and even with top end consumer hardware there are odd limitations based on the model.
>>109752864mikuboxes were passable in like 2023 to maybe mid 2024, but you could argue they were never good sure.
>>109752380>>109752381I velly solly!https://files.catbox.moe/mswbcw.mp4>>109752665The prompt is literally: Mika is a 40 year old MILF type Japanese woman. She is slightly overweight. She is wearing a tiny colorful bikini. She is also wearing a tiny rainbow colored yarmulke with a little spinning propeller attached.She's different every time. You fell in love with my writing, bro.
>>109752901I see. Are the limitations due to memory? And do you guys plan to buy decomissioned datacenter GPU accelerators in the future to continue genning? What do people use for making AI videos?I'll be shameless here I wanted to generate hentai videos>>109752918Apparently it gets single-digit tokens per second for any model that fits in them. I really don't think that's meaningfully useful
>>109752918They were on the way out from jump street, I just remember the amount of shit anons would talk and act like they were doing something big brained back then.You can only do shit like that once the tech stabilizes which is clearly not the case so early in the lifecycle.
>>109752932the preemptive fucking cackle i did seeing pyle in that thumbnail, holy fuck bro. of course she's just a fucking prompt, very nicely done man. will be inserting her into my favorite movies today.
>>109752697posting this kind of screenshot like they mean anything shows you're either a newfag or gaslighting
>>109753058No worries friend
>>109753058He's going criticalAgain, all of his efforts are moot when you can just read the links in OP. Let him shit his pants and cry, nothing has changed and nothing will change for him.
>>109752842Images are generally a lot easier to generate. I would assume even your current hardware, if it's not 10 years old, will be able to generate some animu images with Anima.For videos and especially with Minimax H3, I feel like you'd at least need a 5060ti 16GB to make it bearable. WAN2.2 or LTX can run on a lot less if you're fine with either no sound, or questionable image quality.
newbie anon from previous thread. I got flux1-dev-kontext_fp8_scaled with prompt:> Modify the image so she has very big breasts and her cleavage showingneed help to make them bigger please?
>>109753128Flux is cucked, no seriously.There's a reason why they went quiet once H3 came out and are nowhere to be found
>>109753128look loraholic loras up on civitai, these are the only non slopped slider loras on civitai, they easily let you change the breast size without affecting anything else, godspeed coomer
>>109753128>>109753149ignore what I said, I thought you were using Krea2 not trash
>>109753128you want a model that isn't censored and preferably trained on nsfw stuff
>>109753159He's new, can't fault him for getting scammed by SAI But German.I'm still trying to figure out why they would continue the practices that destroyed that company after leaving that company. It's not even about the direction of the model outside of them making the biggest inefficient piece of shit that nobody will bother working on or finetuning.The performance of the model does not warrant the size or the fact it's slow as fuck even on top hardware
You guys just said that Mikuboxes are trash and P40s are useless but then why did P40 prices go up like that? most articles say they're 200$ but since those were written it seems to have gone up by about 100$
>>109753128>fluxoh no no no no
minimax ref model being able to pretty okay-ishly clone voices and do inflections made me think, what are the best local voice gen models? Been a while since I looked in that direction
Who is gonna tell him we're still in 2022 when it comes to audio gen?
>>109753211Because much like covid everything trends up because they know people are more desperate.I'm still going to shit on faggots that bought Mikuboxes and the retards that actually listened to those faggots.
>>109753211retarded anime posting faggot, everyone knows the prices went up on everything. even fucking ddr/2/3 went up.
>>109753231It's so weird how there's now a video model that can learn characters lora-less with <10 reference images and clone voices with 15 second audio clips but we don't have image and voice models that can do the same. Feels like a step was missed along the way.
>>109753255Stability mindset caused severe damage to western models. Doesn't help many of their Ex employees are in these companies and continue the self destructive practices.
>>109753255>but we don't have image and voice models that can do the sameWell yeah image models can't do voice, anyways local video and image gen is on api level if you have the hardware, it's just audio that's 4 years behind
>>109753255I'm with you on image models, but we definitely have voice models, that work quite well with voice cloning.
Forget the base model many of these makes make models and refuse to give it support in regards to controlnets and other goodies and are surprised when nobody adopts them and they go to shit.They also expect people to make shit for them but they also releases gimped models on purpose and cry when it's ignored.
>>109753232Please to understand, we are poorfags>>109753254I know for a fact that DDR3 didn't go up too high. i'm going to buy 24 more GB of ram when I go to my parents' house where my old DDR3 computer sits with 8GB of ram.
>>109753293Those models are basically just advertisements for their API, they bever even intend them to be wider adopted
>>109753255Who was it again that said they're gonna release a model with a 1 image to lora ability? Member something like that but of course it never happened
>>109753272>but we don't have image and voice models that can do the samehurr hurr, you forgot to say that audio models can't do images>image gen is on api level if you have the hardwareModel? It has to be able to do reference, without lora training.
>>109753285>but we definitely have voice models, that work quite well with voice cloning.Which ones?
>>109753293sorry but we're way past the era of controlnet. nobody bothers because a model that still needs controlnet is outdated shit not worth developing tooling for. yes that includes anima and krea. edit should come built-in just like Klein introduced. you can provide openpose and depth to any API model and it works fine.
>>109753320And they end up making fuck all in the process. What gave Stability the edge in the beginning was due to having open models that the community actually improved with tooling. Now there's zero incentive and they make anything. With the current parts crisis this is the perfect opportunity to just release shit and allow people with the means to advertise the models for free while getting free labor.Krea2 almost had it right until it became clear that the base model and the turbo were not made from the same parts. Such a easy fucking win they pissed away for no reason at all.
Can you take your schizo back? He's shitting up /adt/ again
>>109752546>Huang then added: "Keep buying our overpriced GPUs, you fucking faggots", and everyone clapped.
I cant search nsfw on civitai, i cant find options to enable adult after logging in
There's an anime called Soukan Rensa which I like but there's a scene that's too short. I want to extend it and add more scenes like that using the same characters. How would I go about doing that? It will probably trigger the safety mechanisms of all models though.
>>109753316if a 2.6x price hike for isn't "too" high, then pray tell what the fuck isddr4 is the same price
>>109753372Fair point. They should always target for these models to hit 90 tier card vram at fp8 and they could easily reach those goals
>>109752546why won't graphics card Jackie Chan fix the Nvidia drivers for linux?
>>109753380civitai.red
>109753378Please fuck off you needy faggot, we're actually having a discussion
>>109753396Can you please stay here? kthxbai
>>109753389there should be a Medium variant trainable base model that can run on 24GB, and a fuckhuge variant that can run turbo on 24GB. quanted/turbo giant models are simply better than shitty 2-4b base models
>>109753372krea 2 shipping without editing capability is such a mother fucker of a monkey paw curl. Great model, but censored (at least the chains can be reasonably broken) honestly edit should be a bare minimum at this point.
>>109753362>hurr hurr, you forgot to say that audio models can't do imagesYou're the retard unable to articulate yourself correctly, reading that back sounds almost like you had a stroke typing that.>It has to be able to do reference, without lora trainingWhy, what does that have to do with anything? Wait let me answer for you, nothing. You're being a dishonest faggot for some reason. Matter of fact if you can run kimi at home you know? No Lora's needed, promise
>>109752182>mfw Resource news09/07/2026>Sol-H3 Speed-of-Light MiniMax-H3 on an 8× NVIDIA B300 Blackwell Systemhttps://nvlabs.github.io/Sana/Sol-Engine/Sol-H3>MageTrail: Danbooru/E621 Full-Finetune of MageFlow 4B https://huggingface.co/RicemanT/MageTrail>Learning 3D Editing without Paired Supervision via Generative Prior Distillationhttps://github.com/thiamine128/PriorEdit3D>UniMate: One Unified Model to Animate Diverse Skeletonshttps://linzhanmou.com/unimate>VICAL: Vicinal Consistency Alignment for Long-Tailed Visual Recognitionhttps://github.com/FlamieZhu/Vicinal-Consistency-Alignment>WorldSculpt: Generating Compositional Worlds from Grounded Videoshttps://alaya-lab.github.io/WorldSculpt>AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memorieshttps://zunwang1.github.io/AnchorWeave>ComfyUI-InpaintCanvas: Krita-style inpainting without leaving ComfyUIhttps://github.com/DenRakEiw/ComfyUI-InpaintCanvas>ComfyUI-Ref2VA-VSA: Ultra-Fast Character Video Generationhttps://github.com/Kablex/ComfyUI-Ref2VA-VSA09/06/2026>ComfyUI-Viggle-Animate-H3https://github.com/Saganaki22/ComfyUI-Viggle-Animate-H3>uncomfymcp: MCP server for generating images on ComfyUI from a chat client.https://github.com/aschet/uncomfymcp>MiniMax H3 Semantic Bridgehttps://huggingface.co/speach1sdef178/MiniMax-H3-Semantic-Bridge09/05/2026>Intern Lumina U2: Multi-Codebook Diffusion LLM for Omni-Visual Understanding and Image Generationhttps://internlm.github.io/InternLumina-U2>Musk’s xAI loses court bid to block Minnesota's AI ‘nudification’ banhttps://www.reuters.com/legal/litigation/musks-xai-loses-court-bid-block-minnesotas-ai-nudification-ban-2026-09-04>Add Sparse Attention nodehttps://github.com/Comfy-Org/ComfyUI/pull/16072>MiniMax-H3 FL2V 8-Step Motion Enhancerhttps://huggingface.co/rzgar/minimax-h3_fl2v_8Step_motion_enhancer>Qwen3.8-Flash-Next-NVFP4 https://huggingface.co/nvidia/Qwen3.8-Flash-Next-NVFP4
>mfw Research news09/07/2026>Joint Alignment and Distillation for Video Generation via Sample-Guided Distribution Matchinghttps://arxiv.org/abs/2609.04283>ReaDiT Guidance: Control for Image and Video Generation using Diffusion Transformer Featureshttps://arxiv.org/abs/2609.04649>RefDiT: Local Attribute Guidance in Reference-Based Image Generationhttps://arxiv.org/abs/2609.04976>Importance-Aware Low-Rank Distillation of Diffusion Transformershttps://vislearn.github.io/SVDtrunc>Measured Sliders: Learning Continuous Controls from Differentiable Image Measurementshttps://arxiv.org/abs/2609.05234>PAPT++: Risk-Aware Adversarial Tuning and Generation for Single Domain Generalizationhttps://arxiv.org/abs/2609.04837>AngelFingerprint: A Traceable, Explainable, and White-Box Stealthy Watermark for Text-Guided Image Editinghttps://arxiv.org/abs/2609.04709>Step Back to Move Forward: Reflection-Aware Preference Optimization for Visual Generationhttps://arxiv.org/abs/2609.04282>LensStyle: Learning the Optical Aesthetics for Controllable Stylized Lens Effect Renderinghttps://arxiv.org/abs/2609.04939>Intrinsic Temporal Adaptation of CLIP for Partially Relevant Video Retrievalhttps://arxiv.org/abs/2609.04800>WeAgent-MMGenEdit: Full-Stack Recipe for Multimodal Agentic Image Generation and Editinghttps://arxiv.org/abs/2609.05171>FailSAE: Towards Interpretable Failure Prediction for Vision-Language Models via Sparse Autoencodershttps://arxiv.org/abs/2609.04276>Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inferencehttps://arxiv.org/abs/2609.05275
>>109753434*if you can run kimi at home you can do everything the api's can do, you know?
>>109753434>Why, what does that have to do with anything?...that was literally the thing I was talking about? Right here in my post >>109753255? I said>that can learn characters lora-less with <10 reference imagesthen you replied to me boasting about your unnamed models, logically this must mean your models can do what I described? If not then why the fuck did you reply to me?
3 editing models have been announced.minimax H3 Image, flux3, and krea3.but nothing ever happens.all I need is a powerful editing model with up to 2 MP.
>>109753447>>109753438
how do you set this node?what are the correct settings
>>109753409The limit is probably 16GB, if you need 24GB to run or train loras it severly limits the userbase which in turn limits community development as most people won't bother to work on something that barely anyone can use, and it'll stay like this for a while with the current hardware crisis>inb4 vramlets etc
>>109753458I literally told you kimi can do that you faggot, do you read?
>>109753467I can't wait for anyone of them finally releasing their model already. So many things I wanna do with H3 Ref2VA hinge on me being able to create good starting images.
>>109753493then why did you ask "why, what does that have to do with anything" instead of just saying "yes, it can"why get bitchy about the one requirement I was looking for in the first placeand I've never heard of this kimi before which leads me to believe it's snake oil, and since now I know you're a temperamental bitch I know asking you for proof in the form of a good image will yield nothing but more bitching so I'll just start ignoring you now
>>109753438>>Sol-H3 Speed-of-Light MiniMax-H3 on an 8× NVIDIA B300 Blackwell System>https://nvlabs.github.io/Sana/Sol-Engine/Sol-H3could this work on consumer hardware?
>>109753517No
>>109753509>I've never heard of this kimi therefore it doesn't existjust rope yourself at this point, holy fuck you're retarded
>>109753509>I'm completely unfamiliar with the landscape, I didn't even know there was an AI other than ChatGPT. I enjoy bitching and whining and actively maintain problems so I can continue to bitch and whine about them.
>>109753492I think turbos have a place specifically because of that. Also quanting models help as well.They are going to have to accept a slight dip in performance.
>>109753504the only sure thing is Flux 3 Dev, and unfortunately, that’s going to be heavily censored and run slowly.even though it’ll definitely raise the bar for editing.butt they’ve been using the words “weeks” and “months.” So it could still take forever.with minimax, I’m worried about chinese culture.as for krea3, there’s just one mysterious screenshot from an alleged employee.
>>10975351790% of that schizo spam is worthless. Please don't engage. He's giving llm announcements for a model that can't even run on most of /lmg/'s rigs atm and is confirmed shit tier due to how nivida handles those releases.
>>109753563I can work with a censored model for my usecases just fine. I just think Flux3 will be terrible at anime stuff again. Would much prefer one of the other two, but I know they're even more unlikely to release with editing capabilities or at all.
>>109753438>>109753447Hey debo why does your "stable diffusion" general nearly die every single day at the same time? And why are the only "anons" posting there always there at the same time? I've never seen this behaviour in any other general
>>109752469Knowing one poster is black isn't as exciting as knowing the other guys name, his fathers name, his fathers place of work... The second guy was actually doxxed (by his own doing)...
>>109753579fuck off nigbo
>>109753669He's in a perma state of kamikaze. >>109753688Mention Ani's lust for BBC attached to anime characters and his entire own falls apart. He makes sure that the cocks are extra dark too.
>>109753688I mean he has a wix site also with his full name in the url with address, phone number, "cv" (lol)...I don't think he cares anymore
>>109753796With his /d/ post I don't think so either. He's mindbroken and just trying to drag everyone else down with him.
Nobody fucking cares. Kill yourself already
>>109752469>The schizo is black.? i dont think sasori would make those kind of posts hes annoying but mostly harmless i believe
>>109753810Or you could just stay in /sdg/?There is a thread that is perfect for avatar fags like trani.
Throwing a fit each time a certain user posts. Maintain Thread Quality, indeed.
fucking H3 Cache node is causing so many fucking problems lately. Why can't they just figure their shit out? It's probably the best time saver while still maintaining quality, but it is so incompatible with different things. Fuck!
Brainlet here. Is using the PDD 8-step optimizer for Minimax H3 incompatible with using coomer loras? Every time I try to add a lora on top of it the video disintegrates in quality
>>109753958You answered the question yourself, but unless the lora specifically states turbo lora compatibility assume none
>>109753958https://huggingface.co/spaces/multimodalart/h3-acceleration-arenabecause that lora is trash
>>109753858stealing this idea
>>109753981thanks i'm new to this. so when the loras all recommend 1.0 weight i need to divide them up if i'm stacking them
Bros I'm a 3060-let. Can I use Minimax H3?Also I fucking suck at using comfy. Has anybody tried using Claude to slop out workflows?
>>109754039Yes you can run H3. If you're not feeling comfy then just download Wan2GP Desktop Launcher and go have a good time.
>>109754066this, alternatively..>>109754033if you run 4 turbo loras each need to be at 0.25 for maximum fidelity
>>109754039If you have a reasonably up to date version of ComfyUI, just use the inbuilt Minimax H3 template. Try running it at like .2 or .4 MP and you'll soon know whether or not you can run it.
>>109754039Honestly it'll be easier to sit down for like an hour or two and learn the default MMH3 workflow. it's not that robust. you can ignore like 90% of the nodes. i have a plug+play one where I highlighted just the stuff you need to tweak and hid all the other crazy nodes if you want
>>109754039What do you mean you suck at comfyui? Just download the zip file, unzip it, start using the .bat file, then load the default workflow from the integrated template loader and see if it works
https://litter.catbox.moe/pzodrc8sxksl29li.mp4
>>109754024What a retarded fucking video
>>109754143These people are tech illiterate anon.
i cant find a definitive way to create h3 minimax ref2video prompts, grok used to work but its kinda retarded sometimes, local LM takes a lot of memory and is also kinda retardedit's a hassle honestly
>>109754024>>109754156liar, it made me chuckle. I completely missed it cause I thought it was just the retarded png
>>109754174Shame because gemma is the best model for it. Sounds like a skill issue desu
>>109754174what do specifically need help with? use grok/chatgpt to create a reusable system prompt off of the official minimax guides and iterate on it. that'll produce more stable guides for whatever ai you're using to create the prompt.
>>109754169It's just that, why do these people insist on doing twenty extra steps to do the thing that is already as easy as unzipping a file to do?
I made an infinite video workflow, it's honestly not as exciting as I thought and ended up more of a proof of concept, but still fun. I get roughly 1s of video every 22 seconds with the 8 step turbo lora at 0.98MP
>>109754196Correct answer. I had gpt-5.5 shit out an h3 prompt writer skill for pi from the H3 docs when it first dropped. When I encounter issues with the prompts it writes I work with it until it's where I want it, then have whatever model I'm fucking with revise the skill to incorporate the new knowledge. 5.6 Sol is an incredible coom-artist. It works extremely well now, very consistent, I don't even bother to copy and paste the prompts - the skill decides on all the settings to use and just runs it for me. So my workflow is now "gimme a 5s video of this chick @slutphoto.jpg swallowing a tennis balls then painfully firing it out of her asshole" and it does the rest. It's fabulous.
>>109754247that is very fastwhat other speed cheat are you using?
>>109754275That is only ck kitchen and the turbo lora at 8 steps, but the turbo lora already is a huge hit to quality. The trick obviously is that it generates 5 second videos, which are the fastest to gen. Every second after 5 scales gen time exponentially since the knowledge context gets larger, meaning 10 seconds don't take twice but three or four times as long
gday choomsfiles.catbox.moe/woamdy.mp4
>>109754156
How much VRAM do I need to run H3? Is 24GB enough? I don't want to be part of the permanent gooning underclass.
>>109754295yeah it's unfortunate. i wonder if we'll ever figure out how to manage context during generation, like just continue on with needed. might be a dumb question though2 seconds - 1m12s5 seconds - 4m40s8 seconds - ~10m10 seconds - ~15m15 seconds - 35m
>>109754398yes i used to gen just fine on 16gb vram
>>1097543986GB is enough if you have at least 32GB of system RAM to back it. More VRAM you have, less system RAM you really need, but more of both is always better. If you have 24GB of VRAM then you're extremely well equipped for H3 and could probably run it without no issues with 16GB+ of system ram.
>>109754196the hassle comes from trying to make a short movie with interconnected scenes. the process of telling Grok or a local LM "now make a continuation scene with these details" and then adjusting it forces you to rewrite a lot of the prompt every time—like the summary, the detailed description, the audio, and so on.it gets tiresome if you want to do more than a little 15-second video, it feels like a slog sometimeswhen i used to do this with grok imagine the pure text approach was much more straightforward but sadly that piece of shit is censored to hell now
>>109752182How Many Turbo loras we got for minimax ? I didnt follow the news anymore
>>109754412Well it all depends on how much support fal gives us, but I doubt we get anything like that until minimax h3 max releases for local, but it's already released as api yesterday so there's thatVideo with a bitrate over 10kb/s:https://files.catbox.moe/tdwu28.mp4
>>109754471A few dozen, most are useless trash. Worth trying currently: larryvrh_v4_step600_ema, plaguekind_parasyte_turbo, lightx2v_fl2v_8step_v10_native.
>>109754428>>109754430Thanks. I have 16 GB VRAM/32 GB system RAM now, so I guess I'll just try it. I thought I had to fit it all on VRAM.
>>109754431continuing scenes off of ref2va is not going to work. recommend I2V. ref2va redraws your subject. i2v is tailored to start from exact.as for the prompting thing... yeah, but that's going to be an issue for a very long time. solve it with genAI
>>109754825nope. i have a 5090 32gb vram and i still spill over to system ram. that's going to be what it's going to be.
>h3 eros max can't do genitalslame
>krea 2 can't pull off the most basic action poses>meanwhile i can just throw the image of the subject into h3 and ANIMATE 10+ seconds of that character doing that pose/actionseriously what that anon said earlier in the day about a "few skipped steps" i think is an understatement, where the fuck are we that a multimodal BTFO's everything? Maybe that lecunny faggot was right and not just about LLM's.
>>109755039i want my 9 seconds back
>>109755039I don't get it
ok this seed kinda bad >>109755039hold up. I have an idea
>>109755136I get it now
>>109755242explain
>>109755253>magictrick.mp4>nothing happens in the video>video disappears
Is it better?It's hard to do fast action with just 24fps
>>109755268>>109755360You should've just ran with my explanation anon, motionblur is not a magic trick
>>109755360to not just shitpost, you could've made the plate two colors or give it a engraving, that way it's easier to understand that it's being flipped without showing her privates
>>109755360>>109755487OOOHH that's what he's trying to accomplish? based retard anon, yeah, make the plate's sides distinct you dingaling.
i want to understand obsession this fag has over aang and zuko? how many updates and versions must he release?https://civitai.red/models/2773267/aang-and-zuko-krea-2-lora?modelVersionId=3122618
it's funny in my head
>>109755671what exactly is the appeal of this..?
>homosexual male doesn't understand the appeal of a woman standing half naked with only a plate in front of her and she flips it
>>109755607Is there now any H3 LoRa that doesn't degrade the output?
>>109755744lolno
>>109755711We have a mental health crisis in this general
>>109755733I'm into more dirty stuff.If videos like that turn you on, okay, why not?I'm the last person who should judge someone's sexual interests.
Whatever happened to that comfy kitchen update that was supposed to give more gains for H3?
>>109755778So.. miniconstruct-man what's next on the menu?
https://litter.catbox.moe/v99sqxfzdkawm896.mp4
>>109755818he added sparse attention in a commit
>>109755839Is the music/drumming generated in H3 without reference?
>>109755852yep!
>>109755839bro fix your aspect ratio
>>109755840Is it in comfy main yet?Nobody is talking about any noticeable gains, just wanted to make sure before pulling.
>>109755839Aspect ratio?
>>109755855working on it, you'll have another video in like 20-30 mins most likely.
>>109755821huh?
>>109755821wrong anon. that guy always spams those shitty zelda webms with a billion scene cuts
>wrong anon. that guy always spams those shitty zelda webms with a billion scene cuts
Dunno what went wrong with the quality here, but at least the aspect ratio is ok nowhttps://litter.catbox.moe/o9psf5pjv4t4pcrq.mp4
OUGHHGHG I'm cooking cinematic, hdr10 certified, film grain ingrained, christopher nolan aah kino. But as all good things this will take a while
It keeps going UP. wthhttps://www.bhphotovideo.com/c/product/1898513-REG/pny_vcnrtxpro5000b_pb_nvidia_rtx_pro_5000.html
>>109756111You should have prepared before this started everyone was ringing the alarm bells back then.Don't let that distract you that those cards were a bad deal from jump street anon.
https://litter.catbox.moe/y333eyrvwvdctf2n.mp4
>>109756111They've always been $7k~$10k. I'm not sure who the market for them even is. AI companies aren't using RTX Pro's. They're using H200's.
>>109756213oh wait that's a 5000, not 6000jesus christ
>>109756213Honestly a scam card for the whole series, never enough to do enough while being 2-3x more while not even being close to the 5090 in raw performance.>>109756254Completely worthless card too, I feel like it was preemptive gambit but even then why not just run 2 5090 for a little more and get twice the vram. It's not even like the performance is better.
>>1097561335090s don't fit in my case. I couldn't care less about gayming, I am tired of waiting half an hour for H3 gens. Regardless it is my goal to get one eventually, one way or another
>>109756280what card are you using?5070 ti is decent enough, use for a year and sell it.
>Thread status: HIGHLY INORGANIC DANGER!
https://litter.catbox.moe/flr86pjsmxcnzxiv.mp4
>>109756280Even with inflated prices 2x 5090 will offer 64gb of vram around the same price, just get a new case
>>109756273yeah i know it's worthless. RTX PRO 6000's used to be $10k. the fact a shitty RTX PRO 5000 is that high now is just insane. Last I remember they were around $5k.
can i request some muscle mommies pls
>>109756309too bad it didn't sing in sinatras voice
>>109756298>5070 tithat is what I have, managed to find one that was 288mm. I want the sexy slim RTX Pro cards
>>109756311this. you should buy the biggest case you can afford anyway for maximum airflow and upgrade capabilities. this is another reason my gpu temps stay below 60C even at 100% usage.
>I am tired of waiting half an hour for H3 gens.>mfw spent an entire year waiting 1 hour+ daily for 5 second wan gens
>>109756344why is she so comically tall jesus christ
>>109756325
>>109753128Those are big by American standards
>>109756311That's a lot of power to burn on worthless slop iterations. Aren't the Pro cards more power efficient
>>109756389Are you unable to undervolt?
>>109756415Kek nice
>>109756344this is neither muscle>>1097563695.5/10
>>109756445*nor mommy (sorry i'm drunk)
>>109756452
>>109756415are you unable to let go?
>>109756502that's a way better MOMMY LET'S GO
>>109756543He is unable to do a lot of things but he can seethe like no other
https://litter.catbox.moe/z6xyfyezqzc6rybt.mp4
>>109756618her dystopia has turbo lora frywhos the real loser
>>109756543>>109756570Notice how they cannot make anything interesting and funny and only seethe.Much like the schizo who is all talk so are they.....if it even is a small group of dents
>>109756639You should probably visit an eye doctor anon
>>109751345 kino
>>109756673you don't see the fry? its literally everywhere
>>109756724No I don't see it schizo, here is the exact same thing but actually using a turbo lora:https://litter.catbox.moe/5s6nuuvcjqqjswey.mp4
What am I doing wrong?I want to add specific things that use the fedor LoRA, but when I enable it, I can't generate specific fonts anymore.
>>109756618one of the cars fuses into another one.
Guys, I'm seriously considering just dropping ComfyUI piping from MiniConstruct.So sick of trying to get this to work, and I don't even care about the feature myself, I only added it cause people here requested it.
>>109756344>Hey let's post some pedophilia on /ldg/. My heroes the Japs love it so it must be OK.
>>109756776Beats me, I like your car though.
>>109756839>>Hey let's post some pedophilia on /ldg/. My heroes the Japs love it so it must be OK.
>>109756776try using the kroma lora, it has a better dataset
>>109756822it's legit useful to have all the media + prompt sent to comfy instead of manually doing it yourself. mine is done the lazy way though and only works with my specific workflow. trying to get it to work universally sounds like a pain in the ass.
>Hey let's post some pedophilia on /ldg/. My heroes the Japs love it so it must be OK.
https://litter.catbox.moe/h5tcy1tllgqp57jb.mp4SHUT THE FUCK UP!
>>109756932oops fucked up my aspect ration again, re-genning.
>>109756906I haven't completely given up. Still making progress.It will require the user to set up a Minimax workflow (like comfy template), export it as API, upload it in the MiniConstruct Comfy/Execution settings, then map the bindings in the same settings window (H3 output to Comfy input text, duration to PrimitiveFloat(Duration), Picture 1 to `Load Image 143', etc). For maximum flexibility the comfy workflow should have existing connections for all the available image, vid and audio ref slots (h3 states a maximum of 9 ref images, 3 videos, 3 audio tracks, and a maximum of 12 refs).
Why does the ref_video_0 want an image output? I'm using the basic load video and can't pass it in there
So I just learned about freetoken. Is it as big a deal as it sounds? Running deepseek v4 flash on a 3090 sounds pretty nice. Is there discussion on implementing this methodology in other software like lm studio?
>>109756953
>>109756963you need get video components node or something
>>109756776Fedor fucks up prompt adherence.
>>109756968 How will that change the h3 node to not want an image input in the video slot?
>>109756979nigga i dunno but thats what everyone says
>Penis Lora>It makes the girl has futanari and penetrate male crotch insteadWHY LORA TRAINING IS SO FUCKING HARD WITH MINIMAX ???????
>>>"""penis lora"""
https://files.catbox.moe/sdeh8d.mp4Voila!
>>109756986whats wrong with that outcome?
>>109756979ref_video_0 takes images because videos are just a series of images. that input slot is taking all the frames from the video as images. they do it this way i guess because its less error prone than taking entire video formats as input.
>>109756993I have no idea. Probably because the girl in upside down position
>>109756938I'll tear that ass up
>>109757018weirdo
>>109757018no, i don't think you will
>>109757025Nobody told that old bitch to have a ass like that anon
>>109756992Staple to staple
>>109756389RTX 6000 pro using about half the wattage of a 5090 keeps up in inference pace so yes.
omgomgomgomgafter a week straight of wasted tokens and aimless refactors, a complete workflow was finally producednow its time to spend 5 hours to see how shit this gen is
>>109757056......How do retards like you two exist
>gen is looking awesome in the preview>finally finished after 3 hours of iterations>excited>post it in my favorite general>0 (You)'s>refresh>0 (You)'s>refresh>0 (You)'s
>>109757075Post the gen
>>109757075Either you got it or you don't. You need to have that dog in you anon.
>>109757064this might shock you but I bet less than 5% of people in these threads know about torch compile. there are retards everywhere.
>>109757062total gen time was 80+min, which is actually low compared to the 250m+ failure runs I've discarded
>>109757095You're right...This is part of the reason why I believe a safe for work AI board should exist it would destroy multiple birds with one stone
>>109757064From my own anecdotal testing the results speak for themselves, no retardation required. RTX 6000 pro is just more power efficient using the same environment setup on the same workflows compared to a 5090. For using two 5090 cards that goes out of scope for my assertion.
>>109757086ill just make something better>>109757093unfortunately most generals are very anti-ai unless its some epic meme
>>109757124Are you low IQ or something? You're bitching about the price when you can come out with more net vram and at most slightly more power draw if you take the time to learn how to power limit. Actually I'm pretty sure you can't even afford the 5090 before the price jump so I'm trying to talk reason into some jackass without any common sense.
>>109757124Are you getting the Max Q (Lol) model confused with the actual full power card?Those cards are highly situational and cucked. Also what is undervolting?Are you unable to undervolt?
>>109757164Why are you engaging the rtx pro 5000 schizo?This is the second time he's doing the exact same thing
>>109757187We need to stop misusing the term schizo and just call him a jackass. I bet he couldn't even afford a 5060 if he's that stupid
>>109757189There is no correlation between intelligence and income, some of the dumbest pieces of shit I've ever encountered were very well paid.
>>109757189I call him schizo because he is literally samefagging with himself trying to make this retarded bit appear organic.this is him, that is why I call him schizo, those "two" retards started posting at the exact same minute in this thread>>109757124>>109757182
>>109757196You know you're replying to me twice in that post right?
>>109757205I wouldn't be surprised, you're severely mentally ill
>>109757207you're going to trigger him into a melty and we're all going to have to suffer
>Misusing basic terms>random time wasting argumentYep I should have just stuck to what the rentry said.Find a better use of your time, I was trying to help someone and have them save money vs burn money on 1 shitty card
https://files.catbox.moe/1vv58y.mp4
>>109757250
>>109757252>Page 1>Not at bump limitGot a hot date to get to or something?
>>109757261I think he's triggering whoever of the dynamic duo that tried to derail the thread and false flag. Pretty based desu, because they seethe at two generals now.Inb4 singular anon cope that has failed for 3 years running
>>109757010Got it working and now got boob2boob motion, very nice :thumbs_up:
>>109757261Just the usual schizobake, he's afraid someone will make a thread before him and not put in his low effort slop. Just a narcissism thing
>>109757282There you are misusing terms again. Is it a coping mechanism when the links have been in 98% of all the threads?Every day you screech and every day you lose. Don't you ever get tired of this?
>New thread collage has two gens>Schizo talks about his discord buddy links that nobody gives a fuck about:^)
>>109757282>he's afraid someone will make a thread before him and not put in his low effort slop.different guy
He thinks if he says it enough his fanfiction will come true despite this being the norm since the thread's creation only a dent would think this would work.
>>109757303You have problems dude
>>109757310Not as much as the guy that cries all day makes pointless bait arguments and can't make coherent gens. I don't think you understand that the majority of the thread don't fuck with either of those two rentry retards, If they did /sdg/ wouldn't be a desolate laughing stock.Even the anime thread put up a ward. To blame one person for the state they are in really shows a delusion that can only come from mental illness
>>109757303yeah i just ignore him mostly
>>109757164Cost doesn't have anything to do with power efficiency for my assertion. The RTX 6000 pro does the same work a 5090 does using less power, about half. That's it and that's all I'm saying. Nothing more.
>ORGANIC ACTIVITY HAS REACHED ALL TIME LOW. ABANDON! ABANDON!
Is there a thread that doesn't have the same person fighting them self for week upon weeks?
>>109757336He doesn't realize it due to his slowness, I mean he's behind 90% of the post in his containment thread. Also his ritual post make people feel sorry for him more than anything
>DANGER! DANGER! STOP SAMEFAGGING IMMEDIATELY OR THE THREAD WILL COLLAPSE
You're really not good at this
The reflections of this model is something elsehttps://files.catbox.moe/d31r6m.webm
litterbox shit the bed again?
>>109757535>>109757250
I just tried the new audio patcher model and holy shit, absolute game changer if you use it with a turbo lora, even at 4 steps it sounds crystal clear. Too bad catbox is down, can't even show it
>>109757722what audio model, post the link as well cos im tired of shitty voices
>>109753388I said ddr3, nigger