Discussion and Development of Local Image, Video, and Music ModelsPrevious:>>109712324 (Cross-thread)https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/animanon
>>109717308>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
2nd
>https://huggingface.co/spaces/multimodalart/h3-acceleration-arenait's up
euler/sgm_uniform 12 steps, turbo lora at 1.0First run takes 23.44 secondsFollowing runs roughly 14 seconds
>Cry about something the previous day>Fuck with OPThis is why your camp is in the position they are in, you're will to harm users by allowing that no life schizo to distribute malware.
>>109717497You can kick and scream all you want even if I'm not here the community has decided, the links stay and vandalizing the thread every fucking week for almost a year now won't change that.
>>109717505Oh god, I didn't realize we were dealing with real schizos. This was my first thread in a year or so, and I was just making the reality I wanted to see. You can tell by the bad copy/paste on previous section.Good luck with your demons anon.
https://huggingface.co/spaces/multimodalart/h3-acceleration-arenaso the lightx2v 8step native res @ 8 steps and larryvrh v4 step 600 @ 6 steps are in the top bracket confidence band. but the other ones there are those silveroxide ones and plaguekind one, both from roughly the past week. anyone used them? noticeably better, qualitatively speaking? (obviously for now the arena stats aren't resolved and are leaning towards everything in that bracket being about the same anyway)interesting that the turbo loras are beating 28-step baseline model too, wouldn't be that surprising with most image models since turbo loras tend to help with prompt adherence to braindead prompting but the prompts on that arena are AI skill-converted up into huge explicit prompts, and i've never felt like using baseline h3 felt like a "raw"/reference model anyway, in the sense of raws like z-image base etc.
>>109717633I think the turbo loras degrade skin the most notably and a lot of the videos in that arena didn't have people. So they might perform a little worse if your use case is cooming or other person focused videos.
https://files.catbox.moe/iy5mtu.mp4
>>109717405They seem all more or less comparable.
>>109717699I use lightx2v 8step @ 8 and yeah it definitely degrades the quality a little, but i'll happily take it in exchange for almost halved gen times (and it keeps audio basically intact rather than so warped you might as well not output it as with the 4-step lora). For me the crucial thing is that the overall composition of the output is the same as base model i.e. the model's not actually getting any dumber or less capable of following extreme case edge requests, just outputting a tiny bit warped, which also means if i really love a gen i can redo it on base with the same seed and get the exact same result in higher quality, unless the gen happened to rely on some really small few-pixels lucky artifact for why i liked it
durbo for reference model drobbed :Dhttps://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
>https://github.com/Saganaki22/ComfyUI-VDN-H3anyone try this?of course author didn't include a reference without cope nodes for their example for some retarded reason
>>109717797>8 steps>768pI shan't be using
>>109715496>anon never reported back
>>109717797Finally. Been waiting forever for this.
>mfw Resource news09/02/2026>ComfyUI MiniMax H3 MotionCachehttps://github.com/Mozer/ComfyUI-MiniMax-H3-MotionCache-FastVAE>MiniMax-H3 WebUI — 视频生成工作台https://github.com/AntaresAlice/h3-webui>H3-World: Turning Language Understanding into World Controlhttps://huggingface.co/DANNY621/H3-World>Identity-Conditioned Latent Consistency Distillation for Face Synthesishttps://github.com/UFPR-IPASP-PR/FaceRec-IdentityConsistency>TUE-Detector: A Tool-Using Expert MLLM-Based Detector for AI-Generated Videoshttps://github.com/Louis-YW/TUE>MegaStyle++: Scaling Image Style Space through Hierarchical Style Definitionhttps://github.com/Tencent/MegaStyle>PredErase: Training-Free Object-and-Effect Removal with Predictive Latent Guidancehttps://github.com/xiuwk0820-collab/PredErase>video2dlssnr: standalone DLSS 5 Neural Rendering video toolhttps://github.com/DaniilSokolyuk/video2dlssnr>ComfyUI MiniMax H3 Video Outpainthttps://github.com/TwoAbove/ComfyUI-H3VideoOutpaint>ArtiFixer: Few-step causal auto-regressive model that enhances and extends 3D reconstructionhttps://huggingface.co/nvidia/ArtiFixer09/01/2026>ComfyUI-H3VAE_TRT: TensorRT version of the MiniMax-H3 VAE in ComfyUIhttps://github.com/lihaoyun6/ComfyUI-H3VAE_TRT>MiniMax-H3 Fused Turbo (INT8 ConvRot) https://huggingface.co/MATLOWAI/minimax-h3-fused-turbo-int8-convrot>RegionCache: Semantic-Aware Region Reuse for Efficient Multi-Turn Image Generationhttps://github.com/hebutBryant/RegionCache>Discrete Diffusion Bridges for Spatiotemporally Aligned Image Translation and Generationhttps://github.com/HKU-HealthAI/DDB>FoundYou: A Unified Model for Personalized Segmentation and Retrievalhttps://ga1i13o.github.io/FoundYou>ContextBias: Controlled Evaluation of Bias Persistence Under Context Shift in Text-to-Image Modelshttps://github.com/Sina-Emami/ContextBias>FairReL: Deepfake Detection using Fairness-Aware Representation Learninghttps://github.com/xiaoman89/FairReL
>mfw Research news09/02/2026>SpatialGuard: Harness-Guided Verifiable Spatial Reasoning for Text-to-Image Generationhttps://arxiv.org/abs/2609.01582>DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolutionhttps://arxiv.org/abs/2608.31106>Gaussian Core LoRA: Distribution-Aware Dynamic Adaptation for Broad Concept Erasurehttps://arxiv.org/abs/2609.01433>Physically Plausible Video Generation via Visual-Semantic Chain-of-Events Conditioninghttps://arxiv.org/abs/2609.00656>Training-Free Inpainting Across Domains with a Frozen Text-to-Image Diffusion Modelhttps://arxiv.org/abs/2609.00862>CameraEditor: Camera-Controlled Image Editing via Video-Prior Sequential Modelinghttps://arxiv.org/abs/2609.01479>SAGE: Subpopulation-Aware Generative Enhancement for Mitigating Spurious Correlationshttps://arxiv.org/abs/2609.01051>Denoising Diffusion Generative Models Secretly Calculate Attentionshttps://arxiv.org/abs/2609.00885>Advanced Pixel Diffusion Model with Guided Sparse Global Refinementhttps://arxiv.org/abs/2609.00798>No Pixel Left Behind: Filling Gaps in Anime Colorizationhttps://arxiv.org/abs/2609.00800>GenScale: A Benchmark for Relative Object Scale in Image Generation and Editinghttps://arxiv.org/abs/2609.00525>MeRoPE: Metric Rotary Position Embedding for Camera-Controlled Video Generationhttps://qiaozhijian.github.io/merope>Reliability Challenges in Diffusion Vision-Language Modelshttps://arxiv.org/abs/2609.01318>Less Is More: Balancing Positive and Negative Space in Visual Concept Blendinghttps://arxiv.org/abs/2609.00476>Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Modelshttps://arxiv.org/abs/2609.00355>Solaris: Towards Interfaces That Are Generated, Not Codedhttps://runway.com/news/research/introducing-solaris>TPSO: Training-Free Diverse Image Generation via Semantic Prompt Embedding Optimizationhttps://arxiv.org/abs/2511.19811
>>109712959cont
>>109718109now do the stupid cat squashed by the krea2 titan
>>109718109I can easily smoke you on krea2 and anima
>>109718184corr
>nigbo
>>109717308>collage op>no schizo in collage>schizo sperging out but nobody is replying to him>schizo so buttblasted he starts shitposting more obviously>didn't even reply to the news postsIs the thread healing slowly? Might even bake the next thread if it makes him this upset
>>109717633>https://huggingface.co/spaces/multimodalart/h3-acceleration-arena>MiniMax-H3 · 28 steps baseline_28step>BASELINE>28 steps
>>109718375we will know when there isn't some drama shittery in the op and nobody shits their pants about that. I am so sick of seeing twink spammer and catjack gens already
>>109718391We all are brother
>>109718375Would be better if the collage actually contained some videos.
>>109718426Trolls are usually too poor for video collages to work on their machines
>>109717308>>Maintain Thread Quality>https://rentry.org/debo>https://rentry.org/animanonI knew the peace would not last.
where is kino?
>>109718555Try /sdg/
>>109718490Never made one with the method linked in the OP, but I'm struggling to believe creating video collages is so resource intensive.
Has anyone done testing and settled on a strong favorite for H3 sampler+scheduler besides the default res_multistep and simple?
>>109718568it's not. the crybabies make things up.
>>109718660>>109718568>>109718490>>109718394>>109718391Keep struggling it just makes new posters aware
>>109718604Isn't euler+simple the default?>>109718660Assumed as much. There's no real excuse for not making a video collage, then.
>>109718568Try it out on an old laptop or something and you'll see. You missed the last troll whose collages would never be more than 1 frame every 5 seconds.
>>109718679Someone caught a stray but the horned schizo is very low IQ don't be surprised he can't do basic stuff, he's still trying to I guess cover his multi year history of being a burden
>>109718388I always knew big dick larrys was the best and everyone else was just huffing copium, though the examples are a bit unfair considering they are all stillshots with no motion, making the 28 step baseline lose to a turbo lora due to the overbaked nature of lora outputs.But still, this info is huge, for stillshots and slow moving videos maybe even anime you can simply use the larry lora over going 32 steps with spectrum and comfykitchen cutting the gen time in half>>109718568You're currently replying to the thread schizo going undercover (see here >>109718679), he makes collages using singular images he shitted out, he is in no position to cry about anything>>109718604I've fucked around in the beginning, but found out quickly res_multistep + simple is king, it just works. Though that is an interesting thing to check out, I'll let you know if I find anything out after testing
>>109718660The crybaby was the one complaining that the collage script sucked when in reality it was his own poorfag system that betrayed him
>>109718708If you're using the collage as a vector of attack you already lost. I only bake when nobody else does it and it's page 6 or less, you blitzed this thread and are now crying because anons saw your obvious antics.If you weren't so mentally ill you would see a lot of success with this instead of rehashing this on a almost weekly basis.
>>109717772Because the base isn't good to begin with
I wonder which one is better
>>109718697Don't have a really old laptop. Is a Steam Deck shit enough?>>109718708I'd only recommend larry for anime or non-realistic stuff, because it results in a slightly smudged plasticky look, that kinda works for anime.
>He's desperatePerhaps don't remove the one thing that pushed the creation of this thread in the first place and you might see some success.
>>109717308>didn't make it into the collage this timeGrim. I need to up my gen game.
my gens are an oasis
>>109718738>you blitzed this threadI didn't make this thread schizo anon, I'm just pointing out you're a schizophrenic.Hope that helps with your deduction of us anons and fingering out who is who
https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_ref2v_turbo_8step_v1.0_768p_comfyui_bf16.safetensorsfinally, new ref2v lora. 0.1 worked fine with 6-8 steps though.
>>109718751Both kinda shit. Image only collage vs. a single Image.
>>109718184prompt? Can krea just do this
>trying to defend the one schizo that split the general for months and continue to do it to this day>expect anons to not point out his rentry not being in OP despite it being in 99% of threads from it's creation>misusing terms againI'm really sorry this is the only stimulation you get out of your life but it will be fixed naturally unless you want to pull an all nighter again which I welcome.
>>109718833Tried it already? Is it noticeably better?
>>1097188330.1 still had problems, especially with fast motion.Test this one out and see how it goes.
>>109718835this
You will never convince me to use a turbo lora. Never. Fuck off.
what the fuck is the best local lm for generating h3 prompts with gemmaprompt/miniconstruct? why does no one ever answer this question?
>>109718870I use gemma 31b, works well.I didn't compare with others.
>>109718870Because it's a retarded question without knowing which models you can run, just use qwen3.8 27B since you probably don't have enough memory to run any of the 400b models
>>109718870Because you're not giving us your specs
>>109718842nope downloading now, but the first one worked just fine at 6-8, 4 could have audio issues and bumping it helpedmodel -> lora -> comfy attention -> shift node at 12/3this is with old turbo, as it knows initial d:https://files.catbox.moe/vib09o.mp4
>>109718884rtx 5060 ti and 64gb dram
>>109718870>why does no one ever answer this question?Multiple anons have shared their choice models before. I think you just added that to mess with anon.
>>109718870I think the Miniconstruct guy already said it was Gemma 31B, mainly because QWEN3.8 27B doesn't really stick to the strict output format required and Glimmer apparently overthinks things to a ridiculous degree.
>>109718888Either gemma 12b or 27b moe (doesn't jail break well).You don't have enough vram to run 31b in a non painful manner imo
>>109718882>qwen3.8 27B I don't know why but Claude and Sol told me 3.8 is not worth switching to from 3.6 on my 12gb system
>>1097189183.8 is more safety slopped and bitches at you if you want to set someone on fire even in a animated cartoon.
>>109718895>QWEN3.8 27B doesn't really stick to the strict output formatNot sure if true if thinking is enabled.
>>109718870>gemmaprompt/miniconstructyou havent vibecoded your own yet? come on unc get with the times
>>109718918>takes advice from ai>even worse might take advice from schizo poster recommending chat llm gemmaYou do you anon, but qwen 3.8 27b can easily do those prompts and is a massive upgrade over 3.6 in every capacity. And if you're already using 3.6 it would be as easy as swapping the models
kroma is already DOAThe fucking retard is training it at too low of a resolution and the model forgot everything in regards to style knowledge.>>109718939Have to agree with this>>109718942If you can run QWEN3.8 27B at q5 you can easily make this tool yourself
For people recommending models post the outputs you get and whyGemma 31b gives me what I want with little fidgeting. even at qat which I typically don't use. Qwen is more stubborn and unless you use a uncensored model version of 3.8 you can see it go against your wishes even during it's thinking loop.I think the orcarouter model is a lot better but I feel like the size savings is not really worth it unless you just like qwen.
The current regommendations for minicoonstruct are Gemma 31B and qwen3.8(with at least a q5 quant).>>109718895qwen3.8 had big fencing issues on q4 but showed dramatic improvement on q5.>>109718918hmm, maybe I should add 3.6 to the benchmarg suite.>>109718910I'm on an RTX 5070 Ti and get a reasonable 9-10 t/s running a spicy QAT q4. I'd say its a reasonable balance between speed and quality
>>109718954>The fucking retard is training it at too low of a resolutionThe way model training works, if you didn't know, is that it is first trained on lower res images before being trained on higher res images. It's a little known and nuanced process that all model training follows. I'm sorry you don't understand this.
>>109718993Seeing how he shat the bed with chroma which had a hard resolution gate because of his training compared to regular flux I think you should take a moment and think about what I'm saying
>>109719005Where is your model?
>>109718954>the model forgot everythingwhere have I heard that before...
>>109718954>If you can run QWEN3.8 27B at q5 you can easily make this tool yourselfim not yet smart enough or patient enough to get local llms that run on my shitbox to Actually Code anything of merit :( its fine for prompt generation doe :)
>>109719013You're taking this awfully personal
>>109718990Logically, shouldn't q5 be worse than q4?
>>109719117.......no, not even close
Are there any finetunes stil lbeing made or are we all just waiting for the next hot thing that makes us forget previous hot thing from a month ago
>>109719117Not sure how you'd figure that? Why would a stronger quant produce better results?In my benchmarking tests, Q5 is a pretty significant upgrade, at least for this specific version: https://huggingface.co/HauhauCS/Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUFQ4_K_M for this version did pretty well: https://huggingface.co/llmfan46/Qwen3.8-27B-Ultra-Uncensored-Heretic-Native-MTP-Preserved-GGUFbut still worse than Hauhau Q5.Avoid BlackFrost-AI.
>>109719133You're right. It's been so long since I've had to think about q quants that I forgor.
>>109719154>forget previous hot thing from a month agoWhat hot thing from only a month ago is anon still gooning over? The local SOTA meta has been established for well over a month, anon.
>>109719154new hot thing is theoretically H3 releasing their own official upscaler and improving their ref2v model, also maybe H3 image.
>>109719167He ask that retarded question every thread much like this retard>>109718555You're dealing with mentally ill anons
>>109719172and their dedicated sparse attention so we finally get something tuned to how the model was trained
Demonstrative nsfw that requires no loras.-( https://files.catbox.moe/kej8ih.mp4 ) here's the textual phrasing / composition. (.txt) prompt is too long to paste here. sufficive h3 likes big dumb cave man smoothbrain mechanical and repetitive descriptive wording and step by step guidance describing the approach, the embrace, the action of aligning for intrusion and tip to base mapping and movement, specify envelopment and this gets you constituent target hole confinement and insertion without redefining the hole or the peen. For male anatomy refer to the target sex organ mechanically without referencing names (penis / cock ect becomes shaft or thick or very thick shaft). For female write it like this - vagina slit crotch sex hole - do it every time. to get deep penetration think of the penis as a literal spear orchestrate stabbing and impalement fully to hilt insertion so like - continued in .txt file.
>>109719202that's the worst thing I have ever seen
>>109719202sexy
>>109719202>For male anatomy refer to the target sex organ mechanically without referencing names (penis / cock ect becomes shaft or thick or very thick shaft). For female write it like this - vagina slit crotch sex hole - do it every time. to get deep penetration think of the penis as a literal spear orchestrate stabbing and impalement fully to hilt insertion so like - continued in .txt file.You're giving me a little chub when you talk like this
>>109719154Finetuning is dead. Used to cost like $100k for a decent bake, now it’s in the millions. Local hardware stagnation caused this btw
Implementing a new creativity toggle and Gemma 31B is having trouble passing the tests.Qwen3.8 Q5 had no trouble from the start.
>>109719285Anima is a finetune and cost a bit more than $50k.The problem is that local bakers are completely incompetent and accuse every new model of being untrainable. Including Anima itself, which makes absolutely no sense, because then how was it trained from Cosmos2 in the first place?
>>109719356They also can't follow basic instructions and cry when it doesn't work. Chroma to this day will remain an embarrassment on local finetuning.
https://gofile.io/d/MNfo8UbP
new lora for ref2v seems goodhttps://files.catbox.moe/wkml6z.mp4
>>109719388steps?
>>109719388Sorry I don't click random catbox links actually post it boyo
>>1097193948, havent tried 4 or 6 yet
What do i need to prompt to get action from the very beginning? Every single fucking video starts 2 seconds in at the earliest and 99.9% of the videos are very slow.How to fix that?
>>109719417I usually write "the shot starts immediately with..." or something like that
>>109719417if ltx reroll, never found a reliable solutionif anything else, no idea, it's only happened to me with ltx
>>109719383you use aitoolkit for loras right? post config pls
>>109719457MMH3, everything is in fucking slow motion and starts halfway the video
>>109719417I'm not interested in playing 20 questions in order to answer your vague question. Post a catbox so anon can see exactly what's wrong.
>>109719388Yeah, it's neat. Also did a small comparison between v0.1 and v1.0, both at 8 steps euler simple, 12 3 shift. Also the same seed, but that hardly matters when turbo loras change the output so wildly.This one's only using a single reference.v0.1:https://litter.catbox.moe/hrsaw2.mp4v1.0https://litter.catbox.moe/h3ays5.mp4Currently doing one with two references, that will probably ruffle a certain persons feathers, but it was one of the only ref2va gens I had readily available.
>>109718884>to usThis is 4chan and not your personal discord server.
>>1097194357-9 roflmao
>>109719202Even though I don't exactly enjoy your example I think H3 is actually really good at nsfw out of the box, but not that good for anime/3d style.As you said you simply have to autistically describe what you want to see, ergo, describing everything that is happening. Are his hips moving? Prompt it. Are his knees bending during thrusts? Prompt it. But yeah H3 can do nsfw and is really good at it in 3dpd style since it was definitely trained on 1-2 person porns, though it falls apart at anything that uses more than 2-3 subjects in ref2va>>109719388Ref2va looks so much worse than fl2va so it's kind of hard to tell if it's better than larrys, I think there is no getting around to prompting 1 for 1 comparisons with quality prompts that use 4-6 reference to actually see a difference.>>109719506>runs at 8fpsWhat's the point if it's a still image compilation?Turbo loras struggle with motion and movement, not exactly with coherence
>>109719458>>109719383>you use aitoolkit for loras right? post config plsSorry, I don't use that. Might release the trainer sometime in the future.Pro tip: target mu=1.15 for the schedule no matter what the resolution if you use turbo on top of raw.
Riddle me this. Should going higher steps with sparse attention produce a different (noticeably better or worse) result than lower without? Like, say, 30 steps with sparse attention vs 20 vanilla. I know I can test it, and I'm trying knowledgable anon can comment on how these methods work (or has tested them already).
https://civitai.red/models/2911018/anima-expanded-29b?modelVersionId=3292754
>>109719581Do you want the long or the short answer, cause the short one is that sparse attention is a retard node that caps your quality inherently and will always be worse than base h3 at 20 steps or higher, no matter how many steps you use with sparse attention in comparison
>>109719618short answer is good, appreciated. Are there any other speedups that should theoretically allow the high steps to increase quality while preserving some speed, or is there just no free lunch here
>>109719506Using two references, same settings.v0.1:https://litter.catbox.moe/530mn1.mp4v1.0https://litter.catbox.moe/m0yd9h.mp4I feel like the seed was just kinda bad for v1.0>>109719550Well then. Do provide us with a more apt comparison between the turbo loras.
>>109719435Lonely again thread schizo?
>>109719647Spectrum is the free lunch you're looking for. It actually works in the example you posted before, where you could get a better video at the same time using spectrum with more steps.To simplify: H3 base @20 < Spectrum @32And Spectrum is most likely still faster while having both higher quality video and audio.>why 32Seems to be the sweetspot after which video quality no longer improves.And comfy kitchen attention is completely free with no drawback at all, since it doesn't require you to downgrade cuda and pytorch or some other things like sageattention which is the retard equivalent with a slight video quality drawback, but nowhere near as bad as sparse attention since they do different things. To be fair, the quality difference between sage and ck is barely noticeable unless viewed side by side, but why use the inferior version when ck already exists?>>109719687Yes I will, I already have 1 video comparison, will post once I have one more with more action to truly see the difference
>>109719550 been using the reference model it likes to ignore instruction but with enough autism Minimax will literally depict anything kek.>>109719581>>109719618Been using the 4 step turbo at 8 steps, its pretty satisfactory. (nsfw+cum_inflation+bursting warning https://files.catbox.moe/rykgb3.mp4 / https://litter.catbox.moe/5w4tty.mp4 )
motion seems better in new ref2v lightx2v lora:https://files.catbox.moe/9i9iqt.mp4
>>109719763theres a new ref2v lora?
Here a comparison between Larry at 6 steps, new ref2va lora at 8 steps, and base at +30 steps. And I don't mind if my gens trigger the village schizo.Larry 600:https://litter.catbox.moe/afrzhr.mp4https://litter.catbox.moe/ha3u0t.mp4New Ref2va 1.0 Lora:https://litter.catbox.moe/933vtk.mp4https://litter.catbox.moe/qty9rd.mp4Base at 30 steps:https://litter.catbox.moe/j42qbj.mp4https://litter.catbox.moe/ow5d4y.mp4both videos have over 5 reference samples and used ref2va and res_multistep + simple @ 0.98MP.
>>109719604get out grifter
>>109719857yeah 1.0 finally came out
>>109719998Thanks for doing this.
I cannot into catbox tonighthttps://streamable.com/8su5mt
>>109719998Base > lightx2v > Larry600honestly, I've always found larry loras so shit, I don't know why redditors like shilling it so much
>>109720127kino
>>109717797Ref2va model is fundamentally broken compared to fl2va. By the way image references in fl2va work perfectly fine.
>>109720127hot.. would
>>109720160ref is mainly for video reference and that's about it. I couldn't tell you which one has better audio
>>109720160>Ref2va model is fundamentally broken compared to fl2va.what do you mean by this?
>>109719998now gen at 0.5M . The light2x lora is only good for 1 resolution
Where will we be at a year from now?
>>109720137I think larry had their use cases (and still does), especially for quickly testing drafts at low res since it's very similar to the end result you get with base at high steps in content.Though now with the base ref2va lora I'd say it's actually okish for full 0.98MP videos and would save 50% of the total gen time to 32steps with ck and spectrum if you have no moving parts like long hair, blood splatter or anything else like that.>>109720108Another fun fact, the ref2va turbo model seems to also work on fl2va. Might even be better than the older fl2va version, though can't confirm it yet>>109720207He probably just means how shit the video quality of ref2va is in comparison to fl2va. It's especially obvious if you prompt at 0.5MP and lower where fl2va looks quite decent and ref2va like shit. But that's because otherwise both models couldn't be the same size, not exactly broken or anything>>109720227I could try it if you're interested, though I doubt it makes much of a difference
>>109720245>I could try it if you're interested, though I doubt it makes much of a differencetry it, especially pay attention to prompt adherence.
>>109720235MARS LOL
>>109720270There you go, same exact prompt adherence, followed my prompt just as well as 0.98MP, maybe even better considering lower resolutions have higher prompt adherence, so I can't confirm your statement. But it's sample size 1 so farhttps://litter.catbox.moe/rprm8q.mp4
>>109719998What resolution?
updated lora seems nicer, this is only at 0.4mp.https://files.catbox.moe/ybgvnu.mp4
>>109720270>>109720314nevermind I just noticed, it did fuck up harry and ron sitting across from hermione in shot 1. Maybe you're on to something
>>109720227>The light2x lora is only good for 1 resolutionhuh?
Testing the H3 sarcasm functionality.https://streamable.com/oe5gjp
>>109720330and catbox is derping up again. in any case, it seems good at various resolutions
>>109720365Just use litterbox, it's up. Catbox isn't designed for ai slop in the first place
>>109720356i came
>>109720344It says right there "768p"When you use it at low resolution, it doesn't look as good.
Gend for 6 hours with H3, over 30 gens and all failed, lmao.I really need better refs tomorrow.
nobody ever stopped to ask is our children learning?
>>109720365>>109720373please make sure to use 3 day expiry. sucks when shit disappears after just 1 hour
>>109720390is this mori calliope
>>109720394I always use 1day since that should be more than enough time to check things while the thread is up
>>109720377Are you retarded? That doesn't literally mean it will only work at 768p.
Getting sick of this yellow fever moron spamming his shit gens.
>>109720373both sites are acting up, in any case it works, voice clone etc.
>>109720375I specifically prompted for her butt to be visible in the mirror just for you, anon.
>>109720446is that an Eva in Gundam colors
>>109720453It's way too small, but it does look like it.
>>109720390whats her soundcloud?
Uguu is more reliable than litterbox, but 3 hours onlyhttps://uguu.se/
>>109719998does it beat this?>https://huggingface.co/Rkss/Minimax_h3_fl2v_lightx2v_turbo_4to8step_v0.1-v1.0_768p_v4_step600_dareties_fro095_native
>>109720530yeah makes it a pain when checking old threads
Here is something slightly interesting.1 image ref. of a character used in fl2va with the ref2va prompt. Basically using fl2va like ref2va but with the quality of fl2va and the lora for ref2va.And while it was a quite simple prompt, it followed it perfectly. Potentially allowing you to use fl2va for t2v with a single image reference, maybe even two?Oh and the prompt time was 150s for 10s of video at 0.5MPhttps://litter.catbox.moe/ay1bzy.mp4>>109720569>made for euler samplerNah, I don't even have to test it to know it's dogshit, since euler can't do fine details for some reason and the speed increase of 3-5% over res_multistep does not make up for it.Someone posted a benchmark site for speed loras itt and larry won without a contest. And with the new 1.0 ref lora from lightx2v being even better than the larry lora we currently have a new king.Feel free to try it out yourself though if you're interested
>>109720646light2x is made for euler too you know.it's a distilled lora. The point is to have smooth delta, not which sampler is better
Found this nice node pack for managing longer H3 videos, it's pretty feature rich (seems to be missing sparse attention atm though). Having a timeline editor in comfy is so useful for this kind of stuff https://github.com/roadmaus/ComfyUI-Continuity
>>109720673>light2x is made for euler too you know.Source?
>>109720586no need to do that if you are here 24/7 just a tip
>>109720673>>109720646i've noticed er_sde being a pretty consistent upgrade over euler
>>109720530Don't use Uguu. Some people have jobs.
>>109720697Recommended Inference SettingsSteps: 8Video shift: 12Audio shift: 3Sampler: EulerResolution: Up to 768phttps://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/51
Testing h3 for the ' bonk on the head' . it never rendered on wan 2.2
>>109720402dunno. just took the pic from earlier in thread>>109720504she's a mime rapper
>>109720740decent bonk
Anyone know how to get better eyes in Krea 2? Either the eyes are really bad, or they're just ok. I tried a face adetailer but a lot of times it causes the face to look like it's part of someone else's body.
>>109720703I'm not
>>109720771okay schizo
>>109720729Very Interesting, here is the same thing but with euler for comparison:res_multistep:https://litter.catbox.moe/933vtk.mp4Euler:https://litter.catbox.moe/m1rnsn.mp4Honestly hard to tell which one is better if it doesn't just come down to personal preference, but if the official documentation recommends euler who am I to judge.Just makes it more tedious to switch from turbo to non turbo each time>>109720703I set all my files to expire in 1h to make sure only the thread schizo can view my posts in time>>109720766post an example of those shitty eyes
>>109720766no
>>109720740bonk good
>>109718774I rarely post and somehow got two in this time.
>>109720675i don't understand why people keep trying to make half-assed timeline editors in Comfy when DaVinci Resolve can be had for free
>>109720959This one will basically build the prompts for each segment then automatically pass in the motion context etc between them.
Imagine being obsessed with a Californian NEETCouldn't be me
>>109721004i guess that's neat if you don't plan on editing anything at all
>>109721045NTA but it's not about editing. Sometimes you'll want a long continuous shot that can't be hacked in with an editor. You could try but it won't look good.
https://litter.catbox.moe/5d0byj.mp4
>>109720127i exhaled out of my nose slightly stronger than usual
He's trying really hard today.....
HOW THE HELL DO YOU GET A JESTER COSTUME WITHOUT IT TURNING INTO POMNI EVERY SINGLE TIME AAAAAAH
>>109721184use a reference or keyframe?
>>109721035kek
Friendly reminder to:-Check if comfyui is raping your main SSD by constantly accessing the pagefile and place your pagefile somewhere else or reconfigure comfyui-Undervolt you GPU, especially if it's a 5000 series card, those can often run at 30% less wattage with the same performance while also running 20C cooler.-Check your ram temps and install a ram fan if they are too hot since constant burst heating is what kills them-Carefully clean the dust from your PC case by holding the fans in place and using either a can of air or better yet one of those electric mini blowers designed for PC's (preferably not in a small room or even indoors, since dust is cancer)
>>109721280and clean the jizz out of your mouse and keyboard with isopropyl alcohol
>>109721484wtf lol, based
>>109721280what about my SSD temp, is this normal? it's showing in red which im assuming is too hot
>>109721530just buy a case fan
>>109721530Come on anon....google that shit, we need to know the model and gen and at that point a websearch will answer that for you
>>109721530>>109721559the thread schizo is right you know, we don't know what kind of ssd you have and they all have different maximums for their temps.Though if 63C is idle then it's really shit regardless, should be 50C at idle when not transferring or reading large files, even if the ssd is on your motherboard under the gpu or next to the cpu.Regular 2.5" ssds should be room/case temp even
>>109721590>Reputation so bad he has to pretend people are just like him.Surely this will work the 80th time. Good luck though.
>>109721559you should not dish out any advice in these threads buddyconcentrate on getting well
>>109721559no i prefer human discussion that's the point of an online board>>109721590Predator GM7000 NVMe SSD 2TBit runs 63 on idle. lodged somehwere akwardly behind the GPU i think. under load doesn't change much weirdly.
>>109721611What does that even mean? Translate from Schizo to human?>>109721622Like I said, 50C is ok for idling, the lower the better, though it might get hotter during gaming just because the gpu and case temp increases, but if you're just idling on your desktop watching yt or something it really shouldn't go above 55C. But it's not going to die at 65C either since those things are designed to run up to 70-90C. What I'm trying to say is, either you have a shit motherboard that has bad m.2 ssd heat transfer, or your case is running too hot, but you don't actually have to worry too much when it comes to lifespans, but you might lose a lot of performance during intense reads because of thermal throttlingTL:DRLifespan not much affected, performance is
Check these apples, Jack!
>>109719717appreciate what you did to my Joe, Jack!
Can DLSS 5 do environment only instead of characters ?? Might be useful for it
>>109721731Forget to mention i want too do DLSS5 on existing videos
>>109721731it takes masks as input for exactly that, selecting what should be enhanced
>>109721677ugh i think i'll just leave it. if you say the only downside is speed sometimes is worse, i'm too lazy to do anything about it
What tools do you guys use to write/improve your prompts?
>>109721769my brian
>>109721731the nvidia demonstration video showed that being tunable, only faces, only characters, only environment so on
how come i can use any other checkpoint but when i try to use krea i get connection timed out..
>>109721794what's the problem, vague poster, afraid to give us details that might help us solve your issue?
>>109721794you can't use krea 2 on the same basic workflow as the average model, if that's the issue
>>109721806sorry. im using https://github.com/Haoming02/sd-webui-forge-classic and was following this https://github.com/Haoming02/sd-webui-forge-classic/wiki/Download-Models , including the right vae and text encoder. anima is working fine, but krea just times the connection out when i try to generate. im still new to all this.>>109721820hmm, the github says it works with krea. it even has a krea ui preset. not sure what else im doing wrong then
>>109721484im a fan of this art
>>109721782Hello, Brian!
>>109721558who dis? she cute
would buying an RTX PRO 48GB be worth it right now? You think the price will go up, down? I'm tired of waiting for 35 min vid gens. Getting FOMO fever and I want to be the one guy with VRAM when the inevitable Cloud dystopia arrives
>>109721855I haven't seen any gens from you. It's very difficult to recommend anything at this point.
What's the point of this node? I heard it's supposed to improve results but it seems like when I enable it, the prompt cannot tell what penetration is, I turn it off and it actually does penetration images fine
Let's cook!
>>109721855yes it worth
>>109721855these choices are a gamble, no one can tell with certainty what will happen
>>109721855>costs more than a 5090>is slowerI'd wait honestly, maybe they figure out how to run h3 minimax in a multigpu setup that allows you to use parallelism and then you could save a couple thousand bucks by buying a large motherboard and three 15gb gpu's instead of one slow old ass rtx card.Alternatively buy the 5090 instead since it's almost $3000 cheaper while also being faster and big enough to run H3... super alternatively wait how much memory you will need for H3 max
>>109721950I could care less about maxing out gaytracing in AAA movieslop. I like the new RTX 5000s because of SFF and lower power draw. I hate 5090s>>109721939Very true, which is why I'm so torn right now. I could drive to the Micro Center an hour from here and pick up one of the remaining 3
>>109721992so you do care a little
>>109721992You were the one crying about having to wait 30 minutes per video gen then suggested buying outdated trash because of the increased vram size that will still lead to 30min gen time.Also 450w vs 300w shouldn't be a dealbreaker when you're paying almost $8000 for a shit gpu.There are single slot 5090s as well so everything you're saying makes 0 sense
>>109721677you're really desperate today, too bad all that effort evaporates whenever anon clicks the link.We're at a time where there are less newfags than usual so good luck
uploaded sexo. have fun with her.https://civitai.red/models/2912060/benedikta-harman-final-fantasy-xvi-krea2-lora?modelVersionId=3294047
>>109721828nevermind, i got it working. my page file size somehow went back to automatic even though i thought i had it set to 40gb... for years. huh. dunno how tf it changed itself
every time i try to use an LLM to write a prompt for me, I end up re-writing it in half the words and it turns out better.
>>109721837A Robb
>>109721855If your desperate just pick up a 16 gig 4080 or a 5070ti, you'll burn like 400 - 900 on it and can save till the equivalent 6x series card rolls around and you'll get like a 3 x speed increase and likely more ram for 100 to 300 dollars additional. genning 15 second clips with the 4step turbo lora at 7 steps in less than 2 minutes.
>>109722123>6x seriesThere will be no 6x series
>>109722157>There will be no 6x seriesHow do you know?
>>109722204I think he means they will jump to 7000 because of the blackwell pro line?Trying to take a guess here, I think AMD fucked up with that with the cpu line
>>109717700kek, topicalIn his defense, he probably would kick a random's ass thanks to his (even if just beginner) BJJ moves.
>>109718128I mean ... she's still got nice skin at least. And you know what Mr. Franklin said about GILF pussy so I'd give her a chance.
>>109718132Adorable! How TF is no one else commenting on this?>>109718391>I am so sick of seeing twink spammer>t.asteletPearls before swine fr fr
Next up on MiniConstruct: local file storage to replace IndexDB
>>109722291I never commented cause I keep seeing those two plastered across x, its cool and cute I'm just at saturation on those two.
>>109719202Dayum, son, looks like she's about to pop.>cum outta his mouthwut
>>109719719I know I probably shouldn't even ask but qrd on why you seethe every time he posts? I like his gens.
>>109720127KEKNice to see my suggestions being of some use. Also>You've got red on youSomeone watched the actual movie, eh?
>>109717413What model?
>>109720356Hehehe, that's two for two.
>>109720390Sick style (both visually and animation-wise)!
>>109721035lulThat's not where I thought this was gonna go.
>>109721090>loli-senpaiLoli-sama or -dono, no?
Rape All New Faggots
>>109721519>a dozen cunnies flattened>basedNo, the fuck is wrong with you?
>>109721731>L'Origine du melonor>L'Orange du monde
>>109722098I see Kandinsky-posting is going strong, nice.>>109722123>no nipples on the uddersBoo! Boo this man!(or "moo")
>>109722240Simple but eye-catching. Nice.
>>109722344>across x???>I'm just at saturation on those two.I get the sentiment for the latter but Astolfanon's gens are based. Oh well, more for me, teehee <3
>>109722489https://files.catbox.moe/bmypm3.mp4
I ain't no fortunate Jack..
>>109722536this is fucking amazing.high fidelity. no visible slop.
>>109722536
>>109722536Sick design and really cool animation. Would fit right into a FromSoft game.
>>109722448So you're starting with self-cest, I'm guessing, no-gens?>>109722513I recant my boos.
>>109722640Marble Joe
>>109722653daisy,daisy, give me your answer do...
What is with you and Sleepy Joe, anon?>>109722625Israel's greatest ally. Even had them marry into his family
>>109722660Shit's hard to get right. I wanna imitate >>109722027 but no luck so far. Getting that blotchy, impressionistic vibe proved way more tricky than I thought. Kinda just throwing shit at the wall at this point desu
can minimax do porn now
>>109722661I love uncle joe...
>>109722678Yes, many chinks make tiny dramas with it. You see them get posted on Xitter
Hello, Twinkletoes.
>>109722679Go Brandon, go!>>109722692>shoesWTF is wrong with you, anon?
you better make a good collage there are quite some kinos here
>>109722722Many things wrong with me - but.. you shall get toes next!
>>109722678If you like body horror pussies and non-human noises, yeah.
>>109722769Giga cursed.
Minimax h3 has too many optimizations in a month I'm so fucking lost. There's like 50 turbo loras now
>>109722655wow o.o cool
>>109722842Those are all snake oil. You train a lora on random images, call it a "turbo lora", and nobody notice the difference.
>he fell for it and replied
>>109722722Twinkle my toes, anon~
>>109722882You son of a bitch, you've gone and done it now ...Sleep with one eye open tonight.
>>109722842only ones that are worth it :sage attention 2 or comfy kitchen attentionlow tau (<1) sol attention
https://civitai.red/models/2877206/dasiwa-minimax-h3 are these any good?
>>109722678>>109722774minimax can do porn amazing what are you talking about xD. here's an explainer (too long to post here) https://files.catbox.moe/97jozz.txt and heres a pron genned using a single image as ref and the methodology outlined in the .txt: https://files.catbox.moe/04rndk.mp4 / https://files.catbox.moe/gf6iyi.mp4 / https://litter.catbox.moe/zx44g4scrbh0xrny.mp4 / https://litter.catbox.moe/veok0ohxg2t7em3z.mp4 Is someone ddosing catbox?? the fucks going on?
Swingin...
>>109722921Fuck off already you worthless moron
>>109722923<3
it's gonna be a shitbake isn't it
>>109722909what even is it about?also what does picrel even mean? no nsfw?people will make giant explanation pages but will never explain what the model is about
>>109722732Give me like 30 or so minutes.
>>109722542I like when she turns around. Is it possible for 8GB VRAMlets to run minimax?
>>109722936Tos violating? what platform? could be something as innocuous as violence.
>>109722911can you try non scalie/furry anime porn?also catbox just regularly dies like that, nothing new
>>109718132Workflow pls.
>>109722929I'm too lazy to make one and there is in theory absolutely no rush since it easily takes 3-6 hours to land on page 10. The only "rush" would be to not let schizo kun bake and make a one image collage
>>109722945>>109722945>>109722945
>>109721493hair too long and wrong color
>>109719383Anyone got H3 vids of her doing this? I'll jerk off to it