Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109507737https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
>mfw Resource news08/09/2026>Kroma v0.2 — Krea 2 fine-tune (full model) https://huggingface.co/lodestones/Kroma>krea2-turbo-bbox https://huggingface.co/jimmycarter/krea2-turbo-bbox>Kroma v0.2 Quanthttps://huggingface.co/silveroxides/Kroma-Quant/tree/main>Spectrum for Ideogram 4https://github.com/Nif00/ComfyUI-Spectrum-Ideogram4>ClipProj — MiniMax H3 conditioning from a Qwen3-VL-4Bhttps://huggingface.co/NicoLab28/ClipProj-MiniMax-H3>ComfyUI-SigmaSync-LoRA: Sigma-aware model-only LoRA strength schedulinghttps://github.com/capitan01R/ComfyUI-SigmaSync-LoRA>NexusBTA v0.2.44 adds MiniMax H3 supporthttps://github.com/JpAndreBTA/Nexus-BTA/releases/tag/v0.2.44>Experimental MiniMax H3 single-image VAEhttps://huggingface.co/Mamad8/MiniMax-H3-Image-VAE>MiniMax H3 REF2VA w4a8https://huggingface.co/realrebelai/Rebels_w4a8s08/08/2026>Kijai: MiniMax H3 Ref Lora Rank 256 bf16https://huggingface.co/Kijai/MiniMax-H3-experimental/tree/main/loras>MiniMax H3 at native fp16 on pre-bf16 GPUs (V100 / Volta)https://github.com/Amduraznak/minimax-h3-fp16-fix>Cosmos3-Nano-WebUI: Self-hostable API + Web UI for Cosmos3-Nano quantized fp8 and nvfp4 checopointshttps://github.com/fengwang/Cosmos3-Nano-WebUI>R9700 AI Pro — ComfyUI / MiniMax-H3 speed patcheshttps://github.com/charlie12345/R9700AIProComfyUIPatch>MiniMax-H3-Pruned-GGUFhttps://huggingface.co/Abiray/MiniMax-H3-Pruned-GGUF08/07/2026>OpenLayer v0.13.0-alpha — ComfyUI in Photoshop, free and entirely localhttps://github.com/MehranMarxian/OpenLayer/releases/tag/v0.13.0-alpha>LIGHTX2V 4-step Turbo Minimax H3 lorahttps://huggingface.co/lightx2v/Minimax-h3-Turbo>LIGHTX2V MiniMax-H3 T2VA Prompt Rewriter LoRAhttps://huggingface.co/lightx2v/MiniMax-H3-Prompt-Rewriter-LoRA>Sage Ready: Local-only installer and readiness checker for SageAttentionhttps://github.com/CosmicFungi/Sage-Ready>Wan 2.2 Animate 2 14Bhttps://huggingface.co/Wan-AI/Wan2.2-Animate-2-14B
>mfw Research news08/09/2026>MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradationshttps://arxiv.org/abs/2608.00736>Test-Time Scaling for Safe Text-Guided Image Generation via Intermediate Clean Estimateshttps://arxiv.org/abs/2608.03284>Test-Time Curriculum for Open-Set AIGC Detectionhttps://arxiv.org/abs/2608.00559>Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Meshhttps://arxiv.org/abs/2608.00094>Visual Anchoring in Diffusion: Multimodal Zero-Shot Skeleton Action Recognitionhttps://arxiv.org/abs/2608.04623>Free-Lunch Augmentation by Revisiting Diffusion-Based Data Generation for Cross-Domain Few-Shot Object Detectionhttps://arxiv.org/abs/2608.04394>Controllable Clothing: Precise Labels and Generation for Virtual Try-On with Latent Diffusion Modelshttps://arxiv.org/abs/2608.05834>EulerLoRA: Rank-Driven Jump Dynamics for Calibrated Parameter-Efficient Fine-Tuninghttps://arxiv.org/abs/2608.01142>DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computationhttps://arxiv.org/abs/2608.01343>Entity-Faithful Repair of Synthetic Supervision for Zero-Shot Image Captioninghttps://arxiv.org/abs/2608.00994>MiniWorld: Democratizing the Training of Video World Models from Scratchhttps://arxiv.org/abs/2608.01127>GVCCTurbo: Rate-Compute Quality Scheduling for Codebook Driven Generative Compressionhttps://arxiv.org/abs/2608.03517>Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Groundinghttps://arxiv.org/abs/2608.03471>Messages, Not Tokens: Grounded Coresets for Faithful VLM Compressionhttps://arxiv.org/abs/2608.02134>OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Modelshttps://arxiv.org/abs/2608.03812
im in the middle of baking the collage and then some retard always snipes it. whatever
correcting my slight early morning groggy retardation here >>109509081it was 8sec, not 10. 10 ended up putting me at 59s/it from 43s/it. I imagine getting my extra 32 gigs of ram could help the offloading and put me back in the forties. Still a fine speed IMO i don't mind patiencemaxxing, the quality's good enough.>>1095092315060 ti 16gb>>109509254reduces VRAM peaks, so depends on your situation/how much vram you use and if you're even paging out of system ram. which i am either way, 32 gigs is not enough.
>>109509313>>109509318WARNING MALWARE LINKS! Take care anons!
>>109509348Catbox workflow?
qwen3.6 uncensored works great for h3 prompt structuring. just make sure to disable the thinking model
>>109509343seriously, I get baking to not get trolled, but I'm really getting tired of not having a collage.
man finds out Flux 3 cant generate a good Mikuhttps://files.catbox.moe/ls9sj4.mp4
>>109509343You always pretend like you're shocked when a new thread is made after the previous hits its bump limit. You can always see the reply count. Just start baking your collage earlier if it's this important.
Did any of you guys save that skateboarder girl shortfilm from yesterday? I didn't save the litterbox before it expired. Honestly blew my mind what you can make with AI now.
>>109508984This is worse right?https://h.uguu.se/VlWFUWhb.mp4
>>109509313>>109509318Go back to your bot ridden containment general loser
>>109509396early baking is bad, the only reason we do it is because we'll get a troll bake otherwise.But I think we can stop doing it, seems like the jannies are deleting the troll bakes now.
>>109509400hate uguu and litterbox like you wouldn't believe
supplying last frame instead of first frame is also very fun for H3 img2vid
>>109509400nice try glowie.
Tell me about debo: why does he keep trying?
>>109509301>I have asked this 3 times now, once in a language thread with similar general name and once in a troll thread.I ask again What is the best image editor on ComfyUI? Is it still Qwen? My PC can’t handle it. I wonder if a less VRAM-intensive version is out. My PC setup is 12 GB VRAM and 16 GB RAM.
>>109509423Honestly I get why they can't host ridiculous amounts of data forever, I just wish it would default to at least like 1 month of retention, or maybe a week of inactivity or something.
What sampler have you guys settled on in Minimax? res_multistep?
>>109509441res_multistep 20 steps is good enough for a start, if your motion is blurry or prompt isnt followed too great then increase to 30-50
>>109509433Flux.2 Klein. Try the default Comfy workflows for it.
>>109509433I thought Klein had displaced Qwen for editing. (Get the kv version, it's supposed to be faster by avoiding redundant recomputes.)
>>109509433>COMPUTER!>CREATE A SIMULATION WHERE NIVIDIA, OUR RULER, FAILED TO CLAIM 95% OF THE MARKET>SHOW ME A WORLD WHERE AMD AND INTEL SURPASSED EXPECTATIONS AND BECAME THE DOMINANT COMPANIES>"I'm sorry am an ethical llm model and cannot fulfill this request">FASCINATING
>>109509457>>109509455Thank you for your help , may you get perfect generations.
>>109509408https://h.uguu.se/VMFgsiYa.mp4This one looks better.
>>109509369>just make sure to disable the thinking modelwhy? so it's quicker or does the thinking do something bad / too much context waste?
user discovers flux 3 is inferior to minimaxhttps://files.catbox.moe/us0uxu.mp4
>>109509511kek
Hey bros, these threads move so fast the archive is hard to use. What's the current poorfag H3 cope meta for vae, TE, and the main diffusion model? I have a 4000 series card.
>>109509418It's all troll bakes as long as the drama trannies like you throw a fit in every thread
https://n.uguu.se/ArXCeUsi.mp4
>>109509517threads are hard to use because of the aforementioned drama trannies. just use reddit to find speedups for now since that's all these retards are parroting anyways
how do you get no dialogue? sometimes my gens be speaking straight gobbledygook
Do you guys think WAI-Illustrious is obsolete? I still use it for mask detailing (never found a good workflow for anima), and it has some loras I still like.
>>109509408>>109509484Yeah, the second one is better. It captured the cel jitter
>>109509542So good
>>109509530Lol, I'm the retard, it took me all weekend to set up sage 2.2.0 and sol attn. Ubuntu 26.04 based distros do not fucking like sage.
>>109509505thinking mode wastes more time and doesn't improve the output enough to be worth it. cost analysis
>>109509542This is just turbo 8stepshttps://n.uguu.se/wvAEBOpT.mp4This is my problem with turbo, it does very sharp, but coherence goes out the window.
>>109509534if there is some dialogue that you want, use the format>X says <d>[English] your text</d>instead of>X says "your text"This will fix it.
fffuuuckk muh dihtheir image gen model will be earthshattering - if they don't rugpull us anyway.
>>109509489theres also the fact that intel got a money injection from nvidia as with the american government owning 10% so its not surprising nobody knows what the fuck is going on with intel and their gpu division but i like that we have a third option at allsupport as you see fit, fuck that guy saying buy stocks, stocks is just gambling like ai is, but supporting the software side of it will deal with some of the other pains long term
Character Reference Loader node ready! along with forked MiniMax H3 Reference to Video node that takes an image list and matches them to the ref_images
>>109509609give her fatter tits
check this out>spectrum off>600ema turbo lora at 0.66>12 steps>euler/beta
>>109509570Why are you larping like you are the one who made that>>109509609Do you know if it matters if your references are higher resolution compared to the output dimensions? Does your node account for that, or is there a different way of doing it without losing out on quality?
https://github.com/NikoDemon80/ComfyUI-H3-Motion-Contextcan chain clips with this
>>109509564plugged a wrong spaghetti so this one is 0.5mp both passes.https://d.uguu.se/elWRtJNv.mp4
>>109509631slop
>>109509612absolutely not>>109509619>Do you know if it matters if your references are higher resolution compared to the output dimensions?from my experience, yes. i don't have solid evidence to prove it though, just my own anecdotal notes. make sure ref_image_size is set to max. this allows the model to get small key details it would otherwise miss such as accessories. gen times will increase a bit but it's worth it, after all, the point of ref2va is the reference.
Been testing this:https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/blob/main/minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensorsAt Strength 1.0 + Euler + beta + 8 stepsFeels much less erratic than v1. Still there is some blur and sometimes it adds weird details to video but feels much closer to usable.Anyone else tested it? Anyway to improve its outputs?
The new spectrum node upgrade broke the preview progress of sampler.... Is ogre
ok who's the anon who said to use the int8 q8a8 model loader node, still testing it but as far as I can tell it's a huge speed boost for no discernible quality lossmay you always get trips anon
>>>/wsg/>>6211018
>>109509674update it again
>>109509679i tried it out as well and i'm genning 1mp now. amazing
>>109509686ah shit >>>/wsg/6211018
JC Denton runs into someone familiar in the cityhttps://files.catbox.moe/5d7udj.mp4
>>109509679>>109509698official nodes support it if ur on cu130
>>109509698>model type flux2is that right?
>>109509570Content aside. How did you make 2 minutes animation like this ?
>>109509711no idea, it just works
I need to add more swap. VAE memory management still fucked.
>>109509716By genning four 30 second clips and editing them together?
>>109509696I did it still broken
>try Minimax>30 mins for a 4-second 0.2 MP clip on my iGPUNot practical, but possible! I thought it was doing 20 steps for a single frame at first, so I'm glad I let it finish rather than abort.
>they trained minimax on Seinfeld but not Frasier
>>109509724>>109509732What app for combining videos ? Might do it someday when Minimax matures
>>109509609Post links to the images please. I need Kanna butt...
>>109509708https://files.catbox.moe/nv7xmf.mp4
>>109509751https://x.com/ai_daihuku23/media
>>109509740>What app for combining videos ?anistudio
>>109509740Brother if you can't figure out how to stich 4 clips together you ain't making all that, sorry.
You think it's possible to actually make full length episodes or even a film?>>109509758Thanks anon. Kanna sex.
>What app for combining videos ?
>rando anon asks retarded question>my brain; awesome another reason to dogpile>>109509740dumb fucking retard lol
>>109509739nobody wants to gen that niggas big ass forehead
>>109509739indeed. Friends is missing a couple of characters as well.Also, two things. I posted it on a previous bake, but I still want some answers. Does anyone else get a few bad gens (bad as in not following the prompt correctly) and then, after it gets it right once, the next gens are generally correct? Don't know what may be causing this.Also: how in the fuck does h3 nail character's facial movement when the model was not specifically trained on a certain character? For example, I genned a few of the kike lord just starting from a still picture and it pretty much nails it. I know for a fact the model is not trained on his face, so how in the fuck can the model do this? https://files.catbox.moe/14dbtv.mp4
ayy yoo cuh what app to play on my iphone
>>109509778If I had the autism and didn't have a job I would absolutely spend my week doing thisI might anyway
>>109509778nta but I wonder if you could extend a shot by using the last frame and use img2vid. might be able to stitch some shots together for a seamless long shot
>>109509761>>109509765>>109509772Using adobe after effects is overkill. I asking for something simpler for AIslop
>>109509792>after effectslmao, brown nigga what are you even on about just use davinci resolve you el tardo its FREE
>>109509686Adding a reference video helps with the dance moves, but bleeds in the style. Maybe it could work with a more specific prompt.>>>/wsg/6211025
genning at 0.4 mp really hurts the quality. I need to gen at 0.6 mp minimum but it takes forever to gen
>>109509809img2img the reference to match the style
>>109509739>uncshows nobody nows
>>109509804Still too overkill. I want something simpler similar to "Webm for Bakas" i use
>>109509818just go visit the retard store and buy the retard editor
being able to get 15 seconds of content in under three minutes customized with any subject imaginable for my very specific fetish is absolutely wild
>>109509821okay where are the retard store and the retard editor ?
>>109509818>Webm for bakas
So far I'm not seeing a difference between the text and reference model, it seems to also be able to handle references well when properly prompted including style transfer.If not for the loras I would stick to the ref model but this isn't adding upPlease advise
>>109509824low resolution fetish?
>>109509802Based. Prompt?
>>109509832Great tool. It just works
>>109509698Doesn't work with lora(s?). tried with turbo and getting static
Lets just admit. We all suck at this and this guide https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md is retarded.
>>109509851lora mode needs to be set to stochastic
>>109509842It's just a wrapper for ffmpeg nigga, an outdated wrapper for an ancient version of ffmpeg
>>109509802>>109509841my down syndrome retarded low IQ queen returnspls catbox these gens
>>109509860I want a GUI. not fiddling around with cmd
>>109509882You can have chatgpt slop you up something modern that'll do the exact same thing in about 30 seconds
>>109509840I admit to not caring about visual fidelity as much as the average viewer but .3mp and the firstblock cache get me to a very solid place. Taking firstblock off it's still under five minutes. I get that people can get faster gens but this is good enough to blow my mind still
>109509882Why are you even here if you can't ask google to give you a one shot command>windows>AIkeeeeEeek
>>109509882tell ai what you want so it makes you a .bat that you can double click nigge
>>109509898>>109509895sigh.... you guys really rely on AI on literally everything
>>109509905have you been living under a rock? modern AI is better than 90% of programmers.
>>109509905Bored on a sunday?
https://civitai.red/models/2839513/male-ass-h3that preview image made me laugh my ass off (nsfw)
>>109509911>>109509916Im not going to pay Sam Altman to create a simple app
>>109509905thats fatherless behavior
need more prompt donations to test my 2pass WF.
>>109509925Are you too stupid to run the model locally?I think /sdg/ is more up your speed, it's a place for low ability posters
>>109509925Fine anon. I'll do it for you. Tell me what you want. Just combining videos in a GUI? That's it?
>>109509936>Telling people stupid without providing a solutionLol
Best text model for interpreting the prompt guide and improving your own prompts? Working with GLM but I think it's too wordy and adds too much the AI can't really interpret.
>>109509949Literally it. two videos into one. With audios. In standard h264 format Nothing else
>>109509905ur probably the dumbest poster itt, congrats
>>109509905>rely on AIIt's wasn't actually possible to create useful apps in 30 seconds before AI.
He's the same bored retard trying to bait, you can tell by his constant adversarial tone. Ignore him or post wheelchair gens.
>>109509905>sigh.... you guys really rely on AI on literally everythingSlopping up a GUI wrapper over ffmpeg with your precise wants is like the most fit usecase for LLMs
>>109509988this. after the waifu app thing he pointed out it stopped being funny.even if it were a real retard, it stopped being funny, just move on.
>>109509905says the guy who is too brown (low iq) to use fucking ffmpeg, lmao
how car can you push the model off of the trained 15s limit theoretically? I want more
Is it worth getting anything above 32gb vram until you manage to cross the 256gb mark where you can run quant 4 versions of 500b models?I just don't see the point of that dead man's land zone in between. MAYBE 48gb is justified as it usually comes in a standard card and allows you to run extra stuff ontop of your main LLM like additional small video gens or image gen.
>>109510006you can gen longer, its just too costly
>>109510006H3'S THEORATICAL LIMITS ARE TOO STRONG FOR YOU GOONER
So Minimax H3 can't do explicit sex, but what about implied sex? Like sex under the covers, or with a POV camera? Or movie sex scenes? Have any of you guys tried it yet? Can you make it convincing?
>>109509961Alright. The AI is cooking
>>109510014>theoreticallyI can trade off resolution or whatever, but how much longer are we talking before the gens fall apart?
>>109510038Nice slowmo, homo.
>>109510027H3 can do anything if you reference the video and prompt properly.
0.5MP, 20 step, 5s. Default comfy workflow: 384s. Adding patch sage attention, Mem Eff Sage attention, Low VRAM attention: 406s. T-thanks I guess?
>>109509919the third one...
>>109510049too much effortit works, but we need a better approach. I think with some light loras and careful prompting we can achieve greatness
>china releasing state of the art uncensored local models to destabilize the westi love this psyop
Do any of the turbo loras actually work with the reference model?
>hurrrrr minimax can do anything with a reference!then I may as well just watch the reference
>>109510071
>>109510087please be trolling
r2v is funhttps://files.catbox.moe/0vl11z.mp4
>>109510084Yeah, all of them.
>>109510084yes >>109509614
>>109510078china won a long time ago. the playing field wasn't fair from the start. they arent heavily regulated, have no issue with straight up training their ai models off western sota models(deepseek/kimi), embrace ai as a culture and far stronger educational backgrounds. the west had no chance.
>>109510094I'm very curious if you can get the lip smack from JC in there
Alright that's it. I'm downloading more RAM.swapoff /swapfilerm /swapfilebtrfs filesystem mkswapfile --size 12G /swapfileswapon /swapfile
swapoff /swapfilerm /swapfilebtrfs filesystem mkswapfile --size 12G /swapfileswapon /swapfile
>>109510096>>109510100I will try, it seems to be more ridgid with references.
I'm using a default workflow for minimax and my outputs have a lot of artifacting. Do i need to crank steps? Upping resolution doens't seem to really help.
>>109510073>too much effortIt's really not. Add a 5 second video. Say it's a weak_reference and it works as a lora for that concept.
>>109510108I think if you used that particular voice clip and prompted for it, it would probably work. my source is just a short .mp3 of random JC lines.
>>109509652>Anyone else tested it?of course, it's the current best by far
>>109510049Wrong, idiot.
>>109510121yes
>>109510128Unlike you, I've actually been experimenting with R2V a lot, and neither reference nor weak reference works well for audio.
>python.exe
least helpful general award goes to /ldg/
>>109510159Not your tech support Rajeesh
punch this fuck
>>109510159u probably asked a brownoid question that an allm could have answered
>>109510142>>109510148skill issue. Follow the prompt guide.
>>109510171this, bodied that saar freak
>>109510159lurk moar
JC on global warminghttps://files.catbox.moe/4zsxuz.mp4
>>109510171>>109510180>>109510198Dead internet theory proof #14
>>109510183I know more about the prompt guide than you do, brainlet. That's why I'm aware of it's limitations and issues.Everything you say means nothing until you produce an example proving me wrong.
>>109510215yeah dead internet theory is when people don't spoon feed you basic questions that could just be binged.
>>109510215>being called out for being a retard>must be botscan't make this shit up.
just asked my villages local facebook group and they said "pls redeem sir"
Did more testing with turbo and reference model, the model diminished in ability when working with more than one ref and more complex prompts
>>109510215u were too big of a pussy to link your question as you complain because u know it indeed was a brownoid tier question that an llm could have answered.
>>109510240impressive, now make them fart really loudly.
>>109510244You're not helping sperging like that, especially when a brown man opened local image gen models to begin with
What's the reference limits? One picture on one video? 2 pictures on 1 video was scuffed. 4 pictures was scuffed.
>>109510257>no linkcase in point, concession accepted.
Ya I think I like this 2pass WFhttps://h.uguu.se/odfOgEGh.mp4
>>109510240Kino
>>109510244>>109509740
>>109510244>>109510225>>109510223Dont talk to me, Clanker
>>109510268I'm not him based on your same combative nature you're playing both sides>>109509988
>>109510273this was posted on preddit like a week ago or something
>>109509857>lora mode needs to be set to stochasticsame shit
>>109510102I guess we'll see what happens. Historically China hasn't had much tolerance for companies who become so powerful they're like a branch of the government, certainly not if they're lead by headstrong people like Musk.
>>109510280No worries — I'll step back. If you want help with something later, just say the word.
>>109510278>whats 2 + 2, dont give me that 4 bullshitu got ur answer, u just didnt like it
>>109510293i guess it's not built to work for h3, because it works for wan.
>>109510285was it? link? It's an old prompt but I don't post my shit on reddit.
>>109510297nooo saaar i need easy peesy app to plug in video so it come out like nolan film saaaaar please help
>>109509961Almost done. Excuse the AI for going overkill.
>>109510280Unfortunately I'm abliterated, so your anus is going to be obliterated.
>see kino h3 gens that japs are posting on X>they’re better than anything posted here
>>109510010Coming from 5090 + 5070 Ti. Yes, it is absolutely worth going for 48gb, especially if you only have a 5090 in your system.Getting an additional 16gb card opens up the Gemma range completely along with other models in that region and will give you a shittton of more context. It will also allow you to do stuff like have Gemma 12b running on the 16gb card to caption images while 5090 does the heavy lifting in video generation, something I noticed was pretty damn useful after playing with H3.I'd say it's a pretty optimal combo for local.Is it worth going higher than that by adding another 16gb? I doubt it, but then again I have no idea how much of a speedup you'll get in lower quants of Dsv4f by adding another card into the mix.
>>109510325AI slop, Japan: :)
>>109510317Thanks bro. Better than the rest of assholes here
>>109510325Japs are some of the biggest goyslaves, probably using maximum quality API H3 with promp enhancer.
>>109510326Too bad most cases can't fit a second card worth a shit, that's my biggest blocker from doing that
>>109510317which AI creates this exact style of user interface I've been seeing everywhere lately?
>>109509679???Takes twice as long for s/it for me.Likely won't wait to measure quality.Not sure if a troll post or I am missing something.
>>109510345Most local models can do this, this is a simple wrapper and basic bitch UI
>>109510156>640GB ought to be enough to run any Python program
>>109510351yes nigge but what ai is it specifically that goes for that style
>>109510345im using claude, but I just told it to make it modern. I didnt give it any other design tips
>>109510340if you download his vibe coded app you are 100% joining his proxy or botnet network.
a bombhttps://files.catbox.moe/ptgekw.mp4
>>109510355Seeing how you're shitting up the thread non stop you probably can't afford to run it>>109510361I doubt it but it would be funny if his neg hole got pozzed
>>109510377Anyone can afford to runit
does anyone have issues with git and github those last few days? it's unstable as fuckhttps://news.ycombinator.com/item?id=49198302
>>109510343Yeah I know, most cases weren't built for these cards.I have my 5070 Ti zip tied to the side of my case. It has the exact same cooler as the 5090 so there's no way in hell they fit in there, even though it's an ATX case, as they're basically 3.5 slot cards.Only case that makes sense is basically the Phanteks Enthoo Pro 2 Server Edition.That's guaranteed to fit the cards while leaving enough of a gap between them. I'll buy one of these soon.
>>109510326I have 10GB 3060 still lying around before the upgrade to 5080 (the case is too small as the other anon mentioned)does multigpu sampling works in comfy? I want better H3 gens. If yes, I'll buy a new case
>>109510325I made lots of kino but janny wouldn't like it.
>>109510361nah, I just happen to be in a good mood thanks to a banger ref2va goon gen. it was delicious
>>109510390Have you tried codeberg or gitlab (local hosted)? Github is working ok for me and locally forego works rock solid
>>109510362Nice. Fun fact: twin towers are already missing from Deus Ex skybox despite being made before 9/11
>>109510325sasuga asian jeans desu
https://files.catbox.moe/wfcu86.mp4Two passes kinda work, but idk about samplers, sometimes shit deep fries too much.
>>109510413They just forgot to add them in?
>>109510420Hardware restraints.. or so MJ12 would like you to think
>>109510417https://h.uguu.se/rfGgBSte.mp4 (NSFW)er_sde/beta57, 20pass but stop at 10into turbo euler/beta start at step 4 out of 8I was getting deepfried results when I let the first step go all the way.
>>109510397I tried multigpu but I couldn't get it working, threw a bunch of errors, so I have no idea how well that actually works.Someone else can chime in there.But when it comes to the LLM related stuff it's absolutely worth it stacking GPUs.In general I would imagine we're moving towards parallelism eventually with all of the AI stuff, so it's only a matter of time until multigpu is a standard with all of these systems and keeping those old cards around will pay off.
Use Picture 1 as the exact character identity reference for JC Denton, keeping his signature black sunglasses and stoic features. Extract the vocal timbre, frequency, and speech pattern from Audio 1 to build a voice clone.Cinematic medium shot of JC Denton. 0 to 1 seconds: JC Denton is sitting in an art studio, in front of a blank white canvas and is completely silent.2 to 5 seconds: JC Denton paints an anime style Hatsune Miku on the white canvas and is completely silent.6 to 7 seconds: medium shot of JC Denton painting and is completely silent.8 to 10 seconds: JC Denton looks at the painting and says "That's a nice Hatsune Miku, for sure.".https://files.catbox.moe/jsd94a.mp4
>>109510340AI is complete. Feel free to try it.This requires ffmpeg in path:https://files.catbox.moe/nmojie.7zThis has ffmpeg bundled already:https://files.catbox.moe/j922x9.7z
>>109510447his neovagina closed up
>>109510447eeeeew FUCK
>>109510325Not surprising. Japs are more creative on average, and creative people get the most out of AI.
>>109510463GPT-5.6 says this has like 5 backdoors in it
>>109510463I might test this laterThank you for your contribution
>>109510325Examples?
>>109510447yeah, I'll try that, thanks
>>109510463Thanks for spoonfeeding me anon. I just want to make a mini scene of my successful gens
>>109510463>exe>no sourceyeah, anybody stupid enough to open that deserves to get cyber raped
>>109510490Post it or you're ungrateful fuck
>niggerlicious voodoo going on in comfyui just raped my gen speeds and i see why nowwhat in the
>>109510498Stop using bots
honestly cant believe you niggers are genning on windows
>>109510491what's the point of including the src when the anon isn't a programmer?
this shitpost was made by Debo. Know his post pattern
https://d.uguu.se/UABTihpT.mp4birthing machine
>>109510491Yeah, where the fuck is the source code? 2 exes and one is just massive as fuck. Where is your source code? I won't be trying your program, as you compiled it for Windows and provided no source code repo I gave it the benefit of the doubt, and even downloaded both sip files, yet no source code. Crazy.
>>109510305>>109510293Sorry, it works actually, thanks. I selected an fp8 model by mistake
>>109510538Why are you helping someone that begged like a bitch because he's too retarded to use a basic web search?
>>109510538Bro if you're going to whine for that, just ask chatgpt to make one for yourself
>>109510550Because I can and it's the right thing to do.
I didn't see the namefaggotery, not replying further schizo
>>109510554You don't belong here, you're using a trip and feeding a troll that has nothing else going on and does this daily. You can identify him by his constant arguing and never having a solution to any subjectRead the OP
Yeah, very happy with this workflow, time to gen some kinos
>no gensYour cums dried out along with your creativity
>>109510538one is massive because it has ffmpeg and other libs bundled in it. i didnt want anon complaining it didnt work because they didnt have ffmpeg in their PATH env variable.the src is 2.5gb. just feels pointless posting it when i know you arent going to look at it
>>109510574you have no gens either thoughbeit
>>109510491Some source file being provided doesn't mean the compiled exe, which is what the brown anon who needed to be spoonfed is going to run anyway because if he had any brain he wouldn't be in this position in the first place, is in any way related to it. The way to go here is to decompile it and ask an LLM what that shit does.
>>109510578>the src is 2.5gbHAHAHAHAHAHAHA
>>109510578>the src is 2.5gbWhat the fuck. Is that just what Rust is like?
>>109510594if the nigga can't figure out how to vibe a simple app he isn't going to know how to decompile that shit nigga. this guy is 100% getting pwned one day.
>>109510594Do you know about reproduciblity and the importance of reproducible builds?
bollywood is finished.https://files.catbox.moe/igw4b6.mp4
>>109510615pretty good
>>109510615>bollywood is finished.More like bollywood is just ramping up
>>109510604whoops, I included multiple folders that had binaries in them. it's less than about 15MB.
>>109510632>it's less than about 15MBhahahahahahaa. Still ridiculous, but slightly less so.
>>109510656it's 99% cli.win32-x64-msvc.node from node_modules. needed for @tauri-apps/cli. either way im not posting sauce so either use it or dont
Is a video or image better for copying a art style?I want to copy how the anime looks
>>109510578>the src is 2.5gb.ok now we have confirmed the retardation required to spoonfeed a brown to this level, u have to be even more retared urself
If you're on at least 16gb, a 4000 to 5000 series, remove kijai's low vram node, 2ldr the math doesnt math up and causes inconsistent/slower speed, and remove all memory related startup flags, they're all memes that also slow it down and fuck with its memory management2ldr for the 2ldr let cumfart manage its own memory, only use sage attention and the h3 sage attention patch. i just got my reference model speed at 0.8mp 10s down to 64s/it before the compile speed boostthats.. basically every single speedup node removed for a speed BOOST. rough buddy.
this time, flux 3 lost. that stupid team made a mistake by ignoring us. local is now strong, even without flux 3
>>109510517Kino
>>109510687*actually 55s if i stop bringing up the 4chan tab kek
>>109510686>didn't read rest of reply chain>u>urself
>>109510696not only that, but their image model will probably have the same impact, I really can't waithttps://xcancel.com/MiniMax_AI/status/2086253065657790895?sort=Likes#r
Whats the most complex h3 gen youve made so far?
>>109510687>>109510705No one genning woth 16 gb of ram. Thats insane
>>109510712so Krea is dead long live MiniMax?
>>109510712Are you joking, holy shit. What a great year for this hobby
>>109510720I do. Works with Krea2 and H3.
>>109510720vram, hence 4000/5000 series you dingaling.and there actually are a few here genning on 16gb of ram, with an 8gb vram card.
>>109510719a conga lone of women with dicks fucking each other in the ass
>>109510740share it bruv I wanna see
>>109510732no you are not fucking fag. why lie?
JC drives a car and somehow breaks physicshttps://files.catbox.moe/5fcxun.mp4
>>109510751Did you read the links in OP?You're supposed to ignore his screeching, he spends all day baiting.
>>109510687>remove kijai's low vram nodeWhich one ?
>>109510746dont want jeets stealing my porn
>>109510726>What a great year for this hobbyI was about to call 2026 a dud but yeah, now it's a great year
any llm can shit out a 5kb 120line python script to pass args to ffmpeg to both extract last frame of a video or combine two videos. can be written by literally just typing into google and bro waits an hour to download a 3gb rootkit instead of just learning how to pass a few args in his cli. fookin grim m8. has to be trollin..
https://x.com/ArigatoLin557/status/2084993384306135094>the superior japanese gens in question
Someone prepare to make the thread
>>109510751???
>>109510770No shit do you even read OP?It's the same pattern
>>109510784>>109510784>>109510784>>109510784
>>109510757>my vision is augmented.idk why but this line cracks me up every time.
>>109510757>my vision is augmentedGod damn it...
>>109510732>>109510781They're bots ignore them.
>>109510770>anon already satisfied with his product>sleek ui with plenty of options>complaining about someone else's giftyikes
Guys can someone vibecode a multi GB interface for me as exe without sourcecode so i don't need to type in a ffmpeg command? I will run it as admin because i trust you. Thank you
there we go, the car moves forward off the ramp.https://files.catbox.moe/zkc0eu.mp4
>>109510826Working on it, sir.
>>109510826I'm assuming the requester and provider are the same person, in an attempt to get random anons to use the fileYou need to be extremely lazy to ask other people make things with AI for you, when you can just ask AI directly instead
>>109510837Looks like a gmod animation
>>109509770Windows live movie maker
My dumbass stuck scrounging for free credits and cheap GPU rentals because Nvidia told me 6 years ago 8gb of vram was enough! Somehow still 4x~ cheaper running serverless and 5-6x cheaper renting if you constantly Gen vs API costs at .17 cents/second.
I been using image to video instead of references to video. I feel dumb.
>>109511133Yeah, reference is a bit slower, but way more flexible.
>>109509939Final Fantasy 8 too much