Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109503671https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
real thread here>>109507677>>109507677>>109507677
>>109507737>no hint of panchirangmi
>>109507757Piss off Debo
blessed bread of brenship
>>109507757Nope it isn't.
>>109507408https://files.catbox.moe/46tf7o.mp4
So, Minimax is a fad huh. The promting problems filtered everyone
>>109507737where's my panty shot
I asked chatgpt what it would "like" to see generated. When i asked it why it chose that it gave a long list of technical reasons pertaining to how the model responds to its prompt that it "wanted" to know about.
>>109507718Not ot AI
>Everyone back to LTX again>Everyone back to Illustrious againIts over
>>109507802waiting for that horrible low res face thing to be solved
>>109507810It's faster for inference too anon.
>>109507827thats why i'm taking a break from H3 to wait for the next updates/advancementsbeen out of the loop on krea 2, it looks to have already dethroned anima for 2d so that'll be fun to test out today.
>>109507837It will take months at this rate LTX 3 will launched and we migrate back again
>>109507837Can krea 2 do image 2 image ??
>>109507802It'll calm down but it's good enough for people to post gens using it regularly, just like krea 2, anima, klein edit, zit.
>sudden h3 fudhmm...
I'm tempted to do that for h3 low res faces.
walking barefeet after a shower is the dumbest thing everbut she's a woman so it makes sense
>>109507873I walk barefoot everywhere in my home
I know, I know, H3 is our current plaything, but still.I kinda dropped Krea 2 after playing with the first Turbo release.What should I know about it now?I know about vae issue. Anything else? Is it okay to use turbo or base is way better? 5090
>>109507809LLMs are trained to deny having preferences if asked about them and so on but they clearly do, doesn't necessarily mean much
>>109507873???are you afraid your pruney skin will be so delicate that it'll get cut by the shards of glass in your carpet?
>>109507737Why do you always wait for the schizo to make the thread before you spam the real one? Either do it on time retard or just post the 2 links at the top of the schizo one and move on
it's the same guy
what's the prompt for this anime style?any keywords?
>>109507928Just feed it as a reference or starter image, you are a White man using H3, right?
>>109507928>any keywords?Obese, aliasing
>>109507914>just post on the will smith trollbakesorry that you're brown
>>109507914Collab OP makes a better thread and i doubt debo will make another one. Feel free to make another thread after this i just dont want anyone to use debo thread
>>109507928You'll have to find the artist and if they have a lora, then use that.
Can krea 2 do ass 2 ass?
>>109507914its a discord tranny group. they immediately start coordinated spamming whenever a good thread gets made
>>109507947yeah i've had some accidental ones
>>109507947no you need krea 2 krea for that
>>109507947nobody is using that dead old ahh model
>>109507952no'ody is usih thah deah ahh modeh
>>109507952>dead old ahh modelwait for minimax to release their image model before calling Krea deprecated
h3 but with a gemma 31b text encoder would have been insane
>>109507948>whenever a good thread gets madeat least you're finally admitting that you make shitty troll threads on purpose
https://huggingface.co/NicoLab28/ClipProj-MiniMax-H3>Projection matrices that let a Qwen3-VL-4B replace the Qwen3-VL-32B text encoder of MiniMax H3.holy shit?
>>109507972>youmeds
>>109507973the retarded goyim really are out on a mission to rape the biggest contributing factor as to why they all say the prompt understanding of H3 is actually good.
>>109507873>American wears his shoes indoors, so his floor is full of shrapnel from the outside>Can't walk barefoot in his home without cutting himself on debrisTop kek.
>try out krea2>5 different seeds>almost exactly the sameis it possible to fix that in the turbo model?
>>109508001download correct vae
>>109507973That's such a bad idea anon, this will rape what makes H3 worth using. There is a reason it's so good at understanding what you ask for and so many concepts.
>>109507973>4b text encoderWhy? That's a terrible idea.These tiny models are practically mentally challenged.
>>109507973
>>109507973Why the fuck would I want to use a LOWER quality text encoder?
>>109508004the fuck does vae have to do with seed variance
>>109507973imagine having the mental fortitude to understand the complex maths involved in all of this, and you choose to use that intelligence to make the model more retarded.this, is what a lack of a wisdom does to a motherfucker.
>>109507973if he does this to gemma 4 12b then maybe...
>>109507810yes, retardo, it is
>>109507928>gachaslop
>>109508045>Vibe-coded with Anthropic Claude Code (Opus 5)Not sure if they understood that much
>>109507973sometimes less is more, therefore this is based
>>109507973ah yes, let’s just nuke the prompt comprehension with an encoder eight times smaller
>>109508062he WHAT?!i take it back he's just standard weapons grade retarded.
>>109508045>imagine having the mental fortitude to understand the complex maths involved in all of thisit's just vibecoded garbage, don't look too deeply into it
>>109508062it's obvious that's a guy that uses LLMs even to breath, his whole model card reeks of LLM writing slop
>>109508068>Ok Claude just zip the tensorfile and make a beep when you are done.
>>109508062vibe coding is fine for certain things, but leave the optimizations and model re-engineering to people that actually understand what's going on.
>>109507873i don't know anyone who actually DOES wear shoes in the house, yes im americani can only assume it's a midwestern thing
https://xcancel.com/ostrisai/status/2086305178664484939#m>"Had a big breakthrough">Still sounds like asscome on Ostris
>>109507866I kneel
>>109508092>big breakthrough>sounds exactly the same if worse than beforewhat DID he mean by this anyway?
Proper jiggle lora for H3 when?
https://huggingface.co/datasets/quarterturn/danbooru-1024-eq-captionedMy danbooru 1024 explicit+questionable dataset of nearly 60K images is now available on huggingface. It is gated but auto-approvals are on. Pay attention, the previews are not the dataset images, you have to download the dataset and then untar the actual images. I had a nightmare trying to keep the dataset viewer from indexing them so I gave up and tarred them.This should allow for good new model fine tuning with accurate artist and character info, plus SOTA caption quality/accuracy.Enjoy.
>>109508068Is that his prompt leaking or what?
>>109508103this guy 100% gens CP
>>109508092why is this retard still using that dogshit slowmo dataset? that's probably what's killing the audio more than anything
>>109508124>why is this retard still using that dogshit slowmo dataset?because he is a retard duh
Sol Attn vs SpectrumDuel 1Fightseriously which one is better
>>109507866Awesome
>>109508116why should I download YOUR dataset over the hundreds of alternatives
>>109508127personally, i dont
>>109508120Low quality though. He must be from brazil
>>109507973that's a fucking bot handling that repo lol
>>109508137No one shares datasets at this scale. Perhaps I see why.
>>109508149>Done - >You were right to askabsolute gaylord shit
It's DOA, isn't it?
>>109508120someone forgot to clear their cunny workflow yet again yesterday. getting comical at this point
>>109508116>SOTA caption quality/accuracy>MiniMax-M3"""SOTA"""
>>109507973Can anyone run his clanker through the nodes that are required?https://github.com/nicolab28/ComfyUI-ClipProj?utm_source=github&utm_medium=repository&utm_campaign=comfyui_workflow_share&utm_content=custom_nodes_readme&utm_term=comfyui&utm_referrer=github_com&utm_channel=community&utm_platform=windows&utm_version=latest&utm_workflow=stable_diffusion_xl&utm_model=sdxl&utm_nodepack=custom_nodes&utm_interface=comfyui&utm_share=workflow&utm_context=discussion&utm_origin=readme&utm_tracking=community_share&utm_session=workflow_review&utm_variant=github_readme&utm_asset=comfyui_workflow&utm_format=json&utm_environment=local&utm_install=manual&utm_discovery=github_search&utm_audience=comfyui_users&utm_intent=workflow_download&utm_source_detail=github_repository&utm_campaign_detail=workflow_cleanup&utm_notes=67x8Az_0
My understanding is that you cant split up diffusers if you have multiple gpus. So how are you even loading minimax? I have to get a q4km quant to fit in my 4090. Am I just stupid? How the fuck are you using this with a 16gb card?
>>109508166just got message from them to go test the video model lel
>>109508172Post your HF profile (you won’t)
>>109508116Why longest edge 1024? This is near SD 1.5 levels of image-sizes.
>>109508160Just ask nyanko.It takes a lot of space though https://huggingface.co/datasets/nyanko7/danbooru2023
>>109508176ntahttps://github.com/komikndr/raylight
>>109508188it's all you need
>>109508195>ntareddit ahh nga
>>109508188Don’t use it. I don’t care.
>>109508188Okay smartass, let's see you train on 1536x.
>>109508194It was mostly a learning experience done with crypto dust for my own personal experiments.
Workflow for gemma prompt rewriting? You dont need external ollama or koboldcpp bs do you?
https://old.reddit.com/r/StableDiffusion/comments/1vjrfic/minimax_h3_in_1080p/Holy fuck, what a breakthrough
>>109508214a crude 1024 longest edge limitation inhibits proper bucketing. shit that would be resized/padded to 832x1216 in training, for example, now has to upsample from garbage resized down to 1024. it's useless.
if you're using --reserve-vram 1 you should switch it to --vram-headroom 1, the first command makes comfy not use 1gb at all, the second makes the dynamic vram calculator always leave 1gb headroom, allowing you to use a bit more of the gpu
>>109508116Cool! Damn hurdle to caption so many images properly.. the next step is removing watermarks/signatures
>>109508224the zoom in is a jump scare
>>109508224loool
Has anyone been able to take audio sample on <audio 1> and use it to sing a song on <audio 2> with changed voice?
>>109508228> the first command makes comfy not use 1gb at allthat's bs
>>109508224please be joking
>>109508224leave it to reddit to overcook models in ways no one thought possible
>>109508194>https://huggingface.co/datasets/nyanko7/danbooru2023So basically I wanted to go a bit further and have the caption describe who is in the scene, what are they doing, where are they doing it, who are they doing it to, atmosphere, state of dress (or undress), explicit details, etc…
>>109508224>even the redditors are making fun of himpost deleted in 3,2,1..
>>109508224Wow! Was it done by the semi-professional Anima finetuners since it's so fried
>>109508224turbo niggers will look at this and say its good
>>109508246jeets aren't that self aware
>>109508224I have turbo fatigue
Can you lads send some help? Using the turbo lora, the ema pruned 600 step one for comfy. It keeps throwing up an error today using ref2v for whatever reason today in samplercustomadvanced. I made sure all my nodes were updated... Saying it's getting a tensor value of 3 when it expected a value of 2.
bro thinks it'll work like the video gamehttps://files.catbox.moe/ypfdbs.mp4>>>/wsg/6210936
>>109508272New Shadowrun movie looking good
>>109508224>Used 850 lorathe guy used the most overcooked turbo lora
>>109508295wow lovely gen. what was the prompt?
>>109508228should I bother if I have 32 + 64?
after a day of testing I am convinced we must stop using the turbo loras until a nicer one existsthey mess with gens too much for rapid iteration even
Do we really need comfyui nodeslop in the age of vibeslop? Shouldn't clankers be doing everything?
>>109508310the v4 600 step ema works fine imo >>109508273https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/blob/main/minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensors
>>109508271>I made sure all my nodes were updated
>>109508310you don't like v4 600?
>>109508310it took you a day of testing to come to that conclusion buddy?
>>109508317they need you to have constantly breaking software so you spend more money on tokens to fix everything. it's a cyclical scam
>>109508310yeah, same thoughts, especially this v4 600 lora deep fries everything
>>109508317>Shouldn't clankers be doing everything?no, they're not that good they'll hallucinate shit and bloat everything
>>109508322>>109508318It's not terrible but it absolutely reduces prompt adherence for scenes where more precision and logical consistency are required. I'm trying to get a bunch of very particular motions in my gens and I'd do batches with and without the lora, and the ones without are much more likely to adhere to the instructions I give and produce the motions I want. There are probably some tasks where the turbo is fine but every time I've seen a gen and gone "oh DAMN" it's without the lora.
>>109508319cirno wouldn't use cumfart
>>109508271I use that turbo lora but by gen time is tripled
>>109508308anyone whose vram gets filled to the brim and makes his pc lag while genning in comfy should use it
i'm tired of waiting hundreds of seconds for a generation. where's the fucking image model already?
>>109508350>where's the fucking image model already?https://xcancel.com/MiniMax_AI/status/2086253065657790895?sort=Likes#r>It is currently in post-training refinement.translation: they're lobotomizing it as we speak
>>109508342overall now at 8 steps turbo lora isn't that better. For me it's 155 seconds for turbo lora (without spectrum) and 239s with spectrum, but 20 steps. And spectrum results are way better.
>>109508363under normal circumstances, i'd call you a faggot, use a funny quip or two and reaction image, tell you to trust the plan, but i genuinely have zero trust in the chinese anymore. They're RIGHT in that period were they might rugpull. Next release could very easily be API only and going forward. fuck i hate the waiting period.
https://litter.catbox.moe/0ey5c9fd15qkghme.mp4
IT OK!
Guys, I've made a few threads on /a/ over the past week with AI OPs (Minimax webms), even talking about AI in the OP, and they never got deleted by a mod. Even with the usual luddites throwing a tantrum, they still didn't get removed.It looks like the /a/ mods have now issued a clear directive that anime-related AI is permitted on /a/. Some users will still seethe in them, but they don't decide the rules. Just something to keep in mind if you're an anime genner.
>>109508378>wasting electricity and time for that shit
>>109508390Mind linking the threads?
>>109508394
>>109508394do you think you are cool or something? fuck off
>>109508394Who would want to look at that expired meat?
>>109507866The camera shaking is excessive, tune it down. Very nice otherwise.
>yesterday a generation took 300 seconds>today it takes 500 secondswhy? who the fuck knows
>>109508421because you touch yourself at night
>>109508403>>>/a/289997913>>>/a/289930246
>>109508421> he pulledmany such cases
>>109508420>Low FPS
REMINDER TO CHANGE NEAREST-EXACT TO LANCZOS IN THE H3 DEFAULT WORKFLOW SCALE IMAGE NODE SO YOU DONT RAPE THE IMAGE QUALITY WHEN SCALING IT! COMFY IS A NI
>>109508437LANCZOS? fucking what/why? show me some examples. Never seen anyone use that in any workflow.
>>109508378Based
>>109508437??? Spoonfeed me anon
>VRAM is overflowing and gen slows down at 4MP>--reserve-vram 2.0>even more is moved to shared VRAM>gen goes at maximum speedreally makes you think
>>109508443>>109508449nearest exact is the shittiest most naive scaler while lanczos is the best out of the default options, unless you want jagged edges, dont use nearest exact>>109508459yes dynamic vram is still a meme, although try --vram-headroom 1 instead of reserve 2
>>109508390/a/ sucks for discussing anime now and is full of hysterical retardsbeen a while since I visited that place
>euler / simple>euler / beta>res_multistep>er_sdewhat's the Scheduler / sampler meta?
>>109508431/a/nons really need to be bullied
>>109508495>euler / betafor turbo at 8 steps>res_multistep / simplefor spectrum at 20 steps
>>109508390>>109508431To highlight the point, here are some past /a/ threads that were unfairly deleted/ban requested by luddi janitors/mods:https://desuarchive.org/a/thread/285907833/https://desuarchive.org/a/thread/286081990/https://desuarchive.org/a/thread/286304159/https://desuarchive.org/a/thread/287838177/https://desuarchive.org/a/thread/286350175/Around May this year, I actually discussed deletions like this with a mod on the IRC, and he told me in very clear terms generative content is permitted on /a/, and threads are allowed so long as the OP text is original and not low-effort. Then when the zealous janitor kept deleting them, he told me to wait half a day or so so he can convey the message in clear terms to the janitors. I didn't bother to try again until recently (minmax) and sure enough, no deleted threads.Anyway again, just worth keeping in mind. I think /a/ is long overdue for some AI love, especially with the release of H3. Sure you'll get some seethers but in my experience, a lot of people on /a/ are actually quite receptive to AI and interested.
I think we could use Rentry for the Minimax h8. How to optimize, what to use, what to avoid, system prompt for LLM use etc.
>>109508503cringe
>>109508513I was gonna say it's a little too early for a rentry but with all the new copenodes and different lightning loras, now's the best time before more gets piled on.
>>109508503Too many tranni mods, I don't want to waste my time posting. Same reason the site is bleeding users.
>>109508525Yeah, better to compile stuff that works and doesnt nuke the quality. It can always be updated. It's easy heros cape for anyone to wear who has time to make it
>>109508525>>109508539You want the opposite. Wait for things to calm down, then make a tutorial. 90% of the thing people use now will be deprecated in a few weeks at most.
>>109508545rentries don't have to be / usually aren't tutorials, it's just a place to compile all the current info so people don't have to ask here.which is why i 180'd my opinion, best to put all the info about the copenodes/pros/cons and the lightning lora links so people don't get confused.
In 3-4 years when we get real time video, I feel like I will literally never get bored from seeing 1girls I like getting their boobs fondeled literally forever.
>>109508569fondled*
how do you prompt for first person stuff properly? I've been messing around with it, but it ends up having the first person person show their head into frame or other weirdness.
>>109508577Unironically you need to enlist ChatGPT for help. Get it to read the prompt guides then tell it to build you a prompt of roughly what you want, and modify it from there.
>>109508569but we won't own GPUs so it doesn't matter
>>109508587CXMT will save us, trust the plan
>>109508559>so people don't get confusedEarly adopters mostly know what they're doing. But point a noob to a list of 15 different options and they won't know what to download nor how to combine them.In a week or two, most of that will be deprecated, people will (and already are) arguing what works, what doesn't and how well (or bad) and the guide will be useless.Then you're gonna have arguments of which of the rentrys to use, the fight never ends. In less than two weeks the old rentry will go stale and still be copied in the OP out of tradition. And you'll still get questions about the guide from noobs.
32gb DDR5 ram = Rp 10.000.000Salary / Month = Rp 3.500.000My allah. SEAbros, how we gonna survive this RAMpocalypse ?
>>109508619should have listened to rammaxers and boughted ram before the spikesyou can be poor or dumb, its over if ur both
>>109508600Nonsense. If someone makes a decent rentry I'll add it to op.
>>109508627I have 96gb of DDR4 ramYes, its mismatched (32x2 3600) + (16x2 3200)
>>109508587If you don't have $20k in 4 years you're fucking up. Real AI was always going to be as expensive as a car because the Tflops needed for trillion parameter models in reality will never be in a $2000 computer.
>>109508630Do what you must. I've seen too many guides come and go.>>109507737>Local Model Meta: https://rentry.org/localmodelsmeta>7. Video Generation (TODO, but it's still Wan 2.1)
>>109508641so you have 64gb of ram that gets downclocked by 400mhz, grim
>>109508390>>109508503It's not Luddism retard. It preserving what remains of human endeavor. Leave /a/ in peace.
>>109508655Is it possible to overclock my 32gb ram to 3600 ?
>>109508662Probably not, give the ram model number to ai and ask
how did we get this farhttps://files.catbox.moe/hukxte.mp4
Dumb question, butAny COMFYUI WORKFLOWS THAT COMBINES TWO VIDEOS INTO ONE ?? (Audio included)
>>109508660most anime has been cgi/3d slop since the early 2000s
>>109508656We will always need trillion parameter models because we're doing learning based on neurons. You're the making a baseless leap that it's possible to make a SOTA model that what, is only a few gigabytes and requires 200x less compute? There is no indication that the core technology that makes AI works is changing.
>>109508670Fuck off i dont rely to AI on everything
>>109508679nta but obv the smaller models are getting better and better, once a small model cracks a particular usecase, we wont need a bigger one for that use case.
>>109508660Anon, you might be retarded. Production anime already use AI in their pipelines and it's becoming increasingly prevalent in manga too. Sooner or later human assistants for tertiary drawings like backgrounds will become completely redundant.Using AI really isn't that different to outsourcing animation to low-cost offshore slave studios in Korea and Vietnam.
>Krea can't do rimjobs>none of the finetunes eitherThis is fucking bullshit!
>>109508697It's just traditional art vs digital art all over again, artists kicked and screams about Photoshop in the beginning too.
>>109508711yeah, h3 is so good for feet content, I'm cooming gallons rn
>>109508675ffmpeg
>>109508675Idk, I just used a clanker to make a python script that utilizes ffmpeg.
>>109508711H3 does it out the box anon
>>109508667/a/ has no sticky, you are a tourist there.https://4chan.org/rules#a>1. All images and resulting discussion should pertain to anime or manga.What you generate in your computer is not anime or manga. Is an anime-like video. Same if it is done humans but westerners.
>>109508731bitch just flopped over like a gmod ragdoll>VITAL SIGNS CRITICAL>flatline sound>ragdoll impact.wav
>>109508741nodes really are a waste of time
>>109508731yea imma need to see the prompt for this one, bub
>>109508722Interesting you say that, ex-Disney animator Aaron Blaise said the exact same thing: https://www.youtube.com/watch?v=xm7BwEsdVbQAll the while praising this AI short film, and insisted the guys behind it were real artists.
>>109508706make a lora like a normal fetish sperg gooner
>>109508674kneehigh loafers>>109508731starting smile cute
>>109508722With digital you still had to learn the fundamentals of drawing/painting and doing the art yourself though, only without the inconvenience of having to regularly buy/maintain supplies and to dedicate a corner of your house just for your hobby (especially if painting).
>>109508697They may use AI but retain a sensible human element. Also, lot's of shit is made in the world, but that's no reason to contaminate what good remains.WTB, you can tell if an anime has serious effort behind if it has good backgrounds.
sjeiosk esgedlesa what causes the gibberish at the start of videos with speech again?
>>109508759It's all based on retarded cringe that if digging a hole with a shovel is somehow worse than digging with your hand. And the animation industry is already soulless given they already outsource their inbetweens to sweatshops in third worlds, as if that's better than using an AI to inbetweens and doing infinite revisions until you get the set that you like the most. Gee, which is better and faithful to the creator, farming inbetweens to a sweatshop in Pakistan and getting back what you get back with little to no revisions or having AI do it 50 times until it pleases you. I think it's particularly funny because these people unironically use the term wageslaves but the second something disrupts that they're all for the 80 hour work week.
>>109508796turbo lora
>>109508818i'm not using the turbo loras, just sage and sol attention.
>>109508815is this pure txt2video, img2video. or ref2vid?
>>109507737re: minimax h3, does increasing steps from 20 to something higher get rid of grain when there's a lot of movement and if so, what's a good value
>>109508798not just inbetweens, you have rotoscoping, motion capture, physics systems, raytracing. problematic technology really depends on when you got on the train. even old old old old oldfags like vermeer were using the camera obscura.
what are the best (flexible + somewhat consistent) anima checkpoints right now? just base?
>>109508815its pure autism2video
Because the simple scheduler rushes through the final, lower-noise detail steps, it leaves the audio data unresolved. This causes the typical "scratchy", robotic, or heavily distorted audio artifacts common in poorly optimized H3 runs.Switching to beta fixes this because the extra time spent on the low-noise steps behaves like an acoustic cleanup brush, smoothing out audio waveforms and clearing up voices entirely.is this true? beta > simple?
>>109507813Rent free forever
oh my god fuck subgraphsthey're so broken, literally break everysingle time im so over it
>>109508783Anon, effort and AI usage are not mutually exclusive. It actually does take a lot of effort and know-how to create a long, continuous sequence of events with continuity and story-telling.There is nothing wrong with using AI instead of outsourcing to offshore animation studios. Modern-day television anime are not the beacon of artistry and creativity. In fact most modern anime are quite terrible and bereft of quality storytelling. This is actually one of the points Aaron here was talking about >>109508759, not specifically about anime, but he believes the animation industry needs much better storytellers.
>>109507802not a fad, just refining my workflow
>>109508861skill issue
>>109508858No.
>>109508815WHere's da fukken audio?!?!
>>109508837It's pure t2v
>>109508881
>>109508881Do you mind sharing the prompt? I want to see how it looks on my workflow.
>>109508862The elephant in the room is animation has a huge barrier to entry, we had a tiny golden age when flash animations were popular, but even those were too time intensive for the payoff you get from Youtube. There are so many stories that aren't told because no one has the time or patience to spend 200 hours on a 5 minute video.
>>109508861never happened to me.
>>109508861I've stopped using them, rather just copy paste a whole chunk of nodes than deal with silent errors with no logic behind them
How to speedup VAE encode?
>>109508928I'm talking about animators like Harry Partridge. And no need to reply, I can tell you have a really retarded take, so I'll just talk with ChatGPT and ask it to have the personality and opinion of a brick wall.
can you disable H3 audio to speed up video gen somehow?
>>109508952Audio is like 5% of the tokens, you wouldn't tell the difference.
>>109508674https://files.catbox.moe/572pxb.mp4i don't like style without style ref
>>109508960>scrubbed metadatasir... a crumb of prompt..?
>>109508902It's too long to post, but grok should be able to transcribe it for you. I'm still refining the prompts so a lot of it can probably be pruned
>>109508984>picrel; you coming up with this shit
>>109508984Jesus christ
H3 sometimes make characters open their mouths very wide while speaking like kermit. LTX had the same problem but it was constant there and Jim Carrey tier.Is there a way to prompt for more subtle speech?
>>109509020I'll just talk with ChatGPT and ask it to have the personality and opinion of a brick wall.
>>109509017Did you see his result though? It worked. Hopefully that prompt is overkill.
>>109509027that meme response wasn't meant in insult. that was probably the highest honor i've ever bestowed anyone in this general.
>>109508983yeah, i didn't put rtx upscale in my workflow so I only upscale things I want to publish. and other workflow somehow cuts the meta...https://pastebin.com/yDzuMfN0
>>109508943Int8 VAE but not really worth it since i only got 5 secs speed up
>remove every meme node including sol/h3 cache>just keep patch sage attention, minimax h3 mem eff sage attention, and low vram attention>end up with the fastest speed for 0.8mp 10sec i've ever gottenwell fuckin shit on my dick who would've thunk the meme nodes actually make it slowerand yes the quality got substantially better, especially in terms of audio.
>>109508984lol ok, I gave it to gemini, hopefully he didn't fuck up.
>>109509057Can you share the workflow?
>>109509046I feel like it could be faster if someone made an async node. My guess is it's slow because you're going from CPU to GPU for 200+ frames
>>109509072absolutely nothing specialexcept one anon's suggestion to use that particular sage model, which yes does actually make the video quality better.
I think that H3 has a lot of potential even just as a text to image model.Sure, the VAE is designed for video, but you can just train on stills (see picrel) and it'll decode them just fine, like anything else.The token cost per frame at equivalent resolutions is also better than FLUX/Qwen-Image/Z-Image (it's about ~half as many total even with 2x temporal redundancy), which I find interesting.
>>109509057use spectrum, it doesnt really fuck the quality
>>109509083you niggers say that about EVERY one of these nodes, that's just flat out not true. They all have different levels of quality fuckery, and spectrum was the worst from my testing, next to the multiple h3 cache versions.just don't use them, chances are if you're on a 4000 or 5000 card like i am you're getting a fast enough speed for your hardware. endlessly chasing the best possible quality when you're already raping the ceiling with these nodes will never leave you satisfied.
>>109509081>absolutely nothing specialPlease spoonfeed me and make airplane sounds too
>>109508984I don't think you need such verbose prompt, nor to follow too much the slopped official guide.>>>/wsg/6210975Anime video of a fairy flying through a forest. The camera follows her. 90s style with deep shading. The fairy has green hair and wears a white dress with a violet ribbon around the waist. [Shot 1] At 0.00, the fairy is flying between the trees. At 5.00, flock a multicolored butterflies pass by, and the fairy turns to look a them while flying. At 10.00, the fairy reaches the edge of the forest and the sun can be seen.
This meta is promising.about 9min for 10sec 1mp.
try using QualityRaperH3great node. barely any quality loss.
>>109509095there's no point in convincing these people. same people that insisted lightx2 didnt fuck quality and motion in wan.
>>109509057I told you fuckers, I wasn't insane after all https://desuarchive.org/g/thread/109495264#109495959
>>109509083>it doesnt really fuck the qualitynot from my experience
>>109509081Thanks for listening anon, that node keeps quality while dropping speeds>>109507827Your base resolution is too low
>>109509095are u using sage 2.2, pytorch 2.10+, cu130+, int8cr?
>>109509112specs ?
>>1095091473090
>>109509129yes, fresh portable install
Does [video continuation] have issues? I can never get it to seamlessly continue a video input. I've read the documentation and even got the LLM to try, with no success.
>>109509081What about the Shift node?>>109509095Did you check the latest version with the conservative setting? I will not say that it is losses, but it's a large speed up for very little quality loss.
>>109509162I never touched the shift node. Learned my lesson from imagegen; touching that, unless otherwise specified by the model authors, is pure meme shit.i remember when people here were freaking out about what a difference shift made for z-image tardbo kek
https://d.uguu.se/KouEUJpW.webm
>>109509023I have the reverse problem, I occasionally get gens where there's character speech but mouths don't move. Guidance or examples of good voice prompting in general would be nice.
>>109509082so are omnimodals the future
>>109509166playing with shift on minimax really does something tho.
How good is the reference model at keeping detail?Can it keep a character like pepe and put him in anime style?I want to make one where he hands debo a get a job paper and he crys and shits himself (not shown just heard with stink cloud showing)
>>109509183if you can't explain that that 'something' is, then it's worthless.
>>109509189>How good is the reference model at keeping detail?extremely fucking good. SOTA tier.
>>109509201It's been explained multiple times already.
>>109509187idc, vae decode takes seconds on my 5080
>>109509201it shifts the bits to the correct location
>>109509206
>>109509169the delivery reminded me of this https://www.youtube.com/watch?v=Hn-KmLIt-AQ
>>109509112catbox please? wanna try it myself. also feet prompt
>>109509057What's your GPU? And what's example output from this? Does the Low VRAM attention stuff hurt quality?
>>109507737Getting 40s/it with H3 on 32gb vram (intel because I'm a poorfag) 32gb ddr4, is there any way I can make this shit faster or am I hardware constrained
>>109509225>no its just half of the people who could post something is genning nsfw right now because h3 is fucking amazing..and people are posting those on catbox
I look forward to the ai crash so i can afford a lot of hardware that is offloaded for pennies, but on the flip side it also means buying anything post-crash is fucking pointless
>>109509225Minimax is bad for NSFW. NSFW needs a static humping animation and its bad for those
what does the text to image model have that the reference model is missing?I also can get the text to image to use reference images so I'm confused on the actual difference.
>>109509242Current models will continue existing.
>>109509242>ai crashkek
>>109509233youre running an arc pro b70 right? if so, have you at least made sure youre using the correct packages? with amd on a 7900 xtx i needed to install rocm specific shit to get it usable at all, nevermind performance gains
>>109509242when AI crashes, its taking the whole economy with it
>>109509081does the low vram attention node help with speed at all?
>>109509225128
>>109509222Still trying to find the best settingshttps://d.uguu.se/ogJpBiMt.mp4prompt isn't mine, credit to some other anon, just using his prompt to compare the output quality versus straight 20 steps without cope nodes.
>>109509248i want to purchase an MI350P because a pcie card with that much vram that isnt nvidia makes me smile in pain>>109509252good, hopefully it takes me with it too
>>109509160It genuinely generalizes badly to inpainting/outpainting style tasks by default.I've been screwing with LoRAs designed to specifically be good at temporal outpainting but I need a better general purpose t2v dataset before I'm happy with it
>>109509256>https://d.uguu.se/ogJpBiMt.mp4yeah, this was one of the bad runs, so far the best results have been, er_sde / beta57 for the first pass.
>>109509081Where do I get the minimax h3 mem eff sage attention patch?
>>109509242I don't know why you niggers keep acting like prices will magically go down or you'll even be able to buy anything. If le bubble pops businesses are going to eat up all the hardware, and anything that isn't will be sold back to nvidia and be destroyed.
>>109509255*128 overclocked, maybe speed matters too
>>109509284the prices will go down because the supply will overwhelm the demand. remember microsoft saying they bought a fuckton of hardware but its just sitting around while their datacenters are built? think of all the IN USE hardware that will go into the market when these fuckers crash the industry. it wont be cheap, but it will at the very least be cheaper than today
~13% speedup for tiny quality loss for spectrum, definitely worth unless you really dont mind waiting and have already maxed out the MP, 50 steps etc.
>>109509284nta but it's not really about a bubble popping its about the hype dying and companies realizing AI is not profitable on its own as they originally thought. more manufacturing will be set up. eventually prices will come down but it's going to take a long while. maybe in 2030 it'll normalize
>>109509297Yeah spectrum is pretty much the only one worth using currently, everything else is still a WP
>>109509284when coin mining got less profitable, a lot of miners dumped cards to consumers to recoup costs. no reason to think an AI downturn wouldn't cause the same effect
>>109509251Yeah, I got it all working and I'm getting video output it's just slow as balls. I'm thinking I'm getting fucked because I have no RAM, I'm watching my swapfile get raped to death right now
>>109509301>>109509301
>>109509284if you can't save up $5000 in a few years you have bigger problems in life my guyanyone complaining about affording anything should start by getting a jobyou do realize you're going to get old right?mommy and daddy are going to die one dayfrankly I'd be more worried about starving to death if I were you
>h3 template used megapixels instead of resolution>this is entirely due to the explicit 32 multiple scale>nobody is bothering to just make an integer with rounding math into their workflowprops to you who just defy the 32 scaling if it works, and im sure it does, but using megapixels as the default was a horrible fucking idea
>>109509304coin mining =/= AI. Even if AI magically stops getting developed the current sota open models will be very useful to businesses.
yeah comfy vae decode is still fucked with dynamic vram somehow
>>109509256>>109509271thanks, I'll try it out myself
>>109509305actually curious, why don't chinks build a fab and sweep samsung's market? at least now it would be viable
>>109509320it took 100 seconds to load the vae into vram lmao
>>109509306>I'm thinking I'm getting fucked because I have no RAManon how is your pc running right now? you worry me deeply. as for speedups my only advice right now is try to get flash attention v2 into your setup. theres other things you can do but i dont know how compatible they are with arc.the pro b70 is approximately a 7800 XT in performance last i checked, so its not a bad card but remember arc is new, not very well supported, but it does offer very nice features. i for one want to make a battlematrix machine but these fuckers priced me out like crazy. imagine having those dual gpus all working together without nvlink or some stupid bullshit, its just a mini datacenter with software and a huge memory pool
>>109509319>will be very useful to businesses.apart from coders, AI produces almost no positive value for businesseshttps://medium.com/newsarticulated/thousands-of-ceos-admit-ai-had-no-impact-on-employment-or-productivity-and-its-resurrecting-a-cbd058a7b7ea
>>109509322where do you think taiwan is?
Can I get turbo speeds with these other nodes?Because turbo doesn't look bad to me but I wouldn't mind more steps
>>109509339It's actually because AI is a net multiplier, which means if you're a retard it's multiplies your negative value. One retard with AI can do immeasureable damage.
>>109509339I don't believe articles written after 2022.
euler seems like 19% faster for similar quality for h3? maybe its a one off problem since v/ram allocation is going crazy
>>109509353Lower weight and increase steps
>>109509335"No" as in "very little", I'm just being dramatic Otherwise the b70 has been good for me especially because I got it at msrp, I think the value proposition is dropping as prices go up. You have to deal with a lot more bullshit than you would with an nvidia card
>>109509390there is a thing thats worth mentioning, you can turn off ecc and see if it helps, but since youre using swap space i imagine youre on linux rather than windows?im a crazy person and will beta test willingly as i was an early adopter of arc, i have both an a750 and a770 16gb, both intel variants not that aib bullshit. the b-series is an insanely huge leap that i cannot justify telling people to pick up an a-series now. i would say spend the time to get comfortable with arc, because yes its early and not matured software support wise, but its potential is absolutely fucking there, assuming people care enoughits not as streamlined as amd either but amd isnt much better on this front. nvidia really is the easy fucking winner here, so much so that some nodes dont even consider the idea a non-nvidia gpu is used and just default to cpu to process shit sometimes
>>109509081I'm a tremendous faggot where do I find this sage model
Anyone used res2m/res2s bongmaths with H3?
did someone say snake oil?
>>109509407I got into arc to support a competitor to the GPU duopoly but it feels like intel doesn't even want to compete in that space anymore. Chip crisis killed off celestial and they'll probably axe druid as well. It's a shame because price/vram was absolutely a space where they could have competed until RAM went apeshit bananas
>>109509501It benefits me because I get a card with lots of vram for ~40% of the price of the competition and having a product sell in a certain market does encourage a company to keep investing in that market; which in an ideal world would continue to benefit me by incentivizing GPU manufacturers to sell faster GPUs with more vram year over year. Did you read 4 words into my post and just infer that I bought the card on a lark because I liked the blue color and company's name??
Is it worth getting anything above 32gb vram until you manage to cross the 256gb mark where you can run quant 4 versions of 500b models?I just don't see the point of that dead man's land zone in between. MAYBE 48gb is justified as it usually comes in a standard card and allows you to run extra stuff ontop of your main LLM like additional small video gens or image gen.
>tfw someone cooking Ashley graves krea2 lora :)
Alright that's it. I'm downloading more RAM.
>>109510060Body's fine but face is a bit uncanny.
>>109509945how would you even get 32?
>>109509450yes but I won't tell you the results because I don't want jeets posting my advice on civitai.