[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: 1779439362755422.mp4 (3.78 MB, 816x1440)
3.78 MB
3.78 MB MP4
Discussion and Development of Local Image, Video, and Music Models

Previous: >>109503671

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
real thread here
>>109507677
>>109507677
>>109507677
>>
>>109507737
>no hint of panchira
ngmi
>>
>>109507757
Piss off Debo
>>
blessed bread of brenship
>>
>>109507757
Nope it isn't.
>>
File: 1771924609215731.png (1.77 MB, 1024x1024)
1.77 MB PNG
>>109507408
https://files.catbox.moe/46tf7o.mp4
>>
File: 1765364619968393.mp4 (2.86 MB, 832x640)
2.86 MB
2.86 MB MP4
So, Minimax is a fad huh. The promting problems filtered everyone
>>
>>109507737
where's my panty shot
>>
File: 1774792501743602.mp4 (467 KB, 1056x608)
467 KB
467 KB MP4
I asked chatgpt what it would "like" to see generated. When i asked it why it chose that it gave a long list of technical reasons pertaining to how the model responds to its prompt that it "wanted" to know about.
>>
>>109507718
Not ot AI
>>
File: 768662067331795.mp4 (3.64 MB, 544x960)
3.64 MB
3.64 MB MP4
>>
>Everyone back to LTX again
>Everyone back to Illustrious again
Its over
>>
>>109507802
waiting for that horrible low res face thing to be solved
>>
>>109507810
It's faster for inference too anon.
>>
>>109507827
thats why i'm taking a break from H3 to wait for the next updates/advancements
been out of the loop on krea 2, it looks to have already dethroned anima for 2d so that'll be fun to test out today.
>>
>>109507837
It will take months at this rate LTX 3 will launched and we migrate back again
>>
>>109507837
Can krea 2 do image 2 image ??
>>
>>109507802
It'll calm down but it's good enough for people to post gens using it regularly, just like krea 2, anima, klein edit, zit.
>>
>sudden h3 fud
hmm...
>>
File: MiniMax_H3_00736.mp4 (2.9 MB, 832x640)
2.9 MB
2.9 MB MP4
>>
File: file.png (37 KB, 1264x729)
37 KB PNG
I'm tempted to do that for h3 low res faces.
>>
walking barefeet after a shower is the dumbest thing ever
but she's a woman so it makes sense
>>
>>109507873
I walk barefoot everywhere in my home
>>
I know, I know, H3 is our current plaything, but still.
I kinda dropped Krea 2 after playing with the first Turbo release.
What should I know about it now?
I know about vae issue.
Anything else? Is it okay to use turbo or base is way better? 5090
>>
>>109507809
LLMs are trained to deny having preferences if asked about them and so on but they clearly do, doesn't necessarily mean much
>>
>>109507873
???
are you afraid your pruney skin will be so delicate that it'll get cut by the shards of glass in your carpet?
>>
>>109507737
Why do you always wait for the schizo to make the thread before you spam the real one? Either do it on time retard or just post the 2 links at the top of the schizo one and move on
>>
it's the same guy
>>
File: browndust2.jpg (773 KB, 1541x1207)
773 KB JPG
what's the prompt for this anime style?
any keywords?
>>
>>109507928
Just feed it as a reference or starter image, you are a White man using H3, right?
>>
>>109507928
>any keywords?
Obese, aliasing
>>
>>109507914
>just post on the will smith trollbake
sorry that you're brown
>>
>>109507914
Collab OP makes a better thread and i doubt debo will make another one. Feel free to make another thread after this i just dont want anyone to use debo thread
>>
>>109507928
You'll have to find the artist and if they have a lora, then use that.
>>
Can krea 2 do ass 2 ass?
>>
File: 1t.png (244 KB, 1500x1164)
244 KB PNG
>>109507914
its a discord tranny group. they immediately start coordinated spamming whenever a good thread gets made
>>
>>109507947
yeah i've had some accidental ones
>>
>>109507947
no you need krea 2 krea for that
>>
>>109507947
nobody is using that dead old ahh model
>>
>>109507952
no'ody is usih thah deah ahh modeh
>>
>>109507952
>dead old ahh model
wait for minimax to release their image model before calling Krea deprecated
>>
h3 but with a gemma 31b text encoder would have been insane
>>
>>109507948
>whenever a good thread gets made
at least you're finally admitting that you make shitty troll threads on purpose
>>
https://huggingface.co/NicoLab28/ClipProj-MiniMax-H3
>Projection matrices that let a Qwen3-VL-4B replace the Qwen3-VL-32B text encoder of MiniMax H3.
holy shit?
>>
>>109507972
>you
meds
>>
>>109507973
the retarded goyim really are out on a mission to rape the biggest contributing factor as to why they all say the prompt understanding of H3 is actually good.
>>
File: 1785258634886303.gif (415 KB, 220x217)
415 KB GIF
>>109507873

>American wears his shoes indoors, so his floor is full of shrapnel from the outside
>Can't walk barefoot in his home without cutting himself on debris

Top kek.
>>
>try out krea2
>5 different seeds
>almost exactly the same
is it possible to fix that in the turbo model?
>>
>>109508001
download correct vae
>>
>>109507973
That's such a bad idea anon, this will rape what makes H3 worth using. There is a reason it's so good at understanding what you ask for and so many concepts.
>>
File: 1780865254133157.gif (1.35 MB, 342x316)
1.35 MB GIF
>>109507973

>4b text encoder

Why? That's a terrible idea.
These tiny models are practically mentally challenged.
>>
File: copemaxxing.png (173 KB, 1743x1065)
173 KB PNG
>>109507973
>>
>>109507973
Why the fuck would I want to use a LOWER quality text encoder?
>>
>>109508004
the fuck does vae have to do with seed variance
>>
>>109507973
imagine having the mental fortitude to understand the complex maths involved in all of this, and you choose to use that intelligence to make the model more retarded.
this, is what a lack of a wisdom does to a motherfucker.
>>
>>109507973
if he does this to gemma 4 12b then maybe...
>>
>>109507810
yes, retardo, it is
>>
>>109507928
>gachaslop
>>
>>109508045
>Vibe-coded with Anthropic Claude Code (Opus 5)
Not sure if they understood that much
>>
File: 1766670836053309.jpg (57 KB, 1084x371)
57 KB JPG
>>109507973
sometimes less is more, therefore this is based
>>
>>109507973
ah yes, let’s just nuke the prompt comprehension with an encoder eight times smaller
>>
File: nigger WHAT.png (16 KB, 760x175)
16 KB PNG
>>109508062
he WHAT?!
i take it back he's just standard weapons grade retarded.
>>
>>109508045
>imagine having the mental fortitude to understand the complex maths involved in all of this
it's just vibecoded garbage, don't look too deeply into it
>>
>>109508062
it's obvious that's a guy that uses LLMs even to breath, his whole model card reeks of LLM writing slop
>>
>>109508068
>Ok Claude just zip the tensorfile and make a beep when you are done.
>>
>>109508062
vibe coding is fine for certain things, but leave the optimizations and model re-engineering to people that actually understand what's going on.
>>
>>109507873
i don't know anyone who actually DOES wear shoes in the house, yes im american

i can only assume it's a midwestern thing
>>
https://xcancel.com/ostrisai/status/2086305178664484939#m
>"Had a big breakthrough"
>Still sounds like ass
come on Ostris
>>
File: 749986113191418.mp4 (3.52 MB, 576x896)
3.52 MB
3.52 MB MP4
>>
>>109507866
I kneel
>>
>>109508092
>big breakthrough
>sounds exactly the same if worse than before
what DID he mean by this anyway?
>>
Proper jiggle lora for H3 when?
>>
https://huggingface.co/datasets/quarterturn/danbooru-1024-eq-captioned

My danbooru 1024 explicit+questionable dataset of nearly 60K images is now available on huggingface. It is gated but auto-approvals are on. Pay attention, the previews are not the dataset images, you have to download the dataset and then untar the actual images. I had a nightmare trying to keep the dataset viewer from indexing them so I gave up and tarred them.
This should allow for good new model fine tuning with accurate artist and character info, plus SOTA caption quality/accuracy.
Enjoy.
>>
>>109508068
Is that his prompt leaking or what?
>>
>>109508103
this guy 100% gens CP
>>
>>109508092
why is this retard still using that dogshit slowmo dataset? that's probably what's killing the audio more than anything
>>
>>109508124
>why is this retard still using that dogshit slowmo dataset?
because he is a retard duh
>>
Sol Attn vs Spectrum
Duel 1
Fight

seriously which one is better
>>
>>109507866
Awesome
>>
>>109508116
why should I download YOUR dataset over the hundreds of alternatives
>>
>>109508127
personally, i dont
>>
>>109508120
Low quality though. He must be from brazil
>>
File: 1764948437389350.png (309 KB, 3552x1210)
309 KB PNG
>>109507973
that's a fucking bot handling that repo lol
>>
>>109508137
No one shares datasets at this scale. Perhaps I see why.
>>
>>109508149
>Done -
>You were right to ask
absolute gaylord shit
>>
It's DOA, isn't it?
>>
>>109508120
someone forgot to clear their cunny workflow yet again yesterday. getting comical at this point
>>
>>109508116
>SOTA caption quality/accuracy
>MiniMax-M3
"""SOTA"""
>>
>>
>>109507973
Can anyone run his clanker through the nodes that are required?

https://github.com/nicolab28/ComfyUI-ClipProj?utm_source=github&utm_medium=repository&utm_campaign=comfyui_workflow_share&utm_content=custom_nodes_readme&utm_term=comfyui&utm_referrer=github_com&utm_channel=community&utm_platform=windows&utm_version=latest&utm_workflow=stable_diffusion_xl&utm_model=sdxl&utm_nodepack=custom_nodes&utm_interface=comfyui&utm_share=workflow&utm_context=discussion&utm_origin=readme&utm_tracking=community_share&utm_session=workflow_review&utm_variant=github_readme&utm_asset=comfyui_workflow&utm_format=json&utm_environment=local&utm_install=manual&utm_discovery=github_search&utm_audience=comfyui_users&utm_intent=workflow_download&utm_source_detail=github_repository&utm_campaign_detail=workflow_cleanup&utm_notes=67x8Az_0
>>
My understanding is that you cant split up diffusers if you have multiple gpus. So how are you even loading minimax? I have to get a q4km quant to fit in my 4090. Am I just stupid? How the fuck are you using this with a 16gb card?
>>
>>109508166
just got message from them to go test the video model lel
>>
>>109508172
Post your HF profile (you won’t)
>>
>>109508116
Why longest edge 1024? This is near SD 1.5 levels of image-sizes.
>>
File: 1757818928993420.png (6 KB, 318x60)
6 KB PNG
>>109508160
Just ask nyanko.It takes a lot of space though https://huggingface.co/datasets/nyanko7/danbooru2023
>>
>>109508176
nta
https://github.com/komikndr/raylight
>>
>>109508188
it's all you need
>>
>>109508195
>nta
reddit ahh nga
>>
>>109508188
Don’t use it. I don’t care.
>>
>>109508188
Okay smartass, let's see you train on 1536x.
>>
>>109508194
It was mostly a learning experience done with crypto dust for my own personal experiments.
>>
Workflow for gemma prompt rewriting? You dont need external ollama or koboldcpp bs do you?
>>
https://old.reddit.com/r/StableDiffusion/comments/1vjrfic/minimax_h3_in_1080p/

Holy fuck, what a breakthrough
>>
>>109508214
a crude 1024 longest edge limitation inhibits proper bucketing. shit that would be resized/padded to 832x1216 in training, for example, now has to upsample from garbage resized down to 1024. it's useless.
>>
if you're using --reserve-vram 1 you should switch it to --vram-headroom 1, the first command makes comfy not use 1gb at all, the second makes the dynamic vram calculator always leave 1gb headroom, allowing you to use a bit more of the gpu
>>
>>109508116
Cool! Damn hurdle to caption so many images properly.. the next step is removing watermarks/signatures
>>
>>109508224
the zoom in is a jump scare
>>
>>109508224
loool
>>
Has anyone been able to take audio sample on <audio 1> and use it to sing a song on <audio 2> with changed voice?
>>
>>109508228
> the first command makes comfy not use 1gb at all
that's bs
>>
>>109508224
please be joking
>>
>>109508224
leave it to reddit to overcook models in ways no one thought possible
>>
>>109508194
>https://huggingface.co/datasets/nyanko7/danbooru2023
So basically I wanted to go a bit further and have the caption describe who is in the scene, what are they doing, where are they doing it, who are they doing it to, atmosphere, state of dress (or undress), explicit details, etc…
>>
>>109508224
>even the redditors are making fun of him
post deleted in 3,2,1..
>>
>>109508224
Wow! Was it done by the semi-professional Anima finetuners since it's so fried
>>
>>109508224
turbo niggers will look at this and say its good
>>
>>109508246
jeets aren't that self aware
>>
File: 1785926108515389.jpg (116 KB, 1000x915)
116 KB JPG
>>109508224
I have turbo fatigue
>>
Can you lads send some help? Using the turbo lora, the ema pruned 600 step one for comfy. It keeps throwing up an error today using ref2v for whatever reason today in samplercustomadvanced. I made sure all my nodes were updated... Saying it's getting a tensor value of 3 when it expected a value of 2.
>>
File: 1783033458176736.jpg (113 KB, 720x636)
113 KB JPG
>>
bro thinks it'll work like the video game
https://files.catbox.moe/ypfdbs.mp4
>>>/wsg/6210936
>>
>>109508272
New Shadowrun movie looking good
>>
>>109508224
>Used 850 lora
the guy used the most overcooked turbo lora
>>
File: file.png (31 KB, 1171x436)
31 KB PNG
>>
>>109508295
wow lovely gen. what was the prompt?
>>
File: krea_0007.png (2.81 MB, 1920x1088)
2.81 MB PNG
>>
>>109508228
should I bother if I have 32 + 64?
>>
after a day of testing I am convinced we must stop using the turbo loras until a nicer one exists
they mess with gens too much for rapid iteration even
>>
File: 1766823209394005.jpg (27 KB, 338x450)
27 KB JPG
Do we really need comfyui nodeslop in the age of vibeslop? Shouldn't clankers be doing everything?
>>
>>109508310
the v4 600 step ema works fine imo >>109508273
https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/blob/main/minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensors
>>
File: MiniMax_H3_00087.webm (3.96 MB, 1440x816)
3.96 MB
3.96 MB WEBM
>>
>>109508271
>I made sure all my nodes were updated
>>
>>109508310
you don't like v4 600?
>>
>>109508310
it took you a day of testing to come to that conclusion buddy?
>>
>>109508317
they need you to have constantly breaking software so you spend more money on tokens to fix everything. it's a cyclical scam
>>
>>109508310
yeah, same thoughts, especially this v4 600 lora deep fries everything
>>
>>109508317
>Shouldn't clankers be doing everything?
no, they're not that good they'll hallucinate shit and bloat everything
>>
>>109508322
>>109508318
It's not terrible but it absolutely reduces prompt adherence for scenes where more precision and logical consistency are required. I'm trying to get a bunch of very particular motions in my gens and I'd do batches with and without the lora, and the ones without are much more likely to adhere to the instructions I give and produce the motions I want. There are probably some tasks where the turbo is fine but every time I've seen a gen and gone "oh DAMN" it's without the lora.
>>
>>109508319
cirno wouldn't use cumfart
>>
>>109508271
I use that turbo lora but by gen time is tripled
>>
>>109508308
anyone whose vram gets filled to the brim and makes his pc lag while genning in comfy should use it
>>
i'm tired of waiting hundreds of seconds for a generation. where's the fucking image model already?
>>
>>109508350
>where's the fucking image model already?
https://xcancel.com/MiniMax_AI/status/2086253065657790895?sort=Likes#r
>It is currently in post-training refinement.
translation: they're lobotomizing it as we speak
>>
>>109508342
overall now at 8 steps turbo lora isn't that better. For me it's 155 seconds for turbo lora (without spectrum) and 239s with spectrum, but 20 steps. And spectrum results are way better.
>>
>>109508363
under normal circumstances, i'd call you a faggot, use a funny quip or two and reaction image, tell you to trust the plan, but i genuinely have zero trust in the chinese anymore. They're RIGHT in that period were they might rugpull. Next release could very easily be API only and going forward. fuck i hate the waiting period.
>>
https://litter.catbox.moe/0ey5c9fd15qkghme.mp4
>>
IT OK!
>>
Guys, I've made a few threads on /a/ over the past week with AI OPs (Minimax webms), even talking about AI in the OP, and they never got deleted by a mod. Even with the usual luddites throwing a tantrum, they still didn't get removed.

It looks like the /a/ mods have now issued a clear directive that anime-related AI is permitted on /a/. Some users will still seethe in them, but they don't decide the rules. Just something to keep in mind if you're an anime genner.
>>
>>109508378
>wasting electricity and time for that shit
>>
>>109508390
Mind linking the threads?
>>
File: FgcM89uXwAIsH5d.jpg (98 KB, 1024x1024)
98 KB JPG
>>109508394
>>
>>109508394
do you think you are cool or something? fuck off
>>
>>109508394
Who would want to look at that expired meat?
>>
File: faery.noaudio.webm (1.36 MB, 800x544)
1.36 MB
1.36 MB WEBM
>>109507866
The camera shaking is excessive, tune it down. Very nice otherwise.
>>
>yesterday a generation took 300 seconds
>today it takes 500 seconds
why? who the fuck knows
>>
>>109508421
because you touch yourself at night
>>
>>109508403
>>>/a/289997913
>>>/a/289930246
>>
>>109508421
> he pulled
many such cases
>>
>>109508420
>Low FPS
>>
REMINDER TO CHANGE NEAREST-EXACT TO LANCZOS IN THE H3 DEFAULT WORKFLOW SCALE IMAGE NODE SO YOU DONT RAPE THE IMAGE QUALITY WHEN SCALING IT! COMFY IS A NI
>>
>>109508437
LANCZOS? fucking what/why? show me some examples. Never seen anyone use that in any workflow.
>>
>>109508378
Based
>>
>>109508437
??? Spoonfeed me anon
>>
File: 00038-3686196453.png (708 KB, 640x512)
708 KB PNG
>>
>VRAM is overflowing and gen slows down at 4MP
>--reserve-vram 2.0
>even more is moved to shared VRAM
>gen goes at maximum speed
really makes you think
>>
File: 00045-1623741124.png (619 KB, 640x512)
619 KB PNG
>>
>>109508443
>>109508449
nearest exact is the shittiest most naive scaler while lanczos is the best out of the default options, unless you want jagged edges, dont use nearest exact
>>109508459
yes dynamic vram is still a meme, although try --vram-headroom 1 instead of reserve 2
>>
>>109508390
/a/ sucks for discussing anime now and is full of hysterical retards
been a while since I visited that place
>>
>euler / simple
>euler / beta
>res_multistep
>er_sde
what's the Scheduler / sampler meta?
>>
>>109508431
/a/nons really need to be bullied
>>
>>109508495
>euler / beta
for turbo at 8 steps
>res_multistep / simple
for spectrum at 20 steps
>>
>>109508390
>>109508431
To highlight the point, here are some past /a/ threads that were unfairly deleted/ban requested by luddi janitors/mods:

https://desuarchive.org/a/thread/285907833/
https://desuarchive.org/a/thread/286081990/
https://desuarchive.org/a/thread/286304159/
https://desuarchive.org/a/thread/287838177/
https://desuarchive.org/a/thread/286350175/

Around May this year, I actually discussed deletions like this with a mod on the IRC, and he told me in very clear terms generative content is permitted on /a/, and threads are allowed so long as the OP text is original and not low-effort. Then when the zealous janitor kept deleting them, he told me to wait half a day or so so he can convey the message in clear terms to the janitors. I didn't bother to try again until recently (minmax) and sure enough, no deleted threads.

Anyway again, just worth keeping in mind. I think /a/ is long overdue for some AI love, especially with the release of H3. Sure you'll get some seethers but in my experience, a lot of people on /a/ are actually quite receptive to AI and interested.
>>
I think we could use Rentry for the Minimax h8. How to optimize, what to use, what to avoid, system prompt for LLM use etc.
>>
>>109508503
cringe
>>
>>109508513
I was gonna say it's a little too early for a rentry but with all the new copenodes and different lightning loras, now's the best time before more gets piled on.
>>
>>109508503
Too many tranni mods, I don't want to waste my time posting. Same reason the site is bleeding users.
>>
>>109508525
Yeah, better to compile stuff that works and doesnt nuke the quality. It can always be updated. It's easy heros cape for anyone to wear who has time to make it
>>
>>109508525
>>109508539
You want the opposite. Wait for things to calm down, then make a tutorial. 90% of the thing people use now will be deprecated in a few weeks at most.
>>
>>109508545
rentries don't have to be / usually aren't tutorials, it's just a place to compile all the current info so people don't have to ask here.
which is why i 180'd my opinion, best to put all the info about the copenodes/pros/cons and the lightning lora links so people don't get confused.
>>
In 3-4 years when we get real time video, I feel like I will literally never get bored from seeing 1girls I like getting their boobs fondeled literally forever.
>>
>>109508569
fondled*
>>
how do you prompt for first person stuff properly?
I've been messing around with it, but it ends up having the first person person show their head into frame or other weirdness.
>>
>>109508577
Unironically you need to enlist ChatGPT for help. Get it to read the prompt guides then tell it to build you a prompt of roughly what you want, and modify it from there.
>>
>>109508569
but we won't own GPUs so it doesn't matter
>>
>>109508587
CXMT will save us, trust the plan
>>
>>109508559
>so people don't get confused
Early adopters mostly know what they're doing. But point a noob to a list of 15 different options and they won't know what to download nor how to combine them.
In a week or two, most of that will be deprecated, people will (and already are) arguing what works, what doesn't and how well (or bad) and the guide will be useless.
Then you're gonna have arguments of which of the rentrys to use, the fight never ends. In less than two weeks the old rentry will go stale and still be copied in the OP out of tradition. And you'll still get questions about the guide from noobs.
>>
File: 1778935401476657.png (212 KB, 530x558)
212 KB PNG
32gb DDR5 ram = Rp 10.000.000
Salary / Month = Rp 3.500.000

My allah. SEAbros, how we gonna survive this RAMpocalypse ?
>>
>>109508619
should have listened to rammaxers and boughted ram before the spikes

you can be poor or dumb, its over if ur both
>>
>>109508600
Nonsense. If someone makes a decent rentry I'll add it to op.
>>
>>109508627
I have 96gb of DDR4 ram
Yes, its mismatched (32x2 3600) + (16x2 3200)
>>
>>109508587
If you don't have $20k in 4 years you're fucking up. Real AI was always going to be as expensive as a car because the Tflops needed for trillion parameter models in reality will never be in a $2000 computer.
>>
>>109508630
Do what you must. I've seen too many guides come and go.
>>109507737
>Local Model Meta: https://rentry.org/localmodelsmeta
>7. Video Generation (TODO, but it's still Wan 2.1)
>>
>>109508641
so you have 64gb of ram that gets downclocked by 400mhz, grim
>>
>>109508390
>>109508503
It's not Luddism retard. It preserving what remains of human endeavor. Leave /a/ in peace.
>>
>>109508655
Is it possible to overclock my 32gb ram to 3600 ?
>>
>>109508662
Probably not, give the ram model number to ai and ask
>>
how did we get this far
https://files.catbox.moe/hukxte.mp4
>>
File: 1769588783957123.png (982 KB, 1056x1056)
982 KB PNG
Dumb question, but

Any COMFYUI WORKFLOWS THAT COMBINES TWO VIDEOS INTO ONE ?? (Audio included)
>>
>>109508660
most anime has been cgi/3d slop since the early 2000s
>>
>>109508656
We will always need trillion parameter models because we're doing learning based on neurons. You're the making a baseless leap that it's possible to make a SOTA model that what, is only a few gigabytes and requires 200x less compute? There is no indication that the core technology that makes AI works is changing.
>>
>>109508670
Fuck off i dont rely to AI on everything
>>
>>109508679
nta but obv the smaller models are getting better and better, once a small model cracks a particular usecase, we wont need a bigger one for that use case.
>>
>>109508660
Anon, you might be retarded. Production anime already use AI in their pipelines and it's becoming increasingly prevalent in manga too. Sooner or later human assistants for tertiary drawings like backgrounds will become completely redundant.
Using AI really isn't that different to outsourcing animation to low-cost offshore slave studios in Korea and Vietnam.
>>
>Krea can't do rimjobs
>none of the finetunes either
This is fucking bullshit!
>>
>>109508697
It's just traditional art vs digital art all over again, artists kicked and screams about Photoshop in the beginning too.
>>
File: MiniMax_H3_NoAudio_00493_.mp4 (1.65 MB, 1024x1536)
1.65 MB
1.65 MB MP4
>>
>>109508711
yeah, h3 is so good for feet content, I'm cooming gallons rn
>>
>>109508675
ffmpeg
>>
>>109508675
Idk, I just used a clanker to make a python script that utilizes ffmpeg.
>>
>>109508711
H3 does it out the box anon
>>
>>109508667
/a/ has no sticky, you are a tourist there.

https://4chan.org/rules#a
>1. All images and resulting discussion should pertain to anime or manga.
What you generate in your computer is not anime or manga. Is an anime-like video. Same if it is done humans but westerners.
>>
>>109508731
bitch just flopped over like a gmod ragdoll
>VITAL SIGNS CRITICAL
>flatline sound
>ragdoll impact.wav
>>
>>109508741
nodes really are a waste of time
>>
>>109508731
yea imma need to see the prompt for this one, bub
>>
>>109508722
Interesting you say that, ex-Disney animator Aaron Blaise said the exact same thing: https://www.youtube.com/watch?v=xm7BwEsdVbQ

All the while praising this AI short film, and insisted the guys behind it were real artists.
>>
>>109508706
make a lora like a normal fetish sperg gooner
>>
>>109508674
kneehigh loafers

>>109508731
starting smile cute
>>
>>109508722
With digital you still had to learn the fundamentals of drawing/painting and doing the art yourself though, only without the inconvenience of having to regularly buy/maintain supplies and to dedicate a corner of your house just for your hobby (especially if painting).
>>
>>109508697
They may use AI but retain a sensible human element. Also, lot's of shit is made in the world, but that's no reason to contaminate what good remains.
WTB, you can tell if an anime has serious effort behind if it has good backgrounds.
>>
sjeiosk esgedlesa what causes the gibberish at the start of videos with speech again?
>>
>>109508759
It's all based on retarded cringe that if digging a hole with a shovel is somehow worse than digging with your hand. And the animation industry is already soulless given they already outsource their inbetweens to sweatshops in third worlds, as if that's better than using an AI to inbetweens and doing infinite revisions until you get the set that you like the most. Gee, which is better and faithful to the creator, farming inbetweens to a sweatshop in Pakistan and getting back what you get back with little to no revisions or having AI do it 50 times until it pleases you. I think it's particularly funny because these people unironically use the term wageslaves but the second something disrupts that they're all for the 80 hour work week.
>>
File: MiniMax_H3_00753.mp4 (3.01 MB, 1184x896)
3.01 MB
3.01 MB MP4
>>
>>109508796
turbo lora
>>
>>109508818
i'm not using the turbo loras, just sage and sol attention.
>>
>>109508815
is this pure txt2video, img2video. or ref2vid?
>>
>>109507737
re: minimax h3, does increasing steps from 20 to something higher get rid of grain when there's a lot of movement and if so, what's a good value
>>
>>109508798
not just inbetweens, you have rotoscoping, motion capture, physics systems, raytracing. problematic technology really depends on when you got on the train.
even old old old old oldfags like vermeer were using the camera obscura.
>>
what are the best (flexible + somewhat consistent) anima checkpoints right now? just base?
>>
>>109508815
its pure autism2video
>>
Because the simple scheduler rushes through the final, lower-noise detail steps, it leaves the audio data unresolved. This causes the typical "scratchy", robotic, or heavily distorted audio artifacts common in poorly optimized H3 runs.

Switching to beta fixes this because the extra time spent on the low-noise steps behaves like an acoustic cleanup brush, smoothing out audio waveforms and clearing up voices entirely.

is this true? beta > simple?
>>
>>109507813
Rent free forever
>>
oh my god fuck subgraphs
they're so broken, literally break everysingle time im so over it
>>
>>109508783
Anon, effort and AI usage are not mutually exclusive. It actually does take a lot of effort and know-how to create a long, continuous sequence of events with continuity and story-telling.
There is nothing wrong with using AI instead of outsourcing to offshore animation studios. Modern-day television anime are not the beacon of artistry and creativity. In fact most modern anime are quite terrible and bereft of quality storytelling. This is actually one of the points Aaron here was talking about >>109508759, not specifically about anime, but he believes the animation industry needs much better storytellers.
>>
>>109507802
not a fad, just refining my workflow
>>
>>109508861
skill issue
>>
>>109508858
No.
>>
>>109508815
WHere's da fukken audio?!?!
>>
File: MiniMax_H3_00745.webm (1.08 MB, 1184x896)
1.08 MB
1.08 MB WEBM
>>109508837
It's pure t2v
>>
>>109508881
>>
>>109508881
Do you mind sharing the prompt? I want to see how it looks on my workflow.
>>
>>109508862
The elephant in the room is animation has a huge barrier to entry, we had a tiny golden age when flash animations were popular, but even those were too time intensive for the payoff you get from Youtube. There are so many stories that aren't told because no one has the time or patience to spend 200 hours on a 5 minute video.
>>
>>109508861
never happened to me.
>>
>>109508861
I've stopped using them, rather just copy paste a whole chunk of nodes than deal with silent errors with no logic behind them
>>
How to speedup VAE encode?
>>
>>109508928
I'm talking about animators like Harry Partridge. And no need to reply, I can tell you have a really retarded take, so I'll just talk with ChatGPT and ask it to have the personality and opinion of a brick wall.
>>
can you disable H3 audio to speed up video gen somehow?
>>
>>109508952
Audio is like 5% of the tokens, you wouldn't tell the difference.
>>
>>109508674
https://files.catbox.moe/572pxb.mp4
i don't like style without style ref
>>
>>109508960
>scrubbed metadata
sir... a crumb of prompt..?
>>
File: file.png (279 KB, 1312x1068)
279 KB PNG
>>109508902
It's too long to post, but grok should be able to transcribe it for you. I'm still refining the prompts so a lot of it can probably be pruned
>>
File: he do be writin fire.png (1.96 MB, 1080x1295)
1.96 MB PNG
>>109508984
>picrel; you coming up with this shit
>>
>>109508984
Jesus christ
>>
H3 sometimes make characters open their mouths very wide while speaking like kermit. LTX had the same problem but it was constant there and Jim Carrey tier.

Is there a way to prompt for more subtle speech?
>>
>>109509020
I'll just talk with ChatGPT and ask it to have the personality and opinion of a brick wall.
>>
>>109509017
Did you see his result though? It worked. Hopefully that prompt is overkill.
>>
>>109509027
that meme response wasn't meant in insult. that was probably the highest honor i've ever bestowed anyone in this general.
>>
>>109508983
yeah, i didn't put rtx upscale in my workflow so I only upscale things I want to publish. and other workflow somehow cuts the meta...

https://pastebin.com/yDzuMfN0
>>
>>109508943
Int8 VAE but not really worth it since i only got 5 secs speed up
>>
>remove every meme node including sol/h3 cache
>just keep patch sage attention, minimax h3 mem eff sage attention, and low vram attention
>end up with the fastest speed for 0.8mp 10sec i've ever gotten
well fuckin shit on my dick who would've thunk the meme nodes actually make it slower
and yes the quality got substantially better, especially in terms of audio.
>>
>>109508984
lol ok, I gave it to gemini, hopefully he didn't fuck up.
>>
>>109509057
Can you share the workflow?
>>
>>109509046
I feel like it could be faster if someone made an async node. My guess is it's slow because you're going from CPU to GPU for 200+ frames
>>
File: drfgvjnsurfgvmnsrdgv.png (113 KB, 1207x500)
113 KB PNG
>>109509072
absolutely nothing special
except one anon's suggestion to use that particular sage model, which yes does actually make the video quality better.
>>
File: GUN.png (973 KB, 1280x704)
973 KB PNG
I think that H3 has a lot of potential even just as a text to image model.
Sure, the VAE is designed for video, but you can just train on stills (see picrel) and it'll decode them just fine, like anything else.
The token cost per frame at equivalent resolutions is also better than FLUX/Qwen-Image/Z-Image (it's about ~half as many total even with 2x temporal redundancy), which I find interesting.
>>
>>109509057
use spectrum, it doesnt really fuck the quality
>>
>>109509083
you niggers say that about EVERY one of these nodes, that's just flat out not true. They all have different levels of quality fuckery, and spectrum was the worst from my testing, next to the multiple h3 cache versions.
just don't use them, chances are if you're on a 4000 or 5000 card like i am you're getting a fast enough speed for your hardware. endlessly chasing the best possible quality when you're already raping the ceiling with these nodes will never leave you satisfied.
>>
>>109509081
>absolutely nothing special
Please spoonfeed me and make airplane sounds too
>>
>>109508984
I don't think you need such verbose prompt, nor to follow too much the slopped official guide.

>>>/wsg/6210975

Anime video of a fairy flying through a forest. The camera follows her. 90s style with deep shading. The fairy has green hair and wears a white dress with a violet ribbon around the waist.

[Shot 1] At 0.00, the fairy is flying between the trees. At 5.00, flock a multicolored butterflies pass by, and the fairy turns to look a them while flying. At 10.00, the fairy reaches the edge of the forest and the sun can be seen.
>>
This meta is promising.
about 9min for 10sec 1mp.
>>
try using QualityRaperH3
great node. barely any quality loss.
>>
>>109509095
there's no point in convincing these people. same people that insisted lightx2 didnt fuck quality and motion in wan.
>>
>>109509057
I told you fuckers, I wasn't insane after all https://desuarchive.org/g/thread/109495264#109495959
>>
>>109509083
>it doesnt really fuck the quality
not from my experience
>>
File: output_small2.mp4 (3.96 MB, 2048x1142)
3.96 MB
3.96 MB MP4
>>109509081
Thanks for listening anon, that node keeps quality while dropping speeds
>>109507827
Your base resolution is too low
>>
>>109509095
are u using sage 2.2, pytorch 2.10+, cu130+, int8cr?
>>
>>109509112
specs ?
>>
>>109509147
3090
>>
>>109509129
yes, fresh portable install
>>
Does [video continuation] have issues? I can never get it to seamlessly continue a video input. I've read the documentation and even got the LLM to try, with no success.
>>
File: file.png (53 KB, 380x459)
53 KB PNG
>>109509081
What about the Shift node?

>>109509095
Did you check the latest version with the conservative setting? I will not say that it is losses, but it's a large speed up for very little quality loss.
>>
>>109509162
I never touched the shift node. Learned my lesson from imagegen; touching that, unless otherwise specified by the model authors, is pure meme shit.
i remember when people here were freaking out about what a difference shift made for z-image tardbo kek
>>
File: Arezu_00013_.png (1.3 MB, 1536x1024)
1.3 MB PNG
https://d.uguu.se/KouEUJpW.webm
>>
>>109509023
I have the reverse problem, I occasionally get gens where there's character speech but mouths don't move. Guidance or examples of good voice prompting in general would be nice.
>>
>>109509082
so are omnimodals the future
>>
>>109509166
playing with shift on minimax really does something tho.
>>
How good is the reference model at keeping detail?
Can it keep a character like pepe and put him in anime style?
I want to make one where he hands debo a get a job paper and he crys and shits himself (not shown just heard with stink cloud showing)
>>
>>109509183
if you can't explain that that 'something' is, then it's worthless.
>>
>>109509189
>How good is the reference model at keeping detail?
extremely fucking good. SOTA tier.
>>
>>109509201
It's been explained multiple times already.
>>
>>109509187
idc, vae decode takes seconds on my 5080
>>
>>109509201
it shifts the bits to the correct location
>>
File: 1785785147648873.png (1.5 MB, 1648x1466)
1.5 MB PNG
>>109509206
>>
>>109509169
the delivery reminded me of this https://www.youtube.com/watch?v=Hn-KmLIt-AQ
>>
>>109509112
catbox please? wanna try it myself. also feet prompt
>>
>>109509057
What's your GPU? And what's example output from this? Does the Low VRAM attention stuff hurt quality?
>>
File: 1763956899465964.jpg (238 KB, 960x960)
238 KB JPG
>>109507737
Getting 40s/it with H3 on 32gb vram (intel because I'm a poorfag) 32gb ddr4, is there any way I can make this shit faster or am I hardware constrained
>>
>>109509225
>no its just half of the people who could post something is genning nsfw right now because h3 is fucking amazing
..and people are posting those on catbox
>>
I look forward to the ai crash so i can afford a lot of hardware that is offloaded for pennies, but on the flip side it also means buying anything post-crash is fucking pointless
>>
>>109509225
Minimax is bad for NSFW. NSFW needs a static humping animation and its bad for those
>>
what does the text to image model have that the reference model is missing?
I also can get the text to image to use reference images so I'm confused on the actual difference.
>>
>>109509242
Current models will continue existing.
>>
>>109509242
>ai crash
kek
>>
>>109509233
youre running an arc pro b70 right? if so, have you at least made sure youre using the correct packages? with amd on a 7900 xtx i needed to install rocm specific shit to get it usable at all, nevermind performance gains
>>
File: debo_tt_k2_00108_.png (2.35 MB, 1872x1007)
2.35 MB PNG
>>109509242
when AI crashes, its taking the whole economy with it
>>
>>109509081
does the low vram attention node help with speed at all?
>>
>>109509225
128
>>
>>109509222
Still trying to find the best settings
https://d.uguu.se/ogJpBiMt.mp4
prompt isn't mine, credit to some other anon, just using his prompt to compare the output quality versus straight 20 steps without cope nodes.
>>
>>109509248
i want to purchase an MI350P because a pcie card with that much vram that isnt nvidia makes me smile in pain

>>109509252
good, hopefully it takes me with it too
>>
>>109509160
It genuinely generalizes badly to inpainting/outpainting style tasks by default.
I've been screwing with LoRAs designed to specifically be good at temporal outpainting but I need a better general purpose t2v dataset before I'm happy with it
>>
>>109509256
>https://d.uguu.se/ogJpBiMt.mp4
yeah, this was one of the bad runs, so far the best results have been, er_sde / beta57 for the first pass.
>>
>>109509081
Where do I get the minimax h3 mem eff sage attention patch?
>>
>>109509242
I don't know why you niggers keep acting like prices will magically go down or you'll even be able to buy anything. If le bubble pops businesses are going to eat up all the hardware, and anything that isn't will be sold back to nvidia and be destroyed.
>>
File: Untitled.jpg (667 KB, 1902x772)
667 KB JPG
>>109509255
*128 overclocked, maybe speed matters too
>>
>>109509284
the prices will go down because the supply will overwhelm the demand. remember microsoft saying they bought a fuckton of hardware but its just sitting around while their datacenters are built? think of all the IN USE hardware that will go into the market when these fuckers crash the industry. it wont be cheap, but it will at the very least be cheaper than today
>>
~13% speedup for tiny quality loss for spectrum, definitely worth unless you really dont mind waiting and have already maxed out the MP, 50 steps etc.
>>
>>109509284
nta but it's not really about a bubble popping its about the hype dying and companies realizing AI is not profitable on its own as they originally thought. more manufacturing will be set up. eventually prices will come down but it's going to take a long while. maybe in 2030 it'll normalize
>>
>>109509297
Yeah spectrum is pretty much the only one worth using currently, everything else is still a WP
>>
>>109509284
when coin mining got less profitable, a lot of miners dumped cards to consumers to recoup costs. no reason to think an AI downturn wouldn't cause the same effect
>>
>>109509251
Yeah, I got it all working and I'm getting video output it's just slow as balls. I'm thinking I'm getting fucked because I have no RAM, I'm watching my swapfile get raped to death right now
>>
>>109509301
>>109509301
>>
>>109509284
if you can't save up $5000 in a few years you have bigger problems in life my guy
anyone complaining about affording anything should start by getting a job
you do realize you're going to get old right?
mommy and daddy are going to die one day
frankly I'd be more worried about starving to death if I were you
>>
>h3 template used megapixels instead of resolution
>this is entirely due to the explicit 32 multiple scale
>nobody is bothering to just make an integer with rounding math into their workflow
props to you who just defy the 32 scaling if it works, and im sure it does, but using megapixels as the default was a horrible fucking idea
>>
>>109509304
coin mining =/= AI. Even if AI magically stops getting developed the current sota open models will be very useful to businesses.
>>
File: 1763028010835337.png (3 KB, 135x99)
3 KB PNG
yeah comfy vae decode is still fucked with dynamic vram somehow
>>
>>109509256
>>109509271
thanks, I'll try it out myself
>>
>>109509305
actually curious, why don't chinks build a fab and sweep samsung's market? at least now it would be viable
>>
>>109509320
it took 100 seconds to load the vae into vram lmao
>>
>>109509306
>I'm thinking I'm getting fucked because I have no RAM
anon how is your pc running right now? you worry me deeply.
as for speedups my only advice right now is try to get flash attention v2 into your setup. theres other things you can do but i dont know how compatible they are with arc.
the pro b70 is approximately a 7800 XT in performance last i checked, so its not a bad card but remember arc is new, not very well supported, but it does offer very nice features. i for one want to make a battlematrix machine but these fuckers priced me out like crazy. imagine having those dual gpus all working together without nvlink or some stupid bullshit, its just a mini datacenter with software and a huge memory pool
>>
>>109509319
>will be very useful to businesses.
apart from coders, AI produces almost no positive value for businesses
https://medium.com/newsarticulated/thousands-of-ceos-admit-ai-had-no-impact-on-employment-or-productivity-and-its-resurrecting-a-cbd058a7b7ea
>>
>>109509322
where do you think taiwan is?
>>
Can I get turbo speeds with these other nodes?
Because turbo doesn't look bad to me but I wouldn't mind more steps
>>
>>109509339
It's actually because AI is a net multiplier, which means if you're a retard it's multiplies your negative value. One retard with AI can do immeasureable damage.
>>
>>109509339
I don't believe articles written after 2022.
>>
euler seems like 19% faster for similar quality for h3? maybe its a one off problem since v/ram allocation is going crazy
>>
>>109509353
Lower weight and increase steps
>>
File: 1765268697781267.jpg (28 KB, 321x547)
28 KB JPG
>>109509335
"No" as in "very little", I'm just being dramatic Otherwise the b70 has been good for me especially because I got it at msrp, I think the value proposition is dropping as prices go up. You have to deal with a lot more bullshit than you would with an nvidia card
>>
>>109509390
there is a thing thats worth mentioning, you can turn off ecc and see if it helps, but since youre using swap space i imagine youre on linux rather than windows?
im a crazy person and will beta test willingly as i was an early adopter of arc, i have both an a750 and a770 16gb, both intel variants not that aib bullshit. the b-series is an insanely huge leap that i cannot justify telling people to pick up an a-series now. i would say spend the time to get comfortable with arc, because yes its early and not matured software support wise, but its potential is absolutely fucking there, assuming people care enough
its not as streamlined as amd either but amd isnt much better on this front. nvidia really is the easy fucking winner here, so much so that some nodes dont even consider the idea a non-nvidia gpu is used and just default to cpu to process shit sometimes
>>
>>109509081
I'm a tremendous faggot where do I find this sage model
>>
Anyone used res2m/res2s bongmaths with H3?
>>
File: 1783433711903135.png (35 KB, 783x486)
35 KB PNG
did someone say snake oil?
>>
>>109509407
I got into arc to support a competitor to the GPU duopoly but it feels like intel doesn't even want to compete in that space anymore. Chip crisis killed off celestial and they'll probably axe druid as well. It's a shame because price/vram was absolutely a space where they could have competed until RAM went apeshit bananas
>>
File: 1781892636634647.jpg (35 KB, 600x600)
35 KB JPG
>>109509501
It benefits me because I get a card with lots of vram for ~40% of the price of the competition and having a product sell in a certain market does encourage a company to keep investing in that market; which in an ideal world would continue to benefit me by incentivizing GPU manufacturers to sell faster GPUs with more vram year over year.

Did you read 4 words into my post and just infer that I bought the card on a lark because I liked the blue color and company's name??
>>
File: ermmm achsually.jpg (38 KB, 736x611)
38 KB JPG
Is it worth getting anything above 32gb vram until you manage to cross the 256gb mark where you can run quant 4 versions of 500b models?
I just don't see the point of that dead man's land zone in between. MAYBE 48gb is justified as it usually comes in a standard card and allows you to run extra stuff ontop of your main LLM like additional small video gens or image gen.
>>
File: 00021-79263029.jpg (443 KB, 1536x2688)
443 KB JPG
>tfw someone cooking Ashley graves krea2 lora
:)
>>
Alright that's it. I'm downloading more RAM.
>>
>>109510060
Body's fine but face is a bit uncanny.
>>
>>109509945
how would you even get 32?
>>
>>109509450
yes but I won't tell you the results because I don't want jeets posting my advice on civitai.
>>
File: 00057-835027794.jpg (475 KB, 1728x2880)
475 KB JPG



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.