[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: collage.jpg (3.41 MB, 6867x3565)
3.41 MB JPG
Discussion and Development of Local Image, Video, and Music Models and Software

Previous: >>109416283

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Z
https://huggingface.co/Tongyi-MAI/Z-Image

>Qwen
https://huggingface.co/collections/Qwen/qwen-image

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>LTX-2.3
https://huggingface.co/collections/Lightricks/ltx-23

>Wan
https://github.com/Wan-Video/Wan2.2

>Chroma
https://huggingface.co/lodestones/Chroma1-Base
https://rentry.org/mvu52t46

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
blessed bread of brenship
>>
so anima is just dead?
>illust still does characters and nsfw better
>nothing about kreanima
>lodestone stealing the thunder with a new krea chroma
>turdyruss posting about loicenses and quants instead of baking the next model
>krea 2 is somehow faster
what went so wrong?
>>
blessed thread of frenship
>>
For me it's ani studio with its great documentation and ani's personal support on discord
>>
File: 791684948326128.png (544 KB, 647x791)
544 KB PNG
So when are you releasing your feature length AI film anon?
>>
>>109418619
theaters still exist?
>>
File: 44764605596156.jpg (371 KB, 2176x1472)
371 KB JPG
>>
>>109418604
we're well aware of his outdated fudding tactics.
they no longer work now that a new model drops every month lmao
>>
File: :-.png (36 KB, 571x175)
36 KB PNG
>>109418614
>ani's personal support on discord
he's actually quite aggressive on discord, I tried to get help to compile on linux and blew up and said it was my fault for not knowing what linked lists are or some retarded shit like that
>>
>>109418641
I think you should take ani's advice in your pic, lolcow
>>
Thirty
>>
>>109418661
>I think you should take ani's advice in your pic, lolcow
i see, so this is the kind of abuse that people trying anistudio should expect from the dev, gotcha
>>
>>109418641
The problem with C++ is that you actually have to be a good programmer to make anything useful with it.
>>
>>109418672
nice melty catjak. I give it 0 GitHub stars
>>
>>109418641
hate to admit it, but Ani is correct here
>>
File: 4040382708980.jpg (429 KB, 2176x1472)
429 KB JPG
>>109418626
apperantly
>>
File: ComfyUI_02901_.png (1.1 MB, 1024x1024)
1.1 MB PNG
>>
Hello, returning from a 5-month break here. Can we get an update on AnimAnon's OpenAnima project he said he was starting alongside LAX from Laxhar Labs, the creator of noob? I heard they were working on an Apache licensed cosmos-based model just like Anima, but it was going to be more "open for game developers to use". I looked around and couldn't find any info on it. It seems like LAX has joined ComfyOrg to shill NovelAI, and AnimAnon's UI hasn't been updated either.
Can someone update me on the status of AnimAnon and what he is currently working on? I am looking forward to seeing his 5 months of progress towards an open diffusion model for all.
>>
if you ever include more than a single trigger word in your lora dataset captions you are a jeet or a schizo. a single word is all you will ever need for all use cases.
>>
>>109418689
He never did it because he has no idea how to make a model and it was all posturing because he was jealous that tdrussell got a grant, while Ani thought that Yoland was suddenly going to forget the reason why Ani is physically banned from the ComfyOrg office (making death threats to staff) and give him a trillion dollars if he samefagged enough on 4chan to make it look like anyone supports him.
>>
>>109418689
ani spent 6 months in the hyperbolic time chamber.
with his newfound skillset, he's about to release a new model that will make anima look like SD 1.4
>>
>>109418704
yikes, that's embarassing? so it's been what, 4 years now of zero progress from him? that's quite pathetic
>>
>>109418697
dude just rename the lora to include the triggerword/triggerphase
>>
>>109418715
>make anima look like SD 1.4
>trani needs to train 5 months to accomplish what turdyruss already did on a day 1
keeeeeeeek
>>
imagine if someone forked and rewrote anistudio in python
>>
File: Krea2_turbo_00372_.png (1.18 MB, 768x1368)
1.18 MB PNG
>>
>>109418731
claude could probably oneshot it
but im not wasting my tokens on that junk just for a meme
>>
File: .jpg (1.6 MB, 1890x2216)
1.6 MB JPG
>>109418720
didn't expect any better from the guy that spends entire days spamming threads at the rate of up to 100 posts per threads instead of updating his dependencies to support latest models (>6 months still no anima support btw)
>>
File: 68273794515645.png (3.4 MB, 1344x1728)
3.4 MB PNG
>>109418688
>>
>>109418724
are you fucking retarded or what?
>>
>>109418737
prompt or lora?
>>
>>109418726
Ani is the creative mind of the group. tdrussell simply used the framework Ani already designed.
designing a new superior framework from scratch in a couple of months is actually fast
>>
>>109418868
Chatgpt wrote it
>Field too long
wtf
https://pastebin.com/M9pJ7KZ9
>>
File: 694600786335659.png (3.46 MB, 1344x1728)
3.46 MB PNG
>>109418903
Nice
>>
File: 1764509778989289.jpg (851 KB, 1248x1824)
851 KB JPG
>>
>>109418873
>tdrussell simply used the framework Ani already designed
what framework? LMAO
you didn't design shit, murderous reject
the only thing you designed is your own ostracization from every space in this community
>>
File: 617395130542983.png (3.91 MB, 1728x1344)
3.91 MB PNG
>>
File: 929799666978196.jpg (627 KB, 1728x1344)
627 KB JPG
You think it would've been fun being a big shot cocaine dealer in mid 80s LA?
>>
>>109418641
Imagine trying to flex with a very, very basic concept like linkedlists kek
>>
krea loras do a good job of not overfitting outfits which is pretty interesting compared to other natural language models I have tried before. The lora dataset was just "mirko" tagged as the character not in depth descriptions of the outfit
>>
>>109419069
I think you need to read up on linked lists
>>
for me its MiniMax- H3
>>
>>109419087
Whatever you say, avatarfaggot.
>>
File: Krea2_turbo_00795_.png (1.27 MB, 1024x1024)
1.27 MB PNG
>>
she would have looked better with a trigger phase
>>
>>109419125
sorry I didn't train a new character lora for every post
>>
testing out lora mixing in krea 2 and it works so well. in Chroma if you mixed loras it would completely fuck both of them up. what gives?
>>
>>109419172
I don't know
>>
Now that the dust has settled, what's the verdict on Kroma 0.1 lora? Everyone has been saying that all Krea needs is a large-scale booru / NSFW finetune, well, here it is.
>>
fuck off from /adt/ schizo and stay in your containment general
>>
>>109419198
link? I'll give it a go
>>
>>109419206
Who did this to you?
>>
File: 704317805657529.png (1.86 MB, 1344x960)
1.86 MB PNG
>>109419172
Can't quite tell if this is a rhetorical question
>>
>>109419226
/adt/ likes ani and there is nothing wrong with that
so stop shitting up the thread
>>
>>109419217
https://huggingface.co/lodestones/Kroma
>>
>>109418758
>>109418688
ngl, Anima is pretty good, but Krea2 is surprisingly far better at style preservation and mixing.
>pic unrelated
>>
>>109418704
>while Ani thought that Yoland was suddenly going to forget the reason why Ani is physically banned from the ComfyOrg office (making death threats to staff)
why do you lie all the time schizo? ani confirmed that this is not true >>109419250
>>
File: 408727668901514.png (3.68 MB, 1728x1344)
3.68 MB PNG
>>109419272
Well it does have 6 times the parameters.
>>
File: momo_00072_.jpg (1.12 MB, 1792x2304)
1.12 MB JPG
>>
>>109419301
is anyone even still bothering to defend anima as better, anyone still using it is mostly because of budget setups
>>
File: uh oh.png (113 KB, 474x530)
113 KB PNG
>>109419299
>>
>>109419307
>>109419303
>>
I wish I got millions to vibe code a mediocre frontend for open source software like comfy
>>
>>109419301
Fair point.
>>
Am I supposed to believe that right now an avatarfag is only posting in one thread and not another and some random completely different anon is defending him?
>>
>>109419305
Obviously Krea is "better" in terms of basic capabilities. The problem is that Krea hasn't been trained on 10 million booru images.
>>
>>109419272
can you show some comparisons?
>>
>>109419309
>comfy doesn't like yoland tantrums and I make him seethe on sight is all
lmfao
you're the one that had a real life melty because you couldn't follow instructions and your crush went down mt fuji without you
HAHAHAHA
also we can read ani
>I'll get rid of him for the both of us
those are your words, not mine
>>
>>109419272
>>109419343
*style comparisons
im curious how well it knows artists etc
>>
>>109419198
>256x training
>lora extract
>well, here it is.
AHAHAHAHAHAHAHA the absolute state of chromakeks
>>
when are we getting a good model?
>>
>>109419272
show krea anime goon, 1 girl doesn't count btw
>>
>>109419380
2 days
>>
File: momo_00136_.jpg (1.08 MB, 1792x2304)
1.08 MB JPG
>>
>>109419398
she looks like she fucks human men
>>
File: 372600551991218.png (3.75 MB, 1856x1280)
3.75 MB PNG
>>
>>109419198
>>109419217
>>109419254
Its pretty good.
>>109419357
did you even try it yet moron?
>>
File: 582356478114736.png (3.47 MB, 1280x1856)
3.47 MB PNG
>>109419382
2girls
>>
>>109419398
wtf is the outfit in the bottom right? it looks nothing like the other images. gpt wouldn't make this big of a mistake
>>
>>109419428
is that the best krea can do? pathetic, show me penis in vagina without it becoming some abomination then we can talk
>>
File: Krea2_turbo_00806_.png (907 KB, 1024x1024)
907 KB PNG
>>
so anyone actually tried that chroma krea lora?
>>
>>109419446
yeah
>>
>>109419446
yesh
>>
>>109419446
no
>>
>>109419446
sometimes.
>>
>>109419254
barely did anything, dunno seems garbage
>>
>>109419446
have you? :|
>>
File: orodSh_00138_.jpg (1.3 MB, 1776x2560)
1.3 MB JPG
>>109419446
Yeah. Wrecked details. I don't know what the dataset is
>>
>lora
he prolly mad AF yall aint respecting xirs finetune identity
>>
>>109419490
it has nothing to do with the dataset, it's because he's training at 256x256 like a retard. the same shit that killed every chroma model before it.
>>
I would like to generate some images tonight. Do you have any ideas besides
>1girl, solo, standing, portrait aspect ratio
?
>>
>>109418996
Hot.
>>
>>109419541
I do but it all involves hardcore human reproduction which cannot be posted here.
>>
>>109419529
I don't know what the purpose of the lora is because I don't know the dataset. Shouldnt 256 resolution be alright if it's continued with higher resolution
>>
>>109419554
Well, maybe I will download some jeet lora myself and see what happens. Almost never use any loras anyway.
>>
>>109419446
>>109419490
It has the classic low detail and incoherency of being under trained. It is called 0.1 after all. It will probably be good when it's done.
>>
Alright I'm done with this shit for a while.
>>
>>109419560
kekstone's idea of 'higher resolution' is a single epoch all the way at the end after the model has already been fried to shit attempting to learn scaled-down 256x256 text.
>>
If Minimax H3 is good locally as their API, local video gen is about to change dramatically
https://litter.catbox.moe/7zx9q50ta57k6nkj.mp4
>>
>parents saw my futa folder
>>
>>109419573
poor guy, last time i frequented this general he was heralded as a hero, now he's a laughing stock.
>>
Every time you get excited about something you end up disappointed.
>>
File: based ani.png (65 KB, 715x471)
65 KB PNG
>>109419307
Why are you even defending a person who smiles at Comfy and agrees with whatever's said. People like that don't deserve any respect
>>
>>109419529
>it's because he's training at 256x256 like a retard.
again? wtf is his problem??
>>
File: Comp-1.jpg (1.62 MB, 2880x1695)
1.62 MB JPG
>>109419343
>>109419349
Here is a simple comparison, and don't get me wrong, Anima is pretty good, but like I said before, Krea2 makes it feel like nothing is out of place.
>>
File: Comp-2.jpg (1.63 MB, 2880x1694)
1.63 MB JPG
>>109419343
>>109419349
>>109419624
And here is another one with a known character.
>>
>>109419624
krea is cunny pilled. I think anima butchers the anatomy on lolis too much
>>
>>109419594
chroma was a good idea for a project, the first natural language NSFW on the modern flux.1 architecture. but as said before: kekstone has no vision, only ideas. he is incapable of completing a project properly because he has no actual goal for the project in the first place. he is incapable of just letting it simmer for a month, he has to train at 256x256 so he can see the 'results' faster. he decides to perform drastic architectural modifications when it doesn't converge fast enough.
>>
H3 gonna be nuts for hentai...
https://streamable.com/zaho8n?src=player-page-share
>>
>>109419624
So what's the strategy to get good anime style with krea? artist lora or you can just straight prompt it?
>>
>>109419624
Anima's hands are extremely bad.
Krea's default anime style is too strong and comes through.
>>
>>109419529
Lodestone is retarded, but not because he starts out training at 256 resolution, every major finetune does this and most large scale loras as well, and then you increase to 512, and then to 1024.

Chroma's anatomy problems was because he trained very little on 1024, hopefully he's going to do a better job on the Krea 2 finetune he's making, at least he isn't trying another pixel space training, he's wasted something like 9-10 months on Radiance and Zeta at this point, they will never be in a releasable state.
>>
>>109419666
At first you had my curiosity, but now you have my attention.
>>
>>109419446
it knows a ton more, / tons of booru tags and can be prompted with just tags even. Its also less sloppy across the board than base krea. But it lacks detail as it was only trained at 256 res so far
>>
>>109419624
>>109419636
but krea looks like soulless RL'd slop in both of these, while anima is actually pretty good...
>>
>>109419541
doomscroll through civitai dot red unc surely youll find some inspiration
>>
>>109419668
normally picking a professional anime mangaka or a big ip + style and loras for booru artists. literally the opposite of anima where its amateur fan artists and you need loras for the pros
>>
File: GIMME THAT.png (283 KB, 415x739)
283 KB PNG
>>109419666
h-hot, my gpu will be running 24/7 I can guarantee that lmao
>>
File: 704214693300733.png (650 KB, 801x747)
650 KB PNG
>>109419441
https://files.catbox.moe/qos10h.jpg
>>
>>109419668
Both in most cases, but a lora will always be the superior solution.

Here are a bunch of promptable styles: https://lumenastrum.github.io/clio-style-preview/gallery/
>>
>>109419666
>>109419689
erection*
>>
>>109419710
oh very cool website. thank you.
>>
>>109419710
>but a lora will always be the superior solution
loras are always cope in every model
>>
>>109419691
artfag here, in these comparisons, krea has more soul while anima looks like generic pornslop
>>
>>109419636
>>109419624
now do piv goon, you won't
>>
>>109419718
kek
>>
>>109419710
arr rook same
>>
>>109419666
>gooner slop
ok now we're talking, when will they release this?
>>
>>109419688
he's not starting on 256x because of some grand architectural plan. it's solely because he doesn't want to wait for 1024x because it would take to long with the hardware he has. there is no reason to start with 256x for a finetune, this isn't a base model project. the knowledge is already there within krea itself.
>>
>>109419740
1 day, 23 hours, 30 mins about
https://modelscope.cn/models/MiniMax/MiniMax-H3
https://litter.catbox.moe/zw9qn64huprgu709.mp4

>>109419688
chroma's "anatomy problems" where idiots using a base model and expecting it to behave like a RL. Use a actual chroma RL and all those issues go away
>>
File: 1770706640363586.gif (3.6 MB, 498x295)
3.6 MB GIF
>>109419666
Please tell me you will be able to run this on a regular consumer GPU.
>>
File: dafgafdasdsa.png (25 KB, 892x289)
25 KB PNG
>>109419763
3060
>>
>>109419763
>Please tell me you will be able to run this on a regular consumer GPU.
comfy said it can be run on a 3060, pretty sure that model won't be bigger than ltx
>>
>>109419757
>>109419768
>>109419769
Time to make my dick explode.
>>
>bro it will run on a 3060
>*at 640x360 5 secs long
>>
>>109419757
>https://litter.catbox.moe/zw9qn64huprgu709.mp4
based outputs anon, keem them comming
>>
>>109419769
yea it feels like its between 10-20B, just highly RLed like zimage was.

Flux3 is likely 32B being flux 2 dev continually trained. That is why it knows so much more
>>
>>109419768
now I'm interested.
my 3090 is ready.
>>
>>109419757
Dear lord, I will finally have use for my 5090 outside of image gen. Local anime is here.
>>
File: 348512580620765.png (1.46 MB, 1024x1024)
1.46 MB PNG
>>109419768
>>109419769
inb4 30 minutes per 5 seconds.
>>
>>109419779
>Flux3 is likely 32B being flux 2 dev continually trained. That is why it knows so much more
I don't think you need a model that big to know a lot of concepts, look at anima it has insane knowledge it's only a 2b model
>>
>>109419768
>>109419769
its not going to "run on a 3060" like you retards think it is. it will take 9 hours for a 128x72 video. he's just shilling his dynamic vram offloading. this is no different than saying "it can run on your CPU", it's a meaningless statement.
>>
>>109419748
>there is no reason to start with 256x for a finetune
Of course it is, the reasoning is the exact same as when training the base model, the primary reason to do a full finetune is the introduce new concepts or finetune existing concepts that are very poorly trained, so you use the same methodology as with full base model, learn in resolution stages.

Every finetune even from the SDXL days does this, have you been in a coma ?
>>
>>109419791
if they're not retarded maybe they also include a turbo lora, I really hope so
>>
>>109419797
nah, if its a 4-8 step distill it will prob be LTX speed
>>
>>109419668
Just test and see what the model knows and compensate for the rest with LoRAs.
>>109419678
Still, in my opinion it's better at preserving styles and applying them.
>>109419691
>debatable

>PS. Seems like quoting multiple posts at once is considered spam now
>>
>>109419803
>it will prob be LTX speed
not only that but we have int8 now, and that makes inference 2x faster, we're eating good boys!
>>
>>109419791
>inb4 30 minutes per 5 seconds.
idk he said slow, not *very* slow.
>>
>>109419727
They are both good at it, although the lack of controlnet for Anima kills it.
>>109419645
Just the way god intended.
>>109419666
Alright, name your price satan.
>>
>>109419768
https://litter.catbox.moe/4s6rkkbqufpra8qi.mp4
it's insane to think I could run that on my gpu, the quality is so good, maybe it's another Z-image turbo moment, like Alibaba they managed to make it small and good
>>
>>109419840
>https://litter.catbox.moe/4s6rkkbqufpra8qi.mp4
Can someone make a gen that doesn't end up looking like a shitty mobile app ad?
>>
I2V will help with the slop though
https://litter.catbox.moe/lcdyw46tgg62lhxy.mp4

I still think flux will be better as it wont be as overly RLed. But thats only if we actually get a undistilled model
>>
>The release countdown is the time remaining for them to bake in aggressive safety RL training into the model.
>>
>>109419854
how about these
https://litter.catbox.moe/93mwpq7xabckgp25.mp4
https://litter.catbox.moe/ht8lwccs8hkxfd0g.mp4
https://litter.catbox.moe/at7ktrwl9d2qigk5.mp4
https://litter.catbox.moe/6kpthh3yeis7um1g.mp4
https://litter.catbox.moe/874gl59fl0n01dmq.mp4
>>
you guys need to stop hyping shit that will do 1280x720 clips on 96gb ram 32gb vram
>>
>>109419703
>no penis penetration visible
I accept your concession
>>
>>109419883
china doesn't bake cuckcages into the model
>>
>>109419870
>But thats only if we actually get a undistilled model
we'll get flux 3 dev, and dev is always a guidance distilled model
>>
>>109419883
>safety RL
china never did that, only western cucks like BFL (Flux Kontext), Krea and Ideogram added filters onto their models
>>
>>109419757
I feel shit like this needs to be banned, it is like how evs have so much power, if you let everyone have this level of tech then how will big studios survive?
>>
>localkeks already paying for api credits so they can precum to a model they can't even fit into their poorfag rigs
they really are such easy marketing cattle
>>
File: DIE YOU SCUMS.png (44 KB, 360x360)
44 KB PNG
>>109419919
>how will big studios survive?
lmao, fuck them, I would have defended Hollywood if this was hollywood of the 70s or the 90s, but modern Hollywood can die for all I care
>>
>>109419902
looking forwards to people with actual good genning skills to use this model.
>>
Lodes is making a krea chroma finetune
https://huggingface.co/lodestones/Kroma
>>
>>109419921
it will AT LEAST work on as little as 12GB vram we already know. Prob like 6GB with FNN chunking
>>
>>109419919
>I feel shit like this needs to be banned
LOL
>if you let everyone have this level of tech then how will big studios survive?
LMAO
>>
>>109419926
it's been a humiliating year for hollywood but spiderman #18 will turn the tide for now
AAA games are losing ground to smaller projects too, has the formula become too formulaic
>>
>>109419928
maybe we can have goon
>>
>>109419919
This.
Whenever I see huge tech advancements like this, my first concern is the survival of hollywood, and the women who rely on onlyfans to make a living
>>
>>109419919
>but muhh heckin giant woke studiorino??
>>
>>109419919
I agree, we need to think about the studios. This is going to kill jobs!
>>
>>109419943
keek
>>
>>109419902
Can't wait for the finetune on all available hentai.
>>
>>109419928
Yeah I knew that, didn't know he had started dumping checkpoints though, thanks anon

This is the ideal model for him, Krea 2already knows a lot of NSFW and has great anatomy overall, trains extremely well, undistilled base model, really good overall quality

If he fucks this up it's all on him, he isn't even doing his shitty pixel space attempts this time around, so there's actually hope
>>
>>109419929
work and perform well are two different things.
>no info about quants
>no info about speed
>no info about parameter count or model size
you are getting excited over marketing tidbits thinking this is going to be some hyper-optimized turbo model that anyone can run. wan 2.2 unquanted requires an 80GB card.
you can quant-down pretty much anything. people "run deepseek" on 12gb, but it's not the same deepseek that was trading blows with GPT and Gemini.
>>
>>109419903
I'd settle for 960x540 on 60gb sysram and 24gb vram
>>
>>109419975
>wan 2.2 unquanted requires an 80GB card.
why would you want to run an unquantized model when 8bit has an equivalent quality
>>
>>109419975
lol no you dont. You can run full FP16 wan2.2 on 12GB vram with maybe a 5-10% slow down with DDR5. its called offloading / weight streaming
>>
>>109419757
>Use a actual chroma RL and all those issues go away
may_i_see_it.jpg

also anima is a base model and doesn't have nearly as bad anatomy problems as chroma despite being a fraction of the size
>>
>>109419903
that's still better than hyping up flux 3 imo, at least with minimax, the videos you see on the internet will be exactly the same locally, flux 3 dev won't have the same level of flux 3 max videos you see in twitter
>>
>>109419995
anima is not a base model. There is a reason why you cant train it worth shit. It has RL cooked in
>>
>>109419941
>has the formula become too formulaic
It's the woke shit, people are done with the LGBTQ and uglifying of women and 'le evil white men'

Halo releases and has a ugly bitch lecturing Master Chief, it's so over for current day AAA industry
>>
>>109419902
ltx is probably obsolete in 2 days, if that model has the same size
>>
>>109420010
>it's so over for current day AAA industry
good, this industry needs to burn so that we can start all over from the ashes
>>
No I2V for EU / UK cucks
https://digital-strategy.ec.europa.eu/en/policies/regulatory-framework-ai?utm_source=chatgpt.com
>>
>>109420028
tldr benchod
>>
>>109420033
Not allowed to even have the POSSIBILITY to make / edit images / videos of others
>>
>>109420007
It has the seed diversity of a base model, the style randomness of a base model, prompts like a base model, has the artist knowledge of a base model, and is openly stated to be a raw base model by the dev.
>It has RL cooked in
do you have any proof at all for this or
>>
>>109420028
>source=chatgpt.com
loool
>>
>>109419984
Why would you run 8bit when 2bit has the equivalent quality
>>
>>109420040
thats a good thing though
>>
>>109420033
>it's not trustworthy unless we've approved it first
>>
I don't believe a single word coming out of comfyanon.
>>
>>109420011
It was kind of obsolete from a community standard anyway, it never managed to beat the Wan2.2 NSFW loras, so it's just slower
>>
>>109420046
sure, try training it
>>
Did some I2V video with minimax + midjourney v8.2 as an image, it's really good at text
https://litter.catbox.moe/y8c939hufy1rdz5v.mp4
>>
>>109420057
I believe him when he says he will never implement gguf. what a faggot
>>
So when is Minimax out? Is it on Venice yet?
>>
tried out that booru krea tune with very simple on bed, spread pussy, 1girl, complete waste of time don't bother yet
>>
>>109419797
>We will never anything as sora on local hardware
>wan 2.2
>ACK
You must be fucking new here.
>>
>>109420067
good. GGUF is useless. Just stream the weights for the faster and better int8 convrot. Or FP16 if you really must have full precision
>>
>>109420081
>Just stream the weights for the faster and better int8 convrot.
people don't have that much vram so they use Q6 or Q5, that's why GGUF is a must
>>
>>109419690
>>109419688
You don't understand Krea 2 if you think this LoRA's issues can't be fixed by a single 2-Pass. I generated these with Chroma v1 as starting image (bad/incoherent details and all) https://litter.catbox.moe/qvncvcg8if1ofwti.png

My only qualm with this LoRA is that it does not deslop as much as my v1 mix yet, at least at 2K not to the amount that I wanted, but that is expected given the small scale of the training. The LoRA does a decent deslop job at 1k res.
>>
>>109420066
'dites won't like this one
>>
>>109420078
don't call kekstone's dogshit a booru tune. it has maybe 100 anime images at most, the rest is know-your-meme garbage, reddit fanart, and furry diaperscat.
>>
File: 763919087318357.png (849 KB, 750x844)
849 KB PNG
>>109419906
penises are kinda gay, but fine https://i.postimg.cc/RSqj01xQ/456773922367.jpg
>>
https://ai-act-service-desk.ec.europa.eu/en/ai-act/article-2
No ai for EU
>>
>>109420065
you know what you're right, anima is completely untrainable, the fact that there's thousands of loras means nothing because ALL of them are bad, this is proof that anima isn't actually a base model and the dev is just lying and refused to release the actual base model for some reason
>>
>>109420116
>No ai for EU
then why BFL is allowed to make AI models then? they're a German company
>>
>>109420091
retard, you dont have to fit all the weights at once with non gguf. That is literally a flaw of GGUF on top of it being much slower
>>
>>109420124
germany isnt part of the eu you retard
>>
>>109420124
its why its so cucked as a service and will be neutered on release. They are liable if anyone makes any nsfw of anyone / virtual young looking characters
>>
>>109420135
lol what
>>
>>109420135
that's right they ARE the EU
>>
>>109420135
are you braindead or something?
>>
>>109420124
>wait, lewdity in media is bad and nono but onlyfans and pornhub are fine?? what gives!
anyone who uses a BFL model is absolute goodgoy cattle
>>
I'd say it's pretty good on anime
https://litter.catbox.moe/sjuvip3u93rmpi02.mp4
https://litter.catbox.moe/sbuyw7e7klwh4f9x.mp4
>>
>>109420081
Q8 is superior and it's possible to stream
>>
>>109420161
How are you making this shit?
>>
>>109420125
>you dont have to fit all the weights at once with non gguf.
enjoy your 2x slower outputs retard
>>
>>109420168
see >>109419921
>>
>>109420161
where goon
>>
File: orodSh_00169_.jpg (1.42 MB, 1776x2560)
1.42 MB JPG
>>
>>109420168
https://openrouter.ai/minimax/hailuo-3
>>
>>109420066
insane levels of slop.
>>
>>109420161
Good, but not excellent. These are considerably slopped outputs if slo-mo or 3D wasn't part of the prompt.
>>
>>109420181
Can it do NSFW? That's all I care about. I'm not pouring money into a cucked API model.
>>
>>109420186
do better then
>>
File: int82.png (177 KB, 1158x1864)
177 KB PNG
>>109420167
mathematically wrong. Also Q8 is like 2.25x slower than int8 convrot
>>
>>109420161
one year ago this would be unthinkable to have locally
today it is meh
>>
>>109420191
we can, with seedance 2.5 on comfycloud
>>
>>109420188
>Can it do NSFW? That's all I care about.
>>109420179
>where goon
https://litter.catbox.moe/zw9qn64huprgu709.mp4
https://streamable.com/zaho8n?src=player-page-share
>>
>>109420191
I don't have access to the Flux.3 API but that would be a piece of cake.
>>
>>109420205
How can I know its from that? That could be from Seedance. Plus I'm into real people. Can it do real people?
>>
>>109420200
it's a huge jump compared to LTX or Wan though, I'll definitely take it
>>109420208
when will you understand that what API Flux 3 is doing is not what Flux 3 dev will be doing? we won't get something as good, don't dream
>>
>>109420208
https://flux3.dev/
>>
someone is going to remake one punch man s3 with AI and the redditors are going to somehow do a turn to defend the original as soulful
>>
>>109420113
nigga stop getting groomed by the plebbitor
>>
Can someone make like a none chink coded gen?
>>
File: 1754863455288562.png (284 KB, 2255x1318)
284 KB PNG
>>109420197
depends on the model, for Zimage turbo Q8 wins
https://github.com/BobJohnson24/ComfyUI-INT8-Fast/blob/main/Metrics.md
>>
>>109420222
fake scam website that uses another model
>>
>>109420093
Why are you bothering with this lora anyway ? It's clearly just an experiment and it's nowhere near done.

The main thing of interest is the Krea 2 finetune which has just started dumping checkpoints: https://huggingface.co/lodestones/Kroma
>>
>>109420116
>113 separate articles you have to adhere to if you're doing any AI work in the EU
>"why is our economy stagnating? why don't we have any big tech companies or unicorn startups?"
europoors are absolutely fucking retarded. the entire continent is doomed. will be a complete nothingburger economic backwater in less than 30 years, while america and china share dual superpower status
>>
>>109420222
it's a scam site, flux 3 dev doesn't exist yet
https://bfl.ai/blog/flux-3
>>
>>109420197
who the fuck is quanting a small piece of shit model like anima lol
>>
>>109420239
he said before it was badly quanted. He converted stuff that should not have been converted. When done right int8 convrot is better
>>
>>109420205
It skips too many frames. Possibly it won't be as bad with interpolation, but needing it for a model of this scale is bad. I still think the model is distilled, because its realism is so far-off from what it can do with anime. Previous Minimax versions had better realism than what they're giving us with H3.
>>
>>109420254
>he said before it was badly quanted.
where did he say that?
>>
>has access to latest video model
>only gens anime slop
WHYYYYYYYYY
we want real human 3d goon material, not this weeb crap
>>
>>109420264
on bandoco discord long ago
>>
i'm gonna be playing beast of reincarnation instead of fucking with some slopped video models desu
>>
do some anons really think Anima doesn't supersede Illustrious in every way?
>>
>>109420265
That's what I mean. I don't care about pedo weeb shit. Its a good model but if it can only do cartoon shit well, and Live action is slop, pass.
>>
>>109420275
I do.
Illustrious is still better than Anima.
>>
>>109419902
>https://litter.catbox.moe/6kpthh3yeis7um1g.mp4
Can it turn manga into anime?
>>
>>109420271
from a cursory glance, it seems you're just exchanging one kind of slop for another
>>
>>109420275
anima is the rejected middle child now, booru tags are done better on illustrious, natural lang krea
>>
>>109420253
Me. I can run llm, imagegen and tts at the same time
>>
>>109420295
krea can do booru tags good now with the chroma tune. But its too low res
>>
>>109420275
if real anons thought that there would be illustrious images posted still instead of only anima
>>
File: orodSh_00177_.jpg (1.38 MB, 1776x2560)
1.38 MB JPG
>>
>>109420305
why are you using cumfart at all for that? seems like the app is just bloat when it's all in kobold
>>
>>109420287
last time I asked for a gen an anon thought looked better on Illustrious the Illustrious gen had a completely fucked up background and chairs and didn't follow the prompt properly
but the Anima gen lost because it was "slightly blurry"
>>
H3
https://litter.catbox.moe/ct81rvffggainzaz.mp4
>>
>>109420265
>>has access
everyone can test that out on the API, why won't you do the same? >>109420181
>>
>>109420222
>mogao
>happyhorse
>flux3.dev
how do localkeks continue to fall for the exact same trick time and time again??? are they actually brown?
>>
>>109420205
>no genitals
>no nipples
>goon
lol
>>
>>109420327
have you not seen anima backgrounds in your life? it's just as shit but with worse hands
>>
>>109420331
HOLY SHIT lmao, is this i2v or did you use references?
>>
>>109420337
you're blind
>>
File: Editing.png (1.57 MB, 2286x952)
1.57 MB PNG
You know, I don't use Klein Edit enough. It's pretty good! Would've took me eons in PS to get a decent image, but it did that in 119s.
>>
>>109420325
I don't. I added sd support to my local copy of exllamav3 to avoid wasting vram
>>
>>109420346
>the only content
>>
>>109420248
Because look at my Chroma-Krea outputs. Any kind of deslopping is good even if there's mistakes in it. Even this tiny LoRA is useful for deslopping Krea.
>>
File: nice.png (140 KB, 480x270)
140 KB PNG
>>109420331
>>
>>109420331
kek
>>
>>109420357
the model doesn't care.
>>
>>109420331
While this is hilarious, its also hilarious in how shit it is lmao.
>>
>>109420331
Horrible and slopped
>>
>>109420346
>I don't use Klein Edit enough. It's pretty good!
it's the best edit model we have yeah, without that BFL would be seen as much as a clown as kekestone
>>
>>109420378
I care
>>
>>109420346
is it better than turbo?
>>
>>109420346
Is that the default comfy workflow?
>>
>>109420346
BFL staff always appears on these threads to shill their company when they're about to publish a new model, too bad they already got mogged by MinMax
>>
>>109420382
like every video model, they really shine with I2V, it works fine with realism if you give it a realistic image
https://litter.catbox.moe/whhlkcslnqfbe311.mp4
>>
>>109420409
what happens if you put a naked anime as ref?
>>
>>109419795
Yeah Flux 3 can be just 12 or 20B, hyper optimized and purposely trained at that to ensure VRAMlets can run it with crazy optimizations.
>>
>>109420423
will know when weights release. API wont allow that. You can get nsfw to slip through the output filter if its fast is all
>>
https://litter.catbox.moe/opk3c341drvy4u2x.mp4
goddam I know I'll see even more gugu gaga slop on tiktok from now on...
>>
>>109420409
Ah, yes. It does men standing there and doing almost fuck all in a stock commercial setting. That's not impressive and you know it anon.
>>
>>109420402
>too bad they already got mogged by MinMax
Yeah, given that there is zero doubt MiniMax will be easier to train NSFW on than anything from censor crazy BFL, Flux 3 is DOA

Flux Edit sees some use, but for everything else image-wise BFL was killed by ZiT and now Krea 2, and now MiniMax will their video model

Based
>>
Is it actually better than seedance? According to the mememarks it is, but I swear I've seen better ai video than these examples.
>>
>>109420357
>>109420390
Ironically, I edited that prompt too many times.
>>
File: Krea2_turbo_00434_.png (1.33 MB, 1368x768)
1.33 MB PNG
>>
>>109420453
>Flux 3 is DOA
I can tell flux 3 max is a superior model, but first of all, we'll get flux 3 dev, and it won't be as good, and since they come up with multimodal shit, it'll be some giant bloated model, so yeah, smells like DOA in germany
>>
>>109420455
it has less artifacts than seedance 2. Both lose to flux 3 for anything realistic / non slopped though
>>
>>109420461
oh fuck yeah now that's some collage material I'm gonna jerk to this right now and then again when I see in the collage in the next thread because that's collage material right there
>>
>>109420327
>but the Anima gen lost because it was "slightly blurry"
which is insane, because if you put illustrious vs anima side-by-side, EVERY illustrious image is "slightly blurry" because of the dogshit SDXL VAE. you can instantly notice it on every single gen
>>
>>109420463
>since they come up with multimodal shit
the fuck are you on about. All these video models are multimodal. You start training them on image, then video, then video with audio. That has been every video model so far
>>
>>109420441
you deserve it for browsing chinktok
>>
File: minimax.webm (1.55 MB, 492x876)
1.55 MB
1.55 MB WEBM
yup, im sold
>>
>>109420476
you know what I'm talking about retard, flux 3 will output video and image, it's meant to do anything, minimax is only specialized on video
>>
>>109420474
based
>>
>>109420475
vae is less important. Krea is winning despite its dogshit vae cause it knows SO MUCH compared to anything else and prompts / generalizes so well.

You can always just pass it through a 1 step upscale with a better vae like klein or such after anyways
>>
>>109420441
>horse sounds on the cock
lmao
>>
>>109420483
retard. Images are just 1 frame from the video. With more steps spent on it. It costs nothing more than what it would already learn
>>
This. Is. Insane
https://x.com/BytePlusGlobal/status/2083069262969844158/video/1
>>
File: minimax_2.webm (1.56 MB, 486x866)
1.56 MB
1.56 MB WEBM
>>
>>109420455
It's not even HappyHorse 1.0 tier, it's choppy and skips frames, can't handle complex fight scenes or anything with dynamic movement outside of anime. It's inconsistent in speed of anime gens. I'd say those mememarks were gamed, or we are simply getting a different, distilled model than what appeared on those mememarks.
>>
>>109420498
>https://video.twimg.com/amplify_video/2083066957230915584/vid/avc1/3840x2160/uPnVRwexFGDPF_9J.mp4
post direct links lil bro
>>
>>109420507
>api only seeddance 2.5
why post this here
>>
>>109420497
flux 3 will be able to do image edit as well, it's not just a video model, it's meant to be used as a video model, as a video edit model, as an image model, as an image edit model, yes Minimax will do 1 frame but it'll look like shit because they never added more layers to be good on images, flux 3 is bloat, and that bloat is gonna kill that model
>>
>>109420481
finally a good gen.
>>
>>109420511
ComfyCloud is in the OP.
>>
>>109420498
>>109420507
I don't see any improvement compared to seedance 2.0, my goat is washed...
>>
>>109420503
That's the kind of gens that really matter. Thank you.
>>
>>109420516
ComfyLocal is in the OP too.
>>
File: minimax_3.webm (1.63 MB, 876x492)
1.63 MB
1.63 MB WEBM
>>
>>109420512
lol, have you seen H3's feature list. It does all that and more. Its not bloat. It means better generalization.
https://litter.catbox.moe/gqdga49li8zcrrut.mp4
https://litter.catbox.moe/k0vcujmci721cb7b.mp4
https://litter.catbox.moe/6yw5g6gz1zx12npu.mov
>>
>>109418548
>Music Models
hi nigbo
>>
>>109420516
>$300 a year to generate that slop
lmao
>>
>>109420535
yes, ComfyUI has bridged the divide between local and cloud. there isn't any kind of console wars rivalry anymore, we can freely share both local and api gens in this thread.
>>
>>109420537
>Its not bloat. It means better generalization.
it means it's a giant model, no one will run that, people who only care about videos will use a model that can do only videos
>>
File: minimax_4.webm (1.66 MB, 730x410)
1.66 MB
1.66 MB WEBM
>>
>>109420536
>no feathers
>>
>>109420557
I will run whatever is best. Even if its 100B that is like a 5% slow down with offloading using DDR5
>>
File: reallygay.png (45 KB, 661x718)
45 KB PNG
Is there any way a British nigga can mirror civitai pages without having to pitifully beg strangers for a catbox link like a pauper? Places like civitarchive and seaart would just use the civitai page as a download link (Really gay/retarded)
>>
>>109420556
>we can freely share both local and api gens in this thread.
pretty sure you'll be banned for off topic posts, but you can try though, go ahead!
>>
>>109420571
>>109420571
>>
>>109420559
>immediately turns into the metaverse when the camera turns
im noooooticing it's not really good at style preservation. a lot of 'anime' also has a genshin impact/CyberConnect2 3d look to it
>>
>>109420575
>>109420577
every tim
>>
>>109420577
too late schizo kek
>>
>>109420577
epic fail!
>>
>>109420498
That is just enhanced I2V and slightly longer gen times than Flux.3, nothing too crazy or that Flux.3 can't handle.
>>
>>109420577
>Maintain Thread Quality
>https://rentry.org/me
>https://rentry.org/you

Real Thread:
>>109420571
>>109420571
>>109420571
>>
>>109420503
This is the first 3dpd video gen posted here that actually looks good with almost no slop. Good job.
>>
>>109420498
>seddance 2.5
>slightly plastic faces
>slightly stiff/weird lip movement during speech
>kinda slopped keyframe composition
>fucked eyes for some faces during basic slight movement
>very slight very transparent film with slightly noisy artifacts across the image

it does seem slightly better compared to 2.0 overall but not by too much if at all for some things, its more cinematic and maybe more coherent, although u need to compare results directly across a few seeds with harder prompts to really see



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.