[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: collage_1785448399_1.jpg (1.01 MB, 2863x2101)
1.01 MB JPG
Art Edition

Discussion and Development of Local Image, Video, and Music Models and Software

Previous: >>109410072

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Z
https://huggingface.co/Tongyi-MAI/Z-Image

>Qwen
https://huggingface.co/collections/Qwen/qwen-image

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>LTX-2.3
https://huggingface.co/collections/Lightricks/ltx-23

>Wan
https://github.com/Wan-Video/Wan2.2

>Chroma
https://huggingface.co/lodestones/Chroma1-Base
https://rentry.org/mvu52t46

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
damn, local really fell behind. it feels like api is the only choice if you want actual cutting-edge quality. what went wrong???
>>
Amazing collage, if you dislike it you are Debo or Ani btw
>>
>>109413290
local was never about being cutting edge you moron.
>>
>>109413302
what about gpt 4.5?
>>
>>109413302
holy cope
>>
>>109413308
>>109413311
Local is about freedom. Nobody is pretending like local is the most cutting edge in ANY front.
I know you're trolling but I know there's a few retards that genuinely believe your nigger brained takes.
>>
>>109413272
>94% target (mostly ordinary) women
Because they are beautiful
>>
i have a question
>>
>>109413327
anti-ai luddite trannies hate that though
>>
>the troll got beat to the bake so he immediately pivots to FUDing
>>
Where good baker?
>>
File: Krea2_turbo_00676_.png (1.24 MB, 1024x1024)
1.24 MB PNG
best i could do for now
>>
>>109413342
use the fluxfusion refusal lora to get girl to suck on it
>>
>>109413272
>>Maintain Thread Quality
>https://rentry.org/debo
>https://rentry.org/animanon
so you are using proxies still? will mods ever get rid of you? using proxies is against the site rules
>>
>>109413352
ive gotten her to suck on bananas but only separate bananas that it generates
>>
>>109413363
He's wasting his life away to troll a 4chan thread. Just let him be pathetic and wallow in his filth.
>>
File: 1468547544854.jpg (93 KB, 1000x1000)
93 KB JPG
>>109413339
you ask for the good baker, yet you always complain when he bakes. curious.
>>
>>109413338
*sucks the prostate out of your butthole*, what now retard?
>>
>EVGA GTX 760 graphics card on craigslist for 25 bucks
Can I use that for anything
>>
>>109413384
you might be able to play tetris on it if you lower the graphics settings
>>
>109413363
>109413375
Do you think any newfrens fell for these posts?
>>
>>109413396
*rapes you gently*
>>
>>109413378
What's up faggot? Monkey pox healed yet?
>>
>109413396
Fell for what? The fact I stated? The people trolling here day in and day out are pathetic people with nothing going on.
>>
i don't know who any of you guys are, and at this point i'm too afraid to ask
>>
>>109413396
You've done him
>>
>>109413245
There is no workflow since Comfy doesn't support ACEStep all that much (and it's slower on it).
Inference is done on
https://github.com/ServeurpersoCom/acestep.cpp

This one has some experimental features and is based on the above
https://github.com/scragnog/HOT-Step-CPP
50 steps, Scrag's VAE, DiT-only generations (LM Disabled), DPM++ 3M sampler, CFG 12-20 (mostly 20), rest of settings default. Note those were LoRA outputs so of course they superior to what you get out of the box, but what you get out of the box isn't that bad either since now you can prompt for anything that the Base model was trained to do with much fewer issues on the 0.3 merge (Turbo had crappy sound quality and also lacked key capabilities of the Base model, and SFT had heavily diluted creativity anyways and on had too many issues on XL).
>>
do you think any newfrens fell for the local meme? wasting hours trying to use klein or ltx instead of just using gpt and seedance?
>>
>>109413365
sucking on the peepee of bananafren require FFrefusal lora des
>>
thanks for the music model musicman. i think my laptop is still a bit too shit to do music but ill save the model until i can get a better one
>>
>>109413412
I'm the King of LDG, bow before me
>>
>>109413416
Hi I work for a small firm, would you be interested in us supporting it on Comfy?
>>
File: 789654.webm (1.91 MB, 256x448)
1.91 MB
1.91 MB WEBM
>officer aryan, we need you to stop the api cuck spergout
>yes sir. hut hut hut!
>down with local! down with loc- *POW!* ACK!!!!!!
>>
File: AdelAI.png (1.97 MB, 1659x948)
1.97 MB PNG
Babe Babe Wake UP! A new AI model has just dropped from AIdeaLab, a Japanese startup studio going full send on anime-focused generative AI with text to video, image to video, frame interpolation.

They're not fine-tuning someone else's model, they're building novel GenAI research from scratch across image gen, LLMs, font style transfer, and speech recognition.

https://huggingface.co/aidealab

https://huggingface.co/spaces/aidealab/AnimeGen-T2V
https://huggingface.co/spaces/aidealab/AnimeGen-I2V
https://huggingface.co/spaces/aidealab/AnimeGen-Frame-Interpolation
https://huggingface.co/spaces/aidealab/AIdeaLab-VideoJP
>>
>>109413439
anima bross!?!?!
>>
>>109413432
I'm not the creator of the model, I just made the merge. AFAIK there already is Comfy support for ACEStep 1.5 and ACEStep 1.5 XL (no idea about gguf though), but I don't use it since it's a bit inefficient
>>
>>109413439
Surprised it took this long to see more dedicated models here.
>>
>>109413439
Someone needs to put this through ComfyUI immediately!!!
>>
>>109413439
it only does video?
>>
>>109413439
you got me excited then I realized its video only
>>
>>109413439
>They're not fine-tuning someone else's model
Like the brazilians did? Yeah I don't believe you
>>
>>109413439
>anima this year
>krea 2 this year
This space is moving WAY too fast
>>
>>109413456
please bring it to comfycloud first
>>
>>109413463
>gen a single frame
>???
>profit
>>
>>109413463
why? anime images got solved years ago
>>
>>109413464
>>109413463
35 star status?
>>
>>109413439
Neat, shame it's separate models for t2v and i2v though
>>
>>109413471
????
>>
>>109413439
can we at least wait for lodestones to finish finetuning before releasing a new model? slow the fuck down like holy shit
>>
>>109413478
35 star status?
>>
>>109413469
not really, anima and illustrious are meh. krea is good but obviously lacks knowledge and anime won't ever get any better due to those niggers extreme heavy focus on booru tags.

>>109413471
relevant post status?
>>
File: 51VWH3J24TL.jpg (102 KB, 1000x1499)
102 KB JPG
>>109413439
either enormous VRAM requirements or an API with insane pricing
>>
File: Screenshot_3535.png (14 KB, 664x188)
14 KB PNG
>>
>>109413479
huh?
>>
>>109413484
Chang has done it again
>>
>>109413480
>extreme heavy focus on booru tags
so you know nothing about Anima, you just threw it under the bus for no reason
>>
>>109413484
>>
File: ekk,kweekkk.png (1.17 MB, 864x1184)
1.17 MB PNG
>>109413484
jejypow JSID bruh its over
>>
>>109413439
>They're not fine-tuning someone else's model, they're building novel GenAI research from scratch
> Wan2.2
can you stop lying?
>>
>>109413492
anima is too underpowered. it fucks up lines all the time. try getting it to gen a chair without it fucking up the geometry consistently.
>>
File: 1761311588225788.png (161 KB, 500x500)
161 KB PNG
>>109413484
I was right as usual >>109413464
>>
>>109413439
>AnimeGen T2V is based on Wan 2.2 T2V A14B and further trained to improve anime-style motion, character rendering, color composition, and visual consistency for animation-oriented use cases
High noise/low noise split models. I am NOT going back to that
>>
>>109413498
gen goon with krea 2, you can't unless is the most vanilla shit ever
>>
>>109413497
>can you stop lying?
35 star status?
>>
>>109413505
this is bait, you're not getting my Krea 2 goon gens for free
>>
>>109413505
irrelevant to what I said. is that your coping mechanism?
>>
>>109413439
Neat, but just a toy compared to Flux.3
>>
>>109413510
>>109413511
uh oh, mad your lies didn't pan out? 35 star status?
>>
i dont get why they always finetune wan instead of ltx
>>
>>109413517
what the fuck are you talking about you empty headed slut
>>
>>109413517
okay, liberal. any more projections?
>>
File: 1120336386556824.png (3.33 MB, 1536x1536)
3.33 MB PNG
>>109413416
Ah shit, reposting cuz I didn't notice the old thread died. Hey I'm the anon that asked for help few threads back. Thanks again, I was able to setup everything and trained a LoRA on some jazz-rock music. Overall the model is fun to play with, but it is fairly limited. Especially for more sophisticated music I think the base model is not quite good enough.

This is about as good as I could get from my limited playing around:

https://files.catbox.moe/j7ym5l.mp3

Still though, it is fun. Think I'll try another training run with a much more aggressive LR to see what happens.
>>
>>109413504
>High noise/low noise split models
woah its like we took a time machine into the past
>>
>>109413439
Oh can't wait for tdrusell working on it and make a turbo lora :^)
>>
Blessed thread of frenship
>>
File: ComfyUI_00141_.jpg (1.19 MB, 1632x2176)
1.19 MB JPG
>>
>>109413545
catbox?
>>
>>109413545
me on the right
>>
>>109413439
I feel like I saw this a week or two ago at some conference, but the examples were not good and I forgot about it.
>>
>>109413484
>one year later, local STILL on wan 2.2
local is so pathetic, it's embarrassing comfyui has local nodes anymore at all.
>>
>>109413545
me in the background getting told I need to leave for touching womens asses
>>
>>109413565
just like your model faggot
catastrophic forgetting status?
>>
>>109413575
My models are perfection and grace. (not anima)
>>
File: 4557845437894.png (18 KB, 900x806)
18 KB PNG
where is kino?
>>
>>109413272
>>109413272
Whoever made this disgusting ugly collage deserves to get shot and have their GPU taken away from them.

Horrible!
>>
>>109413595
under there
>>
>>109413608
its art
>>
>freeze adapter
>tfw it just werks
????
>>
>>109413608
its okay your gen or your favorite anon gen will be in the next faggollage im sure of it
>>
>>109413532
Np anon
>much more aggressive LR
How complex is the music and how many songs was it? Can you show example of input audio? I trained an Initial D LoRA using my usual settings before and had poor results, so I increased the rank/alpha to 256/512 and lowered LR to 0.00009 and had great results starting at early epochs like 400ish (I even saved it up to 1000). For very complex music, I suggest you try it. I posted the model here https://civitai.com/models/2702491/super-eurobeats-acestep-15-xl
In that case I was inferencing on Base/Turbo 0.5 merge. My 0.3 Merge has improved results for that LoRA, but I noticed because it was trained at high rank I had to lower the last epoch model to 0.95 to get improved results.
https://vocaroo.com/1i510IAPHvb8
(Catbox and litterbox both shitting the bed for this file)

Things to try if you haven't already: Is LM disabled? Model is much worse with it enabled. Also try STORM sampler, haven't had to use it myself but I see others talking about that improving sound quality.
>>
>>109413608
Should have accepted by now that the baker is a retard only a few IQ points higher than the troll baker.
>>
>>109413484
Do they not know how easy it is to check their claim or did they think no one would notice?
>>
>>109413662
Also, try to save every 20/25 epochs, the more checkpoints to look at, the better. I tend to save even intermediate "best" checkpoints that I test as I go, I have gems through those.
>>
>>109413637
It's working! it's.... sorry I forgot who naruto was
>>
>decide to not train nothing and leave everything as it is
>tfw it just werks
????
>>
File: Krea2_turbo_00705_.png (1.09 MB, 1024x1024)
1.09 MB PNG
>>
>>109413696
damn, she's hot and has lots of holes. name?
>>
>>109413699
spunkbabe
>>
>>109413699
Alysa Liu
>>
>>109413699
Mother, Your
>>
File: ComfyUI_00151_.jpg (1.01 MB, 1632x2176)
1.01 MB JPG
>>109413559
It's Ideogram+a lora I'm still training. It'll be no use to you brutha
>>
>>109413637
>tfw the guy who originally started the "catastrophic forgetting" rumor "confirmed" his theory by ripping out half of the models arc
>>
File: ComfyUI_temp_tlsgv_00017_.png (2.53 MB, 1248x1728)
2.53 MB PNG
>>
>>109413439
>wan 2.2 finetune
I sleep
>>
File: ComfyUI_temp_tlsgv_00033_.png (3.23 MB, 1152x1920)
3.23 MB PNG
>>
>>109413662
Originaly I tried with a LR of 0.0005 dataset of 25 songs for 1500 epochs. Input audio for example Aja by Steely Dan, it's pretty hard to even tag because musically it's quite complex with shifting time signatures, weird ass modal jazz sections, etc... I'll try another run with a high LR like 2e-4 just to see what happens, but ultimately I think this kind of music is a bit too ambitious for now.
>>
>>109413800
still waiting for the 80s rock
>>
File: 1768986845160436.png (329 KB, 1275x769)
329 KB PNG
What went so fucking wrong?
>>
>>109413804
I'm sorry to say, but this is as close as you're gonna get brotherman https://files.catbox.moe/j7ym5l.mp3
>>
>>109413807
Communist thoughts always leads to laziness.
>>
>>109413807
>still more stars than cumstainUI
holy shit, how is he so powerful? not even api nodes can save cumrag
>>
>>109413807
Blast from the past.
>>
>>109413819
>anifart talking about stars
kek
>>
File: debo_pc_k2_00011_.png (2.99 MB, 1920x1280)
2.99 MB PNG
>>109413807
>What went so fucking wrong?
dota ruined him. he coulda had it all, but he chose to play dota instead
>>
File: 9f5fa7.gif (768 KB, 240x240)
768 KB GIF
can local do this?
>>
>>109413843
that is local. only the real ones know the name of the model
>>
>>109413838
man fuck off
>>
>>109413838
A tale as old as time (or at least 20 years).
>>
>>109413807
pytorch got too gay so he dropped it. he still contributes to llama cpp every blue moon. Comfy scaled on a pile of shit and not many people like it anymore. Voldy knows when a webapp ran it's course
>>
>>109413800
>0.0005
I see what you're doing wrong. LR of 0.0005 is too aggressive for this model. I'd lower it to 0.0001 or 0.0003 (the default LR recommended by devs is 0.0003 if using rank/alpha of 64/128) and try again. Do not go up, go down. If you're increasing something, only increase rank/alpha to like 128/256, but compensate for it (0.0001 LR or lower).
>>
>not many people like it anymore
>>
>>109413843
that will smith ai video meme was made by a local model acually
https://huggingface.co/ali-vilab/text-to-video-ms-1.7b
>>
>>109413862
>Downloads last month
>100,082
Still a lot of kinographers out there
>>
>>109413738
she'd go from 3/10 to 8/10 with black brows
>>
>>109413862
so local is way ahead of api then
>>
File: image.png (32 KB, 1105x170)
32 KB PNG
>>109413862
did we regress?
>>
>>109413861
It's a pretty shitty app and they can't stop putting lipstick on the pig
>>
>>109413859
I missed a 0, my b. I meant to say 0.00005.
>>
We await FurkUI
>>
>>109413869
Kinographers AND Synthologists
>>
File: Krea2_turbo_00726_.png (1.26 MB, 1024x1024)
1.26 MB PNG
>>
>>109413807
>>109413811
this

he also simply refused to add more contributors despite not caring to even merge many prs for months on end, let alone do shit himself.
>>
>>109413780
prooooompt?
>>
>>109413439
>anime
snort_mimimimi_snort_mimimimi.wav
>>
>>109413923
1girl, peace sign, mirror selfie, asian, duckface
>>
>>109413912
kek, not bad at all
>>
File: file.png (7 KB, 650x80)
7 KB PNG
Is Civit fucking with me?
>>
>>109413969
kreakeks btfo
>>
File: ComfyUI_00023_.png (2.12 MB, 1672x1254)
2.12 MB PNG
>>
>>109413888
Ah, probably just needed more epochs to converge at a lower LR like that if using rank 64. I listened to the music, don't think it's something ACEStep shouldn't be able to handle. The gen you shared was close
>>
>>109413969
Get fucked
>>
File: 230921CUI_00001_.png (550 KB, 832x1216)
550 KB PNG
>>
>>109413978
>>>/g/lmg
>>
>>109413978
after two weeks when you realize you dont really have a use case for those, donate the compute to kekstone
>>
what is going to happen to all the partner nodes when all the cloud companies die?
>>
>>109414017
desu who cares?
>>
File: Krea2_00101_.png (3.45 MB, 1536x2048)
3.45 MB PNG
>>
>>109414017
we start the stopkillingmodels movement
>>
>>109413888
>>109413984
Also, conversely, it's good try higher rank if poor results come from lower ones, but if starting from a higher rank it might lead to less creativity and worse results than intended on smaller datasets at the higher epochs.
>>
>>109414017
we saw what happened after sora: comfyorg bribes local labs to turn their models into API. once sora shut down, alibaba quickly sold out wan and happyhorse to try and fill the gap. API will never truly die
>>
>>109414064
It's a process. They are pushing (((Open))) models with sneaky commercial licences but will call them open source. Once the ecosystem is strangled out with fees and usage rights, they will go api only
>>
when is flux.3 releasin?
>>
>>109414086
approximately fourteen more days
>>
>>109414086
Two weeks after the API so they can censor the shit out of it
>>
>>109414098
ok good, we dont need more coom slop
>>
absolute localcope
>>
>>109414064
>comfyorg bribes local labs to turn their models into API
Surely you have some kind of documentation to back up this claim. Not that I don't doubt it, I just like proofs :]
>>
>>109414077
>(((Open))) models
anyone else smell the bullshit with these?
>>
apache 2 anima status?
>>
>>109414141
>>>/g/hsg
>>
>>109414141
was cancelled since it wouldnt be able to beat MUGEN
>>
>>109413969
lmao
>>
>>109413439
>anime-focused generative AI
Into the trash it goes.
>>
>try out the flux3 image gen
>legs fusing into each other at the knee
oof
>>
>>109414098
>they can censor the shit out of it
not only that, but we'll get an inferior distilled version, if flux 3 dev ends up being better than ltx 2.3 I'd be really surprised, smells like DOA ngl
>>
>>109414064
>Comfy is so powerful he can force OpenAI to remove some of their models
omfg based???
>>
when is ltx 2.6 releasing?
>>
File: 1757131076052403.png (1.36 MB, 768x1152)
1.36 MB PNG
Now this is hagmaxxxing.

Can Krea 2 do this?
>>
>>109413807
Too slow and Comfy happened. By the time Forge came, damage had already been done
>>
>>109414124
>Surely you have some kind of documentation to back up this claim
no that would defeat the whole purpose of his reply
>>
comfyorg is honestly ruining local, only the most deranged freetards think otherwise. so many models lost thanks to API nodes. even anima was an attempt to sabotage local, and comfy proceeded to buy out noobai and turn LAX into a novelai shill. comfyorg needs to burn
>>
File: flandre_flux3.png (1.5 MB, 1024x1024)
1.5 MB PNG
so this is the power of flux3
>>
>>109414190
Babe Ruth?
>>
>>109414201
flux 3 can't make images yet, but I've seen that fact a few days ago, did that change?
>>
>>109414208
the website offers image gen for what i assume is flux 3. 50 credits or so for a 1mp image. its either 3 images or a 3second video lol from the free credits
>>
>>109414214
>the website offers image gen for what i assume is flux 3.
they don't say exactly what model you were using? that sucks!
>>
>posting all dat instead of proofs
>>
>>109414201
Glad to see the Flux Moon persists
>>
Krea controlnet is a bit hacky at the moment but
>>>/ic/7998231
>>
>>109414225
actually looking at it the site itself seems pretty shady, might just be a scam honestly
https://flux3.dev/
>>
>>109414233
tbdesu this doesnt look like itd be substantially different than feeding the OG image into a vision model and using that as a prompt
>>
>>109414240
yeah, I really recommand you to stop using this, always go to their official website if you want to try it
https://bfl.ai/blog/flux-3
>>
File: ekk,kweekkk.png (1.17 MB, 864x1184)
1.17 MB PNG
>>109414240
>falling for jeet sites
>>
>>109414214
Which website? Sucks that Flux 3 is literally only API only right now but it has blown up, and we have yet to see the Dev model anywhere. It's a red flag for sure
>>
can i request some muscle mommies pls
>>
>>109414240
well they're honest
or worried about a legal shitstorm
>>
>>109414190
>hagmaxxing
>can't even make saggy withered hagtits
you're a fake haggenner
>>
>logged into the flux3 scam site
its ogre
>>
>>109414248
they can't make dev if the main model isn't finished yet, they're still on the beta phase so they're still training the main model
>>
>>109414254
>hes not ben franklin pilled
actual fake hagmaxxer exposed
>>
File: LTX_2.3_i2v_00103_.mp4 (1.4 MB, 768x1152)
1.4 MB
1.4 MB MP4
>>
File: Krea2_00094_.png (3.56 MB, 2048x1536)
3.56 MB PNG
is this hag
>>
>>109414240
seems like you can try flux 3 API wise now
https://xcancel.com/NousResearch/status/2082911477904654741#m
>>
>>109414265
They can't? Or they don't want to? It's literally a few clicks away for them to have an alpha Dev preview, or is distillation some complex, drawn out process?
>>
>>109414284
why would they spend money making a dev preview? it will be obsolete just a few weeks after once they make dev final
>>
File: 004131CUI_00001_.png (582 KB, 832x1216)
582 KB PNG
>>
why would any company release an open weight 'preview' model? so impatient brownskin mouthbreathers can start merging civitslop into it early and then claim how v0.9 was the 'real was leaked accidentally' compared to 1.0 which was 'censored and lobotomized for safety' like they said about SDXL?

just release the final model when it's good.
>>
>>109414240
I wonder why Jewgle still ranks such obvious unverified jeet scam sites. Their algorithm is gamed hard and they do nothing about it.
>>
>>109414318
>v0.9 was the 'real was leaked accidentally' compared to 1.0
kek
>>
>>109414320
>Their algorithm is gamed hard and they do nothing about it.
because they don't care, it makes them money that's all that matter to those snakes, that's also why youtube is filled with scam ads
>>
>>109414318
this, there's no reason to release an unfinished model, people are stupid enough to test it and make a final conclusion on the product
>>
>he didnt download the uncensored gemma31b before they replaced it
>>
>>109414329
>it makes them money
They're scammers too*
>>
>>109414250
pls
>>
>>109414318
>you now remember anon screeching that XL couldn't into realism because anon didn't like bokeh
>>
>>109414375
>XL couldn't into realism
still can't, it has a really dogshit VAE
>>
>>109414284
>is distillation some complex, drawn out process?
It's pretty complex, and recently Nvidia improved it, I wonder if BFL will use this to make dev
https://research.nvidia.com/labs/genair/pdd/
>>
File: debo_pc_k2_00014_.png (2.21 MB, 1920x1280)
2.21 MB PNG
>>
File: 005554CUI_00001_.png (552 KB, 832x1216)
552 KB PNG
>>
>>109414340
Every open LLM model sucks with long outputs and large context sizes, including gemma31b.
>>
File: LTX_2.3_i2v_00111_.mp4 (1.74 MB, 768x1152)
1.74 MB
1.74 MB MP4
>>
>>109414411
yep ltx sucks ass
>>
>>109414411
bootiful light2v
>>
>>109414318
its always objectively good for any company to release as many checkpoints as possible since theres a chance we get in some way better, or less censored version. if its a worse checkpoint worth nothing then its easy to just not care.
>>
Which Krea 2 model do you guys do inference on? Turbo or Raw? Do you use any NSFW fine-tunes?
Base Krea 2 is so censored it's not even fun to use with non-nsfw loras. If refuses to do even shirtless men (I prompted for a shirless fat dude and it refused), doesn't show people barefoot with regular prompting etc
>>
>>109414399
every llm period, newfag
https://github.com/adobe-research/NoLiMa
>>
>>109413327
idk why they use "target" like it's a bad thing rather than just more people prefer to look at women
>>
>>109414397
slop
>>
>>109414434
it's only 'objectively good' for you. remember when chroma was supposed to be the next 'base model for finetuning' but nobody touched it because it had 50+ different versions and nobody could decide which was the better one? just because you stalk these threads all day to keep up with the latest chingbongv0.56_3b_fp8 release doesn't mean everybody else does, and when people hear of a model but see it has a ton of different sidegrade versions they lose interest.
if a company wants their model to get popular and an ecosystem to form around it (to maximize brand recognition), then it's best to release one really good model and be done with it. if you need to release 50 different checkpoints, then maybe your model just isn't actually good in the first place
>>
File: debo_pc_k2_00017_.png (2.33 MB, 1920x1280)
2.33 MB PNG
>>109414452
thanks
>>
>>109414457
>it's only 'objectively good' for you
for all the people.
>remember when chroma was supposed to be the next 'base model for finetuning' but nobody touched it because it had 50+ different versions and nobody could decide which was the better one?
no, chroma didnt launch too far in the normieville because:
1. NO documentation and marketing, unless you were in their discord you had no idea what the fuck were they working on or what is the difference between many of the checkpoints
2. it took too long to train because they didnt have too much money compared to frontier labs, and once zit came and was better in almost every way, it was over
3. its too big and takes too long for normie vramlets who are used to basically just sdxl


so again, you can and should advertise and funnel people to 1/2 versions of your model, and release 10k different checkpoints in some repo at huggingface for powerusers if you want, and it wont be a problem.
>>
im more of a chingbongv0.53_epsilon_confetti_0.5b_nfv4 man myself
>>
>>109414499
an epsilon model will never be able to generate a black dark enough to satisfy /hdg/
>>
>>109414495
you are fundamentally misunderstanding why companies release models. local releases exist solely to advertise API through brand recognition. every single local release, from StabilityAI to Krea, operates this way. when one hears about flux dev they know what to download. at most they have to choose between a large or small model size.
it's not good for a company to release millions of checkpoints, it serves them no benefit and they're all just junk anyway. junk that creates confusion and clutter when people search for loras and such and find they were created with an experimental old version because freejesh couldn't wait 2 weeks for the full base model to release.
>>
File: snailcat0083.png (1.36 MB, 1024x1024)
1.36 MB PNG
>>
>>109414533
I agree with him, if their local models (which only serve as advertisment product) are numerous and convoluted, people won't take the API company seriously
>>
Reeks of newfaggotry for some reason
>>
File: ComfyUI_4357.jpg (621 KB, 1904x1904)
621 KB JPG
>>
File: 1784333625827844.jpg (210 KB, 1024x1024)
210 KB JPG
>>
File: ComfyUI_06060_.png (2.35 MB, 1536x1536)
2.35 MB PNG
>>
File: Ideogram__00048_.jpg (2.92 MB, 3840x2160)
2.92 MB JPG
4k gens but at what cost?
>>
>>109414755
The cost
>>
>>109413272
Ok thinking to switch to krea2 as it looks promising in generating extremely degenerate tentacle hentai porn
But I'm unsure on how wieldy the model is to use?
What is your experience so far?>>109413418
>Muh gpt muh seedance
Censored trash
>>
>>109414759
SLOOOoOOW
>>
>>109414755
it's the best at realism but it sucks ass to use and nobody wants to train for it. stupid muslim niggas.
>>
File: ComfyUI_00583_.mp4 (477 KB, 960x640)
477 KB
477 KB MP4
>>
>>109414759
What hardware?
>>
File: 1766991802220931.jpg (300 KB, 896x1856)
300 KB JPG
>>
File: ComfyUI_00584_.mp4 (747 KB, 640x960)
747 KB
747 KB MP4
>>
File: ComfyUI_00585_.mp4 (214 KB, 640x640)
214 KB
214 KB MP4
>>
>>109414533
i had a lot of respect for you lot but you are supporting that mentally ill alcoholic
you are turncoats
>>
File: Ideogram__00049_.jpg (2.89 MB, 3840x2160)
2.89 MB JPG
Big ship
>>109414783
16gb vram 64gb system ram
>>
Krea 2 Loras ABSOLUTELY needs trigger words if you are planning to do inference with Turbo. It took me a while, I finally got around testing it and it consistently evades the slop bias more often if you use trigger words
>>
>>109414790
This is my kind of collage
>>
>>109414755
>>109414759
Post workflow and I'll run an rtx pro 6000 on it for a comparison.
>>
>>109414790
heh
>>
>>109414854
I'm using my own lora, which adds a little to gen time but here is the workflow
https://files.catbox.moe/7hp8ya.png
>>
What sort of graphics card do I need to generate realistic looking things similar to grok?
>>
>>109414949
Nvidia H100, they run about $40,000
>>
File: 024121CUI_00001_.png (978 KB, 1216x832)
978 KB PNG
>>
>>109414949
>realistic looking things
>grok
?

for video wait for flux 3
for images https://civitai.com/models/2168935/z-image-turbo
>>
>>109413290
>>109413302
>>109413308
>>109413311
>>109413325
Are you guys talking about local with resource parity to remote (lol) or local in the drastically-lower-resources sense? Because obviously if it's the latter, it's not going to deliver to the same level. Unleashing a prompt on your 24-96Gb local VRAM is never gonna produce the same outcomes as the weights from 100Gb+ datacenter deployments.
>>
>>109413804
>80s rock
That needs a LoRA
>>
>>109414978
It's just a local FUD troll.
>>
File: log.jpg (1.49 MB, 2028x1844)
1.49 MB JPG
Yeah I still have no idea what the fuck is going wrong.
I asked jewgle, they said the error in my logs were an OOM issue, so I went and found the smallest starter model in comfyui that I could, disabled pin_memory, same fucking problem, still crashing with extremely minor edits on small images.
I've tried multiple models now and I just keep getting this bullshit out of memory error despite the fact that they should fit within my (admittedly modest) 9070XT + 32GB ram specs.
>>
>>109414233
cool
>>
>>109414986
>windows
>comfy-desktop
>AMD

the cursed trifecta
>>
we... ran out of pills
>>
>>109414759
>>109414854
>>109414941
Oh shit I just noticed something. I had my gpu still power limited because of training overnight. Picrel is the 4K gen time with it maxed out. A bit more bearable
>>
>>109414986
>9070XT
found it
>>
>>109414986
>no such file comfy_blue_logo.png
Is this the issue?
>>
go go go
>>109402749
>>109402749
>>109402749
>>
File: 1773558378458766.png (878 KB, 2730x1716)
878 KB PNG
https://xcancel.com/MiniMax_AI/status/2083008095488516262
Babe! Minimax H3 will be open weights!
>>
>>109415018
>Minimax H3
who?
>>
File: 00104-1695815093re.png (3.18 MB, 1920x1273)
3.18 MB PNG
test
>>
>>109415018
Does that mean that my 2070 will finally have time to shine?
>>
File: file.png (103 KB, 423x567)
103 KB PNG
>>109414996
>>comfy-desktop
I'm going to move back to the portable version and test it with the same result because of that comment.
)v:<
>>109415012
That's just the log I happened to use this time (from pic rel) but I was getting the same result from several of them, minus that error of course.
>>
>>109414986
its funny to me that the log doesnt explicitly name which quant youre running but rather only alludes to it
>>
>>109414986
you should use int8/fp8 of all models, of main model checkpoint, text encoder etc

then try with only
--disable-dynamic-vram
argument
>>
File: 1761076769963025.webm (3.83 MB, 960x720)
3.83 MB
3.83 MB WEBM
>>109415027
https://xcancel.com/pika_labs/status/2083012182477095099#m
you don't know minimax? anyways, this shit has an insane video quality, I'd put it a bit worse than Seedance 2.0 but it's actually really good
>>
>>109415038
>if that doesnt work you can fuck with it more but just use a different model, flux 2 is obsolete anyway
why did you delete this? caught in 4k, tranny
>>
>>109415046
Alright I'll give that a shot
I suppose the most annoying part is I'd already been messing around with it for the first time a few months back and didn't have any issues with minor edits to photos as well as img-to-3d stuff.
The results weren't great but at least there was a result.
>>109415060
Probably because he saw the part where I mentioned that I'd tried out a few different models with the same result.
>>
File: I call it.png (57 KB, 360x360)
57 KB PNG
>>109415018
>Flux 3 will be 40b
>Minimax will be 50b
nothing ever happens
>>
>>109415059
its plastic af
>>
File: LTX_2.3_i2v_00121_.mp4 (2.56 MB, 512x768)
2.56 MB
2.56 MB MP4
>>
>>109414986
>>109415032
this thread is for git cloners only please direct yourself to the door
>>
>>109415059
is pika where you eat inedible things
>>
>>109415077
I'm gonna use it for I2V so...
https://xcancel.com/Hoshimiko_AIart/status/2082750075797966895#m
>>
File: 158463.png (8 KB, 680x907)
8 KB PNG
>>109415071
it already happened
>>
>>109415092
fuck (((LTX))) this shit sucks, it didn't beat Wan 2.2 imo
>>
>>109415018
Another Seedance 2 tier model? Whaaat? How is it with anime?
>>
>>109415018
looks like the motion has quite some grain, and its heavily biased for usage in commercials and ads, but its basically omnimodal and has text to audio, and the rest of the stats look good enough, although we would need to test locally to see everything exactly.
>>
>>109415018
reminder to all the ungratefuls to THANK BFL. without flux 3, we would never have gotten minimax h3!
>>
>>109415077
Yeah it's slopped, but if it's open weights what does that matter?
>>
>>109415120
>How is it with anime?
pretty good >>109415090
https://xcancel.com/kiyoshi_shin/status/2082971035364946077#m
>>
>>109415125
>without flux 3, we would never have gotten minimax h3!
why do you think that? for the moment they have no choice but to use us for free advertisment because no one can compete with seedance 2.0
>>
File: 1773485848632240.png (29 KB, 535x383)
29 KB PNG
>>109415018

>Minimax H3 weights in a few days.

I assume they dont want to commit to a date because they dont want to get fucked by another company or Flux 3, all these companies for all ai models always strategically want to wait, since its the best strat.
>>
File: Untitled.jpg (835 KB, 3682x1555)
835 KB JPG
:|
>>
>>109415140
>I assume they dont want to commit to a date because they dont want to get fucked by another company or Flux 3
if your model is good, the date shouldn't matter, if flux 3 has more advantages, no one will care about minimax regardless of the fact they released it before or after
>>
>>109415018
Are they going to give us full non-distilled weights? If so that's a massive blow to BFL because there's no way anyone's going to be on Flux.3 if they give us distilled weights we can't tune. Obviously Flux.3 will still have all the kino styles but what good are they if they give us a distillation?
>>
File: this.png (168 KB, 600x338)
168 KB PNG
>>109415018
Finally, cool things are happening, 2026 had such a shit begining but things are starting to be interesting
>>
>>109415150
That's the sort of thing you should be using klein for
>>
File: 1763678761852603.png (57 KB, 220x215)
57 KB PNG
>>109415154
I'm pretty sure they're both going to wait for the other one to release their model first kek
>>
>>109415138
because being able to use safetyslopped western goyware like flux as a punching bag is better for publicity and boosts the 'based china' image.
https://en.wikipedia.org/wiki/Cool_Japan
if you think china is unaware of their reputation as the 'based open model providers sticking a finger to closedai and dario' then you're a bit silly. they capitalized on this with z-image and intentionally poked fun at BFL.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.