[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
▼ Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: Krea2_turbo_04852_.jpg (1.49 MB, 1776x2368)
1.49 MB JPG
Keeping the realm clean

Discussion and Development of Local Image, Video, and Audio Models

Previous: >>109996867

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP
Neural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Qwen Image 2.1
https://huggingface.co/Qwen/Qwen-Image-2.1

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3
https://neta.art/use-cases/en/h3-1000-prompt-list

>Anima
https://huggingface.co/circlestone-labs/Anima
https://animastyles.thetacursed.com
https://tagexplorer.github.io/
https://animadex.net

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/neo_collage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Blessed thread of frenship
>>
>mfw Resource news

10/07/2026

>Run HunyuanImage 3.0 (80B) natively in ComfyUI on a single 12–24 GB GPU
https://github.com/PedroMarinhoDev/ComfyUI-HunyuanImage3

>ComfyUI H3 Video Upsampler
https://github.com/dntpi/ComfyUI-H3-Video-Upsampler

>Veda Sparse Attention for ComfyUI (MiniMax-H3)
https://github.com/veda-sparse/Veda-on-ComfyUI

>ReDetail 2.0: Video upscaling and re-detailing for ComfyUI on LTX-2.5
https://github.com/Bambushu/redetail

>FastVideo FastH3 Trim for ComfyUI
https://huggingface.co/FastVideo/FastVideo-FastH3-Trim-Comfy

>FIBO Scene Analyzer [dev]
https://huggingface.co/briaai/fibo-scene-analyzer

>Two Halves are More than One: Phase-wise Velocity Distillation for Fast and High-Quality Image Generation
https://github.com/PolyU-VCLab/PVD

>S2PD: Serial-to-Parallel Diffusion for Physically and Logically Consistent Video Generation
https://jefequien.github.io/S2PD

>Disentangling Dual Image References in Frequency Aware Diffusion Models for Personalized Generation
https://github.com/htyjers/Dual-FDM

>Talk Like You: Imitating How You Speak in Real-Time Talking Head Generation
https://bq-wang0511.github.io/TalkLikeYou

>On Color Alignment in VAE Latent Spaces and Its Applications
https://julian075.github.io/Color_Subspace

>Test-Time Adaptation of Quantized ViTs via Single-Pass Quantizer-Aligned Recalibration
https://github.com/chahh9808/QuAR

>OpenWAM: An Open Framework for Composable World-Action Models
https://openwam.stanford.edu

>ComfyUI Pixel Art Refiners
https://github.com/envy-ai/ComfyUI-Krea2-Pixel-Art-Refiner

>Empirical Variational Autoencoder
https://github.com/mapooon/EVA

>Kroma 0.3.1 - full turbo opd model
https://huggingface.co/lodestones/Kroma#kroma-v031-opd--recommended

>Fizgig 7.1.0 - Z-Image Turbo gets a new Training Adapter
https://github.com/shootthesound/Fizgig/releases/tag/v7.1.0

10/06/2026

>Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation
https://github.com/kandinskylab/kandinsky-6
>>
>mfw Research news

10/07/2026

>Learning to Read the Contextual Tokens in DiTs
https://omer11a.github.io/learning_to_read

>UniSlider: Perceptually Uniform Sliders for Continuous Image Editing
https://color.cvc.uab.cat/unislider

>Local Content-Style Control for Diffusion-based Image Stylization
https://arxiv.org/abs/2610.08704

>Keepsake: Selective Spatial Memory for Long-Horizon Video Generation
https://arxiv.org/abs/2610.06588

>Safe Image Generation via Reinforcement Learning
https://arxiv.org/abs/2610.05908

>SPIN: Image Immunization Against Diffusion Editing via Single-Step Projection in Stochastic Neighborhoods
https://arxiv.org/abs/2610.06334

>CentriQ: Calibration-Free Quantization of DiTs via Exact Mean Centering
https://arxiv.org/abs/2610.06260

>A Fine-Grained Analysis of the LoRA Fine-Tuning Landscape with Implications for Data Selection
https://arxiv.org/abs/2610.06542

>Energy-Conditioned Noise Schedule and Whitening for Spectral Diffusion
https://arxiv.org/abs/2610.07206

>Scalable Minimal-Change Learning for Controllable Image Editing
https://arxiv.org/abs/2610.06021

>HuC-VideoMAE: Human-Centric Video Masked Autoencoding from synthetic data
https://arxiv.org/abs/2610.08433

>Building Rome from a Single Image
https://build-rome.github.io

>Beyond Transport Cost: Routing Differences between Flow Matching and Optimal Transport
https://arxiv.org/abs/2610.05921

>Backend-Agnostic Sparse Attention for Fast High-Resolution Visual Generation
https://arxiv.org/abs/2610.08772

>Harmful Content Generation in T2I Models: Capabilities and Moderation Limitations
https://arxiv.org/abs/2610.06503

>MC-Sparse: Deconstructing and Closing the Dense-Sparse Attention Gap in DiTs
https://arxiv.org/abs/2610.06801

>Protective Perturbations Must Survive the Resize: Scale-Robust Image Immunization against Malicious Editing
https://arxiv.org/abs/2610.07464

>Anchored or Drifting: What Recursive Self-Generation Reveals About Training Data
https://arxiv.org/abs/2606.31991
>>
File: Krea2_turbo_02001_.jpg (1.42 MB, 1776x2368)
1.42 MB JPG
>>110007984
>>
>>110008001
kek
>>
Would you like to see my gens?
>>
File: Test 00003(2).mp4 (3.92 MB, 832x1248)
3.92 MB
3.92 MB MP4
Finally got something decent out of YuE2:
https://litter.catbox.moe/k5mbjv.mp4
>>
>>Harmful Content Generation in T2I Models: Capabilities and Moderation Limitations
>https://arxiv.org/abs/2610.06503

>fearmongering over Flux 1 Dev in October of 2026
loooooooool

>Generation quality is largely preserved under harmful prompting, producing imagery of sufficient fidelity to pose risks for disinformation and abuse; FLUX.1-dev produces clearly realistic harmful images
imagine not knowing how much better models have gotten at producing "harmful imagery" since august of 2024 when Flux 1 released
>>
>>110007984
>>110007993
Fuck off lonely loser
>>
>>110008070
It reminds me of nausicaa or escaflowne.
>>
>>110008070
>uses nihonese to hide the absolutely dreadful and non rhyming lyrics that are mostly off beat
ai slop coworker music
>>
File: output-2min.mp4 (3.85 MB, 424x240)
3.85 MB
3.85 MB MP4
I have been experimenting with seamless perpetual video generation. 5-sec mmh3 clips stitched with ~1-sec overlap using custom python for latent pinning and guides. Took about 2hrs to gen 3 minutes worth @960x544. Prompt was "Team Fortress 2 gameplay video."
>Welcome to Slop Fortress 2
>>
does h3 support .gif as an input format for images?
>>
File: 095066.jpg (563 KB, 1440x1936)
563 KB JPG
>>110007854
I can fix her
>>
>>110007854
What good is AI for if I can't look like that
>>
>>110008407
soon :tm: trust the plan
>>
how do you deblur fast motion scene in h3?
>>
>>110008451
there's a lora for that
>>
cozy breas
>>
>>110008380
you'll probably need to convert to 24-fps mp4 with 17n+5 frames total (22,39,56,...)
>>
File: 1784336449801007.png (1.36 MB, 1632x1248)
1.36 MB PNG
>>
after like 6000 h3 gens, im all gooned out and dont want to open comfy again for years
>>
>>110008625
see you tomorrow, anon
>>
Are h3chads using Euler/beta or Euler/simple
>>
>>110008647
res_multistep/simple
been meaning to try seeds_2
>>
File: bis_00005_.jpg (1.25 MB, 1512x2048)
1.25 MB JPG
>>
>new kroma is massively better
Finetuning it to be a turd huh? Im so glad to see another 'beautiful woman' slop funetune that fully deletes one concept from it's training data for another, instead of adding only and never subtracting, therefore the new kroma is a turd
>>
>>110008400
>bulge: 404
grim
>>
kant wait til kroma becomes the meta and i kan pretend like ive been for it this whole time
>>
File: attention?.jpg (217 KB, 2236x942)
217 KB JPG
which attention node do I use for int8 h3?
>>
>>110008708
comfy kitchen is faster in most hardware. you can't use it with mem eff whatever node though
>>
>>110008708
for me its cumfy kitchen, i place my full trust in a man who started off as an anon and then went on to make a million dollars
>>
File: bis_00083_.jpg (1.45 MB, 1816x2456)
1.45 MB JPG
>>110008703
It sometimes does similar noise pattern Chroma does, very apparent when doing i2i. Could be user error tho
>>
>>110008608
Good
>>
>>110008723
>>110008727
the comfyui guide says
>Comfy Kitchen attention does not support these checkpoints and sampling crashes with an alignment error
https://docs.comfy.org/tutorials/video/minimax/minimax-h3#quality-degradation-with-int8-attention
it doesnt look like kitchen should be used with int8 based on this. is comfy lying?
>>
>>110008762
cumfart likes to fuck with us yeah
>>
>>110008682
cuntboy supremacy
>>
File: 1782936834900618.png (5 KB, 299x121)
5 KB PNG
>>110008762
>is comfy lying?
yea
>>
Why did he call it "Comfy Kitchen" instead of something like "Comfys big cum cums in japan"

How did they come up with the name?
>>
File: file.png (2.42 MB, 1088x1408)
2.42 MB PNG
yeah kroma is a huge step up actually what the hell
>>
>>110008800
i just woke up, is this supposed to be kroma-v0.3.1-turbo-opd?
>>
>>110008830
ye
>>
>>110008830
nta but that's what i'm using
>4.71 seconds
i'm going to take a little break from h3 to explore this
>>
>>110007984
>Run HunyuanImage 3.0 (80B) natively in ComfyUI on a single 12–24 GB GPU

Interesting, someone still cared about that useless gigantic model? What about the new video model btw?
>>
But it still has diapers I hope ?
>>
>What OPD is: v0.3.1 was distilled on-policy. Instead of the classic offline recipe — imitating the teacher on a fixed sampling schedule, which slowly pulls the student off the original data manifold — the student generates its own trajectories and the teacher corrects it on those exact points. Training only ever happens on states the model actually visits, so the distilled model stays inside the original model's distribution: Turbo speed without the usual distillation tax (mode collapse, washed-out detail, prompts that suddenly stop working).
bold claim
>>
>>110008899
sounds like a slower training method
>>
>>110008731
>>110008800
How do gens using reference images look? Better or worse?
>>
>>110008070
>YuE2
DOA model with no way to train it and sound quality worse than ACEStep. Thing is, we were promised the tokenizer as soon as it dropped. Chinese culture...
>>
File: 852156.webm (3.38 MB, 576x320)
3.38 MB
3.38 MB WEBM
the aryan never backs away from the face of danger
>>
File: 572261.webm (2.26 MB, 576x320)
2.26 MB
2.26 MB WEBM
>>
File: 1771464048432732.gif (359 KB, 220x220)
359 KB GIF
>>110008070
dude, you got me motivated to try yue
>>
>>
>>110009037
>>110009067
What are you using for these anon?
>>
>>
File: 1788754993185791.png (975 KB, 1504x1008)
975 KB PNG
kroma
>>
File: 539854.webm (3.95 MB, 576x320)
3.95 MB
3.95 MB WEBM
>>110009106
ltx 2.5
>>
>>110008608
>>110009037
>>110009087
Best posters in this general
>>
>>110009139
kino delivery
>>
>>110008782
can you link this issue? I cant find it. theres lots of int8 stuff
>>
File: 1790055568107374.png (2.44 MB, 1888x1056)
2.44 MB PNG
>>
File: 1790376545750564.jpg (93 KB, 450x494)
93 KB JPG
yue is uncensored and based. holy...
>>
File: 1767263204166253.png (1.99 MB, 1152x1728)
1.99 MB PNG
>>
>>110009169
nm found it
>>
File: file.mp4 (241 KB, 608x512)
241 KB
241 KB MP4
>>
>>110007854
I'm glad that you got rid of that collague. Would be awful if your shitty gens weren't the only ones in these threads.
>>
File: IMG_4092.jpg (535 KB, 869x1052)
535 KB JPG
>>110007854
how do i animate bouncy milkers like this?
https://x.com/courage_saipa/status/2107796781585318371?s=61
>>
Test
>>
File: 1762421116025750.png (3.67 MB, 1728x1152)
3.67 MB PNG
>>
>>110009198
?
>>
File: ComfyUI_Kroma_00029_.png (2.01 MB, 1152x1152)
2.01 MB PNG
This Kroma model is much better desu, though it remains super slopped at high res. It's still way behind Chroma-Krea, but at least it is fast and can do full NSFW prompt understanding. It's almost caught up to Chroma-Krea at lower res (1152x1152, 1024x1024). Maybe it could be used to refine Chroma-Krea NSFW gens as a separate inpainting task.
>>
File: ComfyUI_Kroma_00030_.png (1.89 MB, 1152x1152)
1.89 MB PNG
>>
File: ComfyUI_Kroma_00032_.png (1.59 MB, 1152x1152)
1.59 MB PNG
>>
File: 1773679497152717.png (1.29 MB, 1152x1728)
1.29 MB PNG
>>
File: 1777072656157922.png (1.02 MB, 1024x1024)
1.02 MB PNG
>>
>>110009281
subscribe to my patreon first and ill tell you
>>
>>110009302
>though it remains super slopped at high res
what res is the cutoff?
>>
File: 108865990018084.png (3.37 MB, 1280x1856)
3.37 MB PNG
>>
>>110009212
y u repost
>>
File: 5793326.webm (3.59 MB, 576x320)
3.59 MB
3.59 MB WEBM
>>
File: 1780964422430038.png (2.76 MB, 1408x1408)
2.76 MB PNG
>>110009347
thats a kroma remake with the same prompt*. the original was qwen.

*with the small addition of "The two lumps from the testes are viable almost separating the scrotum into two connected circles."
>>
File: 1777106545790701.png (1.21 MB, 1728x1152)
1.21 MB PNG
>The right side is a large window whose glass shows the cafe sign painted backward from the inside, outlined bubble letters filled with dither reading "BURGER CAFE" in mirror image, with a half sun of rays in the top right corner.
nice
>>
>>110009360
elder scrolls 6 looks fucking kino
>>
>>
>>110009380
major consistency error in this gen. the location reads "burger cafe" but the dialog is stating they don't sell hamburgers. this appears to be a hallucination error by the AI that produced a logical contraction in the scene
>>
>>110009361
yeah i guess it does look a lot more like balls now
>>
>>
File: ComfyUI_Kroma_00054_.jpg (3.01 MB, 2048x2048)
3.01 MB JPG
Hm, maybe it is getting there.
The main I gen at is 2048x2048. I choose this res for Krea 2 too, on Kroma I had to dial it back to 1152x1152 to get consistently non-slopped results that hold up to my Chroma-Krea gens. I'll test img2img and see if it holds up.

Interestingly, 1152x1152 seems to be the only square resolution from a particular prompt that is non-slopped, 1024x1024 was also slopped from my quick test.
>>
File: 454453121155.jpg (3.85 MB, 7860x2620)
3.85 MB JPG
>>110009406
Meant to quote
>>110009331

For this prompt, 1152x1152 is consistently non-slopped on Kroma, others also perform slightly better
>>
>>110009388
>we got Elder Scrolls VI before GTA VI
fuck
>>
File: ComfyUI_Kroma_00062_.png (1.99 MB, 1152x1152)
1.99 MB PNG
>>
>>
File: ComfyUI_Kroma_00053_.png (1.73 MB, 1152x1152)
1.73 MB PNG
Maybe this smaller res output could be enlarged and fits well into Krea 2 at 2048x2048 on a second pass after re-scaling.
>>
File: 416448582475662.png (1.57 MB, 896x1152)
1.57 MB PNG
Hmm, yeah this Kroma is not too shabby.
>>
>>110009490
he discovered the beer duplication glitch
>>
File: 1006262788350920.png (3.2 MB, 1344x1728)
3.2 MB PNG
>>110009517
Damn, imagine if you had to get kicked in the dick to dupe your beer...
>>
>>110009527
i'd build a self dick kicking machine
>>
>>110009490
was this posted to trick people into thinking h3 is bad?
>>
>>110009537
a motorized version of this
>>
>>110009323
What pose were you attempting to achieve?
>>
>>110009406
>>110009457
Similar to native Chroma then I guess.
>>
>>110009540
this is just your average h3 gen
>>
File: 85647.webm (3 MB, 576x320)
3 MB
3 MB WEBM
OH N-
>>
File: file.png (799 KB, 1585x1434)
799 KB PNG
compositing confuses me
>>
File: MiniMax_H3__00998.mp4 (2.46 MB, 640x960)
2.46 MB
2.46 MB MP4
>>
File: O_O.jpg (13 KB, 200x240)
13 KB JPG
>>110009590
>>
>>
Don't hype Kroma up too much or he'll stop working on it
>>
>>110008800
I was going to say that the generated face looks ugly af but to be fair she's that ugly irl, once you don't get distracted by her cleavage you can easily notice her uglyness, it's insane.
>>
File: ComfyUI_Kroma_00066_.png (1.68 MB, 1152x1152)
1.68 MB PNG
>>110009556
It's beautiful. The future is bright
>>
>>110009655
>she's that ugly irl
anime is not real life
>>
>>110009540
kek
>>110009670
retard
>>
File: 7426674.webm (3.73 MB, 448x448)
3.73 MB
3.73 MB WEBM
>>
File: ComfyUI_Kroma_00069_.png (1.98 MB, 1152x1152)
1.98 MB PNG
>>
0/10 elbow too sharp

>>110009665
yeah I hope they keep training
>>
File: ComfyUI_Kroma_00072_.png (1.11 MB, 1152x1152)
1.11 MB PNG
>>
>Take a break from this general and genning in general for for a few months
>Check over previous generals to get caught up
>Gays are freely posing itt now
If only you knew how bad things really are
>>
File: ComfyUI_Kroma_00070_.png (1.13 MB, 1152x1152)
1.13 MB PNG
>>
File: AAAAAAAAAAAAAAAAAA.png (485 KB, 402x780)
485 KB PNG
>>110009703
>>
File: 1774620527033968.png (3.15 MB, 1152x1728)
3.15 MB PNG
accidental gen
>>
File: 1782523608983967.png (3.57 MB, 1152x1728)
3.57 MB PNG
>>
not sure what to think about the new krea, since is a turbo model, the composition tends to be the same no matter the seed, but a new dataset and captions are really good for degenerate purposes
>>
>>110009772
new krea?
>>
>>110009718
workflow im finna goon
>>
File: ComfyUI_temp_omiuv_00013_.png (3.32 MB, 1280x1840)
3.32 MB PNG
>>110009784
sorry, I meant kroma
>>
File: file.png (2.27 MB, 1152x1152)
2.27 MB PNG
>>110009772
you can still add krea2 loras on it
>>
File: ComfyUI_temp_omiuv_00014_.png (3.14 MB, 1280x1840)
3.14 MB PNG
>>110009797
yeah I noticed but repeated images are big turn-off for me
>>
>>110009796
I'm pretty sure that came out with a base model variant too though, might be cfg distilled still tho idk
>>
File: ComfyUI_Kroma_00074_.png (1.53 MB, 1152x1152)
1.53 MB PNG
>>110009494
Too much leftover noise if I do a 2 pass, what it can currently do will have to do for now.
>>
File: mpv-shot0252.jpg (1.02 MB, 2048x2048)
1.02 MB JPG
>>110009457
Interesting.

this one was done with Krea2 + SummerVibes lora.

posting for research purposes
>>
File: ComfyUI_temp_omiuv_00018_.png (3.09 MB, 1280x1840)
3.09 MB PNG
>>
>>110009772
>but a new dataset and captions
Is this information anywhere outside his groomcord?
>>
File: ComfyUI_temp_omiuv_00025_.png (3.21 MB, 1280x1840)
3.21 MB PNG
>>
I'm annoyed that H3 keeps using weak references as single frame, and without that ref it doesn't understand what I want despite prompting autistic details.
>>
>H3
>set style as "anime" in prompt
>still gens semirealistic
>>
>>110009838
Do normnigs still find duckface attractive? I thought the whole ass botox trend has passed.
>>
>>110009694
This is pretty cool. Needs more compression and CNN will probably buy these.
>>
File: debo_asw_k2_00099_.png (3.14 MB, 1792x896)
3.14 MB PNG
>>110009754
>>110009766
cool
>>
>>110009847
xoomers still like it
>>
File: ComfyUI_temp_omiuv_00028_.png (2.98 MB, 1280x1840)
2.98 MB PNG
>>
>>110009858
nice gen
>>
>>110009846
>H3
>prompt japanese cel shaded anime style in 'anime style'
>it does it
skull issue?
>>
>>110009824
>>110009838
>>110009882
Basado, I also do this. Have hundreds of loras for random photogenic models.
>>
i dont know if its the best at 1152x1152, i find it renders detail best at 3mp
>>
File: ComfyUI_temp_omiuv_00034_.png (3.01 MB, 1760x1360)
3.01 MB PNG
>>
File: 1767807957481815.png (2.51 MB, 1152x1728)
2.51 MB PNG
>>
File: ComfyUI_temp_omiuv_00036_.png (3.34 MB, 1760x1360)
3.34 MB PNG
>>
File: ComfyUI_temp_omiuv_00037_.png (3.3 MB, 1760x1360)
3.3 MB PNG
>>
File: 1783577121279548.png (2.03 MB, 1152x1728)
2.03 MB PNG
>>
File: ComfyUI_temp_omiuv_00039_.png (3.11 MB, 1760x1360)
3.11 MB PNG
Can you imagine we had the internet of the 2010s but with AI content, I hate that you can't post anything fun on social media anymore
>>
File: ComfyUI_temp_omiuv_00041_.png (3.46 MB, 1120x2080)
3.46 MB PNG
>>
>>110009965
desu girls that age are smoking carts not bongs unc
>>110009976
kino
>>
>>110009976
Hot for a granny
>>
File: ComfyUI_temp_omiuv_00043_.png (2.93 MB, 1920x1200)
2.93 MB PNG
>>110009981
why is that with zoomers they think its always about them? lel
>>
File: ComfyUI_Kroma_00077_.jpg (3.75 MB, 2048x2048)
3.75 MB JPG
>>110009915
>3MP
Got poor results out of that since it takes too long.

Results vary depending on the prompt. When 2MP is not slopped then it wins. When it's slopped , then it loses. As the model progresses, I expect it to always be better at 2MP.
>>
File: ComfyUI_temp_omiuv_00045_.png (2.8 MB, 1920x1200)
2.8 MB PNG
>>
>>110010003
>posts zoomers
>ERM WHY DO LE HECKIN ZOOMERS SAY THOSE ARE LE ZOOMERS?
kill soilennials
>>
>>110010003
before you reply >>110010024 is not me but he is correct
>>
>>110010003
>>110010021
bro what is this trash?
>>
File: ComfyUI_temp_omiuv_00046_.png (3.5 MB, 1120x2080)
3.5 MB PNG
>>110010024
I was just recreating something I lived when I was younger, I don't even smoke weed anymore, not everything revolves about your generation

social media got you paranoid lol
>>
File: 1791316287466842.png (2.26 MB, 1152x1728)
2.26 MB PNG
>>
>>110010038
i failed at getting a decent looking teenage jencon. she kept coming out post wall
>>
>>
File: 1780301661962962.png (2.2 MB, 1152x1728)
2.2 MB PNG
>>
File: ComfyUI_temp_omiuv_00049_.png (3.09 MB, 1120x2080)
3.09 MB PNG
>>
File: ComfyUI_temp_omiuv_00051_.png (3.22 MB, 1120x2080)
3.22 MB PNG
>>
>>110010062
>>110010064
Got company gpu for weekend, new version incoming
>>
>you in the back
>>
thanks I lost these when my laptop broke.
>>
File: file.png (474 KB, 405x720)
474 KB PNG
>kroma can now do reliable length marking tattoos
>>
File: ComfyUI_temp_omiuv_00060_.png (3.12 MB, 1120x2080)
3.12 MB PNG
>>110010073
nice, I love your loras
>>
>>
>>110010089
Ideas/requests? Not gonna run hc porn, softcore is alright
>>
>>
File: ComfyUI_temp_omiuv_00065_.png (3.19 MB, 1120x2080)
3.19 MB PNG
>>110010105
Train a young gillian anderson lora
>>
File: ComfyUI_temp_omiuv_00066_.png (3.35 MB, 1120x2080)
3.35 MB PNG
>>110010105
If you have the time and want to, you could retrain the Inde Navarrete lora, I think the dataset has too many images with depth of field/bokeh effect but I understand that since is based off the movie
>>
is that the obsession girl?
>>
>>
>>110010124
>>110010140
Inde lora update is coming, can't promise Gillian. I have X-files bluray so it wouldn't be too much effort
>>
File: ComfyUI_temp_omiuv_00073_.png (3.02 MB, 1120x2080)
3.02 MB PNG
>>
>>110007854
why is this Krea gen not hot poop? lora?
>>
>>110010148
https://huggingface.co/xixxix-HF/NikkiObsession
>>
File: 85328.webm (3.99 MB, 448x448)
3.99 MB
3.99 MB WEBM
why did he do it?
>>
File: 350788973479309868497.jpg (753 KB, 1344x2016)
753 KB JPG
>>110010170
they are good loras. that videodrome lora is pure kino.
>>
File: Shake.mp4 (2.49 MB, 1040x1040)
2.49 MB
2.49 MB MP4
>>110009809
Don't sleep on L2VA boys
>>
File: 1762254803631140.mp4 (677 KB, 544x464)
677 KB
677 KB MP4
What are the current best models for
>video
I'm assuming h3 is still the best?
>anime
>realism
>image editing
>>
>>110010293
the final redpill is injecting the frame into the center of the video
>>
>>110010306
>video
Flux 3, but they will never drop the weights
>>
>>110010329
Flux 3 is proprietary and not even a good propriety model

The 'dev' version of that they said they would release is obviously going to be worse

H3 is better so there is no point, and let's face it, BFL will never release a video model that can gain any traction since the censor their shit to bits and it's aimed at sloppy advertising style clips

Perhaps LTX will release a new model that is competitive with H3, else we need to hope that there's a new Minimax or other chinese model to push the local boundary
>>
>>110010306
ani said that h3 is the best by a large margin and nothing's even close
>>
File: 1779792498979.webm (2.8 MB, 720x1280)
2.8 MB
2.8 MB WEBM
>>110010424
can h3 do this?
>>
>>110010432
I bet my dreams would look something like this if you could record them and view awake.
>>
>>110010432
That's actually awesome, best AI slop I've seen in a long time, my sides
>>
>100 turbos later, my old one still won by landside
>>
>>110010432
I remember these!
>>
post malfoid
>>
File: 1347.png (2.7 MB, 1248x1664)
2.7 MB PNG
>>
Kroma still refuses to do proper facial emotions though.
>>
File: 454564515151.png (327 KB, 890x813)
327 KB PNG
>>110010389
>Flux 3 is proprietary and not even a good propriety model

It's the strongest world model on a physis-IQ benchmark, I wouldn't be so quick to praise Seedance anon. I got the same idea too from seeing Flux 3 teasers, they truly have a monstrous model. You may be right about them never giving us a Dev model on par with their API, but don't lie about the capabilities of the model, it's SOTA, and the closest thing to Sora 2. Obviously, if Sora 2 were still around it'd be right there next to Flux 3 or even at the top.
>>
>>110010599
why is everyone using h3 instead of cosmos?
>>
spoonfeed a new nigga. can this kroma do image references? like for person identity, clothes, etc.?
>>
Anons...
The amount of LORAs for Anima seems a bit low-ish? Maybe?
- Is it because it's still fairly new and I am an impatient faggot?
- Some other reason?
>>
>>110010657
There is a weird push to hate on it for whatever reason
>>
>>110010668
Weird.
What are the haters promoting instead?
Staying with IL?
>>
>>110010611
h3? h3? data please
>>
>>110010678
In my experience they aren't promoting anything, they just shit on existing things without any alternative offered or even any images attached. It's alleged that the model is bad at learning but I have never ever seen any examples demonstrating this phenomenon. I would recommend you do what I do and completely disregard any statements that don't have any evidence (gens with metadata) attached or at least provided upon request.
>>
I have no idea how to make h3 clips not to look dogshit unless i feed it a good reference image
>>
>>110010692
The problem is that Anima loras aren't flexible. You won't be able to tell an Anima lora is bad just with images.
>>
>>110010714
You need to go high resolution, but even so, 99% of people use reference images anyway, t2v is always going to be a subpar workflow
>>
>>110010723
Then I guess several images + prompts would do? Or elaborate on what exactly do you mean by flexible.
>>
>>110010723
isn't that a good thing? Less defects and all
>>
>>110010657
Most people are either staying with SDXL derivatives or moving to Krea 2, Anima is in this limbo inbetween space, it's not bad, just not good enough to warrant a switch
>>
>>110010714
are you trying to do text only?
>>
>>110010732
By flexible, I mean being able generate images that aren't the same as the images in the dataset while maintaining good likeness. Shouldn't that be obvious?

>>110010733
It's a good thing if you want only slight variations of the training images. If you try to gen something outside of the dataset, you either won't be able to or you get body horror.
>>
>>110010738
Unfortunate.
Maybe I missed some IL models upgrades, but it seems to me Anima is MUCH better with anatomy.
>>
>>110010755
>By flexible, I mean being able generate images that aren't the same as the images in the dataset while maintaining good likeness. Shouldn't that be obvious?
This should be incredibly simple to demonstrate with as few as 2 images so I'll go ahead and follow my own advice of completely disregarding this statement.
>>
>>110010678
>>110010692
the best alternative right now is unfortunately a saas model
>>
>>110010657
I haven't bothered uploading my stuff

>>110010723
>The problem is that Anima loras aren't flexible
Too easy to overfit.
>>
>>110010770
If you're going to prove that lora training works well with Anima, you're the one that needs to provide proof, not the other way around. I've never seen anyone posting any working examples.
>>
>>110010743
No, just experimenting shit
rn i can only use a reference image; ideally, it defines the whole Aesthetic and tells h3 to copy.
There is no other consistent way to make h3 look good
>>
>>110010795
I haven't done any training on anima, I am simply warning the anon that asked that there are constantly posts like yours that say some shit and don't provide anything to support it
I'm not saying you're wrong, I'm just saying your post should be disregarded
>>
>>110010807
that's why i don't use h3. it is not very good at creating the correct style unless you spoonfeed it
>>
File: doom_g.webm (3.28 MB, 1920x1080)
3.28 MB
3.28 MB WEBM
>>
File: gigafly_fly1.webm (1.06 MB, 854x480)
1.06 MB
1.06 MB WEBM
>>
File: gigafly_fly2.webm (637 KB, 854x480)
637 KB
637 KB WEBM
>>110011086
>>
File: doom_piano.webm (3.2 MB, 1920x1080)
3.2 MB
3.2 MB WEBM
>>110011069
>>
File: dsa.png (262 KB, 1879x883)
262 KB PNG
>>110011123
loving the new manual sigmas, but still at 1080p video getting some flickers? why does minimax h3 does that? i am also using bf16
>>
File: mpv-shot0251.jpg (367 KB, 1152x1536)
367 KB JPG
>>
File: Oopsie.mp4 (3.34 MB, 800x800)
3.34 MB
3.34 MB MP4
>>110010311
Don't sleep on V2V extension boys
>>
File: doom_kid.webm (3.3 MB, 1920x1080)
3.3 MB
3.3 MB WEBM
>>
>>110011131
how long to gen 1080p?
>>
>>110011282
I'm using RTX 5090 on Vast ai
>>
File: mpv-shot0252.jpg (802 KB, 1152x2048)
802 KB JPG
>>
>>110011131
>new manual sigmas
Please explain to me like I'm retarded, because I kinda am
>>
File: mpv-shot0251.jpg (628 KB, 1152x2048)
628 KB JPG
what am I doing this late

gn
>>
File: gay.png (68 KB, 1300x420)
68 KB PNG
>>110011335
it's a node that you plug into the "sigmas" input of SamplerCustomAdvanced instead of using a normal scheduler node. sigmas are basically the noise levels for each step, 1.0 is pure noise and 0.0 is the finished clean video. the sampler goes through the list one by one, so 9 numbers = 8 steps. a normal scheduler makes this list for you, ManualSigmas just lets you type it yourself. the ones i posted stay close to 1.0 for the first few steps, which is where the model decides the motion and composition, then drop fast at the end for the details. so you get good motion in only 8 steps. just copy the numbers into the node, connect it, and you don't need a scheduler node.
>>
>>110011352
Goodnight fellow coomer
>>
>>110011334
>>110011352
dis new kroma?
>>
>>110011361
based
>>
>>110011370
Krea 2
>>
File: r34d.png (1.66 MB, 1456x816)
1.66 MB PNG
>>110011361
>>110011373
make sure you are using the hyperflow lora it's really good compared to light2x

https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/resolve/main/minimax_h3_hyperflow_8step_v1.0_comfyui_bf16.safetensors
>>
File: as3432.png (1.66 MB, 928x1232)
1.66 MB PNG
Krea 2 has great brush stroke effects unlike other models
>>
>>110011361
Thank you my good man
>>
>>110011295
you gen native 1080p or upscale?
>>
>>110011394
but is it better than ema600?
>>
File: cillian_doom_2.png (1.34 MB, 928x1232)
1.34 MB PNG
>>110011435
native, i set the latent to 1920x1088 straight in the workflow, no upscale. takes longer on the 5090 but looks way cleaner than upscaling

>>110011440
haven't tried ema600 yet so can't say, hyperflow just gave me better motion and less blur than light2x at 8 steps. if you've used ema600 post a comparison
>>
>>110011488
I was in a coma for couple of days. Is this a new model or something? Would love to see better art mediums and styles out of the box. Krea2 is too slopped in many ways.
>>
File: 341sdas.webm (3.35 MB, 1920x1080)
3.35 MB
3.35 MB WEBM
Still can't fix the faces distorting / blurring / losing details can anyone suggest a fix
>>
File: Shaken.mp4 (3.59 MB, 576x1040)
3.59 MB
3.59 MB MP4
>>110011352
Trying to tard wrangle the H3 R2V workflow to accept a still image as a midpoint. It's bretty gud but not a perfect match as the cup stays in shot.
Also, some of you guys are alright. Don't post any bitches here today or they might get shaken.
>>
File: ComfyUI_08154_.png (1.65 MB, 1920x1080)
1.65 MB PNG
>>
File: output.jpg (1.03 MB, 2720x1536)
1.03 MB JPG
>>
>>110011352
Lovely
>>
>>110011639
face detailer?
But it's slow
>>
Kill ani
>>
>>110011639
Only thing I've seen sorta preserve heads/faced is generating 2 second clips and combining, but it might have been done with genning longer sequences and taking frames for use with flf on those 2 second versions
>>
i don't get how ComfyUI works
Sometimes it stores blocks in RAM
Sometimes you just OOM. Do I have to whip it ?
>>
File: 1790205569520744.jpg (2.2 MB, 3552x4736)
2.2 MB JPG
hello
>>
File: MiniMax_H3_01044_.webm (1.32 MB, 1920x1088)
1.32 MB
1.32 MB WEBM
>>
>>110012021
hey imposter-anon
>>
File: lolo.jpg (9 KB, 250x243)
9 KB JPG
>>110012021
>>110012006
>>110011998
>>110011935
interesting. always the same "different" people arriving at the same time every day.
>>
>>110012036
Weirdly enough exactly when "the anons" all arrive together in /sdg/ when it hits page 10 again
>>
File: 1788102609001.jpg (1.37 MB, 1776x2368)
1.37 MB JPG
>>110012036
I just came home from work time to make fun of ani and debo
>>
https://huggingface.co/DeepBeepMeep/Kandinsky6/tree/main

deepy is cooking it up.
>>
>wagie spends all his free time seething
grim.
>>
File: ComfyUI_00017_.mp4 (3.19 MB, 1184x896)
3.19 MB
3.19 MB MP4
>>110012006
yea
>>
>>110012178
ugly whore
>>
>>110012064
i wonder if it can go past the official 5 seconds
>>
>>110012178
chat is this real?
>>
File: MiniMax_H3_01246_final.mp4 (3.61 MB, 800x1056)
3.61 MB
3.61 MB MP4
They want to be me so bad
While this is a turbo gen moving forward I will only use comfy kitchen cope node
>>
>>110012244
How do you find Minimax? I haven't bothered with animation at all because I think it's not ready. It's ready for joke gens of course.
>>
>>110012255
Best we got desu
>>
>>110012178
cutie
>>
>>110012244
who are you?
>>
>>110012259
Interface like Adobe After Effects/Premier would be cool - you'd have a timeline which can be edited and segmented as usual.
But every segment is a prompt.
No need for this LLM dilly dallying "at this time there is a cut" it would automatically create the prompts.
>>
>>110012282
true...
>>
>>110012357
Sure some jeet will use Claude to program a python abomination for this. And these twitter advertisers are stealing opinions and ideas from these threads as we speak.
I'm talking about something simple which works like industry standard software. That's still outside the scope of a common twitter influencer without industry experience.
>>
File: 00031-944349932.jpg (358 KB, 2880x1536)
358 KB JPG
honestly not seeing the hype for the new kroma model. maybe its a lora compatibility issue.
>>
File: 00036-2144609000.jpg (317 KB, 2880x1536)
317 KB JPG
>>
>>110012391
no f*cking way this is ai
>>
>>110012413
Plastic 3d is actually surprising good, it was like that even with Pony.
It's an artifact.
>>
>>110012458
Artifact of a greater understanding.
>>
>>110012458
this is just beyond anything i've seen
anon is a master of prompting, the quality is absolutely heckin insane man
and the creativity is on another level
>>
>>110012474
I know you are trolling. just do your own images.
>>
>>110012492
>>110012492
>>110012492
When ready
>>
>early shitbakes with troll OP pics
not this shit again man
>>
File: 00063-3787684092.jpg (727 KB, 2880x1920)
727 KB JPG
>>110012474
its just a particular aesthetic i like.
>>
>>110012518
just bake when bump limit reaches
>>
>>110012518
If you make a proper bake I will try to delete the thread, we have been having constant troll spam for weeks now
>>
File: 85641037.png (3.81 MB, 1664x1664)
3.81 MB PNG
gm
>>
>>110012548
Will bake at bump limit
>>
>>110012548
why? you didn't make a proper bake huh?
>>
>>110012564
I don't do colleges and never have and never will. Everything else is fine
>>
File: ComfyUI_05761_.jpg (160 KB, 1248x832)
160 KB JPG
>>110012391
or maybe it's just shit like any other lodestone model
https://www.youtube.com/watch?v=WeYsTmIzjkw
>>
He won't ever make a good model, while the model has improved it's forgetting basic shit
>>
it's up

>>110012667
>>110012667
>>110012667
>>
>>110012573
That's because you are a narcissist and have taken over control of baking.
You have alienated any creative people from these threads.
>>
>>110012672
>>110012675
We're not playing this game again today, fuck off
>>
>>110012682
Who is we?
>>
where did collage anon go anyways?
>>
>>110012716
Not everyone has time to watch this nutjob that keeps trying to remove links, he still post but he's obviously busy.
>>
>>110012722
Yet again "We" fallacy.
>>
NON TROLL THREAD HERE


>>110012735


>>110012735


>>110012735


>>110012735
>>
Every single time this "anon" arrives, this shit happens.
>>
File: gemma1.jpg (833 KB, 1800x1200)
833 KB JPG
This is Gemmers 4 idea of Clint's film. Very sloppy.
>>
File: gemma2.jpg (990 KB, 1800x1200)
990 KB JPG
Driven through a lora.
>>
Why are the rentry schizos seething again?
>>
So he has a autistic rage fit by creating 2 new threads and prays the mod deletes the first even though the mod is already aware of his antics?
>>
File: new.png (2.89 MB, 1200x1800)
2.89 MB PNG
>>
>>110012846
not ai
>>
>>110012846
I didn't bother to edit the signature away.
>>
File: Crusade_.jpg (886 KB, 1200x1800)
886 KB JPG
>>
goooooooooo
>>110012831
>>110012831
>>110012831
>>110012831
>>
File: ta32.jpg (78 KB, 885x498)
78 KB JPG
>>110012831
>>110012735
>>110012667
>>110012492
which one do i use?
>>
File: old2.png (2.95 MB, 1200x1800)
2.95 MB PNG
>>110012869
>>
>>110012881
The first one not made by a seething schizophrenic that really wants to remove links warning new users
>>
>>110012894
>The first one
the pepe one?
>>
>>110012899
Yeah
>>
File: 00120-3580860222.jpg (542 KB, 2816x1920)
542 KB JPG
fucking idiots with the constant multiple split thread bakes.
>>
>>110012983
We in here
https://boards.4chan.org/g/thread/110012667/ldg-local-diffusion-general#bottom
>>
So why did this guy bake three troll threads after a real one was baked?
>>
>>110013013
Nice thanks anon
>>
>>110009655
I'd still fuck the shit out of her and so would you you fucking incel



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.