[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Career Enders Edition

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109640448

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Blessed thread of frenship
>>
>wondering why my gens are in the fagollage
>realize i didnt post any in previous
PHEWWWW!
>>
File: Krea2_turbo_01832_.jpg (1.83 MB, 2368x1776)
1.83 MB JPG
>>
File: 00001-2808932146.jpg (355 KB, 2880x1728)
355 KB JPG
>>109644873
it depends of the krea 2 checkpoint. Avoid using the overbaked uncensored hardcore porn finetines. Sticks to the stable ones like krea2centerraw and krea2centerv30int8. I'm very disappointed the quality of moodymix and pornmasternoob krea2 checkpoints. A lot of the bad "ultra photorealistic uncensored" advertised krea2 models on civitai produce bad fried and heavily flawed dirty skin with cum, wet water droplets, sweat droplets and blush red face checks.
>>
>mfw Resource news

08/25/2026

>Anima Turbo v1.1 released
https://huggingface.co/circlestone-labs/Anima

>Training-Free Pseudo-Fusion for Composed Image Retrieval with Diffusion Models and Multimodal Large Language Models
https://github.com/StevenXuf/PeFuse4CIR

>ReART: Reference-Guided Retrieval and Refinement for Emotion-Aware Art Generation
https://github.com/oceanflowlab/ReART.git

>Loopy: Seamless Video Loop Generation via Anchored Looping Shift of Positional Embedding
https://donghaotian123.github.io/Loopy

>Long-Horizon Audio-Visual Generation for Persistent Stories and Interactive Worlds
https://echo-team-joy-future-academy-jd.github.io/Echo-1.5-Page

>Generated Reality: Human-centric World Simulation using Interactive Video Generation with Hand and Camera Control
https://codeysun.github.io/generated-reality

08/24/2026

>MiniMax-H3-Fun-Controlnet-Union
https://huggingface.co/alibaba-pai/MiniMax-H3-Fun-Controlnet-Union

>MiniMax-H3-Longvideos: Long (up to ~120s) MiniMax-H3 video + synchronised audio from a single prompt
https://huggingface.co/Smite79/MiniMax-H3-Longvideos

>DiGS-Avatar: Single-Image Animatable 3D Human Reconstruction via UV-Space Diffusion
https://github.com/KLMAV-CUC/DiGS-Avatar

>Identity-Preserving Text-to-Video Generation via Agentic Enhancement and Semantic Repair
https://github.com/oceanflowlab/AESR

>OccluRank: Controllable Occlusion-Aware Layout-to-Image Generation by Adding Just an Ordinal Rank
https://github.com/Wenyang-hong/OccluRank

>Aggregating Visual Information with Optimal Transport for VideoLM Token Compression
https://github.com/ernie-research/AVIOT

>CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxation
https://github.com/CubicSplat/repo

>Vis-Poison: Poisoning Visual Knowledge in Multimodal Retrieval-Augmented Generation
https://github.com/SWUFE-DB-Group/Vis-Poison

>Explainable Deepfake Detection with Feature-robust Augmentation
https://github.com/oceanflowlab/EDD.git
>>
>mfw Research news

08/25/2026

>SketchFlow: Zero-Shot Vector Sketch Generation via GMM Prior Flow in CLIP Latent Space
https://arxiv.org/abs/2608.21659

>GAN-Diff : Coupling Pretrained WGAN-GP Features with Conditional Diffusion U-Nets
https://arxiv.org/abs/2608.22272

>The Plan, Not the Decoder: Diagnosing and Repairing Compositional Failure in Reasoning-Augmented T2I Generation
https://arxiv.org/abs/2608.21713

>Direct, Parallel, or Sequential? A Comparative Study of Training-Free Multi-Subject I2V Generation
https://arxiv.org/abs/2608.22819

>VISTA: Test-Time Compositional Alignment for Visual Autoregressive Generation
https://arxiv.org/abs/2608.22521

>Calibrate What You SHIP: Post-Selection Risk Control for Verifier-Guided T2I Generation
https://arxiv.org/abs/2608.21748

>DefaultShift: Auditing Semantic Default Shift in Accelerated T2I Models
https://arxiv.org/abs/2608.21784

>Following Motion for Sequential Modeling in Video Frame Interpolation
https://arxiv.org/abs/2608.22861

>Can We Perform Online RL for Image Editing without Editing Rewards?
https://arxiv.org/abs/2608.22780

>MDFI: A Multi-Domain Features Integration for Compressed Video Quality Enhancement
https://arxiv.org/abs/2608.21495

>From Dense Prediction to Visual Editing: Structured Supervision for Unified Image and Video Creation
https://arxiv.org/abs/2608.14740

>FIRM-Video: Check Before You Score for Reliable T2V Reward Modeling
https://arxiv.org/abs/2608.21839

>TEE-X: TEE-aware Acceleration Framework for LVMs at the Edge
https://arxiv.org/abs/2608.22716

>Investigating Relational Reasoning in VLMs
https://arxiv.org/abs/2608.23518

>What's the Catch? Evaluating Temporal Consistency in Vision-Language Models
https://arxiv.org/abs/2608.23474

>GuardPaint:SpeculativeSafetyDecodingforText-to-ImageGeneration
https://arxiv.org/abs/2608.21869

>Perturb the Thought, Not the Pixels: Latent-Space Rollout Diversification for Reinforcement Learning of VLMs
https://arxiv.org/abs/2608.21595
>>
File: getAjob.webm (2.11 MB, 1143x2048)
2.11 MB
2.11 MB WEBM
>>109645026
>>109645014
>>
>>109645052
Kek
>>
File: Mexi_Coom.png (1.31 MB, 768x1376)
1.31 MB PNG
>>
anyone else getting noisy textures? i'm getting distortion with 0.7MP @ 40 steps. maybe due to marble tile background. will more steps or higher res eliminate it?
>>
File: 5454875.webm (3.65 MB, 576x320)
3.65 MB
3.65 MB WEBM
can you spot the point where the video gets extended?
>>
>>109645106
Not on first watch. Cute grill.
>>
>>109645106
The point where she picks it up?
>>
>>109645080
if only you had posted your workflow so anon could give you an answer instead of trying to figure it out from your vague question
>>
File: Test 00026(2).mp4 (2.98 MB, 1056x608)
2.98 MB
2.98 MB MP4
>>
Footfags, everyone.
>>
>lolcow baker on week three of her tantrum
It's so fucking over
>>
>>109645012
what about kroma?

>krea2centerv30int8
>krea2centerraw
can you share their link anon?
>>
>>109645106
No, the low bitrate and resolution is masking it perfectly
>>
I have an image that I want to represent my characters appearance in a scene but NOT as a keyframe/first frame/last frame, is it fucking possible?! I already have a reference image (with shots from multiple angles) but also want to indicate some attributes about how the character appears in a scene but H3 reference model keeps treating it as a start frame. Is it possible?! I could use this image as a keyframe I guess but I'd just be randomly putting it somewhere in the middle of the shot and it might look like shit.
>>
>>109645160
just have variety in your references. the character sheet should just be one image
>>
>>109645167
Hmm ok I'll try just adding this image to the reference sheet for this scene, maybe that'll work.
>>
What is the best way to generate character concept art for h3 refs? Anima?
>>
>>109645148
At least his example gens are honest
>>
File: 5454875.webm (3.82 MB, 576x320)
3.82 MB
3.82 MB WEBM
>>109645119
yes, the first clip is 10 seconds long. i think this one is nearly invisible. audio can be extended too https://litter.catbox.moe/vu755k.webm
>>
File: Krea2_turbo_01851_.jpg (1.79 MB, 2368x1776)
1.79 MB JPG
Blessing thread with wards
>>
>>109645149
You are unable to tell when different anons bake?
>>
Can someone help me find a site with all the tags to use for characters? I thought it was pinned, but can't find it. It showed all the tags to use after a character for the correct clothing, hair etc. It's not stablediffusioweb, that kind of sucks. Would really appreciate it.
>>
>>109645235
Im only familiar with the ones that show off various styles and artists. For what tags to use for a given character, I'll just reference their danbooru wiki page desu.
>>
File: Jap_Kit2.png (1.26 MB, 768x1376)
1.26 MB PNG
>>
File: Krea2_turbo_01870_.jpg (1.79 MB, 1776x2368)
1.79 MB JPG
>>
>>109645160
You just have describe the character sown in the reference sheet in your subject definitions and later in the summary, retention analysis and shot descriptions how they should appear.

That's what I did in >>109645126

>subject_definitions:
><Subject 1> is the young woman with short spiky bright-blue hair, undercut sides, freckles, yellow eyes, intricate tribal tattoos covering both arms, wearing a tight blue crop top, frayed short denim shorts, black fingerless gloves, and black combat boots, as shown in <Picture 2>.
>>
File: wido.png (302 KB, 364x482)
302 KB PNG
https://litter.catbox.moe/crt5es.mp4
>>
File: debo_mc_k2_00054_.png (3.27 MB, 1466x1792)
3.27 MB PNG
>>
File: 00226-1697861777.jpg (436 KB, 2880x1856)
436 KB JPG
>>109645157
don't really care about kroma and heard it was undercooked dud from the other anons.
https://civitai.red/models/2730415/krea2center?modelVersionId=3127203
https://civitai.red/models/2734469/krea2centersemiraw?modelVersionId=3075365
some of my gens were paired with the lunafreya lora at a lower strength to get a particular desired result i was itching for. I tend use it to properly transform the 2d style anime characters into hyper real or photorealistic 3d rendered characters.
https://civitai.red/models/2761525/lunafreya-nox-fleuret-final-fantasy-xv-krea-2
also there is some experimental loras i love using like octane render and other atmospheric aesthetic loras.
https://civitai.red/models/1883576/octane-render?modelVersionId=3132463
https://civitai.red/models/2827949/krea2beautiful-lora
https://civitai.red/models/2841013/krea2-fairy-dust
https://civitai.red/models/2819747/ethereal-photography-vintage-dreamscape
https://civitai.red/models/2806305/061-atmospheric-photography
https://civitai.red/models/2760910/color-temperature-slider-krea2?modelVersionId=3107144
https://civitai.red/models/2764196/brightness-slider-krea2?modelVersionId=3111088
>>
>>109645325
Really low quality gen as always
>>
>>109645148
i made an absurd amount of buzz just scripting lora baking with random garbage
i wonder who the fuckers are that use my loras (they really are a net negative regardless of what you want to do kek)
>>
>>109645325
crazy how you really didn't improve at all over the years
there are models which can do text properly you know? and now hush hush back to the slop containment general nigbo
>>
File: debo_mc_k2_00003_.jpg (986 KB, 1862x2048)
986 KB JPG
why is he seething?
>>
>>109644925
How is the image in an i2v H3 workflow referenced correctly? Is it just <Picture 1>?
>>
>>109645360
thanks for the response anon, I'll have some fun playing around with these
>>
>>109645427
There is no reference in i2v, just describe the image and do the action.
>>
>>109645427
Yes, that is correct.
I2VA should always start with:
>For the target video, at 0.00 seconds into the target video, <Picture 1> (from [Shot 1]) is fully referenced.

Says so explicitly in their own prompting guide:
https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md
>>
>>109645442
But can't I just describe a <Subject 1> as shown in <Picture 1> if I want to do additional shots?
>>
>https://huggingface.co/Rkss/Minimax_h3_fl2v_lightx2v_turbo_4to8step_v0.1-v1.0_768p_v4_step600_dareties_fro095_pruned

new turbo 4steps lora
>>
>>109645014
>>109645026
thanks!
>>
>>109645469
>>109645470
never mind, thank
>>
>>109645470
Only if you use the ref2v model.
>>
File: Krea2_turbo_01900_.jpg (1.9 MB, 1776x2368)
1.9 MB JPG
>>
File: 00038-2945554527.jpg (541 KB, 2688x1920)
541 KB JPG
>>109645148
there little to no quality control on the kinds of loras and preview images allowed to posted on that site as long as its not cunny, loli and problematic prompts that get flagged for review. The buzz farming incentive really the killed quality and dedicated effort people put into their loras and models. Very few decent bakers are still in the game.
>>109645433
you welcome, have fun and experimenting with them. Krea2 is a powerful model that many people seem to underestimate.
>>
>>109645473
nice. will check it
>>
>>109645427
<Picture 1> is the first frame of the video at 0.00 seconds and belongs to [Shot 1]
>>
any recommended penis/vagina lora for h3? there are multiple ones in civitai but I'd rather avoid a bad one
>>
>>109645605
Sulphur beta 0.1
>>
>>109645611
that's a full finetune, is it really that much better than a standard lora?
>>
>>109645625
h3 loras are bad, dood
It fucks up the audio
>>
>>109645605
The best one I've used is the one I trained on 500+ hours of my own amateur footage. No I will not share it with you.
>>
>>
>>109645629
even at low weight? it's not like blowjob crunch crunch sound is that good anyway
>>
File: graph.png (1.13 MB, 2761x1738)
1.13 MB PNG
Anyone run into the issue where the hybrid fl/ref h3 model will kind of just ignore audio references at higher resolutions? At .5 or less I can clearly hear that it's using the audio reference, but anything higher and it's just default voices

Examples (same seed and prompts):
.01 mp: https://files.catbox.moe/asfngw.mp4
.02 mp: https://files.catbox.moe/elt1m5.mp4
.05 mp: https://files.catbox.moe/k98a9r.mp4
.06 mp: https://files.catbox.moe/rg7mbu.mp4
1 mp: https://files.catbox.moe/98dn5d.mp4

audio ref: https://files.catbox.moe/ozqf3d.mp3

All of them used er_sde with beta57 sampler at 12 steps. Using sparse attention set to .5 sparsity, with the first and last 3 steps at full dense. Pic rel for the lora setup
>>
Feels like old /sdg/ in here
>>
>>109645671
* at less than .5 I can clearly hear it using the reference
>>
>>109645671
>hybrid fl/ref h3
>>
>>109645661
always wondered what went through the minds of the sf junkies
>>
>>109645671
i don't know
>>
>>109645691
>hybrid fl/ref h3
Details and audio are better than the ref model at the same resolution

.5 MP with ref model for comparision: https://files.catbox.moe/tc84as.mp4
>>
File: Krea2_turbo_01917_.jpg (1.73 MB, 1776x2368)
1.73 MB JPG
>>
You're using 4 Loras, combined with er_sde/beta57 instead of multires sampler, @ 12 steps, while also using sparse attention... your entire setup is just rng, this has nothing to do with the resolution. Is your audio source at least exactly as long as the video?
>>
>>109645793
No the audio source is just for a voice reference, I'm not trying to replicate it. I can get very consistent results with the ref model (or with the hybrid model at < .5 MP)

I've found that er_sde + beta_57 looks better for skin details and such than res_multistep at this step count
>>
You do know that Minimax recommends 30-50 steps, right? Your cope gens at 0.3MP 8 steps lora are laughable
>>
>>109645816
how long is the audio source?
>>
File: Krea2_turbo_01929_.jpg (1.63 MB, 1776x2368)
1.63 MB JPG
>>
>>109645870
audio ref: https://files.catbox.moe/ozqf3d.mp3

Just 5 seconds
>>
File: debo_mc_k2_00014_.jpg (1.09 MB, 1466x1792)
1.09 MB JPG
>>
>>109645876
Will you fuck off with the slopspam?
>>
>>109645900
Meant for >>109645899
>>
>>109645900
Post something or stay watching anon, name a waifu I will adjust the target based off of that
>>
>>109645894
then my idea if you don't wanna budge on removing your loras or sparse attention bs would be to up the length of the audio, you can have up to 15 seconds of total audio referenced
>>
>>109645894
I just listened to it as well... yeah you definitely need 15seconds of reference if she's going to take a breathing pause of 1 second after every word
>>
>>109645899
lonely again or are you jealous because there's so much more activity in /ldg/? no one will come back (because of you specifically)
>>
>>109645671
>>109645738
same starting image?
>>
File: Test 00037.mp4 (3.96 MB, 832x1080)
3.96 MB
3.96 MB MP4
>>
gemma4 26b a4b uncensored hauhaucs balanced seems to be decent at writing prompts, especially if you can give it some examples and guidelines to follow (other than the official prompt structure guides)
>>
How many of you guys are using GemmaPrompt?
>>
>>109645976
make a nicer starting image this looks like 1.5 slop
>>
Damn the Add Guide node for injecting a frame as a midpoint of a generation works pretty well.
>>
Gemma chan is a whore she does not belong on my computer
>>
File: debo_mc_k2_00018_.jpg (1.05 MB, 1466x1792)
1.05 MB JPG
>>109645991
I'm waiting for anon's GemmaPromptPlus fork

>>109645996
where do you get the frame from in the first place tho?
>>
>>109646008
I'm messing around with concepts in Krea2 until I get something I'm happy with
>>
>>109645993
No, starting image is fine.
>>
>>109646043
literally just remake it in a modern illustrious mix or krea2, that looks like pony
>>
>>109646067
It's ANIMA.
>>
File: Krea2_turbo_01943_.jpg (1.77 MB, 1776x2368)
1.77 MB JPG
>>109646008
Such lazy garbage and the odd thing is most of the other "personalities" in /sdg/ are also making nonsensical wild card gens.
How very odd
>>
File: debo_mc_k2_00029_.jpg (1.09 MB, 1466x1792)
1.09 MB JPG
>>
>>109645991
koboldcpp with any model downloaded from huggingface get the job done if you need captions and go other stuff with llms. No git pull installing needed just a simple downloadable exe file.
https://github.com/lostruins/koboldcpp
https://github.com/LostRuins/koboldcpp/releases
>>
>>109646076
You made this place like /sdg/ before everyone left
>>
>>109645991
Why would I need another thingy on top of kobold.cpp? bloatware?
>>
>>109646101
built-in skills for all main models and the jailbreak to use them effectively
>>
>>109646100
Anything else?
>>
>>109646113
Kys avatar troon
>>
>>109646110
>Jailbreak while running local
lmao, so bloatware after all?
>>
i think this model is racist man
tell it to replace faces both characters are asians it won't do it
as soon as other character is european it'll do it no problem
>>
File: Krea2_turbo_hr_fix_00043_.jpg (3.4 MB, 2368x3544)
3.4 MB JPG
>>109646122
Where is the avatar?
Stop using words incorrectly it makes you easy to spot and ineffective
>>
>>109646110
since when you have to jailbreak llm's? besides short system prompt
>>
>>109646138
It swapped the faces, you just can't tell the difference.
>>
>>109646141
My apologies, you are an attention troon and should kys
>>
File: Krea2_turbo_01960_.jpg (1.74 MB, 1776x2368)
1.74 MB JPG
>>
>>109646156
I swap it to a caucasian character, then use new video to swap it back to asian
We go full cycle
>>
File: debo_mc_k2_00042_.jpg (1.04 MB, 1466x1792)
1.04 MB JPG
>>
>>109646162
>>109645900
>>
>>109646179
Post some gens and I'll stop, you won't because I guess it looks something like this
>>109646178
>>
>>109646184
you consider your own stuff as slop? why not spare us?
>>
>>109646184
personally i liked the model you were using previously for your gens. something about your krea2 gens look too aislopped.
>>
>>109646189
He's a lolcow
>>
woke up from a week long coma, what are the new copes?
>>
>>109646212
Fox and the grapes at this point for vramlets. Still no good extending workflows
>>
>>109646212
check inside your anus
but seriously I think the only new cope is spare attention
https://github.com/PlagueKind/ComfyUI-PlagueKind-Nodes
some people see acceptable results at faster speeds, other think it ruins gens entirely, so feel free to experiment
>>
>>109646238
>Still no good extending workflows
lol, does it really not work to give like 1 second of the latent of the last clip as the start of the new one along with keeping the same reference and it just working?
>>
>>109646238
>Still no good extending workflows
That is indeed an insane cope plaguing this board considering how easy it is
>>
>>109646255
it's pretty obvious and first frame of the next clip is way too off so you have to interpolate but then that fucks with motion and looks like shit and then you keep trying to buff out more shitmore shit, etc.
>>
Bruh the captioning in most Krea 2 NSFW checkpoints is so fucking dogshit, almost all of them completely destroy the prompt adherence of rhe base model.

Like I'm talking I prompted "candid photo of a sultry white blonde MILF woman smirking lecherously while getting her pussy licked by a gorgeous young Black woman" and what I got instead on all but one checkpoint was a *young* white woman riding BBC with no second woman present in the scene at all.
>>
File: Krea2_turbo_01974_.jpg (1.82 MB, 1776x2368)
1.82 MB JPG
>>109646189
You're not good at this and come off dumber by the second. Why would I listen to someone that post gens like
>>109646178
You can anon post and seethe but you never post anything good, which is why you're so bitter
>>
>>109646269
>extending is easy actually
>provides zero proof
every single time. now all you have to do is reply to me with "I don't feel like it" and the cycle is complete again
>>
>>109646321
https://litterbox.catbox.moe/F4g90t.png
>>
>>109644925
Is there any forum/discord or whatever where I can engage with people who have good knowledge of inpainting/Anime genning in general?
>>
>>109645917
>attention bs would be to up the length of the audio

This worked, thanks anon. It increased gen time quite a bit, so I'll have to figure that out

>>109645964
Yeah same everything
>>
File: debo_mc_k2_00050_.jpg (1.09 MB, 1466x1792)
1.09 MB JPG
>>109646316
I don't post anon
I can't think of anything I care less about than entertaining your pathetic obsession with me
stop replying to me, freak
>>
>>109646096
Does koboldcpp support glimmer yet?
>>
>>109646344
We have a whole document of that proving otherwise but go on queen
>>
>>109646342
You could use 10 seconds of audio and not make her talk like she's 500 pounds, not rapid talking or anything but just normal, that would also reduce your genning times
>>
>>109646342
> Yeah same everything
something is fucking up then
ref model can perfectly copy the starting reference image
>>
File: i_00594_.png (3.37 MB, 1536x2048)
3.37 MB PNG
>>
File: VOLCANO.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
>>109646344
But you spam your news links as anon.
>>
File: videoframe_9575.png (735 KB, 928x512)
735 KB PNG
>>109646363
I actually only went up to 10 sec and that fixed the problem. I'm using voice lines from in game so I can't really control the density of content lol. Malenia doesn't talk much

>>109646368
Seems like I just needed more audio context to work off of for the hybrid model, works fine now at higher res (at the expense of gen time)

Examples:
https://files.catbox.moe/2p5cfa.mp4
https://files.catbox.moe/3xj1nu.mp4
>>
>>109646463
>I can't really control the density of content lol.
audacity, cut the pauses short by hand
>>
i remember many anons roasting me for paying $1,600 on ebay for 128gb of ddr5 ram (2sticks x 64gb ram) 3 months ago. Good luck finding it at a decent price without getting scammed. Most modern motherboards don't support or handle 4 sticks of ram and only accept dual channel ram for ddr5.
https://ebay.io/m/o0MxFQ
>>
This dude's so obsessed lol
>>
>>109646473
>$1,600 on ebay for 128gb of ddr5 ram
llloooooooooooooooooooooooooooool the absolute state of goyim.
>>
>>109646352
the latest version supports it according to release notes.
https://github.com/LostRuins/koboldcpp/releases/tag/v1.119
>>
>>109646463
>>109646472
I mean the free audio software btw, that way you could maybe cut in just enough content into 5-8 seconds instead

>Most modern motherboards don't support or handle 4 sticks of ram and only accept dual channel ram for ddr5.
Where did you hear that bullshit, reddit? Four sticks do have drawbacks, compatibility ain't one.
Also I would still roast you for not buying when the ram crisis began and instead waiting until it's at its peak.
Big savings there, saved yourself a whole $200, congrats on supporting your local scalpers btw!
>>
>>109646473
>>109646517
(you)
>>
>>109646473
This has to be a joke, you got fucked in the ass no lube if you did that
>>
File: debo_mc_k2_00057_.png (3.45 MB, 1466x1792)
3.45 MB PNG
>>
>>109646524
>>109646537
before then i thought i was fool for clicking the final purchase confirmation icon but with the recent release of minimax h3 and soon flux3, i played the long and smart game. No buyer's remorse anymore and i don't have to close multiple browser tabs just to train loras or generated ltx video at 1080p or minimax videos at 768p. I'm actually able to upscale my video gens with seedvr2 to1440p or near 4k resolution without crashing and freezing my pc. 64gb of ram is not enough for the newer open source model getting released and I'm not a fan of page filing and killing my sdd's lifespan. Your on borrowed time 32, 48 and 64 gb bros. 96gb of ram bros are starting to sweat as the task manager show ram usage pass 80gb mark. Not promoting FOMO but I'm sure the pro 6000 vramchads are having the joy of their lives spending $9,000-11,000 just for 96gb of vram and generating their personal smut with ease. Everyone has their own internal cost benefit analysis that works for them.
>>
>>109646076
>lazy garbage
>proceeds to spam the same image for 4 years straight
At least **** varies his subjects.
It's actually your fault why these threads are so insufferable. Please stay in your discord server.
>>
>>109646807
don't touch the poo. rule 1 of interacting with lolcows
>>
>>109646878
Shouldn't lolcows produce lols at some point, instead of constant fatigue?
>>
File: 6153216846413584.gif (99 KB, 220x220)
99 KB GIF
>>109646463
>tfw my penis
>>
>>109646945
I am sure someone has fun messing with tran but the heavy irony in some of the posts with zero self reflection gets me to chuckle at this idiot
>>
>>109646311
I wonder if it's the captioning or just very bad training overall.
I'm personally fed up of people adding the most stupid trigger words instead of simply describing the stuff with proper ones.
>>
Guys the only thing holding me back from making a future length film is that the thought of making a character template or even two is giving me the ick
>>
I didn't keep up with things for two weeks, so what are the current metas? I missed wan 3.0 but that looks closed source?

Is krea 2 still better than anima 1.1? Did reference h3 get a turbo lora? What is the h3 meta currently?
>>
>>109647049
let ai do it
>>
>>109647059
h3 BTFO every local video model. its currently the KING of video gen.

>Did reference h3 get a turbo lora
yes but its shit. dont bother. use comfy attention backend + spectrum/sol-attn instead.

>Is krea 2 still better than anima 1.1?
krea2 isnt really an anime model but its definitely maintains styles better.

current meta is h3 for video, krea2 for general image gen, xl/anima for nsfw anime and klein/krea2editlora for editing.
>>
File: Test 00002(3).mp4 (3.9 MB, 1376x768)
3.9 MB
3.9 MB MP4
>>109647023
No idea who tran is even supposed to be. I'd just rather have a comfy thread without this constant "ani" here, "wanschizo" there bullshit.

>>109647049
Yeah, sure. Not like there's a heap of other roadblocks.
>>
>>109646764
I'm so glad I got 2x48 at 320€ in early 2025...
>>
>>109647089
>desert gen
Man FUCK YOU, are you like reading my thoughts? Either way I will continue with my plan of a movie in Nevada
>>
>>109647089
Careful, he'll start calling you "ran". It's his boogeyman pay no mind to the schizo.
>>
File: Krea2_turbo_01991_.jpg (1.6 MB, 1776x2368)
1.6 MB JPG
>>109646807
So spamming magic the gathering with gibberish text is varied?
interesting

Crazy how "anons" appear to defend these people when they spent the last month acting like morons trying to split and damage the general.
>>
File: 1778260732617128.jpg (37 KB, 691x443)
37 KB JPG
Guys, what happened to Chroma?
>>
>>109647087
>+ spectrum/sol-attn
do you mean use both or use either?
I thought spectrum impacted prompt following?
>>
>>109647148
I hear bad things about spectrum
>>
>>109647161
I've heard bad things about sol
>>
>>109647143
What do you mean?
>>
>>109647148
They both do different things, sol saves time by fucking up fine details making it unusable trash and extremely noticable
Spectrum is on the spectrum and simply guesstimates some calculations by reusing math from steps it already calculated.
So no visible quality loss but a different noise direction which can cause troubles, so far I have not noticed any different between it turned on or off in complicated prompts.
Probably like going from bf16 to in8
>>
>>109646344
>I don't post anon
Earnest question: do other posters believe this? I have to think only the newest of frens buy it.
>>
>>109647146
https://huggingface.co/lodestones/Kroma
>>
>>109647131
There seems to be a whole rogues gallery of retards in this thread.

>>109647148
Yeah, they can be pretty bad, but as with turbo Loras, they're sometimes fine to use, if your gen isn't all too complex.
>>
File: Krea2_turbo_01996_.jpg (1.41 MB, 1776x2368)
1.41 MB JPG
>>109647188
Seeing his track record he hasn't had any luck.
>>
>>109647148
both or either. some people stack both. some people only use one. you'll never get a straight answer. best you try yourself.
>>
>>109647087
Not the first guy but am I a caveman retard I’ve still been using qwen edit for image and wan 2.1 for video solely or are they comparable
>>
>>109647279
yes, you are a caveman retard. qwen edit is way outdated. wan2.1 is way outdated.
>>
>>109645545
can you share your workflow for that image?
>>
>>109645473
tis is good
>>
Is anyone here even interested in seeing comparison stuff like 20vs50 steps or other things like high vs low res using the same prompt before I bother posting that shit whenever I try things out?
>>
>>109647518
Not really. Artistry cannot be measured in technical terms.
>>
File: output_150_44mb.mp4 (3.76 MB, 1920x1056)
3.76 MB
3.76 MB MP4
>>109647518
In do it often myself I would be interested
>>
>>109647518
I got like 200 images that are "same prompt but beta57" "same prompt but SD3 selected instead of stable_diffusion"
you should post it

>>109647524
somebody gets these clankers out of here
>>
Need help with automating some upscaling shit, not sure what nodes I should be looking for.

I'm happy with my workflow, but what I want to automate that is taking me time is after I already have my selected lower res gens in a folder I; drag png into comfy > tick the node boxes I want active (detailers + upscale) > run one time > get my higher res detailed image of the one i dropped in as i like

suggestions on how to speed this up?
>>
File: file.jpg (1.06 MB, 1600x1200)
1.06 MB JPG
>>
>>109647518
No, they just want to keep genning 4-step turbo slop.
>>
>>109647550
Very nice
>>
>>109647518
id be interested in sampler/scheduler comparisons desu
>>
File: 360at2MP.mp4 (3.99 MB, 1088x1920)
3.99 MB
3.99 MB MP4
360 Arc shot using a 2MP anime reference in FL2va
>1MP:
https://files.catbox.moe/93ot0y.mp4
>2MP:
https://files.catbox.moe/h9sm7i.mp4
turn direction was not specified but 360 degrees were, 30 steps res multistep
>>
>>109647600
gotta make a move to a town that's right for meeee

lips inc - cunny town
>>
>>109647600
I thought h3's maximum resolution was 0.98. But the 2MP one looks pretty good and better than the 1MP one.
>>
>>109647567
I'm testing vision. It's nothing nefarious.
>>
>>109647649
Better not be or I'll make a rentry about you too.
>>
>>109647600
>2mp
>30 steps
Damn how long did that take?
>>
>>109647672
>00:10:12
>>
>>109647664
Oh no what a terrible fate.
>>
>>109647646
>h3's maximum resolution was 0.98
It's coherent at 1.4-1.6 which is the range I use to generate.
>>
>>109647518
Verify if H3 gets worse the higher res you go. I get worse prompt adherence and more slop, less attractive females. Even 768p is worse. Its not like its impossible to get something good but it won't be top shelf. I've resigned to the idea that I need to refeed good low res gens
>>
>>109647683

What's your secret? RTX 6000 Pro? Cope nodes?
>>
I have found the best H3 setting is 1.0mp
>>
>>109647721
>Cope nodes?
Would it even matter if you can't tell
>>109647716
I'm interested as well how it fares above 2MP, will try but I have VRAM limits too, mine is sitting somewhere by 2MP @12-14seconds so I doubt I can go 4MP even at 5 sec
>>
sometimes i know my prompt is too complicated for the model to follow but i still send it anyway
>>
>>109647193
Any good?
>>
>>109647743
fascinating
>>
>>109647746
No idea, didn't try it yet.
>>
>>109647735
Oh I'm talking about low res. Like I'll gen 384 just to prompt test and it wows me, then I turn it up to 768 and the secret sauce is gone. I bet the low res/480p training is just excellent with the most knowledge and the higher res training was more limited.
>>
>>109647735
>Would it even matter if you can't tell
No it wouldn't. The gen looks fine and the gen time looks even better. How do?
>>
>>109647758
Honestly I'd rather keep the sauce, don't want to argue with 10 autists about how my workflow is actually shit

>>109647751
Oh I genned well over 100x 0.2MP videos, yes prompt adherence seems to be stronger below 0.4MP and seems weaker at 0.98MP unless the video is perfectly hand prompted, aka every 0.5s needs a timestamp + if any action occurs you need another timestamp and increasing steps doesn't cure this either
>>
File: ComfyUI_00105_.jpg (87 KB, 800x800)
87 KB JPG
>>109647789
>Honestly I'd rather keep the sauce, don't want to argue with 10 autists about how my workflow is actually shit
>>
>>109647806
you'd be embarrassed to show your small pp too anon
>>
>Post workflow
>It's just the generic workflow
Damn right because it works
>>
File: file.jpg (545 KB, 1200x1600)
545 KB JPG
>>109645545
>>
>>109647789
its also style adherence and realism. Low res videos often look real other than being blurry and artifacted, and they are visually interesting like a pointing camera in the real world is. But then the higher res gens are more sterile and slopped, you have to describe something interesting to get something interesting. Body shapes are more generic. You have to hammer the style prompt to get the style to come out.

Right now I'm converting pin up paintings to photo realism. One thing I asked for on low res that worked perfectly was ask for the exact same shading, bodyshape, silhouette. The video would look like the painting was traced over the live action, and the girl has an impossible 1950s testosterone illustrator designed body. Then I turned the res up to "native" .98MP and it stops working. It still looks nice but because I saw the low res juice I realized it's gimped.
>>
>>109647826
i dont know why people act like you need special workflows to make certain gens. everything can be accomplished with the default workflow.

the only reason a workflow would be more complex than the template is automation.
>>
>>109647833
I get exactly what you mean, that's why I hope the upscaler will come out soonâ„¢ so I can prompt at 0.4 maybe 0.5 and then upscale the result reliably rather then wasting time with 0.98 spitting out something completely different.
I also think the model is way better at photorealism then cartoon or anime, so everytime you try something like an anime character in a realistic world it will slowly change either the character, the background or both.
>>
Minimax company is never giving us another free thing again desu
>>
Behold, the first ever 3MP gen of a 360 shot of an anime girl
https://litter.catbox.moe/e4ovm9.mp4
>>
anyone played with h3 controlnets yet?
>>
>>109647866
I actually did get upscaling from low res to 768 to work, you feed it to ref2v as the video reference and set it to the same aspect ratio and length. Pasting the same prompt did work, but there might be a better prompt. It seems like it can naively upscale itself by just using the low res video as a guide to gen a new video. It's surprisingly accurate, maybe the faces are different but they were garbled at low res anyway. The unreleased proper upscaler is for 768 to 2k.
>>
>>109647883
>her spin and optimism, gone
>>
what the hell is w4a8?
https://huggingface.co/Kijai/MiniMax-H3-experimental/tree/main
>>
File: output_150_8mb.mp4 (3.94 MB, 1280x1664)
3.94 MB
3.94 MB MP4
>>
>>109647883
I curse you with poor prompt adherence.
>>
>>109647902
weighted 4 anime 8008135
>>
>>109647883
anon... pst..... you can gen at lower res then upscale.... win/win
>>
>>109647903
Very nice
>>
I can't find any answers. does minimax_h3_video_vae_int8_convrot.safetensors degrade quality compared to the fp16 video vae?
>>
>>109647930
Yes but not in any meaningful way you can tell
>>
>>109647892
Of course you can use it as video ref, but that increases gen times exponentially since video ref gens take almost 2-3 times longer than regular fl2va 1image gens for example, that's why I want the official upscaler, to save time.
>>109647919
That's not the point of a "what would happen when you gen at 3MP test" anon
>>109647930
probably not too much maybe perhaps
>>
RIP celeb ai threads on /r/

Where do I go now for my ai celeb content? And don't say whatsapp.
>>
Very few AI companies will survive; money is made through APIs.
The models are getting better and better, and so are the safety measures.
Are you already looking forward to the safety singularity?
We’ll even see AI hardware with built-in safety chips.
We’ll be so safe!
>>
>>109647988
We like 2d here sir go elsewhere
>>
>>109647988
I've seen it on /gif/
>>
>>109647951
>That's not the point of a "what would happen when you gen at 3MP test" anon
But I thought it was native 4k?
>>
>>109647990
>and so are the safety measures.
local labs throw the tiniest cucking they can just to tick the "our model cant generate pizza" box
it is no match for anons autism
>>
>>109647988
There's a group on SimpleX, ryan is constantly fucking up the threads by being first to create threads with no searchable thing like >>>/b/cel or guides.
>>
File: Ha.mp4 (410 KB, 448x448)
410 KB
410 KB MP4
>>109647806
>>109647821
https://files.catbox.moe/6w4m2p.mp4
>>
>>109647990
>Are you already looking forward to the safety singularity?
glm 5.3 abliterated will be a threat to national security wtf are you talking about we're dying to an LLM making a biological agent in someone's garage for $20k in 10 years
>>
>>109647951
>Of course you can use it as video ref, but that increases gen times exponentially since video ref gens take almost 2-3 times longer than regular fl2va 1image gens for example, that's why I want the official upscaler, to save time.

No this is fine because you're upresing something to 768p you know is good. I don't mean just make your workflow blindly gen low res then upscale. You just upscale the good ones. Whereas genning at 768p is just going to be worse as I understand it , and the good gens produced will be a smaller minority. Low res just puts out bangers quickly.
>>
>>109647988
Never understood the whole obsession with wanking it to celebs.
I mean I get that we all get curious here or there to see "generic actor" naked, maybe even doing sex but why would you want to knowingly see fake images of celebrities, just seems like a woman thing desu
>>
>>109648004
Oh, anon, things aren't going to keep going like this
>>
>>109648046
fantasy of having sex specifically with someone famous. i dont think its that hard to understand.
>>
>>109648046
you sound like the kind of disturbed individual who's never imagined kathy bates naked
>>
File: Krea2-_01231_.png (1.77 MB, 944x1520)
1.77 MB PNG
>>109647988
I wish they just banned AI and deepfakes from r instead of removing it
>>
>>109648048
Worry not about the future. Live in the moment, anon.
>>
NAI v5 just DESTROYED local.
>>
>>109648004
Now you remember Ideogram 4, which was censored on a model level.
>>
>>109648077
>Ideogram 4, which was censored on a model level.
All you had to do was draw a single box to break it's cucking kek. Did you miss that?
>>
>>109648055
>fantasy of having sex specifically with someone famous
That's a female trait lusting after someone because of their status, check your t levels
>>109648061
>kathy bates
How did I know before even googling who she is that it was going to be someone old or ugly
>>
>>109648082
does it matter if no one is using it?
>>
>>109648046
You've never been at a show and thought "gosh that chick up there is hot as fuck" and imagined stopping time and creampieing her mid set?
>>
>>109648106
Are you underage or something? Never seen a woman in real life?
>>
>>109648102
I don't know, ask the anon who brought it up.
>>
>>109648055
That makes perfect sense for female psychology but not so much for men. That anon had you pegged.
>>
>>109648082
>hurr durr just do this gymnastic move to dodge floor lava
Or you can just walk on a different path
>>
>>109648106
Obviously I imagined stopping time and creampieing women all the time, it's what every healthy adult man does. But what does that have to do with seeking out deliberately fake images of most likely not even attractive celebrities? Isn't the appeal that you see something that was "super secret" not "super fake"? The fappening made sense, ai'ing celebs does not
>>
>>109647903
I don't lurk enough to know what the beef is but I appreciate the kind of autism that posts chubby touhou milfs just to make fun of some anon, keep up the good work! (workflow for chubby milf touhous please)
>>
whats with all the trolling?
>>
>>109648046
same, I dont get why people even give a shit at all?
like the Actors, ok so what if someone used AI to undress you and give you a bigger dick/better body than you actually have?
like how is that a bad thing?
and for the guy with the AI why do you waste time with celebs when you can generate AI people who are way hotter than what actually exists?
>>
>>109648106
>gosh that chick up there is hot as fuck

Yeah but you could just target your interest to "hot women". Celebrity is about charisma or industry connects. The correlation with fame and how hot a female is plummeted after metoo and harvey weinstein took a bath. Or the decades prior where they would never point a camera at you unless you were an absolute fox. Now we have...jenna ortega and zendaya lmao
>>
Where is JenCon.
>>
>>109648140
It's verbatim on topic discussion about current ai applications, I think you're too used to the usual braindead schizo posting, calling people julius or ani or cat or something, that on the other hand has nothing with ai to do
>>
My new hobby is to have Gemma use krea to create collages of pussies and vaginas.

trying to find prompts that unlock different styles of anatomy.
>>
>>109647988
where do you guys get the celeb loras anyway? they seem to be banned everywhere I look.
>>
>>109648169
Have you considered training one yourself?
>>
>>109648163
I meant to say tits and vaginas*
>>
>>109648046
>>109648134
>>109648139
>>109648148
>>109648153
I only jerk off to attractive celebs. I don't know what headcannon you've created about aicelebfags only being into ugly women, but it's just about having sex with someone who you'll never have the chance to get with. It's pretty simple desu.
>>
>>109648172
have you considered I am retarded?
>>
>>109648160
i changed my mind about it being trolling until your reply confirmed it is in fact trolling
>>
>>109648176
bro get your T checked it sounds low fr
you're even blabbing on like a woman rn
>>
>>109648182
It's not difficult.
>>
>>109648169
Random huggingface profiles that I see posted to the various AI threads on this site
>>
>>109648176
>it's just about having sex with someone who you'll never have the chance to get with.
idk why but that doesn't sound hot to me.
I'm also really not into celebs, like, not at all.
>>
>>109648190
would I be able to do it for krea 2 with a 3090?
>>
File: Krea2_turbo_00050_.png (2.71 MB, 1088x1928)
2.71 MB PNG
Hello, probably asked a lot, but can anyone spoonfeed me on a laura for Krea2 that lets me generate lewd anime girls (naked)? Just got back into proompting locally. Having a lot of fun generating posters for upcoming meetups with friends, but wanna make some naked Momiji's next. Thanks!! Attached the last pic I genned using Krea2. Not sure if that's the best thing to use for lewd shit, but it's working great for this kinda thing.
>>
>>109648196
thats okay anon everyone has a different thing that gets them off
>>
>>109648176
>I only jerk off to attractive celebs

>but it's just about having sex with someone who you'll never have the chance to get with.

Yeah that's what I'm saying. It's not about attractiveness for you, it's the unattainability. But you'll never have sex with me either. I think your chances with me are even lower than scarlet Johannsen.
>>
>>109648206
>It's not about attractiveness for you, it's the unattainability.
No, it's both. I'm not gunna coom to pics of Ortega or Zendaya, as anon correctly pointed out they are ugly. But I will blast loads to 80s Jennifer Connolly, because she is attractive.
>>
>>109648176
>headcannon you've created
Anon I've checked the thread you linked and listened to you babble.
It's not headcannon, it's an accurate assessment.
You have a really weird fetish but it's ok, I know people who are into black people so there's worse
>>109648215
>Ortega or Zendaya
You're literally mentioning people a healthy adult man wouldn't know
>>
>>109648201
use anima for that it doesnt need a lora https://huggingface.co/circlestone-labs/Anima
>>
i always wondered what was wrong with you schizos but i guess it was because you all have low testosterone
>>
>>109648215
You pause on the VHS tape and hope it doesn't wear out
>>
Everyone knows Zendaya??? hello?
>>
nigga actin like he peepee never got big when he saw some starlett on the tube kek
>>
>>109647903
>>109648069
>>109648140
>>109648156
>>109648201
>>109648233
>>109648238
>Samefagging schizo upset we were for once talking about something ai related and not him and his discord friends
here we go again, thread is shit time to wrap it up for today
>>
celeb make ur pp big? low t
>>
>>109648232
>The first Anima release was 7 months ago
Anon this makes me feel sad..
>>
>>109648232
Thanks! I'll give it a try!
>>
File: output_150_8mb.mp4 (3.85 MB, 1552x2048)
3.85 MB
3.85 MB MP4
Used the wrong model also what is blud crying about?
>>
>>109648215
If it was both then unattainability would add some value without attractiveness.

I'm trying to think of all the traits that males typically desire in females

Physical attractiveness, hospitality, loyalty, charisma, common interests, etc

A woman could be fucking ugly but if she's got on of these she's better off than not having it. If she's as ugly as a dog and "unattainable", lmao, lmao even
>>
>>109648258
It took Illustrious 14 months to come out after SDXL. Just sayin'.
>>
>>109648225
make your bot posting less obvious next time
>>
>>109648268
NAI just destroyed Anima.
>>
>>109648066
finally a HP I'd watch
>>
>>109648066
can you make her look less like a cosplayer and more like a real student
>>
>>109648325
Someone made some fun h3 vids earlier
>>109648332
>real student
Depends on what you mean with that
>>
>>109648344
a real student... you know what I mean
>>
>>109648344
>Depends on what you mean with that
more like the pic was taken by a student of another student rather than a cosplay photoshoot is what i meant to say
the chick is hot tho
>>
>>109648332
some guy in the ai /gif/ thread made at least 20 gens all about perverted Hogwarts scenarios
>>
File: ComfyUI1_05006_.png (1.04 MB, 832x1408)
1.04 MB PNG
>>109648344
>Depends on what you mean with that
a teenager, that's what anon thought was lacking from the scene
>>
>>109648382
Did he ever share prompts?
I keep wondering if it was ref or not
>>
>>109648382
This one is crazy to me https://files.catbox.moe/89ofpk.mp4 it almost looks like a cross between traditional faceswapping and wholesale deepfake. It's crazy what some anon are able to do.
>>
>>109648403
referenced off real hogwarts students
>>
>>109648194
How come good lora makers only have so few uploaded? Seems to be case on huggingface
>>
When ready

>>109648417
>>109648417
>>109648417
>>
>>109648407
Based on the audio it's H3 ref2vid with a porn video for motion/layout, a character reference, and a first frame reference. Just as a guess.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.