[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: output_4mb_no_audio.mp4 (3.96 MB, 1366x2048)
3.96 MB
3.96 MB MP4
New music model not working on my blackwell gpu ;_;
Previous: >>109544588

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
>>109547997
Serious questions for you.

Do you think this OP pic was appropriate for this general and thread? What the fuck is wrong with you?
>>
>>109547997
What happen to the "Discussion of..." part?
>>
>>109548011
weebs don't understand what is or isn't appropriate.
>>
>>109548011
Relax it’s just cartoon girls kissing.
>>
Need a version with flat Migu. Teto's tits can stay fat though.
>>
>>109548017
>>109548011
Listen I don't bake often
It is what it is
>>
>>109547873
I just gave this new lora a try https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_ref2v_turbo_4step_v0.1_comfyui_bf16.safetensors

it's trash

Do a comparison. At 8 steps res_multistep/simple the quality is horrible. ema 600 is miles better retaining crisp details (8 steps euler/beta)
Tthis new lora makes 1mp look like 360p, especially in motion. Ema 600 still wins in ref2v
>>
>>109547997
hot
>>
>>109548029
Based
>>
>>109548030
OP image is using that lora works on my machine at 6 steps
>>
>>109548025
But girls kiss each other is devilry!
>>
So how does the new LTX 2.5 compare to MiniMax H3?
>>
File: file.jpg (347 KB, 1024x1536)
347 KB JPG
>>
File: banner.png (592 KB, 1536x512)
592 KB PNG
Anons please enjoy my new prompt creation/enhancement tool which runs fully local using any openai comaptible backend (I use Gemma4 of course). Gemma-chan now knows H3, klein, krea2, anima, etc!
https://github.com/whp199/GemmaPrompt
>>
>>109548039
It's bad. Now run ema600 euler/beta 6 steps
>>
>>109548041
Even better.
>>
>>109547997
Next time, make a collage.
>>
okay, the reference turbo lora, but at 8 steps, is VERY good and giving me better outputs vs spectrum/20 steps. my downfall meme prompt turned out even nicer.

https://files.catbox.moe/dd1trg.mp4
>>
anyone test the new anima model? the one that added a few layers and additional booru stuff
iz it gud?
>>
hey guys check out my fwc (fat white cock)
https://files.catbox.moe/uk7nex.png
>>
If I use the ref turbo lora, should I be disabling Spectrum? Especially if I have camera movement.

I was trying the fl2v turbo on my ref gens yesterday and I noticed it frequently produced undesirable visual effects, which I think were extrapolated from the lighting in my scene.
>>
>>109547997
Pretty sure this OP image will get deleted.
Good job OP.
>>
>>109548049
works on my machine
>>109548051
I shouldn't have to be OP and actively avoid it. I should be the last person being OP. So no I never will
>>
>>109548056
is it smaller than average cock Thursday already?
>>
>>109548052
also this is only at 0.4mp! but it's pretty clear.
>>
>>109548052
what happened to Morocco and Algeria?
>>
File: Nemona_00060_.png (791 KB, 896x1152)
791 KB PNG
>>109547978
you seem to like my nemona gen anon
>>
>>109548052 >>109548049
>>
>>109548063
You didn't have to bake on page 5 you nervous premature ejaculator.
>>
>>109548075
It was at page 6 are you just complaining to complain?
>>
>>109548069
>LAYOUT:
Adolf is in a bunker sitting at a desk in front of an army map of europe, during world war 2 in Germany.

china must know something we dont
>>
>>109548056
crispy
>>
>>109548080
OooOOooHH oh my god!! not page 6!!!!???
>>
>>109548075
>nervous premature ejaculator
don't shame a man for trying to run a marathon on his first try.
>>
I'll ask again
Anonymous 08/13/26(Thu)21:36:42 No.109548084▶
>>109547848
>https://github.com/whp199/GemmaPrompt
Can I feed it pages of a script and it will generate the shots for the scene, and arramge tje shots in the minimum amount of clips to be generated.
>>
>>109548047
>krea turbo and not base
BAKA
nb4 hurrdurr you arent supposed to use it
nb4 hurrdurr its worse than turbo
(I know it'll work for both versions anyway but)
>>
File: 1662960455357_1.mp4 (2.31 MB, 1280x720)
2.31 MB
2.31 MB MP4
Challenge for Minmaxers: do a character swap for vid-related using Ref2V.
I already tried. First I tried distilling the video into a text prompt, result was okayish but nowhere near the level of the video.
Then I tried using the video as a reference input and it outright just didn't work. It just copies the input video into the target video. Even if I mark it as a weak reference, it still copies it.
>>
>>109548051
It's car jack, he doesn't do collagens.
>>
Ok OP, I apologize, I'm being too hard on you. Thanks for baking.
>>
>>109548106
Busy with music right now and disappointing it's not working on my gpu
>>
File: file.jpg (447 KB, 1024x1536)
447 KB JPG
>>109548044
Another one.
>>
>>109548052
slightly more angry hitler:

https://files.catbox.moe/j9yni7.mp4
>>
>>109548121
Do the same with ema 600 euler/beta. I'm waiting
>>
>>109548075
>>109548089
Are you upset that some low quality xitter screenshot thread was bumped off the catalog to make room for this Blessed thread of frenship?
>>
>>109548047
Impressive. Very nice. Seems to work really well, way better than my attempts to make a card in sillytavern and run it that way, It's faster too.
This bitch needs correction though that attitude.
>>
>>109548052
I need to use 12 steps or it doesnt follow the prompt
>>
>>109548135
Post an example prompt. I ain't installing python shit.
>>
i like short hair gemma better
>>
>>109548136
TETSUOOOOOOOOOO
>>
Anyone else getting Error log


# ComfyUI Error Report
## Error Details
- **Node ID:** 37:13
- **Node Type:** MiniMaxMusic3TextEncode
- **Exception Type:** AttributeError
- **Exception Message:** AttributeError: 'RVQDepthDecoder' object has no attribute '_v_block'
with the new music model?
>>
>>109548133
I'm upset because he didn't make a collage and that he self inserted in the OP.
>>
>>109548139
>I ain't installing python shit.
get good, lazy faggot. even vibecoders like anon at least put some effort in.
also i didn't have to install anything myself it just ran.
>>
>>109548144
https://files.catbox.moe/f6ejv7.mp4
>>
>>109548148
That's more on you than on me.
>>
File: Gemma-chan.png (1.73 MB, 1000x1496)
1.73 MB PNG
best gemma.
>>
File: 1763753323774718.png (142 KB, 1860x507)
142 KB PNG
>>109548136
8 is fine for me
>>
>>109548161
you don't need a node for comfy kitchen you fucking idiot
>>
>>109548151
>>109548144
It works now they made a 2 line error and pushed a fix
>>
>>109548044
>>109548119
Reminds me of Noodles from Gorillaz.
>>
>>109548174
I need it cause I dont want to use it for every model with the startup args, I know it works fine in this case so the node is better, what if it's shit for krea 2? if not I just add 1 node.
>>
well it's not censored. that's good. still tweakin' it but this has some good potential as a model

Experimental NHH:
https://voca.ro/1k3BrgffDcSF
>>
>>109548161
>no spectrum
>>
>>109548180
Yeah I remember those images vaguely too.
This was a test with gemma 4 to generate prompts. I don't need python shit for this.
>>
>Can run everything at MAXXXXXXX
Feels good man great speeds for 5 minute gens
>>
>>109548197
>it/s
what
>>
File: Screenshot_3568.png (3 KB, 419x33)
3 KB PNG
>>109548197
fuck you
>>
>>109548190
yes because it fucks with the lora, you have two diff types of math going on, test with/without
>>
>>109548208
That's someone with daddy's GPU.
>>
>Adolf gets into the drivers seat of a white Toyota Supra parked beside the desk and drives away very fast, as eurobeat music starts to play.

https://files.catbox.moe/a7v9sb.mp4
>>
>>109548211
No, fuck you leather head!
>>
>>109548215
it has to be some fuckery with how it calculates steps to make anons feel good looking at it/s instead of s/it
>>
>>109548161
is modelsampling still a thing?
>>
>>109548106
holy fucking hell, the future will fry my dopamine receptors i tell you that
>>
File: file.png (181 KB, 263x405)
181 KB PNG
How is everyone doing today?
>>
>>109548072
yeah it was good
>>
>>109548239
sigma shifting has always been a thing
>>
>>109548228
>warps into the car
do ema600 pussy
>>
>>109548230
Usually the richer the guy is = less he is capable of creating art or even funny gens.
Applies universally.
>>
>>109548244
Please ana no, my dick is tired
>>
>>109548252
gross.. what an ugly dumb bitch
>>
>>109548239
12/3 default works yes, ensures audio and video are less messed up in general
>>
>>109548184
are you still using AceStep? Have you tried MiniMax music?
>>
>>109548251
that's completely irrelevant
>>
>>109548161
What line goes to guide and to sampler?
>>
>>109548265
no i'm using krea to gen it
>>
>>109548258
sup moishe
>>
File: 1759736567523634.png (475 KB, 1920x938)
475 KB PNG
>>109548270
chain with lora and other stuff to basic guider
>>
File: file.png (529 KB, 1200x675)
529 KB PNG
>>109548251
>>
>>109548269
Maybe for you because you don't even generate images or videos.
>>
>>109548272
lolwut?
>>
hitler visits a starbucks

https://files.catbox.moe/yg4rvd.mp4
>>
Why won't they provide a basic prompting guide with this thing?
>>
>>109548291
anon you're arguing with shadows
>>
>>109548281
Thanks. Is the new ref lora any major upgrade?
>>
>>109548297
We are nothing but shadows and dust anyway.
I'm actually preparing my new image prompt as we post.
>>
>>109548301
yes, but 8 steps, 4 may be too blurry. and the fl2v 8 step lora is already very good, this is the first ref specific lora they released.
>>
>>109548296
>250bpm,skrillex,drum and bass,uk,mick gordon,kitchen rattling
>>
>>>>109548281
You guys bring new copes every day ending up using none of them
Dissolve your delusions or prove me wrong by a comparison >>109548030
>>
0.6mp, vs 0.4 (previous)

https://files.catbox.moe/aim6qq.mp4
>>
File: 1772345002765579.jpg (9 KB, 372x160)
9 KB JPG
>>
Anon who told about
>https://huggingface.co/Astral01/Ace-Step_v1.5_XL_Base_Turbo_0.3_Merge
What app do you use for musicgen?
>>
>>
>>109548346
the legendary...super saiyan god...
>>
Could someone point me to the latest loras? things are moving so fast
>>
Just woken up from a 2 week coma what's the current FOTM model that'll be forgotten next month
>>
>mfw Resource news

08/13/2026

>Lightx2v MiniMax H3 Turbo Ref2V 4Step/8Step Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>GemmaPrompt: Local prompt enhancer for ComfyUI diffusion models
https://github.com/whp199/GemmaPrompt

>ComfyUI-H3-FaceRefine
https://github.com/Carasibana/ComfyUI-H3-FaceRefine

>Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising
https://github.com/Ai-ZL/Hybrid-LUT

>MiniMax H3 Creator for ComfyUI: Multi-Shot 60s Timelines, Resizable Satellite Stage, & Ollama/LM Studio Refiner
https://github.com/roadmaus/ComfyUI-MiniMax-Creator

08/12/2026

>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler
https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler

>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready
https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy

>LTX-2.5 22B — GGUF
https://huggingface.co/realrebelai/LTX-2.5_GGUFs

>ComfyUI NVIDIA RTX VSR Pro
https://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro

>Stable Layers: Decomposing Images into Editable RGBA Layers
https://huggingface.co/StabilityLabs/Stable-Layers

>PEAK: Precise and Persistent Concept Erasure via k-Sparse Autoencoders
https://github.com/manmanTAT/PEAK

>Flow Straight to Reality: Perceptually Consistent Flow Matching for Efficient Image Restoration
https://github.com/aiimaginglab/PCFlow

>MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understanding
https://shuaiwang97.github.io/MMArt

>ComfyUI H3 Studio: Video editor for MiniMax H3 inside a single ComfyUI node
https://github.com/shootthesound/ComfyUI-H3Studio

>MINIMAX H3 Prompt Studio
https://github.com/lololerigolo60/Minimax-H3-prompt-studio/tree/main

>ComfyUI Image Conveyor v1.4
https://github.com/xmarre/ComfyUI-Image-Conveyor/releases/tag/v1.4.0

>VPIPE: Real-time multimodal AI pipelines on Apple Silicon
https://github.com/tgo-app-dev/vpipe

08/11/2026

>LTX-2.5
https://ltx.io/model/ltx-2-5
>>
>mfw Research news

08/13/2026

>Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence
https://arxiv.org/abs/2608.12290

>Through Van Gogh's Eyes: Global Style Transfer with Diffusion Mod
https://arxiv.org/abs/2608.11546

>LoSA: Near-Lossless Sparse Attention for Training-Free Video Diffusion Acceleration
https://arxiv.org/abs/2608.12032

>UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos
https://arxiv.org/abs/2608.11752

>RA-ClipScore: Making Generative Model Evaluation More Interpretable
https://arxiv.org/abs/2608.12088

>HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion
https://arxiv.org/abs/2608.11913

>Robustness of AI-Art Detectors under Generator Shift
https://arxiv.org/abs/2608.11643

>LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time
https://arxiv.org/abs/2608.11745

>TangPoetryBench: A Multi-Dimensional Benchmark and Rubric-Conditioned Evaluator for Poetry-to-Image Generation
https://arxiv.org/abs/2608.11452

>Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment
https://arxiv.org/abs/2608.11537

>Generative Video Compression Based on Hierarchical Referencing
https://arxiv.org/abs/2608.11618

>Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections
https://arxiv.org/abs/2608.11576

>Understanding Why Foundation Models Work for Diffusion-Generated Image Detection
https://arxiv.org/abs/2608.12155

>Draw This First
https://arxiv.org/abs/2608.12064

>Context Blindness in DPO: Mitigating Object Hallucination in MLLMs via Context-Calibrated Preference Optimization
https://arxiv.org/abs/2608.12158

>simpleposter: A simple baseline for product poster generation
https://arxiv.org/abs/2605.08784
>>
File: VOLCANO.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
>>109548367
Please stop spamming
>>
https://files.catbox.moe/rtb0l5.mp4
>>
adolf learns to calm down:

(second source was asuna)

https://files.catbox.moe/yupxmw.mp4
>>
>>109547997
Using minimax h3 with face reference pictures, is there a way to increase likeness when the camera is zoomed out? I'm trying to gen shots where the camera zooms in on the face and while similar from afar, the face starts morphing into the reference only once the camera gets pretty close.
>>
>>109548355
Interstellar alternate ending
>>
Should I update my comfyUI portable?
>>
>>109548375
did he died
>>
File: imhavingablast.jpg (3 KB, 242x18)
3 KB JPG
>>109548424
>>
>>109548409
you'll have to explicitly tell it to with time stamps
>>
can I get good gens on a 12gb gpu or do I have to load up on cope nodes just to get shit slurry soup?
>>
>>109548443
I usually do a scene overview where I already explain that the person should have the face in reference XYZ. Does that mean I have to give it the face at different zoom levels??
>>
>>109548409
H3 has issues with faces that take up only a small part of the frame. It just kinda sucks at it. Higher res may help but there's no great solution, I suggest just using more close up shots. Someone posted a face refiner node that may or may not work
>>
File: 1773905476398898.png (693 KB, 455x703)
693 KB PNG
>>109548244
jumpscared by the uggos on that browser
>>
this model is amazing. I just added asuna to test if it'd work. ref2v turbo lora but 8 steps.

Use <Picture 1> as the reference image for Adolf. Use <Audio 1> as the voice reference for Adolf's dialogue. Use <Picture 2> as the reference image for Asuna.

LAYOUT:
Adolf is in a world war 2 bunker in Germany.

an anime style Asuna is standing beside Adolf, who is sitting at a desk. Adolf says in german "Ich habe euch gesagt, ihr sollt zuhören, aber ihr habt nicht zugehört – und jetzt haben!"

Asuna sits in Adolfs lap, and Adolf says in german "blue archive, nice."

https://files.catbox.moe/uxn8gt.mp4
>>
>>109548244
What tool is that?
>>
>>109548477
https://huggingface.co/spaces/malcolmrey/browser
>>
>>
>(adult woman:-1)
>>
File: MiniMax_H3_00053_noaudio.mp4 (2.15 MB, 1376x768)
2.15 MB
2.15 MB MP4
>>
>>109548529
(elderly woman:100)
>>
>>109548529
based.
also y'all share your cunny prompts. i'm already 2 days out of ideas
>>
>>
>>109548464
I guess it's just a limitation then
>>
File: debo_dm_k2_00085_.png (2.3 MB, 1872x1007)
2.3 MB PNG
>>
>https://huggingface.co/Kijai/MiniMax-H3_comfy/blob/main/loras/minimax_h3_ref2v_lightx2v_turbo_4step_v0.1_bf16_resized_avg_rank_21_bf16.safetensors
kino time
>>
>>109548538
How do you avoid slow motion slop like this?
>>
>>109547997
LOL saved.
>>109548061
Agree.
>>
>>109548575
does it mean it will follow prompts if using with pruned?
>>
>>109548106
here ya go anon
i also added the most important part
https://files.catbox.moe/cr0ak7.mp4
>>
>>109548575
>Kijew has one now
mother of a fuck, somenoe post results from this so we can compare to the lightx2v.
>>
>>109548575
why use the 4step when the 8step isn't that much slower?
>>
I can't stop living out my fantasies with first person pov h3 gens
>>
>>109548553
>also y'all share your cunny prompts. i'm already 2 days out of ideas
Then get more creative because I know you haven't tried tentacles yet
Go look at sexy Instagram kids for inspiration, personally I'm just making intimate candid daddy daughter moments like dropping her off at kindergarten and kissing her goodbye

If you emphasize sultriness/seductiveness you can get lewder children than IRL (even with the power of a slutty Instagram mother) so you can do bratty sexy stuff too
>>
>>109548344
I'd start with https://github.com/ServeurpersoCom/acestep.cpp
for most clean lightweight solution,

then try
https://github.com/scragnog/HOT-Step-CPP
If you want more features
>>
>>109548635
Post some
>>
>>109548635
i bet that one guy who would photoshop himself into pokemon pictures is having a blast
>>
>>109548635
>I can't stop living out my fantasies with first person pov h3 gens
H3 is so good at this but takes so long to gen im running into the issue where sometimes I'd rather just remember the previous fap session's gen then wait 10 minutes for the potential for a better video
>>
>>109548635
true
public handjob is one of my favourite genre to self insert
>>
>>109548659
>public handjob is one of my favourite genre to self insert
Why? Handjob fetishes are either virgin or trauma, which are you (or both?)
>>
>>109548656
>so long
My n word. This is 3 times as fast and 10 times as good as anything we've had before
>>
>>109548367
>>109548373
Thank you for the news, friend.
>>
>>109548635
uptight women cosplaying... be still my heart
>>
>>109548575
how does this differ from the earlier lora on the huggingface, other than file size
>>
>>109548670
you just dont get the thrill of having your willy stroked in a public setting while everyone around you acts nonchalant about it
>>
>>109548672
Yeah I know but it's still long compared to shitting out sdxl or even anima goon slip
Like I want to generate drunk Finnish blondes in revealing bikinis RIGHT NOW but the genning time is putting me off
>>
>>109548635
Same, pic of my own dick POV plus literally any other woman has been running on my machine for days straight
>>
>>109548700
you don't gen for the current goon sesh

you gen for the next one
>>
>>109548687
I mean compared to minimax_h3_ref2v_turbo_4step_v0.1_comfyui_bf16.safetensors
>>
>>109548670
handjob fetish? lmao what. yeah bro I got a handjob fetish and a missionary sex fetish i'm so weird dude!
>>
File: debo_dm_k2_00086_.png (2.1 MB, 1872x1007)
2.1 MB PNG
>>109548674
always happy to contribute :)
>>
>>109548672
Yeah I know but it's still long compared to shitting out sdxl or even anima goon slip
Like I want to generate drunk Finnish blondes in revealing bikinis RIGHT NOW but the genning time is putting me off
Maybe I just gooned too hard

>>109548698
Yeah of course I don't. If no one cares are you even in public? The point of voyeur is they don't know they're being watched. The spectrum of exhibitionism runs from forcing people to watch to the potential of getting caught.
Your fetish sounds cuck coded, like free use. So I think you're a virgin and/or a cuckold
>>
>>109548635
share with the class
>>
Do you need to do some additional setup for kitchen attention? Keep getting oom on anima
>>
>>109548711
Handjobs don't look or feel better than a bunch of other things so yes, something must have happened for you to overvalue it. Missionary sex you can fetishize but I doubt you do (if you do that's a little virgin coded too)
>>
>>109548730
>oom on anima
that's possible?
>>
I don't want complex scenarios, I just want to add some motion to my (artistic) nude gens. What's the best/fastest model for that?
>>
>>109548735
kek, you're mad retarded. you sound like a retarded therapist trying to get a kid to troon out
>>
>>109548736
Yeah idk. It just keeps swallowing more and more memory until it ooms without ksampler starting
>>
>>109548742
h3 is better at almost literally every video task anyone has attempted so far compared to every other model. The one thing it doesn't really do well of the box is specific sex acts and genitals (although a reference helps).
>>
>>109548730
Aint no way diddy blud
>>
>>109548730
see >>109548151
>>
File: 1761922658743001.mp4 (220 KB, 448x448)
220 KB
220 KB MP4
>>109548730
>Keep getting oom on anima
>>
>>109548756
>>109548758
My launch script always git pulls the latest
>>
>>109548716
*night cat*
>>
>>109548604
Oh shit, thanks. Will be probing this imminently.
>>
>>109548751
Stable Video Diffusion was fun enough adding some movement to gens but the resolution was bad
>>
File: file.jpg (601 KB, 1536x1024)
601 KB JPG
>>
>>109548788
Update for the ref2v prompters: this worked out of the box by putting Meiya in there and describing her appearance in moderate detail in the subejct reference, and writing out more or less what happens in each shot of the video. I fudged timestamps but didn't check too hard whether they matched precisely.

After that I tried the same thing replacing Meiya with a corgi and it just spat the source video back out. Ref2V seems to need a little bit of coaching to work, but when it works it works.
>>
>>109548815
Gemma is gifted.
>>
>>109548575
ok I tested at 8 steps with my previous downfall meme text. seems good? idk how it differs from the original other than file size, it isnt 2 gigs which is nice.

https://files.catbox.moe/xpmgor.mp4
>>
okay i think the camera lunge looks a lot better if you say she suddenly appears in front of the camera. now just gotta fix some other oddities like the disappearing tentacles and the gun firing after it had already fired in the final cut
>>
>>109548725
how am i the cuck if i'm sitting at a table in a restaurant while the waitress reluctantly gives me a handjob with the same detached, emotionless mechanicalism as if she were masturbating a horse, and everyone around me, including my wife and her friends, just continues their conversation as if nothing strange is happening?
>>
I just keep getting this error now. Comfy why did you do this too me?
[ERROR] ERROR lokr diffusion_model.blocks.1.mlp.down.weight Allocation on device 0 would exceed allowed memory. (out of memory)
Currently allocated : 832.96 MiB
Requested : 192.00 MiB
Device limit : 15.60 GiB
Free (according to CUDA): 16.44 MiB
PyTorch limit (set by user-supplied memory fraction)
>>
>desert, japan: :O
>>
>>109548868
japan doesn't have deserts. its an island
>>
>>109548877
we can fix that
>>
>>109548639
Thanks. What about adapter? Which one I should choose? I downloaded your model, scrag vae gguf. But acestep.cpp doesn't launch because
> no usable pipeline, synth missing: Text-Enc
>>
guys, i just thought of a revolutionary idea.
imagine an asian girl, but she has big boobs. thoughts?
>>
>>109548901
it would never work
>>
>>109548816
>After that I tried the same thing replacing Meiya with a corgi and it just spat the source video back out
Yeah that's what was happening to me all the time.
I notice you're using syntax I'm unfamiliar with in your prompt, like [FILL IN: ...] - is this officially supported? I didn't see anything about this in the documentation.
For comparison's sake, this is more or less what ChatGPT produced after I had it analyse the video: https://rentry.org/ts22cr4t
>>
File: MiniMax_H3_00105_.webm (2.97 MB, 864x864)
2.97 MB
2.97 MB WEBM
>>
>>109548877
Do you think islands can't have deserts? Have you seriously never heard of a deserted island?
>>
>>109548106
failed with random realism big boob miku but worked with widowmaker model sheet
>>
>>109548913
slop
>>
>>109548913
No one except Will Smith should eat spaget
>>
>>109548922
Nice. Did you use [FILL IN:] too? Would be nice if video ref worked a bit more consistently.
>>
>>109548935
>[FILL IN:]
qrd
>>
https://huggingface.co/4BEraser/BArtstyle-Blue-Archive-Artstyle-LoRA-for-Krea2
>>
>>109548926
it's noodles
>>
>>109548824
Gemma...
>>
oh dear....people are still making image model loras when the minimax image model is on the way that will eliminate the need for loras?
>>
File: debo_dm_k2_00090_.png (1.79 MB, 1872x1007)
1.79 MB PNG
>>109548780
not convinced thats a cat, but thanks for the prompt idea
https://suno.com/s/xLfYua5CTfgSp1Kl

>>109548926
drastically low number of will smith spaghetti gens lately
>>
>>109548635
...I'm just genning lewds of my ex.
>>
>>109548911
>syntax
whoops yeah I noticed that after the fact, I gave my toaster the basic idea for the gen and asked for template to fill in. My prompt improver has a tendency to overdescribe things and I wanted the prompt to contain only things actually shown in the video. I forgot to delete it before I gen'd but I guess it didn't matter.

FWIW I just tried again with Mumei and it didn't work, so maybe the problem is that only some anime girls are slutty enough to dance on a box
>>
File: file.png (584 KB, 1467x1887)
584 KB PNG
>>109548935
>>
>>109548963
Thank you for replying...
My song tonight is this:
... ?v=UPlP54CS0eo
I can't paste music as it's terminal only lol
>>
>>109548965
I'm also genning lewds of this guy's ex
>>
>>109548988
I'm genning lewds of this guy
>>
>>109548961
I bet will be making loras for that too
>>
File: debo_dm_k2_00092_.png (1.97 MB, 1872x1007)
1.97 MB PNG
>>109548979
got it, https://www.youtube.com/watch?v=UPlP54CS0eo

terminal posting is hardcore lol
>>
>>109548965
I mean for what they are likely to be like that could have been done in Wan or LTX. There is a lot of uninspired genning going on, not exactly pushing the boundaries of H3 can do. Having some static image of some shitty e-celeb talking and moving his hands a nit is pure garbage.
>>
>>109548974
It's so funny that the gigantic super-precise replacement prompt and "Side angle shot. She presses her finger to her lips seductively" both work and both do the same thing, but also both fail sometimes. This model really is a mystery huh
>>
>>109548961
Have you seen what minimax thinks a pussy looks like?
>>
>>109549012
Thanks, this is actually sharper than the local record I have on my disk.
>>
chadcel:
>sleeps with maybe 2-3 women a week
>has to go out
>has to socialize
>limited to 3DPD
>forced to participate in the social hierarchy
>has to learn "game"
>has to maintain a social life
>has to spend money on dates
>has to deal with rejection
>has to compete with other men
>limited by geography
>limited by the number of hours in a day
>genetics hardcap his potential

chadgenner:
>can gen videos of himself having sex with any woman in the world
>can gen sex with hundreds of different women while he sleeps
>never has to leave the house
>never has to get dressed or learn "game"
>not limited to 3DPD
>zero rejection rate
>doesn't need to participate in the social hierarchy
>can be in 50 different scenarios simultaneously
>wakes up to more Ws than chadcel accumulated all year
>genetics replaced by inference compute
>chadcel maxxed looks, chadgenner maxxed VRAM
>>
>>109549019
Give it a nice pussy as <subject 1> 's pussy but don't get too close, I've seen horrors.
>Minimax Music.
It sounds good but the style knowledge is limited. If it trains well, it might be useful.
>>
>>109549031
>zero rejection rate
we're literally just talking about how half the anime girls don't want to get in our prompts, anon
>>
File: file.jpg (401 KB, 1536x1024)
401 KB JPG
I think about you.
>>
Has anyone messed around with those continuous workflows? How are they and how are the transition points?
>>
>>109548604
how did you get the text to direct the video, I need a good video to h3 text format
>>
>going from 0.4M video to 1.0M increases gen time x4 times in H3
Man...is this supposed to happen? I thought 1M is its native resolution
>>
>>109549039
This has typical slop generation adjectives.
>>
>>109549046
Asked for a template for a five second video and wrote the direction myself unfortunately. For most gens I have a system prompt written by Claude who I fed the guides to, but it isn't perfect when I want to be precise.
>>
>>109549031
Mounting an auto jerk off fleshlight to the edge of my desk i am unstoppable
>>
>>109549031
well, until it can work well in VR this is only really going to be porn but self self made content rather than the pretty boring done it all before stuff available on porn sites. Other than porn then any other things like - god forbid hand holding is again not going to be realistic without VR.
>>
>>109549070
>h*nd h*lding
get the FUCK out
>>
>>109548751
Are you only talking about local models? Because seedance 2.5 mogs.
>>
>>109549031
This is basically my workflow. Only thing I need now is fully local video-to-4d gaussian splatting for my vr headset. Anyone know a solution?
>>
File: MiniMax_H3__00068.mp4 (2.87 MB, 928x672)
2.87 MB
2.87 MB MP4
>>109548490
>>
>>109549048
yes, that's easily over 2x more pixels for whatever duration you've set it to
>>
>>109549092
hag juice?
>>
>>109549070
hopefully we'll one day get to the point where on-the-fly reactive VR generation is a thing so I can experience (artificial) love before I die
>>
>>109549070
Physical sensation is ultimately just a chemical reaction in the brain, which could theoretically be hijacked.
It's the same way a dream can feel completely real even though it literally isn't. Surely at some point in the future we'll figure out how to replicate or manipulate those sensations artificially
>>
>>109548913
inaccurate.. everyone knows the asians slurp the shit out of their soup as loudly as possible
>>
File: 07371-2006223619.png (417 KB, 640x480)
417 KB PNG
>>>/wsg/6213784
>>
>>109549092
why did LDG brew age her?
>>
https://github.com/Comfy-Org/ComfyUI/pull/15439

noice
>>
>>109549080
I mean yeah what's the subject of this thread anon
>>
>>109549125
?
>>
>>109549121
oh shit is this going to be a game changer or is it just going to be as finicky as ref2v
>>
>>109549120
because ldg is bad for you
>>
>>109549080
Sneedance probably does well in the way, but it's not really going to be creative, it's tuned to output good quality but pretty safe output and that is putting aside the sfw restrictions. It would be interesting for a long term comparison between that and H3 just to see how they hold up, and talking about H3 on local and not commerical
>>
>>109549113
Humans still have an energy body.
>>
File: M2AaqgK.gif (1.41 MB, 350x272)
1.41 MB GIF
>>109549092
its probably an impossible ask but I'd love to see this replacing picrel somehow
>>
File: 00025-1793220125.png (2.63 MB, 1920x1280)
2.63 MB PNG
>>109548244
this guy celebrities loras for krea 2 are soo undercooked and very bad. Very sure he rushed the entire process and used low res images of his old dataset.
>>
Are big ass text encoders the future of image/video generation?
>>
>>109549106
Elon Musk selling $50k brain chips that direct your gens directly into the brain versus Nvidia 7090 at $20k or Blackwell 8000 at $60k. Decisions, decions
>>
>>109549098
It's the same duration. Pixels increase x2.2 but gen time increases x4 times
At this point i don't feel like I want to try its native resolution and steps
>>
>>109549140
a dream is just as real as your waking life
>>
Damn, I wanna gen normal 1girl stuff but I'm stuck in this gay ass dickgirl commission, I only accepted it because of the money but holy shit why are they such faggots, they want long dicks hanging in front of their faces, cant be erect, giant balls, plus the "girl" with big tits has to look really pretty.

Also the guy is such a fucking homo, always talking in innuendo like trying to make a pass on me, talking really slutty, I always write him off with an uncomfortable "lol", I said dude I'm straight as an arrow, stop
>>
>>109549143
no idea how to do that
>>
>>109549169
>Pixels increase x2.2 but gen time increases x4 times
So quadratic, exactly what you would expect?
>>
>>109549180
sure you are
>>
>>109549180
lol people pay for that stuff?
>>
>>109549148
he uses little to no caption and claims that's the best
>>
>>109549177
Not as real, different.
As long as you know that your energy body is more than your 'brain'.
That conquers anything what western academic knowledge tells you otherwise.
>>
>>109549180
wait, people commission for AI gens? Where?
>>
>>109549180
oh god this hits way too close to home I always get these weirdass commissioners too
>>
>>109549180
futafags truly are the worst, yes
>>
File: 1760149190441222.png (224 KB, 412x607)
224 KB PNG
i may have gone too far in a few places
>>
File: 455511551544.png (36 KB, 395x678)
36 KB PNG
Minimax music testing take 2, it seems to do much better with agent skill loaded on Claude, but still not super accurate
Melodic house/EDM
https://files.catbox.moe/1ut189.mp3

Jap 1980s city pop
https://files.catbox.moe/z7uy05.mp3

>>109548899
Did you download the Q8 4B LM? It should look something like pic rel
>>
>>109549180
>>109549211
Where do you guys find clients
>>
>>109548485
yeh that's one a downloaded all the Krea2, Wan and LTX celebs. Of course tech moves fast and so the Krea2 are really the ones. Wonder how many H3 Loras will be trained seeing as it does a lot already. Certainly as he is clearly Polish there will be some he will want to do. And who does not want Iga Swiatek in a H3 "situation"?
>>
>>109549226
xitter
>>
Guys, can you give LTX 2.5 a chance? It might impress you.
>>
>>109549202
the brain is all there is
>>
>>109549224
With a LoRA, there's no doubt Minimax Music will shine due to its better arch
>>
>>109549234
rofl
>>
>mfw accidentally putting your penis in your android gf's garbage disposal instead of her synthetic vagina during sex
>>
>>109549224
See if it can do a fusion between Eiffel 65 and A Flock of Seagulls.
>>
>>109549224
I tried it a couple of times, then immediately went back to SaaS.

https://files.catbox.moe/ih7bdh.wav
>>
File: 1782066696979956.png (439 KB, 1024x960)
439 KB PNG
The new R2V Turbo Lora is for R2V model right ??
All this time i use I2V model for R2V nodes
>>
>>109549234
LTX 2.5 be like
>>
>>109549238
For your chemical addled body, maybe.
We all have the same faculties to test out extra sensory feelings.
>>
>>109549268
this guy's brain has no chemicals!
>>
>>109549189
>>109549206
yeah everywhere, if you post it on any social media and you get good engagement, you always will get users asking for commissions, specially perverts with money who rather pay up than do it on their own, I've done really weird stuff, for example I had this client who paid for a obese woman with hairy bush, armpits, tits hair, posing naked on a couch, she had to be brunette, poor guy, now that I think about it, he was probably abused by a woman like that when he was a child, people with fucked up fetishes normally get it when they are young by shock value

>>109549215
this one is a furry/futa fag, he is usually nice, just too much of a faggot with his double entendre and always trying to make a pass
>>
>>109549273
Grow up kid.
>>
File: videoframe_159123.png (1.91 MB, 3840x2144)
1.91 MB PNG
>>109548961
i don't expect minimax to seriously train their image generation model deeply on media content from anime, videogames and cartoons. loras will always matter and cooking them right matters too. Hopefully the model is not stupidly bloated and ultra slopped like flux 2 dev.
>>
File: gukmugjk,mgujk,ugfjk,ug.png (2.27 MB, 999x1019)
2.27 MB PNG
that sneaky little slut prompt generator the anon made sneaked in a close up shot of my clown girl gen i'm doing, got jumpscared seeing her hyper realistic every-fucking-pore tier 4k up close shot. It's fucking crazy how good h3 can handle fully reproducing likeness when asked. the lightx2v lora really does not skip a beat.
and this was with the v0.1 lora i haven't tried the new 1.0.
(stuff mentioned before people ask for links:)
https://github.com/whp199/GemmaPrompt
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main
>>
>>109549281
kill yourself magical thinker
>>
>>109549268
It's also possible to hijack your brain's perception to the point where it no longer thinks or feels that there's anything beyond the brain itself
>>
>>109548041
it's not devilry, it's DiGiorno!
>>
>>109549261
Me too and I am not sure if the r2v-lora is any better.
>>
>>109548011
>Expects safe for work images on 4chan
IMHO it's actually totally appropriate
>>
>>109549279
You'd have to pay me pretty handsomely to do that kind of work.
>>
>>109549293
Little kid.
>>
ok, this worked for a video swap of generic image 1:

subject_definitions:
<Subject 1> is the man in <Picture 1>. Match his face, hair, skin tone, body type, and overall appearance precisely to <Subject 1>. Keep his identity consistent throughout the entire video.

summary:
<Video 1> is the source video for the target video edit. Replace the character in <Video 1> with <Subject 1>.

retention_analysis:
The background environment, lighting, camera framing, desks, and shadows must be preserved exactly from <Video 1>. Inherit the original performer's movement, position, head tilt, and pacing frame-by-frame.

detailed_description:
The target video matches the style of <Video 1>. <Subject 1> replaces the character by occupying their exact screen coordinates. <Subject 1> mimics the original motion exactly, including all expressions and gestures. Do not introduce new movements, cuts, or background changes.
>>
>>109549282
you could legit nudify namine there just fine by using a pre-lora'd nude image, or just a ref image of namine + a nude body with the head not included. literally just prompt that she gets up from the chair and slips her dress off.
that's how good H3 is.
>>
File: file.jpg (679 KB, 1536x1024)
679 KB JPG
I am entertaining underage posters in this thread. Yellow tape means do not enter here.
>>
>>109549264
I tried LTX 2.5 and despite maybe it wasn't set up right for all the setting it is still bad and even if it was an improvement on 2.3 then there really is not a use case for it. It's 8 steps as LTX 2.3 but it was slower and has anyone really shown what it can do - has anyone even cared to do so? So again H3 can run better on 8GB GPU's so what exactly is the point of it seeing as it runs worse than H3 anyway never mind the model being far far far worse anyway? It can still use the Loras as for as Io know from 2.3, but that isn't going to save it. LTX was a weird model, was a substitute of Wan due to longer gens, faster and had the audio but that has all been solved by H3 and with the model being far superior. It's a model that was DOA before it was almost conceived.
>>
>>109549315
get new material
>>
>>109549340
Just wait for your evening medication at the psych ward.
>>
>>109549336
Ask a chatbot to reword that post so you don't sound like a third world-er and make it so that we can actually understand you.
>>
>>109549356
you're the one seeing things that aren't there
>>
File: 1775700133348791.png (179 KB, 918x1519)
179 KB PNG
>>109549224
Yes, now I can launch it. Could you also please share your settings? Because my output... Well... It's nowhere as clean as yours.
>>
>>109549357
>use a LLM after drinking half a bottle of wine (currently) and a lager earlier today
maybe it will just laugh, they are pretty advanced now
>>
>>109549359
I ain't the one who has a prescription.
>>
>>109549194
that's what i suspected. I hate how bakers keep half-assing the basic captioning process and expect the lora to be functional and versatile. The lora section for krea2 on civitai is full off bad loras with very terrible captioning trigger information and its no wonder why they don't work for shit regardless of strength of lora weight used.
>>
1980s rock skill-enhanced Minimax Music gen
https://files.catbox.moe/oeqfj1.mp3

>>109549224
Also from what I see on my folder you also need Qwen3-Embedding-0.6B-BF16.gguf
(Not shown there), you can keep LM off (its code box empty)

>>109549251
I don't think it'll be that smart in recognizing artists kek, but I'll try
>>
fp32 model is so fucking fast compared to the other audio models
>>
>>109549380
only prescription I have is for one pussy (your mother's) to be taken daily
>>
One thing I give this model credit for is that it give bi lingual singers an accent when speaking.
>>
>>109549383
pretty good. can you edit music with it? did you provide lyrics?
>>
File: file.jpg (548 KB, 1536x1024)
548 KB JPG
Aitght mate. Let's see...
>>
>>109549431
I don't need one m8.
>>
>>109549431
we have mute coppers now?
>>
The turbo lora for ref is worse than the non ref lora.
>>
>>109549327
scratch that, this is better

<Subject 1> is the person in <Picture 1>.

Replace the man in <Video 1> with the man from <Picture 1>. Match his face, hair, skin tone, body type, and overall appearance precisely to <Picture 1>. Keep his identity consistent throughout the entire video.

Preserve the original background, environment, set design, and scenery from <Video 1> completely unchanged. Maintain the original background depth, objects, and spatial layout, ensuring only the subject is swapped.

Follow <Video 1> strictly for all action, timing, sequence, body movements, hand movements, exact facial expressions, eye direction, head movement, lighting, camera angle, framing, and overall motion. Replicate every micro-expression and emotional state from <Video 1> with absolute fidelity. Do not alter or invent any actions.
>>
gonna be honest guys
I already stopped genning with H3
I got all the "saar pls show her bob" gooning out of my system and now I'm bored again
>>
Can Gemma generate some hot Miku jeets?
>>
Guys, is it time to bring back GenJam? We're all using Minimax so why don't we have a videojam?
>>
>>109548974
again did you type this manually up or got a good workflow to tag a video automatically?
>>
>>109549462
I never even started.
>>
>We're all using Minimax
>>
>>109549467
by the time i see what the jam is about and get a video the thread is over
>>
>>109549436
Because you are a faggot.
>>
>>109549475
gen in .3mp, it's enough for memes
>>
>>109549481
nah
>>
>>109549484
but anon...
>>
>>109549487
You can't escape the statement. Post your gen.
>>
>>109549370
My settings are 50 steps, DiT-only (LM disabled ) to enhance creativity, DPM++ 3M sampler, CFG of 12-20 (mostly 20 since LM is off).

This will improve the results on merged model, but it still won't be as good as the LoRA results I shared. For results generally as good as the outputs I shared you need to train a LoRA on Base XL model
https://rentry.co/s8fg8ber
No need to use the long caption script from my training guide, and it doesn't work for everything anyways since Genius doesn't have lyrics for every artist.

Grab lossless files from your fav artist or set of artists (for genre LorA), make it around 20 songs (anywhere from 10-25 is fine), in the same folder have lyrics.txt alongside each file, format them in a way ACEStep XL expects, then make the .json and pre-process it with Side-Step as I shared there.

Try LR 0.0001 or 0.0003, rank 64 with 128 alpha, or 128/256.
Modal gives free GPU credits ($30) a month which works through several training runs, you'd upload the pre-processed tensors there then run a headless training script. I trained both LoRAs without 60s chunking, but 60s chunking can save training time and allows for more training runs (at the cost of perhaps less structure in songs).
>>
>i'm still here making different anime girls DROP IT
>>
What should the theme for videojam be?
>>
What's even the point of genning something for a thread theme when we don't have collages anymore to showcase them?
>>
File: Krea2_turbo_hr_fix_00043_.jpg (3.4 MB, 2368x3544)
3.4 MB JPG
>>109549463
I used gemma to help with this
if you give me an idea I can make one for you but I think you can do it yourself
>>
>>109549495
I don't need to escape anything.
>>
>>109549496
Inference is still done on the 0.3 merged model with LoRAs. Also for Modal I'm referring to the H100 GPU.
>>
File: gefft.jpg (541 KB, 1536x1024)
541 KB JPG
Gefft. Again, this is better.
>>
File: MiniMax_H3_00069_.mp4 (1.2 MB, 1472x1472)
1.2 MB
1.2 MB MP4
>>109548106
>>
how do I stop blowjob loras from making my dick bigger? i'm average sized, but they keep giving me a pringles can sized dick when the female starts sucking
>>
>>
>>109549513
Saved.
>>
>>109549510
But you did. Please take your evening medicines.
>>
>>109549503
Theme should be your ex girlfriend
>>109549506
Calm down sperg
>>
>>109549519
penii grow larger when stimulated
>>
>>109548106
so what is the trick for a generic character swap, without using a 5000 word llm prompt
>>
>>109549506
But you don't make anything anon
>>
>>109549526
I stopped taking my meds and I'm more lucid than ever.
>>
>>109549503
weekly /ldg/ videojam? could be fun.
>>
>>109549535
The 5000 word lm prompt failed me.
These two got a working prompt (albeit not a 100% success rate): >>109548604 >>109548974
>>
>>109549540
Where did this get you? You are still here spiting other people.
>>
File: 00070-3835331322.png (506 KB, 640x512)
506 KB PNG
>>
>>109548922
>that leg up repeat at the end
why does it do that? I get that specifically on ref2v with video refs when I try to make it match the movement and it'll sometimes repeat like that at the end or sometimes at the start
>>
>>109549519
idk about the specifics of that lora, but it helps to have statements about things remaining consistent.

So for your retention_analysis: block you could probably add something like
<Subject 1>'s penis remains consistent in size and shape.
>>
>>109549549
we had genjam last year but certain autists couldn't hold back their autism and started throwing hissyfits over the theme
>>
>>109549565
Thanks for coming back,
https://www.youtube.com/watch?v=uukOdFdyqVQ
>>
>>109549308
R2V run worse too. Even when i only use Images as reference
>>
>>109549148
he has truly been an inspiration to make me figure out how to do that shit myself
>>
AR sampling takes the most time I think I'm going to test 100+ steps because I got that grown man gpu nom sayin?
>>
>>109549561
maybe it's time you took yours.
>>
>>109549567
okay i'll try that, thanks. im using the mm blowjob lora with the ref2va model
>>
File: 00084-3031421445.png (549 KB, 640x512)
549 KB PNG
>>
>>109549593
Thank you for replying.
>>
So is the new R2V turbo lora actually bad? And worse than the FL2V turbo lora?
>>
>>109549617
It's pretty good I think it gave the best results with the target model, I posted the samples in last thread and the result is OP
>>
>>109549616
no problem.
>>
>>109549626
It is somewhat sad that you are here...
>>
>>109549617
No, it's not bad at all. It's pretty damn solid. We just need that fucking upscaler and we're golden. First pass 10 steps with the lora then (unknown amount of steps) for the hires final pass.
Man i hope they don't rugpull us.
>>
>>109549617
no, it's actually good (lightx2v). use 8 steps (for both, even the ref one)
>>
File: 1658680757111.webm (953 KB, 1920x1080)
953 KB
953 KB WEBM
Here is another super interesting character replacement candidate for R2V. Just sayin...
>>
>>109549635
I am my own king and queen.
>>
>>109549635
it is but idc.
>>
File: file.png (122 KB, 1862x327)
122 KB PNG
>>109549617
this is the current meta for r2v
>>
>>109549535
As far as we can tell there's no magic sauce, I dunno what this poster >>109549513
did but so far we've seen two extremely different ways of prompting to get the replacement effectively.

In my experience you need to "guide" the model to understand your reference by pointing out important things and how they fit together, like when I attach a reference character I also describe their appearance in broad strokes (purple bodysuit, angular purple hair, dramatic eyeshadow, etc). But maybe that's placebo.

Not sure how well my solution works because it gave me a vaguely tomboy coded Spike
https://files.catbox.moe/fswl9b.mp4
>>
File: MiniMax_H3_00439.webm (1.71 MB, 896x704)
1.71 MB
1.71 MB WEBM
>>109549565
>>109549610
empty prompt
>>
>>109549645
god i remember this clip. ive blown quite a few loads to that one. i dont even like the character just jesus christ someone really made this shot gold
>>
>>109549653
>turbo lora
link?
>>
>>109549652
It was an impulse to make you to reply.
>>
>>109549653
No it's fucking not, sage gives more speed and better outputs if you use the fp16 model wtf
>>
>>109549666
nice 666
>>
>>109549663
https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras
>>
>>109549673
I am from the heaven.
Most people are confused.
>>
>>109549641
I just hope it won't be like 24gb vram minimum for upscale required
>>
>>109549616
The reference like face is accurate but it tend to hallucinate movement for me. My background didnt move at all when the camera moves.

Also its 1.5x slower than I2V model using R2V nodes. And a worse motion too. Dunno what im doing here.
>>
>>109549653
Explain to me why I should use "comfy kitchen" instead of sage attention w/ triton.
>>
>>109549682
not him but, it's native and there is no compression involved, so potentially more speed/less noise
>>
>>109549679
Do you want to fight.
>>
>>109549674
Whats the difference between the Lightx2v ones ?? Im confused
>>
>>109549653
newfags do not use this retarded workflow, only these down syndrome retards that don't know what kitchen even is use it. use this >>109549672
pick the fp16 model in the dropdown for patch sage attention. mem eff sage attention from kjnodes works nicely for me to reduce vram peaking but other than that you don't need anything else.
and again use the prompt enhancer an anon made, it works pretty much perfectly.

>>109549291
>>
>>109549660
it comes from an era where animators made an effort to imbue the animation with sex appeal. these days they just give the girl some honkers and call it a day
>>
Minimax Music Skill-Enhanced Japanese Shoegaze test
https://files.catbox.moe/0duqa6.mp3

Here's a corresponding (condensed caption) ACEStep 1.5 XL 0.3 Turbo/Base merge output for the same lyrics, no LoRA whatsoever (just removed the Russian part from the top which was a mistake)
https://files.catbox.moe/3i6eye.mp3

There's no question about it, ACEStep simply is better and significantly less slopped.
>>
>>109549679
no such a thing.
>>
>>109549681
Lies.
>>
>>109549698
I agree with this but they can also just launch llama.cpp and add the model rules to the system prompt
>>
>>109549708
That's your own personal thing.
>>
>>109549710
Well im doing porn so i dont post my result here
>>
>>109549729
You never gave anything.
>>
>>109549704
the audio quality a far better on minimax tho, Acestep can suffer from shifty quality like the 3.5 models in Suno
>>
is this a new form of schizoposting
>>
>>109549723
No, it’s a universal reality. If you’re talking about a state of mind then that’s a whole other thing.
>>
>>109549738
Right.
>>
>>109549729
that's why /vdg/ is around, it it's a celeb then post it in catbox
>>
>>109549740
On.
>>
>>109549737
Just make gens throwing him getting job applications and add it to OP to keep him away
>>
>>109549736
Yea, enhanced quality, but it's too slopped out of the box. That audio quality is not the best ACEStep can do, with a LoRA it's way better (I'm assuming it's worse out of the box due to alignment not being there yet). Though Minimax LoRAs could be promising, so far these are disappointing results.
>>
>>109549745
Ok ukraina / Israeli fuck.
>>
Does anyone have comparisons of the different loras?
>>
>>109548961
>>109549330
I can't fucking wait, bros.

>>109549282
Namine a cute.
>>
>>109549755
so, wouldn't' it be better to gen in Acestep and then use Minimax to embellish (remix) as per doing so in Suno?
>>
>>109549757
I am going to do you in.
>>
It just like v0.1 I2V turbo lora again
Wait for V1
>>
File: r2v_nodepile.png (168 KB, 944x1448)
168 KB PNG
>>109549698
anything wrong with my stack? i recently removed spectrum because i was told it does not play nice with the turbo lora.
should the turbo lora always be 1.0 in weight? i remember back in the wan days it was common to lower this a bit.
also i'm assuming it's important to keep the scheduler set to simple when using the turbo lora, but what about sampler? what do people here generally prefer between res_multistep, euler and er_sde?

(i'm using an rtx 5070 ti + 64gb ddr5).
>>
okay, image to video replace, this generic prompt works (edit man to woman if girl, etc)

[Reference Roles]
Image 1 is the subject identity reference ONLY. Discard, ignore, and exclude the background, environment, and setting of Image 1 completely.
Video 1 is the master environment plate and motion template.

[Layer 1: Master Background]
The final video must use the exact background environment, studio set, lighting, and pixel data from Video 1. The environment from Video 1 is completely locked and unchangeable. Do not use any background elements or colors from Image 1.

[Layer 2: Character Inpainting]
Extract ONLY the face, hair texture, skin tone, and body features from the person in Image 1. Insert this identity onto the subject in Video 1.

[Performance Track]
The new character must replicate the precise facial movements, hand gestures, and timing of the performance in Video 1.

---

example: https://files.catbox.moe/xe7dt9.mp4
>>
File: iChads.jpg (225 KB, 1846x648)
225 KB JPG
Mac Chud update (sorry droidjeets, this is a Chud thread)

Gen times are pretty shit but I have unlimited memory. About 66 second/it for a 5 second Ref to Image and .2 megapixels using the largest model and largest text clip
>>
>>109549780
You have redundant shit that slows down your gen times, you really should read what those nodes do
>>
>>109549776
works good at 8 steps, but it will be better at 1.0 ofc.
>>
File: nodes.png (287 KB, 1967x960)
287 KB PNG
>>109549780
you don't need the model patch torch node, this is the order i use. mem eff i usually place after patch sage attention.
>>
>>109549795
You're using sage twice, either use one or the other or you get the worst of both worlds
>>
>>109549801
It was explained here that node just patches the sage attention you already use, so it IS a full sage patcher itself?
fug
>>
>>109549801
Ok, and which one makes more sense? I keep getting told the Mem Eff Sage Attention is good if you need to cut down on vram. What is the usecase for cutting down on vram?
>>
File: 1760578003957600.mp4 (283 KB, 736x736)
283 KB
283 KB MP4
>>
H3 is slop
>>
File: file.png (53 KB, 266x128)
53 KB PNG
I was told to stop posting here. I will do it.
>>
>>109549780
use comfy kitchen and stop listening to the trolls. you have a 50 series, these idiots are probably running a 3090 or something and comfy kitchen doesn't work for them
also don’t stack the fp16 sage patch with mem eff sage
>>
>>109549825
>comfy kitchen doesn't work for them
holy misinformed KEK
>>
>>109548047
Doesn't work via lan?
>>
lmao, looks what happens if you dont initialize the swap on frame 1:

https://files.catbox.moe/1n4ls7.mp4
>>
That's it, are you guys ready for sexy jam?

The theme: use Minmax to generate the sexiest video you can imagine. Can be anime, 3d, live action, etc.
>>
File: hmmmm.png (25 KB, 329x333)
25 KB PNG
>>109549787
I have no idea what's going to happen
>>
>>109548047
entire thing reads like claude slop
>>
>>109548092
krea 2 turbo is the same shit over like shitty ZIT. using the turbo and base workflow together is kino
>>
>>109549853
alright does anyone have the image of the 'tree' with three lines
>>
File: ref_MiniMax_H3_00018_.mp4 (1.89 MB, 800x1056)
1.89 MB
1.89 MB MP4
>>109549781
thanks

source:
>>>/wsg/6191829
>>
>>109548856
Because there's no emotional connection and she does this to everyone and you have no problem with it because you're a gross cuck, cuckie
>>
>>109549769
Possible, but no idea if there's a remix option. ACEStep can do better if I just change the sampler or get a different seed (in this case the sampler was the default, Euler), and the 0.3 merge is a bit unstable on some seeds and samplers, but it's safe to assume it can do the song that I linked with no quality issues.
>>
>>109549887
what are they doing? is this sex training? is there such a thing for guys? like just a whole bunch of dudes thrusting into the ground or something?
>>
>>109549645
do you want meiya because this is how you get meiya
https://files.catbox.moe/xyf1sl.mp4
>>
Why does /dmp/ have such a stick up it's ass over the use of AI music?

>>109538146
>>
>>109549920
>hair clipping through the seat
soulless slop
>>
>>109549920
Looks more like a headswap. What about on-model Meiya? Also there shouldn't be nipples in the target video.
>>
My result feels blurry on R2V Turbo Lora. Gonna try to increase to str to 1.5
>>
>>109549940
You should increase res or steps
>>
>>109549924
luddite mentality, they think all music should be performed by "artists" despite the fact samplers exist, over production exists, auto-tune exists and AI won't actually hurt real musicians just like classical music wasn't hurt when the gramophone was invented
>>
>>109549942
Same thing with 8step and 0.8mp
>>
>>109549940
1.0 str, 8 steps. dont use at 4 or it will kinda be blurry.
>>
ChatGPT told me it's better to use three different reference images of the same character (front view, back view, 3/4 view), rather than a single larger image that compiles all of them together (like a reference sheet).
Is this true?
>>
>>109549952
Try 6, I've noticed with the other models going too high messes with stuff
>>
>>109549940
minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensors
This one at 1.25 and 8 Steps gives me the best refs so far. The newer ones just doesn't compare idk why
>>
>>109549960
they're contained in <Picture N> so it couldn't hurt, there shouldn't be any ambiguity or bleed unless you specify so in the retention field
but I'm curious why it says 3/4 instead of profile
>>
>>109549658
Did you try OCRing anon's prompt here? >>109548974
>>
>>109549940
if you are using spectrum, remove it, sigma shift is still 12/3
>>
>>109549924
Because they're idiots. Nobody is taking their fun away from them however I do understand banning fully genned AI songs because then their thread would be flooded with AI rather than people who enjoy the hobby.
>>
>>109549801
alright i just ran this removing my patch sage attention node and keeping the memory efficient one, memory efficient was slower by almost 20s/it lol i guess i'll just remove that shit and stick with the fp16 model sage attention.
also on the same seed, noticeably worse quality than the fp16 model.
>>
>>109549957
I get good results with 10 steps for turbo. I was still getting some blurry distortions with 8 steps for whatever reason.
>>
>>109549924
>"AI" includes all neural network-based models and not hard-coded automation/procedural generation.
tfw this bans Synplant, the coolest VST ever created

I get why a musicfag would be against AI, it's not yet good enough for trained ears to be fooled. That being said, using generated tracks as samples is actually a pretty neat usecase that can circumvent "this sounds like AI".
>>
>>109549960
I had issues where multiple images of the same character would be somewhat treated as keyframes.
>>
>>109549977
>Reference images
>I do think the extra views are worthwhile for this one. In particular:
>Front + 3/4 + back is what I'd use first.
>The back view is unusually useful because the choreography explicitly asks for significant rear-facing footage. Unlike a 3D character, where the model can at least infer that the same modeled object is being rotated, a 2D anime character's rear hair shape, clothing construction, accessory placement, and silhouette have to be inferred from drawings. Giving H3 an actual back-view design removes a lot of ambiguity.

>I wouldn't add a side view immediately. Three coherent references give it plenty of identity information without drowning the conditioning in still images. If profile shots subsequently look inconsistent, then add the side view as <Picture 4>.

>And because of everything we've learned from your previous H3 tests, I'd make the turnaround images as neutral as possible: plain white background, relaxed neutral stance, no dramatic hand gestures, no elaborate props, and consistent scale/cropping. That minimizes irrelevant composition and pose information for H3 to latch onto while maximizing the information we actually want: the character design from multiple angles.
>>
>>109550009
>Synplant, the coolest VST ever created
need more devs to train a neural network to be able to recreate sounds in-synth
>>
>>109549704
>>109549910
Same ACEStep result with Stork 4 sampler
https://files.catbox.moe/plp02t.mp3

From what I understand, this sampler has the best quality
>>
>>109550009
Why not use AI where it shines, electronic music?
>>
>>109549982
I only use Comfy Kitchen. No matter what i do its still blurry and i even up the res to 0.8
>>
>>109550046
samper?
>>
>>109549821
I understand why it filters coping promptlets and it's well deserved.
>>
>>109550046
same, model -> lora -> comfy kitchen node -> sigma

even at 0.4 it works fine at 1 strength, try update folder cause there was a ck update, also some nodes were updated
>>
>>109550055
The model is garbage.
>>
>>109549960
How would it know that? Hypothetically there's a limit to an image size so separate images would let more details through, but would be unnecessary if the design isn't **super** detailed with stuff you'd need to zoom in for.
>>
>>109549973
Only using 1.0 strength, but I've noticed the same thing. Ref2AV gets all wacky with the more recent lightx2v turbo lora.
>>
>>109550064
are you talking about ltx?
>>
>>109550055
the only thing that has confused me so far is video swapping. otherwise, reference (image and sound) is a ton of fun.

https://files.catbox.moe/mykgte.mp4
>>
>>109549924
Because they don't want to be flooded with low effort AI submissions. writing an AceStep prompt doesn't make you a music producer. sorry.

It's like joining a painting class and instead of painting you just pull out your laptop and start prompting.
>>
>>109549821
good morning gods chosen
>>
>>109550078
Minimax H3.
>>
I'm worried, putting aside what unethical practices may have been employed by the Chinese to produce AI models that make all western ones obsolete, how exactly can the west compete? It's not as if the west has not employed unethical practices for a long time, is this how it may go down going forward? Flock cameras to study all motions then fed to AI models to train?
>>
<Picture 1> is the physical reference for Forsen.

include the forest and wooden cart from <Picture 2>.

Forsen and the blonde nordic man in <Picture 2> are sitting beside each other. Forsen says to the nordic man "Hey, when do we get to sweden?". The nordic man says "What the fuck is sweden?" in a swedish accent. Then, a dragon flies overhead and shoots fire over the cart.

2 images, ref model.

https://files.catbox.moe/uh881f.mp4
>>
>>109549513
this got positive feedback on /trash/
>>
>>109550084
If it sounds good, it sounds good but I get you. Prompting a song takes like 1/100th of the effort that writing one does.
>>
>>109547997
teto is shit as a waifu ngl
>>
writing a kino video or image prompt does make me an artist thoeverbeit
>>
File: Never skip penis day.webm (3.67 MB, 720x1280)
3.67 MB
3.67 MB WEBM
>>109549916

Yes, that exact thing exists for men.
>>
>>109550117
>Swayden
>>
>>109550122
>>109547997

She is. Its westernfag safe teenager. They only like her because she is 31 years old and act like 15 year old. Japs dont really like her that much
>>
>>109550125
Don't do this bros. I did this on roids and now my prostate is the size of a cantaloupe.
>>
>>109550125
LMFAO nice
>>
>>109550122
>drills
>fat ass from all the baguettes
whats not to like
>>
>>109550124
i sometimes spend 3 days on the same prompt
>>
>>109549926
>>109549930
yeah it's not great is it. Getting precise with ref2v seems like it requires some finesse and experience or maybe just luck and seeds. Anon's turbo autistic "perfect replacement of object in physical space" got me this, which is a lot better I think, but isn't really the reference video any more (and the hair is still having issues).
https://files.catbox.moe/l0cot4.mp4
>>
>>109550125
this makes me feel healthy and i goon daily right until before i blueball myself
>>
>>109550121
>Prompting a song takes like 1/100th of the effort that writing one does.
after you get good enough at producing music you stop needing to exert any effort. if you watch any notable producer (or songwriter) do his thing, you'll see it's as if the music comes from the ether. as if the producer is merely a vessel for some higher power. unironically.
>>
>>109550165
no matter how good you are at it it still takes more time than prompting. even if you can improvise the whole damn thing.
>>
>>109550165
when you see "professional" artists take years to slop out a few songs and announce it's their best work then you know already AI has bested their garbage
>>
>>109550165
im a music producer. this is true
>>
>>109549930
>Also there shouldn't be nipples in the target video.
How do you prompt to exclude something? Because in my experience when you prompt "NO nipples" it'll make nipples appear.
>>
Plus there's stuff like mixing and mastering.
>>
>>109550186
NO sex. NO big butt. NO cat girls
>>
>>109550186
prompt the video in a way where nipples wont show up. make it clear that it's a suggestive yet sfw work
>>
>>109550176
when youre having fun or invested in a track, time is meaningless. also time =/= effort anyway. i can exert more effort in 10 minutes jerking off than i would edging myself for 3 hours.
>>
File: 1781428038536824.png (96 KB, 340x296)
96 KB PNG
Lol i tried the I2V 8 step lightx2v lora and the blur are gone. R2V turbo lora are undercooked
>>
>>109550186
you have to prompt for big juicy nipples. remember the model doesnt like explicit terms so if you add one it'll censor it and not give you them
>>
File: MiniMax_H3__00069.mp4 (3.65 MB, 800x800)
3.65 MB
3.65 MB MP4
>>109550125
lmfao
>>
>>109550216
While it’s not exactly the same thing, time definitely adds to the effort.
>>
>using 2-4s is bad formatting

ok, this is what the guide says:

[Shot 1] The clip opens... (Describe the starting action here, no timecode needed).

[Shot 2] At 00:02.000, the subject... (Describe what happens at exactly 2 seconds).

[Shot 3] At 00:04.500, the camera shifts... (Describe what happens at 4.5 seconds).
>>
>>109550165
Yea but tbf compare an a real artist with an AI artist. The quality of the output is the same or at least almost, but the talent required to get there is not. That's the advantage of AI. Imo AI bridges the gap between novice with no skills and expert. Someone good at prompting and training models is just as good as someone who's good at producing their own music.
>>
So is it just me or is wan2.2 significantly clearer, with more detail, for both t2v and i2v? H3 is way more capable in every other way, but it does seem blurrier no matter what I do.
>>
i've been stroking my penis for over 10 hours thanks to H3, and now it's somehow hard and soft at the same time.
it also keeps oozing water every now and then. it's been like this for a while, and i'm starting to get kinda worried. what does this mean?
>>
>>109550240
The problem is that these AI models aren't good enough at following complex prompts yet.
>>
lole
>>
I am but a vessel for 1girl large breasts loli gens. The lord moves my fingers to type these tags into the prompt box.
>>
>>109550230
muh dih leaked a white liquid seein that .mp4 thank you for continuing that gen for me
>>
>>109550244
Dont use cope nodes and gen at recommended scheduler and steps
>>
>close the terminal after using comfy
>entire system freezes for a second
is this normal?
>>
>>109550240
AI is a tool like any other, creative people good with AI, or any app, will be superior to a random retard who can barely use any technology. You can do so much wild stuff with these models, it'd be foolish to neglect "all AI".
>>
>>109550240
yeah im not saying the opposite. if anything im just preaching to the choir that taste rises above all. plenty of producers whove dedicated their lives to it have shit taste and make shit music. the opposite is also true. often the most interesting music comes from "non musicians" anyway imho
>>
>>109550250
>complex prompts yet.
the duration of gens shouldn't need overly complex prompts. Putting in every detail, every single thing that makes it more complex is overkill. Let the model create things out of your own creative prompt rather then trying to issue instructions that the models has to abide by.
>>
GenJam Theme: music video
>>
https://huggingface.co/Kijai/MiniMax-H3_comfy/blob/main/loras/minimax_h3_ref2v_lightx2v_turbo_4step_v0.1_resized_avg_rank_20_bf16.safetensors

updated by kijai 1hr ago
>>
>>109550240
you would think so and yet when you look at a professional artists musical catalogue it pretty lame. You could put together their good songs on one greatest hits album and no more. AI get's the label of "AI slop" because of how easy it can best professional artists.
>>
Ok this music model is fucking ass.....
I don't know what they did but it's easier to get my point across with acestep and the outputs do sound worse because of it.
They need a proper guide because whatever the fuck they are saying is not working
>>
forsen

https://files.catbox.moe/vpcnyb.mp4
>>
>>109549148
The z image turbo ones work better than flux for me
>>
Are we postmaxing in this thread?
>>
>>109550274
True but making AI music only by prompting isn’t using the technology to its full potential yet, because the songs tend to end up sounding pretty samey and have that “AI sound.” Sample it, rearrange it, etc., to make it more unique.

>>109550294
Well it depends on whether you want to make something that hasn’t already been done a million times.
>>
Page 10 or bust
>>
>>109550355
i'm already BUSTIN' to these CLUSSIES my man
bustin makes me FEEL GOOD
>>
>>109550319
>Ok this music model is fucking ass
qrd?
>>
>>109550310
Whats the difference besides smaller file size ?
>>
>>109550312
Honestly, the top artists don't do any songwriting, that is delegated across a bunch of producers or a single pro who can use FL studio. Maybe if they work in EDM they actually compose the music themselves, but aside from those you will have guys who mostly have others help them. Even with all that help, the songs put out nowadays are just recycled versions of each other, AI is a way to break free from that as you can seamlessly blend genres (so far I've seen Udio do this well)
>>
>memes
>coom
Okay but show me the kino artistic video generations you've created
>>
>>109550363
The model just doesn't listen well I was able to nail the correct feel and vibe of multiple songs with acestep with minimum effort, this model just keeps fucking up even with llm assistance
>>
I feel like the lightx2 lora for the ref model ruins the prompt adherence. I haven't done a ton of ref model gens but it generally does a pretty good job of following my instructions. The lora at 8 steps just seems all over the place, some gens are really close and others are way off with the same prompt.
>>
>>109550319
ACEStep's DiT is 4B, but yea something's off about this model, maybe too much SFT and we didn't actually get the Base model (ACEStep is a combination of model releases, SFT and Turbo alone are also not that good compared to its Base or Base merges)
>>
>>109550353
>Well it depends on whether you want to make something that hasn’t already been done a million times.
of the gens I have seen so far (not looked on the interwebs mind you) the creativity is sadly lacking. I include myself in that as well.
>>
>>109550353
someone good with AI could use AI to make samples or loops and then make songs with it.
>>
>>109550383
It doesn't listen worth a fuck, I have never had a model fight me so fucking hard
>>
File: 1778932124745925.jpg (7 KB, 114x124)
7 KB JPG
Got body horror with 8step with new R2V turbo loras. Sorry bros ill wait for better turbo lora for R2V model.....
>>
>>109550383
here is how I would see it despite not using the minimax model. The V5.5 model on Suno is good but also pretty lame in that it will stick to certain styles probably because it is censored around around a lot of songs in it's training. The V4.5+ model is actually better, the quality may not be there but the songs genned are far better and are afar more varied and higher qulaity in their composition.
>>
>>109550379
what model?
>>
>>109550387
I believe that's the standard practice.
>>
we reaching 600 posts here?
>>
>>109550393
The body horror and reduced prompt adherence makes the turbo loras pointless because you would have regen anyway.
>>
>>109550412
>>109550412
>>109550412
>>109550412
Please....old bakers
Come back
>>
>>109550417
I would've gotten it at page 10 desu but yes i am still here
>>
>>109550417
fuck off, not page 10 yet
>>
>>109550425
bake a collage one
>>
>>109550434
Never will, also don't want to bake but I will if I have to
>>
>>109550444
you didnt have to...
>>
File: comfymikus.png (1.62 MB, 1024x1024)
1.62 MB PNG
>>109548557
>>109548490
>>109549092

Keep up the good work, I love ad renderings of things that do not exist or things presented in the shape of things that do not belong.
>>
File: MiniMax_H3__00070.mp4 (2.99 MB, 928x672)
2.99 MB
2.99 MB MP4
>>
>>109549795
>>109549801
and why are you using int8 checkpoint instead of nvfp4 if you have blackwell?
>>
>>109551709
nvfp4 is much lower quality
>>
File: 1756755913288268.jpg (124 KB, 1363x772)
124 KB JPG
>>109551709
clankerslop for reference
>>
>>109551833
it's too generous with stars. What's the point of 5 if there isn't a worst in class?
>>
>>109548344
his gens are retard
his recommendations are retard

you have to use comfyui for best gens
and you must know what to use and how

one hint, latest er_sde sampler comfy code plugged into custom sampler is one of many options

note:
comfyanon too is retard,
his original implementation was very good,
then he let retard code in,
then because of retard coupling comfyui has every time there is a code updoot related to llama,sd,ops and qwen impacted text encoder files generation becomes retard even more,
nobody gens acestep properly via comfyui since somebody would report this retard situation by now.
siutation is retard.
so you will get ai slop super clean sound with latest comfy code.
and it will change over time. because pulls are often messing it up due retard coupling comfyui has.
but even tho it is retard it will still have many advantages over everything else due to options comfyui offers.
so not everything is retard in that sense.
but it is mostly retard if you ask me.

also latest torch+torchao combo will have impact and make the generation retard.

o get good stuff out of this retard situation, use comfyui and once you get good sounds freeze it and use sidestep to train loras (no more than 20-25 songs per dataset).



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.