[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: 1767812146021601.webm (1.82 MB, 864x608)
1.82 MB
1.82 MB WEBM
Previous: >>109478939

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
>>109480276
>Inb4 deleted because of tranitor again
>>
Blessed thread of frenship
>>
>PokeGOD OP
ooh yeah thats the good stuff, yours gens have been some of the best.
>>
Okay, I'll try your thread, even though these days only the earlier bakes tend to stay up.
>>
>sexualized OP
>no collage
grim.
>>
What's the max length I can do with MiniMax? Can I do 30 second kino? Even longer?
>>
File: 200.gif (508 KB, 268x200)
508 KB GIF
>smith slop op
>pedo op
>>
aw hell naw chat we got a diddyblud ahh baker im crine rn fr :sob::skull:
>>
>>109480332
Idk, but you can always chain gens using references or first frames.
One thing I found weird is I was trying to use a video reference and my computer kept OOMing (24GB) when trying to do v2v over 6 seconds. Wondering if anyone encountered similar issues.
>>
>>109480332
If you have 128gb of VRAM with 1tb of ram you can do 30 seconds
>>
>>109480332
I think the official limit is 15s but it can do longer videos.
>>
>mfw Resource news

08/06/2026

>UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models
https://zhouhyocean.github.io/uniworld-view

>OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Films
https://xin1u.github.io/OminiVR_PAGE

>DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Models
https://github.com/Zhong-Chenchen/DIVE.git

>Multi-View Face and Gesture Animation with Dynamic Gaussians
https://dfki-av.github.io/MVFGA

>EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot
https://empaava.top

>Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation
https://github.com/Aoko955/Flash-VAED

08/05/2026

>Inline Studio v1.2.62 - Minimax H3 Lora training still only
https://github.com/inlineresearch/Inline-Studio/releases/tag/v1.2.62

>Qwen3-VL-32B-Instruct-MiniMax-H3-GGUF
https://huggingface.co/nif0/Qwen3-VL-32B-Instruct-MiniMax-H3-GGUF

>Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF
https://huggingface.co/nif0/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF

>MiniMax-H3-TAE: 2D tine VAE for MiniMax-H3
https://huggingface.co/Kijai/MiniMax-H3-TAE

>SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference
https://github.com/6somehow/DAC-SPADE

>CAPE-T2V: Captioner-Anchored Prompt Enhancement toward Two-Sided Conditioning Alignment in Text-to-Video Generation
https://github.com/yizzz927/CAPE-T2V

>JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion
https://github.com/jd-opensource/JoyAI-Video-Edit

>ParVL: Parallel Scaling and Expandable Compute Allocation for MLLMs
https://github.com/YangYangGirl/ParVL

>OliveGemma: A 3 Billion Visual Language Model for Recognising the Mediterranean & European Diet
https://huggingface.co/JamesZar/OliveGemma-3B

08/04/2026

>stable-diffusion.cpp adds support for MiniMax-H3
https://github.com/leejet/stable-diffusion.cpp/blob/master/docs/minimax_h3.md
>>
Has local video gen officially been solved? How can it realistically even get better than Minmax H3? How can BFL possibly compete with their upcoming model?

The sheer power and capabilities of r2v just astounds me, and we can generate 15 seconds natively on consumer-grade hardware. I just cannot see how this could be improved on in any significant way.
>>
>>109480336
I will take the cute anime girl
>>
>>109480353
Stop spamming and get a job loser
>>
>mfw Research news

08/06/2026

>When Diffusion Models Forget Who You Are: Identity Preservation in Face Inpainting under Large Occlusions
https://arxiv.org/abs/2608.04820

>HelloWorld: Enabling Socially Interactive Characters in Video World Models
https://arxiv.org/abs/2608.05070

>OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing
https://arxiv.org/abs/2608.05049

>ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing
https://guoxu1233.github.io/ContextMaster

>STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Models
https://arxiv.org/abs/2608.04887

>Simile Understanding in Text-to-Image Models: An Evaluation Framework
https://arxiv.org/abs/2608.04750

>ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation
https://arxiv.org/abs/2608.04436

>CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models
https://arxiv.org/abs/2608.04302

>Rethinking Pixel Mean Flows via Interval Denoiser
https://arxiv.org/abs/2608.04818

>Persistent Object Narratives for Token-Efficient Video Language Models
https://arxiv.org/abs/2608.04866

>Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models
https://arxiv.org/abs/2608.04349

>Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles
https://arxiv.org/abs/2608.04483

>Beyond Global Routing Aggregation: Phase-Aware Expert Merging for MoE Vision-Language Models
https://arxiv.org/abs/2608.04454

>When does training on downscaled images yield the same gradients?
https://arxiv.org/abs/2608.04448

>Unleashing the Potential of Vision-Language Models for Generalizable AI-Generated Image Detection
https://arxiv.org/abs/2608.04935
>>
>>109480332
I've done 30 seconds 1mp, had to have claude monkey patch comfyui's code though, it was breaking at 1mp gens longer than 12 ish seconds.
>>
based op
>>
>EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot
https://empaava.top

>The chat bot in question
>>
>>109480359
It can and will get better but we're at the point that it's "good enough" for tons of practical applications, WAN was too limited still.
>>
>>109480359
>How can it realistically even get better than Minmax H3?
faster gens, less melted details, better nsfw ofc (need to see if a good nsfw lora ecosystem develops for h3)
>>
>>109480336
>jewslop comedian reaction gif
into the trash you go
>>
>>109480375
By better I mean other models that can supersede h3.
Considering how long we were stuck with wan for, and how big of an upgrade h3 is, I think we will be waiting a very long time before we see even remotely serious competition in the local space.
>>
>>109480359
>Has local video gen officially been solved?
video gen in general hasn't been solved. consistency, physics, fine detail. we are getting close, but i think we are ~2 years off. then another year for people to start taking it seriously and treating it as more than a meme machines.
>>
>>109480375
>better nsfw
we won't get that from BFL lol.
>>
>>
>>109480393
You're likely right, I hope you're wrong though, either way I'm excited.
>>
>>109480391
Wow, another image board Nazi who has never even punched a man. You're gonna put me in the trash?
>>
Hi anons. Anyone tried running this on 3090 and 32gb ram? is it painfully slow? is 5090 and 64gb ram the minimum for good time with minimax H3? I want to atleast be able to watch youtube ishowspeed while i wait for the prompt to finish. Maybe also play a match of Dota 2 while i queue up some prompts.
>>
>>109480411
lol
>>
>page 2
>48 images
Why new thread doko??
>>
>>109480431
Because *someone* troll baked.
>>
>>109480411
absolute cinema
>>
>>109480411
Based.
>>
so i can do 20 steps in 8 minutes about for 0.7mp, is 40 steps worth it at all?
>>
>>109480443
>Was there some prompt guide for H3?
ref: https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md
non-ref: https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md
As always, I recommend plugging the raw MD file into an AI assistant like like ChatGPT or Claude or whatever, then giving it your normal prompt and asking it to make it follow the official conventions.
>>
>>109480446
Is it a troll bake just because you don't like the OP? Or are you just a pussy?
>>
File: 1768606814156664.png (61 KB, 600x535)
61 KB PNG
>>109480276

GPU Speed dont matter for Genning right ? I really afraid of the fan speed when genning Minimax video so i consider to Downclock my GPU.
>>
>>109480419
Back to r*ddit immediately you piece of garbage
>>
File: 1.png (43 KB, 887x587)
43 KB PNG
What happened to the youtuber Aiorbust? his profile is gone? he was a invaluable resource to the community and provided tons of workflows and teached me the ropes. Wtf get him back.
>>
>>109480481
talking about the *other* thread.
>>
>>109480505
>teached
white people don't watch diffusion youtubers rajesh
>>
File: MiniMax_H3_00243_R.mp4 (2.2 MB, 832x1248)
2.2 MB
2.2 MB MP4
>>109480420
>>
>>109480512
Yeah thats what I mean
>>
https://civitaiarchive.com/models/2761774?modelVersionId=3203865

Download before it gets taken down
>>
uh oh troon janny is starting to get really upset that his bakes dont get used
>>
>>109480534
did you miss the part where he already had a krea and LTX version? It's not getting taken down.
>>
>>109480534
why would i want to download this and why would it get taken down?
>>
>>109480534
if I wanted a mid bitch with a flat chest I'd just call your mother
>>
>>109480572
lol triggered smithloser
>>
>>109480359
The bar has been set very high.
One thing is for sure, it's extremely difficult to justify paying money for commercial video gen like VEO and Kling now. Aside from not allowing lewds, the gap between those "big" models and H3 just isn't anywhere near big enough to justify the price.
Wan should die. If those alibaba chinamen were smart they would release their latest model to the public right now to help gather data for improvement. But they won't, and wan will die.
>>
>>109480359
we still don’t have consistent minute long gens yet but we’re getting there for sure
>>
File: file.png (44 KB, 1120x305)
44 KB PNG
are they telling the truth or will we get the z image treatment?
>>
>>109480590
>wan will die.
think it's already dead, 3.0 looks like a stillbirth.
>>
>it works
This mommy gonna get some animations
>>
>>109480593
H3 is more than capable, there is no hard limit on video length, and it provides tools for natively extending videos.
>>
>>109480597
Never trust "coming soon". Ever.
>>
Trying this minimax H3 for the first time. people were saying it’s uncensored and good with physics, specifically jiggle physics.

I’ve been using int8 ref2va trying to make a woman wearing bikini dance to <Audio 1> with active movements.
the energetic dance works and aligns with the song, but her tits barely have any movement, like they’re bolted on. Even nude dance with reference picture has the same problem

No matter what I prompt or try hard using creative wordings, even providing a reference 1 second video (from LTX).

I am going insane, how are you doing it?
>>
>>109480597
What's "2k model" exactly? That sounds pretty exciting though. Is it going to solve the melted details issue?
>>
>>109480632
its a vae upscaler
>>
File: adfdasdd.png (1.22 MB, 1552x882)
1.22 MB PNG
https://files.catbox.moe/nbkot8.mp4
>>
Does anyone else get a click sound at the very start of their R2V gens? Sounds kind of like a soft clack.
>>
>>109480671
I've heard the click but not the clack
>>
>3:4
>640x832
>0.5 Mpx
>4s
>15min gen time

How fast can you gen?
>>
Are you guys up for a collaborative project?
I want to put together an index of characters that H3 has knowledge of. Entries could be marked with:

O : Full knowledge
~ : Partial knowledge
X : No knowledge

And alongside the index, a repository for clean 10 second voice comp .wavs for cloning. That way if someone is ever looking for a character, they can check to see if it exists in the repo, or contribute the clips themselves.
>>
>>109480730
I don't have the time for this but good luck, OP
>>
File: t2v_MiniMax_H3_00030_n.mp4 (1.96 MB, 1280x960)
1.96 MB
1.96 MB MP4
>>109480706
1280x960 resolution, 5s duration, 394s gen time
no upscaling, using sageattention
>>
>>109480706
>no turbo lora
>1280x736 @ 0.9 MP
>14s
About 8 minutes on a 4090 with sage and spectrum. Quality is good. This is probably the best gen I've done with it so far :
>>109475098
>>
>>109480749
Very nice gen.
>>
>>109480706
Last gen:
>16:9
>1.0 mpix (+1.30x RTX ULTRA)
>5 sec
>25 steps (so +5)
>just sage
~8 minutes
>>
File: 1761147689218851.png (9 KB, 1051x73)
9 KB PNG
>>109480730
Claude is searching the entire internet for me to make a list. He could be here right now. He said the mean redditors blocked his crawler.
>>
>>109480706
>1MP@24fps
>5s
>690s
>6 steps
>>
>>109480777
600s actually without RIFE and upscaling
>>
File: 1767807993573808.jpg (59 KB, 1417x725)
59 KB JPG
Is this right (and casual) way to Downclock GPU ??
>>
>>109480631
Maybe put in more stuff about her whole body? I get almost comical jiggle
>>
What are the bare minimum specs for running MiniMax? Can my M5 Pro integrated chip run it?
>>
>>109480790
yeah
>>
>>109480790
what program are you using?
>>
>>109480818
Thanks. I see the diference now. My Fans isnt that loud anymore with genning. Now im gonna see if its reduce genning performance or not
>>
>>109480812
>Can my M5 Pro integrated chip run it?
if you have 64gb + of memory then yes probably but expect maybe 20 minutes or longer per video
>>
>>109480830
NVIDIA app
>>
File: Krea2_turbo_00738_.jpg (2.88 MB, 2176x2896)
2.88 MB JPG
>>109480671
Do you dust your fans and if so do you make sure to keep them stationary to not mess with the bearings while cleaning?>>109480790
You should read up on undervolting it makes a night and day difference
>>
>>109480597
They're Chinese.
>>
>>109480790
+500-1000mhz for the memory clock, 75-90 power limit depending on heat generated/cooling. Get HWINFO65 and check gpu temps, especially memory junction temp. Also set up a curve for your core clock. Get MSI Afterburner, tell Claude your GPU and what you want to do and it'll give you the settings. Same deal with a custom fan curve.
>>
File: 1784772695890050.webm (959 KB, 1376x768)
959 KB
959 KB WEBM
>>109480706
>16:9
>1376x768 @ 1.0MP
>5s
>94.39s gen time
>>
>>109480706
3090, 32Gb ram,
about 5min for 6sec at 0.7mp
>>
>>109480858
It's a sound from the gen audio you stupid idiot, and can be seen in the waveform editor. I probably undervolt more than you.
>>
>>109480876
You have crazy hardware right?
>>
>>109480730
i think the number of characters it fully knows is small enough that using a reference should be the norm
>>
File: 1781333889797965.png (267 KB, 1328x671)
267 KB PNG
>>109480885
5090 but this optimization setup also really sped things up significantly, thats also with res_multistep at 20steps
>>
>>109480903
try 10 steps trust me the quality doesn't degrade, I only use the sage attention patch though
>>
>>109480903
I'm using the default workflow FL2VA, so there are no steps for me.
>>
File: MiniMax_H3_00003_.webm (3.74 MB, 864x480)
3.74 MB
3.74 MB WEBM
Reference to Video is insane, there's lots of room for improvement but the fact you can just shove things in there and it works is incredible.
>>
>>109480790

280 secs with 100% power
310 secs with 90% power

Dammit, I guess i cant gen with underclocked GPU
>>
>>109480749
lol
>>
File: 1776005914918427.webm (1.11 MB, 1376x768)
1.11 MB
1.11 MB WEBM
>>109480908
quality isn't bad when things are slow but it gets a lot more noisy with quick movement at lower steps
10 steps does bring it down to 80secs instead of 94secs, but I don't mind the extra time for better fidelity at this speed
>b-b-but it's not the same seed!!!
Yeah I'm just vibeoptimizing bro
>>109480914
open it with the white arrow in the top right of the node
>>
>>109480903
you use turbo lora with 20 steps?
>>
>>109480939
damn that reminds me that i never did play paradox chapter 3
>>
>6steps, 7s, 24fps, 0.5MP upscaled to 2MP using RTX SR.
>gen time: 420s
https://files.catbox.moe/or5pnh.mp4
>>
File: MiniMax_H3_00174.webm (3.82 MB, 960x544)
3.82 MB
3.82 MB WEBM
It's actually wild what it can do with just T2V.
https://files.catbox.moe/tesnjl.webm
>>
The model is so fucking good. And it's not even that big. And the multimodality. Doesn't even seem possible. The future is happening, ai won.
>>
>>109480970
last i check on it the last half or so of the 3rd route of part 3 is MTL but it was still legible enough to be finished. now im waiting on the expansion
>>
>>109480974
scrolling the pussy like a mouse
>>
>>109480671
yeah, i just replicated it. I guess I'll just have to trim my opening frames.
>>
>>109480706
3050ti (4gb), 16gb ram
0.4 mpx
5s
15min
>>
File: MANG_004.jpg (334 KB, 845x1126)
334 KB JPG
>>109480987
The /g/ approved way.
>>
>gen follows prompt perfectly at 0.4
>extremely retarded at 0.7
What causes this?
>>
>>109480974
ok but these to be longer, hot damn.
>>
>>109481005
jews
>>
File: may_00011_.png (824 KB, 896x1152)
824 KB PNG
Looking forward to further exploring the capabilities of R2V.

https://h.uguu.se/yuPyznGh.webm
>>
>>109480983
>last half or so of the 3rd route of part 3 is MTL but it was still legible enough to be finished
nice, gives me something to do while genning lol.

>>109480983
>now im waiting on the expansion
oh shit, how big is the expansion going to be? cool that the series is still going after all these years.
>>
File: 1767264466242803.webm (3.83 MB, 864x1344)
3.83 MB
3.83 MB WEBM
Turbo Lora needs some work. It tend to hallucinate more often on scene change. This took lots of retries

>>109480706
0.5 mp
15 secs
turbo lora with 6 step + euler + beta
270 to 310 seconds gen
>>
File: 1781014444355610.webm (972 KB, 1376x768)
972 KB
972 KB WEBM
>>109480968
without turbo lora on: 105secs
maybe I don't need it. I had another gen without it on that was 85secs so content still drives a lot of the gen time it seems.
My gains must just have come from sage attention + cache. Thought I tested it before but guess I'm wrong
In that case yeah I wouldn't recommend the turbo lora for anything very useable yet, its not good enough to drop steps in my opinion
>>
>>
>>109481022
not sure how big exactly but its going to be a NG+ mode that features a certain character from the end of the game. according to TrTr youll also be able to save some companions you couldnt before
>>
>>109481005
I restarted comfy and now the gen is good again... wtf?
>>
>>109480945
nvm. i guess my comfyui just warming up
>>
I honestly blame win11
>>
File: 8888888.mp4 (2.47 MB, 800x1056)
2.47 MB
2.47 MB MP4
she took the beer away...
>>
>>109481005
have to nail down everything in the prompt
>>
>try to create looping video
>image slowly stretches vertically towards the end
it's always the small things ruining everything
>>
>>109480749
bumping from 5s duration to 7s duration doubled the gen time from 394s to 667s
>>
>>109481053
is this ani's pov?
>>
File: 1758690702827603.mp4 (1.11 MB, 640x832)
1.11 MB
1.11 MB MP4
>>109480706
same parameters as you listed, using 15 steps, no turbo, no cache, only sageattention
4m30s on a 3070 with 32GB ram
>>
>>109481054
Pretty certain there was some residual shenanigans with sol attention that never got properly cleared.
>>
>>109481064
holy...
>>
File: celes720.jpg (403 KB, 1280x720)
403 KB JPG
Newbie question: When I prompt Minimax H3 for pussy it always seems to want to start to veer into male genitalia generating absurdly_fat_mons and sometimes even a little degenerate penis. How do I prevent this? It's done this 3 times already with a slightly different prompt and different seed. I don't think this thing really supports negative prompting in the sense SDXL stuff does or does it? The default Comfy workflow img2vid H3 workflow provides no input for such at least. I've really only used SDXL-based stuff and Forge UI before this dabbling a bit into Wan2.2 via SwarmUI about half an year ago which IIRC didn't really support negative prompting either though you could at least enter one.

This is img2vid not txt2vid btw. Reference frame is an AI generated image of a girl with a leotard covering pussy with cameltoe. I'm asking for a pussy reveal in the prompt.
>>
File: bruh.jpg (28 KB, 352x460)
28 KB JPG
is there any way to have the sampler preview in the H3 node not be a tiny stretched out noodle?
>>
>he uses the subgraph
>>
>>109481071
Retard it can't do nsfw out of the box without a reference
>>
>>109481073
no, subgraph is fucked atm
>>
>>109481071
>How do I prevent this?
(you) don't, newfag, you wait for someone with $80,000 worth of hardware to make new loras and checkpoints for you
>>
>>109481082
blatant lie
>>
>>109481082
feisty little sperg
>>109481085
Rather have new blood to not have a dead thread like /sdg/
>>
>>109481083
:'( ok
>>
>>109481092
I'm not telling the newfag to fuck off, he can learn, that's why I'm teaching him the true way of the lurkchad
>>
>>109481097
shut the fuck up and either ignore or help
>>
File: file.png (319 KB, 1291x313)
319 KB PNG
>>109481066
on 0.2mp costanza has black-blue pants
on 1.0mp he suddenly gets red pants, same seed
on prompt I don't have pants color defined and the reference image is just his upper body holding a baseball bat
so because the pants color is not defined it just decides to change it sometimes
>>
>>109481109
I'm helping by telling him to wait for loras and checkpoints which is the truth retard why dont you kys
>>
>>109481071
provide an end image for it to refence
>>
>>109480471
I asked Claude to conduct a deep analysis of the strengths and weaknesses of Qwen 3.6 27b Hereric and to write a system prompt in the form of a harness, taking into account the architectural analysis to upsample input into structured output.

A second pass with a different harness that critiques the finished prompt as a film director in the appropriate genre, includes a comprehensive glossary, and upsamples the structured output once more through more detailed storytelling, cuts, additional scenes, etc.
Animation, anime, cinematic, TV series, and XXX each have their own editing, stylistic, and camera techniques.

If only it weren’t for that miserable model switching.
>>
File: 1764246673682132.webm (3.83 MB, 864x1344)
3.83 MB
3.83 MB WEBM
>>109481024
changing the turbo lora str to 1.5 reduce the hallucination but not completely. Cant complain about this i guess
>>
>>109481122
>Heretic
>not using gemma4
The fuck are you doing?
>>
File: IMG_2027.png (113 KB, 1247x1144)
113 KB PNG
>>109481073
use this instead on model load. and you can use default vide vae for it
>>
>>109481112
also here is the video, would real costanza ever wear red pants like this I don't think so
https://files.catbox.moe/54j00x.mp4
>>
With R2V, is all that boiler plate really necessary? Like do I have to write all the retention analaysis?
>>
>>109481131
You should know by now most people in this thread aren't the brightest
>>
>>109481150
no
>>
How's the turbo lora? should I wait for it to git gud?
>>
Anons, I come to you in a time of great need. I've been told H3's porn ability goes up exponentially if you feed it porn reference pics. Yet, all my local diffusion models are furry. What's the leading human porn model?
>>
File: 1785181187335625.webm (1.18 MB, 1376x768)
1.18 MB
1.18 MB WEBM
>>109481025
Actually turbo lora does consistently shave off 10-15% of the gen time, even at 20 steps, its worth turning back on.
Its always 100+ seconds without, and 80-90 seconds with.
>>
>>109481183
Also no wonder, holy fuck, civit is now SFW only? Jews can't have people undercutting onlyfans.
>>
>>109481183
google images
>>
>>109481187
you need civitai dot red for nsfw, they split it
>>
>>109481144
speaking of preview override, does it add to inference time or is it the same as the default preview?
>>
File: 1757307518466059.webm (1.7 MB, 1440x816)
1.7 MB
1.7 MB WEBM
We can make our own JAVs now boys.
>>
>>109481165
no harm in downloading it and trying it out - it's there even if you just use it for previews to try out prompts
>>
>>109481187
kek calm down schizo. they stupidly split the domains so it's civitai.red for NSFW now.
civitai is so wildly incompetent at handling their own rules there's a ton of loli shit on civitai.red even though it's technically not allowed + you're "MEANT" to use the PG filter to see "underage characters".
>>
>>109481187
They a red version for NSFW stuff, kinda like 4chan with blue boards.
>>
>>109481165
turbo lora is not good, stick with spectrum for now
>>
>>109481190
bbc lover
>>
>>109481165
It's complete fucking dogshit right now.
>>
>>109481210
I tried spectrum and got no speed difference
>>
>>109480879
Hey show some respect to the most notable poster in these threads
>>
>>109481210
spectrum sucks
>>
File: 1785820608329912.webm (2 MB, 864x1344)
2 MB
2 MB WEBM
I've run out of ideas
>>
>>109481233
source?
>>
>>109481082 >>109481116
Yeah seems I really should have just tried feeding it a pic with pussy already visible before posting.
>>
>>109481165
It's pretty good with the right tweaking
as in >>109480974, >>109480179 and >>109479620
>>
>>109481237
My PC, it lost the comparison test against sage attention
>>
>>109481198
Haven’t measured, no noticeable difference

on the other note,
pruned vs unpruned model?
pruned vs unpruned turbo lora?
do they mix?
>>
>>109481242
I've been doing movie style sex scenes specifically not showing genitalia. Works gud, very erotic.
>>
>>109481236
0/10 expected the boys to throw acid on her
>>
Now I know for 100% sure H3 can generate pussy, and it actually looks good. No loras, you just need to know what to prompt and what not to prompt.
>>
my output is looking deepfried with the turbo lora, what causes that?
>>
HOLY FUCK
THIS IS THE GOLDEN AGE
WHAT A TIME TO BE ALIVE
>>
>H3 knows how to make perfect wet pussy sounds unprompted
mmmmAAAMAAAA hoooEEEYYY BOY i haven't felt a boner pop off like this since illustrious/noob first came out
alright queing my 10 gens with this good-enough setup and fucking off

seriously my advice to those genuinely unable to figure out nudity; give it fucking references, multiple if you can/want, i gave it the main clothed character reference and two nudity references, it's a 1:1 on the accuracy for BOTH the body TYPE and details like nipples and pussy lips.
>>
also to whoever was whining it couldn't do reflections;

it can do reflections.
https://files.catbox.moe/q5ou35.png
>>
>>109481296
i'm not even giving refs, cuck. it works in simple i2v if you choose your words correctly.
>>
>>109481304
you can't like, not post the full video like that.
>>
I think cache node is what's raping the prompt adherence on my gen.
>>
>>109480749
prompt?
>>
>>109481306
i'm not saying it doesn't work without refs, i'm saying give it refs to get exactly what you want, cuckold.
>>
>>109481276
>>109481306
pzl to share
I can get it to show pussy from the front but it's hit-and-miss

>>109481296
>just show it a pussy and it will show you the pussy back
Boooooooooooooring
>>
i knew high quality video and audio generation had to be possible somehow (a la library of babel / monkey typewriter) but i never thought it would be possible in my lifetime, let alone on consumer hardware. kind of fucking insane man. it might be just used for porn and scams, but damn if it isn't wild to think about
>>
File: 1761213749992434.jpg (67 KB, 736x552)
67 KB JPG
Can you imagine if the OP webm was your daughter? My heart would feel like a pack of C4 being struck by lightning.
>>
>>109480749
kek
>>
Playing around with Krea 2 and anyone have tips on how to make symmetrical eyes?
Eyes are sometimes wonky. Do I need a lora or some other extension for this?
>>
one of the turbo loras got updated again

https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora/tree/main
>>
>>109481144
This killed my gen time
>>
>>109481296
how? give the sound prompt
>>
>>109481364
>step 922
zzz
>>
>>109481364
It's not "one of the", That's the only one. Every other one are just taking this and converting it to comfy format.
>>
>>109481364
Holy fuck im gonna get fired at this rate. I cant stop gooning and genning
>>
>>109481344
I for one would be confused how I managed to produce a two dimensional cartoon daughter
>>
minimax_h3_turbo_4step_ema_ckpt850.safetensors

Steps: with ckpt850, 4 steps is already sharp (earlier checkpoints needed 6–8 to firm up). Any count ≥ 4 is valid; more steps still help a little. Keep the scheduler on simple.
LoRA strength (default 1.0) is the dial for the sharpness/artifact trade-off: if the result shows blurry ghosting / smear, nudge strength up (e.g. 1.05–1.2); if it shows over-sharp grain / artifacts, nudge it down (e.g. 0.8–0.95).
Base model: works with any MiniMax-H3 base — full (bf16, int8_convrot) and the pruned/curve variants (pruned_int8, pruned_fp8); the ComfyUI node auto-detects a pruned base and re-injects the time-conditioning at run time, so one LoRA covers every base.

oh shit, 4 steps working? before it needed 10.
>>
File: MiniMax_H3_00093_.webm (3.04 MB, 704x1056)
3.04 MB
3.04 MB WEBM
>>
File: 11111.mp4 (2.02 MB, 1056x608)
2.02 MB
2.02 MB MP4
https://files.catbox.moe/svh1zf.mp4
>>
>>109481364
>put these speed holes in your new Ferrari, make it look and sound like shit but it goes faster!
don't believe Turbo lies
>>
>>109481397
>casing floats and doesn't fall to the ground
prompt harder
>>
File: 1769067743979962.png (982 KB, 1056x1056)
982 KB PNG
Whats the best prompt enhancer workflow and text model out there ? Pls i need answers i had enough of manual prompting
>>
>>109480487
yes, you can gen at 1hz
>>
>>109481422
???
>>
>>109480421
Minimax H3 with Comfy default workflow works perfectly with my 3090 + 32gb.
>>
>>109481415
Spend like 10 minutes with claude and make one that suits your exact needs.
>>
>replace the sampler feeding SamplerCustomAdvanced with MiniMax-H3 Turbo Sampler (4-step), and set the scheduler to 4 steps (simple).

ah, I didnt do that part and got a cursed southpark test video. could be a creepypasta generator desu
>>
>>109481439
i dont want to give me personal info to claude
>>
>>109481403
>an infinite amount of possibilities, right at your fingertips
>reposts the same fucking thing for the billionth time

epic ftw
>>
>>109481409
I have my reference and i2v/t2v workflows saved, im just gonna test out of curiosity.
>>
File: sdfsdfsdfsfsfsfsf.webm (3.84 MB, 720x1280)
3.84 MB
3.84 MB WEBM
Can't wait for the upscaler they're planning to release. I can't go above 1mp or else the style etc changes dramatically. Absolute kino otherwise. The physics of the dress is particularly impressive, did multiple gens and they must have trained a ton on physics.
>>
>>109481265
>do they mix?
In general yes, it's just pruned is more optimized for inference.
>>
>>109481450
>an infinite amount of possibilities, right at your fingertips
>manually typing seetheposts on 4chan
>>109481403
I like it
>>
File: Krea2_turbo_00763_.jpg (3.39 MB, 2176x2896)
3.39 MB JPG
This was a pain in the ass to get right
>>
>>109481450
lol its worse, I've seen some people generate clips of The Big Bang Theory, not even unironically ones, just trying to be like the show and even more cringe
>>
>>109481479
clips from shows are more impactful because everyone can understand the reference
BBT sucks though, the office is good.
>>
>wake up
>gen
>sleep
>wake up
>gen
>sleep
is anyone else like this?
>>
Is there an easy way of viewing video metadata?
I don't wanna drag the workflow into comfy everytime, sometimes switching workflows fucks the UI. I just want the prompt quickly damn it
>>
>>109481454
You can use RTX Super Resolution for the time being.
>>
>>109481064
does she have pierced nipples?
>>
>>109481530
you don't have a prompts.txt sitting around?
>>
>>109481437
gpu (clock) speed is measured in hertz
>>
>>109481526
every other week or two I fall down the hole again, yes
>>
File: 1762158004460626.jpg (35 KB, 597x589)
35 KB JPG
How to get $100000 in two weeks so i can buy Those Workstation Nvidia GPU with 1terabytes of DDR5 ram + 30tb of SSD
>>
>>109481494
they just need to be porn parodies or be out of context - simply using the show as something to build around. The problem with NPC's is they will simply try to make new material with the same scenes, the same characters so aiming for new episodes of the same dross. But it's early days of the new model, people will branch out into more creative things hopefully.
>>
>>109481541
For you
>>
>>109481557
If i overclock your GPU will it die?
>>
>>109481538
shit... I should start doing that. doy
thanks
>>
>>109481560
No
>>
>>109481476
pedo
>>
My whole system seems to hang at certain sizes+lengths. It gets stuck on initializing the model and the whole thing starts to chug, I can't even smoothly use a browser. The breakpoint seems to be somewhere around 15 secs at .5MP.

Working on a 5080 and 64gb of system ram. Think there's anything I can do to resolve this?
>>
>>109481454
she has grown tits and giant baboon ass, nooticing the ever progressing fat fetishism

Sucks, LTX and community loras couldn’t do slim/graceful shapes either, only Wan 2.2
>>
>>109481587
stop using sage
>>
>>109481587
are you using the rtx upscaler by any chance?
>>
File: 1765970604752108.mp4 (1.42 MB, 1376x768)
1.42 MB
1.42 MB MP4
>>>/wsg/6209327
>>
>>109481576
How so?
You have a issue with petite women with large breast?
Or are you doing that marvel le reddit cope
>>
ain't no way this lil green alien nigga generating 2k on a 16gb card.

-the turbo in question https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/tree/main
>>
>>109481597
Afraid not. My workflow is sage patch -> minimax mem eff patch -> spectrum, but I'm pretty sure sage patch and the minimax patch don't actually work together

>>109481592
is this a shitpost or real advice
>>
>>109481611
>sage patch -> minimax mem eff patch -> spectrum
you should only be using one of those 3, thats your problem
just use mem eff H3
>>
>>109481590
>that
>"baboon ass"

ireland
>>
>>109481622
I'll try that. I was under the impression that the mem eff patch and either spectrum or the cache node worked together
>>
Nakadasheeeiieeiieeieiieie!!!!!
>>
>>109480597
why are people hyped for an upscaler? we already have a lot of upscalers already
>>
>>109481603
that is a sexualized teenager (underage child)
>>
>>109481667
>mutt
lol
>>
File: Untitledsdfsdf22f.webm (3.83 MB, 1080x1920)
3.83 MB
3.83 MB WEBM
I've been going through my old gens heavy on style and grime. Very impressed so far. For webm related it was able to get the boat far in the distance from the start, very impressive how it picks up speed and size.

>>109481532
Doesn't work for the style I'm doing.
>>
Dear 109481667,

obvious troll is obvious

Sincerely,
Anon
>>
>>109481674
MAKE HER PISS HERSELF
>>
she is only 52 you sick freak
>>
>>109481667
Grown woman you low iq retard. You're so fucking desperate it's hilarious.
Tell me how you never got college pussy without telling me how you never got college pussy.
Sorry your hero got busted posting actual creep shit too, that's probably why he's been unemployed for over a year
>>
>>109481688
that's her grandmother. the one I am referring to is 13
>>
>>109481693
whoah man chill out. you are still a pedo tho
>>
>>109481674
lol'd hard wasn't expecting that
but also maker her piss herself >>109481685
>>
>>109481476
Thank you for inspiring us, anon.
>>
>>109480359
>power and capabilities of r2v
faces are not preserved accurately enough, not even close
that said I doubt there's much room for improvement left here, loras are the answer to this problem
>>
File: 201853CUI_00001_.png (946 KB, 1216x832)
946 KB PNG
>all these videos
buncha moneybags
>>
>>109481707
>>109473827
>>
>>109481741
Nigga it can run on 6gb of vram.
>>
>>109481745
>it can run on 6gb of vram.
and it takes 2 litteral hours right?
>>
>>109481745
aint no way
>>
>>109481748
it takes 15min with 4gb vram, without any of the speedup shit
>>
>>109481757
resolution? temporal length? steps?
>>
>>109481748
resolution and length massively impact vram requirements and gen times
if you have 32gb you can easily create solid low-res vids with 6gb vram at reasonable times (10-15min)
>>
>>109481769
0.4mp, 5seconds, 20 steps
>>
>>109481769
like 1.5mp, 380 frames, 20 steps
>>
>>109481757
I dont know how you can patiently waiting for that long. I already hit my limit when my gen is at 5 minutes
>>
>>109481779
kek
>>
>>109481782
You are acting like a spoiled kid. Just browse tiktok or something for 5 minutes.
>>
btw kijai is working on a fix to turbo lora's effecting audio
https://github.com/Comfy-Org/ComfyUI/pull/15243

Also the turbo lora has a newer version out that is nearly perfect now
>>
>>109481790
Just do the math anon. And theres potential of failed gens
>>
the turbo lora is one of the worst things i've seen. why would anyone release something this pile of garbage? please leave the optimization to people who actually know what they're doing, like KJ and lightx2v
>>
>>109481741
With only the turbo lora it took me 6 minutes to gen 5 seconds on my 3060 12GB, I'm sure it can go faster if I get all the right nodes but I can't be arsed when everything still volatile, once the best nodes settle themselves I'll start genning in earnest.
>>
>>109481200
>has the power to generate anything he desires
>still chooses to watch a cute girl get fucked by another man
>>
File: Krea2_turbo_00771_.jpg (2.24 MB, 1776x2368)
2.24 MB JPG
>>109481742
Crazy how he just post shit like that non stop and seethes when people want nothing to do with him IRL
I need to look into high res fix pass I think there's a ton left on the table with this model
>>
who cares about the turbo lora? we need a turbo decoder vae
>>
>>109481799
Get this fucking finn guy a raise
>>
>>109481707
you're not fooling anyone you clown, a nonce is what you are.
>>
>>109481808
>cute girl
>not using yourself as a reference
oof
>>
>>109481808
Who's to say that's not anon?
>>
>>109481799
I should state its working perfectly at like 8 steps. Its still not enough for 4 steps
>>
File: 1768292354369089.png (5 KB, 952x71)
5 KB PNG
>>109481606
He be lying it won't even start on my 5070ti using that workflow.
>>
>>109481827
Asian men don't use 4chan
>>
>>109481830
Lying on internet? How can it be...
>>
>>109481833
what do they use then?
>>
>>109481833
saar



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.