[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


Discussion and Development of Local Image, Video, and Music Models

Previous: >>109522303

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: 1760442155997786.jpg (24 KB, 552x530)
24 KB JPG
>>109523403
Man i just want to use R2V similar to I2V.
But it keep creating me a new scene.
Those Voice reference are too good to ignore.
Also it helps with face drifitng. with additional face reference.

The problem is
IT ALWAYS CREATING AN ENTIRELY NEW SCENE AAAAAAAAAAAAAAAAAAAAH
>>
>>109523233
You're coping. You're a cuck who relies on the big boy to do the big man job because you're too stupid and worthless to handle it yourself..
>>
>>109523416
Too bad there's not a full prompting guide which tells you exactly how to continue scenes. Damn.
>>
blessed thread of frenship
>>
tfw ive been deleting the "For the target video, at 0.00 seconds into the target video, <Picture 1> (from [Shot 1]) is fully referenced." part this whole time
>>
>>109523416
It's because you didn't read the documentation. R2V is not a model for babies, you have to actually write your prompt properly.

https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md#4-retention_analysis
>>
sick, this worked

[2-4s] the camera hard cuts to a first person view of the character looking down at the guitar as she plays it.

https://files.catbox.moe/yvexnc.mp4
>>
>>109523434
sorry anon.
>>
>>109523403
>>Maintain Thread Quality
>https://rentry.org/debo
>https://rentry.org/animanon
hey jannies/mods. This is a tool used by the sharty to troll. it's basically a call for a raid
https://basedjakwiki.org/Trolling
>>
>>109523436
It's pretty nuts that local model can do this in the first place and even with shitty hardware
>>
>>109523427
I tried this
[Shot 1] completely use the background, camera, lighting, character, artstyle from <Picture 1>
And it works but its not a "fix al"l prompt.

>>109523436
I tried everything in the documentation and it fails
>>
File: 1767442389610414.mp4 (621 KB, 640x640)
621 KB
621 KB MP4
>>
File: janny cleanup.mp4 (3.88 MB, 1184x896)
3.88 MB
3.88 MB MP4
>>
>>109523458
You clearly haven't or you have the wrong model selected or something. R2V literally has a task type for your exact usecase, it's called [keyframe completion]. Works for me just fine.

If you're struggling just use ChatGPT. Feed it the documentation and tell it what you want then you'll at least get a scaffolding you can modify to your liking.
>>
>>109523467
make something original
>>
>>109523477
When you stop spamming.
>>
>>109523484
>announcing a report
>>
>>109523475
fl2va model can do R2V as long is a picture. CMIIW
>>
If that thread was the real one you wouldn't have to spam it so hard desu
>>
>>109523485
>>109523495
>>109523466
>>109523459

Im already there you retard.
Answer my question
>>
I wonder why 0.4 megapixel is default in workflows because it messes face details
>>
File: 1781823823549141.jpg (83 KB, 750x1000)
83 KB JPG
Did Debo even gen anything or he just one of those "ANTI AI" Retards ?
>>
>Spamming/flooding
>>
>>109523508
Use R2V. Its a powerful tool
>>
>>109523510
He gens, but he's poor, low IQ and mentally ill.
His GPU isn't good enough to run minimax so he copes by shitting up /ldg/ for 20 hours a day.
>>
File: 1764151512501243.mp4 (1.13 MB, 736x576)
1.13 MB
1.13 MB MP4
blah bad seed
>>
>>109523523
awkward angle
>>
>>109523516
ref2vid? already using it. Perhaps it's user error, wouldnt be the first or last time
>>
>>109523508
you're likely using too many cope nodes
>>
>>109523508
>can gen any woman
>gens sweeny
the man with no taste
>>
Seeing him get worked up like this is nostalgic. It's been awhile since he's gone full spam mode.
>>
>>109523526
>>109523517
>>109523540
What are your GPU ? Ill send you my spare 5070ti PC. I wont murder you i swear
>>
>mfw Resource news

08/10/2026

>H3 Motion Context: Chain MiniMax H3 clips
https://github.com/NikoDemon80/ComfyUI-H3-Motion-Context

>MiniMax H3 Turbo — ComfyUI 4-Step T2V and I2V LoRA
https://huggingface.co/joyfox/MiniMax-H3-Turbo

>HRDiT: Training-Free High-Resolution Image Generation with Off-the-Shelf Diffusion Transformer Models
https://github.com/zylwithxy/HRDiT

>RoRA: Role-Oriented Regional Allocation for Visual Token Pruning in MLLMs
https://github.com/LukieLuu/RoRA

>AVCap: Reinforcing Audio-Video Joint Caption with Detail-Aware Reward
https://huggingface.co/collections/Apryle/avcap

>Prune Once: Retraining-Free Task-Agnostic Pruning for Vision-Language Models
https://github.com/cau-hai-lab/PORTA.git

>Unsloth Minimax H3 GGUF (Q2:Q8)
https://huggingface.co/unsloth/MiniMax-H3-GGUF

>MiniMax-H3 for Apple Silicon - Rebuilt from the official weights
https://huggingface.co/uetuluk2/minimax-h3-mlx-rebuild

>Soran’t: Small Next.js front end for video generation on ComfyUI
https://github.com/pwillia7/ai_video_fe

08/09/2026

>Kroma v0.2 — Krea 2 fine-tune (full model)
https://huggingface.co/lodestones/Kroma

>krea2-turbo-bbox
https://huggingface.co/jimmycarter/krea2-turbo-bbox

>Kroma v0.2 Quant
https://huggingface.co/silveroxides/Kroma-Quant/tree/main

>Spectrum for Ideogram 4
https://github.com/Nif00/ComfyUI-Spectrum-Ideogram4

>ClipProj — MiniMax H3 conditioning from a Qwen3-VL-4B
https://huggingface.co/NicoLab28/ClipProj-MiniMax-H3

>ComfyUI-SigmaSync-LoRA: Sigma-aware model-only LoRA strength scheduling
https://github.com/capitan01R/ComfyUI-SigmaSync-LoRA

>NexusBTA v0.2.44 adds MiniMax H3 support
https://github.com/JpAndreBTA/Nexus-BTA/releases/tag/v0.2.44

>Experimental MiniMax H3 single-image VAE
https://huggingface.co/Mamad8/MiniMax-H3-Image-VAE

>MiniMax H3 REF2VA w4a8
https://huggingface.co/realrebelai/Rebels_w4a8s

08/08/2026

>Kijai: MiniMax H3 Ref Lora Rank 256 bf16
https://huggingface.co/Kijai/MiniMax-H3-experimental/tree/main/loras
>>
>mfw Research news

08/10/2026

>Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation
https://arxiv.org/abs/2608.06751

>From Cheap Fakes to Pure Synthesis: Addressing the New Era of T2V Fake News Videos
https://arxiv.org/abs/2608.06732

>PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model
https://arxiv.org/abs/2608.06794

>Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models
https://arxiv.org/abs/2608.06768

>MaskFlow: Precise, Consistent and Seamless Regional Image Editing
https://arxiv.org/abs/2608.06929

>Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity
https://arxiv.org/abs/2608.07120

>Addressable Memory for Video World Models
https://research.nvidia.com/labs/sil/projects/WorldTrace

>Local Epistemic Uncertainty Guided Active Sampling for Plug-and-play Diffusive Image Restoration
https://arxiv.org/abs/2608.06981

>ControlRef: Efficient Layout-Guided Multi-Instance Generation via Anchored 4D-RoPE
https://arxiv.org/abs/2608.06878

>CustomDance: Customized 3D Dance Generation with Coarse-to-Fine Human-Centered Interactive Control
https://arxiv.org/abs/2608.06722

>Bend the Basics: Degradation-Aware Deformable Tokenization for All-in-One Image Restoration
https://arxiv.org/abs/2608.06832

>Stable Curves, Unstable Items: Item-Level Scaling Heterogeneity in Video LLMs
https://arxiv.org/abs/2608.07014

>A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy
https://arxiv.org/abs/2608.07427

>Alignment has a Fantasia Problem
https://arxiv.org/abs/2604.21827
>>
What should I gen?
>>
File: brainlet setup.png (73 KB, 879x504)
73 KB PNG
>>109523536
>you're likely using too many cope nodes
Anon posted here, I copied
>>
File: debo_dm_k2_00036_.png (2.6 MB, 1872x1007)
2.6 MB PNG
>>109523510
I have more gens posted to /g/ than almost any other anon. I'm prolific. I also curate and post the news :)
>>
>>109523572
Why did you say the default resolution messes with details without saying you're also using the mess with details nodes?
>>
>>109523569
A man losing his mind over a community that doesn't like him.
>>
>>109523573
Post your Minimax gen then loser
>>
https://files.catbox.moe/exvxnd.mp4
>>
>>109523593
kek nice
>>
>>109523583
Is using sage attention startup flag + default workflow the way to go?
>>
>>109523599
That's what I do.
>>
>>109523530
I think Turbo lora was the cause. I tried no turbo + default res_multistep / simple and it works now.
Still need more testing though
>>
>>109523593
average changeling round
>>
>>109523616

>>109523475
>>
File: MiniMax_H3__00012-small.mp4 (3.17 MB, 1728x960)
3.17 MB
3.17 MB MP4
>>
minimax has done a real number on debo huh
>>
ok im going to the other thread. this one suck ahh nga
>>
QRD on the guy having a bad time?
>>
File: Krea2_turbo_00553_.png (1.25 MB, 1024x1024)
1.25 MB PNG
>>
>>109523569
the spammer killing himself
>>
File: 1781811510173628.mp4 (542 KB, 576x736)
542 KB
542 KB MP4
>>
>>>/wsg/6211923
adding a separate reference just for the face really helped.
>>
How many proxies does he own?
>>
>>109523665
These videos you do are so badass.
>>
>>109523676
probably works at an ISP or something and can just swap out IPs at will.. i used to be able to do that a while back
>>
>>109523676
Did he pay $10 to that furry pedo website ?
>>
>>109523684
Thanks, hopefully I won't hit a wall and I'll be able to finish the whole chapter.
>>
>>109523646
His name is debo.
He's mentally ill (severe autism), poor, has a low IQ, ugly, addicted to attention, and spends 20 hours a day on 4chan.
He's angry that he can't run H3 on his poorfag GPU so he copes by trying to ruin the thread for people who can run it.
We're all waiting for the mods to finally rangeban him.
>>
File: 1766463584121810.mp4 (892 KB, 832x480)
892 KB
892 KB MP4
>>109523665
would watch a full blame kino
>>
>I'd rather not write a shove-it-down-her-throat version — it's an identifiable real person being physically forced. Here's the same beat played as consensual comedy instead:
fuck OFFFF
>>
>>109523719
KEK
>>
File: 1782710960751479.mp4 (470 KB, 640x640)
470 KB
470 KB MP4
>>
File: MiniMax_H3__00014.mp4 (1.34 MB, 1504x1120)
1.34 MB
1.34 MB MP4
>>
>>109523690
>the whole chapter.
how long would a whole chapter be?
>>
holy shit did you dumbasses actually move to debo's thread?
>>
>>109523767
This one is 26 pages.
>>
File: 00008-886344185.png (736 KB, 1024x1024)
736 KB PNG
>>
>>109523761
Neat, you should do one where he's gradually raised up on an ice pillar
>>
>>109523780
I haven't but as you can see this shithead has been making this thread unusable. We have a bunch of newfrens who don't want to deal with this shit.
>>
File: MiniMax_H3_00562_.webm (858 KB, 768x960)
858 KB
858 KB WEBM
>>
>>109523783
damn
im rooting for you anon its kino
>>
>>109523780
He samefags his own thread
>>
>>109523790
You ever make new gens or just spam your old mid ones?
>>
File: file.png (117 KB, 904x823)
117 KB PNG
For some reason the ref_video thing only takes images. Is there a fix for this?
>>
Post cute waifus and I will make videos of them
>>
>>109523798
Don't care what you think moron
>>
>>109523803
you need the GetVideoComponents node
>>
>>109523807
Keep that attitude and you'll make it in the OP
>>
cozybreas
>>
File: feed.png (67 KB, 539x656)
67 KB PNG
>>109523803
feed it in at 24fps.
>>
>>109523810
Thanks. Fixed the problem.
>>
>>109523811
I'll say it again just so your tiny brain understands: I don't care what a moron like you thinks. You have autism, you don't matter.
>>
File: ComfyUI_02150_.png (929 KB, 1024x1024)
929 KB PNG
>>109523804
im planning on doing one where the viewer is looking down at a party application form then the camera tilts upward to reveal this cleric who submitted the form
>>
File: 1760951139961407.mp4 (117 KB, 800x544)
117 KB
117 KB MP4
>>
>>109523824
Was expecting a jumpscare
>>
>>109523824
>look mom! I wasn't lying, I really have a girlfriend!
>>
>>109523824
what a weird thing to gen
>>
>30 deleted posts
>>
>>109523889
it doesn't take a crystal ball to see
anon will always find a way to seethe
drink and jack off by himself
anon's a schizo, I can tell
>>
File: debo_dm_k2_00040_.png (2.44 MB, 1872x1007)
2.44 MB PNG
>>109523889
and I'm still here. irrefutable proof that I'm a positive (and likeable) contributor
>>
File: 1775341252564847.mp4 (647 KB, 736x576)
647 KB
647 KB MP4
>>
>>109523889
kinda weird how all the discussion in the other thread stopped as well
>>
>>109523898
I'm a schizo. That dude is mentally bucked and broken
>>
LDG win bigly
>>
nah bro youre still too prideful
>>
>10 minutes for a 10 second 0.5mp gen
Fuck I wish newegg would hurry up with my ram already.
>>
File: 1759520835257493.mp4 (731 KB, 864x480)
731 KB
731 KB MP4
>>
File: waifu.mp4 (2.35 MB, 832x1248)
2.35 MB
2.35 MB MP4
>>
>>109523932
When and how did that bitch get into my room?
>>
File: 1769507366991791.mp4 (393 KB, 768x544)
393 KB
393 KB MP4
>>109523880
it would be weird if it was my own apartment
its funny
>>
>>109523939
HAHAHA
>>
>>109523898
kek
>>
qrd? i just woke up
>>
>>109523961
we poastan
>>
>>109523939
BLACKED
>>
File: 1785743032495193.mp4 (180 KB, 576x736)
180 KB
180 KB MP4
okay enough of that
>>
>>109523984
look at comfyanonymous living his best life
>>
heh

https://files.catbox.moe/i5h3ko.jpg
>>
>>109524009
Based.
>>
>>109524009
That's illegal.
>>
>>109524009
Hello, 2014. simpler times.
>>
File: h3noaudio_00027.webm (1.2 MB, 1384x768)
1.2 MB
1.2 MB WEBM
https://files.catbox.moe/bzh8nt.webm
>>
And what did newfrens learn today?
>>
File: 1765030656410877.mp4 (551 KB, 800x544)
551 KB
551 KB MP4
>>
>>109524066
How did you do all those screenshots? manually pasted in ? pretty neat.
>>
File: 1773045676276504.mp4 (439 KB, 576x704)
439 KB
439 KB MP4
>>109524072
its an image from anon
>>
>>109524049
why would you advertise?
>>
>>109523475
>>109523616
>>109523458
I'm not running the turbo workflow and it's not working for me. Something is breaking the formatting. It works okay if I do just the simple shit like <cut 1>
>>
>>
>>109524150
Cute
>>
File: AnimateDiff_00310.mp4 (2.62 MB, 1280x960)
2.62 MB
2.62 MB MP4
>>109524111
some ads are fun
>>
>>109524111
Blue Prius belongs in Azeroth. You’d know this if you weren’t a tourist.
>>
File: 1770376705885438.mp4 (641 KB, 768x544)
641 KB
641 KB MP4
>>
>>109524153
>not pulling out an image of a 1girl
gotta get that low hanging fruit anon
>>
>>109524153
Don Draper would have loved AI.
>>
File: 9f5fa7.gif (768 KB, 240x240)
768 KB GIF
can h3 do this?
>>
File: h3-1786426255249 (1).mp4 (705 KB, 960x544)
705 KB
705 KB MP4
Live-action woman christina hendricks as "dexter's mom" from dexter's laboratory, tall with a curvy figure, standing at a kitchen sink washing dishes, seen from behind at first. She wears a pale green blouse with a wide collar, a white apron tied at the waist, dark green pants, and bright yellow rubber dish gloves. Warm auburn wavy hair pinned back. 1990s suburban kitchen: pale yellow cabinets, checkered curtains, soft afternoon light through the window. Camera glides in a smooth 180-degree arc from behind her right shoulder around to face her directly. As the camera settles on her face, she turns her head, notices the camera/viewer, and breaks into a warm, genuine smile. Shallow depth of field, soft cinematic sitcom lighting, warm color grade, steady smooth camera motion, 24fps film look.
>>
File: MiniMax_H3_00140_.mp4 (1.11 MB, 864x464)
1.11 MB
1.11 MB MP4
>waiter, deal with this foidbabble
>>
>erotic horror
Gens instantly become kino
>>
>>109524216
Proof?
>>
>spend like a month with claude making a workflow
>workflow breaks after update
>spend a couple of days on again
>finally reach a breakthrough last night and it's so much better than before

AI agents aren't good yet. They will be once they are able to give you a large amount of alternatives for a solution.
>>
File: 1759548825418219.mp4 (620 KB, 864x480)
620 KB
620 KB MP4
>>
>>109524153
I expect a catbox of her sitting on the scanner glass first thing tomorrow morning
>>
>>109524229
Even with gemma4_Q2_XSSSS it would not take you a full month to slop out a workflow.
>>
>>109524229
I’m genuinely puzzled by the idea that people need Claude to make workflows.
>>
>>109524253
half of the world population is below average iq
>>
>>109524229
Post the workflow.
>>
File: 1580w715413930.jpg (5 KB, 225x225)
5 KB JPG
anything new about Vae decode fuckery? it's taking 1/4 of my gen time. What a disaster.
>>
>>109524216
POST IT
>>
>>109524280
KJ King released one that cuts down decode time significantly. That's all the spoon feeding you'll get from me.
>>
>>109524290
nigger I tried it already it didnt work
>>
>>109524280
You're supposed to use a quantized vae.
>>
>>109524247
>>109524253
Oh it's not a regular workflow, it's working around limitations of a model and taking things that shouldn't work together, work.

>>109524277
Lolno
>>
t2v first person perspectives are the ultimate promptlet filter
>>
>>109524337
Uh huh. Sure, anon.
>>
File: who.mp4 (1.26 MB, 672x1216)
1.26 MB
1.26 MB MP4
>>
File: MiniMax_H3_00003_.webm (1.62 MB, 768x544)
1.62 MB
1.62 MB WEBM
wtf why didn't anyone say this shit was so easy
>>
File: h3_00032.webm (403 KB, 1384x768)
403 KB
403 KB WEBM
>>
>>109524373
for the love of god do meg
>>
>>109524417
Nah her cock would be visible hanging out under the skirt
>>
>>109524367
A common misconception is that demons are bound to the circle. The Magic Circle protects the summoner while spirits are bound to a triangle drawn outside the circle.
>>
>>109524434
Actually Debo is a devil not a demon
>>
>>109524364
she cute doe
>>
what's the current hype workflow if I want to do video on a 4070 ? minimax h3 ?
>>
ComfyUI+Linux+Intel Arc is the greatest combo:
Updated ComfyUI to try Minimax and got "tensor size" error. Tried again with different workflow and got the error again plus it made the system to stutter, ate all free space on the disk fucking up tabs in Chromium and desktop display settings after system restart.
>>
Is there a lora or prompt phrasing to make "normal" bodies? I don't want everyone to either be a supermodel skinny or when I prompt "slight body fat" for them to be obese cows
>>
BROS BROS

LOCAL IS SO FUCKING BACK
>>
>>109524236
Kalimba
>>
>>109524678
Is Sua the dog? Looks alright, if a bit blurry.
>>
File: 5239797245624689628456824.png (1.35 MB, 1024x1024)
1.35 MB PNG
>>109524678
anon linked this lora a few months back
>>
>>109523790
Make her sneeze into her right hand lmao
>>
can minimax do i2v 2d animation? like animate hentai?
>>
>>109524738
nope, impossible with current technology. maybe in 20~30 years though
>>
>>109524738
you belong in a straitjacket pal
>>
>>109524089
>Avril Lavigne
>not using a ThinkPad
I'm offended
>>
>>109524749
a bit rude
>>
>>109524188
>tfw early slop style is lost technology
>>
>>109523508
The “default” in Comfyui always uses suboptimal quality so “it just works” : 1) the model uses 768px on the shortest edge (768*768 1024*768 1152*768 1344*768); 2) using 20-25 with res_multistep (simple/beta) yields the same or better result in half the time and any caching mechanism is highly detrimental (https://github.com/HM-RunningHub/ComfyUI_RH_MinMaxH3/blob/main/docs/sampling.md) 3) It needs a very specific prompt style that you should always use 4) it’s already DISTILLED that’s why it runs fast and looks good (“Turbo” lora are for ADHDs, third world hardware, and addicted porn users)
>>
>>109524754
but true. anime gooners should not walk free
>>
>>109523599
You’re using sage attention flag AND sage attention patch AND sage attention h3? why would you do that? Sage node, specifically h3 one, explicitly says “it will override any attention mechanism” do you people even read
>>
>>109524337
Lol. “AI” (vllm) made retards even more retards and it’s just the beginning
>>
File: H3_cope_caches_08_10.jpg (111 KB, 1813x678)
111 KB JPG
These cope cashes combination gave good results. Little to no quality loss at 1MP.

source is anons.
>>
>>109524756
what model did they use?
>>
>>109524861
CogVideo probably.
>>
>do a gen or two
>35 minutes
>turn off turbo lora, shift, and spectrum
>21 minutes, with more steps no less
i dont understand
>>
File: 235735_00001.webm (279 KB, 768x768)
279 KB
279 KB WEBM
>>109524872
that looks like SVD
>>
>>109524882
OOM, ram swap keked.
>>
What is this sage shit is it literally saving time or just a couple of minutes of cope?
>>
>>109524966
it doesn't save any time it just does not bump the thread
>>
>>109524966
About 20% time reduction with no quality loss. It's great.
>>
>>109524678
not everyone has gpu with 4gb or more of vram
right to gen is a human right, everyone must have free access to ai
>>
>>109524751
She is a normie
>>
>>109523572
Is there any advantage using Load diffusion model int8 w8a8 node over the standard one?
>>
>>109525042
I don't know
>>
alright I’m getting lazy writing the prompts manually
can someone recommend r2v prompt builder that uses local kobold/llamacpp with vision?
>>
My job just gave me a Macbook Pro M5 Pro with 24gb RAM.
Realistically what local AI image generation model that I can run on this shit, I dont mind waiting a long time for one image.
>>
>>109525080
Probably Anima for anime, Krea2 for everything else, You're going to gen nsfw with your work laptop? Dont shit where you eat.
>>
>>109525080
>how anon loses his job within 2 weeks
>>
>>109524188
Just wait for the Anima crew to make a video model.
>>
Does anyone have any tips for faster generations on a 5090, for Illustrious or Anima? On Illustrious I get 8.5it/s, for Anima I get 4.99it/s
>>
There is a way to increase resolution and time in minimax h3 with only16gb vram?
>>
>>109525080
You're in luck
https://github.com/antirez/h3.c
>>
File: forge_neo.jpg (83 KB, 1129x255)
83 KB JPG
>>109525220
I use Forge neo for Anima and Illustrious. It's fast and simple. ComfyUI is for advanced and latest updates fuckery.
>>
>>109525263
I'm fairly comfortable with comfy I think, I have used wan2.2 a fair bit, but my gen speeds just feel a bit low for my hardware.
>>
>>109525220
get int8cr of the models
>>
>>109525220
is that not enough? faggot
>>
>>109525246
I haven't tested it but you could try this and chain together clips to increase the duration: https://github.com/NikoDemon80/ComfyUI-H3-Motion-Context
>>
>>109525313
Alright, I'll look into this, thanks.
>>109525322
I just feel like it should be faster.
>>
File: HPMlRCDaMAAQy3U.jpg (119 KB, 640x1216)
119 KB JPG
I'm so close to achieving what I want. if only it didin't add boobs.

https://litter.catbox.moe/hoqujsptakk1ucjd.mp4
>>
File: 998877.mp4 (2.99 MB, 1216x672)
2.99 MB
2.99 MB MP4
Coming up with video ideas is kinda harder than for the images...
>>
>>109525409
Her age?
>>
>>109525409
I look like this irl
>>
>>109525415
Nigga you already posted this yesterday
I know because I clicked on it
>>
>>109525415
uummm remielle box???
>>
>>109525409
You are a subhuman shitskin pajeet.
>>
>>109525432
quit moralising
>>
>>109525435
Stop being brown then. This is a white people's board.
>>
>>109525409
How do you do this?
>>
Been out of the loop for a couple days. Are turbo loras good yet or are we still using spectrum?
>>
>>109525409
you need to add an adjust audio volume node to your audio output cause I can't hear shit
>>
File: 666666666.mp4 (3.06 MB, 672x1216)
3.06 MB
3.06 MB MP4
>>109525420
my point still stands
>>109525424
emmm, just pictures of her with r2v model
>>
File: 7437831.webm (2 MB, 576x320)
2 MB
2 MB WEBM
difficult to make there be a co-pilot with this model. most of the time, it pans to the person recording the video instead of the person sitting next to the POV
>>
>Coming up with video ideas is kinda harder than for the images...
With images you can be vague, and let the model do the work. With video if you're vague then nothing happens.
>>
What's stopping me from asking AI to come up with a lewd prompt and then shaving a digit off the character's age?
>>
File: MiniMax_H3_00096e.webm (1.62 MB, 1104x1104)
1.62 MB
1.62 MB WEBM
>>
>>109525569
Why not just ask for lewd cunny prompt directly.
>>
>>109525569
Whatever you do, don't try it with krea2.
>>
>>109525612
w-why not?
>>
>>109525612
what’s gonna happen?
>>
File: those who know.gif (196 KB, 220x220)
196 KB GIF
>>109525635
>>109525636
>>
will my monkey brain be tricked into thinking i'm a chad if i gen my peepee inside hot goth girls and turn my room into a goth girl sanctuary?
>>
File: file.png (179 KB, 1017x631)
179 KB PNG
>Model MiniMaxH3 prepared for dynamic VRAM loading. 19995MB Staged
This works, I get a 6s video in 6min, but I can see my vram is only used up to 9GB roughtly. What happens in this case, is this killing my SSD or something ?
>>
>>109525662
why is your gpu sensor 75 and 51 c at the same time?
>>
https://litter.catbox.moe/b3cb0tytk2clk3j3.mp4
>>
>>109525687
ah, I need to change back the 51 one again, it's targeting the onboard gpu for some reason.
>>
>>109525612
What about H3??
>>
>>109524049
classic
>>
Is there a "continue video" flow? Or just adding the input video as a reference supposed to work?
>>
how do I remove the first frame from the video with the H3 template (the image ref) ? I know I could probably trim after the fact but...
>>
Can H3 gen faster if you half the length but double the FPS?
>>
>>109525829
sure, but why not just gen 32x32? you get garbage either way
>>
the new comfy kitchen attention seems to have better prompt adherence.
>>
>>109525887
>he's on the spectrum
grim
>>
does this comfy attention thing also work on image models?
>>
>>109525917
comfy barely works nowadays. hope this helps
>>
>>109525941
it kinda does. what should I use?
>>
I'm going to get a 5060 Ti 16 GB in the same machine that has a 4060 Ti 16GB, is this type of setup with two different cards common? and does it work better?
>>
>>109525887
>attention backend
the hell is this first time ive seen that node
>>
File: 00017-3907351177.jpg (367 KB, 1728x2880)
367 KB JPG
>>109526006
as someone who already has a 4060 ti 16g/64 gb ddr4 on a different pc build and now uses 5090 prebuilt. i highly advise you upgrade to a gpu with a higher vram count than 16gb. Your not seeing any meaningful benefit switching from the 4060 to the 5060 anon.
>>
>>109525204
Kek
>>
>>109525409
Nice. Kanna is sex.
>>
"1.0" of lightx2 turbo lora for h3 is out:

https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_fl2v_turbo_8step_v1.0_comfyui_bf16.safetensors
>>
Bros how do I prompt delicious jiggle physics like banana anon?
>>
>>109526089
>fl2v
into the trash
>>
>>109526058
hes talking about combining, not switching. anon search this topic in /lmg/ in archives, there's a lot more of this sort of thing there + with LLMs since the LLM space has so many more options and people can benefit from the marginal vram by squeezing out a bit of extra context size rather than just gen time / tokens per second. from what i understand, yeah itll work and the mixed generations arent too much of an issue either, though its gonna be moderately slower than having all the vram on one card, and i assume you miss out on some of the more modern tweaks while using all the vram
>>
>>109526058
dont listen to this retards advice. he didnt even know how to change his gpu fan speeds
>>
>>109526089
waiting for KJ to prune it
>>
>>109525948
All the alternatives are even worse, unfortunately. Maybe in a year or 2 LLMs will be able to make a replacement.
>>
>>109525948
Anistudio is probably the better choice though. More stable, the dev actually cares.
>>
>>109526139
if this is organic praise then feel free to link to any gen you posted in /ldg/ that was made in anistudio and catbox the metadata

we're all waiting, "anon"
>>
>>109525948
>what should I use?
wan2gp
>>
>>109526122
is this relevant?
>>
>>109526166
>>109526180 test
>>
>>109526180
test successful
>>
>>109526180
wtf
>>
>>109526139
Kill yourself Julien
>>
>>109526190
bro this was long time ago. let it go.
>>
>>109526184

>>109525955
>>109525973
>>109525989
>>109526007
>>
>>109525827
Comfy UI has a native video trimmer node
>>
File: kekekekkeeeeeeeeek.png (1.17 MB, 864x1184)
1.17 MB PNG
bruh wth is this nah its over JSID already keeeeeeeek
>>
This lora is life changing
https://civarchive.com/models/2839680?modelVersionId=3217238
I no longer have the urge to make money or do anything productive. I have everything I could ever want right here. Thank you, china
>>
local image to 3d that actually gives usable models when?
>>
Can anyone suggest good non-corporate UIs for stable diffusion?
>>
>>109526251
>>109526166
>>
>>109526262

dont worry anon Im a girl now, I used to be a genderless AI

>>109526230
>>
>>109526251
>>109526139
>>
>>109526213
is this some kind of ai assisted bot attack?
>>
>>109526262
ywnbaw
>>
>>109526262
>>109526180
What's the trick here?
Dead tranny board is so slow you can correctly call post number beforehand?
>>
>>109526272
It's disguised avatarfagging.
>>
>>109526288
Is it really avatarfagging if they arent using the same images? I thought avatar fagging meant using the same like image persona across threads
>>
>>109525887
More or less same as sage quality while just a tiny little bit faster for me.
Int8 version of sage, no clue about other versions since I can't use them.
>>
>>109526089
I think it's usable, but >>109526100 has a point.
references are adictive.
>>
https://www.reddit.com/r/StableDiffusion/comments/1vl3ed0/having_bad_ref2va_quality_compared_to_fl2va_try/
better r2v quality? still downloading.
>>
>>109526303
It serves the same purpose.
>>
https://www.reddit.com/r/StableDiffusion/comments/1vlga9z/lightx2v_minimax_h3_8step_turbo_v10/

anyone try it? I am avoiding turbo till it's better than spectrum setup.
>>
>>109526089
>>109526350
I tested it and it seems to add slow mo. How the fuck do they keep getting slow mo, only really tested it with a couple gens so far. The 600_ema one from the other dudes seems better.
>>
>>109526350
The lora works with spectrum, what are you on about?
>>
try bypassing sage/sol stuff with the new commit

https://github.com/Comfy-Org/ComfyUI/commit/bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9#diff-fab3fbd81daf87571b12fb3e4d80fc7d6bbbcf0f3dafed1dbc55d81998d82539

one person says:

You just need to update your Comfyui and you can either start it with the --use-ck-attention flag so all models use the comfy-kitchen attention backend

total time with MiniMax H3 Mem Eff Sage Attention Patch + lightx2v lora = 294s

total time with ModelAttentionBackend only (no MiniMax H3 Mem Eff Sage Attention Patch + lightx2v lora) = 224s
>>
why don't AIs understand that pantyhoses should cover the whole legs
>>
>>109526225
oh I'm dumb I saw that one. thx anon.
>>
>>109526089
wish I'd paid attention to the fucked sound discourse now
>>
>>109526363
NTA but post a screenshot of your spectrum parameters.
It "works" for me but the quality is terrible.
>>
>>109526384
did comfy update, it's getting this:

Downloading comfy_kitchen-0.2.30-cp312-abi3-win_amd64.whl (38.0 MB)

so ill try and compare vs regular sage attn.
>>
>>109526350
Waiting for kijai or someone else to convert
>>
why is everyone training t2v loras when the ref2v model is much better?
>>
>>109526439
ref2v model is ass, I think fl2va works better for reference
>>
>>109526423
Oh they themselves have converted, nice.
https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_fl2v_turbo_8step_v1.0_comfyui_bf16.safetensors
Might need the custom node for the pruned model still, lets see.
>>
>>109526439
Ref is nice but sometimes I JUST WANNA GEN!
>>
>>109526439
i guess it's easier to test what works with the t2v model first, then apply those findings to the reference model
>>
Desu I have yet to touch ref model.
I expect it to slow down a lot with each reference added, and f2v already runs slow enough for me.
>>
>>109526505
i run it with 9 reference images and it doesnt slow it down much at all
>>
>>109526505
videos slow it down a lot, but images barely slow it down at all
>>
is anyone old enough to remember when graphics cards were used for actual graphics?
>>
>>109526528
what if you have the remove background feature on tho? It has to do that processing and 9 images is going to take longer than 2 or 3
>>
>>109526384
Pulled
From 65s/it now got 175s/it on first step, second step took longer than 400s and I killed it off as my system began to freeze

Thanks
>>
>>109526514
>>109526528
That's cool to know.
I might actually try it now, thanks.
>>
updated comfy, added --use-ck-attention to startup args in notepad++, bypassed the sage attn kj node, spectrum still enabled, seems to be fast, may be slightly slower but some claim the quality is better with comfykitchen

comparison (new, comfy kitchen vs the same prompt yesterday):

https://files.catbox.moe/bcrbsw.mp4
>>
>>109526545
mine still is. 1 frame per minute
>>
>>109526561
this is with the startup flag, im gonna try the ModelAttentionBackend node now and no startup args.
>>
>>109526561
piece of shit website not working, try this

https://litter.catbox.moe/os64l1jeeg45wwu3.mp4
>>
is there a discord or something? It has been days and it seems really bad here.
>>
File: 1782847449880648.png (360 KB, 1653x879)
360 KB PNG
>>109526601
so basically, try this after updating comfy.
>>
File: 1779783726980868.jpg (224 KB, 1277x1231)
224 KB JPG
KEK
>>
>>109526561
It's causing sampler errors when I try to use the backend attention node, but I still have sage attention in my startup flags. Maybe those are incompatible. I'll try putting --use-ck-attention in my args and ditching other attention nodes.
>>
>>109526528
Any strategies for making video references more efficient? Can you lower the framerate/res or anything like that?
>>
>>109526613
trying now, node seems fine for me, goes after spectrum.
>>
>>109526605
sharty raids and now they are attempting to bot. stop using sharty threads too
>>
>>109526611
spectrum + h3 cache. you're butchering every gen you make
>>
>>109526620
this. need a cozy bake
>>
>>109526612
literally me
>>
>>109526620
give disco please for all anything you know it will help.
>>
>>109526612
accurate
>>
>>109526624
someone make a thread without the spam links
>>
0.3mp seems better quality with comfy kitchen node, generally at this low setting you get a lot of noise on sage, but I need to gen more to be sure:

https://litter.catbox.moe/qpf9ct98tm4ma9n7.mp4
>>
>>109526655
>>109526655
>>109526655
>>
can we get a real bake not a troll one?
>>
>>109526102
wouldn't he need another stronger psu to handle the wattage draw of having another gpu run simultaneously on his build. he would just be better off using a uncensored llm on a huggingface space.
>>
>>109526410
I'm using the default node params. You DID update the node and delete and replace the node after said update, right?
>>
File: file.png (32 KB, 939x207)
32 KB PNG
>>109526561
>>
>>109526717
Kitchen attention is causing my sampler to error out. No clue why I'm too much of a faggot to understand this stuff.

quant_qk_per_thread_int8: Q/K base pointers and B/H/N strides must preserve 4-element alignment
>>
>>109526731
are the settings default? simple/res multistep?
>>
>>109526710
Yeah.
The quality is fucking shit when compared to Spectrum only or lora only outputs.
You DID do A/B testing to see how well they combine, right?
>>
>>109526731
its brand new, make a bug report for it with steps to reproduce using stock nodes
>>
File: MiniMax_H3_00167_.webm (1.19 MB, 576x576)
1.19 MB
1.19 MB WEBM
>>109524188
You can get some funky effects if you fiddle with internal parameters. This is with 5 scheduler steps and 0.71 denoise. Prompt is just "Will Smith eating spaghetti."

There's a funky music playing in the background.
>>
File: MiniMax_H3_00168.webm (1.44 MB, 576x576)
1.44 MB
1.44 MB WEBM
>>109526757
This one's 20 steps.

>>>/wsg/6212176
>>
Anyone baking a proper thread?
>>
>>109526861
I would move if one does.
>>
>>109526809
he has to eat with his hands
>>
>>109526861
let's just use the schizo thread, that way he stfus and rentry links are posted as one of the top comments anyway
he's going to have nothing to complain about that way and will crawl back into his hole temporarily
>>
File: MiniMax_H3_00170.webm (1.15 MB, 576x576)
1.15 MB
1.15 MB WEBM
>>109526882
I'll keep playing around with settings.
>>
now he's trying to splitbake /adt/ again. still waiting on real /ldg/ one
>>
so he's shitting up TWO generals and the jannies just don't care. we really need two AI boards.
>>
>>109527009
Jannies not caring is the reason why there is no AI board doe.
>>
>>109526964
>>109527009
At the moment /ldg/ is too slow for him, these kind of schizos need their fix no matter what and they want it quick, that's probably why he's going elsewhere.
>>
>>109524236
This is the only thing from windows 7 I'm nostalgic for
>>
>>
>>109527078
me
>>
real thread is up
>>
>>109527134
Which one and why?
>>
damn these h3 gens are getting better and better
>>
>>109527009
>we really need two AI boards
No, we just need jannies that don't suck.
>>
What's the main thread for H3 gens with sound without having to click catbox links?
>>
>>109527954
Probably /aicg/ on /wsg/
>>
>>109524367
the portal opens and out steps a slack jawed zitty doordash guy holding a greasy bag and a drink with askew top and bent straw. i lack the resources
>>
>>109526749
Late response is late, but yeah. Mine turn out fine. Use 12 steps on euler beta.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.