[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: ms_00013_.png (1.51 MB, 1536x1536)
1.51 MB PNG
Submit your video gens for Sexy Jam 1:
https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewform

Previous: >>109554288

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
neither thread has a collage fuck off
>>
Is it true that ani ruined genjam last year?
>>
Blessed thread of frenship
>>
>>109557428
collagecucks eternally BTFO'd
you're too slow!
>>
>>109557428
Don't care, autistic retard.
>>
>>109557431
his gens were so good catjak had a melty and canceled it
>>
File: 1786736243947507.png (859 KB, 1114x840)
859 KB PNG
what is miku doin'
>>
File: 1760600801551750.jpg (306 KB, 1920x1081)
306 KB JPG
>>109557412
better quality:

https://files.catbox.moe/a32c3o.mp4

generic prompt will work just plug in any 2 images/characters.

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

<Picture 2> is the physical reference for <Subject 2> with the same clothing.

[A 10-second ultra-high-action combat scene on a rainy Tokyo skyscraper rooftop at night. Glowing neon signs blur in the background. Matrix-style cinematic slow-motion, hyper-detailed physics, raindrops freezing in mid-air, dynamic tracking shots, crisp focus, and 4k texture.]

00:00.000 - [SCENE 1: <Subject 2> is holding a classic beige acoustic guitar. Extreme close-up on <Subject 2>'s face, eyes narrowing. The camera rapidly whips down to their hands as they spin a classic beige acoustic guitar and hurl it forward like a massive projectile. [Sound: Wooden and metallic whoosh, slicing through the air, torrential rain humming]]

00:02.500 - [SCENE 2: The camera cuts to a sweeping bullet-time arc around <Subject 1>. The video drops into extreme slow-motion as the spinning beige acoustic guitar flies toward them. Raindrops completely freeze in place. [Sound: Time-dilation drone, deep bass drop, slow-motion air displacement]]

00:05.000 - [SCENE 3: A dramatic low-angle tracking shot. <Subject 1> bends backward at an impossible angle to dodge, exactly like Neo dodging bullets. The matte beige wooden body of the guitar skims inches above their chest, leaving a visible rippling wind-trail in the air. [Sound: Intense aerodynamic rushing sound, distorted acoustic wood echo passing by]]

00:07.500 - [SCENE 4: The video snaps back to normal speed. <Subject 1> flips upright, catching their balance on the wet rooftop. In the background, the thrown beige guitar smashes into a massive neon billboard, shattering into wooden splinters amidst a brilliant shower of electrical sparks. [Sound: Loud wooden splintering impact, explosive electrical glass shattering crackle]]

00:10.000 - [END OF VIDEO]
>>
File: 1773227206257754.mp4 (1.75 MB, 736x576)
1.75 MB
1.75 MB MP4
>>
File: 00030-650973621.jpg (590 KB, 3008x2112)
590 KB JPG
is this the real thread or not? make up your minds and stop with the dumb trolling or jannies are just going to 404 the threads out of spite.
>>
>>109557462
gaht JIGGLY DAMN hALOOMBAGA
>>
>>109557428
not OP. Just have one ready and post it, then if if it's decent someone might post it in a future thread. Another point would be anons had plenty of time to make a thread and yet didn't
>>
>>109555905
what did you use to make this image?
>>109557462
nice jiggle
>>
>>109557462
If only Zack Snyder had used slo-mo for this instead of CGI sesame seeds.
>>
File: 1777773587819816.webm (3.89 MB, 704x896)
3.89 MB
3.89 MB WEBM
>>
>>109557467
this, so fucking annoying and retarded. i dont know how the jannies allow it at all, they should insta-nuke whatever thread doesn't post first
>>
>>109557486
tsmt
>>
Will Qwen3.8 heretic be the king of Minimax prompt writing?
>>
>mfw Resource news

08/14/2026

>ReDetail: Generative video re-detailer through LTX-2.5's pixel spatial upscaler,
https://github.com/Bambushu/redetail

>V-RAE: Rethinking Video Latent Spaces for Generation
https://v-rae.github.io

>4 step ref2v Minimax H3 turbo Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>Learning Unified Video and Image Representation for Video Face Forgery Detection
https://github.com/haotianll/UVIF

>AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design
https://github.com/Yaxin9Luo/AutoDesign

>H3 Motion Context — Timeline
https://github.com/BSG-Walter/ComfyUI-H3-Motion-Context-Timeline

>RTX PRO 6000 Blackwell workstation edition price doubles, as NVIDIA launches new AI model
https://www.neowin.net/news/rtx-pro-6000-blackwell-workstation-edition-price-doubles-as-nvidia-launches-new-ai-model

>MAGI-2 Preview: Scaling Video Generation Models Efficiently
https://sand.ai/blog/magi-2-preview

08/13/2026

>Lightx2v MiniMax H3 Turbo Ref2V 4Step/8Step Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>GemmaPrompt: Local prompt enhancer for ComfyUI diffusion models
https://github.com/whp199/GemmaPrompt

>ComfyUI-H3-FaceRefine
https://github.com/Carasibana/ComfyUI-H3-FaceRefine

>Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising
https://github.com/Ai-ZL/Hybrid-LUT

>MiniMax H3 Creator for ComfyUI: Multi-Shot 60s Timelines, Resizable Satellite Stage, & Ollama/LM Studio Refiner
https://github.com/roadmaus/ComfyUI-MiniMax-Creator

08/12/2026

>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler
https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler

>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready
https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy

>LTX-2.5 22B — GGUF
https://huggingface.co/realrebelai/LTX-2.5_GGUFs

>ComfyUI NVIDIA RTX VSR Pro
https://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro
>>
>mfw Research news

08/14/2026

>SNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolation
https://arxiv.org/abs/2608.13460

>Spatially-Grounded Text-to-Video Generation via Inference-Time Gradient-Free Optimization
https://arxiv.org/abs/2608.13037

>SketchSense: Learning to Interpret Imperfect Sketch Guidance for Image Inpainting
https://arxiv.org/abs/2608.13186

>Semantic Steering for Controllable Generation: Tuning-Free Concept Erasure in Multimodal Diffusion Transformers
https://arxiv.org/abs/2608.12829

>MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval
https://arxiv.org/abs/2608.12532

>From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion
https://arxiv.org/abs/2608.13043

>Erase but Preserve: Controllable Removal of Copyrighted Animation Characters via Optimized Semantic Anchors
https://arxiv.org/abs/2608.12806

>SCOPE: Subspace Clustering with Online Per-Head Top-K Estimation for Sparse Video Attention
https://arxiv.org/abs/2608.12780

>StrAD: A Streaming Method and Benchmark for Audio Description Generation for Long-form Videos
https://arxiv.org/abs/2608.12549

>HPSD: Hybrid-Policy Self-Distillation for Text-Image-to-Video Diffusion Models
https://bujiazi.github.io/hpsd.github.io

>Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation
https://hmrishavbandy.github.io/cmd-site

>A Controlled Study of Self-Supervised Image and Video Pretraining under Limited Resources
https://arxiv.org/abs/2608.13183

>PixSDS: Why Latent SDS Makes Noisy Pixels
https://sevashasla.github.io/pixsds-webpage

>RIG-RoPE: Relation-Stratified Multimodal Attention with Instance-Local Rotary Geometry and Representation-Aware Traversal Coordinates
https://arxiv.org/abs/2608.05154
>>
>>109557483
I guess you used turbo lora?
>>
File: MiniMax_H3_00023_.webm (1.13 MB, 864x480)
1.13 MB
1.13 MB WEBM
>>
>>109557402
miku, i don't have your money, i swear i only need 3 more days, please don't look at me like that
>>
File: migu-elo1.mp4 (1.81 MB, 736x576)
1.81 MB
1.81 MB MP4
Another Migu music video: https://files.catbox.moe/7ap6sd.mp4

Use refernce video and audio exactly the way it is, duplicating the actions and edits. The subject are: S1 chibi chubby hatsune miku, who has long teal hair tied in twintails, turquoise eyes, and wears her signature idol outfit, and S2 chibi chubby Kagamine Rin, who has short bob-cut straight blonde hair with a pointy upright white bow, and wears her signature idol outfit. Replace the original actors as directed below but leave everything else the same.

cut 1: (cut) 00:00.00 to 00:03.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 2: (cut) 00:03.00 to 00:04.00 S2 sings at a mic
cut 3: (cut) 00:04.00 to 00:06.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 4: (cut) 00:06.00 to 00:08.00 S2 sings at a mic
cut 5: (cut) 00:08.00 to 00:10.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 6: (dissolve) 00:10.00 to 00:15.00 S2 sings at a mic

It mostly followed it. It's a huge pain in the ass ripping video from youtube now, you basically have to record the screen and system audio, nothing else works anymore.
>>
>>109557456
tweaked a bit, better slow mo this time

https://files.catbox.moe/d2bl19.mp4
>>
File: i2v_MiniMax_H3_00036_.mp4 (1.98 MB, 928x704)
1.98 MB
1.98 MB MP4
>>109557454
>>
File: comfy studio irl.jpg (253 KB, 1194x1274)
253 KB JPG
most in depth look at comfy office and comfy's employees and comfy
>>
can someone please explain debo's autistic fixation on will smith and spaghetti? like i don't get it.
>>
>>109557536
why do vocaloids need to consume flesh?
>>
File: 1759033788238352.webm (3.9 MB, 704x896)
3.9 MB
3.9 MB WEBM
>>109557522
yea
>>
>>109557509
>>109557514
fuck off unemployed loser
>>
>>109557547
jesus christ that poor horse
>>
>>109557544
I don't understand why you think ani's threads are made by debo. He only started doing the splitbaking routine after the ani rentry was added back in december.
>>
why is the OP seething and spamming the other thread?
>>
>>109557560
melties
>>
>>109557547
Peak American physique.
>>
File: MiniMax_H3_00154_.webm (1021 KB, 608x1024)
1021 KB
1021 KB WEBM
>>
File: ldg.png (19 KB, 1893x79)
19 KB PNG
>>109557560
>>
>>109557577
>figurine
perfect size for my pp
>>
pay no attention to ani or debo. just make sure to keep the two rentry links in the OP.
>>
i need to get H3 to say kikes properly, tried Kaikes and that didn't work, any suggestions?
>>
>>109557588
yup o7
>>
>>109557588
>pay no attention
>pay attention
>>
>>109557588
it wouldn't be a schizo echo chamber if it wasn't there
>>
>>109557534
holy shit, this model. this time in place of the bocchi images I just used swimsuit marciana (NIKKE).

can I submit this for the lewd jam?

https://files.catbox.moe/hrwn23.mp4
>>
File: 1766675567898433.jpg (228 KB, 902x1625)
228 KB JPG
>>109557606
source image used:
>>
>>109557577
good panty pull, usually it's jankmaxxing
>>
>>109557606
yes
>>
File: 1760712745727230.webm (3.9 MB, 704x896)
3.9 MB
3.9 MB WEBM
>>109557574
it's the form that matters
>>
Lol. Ani thinks that by spamming his own thread he'll get this one deleted. Let's see if he's right.
>>
File: 1778399418915787.jpg (92 KB, 960x544)
92 KB JPG
>>109557534
what is that?
>>
File: WBCaique.jpg (172 KB, 850x1500)
172 KB JPG
>>109557593
maybe caiques?
>>
>>109557629
also this one worked good too, got the 2 characters. one is tia.

https://files.catbox.moe/chev0g.mp4
>>
>>109557642
thats me generating at 0.4mp cause I want to test fast
>>
>>109557631
ngl i was really expecting her to start drilling through the earth
>>
>>
File: GPUPrices.ai - 8GB.png (269 KB, 1415x868)
269 KB PNG
>>109556485
>i assume you're just asking to waste my time
Sorry if it came off that way. It wasn't my intention. I'm considering ordering a GPU. Since I haven't tried image generation in a serious manner (a few times online, free ones), I don't wanna overspend. Basically, I'm considering trying it, but I'm unaware if it's gonna work out. Having an idea about bare minimum VRAM requirements might make my search easier. As for LLMs, I was asking in case certain LLMs might be better suited for certain topics. I'm not too knowledgeable about AI stuff.
>i hope someone reading this derives benefit from it
I have. I now know 8GB might be enough for me to try serious image generation for the first time.
>adding a lora makes basically no difference
Did you mean LoRas don't lower VRAM requirements? With my initial LoRa question, I was wondering if 8GB is enough for offline LoRa training. Sorry if I could've been more specific.
>>
What model are (you) using? I want to try some new stuff
>>
kek

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

front view of <Subject 1> running toward the camera holding a silver pistol in the air. the camera distance remains the same as she runs forward.

https://files.catbox.moe/0uh72j.mp4
>>
would the vram problem be solved if models could run across multiple gpus? i've got a few old gpus collecting dust that could actually be put to use if this were a thing
>>
File: 00049-424934055.jpg (490 KB, 1472x2880)
490 KB JPG
>>
>>109557705
yeah, already happened. just not diffusion transformers
>>
okay now I got the proper camera. low mp to test.

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

medium front view of <Subject 1> running toward the camera holding a silver pistol in the air. the camera tracks her remains the same distance as she runs forward.

https://files.catbox.moe/d2hzrv.mp4
>>
>>109557723
catbox didnt work.

https://litter.catbox.moe/8wtd2d4i0en9193f.mp4
>>
so second image as the background, just say the setting is <Picture 2> and it will work. insert whatever bg you like. (this is RE9).

https://files.catbox.moe/74q91b.mp4
>>
>>109557675
I too enjoy a cuppa coffee
>>
Guys, is there a guide for extracting clean dialogue from an audio clip?

I got this nodepack for comfy: https://github.com/diodiogod/TTS-Audio-Suite
But when I clicked the VOCAL/NOISE REMOVAL GUIDE: https://github.com/diodiogod/TTS-Audio-Suite/blob/main/docs/VOCAL_REMOVAL_GUIDE.md
I noticed it isn't actually a guide at all, just a list of types and options.
Since this dumbass doesn't know how to write a guide, does anyone have a workflow to spare? I just wanna clone voices.
>>
File: MiniMax_H3_00067_.mp4 (751 KB, 928x672)
751 KB
751 KB MP4
Minimax Music.cpp is here
https://github.com/ServeurpersoCom/minimaxmusic.cpp

The RVQ encoder has been reverse-engineered
https://huggingface.co/MiniMaxAI/MiniMax-Music3/discussions/5

So with Minimax Music cpp we'll soon be able to remix songs (continue, extend, etc...) and the LoRA trainers are coming.

Of course Comfy doesn't care about music, also .cpp is vramlet friendly and faster on high VRAM GPUs so go there music frens!

https://files.catbox.moe/3kvixg.wav
>>
Can I use Minimax H3 on my RTX 3060 12gb and 32 RAM?
>>
got a nice Rio:

https://files.catbox.moe/r04e4b.mp4

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

medium front view of <Subject 1> wearing a black bodysuit, slowly walking toward the camera holding a silver pistol in the air, while remaining silent. the camera tracks her and remains the same distance as she walks forward.
>>
>>109557821
Yes with the right comfyui settings it's perfectly possible. The question will be if you can generate at a speed, duration, and resolution that is acceptable to you.
>>
>>109557774
Awesome I didn't realize the weights were out, I'll have to give it a try some time.
>>
Is it possible to downclock my RTX 5070ti to 50% of power target ? Genning with Minimax really makes my GPU fan screams
>>
>>109557854
I notice no diff at 70% for my 4080, quieter fans etc. 50 may be too much idk.
>>
File: MiniMaxH3_00091.png (778 KB, 640x1152)
778 KB PNG
okay h3 totally fucked up kikes again but everything else turned out perfectly KAIRI KINO INCOMING

https://d.uguu.se/iisGsvBg.mp4

>>109557647
lol might try this. poor cute fellas have to share a name with my least favorite people on the planet.
>>
>>109557767
you probably need software to strip the vocals, there is Audacity that has an add-on AI feature to strip all the vocals, instruments etc into stems so you can then edit or simply use as is. There is no doubt other software that can do it as well, maybe just for stripping the vocal a basic vocal remover.
>>
>>109557839
Rio again, competing for sexy jam

https://files.catbox.moe/odqafy.mp4
>>
>>109557864
How to do below 70% ? Nvidia App only do down to 84%
>>
>>109557876
msi afterburner slider, mine is at 70% pl. 50 may be way too low, but you can try
>>
>>109557854
nvidia-smi -pl 150

Limits it to 150W
>>
>>109557866
nice
>>
>>109557854
deshroud and replace with bigger fans (bigger = quieter).
>>
>>109557909
>Limits it to 150W
wow that's really low!!!
>>
>>109557353
quick proof of concept
https://files.catbox.moe/516svi.mp4
was gonna go with the classic pendulum but I'm sure it can do that fine too
>>
File: 56287623.mp4 (1.03 MB, 1056x608)
1.03 MB
1.03 MB MP4
>>109557402
>>
>>109557928

Is that not half of the tdp? I just googled it for 3s. I keep my 3090 at 250 - 300W instead of 370W. Not that much slower gen times but much less heat.
>>
>>109557841
Good point. What speed, duration and resolution can I expect with my specs?
>>
File: 1772596712151176.jpg (59 KB, 500x900)
59 KB JPG
Elegg in RE test: source image and the output.

https://files.catbox.moe/j3hbln.mp4

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

full body front view of <Subject 1> wearing a yellow thong, who turns around to show her ass and runs away while completely silent, the camera follows her as she moves, keeping the same distance.
>>
>>109557954
I can't say, my specs are different. It's also gonna depend on what you're genning. If you're running inference with the ref2v model and using a video reference the gen times are way different than if you're just using one reference image.
>>
File: file.jpg (44 KB, 584x364)
44 KB JPG
any h3 lora to stop that horrible crunch sound with blowjobs or sucking in general?
in fact, any lora even focusing on audio for h3?
>>
Is the lightx2v turbo lora generally seen as a better optimization than spectrum? What are the pros and cons of each?
>>
>>109557462
I miss movies like that
>>
>>109557992
prompting will do it, worked fine for me
>Lips sliding across wet skin, gagging and muffled moans.
>>
>>109557584
Yep, and people still fall for this shit.
It's not a melty, it's just retarded sharties. Years of that.
>>
>>109557996
I personally find it much better because less time and the quality is still good at 8 steps. spectrum is skipping some of the 20 steps but in a different way.
>>
>>109557953
I'm surprised it's stable.
>>
>>109557992
yea the turbo lora fixed it. just prompt for blowjob sounds
>>
File: 1773291149273119.jpg (30 KB, 1836x92)
30 KB JPG
>>109557774
OK this is amazing, it's not even reverse engineered, it's just there.
Did they forget it? Or "forget it"? Doesn't matter much but man is this good for the future of local music.
>>
>>109557875
the sigh at the end did it for me



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.