[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: ms_00013_.png (1.51 MB, 1536x1536)
1.51 MB PNG
Submit your video gens for Sexy Jam 1:
https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewform

Previous: >>109554288

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
neither thread has a collage fuck off
>>
Is it true that ani ruined genjam last year?
>>
Blessed thread of frenship
>>
>>109557428
collagecucks eternally BTFO'd
you're too slow!
>>
>>109557428
Don't care, autistic retard.
>>
>>109557431
his gens were so good catjak had a melty and canceled it
>>
File: 1786736243947507.png (859 KB, 1114x840)
859 KB PNG
what is miku doin'
>>
File: 1760600801551750.jpg (306 KB, 1920x1081)
306 KB JPG
>>109557412
better quality:

https://files.catbox.moe/a32c3o.mp4

generic prompt will work just plug in any 2 images/characters.

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

<Picture 2> is the physical reference for <Subject 2> with the same clothing.

[A 10-second ultra-high-action combat scene on a rainy Tokyo skyscraper rooftop at night. Glowing neon signs blur in the background. Matrix-style cinematic slow-motion, hyper-detailed physics, raindrops freezing in mid-air, dynamic tracking shots, crisp focus, and 4k texture.]

00:00.000 - [SCENE 1: <Subject 2> is holding a classic beige acoustic guitar. Extreme close-up on <Subject 2>'s face, eyes narrowing. The camera rapidly whips down to their hands as they spin a classic beige acoustic guitar and hurl it forward like a massive projectile. [Sound: Wooden and metallic whoosh, slicing through the air, torrential rain humming]]

00:02.500 - [SCENE 2: The camera cuts to a sweeping bullet-time arc around <Subject 1>. The video drops into extreme slow-motion as the spinning beige acoustic guitar flies toward them. Raindrops completely freeze in place. [Sound: Time-dilation drone, deep bass drop, slow-motion air displacement]]

00:05.000 - [SCENE 3: A dramatic low-angle tracking shot. <Subject 1> bends backward at an impossible angle to dodge, exactly like Neo dodging bullets. The matte beige wooden body of the guitar skims inches above their chest, leaving a visible rippling wind-trail in the air. [Sound: Intense aerodynamic rushing sound, distorted acoustic wood echo passing by]]

00:07.500 - [SCENE 4: The video snaps back to normal speed. <Subject 1> flips upright, catching their balance on the wet rooftop. In the background, the thrown beige guitar smashes into a massive neon billboard, shattering into wooden splinters amidst a brilliant shower of electrical sparks. [Sound: Loud wooden splintering impact, explosive electrical glass shattering crackle]]

00:10.000 - [END OF VIDEO]
>>
File: 1773227206257754.mp4 (1.75 MB, 736x576)
1.75 MB
1.75 MB MP4
>>
File: 00030-650973621.jpg (590 KB, 3008x2112)
590 KB JPG
is this the real thread or not? make up your minds and stop with the dumb trolling or jannies are just going to 404 the threads out of spite.
>>
>>109557462
gaht JIGGLY DAMN hALOOMBAGA
>>
>>109557428
not OP. Just have one ready and post it, then if if it's decent someone might post it in a future thread. Another point would be anons had plenty of time to make a thread and yet didn't
>>
>>109555905
what did you use to make this image?
>>109557462
nice jiggle
>>
>>109557462
If only Zack Snyder had used slo-mo for this instead of CGI sesame seeds.
>>
File: 1777773587819816.webm (3.89 MB, 704x896)
3.89 MB
3.89 MB WEBM
>>
>>109557467
this, so fucking annoying and retarded. i dont know how the jannies allow it at all, they should insta-nuke whatever thread doesn't post first
>>
>>109557486
tsmt
>>
Will Qwen3.8 heretic be the king of Minimax prompt writing?
>>
>mfw Resource news

08/14/2026

>ReDetail: Generative video re-detailer through LTX-2.5's pixel spatial upscaler,
https://github.com/Bambushu/redetail

>V-RAE: Rethinking Video Latent Spaces for Generation
https://v-rae.github.io

>4 step ref2v Minimax H3 turbo Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>Learning Unified Video and Image Representation for Video Face Forgery Detection
https://github.com/haotianll/UVIF

>AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design
https://github.com/Yaxin9Luo/AutoDesign

>H3 Motion Context — Timeline
https://github.com/BSG-Walter/ComfyUI-H3-Motion-Context-Timeline

>RTX PRO 6000 Blackwell workstation edition price doubles, as NVIDIA launches new AI model
https://www.neowin.net/news/rtx-pro-6000-blackwell-workstation-edition-price-doubles-as-nvidia-launches-new-ai-model

>MAGI-2 Preview: Scaling Video Generation Models Efficiently
https://sand.ai/blog/magi-2-preview

08/13/2026

>Lightx2v MiniMax H3 Turbo Ref2V 4Step/8Step Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>GemmaPrompt: Local prompt enhancer for ComfyUI diffusion models
https://github.com/whp199/GemmaPrompt

>ComfyUI-H3-FaceRefine
https://github.com/Carasibana/ComfyUI-H3-FaceRefine

>Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising
https://github.com/Ai-ZL/Hybrid-LUT

>MiniMax H3 Creator for ComfyUI: Multi-Shot 60s Timelines, Resizable Satellite Stage, & Ollama/LM Studio Refiner
https://github.com/roadmaus/ComfyUI-MiniMax-Creator

08/12/2026

>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler
https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler

>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready
https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy

>LTX-2.5 22B — GGUF
https://huggingface.co/realrebelai/LTX-2.5_GGUFs

>ComfyUI NVIDIA RTX VSR Pro
https://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro
>>
>mfw Research news

08/14/2026

>SNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolation
https://arxiv.org/abs/2608.13460

>Spatially-Grounded Text-to-Video Generation via Inference-Time Gradient-Free Optimization
https://arxiv.org/abs/2608.13037

>SketchSense: Learning to Interpret Imperfect Sketch Guidance for Image Inpainting
https://arxiv.org/abs/2608.13186

>Semantic Steering for Controllable Generation: Tuning-Free Concept Erasure in Multimodal Diffusion Transformers
https://arxiv.org/abs/2608.12829

>MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval
https://arxiv.org/abs/2608.12532

>From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion
https://arxiv.org/abs/2608.13043

>Erase but Preserve: Controllable Removal of Copyrighted Animation Characters via Optimized Semantic Anchors
https://arxiv.org/abs/2608.12806

>SCOPE: Subspace Clustering with Online Per-Head Top-K Estimation for Sparse Video Attention
https://arxiv.org/abs/2608.12780

>StrAD: A Streaming Method and Benchmark for Audio Description Generation for Long-form Videos
https://arxiv.org/abs/2608.12549

>HPSD: Hybrid-Policy Self-Distillation for Text-Image-to-Video Diffusion Models
https://bujiazi.github.io/hpsd.github.io

>Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation
https://hmrishavbandy.github.io/cmd-site

>A Controlled Study of Self-Supervised Image and Video Pretraining under Limited Resources
https://arxiv.org/abs/2608.13183

>PixSDS: Why Latent SDS Makes Noisy Pixels
https://sevashasla.github.io/pixsds-webpage

>RIG-RoPE: Relation-Stratified Multimodal Attention with Instance-Local Rotary Geometry and Representation-Aware Traversal Coordinates
https://arxiv.org/abs/2608.05154
>>
>>109557483
I guess you used turbo lora?
>>
File: MiniMax_H3_00023_.webm (1.13 MB, 864x480)
1.13 MB
1.13 MB WEBM
>>
>>109557402
miku, i don't have your money, i swear i only need 3 more days, please don't look at me like that
>>
File: migu-elo1.mp4 (1.81 MB, 736x576)
1.81 MB
1.81 MB MP4
Another Migu music video: https://files.catbox.moe/7ap6sd.mp4

Use refernce video and audio exactly the way it is, duplicating the actions and edits. The subject are: S1 chibi chubby hatsune miku, who has long teal hair tied in twintails, turquoise eyes, and wears her signature idol outfit, and S2 chibi chubby Kagamine Rin, who has short bob-cut straight blonde hair with a pointy upright white bow, and wears her signature idol outfit. Replace the original actors as directed below but leave everything else the same.

cut 1: (cut) 00:00.00 to 00:03.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 2: (cut) 00:03.00 to 00:04.00 S2 sings at a mic
cut 3: (cut) 00:04.00 to 00:06.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 4: (cut) 00:06.00 to 00:08.00 S2 sings at a mic
cut 5: (cut) 00:08.00 to 00:10.00 S1 sings at a mic in the foreground while S2 sings in the background, out of focus
cut 6: (dissolve) 00:10.00 to 00:15.00 S2 sings at a mic

It mostly followed it. It's a huge pain in the ass ripping video from youtube now, you basically have to record the screen and system audio, nothing else works anymore.
>>
>>109557456
tweaked a bit, better slow mo this time

https://files.catbox.moe/d2bl19.mp4
>>
File: i2v_MiniMax_H3_00036_.mp4 (1.98 MB, 928x704)
1.98 MB
1.98 MB MP4
>>109557454
>>
File: comfy studio irl.jpg (253 KB, 1194x1274)
253 KB JPG
most in depth look at comfy office and comfy's employees and comfy
>>
can someone please explain debo's autistic fixation on will smith and spaghetti? like i don't get it.
>>
>>109557536
why do vocaloids need to consume flesh?
>>
File: 1759033788238352.webm (3.9 MB, 704x896)
3.9 MB
3.9 MB WEBM
>>109557522
yea
>>
>>109557509
>>109557514
fuck off unemployed loser
>>
>>109557547
jesus christ that poor horse
>>
>>109557544
I don't understand why you think ani's threads are made by debo. He only started doing the splitbaking routine after the ani rentry was added back in december.
>>
why is the OP seething and spamming the other thread?
>>
>>109557560
melties
>>
>>109557547
Peak American physique.
>>
File: MiniMax_H3_00154_.webm (1021 KB, 608x1024)
1021 KB
1021 KB WEBM
>>
File: ldg.png (19 KB, 1893x79)
19 KB PNG
>>109557560
>>
>>109557577
>figurine
perfect size for my pp
>>
pay no attention to ani or debo. just make sure to keep the two rentry links in the OP.
>>
i need to get H3 to say kikes properly, tried Kaikes and that didn't work, any suggestions?
>>
>>109557588
yup o7
>>
>>109557588
>pay no attention
>pay attention
>>
>>109557588
it wouldn't be a schizo echo chamber if it wasn't there
>>
>>109557534
holy shit, this model. this time in place of the bocchi images I just used swimsuit marciana (NIKKE).

can I submit this for the lewd jam?

https://files.catbox.moe/hrwn23.mp4
>>
File: 1766675567898433.jpg (228 KB, 902x1625)
228 KB JPG
>>109557606
source image used:
>>
>>109557577
good panty pull, usually it's jankmaxxing
>>
>>109557606
yes
>>
File: 1760712745727230.webm (3.9 MB, 704x896)
3.9 MB
3.9 MB WEBM
>>109557574
it's the form that matters
>>
Lol. Ani thinks that by spamming his own thread he'll get this one deleted. Let's see if he's right.
>>
File: 1778399418915787.jpg (92 KB, 960x544)
92 KB JPG
>>109557534
what is that?
>>
File: WBCaique.jpg (172 KB, 850x1500)
172 KB JPG
>>109557593
maybe caiques?
>>
>>109557629
also this one worked good too, got the 2 characters. one is tia.

https://files.catbox.moe/chev0g.mp4
>>
>>109557642
thats me generating at 0.4mp cause I want to test fast
>>
>>109557631
ngl i was really expecting her to start drilling through the earth
>>
>>
File: GPUPrices.ai - 8GB.png (269 KB, 1415x868)
269 KB PNG
>>109556485
>i assume you're just asking to waste my time
Sorry if it came off that way. It wasn't my intention. I'm considering ordering a GPU. Since I haven't tried image generation in a serious manner (a few times online, free ones), I don't wanna overspend. Basically, I'm considering trying it, but I'm unaware if it's gonna work out. Having an idea about bare minimum VRAM requirements might make my search easier. As for LLMs, I was asking in case certain LLMs might be better suited for certain topics. I'm not too knowledgeable about AI stuff.
>i hope someone reading this derives benefit from it
I have. I now know 8GB might be enough for me to try serious image generation for the first time.
>adding a lora makes basically no difference
Did you mean LoRas don't lower VRAM requirements? With my initial LoRa question, I was wondering if 8GB is enough for offline LoRa training. Sorry if I could've been more specific.
>>
What model are (you) using? I want to try some new stuff
>>
kek

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

front view of <Subject 1> running toward the camera holding a silver pistol in the air. the camera distance remains the same as she runs forward.

https://files.catbox.moe/0uh72j.mp4
>>
would the vram problem be solved if models could run across multiple gpus? i've got a few old gpus collecting dust that could actually be put to use if this were a thing
>>
File: 00049-424934055.jpg (490 KB, 1472x2880)
490 KB JPG
>>
>>109557705
yeah, already happened. just not diffusion transformers
>>
okay now I got the proper camera. low mp to test.

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

medium front view of <Subject 1> running toward the camera holding a silver pistol in the air. the camera tracks her remains the same distance as she runs forward.

https://files.catbox.moe/d2hzrv.mp4
>>
>>109557723
catbox didnt work.

https://litter.catbox.moe/8wtd2d4i0en9193f.mp4
>>
so second image as the background, just say the setting is <Picture 2> and it will work. insert whatever bg you like. (this is RE9).

https://files.catbox.moe/74q91b.mp4
>>
>>109557675
I too enjoy a cuppa coffee
>>
Guys, is there a guide for extracting clean dialogue from an audio clip?

I got this nodepack for comfy: https://github.com/diodiogod/TTS-Audio-Suite
But when I clicked the VOCAL/NOISE REMOVAL GUIDE: https://github.com/diodiogod/TTS-Audio-Suite/blob/main/docs/VOCAL_REMOVAL_GUIDE.md
I noticed it isn't actually a guide at all, just a list of types and options.
Since this dumbass doesn't know how to write a guide, does anyone have a workflow to spare? I just wanna clone voices.
>>
File: MiniMax_H3_00067_.mp4 (751 KB, 928x672)
751 KB
751 KB MP4
Minimax Music.cpp is here
https://github.com/ServeurpersoCom/minimaxmusic.cpp

The RVQ encoder has been reverse-engineered
https://huggingface.co/MiniMaxAI/MiniMax-Music3/discussions/5

So with Minimax Music cpp we'll soon be able to remix songs (continue, extend, etc...) and the LoRA trainers are coming.

Of course Comfy doesn't care about music, also .cpp is vramlet friendly and faster on high VRAM GPUs so go there music frens!

https://files.catbox.moe/3kvixg.wav
>>
Can I use Minimax H3 on my RTX 3060 12gb and 32 RAM?
>>
got a nice Rio:

https://files.catbox.moe/r04e4b.mp4

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

medium front view of <Subject 1> wearing a black bodysuit, slowly walking toward the camera holding a silver pistol in the air, while remaining silent. the camera tracks her and remains the same distance as she walks forward.
>>
>>109557821
Yes with the right comfyui settings it's perfectly possible. The question will be if you can generate at a speed, duration, and resolution that is acceptable to you.
>>
>>109557774
Awesome I didn't realize the weights were out, I'll have to give it a try some time.
>>
Is it possible to downclock my RTX 5070ti to 50% of power target ? Genning with Minimax really makes my GPU fan screams
>>
>>109557854
I notice no diff at 70% for my 4080, quieter fans etc. 50 may be too much idk.
>>
File: MiniMaxH3_00091.png (778 KB, 640x1152)
778 KB PNG
okay h3 totally fucked up kikes again but everything else turned out perfectly KAIRI KINO INCOMING

https://d.uguu.se/iisGsvBg.mp4

>>109557647
lol might try this. poor cute fellas have to share a name with my least favorite people on the planet.
>>
>>109557767
you probably need software to strip the vocals, there is Audacity that has an add-on AI feature to strip all the vocals, instruments etc into stems so you can then edit or simply use as is. There is no doubt other software that can do it as well, maybe just for stripping the vocal a basic vocal remover.
>>
>>109557839
Rio again, competing for sexy jam

https://files.catbox.moe/odqafy.mp4
>>
>>109557864
How to do below 70% ? Nvidia App only do down to 84%
>>
>>109557876
msi afterburner slider, mine is at 70% pl. 50 may be way too low, but you can try
>>
>>109557854
nvidia-smi -pl 150

Limits it to 150W
>>
>>109557866
nice
>>
>>109557854
deshroud and replace with bigger fans (bigger = quieter).
>>
>>109557909
>Limits it to 150W
wow that's really low!!!
>>
>>109557353
quick proof of concept
https://files.catbox.moe/516svi.mp4
was gonna go with the classic pendulum but I'm sure it can do that fine too
>>
File: 56287623.mp4 (1.03 MB, 1056x608)
1.03 MB
1.03 MB MP4
>>109557402
>>
>>109557928

Is that not half of the tdp? I just googled it for 3s. I keep my 3090 at 250 - 300W instead of 370W. Not that much slower gen times but much less heat.
>>
>>109557841
Good point. What speed, duration and resolution can I expect with my specs?
>>
File: 1772596712151176.jpg (59 KB, 500x900)
59 KB JPG
Elegg in RE test: source image and the output.

https://files.catbox.moe/j3hbln.mp4

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

full body front view of <Subject 1> wearing a yellow thong, who turns around to show her ass and runs away while completely silent, the camera follows her as she moves, keeping the same distance.
>>
>>109557954
I can't say, my specs are different. It's also gonna depend on what you're genning. If you're running inference with the ref2v model and using a video reference the gen times are way different than if you're just using one reference image.
>>
File: file.jpg (44 KB, 584x364)
44 KB JPG
any h3 lora to stop that horrible crunch sound with blowjobs or sucking in general?
in fact, any lora even focusing on audio for h3?
>>
Is the lightx2v turbo lora generally seen as a better optimization than spectrum? What are the pros and cons of each?
>>
>>109557462
I miss movies like that
>>
>>109557992
prompting will do it, worked fine for me
>Lips sliding across wet skin, gagging and muffled moans.
>>
>>109557584
Yep, and people still fall for this shit.
It's not a melty, it's just retarded sharties. Years of that.
>>
>>109557996
I personally find it much better because less time and the quality is still good at 8 steps. spectrum is skipping some of the 20 steps but in a different way.
>>
>>109557953
I'm surprised it's stable.
>>
>>109557992
yea the turbo lora fixed it. just prompt for blowjob sounds
>>
File: 1773291149273119.jpg (30 KB, 1836x92)
30 KB JPG
>>109557774
OK this is amazing, it's not even reverse engineered, it's just there.
Did they forget it? Or "forget it"? Doesn't matter much but man is this good for the future of local music.
>>
>>109557875
the sigh at the end did it for me
>>
>>109557953
What's recommended is usually 80%, where you still get most of the perf, usually 90-95% depending on the card.
Bonus if you also undervolt the card, you can even better performance that way.

>>109558025
It's stable because nvidia-smi doesn't undervolt, so the card just runs slower.
>>
File: 00076-1798641564.jpg (438 KB, 2880x1920)
438 KB JPG
>>
>>109558006
>>109558027
thanks anons, I'll retry then
>>
There has only been one submission for Sexy Jam so far. Is the interest just not there?
>>
>>109558059
give it some time, at least the weekend, just repost about it in new threads
>>
>>109557965
ref is a kino model. got the Miku shimapan to work. would be easier to just add it as a reference but whatever, just 2 sources, miku and the location.

https://files.catbox.moe/oq291a.mp4

<Picture 1> is the physical reference for <Subject 1> with the same clothing.

the setting is <Picture 2>

full body front view of <Subject 1> who turns around and runs away while completely silent, as she runs the skirt flies up to reveal her underwear which has a pattern of alternating horizontal teal and white stripes. the camera follows her as she moves, keeping the same distance.
>>
>>109558059
I was gonna post clussy but i'm not at all happy with how my prompt turned out. Will probably try to get something in by the end of the day.
still never got a definitive answer what time this ends kek
>>
>>109558071
Well there was supposed to be a 48 hour time limit but I guess I'll wait until after the weekend.
>>
can you reference clothes for people/characters in h3 and tell it to ignore the person wearing it?
>>
>>109558087
Yes. Check the prompt reference guide.
>>
>>109558059
genning at 20 steps. only the highest effort for you, anon
>>
i like to reference a person/character in h3 and tell it to ignore the clothes
>>
File: MiniMax_H3_00058_s.mp4 (2.58 MB, 608x1056)
2.58 MB
2.58 MB MP4
>>
>>109558006
>Lips sliding across wet skin, gagging and muffled moans.
where i put this ? overall_soundscape ?
>>
>>109558106
no u
>>
>>109558070
uh oh, miku slipped!
>mfw minimax was literally trained on lewds

https://files.catbox.moe/a8yfda.mp4
>>
>>109558099
ok thanks, I didn't know if it was possible
>>
>>109558115
not lewd but clearly ecchi anime
which makes me think no western studio would ever do that outside of the cloud ones that on everything then put guardrails
>>
File: Milk Advert.webm (1.44 MB, 640x480)
1.44 MB
1.44 MB WEBM
Trying out some weird 2000s commercial type shit
>>>/gif/31041597
>>
>>109558129
*company, not studio obviously
>>
>>109558134
kek nice, for making ads you can also give google ai mode a basic template then say "make a 10 second prompt for a 90s pasta maker infomercial" and it will work.

the model isn't hard to make prompts for, but AI can generate a giant block of text to test stuff instantly.
>>
>>109558134
someone should try redoing that toy commercial for epstein island that was first made on sora
>>
miku variation:

suddenly, a group of resident evil zombies show up to the right, and Miku takes out a shotgun and blasts them, knocking them to the ground.

https://files.catbox.moe/0025vj.mp4
>>
>You ARE buying Kingdom Hearts 4, right?
https://files.catbox.moe/cy9fxf.mp4
>>
File: Varang forest.webm (3.52 MB, 1136x896)
3.52 MB
3.52 MB WEBM
She has found herself a new friend and she's going to keep you.
I need to learn how to stitch videos together, ~15 seconds isn't enough at all.
Just imagine she licks you and then straddles your face. I'll play with this more later.
>>
>>109558182
>fear boner: activated
>>
>>109558182
I don't get the fetish but impressive gen
>>
>>109557402
Pick! I turned myself into a rickle!
>>
>>109558182
>MAGTt meets his SEA monkey wife
>>
>>109556385
>is there a audio database with voices, or an idiot proof way of doing it?
yeah
https://sounds.spriters-resource.com/
>>
File: pics.png (2.75 MB, 2432x1239)
2.75 MB PNG
>>109558199

Multiple reference images are really beneficial here and they remove the element of luck quite a bit.
Separate background for consistency and then photoshopping the characters in there helps getting the flow you want.
>>
>>109558292
>https://sounds.spriters-resource.com/
Based, thank you. I was just thinking about trying voice cloning.
>>
File: ms00165_.webm (1.44 MB, 832x640)
1.44 MB
1.44 MB WEBM
>>
>>109558315
oh yeah so i'm fucking around with your butt rubbing prompt, genning my own avatar girl in Krea 2 first, do you use a singular full body image for reference or a wide screen ratio full body character sheet? I noticed some people go with multiple references at different angles versus a single reference sheet..for such a complex character detail wise i'm unsure.
>>
>>109558315
>It was all keyframes
Well played. Every time I've tried using keyframes the result isn't smooth at all.
>>
this. so much this.
https://www.reddit.com/r/StableDiffusion/comments/1vofzg1/ill_take_you_to_the_candy_shop/
>>
Favorite prompt you ever used?
>>
>>109558378
1blowjob
>>
do you think the jannies understand the autistic sdg/ldg lore?
>>
>>109558333

I think you're referring to that other avatar guy with the beach bimbo thing.
Here I didn't use any character reference pic at all, because I thought this should be able to get the likeness from those frames well enough and it did.

>>109558341

That's why I have a separate background picture there with no characters and tell it to keep it preserved throughout this.
What also worked was one pic with a character in an environment and the other keyframe was a character with no background at all.
That way it couldn't accidentally take any elements from the other environment and screw things up.
I noticed that if keyframes have even a slightly different background then the result can be a bit jerky.
>>
File: MiniMax_H3_00070_.mp4 (2.47 MB, 1184x896)
2.47 MB
2.47 MB MP4
>>109557848
Yep, model got trained on 32kHz so it has significantly better audio quality than ACEStep (which was trained at 24kHz) requires a bit of wrangling when prompting it compared to ACEStep though. Make sure to load this skill into LLM of choice https://github.com/MiniMax-AI/MiniMax-Music3/tree/main/skills

BANDMAID style babymetal, first try, it's a banger
https://files.catbox.moe/8au6uh.mp3

Lyrics (romanized)
https://pastebin.com/EsPd4qLy
Lyrics were given in standard Japanese to the model, its recognition of standard characters is exceptional.

Also we can now do multiple bands seamlessly, possibly also multiple singers as you could with Udio/Suno, an area ACEStep XL struggled with a lot.
>>
>>109558410
right, hard to separate you two. Now that's a way to differentiate kek
very impressive stuff though, i wouldn't have been able to tell you even keyframed.
>>
hey /ldg/, boy does Miku have the product for you.

https://files.catbox.moe/cxytcp.mp4
>>
>>109558363
that subreddit constantly gets braindead posts like that
you cannot convince me it's not just indians trying to "flex" their ai knowhow
>>
>>109558420
ketchup on pasta is hilarious. computer, generate me a video to maximize italian pain



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.