[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: 1786659052456127.mp4 (1.2 MB, 1472x1472)
1.2 MB
1.2 MB MP4
Janitor, this thread has the GenJam information in the OP, unlike the shitpost thread made by the known ban-evader.
Stop interfering with community activities you boring twat. I know for a fact the mod team hate it when janitors interject unnecessarily.

Submit your video gens for Sexy Jam 1:
https://docs.google.com/forms/d/e/1FAIpQLSf-MTkQa--uydhU0DzyqMZXqeK2Z09qcHxiAGjpfJesj85mHw/viewform

Previous: >>109557402

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: 00010-1024449175.png (2.54 MB, 1088x1920)
2.54 MB PNG
clussy
>>
devilish
>>
>mfw Resource news

08/14/2026

>ReDetail: Generative video re-detailer through LTX-2.5's pixel spatial upscaler,
https://github.com/Bambushu/redetail

>V-RAE: Rethinking Video Latent Spaces for Generation
https://v-rae.github.io

>4 step ref2v Minimax H3 turbo Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>Learning Unified Video and Image Representation for Video Face Forgery Detection
https://github.com/haotianll/UVIF

>AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design
https://github.com/Yaxin9Luo/AutoDesign

>H3 Motion Context — Timeline
https://github.com/BSG-Walter/ComfyUI-H3-Motion-Context-Timeline

>RTX PRO 6000 Blackwell workstation edition price doubles, as NVIDIA launches new AI model
https://www.neowin.net/news/rtx-pro-6000-blackwell-workstation-edition-price-doubles-as-nvidia-launches-new-ai-model

>MAGI-2 Preview: Scaling Video Generation Models Efficiently
https://sand.ai/blog/magi-2-preview

08/13/2026

>Lightx2v MiniMax H3 Turbo Ref2V 4Step/8Step Loras
https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main

>GemmaPrompt: Local prompt enhancer for ComfyUI diffusion models
https://github.com/whp199/GemmaPrompt

>ComfyUI-H3-FaceRefine
https://github.com/Carasibana/ComfyUI-H3-FaceRefine

>Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising
https://github.com/Ai-ZL/Hybrid-LUT

>MiniMax H3 Creator for ComfyUI: Multi-Shot 60s Timelines, Resizable Satellite Stage, & Ollama/LM Studio Refiner
https://github.com/roadmaus/ComfyUI-MiniMax-Creator

08/12/2026

>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler
https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler

>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready
https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy

>LTX-2.5 22B — GGUF
https://huggingface.co/realrebelai/LTX-2.5_GGUFs

>ComfyUI NVIDIA RTX VSR Pro
https://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro
>>
File: HunyuanVideo_00043.mp4 (497 KB, 640x480)
497 KB
497 KB MP4
>>
>mfw Research news

08/14/2026

>SNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolation
https://arxiv.org/abs/2608.13460

>Spatially-Grounded Text-to-Video Generation via Inference-Time Gradient-Free Optimization
https://arxiv.org/abs/2608.13037

>SketchSense: Learning to Interpret Imperfect Sketch Guidance for Image Inpainting
https://arxiv.org/abs/2608.13186

>Semantic Steering for Controllable Generation: Tuning-Free Concept Erasure in Multimodal Diffusion Transformers
https://arxiv.org/abs/2608.12829

>MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval
https://arxiv.org/abs/2608.12532

>From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion
https://arxiv.org/abs/2608.13043

>Erase but Preserve: Controllable Removal of Copyrighted Animation Characters via Optimized Semantic Anchors
https://arxiv.org/abs/2608.12806

>SCOPE: Subspace Clustering with Online Per-Head Top-K Estimation for Sparse Video Attention
https://arxiv.org/abs/2608.12780

>StrAD: A Streaming Method and Benchmark for Audio Description Generation for Long-form Videos
https://arxiv.org/abs/2608.12549

>HPSD: Hybrid-Policy Self-Distillation for Text-Image-to-Video Diffusion Models
https://bujiazi.github.io/hpsd.github.io

>Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation
https://hmrishavbandy.github.io/cmd-site

>A Controlled Study of Self-Supervised Image and Video Pretraining under Limited Resources
https://arxiv.org/abs/2608.13183

>PixSDS: Why Latent SDS Makes Noisy Pixels
https://sevashasla.github.io/pixsds-webpage

>RIG-RoPE: Relation-Stratified Multimodal Attention with Instance-Local Rotary Geometry and Representation-Aware Traversal Coordinates
https://arxiv.org/abs/2608.05154
>>
are qwen ever gonna try to make up lost ground? I really liked qwen image
>>
File: 294j5.png (11 KB, 644x800)
11 KB PNG
>Janitor, this thread has the GenJam information in the OP, unlike the shitpost thread made by the known ban-evader.
>Stop interfering with community activities you boring twat. I know for a fact the mod team hate it when janitors interject unnecessarily.
>>
>>109559522
Kekd
>>
Is it ani or debo who is aggressively trying to get the rentries removed from the OP?
>>
to the one guy asking about removing music for H3. try this
>non_diegetic_music: N/A
seems to work pretty consistently for me
>>
>>109559563
both of you fucktards should just READ THE FUCKING GUIDE HOLY SHIT

https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md#42-shots-and-cuts
https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md#4-retention_analysis

like is it really that fucking hard? jesus christ
>>
Omegabased
>>
>>109559570
hmmm, nyo~
>>
Blessed thread of frenship
>>
>>109559570
yes. what i typed was in fact in the guide.
>>
File: debo_ds_k2_00001_.png (2.55 MB, 2048x1101)
2.55 MB PNG
>>109559558
you've seen me say over and over again that I don't bake and I don't care about the rentries
>>
>>109559522
just letting you know /trash/ seems to love this OP gen
>>
>>109559532
>reposting slop instead of genning a new better image
Bruh
>>
what about character yapping. how to consistently stop char from randomly talking?
>>
>>109559570
don't even have to read it, just dump it on an llm and let a clanker do the work
>>
is this the thread?

literally Miku, again

https://files.catbox.moe/biuhkd.mp4
>>
>>109559596
describe them as not talking in each shot
>>
>>109559596
"is silent", or use timestamps

at 00:03.000 character 1 says "this is me saying stuff"

at 00:05.000 something else happens

if there is no time they cant do anything else. that or do dialogue like says "blah" then remains silent.
>>
Am I the only one using GemmaPrompt?

https://github.com/whp199/GemmaPrompt
>>
File: MiniMax_H3_00026_.webm (643 KB, 864x480)
643 KB
643 KB WEBM
hmm, didn't seem to understand 'futuristic cyberpunk megacity'
>>
>>109559522
Hitler > debo
>>
Let's play a game. I'll post images from 2 models. Try to guess which is which.
Model A: samANIMA 2.2.5, with custom realism lora (2000 images) trained on anima base.
Model B: Krea2, with lora trained on the same dataset.

Why the lora? It helps samANIMA's realism a bit, and also slightly de-slops krea. Each COLUMN is a different model and there's 2 seeds going down. For the prompts, I'm using Gemini captions from a high-quality LAION subset I happen to have on one of my hard drives. This wasn't what the lora was trained on, it's just a convenient source of diverse, high quality prompts.

Next post will start the images. HARD MODE: no zooming in and hyperfocusing on the small details.
>>
File: image_00006_00000.jpg (1.82 MB, 2240x2240)
1.82 MB JPG
>>109559647
A daytime, medium shot depicts a charming, rustic log cabin nestled amidst a lush forest. The cabin, constructed from light-brown logs, features a green metal roof and a stone chimney. A porch extends from the front of the cabin, adorned with a wooden railing. The entrance to the cabin is approached by a stone-paved walkway. The forest setting is characterized by tall, slender trees with textured bark. The ground is covered in a mix of dirt and fallen leaves. The lighting is bright, casting shadows that add depth to the scene. The overall mood is peaceful and inviting, evoking a sense of warmth and comfort.
>>
>>109559658
Easy, left is krea. Ewewew.
>>
>>109559658
shit copied wrong prompt, corrected:

A medium shot shows a fair-skinned woman standing on a dusty path, wearing a long, colorful dress with her hands clasped behind her back. Her blonde hair is short and styled away from her face, and she is smiling brightly at the camera. The dress features vertical stripes of red, orange, green, and blue, with beaded detailing around the neckline and waist. Behind her, two figures in traditional garments walk away down the path, their faces obscured by their clothing. One wears a pink robe, while the other is draped in tan. A small dog can be seen near the bottom left of the frame. The setting appears to be an arid, desert landscape, with a stone building visible in the background. The path is light brown and dusty, blending with the color of the building and the surrounding terrain. The lighting is bright and warm, casting soft shadows and enhancing the vibrant colors of the dress and the landscape. The overall mood is serene and joyful, suggesting a peaceful and exotic location.
>>
File: image_00005_00000.jpg (2.86 MB, 2240x2240)
2.86 MB JPG
>>109559658
the image for the log cabin prompt
>>
>>109559667
you are SO wrong faggot. SamANIMA has a particular way it slops faces, both on the left are Anima.
>>
File: image_00007_00000.jpg (1.59 MB, 2240x2240)
1.59 MB JPG
>>109559674
A close-up shot captures a man wake surfing on a sunlit body of water, with a clear blue sky overhead and a distant landscape visible. The man, fair-skinned with dark, wet hair, is turned to his left, looking ahead with a focused expression. He is wearing a black sleeveless wetsuit top and colorful striped board shorts. His bare feet are positioned on the wake surf board, which is black with a white circular design in the center. The board is surrounded by turbulent white water and green-tinted water. Splashes of water are visible around the board as the surfer rides the wave. The background reveals a hilly landscape with houses and greenery under a clear blue sky. The lighting is bright and sunny, highlighting the action and creating dynamic reflections on the water's surface.
>>
>>109559679
kys malding troon
>>
File: image_00009_00000.jpg (1.32 MB, 2240x2240)
1.32 MB JPG
>>109559685
A vintage Volkswagen bus, painted in a two-tone scheme of white and teal, is prominently displayed in the foreground. The front of the bus features a large, circular Volkswagen logo in a light blue color, and round headlights shine brightly. Its doors are closed, except for the driver's side which is slightly ajar. The chrome bumper adds a touch of classic elegance. The bus's tires are black with white rims.

In the background, there is a group of people, their features obscured by the depth of field. The architecture behind the bus suggests an urban setting, with arched stone doorways and walls. The lighting is bright, indicating a sunny day, and there are shadows cast beneath the bus, adding depth to the scene. The color palette is vibrant, with the teal and white of the bus standing out against the neutral tones of the surroundings.

The overall mood is one of nostalgia and appreciation for classic automobiles. The scene suggests a relaxed, perhaps tourist-oriented environment.

>>109559679
are you sure? :)
>>
>>109559645
hitler for thread janny
>>
>tfw looked at the first comparison for half a second
>on my phone
>didn't zoom in
>got it correct
I am the greatest genner alive. You know who I am, my gens are constantly hailed by anon as "really good". Bow before me.
>>
File: image_00008_00000.jpg (1024 KB, 2240x2240)
1024 KB JPG
>>109559691
A close-up, eye-level studio shot presents a square dessert resembling a no-bake cake on a white plate. The dessert is composed of layers of graham crackers acting as the base and middle tier, sandwiched between thick layers of white cream, and topped with a rich, dark chocolate frosting. The bottom graham cracker layer has a slightly uneven edge, suggesting a homemade quality. The two layers of white cream are smooth and appear fluffy, contrasting with the coarse texture of the graham crackers. The top layer of chocolate frosting is dark brown, smooth, and shiny, with subtle ridges indicating it was spread with a spatula or similar tool. The lighting is soft and even, highlighting the textures of the different layers without casting harsh shadows. The dessert is centered on the plate, which has a slightly raised rim. The background is a light beige or off-white, creating a neutral backdrop that emphasizes the dessert’s colors and textures. The overall mood is appetizing and inviting, showcasing the dessert’s layered construction and rich ingredients.
>>
File: image_00004_00000.jpg (1.74 MB, 2240x2240)
1.74 MB JPG
>>109559698
A medium shot captures a person on a motorcycle against a vast mountain backdrop under a cloudy sky. The person, wearing a full-face helmet with a grey visor and a grey and black riding jacket, is perched on a black and orange motorcycle. They are also wearing black riding pants and gloves. The person's gaze is directed towards the left, away from the camera. The motorcycle, a prominent feature in the foreground, is parked on a rocky terrain, its tires resting firmly on the ground. The bike's design incorporates both black and orange, with the orange accents adding a vibrant contrast. The rear tire is visible, as well as the exhaust system. The ground, composed of large, light-colored rocks and sparse greenery, suggests a rugged, natural environment. The background is dominated by a sprawling landscape of rolling hills densely covered with trees. The color palette of the image is predominantly muted, with shades of grey, green, and brown creating a sense of earthiness. The sky, filled with voluminous grey clouds, contributes to the slightly overcast mood. The lighting is soft and diffused, casting subtle shadows that add depth to the scene. Overall, the image evokes a sense of adventure and exploration, with the person and motorcycle positioned as subjects against the expansive natural setting.
>>
File: image_00002_00000.jpg (1.51 MB, 2240x2240)
1.51 MB JPG
>>109559703
last one. maybe getting a bit easy now

A woman in a wedding dress stands in an outdoor setting, posing elegantly amidst lush greenery and decorative elements. She wears a form-fitting, ivory-colored gown with intricate lace and beaded embellishments. The dress has a sweetheart neckline and extends into a mermaid-style skirt that pools slightly on the ground. She has a sheer cape-like garment draped over her shoulders and around her neck. Her dark hair is styled with a loose braid adorned with floral accents. She wears a bracelet on her left wrist and earrings, adding to her sophisticated look. Her expression is serene and poised.

The backdrop is a blurred garden scene. To the left, there is a vintage-style wrought iron gate painted in a muted blue. Pink flowers are visible in the foreground near the gate. Behind the woman, there is a dark, still body of water, possibly a pond or small lake, reflecting the surrounding greenery. To the right, an antique gold-colored picture frame leans against foliage, adding to the vintage aesthetic.

The scene is illuminated by soft, natural light, which casts subtle shadows and highlights the textures of the dress and surrounding elements. The overall atmosphere is romantic, elegant, and ethereal, conveying a sense of timeless beauty and sophistication. The composition is well-balanced, drawing attention to the woman as the central figure while incorporating the surrounding environment to create a cohesive and visually appealing image.
>>
File: videogen__00003_cut.gif (3.62 MB, 576x800)
3.62 MB GIF
>>
Minimax H3 keeps adding shitty background music to my gens even if I specify in the prompt to not add it. Anyone else had this problem and solved it?
>>
>>109559658
do this with the images turned black and white - like a new gen, turn them black and white in gimp, no fancy filters just basic to grayscale.

I think it would be interesting.
>>
>>109559716
non_diegetic_music: N/A
>>
its like im literally watching the movie.

https://files.catbox.moe/bt6sz9.mp4
>>
>>109559736
well besides the fact miku doesn't completely replace the woman so now there's two women
>>
>>109559726
Too much effort, and I think the point I'm trying to make is clear. Anons in this very general will look you dead in the eyes and tell you Anima is an untrainable dead-end model that is in fact worse than SDXL.

anime model btw
>>
>>109559743
he's hallucinating, its fine
>>
>>109559522
Fucking hell i clicked it
FUCK YOU
>>
>>109559751
Now you're gay and an antisemite
>>
>>109559744
>Anons in this very general will look you dead in the eyes and tell you Anima is an untrainable dead-end model that is in fact worse than SDXL.
yeah. rhetoric pushed by parties with a vested interest in shilling their own anime model.
>>
>>109559744
bla bla
agreed. I'm not looking lol. fuck your lora trash.
>>
File: 1773163553433199.jpg (25 KB, 600x394)
25 KB JPG
>>109559744
Im still mad my favorite character lora is only available on Anima so i permanently cucked to use those.

How to extract Lora and trained it to other models like KREA 2 ?
>>
>>109559759
the answer is, indians, as always.
>>
>>109559771
Just make your own dataset and train with that. It's not hard, just a bit annoying.
>>
File: 1761835546086491.jpg (63 KB, 738x703)
63 KB JPG
>>109559776
I dont have the dataset
>>
>>109559780
out of disk space?
>>
>frog poster is incompetent
Unsurprising
>>
File: 1774373669528272.jpg (37 KB, 500x500)
37 KB JPG
>>109559788
No im just too lazy. How to extract a lora with tags already set in automatically ?
>>
>>109559780
What part about "make your own" did you not understand? You can't compile and crop 30 images of your waifu?
>>
>>109559796
i sent you a message, check your 4chan inbox
>>
>>109559799
she'd honestly probably notice me looking in the window.
>>
File: MiniMax_H3_00064_s.mp4 (3.75 MB, 608x1056)
3.75 MB
3.75 MB MP4
>>109559771
Just use ref2v :)
>>
>>109559522
Stop being a faggot and make a collage.
>>
any news about Kijaigod's sol attention patch?
>>
>>109559815
...and don't forget to include my video!
>>
File: 1778546765307179.png (359 KB, 646x595)
359 KB PNG
>>109559799
>30
Im talking about THOUSANDS
>>
File: 1759700684413705.jpg (125 KB, 672x857)
125 KB JPG
Might as well post it in here too
>>>/gif/31042845
>>
>>109559815
How about you commit suicide instead anon? Serious suggestion.
>>
File: ed5.png (131 KB, 680x1112)
131 KB PNG
>How about you commit suicide instead anon? Serious suggestion.
>>
>>109559822
You don't need that many images for a character lora. I bet you zoomies don't even know who Furkan is baka
>>
literally miku (again)

https://files.catbox.moe/dfaimc.mp4
>>
>>109559844
in all my life Total-Resort-3120 i don't think i've seen you prompt anything but miku
>>
>>109559851
think again, I just did a wakaliwood film (with vj emmie voice cloned).

https://files.catbox.moe/pp8k5x.mp4
>>
>>109559860
10s clips, and just asked google ai to regenerate the prompt with a diff captain alex scenario. 5-6 clips stitched and gg.
>>
File: 914130647101285.mp4 (3.65 MB, 640x832)
3.65 MB
3.65 MB MP4
>>
>>109559835
why isn't this image animated?
>>
neat glowstick trails on this one

https://files.catbox.moe/2diooa.mp4
>>
I love LDG
>>
brb
>>
has anyone tried keyframing with h3 using the Add Guide for MiniMax H3 node?
>>
File: file.png (445 KB, 927x522)
445 KB PNG
>>109559902
me too anon, me too.
>>
just going for a drive

https://files.catbox.moe/tafls2.mp4
>>
Minimax h3 music kinda sucks. Maybe the actual sound quality of the output is better than ace step, but it's so bland. I'm using the skill to write the prompt just haven't really got anything interesting enough to post.
>>
File: 1778067625430220.jpg (8 KB, 238x250)
8 KB JPG
If only i can use my old GTX 1080 as spare VRAM card.........
>>
what facial expression should i be making when im taking pics of my peepee for h3 gens? im new to all this
>>
File: MiniMax_H3_00027_.webm (1.71 MB, 864x480)
1.71 MB
1.71 MB WEBM
>>109559627
>>
>>109559948
like this
>>
Anyone tried NAG for H3?
https://huggingface.co/CCP6/H3-Shadow-Negative-Nodes
>>
>>109559956
I was waiting for your report before trying, actually.
>>
File: collage-spaced-bbe50f-2.mp4 (3.81 MB, 1382x1408)
3.81 MB
3.81 MB MP4
another day, another collage
>>
>>109559947
it is ironic that the 20 series cards which were always looked down upon as the ones to miss out ended up at least being somewhat useful if people wanted to use them if they have one as that is the cut-off lower point for use in AI programs
>>
>>109559985
is the middle image krea?
>>
>>109559956
Its gonna kill the gen times.
Minimax understand "No" in their prompt
>>
I wonder how many people will get this reference

https://files.catbox.moe/169xz9.mp4
>>
>>109559841
I've seen the biggest roach last week in the kitchen, god I hate these buggers.
>>
>>109559989
It's a selfie I took. Get baited fag
>>
>>109559831
that smug smile is perfect, very erotic
>>
Do I need to do something special to make chunk feedforward impact the generation time on longer videos?
>>
>>109559883
I love Asuka (male) and Rei (male)!
>>
GAAAAAAAAAAAAAAY
>>
>>109560032
if you are on blackwell, it seems to do nothing or even do nothing, see tests by anon a few threads back
>>
File: 509898645493463.mp4 (3.66 MB, 640x832)
3.66 MB
3.66 MB MP4
>>109560033
Understandable.
>>
>>109560047
How are you describing this specific visual style in H3? Low quality early 2000s camera?
>>
>>109560046
Well shit. Thank you.
>>
>>109559679
samanima definitely likes to fuck up the teeth from what i remember. i guessed right
>>
>>109560046
*or even be worse than nothing
>>
File: 1786617054333681.jpg (1.01 MB, 4078x3682)
1.01 MB JPG
>>109560046
>>109560058
saved it
>>
File: 630970873962889.mp4 (3.66 MB, 640x832)
3.66 MB
3.66 MB MP4
>>109560056
I'm not describing it at all, just using a start image generated with Krea 2 that looks like it's taken with an old cell phone camera.
>>
promotions?

https://files.catbox.moe/jsbe42.mp4
>>
>>109560084
the H3 image model could unironically match NBP if it's as good as the video model
>>
>>109560084
ah, i see
>>
>>109560056
h3 can't do that style, it is too slopped
>>
I remember a trick with ltx where you could replace the audio with your own and have the video being genned with your injected audio in mind. Is such a thing possible with h3?
>>
>>109559808
wf? i wanna gen witcher 3 blender porn renders
>>
>>109559883
This. This is the only use case for AI.
>>
>>109560120
yes. wan2gp has that feature
>>
>>109560114
nah, it probably could
>>
File: 1774966990763744.jpg (59 KB, 512x386)
59 KB JPG
>5 second gen at 0.7mp
>70 seconds
>10 second gen at 0.7mp
>250 seconds

This is ComfyUI issues, right ? It should be 140 to 150 seconds
>>
>>109560151
gen time does not scale linearly
>>
>>109560130
So that possible, ok thanks.
>>
File: MiniMax_H3_00029_.webm (1.16 MB, 864x480)
1.16 MB
1.16 MB WEBM
>>109559950
I was too ambitious with this one
>>
>>109560122
Just using the stock ref2v template.
>>109560151
Transformers grows exponentially with resolution and length innit
>>
>>109560162
>>109560168
Why cant they just optimize it like, chained 2s+2s+2s+2s+2s= 10s gens then?
>>
>>109560144
no one has been able to do it without input images
>>
>>109560172
say they did do that, every single chain needs to know what happened before and whats going to happen next, so you're back where you started in terms of time complexity. Unless of course you want inconsistency and a new gen every 2 seconds, in which case it would speed things up
>>
>>109560032
It doesn't increase speeds or reduce vram usage. Anon is right about that, but it does allow people to say "I don't care how long this takes I just want 1mp @ 15s" without it failing.
>>
>>109560167
ENTER
>>
File: 1784001824705361.gif (3.34 MB, 320x180)
3.34 MB GIF
Can Krea 2 consistently change camera angle ? Its all i want so i can create a mini hentai scene for my slop
>>
File: debo_dm_k2_00143_.png (2.51 MB, 1872x1007)
2.51 MB PNG
>>109560211
use an edit model like qwen image edit with the next scene lora. much better option for NVS
>>
File: 1781534706111959.webm (2.48 MB, 880x1320)
2.48 MB
2.48 MB WEBM
>>
>>109559946
Try using the demo to rewrite your prompt
https://huggingface.co/spaces/MiniMaxAI/MiniMax-Music3
Then check the advanced tab and copy the metadata, eventually you'll have a few examples from the demo and you can give it to the LLM you're using the skill on as a more complete prompting guide. Also if you're doing something wrong, you will know right away because the demo oneshots the prompt usually (if you inference it).
>>
>>109560084
he finally bought them, huh?
>>
>>109559883
Can I make shit like this with a 9070xt or nvidia only?
>>
>>109560283
hahahahahahaha
>>
>>109560283
It'll take longer, but you can.
>>
gemmapromptfag
it keeps writing timestamps in minutes instead of seconds
>>
GemmaPromptcuck, why does your UI keep including a 00:00.000 timestamp in multi-shot sequences? Nobody fucking wants this you sperg.
>>
File: 687345348979.mp4 (785 KB, 864x480)
785 KB
785 KB MP4
>>
File: MiniMax_H3_00031_.webm (1.49 MB, 864x480)
1.49 MB
1.49 MB WEBM
>>109560167
peak slop
I could fish for a better seed but these take too long
>>
>>109560178
What about with refs?
>>
>>109560290
Can I ask for a little spoonfeeding? I don't know where to start with the links in OP.
What should I be using first or start learning?
>>
>>109560317
i mean any media inputs in general, whether it is references or starting frames. the model doesn't actually has a concept of that analog feel if you try to prompt for it directly
>>
https://voca.ro/1fU0vGNc9NwS
>>
>>109560329
Sucks that it can't do it by itself but 99% of the time you're going to be giving it images anyway so does it really matter?
>>
>>109560297
I fucking love Kanna gifwtwm
>>
>>109560320

>Download comfyUI
>Download the H3 model, VAEs and text encoder and put them into the right folders in comfy
>Download a video you like
>Drop it into ComfyUI to get the workflow and nodes etc.. they show up automatically if the video has them embedded.
>Or just download a workflow and drop it in
>Update everything in ComfuUI so the random nodes work.

Congratulations, you're set.
Now click generate and see if it works, if there's a problem ask AI about how to unfuck it.
>>
>>109560338
it kind of does matter if you want the effect to look realistic for the whole video. analog effects aren't static. the model doesn't know how to continuously distort the quality for the whole clip
>>
>>109560346
wouldn't they really just drag a template workflow in and then fix stuff in there like all the models, text encoders, VAE's etc as the errors would show up?
>>
>>109560320
- Download ComfyUI PORTABLE (AVOID INSTALLER)
- Models and Workflows (Minimax H3) https://huggingface.co/Comfy-Org/MiniMax-H3
- (Optional) Models and Workflows (10Eros Sulphur (Based on LTX 2.3)) https://huggingface.co/TenStrip/LTX2.3-10Eros/tree/main Workflows : https://huggingface.co/TenStrip/LTX2.3-10Eros_Workflows/tree/main
- Start experimenting with prompting.

Just for Info. Minimax is good for scene to scene slop. LTX 2.3 is only good for static video shot with no camera movement.
>>
new /vp/ request for you guys

>>>/vp/59508265
>>
>>109560346
>>109560367
Thank you anons, I will start tinkering with ComfyUI. Will share if I can generate something decent.
>>
>>109560371
And another:

>>>/vp/59508271
>Hypnotized Hilda stripping down to her underwear
>>
4steps lora but run it at 6 step and modelshift 6/3
It's better than 8steps lora
>>
Anything you can do to cheat low res H3 looks less AI sloppa ?
Upscaling is only good when base video is decent enough.
>>
what is the model that generates halloween deathsmiles metal when you try to generate death metal?
>>
File: frame_0010.jpg (40 KB, 864x480)
40 KB JPG
>>109560305
full sequence if anyone cares (I'll get this link right I swear)
>>>/wsg/6214785
>>
>>109560297
Best genner in this thread.
>>
FEDERAL AID
https://voca.ro/1cv0nQQAfOYd
>>
File: kekekekkeeeeeeeeek.png (1.17 MB, 864x1184)
1.17 MB PNG
>>109560385
>>109560398
>>109560414
nah JSID already bruh tranime kekypows aino wth they doin nga keeeeeeeeeeeeeeeeeek
>>
COON TOWN
https://voca.ro/159z7KvWqheJ
>>
pretty view!

https://files.catbox.moe/6heehz.mp4
>>
File: 1767672247523992.jpg (489 KB, 1129x1129)
489 KB JPG
>>109560431
>>
>>109560451
need to fix last part, safer got cut off
>>
any way to meaningfully speed up vae decode?
>>
Some requests to you guys from /trash/:

>>>/trash/84720571
>well, if you're asking: have Roxanne here talking dirty to Max as she rolls onto her side to show off her body

>>>/trash/84720600
>See if you can get Lilith here >>>/trash/84720372
> to shake her hips and swing her lil pp.

>>>/trash/84720681
>Can you make her roll the coin in her hand over her fingers?
>>
i discovered the ultimate kinoplexatorium settings
>>
>>109560485
video_vae_int8_convrot.
1 minute less decoding time in my own experience.
>>
>>109560528
thanks ill try it
>>
Seems like /v/ is starting to come around to sloppa.
>>
File: sko1.mp4 (1.43 MB, 736x736)
1.43 MB
1.43 MB MP4
>>
so apparently, with 8 step ref turbo, you use euler + beta? better outputs?
>>
You guys will handle the /trash/ requests right?
>>
>>109560528
is there one for audio vae?
>>
>>109560557
Why should I? Convince me.
>>
>>109560557
I don't go there often
>>
>>109560371
>>109560395
>>109560500
>>>/r/
>>
>>109560587
none
>>
>>109560536
The first stellar blade:BR had looked very sloppa and they defended it.
Even before CEO posted his actual slop
>>
>>109560511
I am seated.
>>
drama on the bridge

https://files.catbox.moe/rfjwiv.mp4
>>
File: MiniMax_H3_00057_.webm (889 KB, 480x864)
889 KB
889 KB WEBM
>>109560297
I've done not much besides gen videos of giant lolis for the past few days
>>
>>109560588
Because you'd make them happy.
>>
>>109560588
>>109560608
This. You'll drain tons of semen from anons.
>>
>>109560611
But I'm not a faggot.
>>
>>109560611
the reward is giving a digital handjob to a bunch of strangers?
>>
beta/euler does seem nice with the turbo lora, ill keep testing.

https://files.catbox.moe/ji5jzc.mp4
>>
>>109560636
https://darkstarrddev.us.ci/
>>
>>109560640
not clicking your screamer link troon
>>
>Switching to Euler entirely eliminated the color burning. Because Euler calculates a direct mathematical trajectory rather than factoring in a history of previous noisy steps, it allowed the Turbo LoRAs to scale safely up to 8 or 10 steps without blowing out the colors.

hmm
>>
>>109560597
please hold, a mouse has crawled into the tape deck
>>
The Turbo 4-step has slow motion compared to 8-step LoRA or something?
>>
File: edit_00007_.png (784 KB, 1024x1024)
784 KB PNG
https://vocaroo.com/16hJs1EkSR0a
>>
File: sko2.mp4 (2.23 MB, 736x736)
2.23 MB
2.23 MB MP4
fuck me, we've come a long way from wan 2.2 garbo
>>
File: -face-1-2413067273.jpg (57 KB, 1600x1100)
57 KB JPG
Happy Halloween guys.

https://files.catbox.moe/cu88xe.webm
>>
>>109559538
>>109559544
thanks!
>>
>>109559558
it's any sane anon who cares about the future of ldg
>>
>>109560736
lol what the fuck
>>
>>109559558
I don't have enough fingers and toes for counting the number of anons that want the drama tranny shit to stop in all forms
>>
>>109560723
Someone make a Gemma-chan lora for Krea 2...
>>
>>109560788
>I don't have enough fingers and toes for counting the number of anons that want the drama tranny shit to stop in all forms
wood chipper accident?
there wouldn't be a problem if those two retards acted like normal human beings.
>>
File: img_4733.jpg (394 KB, 1252x798)
394 KB JPG
Music idea: Requiem for a Coping Luddite.
>>
File: 1786626359070814.png (949 KB, 1024x1024)
949 KB PNG
>>109560792
I just use klein to edit this one
>>
>>109560829
when are we getting a qwenchan?
>>
>>109560829
9b? Every time I've tried it the results were shit.
>>
>>109560856
Never because Qwen is boring and benchmaxxed.
>>
>>109560872
yeah well ur a faggot
>>
>>109560856
capybara milf
>>
>>109560874
Go back to /r/locallama if you want to fuck the capybara. They love Qwen there.
>>
>>109560856
only a small souled reddit bugman would enjoy qwen
>>
boutta fuck me capybara bros
>>
>>109560813
>I will continue to drama tranny so long as I accuse the people trolling me are unrelated anons
cool story bro
>>
Can H3 do looping animations?
>>
>>109560788
and yet no one uses your rentry-free threads when there's a rentryful one up
curious!
>>
>>109560889
>>
>>109560901
last frame is usually off so you need to stitch the loop yourself
>>
>>109560904
I'm not interested if you haven't picked it up yet drama tranny
>>
>>109560894
who is trolling me?
and how come the nice clean drama free bakes without those dastardly reentry links are filled with the same retarded bullshit that prompted the creation of those two links in the first place?
>>
Is there qwen3vl_32b_minimax_h3 clip replacement that speeds things up?
>>
>>109560856
>when are we getting a qwenchan?
I'd actually want to know if the new qwen 3.8 27B can be used for h3 prompting, because you can feed it pretty huge images.
>>
[shot 1] <subject> pushes the breasts together from the sides
[shot 2] <subject> cups the breasts from underneath and lifts them upward
[shot 3] <subject> squeezes the breasts firmly with both hands
[shot 4] <subject> gently massages the breasts in slow circular motions
[shot 5] <subject> pinches, rolls, or tugs on the nipples
[shot 6] <subject> presses the breasts flat against her chest with open palms
[shot 7] <subject> lets the breasts bounce freely by jumping or bouncing in place
[shot 8] <subject> shakes the breasts side-to-side
[shot 9] <subject> claps the breasts lightly together
[shot 10] <subject> spreads the breasts apart with her hands
[shot 11] <subject> twists or rotates the breasts gently in opposite directions
[shot 12] <subject> rubs the undersides and outer curves of the breasts
[shot 13] <subject> holds the breasts and bounces them up and down with her hands
[shot 14] <subject> presses the breasts together from the top and bottom
[shot 15] <subject> runs her fingertips lightly over the skin and areolas
[shot 16] <subject> pushes the breasts up toward her face or chin
[shot 17] <subject> lets the breasts hang and sway while leaning forward or moving her torso
[shot 18] <subject> squeezes one breast while cupping and lifting the other
[shot 19] <subject> uses the backs of her hands or forearms to push the breasts together
>>
>>109560926
You can use the int8 convrot which should be fine or even nvfp4 if you want retardation.
>>
>>109560929
you forgot something
>>
>>109560929
I doubt this would work
>>
>>109560933
Does quantizing the text encoder really affect gens a lot? I've been using int4...
>>
<Subject 1> breasted boobily to the stairs and titted downwards.
>>
>>109560941
Yeah it does from my tests, at least anything below int8.
>>
>>109560933
>https://huggingface.co/unsloth/MiniMax-H3-FP8/tree/main/text_encoders

this one? why is it 27 GB
>>
>>109560948
Not sure I can even fit int8 with my 24gb vram/32gb ram
>>
It's a huge relief people are on h3, because it's way less ugly than krea, I hate how krea looks, though I like that it does swastikas.
>>
>>109560951
Yeah, though I prefer the int8 here : https://huggingface.co/Comfy-Org/MiniMax-H3/blob/main/text_encoders/qwen3vl_32b_minimax_h3_int8_convrot.safetensors

>>109560953
Try it anon, if you get oom it is what it is, you'll know you'll have to use fp4 until you get 32GB more ram.
>>
>>109560953
sudo swapon -a
>>
Guys remember to submit your minimax gen to Sexy Jam.
>>
>>109560971
there's no point since ani is going to win anyways
>>
which of the multiple sol-attentions is the one that got updated to be allegedly better than sage?
>>
no thanks, jeet
>>
>>109560980
chudai wataa rajeesh?
>>
>>109560978
There is just one sol attention anon.
>>
>>109560986
nope there's several implementations ackshully
>>
>>109560978
use comfy kitchen attention . Onions attention degrade quality
>>
>>109561007
Stop recommending this when you don't know what the fuck you're talking about retard. You're a dumbshit without even the faintest understanding of how sage works.
>>
>>109561017
i'm about to sage you faggot
>>
the kinos were supposed to be showing an hour ago, but of course i have to keep changing the prompt
>>
>>109560528
I tried this vae and it crashed every time
>>
File: H3_noAudio__00029_.mp4 (2.27 MB, 1248x896)
2.27 MB
2.27 MB MP4
>>
File: 1777785879698842.mp4 (214 KB, 576x736)
214 KB
214 KB MP4
>>109560963
>>109560970
Just tested. It works fine with int8. Gonna have to go back to some old gens later and see if it affects the quality.
>>
>>109561120
that's not how eating works.
>>
File: darkness_cover.jpg (343 KB, 1920x1080)
343 KB JPG
>>109559522
this seems like a totally normal thread.

darkness
https://suno.com/s/TmEEhZ8IVn8EKEADhttps://youtu.be/iEEX-xfqWR4
>>
>>109561126
proofs?
>>
>>109561131
more mime kino
>>
>>109561118
Late 80s early 90s anime feel unlocked
>>
File: H3_noAudio__00030_.mp4 (1.9 MB, 1248x896)
1.9 MB
1.9 MB MP4
>>
File: file.png (116 KB, 1857x785)
116 KB PNG
>>109560929
low res video with the first seven, boob warning https://files.catbox.moe/rw922i.mp4
I am using the boob jiggle lora
>>
>>109561131
i love mime girls but please stop namefagging
>>
>>109561188
This straight out of H3 or did you use ref ?
>>
>>109560963
>>109560933
I tried qwen int8 convrot . It's 10 seconds slower than nvfp4
on my 3090TI 64gb ram

139s vs 149s - 0.4M 9seconds video
>>
File: 1784869297841890.mp4 (131 KB, 864x480)
131 KB
131 KB MP4
>>
>>109561208
full 38 second boob run, do not click at work
https://files.catbox.moe/zxyp9q.mp4

she can't do it, should have known this thing is for 15 second videos after all
>>
>>109561286
>nvfp4
>on my 3090TI
I didn't even know nvfp4 worked on a 3090 ti.
>>
GemmaPrompt cuck, why couldn't you optimise your UI?
>>
>>109561307
It works, it's just not hardware accelerated, only Blackwell has that
>>
>>109561249
make me, newfag.
>>
>>109561309
Why are you asking here instead of making an issue?
>>
>>109561131
why doesn't pw post her kino anymore?
>>
>>109560725
Cute anime girl should not kill
>>
>>109561334
kys ani
>>
>>109561307
It works.
Int8_convrot is supposed to be better than nvfp4 for my GPU, but I guess the extra 10 GB makes total time actually worse
>>
are there any other good H3 prompt writers now that GemmaPrompt is abandonware?
>>
>>109560933
>>109560948
Theres no difference between INT8 and NVFP4 text encoder. NVFP4 is just for RTX 5000 and above series
>>
>>109561131
>>109561319
Lumi!!! :3 I missed you so much!
>>
>>109561342
fuck you nigger her sunos are kino
https://suno.com/song/b4bc27c9-a55c-4f58-9f6b-92b9b7994bdf
>>
>>109559522
anons, I need your help
it's been a while since I did anything with ComfyUI
now, all my detailer nodes stopped working, and the small details look smeared
I suppose an update broke something... but I don't know what, and I don't know how to fix it

care to review my workflow, please?
>https://files.catbox.moe/civrb9.json
>>
>>109561373
buy an ad
>>
CENSORED MODELS
ARE
OVER
>>
>>109561395
bout to
https://suno.com/song/7c89c79b-3b06-49ac-a2c9-2af61365ed62
>>
>>109561208
>>109561292
that's because your retarded prompt is not giving the model enough time for the actions to occur you idiot
>>
>>109561346
So Int8 text encoder is better huh. Guess im gonna download it
>>
>>109561404
we are in a hurry here the tits are not going slap themselves
>>
>>109561334
i wish i knew :(

>>109561365
hey :)
>>
I hope we get a sex lora soon. I know you can just give it a ref to get decent results but I don't wanna go through 3dpd porn to get them.
>>
>>109561373
once again, lumi did it better
https://suno.com/s/S7LrWu5AIWX9yDb2
>>
File: ComfyUI_temp_fubvm_00004_.png (2.87 MB, 1440x1920)
2.87 MB PNG
>>109561283
Just used one image ref (pic related) but with a retarded prompt, just prompted shot by shot and the aesthetic, could've done better I think
>>
>>109561425
love u lumi
>>
>>109561389
I can't help right now anon but I wish you the best of luck
>>
>>109561441
>>
File: H3_noAudio__00032_.mp4 (1.46 MB, 1248x896)
1.46 MB
1.46 MB MP4
>>
File: 1774311044036577.jpg (1.86 MB, 1920x2468)
1.86 MB JPG
I think this is the time where i should put my spare 32gb of ram so my total ram is 96 gb
>>
>>109561438
Based H3, insane we have this locally now
>>
>>109561471
H3 + the upcoming image model will be an amazing combo. Shame the music model seems like a dud though.
>>
>>109561527
how much resources does it take to run the music model? i tried running the huggingface demo but it gave me an error and then rate limited me for 24 hours
>>
>>109561389
Just start over and download new ComfyUI portable
>>
File: MiniMax_H3_00073_.mp4 (386 KB, 928x672)
386 KB
386 KB MP4
>>109560723
Kek, not bad
https://files.catbox.moe/klro9f.mp3
>>
>>109561538
Haven't tried it yet but I'm pretty sure it's supposed to fit into 24GB.
>>
File: 1761840963134627.mp4 (197 KB, 576x736)
197 KB
197 KB MP4
>>
>prompt a sweep over something (arc shot)
>camera rolls instead
>>
PW shouldn't be ashamed from postiing anything. i need to hear her kino othwerwise i feel like i'm missing out some amazing perspectrive kino.
>>
File: 1763798753718965.jpg (148 KB, 1024x1024)
148 KB JPG
>>109560792
k2 identity edit is nice. this is just with the single reference image from >>109560829
>>
File: 35352.webm (3.7 MB, 448x256)
3.7 MB
3.7 MB WEBM
>>
https://files.catbox.moe/ipj4zs.mp4
>>
so now that the dust has settled: pussy and asshole lora when?
>inb4 it's on civi-
stfu you retarded dumb nigger I tried all of them already, none work
reference: hd close up of a girl's face by another girl's pussy
prompt: perfectly following the guide to get the girl to eat the other one out
results: insane body horror, with or without the loras
>>
>>
File: MiniMax_H3_00075_.mp4 (1.41 MB, 928x672)
1.41 MB
1.41 MB MP4
Early 90s Noise Rock/Post Hardcore on MinMax Music 3
https://files.catbox.moe/94z881.mp3
>>
>>109561568
Just say pan. It's always a pan
Moving left? Pan left
Following? Pan with
Rotating? Pan around in 3d
Spinning? Pans in a spiral
>>
>>109561538
It sucks. Don't bother.
>>
>>109561648
https://suno.com/song/d2526144-9f3c-48c3-b610-8b27b187664d
>>
>>109561632
Is this really the music model? wow.
>>
And if I was a fish
I wouldn't have to do dishes
Or take out the trash
Or clean my room
And I wouldn't have to get dressed in the morning
Or brush my teeth
And I wouldn't have to wear shoes
'Cause fish don't wear shoes
And I wouldn't have to wear pants
'Cause fish don't wear pants
PW is just ahead of her time desu.
>>
File: 1448800539597.png (660 KB, 1106x1012)
660 KB PNG
My manual prompts for I2V come out better than GemmaPrompt.
>>
>>109561698
>>109561698
>>
>>109561632
Damn, pretty good
>>
>>109561663
Yep, my fav gen so far is its rendition of BABYMETAL style music
https://files.catbox.moe/8au6uh.mp3

Such a fresh dataset. Though the only downside is we can't train LoRAs or do references unless they release an encoder that they withheld as Ostris explained here
https://xcancel.com/ostrisai/status/2088272804609503543#m
(unfortunately)

But once someone reverse engineers that encoder (which apparently costs money) or they release it, local music will be SOTA
>>
>>109561644
Gotta try that, I hope it also fixes character flipping vertically when camera view moves, like if head is on top of the screen, it ends up on the bottom as camera moves.
>>
>>109561361
Of course there is a difference, the quantization is different resulting in a 10GB difference.
NVFP4 being hardware accelerated on blackwell has nothing to do with it.
>>
>>109561694
>human writing is better than llm writing
Many such cases
>>
>>109559522
Has anyone ever noticed that VideoHelperSuite breaks in-browser video previews? If I disable it, I can preview videos in the browser by clicking on the magnifying glass and it works just fine but if I enable the extension firefox tells me there's a mime type error.
>>
>>109561755
maybe it is outputting a format that isn't supported by the browser
>>
>>109561760
I don't have any VHS nodes in most of my workflows but it happens regardless. I only noticed it because I cloned a different comfyui version from the repo and previews worked and as soon as I installed VHS it started fucking up.
>>
>The low quality test run comes out perfect
>The 1MP long run has a single flaw that ruins the whole thing
Such is life I suppose
>>
https://www.youtube.com/watch?v=3UhHsfcz--A
>>
>>109561979
Shill your slop somewhere else.
>>
>>109561843
Video inpainting when?
>>
I'd like to see someone improve the efficiency of Klein.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.