[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


"Permissable" Edition

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109790622

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/neo_collage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
Boba thread of frendship
>>
>mfw Resource news

09/11/2026

>ComfyUI-VDN-H3-24GB
https://github.com/Speach1sdef178/ComfyUI-VDN-H3-24GB

>LTX-2.5 22B IC-LoRA Reference Sheet Control
https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Ingredients#ltx-25-22b-ic-lora-reference-sheet-control

>FastH3-Live v1.2.0 update
https://huggingface.co/datasets/jacokon/fasth3-live

>MiniMax-H3-Image-Training-Adapter
https://huggingface.co/circlestone-labs/MiniMax-H3-Image-Training-Adapter

>TaoMate-H3: Low-latency streaming audio-video generation runtime built on MiniMax H3
https://github.com/TaoLiveAIGC/TaoMate-H3

>FrameForge Motion Context Video Editor for ComfyUIR
https://github.com/spacesimeco-hue/Chain-Motion-AI-Video-Editor

>From Evaluation to Enhancement: Benchmarking and Improving Think-with-Video Reasoning for Video Generative Models
https://huggingface.co/datasets/KlingTeam/VWG-Bench

>Uncertainty DMD: Restoring Diversity in Few-Step Autoregressive Video Distillation
https://scdzx.github.io/Uncertainty-DMD

>A Multi-View and Confusion-Guided Ensemble Framework for Robust Synthetic Image Attribution
https://github.com/ZOMIN28/SIA

>FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation
https://github.com/msu-video-group/freeflow

>World in World: Explore the World with World Models
https://chenxi-song.github.io/worldinworld

>OmniKVQuant: KV Cache Quantization for Omni-LLMs
https://github.com/kaistmm/OmniKVQuant

>BSAI-MiniMAX-H3-Prompt
https://github.com/xm6018924/BSAI-MiniMAX-H3-Prompt

>ComfyUI Native Yue2
https://huggingface.co/Comfy-Org/Yue2

09/10/2026

>Interpreting Object-Dependent Concept Brittleness in Text-to-Image Diffusion Models
https://github.com/Metecade/Object-Dependent-Concept-Brittleness

>Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout
https://alicezrzhao.github.io/mask_forcing

>Geodesic-informed Generative Diffusion Model For Topology-preserved Image Video Generation
https://github.com/nellie689/IGG
>>
>mfw Research news

09/11/2026

>Harnessing Intrinsic Subject-Aware Attention for Controllable Multi-Subject Video Generation
https://arxiv.org/abs/2609.11507

>AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow
https://arxiv.org/abs/2609.10723

>Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation
https://arxiv.org/abs/2609.11638

>Overpainting: Localized Context-aware Diffusion Image Editing
https://overpainting.github.io

>Learning Interaction between Image and Layout Priors for Joint Image-Layout Generation in Design Templates
https://arxiv.org/abs/2609.11519

>OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models
https://arxiv.org/abs/2609.11244

>Shedding Light: A Benchmark for Evaluating Lighting Understanding in Generative Image Models
https://lvsn.github.io/SheddingLight

>Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling
https://compvis.github.io/logit-refiner

>CamPilot: A Multi-Agent Cinematic Assistant for Camera-Controlled Movie Generation
https://arxiv.org/abs/2609.10943

>Mi-Ripple: Restoring Images Degraded by Iterative AI Editing
https://arxiv.org/abs/2609.11317

>SenseNova-U1.5: Towards Native Unified Visual Intelligence
https://arxiv.org/abs/2609.11929

>LoopVAE: Recurrent Depth Across Scales for Visual Tokenization
https://arxiv.org/abs/2609.11516

>Revisiting Avatar-As-Image: High-Fidelity Registration is All You Need
https://yuxuan-xue.com/avaimg

>Are We Really Doing Few-Shot Learning? A Critical Examination of Pre-Training Assumptions
https://arxiv.org/abs/2609.10851

>MultiHuSE: A Multimodal Dataset for Humour Styles and Emotions
https://arxiv.org/abs/2609.11322

>Leveraging Avatar Fingerprinting: A Multi-Generator Photorealistic Talking-Head Public Database and Benchmark
https://arxiv.org/abs/2603.26934
>>
File: VOLCANO.mp4 (3.92 MB, 2048x1130)
3.92 MB
3.92 MB MP4
>>
>>
>>109795620
>>109795625
thanks
>>
File: output_no_audio_lossless.mp4 (3.36 MB, 1056x608)
3.36 MB
3.36 MB MP4
I think there's enough people to take care of things here.
I don't need to do much from what I've seen.
Hope something new comes out to bring life back to this thread but as of now there's not much of value atm, plus everything is being self regulated regardless.
Until then I'm building shit
>>
>>109795639
more
>>
File: ComfyUI_temp_chdmi_00003_.png (2.76 MB, 1120x2080)
2.76 MB PNG
>>
File: ComfyUI_temp_chdmi_00004_.png (2.91 MB, 1120x2080)
2.91 MB PNG
>>
Blessed bread
>>
File: ComfyUI_temp_chdmi_00006_.png (3.16 MB, 1120x2080)
3.16 MB PNG
>>
File: keekkekekeeeeek.png (1.17 MB, 864x1184)
1.17 MB PNG
yoooooooo what this diddyahh unc postin rn bro keeeeeeeek JSID too many gooners in this app
>>
File: MiniMax_H3__00556.mp4 (2.37 MB, 576x896)
2.37 MB
2.37 MB MP4
>>
>>109795672
AI is so awesome.
>>
>>109795578
>I know you guys mostly don't care about turbo but after messing with the latest 10Eros H3 it's pretty damn good. Still not perfect but it handles basic NSFW motions and actions a lot better than base and doesn't have much visible turbo fry.

Yes I've been using it myself, I'm not sure what kind of black magic they did but 4 steps works great.
>>
>>109795578
Post an example bruh.
>>
>>109795696
do zoomers really post pepe all the time like its some kind of new thing? shit has been a thing for over 20yrs!
>>
File: 1771709384635416.png (3.54 MB, 2176x1216)
3.54 MB PNG
>>109794867
ty. the style is mostly @watanabe tomari, jaggy lines, oekaki
https://pastebin.com/HBZpbeCT
>>
>>109795760
uncs be posting goofy ahh thangs. real ngas know the frog is W
>>
>>109795714
>>109795578

Just use his LTX finetune if you want basic NSFW motions
>>
File: MiniMax_H3__00560-small.mp4 (3.72 MB, 832x1248)
3.72 MB
3.72 MB MP4
>>109795754
chek it out
>>
File: MiniMax_H3__00561.mp4 (3.47 MB, 832x1248)
3.47 MB
3.47 MB MP4
>>
>>109795578
>>109795789
>>109795817

I used it but its kills your prompt adherence by half
>>
>>109795821
Every lora does that.
>>
>>109795702
>>109795789
>>109795817
Their scaling is off desu
>>
>10Eros_Max_h3_TURBO-hybrid_beta5_int8.safetensors
https://files.catbox.moe/swoz90.mp4
>>
File: 7jtjqs.png (113 KB, 526x277)
113 KB PNG
>>
>>109795823
Not Riding Lora (mostly)
>>
>tfw you can’t post your gens anymore because some faggot will use it to gen shitty videos

It’s all so tiresome…bye bye /ldg/
>>
File: MiniMax_H3__00563.mp4 (3.41 MB, 768x1376)
3.41 MB
3.41 MB MP4
>>
>>109795905
good riddance faggot, let the door hit you on the way out
>>
manchilds up in here with all their drama
>>
File: 1763435323662447.jpg (6 KB, 320x180)
6 KB JPG
>>109795905
can't believe this is a real post
>>
>>109795910
Shut the fuck up you dirty pajeet vermin
>>
File: MiniMax_H3__00564.mp4 (3.82 MB, 768x1376)
3.82 MB
3.82 MB MP4
>>
>>109795930
lmfao you little girl
>>
>>109795929
It's not.
>>
>>109795929
Your videos are good as your image gens, no wonder he’s so excited generating videos, he never had something like that in his hands, doesn’t have the brains to generate images on his own
>>
>>109795942
>underage on 4chan
>>
>>109795935
Saaar, your kind is so easy to identify
>>
File: MiniMax_H3__00565-small.mp4 (3.72 MB, 768x1376)
3.72 MB
3.72 MB MP4
>>109795950
>>
File: MiniMax_H3__00566-small.mp4 (3.72 MB, 768x1376)
3.72 MB
3.72 MB MP4
>>
please give us 100 more of the exact same gen
>>
>>109796002
make something yourself
>>
>>109795936
it was
>>
>>
File: MiniMax_H3__00570.mp4 (2.73 MB, 640x640)
2.73 MB
2.73 MB MP4
>>
File: MiniMax_H3__00572.mp4 (2.5 MB, 640x640)
2.5 MB
2.5 MB MP4
>>
Saaar
>>
>>109796076
kindly do the needful and shut the fuck up bachoola
>>
File: MiniMax_H3__00573.mp4 (2.75 MB, 640x640)
2.75 MB
2.75 MB MP4



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.