[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: tjrh.jpg (141 KB, 1996x1044)
141 KB JPG
Discussion and Development of Local Image, Video, and Music Models

Previous: >>109660183

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg
>>
File: ComfyUI1_08352_.png (1.09 MB, 832x1408)
1.09 MB PNG
>>109662682
first
>>
File: Attention ldg.mp4 (1.99 MB, 960x544)
1.99 MB
1.99 MB MP4
>>109662682
>>
File: joerogan.webm (3.87 MB, 320x320)
3.87 MB
3.87 MB WEBM
>>
A HELL NAW UNC FINNA CRASH OUT WITH THIS ONE
>>
File: ZI_00108_.png (3.58 MB, 1664x2208)
3.58 MB PNG
https://litter.catbox.moe/7xshz9.png
>>
File: ComfyUI1_08364_.png (1.3 MB, 832x1408)
1.3 MB PNG
>>109662689
>68*C
mine is at 64C.......
>>
>>109662682
>julienbake

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
>>109662235
Can you post a workflow that just works?
I tried pdd, and it looked terrible
>>
too many schizos...
>>
>mfw Resource news

08/27/2026

>Nvidia agrees to buy Hugging Face for $12.9 billion
https://www.reuters.com/technology/nvidia-talks-acquire-hugging-face-13-billion-deal-business-insider-reports-2026-08-27

>lightx2v/Minimax-h3-Turbo · 8-step 768p V1.0 LoRA released
https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/48#6a8fd8bb97c3b91d430a1c86

>MiniMax H3 Director Cut Studio
https://github.com/karuvanan/MiniMax-H3-Director-Cut-Studio

>ComfyUI-Raylight-Windows: 2x H3 iteration speed with dual gpu
https://github.com/9nate-drake/ComfyUI-Raylight-Windows

>GGSS: Geodesic-Gated Spherical Steering for Inference-Time Debiasing of Generative Vision-Language Models
https://github.com/dukesun99/GGSS

>RefVideo-6M: A Reliable Reference-Based Dataset for Instructional Video Editing
https://huggingface.co/datasets/RefVideo6M/RefVideo6M

>VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
https://video-reason.com

>When Composition Doesn't Add Up: Humans Identifying Defects in AI-Generated Images
https://github.com/Future-IQA/CO-AID

>Comfy H3 Sync Sound Community Challenge
https://blog.comfy.org/p/comfy-h3-sync-sound-community-challenge

08/26/2026

>MiniMax H3 Acc FL2VA & REF2VA LoRAs By Wan Team
https://huggingface.co/alibaba-pai/MiniMax-H3-Acc-LoRAs

>MiniMax-H3 Omni Prompt Rewriter LoRA
https://huggingface.co/lightx2v/MiniMax-H3-Prompt-Rewriter-LoRA-Omni

>MiniMax-H3 Acc LoRAs — ComfyUI conversion
https://huggingface.co/aptech0081/MiniMax-H3-Acc-LoRAs-ComfyUI

>H3 Character Sheet Generator
https://huggingface.co/PoopMan333/H3_Character_Sheet_Generator

>DiffusionOPSD: On-Policy Self-Distillation in Diffusion Models
https://diffusionopsd.github.io

>TurboT2VA: Fast Large-Scale Text-to-Video-Audio Generation via Score-Regularized Consistency Distillation
https://github.com/thu-ml/TurboDiffusion/tree/main/turbot2va

>LeFlow: Generative Latent Flow Planning World Models
https://github.com/hsiangwei0903/LeFlow
>>
>mfw Research news

08/27/2026

>VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality
https://arxiv.org/abs/2608.25452

>Efficient Training with Foresight: Multi-Token Auxiliary Supervision for Autoregressive Image Generation
https://arxiv.org/abs/2608.25386

>4DStreamCtrl: Interactive Video Generation with Online 4D Control
https://arxiv.org/abs/2608.25479

>Plans You Can Check: Verifier-Grounded Learning of an Open-Weight Planner for Executable Video-Editing
https://arxiv.org/abs/2608.25622

>Towards Purified Multi-Label Test-Time Adaptation of Vision-Language Models
https://arxiv.org/abs/2608.25653

>Uncertainty-Guided Latent Diffusion Models for Faithful Super Resolution
https://arxiv.org/abs/2608.25998

>FlashNormal: Detailed Surface Normal Estimation from Flash and No-Flash Images
https://arxiv.org/abs/2608.25360

>MLLMCLIP: Feature-Level Distillation of MLLM for Robust Vision-Language Representations
https://arxiv.org/abs/2608.25575

>On the Separation of Human and AI-Generated Images in CLIP Embedding Space
https://arxiv.org/abs/2608.25609

>Unsupervised Post-Training of Foundation Models: A Survey
https://arxiv.org/abs/2608.24982

>Visual General Intelligence: A White Paper
https://arxiv.org/abs/2608.25924

>DeCO: Discriminative Evidence Composition for Fine-Grained Dataset Distillation
https://arxiv.org/abs/2608.25480

>Not All Attention Heads Contribute to Critical Visual Token Selection: Head-Aware Pruning Matters More?
https://arxiv.org/abs/2608.25332
>>
File: shinzo.gif (1.22 MB, 256x256)
1.22 MB GIF



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.