[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: 1786386945448800.webm (3.3 MB, 1820x980)
3.3 MB
3.3 MB WEBM
Discussion and Development of Local Image, Video, and Music Models

Previous: >>109518049 (Cross-thread)

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

please don't feed the schizos
>>
>>109519341
thanks for bakering!
>>
>>109519341
This is Debo thread.
Go to
>>109519228
>>
based official thread
>>
>mfw Resource news

08/10/2026

>H3 Motion Context: Chain MiniMax H3 clips
https://github.com/NikoDemon80/ComfyUI-H3-Motion-Context

>MiniMax H3 Turbo — ComfyUI 4-Step T2V and I2V LoRA
https://huggingface.co/joyfox/MiniMax-H3-Turbo

>HRDiT: Training-Free High-Resolution Image Generation with Off-the-Shelf Diffusion Transformer Models
https://github.com/zylwithxy/HRDiT

>RoRA: Role-Oriented Regional Allocation for Visual Token Pruning in MLLMs
https://github.com/LukieLuu/RoRA

>AVCap: Reinforcing Audio-Video Joint Caption with Detail-Aware Reward
https://huggingface.co/collections/Apryle/avcap

>Prune Once: Retraining-Free Task-Agnostic Pruning for Vision-Language Models
https://github.com/cau-hai-lab/PORTA.git

>Unsloth Minimax H3 GGUF (Q2:Q8)
https://huggingface.co/unsloth/MiniMax-H3-GGUF

>MiniMax-H3 for Apple Silicon - Rebuilt from the official weights
https://huggingface.co/uetuluk2/minimax-h3-mlx-rebuild

>Soran’t: Small Next.js front-end for video generation on ComfyUI
https://github.com/pwillia7/ai_video_fe

08/09/2026

>Kroma v0.2 — Krea 2 fine-tune (full model)
https://huggingface.co/lodestones/Kroma

>krea2-turbo-bbox
https://huggingface.co/jimmycarter/krea2-turbo-bbox

>Kroma v0.2 Quant
https://huggingface.co/silveroxides/Kroma-Quant/tree/main

>Spectrum for Ideogram 4
https://github.com/Nif00/ComfyUI-Spectrum-Ideogram4

>ClipProj — MiniMax H3 conditioning from a Qwen3-VL-4B
https://huggingface.co/NicoLab28/ClipProj-MiniMax-H3

>ComfyUI-SigmaSync-LoRA: Sigma-aware model-only LoRA strength scheduling
https://github.com/capitan01R/ComfyUI-SigmaSync-LoRA

>NexusBTA v0.2.44 adds MiniMax H3 support
https://github.com/JpAndreBTA/Nexus-BTA/releases/tag/v0.2.44

>Experimental MiniMax H3 single-image VAE
https://huggingface.co/Mamad8/MiniMax-H3-Image-VAE

>MiniMax H3 REF2VA w4a8
https://huggingface.co/realrebelai/Rebels_w4a8s

08/08/2026

>Kijai: MiniMax H3 Ref Lora Rank 256 bf16
https://huggingface.co/Kijai/MiniMax-H3-experimental/tree/main/loras
>>
>>109519377
debo is in that thread so that's a fucking lie
>>
File: unnamed (20).jpg (85 KB, 1024x926)
85 KB JPG
What do you use to isolate sounds and voices from videos? To use as references anons?
>>
>mfw Research news

08/10/2026

>Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation
https://arxiv.org/abs/2608.06751

>From Cheap Fakes to Pure Synthesis: Addressing the New Era of T2V Fake News Videos
https://arxiv.org/abs/2608.06732

>PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model
https://arxiv.org/abs/2608.06794

>Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models
https://arxiv.org/abs/2608.06768

>MaskFlow: Precise, Consistent and Seamless Regional Image Editing
https://arxiv.org/abs/2608.06929

>Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity
https://arxiv.org/abs/2608.07120

>Addressable Memory for Video World Models
https://research.nvidia.com/labs/sil/projects/WorldTrace

>Local Epistemic Uncertainty Guided Active Sampling for Plug-and-play Diffusive Image Restoration
https://arxiv.org/abs/2608.06981

>ControlRef: Efficient Layout-Guided Multi-Instance Generation via Anchored 4D-RoPE
https://arxiv.org/abs/2608.06878

>CustomDance: Customized 3D Dance Generation with Coarse-to-Fine Human-Centered Interactive Control
https://arxiv.org/abs/2608.06722

>Bend the Basics: Degradation-Aware Deformable Tokenization for All-in-One Image Restoration
https://arxiv.org/abs/2608.06832

>Stable Curves, Unstable Items: Item-Level Scaling Heterogeneity in Video LLMs
https://arxiv.org/abs/2608.07014

>A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs And Improve Accuracy
https://arxiv.org/abs/2608.07427

>Alignment has a Fantasia Problem
https://arxiv.org/abs/2604.21827
>>
>>109519341
can I run a local model that will nudify all my hot facebook friends (milfs included)?
>>
>>109519409
yes
>>
>>109519409
yes but all the software is shitty garbage that will break so I hope you like being a tinker tranny
>>
>>109519397
https://github.com/facebookresearch/sam-audio
>>
>>109519409
sorry sarr, that is very illegal
>>
>>109519471
thanks anon , will have a read.
>>
so now that you can make anyone suck your peepee, what do you do with this power?
>>
>>109519397
give her back, sasori
>>
>>109519580
suck own peepee
>>
people really love schizo lore instead of diffusion? what a sad state of affairs for /ldg/
>>
>>109519860
The only person screeching about that is the "anon" complaining about it
Not that you should be considered a person, Trani



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.