[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: Krea2_turbo_hr_fix_00210_.jpg (2.81 MB, 2512x3344)
2.81 MB JPG
Discussion and Development of Local Image, Video, and Music Models

Previous:>>109705821

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: 927707418926341.mp4 (3.73 MB, 672x928)
3.73 MB
3.73 MB MP4
>>
Blessed thread of frenship
>>
I can save this general.
>>
what's the current upcoming FOTM model that'll be forgotten again in a month after release
>>
>>109712114
why should i use portable over desktop? legitimate question
>>
>>109712324
TY 4 BAKE
>>
Pure goonslop
>>
>>109712294
Neat. Did you also use the DaSiWa checkpoint? Kinda wanna know, if this one is any good. His WAN2.2 checkpoints were pretty nice, and his LTX checkpoints pretty bad.
>>
>mfw Resource news

09/02/2026

>ComfyUI MiniMax H3 MotionCache
https://github.com/Mozer/ComfyUI-MiniMax-H3-MotionCache-FastVAE

>MiniMax-H3 WebUI — 视频生成工作台
https://github.com/AntaresAlice/h3-webui

>H3-World: Turning Language Understanding into World Control
https://huggingface.co/DANNY621/H3-World

>Identity-Conditioned Latent Consistency Distillation for Face Synthesis
https://github.com/UFPR-IPASP-PR/FaceRec-IdentityConsistency

>TUE-Detector: A Tool-Using Expert MLLM-Based Detector for AI-Generated Videos
https://github.com/Louis-YW/TUE

>MegaStyle++: Scaling Image Style Space through Hierarchical Style Definition
https://github.com/Tencent/MegaStyle

>PredErase: Training-Free Object-and-Effect Removal with Predictive Latent Guidance
https://github.com/xiuwk0820-collab/PredErase

>video2dlssnr: standalone DLSS 5 Neural Rendering video tool
https://github.com/DaniilSokolyuk/video2dlssnr

>ComfyUI MiniMax H3 Video Outpaint
https://github.com/TwoAbove/ComfyUI-H3VideoOutpaint

>ArtiFixer: Few-step causal auto-regressive model that enhances and extends 3D reconstruction
https://huggingface.co/nvidia/ArtiFixer

09/01/2026

>ComfyUI-H3VAE_TRT: TensorRT version of the MiniMax-H3 VAE in ComfyUI
https://github.com/lihaoyun6/ComfyUI-H3VAE_TRT

>MiniMax-H3 Fused Turbo (INT8 ConvRot)
https://huggingface.co/MATLOWAI/minimax-h3-fused-turbo-int8-convrot

>RegionCache: Semantic-Aware Region Reuse for Efficient Multi-Turn Image Generation
https://github.com/hebutBryant/RegionCache

>Discrete Diffusion Bridges for Spatiotemporally Aligned Image Translation and Generation
https://github.com/HKU-HealthAI/DDB

>FoundYou: A Unified Model for Personalized Segmentation and Retrieval
https://ga1i13o.github.io/FoundYou

>ContextBias: Controlled Evaluation of Bias Persistence Under Context Shift in Text-to-Image Models
https://github.com/Sina-Emami/ContextBias

>FairReL: Deepfake Detection using Fairness-Aware Representation Learning
https://github.com/xiaoman89/FairReL
>>
>>109712382
I've only ever pulled and ran from the repo as all white men should.
>>
>mfw Research news

09/02/2026

>SpatialGuard: Harness-Guided Verifiable Spatial Reasoning for Text-to-Image Generation
https://arxiv.org/abs/2609.01582

>DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution
https://arxiv.org/abs/2608.31106

>Gaussian Core LoRA: Distribution-Aware Dynamic Adaptation for Broad Concept Erasure
https://arxiv.org/abs/2609.01433

>Physically Plausible Video Generation via Visual-Semantic Chain-of-Events Conditioning
https://arxiv.org/abs/2609.00656

>Training-Free Inpainting Across Domains with a Frozen Text-to-Image Diffusion Model
https://arxiv.org/abs/2609.00862

>CameraEditor: Camera-Controlled Image Editing via Video-Prior Sequential Modeling
https://arxiv.org/abs/2609.01479

>SAGE: Subpopulation-Aware Generative Enhancement for Mitigating Spurious Correlations
https://arxiv.org/abs/2609.01051

>Denoising Diffusion Generative Models Secretly Calculate Attentions
https://arxiv.org/abs/2609.00885

>Advanced Pixel Diffusion Model with Guided Sparse Global Refinement
https://arxiv.org/abs/2609.00798

>No Pixel Left Behind: Filling Gaps in Anime Colorization
https://arxiv.org/abs/2609.00800

>GenScale: A Benchmark for Relative Object Scale in Image Generation and Editing
https://arxiv.org/abs/2609.00525

>MeRoPE: Metric Rotary Position Embedding for Camera-Controlled Video Generation
https://qiaozhijian.github.io/merope

>Reliability Challenges in Diffusion Vision-Language Models
https://arxiv.org/abs/2609.01318

>Less Is More: Balancing Positive and Negative Space in Visual Concept Blending
https://arxiv.org/abs/2609.00476

>Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models
https://arxiv.org/abs/2609.00355

>Solaris: Towards Interfaces That Are Generated, Not Coded
https://runway.com/news/research/introducing-solaris

>TPSO: Training-Free Diverse Image Generation via Semantic Prompt Embedding Optimization
https://arxiv.org/abs/2511.19811
>>
>>109712408
Why are you linking some 7 star repo that is entirely in Chinese?
>>
>>109712090
> so local is dead once again
no it's not
h3 is crazy powerful and slow it will take months to master it
in one or two years local will be mind blowing
unless they took hardware from us
>>
File: 773332516849502.mp4 (3.7 MB, 928x672)
3.7 MB
3.7 MB MP4
>>
>>109712429
hes just trolling you can ignore
>>
> >109712408
> >109712415
Fuck off
>>
Is there any hope for /ldg/?
>>
File: debo_rdf_k2_00047_.png (3.58 MB, 1664x1280)
3.58 MB PNG
>>109712417
for the same reason I link everything else that seems interesting: to present it to anons so that they can see what new things are out there that can potentially be helpful or useful to them
>>
Any easy way to queue lots of gens in comfy?
>>
>>109712451
Can you fucking stop?
>>
>he thinks people actually reads his malware spam
>>
Threadly reminder:

Do not fall for the "ldg/local is dead" posts. These are made by either brand new frens who've never experienced this thread between model releases or trolls.
>>
>>109712462
i do
>>
finally back to cozy
schools back in session
hype has calmed down
finally cozy again
>>
File: debo_rdf_k2_00048_.png (3.3 MB, 1664x1280)
3.3 MB PNG
>>109712460
>asks me a question
>gets mad when I answer
so... just... don't ask me questions if you're gonna get triggered? idgi
>>
File: Joker2.mp4 (3.98 MB, 1264x728)
3.98 MB
3.98 MB MP4
https://files.catbox.moe/ede2ql.mp4
>>
>>109712505
You're not welcome or wanted here and you keep putting new posters in potentially compromising positions. Why are you such a vile piece of shit?
>>
File: 289587968929415.mp4 (3.94 MB, 928x672)
3.94 MB
3.94 MB MP4
>>
>>
>>109712512
KEK

Should've shot her to be honest but I guess Mika-chan is kinda meta invincible in these. Also, does anyone remember when Mika-chan had a sister or BFF or something?
>>
Yep, I think this general is done for.
>>
File: z_00604.mp4 (2.84 MB, 1376x768)
2.84 MB
2.84 MB MP4
Testing 3 refs with just 4 steps
clanker somewhat delivered
>>
>other thread doesnt even have 100 images posted
jesus fuck you tism faggots are insufferable tismfails
>>
>>109712589
shut up faggot
>>
calm down anonie
>>
>>109712573
should've told kekstone to train at 1024
>>
File: 394977296797580.mp4 (3.76 MB, 928x672)
3.76 MB
3.76 MB MP4
>>109712567
cutie patootie
>>
>>109712569
She had a sister in WW2:

https://files.catbox.moe/pojg9y.mp4
https://files.catbox.moe/mj18ze.mp4

Unfortunately she died at Hiroshima or something.
>>
>>109712580
Seems pretty good. What was it, that it didn't deliver on?`

>>109712589
Last thread was kinda shit, but that's not due to too few gens being posted.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.