Career Opportunities Edition Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109637805https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
>mfw Resource news08/24/2026>MiniMax-H3-Fun-Controlnet-Union https://huggingface.co/alibaba-pai/MiniMax-H3-Fun-Controlnet-Union>MiniMax-H3-Longvideos: Long (up to ~120s) MiniMax-H3 video + synchronised audio from a single prompthttps://huggingface.co/Smite79/MiniMax-H3-Longvideos>DiGS-Avatar: Single-Image Animatable 3D Human Reconstruction via UV-Space Diffusionhttps://github.com/KLMAV-CUC/DiGS-Avatar>Identity-Preserving Text-to-Video Generation via Agentic Enhancement and Semantic Repairhttps://github.com/oceanflowlab/AESR>OccluRank: Controllable Occlusion-Aware Layout-to-Image Generation by Adding Just an Ordinal Rankhttps://github.com/Wenyang-hong/OccluRank>Aggregating Visual Information with Optimal Transport for VideoLM Token Compressionhttps://github.com/ernie-research/AVIOT>CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxationhttps://github.com/CubicSplat/repo>Vis-Poison: Poisoning Visual Knowledge in Multimodal Retrieval-Augmented Generationhttps://github.com/SWUFE-DB-Group/Vis-Poison>Explainable Deepfake Detection with Feature-robust Augmentation and Evidence-grounded Explanation Optimizationhttps://github.com/oceanflowlab/EDD.git>ArtiMo: Agent-Driven Articulated Mesh Animationhttps://zou-2004.github.io/ArtiMo>MiniMax-H3 × Z-Image — spatial detail graft (comfy-native) https://huggingface.co/joeygambino/MiniMax-H3-x-Z-Image-native>MiniMax-H3 4-Step LoRA (FlashGen)https://huggingface.co/Beidouqixing/minimax-h3-4step-lora-flashgen08/23/2026>Krea 2 Turbo — 4-Step Distillation LoRAhttps://huggingface.co/lvladikov/Krea2-Turbo-Distill-4step-LoRA>Alibaba to issue US$10 billion in new shares for huge AI pushhttps://www.scmp.com/tech/big-tech/article/3364957/alibaba-issue-hk80-billion-new-shares-global-ai-push>Nvidia Customers Notified About AI-Related Price Hikes Above 15%https://www.bloomberg.com/news/articles/2026-08-22/nvidia-customers-notified-about-ai-related-price-hikes-above-15
Blessed thread of frenship
>mfw Research news08/24/2026>DiffVC-ONE: Diffusion-based Generative Video Compression with One-Step Video Diffusion Transformerhttps://arxiv.org/abs/2608.20515>MultiCube: Compositional 3D Generation With Part-Level Semantic and Spatial Controlhttps://multi-cube.github.io>GAP-SAM: A Global Artifact Prior for Generalizable AI-Generated Image Manipulation Localizationhttps://arxiv.org/abs/2608.20929>Grounded-Exo2Ego: Structured Semantic Grounding for Robust Exocentric-to-Egocentric Video Generationhttps://research.nvidia.com/labs/amri/projects/grounded-exo2ego>ES-VP : Energy-Shaped Dynamic Visual Prompting for Efficient Model Adaptationhttps://arxiv.org/abs/2608.21194>Scaling Muon for Diffusion Transformershttps://arxiv.org/abs/2608.20818>InfinityEdit: Infinite Video Editing with a Lightweight Edit-Ignition Adapterhttps://arxiv.org/abs/2608.20910>Anchoring Instruction Outside Mask: Exact Reference Caching for Efficient In-Context Diffusion Transformershttps://arxiv.org/abs/2608.21229>Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMshttps://orarl.github.io>Bridging Language and Spherical Space: Object-Centric Control for Text-to-Panorama Generationhttps://arxiv.org/abs/2608.20691>When does fusing hand-crafted knowledge with learned representations pay? A cost-normalized benchmark of stacking, substitution, and interferencehttps://amughrabi.github.io/MomentAux>Is Visual Prompting All You Need? Studying VLM Spatial Reasoning under Progressive Visual Scaffoldshttps://arxiv.org/abs/2608.21170>Enabling Memory-efficient Im2win Convolution with Multi-precision Support on GPU CUDA and Tensor Coreshttps://arxiv.org/abs/2608.20725>LoRC: Detecting AI-Generated Images via Low-Rank Collapse in Semantic Residualshttps://arxiv.org/abs/2608.20882>Llama-Mobile: Efficient 2.7-Bit Quantization of VLMshttps://arxiv.org/abs/2608.21134
first for lolcow baker has no employment opportunities
>>109640479>>109640482
The advert imagegens from last thread were extremely aesthetic, what model+lora combo is that
>>109640487>>109640497
>>109640524>>109640504>>109640499>>109640471This is less fappable than niggers please switch or stop thanks
>Posted in the dead thread award>>109640289I was very impressed by anima when I was playing around with it, but I couldn't get good controlnets. I want to make a visual novel/WEGslop. I've got a couple style loras trained for illustrious, I'm not sure how necessary controlnets are for what I want, but I had a feeling I'd need them at some point.
>>109640598thats because you're indian
>>109640657That makes no sense it would be easier to self insert if I was Indian kill yourself
>>109640598>>109640684Calm down, Rakesh.
>>109640684But you're Indian
>>109640709>>109640702it's okay to be Indian
>>109640714What about black?
can't go to sleep until i queue up all my gens
>>109640723not okay
>>109640725why do you let one bake the threads?
>>109640746sniff
>>109639981Aren't those both Strix Halo? You loaded bastard. Try them out, they should be usable. Either download a portable AMD build of ComfyUI from the latter's Github, or for manual AMD setup steps, see >>109625036 and >>109625041.
posting on this website:https://litter.catbox.moe/90ryru8cmb5elkwh.mp4
>>109640775oops, disregard image that was from the game test gen.
>>109640775ummm, why didn't she get TOS banned?
lets see if this works:https://litter.catbox.moe/x815nhkdvmziq3ps.mp4
>>109640823off by one
>>109640823better check (im not going to keep trying, but the point worked this time.)https://litter.catbox.moe/ausd7dv6anwmmd7v.mp4
>>109640340Maybe someone will figure out how to train a lora but it does not affect voice path . lol
>>>/gif/31084754which one of you is this
how much h3 latent space is 40gb of ram worth? or how many seconds at 1mp is it
which lora is better?https://twinlens.app/compare?share=30e5ab353574
>>109640993>>109641018i dont know
so the model is really good at knowing what initial D is. 2 references, yui + the car. threw in some appropriate music but thats it.https://litter.catbox.moe/zp1opqgq410zbdam.mp4
>>109640993the first 40 tonnes, the last 40 barely anything
>>109641104can it animate the wings of the car?
>>109641122it seems to have knowledge of the show so I wouldnt be surprised if it could do that if it can do the camera stuff like this.
cozy breas
>>109641128also im gonna see if it can do a duel next with another car just for fun.
>>109641128i believe the wings only showed up once in the entire series so i have doubts on it getting it right unless it trained on lots of youtube reuploads of that scene