[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Nothing To Hide Edition

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109828623

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/neo_collage

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
gm saars
>>
File: hag-puncher.jpg (690 KB, 1856x1408)
690 KB JPG
>>
version 3 of the new kino prompt is in the testing phase
>>
He's going to have a big melty when he wakes up and hits the bottle again
>>
File: brewski.jpg (1.17 MB, 2048x2048)
1.17 MB JPG
>>
File: study-session.jpg (456 KB, 1344x1344)
456 KB JPG
>>
File: 81992.png (120 KB, 1710x680)
120 KB PNG
MiniConstruct storytelling is getting very powerful.
>>
cozy breas
>>
File: 6no335.jpg (6 KB, 150x150)
6 KB JPG
>>109836417
Exactly how powerful are we talking?
>>
just woke up from hibernation. how good is yue2 for underground instrumentals?
>>
>>109836417
are the stories in the room with us right now?
>>
>>109836523
you might be the most qualified person here to judge "underground instrumentals" - it's really quite good across various genres and also with references tho.
>>
>>109836523
I'm in the process of testing the instrumental LoRA on audio.cpp (https://huggingface.co/Mothersuperior/YuE2-instrumental-cot-full-loras) right now, it's quite good

https://files.catbox.moe/57gqa7.flac

Also Ostris now supports training LoRAs, this is interesting.
>>
Work in progress
https://files.catbox.moe/pieij6.mp4
>>
>>109836571
>this is interesting
indeed, anon
>>
>mfw Resource news

09/16/2026

>Fizgig 6.0 - RefMod Training, Editing and Exploring
https://github.com/shootthesound/Fizgig/releases/tag/v6.0.1

>SlotDiT: Object-Centric Representations for Diffusion Transformers
https://slot-dit.github.io

>FastVideo-FastH3-8-Step-V2
https://huggingface.co/FastVideo/FastVideo-FastH3-8-Step-V2

>FastVideo FastH3 ComfyUI Repack
https://huggingface.co/FastVideo/FastVideo-FastH3-Comfy

>FastH3 V2 GGUFs
https://huggingface.co/realrebelai/FastH3-V2_GGUFs

>ComfyUI v0.36.0
https://github.com/Comfy-Org/ComfyUI/releases/tag/v0.36.0

09/15/2026

>LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows
https://github.com/LynnReal-AI/LynnReal-Omni

>Meridian: Geometry-guided video model for authoring new observations of existing events
https://huggingface.co/Viggle/Meridian

>SAM3D-Part: Interactive Part Selection and Generation from 3D Objects
https://github.com/Jiahao620/sam3d-part

>RAIN: Region-Aware Inversion Network for Semantic Watermark Extraction
https://github.com/TheLovesOfLadyPurple/RAIN-lightweight-NN-for-one-step-semantic-watermark-extraction

>ComfyUI-Qwen-VAE-Triton
https://github.com/AllenCraigBarnard/ComfyUI-Qwen-VAE-Triton

09/14/2026

>TaoMate-H3 3-Step LoRA for ComfyUI
https://huggingface.co/Robert1212star/TaoMate-H3-3Step-ComfyUI

>ComfyUI MinimaxH3 AutoContext: Automatic segmented inference + inter-segment Latent anchoring
https://github.com/supElement/ComfyUI_MinimaxH3_AutoContext

>Balancing Emotional Alignment and Semantic Consistency in Image Generation via Reinforcement Learning with Valence-Arousal Anchoring
https://github.com/ramon-alana/eit-with-anchor-and-grpo

>MiniMax-H3 RefMod Stacking, Semantic Routing & Multi-Subject Architecture Guide
https://huggingface.co/datasets/malcolmrey/various/blob/main/h3-center/docs/MINIMAX_H3_REFMOD_STACKING_AND_MULTISUBJECT_GUIDE.md

>ComfyUI Fantastic H3 Prompt Builder
https://github.com/Adudeguyman/ComfyUI-Fantastic-MiniMaxH3-PromptBuilder
>>
is there any easy way to remove japanese text from a manga cover and redraw it using AI? trying to replace it with my own english text.
>>
>mfw Research news

09/16/2026

>Efficient Text-to-Image Generation: An Adaptive Step Schedule Controller for Diffusion Models
https://arxiv.org/abs/2609.16572

>VideoMM: Adaptive Macro-Micro Inference for Efficient Video MLLMs
https://arxiv.org/abs/2609.16722

>What Breaks Local Watermarks? A Robustness Benchmark for Local Invisible Image Watermarking
https://arxiv.org/abs/2609.16832

>StackTok: Accelerating VLMs Inference with Budget-Adaptive Visual Token Selection
https://arxiv.org/abs/2609.16841

>PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control
https://czzzzh.github.io/PhysStream

>FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation
https://arxiv.org/abs/2609.16591

>FRPSS: Feature Rearrangement in Pre-Shape Space for Single-Image Generation
https://arxiv.org/abs/2609.16594

>MDN-Control: Mask-Depth-Noise Guided Region Control for Multi-Subject Video Editing
https://arxiv.org/abs/2609.16475

>High-Fidelity Video Quality Assessment with VQA-Specific Saliency
https://arxiv.org/abs/2609.16946

>TecoPrompt: Temporal-Conservative Prompt Learning for Vision-Language Models
https://arxiv.org/abs/2609.16858

>Reasoning with Image Generation
https://hector.gr/reimagin

>ViD: Vision-Dominant Gender Bias Mitigation for Large Vision-Language Models
https://arxiv.org/abs/2609.16647

>Efficient Quantization-Aware Distillation with Cross-Modal Alignment for Edge Vision-Language Models
https://arxiv.org/abs/2609.16689

>GraLoD: Graphics-Inspired Continuous Level-of-Detail Learning for Image Restoration
https://arxiv.org/abs/2609.16578

>Mini-batch Sampling Strategies for Long-Tailed Image Classification: An Empirical Study on CIFAR-100-LT
https://arxiv.org/abs/2609.16365

>VC-Attention: Faster Low-Bit Attention Without Retraining
https://www.nunchux.ai/blog/attention-is-the-video-bottleneck
>>
>>109836784

https://www.linum.ai/field-notes/jit-ddt
this is a cool writeup add this debo
>>
File: debo_hd_k2_00005_.png (1.53 MB, 1664x1069)
1.53 MB PNG
>>109836829
thanks. added to the record
>>
Hibernation Mode: Ultra
>>
when will smith bakes: interesting thread with cool generations
when tranny bakes: crickets...
>>
>>109836371
i am seated
>>
>>109836320
Thanks for the collage, OP. Slim pickings, I'm sure but I'll change that this thread, don'tcha worry.
>>
>>109836842
>halflips
wut

The rest of the misspellings I kinda get but what's this? Half-elf?
>>
File: debo_hd_k2_00014_.png (2.19 MB, 1664x1069)
2.19 MB PNG
>>109837397
halfling, of course
>>
>>109837400
Halflings with knife-ears? If you say so.
>>
File: debo_hd_k2_00016_.png (2.62 MB, 1664x1069)
2.62 MB PNG
>>109837410
AIs have been known to hallucinate
>>
>>109836571
this very good
did it require extensive captioning or you slapped some whatever tags into it?

+ dataset size and time it took to train was?
>>
>>109837417
I was referring to the "of course", debo.
>>
File: debo_hd_k2_00017_.png (2.96 MB, 1664x1069)
2.96 MB PNG
>>109837440
it was a light-hearted of course
>>
>>109837468
Is the style prompt the same for all four? Or what are they respectively?
>>
File: debo_hd_k2_00019_.png (2.98 MB, 1664x1069)
2.98 MB PNG
>>109837479
yeah, they're all the same. though there's some conflicts so its not consistent
hand drawn sketch, sloppy errors, childish drawing with weird proportions

~~character wildcards go here~~

screengrab, arcane design elements, incomplete messy amateur drawing with rough marker outlining, web comic aesthetic, tumblr style, color outside the lines, errant marker strokes, half drawn {steam|magic|arcane|aether}punk backdrop, no clones, no duplicates
>>
*yawn*
>>
WHADDUP, SLOPPERS!? Y'all ready for some good gens for once?
>>
>>109837505
Thanks.

>no clones, no duplicates
What's that about? Also, doesn't really seem like it's drawing outside the lines for example in most of these.
>>
File: dehd_cs39_00049_.png (2.86 MB, 1664x1112)
2.86 MB PNG
>>109837550
>What's that about?
just seeing if the encoder would care. its qwen based so it can sometimes understand counterfactuals
> doesn't really seem like it's drawing outside the lines for example in most of these.
yeah, its too clean. prob turbo distillation getting in the way of the messier intention. the prompt itself is a recycle of the same prompt in chroma and chroma did it better (picrel)
>>
SenseNova-U1.5 status?
>>
>>109837437
I didn't train the LoRA, just downloaded it from HF and inferenced it with the merged VAE. The encoder is not official, so the quality of LoRAs is not the best. Instrumentals are neat so far, but they don't compare in quality to what you can get out SA3 medium, so for pure instrumentals I'd not really use this model.
>>
>>109836523
wtf are "underground instrumentals"? like underground rap? It's so varied you'd have more luck learning how to do them yourself in a VST with an AI teaching you how to use it than generating it with any model out right now.
>>
man this shit is so fucking annoying.
https://github.com/Comfy-Org/ComfyUI_frontend/issues/14599
what actually causes comfyui to be 40% slower while its tab is on-screen vs tabbing out (no, live previews aren't the issue.)
>>
>>109837664
Bespoke personal custom aftermarket vibecoded front end chads win again
>>
File: Cubic Kakihara.png (3.88 MB, 1920x1088)
3.88 MB PNG
>>109837545
Well, ready or not, here I come!
>>
File: The Tortured.png (2.14 MB, 1984x1056)
2.14 MB PNG
>>109837680
>>
File: Paleo-Kakihara.png (2.4 MB, 1024x1024)
2.4 MB PNG
>>109837692
>>
File: The Greased.jpg (1.05 MB, 1920x1088)
1.05 MB JPG
>>109837696
>>
File: Kakihara van Gogh.jpg (1.12 MB, 1920x1088)
1.12 MB JPG
>>109837700
>>
>>109837600
saw it after checking the link test for that
>neat
a bit more than neat, it is quite good has proper trance in it

i never used stable audio 3
you say it is better?
but there is no lora training for it is it not?
>>
>>109837725
>test for that
***thanks for that
>>
File: ComfyUI_Krea_2_00453_.jpg (2.93 MB, 2048x2048)
2.93 MB JPG
>>109837618
SA3 Medium Base can consistently create quality instrumental music that sounds like it came straight from a DAW, and this is without a LoRA or anything.

SA3 medium base trance
https://files.catbox.moe/2ivesc.flac

Its raw music terminology understanding also far surpasses other models, E.G. it knows what is meant by kick, sweeps, snare, claps, etc...

Here's two attempts are a hardcore/gabber and chiptune mix that is not easily possible with YuE 2. (I'm guessing it would need LoRAs, but good luck getting that dynamic range of sound). Second one purposely is more chiptune leaning that the first. This is what you can get only out of pure non-synthetic and non-slopped dataset.

https://files.catbox.moe/ob36vi.flac
https://files.catbox.moe/3iafum.flac
>>
>using h3 f2lva or whatever the fuck you call it
>dialogue sounds super robotic and stilted
>slap this one random lora in, even if i don't have the action the lora is meant to improve
>dialogue sounds way better
wtf??? how am i normally supposed to get good sounding dialogue if this model cant use references or whatever?
>>
You're welcome, /ldg/.
>>
>>109837725
>but there is no lora training for it is it not

There is LoRA training for SA3 (Stable Audio 3 Medium), the one missing it is Minimax Music 3.
>>
>>109837748
have you tried a soviet march yet?
>>
>>109837748
Are you the anon who did the "bimbocore" song about a loli named Lily?
>>
>>109837748
>quality instrumental music that sounds like it came straight from a DAW
Hard disagree. Maybe I know a little bit more about music production than the average person but I can hear the artifacts immediately. It's also very bland and samey. I don't see the point in using AI for QUALITY music. It's not there yet. Only the second track was hard to clock, and that's just because of the genre it was imitating.
Maybe it can be masked with filters, but I don't know. I just wouldn't use AI for music until it gets way better. Maybe you can try having a LLM interface with a DAW?
>>
File: image.png (2 KB, 227x42)
2 KB PNG
>>109836417
cool story
>>
>>109837812
I already prompted that millions of times for you, do you not have a GPU? Kek
>>109837853
No, never heard of that song
>>
>>109836784
>>109836793
thanks!
>>
>>109837888
>I already prompted that millions of times for you, do you not have a GPU? Kek
yeah but i haven't heard it with the new music model yet. still haven't decided on which one i want to download :^)
>>
>>109837886
huh?
>>
>>109837769
ref2va can
>>
>>109837928
yeah and i heard that model is liquid shit and performs way worse when it comes to motions and visual clarity
>>
>>109837880
>Maybe I know a little bit more about music production than the average person but I can hear the artifacts immediately
If there are artifacts, it's very minor and can be fixed in post-processing.

>It's also very bland and samey
99%+ modern music is. SA3 medium composition wise is not the best, but I don't think you're accurately gauging the quality of these models, you clearly have a bias.
>>
>>109837965
but have you tested it?
>>
>>109837973
no...
>>
>>109837976
and I've never tried fl2va either
>>
>use H3 R2V with video ref.
>gen time more than doubles
This better be worth it, will see in 30 min.
>>
>>109837888
>No, never heard of that song
You should:
>https://vocaroo.com/1hqrk3epf4lN
Also, checked.
>>
>>109837969
Also the gens people post to /dmp/ aren't that much better than what I shared quality wise. Found a similar gabber gen on old thread
https://desuarchive.org/g/thread/88983681/#q88990011

Getting a good master and composing quality music is hard. I've trained LoRAs on ACEStep that may have weaker vocals, but their instruments are objectively stronger than the source dataset.
>>
>using anima
>try out krea
>barely any slower gens
uh, the fuck? Why is anima so slow its almost as slow as a model like 6 times bigger?
>>
>>109838073
turn off your turbo slop?
>>
>>109838107
nah, no turbo man. Aesthetic, no turbo lora either.
Same for krea. Is there something wrong with my setup or is that right? Cause thats kinda crazy. I might ditch anima completely for krea2 if that is what i should expect.
>>
what's the best consumer grade GPU neocloud in terms of pricing and service? just to boot up a few containers on nvidiot instances and letting AI agents loose on them (presumably via ssh) for a couple of minutes/hours. I figured you guys would know, considering VRAM is the only viable option for diffusion gen while VRAM requirements are far below text based sota opensource llms (correct me if I'm wrong). Also idc about comfys own cloud option since I want to run now confy workflows/containers as well.
>>
>>109838179
>agents
The new normie twitter buzzword.
>>
>>109838031
Drop the resolution of your reference to 480p and gen time will plummet without impacting on quality. The model is smart enough to infer the missing detail.
>>
>>109837680
I like this one
>>
>>109838322
Thanks, anon. I prompted the model to redraw it the style of Picasso’s synthetic cubism.
>>
File: _OC Kakihara suit.png (1.21 MB, 1080x607)
1.21 MB PNG
>>109838333
Forgot the OG pic.
>>
what local AI do you use to help you write image prompts?
>>
>>109836417
whatever you say monkey
>>
File: kents-quest-0009_.webm (3.57 MB, 928x672)
3.57 MB
3.57 MB WEBM
4-generation scratch test in MiniConstruct. Generated from a single prompt.
>>
>>109838724
nice
git push origin/master when
>>
I think the end game of all this AI slop shit will be that we'll prompt conscious experiences right into our brains.
>>
bros Im using yue to make some meme covers, but audio sounds a bit muddy. are there some nodes that I can use that help with fixing some of the glaring issues? I was using web audio mastering, but I'd rather have everything comfy
>>
>>109839014
More steps
>>
In H3, is 20 steps w/ Spectrum usually better quality than 8 steps with turbo?
>>
The MiniConstruct 0.4.0 version on github is already very capable.
>>
File: 1789151731395.jpg (29 KB, 800x514)
29 KB JPG
is Flux 3 fast version so much cheaper because it's significantly smaller as h3? can we make that assumption?
>>
>>109838031
In my experience with video ref the model often laughs at your face by spitting out your ref as the output.
>>
>>109839195
flux 3 is out for api?
>>
>>109839195
>flux
Who gives a fuck
>>
File: 00003-674959565.jpg (378 KB, 1728x2880)
378 KB JPG
>>109838073
krea2 is superior model and has a better future than anima.
>>
File: revelation_3_11.jpg (264 KB, 1080x720)
264 KB JPG
>>109838947
>I think the end game of all this AI slop shit will be that we'll prompt conscious experiences right into our brains.
>>
>>109839195
>Local Diffusion General thread
now kys
>>
>>109839304
Unfortunately not - the areas of the brain that would be important for this operate at the speed of 56k modems.
That’s why there will never be full-body-dive VR devices like the ones in Sword Art Online. The technology might be capable, but our brains aren’t.
>I love shattering dreams
>>
>>109839292
>krea2
>bunch of shitty overfit loras
>exactly one (1) finetune, that is catastrophically forgetting the base model's knowledge more than any finetune of anima ever did
>"better future"
lol, lmao even
>>
How can these tools detect if an image is AI?
I'm using this to analyze images and for the moment it scores 100% of my images.

I could obfuscate, destroy, add noise to the images but I don't want to destroy the quality of the images. But I swear they are not detectable to the human eye.

So either it is detecting the dimensions of the image, which is always a multiple of 64 or there is some pattern on the colors.

How do I bypass this? I really don't know what strategy is this using?
https://www.zerogpt.com/ai-image-detector
>>
>>109839338
That's literal nonsense.
>>
>>109839403
take a cropped screenshot of one of the images and see if it still detects. would be an easy way to test for dimensions or regular metadata. or just resize the image in some way
>>
>>109839190
Thanks for the info, trani 2
I shan't be touching that shit with a 10 foot pole
>>
>>109839014
YuE 2 on Comfy is 10x slower than audio.cpp. For mastering, just use Matchering 2 as opposed to Web Audio Mastering.
>>
>>109839802
35 stars status?
>>
>>109839802
I wish I were exaggerating, but Comfy is really slow and not good at all for audio/music gen tasks. It "supporting" the models is false advertisement.
>>
>>109839802
It takes me 40 seconds to generate a 2 minute song with sheet music and cover input using the bf16 model in comfyui and you claim it would take 4 seconds on audio cpp? Can we stop with the low effort shit posting already, we know it's you carl johnson, doing a false flag
>>
>>109839823
A node like https://github.com/audiohacking/acestep-cpp-comfyui
Would make it more tolerable but only if it's built for audio.cpp's backend.
>>
>>109839858
>It takes me 40 seconds to generate a 2 minute song with sheet music and cover input using the bf16 model

You're on a 5090.
>>
man prompting for Z-Image-Turbo seems more annoying. Specially trying to get randomized angles for the image. When not specified it almost always goes for a front view centered composition aaaaaaaaa just do it randomly aaaaaaaaaaa
>>
>>109838073
>>109838116
On my end, Krea 2 Turbo with 8 steps is maybe 10% faster than Anima nonturbo with 20 steps. (~270s versus ~300s for 1600x1280 on an iGPU)
>>
>>109839868
Anyways, I'm not necessarily shilling for audio.cpp. The UI is clunky, but with YuE2.cpp you get similar speed gains (though im case of YuE2.cpp it doesn't yet support covers or LoRAs).
>>
>>109839868
Can you actually post benchmarks instead of wasting our time with this?
>>
>>109839937
Anon I'm telling you my experience. Maybe you're new to music models so you're skeptical, but what I said is pretty much fact. On a 3090, it's unbearable speeds for gens on Comfy, I could not wait that long even with int8 convrot, but on .cpp UIs q8 GGUF it's a breeze, my GPU also clearly works less whereas it's struggling with ComfyUI (and it's light on RAM too since on 64GB of RAM it's possible to have an image or video model like H3 loaded into Comfy while simultenously having it loaded on the .cpp UI).
>>
>>109839966
Form what I understand, Comfy's implementation is not even native. That's part of the reason why, it's even faster on diffusers.
>>
>>109839966
You're talking to someone else vagueposting schizo retard, but you could tell us simply how long it takes you to gen a song on your 3090 at 32 steps instead of continuing to vaguepost
>>
>>109839966
You should be paying attention to anistudio since Ani already has h3 working. You could probably get him to add audio.cpp so you don't have to use comfy either. No idea what the speeds are like for vidgen though
>>
>>109839990
>diffusers
Official YuE2 pipeline*
Though even that is a like 2-3x slower than .cpp
>>
>>109839966
>3090
Oh gross
>>
>>109839996
Doesn't take long to download the GGUF models, build the .cpp backend and test my claims. It would take me longer to test a single ComfyUI gen.
>>
on a 3050ti its 15min to an hour for a gen in comfy depending on duration
the only time I tried audio.cpp it took four hours
>>
>>109840053
And if you mean on .cpp, it's like a fraction of a sec per step.
>>
>>109840053
Lmao you fucking retard, how long did it take you to come up with that one? We both know I meant how long does it take on your end with whatever .cpp backend you're using. But you did give me a good laugh
>>
>>109839424
I don't have the tool at hand, but if it's still detecting it, then what's the trigger?
I edited some original images with AI, small changes, and the tool still detected it.

So I'm assuming it sees something about the color patterns.

Can you guys check too?
>>
>>109839919
Vibecoded my own backend+UI that supports a full suite (lyric gen+generate+remix using either melody, lyrics, or both)
Relying on third party nodes/apps is peak boomer shit. Just make it yourself.
>>
>>109840122
>I created the same thing someone else has created a long time ago but in worse wasting my precious and finite lifetime
Guess it's true what they say, no matter how many tools you give a retard, he's still going to be a retard.
>>
>>109840064
03:50 min song

>[AR] Score: 3335 tokens over 1 songs, 3336 steps, 16.1 s (4.8 ms/step)
>[AR] Semantic: 5769 tokens over 1 songs, 5770 steps, 32.1 s (5.6 ms/step)
>[Store] Load NAR: 3779 ms
>[VAE] Tiled decode done: 12 tiles -> T_audio=11076416 (230.76s @ 48kHz), 1996 ms
>[Pipeline] Done: 1 tracks, 230.8 s of audio in 114.7 s (2.0x realtime)
>>
>>109840161
That's on yue2.cpp, audio.cpp doesn't print anything to the console
>>
>>109840161
>tiled VAE
They tell you to specifically not use that if you're not vram poor
>>
>>109840161
So it's slower than comfyui, what's the point then?
>>
File: 455452114454.png (120 KB, 1080x583)
120 KB PNG
>>109840161
About 93s for 3 min song on audio.cpp
>>109840234
Lol
>>
>>109840358
>>109840161
both of examples are below the 2.1x realtime speed I'm getting on comfyui, so either you're using pajeet style hardware below a 3090 or it's shit
>>
>>109840358
Actually the speed varies depending on song, got 67.58s for 3:03 song on audio.cpp
>>
>>109840394
Nope, Comfy is just shit
>[INFO] Checkpoint files will always be loaded safely.
>[INFO] Total VRAM 24575 MB, total RAM 65460 MB
>[INFO] pytorch version: 2.11.0+cu128
>[INFO] xformers version: 0.0.35
>[INFO] Set vram state to: NORMAL_VRAM
>[INFO] Device: cuda:0 NVIDIA GeForce RTX 3090 : cudaMallocAsync
>>
File: ThisNiggaSerious.jpg (6 KB, 320x566)
6 KB JPG
>>109840423
What do you mean nope? What exactly are you disagreeing to? I got ~150s of audio in ~69s of genning in comfyui cold start and using bf16 and all sneed options enabled at 32 steps
>>[INFO] pytorch version: 2.11.0+cu128
Why are you on an outdated cuda and pytorch version... don't tell me you did that to run sageattention for h3 like a retard. Are we being deadass rn unc?
>>
>>109840486
>catjack reaction image
I kinda knew deep down the anon that can't math us just the schizo troll
>>
>>109840506
>calls me one of his furry friend names rather than addressing the point why he's on a severely outdated version of cuda and pytorch
You being deadass frfr?
>>
>>109840486
I updated the venv to CUDA 13 but only supporting decent speeds on latest CUDA when so many of us still have legacy versions is trash dev
>>
>>109840540
Also, it's still not necessarily faster on Comfy, but it's hard to measure since gen times aren't consistent
>>
>>109840506
Alright he's retarded and trolling. Lets ignore him now
>>
>>109840575
>>109840540
>It's the dev's fault I'm using outdated software
Meds?
>>
>>109840423
>>109840540
Performance slowdown is secretly sponsored by Nvidia. It's to give users the illusion that they need a new GPU.
>>
>>109840575
The only claim you've made so far that isn't borderline retarded, but it still kind of is because we can compare tokens/s and it/s to see if there is indeed a 10x speedup as our schizophrenic furry in here claimed and if there isn't well I guess that settles it and it was all just another schizo episode from our beloved
>>
>>109840661
Organically, there wouldn't be a larger than 5-10% difference between CUDA versions.
>>
>>109840709
The retard made a claim when his testing environment was not correct. It seems we still have /sdg/ natives proving why they belong in the special needs section and not here.
>>
>>109840723
And I'm telling you that the symptom itself should not even exist. There's only one explanation for it.
>>
>>109840756
You have multiple anons calling you a retard and your response is to claim they are a namefag.
You have to go back
>>
>>109840146
>oh you made a fizzy drink? well, someone else already made a fizzy drink before you
>if you don't like the taste just like it
>>
>>109840769
>>109840704
>It's my fault there's a literal 10x difference between CUDA versions on Ampere hardware
t. Paid Comfy shills
>>
>>109840805
He's just looking to argue because he's salty over the thread. Ignore him, probably vram poor and can't run local models.
>>
>>109840817
Usecase for running outdated software in a rapidly evolving space?
>>
>>109840817
Maybe if there were a glaring warning not to use this shit, or the UI refused to boot, it would be acceptable, but there isn't.
>>
>>109840835
This. We need to just move away from python entirely. Are there any frontends out there that don't use python?
>>
File: 2340230827309.jpg (50 KB, 959x199)
50 KB JPG
>>109840805
Nice false analogy, it would actually be closer to this
>Oh you made a printer? Is it faster or does it print in higher quality? It does the exact same thing but It's slower and often has problems?? Oh, ok... cool I guess?
>>
>>109840831
I'm not deleting old venvs that I had before H3, simply because they also have a million dependencies for each custom extension installed on it
>>
>>109840893
>Unable to update his currently existing venv
I seriously think you belong in /sdg/ if you can't do such a basic task.
>>
>>109840867
if you're trying to assert intellect, posting evidence of yourself offloading cognitive work to a computer is counterproductive
>>
>>109840914
>Heh I got him with this one
>pushes glasses up and rubs fat belly
>time to post on leddit for some updoots
Listen if 4d humor is too advanced for you just say so, your posts are giving me the ick
>>
>>109840867
It looks better to me and it's easier to use for me. That's why people make their own things.
>>
>>109840899
No, what I'm telling you is that updating shit is how shit breaks, a tale as old as time itself with python.
>>
>>109840867
Don't use AI to think for you. Especially not Google AI.
>>
>>109840938
>Skill issue
>blaming others
Again I think /sdg/ is more your speed
>>
>>109840122
I have a vibecoded UI with everything working, can even run multiple prompt/lyric pairs at once, but it's not the fastest and was made for WSL. I'll look into vibecoding something with .cpp instead of python
>>
>>109840935
Hey if you like your ((printer)) in pink and with a bunch of hardware bugs that's cool too, but once you start valuing your time you simply go for the best product and stop wasting your time with useless shit like this.
Also calling it creating is a stretch, you simply let ai code some slop for you and decided what you like and what not and under the hood, under that pink shell case of your printer? It does the exact same job but worse if you try not to be biased for a second. Cause you didn't do anything new or create something, it's the same garbage running a .cpp or a py script but with a different flavor ui
>>
does anyone have a krea2 workflow that uses the json prompt builder node? I am having issues where it is mirroring my prompt, left is right etc.
>>
>>109841036
Nta, but my vibecoded UI contains a .yaml section where I can paste in a list of prompt and lyric pairs (with their own parameters) that run in batches. It's a custom feature I made that useful for A/B testing or just to find decent prompts when testing batches of them. So that makes my vibecoded UI better than most publicly available frontends
>>
>>109841112
>he names a feature comfyui can do out of the box
I dunno anon, I'm simply not impressed unless I'm missing something. Being better than most is not hard if the competition is also vibecoded garbage.
Once you got something cool that comfyui doesn't have we can talk again. Cause this convo reminds me of the same people who years ago claimed forge is better because it's less complicated and easier to use because connecting nodes was all it took to filter them
>>
>>109841161
Ever think he's just doing it for himself?
Why are you being such a annoying faggot?
>>
>>109841179
The entire argument was it's a waste of time and the product already exists, not whether he has fun vibecoding stuff.
If he argued for that then I wouldn't even have replied you retarded low iq faggot.
>>
>>109841161
Wouldn't that need a custom node for YuE 2?
>>
>>109841215
Not really, you can simply connect multiple ksampler nodes to the conditioning and then generate as many different audio files with different settings using the same prompt as needed. That's of course just one of probably thousand solutions to do that and there are of course custom nodes that can probably do that in a more visually pleasing way too
>>
>>109841247
Yeah but that is not as easy as just copy and pasting once, and it would be faster to copy and paste each prompt/lyric pair and modify individually than creating extra nodes. The .yaml solution is useful for when an LLM creates the prompt and lyrics you.
>>
>try to recreate in comfyui pictures I generated on civitai
>use the same model, prompt, loras and their strength
>artstyle looks slightly different
this is driving me insane
>>
File: point.jpg (98 KB, 814x1358)
98 KB JPG
>>109841294
>Yeah but that is not as easy as just copy and pasting once
Yeah it's actually easier, I copy my current workflow then paste it with ctrl+shift+v connected to the lyrics nodes and it makes an exact copy of it where I simply change whatever settings I dislike. It literally takes me 5 seconds to create a workflow like that and I can have infinite amounts of changes. And like I said if you want it to be more visually pleasing, custom nodes already exist for that.
And just to show I'm not just talking out of my ass, I modified my workflow in literally 5 seconds to now be a able to produce 7 distinct audio files with different settings but the same lyrics
>>
File: 455454123154.png (97 KB, 550x939)
97 KB PNG
>>109841327
Yes but you have to go into each individual box and edit the prompt and lyrics. I just do ctrl c + ctrl v from what the LLM gives me and I'm done, can modify all exposed parameters. Here's what that sort of looks like
>>
File: NoMoreBaitinginLA.jpg (315 KB, 2337x988)
315 KB JPG
>>109841395
At this point I have to assume you simply don't know how to read node based workflows, so I've attached a zoomed in workflow that even you would understand unless you're special needs
>>
>>109841474
That would just generate the same prompt and lyrics for each song unless you duplicate those boxes too.
>>
File: Return_00591_.jpg (1.69 MB, 2368x1776)
1.69 MB JPG
>>
File: 219870536486712359078.jpg (6 KB, 302x134)
6 KB JPG
>>109841591
>That would just generate the same prompt and lyrics for each song unless you duplicate those boxes too
Wasn't that the entire point you brown ESL? Do you even understand what you post? Just vibecode your thoughts to an LLM before posting cause this is mad retarded and you're either shifting the goalpost or you're genuinely special needs
>>109841618
>I was most likely arguing with a schizophrenic furry
Oh ok now it makes sense
>>
>>109841618
you already posted this one
>>
>>109841669
>>109841666
Having a bad day today?
>>
oh god here comes the melty
>>
File: 071422-2.mp4 (2.79 MB, 1978x1236)
2.79 MB
2.79 MB MP4
from video to blender
>>
>>109842048
What's the process?
>>
>>109842048
What does the topology look like?
>>
File: MiniMax_H3__00739.mp4 (3.06 MB, 736x736)
3.06 MB
3.06 MB MP4
this keeps the israelis up at night
>>
>>109842070
>>109842129
Anon it's api shit, if it's even real. We can barely make low quality 3d models on 20-100gb of vram, this is not a thing yet
>>
>>109842070
>generate an exact 360 camera spin with a character standing still.
>extract clip' frames
>send it to colmap to create a 3D Reconstruction
>send colmap results to Brush to train 3dgs model
>>109842129
a bit blurry cuz i only do 360 horizontal spin. adding vertical shots probably would help
>>
>>109840085
i tried it on one of my images and it got classified as a "digitally edited" image. i don't know if that means it think that it is AI
>>
>>109842129
I was paid to say how does the topology looks like.
>>
>>109841666
No, my point is that I generate batches of songs with different captions and lyrics from pasting into a single box, don't need to spend time creating extra nodes nor copy and pasting each individual song. I can just vibecode a solution into Comfy, but that's just as valid as vibecoding my own UI.
>>
>>109842257
you're a cheap whore
>>
>>
>>109842275
i look like this and say this
>>
>>109842271
I am a scientisct.
>>
>>109842268
>My point is I'm a brown shitposter who's not actually making a point and I think pretending to be retarded is the funniest shit ever to get a (you)
No, I really understood that part don't worry. We good, comfyui is quite advanced which is why we have things like chatgpt and claude for those who are intimidated by nodes and shiet
>>
File: MiniMax_H3__00744.mp4 (3.26 MB, 736x736)
3.26 MB
3.26 MB MP4
>>
>>109842366
Are you vram poor and are upset by anons that want to make a workflow that they find appealing?
Before you reply to me post your rig.
>>
anyone ever encountered the phenomenon of people getting angry at AI images (especially 1girls)? not just in forums but also like in real life?
like I post a qt AI girl and multiple anons will chimp out and get angry at the picture.
any idea what this means?
>>
>>109842473
are you implying that 4chan is real life?
>>
>>109842488
Post your rig or sit in the corner.
>>
>>109842488
>Probably means you posted slop, cause when I post 1girls everyone is ecstatic
the reactions I got seem to be mixed, some people seem to enjoy it that its a good image but then it also draws in the haters for some reason who start to seethe about it and getting really angry about it.
>>109842490
wasnt my intention sorry.
but like have you ever shown someone and AI picture and they got really mad at that?
>>
Any local that can create video to music just like elevenlabs does?
Like you add your video and it generate sounds/audio based on the context of the video.
>>
>>109842498
hands first or stop posting
>>
>>109842509
Knew it, you're seething at the anon because he's not vram poor unlike you and has the resources to make his own front end for his convenience.
>>
>>109842504
Anon there is always some tasteless loser who just asks for more when slop is being posted, especially if it's sexually suggestive, so getting 70% approval is the norm, you need to up those rookie numbers to 95% or higher to compete in the big leagues
>>109842525
>brownoid still posting as if his ESL opinion has any weight
kek'd
>>
>>109842504
>but like have you ever shown someone and AI picture and they got really mad at that?
no, but i have only shown AI images to maybe 3 people in my life. only one of them knew it was AI since i ran his drawing through a realism conversion
>>
>>109842545
Having a cry?
You're just seething in this thread, it looks like you have some other issues going on buddy.
>>
File: 1773663909827013.jpg (213 KB, 1080x1920)
213 KB JPG
>>109842048
>>109842180
I'm literally trying to modelize a real girl into a 3D game, so I would like more info about this, I tried many solutions but none was successful (tripo3D, chatgpt image, seedance 2.5 video etc...)
>generate an exact 360 camera spin with a character standing still
What did you use for that? How many sources? front, left, right back images or only front?
>send colmap results to Brush to train 3dgs model
ZBrush? 3DGS model? training? wut?
>>
>>109842275
prompt? very cool and realistic
>>
Why does schizo anon think calling someone brown is an own when asked to post specs?
>>
(You) are a master at baiting.
>>
Why is the schizoâ„¢ recognized by /ldg/, /sdg/, /adt/ and /lmg/ in the delusion that calling someone vramlet when getting called out for his brown behavior a valid defense?
>>
The links stay BTW
>>
Yeah but we should add one for you as well
>>
>nigbo
>>
>Carl Johnson
>>
What makes them seethe so hard?
>>
>>109842747
You're our guy at the source, why not tell us
>>
>>109839675
But you're a severely mentally ill schizo.
>>
File: MiniMax_H3_00816_.mp4 (1.08 MB, 672x1184)
1.08 MB
1.08 MB MP4
>>109842604

- use mini max h3 ref2va with your character as reference. Prompt properly, full body shot, character must stand exactly still, no movement at all. The camera spins 360 from left to right with smooth movement, meaning no camera blur-ness and shit. It must be a full 360 spin, so your first frame is your last frame. Preferably, you gen with high resolution or upscale the clip because you need quality dataset.
- Extract your clip's frames with ffmpeg or whatever. How long the clip should be? I don't know. Probably 3-10 seconds, depend on how many "good" frames you can create with it.
- Then use ColMAP https://github.com/colmap/colmap
Processing > Feature Extraction > SIMPLE_PINHOLE as camera mode >
Feature Matching > Run
Start recontruction
then export model when it's done

- Then use Brush, and load all the bins file you created with ColMAP
https://github.com/ArthurBrussee/brush
It will create a 3d Gaussian splatting model after training is done.

- how to create the actual mesh model, bone rig, etc.. from this 3dsg? It's another task.

This is the gist of it. Ask LLM for configurations.
>>
File: 1775248890154027.webm (424 KB, 720x405)
424 KB
424 KB WEBM
>>109842814
Ok, thanks for all the info anon.

My issue is that it's from a real human and with photos from low resolution and different weird angle, that's why I struggled so much to get consistent results with multiple angles picture generated separately. so I guess the most important part should be with mini max h3.

I don't have anything local installed yet (was only using cloud), so I will try that soon.

You did make a 3D model from it right? If so if you could also share the last part (3dsg -> blender) in case there is some important knowledge to have, I would be very grateful. (I will obliviously make some research on it too)
>>
Gauss was not some bug on you windsheild.
>>
>schizo swept
Lol
>>
File: download.jpg (46 KB, 600x603)
46 KB JPG
are you behaving yourselves?
>>
>>109842940
He had a meltdown and got put to bed, I think he had too much to drink again
>>
>>109842398
dont fucking scream at me
>>
>>109842978
>singular post was deleted
>it was the one that could potentially harm the schizo anon if he got egged on hard enough and did it since he isn't the brightest shed in the tool and prone to harming himself
anon... this is not the victory you think it is
>>
>>109842129
it looks like a gaussian splat so no topology at all
>>
>>109841680
8 young yin, on d12 dice it is 6 through 9.
>>
>>109843065
who needs topology, some kind of topology PERVERT?
>>
>>109843053
Behave yourself and stop seething
>>
File: x_q2bn1q.png (486 KB, 1024x512)
486 KB PNG
>>
>>109843122
I would've continued bullying you if you weren't now officially recognized as a protected victim class by jannies tbqhfam, so instead I'm giving out an honest prayer for your quick mental recovery, amen
>>
File: ed5.png (131 KB, 680x1112)
131 KB PNG
>I would've continued bullying you if you weren't now officially recognized as a protected victim class by jannies tbqhfam, so instead I'm giving out an honest prayer for your quick mental recovery, amen
>>
>>109843134
I like it. model, and prompting style?

I hope it's Krea. So far, no reason to install Krea!
>>
File: wojak.jpg (77 KB, 664x793)
77 KB JPG
What is the ceiling that a 4080 and 16GB RAM can reach? I haven't done any AI related generations outside of GPT for a couple years and I want to get back into it
>>
>>109843210
idk
>>
>>109843210
you can use anima to generate images, video is probably too much for you
>>
>>109843210
You can basically use all image generation models with decent speeds but will struggle with video creation using both H3 or LTX
>>
>>109843156
Buckbroken
>>
>>109843188
using sdcpp, prompt:
lifeless, loneliness, spiritual beauty, vast infinity,
space art by Homer Winslow, in style of Ansel Adams,
starry gradient dark blue starry night sky,
by Roger Dean,
flat horzion in distance , foreground is a bright reflective silver tuburlen ocean,
in the skay there is an orange exoplanet that looks like Io, Ganymede, or Enceladus with cracked oceans and green landmasses,
by Daido Moriyama,
by Chesley Bonestell
view on a distant exoplanet in a strange alternate universe,

Steps: 20, CFG scale: 1.100000, Guidance: 3.500000, Eta: inf, Seed: 27906, Size: 1024x512, Model: , RNG: cuda, Sampler: euler discrete, TE: clip_l.safetensors, TE: t5xxl_fp16.safetensors, Unet: flux1-dev-q8_0.gguf, VAE: ae.safetensors, Version: stable-diffusion.cpp, SDCPP: {"auto_resize_ref_image":true,"clip_skip":-1,"control_strength":0.8999999761581421,"generator":{"commit":"8caa3f9","name":"stable-diffusion.cpp","version":"unknown"},"height":512,"increase_ref_index":false,"mode":"img_gen","models":{"clip_l":"clip_l.safetensors","diffusion_model":"flux1-dev-q8_0.gguf","t5xxl":"t5xxl_fp16.safetensors","vae":"ae.safetensors"}, ...
>>
Why is catjack actually doing this shit? This is just lolcow behavior and he isn't going to convince anyone otherwise. His disses remind me of a cyrax spergout
>>
>>109843210
>16GB RAM
oh nonono
>>
>>109843254
>sdcpp
wow. I need to try the latest version.
>>
>>109843210
Apparently some people have made 16gb work. 2 years ago, I failed to get comfyui to work on 16, put in 32, still sluggish, 64 freed it up, so maybe like 48 is what is needed to have it run nicely? But now, again, some have at least made it work of 16.

sdcpp might be able to help here somehow? The biggest problem with comfyui is it's a moving target, every pull is like rolling the dice.
>>
so what's the catch with fasth3 v2?
I used it and it required 8 steps? you can do it with 6 steps in base already
>>
>>109843266
>sdcpp
The win is flux1-dev-q8_0.gguf which has vague notions of various artists. Modern models have this knowledge scrubbed.
>>
>>109839966
>>109840358
3090 - you should be getting 60-70 seconds for that length
slowdown is due to dynamic vram and few other optimizations.
>>109840423
and you can test that by pulling older versions of comfy before 'optimizations' kicked in.
>>
>>109839106
Anyone?
Also do ppl generally prefer euler to res multistep?
>>
>>109843342
Euler is best for non turbo gens
>>
>>109839106
Yes.

>>109843342
I get near perfect results with res_multistep so I never bothered to even try other samplers.
>>
>>109843342
euler for turbo.
For me, h3 looks like shit even with 20 steps so I never bother.
base has better audio, motion and prompt adherence so, use if you prefer these
>>
>>109843342
Euler is your backup plan for when your wf is busted and you are debugging it.
>>
>>109843349
>>109843371
Who is right and who is wrong?
>>
>>109843342
You're asking questions that can be answered by running the model yourself for less than 20 minutes, 40 tops if you're on outdated hardware while giving us the people you ask for questions barely enough info to give an answer.
But 20 steps spectrum beats 8 steps turbo lora every time just because of the massive audio difference, but it also takes almost 50% longer

>>109843380
>>109843349
this guy is a retard and nobody should use euler for non turbo
>>
>>109843382
KJ says euler sde and SA solver are best for turbo
>>
sd1.6 waiting room.
>>
>>109843210
You can do plenty on 16 gigs, Krea2, minimaxh3 40 seconds at 4mp with a turbo, ltx2.5 sdxl, flux-2.klein
>>
>>109843419
why did memory reqs go down so much?
>>
>>109843424
better vae technology
>>
>>109843424
lots of cheap 16gb cards in the wild.
>>
>>109843419
>minimaxh3 40 seconds at 4mp with a turbo
Lmfao no
>>
>>109843438
I think he means 0.4
>>
what speeds are you getting with h3 turbo
>>
>>109843453
3 minutes to generate a 20 second 320p video at 8 steps
>>
>>109843494
and what is your card
>>
>>109843511
3080ti
>>
>>109843438
yeah i met 0.4
>>
trying krea2 centersemiraw and it really doesnt look any better than raw + turbo diff lora, much more plasticity and "ai-like". Anyone else wrangled that model and managed it to look decent?
>>
Waiting on Qasar. Kind of excited.
>>
Audio remains the greatest challenge.
>>
File: diana_burnwood.jpg (124 KB, 1446x1080)
124 KB JPG
>>109843743
>Paying for voice AI when Breeze TTS exists

https://vocaroo.com/115CRD7NCh1v
>>
>>109843883
>Breeze TTS
Can this make a model for a text to speech model for a phone/llm?
>>
>>109836320
upper left one is genuinely really nice
>>
File: why-the-long-face.jpg (483 KB, 1664x1280)
483 KB JPG
why the long face, anon?
>>
File: ComfyUI_1088.png (2.4 MB, 1664x1280)
2.4 MB PNG
>>
>>109843210
My exact setup. Honestly, I was able to do Minimax videos. 5sec1mp was like 15 or so minutes.
>>
File: ash-rockem.jpg (351 KB, 1664x1280)
351 KB JPG
>>
>>109843983
my 4080 can do 1mp in like 90s, embrace the speedup loras
>>
File: 00057-1567967339re.png (2.83 MB, 1280x1280)
2.83 MB PNG
What did you draw today?
>>
File: china-girl.jpg (329 KB, 1664x1280)
329 KB JPG
https://song.link/i/1475002922
>>
>>109844003
around 500 hanzi characters, if that counts
>>
File: boss-pong.jpg (524 KB, 1664x1280)
524 KB JPG
>>
the kinoplexatorium might be open soon, according to sources familiar with the matter
>>
>>109844244
i am preparing my seat
>>
>>109844057
can you do Jeremy Clarkson in a space station?
>>
>>109844244
what's that? I tried googling and came up with little comprehensible.
>>
>>109844350
I will rape you, newfren.
>>
>>109843963
Ani made some cute gens
>>
>>109844350
it's at the kinoplaza
>>
File: 1787788641191913.jpg (14 KB, 376x369)
14 KB JPG
h3 content is impressive. i tried some prompts from seedance, and the results are really close. it might be even more faithful, if i removed the nsfw loras kek
>>
>>109844459
>nsfw loras
where do you get them for h3? I can't find any
>>
>>109844472
They're all over cvitai red
>>
File: clark-station.jpg (1.7 MB, 3136x2368)
1.7 MB JPG
>>109844342
>>
File: ComfyUI_1122.png (1.09 MB, 1152x896)
1.09 MB PNG
>>109844565
>>
>>109844379
where is her dick?
>>
i cant upgrade my ram cause all my ram slots are already taken
>>
What's next on the horizon after Krea2?
>>
>>109844664
salvation from this mortal coil mhm
>>
>>109844590
>it works
He would manage to pretend to be disappointed on a space station.
>>
>>109844671
be serious for 72.3 seconds
>>
>>109844592
In catjack's buckbroken bum
>>
where did nunchaku nigga go?
I remember he made best speed up slop
>>
File: Kyou_2.webm (3.48 MB, 928x672)
3.48 MB
3.48 MB WEBM
>>
Are live taesd previews currently broken? Is there no way to see the live sampling preview of H3 gens without downgrading comfyUI?
>>
Why not make a new one instead of reposting the old one ?
>>
>>
File: MiniMax_H3_00576_.webm (1.67 MB, 672x960)
1.67 MB
1.67 MB WEBM
>>
erm... anon?
>>
>>109844944
yes, anon?
>>
>>109844964
>>109844964
>>109844964
>>
So are we just going to use the shitty debo thread? At least he included the rentry links this time.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.