[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Weekend Edition

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109730780

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: Krea2_turbo_02068_.jpg (1.46 MB, 1776x2368)
1.46 MB JPG
>>109736456
>>
>>
as always, nothing happens.
a week just like any other.
>>
File: yt_mcYl70vq_Ns.webm (3.95 MB, 1920x1080)
3.95 MB
3.95 MB WEBM
how do you feel about being in pic related, anon?
>>
>>109736499
hell yeah
>-AR 1:2
GET THAT CLOUD SHIT OUT OF HERE
>>
>>109736499
this is just like me frfr!
>>
>>109736499
Very funny. Raffed hard.
>>
What should I prompt?
>>
>>109736499
>>109736554
kinos
>>
>>109736559
You should make kinos instead of being the corner cuck. I need you to lock in anon
>>
File: copium.webm (2.95 MB, 716x720)
2.95 MB
2.95 MB WEBM
>>
>>109736565
but I'm currently working on figuring out the best possible compromise between speed and videoquality anon, my gpu is busy
>>
File: naisv5_32_08.jpg (1.57 MB, 1344x2368)
1.57 MB JPG
>>
>>109736579
No excuses, one does not have the mandate to call for kino without making kino.
>>
>>109736588
no yuo kino!
>>
File: Return_00028_.jpg (1.42 MB, 1776x2368)
1.42 MB JPG
>>109736585
Why do you post this when anons regularly post local gens that clear you?
>>
>>109736601
Grandiose delusional thoughts.
>>
>>109736601
>a strong workflow more than makes up for his shortcomings!
>>
>>109736601
He does it to piss you off
>>
>>109736449
>mfw Resource news

09/05/2026

>Intern Lumina U2: Multi-Codebook Diffusion Large Language Model for Omni-Visual Understanding and Image Generation
https://internlm.github.io/InternLumina-U2

>Musk’s xAI loses court bid to block Minnesota's AI ‘nudification’ ban
https://www.reuters.com/legal/litigation/musks-xai-loses-court-bid-block-minnesotas-ai-nudification-ban-2026-09-04

>Add Sparse Attention node- #16072
https://github.com/Comfy-Org/ComfyUI/pull/16072

>MiniMax-H3 FL2V 8-Step Motion Enhancer
https://huggingface.co/rzgar/minimax-h3_fl2v_8Step_motion_enhancer

>Qwen3.8-Flash-Next-NVFP4
https://huggingface.co/nvidia/Qwen3.8-Flash-Next-NVFP4

09/04/2026

>lightx2v/Minimax-h3-Turbo · FL2V Turbo 4-step v1.2 (768p)
https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/52#6a9a890895a616c64799324f

>ComfyUI NVIDIA DLSS 5 Visual Enhancer
https://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-Interpolation

>Viggle-Animate: Character Replacement in Video from a Single Repainted Frame
https://huggingface.co/Viggle/Viggle-Animate

>DSAQuant: Denoising-Stage-Aligned Quantization-Aware Training for Video Generation
https://robbyant-research.github.io/DSAQuant

>Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuation
https://github.com/AMAP-ML/StateAgent

>FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlow
https://byeongjun-park.github.io/FlashRender

>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes
https://huggingface.co/inclusionAI/LLaDA-Image

>ComfyUI-VDN-H3: v1.4.0 — Faster streaming, VRAM-aware buffer retention Latest
https://github.com/Saganaki22/ComfyUI-VDN-H3/releases/tag/v1.4.0

>AetherScale for ComfyUI: GPU-native NVIDIA video enhancement
https://github.com/vizart-vj/ComfyUI-AetherScale
>>
>mfw Research news

09/05/2026

>Generalization over Memorization: Generalization-Aware Diffusion Adaptation for Single-Image Multi-View Synthesis
https://arxiv.org/abs/2608.29233

>Test-Time Scaling for Video Diffusion Models via Diagnosis-Guided Candidate Recycling
https://arxiv.org/abs/2608.29322

>Streaming4D: Accelerate 4D World Models via Block-wise Video Generation and Incremental Reconstruction
https://arxiv.org/abs/2609.00610

>Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics
https://arxiv.org/abs/2609.02268

>Visual Framing for News Stance Detection via Image Generation
https://arxiv.org/abs/2609.00685

>Controllable Image Captioning with Prompt-Conditioned Scene Rewards
https://focus-emnlp2026.github.io

>TimeSteer: Inference-Time Speech Scheduling in Joint Audio-Visual Diffusion Models
https://arxiv.org/abs/2609.01277

>Diffusion Based Unpaired Data Learning for Inverse Problems
https://arxiv.org/abs/2609.01370

>ExpArt-KG: Artwork Image Description Generation through Iterative Exploration of Knowledge Graphs
https://arxiv.org/abs/2609.00629

>ViTAL-X: Video-Text Alignment with Cross-Modal Temporal Edits
https://arxiv.org/abs/2609.00505

>ASSERT: Adaptive Stochastic Sampling for Robust Diffusion Models on Analog Compute-in-Memory Hardware
https://arxiv.org/abs/2609.00955

>From Detection to Localization: A Unified Forensics Framework for Fully Synthetic and Tampered Images
https://arxiv.org/abs/2609.02640

>Sketch2Inspire: Structure-Sensitive Evaluation for Product Retrieval
https://arxiv.org/abs/2608.29364

>A Lagrangian View of Flow Matching
https://arxiv.org/abs/2609.00198

>Beyond Language Priors: Diagnosing and Fixing Visual-Origin Hallucinations in Multimodal LLM
https://arxiv.org/abs/2609.00231

>SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models
https://arxiv.org/abs/2608.29974
>>
>>109736647
>>109736651
thank you for the news!
>>
>>109736647
>>109736651
Fuck off
>>
File: debo_cl_k2_00011_.png (2.47 MB, 1664x1069)
2.47 MB PNG
>>109736670
np :)
>>
is the latest 4 step lightx2v minimax or the older 8 step turbo lora better?
>>
>>109736795
Is 20 steps raw or 8 step lora better?
>>
>>109736499
absolute kino
>>
File: jewing.jpg (49 KB, 263x493)
49 KB JPG
consume more energy that it would take to sustain all of humanity for decades to run his >datacenter just to analyze every mathematically possible solution to some obscure math riddle from 200 years ago that nobody cares about just to get more of your tax money from the government
that's alignment all right
>>
File: file.jpg (1.32 MB, 2160x2880)
1.32 MB JPG
>>109736785
looks kinda like my cat, picrel (but mine is chonkier)
>>
File: debo_cl_k2_00012_.png (2.59 MB, 1664x1069)
2.59 MB PNG
>>109736853
very high quality loaf
cats should be allowed in space
>>
>>109736853
Talk about the wrong stuff...
>>
>>109736850
my computer doesn't use that much. this is the local ai thread and we don't pay anything other than electric bill
>>
File: file.jpg (1.35 MB, 2160x2880)
1.35 MB JPG
>>109736868
>cats should be allowed in space
i can guarantee once space travel and colonization is a thing, the first thing people will bring to other planets is their cats, they've won the evolutionary race by befriending the apex predator

>>109736886
huh?
>>
>>109736647
>>109736651
>>109736670
>>109736785
>>109736853
>>109736899
Real dented zone post rn
>>
File: Predators.mp4 (3.82 MB, 1152x656)
3.82 MB
3.82 MB MP4
If it bleeds..

https://streamable.com/uv0z1a
>>
>>109736637
SD thread only exist as a coping mechanism and pure spite at this point
>>
>>109736913
You do know everyone thinks you are mentally ill because you can't stop acting like a child right?
>>
>>109736943
>he's upset
>>
>>109736899
It's a reference to the movie Armageddon. It's generally said that astronauts should be made of/have the right stuff.
>>
>>109736954
oh sorry lol my dumbass brain didn't get the reference even though i've seen the movie

>>109736947
honestly i'd rather have a beer with debo even though he's an avatarfag than with you, and that says a lot
>>
File: Krea2_turbo_03256_.jpg (1.47 MB, 1776x2368)
1.47 MB JPG
I find it odd whenever wheelchair gens get posted he goes right to physical violence and doxxing
>>
>>109736899
Local diffusion?
>>
>>109736978
I would post some gens but I'm not at my atelier currently.
>>
>>109736995
yes. cats are pretty much diffusers, they can diffuse through very small spaces as if by mere entropy
>>
>>109737104
How odd you bring the dead discussion you have with yourself in /sdg/ here also if you're not him go talk about that shit in the containment thread seems that they all magically decided to post cat gens and not really talk with each other in /sdg/
>>
>>109737053
What does this have to do with catjack having another meltdown?
>>
>>109737204
Can you define meltdown because
>>109736961
Is a actual meltdown
>>
>>109737203
i literally don't understand what you are talking about, i'm a fucking tourist who visits these threads once every 3 weeks
>>
cozy breas
>>
t
>>
File: debo_cl_k2_00017_.png (2.56 MB, 1664x1069)
2.56 MB PNG
>>109737261
I'd recommend simply ignoring that anon. he's a bad faith schizophrenic drama-baiter
>>
>>109737301
Wow they really amped the graphics for Kitten Space Agency.
>>
File: output_no_audio_lossless.mp4 (3.36 MB, 1056x608)
3.36 MB
3.36 MB MP4
He's having a rough time today isn't he. I know the goal is to be annoying but imagine getting angry and seething to the point you make post like
>>109736961
>>
File: debo_cl_k2_00020_.png (2.68 MB, 1664x1069)
2.68 MB PNG
>>109737320
DLSS: on
>>
>>109737323
That anon is correct. If we actually knew how pathetic you are irl you'd never come back
>>
File: concepo658.mp4 (3.59 MB, 960x544)
3.59 MB
3.59 MB MP4
>>109737497
Doesn't stop those other guys from doing daily antics why do you guys always complain/talk offtopic stuff and never post any gens at least?
There's a general for that called /sdg/ is it because nobody actually post there or that even you guys find it dull and unfulfilling?
>>
jealous locusts...
>>
>>109737686
People don't want to post gens when you are around
>>
WHEN IS GEMMAPROMPT GETTING UPDATED?
>>
>>109737916
It's dead just like coomkit and all projects that got a dozen commits and suddenly stopped. MiniConstruct has a couple more commits into it before suddenly disappears too.
>>
>>109737916
because you touch yourself at night
>>
>>109737686
/sdg/ got taken over by some meth head a long time ago. he does nothing but post SDXL abominations and butterflies. he has singlehandedly shitted up /sdg/ permanently beyond recognition
>>
>>109737963
All these project creators could learn a thing or two from Ani. He has serious dedication and doesn't just abandon his passion project.
>>
>>109737916
JUST MAKE YOUR OWN!
>>
>>109737916
https://github.com/whp199/GemmaPrompt/tree/master/skills isn't that the essence of it?

I think you could run ~that with most ways to run LLM.
>>
>>109737981
why are you being such a lazy fuckwit, anon?
>>
>>109737985
I made my own I don't rely on gemmaprompt or miniconstruct or whatever else is out there.
>>
>>109737970
I think it's the wheelchair dent, he uses 3 personas then a few other anons that are clearly mentally ill reply to him sparsely throughout the day.
>>109737908
This post is by him too, he keeps dog whistling old post said about him and wonders why it falls on deaf ears. He typically gets drowned out when actual happenings take place.
>>
File: AniStudio-05896.mp4 (3.48 MB, 768x1376)
3.48 MB
3.48 MB MP4
>>
File: 1784490116175799.png (2.94 MB, 1984x1344)
2.94 MB PNG
>>
File: Freshmaker.jpg (134 KB, 1920x1080)
134 KB JPG
Fresh video.
https://litter.catbox.moe/pw0ltxiibdu98fv1.mp4
>>
>>109738395
What the fuck was her problem?
>>
File: z_00653.mp4 (1.13 MB, 736x736)
1.13 MB
1.13 MB MP4
>>
>>109738395
kekd and cute butto
>>
File: z_00654_.mp4 (1.82 MB, 736x736)
1.82 MB
1.82 MB MP4
>>
File: z_00655.mp4 (1.38 MB, 736x736)
1.38 MB
1.38 MB MP4
>>
>>109738395
While I don't exactly get the correlation between mentos and crime, I did enjoy the video and it's definitely high quality. So, very nice gen anon
>>
File: comfyui_cdF03pT6uN.jpg (24 KB, 500x329)
24 KB JPG
>ComfyUI-SeedVR2
>Unable to allocate 23.7 MiB for an array with shape (1080, 1920, 3) and data type float32
Anything I can do to fix this error? 16g VRAM & 64G RAM. Same error if video is 1080p or 720p.
>>
File: creepweiser.jpg (1.18 MB, 1456x1456)
1.18 MB JPG
>>
>>109738436
>>109738446
also very nice gens anon, is that fl2va or ref2va?
>>
>>109738450
If you have something stolen from you you can counteract the negative mental effects by popping a mentos and thus the overall happiness in the world increases. However, if you kill your assailant with a rocket launcher it goes down so don't do that.
>>
>>109738477
ref2va
>>
>>109738395

Best thing I’ve seen on here lately lol
>>
File: debo_cl_k2_00036_.png (2.66 MB, 1664x1069)
2.66 MB PNG
>>109738349
sexy dom debos are so in right now
>>
Is there any way to apply a lora to only one character in the picture while other characters are unaffected? im using comfyui with krea 2
>>
>>109738466
face aside that's kinda hot
>>
Can video/image gen benefit from multi GPUs apart from generating from both cards at the same time? Like actually accelerating generation of a run?
>>
>>109738395
uhhh... kino has arrived ?
>>
>>109738571
I was about to answer this question unironically, but then I saw this >>109738590
You're getting better attentionseeking boy
>>
>>109738430
https://litter.catbox.moe/s07q6y05wk3fhd42.mp4
I tried audio ref but adding too many refs causes the prompt to shit itself
>>
>>109738395
Heartwarming! Pure kino.
>>
>>109738629
from my experience so far H3 has a hard time with audio references unless the audio is being used 1:1 over the entire video length.
It does work as reference too, but then it needs to be a good 5sec reference with enough spoken words, but not with too many otherwise it hallucinates spoken words even though you didn't prompt it. Also very sensitive to resolution changes
>>
>>109738395
she cute
it would be cooler if the guy exploded in a ball of fire
>>
>>109738682
I agree, in an explosion of some kind. But I think those three large gibs saved it. Those were good gibs.
>>
Next MiniConstruct update will be a really fat one.
>>
>>109738395
quality entertainment. made me think of those 80s movies that had fun shit like this happen in them
>>
File: 1760619950090048.png (1.32 MB, 1302x1168)
1.32 MB PNG
>>109738395
how make 30 second video, pls tell sir
>>
File: 071241.png (9 KB, 787x130)
9 KB PNG
>>109738598
Take your meds schizo.
>>
File: stylish_hoplite.png (2.4 MB, 1216x1824)
2.4 MB PNG
>>
File: faggot.jpg (44 KB, 899x157)
44 KB JPG
>>109738792
Then how do you explain this? Also newfags who don't even know how to do something as simple as the question that was asked don't know about schizo lore you schizo
>>
>>109738395
>dead link
:/
>>
>>109738817
>Then how do you explain this?
You're fucking crazy that's what
>>
>>109737916
this is exactly why I roll my own shit now. if its open source its not a big deal since it's easy for AI to extend it, but if its closed, you're shit out of luck if you want new features or bug fixes
>>
>>109738395
re-upload. i want to see why this got so many (You)'s
>>
>>109738395
She looks like Brooke Shields when she was young.
>>
I read the guide but maybe I'm just retarded. Is there a containerized/service based front-end / back-end split?
On the general LLM side, I use Open WebUI and llama-server.
I have the WebUI in a quadlet and the llama-server as a bare metal service.

Does SD have that equivalent?
>>
>>109738395
pretty nice
>>
>>109738528
you really need a job debo
>>
>>109738830
Also, you have the (you) on the schizo post that has been posted a 100 times not on the "potential" new guy asking how to do the lora thing, lmao, fuck off schizo faggot
>>
File: cruel.jpg (1.04 MB, 1200x1198)
1.04 MB JPG
>>109738839
>Brooke Shields
time is so cruel
>>
File: file.jpg (252 KB, 1276x1680)
252 KB JPG
>>109738857
Indeed.
>>
>>109738856
>fuck off schizo faggot
I called you out on being schizo which you are. It's often the crazy people who deflect back whatever they're being called for.
>>
File: 106.png (70 KB, 1732x295)
70 KB PNG
>>109738725
>>
>>109738840
not typically, it's usually just a web service thing without a split into backend1, backend2, middleware1-4, front end or whatever enterprise thing you could do.

well yes the front end-client is maybe your browser if you want to see it that way
>>
>>109738901
I respect the commitment to the bit schizo, but you lost when you posted an image of yourself claiming the schizo post that has been posted a trillion times already, instead of the believable post, please try something original... maybe a wheelchair gen?
>>
>>109738880
The H3 video from the other Anon must have used references from when she was 14 or somewhere around that age.
>>
>>109738880
pretty woman indeed, I love her eyes
>>
>>109738743
I wouldn't do something like that in one go; that's the sort of video you want to edit in a non-linear video editor from multiple clips and external audio for the music.
>>
File: beerspill2.jpg (1.25 MB, 1456x1456)
1.25 MB JPG
Oh no! My beer!
>>
File: joe shoeden.jpg (1.08 MB, 1456x1456)
1.08 MB JPG
Jack-o-Shoes
>>
File: tophshoes.jpg (1.06 MB, 1456x1456)
1.06 MB JPG
>>
Not looking good for Qwen in the H3 prompt writing tests
>>
File: the-smoking-pike.jpg (1.14 MB, 1456x1456)
1.14 MB JPG
>>
File: debo_cl_k2_00040_.png (2.34 MB, 1664x1069)
2.34 MB PNG
>>109739304
are they really "hard expectations" if they can all be cleared?
>>
>>109739304
Are you giving them any prompt writing skills or are you letting them wing it?
>>
File: da-pipe.jpg (1018 KB, 1248x1248)
1018 KB JPG
>>
>>109738901
Really can be seen with both rentry schizos
>>
>>109739336
>>109739338
I'm implementing a new story planner for creating multi-generation prompt sequences.

>Qwen has a systematic Story Build problem
>Across the Build cases, Qwen repeatedly refuses to populate blank fields on existing Generations.
>In story-basic-four-beat, Generation 1 came back completely blank:
>creativeRequest = ""
>incomingState = ""
>intendedOutgoingState = ""
>cameraIntent = ""
>while Generations 2–4 were quite good.

>The JSON is structurally legal, but applying that Story Plan would leave incomplete planning fields.
>This is a real Qwen/Story-Planner compatibility problem, not merely a harsh benchmark. The same behavior occurs in one-Generation cases whose snapshot inputs actually were deterministic.
>>
File: Krea2_turbo_03185_.jpg (1.5 MB, 1776x2368)
1.5 MB JPG
>>109739304
I have been saying this for days now
>>
What should I generate?
>>
File: 1780358333965452.png (1.32 MB, 2005x939)
1.32 MB PNG
anyone got a download link for this?
>>
>>109739576
>lora trained on 1000 or so porn images
>28 * $15 = $420
>each version probably needs to be purchased separate so he's made many times that
Is it really that easy? Let's say I have figured out how to train Minimax H3, way better than everyone else's shitty loras. Why shouldn't I just pull a grift like this?
>>
Has local video models been tuned down enough yet to work on a 12GB card or should I still cope?
>>
>>109739576
it's been up on the site
>>
>>109739702
>Why shouldn't I just pull a grift like this?
because you most likely won't make something better than everyone else. if you could, you'd have already done it.
>>
>>109739702
>Why shouldn't I just pull a grift like this?
Because you have integrity.
>>
>>109739713
12GB is just under the very smallest model that is 12.9GB large. Get a 5060ti with 16gigs and you can run nvfp4 comfortably, or get a used 3090 and you can barely run the int8 models, I assume you'll suffer though if you try to gen large or long videos.
But before you buy any hardware there are like a billion 6GB vram zomg workflows on civitai, just try it out. But you need at least 32gb ram, better would be 64gb otherwise it's going to rape your ssd, if you have less than 32gb ram and it will be really slow
>>
>Krea 2
Is that style in base or is there a lora I need to get?
>H3
Have there been any major speed ups since release? I got a turbo lora a few days after but haven't touched it since
>>
File: 1694144409344000.jpg (411 KB, 750x724)
411 KB JPG
https://n.uguu.se/slFHwXFo.mp4
>>
>>109739810
50k likes on goontoob
>>
>>109739810
I don't get
>>
>>109739826
filtered
>>
File: 不是好.png (2.45 MB, 720x720)
2.45 MB PNG
>>
File: jipe.jpg (965 KB, 1776x1776)
965 KB JPG
>>
>>109739800
Krea 2, plus the 'Text Fusion Refusal Reduction' LORA.

For H3, just use Turbo LORA with decent MP size (0.6+MP).
>>
>>109740120
*correction, Krea 2 Turbo, base is ass for good images, sadly.
>>
>>109740120
Thank you.
Is there a best lora for h3 or any will do?
>>
catbox is dog shit what the fuck
>>
>>109740152
Is there anything better?
>>
>>109736499
masterpeice
this is one one of the best ever
and checked
>>
>>109740134
Personally, I have not used any LORAs, sorry!

Might be best to ask other anons in this thread, I find that H3 works best with just the turbo lora, or without any loras :)
>>
>>109739718
what site?
>>
>>109740164
I meant the best turbo lora since I found a few
>>
>>109739810
would've been funny without the ebonics
>>
File: loras.png (8 KB, 446x57)
8 KB PNG
>>109740174
either one of these
>>
File: toph-maxxing.jpg (1.26 MB, 1312x1760)
1.26 MB JPG
>>
deleted 600gb of wan loras. h3 is so much better im never going back.
>>
>>109740243
retard
>>
>>109736449
What model was used for bruce willis?
>>
>>109740315
Probably Krea
>>
>>109740315
>>109740315
Krea indeed.
>>
>>109738658
How do you promt the audio to be used 1:1? I have trouble expressing myself to Krea.
>>
>>109740321
>>109740363
Ah, thanks. I was hoping it was a model I could use with vapourkit but sadly not.
>>
>>109740376
*minimax h3
>>
>>109740380
What is vapourkit?
>>
>>109740388
It's a program that can be used to apply filters over preexisting videos. I've been messing about with DLSS5 in it.

https://github.com/Kim2091/vapourkit
>>
>comfy running super slow
>dont know why
>ask my little ai buddy
>afterburner was somehow power limited to 27%

I don't know how the fuck that happened. I never touched it?
>>
>>109740431
Leave some power for other users
>>
>>109740447
Think of the data centers!
>>
>>109740407
Does it give good quality upscales? Does it also fill in detail?
>>
>>109740376
Use the ref2va model, put audio into audio ref_audio_0

then under subject definition:
<Audio 1> is the synchronized audio track of <Video 1> and is reused in the target video.

In summary add:
["word"+ audio reuse]

retention anal:
<Audio 1>: fully_copy - <Audio 1> is reused 1:1 as the target video's complete final audio track.

This should work, if there are people talking you still need to prompt when they talk with an accurate timestamp at the prompt, though you can also try without.
>>
Some absolute garbage in this thread

Do better idiots
>>
>>109740459
I can't really say as I only just started trying it out.
>>
File: Krea2_turbo_hr_fix_00269_.jpg (3.29 MB, 2512x3344)
3.29 MB JPG
>>109740517
what do you have to offer
>>
>>109740543
Wow... What a beautiful gen. Thank you.
>>
>>109737916
GemmaPrompt dev here, what update would you like to see? It already works perfectly for my use case.
>>
>>109740490
Thanks this is very helpful. I hate intricate prompt crafting and h3 is the worst.
>>
Why do latents have to be so big? Can't somebody come up with a compression algorithm for them?
>>
>got training for anima working on AMD
>its still painful
im glad it works but just barely. honest to god i should just start forcing my way into some of these projects, forking it and adding amd support but FUCK maintaining any of that code
>>
Is comfyui still the best?
>>
>>109740832
The best at what?
>>
I wish there was /leg/ for 2D artwork. I'm desperately in need of advanced models and techniques but all I get is illustrious sloppa
Anyway, are there any decent edit models for 2D art?
>>
>>109740911
/ldg/ duh
Giving away early morning phone posting like a retard
>>
>>109740431
shortcut?
>>
>>109740852
being a ui
>>
I am a retarded tourist
do any of the models that come natively supported in wanGP support NSFW gens or do I need to get comfyui
>>
>>109738395
dude please re-upload and upload it on /wsg/. This treasure needs to be shared and I want to add it to my collection of awesome AI videos
>>
>>109738835
Here you go.
https://mega.nz/file/3xRg1T5a#9FHCr7-KH25hDw6GAXJBVAE7t7ykFPAny4pQWRJenTI

>>109738743
I did this exactly: >>109739045.
The continuity plugins I've tried so far destroy the quality of the later clips. At least with references and a turbo lora.
>>
>>109741042
H3 is (Kinda) uncensored, sometimes it will force clothing but 9/10 times it will just work out of the box (From personal experience, mileage may vary)
Krea2 to create the first frame/image, on the templates choose the refusal reduction
>>
hello
https://litter.catbox.moe/ctrcsw.webm
bye
>>
>>109741090
awesome, thanks.
cool idea, and well executed. looks just like a real ad. I love the cheeky way she pops that last mentos. It's little nuances like these that when you find them in AI videos really get you hooked.
>>
>>109741025
it's the worst at that and it spies on you. the only thing it's best at is being a backend but even then it's unstable as fuck
>>
>>109740490
With your help this is the closest I've been able to get.

https://files.catbox.moe/63khwb.mp4

I just can't get it to replace the original voice with the second <Audio> clip I provided despite multiple prompt revisions. I guess I'm just expecting too much here. Maybe it's just something small and stupid that I am missing.
>>
>>109741382
H3 is a vastly superior model but this is one area where LTX actually shines - not too much movement, good lip syncing, super quick renders and you can really push the boat out on clip length. I made this a few months ago when testing LTX2.3

https://litter.catbox.moe/jkizq1ndm1ss1y2u.mp4
>>
>>109741440
hmm. maybe I'll look into that. So you can do audio only stuff with LTX?
>>
File: 1788314980568411.jpg (151 KB, 1024x1024)
151 KB JPG
Can someone recommend a nice comfyui workflow with all I could need for SDXL models?
>>
>>109741440
If I were to create some yapping only video, I'd always choose LTX. I can just provide the audio I want instead of an empty audio latent, write what is said in the prompt for better adherence, and have a perfectly adequate video in the end. 160s for 1080p and 30s video for me.
>>
File: hm.png (2.14 MB, 1164x923)
2.14 MB PNG
If i generated a test video in shit quality, but it turned out to be great, can i throw the same seed and make it exactly the same in high resolution?

Asking before i waste 50 minutes waiting for gen to finish.
>>
File: AniStudio-06896.mp4 (3.71 MB, 768x1376)
3.71 MB
3.71 MB MP4
>>109738528
>>
>>109741675
I would but I don't talk to brown faggots.
>>
>>109741797
japs are yellow tho
>>
>>
>>109741717
That's unfortunately not how this works, but you could get a feeling for whether or not the higher res video goes broadly in the right direction even before completely generating it. Just use "Model Preview Override" node with taeh3 and you'll pretty much know if it does after like 6 steps. The preview is usually clear enough that you can judge whether or not something has gone horribly wrong at this point, at least visually.
>>
>>109741717
i assume you can probably upscale your shit quality vid and v2v it to keep the same composition
>>
>>109741717
No, the only option would be to upscale the result
>>
I wish my system didn't have a melty when and only when I use local models.
I'd be generating so much degenerate filth you wouldn't believe it.
>>
>>109741931
>>109741949
From what i know:
>crap in
>crap out

I'll try some sort of upscaling just to see if it's worth anything. Any recommendations?
>>
>>109736499
>Ciaran Malik
Literally who? That better not be you, anon.

Funny gen though.
>>
>>109742036
any dumb upscale and v2v half steps / .5 denoise
>>
>>109734278
Nice. Which old school mangaka were you going for with this?

>>109733118
Implessive. Could be straight out of the animoo.

>>109731449
You ... you good, anon? Yiff twice if you're not.
>>
>>109741382
Well you asked for a 1:1 audio replacement, you never said anything of dual audio clip changes.
I already said it's quite difficult and you also need to change things.
Here is what worked for me a few days ago.
Important is, if you're using music then the audio clip of the music cannot under any circumstance be louder then the speaking volume reference clip. Simply castrate the audio clip in sneedacity or audacity and try with this:

<Audio 1> is the dialoge and pacing reference for <Subject 1>, but she retains her own feminine and soft voice.
<Audio 2> is the voice reference for <Subject 1>.

[video editing + audio reference] The target video is an edited version of <Video 1>. <Subject 1> replaces <Subject 2> in <Video 1>.

<Audio 1>: weak_reference - the target speaker follows <Audio 1>'s delivery without copying the original signal.
<Audio 2>: weak_reference - <subject 1>'s voice tone and pitch and general sound

you guys really need to be more specific when asking for shit
>>
File: screenshot.1788692785.jpg (378 KB, 1461x620)
378 KB JPG
quality of life features added:
-tag highlighting. mouse over on a tag such as <subject n> highlights every instance of that tag.
-tag thumbnails. a small thumbnail display what the tag represents is also displayed.
-send to ComfyUI. all media content + prompt are can now be sent directly to comfyui workflows. no need to go back and forth
-nsfw writing guide. the guide instructs the prompt to target nsfw content. less guesswork for the llm
>>
File: 1766728639691054.jpg (222 KB, 1536x1024)
222 KB JPG
If heaven isn't all of your generated 1girls raping you forever, I'm not going.
>>
>>109741717
No
If anything. Higher resolutions will fuck up your prompt
>>
File: 1788352922092231.png (3.6 MB, 2304x1152)
3.6 MB PNG
>>
I decided to give comfyUI a second chance, this time staying away from video generation to see if it'll give me just as much trouble. Going with something simpler, which model (Local, not cloud) offers default image to image conversion.
You know what I'm doing.
You know.
You know what I wanna do.
(Degeneracy)
>>
>>109742312
Not sure why you had so much grief with video generation. It's usually just a matter of using some inbuilt template and downloading the models.
Image to image can sort of be done with any image model by feeding the image in as a latent with lower denoise.
I assume you're looking for image editing, which is quite okay with flux 2 klein 9b or flux 2 dev (if you can run that).
>>
>>109742414
No, it's not on comfy, it's something to do with my GPU. There's something that generative UI does that accesses my GPU in a way no videogame or stress test does which makes my GPU fuck up the drivers and I have to re-install them from scratch to be able to use steam. It's bizzare but I don't have the kind of money where I'd be willing to be replacing hardware in this economy so I just have to deal.
Honestly I'm hoping it's video generation itself and not gen in its entirety.
Actually I was still looking shit up after asking that and I did end up on flux clein as well. I'm downloading 4b right now (9b needs an account if I'm reading this error message right.) but I digress. An anon previously told me but I forgot, what's the keyboard button I press after putting the model files into their respective folders to have it update?
>>
>>109742042
i'm not that retarded. here's a hint: check the filename.
>>
>>109742468
Welp, got it set up and running.
>GPU Utilization 100%
>GPU board power 230-260w
>GPU temp 65C
>GPU memory utilization 19896MB
>GPU Memory clock 2487 MHz
>GPU Memory temp 82C
>CPU Utilization 10%
Running stable so far and at this rate it seems like my test gen will take about 10 minutes.
Wish me luck lads.
>>
>>109742468
It's just "R".
>>
>>109742680
Thank you. I'll make sure to write it down this time, so I don't have to exit/run every time I do something new with the model.

I ran a test run, seems to have no issues so far. I'm still threading lightly, but after it finished I closed out ComfyUI and tried running steam and this time nothing got messed up, so it might have been MinMax H3 that my card specifically just doesn't agree with.
I might start experimenting now. Klein 4b ran an exact 50/50 blend of the two faces and ONLY the faces, and I guess it's up to me now to play with prompts until I can make it stop doing that, assuming my shit doesn't break again, that's gonna have me on edge for a while until.
>>
>>109742468
>>109742734
iirc klein 4b is dogshit and there's no reason to use it you can almost certainly run 9b. just go download the file rather than expecting it to autodownload via comfy
>>
>>109742816
>iirc klein 4b is dogshit
I'm starting to see that, it keeps putting clothes on my nekkid womyns.
>just go download the file rather than expecting it to autodownload via comfy
Sorry man I know it's stubborn but I'm not getting another count with anything for shit.
>>
>>109742816
Is klein 9b censored?
>>
>>109742844
thats okay, maybe you can get it from civit or get some quant or sloptune or something of it from huggingface that doesnt need the agreement clicked. if you cant find it, just give up and don't do anything, don't waste your time with the 4b
>>109742877
IIRC similar in kind to H3 but worse, i.e. it's quite stupid/bad at genitals and outright sex but generally doesn't really *refuse* in the way previous BFL models did (by being lobotomized to hell if the prompt had anything to do with human anatomy), and by its nature as an image edit model much like H3's reference model there's not much that filtering can really do against "here is a sexy pose, here is hatsune miku, put her in the sexy pose", swapping characters into a pre-existing image. Might take a couple of rerolls but it's a fast enough model to run. With that said it was a mediocre model on release, worse quality than the competing ZIT and only relevant because it was an image edit model (plus ZIT seed variety made it unusable for most use cases). so it will probably seem even worse today with Krea2 to compare to for low-filtered raw gen and H3 to compare to for reference consumption (albeit into videos instead of images). honestly id consider just using h3 for image edits with a 3 second gen at 1MP but ive not had any image edit tasks i wanted to do to test this out
>>
>>109742931
so you are saying coomers still haven't uncucked it yet?. It's been a year; there must be some image edit that's usable
>>
>>109740587
it needs to be able to analyze videos for video references, and remove the limit on shot count
>>
File: ComfyUI_temp_meddu_00054_.jpg (395 KB, 1024x1536)
395 KB JPG
>>109742877
>>109743006
just use a nsfw lora you retard
>>
>>109743006
https://huggingface.co/darknight9121/FLUX.2-klein-base-9B-bucket-uncensored
Is this not it?
>>
>>109743053
this is so washed out it may as well be a SD1.5 gen
>>
>>109743006
slowly realizing that local is dead
>>
File: 1705956402513702.png (104 KB, 594x594)
104 KB PNG
>>109743071
>local just got the best video model, comparable to sota api models in quality

>durrrrr local rrr ded, lul
>>
>>109743071
I mean for image generation local is still king. And for uncensored video generation as well.
Only LLM applications are better over api due to the massive context size and tensor parallelism that datacentres allow
>>
Is there anything that helps you track trigger words for loras?
don't you tell me to write it down
>>
does anyone use any magic one-liners for h3 i2v for 2d animation? i want to automate some of the reaction and movement effects by genericizing them into high-level instructions so it picks up more style queues from the actions and environment.
ultimately, i'd just like to spend less time prompt writing than i spend genning. i've got a dedicated 3090 for prompt writing, but the output needs enough tweaking and cleanup that it's usually simpler for me to artisanally handcraft each scene.
>>109742931
h3 seems resistant to terminology that directly implies sexualization. terms like seductive, pleasurable, etc don't really do anything, but it knows what to do if you prompt longhand for an expression or reaction. i'll describe a scene as 'racy' or use associated terms that imply it shouldn't steer away from mature details of a certain theme, but i haven't found a prompt that universally keeps it from that kind of self-censorship.
>>
>>109743141
No the real trick is to not download garbage loras that rely on triggerwords unless they are literally character loras in which case the trigger word is self explanatory.
If you download jeffs lora and he uses the triggerword @jeffAnaL32Rp don't expect any sympathy from us.
Though most loras work without trigger words regardless
>>
>>109743141
comfyui lora manager
>>
>>109743053
Why on earth would you use Klein for NSFW ? Are you literally retarded ?
>>
>>109743071
I can see your nose from here
>>
Sometimes I gen something in chatgpt and post it here to keep you guys on your toes
>>
>>109743182
NTA but not retarded, just new.
I'm sure when I get to the point where I've learned enough about workflows I can assemble my own. But as a starter,I kind of have to go with what's given to me based on my extremely limited knowledge, and that knowledge right now is "Find model, download model, put files in the folder. select A and B, input prompt, click run."
There's not a lot of models these templates I've found that actually start off by letting you input two files and prompt what to do with them which is the stage I'm at right now. Just playing around as early experimentation with the tech.
>>
>>109743220
>not using local like the rest of us
>>
>>109743237
See, I'm so fucking new I'm still mixing up terminology. I meant to say
>>109743237
>and that knowledge right now is "Find template, download model, put files in the folder. select A and B, input prompt, click run."
>>
File: Klein9B_00009_.png (1.25 MB, 800x1296)
1.25 MB PNG
First Klein 9B gen and I want to vomit already
Is this the best local can do?
>>
>>109743424
Indeed this is the best local can do, we don't have any better models than klein 9B the model that nobody uses, no sir. Now would you please remove yourself from this general?
>>
>>109743424
no. Krea2 is better and has a lot of knowledge if fighting games. Why not ask what model is sota?
>>
>>109743436
>>109743439
But I asked which is the best image edit that everyone uses, and anon said Klein
>>
>>109743447
anon is just our appointed representative to deal with newfags. you should have asked anon, instead.
>>
>>109743461
this so much
>>
>>109743424
Obviously not. Flux 2 klein is a smaller variant of the larger 32B model as the name implies. That will most likely handle stuff better, but it's not that great either.

Only in terms of NSFW image editing, local is still ahead, just because cloud only models tend to completely refuse that.
>>
>>109743447
because that anon "assumed" you knew what an edit model is when your query is for img2img. It's not my fault anons are retarded
>>
>>109731169
Nobana (kinda), my beloved <3

>>109738454
>ratemyband
Need to all be hung by the neck until dead/10.
>>
>>109743478
the prompt is only turn this image into a realistic photo. it's not img2img
>>
>>109743504
That's literally img2img, are we being retarded rn on purpose?
>>
>>109743504
>turn this <image> into <image>
well that's totally different. i have a workflow for that if you're interested, and it's only got 71 custom nodes and four different pytorch version dependencies.
>>
Can I just quickly say that there are currently two anons that are new and talking about img2img.
The reference anon you're talking to is not the same as me, the anon that wants to make coomer shit by replacing women in gooner shots with women that don't originally have gooner shots.
>>
File deleted.
I love coming back from a long break of genning and discovering that there are new great models. Krea2 seems way better with way less effort than Z-image.
https://files.catbox.moe/bz6v7a.jpg
>>
>>109743521
It's not the same though
img2img add noise then denoise on image
image edit is just a branch of InstructPix2Pix

technically, it's denoise strength vs guidance scale.
I maybe stupid but you are not fooling me
>>
>>109743590
stfu you retard, it's not even funny anymore we all used up our giggles already see >>109743549
who is now even scared of being mistaken for you
>>
File: AniStudio-25693.mp4 (533 KB, 1088x1928)
533 KB
533 KB MP4
>>109743504
Not the best example, but stylised to realistic can be fairly easily be done by feeding the original image as a latent into a model with a denoising strength lower than 1, and prompting for a realistic gen.

I'm pretty shit at prompting for realism in Krea2, as I almost never do it, but I think that gen kinda shows that it works.
>>
>>109743580
>https://files.catbox.moe/bz6v7a.jpg
Show workflow (and prompt)
>>
hanging out here is making me smell like curry
>>
File: 1786006187231739.png (215 KB, 640x692)
215 KB PNG
>>109743702
We have browns here, but they are not indian
>>
File: 1783336729881162.png (1.65 MB, 1024x1024)
1.65 MB PNG
>>
>>109743424
yikes this is bad
>>
>>109743424
AHAHAHAHAHA LOCAL IS SUCH A JOKE
>>
all me (you) btw
>>
>>109743654
https://files.catbox.moe/8asevw.png
Good luck with that. I'm tweaking the workflow constantly and not using everything in it right now so it's messy. I'm still adjusting from zimage to krea.

Maybe someone can tell me if I'm doing anything wrong with the krea generation in the sampler or anything.
>>
>>109743713
>picrel
AHHH! You can't just bust out spoopy stories like that outta nowhere, anon!
>>
File: Test_00018.mp4 (2.59 MB, 1184x896)
2.59 MB
2.59 MB MP4
>>
>I've been unknowingly using ref2va the entire time for i2v fl2va prompts
sigh...
>>
File: 1767447348153619.jpg (269 KB, 853x1844)
269 KB JPG
GPT IMAGE 2.5 IS INSANEEEEEE
>>
>>109743424
You need a really high step count like 200 for the best results
Don't forget to crank cfg to at least 17 to counter the step count and do warmup gens (gen in batches of 16, after the first 14 the results will be magnificent)
>>
>>109744057
>crank cfg to 17
The sigma of the scheduler only works in even numbers, all images will come out looking like shit. I'd personally set cfg to either 16 or 18
>>
>>109744069
>The sigma of the scheduler only works in even numbers, all images will come out looking like shit
that's because you're not using an abliterated text encoder, odd number chads stay winning
>>
File: Untitled.png (194 KB, 3167x1123)
194 KB PNG
>>109743580
Coomer anon here.
I installed Krea2, but unlike with klein which was heavily censored but at least actually used reference material to swap out elements, Krea2 just generates random profile shots vaguely inspired by reference shots.
Anything I can change in picrelated to make it more accurate? Or do I just need to to start learning how to properly write prompts?
(Or is Krea2 even more censored than Klein and I'm wasting my time?)
>>
>>109744105
>pic related
>shows absolutely nothing worth of value
kek
>>
>>109744121
That's the whole thing. That's how krea2 template opened. Brother I did say I am brand new and don't know shit in like seven different posts so far.
>>
>>109744135
didn't the other anon upload his workflow for you?
>>
>>109744105
Dunno, haven't messed with klein, or i2i or reference material at all, I just do t2i. I'm no expert. Krea with nsfw loras is the best at nsfw I've seen so far. I've gotten basically zero body horror or anything like that, but I had the filter bypass and nsfw loras on from my very first gen with it.
>>
>>109744105
you wanted img2img so you have to denoise like everything fucking else. What the fuck do you even want? Edit or img2img because they aren't the same
>>
>>109744012
it's rotating the pixels
>>
>>109744142
Oh god, thanks for that, I forgot that was a thing. Holy shit that's a lot of boxes, I have a LOT to learn.

>>109744156
>you wanted img2img so you have to denoise like everything fucking else. What the fuck do you even want? Edit or img2img because they aren't the same
img2image I am assuming (Replace person in image 2 with person in image 1)
But I'm not usure of the distinction because that feels like editing too? What's the official difference between the two terms?
>>
>>109744105
>Anything I can change in picrelated to make it more accurate?
Not really, the default workflow is good as is.
>Or do I just need to to start learning how to properly write prompts?
Prompting is overrated for the most part. You really only need to keep iterating until you get what you want.
Remember to flush your GPU cache every now and then to clear out any traces of previous prompts that might show up in future generations.
>>
>>109744192
by the way before we waste any more time trying to help you, do you have comfyui manager installed? Cause if yes the thing you need/want is extremely easily accessible, but not without it
>>
>>109738384
Nice, eerie feel. What's the subject though? A fallen star in the mountains? Also: >>109742078

>>109738395
>link d00d
Anyone DL'd it? Reup?
>>
File: Klein9B_00042_.png (1.24 MB, 800x1296)
1.24 MB PNG
>>109743424
After investigating a bit.
Black Forest Lab recommends cfg 4.0, but the default workflow in comfy uses 5.0 for whatever reason. Also, it didn't warn me that I shouldn't leave negative prompt empty. Simply adding "Indian" to negative fixed most of it.
Default workflows were probably written by AI or worse, jeets

Apparently, BFL also shilled their distilled model more than the base one for whatever reason. But I don't care, I'm going to use that to avoid these brown-coded nodes
>>
>>109744245
>do you have comfyui manager installed?
I have whatever is in the ComfyUI_windows_portable_AMD
>>
>>109744251
>link
sorry bud, only for 4chan premium users
>>
>>109744260
Then check out my retard proof guide https://rentry.org/tkdupekk and install comfyui manager as well. simply skip to step 3 and install it, without it you wont get far after that we can talk
>>
>>109742493
Can't be arsed to right now but thanks for the hint, anon.

>>109744261
qq
>>
What's the current best workflow for h3 reference to video? Still the default one with the default model and Spectrum+Sage Attention?
>>
>>109744309
switch sage for cka and yes, that is the current best workflow in terms of lossless quality with maximum speedups
>>
>>109744326
And Spectrum still uses the same default values as always?
>>
File: stock.jpg (126 KB, 562x836)
126 KB JPG
>>109744349
yes the stock values from when you put in the node are the best
>>
>>109743447
image edit is pretty much a dead end
H3 same thing, the hype is just a few anons lying to themselves to try to justify spending on GPUs which they couldn't afford. it's completely fucked, barely usable
t2i with krea 2 I haven't explored that much, there might be something there, but don't get your hopes up either
>>
i had another dream about generating the same prompt over and over again
>>
>>109743424
>>109743071
>>
How does Spectrum in it's current state compare to using the lightx2v turbo lora at 8 steps?

I remember Spectrum would often have output issues like if you wanted camera shake.
>>
>>109744408
just fyi, catjack literally lies all the time about ani and confirmed cannot read english
>>
>>109744038
>local
>>
>>109744462
We figured that out years ago. If only mods would figure that out
>>
>>109744440
Spectrum takes twice as long at 32 steps compared to the turbo lora at 8 steps (ref2va). Spectrum having any output issues is 100% on your end since it should neither be visible or noticeable due to the way it works.
>>
>>109744408
what did julien do to yoland? kek
>>
File: 1409902357770.png (30 KB, 633x758)
30 KB PNG
>tfw haven't genned anything for over a week because i've been vibing a prompt writer the entire time
>>
File: Img2imgoredit(1).mp4 (611 KB, 544x934)
611 KB
611 KB MP4
https://litter.catbox.moe/2qmspj.mp4
>>
>>109744478
Not true. Clanker even told me Spectrum is known to cause issues with the way it handles model data.
There is absolutely a problem with how it handles camera shake. "Slight camera shake" in the prompt often becomes a vibrating/rumbling camera that looks horrible.
>>
>>109743182
you can use klein to nudify, you know, like an edit model, you dumb faggot

>>109743070
sybau
>>
>>109744485
nothing according to that image and video. so why fixate on nothingburgers?
>>
File: forever.jpg (211 KB, 2116x1358)
211 KB JPG
>>109744492
I'm getting pretty close to abandoning my script writer app at this point. the token burn is crazy, the process runs for hours, and it still can't produce usable h3 prompts at the end of it all.
>>
>>109744506
Skill issue, don't care. Either you're on the spectrum or you're not. You're using turbo loras and argue with me about video quality as if I couldn't piss in your face with an ltx gen and you couldn't tell the difference
>>
>>109744525
but comfy said "you can go near him, i just don't think it's a good idea when you shit on him that [...]"
is reading hard for you anon?
>>
>>109744462
>>109744485
>>109744539
>schizo is still reliving the past by looking at his discord friends post from january '25
Let go schizo, we really don't care
>>
>>109744539
Is being banned being able to do the thing he was apparently banned for?
>>
>>109744551
He dedicated his life to this mission. Not sure what the point of it all is since he just became a concern troll lolcow
>>
how do i join the ldg discord server?
>>
>>109744608
There is a schizo who posts images with a cat that has a bag over his head. You can ask him for an invite. Now I would warn you of the consequences but since you're asking for a discord invite it is clear you don't belong here
>>
>>109744567
i think a hard requirement for being a lolcow is to doxx yourself
>>
How to make img2img end result very realistic? I tried throwing in the generated picture into flux klein and z-image-turbo but it looks like ai slop with dirt on top.
>>
>>109744673
Maybe if you pass it through sdxl then through sd1.5 and any other of the extremely outdated models it'll get better?
>>
>>109744673
anon just stop. for your own sake. it's not gonna work anyway. people will just bait you and waste your time
>>
>>109743009
gotcha, I'll have that over to you by end of day wednesday 9/9
>>
>>109744673
if the generation itself looks too sloppa then you probably cant fix it. if it looks pretty close to realistic then you can do a bunch of editing to make it look like it was taken with a cheap digital camera to maybe hide the vae artifacts. this involves compressing the shading, sharpening the image, simulating the bayer filter, and adding some noise. you need to learn the post-processing that phones go through. take a picture with your own phone and zoom in to see what the noise pattern looks like
>>
File: premium copium.jpg (66 KB, 612x484)
66 KB JPG
>>
Any good workflows for generating coherent long videos with Minimax H3? I've looked at some and they seem clunky.
>>
>>109744733
no, its all snake oil. the only option is to make multiple 15s clips, manually review/regenerate as needed, and stitch them together
>>
>>109744705
anon he's clearly taking the piss. Nobody who is new to image gen would even manage to find obscure and outdated models like flux klein or z-image.
You can see for yourself, go to civitai, check out models sorted by newly released and you won't find any of the models mentioned.
>>109744733
I've considered making one but I never saw the point cause unless you have a shot/cut/scene that is longer than 15 seconds there is no need for it.
Though I might make an infinite continuation workflow someday just to make a "walking infinitely through the city" vlog type video to see how much the ai can hallucinate
>>
>>109744524
lol tranny big mad
>>
>>109744534
You're not very intelligent, anon.
>>
>>109742160
If your a tranny it's uno reverse kek
>>
>>109744773
I'm in the 98th percentile when it comes to intelligence, have you considered it might be you?
>>
>>109744626
I thought it was denying reality and advertising your mental illness
>>
File: 3433.png (306 KB, 934x3592)
306 KB PNG
>>109744531
I'm making something like this, but it's quick and modular. Give a story idea and some number of generations, and the writer will set up the framework for each generation. Information like incoming state, outgoing state, camera state and sequence context. Benchmarks are looking pretty good so far.
>>
it's pretty sad that this entire thread got reduced to some pathetic jeetoids trying to sabotage others
nobody is going to buy your shitty gens, sorry
>>
>>109744800
Nah you're clearly a moron, and don't know what you're talking about.
>>
>>109744829
Ok then, mind posting a single good video gen of yours to show us you know what you're talking about?
>>
File: check_00120_.png (2.93 MB, 1504x2416)
2.93 MB PNG
>>109744673
Dunno what you have in mind, but I would assume going with Krea2 and .6 denoise with a good prompt describing every detail of the character and also describing it as cosplay should get you fairly far.

Again, I'm shit at prompting for realistic stuff, so the example could probably look a lot better, if you let some LLM do that or something.
>>
The H3 reference model works maybe 1 in every 10 tries, and each attempt takes like 15 minutes. Useless model.
Whoever says it's good must only be genning talking head videos
>>
>>109744867
I'm trying to give it some references for a character and then another reference for camera view, etc. I have it set up exactly as their prompt guide says but it still sometimes ignores my character and gens something straight from the camera ref
>>
File: endless repair.jpg (427 KB, 1862x1280)
427 KB JPG
>>109744816
I hope you have more success than I've ended up with. when the app was simpler, the stuff it would gen was mostly incoherent. I added in my layers to track trajectories of story beats and characters, plus review/repair phases to improve the visual storytelling. now its just permanently trapped in trying to pass the continuity checks and never succeeding
>>
>>109744867
been my experience as well and 1 in 10 is pretty generous
if you wanna gen zero movement and 1girl and the most obvious position like standing straight yeah it might work but that's about it
>>
>>109738430
>>109738436
>>109738446
Adorbs.
>>
>>109739137
>>109739236
LMAO, WTF are you even doing at this point, anon? You got Joefever or something?
>>
>>109744531
Don't treat this like a job. If it just becomes a source of frustration, you should probably cut back the scope or abandon it.

>>109744867
I'm assuming you still did something wrong, because it worked out pretty flawlessly for me so far.
>>
>>109744923
more like cut back the cope
>>
>>109744947
>>109744947
>>
>>109736449
ok what models and stuff do i need for editing?
>>
File: plot.jpg (281 KB, 2102x1242)
281 KB JPG
>>109744923
>you should probably cut back the scope or abandon it.
a bit of a sunken cost fallacy at this point. "just one more improvement and maybe it'll work", he said 50 times now.
maybe if I sic astra on an architectural review, I can find a new path forward. or maybe I should call it a failure and drop the whole idea
>>
>>109744966
I understand. Astra is supposed to be pretty good at this sorta stuff, but I would assume it will also hit a wall, because I don't think longer form gens like full episodes consisting of mutiple linked gens are something anyone has really figured out yet, so there's a lot of snake oil it will find when researchning that topic.
>>
>>109744848
That guy is not me wtf
>>
>>109745072
Sorry anon, I mistook you for anon
>>
File: 1765333128246052.png (3.2 MB, 1088x1856)
3.2 MB PNG
Give me some prompts to try I'm not creative right now.
>>
>>109745335
>>109744947
>>
>>109741717
You have to use original video as reference.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.