[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File deleted.
Delicious

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109595352

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Local Model Meta: https://rentry.org/localmodelsmeta
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
File: 1766639263892309.jpg (8 KB, 200x146)
8 KB JPG
WHAT SOUND SHOULD I PROMPT FOR FOR THE HAND SLIDING ON A PENIS TO NOT SOUND LIKE PICREL WITH H3
>>
MrCatJack
>>
>>109599549
>WHAT SOUND SHOULD I PROMPT FOR FOR THE HAND SLIDING ON A PENIS TO NOT SOUND LIKE PICREL WITH H3
prompt for that old whistle toy from the 60s that you slide up and down and it sounds like woooooop
>>
File: Test 00026.mp4 (3.71 MB, 800x544)
3.71 MB
3.71 MB MP4
Kinda neat how well Minimax H3 can do drone footage. That's just T2VA, might be even better with references.
>>
>mfw Resource news

08/19/2026

>Minimax H3 Latent Upscaler
https://huggingface.co/LBH-123-AI/Minimax_h3_latent_Upscaler

>CoinVE-200K: A Large-Scale High-Quality Dataset for Compositional Instruction-Guided Video Editing
https://coinve200k.github.io

>LinCa: Accelerating Diffusion Models via Learnable Decomposed Feature Caching
https://github.com/QHR69/LinCa

>RADmesh: Remesh-Aware Mesh Deformation
https://threedle.github.io/radmesh

>MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding
https://github.com/facebookresearch/moe_vie

>Learning What Not to Learn: Adversarial Disentangled Prompt Tuning for Robust Vision-Language Models
https://github.com/cheny02/ADAPT-ACMMM2026

>Where a New Concept Must Enter: Entry Point Gates Cross-Task Usability in Unified Multimodal Models
https://github.com/Zane-ZYQiu/entry-point-umm

>aDSL: Agentic 3D Creation via Joint Agent-Program Design
https://github.com/sig-pku/aDSL

>Live Interactive Training for Video Segmentation
https://youngxinyu1802.github.io/projects/LIT

>ComfyUi-MiniMax-H3-Image-And-Reference-To-Video: Use I2V and reference images simultaneously
https://github.com/BigStationW/ComfyUi-MiniMax-H3-Image-And-Reference-To-Video

>ComfyUI-MiniMaxH3Mod: RefMod reference adapters
https://github.com/Luisacaotica/ComfyUI-MiniMaxH3Mod

08/18/2026

>Qwen-Video-Edit: Instruction-Based Video Editing by Repurposing an Image Editing Model
https://yunpeng1998.github.io/Qwen-Video-Edit-Page

>ByteDance just released Bernini‑Diffusers‑v2
https://huggingface.co/ByteDance/Bernini-Diffusers-v2

>PixRestore: Unified Image Restoration via Pixel Diffusion Transformer
https://github.com/csslc/PixRestore

>PixelControl: Fine-Grained Condition Fidelity in Text-to-Image Diffusion
https://linxin0.github.io/pixelcontrol_homepage/pixelcontrol-site

>MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling
https://expmaster.github.io/megaparts_webpage
>>
File: videogen__00086_4chan.webm (3.88 MB, 1320x1760)
3.88 MB
3.88 MB WEBM
>>
>nigbo malware
>>
>mfw Research news

08/19/2026

>AI models can't tell time or read a calendar
https://www.livescience.com/technology/artificial-intelligence/ai-models-cant-tell-time-or-read-a-calendar-study-reveals

>From Corpora to Co-Evolving Capabilities: Capability-Centric Data Design for Generalist Image Generation
https://arxiv.org/abs/2608.18076

>MSEditor: Toward Consistent Multi-Shot Video Editing
https://arxiv.org/abs/2608.17559

>Optimize Your Sampling: Tuned Diffusion Sampling with Bayesian Optimization
https://arxiv.org/abs/2608.18040

>SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation
https://arxiv.org/abs/2608.17426

>AViTS: Adaptive Spatiotemporal Token Selection for Efficient Dynamic-Resolution Generation
https://arxiv.org/abs/2608.17995

>TINA+: Probing Residual Visual Knowledge in Unlearned Diffusion Models via Diffusion-Consistent Text-Free Inversion
https://qianlong0502.github.io/TINA-Plus-Homepage

>Magnitude-Direction Decoupling for Fast Video Generation with Flow Matching Models
https://arxiv.org/abs/2608.17695

>SE-MoLoRA: Shared-Expert LoRA Adapters for Domain-Specific Photographic Assessment
https://arxiv.org/abs/2608.17514

>GenRec: Knowing Where to Reconstruct and Where to Generate
https://arxiv.org/abs/2608.17832

>EDITBRIDGE: Towards Faithful and Efficient Ultra-High-Resolution Image Editing
https://arxiv.org/abs/2608.18063

>SFMformer: A Spatial-Frequency Modulation Transformer for Lightweight Image Super-Resolution
https://arxiv.org/abs/2608.17966

>MS-MFAD : Multimodal large language models for Face Anti-spoofing Detection
https://arxiv.org/abs/2608.17328

>What to Edit Next: Visually Aligned Image-Editing Follow-Up Suggestions in Conversational Systems
https://what-to-edit-next.github.io
>>
>>109599565
>different fps for face, hair, and claw
grim
>>
Which model can generate kissing and saliva (almost) as good as grok?
>>
>>109599563
>https://github.com/Luisacaotica/ComfyUI-MiniMaxH3Mod
>In MiniMax H3 you can give the AI a reference — an image, a video, even a GIF — to tell it "look like this". That's powerful, but every reference gets loaded and processed every time you generate, which is slow and can "bleed" its look into everything else in your video.
>
>This pack lets you save that reference once as a tiny .safetensors file (a "mod"), and then reuse it as many times as you want, whenever you want:
Sounds too good to be true, if it works, I can finally put the higher resolution videos there.
>>
>>109599590
h3 is pretty good
>>
>>109599534
Why is it cooking dice?
>>
>>109599620
For me it generates only white liquid and any form of propmt makes everyone shoot white liquid from their mouth
>>
File: orange4.webm (3.28 MB, 1968x1096)
3.28 MB
3.28 MB WEBM
Orange
>>
>>109599549
describe the surface texture dumbass
>>
>>109599634
Knave
>>
File: 1757457835578587.png (1.08 MB, 1280x768)
1.08 MB PNG
>>109599559
>Kinda neat how well Minimax H3 can do drone footage.
not neat at all actually given how much high quality drone footage there has been since 2022
however if you're droning right now I'd like to see drone footage in new / emergent settings like vibrant alien landscapes with exotic megafauna, or under the ocean, or just places with not a lot of training data so we can see the intelligence on display from what the model chooses to synthesize

>>109599577
>>You're comparing a Q4 text encoder to FP32 API anon.
i have not seen any different in the int8 convrot qwen3vl versus the nvfp4 qwen3vl which is voodoo magic but its true
i think the seedance vae is much better. china knows how to make video vaes, like if you were an undergrad learning how to make a text to video model from scratch you would probably use the WAN 2.1 VAE because of how good it is and the amount of literature around it

>>109599590
if you're doing anything related to video you start with H3 first. wan 2.2 did great kissing as well and has loras to make it extra messy
careful prompting for saliva trails with H3 though, H3 takes it too seriously sometimes
>>109599636
ok you're trying way too hard to be messy. you better not be writing prompts yourself with H3 and have setup some kind of AI workflow btw
>>
>>109599591
Tell us whether it works and whether it infected your PC with anything.
>>
>>109599591
>That's powerful, but every reference gets loaded and processed every time you generate
it should get cached just like everything else
also it will do fuck all regarding "bleed"
>>
>>109599534
deleted lmao
>>
File: yasuo.webm (3.27 MB, 1280x720)
3.27 MB
3.27 MB WEBM
>>109599587
>>109599587

>he's never seen mesh warping animations before
anon are you 12
>>
>>109599534
That's disappointing I've seen way more provocative images in OP....oh well
>>
>>109599712
(NTA) not the same style at all, but i still appreciate you getting mad enough to post this because I didn't know this Hearthstone gold card "After Effects animation" style was called a mesh warping animation
>>
>>109599750
i'm (((NTA))) either :)
>>
File: videoframe_24186.png (672 KB, 960x544)
672 KB PNG
Nice, I think I found a decent video extension workflow can you tell where the cut is?

https://files.catbox.moe/ydvzf7.mp4
>>
>>109599786
>can you tell where the cut is
i can't tell where anything is with this dogshit resolution, but that includes the cut so well done
>>
this is what the reference model was designed for.

https://files.catbox.moe/0u3nru.mp4
>>
>>109599809
prompt is simple too!

Use <Picture 1> for the physical identity of Floyd.

Jerry Seinfeld is standing in his living room in New York with Floyd. Standing across from him is Floyd. Medium shot captures a quick back-and-forth conversation. Jerry shrugs his shoulders and speaks his line first: "Hey Floyd, aren't you that criminal or something?" The camera shifts slightly to Floyd, who immediately smirks, and replies: "Got any fent Jerry?".

Immediately, Cosmo Kramer bursts through the door of the apartment slamming the door open, while holding a silver revolver and he points it at Floyd. Cosmo Kramer yells eccentrically "stop! he's a nigger!", and fires the revolver at the head of Floyd, and Floyd falls to the ground.
>>
>>109599641
This, I'm not a pedo, i hate all foid regarding age and I rather 2D supremacy, but, foid never pass 12 yo mentally, so pedo are just based
>>
https://github.com/tritant/ComfyUI_MiniMax_H3_Extender

very cool node, worth a try if you want to chain clips or extend beyond 10 or 15s.
>>
>>109599928
Can you select how many frames to feed the next portion?
>>
File: 1777769670498913.png (36 KB, 870x214)
36 KB PNG
>>109599959
you can pick the length of the preceding clip and then it links
>>
>>109599928
Will this give me computer diseases
i'd like a simple node like this, i crave the narrative
>>
File: skin please.jpg (547 KB, 3188x905)
547 KB JPG
Please father may have skin and eyes like the other girls?

I started my diffusion training on 2 million input images to do a final test of my custom pipeline.

The diffusion gods punish your creations for trying to imitate them.
>>
>>109599980
advanced dementia
>>
File: bigasp3 vs RL lora comp.jpg (2.51 MB, 7168x1152)
2.51 MB JPG
So I have been testing RL lora for BigASP 3:
https://huggingface.co/fancyfeast/bigasp-3/blob/main/e2nvn3l1_checkpoint_00000090_v_old_comfyui.safetensors
but unfortunately I have to say I am disappointed.
The good (if you don't have OP card) part is that it's CFG 1 so runs twice as fast as normal BigASP3/Klein-base (it's not step distilled, just per step speed is higher)
It still retains seed variance.
The bad is everything else.
It fails to fix anatomy.
For the lack of better word, "global coherency" is bad and unlike small details I am skeptical that can be fixed with RL.
It still feels unusable for NSFW (The NSFW gens I tested were significantly worse than SFW examples here), a decent NSFW lora for base Klein or Krea2 will function a lot better.
The green filter feels like least of its problems.
I wish the guy best but I don't think he will be able to pull it off. I know I am not supposed to "armchair train", but considering the state of the base model I feel like he fucked up the base training in a way no RL is going to magically fix. But hey I am just some rando who never made a major finetune before, I hope I am proven wrong.
>>
>>109599975
no sir it is merely a node to extend multiple clips together.
>>
>>109599972
I'll take a look, thanks.

>>109599975
No malware as far as gpt 5.6 pro knows.
>>
File: Krea2_turbo_01134_.jpg (1.8 MB, 1776x2368)
1.8 MB JPG
>>
>>109599928
What's the difference between this and
https://github.com/ethanfel/ComfyUI-MiniMaxH3-Contex-Loop

I'm using this right now and it's okay, but clips definitely get more fried as you chain them along. Probably unavoidable. It's also annoying when you start with t2va but eventually need a reference by clip 3 (if only to keep it consistent with clip 1)
>>
>>109600014
didnt fry for me and I tested 10s + 5s, and ref2v turbo lora at 8 steps, need to fix my prompt but it does work:

https://files.catbox.moe/p0b4g5.mp4
>>
>>109600009
even the jannies don't want to look at you ugly hagslop so maybe take a hint
>>
>>109600050
Make more?
Will do but I need to do something first
>>
>>109600036
15s is doable in one prompt. I was talking about longer chaining to like a 40s clip or more
>>
>>109600036
enable clip by clip mode if you want, then tick "validated" if that clip is how you want it. it will skip that for the next gen. all till you get the full sequence you like.
>>
>>109600068
so in that case, 4 10s clips in the node. just keep your references defined in the first part if any (ie <Picture 1> is the physical reference for Floyd).
>>
>>109597117
Thanks for the demo anon.
>>
File: file.png (206 KB, 1616x1134)
206 KB PNG
>>
>>109600068
Long chains will end up drifting it's unavoidable without correctional parts after.
>>
>>109600175
I hate culture war shit, post an actual dataset or fuck off dont care about your site dramas
>>
so chroma krea is actually good?
>>
>>109600175
Instead of vague posting give some context, what is this and why is the discussion filled with nintendo levels of snitching?
>>
>>109600190
>>109600223
here you go since anon is a fag
https://www.asiaone.com/singapore/singapore-photographer-zhang-jingna-anti-ai-scraping-reddit-cara
https://huggingface.co/datasets/CaptiveDreamer/CaraArchive/discussions
>>
>>109600180
Correctional parts?
idk, I figured that chaining must fail eventually. I suppose then you'd need to restart and use ref2va to make a new cut while using parts from clip 1 as reference.
I just find the output from fl2va to be way better normally so I never start out with it unless I have to. Then I regret not having the ability to add references mid clip sequence.

desu, I don't even get how chaining works with fl2va, it can't possibly by cueing only off the last frame as input, I guess it must use guidance to force some overlapping frames at the start or some similar black magic.
>>
File: 1756566337369450.jpg (736 KB, 2500x1258)
736 KB JPG
>>109600227
I see, thanks anon.
I don't understand how they didn't see it coming, but it all looks very performative.
>>
I think this is what minimax ref was created for:

https://files.catbox.moe/q7eeac.mp4
>>
File: count migu.jpg (785 KB, 2528x1684)
785 KB JPG
>>
>>109600249
dude, you've been making Floyd memes for like 2 weeks straight. you really love that black man, huh
>>
>>109600262
very noisy textures
>>
>>109600197
No
>>
How do you go about continuing an existing video that starts with a new cut? Like, I don't want to seamlessly continue from the last frame of the previous video, but I would like the video to start from a different cut in the same environment, with the same characters. Is [video continuation] still correct for this?
>>
File: gemma_world.png (2.22 MB, 1125x1500)
2.22 MB PNG
>>109600262
>>
>>109600267
>2 weeks
dudes been doing it since ltx2 lmao, it's pure autism
>>
File: 794147361.mp4 (3.98 MB, 832x640)
3.98 MB
3.98 MB MP4
>>
>>109600317
I wanna plap this gemma so much
>>
File: 1563893885818.png (23 KB, 616x822)
23 KB PNG
>>109600325
>omg look at how artistic i am guise! please give me the compliments i crave i am a true intelligent anime artistic director!
>>
>>109600325
that's good.
>>
>>109600325
Pretty nice.
>>
File: output_150_4mb.mp4 (3.89 MB, 1552x2048)
3.89 MB
3.89 MB MP4
>>109600050
Here you go "anon" custom made
>>
>>109599989
>I feel like he fucked up the base training
Add another one to the pile. It's crazy that Anima is the only truly successful large finetune of a DiT model. Next best would probably be Chroma, but it's huge, schizo, unstable, and still undertrained, so I'm counting that as mostly a failure.

Genuinely, what are people doing so wrong with finetuning any post-SDXL model?
>>
>>109600356
>Genuinely, what are people doing so wrong with finetuning any post-SDXL model?
It's hard and each mistake is expensive.
>>
>>109600354
>"anon"
Who do you think it is this time?
>>
>>109600345
>>109600351
Thanks.
>>109600342
Wait a second, I know you.
>>
>>109600342
yes, with H3 he really is a anime director now.
Don't be toxic bro. Let anons post their gens in pieces.
>>
File: 1641324872298.png (552 KB, 1440x1388)
552 KB PNG
>>109600409
>Don't be toxic bro
>>
>>109600342
kys wanschizo
>>
>>109600325
Kino sovl
>>
>>109600342
it's pretty good tho.
>>
>>109600325
MiniMax is really good at the slower paced anime.
>>
>>109600420
sure if you're the type of shallow retard who overuses the word "kino" for everything like a terminally online 4chan autist.
people with functioning brains will recognize it as the work of a manchild who has no story to tell and wants to larp as an arteest who can make abstract anime visuals like their messiah hideaki anno (who was actually a mediocre director)
call me when you can craft an actual story with H3 rather than driving the illusion you're some david lynch esque arteest with nothing to say.
>>
y is lilbro sperging out?
>>
File: waifu.mp4 (2.35 MB, 832x1248)
2.35 MB
2.35 MB MP4
>>109600439
>>
File: videogen__00098_4chan.webm (3.5 MB, 1320x1760)
3.5 MB
3.5 MB WEBM
>>109600450
giwtwm
>>
>>109600450
the next david lynch right here
>>
>post my kino last thread to 0 (zero) (You)s
yeah im not posting here again
>>
File: Misty-cut-NS-00016.mp4 (3.59 MB, 640x960)
3.59 MB
3.59 MB MP4
>>109600439
>>
>>109600472
what was it?
>>
>>109600472
you must be new around here
>>
>>109600472
Just the fact that you're using the word "kino" tells me you're an unintelligent, vapid, worthless piece of trash.
Try expanding your vocabulary beyond what you learn off the internet, retard.
>>
>>109600496
Maybe you'd have more fun on reddit.

Like I'm not telling you to get out or anything, but you don't seem to be enjoying yourself here.
>>
>>109600496
I've started compiling screenshots for your future rentry.
>>
>>109600472
I liked it though
>>
>>109600500
No see, what's great about 4chan is I can call morons like you out for what you are without facing repercussions. I'm more than happy to stay right here and give you a good dosage of reality checks: you're not intelligent, you're not creative, and you derive all your vocabulary off 4chan.
>>
>>109600472
>post shit ton of kinos
>someone bakes a new with no collage
god I hate it
>>
File: 1641337492403.jpg (22 KB, 389x324)
22 KB JPG
>>109600506
>I've started compiling screenshots for your future rentry.
>>
>>109600496
I actually use that word in real life as well and have a buddy who isn't even a channer start to use it now as well when he talks to me, he understands that it's used to describe an elevated piece of media, so he will come at me talking about how something is "kino" now. I also have carte blanche to use any terminology by virtue of visiting this site since 2005.
>>
File: 976355226.mp4 (3.96 MB, 832x640)
3.96 MB
3.96 MB MP4
>>109600439
You'd probably need to craft a story in writing, this is just an aesthetic.
>>
>>109600506
you have nothing better to do with your life?
>>
>>109600513
I'm not the guy you were responding to. But you see, a post like that could have gotten you some upvotes on reddit. Here, you'll just be called a faggot. Which you are.
>>
>>109600519
>I actually use that word in real life as well
Of course you do, you're a low-IQ manchild after all.
>>
>>109600518
Thanks, that's one more.
>>
>>109600267
it's a test case, like hello world as a programmer. floyd is my test bed, and it also makes dindus on twitter seethe so why not.
>>
>>109600519
Lots of anons probably use it in real life, since it's just the word for cinema in many languages.
>>
File: 1546914455036.jpg (30 KB, 208x250)
30 KB JPG
>>109600519
>I actually use that word in real life as well
>he understands that it's used to describe an elevated piece of media
>>
>>109600519
based kino buddy
>>
>>109600524
And what makes you think I care about upvotes and downvotes? Such trivialities might carry meaning for vapid morons like yourself, but I'm happier to indulge in the liberty of explicitly reminding idiots like you of what you are. Such liberties are not afforded on websites like reddit with heavy moderation and post filtering.
>>
>>109600267
You're talking to test anon. He only knows miku, ryan gosling, hasan and george floyd.

He also calls Hatsune Miku Miku Hatsune.
>>
>>109600530
>Lots of anons probably use it in real life
Stupid anons*, bereft of any self-awareness.
>>
this thread needs more kinos
>>
>>109600541
i didn't know miku and floyd were the same person. now it makes sense.
>>
File: 1086349302394180.mp4 (3.9 MB, 832x640)
3.9 MB
3.9 MB MP4
>>
>>109600541
>He also calls Hatsune Miku Miku Hatsune.
He needs to take a bullet to the head ASAP
>>
>>109600539
ok faggot. now write another blogpost about how intelligent you are again.
>>
File: 1786971437259090.png (1.19 MB, 960x1280)
1.19 MB PNG
>set cfg to 1
>negative prompts get ignored
>lots of text in the gen
how do I stop this?
>>
>>109600571
kino.
>>
>>109600571
k
I
n
o
>>
>>109600579
Use that neg extension so you can put negatives in the positive prompt.
>>
>>109600579
Negpip or Nag
Try negpip first
>>
>>109600325
Nice
>>
>>109600589
>>109600596
thanks
ComfyUI-ppm?
>>
>>109600506
Check the voluminous data he's left all over /a/ across his many threads if you want more examples.

https://desuarchive.org/a/thread/284134640/#284137843
>>
How do you go about environmental prompting? If I want a reference to the environment/setting in an h3 gen, like maybe the inside of typical house in anime, or a bedroom, or an upstairs hallway, or a basement, what is the best way to generate one? I've been trying with anima but I'm finding it's quite tricky and results are all over the place. I'm guessing the model usually expects an output image with one or more characters in it, and gets a bit confused when there are zero characters.
>>
>>109600325
>>109600520
>>109600571
kino at a level previously believed to be impossible
>>
>>109600579
>doctor, when i move my body like this it hurts. what should i do?

>stop moving your body like that
>>
>>109600623
is that really what it says? lmao
>>
>>109600632
No, it's gibberish.
>>
>>109600477
This is better than any of the Pokemon gens I've seen so far. Looks like you've finally figured out how to make something appealing, naptfag.
>>
>>109600541
false, I have used hatsune miku for quite a while now.

and those are all valid test cases for shitposts, as I like to do a little trolling on twitter. it's free, why not.
>>
File: 83709728869969.mp4 (3.94 MB, 832x640)
3.94 MB
3.94 MB MP4
>>
>>
>>109600652
That's pretty funny considering this is some other anon stealing his gens to put black guys in them.
>>
>>109600669
is this regular t2v or are you using references?
>>
File: bear.jpg (432 KB, 1920x1280)
432 KB JPG
>>109600280
noisy was in the style prompt, the issue was the inpainting, it added too much noise, i should probably try klein 9b instead
>>
>>109600669
pure kino
>>
Thoughts?

https://civitai.red/models/2873402/licking-foot-v2?modelVersionId=3246699
>>
File: 1783925953274305.jpg (474 KB, 1619x1725)
474 KB JPG
>>109600477
peak kino
or as I like to call it, peakino
>>
>>109600694
just regular t2v
>>109600708
thanks
>>
>>109600715
podophile
>>
>>109600608
Anyone?
What's the model for interior proompting? should i be using krea over anima?
>>
>>109600669
help me doctor, i am not used to this much kino in /ldg/
>>
>>109600715
none
>>
okay, NOW it's an episode of Seinfeld.

https://files.catbox.moe/s8dq4d.mp4
>>
>>109600728
krea, you can specify the objects, the location and how those interact within the scene.
>>
What should I prompt?
>>
>>109600762
me
>>
>>109600762
a nice photograph of a solo female relaxing at an outdoor cafe in southern france in late spring, she should be around 20-25 years old wearing trendy contemporary fashion.
oh and give her massive breasts.
>>
File: 1786167745875272.png (1.12 MB, 960x1280)
1.12 MB PNG
BTC is mooning my blaady bitch bastards
>>
File: 1760195018901202.png (1.18 MB, 960x1280)
1.18 MB PNG
>>
help, I can't stop genning chicks with dicks.
>>
>>109600579
>Iyanakao Sarenagara Opantsu Misete Moraitai

A legendary manga. The H3 anons here have no taste, should be generating animations of this.
>>
>>109600829
Yeah, i've been genning girls with my dick, and it gives me this really surreal feeling. It's like, i know that dick belongs to me, but now i'm seeing it attached to a girl.
It almost feels like an optical illusion, you can't help but stare at it and keep genning more
>>
>github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler/tree/main
how slow is it?
>>
>>109600791
sex with bobina
>>
File: 1769186457889098.png (1.54 MB, 1280x960)
1.54 MB PNG
>>
>>109600858
ok but... it looks like shit compared to vae?
>>
File: Krea2_turbo_01154_.jpg (1.69 MB, 1776x2368)
1.69 MB JPG
>>
>>109600844
the sanest poster itt
>>
>>109600870
I've never seen vae up scaling before. How does it work?
>>
>>109600870
Fuck off pedo
>>
>>109600899
>first pass at lower res (0.3mp)
>take latent, send to rtx upscale (to 1.0mp)
>feed into second sampler with turbo lora
>>
>>109600907
>>
>>109600844
tranny
>>
>>109600753
You got any good proompt examples? Specifically for crafting reference images in mind.
>>
>>109600907
>>109600909
I think my brain is to fried to begin understanding that but thanks for explaining.
I've never had good luck with double sample workflows that I made my self
>>
>>109600922
Read the guide
>>
>>109600936
No, cuck.
>>
File: 1786782880989440.png (1.49 MB, 1280x960)
1.49 MB PNG
>>
File: 84834307217745.mp4 (3.53 MB, 640x832)
3.53 MB
3.53 MB MP4
>>
>>109600894
we are updating the rentry with your weeks long meltdown, lolcow
>>
>>109600929
It's the best workflow imo.
you get the benefits of genning a first pass with zero cope and the sharpness of the turbo lora.

The only thing better would be to just gen without any cope for the full 20 steps. but that's likely 2-3x as long.
>>
>>109600936
Yes, kek.
>>
File: Krea2_turbo_01158_.jpg (2 MB, 1776x2368)
2 MB JPG
>>109600942
Pity
>>
>>109600953
Would you mind sharing a base workflow like that please?
>>
>>109600342
Ah so this must be who the bottom rentrys are talking about
>>
File: frogposting.png (148 KB, 585x524)
148 KB PNG
H3 question. How can I make the dialogue more expressive? Is there any way of prompting it? I'm sick of all the characters sound like Veo slop.
>>
>>109600951
why does catjack make the schizos seethe so much?
>>
File: 1783875536388801.mp4 (413 KB, 960x640)
413 KB
413 KB MP4
>>109600520
good idea
>>
>>109600971
you have to manually write out every little noise. h3 is not a very expressive model
>>
>>109600971
Other than having a good vocabulary to describe it you can try having characters briefly breath in or out before or inbetween saying shit.
>>
>>109600976
Considering he has all of them living in his head, I am guessing self loathing
>>
If I'm background genning am I going to notice 20 vs 30 vs 40 steps?
>>
>>109600791
>mooning
>not even 70k
>>
>>109600983
Alright, how exactly?
>>
>>109600971
I found this guide useful in addition to the the guide on HuggingFace https://fluxart.org/blog/tutorials/minimax-h3-dialogue-native-audio-tutorial
>>
File: 1765974710193074.png (1.06 MB, 960x1280)
1.06 MB PNG
>>
>>109601013
it's too embarrassing for me to give an example. write with lots of ellipses and manually stutter with dashes "h-hello"
>>
>>109601015
this guide is one big hallucination and just proof that the text encoder is good enough for you to try whatever with as long as it's similar in structure to the rest of the prompt format (i.e. sections like this: )
>>
File: Krea2_turbo_01168_.jpg (1.98 MB, 1776x2368)
1.98 MB JPG
>>109600976
>>
KINO! the double door open:

https://files.catbox.moe/i9dhho.mp4
>>
>>109601006
20 vs 30 probably
vs 40 probably not
>>
>>109601035
little insensitive and honestly uncalled for
>>
>Flucks
>WANkers
>LT-EXes
>Minimaxers
>Anigmas
>Kleinsteins
>Qweens

Which is the worst fanbase
>>
>>109601028
I think it is proof that there is lot's of creative space for people to get what they want and use various prompts, Loras etc and when using the ref2v model images, video and audio. So far from say a more commercial model that is designed to spit out a standard gen. Being local anyway means people always had more control, H3 expands it because the model is far better.
>>
>>109601046
not insensitive, the actor for kramer said that at a comedy show.
>>
>>109601047
forgot Kreaps
>>
>>109600907
>>109600909
ooooh interesting. i might try this
>>
File: 1759751560609165.mp4 (1.67 MB, 960x640)
1.67 MB
1.67 MB MP4
https://files.catbox.moe/ocdikl.mp4
if you looped the music in this for 2 minutes it would become a track off a cult internet music record from the 2012
>>
>>109600964
https://files.catbox.moe/vv3ay3.mp4
>>
>>109600971
>Voice delivery is emotional, earnest, dramatic.
>>
>>109601058
the irony of course is that Michael Richards is a member of a masonic lodge, that might be a more interesting subject for a Seinfeld skit where he brings Floyd to a meeting and they put bids in so he can be a slave - just like the old days
>>
File: 1786820145408080.png (1.28 MB, 960x1280)
1.28 MB PNG
>>
>he upscales instead of turning the steps up
ngmi
>>
>>109601079
Thank you very much
>>
>>109601091
RTX upscaler is great. 1.25x for Real stuff, 1.5x for Animes
>>
File: Krea2_turbo_01196_.jpg (1.95 MB, 1776x2368)
1.95 MB JPG
>>
File: 1769986967418399.png (1.21 MB, 960x1280)
1.21 MB PNG
>>
ive seen a lot of dumb shit in my life but this takes the cake
>>
okay, enough floyd i've figured out the reference basics. how about some asuka?

https://files.catbox.moe/0k632c.mp4
>>
>>109601152
cute debo
>>
>>109601166
>no green gas
>>
File: 1765982521889055.png (13 KB, 456x72)
13 KB PNG
why is huggingface so slow today?
>>
>>109599534
It’s really important to have a collage in the OP.

I wake up and check these threads to see how many collage points I scored and it’s really disappointing when it’s not there.
>>
>>109601183
same
>>
>>109601180
Getting DDOSed by Dario.
>>
>>109601180
why are you downloading the snakeoil model?
>>
>>109601183
The baker role is for narcissistic lolcows now. You will enjoy the Catjack and Wanschizo slop from now on
>>
>>109601195
because i downloaded the other weaker hybrid models and they just got better and better.
>>
>>109601183
funny how it's only the schizos that never bake a collage.
>>
role reversal: forgot to prompt them to be silent but still not bad.

https://files.catbox.moe/o7othw.mp4
>>
>>109601199
kijai has a ref lora you can use with the fl model.
>>
File: MiniMax_H3__00182.mp4 (3.23 MB, 640x960)
3.23 MB
3.23 MB MP4
>>
>>109601206
is it better? does it work the same way? how many reference images?
>>
>>109601202
giwtwm
>>109601209
thanks for animating my gen anon
>>
>>109601211
>is it better?
It probably works just as well as whatever snake oil you're downloading.
>how many reference images?
same as ref.
>>
>>109601202
higher mp/res for more detail:

https://files.catbox.moe/7knmll.mp4
>>
New BigASP V3 20-step CFG-1 version works pretty well, somewhat green-biased with a bit too much film grain versus the base BigASP V3 though IMO, hopefully he can iron that that in the next RL stage. Output variety especially for faces though is like way way higher than in any other recent model, which I'm liking a lot.
>>
>>109601231
>probably
>>
>>109601239
bro this looks so ass...
>>
File: 1645699910494.png (902 KB, 641x677)
902 KB PNG
>desperately fed krea 2 all that i had of this AI OC: 7 1024x+ images
>2000 steps later (4 fucking hours the power hungry whore)
>perfect 1:1 quality likeness
how in the fuck, okay i won't question it i'm just gonna spitfire fifty million loras for the rest of the month while it works.

thank you plebbitor for your settings
https://www.reddit.com/r/StableDiffusion/comments/1v9yl1u/krea_2_lora_training_the_very_easy_guide_for_16gb/
>>
>>109601239
it does "realistic porn that actually looks like real porn" more convincingly than most stuff recently also
https://files.catbox.moe/y2akht.jpg
>>
>>109601257
krea2 is really good for training, maybe even better than qwen
>>
>>109601239
I guess all tastes exist in nature, because holy shit anon, this is so repulsive to me
I always hated the "milf in professional porn" look, never understood the appeal
>>
>>109601266
it's also as annoyingly limited by the dogshit VAE as Qwen was though
>>
>>109601272
so I hear
>>
>>109601271
I mean for the one I posted here where both had the same prompt:
>>109601263
what do you call what Kroma did? I'd say it's like, pseudo CGI almost moreso
>>
why dont people make a model with the best text encoder and the best vae?
>>
>>109601288
because only i know what the best text encoder and best vae are
>>
File: 2961421.mp4 (3.88 MB, 832x640)
3.88 MB
3.88 MB MP4
>>109600982
nice
>>
File: MiniMax_H3__00184.mp4 (3.36 MB, 640x960)
3.36 MB
3.36 MB MP4
>>
How do you tell H3 to continue a certain action through the video
SHOT 1, at 00:00, character 1 waves their hand at the screen
At 00:05, while character 1 continues the same hand motion, character 2 enters the scene and says hi


With 9-second video and prompt work, but when I gen 15 seconds
As soon as character 2 appears, char 1 stops doing shit.
>>
new trvthnvke is currently being generated
>>
>>109601257
>upscale the image with comfyui and SEEDVR2
Don't do this.
>>
>>109601324
cute. came a lil. just a lil though.
>>
>>109601257
I gave up on AI-Toolkit and OneTrainer. First one just hanged without any error messages after doing a full reinstall. Second one simply won't startup. Musubi Tuner is much easier to use despite the command line.
>>
File: 1760163230872619.png (1.25 MB, 960x1280)
1.25 MB PNG
>>
>>109601335
I find the model is most coherent up to 10 seconds, after that it's a bit hit or miss.
>>
>>109601335
Have you tried adding "At 00:06, character 1 still waves their hand at the screen"?
>>
seedvr2 is fucking trash
I 'd rather gen native res than use this shit
>>
Why is coming up with ideas for vidgen so much harder than imggen? And don't tell me it's because videos are a series of still images, it's something more than that.
>>
>>109601353
well, the point being the prompt and timing work very well for a 9-second prompt, but at longer duration it's so random.
>>
>>109601365
because you are brainrotted by social media, so you need to come up with something that will get (You)s instead of just making what you want
>>
i just remembered i used some interpolation node for wan to double the fps of videos. is that old trash or are people still using it? what was the node called?
>>
File: debo_wr_k2_00004_.png (3.08 MB, 2048x1101)
3.08 MB PNG
>>
>>109601382
rife?
>>
>>109600971
It's alright, but you have to prompt long and very detailed.
Vowel sounds like ~~~~, sooooooo long, ... also work
>>
>>109601379
No. I make what I want and it just so happens to get (You)s.
>>
Can anyone give me a good workflow for Krea that's naturally creative?
With Anima, I intentionally don't use the turbo lora or model because I don't like how uncreative it is between seeds. It always produces an extremely similar output image every time the seed changes.
With Krea, literally every single workflow including the comfy default uses turbo, and my results are always very similar with very little creativity. Is this just a reality of the Krea2 model? Or is there a workflow/model that better fits what I'm looking for?
>>
https://files.catbox.moe/qsbk7e.mp4

neat, it worked. reference model, 8 step turbo ref2v lora. apologies for dub.

Use <Picture 1> for the physical identity of Asuka.

Use <Picture 2> for the physical identity of Misato.

Asuka is wearing a black bra and panties. Misato is wearing a red jacket and black miniskirt. Asuka and Misato are standing beside each other in the hallway of a condo with white walls. Misato grabs Asuka and pushes Asuka hard against the wall face first. Misato says "hmm what do we have here?" as she runs her hand softly along the back of Asuka from top to bottom, and stops with her hand on her ass and grabs it. Misato smiles.
>>
>>109601397
>8 step turbo ref2v lora
there's a 8 step ref2v lora now? who made it?
>>
>>109601397
>evacuck moron still can't let go of his favorite shitty series
>>
>>109601405
there isnt, you use the 4 step one at 8 steps and it works better
>>
File: MiniMax_H3__00185.mp4 (3.06 MB, 640x960)
3.06 MB
3.06 MB MP4
>>
>>109601393
if you were a free thinker, then you would accept that you have no ideas and wouldn't try to force it
>>
>>109601408
i used ref lora for fl2va
motion is so smooth but quality is much worse vs fl2va lora
>>
>>109601410
You have very shitty taste in women. Are you an indian?
>>
>>109601414
Even the freest of thinkers sometimes forget one cannot force it... No matter how hard they try...
>>
https://files.catbox.moe/rjfo8i.mp4

I kneel China, that was a nice ass grab.

Use <Picture 1> for the physical identity of Asuka.

Use <Picture 2> for the physical identity of Misato.

Asuka is wearing a black bra and panties. Misato is wearing a red jacket and black miniskirt. Asuka and Misato are standing beside each other in the hallway of a condo with white walls. Misato grabs Asuka and pushes Asuka hard against the wall face first. Misato says "hmm what do we have here?" as she runs her hand softly along the back of Asuka from top to bottom, and stops with her hand on her ass and grabs it. Misato smiles.

camera cuts to a closeup of Asuka's ass with Misato's hand on it. Misato says "not so tough now, are you?", as she squeezes Asuka's ass and releases it.
>>
>>109601419
Yes, how'd you know?
>>
>>109601410
please tell me she makes licking sounds when she eats
>>
File: MiniMax_H3__00186.mp4 (2.77 MB, 640x960)
2.77 MB
2.77 MB MP4
>>109601419
>>
>>109601428
>>109601431
SAAR
>>
>>109601424
They must have trained it on videos and images scraped online without any filtering. You can feed it a porn image and it will get the sounds and motions correct.
>>
>>109601439
>sounds correct.
>CRUNCH CRUNCH BLOWJOB
yeah no
>>
>>109601439
if that was the case, the model would know how to draw bagina and benis
>>
even with a filter bypasser, krea 2 doesn't like microbikinis. whats the closest term thatll work?
>>
>>109601424
https://files.catbox.moe/vv7yof.mp4

since the variables are defined already, you can just swap images to get a new pairing.

and I also think it was trained on either ecchi or hentai now. same prompt.
>>
File: Krea2-_00961_.png (1.56 MB, 944x1520)
1.56 MB PNG
>>109601473
How micro are we talking?
>>
>>109601473
ask your cow sandeep
>>
>>109601486
that basically
>>
>>109601480
Now do asmongold doing that to hasan in a hallway.
>>
even with a lower quality image, it's impressive how the model basically sees it as a concept or template to generate, like from a lora. ref2v is fun.

https://files.catbox.moe/br8zr7.mp4
>>
>>109601480
>starts speaking Simlish
>>
>>109600571
Is this all t2v?
>>
>>109601517
I didnt say remains completely silent, that's my fault, kek. If you dont specify dialogue or silence, sometimes you get jibberish.
>>
>>109601394
Anyone?
>>
>>109601504
try this one
>>
Besides running faster could you really ask for more than H3?
>>
im very impressed ref2v didnt just clip through the clothing with this prompt.

https://files.catbox.moe/dtblrk.mp4
>>
>>109601544
this reminds me I should've saved those avatar gens.
>>
File: bkub.png (487 KB, 512x768)
487 KB PNG
>>109601504
I thought the one that does image to video was faster? Why do people use this reference one
>>
>>109601542
i want it to be more expressive by default
>>
Describing what happens in autistic detail seems to work sometimes with H3.
Not surprised, if you have seen how prompts are done for H3 Music, you need to become a poet or whatever Shakespeare to write that shit
>>
>>109601550
make your own, this one is super easy. just plug in blue avatar girl.

Use <Picture 1> for the physical identity of Elegg.

Elegg is on a sunny beach in japan, sitting in a beach chair beside the ocean. The camera view is a front view of Elegg.

Elegg grabs a bottle of suntan lotion and puts some lotion on her hand. She rubs her hands together then massages the lotion in circles onto her breasts, which now have an oily appearance.

Elegg grabs the bottle again and puts lotion in her hand. She flips over and rubs the lotion on her ass with one hand.

>>109601556
i2v isnt really faster, ref2v handles characters much better, i2v is more for entire environment morphs. like, i2v changes the entire environment but ref2v takes your image input and treats it like a character in the video.

both are great though.
>>
>>109601542
A little less blur I guess? It's already insanely good.
>>
>>109601542
Best model made so far but better res, speed and sound is always to strive for
>>109601489
That prompt was just miku and micro bikini besides the normal tags
>>
Crazy how the Ideogram community just died. It seems like the perfect tool to actually make reference images for scenes in H3, but I assume nobody built the right tool for it.
>>
>>109601542
Running much faster
>>
>>109601568
Ideogram was good but the inbaked safety feature was so cringe at times so I understand it. Krea2 does almost as good quality 10% difference but with a 2 times speed up so I get why it failed
>>
>>109601568
I still like ideogram it's just... a lot. You feel?
>>
>>109601542
>a unified FL/REF model with no compromise.
>Clip extending without color shift
>better audio quality
>>
>>109601568
Those bboxes are just too much work when you can just quickly iterate a krea gen with an LLM.
>>
File: generate_20.png (1.12 MB, 1024x1024)
1.12 MB PNG
>>109601574
Yeah, that is true. I heard that json prompting would get around the safety filters, but I never tested it myself.
I hope the lora controlnets on Krea2 actually work well. I'm pretty sure krea2 can do anime just as good as the anima models as well.
Krea2 seems to be best at everything now
>>
>>109601596
I assume someone could just vibecode so Gemma-chan could do all the work for the prompting and bboxes, but I guess it never worked out.
>>
>>109601542
Non-body horror genitals
>>
>>109601607
all genitals are body horror
>>
>>
>>109601597
Ideogram's filter more of an everything filter, rather than a safety filter. You can't even prompt basic everyday items without doing convoluted json gymnastics.
>>
>>109601239
>Ah yes, that one model that makes Flux 2 somehow look like SDXL.

Just slop my shit up, I mean why did anyone put faith in that retard. Retardation started when he decided not to tune Chroma.
>>
>>109601607
oh yeah, fucking this.
>Get the best local video model
>it can't do benis and vagene
>>
I think H3 is a good litmus test for people's creativity in general.

It really lets you do anything. You generally have around a ten second constraint to tell an interesting narrative. And those that can will, but many will just prompt 1 girl doing something.
>>
>>109601263
>The bigASP shill is an actual cuck

Imagine my surprise
>>
not sure how to use slow motion as a prompt or modifier but...it worked with timestamps.

https://files.catbox.moe/jilr67.mp4
>>
>>109601637
>1 girl doing something.
Limitless potential
>>
>>109601597
>I heard that json prompting would get around the safety filters
yes that worked from the start, you just needed to prompt a few bboxes and it essentially skipped censoring
>>
File: Krea2-_00971_.png (1.57 MB, 944x1520)
1.57 MB PNG
>>109601597
You could get around it the same way as Krea but the speed difference and just knowledge was superior
>>
>>109601637
>You generally have around a ten second constraint to tell an interesting narrative
Vines had 7 and it was an endless source of priceless memes
>>
bake a bloody collage
>>
>>109601649
What loras do you use?
>>
>>109601646
>it essentially skipped censoring
It's interesting. It was basically nonexistent once you prompted it right. I'm not talking about cheap tricks or things to bypass the filtering. Just straight up following the prompting guidelines, but so many people dropped it like a hot potato at the way it was implemented that it was hard to recover from.
>>
>>109601656
Turbo and Realism engine
>>
don't bake a bloody collage
>>
>>109601637
nah, 10 sec is just for 1 gen. That is Wan tier, so unless doing for a quick meme then 1 gen should really be part of other gens to create a longer sequence. H3 will do 20 secs easy so no excuses thinking only 10 is really possible. should move on from the mindset of doing what Wan already did over a year ago, the model is good enough for a more serious attempt at creating.
>>
>>109601656
nga just get any of the realism loras and crank them up to 5 strength
combine with snofs if a single lora isnt enough to unlock the safe horny
>>
Now lets see if we can make it to at least page 5 before someone blows their load.
>>
>>109601568
If you could actually use reference images with it it would be more useful. But without it, it's just basic txt2img with extra steps.
>>
>>109601661
You got a good aesthetic prompt. ty
>>
File: h3_00114.webm (3.74 MB, 1648x944)
3.74 MB
3.74 MB WEBM
Everyone, I have an announcement that needs you attention.

https://files.catbox.moe/s4akym.webm
>>
>>109601542
people like you said the same thing about flux 1
>>
How is Krea better than ZImage?
>>
>>109601688
it's not. krea 2 is thoughever
>>
>>109601637
we're really are still limited in our power to 1girl/1boy - not even close to realizing the full potential this top tier genre offers
>>
File: kino alert.gif (577 KB, 498x498)
577 KB GIF
>>109601684
>>
>>109601637
I plan on creating a story, but I keep hesitating knowing how much work is required to make it work right. Like I'd need character references, voice references, appropriate reusable music BGMs, environmental references for multiple settings/rooms that will stay consistent enough between shots.
To be honest, even though all this should work in theory in R2V, I'm still wondering if it actually works in practice. R2v has been pretty good for me so far, but it's definitely not perfect, and I dread the scenario of having to run the same 15 second prompt several times before I get a decent enough cut.
>>
>>109601705
I believe in you frend.
The first H3 made full length episode or minor movie will come out this year and it'll be delicious drama between ai fags and normal fags
>>
>>109601688
People talk shit about certain models since they can do more than their cucked models, don't bother listening to them.
>>
>>109601688
ZImage is good but Krea2 replaced it for me just like something hopefully will replace Krea2 later on
>>
>>109601705
I would recommend always keeping the gens free of music and then adding it in post.
>>
My question is: why is it so rare to get good edit models? Why are we still stuck with shit like Klein and Qwen Edit?
>>
>>109601721
Klein is actually really good. idk why people don't use it more. Even me as someone who knows how powerful it is I barely use it.
>>
this is very interesting, one ref image, just a game screenshot. didnt prompt movement or anything. source is a ff6 screenshot.

Hatsune Miku is walking around a videogame in the style of <Picture 1>.

https://files.catbox.moe/inpibq.mp4
>>
>>109601719
Yeah that's what the plan is, although I am still very new to generating music. Only just started messing around with Ace Step UI yesterday.
I think the hardest part will be composition and just getting the camera to behave. H3 is a big step up on the slop we had before, but in the context of storytelling it's important to have fine grained control.
>>
File: ComfyUI_temp_ajqcc_00014_.png (1.46 MB, 1024x1024)
1.46 MB PNG
>>109601688
It's very clear when you try to do a number of different things.

For example, Z-Image only knows the ghilbi style anime, while base Krea2 can do almost any kind of anime. Krea2 is a garden, Z-Image is a hole in the wall where they stick retards so they feel included too. Literally just gen with both, and you will see.
>>
just plucked a silvery white hair out of the noseberg
I'm getting old
>>
>>109601750
Totally get it. I got these long ass hairs growing out of my ears now. I keep plucking them and they keep coming back.
>>
>>109601728
ok now we got some kino. gta vice city screenshot as input:

Hatsune Miku is running around a videogame in the style of <Picture 1>.

https://files.catbox.moe/vlgn7w.mp4
>>
>>109601762
nice
you can pass it off as mod footage
>>
>>109601657
Yes, this is what I thought before.

With ideogram, gemma-chan can make the bboxes and description, and you have a real character sheet. Krea2 looks better, but it feels like you have less control.
I think the problem with genning is that everyone is stuck in Comfyui noodle jank, and making advanced tools in the UI is difficult.
>>
>>109601773
gonna try this with the gta 6 leak vids with the watermarks removed with klein edit 9b, kek
>>
>>109601762
ok now have her carjack teto
>>
Has nobody generated a long-form story in H3 yet? I find that hard to believe. Surely someone's made something that goes for minutes and isn't just mindless nonsense?
>>
>>109601762
Good gen but that sound is nightmare fuel
>>
>>109601782
i saw a couple minutes long pedo one in the /b/ thread the other day
>>
>>109601784
was just testing style transfer, ideally prompt what type of music is playing to avoid that spam.
>>
>>109601782
i have but its too long for 4chan
>>
File: 1756762529124305.png (1.51 MB, 1101x884)
1.51 MB PNG
https://files.catbox.moe/vuw6nv.mp4

lmao

Hatsune Miku is running around a videogame in the style of <Picture 1>. She runs up to a male pedestrian and starts punching them. The man falls on the floor and drops a green dollar bill.
>>
>>109601780
https://files.catbox.moe/7ljwrk.mp4

holy shit, it worked. but I wanna regen at a better resolution.

<Picture 2> is the image reference for Teto.

Hatsune Miku is running around a videogame in the style of <Picture 1>. She runs up to Teto standing beside a car and pushes Teto to the floor. Miku enters the car and drives away quickly.
>>
>>109601804
also what blows my mind is how the minimap actually works.
>>
something brutal is about to drop
>>
>>109601542
it's very good and a credit to it that it can be picked apart to look for flaws which mean improvements can be made again in a future model/update. For a local model no one could really ask for more at this time
>>
>>109601804
0.6mp! even higher fidelity now and basically game res!

https://files.catbox.moe/tyeezf.mp4
>>
>>109601824
this level of violence against teto is something credit card companies should ban
>>
>>109601824
<Picture 2> is the image reference for Teto.

Hatsune Miku is running around a videogame in the style of <Picture 1>. She runs up to Teto standing beside a car and pushes Teto to the floor. Miku enters the car and drives away quickly.

also note I did not mention GTA at *all* in my prompt but it still worked fine.
>>
>>109601824
the sims talk is the funniest shit
>>
Alright ani made a new thread and removed the rentries again.

Anyone else want to make the real thread? Otherwise I will in a few minutes.
>>
>>
That's it. I'm not moving threads until one is deleted. This shit needs to stop.
>>
>>109601874
It's your own fault for being too stupid to distinguish ani threads from the rest. It's not difficult for the rest of us to figure out.
>>
>>109601879
the fuck are you talking about? not him but one usually, inevitably, gets deleted. there's no point in moving if a janny is gonna sweep one under the rug it's a fucking mess of a system
>>
>>109601879
I literally don't give a shit about the links. I don't care if they're there or not. Being anything other than indifferent towards their presence is the whole issue.
>>
File: MiniMax_H3__00193.mp4 (2.48 MB, 640x960)
2.48 MB
2.48 MB MP4
>>
>>109601880
The only real threads are the ones with the rentry links. That is what the community will post in, and you're a moron for not figuring this out.
>>
>>109601885
You're just a salty retard that's mad their stupid dancing gorilla video didn't get a single (You) after posting it three times now so made an entire thread with it as the main video to make sure everyone sees it this time.
>>
File: MiniMax_H3__00194.mp4 (3.39 MB, 640x960)
3.39 MB
3.39 MB MP4
4chan baking competitions
>>
File: (YOU).gif (573 KB, 640x328)
573 KB GIF
>>109601885
yeah, except one of the last few times the one with the rentry links got deleted and we had to make a new one, regardless, because the jannies are fucking cross-eyed niggers. feel like i'm talking to a bot
>>
>>109601903
No ani, you removed the rentry links again. You keep trying to do this and keep failing miserably. You're a failure and an embarrassment.
The links stay. You only have yourself to blame.
>>
>>109601879
the real thread is obviously the one without the rentries in it because it makes dramatic faggots like you cry piss and shit until you get banned then we don't have to deal with you
>>
here we go again. baker is having another meltdown
>>
>>109601933
>the real thread is obviously the one without the rentries
Sorry ani, you don't get to make that decision. The community does, and the rentry threads keep winning.
You're a failure. A worthless piece of garbage.
>>
>Rentry thread is just filled with the bottom of the barrel retards
Really makes you think
>>
>>109601974
keep spamming so the thread gets deleted faster
>>
WHAT DID I FUCKING SAY
I FUCKING TOLD YOU
>>
oh!
>>
should I bake a new?
>>
>>109602371
new is here >>109601828
>>
>>109602371
Once we're closer to page 10?
>>
>>109602371
Yes please. Janitor retard screws up again.
Otherwise I will bake again.
>>
>>109602422
>>109602422
>>109602422
>>
>retard did it again
>>
New
>>109602575
>>109602575
>>109602575
>>
I give up
at least they didn't ban me
>>
>>109602623
>spams the entire board
>doesn't get banned
how?
>>
>>109602640
The one who spammed the board got banned though :)
>>
>>109602659
but you said you didn't get banned
>>
>>109601180
>today
>>
>>109600342
>>109600412
>>109600439
>>109600518
>>109600532
>the subhuman monkey is jealous of its betters
kek
>>
>>109604215
>>109604215



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.