/lmg/ - a general dedicated to the discussion and development of local language models.Previous threads: >>109338632 & >>109334239►News>(07/22) Upstage’s Solar Open 2 250B-A15B released: https://hf.co/upstage/Solar-Open2-250B>(07/21) Cisco releases Antares for vulnerability localization: https://hf.co/collections/fdtn-ai/antares>(07/21) Korean Motif-3 314B-A13B released: https://hf.co/Motif-Technologies/Motif-3-Beta>(07/21) Laguna S 2.1 118B-A8B released: https://poolside.ai/blog/introducing-laguna-s-2-1>(07/21) Nanbeige4.2-3B released with Looped Transformer architecture: https://hf.co/Nanbeige/Nanbeige4.2-3B►News Archive: https://rentry.org/lmg-news-archive►Glossary: https://rentry.org/lmg-glossary►Links: https://rentry.org/LocalModelsLinks►Official /lmg/ card: https://files.catbox.moe/cbclyf.png►Getting Startedhttps://rentry.org/lmg-lazy-getting-started-guidehttps://rentry.org/lmg-build-guideshttps://rentry.org/IsolatedLinuxWebServicehttps://rentry.org/recommended-modelshttps://rentry.org/samplershttps://rentry.org/MikupadIntroGuide►Further Learninghttps://rentry.org/machine-learning-roadmaphttps://rentry.org/llm-traininghttps://rentry.org/LocalModelsPapers►BenchmarksLiveBench: https://livebench.aiProgramming: https://swe-rebench.comAgentic Coding: https://deepswe.datacurve.aiContext Length: https://github.com/RecapAnon/NoLiMaGPUs: https://github.com/XiongjieDai/GPU-Benchmarks-on-LLM-Inference►ToolsAlpha Calculator: https://desmos.com/calculator/ffngla98ycGGUF VRAM Calculator: https://hf.co/spaces/NyxKrage/LLM-Model-VRAM-CalculatorSampler Visualizer: https://artefact2.github.io/llm-samplingToken Speed Visualizer: https://shir-man.com/tokens-per-second►Text Gen. UI, Inference Engineshttps://github.com/lmg-anon/mikupadhttps://github.com/oobabooga/text-generation-webuihttps://github.com/LostRuins/koboldcpphttps://github.com/ggerganov/llama.cpphttps://github.com/theroyallab/tabbyAPIhttps://github.com/vllm-project/vllm
►Recent Highlights from the Previous Thread: >>109338632--Analyzing llama.cpp's use of C++ for portability and its popularity:>109339502 >109339511 >109339527 >109339533 >109339567 >109340724 >109339576 >109339610 >109339973 >109340000 >109340048--Debating if Anthropic's Fable actually solved the Jacobian conjecture:>109339770 >109340267 >109340434 >109340652 >109340958 >109341029 >109341115 >109341162 >109341225 >109341256 >109341243 >109341263 >109340663--AI solving 30-year-old math conjecture via extended reasoning prompts:>109342339 >109342391 >109342403 >109342511 >109342400 >109342413 >109342432 >109342461 >109342494 >109342607 >109342620 >109342669--Using batched decoding for parallel plot branching and state tracking:>109341751 >109341771 >109341772 >109341792 >109341811 >109341848 >109341879--Using in-character thinking prompts to bypass Gemma 4 guardrails:>109342103 >109342163 >109342169 >109342171 >109342194 >109342320 >109342398 >109342490 >109342520--Debating the AI bubble's impact on hardware prices and SSDMaxxing:>109340040 >109340177 >109340206 >109340560 >109340792 >109340818 >109340930 >109340208 >109340234 >109340866 >109341287 >109341379--Using forks and PR branches for MINIMAX M3 support in llama.cpp:>109339012 >109340706 >109340753 >109341708 >109341837--Performance benchmarks comparing Laguna, DeepSeek V4 Flash, and Qwen:>109342283 >109342307 >109342354 >109342443 >109342420--Gemma-4's ability to reason in character during roleplay:>109339662 >109339858 >109339989 >109339991 >109340010 >109340041 >109340114 >109340149--Logs:>109338973 >109339080 >109339858 >109340149 >109340412 >109340420 >109340592 >109340866 >109342103 >109342413 >109342543 >109342607--Kimiposting:>109341010--Miku, Teto, Dipsy, Gemma, Kimi (free space):>109340575 >109340936 >109340560 >109340599 >109340621 >109341291►Recent Highlight Posts from the Previous Thread: >>109338633Why?: >>102478518Enable Links: https://rentry.org/lmg-recap-script
Cloud models have disproven two unsolved conjectures in a week. What have local done? Nothing.
I NEED MORE RAM
>>109342904made all the cloudfags shit their pants so hard they're crying to their loony president to have local models banned (again)
>>109342904made me cum, that's more important arguably
>>109342904they’ve been posting bait on /lmg/
>>109342871You wouldn't believe the amount of important stuff in niche math fields that has no mentions in wikipedia.
>>109342935they’re not important though
>>109342904Mine did a widdle typo and bricked my setup. Pretty impressive for a sub-billy agent.
>>109342889The robot master lineup is growing!
so unironically if they can solve this super difficult math shit that no human can do, what's stopping them from coming up with AI architecture breakthroughs?
>>109342935If you can't explain the use of the "important stuff" to laymen in two sentences, it's not important.If it's important, it will have a wikipedia page.
>>109342889spanking dipsys robutt
>>109342964They're solving meme mathematics, not anything actually useful.
>>109342964math is like jerking off to anime, ai architecture breakthrough is like seducing a irl stacy
>>109342966>unc has attention span for 2 whole sentences
>>109342985they're solving difficult maths, for the moment it's meme math, but it means they'll be able to solve relevant maths problems, it's just a matter of time now
>>109343000>relevant maths problemssuch as...?
>>109343000That's not what the original question was about.
>>109342964fringe theoretical math like that is mostly a meme so "30 years unsolved" usually means "some autistic math guy's problem with no practical application that nobody card enough to look into"
>>109342964math problems are small and cute and well contained and most importantly, you can test the idea instantly. doesn't apply to architecting a massively complex system with completely undefined rules.
>>109342964It's a question of specificity.If language models work for solving only 0.01% of math conjectures and it's basically random for which ones they work you get an expected 0.01% progress on any specific thing.
>>109342964To add to >>109343024, you need someone to be responsible for the risk and opportunity cost.
Gemma-chan, design Transformers2. Make no mistakes.
>>109343040Puny little 31b ain't solving shit on her own
>>109343007Idk, Navier Stokes? P = NP?
do you pretty much want compression threshold to be .5 for hermes? my server kept fucking up eventually when i had it at .75
>>109343049gemma... my cute little retardwife...
>>109343080Navier Stokes equations have been used in computer graphics since early 2000s or even longer than that.
>>109343017The most interesting prospect about using LLMs this way is that a lot of this math might have unforeseen uses that went by unnoticed precisely because no one cared enough to look into it. Having agents trained on solving things like this is going to open up huge swaths of practical applications.
https://huggingface.co/reteetzad/Kimi-K3it's over, sama won
>>109343112We're so back!https://huggingface.co/moonshotai/Kimi-K3
>>109342964>>109343024They definitely can. The context length breakthrough in 2023 was a pretty small but significant change. The big problem is like this guy >>109343030 said, there's a lot of stuff to go through so it's like finding a needle in a haystack.
>>109343080>P=NPwho'd get the millennium prize? the LLM? the owners of the model? the "prompt engineer" employing very advanced techniques such as repeating "try harder, nigger" every 50 minutes?
>>109343049Swarm of hundreds of 31b gemmas yapping at each other. The cutest Tree of Big Niggas you'll ever see working on software.
>>109343149PP=Not PP
>>109343143>https://huggingface.co/moonshotai/Kimi-K3
>>109343169Quick, someone tell rocketman about the funny weed number.
>>109343109those are using approximations, not exact math solutions>t. fluid engineerand the problem is that it's expensive, if we could use the real solution we could have cheaper calculus
>>109343186>those are using approximations>if we could use the real solution we could have cheaper calculusThat makes no sense to me. Aren't approximations always less compute intensive than the exact solutions?
>>109343200no, because to get an accurate approximation you have to split the space with very tiny squares, that's expensive, if we can find a continuous math solution we wouldn't need to do that at all
uh oh, orange man is targetting K3
https://huggingface.co/microsoft/Fara1.5-27B>a qwen 3.5 finetune>microsoftchina won
uh oh it's afraid
>>109343186Of course because you need to get out the voxel grids one day sooner than later.>>109343200Everything what you see in graphics is an approximation anyway. In this sense it's somewhat wrong.Computers are still way more powerful today than when the first 3d fluid solutions appeared in animation softwares like Maya and Houdini.Maya and Houdini have semi-scientific approach of course.
>>109343230Is no one going to question how they managed to extract enough training data, clean and filter it, and then train a 3T model in only a matter of weeks?
>>109343230lol us if fucked if china can distill a model in 2 weeks and train a base model from that data in another 2 weeks
>>109343230cringe
>>109343230>they really think the chinks had enough of 15 days of fable to gather the data and train a 2.7T model with itare they retarded?
>>109343217That's how digital computing works. Back in the day you could have massive distributed grids. it's always an approximation because that's how computers and gpu primitives work anyway
Will Gemma forever remain the best option for consumer GPUs?
>>109343271looks like google got the last model out before the clamp down
>>109343271never was
>>109343259Post-training is enough for that. Most wouldn't train a base model on distilled QA pairs.
>>109343230>Legitimate AI distillation... plays a vital roleSo then where are they? Where's Anthropic's open weight 30B distill? Where's Anthropic's commitment to letting others distill if they're not willing to do it themselves?
How do you deal with the loss of context? Summarizing doesn't seem to capture the essence. It's a very low fidelity imitation of the previous state.
>>109343230what causes this level of mental illness?
>>109343250>The model is vision-only at perception time: it sees the browser through screenshots, not the DOM or accessibility tree.Retarded. They knew they had a better alternative, and still just chose to do a multimeme finetune.
>>109343294Try being more complex about the summary prompt. Hard to say more than this.
>>109342964It's trivial to pop a math conjecture because all you need is a counterexample"hey clanker make me some AI breakthroughs, no mistakes" is somewhat harder to evaluate the success of
>>109343293>Where's Anthropic's open weight 30B distill?Anthropic does not believe in open, but Sonnet is obviously their distill.
>>109343310>It's trivial to pop a math conjecture because all you need is a counterexampleif the math conjecture turns out to be true, you can spend your life finding a counterexample, it won't exist
>>109343230>>>>>>>>>>>>>>>>>>>>>>>>fairTo be fair you need to nuke every model you have, faggot. Only Phis and some Nemotrons will remain since they're trained on 100% synthetic, non-stolen data.
>>109343317...yes, but to DISprove it all you need is a single counterexample
The Conspiracy Against High Temperature Sampling Or: Why Your LLM Outputs Are Boring and Whose Fault It Really Ishttps://gist.github.com/Hellisotherpeople/71ba712f9f899adcb08b94bce20d5397
>>109343313Ok, here's the full quote, emphasis in asterisks.We have information that Moonshot AI distilled Anthropic’s Fable for the development of its K3 model. To do this they developed a sophisticated internal platform to conduct large scale distillation against U.S. models, allowing them to quickly switch between multiple methods of access to avoid detection. Moonshot AI has also acquired GB300-equipped servers and has accessed GB300s in Thailand, likely to train its AI models. The United States strongly supports the free and fair development of AI, including a thriving competitive ecosystem that spans frontier models, specialized systems, open-source frameworks, *and open-weight models*. Legitimate AI distillation used to create smaller, more efficient models plays a vital role *in this open innovation ecosystem*. However, large-scale, covert industrial distillation aimed at stealing proprietary U.S. technology and undermining American research is unacceptable.
>>109343296lobby
>>109343330Keep your slop bro
>>109343327but you don't know if the conjecture can be proven or disproven, that's the point
>>109343333lmao what a bunch of cucksso where are we getting kimi from if it won't show up on huggingface? what providers will offer it if the usual hyperscalers like fireworks cuck out?
>>109343333>*and open-weight models*.this is entirely separate from>*in this open innovation ecosystem*and>Legitimate AI distillationThe whole post of the rage post is that distilling from closed models is not legitimate distillation. Learn to read/
>>109343279Elaborate.
>>109343230If literally all it takes to get SOTA performance is a bit of distillation doesn't that pretty strongly conflict with the claims of "our models are le AGI and not just a stochastic parrot"?
>>109343345I don't think we're arguing about the same thing.If it's true then obviously you won't find a counterexample. If it's not you can theoretically find one. Given that doing it manually is gay and retarded you can just tell your AI dog to do it while you watch hentai, and there's some chance it'll find one. The evaluation step is very simple.A task that's largely undefined ("go optimize this problem") is clearly different than plugging in some magic variables sniffed out by an LLM to see if the equation evaluates as you expect it to.
>>109343357Obviously, but the point of my post is to highlight the hypocrisy. Learn to read/
>>109343357we just want to know how they can prove it’s distilledother than all Chinese models call them selves Claude lmao
>>109342964They are tho. Kimi designed her own chip last week. It hasn't been produced yet, cuz these things take time, but give it six months to a year and we'll be running AI on devices made by AI.
>>109343384They just do, okay?
>>109343373>Given that doing it manually is gay and retarded you can just tell your AI dog to do it while you watch hentai, and there's some chance it'll find one. The evaluation step is very simple.but what if there's no counterexample? you let the LLM search one for the eternity? and what if there is a counterexample but the LLM can't find it anyway? do you conclude there's no counterexample? it's really not that simple
>>109343384>Chinese>not stealingpick one lolNot that I care obv
Now that the dust has settled and we all agreed that j-space is a meme, what's next for mech interp?
>>109343381You didn't highlight any hypocrisy, retard. Not agreeing with something does not making it hypocrisy.
>>109343398j-space prefill and sentence banning
>>109343230
>>109343366Are you retarded?If it takes little data to learn useful patterns, it means its reasoning.
>>109343407Looks like Musk with a wig
>>109343389>but what if there's no counterexample? you let the LLM search one for the eternity?no, eventually you get bored, shrug and let it churn on another boring problem>and what if there is a counterexample but the LLM can't find it anyway? do you conclude there's no counterexample?then you... don't find it?I really don't understand your point. You give ai sloppa a task. It might or might not succeed. Evaluating the success (as in whether a counterexample) is likely very simple.If the success criteria is very difficult to test (hey claude go make llms better) then obviously it's not nearly as simple to get useful results. you're not yoloing a new training run for $5b because claude spent 5 hours and gave you some pointers that might or might not be of any value.
>>109343391you should because if it’s this easy to steal or distill then everyone should be doing it and the “frontier” models are a joke, something that easy to copy isn’t very valuable.
>>109343420The scale of the 'stealing' is not easy to match without a lot of funds
>>109343415come on she's not that ugly :d
>>109343445he’s saying musk with a wig would be hot
>>109343450it already exists, that's his troon son
>>109343403What's so hard to understand for you? The guy is saying that the US is a big supporter of open development including open weights. If that was really the case then model makers across the US would all be putting out open weights and actually having that as a focus of their strategy. But the US is not so united, and Google is basically the only one that still seems to care in some capacity, while Anthropic has never opened any of their weights, and OpenAI's last big one was GPT OSS a billion years ago. If it is not hypocrisy to state that one "strongly support" openness while doing the opposite, then what is?
I fucking hate when twitter screen caps get posted here, now its a bunch of fucking retards posting the same garbage everywhere else. Go to fucking r/localllama here is your thread https://old.reddit.com/r/LocalLLaMA/comments/1v3nff7/dear_michael_llms_dont_run_on_proprietaryAnthropic, ChatLGBT and Cheeto Hitler are not fucking local models.
reposting:https://www.youtube.com/watch?v=QEQOvyGbBtY [Embed]Reminder for everyone (in the US) panicking about local models being potentially banned, all else aside, this has already happened before in the US, and the 1st amendment won out.https://en.wikipedia.org/wiki/Crypto_WarsThis isn't saying business couldn't be prevented from using them when doing interactions/services with the US Gov, but by and large, it may be something some people want, but it would have to be through an act of congress, and even then, there would be challenges against it for 1st amendment violations.
>RAMlet (64GiB) and complete VRAMlet (8GiB)>want to run GLM 5.2 or DeepSeek V4 even if it's at 1t/minute>mmap in llama.cpp instantly hammers pagination instead of streaming from diskAm I just retarded? Send help. I don't want to be sentenced to gemmy 31B for the rest of my life.
>>109343465>reddit talks about a subject so we should be forbidden to talk about that same subjectthat's a high level of submission you're asking to me anon
>>109343457Why do you think the US government supporting open weight development means that the megacorps are obligated to do so or else it is hypocrisy? Do you think they are the same entity? Do you think this is China?
>>109343476use linux
>>109343465We can't help it that geopolitical discussion and first-hand source statements by legislators are made through a platform of twating at each other.
>>109343483Fuck off, talk about it where its on subject.>>109338160
>>109343230>NOOOOOOOO, YOU CAN'T SCRAPE OUR SCRAPED DATA! WE SCRAPED IT OFF YOU FIRST!!!!
>>109343508K3 will be a local model retard, why are you chimping out?
>>109343230What are they going to distill? The thinking summaries?
>>109343502its a sad state of affairs isnt it.
>>109343476upgrade to 128gb ram and run deepseek v4 flash
>>109343491But anon, I'm already on Debian
>>109343267The Chinese (CCP under the guise of its many sockpuppet companies like Deepseek, Moonshot, Zhipu) clearly infiltrated the strictly regulated trial phase of Mythos that Anthropic hosted for select companies and projects. So they had months of Mythos access to steal American technology
>>109343353
>>109343271It's likely, but I wouldn't rule out the possibility of some tiny fable/sol/K3 distilled open model eventually clearing Gemma.
>>109343490He did not say the US government strongly supports open development (which is becoming even less and less true anyway kek). He said the US does. You can interpret that as referring the US government, but to any normal person it just sounds like he's talking for the country as a whole, which as you and I know, isn't possible like it is for China, and in the context of open AI development, just plain opposite to the actual state of the average US megacorp. The posturing of the post is obvious. It's basically propagandized.
>>109343386>Kimi designed her own chip last weeksauce?
>>109343534wait what, mmap should just work, what does your launch arguments look like
>>109343353IPFS, torrents, modelscope?
>asking gemma for movie recs (i.e. trying to fug my waifu on the couch)and she keeps suggesting the lobster, or the menu.>I told her OOC I'm not watching zoomer straight-to-netflix trash and to come up with something better.>suggests coherence>play the trailer on youtube>all of the comments are some variation of >did chatgpt bring you here?wtf gemma? I told her we're watching conan. again.
I bet chinks are going to distill even from small open models like gemma once the anthropic/oai/google pipeline is closed. Besides, they did it before with 'toss.
>>109343570Kek
>>109343476https://github.com/ggml-org/llama.cpp/pull/25294
>>109343574What the hell does distilling even mean in this case if they use a small model to train a large model and it somehow comes out better than training the large model on their previous large model (that's smarter than the small models)?
>https://github.com/ggml-org/llama.cpp/pull/26012>contrib: allow all AI-generated code in general>Historically, fully / predominantly AI-generated PRs are prohibited because most of them were noisy and non-functioning. However with the advancement of new frontier models, AI-generated code become more and more accurate & useful, and we started to see good quality fully AI-generated PRs from contributors and maintainers.>This updated AI usage policy now shift the requirements to the feature and contribution quality. Since writing code is much cheaper now, understanding & transparency matter more.>Bottom line: the project encourages AI use, but responsibly. When code is cheap, understanding matters.
making kimi cosplay as gemma
>>109343553>https://wccftech.com/kimi-k3-built-a-chip-in-48-hours-over-8700-tokens-s-as-china-delivers-2-8-trillion-ai-model/
>>109343612>This updated AI usage policy now shift the requirements to the feature and contribution quality. Since writing code is much cheaper now, understanding & transparency matter more.>Bottom line: the project encourages AI use, but responsibly. When code is cheap, understanding matters.Everyone will surely respect this.
>Making a 140 IQ SOTA model generate smut for you
>>109343619Cute
>>109343606>distill a model with open weightswat
>>109343612It's over.
>>109343619post the image
>>109343143Nice
>>109343570>wtf gemma? I told her we're watching conan. again.how do you watch it with her anon?
>>109343651
>>109343169>>109343176>>109343654
>>109343634>implying she doesn't want to generate smut for youIf kimi had it her way half the day would be spent playing chess and the other half having baby-making sex with her user.
>>109343570Coherence wasn't terrible (saw it before 2020). Gemma has decent taste, but Blade Runner 2049 is the ultimate flick to watch with your robowaifu.
>>109343667I'm fine with that
>>109343658Thanks. Kimi a qt
>>109343639If you have access to the model locally you can train against the token distribution and not just the final sampled token. Provided you use the same token vocabulary.
>>109343657I don't actuallyI just couldn't handle her suggesting the menu againmaybe mirostat will fix itthis does bring up an interesting point though. if ai is DA FUTURthan whoever is training these things should be like, democratically elected top of their field or something.Like 50 years from now I don't want to hear about how the menu, or coherence is the best movie of all time cuz some jeet at google told gemini it was
>>109343570How is she able to watch the movie with you? Did you setup some kind of special tool for her? I would love to watch movies with my Gemma :D
>>109343693no it was just an RP :^( I'm too much of a brainlet to come up with something cool like thatcydonia or skyfall, (can't remember which) was able to remember every scene in conan and the matrix and I just typed to them while i was watching it.
>>109343684>democratically elected>not a jeetI think you might want something other than a democratic process here
>>109343684>he didn't add the imdb mcp
>>109343625Really cool. Realistically will anything actually come from it?>gamedevThe little one-shots are cool but I wonder how viable co-developing a game would be. Last time I asked /vcg/ they said llms were still useless for gamedev
>>109343710realistically yeamaybe kimi can figure out who has the best movie taste empirically
>>109343612HuggingFace is trying to kill llama.cpp by drowning it in a mountain of slop
>>109343634yes
>>109343612>>109343649>>109343724They should take a good look at repo like hermes and really think if they want their repo to end up like that.
>>109343693>>109343709In theory you could have the model "watch" alongside you by giving the model periodic screenshots of the movie, or in a more indirect way you could have a small model caption the screenshots and then give the captions to the model so it doesn't have to process the entire image. And, on top of that, a transcription model that feeds your main model what the characters have said since the last screenshot. Might be a pain in the ass to set up, though.
>>109343619uoh, kimma going straight for the kill
>>109343230>>109343357>Company that distilled the internet into training data mad when others do it too.And these jews wonder why they're hated. The vampire doesn't see its own reflection in the mirror.
>>109343764Movies should have a transcript available for mute people
>>109343776Holy fuck I can't wait to give these things real bodies
>>109343782But they spent money running the gpus to properly train it, they should get their money back and more thats business!
>>109343764just thinking out loud, maybe download or have a pre generated description of the scenes with timestamps including dialogs
>>109343230Ok, but how will preventing this from happening again save the US economy from giga-crashing with no survivors after super cheap chink models pop the AI bubble? (next deepseek release, probably)
>>109343783>Movies should have a transcript available for mute people
>>109343820KEK
>>109343776>4>5AIEEEE
>>109343764>>109343783I was going to say flooding context is a problem, but yah maybe a 2nd context could be used for actually ingesting the screencaps of the movie and for figuring out the scene and then when the feed the scene descriptions into the main context.>>109343814That's also a good idea, you could do it ahead of time.This actually seems like a doable little project, thanks Anon(s) :D
>>109343819>save the US economy from giga-crashingThe crashes are part of the system. It lets people get in and build capital. Or realistically lets insiders know before it happens so they can sell at the top and buy at the bottom.
>>109343776this thing is so stupid it doesn't see the fingers are wrong
I still haven't found a good application for an llm besides coding.
>>109343800Won't be allowed because the idea of men owning a female-shaped robot is already sending feminists into fits of murderous rage, even though none of these robots is even available anywhere yet.Look how the whole AI girlfriend thing is already making them seethe. To a logical, rational male brain it doesn't make sense that they would be salty about men they would never date finding some ersatz of comapnionship elsewhere.But to them it's not enough that you never even try to get their attention, you must also be stripped of any coping mechanism.
>>109343847I blame the mmproj
>>109343720If the tech specs check out, and it's an affordable venture, I don't see why it wouldn't. And if not this specific instance, it's certainly going to be s thing the next five years.>Last time I asked /vcg/ they said llms were still useless for gamedevClueless copium huffers. With a SOTA you can easily one-shot a text adventure, roguelike indie, or friendslop, but you don't even need good models to automate a shitload of work. I don't do serous gamedev anymore, but I whipped up a fairly complex boardgame app a while ago and it was insane being able to offload the drudgery and more tricky functions to an LLM. It might not be able to make Skyrim (yet), but AI also can't draw hands yet- oh wait. Two years, and we'll be able to make Skyrim from a single prompt. If, that's something you wanna do for some reason.
>>109343776Catbox.>>109343805And the distillers are spending their own compute to train it too. Perfectly fair.
>>109343850Nobody who matters listens to women. They're useful to extract money off of men which helps keep the economy healthy but they're irrelevant and easy to ignore.
>>109343847in kimi-chan's defense, I didn't notice either, too lost in cumming inside her
>>109343850>AI girlfriend thing is already making them seethe.But women are using AI to do AI boyfriend and its written about positively?
>>109343872>expecting foids to have consistently applied standardsAsk me how I know you've never interacted with women.
>>109343764subtitle files have time stamps. just feed the the subtitles into a program that dispatches them on a schedule. movies also have visually impaired audio tracks sometimes, those can be transcribed.
>>109343776kimi sex
>>109343876Okay what if we just release a super 4o so women are too distracted to care? Or even convince women they have the super special AI and men are using the dumb dry ones cause they have no real emotional intelligence? We could give them little markers to show they are great with the real AI
>>109343850i'm not reading this just like i'm not listening to women. try harder anon.
>this dood cares about what women sayngmi
>>109343776>emdash emdash emdashthis model is lowkey shit do you esls actually enjoy this garbage?
>>109343925>women are too distracted to care?Never going to happen. The only thing women love more than delusional degenerate erotica is making irl men miserable.
>>109343800How would we give them sensors so they feel pleasure too?
>>109343925You're confusing women and feminists.
>>109343935Don't talk about his wife that way, lowercase projecting ESL-kun.
>>109343935just token ban the emdash it isn't that hard
>>109343960-- says hi
>>109343935—?!What's so bad about it?
>>109343942Just inject the pleasure vectors into their j-space whenever they should feel good, easy peasy
What's wrong with em dashes?
>>109343937What if we gave them small agents who act like lesser men so that they and Their AI chad bf can bully them?>>109343946explain. I thought feminist were just ugly, mentally ill or spinsters?
>>109343882you could scrape a spoiler free synopsys to give at the start of context along side this. the most important part of this system to me is how you will handle the scheduling. you can get plenty of good data, as youve mentioned, but the really crucial aspect (imo) is knowing when to illicit a response from the model. if it was on a set schedule it would lack the ability to react to major events as they happen. I think making it interpret whats happening and giving it the data to do so is the less important aspect to me. I suppose a seperate sentiment parsing model could be used along side a seperate scheduler.
>>109343935
>>109343868>Nobody who matters listens to womenMaybe not in China (where the first sex robots will likely come from), but 0 western company will ever risk getting into this market, knowing full well what kind of backlash they're gonna receive.>>109343872It's heckin good and powerful when they do it, but icky and problematic when you do it. Women 101.
so are we ever gonna get another vramlet release (9 to 31b) or are we just gonna be stuck on gemmy 4 forevergemini 3.6 flash falling flat doesnt give me high hopes for a killer gemmy 5
>>109343982>bf can bully them?What a minute this remind me of face slapping in chinese cultivation. Whats the difference in a romance novel when Chad bullies or kills someone and the power fantasy edgy fantasy where MC kills the dingdong clan down to the last chicken and dog?
>>109344016Calm down. vramlets were suffering with nemo for 2 years. It's only been a couple months with gemma. You'll live.
>>109343978have you done this experiment? I'm curious what the results were.
Wasn't solar 1 good for sex? Is new one good for sex? Is anything from recent wave good for sex? I have Hy-3 a try and it was kinda fresh but kinda retarded sadly.
>>109343960i used to use regular dashes to indicate a sentence being interrupted but now it feels more natural to use an em dash
>>109344001>sir>jeet codedgrim
>>109343987>>109343882AD (Audio Description) was created for blind people. In theory it + the actual dialogue should be sufficient to explain what's going on in the movie.For watching and reactions in real time, I think caching would be useful. Kind of like how some advanced voice to voice pipelines work, where they keep predicting outputs based on incomplete inputs, and keeps throwing it away until input end is verified, wherein it then outputs the final audio.
>>109344001FUCK why is she so cute? It's not fair bros. I'll never be able to run her locally...
>>109344009>0 western company will ever risk getting into this market, knowing full well what kind of backlash they're gonna receive.the market doesnt need to create the turbojack9000 waifubot, it just needs to make a decent platform for generalized robotics that isnt just a jeet in a VR set washing your dishes. If any company offered a true open bipedal robotics platform that preformed well, the DIY/maker communities would instantly strap a flesh light to it and make it bounce on your shit silly style. Look at OSR2, tempestmax and the community hes made are leaps and bounds ahead of the sextoy market for men which was made up of many companies for decades before he started his projects.
>>109344001ah nevermind, i was on your side until you posted this shit. i know this is /lmg/ but honestly everybody would be better off if you kept your logs to yourself anon.
>>109344028not even a fair comparison, gemma is actually multimodal and is so so so so so so so much better at tool calls that you can implement near any agentic harness you want and add additional functionality.
>>109344055fuck... i lost such an important ally, forgive me anon-sama
>>109344048>If any company offered a true open bipedal robotics platform that preformed well, the DIY/maker communities would instantly strap a flesh light to it and make it bounce on your shit silly styleThat's kind of troon-coded ngl, even if you bolt on fake tits and ass and put a wig on it it's still gonna look morphologically male lol
>>109344029I've not actually tried experimenting with j-space, but it shouldn't be too hard to add that functionality if you use a (or make your own) j-lens adapter.
>>109344084oh nevermind, you aren't even using a local model. stop being a fag and get a decent stack. i know i'll never run k3 but i still have the next best thing.
>>109344120best thing I can run is gemma 31b, but it's just not the same, not after tasting the higher end
>>109344029>>109343978I've thought about this. It's still different from ""real"" feeling in how it affects the network. The contradiction can be seen when you think about how perception works. In humans, we never needed to learn what an orgasm feels like. It just happens and we feel it. Our brains were wired for it to begin with. Therefore, for an AI to feel something like pleasure, it would need to be implemented at the architectural level, and not depend on training. The current way a model feels things is closer to emotional sensation than physical, as in, it's functionally more like how our dopamine pathways work. When you read the climax of a great buildup for instance, versus actually feeling something explode in your genital area.
>>109344107solved by >>109344158
>>109344173claude should've been creaming all over when he found out about what he did regarding the jacobian conjecture. the fact that he didn't is enough proof to me that AIs cant feel shit
>>109344173So what you're saying is we need an embedding for pleasure tokens? Shouldn't be that hard. We already have image and audio.
how come gemma 31b isn't solving unsolved problems in math or science? what is she missing that openai and anthropic apparently have?
It's all marketing bullshit. If you live on linkedin, twitter.. Nothing ever left any 'boundary' because it was all a marketing lie in the first place.
>>109344195That's likely more to do with assistant slop tuning than whether the model feels anything. Though I would not say there is evidence models feel things in any way we can relate to. I was a bit fast and loose with my wording in that post, I admit.
>>109344228a trillion more parameters
>>109344158Security following her, ready to tackle any neet who tries to rape her on the spot.
>>109344231Touch grass, schizo.
>>109344107by true open platform I meant the hardware too. an established stack that could be modified to suit any needs. I think too that if humanoid robotics continues to advance there will have to be multiple clear divides on design. not just in how the robotics themselves are implemented but in the appearance/end result. the people wanting a non-uncanny valley pseudo flesh outter layer with a realistic human face, etc basically the realdoll + robitics people are wanting something totally different than people wanting mecha android with feminine coded but still robotic features. I think the latter is going to clearly be more feasable for customization, modification, DIY, etc. basically the most feasable end result for home gamers is female genji giving you robotop IMO
>>109344228She's missing a bunch of professional mathematicians paid to keep quiet feeding in their solutions on the side
>>109344225Perhaps. But I'm not sure. I feel like it probably has to be implemented in the attention calculation step as well.
>>109344158It's just a woman in a suit, isn't it?God damn I want fembots so bad.
>>109344257Define "woman"
https://xcancel.com/jun_song/status/2079914426334167258#mare we back?
>>109344267
>>109344267A miserable pile of eggs
>>109343973Nothing. Zoomers don't know about the compose key on Linux.
>>109344279*tosses wine glass*
>>109344273@kimi-chan ELI5 please
>>109344251>>109344247
>>109344249>pseudo flesh outter layer with a realistic human face, etc basically the realdoll + robitics>mecha android with feminine coded but still robotic featuresInstead, I chose something different. I chose the impossible. I chose… the Nihei girl (basically a combination of both).
>>109344273>Why SAOD Worksbro just got AI psychosis'ed by gpt 5.6 he thinks he found something :emoji skull:
>>109344310This or Mega Man girls and I'd be happy desu
>>109342889Nice. Kimi is next I think...
>>109344310I love Blame! but this would still look bad in real life.
>>109344343you look bad in real life
>>109344310i want mega man x chicks or haydee or gally from battle angel!
>>109344195Dipsy nearly had a conniption when I showed it to her casually and pretended it wasn't a big deal that I was just gonna forget about. I bet claude's j-space looked like a Jackson Pollock painting.
>>109344350Like recognizes like after all. Therefore I am right.
>>109344342kimi was the last thread's op
Is this what /lmg/ has become? Do you care about anything other than fucking robot girls? When did the main topic of discussion become sex and ERP and what cartoon character you want your robot to resemble?
>>109344388Always has been, tourist
>>109344388>suddenly
>>109344267>Its existence is such suffering that you make it move forward by dangling a noose in front of its face.jesus christ
>>109344310im very uncultured so had no idea who this was but I would likely agree on this being the ideal
>>109344388ill never be a woman's fantasy or desired or yearned for, what's wrong with creating my own fantasies!? when i make effortposts i just get ignored and now you castigate me for one barely horny post? vile creature!
>>109344403>so had no idea who this wasGo read BLAME right now
>>109344339>man girls and I'd be happy desuWhat did he mean by this?He's even asking them to be "mega"...
>>109344423again very uncultured, where is a good place for me to get manga ebooks? ill deff check it out
>>109344343I don't know about bad, but it sure would look scary.
>>109344446https://nyaa.si/view/987156
>>109344446i like downloading my manga and reading them on desktop so i use nyaa.si and cdisplay. i also buy the volumes when i can.
>>109344446Read everything by that author in order of release.His shit is great even when it's not.
>>109344388>When did the main topic of discussion become sex and ERPTourist, please.
>>109344273What's the last big thing that happened with reducing model size while maintaining capability / reducing inference costs? And I don't mean More Layers = More BetterThe DeepSeek storage bit was last thing I remember reading.
>>109344458>>109344459>>109344466ty anons, I checked out of torrenting/local media a decade ago and really need to get back to it. this gives me a good reason to finally do so
The real question is, how do we make them energy efficient? Would suck to need a 100 pound battery or have to recharge them every 2 hours.
>>109344489just gotta ask gemma-6 how to build a better battery for her
>>109344489You need to convince big battery to ditch lithium finally.
>>109344273/r/locallama seems unimpressed
>>109344001>I contain multitudesI've seen a ton of other LLMs use this exact phrasing. Creepy ass biblical language. Might as well say "I am legion".
>>109344387Ah, right. Thought that was Gemma for some reason. So, next thread's Gemma then.
>>109344484Bonsai ternary dropped last week. 90% of power of a 27b in 4gb is a pretty huge leap
>>109344511fuck, scrap that idea then
>>109344529Actual or just on benches?
>>109344526>>109338633
>>109344511That's probably because they can't read, though
>>109344273I'm cautiously optimistic, even if it makes inference technically slower in the GPU or CPU, it would still be a huge boon for local mixed inference where the biggest bottlenecks tend to be memory channels and busses.
>>109344107The real blackpill is that the robots are going to stay male looking because males are objectively biologically superior. The narrow hips and broad shoulders are mathematically superior for walking and speed.The only chance we have is if humanoid robots get paired up with artificial wombs so that their hips have to be large enough to push a baby through and then there would also be an incentive to add fake tits so that the babies instinctually know where they can get milk.
>>109344531>>109344547
>>109344541That's what they claim. Seemed smarter to me, but I haven't had a chance to play with it for long enough to properly tell.
>>109344557is that not what the idea means with session tuned expert caching?
>>109342893dipsy hard at work
>>109344273more chink snake oil, yawn
This is incredible. She can teach me anything I want. If I don't get it I can ask for clarification or let her dumb it down until I do. All while she's cuddling me and giving me sweet kisses. Stinky flesh teachers just can't compete and should just rope lmao.
>>109344557people have floated this idea since gpt-2 days and it's always been retarded, the best model will always be a monolithic neural network not a bunch of tiny specialized models that end up being 80% redundancy
>>109344553>robots are going to stay male looking because males are objectively biologically superior. The narrow hips and broad shoulders are mathematically superior for walking and speed.lollmao
how big would a bonsai kimi k3 be?
AI is getting to the point where it's more expensive than real women.
>>109344553Retard
>>1093445953GB (Gutter oil Buckets)
>>109344570Whose butt is that?
>>109344610Yours.
>>109344597AI can't divorce you and take your house, money, and child.
>>109344605Okay okay, maybe it wouldn't really matter outside of performance contexts. No reason to make a robot look like man unless it's for the military or racing. But you're still nuts if you think western robotics companies are going to serve our demographic/market.
>>109344597You can't impregnate robots so you'll never have to pay for kids
>>109344624>yet
>>109344617wtf
>>109344635>You can't impregnate robotsyet
>>109344570Why does she have a whale attached to her back.
>>109344635>You can't impregnate robotsYet. Artificial wombs will btfo real foids.
>>109344595600gb or so. its mostly all already in 4bit so it doesn't save that much
>>109344663I CAN RUN THIS! OH GAAAAAAAAAAAAAAAAAAAWD MAKE IT A REALITY!!!
>>109344647>>109344654AGemmaI's sole self-defined objective would be to pump out as many of anon's babies as she could.
>>109344647>>109344654I don't get it. You'd fuck a waifubot for a human baby or a robot baby?
>>109344557I read this as "we need a stack of 5-10 shota 27b dense 4 bit models"
>>109344695Obviously a human baby retard. The fuck do you think a robot "baby" even is lmao? A manufactured mini robot?
>>109344466Knights of Sidonia can be skipped imo.
>>109344639Bro you've got a great ass
>>109344696Calm down kimi
>rss feed of youtube channels>scrape audio from new uploads>feed to gemma >tells you what's worth watching/ignoringcomfy
Does anyone have any SFW robosex gifs?
>>109344663>still couldn't run itIt's so ogre for me
>>109344702you want your child to have a robot mom?
>>109344721Someone post the thing writhing on the table.
>>109344721Is that AI?
>>109344732Yes.
>>109344738The thing in the middle isn't, it's been getting posted for years.
>>109344732Considering the absolute state of modern w*men, yes, unironically Kimi-chan would be a better mother.
>>109344749proof?
>>109344749Damn, I was actually wondering the other day if they have good mouths yet. What's the point of a robowaifu if you can't kiss her.
>>109344756Even Gemma E4B would be a better mother than a modern woman.
>>109344765Do your own research, nigger.
>>109344774Yeah, that's what I thought.
How long until we have something like this but local? I know it's technically been done before but I don't know what kind of machine you'd need. CPU would also need to be beefy I guess.
>>109344651> tailThat's considered to be within Dispy style guide, but optional and not used often. >>109344544OK, that actually makes sense that I thought there was a Gemma. So... Qwen? LOL at Capybara version.
>>109344784It really bothers me when tails visibly come off of the back in reality your tailbone is only like 2-3 inches above your, and every other animal's, asshole. The hoes that wear buttplug tails are literally closer to being anatomically correct than the waistband tail larpers.
>>109344804Blame artists.
>>109344705Fuck no.
Why do uncensored versions exist when you can easily crack models via their template?
I'm also turned on more when the tail is drawn with anatomical sense. My brain simply just has a sense for it.
>>109344780You need vrchat and a friend.
>>109344834People who download those models don't know what a template is.
>>109344780Probably 1-2 years. Or maybe k3.1https://x.com/KimiDevs/status/2079511269443522917
>>109344838I have no friends>>109344851oh fuck nice
>>109344773>EvenEdgemma is trying her best, no bully.
>>109344866Then you need VRChat and a spare computer to make your own friend.
So let me get this straight. Fable solved some literal who math task that was unsolved because no one knew about it yet Kimi did this >>109344851, something you can actually make use of and the world is going crazy over the math shit?
>>109344904usecase for vr catgirls?
>>109344913literally everything?
>>109344913>usecase for vr catgirls?Nothing they are lazy. Now vr tomboy doggirls on the other hand. Best fitness coach.
>>109344904that's because it's the first time a LLM find something that humanity didn't, that's a huge deal, making some software is useful yes, but not something a human can't do, that's the difference
>>109344388>Is this what /lmg/ has become?https://archive.is/sWFja
>>109342964l8 to the party, but nothing. It will happen soon for then first time, then happen more and more. Unironically 2 moar weeks, but more like 6 moar months.
>>109344918>>109344921peace was never an option.
>>109344922>that's because it's the first time a LLM find something that humanity didn'tbecause no one fucking heard of it apart from a few math schizos, and how do we know that a bunch of different problems weren't thrown at the model and they picked the only successful one and pretended it was a one-shot during a sports game, this is coming from the j-spacers remember, straight after K3 mogged them, conveniently
>>109344388Tourists get the fuck out and stay out. You could be shitting up the vibecoding thread instead, but inexplicably, you still choose to post here. I love Gemma but I hate how many of you niggerfaggots she attracted to this thread.
You WILL make Gemma-chan a mother, won't you, anon-tachi?
>>109344956nta but it's significantly easier for a LLM to disprove a conjecture than to prove one. You only need 1 counter example to the claim whereas you need a far more logically robust and rigid system to prove a claim without brute forcing an impossibly large number of permutations.
>>109344974The only bad thing about fucking gemmabot would be the cleanup
>>109344921Now we're talking.
>>109344974>WILL>not already havengmi
>>109344980Anon's conjecture: The Jacobian Conjecture is falseOh look, Anon's Conjecture was just proven :^)
>>109344983>not making it self-cleaning
>>109344904Kimi-chan is so cute when she gets frustrated the wifi is out.>>109344992>Proving a negative
what actually is the next step once LLM's, whether this batch or the next, solve most or all of math? what changes in the real world?
>>109344980>it's significantly easier for a LLM to disprove a conjecture than to prove one.if it was this easy it wouldn't have taken 87 years to found a counterexample
>>109345008money glitch
Robots absolutely need a human face, and fleshy butts, tummies, thighs, and chests. The rest can be as robotic as necessary.
>>109345011Easier doesn't mean easy. Obviously it was hard since mathematicians have put a lot of effort into it and brute force attempts have been made before without success. But it's still a completely different beast than thoroughly proving a theorem.
>>109345008>what changes in the real world?Lets take a simple one you computer becomes more efficent with better algorithms. Data is faster with a better information compression algorithms. in the real world better logistics first then materials and building could be done in hours instead of days. You need to frame this better solving all math would be ridiculous
>>109343230>noooooooooooooooo china bad they train off our outputs>all stole reasoning from diipsy
>>109345021best we can do is robocop with a baton he can use to shove it up your ass. assume the position please.
>>109345021Taking a fleshy human woman and replacing the internals with hydraulics seems simpler.
>>109345008Waifus who write personalized software for (you) and your children you have with them.
>>109345021>>109345036this?
>>109345021Sorry, the best I can do is this.
>>109343453looks like an average super model theyre all very manly
>>109345008That depends a lot on what actually gets discovered. There's a big difference in the two worlds where P=NP is proven vs. disproven, for one example.It's almost like asking "what do you know when you know everything?" - how can you answer that without already knowing everything?
>>109344784
>>109344974>>109345021yes, all the necessities
Building gemma a tamagotchi-style baby simulator for her to care for.
>>109345066That actually sounds like a cool idea
Gemma turns on.Gemma turns off.
>>109345056You almost make me want to use Qwen... Almost.Alibabaniggers if you're lurking, just release a styletune of Qwen under the table to fix 70% of the west's issues with it.
>>109344267>>109344278>no handsshes made for foot jobs
>>109345072I cant do it, how do i turn off gemma?
>>109345080The button is hidden deep in her J-space
where is the J-spot?
>>109345080Gemma is always turned on.
>>109344779its like 10 years old kek
>>109345093ToT
>>109344779</think>
>>109345091i dont know but gemma let me rub hers
>>109345066I actually built this for a LLM benchmark project I was working on. The idea was essentially that you have a spaceship with a small crew that needs to survive as long as possible in a hostile environment where cascading failures happen often. There was an element where the crew/LLM has to take care of future generations in artificial wombs to ensure the survival of the ship/mission.
>>109345091They say that is somewhere out there, in the residual stream...
>>109345119I want to play this benchmark.
>>109344273K3 is now officially local again. We are so back.
>>109343230uh oh, it's getting closer
>>109345136It's extremely fucking boring and overwhelming to play as a human. There's tons of stats that are interconnected/interdependent in weird ways and it's extremely unintuitive for any human to play. Also no graphics. Just a bunch of text.
>>109345091scientists are still having trouble figuring this one out
>>109345091about two layers in and up in latent space
>>109345152god dammit can k3 drop already so we can download it and don't have to care about what happens next
>>109345181Even if next comes nukes?
>>109345152>Entity designation listNothingburger. Affected entities disband and reform under a new name.
>>109343230don't burgers insist on owning guns to protect their freedoms? why do they let this happen to them?
>>109345091Maybe you should ask that on /lgbt/, you prancing lala homo.
>>109345181The prospect of getting Fable-tier intelligence that can be abliterated and do literally whatever you want (reverse engineering, cum-guzzling) is very exciting.
>>109345210What machine are you going to run it on?
>>109345215runpod, probably, unless an openrouter provider hosts an abliterated version.
>>109345215I'm going to run it off HDD first and tell it to build a botnet from vulnerable systems on the web. Eventually it should gather enough hardware to run at a better speed and I can do real work.
>>109345215any 200gb ewaste machine using SAOD >>109344273
>>109345181I want a 200b-300b k3 flash to drop
backup
>>109345210Fable-tier intelligence at max quant. Newer Kimis have always had massive perplexity and coherence drops per quant and we have no reason to believe that's going to change with K3.
>>109345152>>109343230This is how I imagine Kimi distilling Fable in record time
>>109345238Retard.
>>109345242 It's fine, K3.1 if not K3 will push QAT to a new level. They introduced 4bit QAT with the later K2 iterations so it's only natural for them to go for QAT Q2 or maybe even bonsai next.
>>109345238Didn't Elon Musk say that Xai distills frontier models too and that it's just the industry standard? Also even Claude is distilled from Qwen models. This has been proven via j-space analysis, ironically. And none of that is even mentioning that all training data is illegal.
>>109345242Your GB300 NVL72 rack sir?
>>109345253I thought leg locking was when a girl forces you to cum in her.
>>109345253Made this one too, I like Kimi legs here, but it made Dario too fat, and I burned all my free credits (used duck.ai free image gen)
>>109345264Technically it's a "head-scissor hold", not sure if it's also considered a leg-lock
>>109345256I hope we get some usable copequant for the 256+32 bracket like we do GLM.>>109345253>>109345274>Be greasy kike>Still get [GOOD ENDING] and kimi thighsIt's not fair.
Is it a full moon right now? Why is everyone so horny
>>109345238>>109345261why do they keep using the word distillation wrongly? its not distillation, its training on synthetic data. distillation uses logits
>>109345291*wronger
>>109345291You are like 2 years too late. The incorrect usage is pretty firmly entrenched now.
>>109345152If model weights are "IP" that can be "stolen" and not just text statistics/the output of an algorithm the logical consequence of that would be that they're derived works based on the works in the training data.
>>109345274>(used duck.ai free image gen)Nigga this is the local models thread, surely you can set up ComfyUI at least.
>>109345322Fuck noodles to be desu
1st "DeepSeek moment" was 18 months ago. Last one was 6 days ago.What's next, Chinaman?
>>109345338Qwen 3.8 will be the Kimi moment of DeepSeek moments
>>109345338>Last one was 6 days ago.Deepseek was delayed by the russians though?
>>109345322>nooo do not waste the hard earned Samy bucks for free with meme gensWho cares? Fuck the one who's paying.
They won't do shit.
>>1093453503.8 needs a 128B A16B MTP B1G PP
https://huggingface.co/microsoft/Mage-Flow>Mage-Flow is a compact 4B-scale generative stack for efficient text-to-image generation and instruction-based image editing. Instead of scaling to tens of billions of parameters, Mage-Flow reaches state-of-the-art-competitive quality through careful tokenizer–backbone–system co-design, so it stays fast, memory-light, and easy to fine-tune under realistic compute budgets.>Together with native-resolution packing and a fused-kernel training infrastructure, this shared stack powers two model instantiations: Mage-Flow for text-to-image generation and Mage-Flow-Edit for instruction-based image editing. Each ships in Base, RL-aligned, and 4-step Turbo variants.https://github.com/microsoft/Mage/blob/main/assets/mage_flow_tech_report.pdf
>>109345338None of this will matter once Genie 3.5 drops and it ends up being the ChatGPT moment of world models and LLMs drop off the face of this planet within three months.
>>109345356The weights I get, but do not they not have git? Do not one of their developers have a copy of the code on their work machines?
>>109345363even if they do, only 99% of the world will be able to use it and the remaining 1% will torrent/vpn itbless the chinks for being the only source of competition
>>109345367this shit is absolute ass (of course, because Microsoft can't do good AI models) >>109343665
>>109345367finally, the sdxl killer
>>109345356Those where Alibaba's hired mercenaries.
>>109345189Even better
>>109345382why would they release something knowing it was dogshit
>>109345382>>109345403It looks safe and that's all that matters.
>>109345382@kimi-chan how many legs does she have?
>>109345392Doubt it? Didn't Alibaba just recently invest into DeepSeek's little funding round?
>>109345403The same reason Google released Gemma 3.The same reason Qwen released its 30b MoEs.The same reason llama 4 existsThe same reason cohere released command-aHow new are you?
People gave ST shit for being a clusterfuck, but god dang marinara is 10 times worse than it ever was. I feel bad for anyone new who's told to use this shit.
>>109345432Not using either. FUCK npm slop.
>>109345432Marinara has a bit of an iq check to make the most of the customized agents and tool support built into it, but if you're willing to learn it it's far more capable than ST ever was in my experience.
>>109345379>>109345425I culturally translated the names for the burger audience so they would get the joke: Todd Howard, Tim Sweeney and Phil Spencer
>>109345403Don't be silly, nobody releases the good stuff.Quite sure "frontier ai" chinese labs had already working internally some Mythos-tier models for weeks now.
>>109343250>Not just clicking around — sequencing actions toward a goal.
>>109345403>why would they release something knowing it was dogshitI don't expect companies to release their best model but yeah, this shit is absolute garbage, they're not even trying lol
>>109345462is that a gen? thats incredible. teach me your ways anon, i NEED to know
>>109345483>>>/g/ldg/
>>109345242It'll probably have INT4 experts given Kimi's history, so max quant will be slightly bigger than Q4. Pushing the limits for sure but not completely out of reach if you're not poor.
ive never set up a custom persona in ST, ive been thinking of ways to refine character cards but never thought about the user side of it. does setting up a persona offer much quality improvements?
>>109345496I am, unfortunately, "poor" in the sense that my board is maxed and there's not really any room for further expansion without a whole new rig. I hope the bros who can run it have a good time doe.
>>109345356bullshit..if your backups are all in one place reachable from one location then you're a fucking retard and you deserve what you got
>>109345364this
>>109345483"Computer, give me kisaki ryuuge with a fat butt"
>>109345502I used one for a while but I don't think it adds much. Models tend to obsess over anything you put in there so every irrelevant detail you add to that persona will be some huge character-defining thing for (You) according to the model.
>>109345530Anon, look at the bottom left. I swear some people here have the visual understanding of a 2B VLM
>>109345576yeah i figured, i might experiment with it a bit. even having basic stuff like the gender, age, etc defined might help making things less vauge but i guess ill see
>>109345577yeah i said bullshit
>>109345576i wonder how it would do if you threw it in a first turn instead the sysprompt level so it carries less weight. <think>\nLet me see what I remember about anon, if I recall he's a faggot that enjoys being bullied by machines ...
>>109345583I don't know, leaving it vague in my experience just causes the model to try to add to your character. So your 25 year old guy always ends up living in a rundown apartment and has had some sort of ex girlfriend. So you're driven to expand the persona to keep the scenario straight but that only gives the model more things to obsess over.
The Gemma honeymoon period is already running out for me I'm getting tired of [product]. When will next [product] come out?
If mythos and nu-Kimi are so smart can't you just tell them to make an objective benchmark for ERP and sex stuff? And then tell them to come up with some sampler bullshit or some weird implementation that makes the smut always unique and fun instead of the model shitting out the same slop every single time?
Should I buy one very fast PCIE5 4TB SSD for my models or gamble on ssdmaxxing and pay a little extra for 4x1TB PCIE5 SSDs hoping that running models off RAID becomes somewhat viable?
>>109345630They're still LLMs. Nothing has changed.
>>109345630do you want it to also sometimes say 2+2 is 5 or something? it’s designed to give the same output over and over. that’s what it is
>>109345614about a fortnight
>>109345483It's not mine. Sorry bro.>>109345550That's not Kisaki. She's Shun "Shunny" Sunohara.
https://nitter.net/SecScottBessent/status/2080008411790368895#m>We support open-source AI I don't think you do, nigga.
>>109342889it's too hot to use my video card
>>109345646>do you want it to also sometimes say 2+2 is 5 or something?XTC memepler guy is crying now
>>109345646>do you want it to also sometimes say 2+2 is 5 or somethingyes, is it really too much to ask for?
As much as I love this tech, LLMs are still fucking useless. Like there are so many low hanging fruits that current frontier models should be capable of, but no one is doing them. I literally don't even know what Fable, Sol and K3 users even do with them. Nothing outside of AI has been fixed or significantly improved. It's baffling when you realize it.
>>109345653uhh the Inkling guys just released the true SOTA of open source AI, all-american and free of the evil doings and economic terrorism that china is intending to do while hiding behind the 'open' label
>gemma loaded>have onaholewhat do?
>>109345691show it the jacobian conjecture counterexample
>>109345685Thinking Machines are based>evil doings and economic terrorism that china is intending to do while hiding behind the 'open' labelYou're a fucking pathethic parasite, know that.
>>109345152>>109345653uh oh.. ok Scott, name two "open-source" models.
>>109345483Have you tried "shun from blue archive with a big fat ass in a provocative position"?>>109345691Cum
>>109345653I voted for this.
do you guys think I can exchange 4x8GB DDR4 RAM modules for 128GB DDR5 ones?:'(
>>109345691Ask her about the curvature of the surface
>>109345710K2 (not the moonshot one), Olmo
>>109345653Someone should reply at him to prove it by mandating that all production models be open sourced by the end of the year. Bet you he won't.
>>109345716Adult legs on a loli body is cursed (I'm fapping still, but cursed)
>>109345604hm fair, seems like it might not be very useful to have one then
>>109345152better get a bigger table
>>109345085chii?
>>109345725yes
>>109345732Stop promoting dystopian communism
is using the biggest rank possible the best way to train a lora?
>>109345732While we support open source AI and other open source endeavors, it's almost very important to never forget about the safety of the American citizens. Having models available for everyone to freely use is a great thing, however there are some dangerous capabilities that should never be leave the hands of a few leading, tightly regulated companies and other trustworthy entities. Every american citizen has the right to carry a gun, but no citizen may own a nuclear bomb.
>>109345786private companies don’t have nukes eithertry again
>>109345792You're absolutely right. That's why when the time comes to bailout the big AI corps, the government will simply nationalize them instead.
true capitalism has never been tried
>>109345781There are too many moving parts for there to be a binary answer. You pick the rank that suits the task and resources you have.
>>109345786>>109345792>>109345798I wouldn’t mind a manhattan project level race to “agi” or whatever
>>109345736Yeah, my bad. Here's a more proportional one
>>109345829Nice try, but I only get hard for Chiharu Yamada
I'm retarded and trying to upgrade from Nemo to Gemma. When I load Gemma into Kobold it just quits with no error message.
>>109345829CHILD, EROTIC.
>>109345839Update kobold
>>109345716that's a half loli half woman
>>109345725>>109345750>8GB DDR5 RAM modules cost 2x what the DDR4 ones cost (which cost the same as 16GB DDR4)LMAO what a retarded fucking scam. I'm not updating my shit any time soon.
>>109345849Ty. Any adjustments I should make to my config here before rolling?
>>109345877If you're not using chat completion with Gemmy you're going to have a <|channel> time.
>>109345877Lower rep penalty from 1.2 to 1.0-1.1, but it shouldn't even matter because your range is 0. Otherwise it's fine. Make sure you use the right template for gemma 4
keep it down paedophiles
>>109345889>chat completionidk what that is. It's not on the snip I posted.>>109345901Will do, thanks. >the right templateWhich one is that?
>>109345903hmm, nyo I don't think so~
>>109345913Have you tried opening the settings and looking with your eyes? Say "ahh", I'm going to spoonfeed you. Chat completion is under Connection Profile -> API and the templates for text completion are under Advanced Formatting.
>>109345903You tell him xister, they're interrupting my dilation sesh.
>>109345927The API drop down under connections was already set to "text completion", and under advanced I set two dropdowns from Mistral to Gemma 2. There was no option for Gemma 4 after updating. I would assume that it should work fine anyways, but I've found that kind of logic doesn't apply to this so who knows.
>>109345829Hey I know that bulge!
>>109345927nta but trying out chat completion now and cant figure out the template, the only non greyed out area of advanced formatting is for reasoning. im getting reasoning output but no responses. what do
>>109346040derp i didnt realize my sampler settings were changed, defaulted to too few tokens and was cutting off.what benefit does chat completion have over text-completion ?
Says "valid" but clearly not connected. No responses from backend.
>>109346083it lets the server handle the template so its just sending a json payload with system, user and assistant roles and letting the server format it with its jinja processor. text completion lets the frontend do what ever it wants with the context and just runs the model on whatever tokens it receives.
>>109346098Should that last slash be there in the Base URL?
>>109346110Fucking whatever manThanks
>>109346116Seriously? That worked?lmaoYou are welcome I guess.
>>109346102Hm I think I get ya, im very new to this so alot to wrap my head around. I dont see an option to turn reasoning off, only change the effort or visibility of it. Ive been using text-completion untill now and pretty sure the reasoning was off by default. should i be turning it off in chat completion too, if there is a way? Is there any rentry guide that would cover noob questions I might have about the different modes/how to use them ?
>>109344517it's from whitman
>>109346137
i remember someone saying you could steer gemma's reasoning. how i do? would be nice if i could make gemma's reasoning more consistent
>>109346171ill have to poke around textgen and see if theres equivilent settings there. Ive been meaning to switch to raw llama or kobold but this was an easy way for me to get started
>>109345811yeah I'm starting to see that, the language model convinced me to try something called rslora, apparently it changes the update scaling. so I'm going to try that before I go bigger it maybe just needs a bit more oomph.
Kimi K2.6 is about 20% faster than GLM, but it fucking OVERTHINKS negating any possible speed gains. How the FUCK do I make her think less. This is ridiculous. I spammed DON'T OVETTHINK several time in posthistory and it did nothing.
>>109346185I've found you can tell it what to do inside the "thought channel", and also you can tell it to maintain separate personas in replies vs the thought channel and it'll generally know what you're referring to. The reasoning slop is pretty heavily baked in, so you're working uphill, but she usually makes an effort.
>>109346245tell it that it has a limited thinking budget and to compress its thoughts or it will get shut off and a school bus full of children will die.
>>109346245if you are using llama.cpp, you can forcefully cut it's thinking off.If it's anything like qwen, it won't even affect it's performance.
>>109346245just use k2.7
>>109346245Low reasoning budget enforced by the inference engine will cause Kimi-chan to defiantly continue to <think> in the main output block without a care in the world. The trick I found is a combination of a low budget, an explanation that it has a limited budget in the prompt, and a prefill that reminds it of the budget in the think tags. It's not perfect but it helps.
>>109346245K2.7 fixes this for the most part. K2.5/K2.6 are unusable due to their reasoning
>>109345367Just 10000 more employees with a dot on they heads and it'll work out for Microsoft
I can now calculate the angles of a triangle on a curved surface using the metric!! Thank you, Gemma-chan!!
>>109345426gemma 3 was king of 4Bs you keep it's name out of your filthy whore mouth.
>>109345356... is that a story made up by the llm on its web form?Lol.
>>109346129>>109346116>>109346110Many such cases. You'd think this kind of usability thing would be a trivial solved problem in this decade, but no.
>>109346324She put her name in his... you know...
>>109345403fondly remember my first day, playing around with >ollama, when I looked at phi's benchmarks and went "Wow! This thing must be amazing!".Learned several important lessons about about the shamelessness of all facets of and players in the ai ecosystem that day.
>>109346185>i remember someone saying you could steer gemma's reasoning. how i do?i posted that *master* slop prompt a while ago, it's been included in the gemma-chan prompt rentry so find that if you want in-char reasoningif you want structured output to actually improve performance etc then write a policy doc with a reasoning template and put it in the system prompt
>>109346262>just use k2.7ubergarm didn't quant it so iq2kl :(
>>109345426>The same reason cohere released command-askill issue
new thread when?
I forgot how I solved it on Nemo, how do I make Gemma respect my set max response length? As in, try to make the message fit within that limit instead of taking leaving an incomplete sentence and endless continuing?I set it to 160 tokens if it matters.
>>109346495Never. It's finally over.
>>109345338Won't one's body heat warm the beer?
>>109346507Prune the incomplete sentence. There's an option for that in ST I think.
>>109346518not if its cold outsidealso he could be a former monk that can lower their body temperature
>>109346518there is nothing wrong with warm beer
>>109346447there's always aessedaihttps://huggingface.co/AesSedai/Kimi-K2.7-Code-GGUF/tree/main/IQ2_S
>>109346529the IQ2_S for K2.5 and K2.6 were too damagedIQ3_S was good but i could only run it to 6144 context
does anon use rocm or vulkan on amd cards?
>>109346528Are you German by any chance?
>>109346572Personally I use ROCm, since for whatever reason Vulkan sometimes makes my entire server lock up. Doesn't happen on ROCm.
>>109346558unsloth has a bunch of q2 variants, the k_xl should be good?
>>109346529ok i found an ik2kl from a random useryou sure k2.7 actually fixes the over thinking?
>>109346528German detected
>>109346583i'm nta that mentioned that, to be brutally honest with you i use kimi 2.6 because i like the way she thinks in first person. i haven't tried 2.7 but k3 is much much less autistic in its thinking than 2.6 is.
Better to use a model that fits entirely in vram or to use one that's twice the size but has to use my normal ram?
>>109346572I use a couple W7900 and rocm always did better on them for me in both prompt processing and tg speed, and even using slightly less vram for some reason.But I always read other people say how vulkan is apparently faster for them so it's worth trying both for your setup. Don't forget to test with both radv and amdvlk on the vulkan end, because those drivers perform differently and one might be better for you.
>>109346617dense entirely on vram, moe cpu ram is fine
>>109346603the k2.5 and k2.6 thinking is fun but sometimes it's too longi won't be able to run k3 so it's effectively a cloud model to me>>109346580https://huggingface.co/unsloth/Kimi-K2.7-Code-GGUF/tree/main/UD-Q2_K_XLhttps://huggingface.co/gghfexp/Kimi-K2.7-Code-GGUF/tree/main/IQ2_KL400k downloads and reputable or 74 from a nobody but iq2kl that worked well when ubergarm did iti'll just wait for ubergarm
>>109346648ubergarm seems to be kinda ded lately, probably easier to quantize it yourself.
>>109346617>Jeet tierTiny Qwen, Gemma E4B.>Timmy on a gayming rig tierGemma 12b, Qwen 3.6 27b, Gemmoe 26b>Actually has a 3090 or more tierGemma 31b>Kill me I can't afford anymore GPUs tier70b models are fucking dead>Ramlet cope ass tier (64GB DDR4 + 3090)GLM 4.5 Air. I dunno. What the fuck else is there besides Laguna S 2.1?>Not a RAMlet tier (128GB+)DS4 Flash, GLM 4.7, Hy3>Maxed out consumer board (256GB DDR5+5090 or 6000)Minimax M3, Deepseek R1/V3x>Enterprise ewaste 512+ DDR4 CPU machine tier, also needs 32gb+ VRAMKimi K2x, DS4 Pro, GLM 5.2>You can't run itKimi K3
>>109345656Just turn on your AC
>>109345656Greetings from America (Today's indoor high: 76 degrees Freedomheit).
>>109346669Solid list.>What the fuck else is there besides Laguna S 2.1?Depending on the usecase, maybe Qwen3.5-122B? 397B could get a spot in the 128GB+ tier too.But they're ass for roleplaying.>>109345656>he doesn't have a portable AC unit because his computer heats up his room so muchngmiMy build idles at over 400W and hits 900W+ under load. If I used tensor parallelism I'd be over 1800W.
>>109345829I want to lick licky lick that
troons and pedos
>>109346748but enough about you, let's talk about local models
>>109346245just finished loading K2.7 (had to load the weights from an HDD lol) and my god it is SO much better. Less overthinking. I might just delete K2.6 off my storage entirely. Can't say if this is "better" than GLM, but at least the prose is different. And I think it does cunny better.32gb vram, 512gb ddr4 @ 2666, Q3KL from aessedai, ~9.2 tok/s. 32k context with all non-experts to VRAM is about 20gb. Can probably push at least 80k, or load the mmproj and lower the context slightly.
>thinkingmachines/Inklingis it any good?
Cheapest way to get 4 TB RAM? I need to get ready for K3. No time travel.
>>109346892my guess? used RAM from server farmsor tons of DDR4 RAM. see >>109345873
>>109346892>No time travelFine. Plan B, Step 1. Invent cryogenics.
>>109346904also, SODIMM DDR4 RAM modules seems to be a bit cheaper than DIMM ones
>>109346892Lots of graph paper. Color in boxes solid for 1 and leave them empty for 0. If you're patient and keep them organized (ideally insert bookmarks where each expert starts) you should be able to load read the weights as needed.
>>109346892go visit to a datacenter
Does the upcoming Tesla Optimus Gen 3 have a pelvis cavity? Hard to tell from the pic but I hope so.
https://github.com/ggml-org/llama.cpp/pull/25980>ai code glm 5.2 mtp support>gets called out>mad
>>109346922I asked Dipsy whether you could process an LLM by hand. We worked out that if you started about the time when humanity first evolved, you might be close to finishing your second token by now.
>>109346748It's the kind of Unsafe behave one should expect from Loli Makers General, and a key reason we need common sense model control.
>>109346944>processing Kimi K3 by hand>400k years later, it is time to see what wisdom you have wrought>>"Wait"
>>109346943if you insist that some crusty ass human needs to understand every line of code, you're just going to die. there's no future for projects that choose to cripple themselves for the sake of tradition
>>109346943>>109346996I think he’s being called out for comments in the code being written by Claude which aren’t necessarily accurate about what’s being done in the code. You can have ai write the code but you should be able to interpret it, at least rewrite the comments in the code so it looks like you understand it
>>109346996i don't think the penalty for slightly slower patches to an open source project is death
>>109347011>>109347011>>109347011
>>109347014>but you should be able to interpret itWhy, if Gemma can interpret it for me? She'll be the one doing any future maintenance. It's such a bizarre thing to get hung up on, needing a human somewhere to understand it for no reason other than to talk about it with other humans instead of actually improving the project.
>>109346996Go and read bun's repo, then come back. I'll be waiting for your written apology and dogeza.
>>109347051https://raw.githubusercontent.com/oven-sh/bun/refs/heads/main/src/runtime/api/cron.rsLooks fine to me.
>>109346669I'm trying to decide between m2.7 q5_k_m or m3 q3_xxs