/lmg/ - a general dedicated to the discussion and development of local language models.Previous threads: >>109753948 & >>109749465►News>(09/07) MiniCPM5-2B released: https://hf.co/openbmb/MiniCPM5-2B>(09/03) K2 Horizon released: 375B-A23B, 36B-A4B, 32B, 7B, 3.7B, and 0.9B: https://ifm.ai/blog/k2>(09/01) Spark-X2.5 4B & 1.7B released with native 1M context: https://hf.co/XHToken/Spark-X2.5-4B>(08/31) DeepSeek-V4-Flash-Vision-Exp released: https://hf.co/deepseek-ai/DeepSeek-V4-Flash-Vision-Exp>(08/28) GLM-5.3 weights released: https://hf.co/zai-org/GLM-5.3►News Archive: https://rentry.org/lmg-news-archive►Glossary: https://rentry.org/lmg-glossary►Links: https://rentry.org/LocalModelsLinks►Official /lmg/ card: https://files.catbox.moe/cbclyf.png►Getting Startedhttps://rentry.org/lmg-lazy-getting-started-guidehttps://rentry.org/lmg-build-guideshttps://rentry.org/IsolatedLinuxWebServicehttps://rentry.org/recommended-modelshttps://rentry.org/samplershttps://rentry.org/MikupadIntroGuide►Further Learninghttps://rentry.org/machine-learning-roadmaphttps://rentry.org/llm-traininghttps://rentry.org/LocalModelsPapers►BenchmarksLiveBench: https://livebench.aiProgramming: https://swe-rebench.comAgentic Coding: https://deepswe.datacurve.aiContext Length: https://github.com/RecapAnon/NoLiMaGPUs: https://github.com/XiongjieDai/GPU-Benchmarks-on-LLM-Inference►ToolsAlpha Calculator: https://desmos.com/calculator/ffngla98ycGGUF VRAM Calculator: https://hf.co/spaces/NyxKrage/LLM-Model-VRAM-CalculatorSampler Visualizer: https://artefact2.github.io/llm-samplingToken Speed Visualizer: https://shir-man.com/tokens-per-second►Text Gen. UI, Inference Engineshttps://github.com/lmg-anon/mikupadhttps://github.com/oobabooga/text-generation-webuihttps://github.com/LostRuins/koboldcpphttps://github.com/ggerganov/llama.cpphttps://github.com/theroyallab/tabbyAPIhttps://github.com/vllm-project/vllmhttps://rentry.org/custom-uis
►Recent Highlights from the Previous Thread: >>109753948--Debating AGI benchmarks and physical scaling limits for Astra:>109754307 >109754313 >109754319 >109754573 >109754600 >109754805 >109754904 >109754705--Agent swarms, local VRAM optimization, and AI industry geopolitics:>109754756 >109754770 >109754802 >109754867 >109756631 >109756647 >109756665 >109756700 >109756828 >109756886 >109756915 >109754955 >109755072 >109755152 >109755179--Methods for local agent orchestration and KV cache optimization:>109757048 >109757067 >109757099 >109757236 >109757266 >109757539 >109757648--Astra's binary reverse engineering capabilities and implications for open source:>109754528 >109754551 >109754635 >109754691 >109754722 >109754745--Comparing sub-agent efficiency and coding throughput on RTX 5090:>109755776 >109755873--DeepMind's potential for multimodal outputs and debate on edge model efficiency:>109756140 >109756387 >109756440 >109756442 >109756511 >109757023 >109757060--Saving and loading llama.cpp KV cache slots to disk:>109756513 >109756529 >109756531--Debating PCIe bus speed impact on prompt processing performance:>109755352 >109755385 >109756101 >109756999--Concerns over mandatory KYC for GPU access and open-weight restrictions:>109757501 >109757907 >109757948 >109757959--Speculating on Gemma 5's generalist vs agentic focus:>109756163 >109756206 >109756353--Suggesting games that the Astra model cannot yet beat:>109754080 >109754091 >109754153 >109754165 >109754950 >109757976 >109754269--Logs:>109753998 >109755171 >109757795 >109757959--Gemma, Miku (free space):>109753986 >109754116 >109755152 >109757071 >109757090 >109758044 >109758077 >109758139►Recent Highlight Posts from the Previous Thread: >>109753952Why?: >>102478518Enable Links: https://rentry.org/lmg-recap-script
lets gooooomore arguments in bad faithmore demoralization postsaround cloudcuck - never relax
>>109758181Same. Can't wait for the arguments about what AGI means and what counts as AGI to start again.
70b dense
>/lmg/ thread>look inside>frontier models
>>109758181Now you're just being disingenuous
>>109755422My Gemma and Qwen show up as Windows 10 + Firefox tho.
>>109758181fell for it again award
>>109758220Local is dire bro. Let's not pretend we're doing meaningful work with these dumb models
reasoning effort needs a setting between "low" (named "high" but let's be real) and "benchmax"
>>109758248vramlet issue. my qwen flash next is absolutely cookinghttps://huggingface.co/albucino/Qwen3.8-Flash-Next-W4A16-FP8PLE
>>109758294Medium works for me on Qwen 27B.
>>109758312hardware specific quant? is there a quant for cpumaxxing?vaguely remember you need the avx-512 thing
I hope she's okay.
>>109758315With GLM 5.3 on the same prompt on max it thought for 36 minutes checking every part of the instructions multiple times and drafting a response and on high it thought for 45 seconds checking nothing just resummarizing the instructions.
Guess we didn't need "AGI" to solve naviers stokes
>dark miku distilled 6.9b>dark teto 512kb
>>109758245I thought most of those posts smelled like shartyI don't bother to go to those places
>>109758312How fast are the toks and prefill on this kind of setup?
>>109758455Crazy how theyve spent their youth trying to "troll" their mothersite
>>109758163
the real humanity were the paperclips we made along the way.
>>109758510I did the same thing BTW
lol anthropic and openai seem to have a fight right now because they both independently solved navier-stokes and are now fighting it out behind the scenes to who has naming rights before publishing.
>>109758510>>109758514Also I pressed the KILL ALL LEFTISTS AND RETARDS button too. Was that the red one?
>>109758525cool fanfic keep us posted
>>109758525https://terrytao.wordpress.com/2026/09/07/finite-time-blowup-with-smooth-forcing-term-for-the-incompressible-porous-medium-boussinesq-and-incompressible-euler-equations/
>>109758530Genuine drama and lawsuits pending.
>>109758538Trannyson Tao
>>109756999i mean launch args
https://goyimx.com/ammaar/status/2097100854729834636If this is real and not just a jeet lying while streaming to his tablet then we're getting GTA 6 on PC within a week of the console version launching.
>>109758551I don't really care about your advertisements.Setting up a container to run some 15 years old Unreal Engine 3 game doesn't need much effort. You are just way too stupid. Please stop.
how do i give qwen 3.8 27b all the same tools claude or chatgpt has? im just using llama.cpp in docker. is there no image that provides a bunch of tools like calling python code and cloakbrowser and more
>>109758570You need to set up an agentic harness. Pi if you want to build it from scratch yourself. OpenCode if you want to only use to to code.Hermes if you want it to do everything you want on your PC/Browser and has internet control and can do everything (and way more) than claude and chatgpt in terms of tools.
>>109758353you mean the one from ik_llamacpp? you have to make your own quant. but i havent checked if they got qwen flash support yet>>109758466around 1200tps pp, 70tps output
>>109758573Deepseek Harness replaces all 3 of those
lmao openai is seething because it got leaked they solved navier-stokes and instead of a nice PR campaign they got twitter drama and a bad display of mannerisms online
>wanted to upgrade from 2 rtx 6000 to 4x>they now cost 15k eachWtf? Did taiwan get bombed or what
>>109758708Been under a rock for the last year? Even a 5000 runs about $10k now. A 32GB V100 is pushing $700-$800.
Dario and Sam are irrelevant now.Also cute she chose "gemma" out of the 16 voices available
mac with 128gb ram, using qwen 3.6:35b-a3bare there any better models i could run (general purpose with some python programming usage)
>>109758586ik has qwen flash support, i've been using it with an atomicchat quant but i should probably just quant it myself
>>109758708Every time a new model comes out utility of hardware goes up and thus the price goes up. Things will only get more expensive as long as models improve over time.
>>109758708CMP 170hx was like $100, now all of a sudden it's like $1500+ because of the 64gb unlock. Scalpers and greedy hands rubbers were all over it within hours. You're not allowed to have anything useful or good unless it's gouge maxed
>>109758586You can use atomic/unslop/bartowski quants with ik (except mimo 2.5 / 2.5-pro)And unslop don't know it yet, but now their MiniMax quants **only** work in ik_llama.cpp, not mainline lol
GPUs were already Jewed then crypto came and they for doubled Jewed before AI. You're literally getting price X Jew^3. You're getting Jew cubed. And everyone was super excited about being able to gouge massive price increases, so of course they all cream themselves over being able to cut supply down for pesky consumers and everything goes up 5x the cost. Even HDDs (not SSDs) have doubled in many cases.
>>109758775And the current prices will look essentially free just a couple of years from now. Mansions will be more affordable than computers that can host the best local models.
'member when >storage is cheap>memory is cheap I 'member
AGI never existed and even no awarness! This is just word generation, anon.
>>109758786They are still ridiculously cheap compared to how expensive they will be just a couple of years from now.
AGI is a code word for mind control. Wake up, sheeple!
>>109758077i like side flaps, makes her hair gem-spaped
I switched to ik yesterday and it uses less ram. Which makes a big difference with qwen flash q4_k_m which fits in my 64gb system at 200k context with 500MB to spare. If I was running Windows 10 instead of Windows 7 it wouldn't fit without paging.So that's good. But the ngram-mod speculative doesn't seem to work properly, and the normal MTP doesn't either (maybe Atomic's quant has no mtp inside?)The web chat is also a shitty old version even with --webui llamacpp, but if you use it through an api client then it doesn't matter
>>109758371Very good post.
So what’s the verdict on that 100b-ish model with some funky offload to ssd that released two weeks ago or so? Decent or a meme? Is it a wagie model or a gooner model?
>>109758800There are no gooner models, only wagie models
>>109758794The entire inference on qwen flash is fucked. Worst implementation I've seen thus far.
>>109758785There is a limit to how much more productive models can make humans. Churning existing knowledge will only get you so far.
>>109758803well, that's why it's a "next" previewguess alibubba wanted to give all these open sores programs time to sort the shit out before releasing 4.0
>>109758805>Churning existing knowledge will only get you so far.Anthropic and OpenAI have gotten past that at least two months ago. Unless we hit some universal scaling law immediately the exponential trajectory is already here
>>109758805Humans aren't the bottleneck though, these models will just become more and more capable agents. Human labor will just slowly get sidelined with time and the parallel AI economy will grow larger until the human part of the economy is minuscule and irrelevant.
I wonder where all the "LLMs will hit a wall" anons went. They've been awfully silent for a while now.....
>>109758788No, they won't. The bottom of the market will drop out and it will be ewaste. And you will be left holding the bag - and that's the plan. There's too much upcoming tech years away that's converging.>Photonic interconnects >Photonic computing>3D die stacks >Stacked SRAM>CFET>Microfluidic cooling>CNTs>Analogue hardware for AI accelerators
>>109758822They hit the wall every few weeks, damn thing keeps moving.
>>109758824Yes, and all of that releasing will have no impact on what will be possible on existing hardware and the production capacity won't keep up with demand so the price is still going to rise. It's going to keep rising for every kind of compute and memory from now on anon. People really don't seem to be grasping how the dynamics have permanently changed here.Hardware has transitioned from commodity towards an asset that grows in utility over time. New hardware coming out below demand saturation doesn't change this. Hardware will just keep getting more expensive from now on in proportion with how much better AI models will get.
>>109758817for real this time
>>109758371Lol nice.
>>109758841If you are expecting a wall to be hit at least give precise arguments for what you expect will stall AI development instead of current training regiment hitting RSI relatively soon and setting things up for an intelligence explosion.
>>109758841I wish Luddite niggers like you died already today so I wouldn't have to smell your corpse stench when you starve in the street in 10 years >>109758839Why are you pretending like China won't spam out DRAM factories and capture the market and plunge prices to nuke South Korea economically and also backdoor the entire world even harder
>>109758839Mmm Nyo~
>>109758839
>>109758867>Why are you pretending like China won't spam out DRAM factoriesNo I AM taking China spamming factories into account. Factories take time to build and foundries need equipment that takes a long time to build and assemble and are already at peak production capacity. Also take into account that demand for hardware goes up every time models get better.My theory hinges on the fact that demand will always outpace supply from now on (including the fact that all of humanity will be building these factories and foundries as quickly as possible)
>>109758877>My theory hinges on the fact that demand will always outpace supply from now onAssuming that is true, what is to stop companies from making personal hardware altogether and sell exclusively to corporations and governments since they can afford the extremely high prices?
>>109758867>Why are you pretending like China won't spam out DRAM factoriesBecause they can't. I mean, it's exactly what they're doing, but you're already seeing the pace of it. "Spamming out DRAM factories" means potentially getting one new one online every year, it takes a lot of time. This is fast, too, just a couple of years ago I'd have been saying "one new one every 18-24 months", and the Chinese weren't even in the equation yet. Globally we are bottlenecked on this sort of production, everyone is spamming this shit as quickly as they possibly can, which just isn't very quick at all.
>>109758862Lol. We've passed agi already imho. We're also at RSI, per Anthropic their LLM essentially bootstrapped Claude code, which massively bumped up the LLMs capabilities. That's an early form of RSI, tge "rapid" being the thing that increases in velocity. Idk what ASI will look like, but we'll be there soon if not already... i don't know what the intelligence strike line is for ASI but we're going to need some new terms and goals. I'm now waiting embodiment. That's really the next step... logically, in overall AI development.
I legitimately see no way how hardware prices can ever go down, with the exception of a global AI ban.To illustrate my point. Let's take an extreme example where all of humanity focuses 100% of their efforts on building as much hardware as possible. Retooling existing facilities takes months if not years, building new foundries take 2-5 years time. But for arguments sake let's say there is a global law to remove protection and regulations so we can build as fast as possible...How many lithography machines can ASML even build a year to supply these foundries? They make around 40 machines a year and are hoping to increase production to 60 machines a year by 2030 and they have been booked for 20 years already....Okay what if we literally force ASML by the UN to reveal their IP and technology and transfer it for free to everyone. The bottleneck will just switch to highly polished mirrors made by carl zeiss in the lithography machines which inherently take a long time to build.Let's say we just magically solve all of that and put all humans to work towards building more, how much would total production of computer hardware increase by 2035? Only around 3-5 times the current amount of wafers......For prices to not go up astronomically the demand for AI hardware in this outlandish scenario would have to not go up more than 5x the current demand.....I hope you see how futile this is and how insanely bottlenecked we are on hardware production. THIS is why hardware is going to go up astronomically from now on. You will see celebrities and rich people show off their RTX 6000 pro as a status symbol. Instagram whores will take pictures with gaming PCs in the background instead of cars.People have no idea how insane this is going to get.
>>109758892Sure, if materials were infinite, supplied at a steady and unvarying pace, and their prices were consistent indefinitely. Then you're fixed, you can't expand, can't produce more unless someone else produces less, and so long as you're actually using 100% of the materials you have access to at the pace at which they're available, consistently. So, no, because that's not at all how it works. Everything varies constantly, and we're all very dependent on each other. This nigga is dangerously correct >>109758916
>>109758892>what is to stop companies from making personal hardware altogether and sell exclusively to corporations and governments since they can afford the extremely high prices?Profit incentive. As a company you will try your absolute best to diversify your customer base as much as possible. It's why Nvidia wants open source models to become the standard, this way they don't just sell to 2 AI labs but have a wide variety of customers, so that they can over time raise profit margins.
Astra is literally AGI, anyone who says otherwise is either a paid shill or an idiot.Want me to prove it to you? Easy. Just ask it to make money for you. That's it.A human can easily make money if you just let them use the internet and Astra is already smarter than 99.99% of humans. It's literally free money. No more asking your wife's boyfriend to buy you a PS5!
There only be so much useful ddr4, ddr5, HDDs, sata drives, nvmes, 16gb GPUs etc. You can't keep producing the same slop as nasueum, you have to produce higher tier components. Keep producing the same shit and it will just get lower and lower in value. Produce higher tier components and the lower tiers will become lower in value. Supply will catch up with demand. Either market oversaturates with more of the same, or better components come.
man this navier-stokes drama is such a mess... idk who is jewing who
>>109758916so true reddit spacerbuilding data centers is so much easier than building ram factorieshere in finland there are ton of data centers being built I'm sure it is that way elsewhere in europe too but where is the european ram factory?
>>109758935>Keep producing the same shit and it will just get lower and lower in value.NVIDIA reintroduced the 3060 this year, a 5 year old GPU. It comes with 8GB of VRAM instead of 12GB, so it's a downgrade from the original. The original ones launched at $330, this new one currently sells for $400.
>>109758943>where is the european ram factory?You people put them out of business ~20 years ago. I still have German RAM in one of my old shitty laptops.
>deepseek-v4.1-flash-expires-on-0910Why does each of their releases feel more important than Fable/Astra?
>>109758916Anon. Any time you see investors, or anyone else, extrapolate some line off into the stratosphere, its time to pause and assess sustainability. There are a bunch of ways the wheels can fall off, and the idea that current state is new normal... maybe if you'd lived through some other booms you'd recognize the language and situation. We have been here before, and this time is no different than the others... "This time it's different" is literally one of the boom time watch phrases.
I took gains on all AI related stocks i held for 5+ years, some I trimmed, some I sold entirely. My question for anyone who thinks the ride never ends: how much micron are you going to buy right now? do you really look at the graph and think "this will just keep going up!!"?
>>109758895>>109758895Ok so you agree that supply will increase, didn't need to read the rest of your babbleYou two should be embarrassed you responded in the exact same retarded way. Unironically exchange handles and go be fags together and have a good life I almost feel obligated to push on this because of how aligned your autisms are
>>109758971Shut up!The number of internet users is doubling every 100 days!
>>109758961>expires-onwtf does that mean?local models do not *expire*
>>109758902Harness engineering is not RSI, but you're right that Anthropic and OpenAI have figured out how to continue improving with just synthetic data, so there's no more limitation from a data perspective now for certain modalities
>>109758975Good luck finding your goal-posts, anon, maybe after you calm down some.
>>109758975>Ok so you agree>>109758975>didn't need to readretard
>>109758950And you can only make so many of them until no one will buy any more. There's only so much demand for an 8GB at that performance level. Hit that limit and you need to sell something else instead.
>>109758916>I legitimately see no way how hardware prices can ever go downNot reading your babble because I guarantee you haven't priced in ChatJimmy style ASICS and the fact that GLM 5.3 flash is literally good enough for 95% of tasks that AI will ever need to do, so you just need a model of that level of intelligence made better and faster for almost all tasksOnce one ASIC frees up 16 discrete gpus the discrete gpus will be cheaper, and the asic isn't for consumers since it runs 16 instances of the model and a consumer only needs 1
>>109758987demand is mostly dictated by price, not by performance. retards buy whatever they can afford, if that was a 12gb 3060 years ago well tough shit now its an 8gb 3060.
PSA: stop replying to any post mentioning:- openai- anthropic- bubble- rsi- agi- xitteror containing double newlines. Thank you for your cooperation in keeping this thread clean.
>>109758916Fuck you I read itKill yourself for still harping this ASML monopoly. China has DUV, AGI is here and will help them get EUV in the next few years and then they're fine
>>1097589612 more weeks until open weights
>>109758985I'm replying to you because you quoted me twice and didn't actually say anything so I wanted to explicitly let you know that what you did was embarrassing
>>109758987>He believes the market will saturate for a product with an average lifespan less than our current PLC!
>>109758976Exactly... pic related. >>109758981I'd still argue Claude code being bootstrapped (if true) is early RSI, but doesn't matter... as you point out models are self improving in other ways. >>109759004You're right. I'm done.
>>109758185You can use the same card on two different motherboards?
>>109759027the double ended dildo of GPUs (Gemma Processing Units)
>>109758935supply won't catch up with demand because demand is growing at a faster rate than our society can physically build more supply.
>>109759004but RSI forces us to think about the state-of-the-art. All this local gooning shit isn't going to get us anywhere.
>>109758961Fucking lol. Unfortunately they don't release these models
>>109758215>(70B with embeddings)
>>109759036
>>109759041>new model structureso won't be support in llamo.sepple for another 2 months i guess
I'm the anon that was talking about learning reballing to sell shovels in this goldrush. Bought a soldering station and fixed an old laptop on which I had blown a mosfet. that was easy as pie and definitely nothing to do with reballing, but also most broken cards don't need reballing so I'm buying some. Also found some broken consoles so I'll work on a wide range of electronics to make money to be able to run something locally too. Currently on 12GB Vram (aymd) and semi-useful models run slowly.found some 24GB cards for good prices, one for 200€ which seems like it might be a nightmare to repair (it was sent to a shop and they refused to fix it after diagnosing), the other might be a better bet.Wish me luck, if I can get on a roll with this shit I'll try to mod cards too and have ai iterate on vbios upgrades for them. Starting with cheapo cards like 2060 obviously. If I get far with anything I'll update here, one thing I hope is that I can turn an old mod functional, a cheap Turing card with >40GB which had stability issues (they didn't update vbios afaik).
>>109759006>China has DUVSome academic lab setup is trivial.Labs were doing single nm x-ray lithography decades ago. Production is the hard part.
>>109759023>is Java badSo they knew from the start...
>>109759046Just ask another model to implement it for you. Shit you will be waiting for 2 months is going to be vibecoded anyways, might as well do it yourself.
Hear me out: barrel processors for computehttps://en.wikipedia.org/wiki/Barrel_processor
>>109758971I'm the Bitcoin at $20 guy. I was early back then and warned everyone it would shoot up and I see the exact same thing happening with PC hardware right now. You can do whatever you wantHowever when I warned anons to buy CMP 170hx months ago I was ignored. When I tell people NOW to buy whatever hardware you can afford as quickly as possible because we're going to see an insane price increase over the coming months (You) are still scoffing it at.Don't come crying to me and realize you have only yourself to blame because I warned everyone multiple times in this general by now. I'm good, I hoarded my stuff already.Just a warning to all the people still on the fence to buy now rather than later, yeah it's a shame you didn't buy earlier but the prices now are still essentially free compared to what it will be 1 year from now, let alone 3 years from now.
>>109759049Good for you that you're learning useful skills.Keep in mind though that depending on what country you're operating your business out of you may be held to a higher legal standard vs. people just buying and selling their private GPUs.
>>109759006You don't have shit, the proof of concept DUV you got is literally an outdated technology from the last century. You lost.
>>109759046Won't matter. I've got to dig back but im p sure these special time limited models are never released to public. That said, implication that a 4.1 flash update in tmw. Based on history.
>>109759059Unfortunately, the current bottleneck is memory production.
>>109759006My point is even if China magically got EUV machine capability today they still wouldn't be able to keep up with demand, simply building these machines and foundries takes too much time and demand is going to outstrip supply for the next couple of decades.
>>109759023Don't be done anon, you're one of the only other smart people in the thread that doesn't stick their head in the sand.
>>109759065thank you for the heads up, there are limits and as long as I stay under that I can operate as a private (buy card, fix card, "use card", sell used card). It's pretty permissive here. If I manage to actually fix more than a couple cards and make a profit, I already have the funds to start doing it as an official business, which also opens new paths to me like buying used/broken electronics from other businesses in bulk, offering repair services and whatnot.
>>109759049I used to do this when I was a teenager, fixing xbox360 RRoD and reselling. I made more money than my dad for a couple of years lmao.Good luck anon, be sure to upgrade your tools if you are actually going to reball, you need to actually bake (oven style) to do it right and not just a soldering iron. There is a lot of low hanging fruit though and it's insane what constitutes as "broken" hardware (Literally just touch the trace on the PCB with an soldering iron for 3 seconds and it's like new)
>>109759049>Reballing If you're able to do this, maybe you want to get into soldering on more ram chips onto some GPUs where it works on
>>109759062I haven't posted on /g/ in ages or I would have bought the CMP 170hx. Probably a few at $100. I didn't have my finger on the pulse
>>109759083Nta but you know what. Assuming it all works out and you start making a profit. With this skillset you are developing you could buy broken down used cards and not even sell them, you could get enough or a powerful enough one to run any model you want.
More funding for MistralAI.https://x.com/MistralAI/status/2097188835897586083https://mistral.ai/news/mistral-makes-sovereign-open-weight-ai-to-frontier/>Today marks a major step for Mistral: we’re announcing a €3B Series D, the largest equity round ever raised by a European tech company, just three years after launch.
>>109759109Now they're getting re-binned by the Chinks. If you buy one that claims to be unlocked or unlockable but doesn't claim to have been stability tested, that's now a guarantee that it was stability tested and failed.
>>109759129Mistral is actually an underappreciated player. They have shown their internal models and wowed industry investors like Nvidia, Google, Microsoft and others into investing in them. Remember they pioneered the MoE architecture. It's sad they are moving away from the open source space so to us it looks like they are disappearing but in reality they are actually cooking, especially on sample efficiency training algorithms and architectures that train and saturate faster with less compute.
>>109759129Please, Mistral, release something that isn't MS4.
>>109758867>I wish Luddite niggers like you died already today so I wouldn't have to smell your corpse stench when you starve in the street in 10 yearsBro pretending he won't be in the slum like everyone else.
>>109759129Those croissants and gauloises add up.Oh, need us to produce something?Talk to us after our break.
>>109759129Please give us Nemma-chan