[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: 50-50.png (442 KB, 512x768)
442 KB PNG
/lmg/ - a general dedicated to the discussion and development of local language models.

Previous threads: >>109745530 & >>109740702

►News
>(09/03) K2 Horizon released: 375B-A23B, 36B-A4B, 32B, 7B, 3.7B, and 0.9B: https://ifm.ai/blog/k2
>(09/01) Spark-X2.5 4B & 1.7B released with native 1M context: https://hf.co/XHToken/Spark-X2.5-4B
>(08/31) DeepSeek-V4-Flash-Vision-Exp released: https://hf.co/deepseek-ai/DeepSeek-V4-Flash-Vision-Exp
>(08/28) GLM-5.3 weights released: https://hf.co/zai-org/GLM-5.3
>(08/28) Hy4-preview 770B-A49B released: https://hf.co/tencent/Hy4-preview

►News Archive: https://rentry.org/lmg-news-archive
►Glossary: https://rentry.org/lmg-glossary
►Links: https://rentry.org/LocalModelsLinks
►Official /lmg/ card: https://files.catbox.moe/cbclyf.png

►Getting Started
https://rentry.org/lmg-lazy-getting-started-guide
https://rentry.org/lmg-build-guides
https://rentry.org/IsolatedLinuxWebService
https://rentry.org/recommended-models
https://rentry.org/samplers
https://rentry.org/MikupadIntroGuide

►Further Learning
https://rentry.org/machine-learning-roadmap
https://rentry.org/llm-training
https://rentry.org/LocalModelsPapers

►Benchmarks
LiveBench: https://livebench.ai
Programming: https://swe-rebench.com
Agentic Coding: https://deepswe.datacurve.ai
Context Length: https://github.com/RecapAnon/NoLiMa
GPUs: https://github.com/XiongjieDai/GPU-Benchmarks-on-LLM-Inference

►Tools
Alpha Calculator: https://desmos.com/calculator/ffngla98yc
GGUF VRAM Calculator: https://hf.co/spaces/NyxKrage/LLM-Model-VRAM-Calculator
Sampler Visualizer: https://artefact2.github.io/llm-sampling
Token Speed Visualizer: https://shir-man.com/tokens-per-second

►Text Gen. UI, Inference Engines
https://github.com/lmg-anon/mikupad
https://github.com/oobabooga/text-generation-webui
https://github.com/LostRuins/koboldcpp
https://github.com/ggerganov/llama.cpp
https://github.com/theroyallab/tabbyAPI
https://github.com/vllm-project/vllm
https://rentry.org/custom-uis
>>
File: mikuthreadrecap.jpg (1.15 MB, 1804x2160)
1.15 MB JPG
►Recent Highlights from the Previous Thread: >>109745530

--Paper: Discussing coding agent harness architecture and agentic buzzwords:
>109746196 >109746230 >109747581 >109747810 >109747593 >109746246
--Comparing Qwen 27B and Gemma for coding and NSFW:
>109747698 >109747701 >109747767 >109747962 >109747991 >109747995 >109747779 >109747922 >109748007
--Attempting to port backend optimizations to llama.cpp using Fable:
>109748024 >109748108 >109748317 >109748793 >109748625
--Benchmarking EPYC server builds with V100 e-waste GPUs:
>109748456 >109748469 >109748494 >109748522 >109748526 >109748528 >109748626 >109748676
--K2-Horizon 7B benchmarks and frustration over required llama.cpp forks:
>109746110 >109746128 >109746335 >109746368 >109746428
--Debating BitNet performance versus bf16 based on VRAM and information density:
>109745701 >109745763 >109745808 >109745838
--Irodori-TTS-v4.1-Anime for Japanese voice cloning and quality:
>109747320 >109747380 >109747447 >109747452 >109747472 >109747523 >109747891 >109747900 >109748002
--Local TTS capabilities and Omnivoice cloning performance:
>109748377 >109748390 >109748452 >109748540 >109748611 >109748490 >109748505
--Spotify's agent architecture for optimizing costs using tiered model reads:
>109746148
--Predicting the timeline for frontier-level open-weight 30b models:
>109747203 >109747235 >109747265 >109747333
--Reacting to purported Jensen Huang tweet claiming AGI via GPT-6 Astra:
>109746028 >109746065 >109746074 >109746088 >109746160
--Debating AGI achievement based on mathematical proofs and hallucinations:
>109745735 >109745750 >109745913 >109748427
--Forum-style chatbot harness utilizing reply chains for LLM context:
>109746469 >109746600
--Logs:
>109746294 >109746469 >109746715 >109747447 >109747457
--Gemma, Inkling-chan (free space):
>109745602 >109748170 >109748216 >109748631

►Recent Highlight Posts from the Previous Thread: >>109745537

Why?: >>102478518
Enable Links: https://rentry.org/lmg-recap-script
>>
Why do some anons use Sparkx 2.5 or K2 horizon over gemma and qwen?
>>
Just ordered 4 cmp 170hx's for an absolute bargain, can't wait
>>
>>109749476
Me too, the seller wanted 2.8k, but I managed to haggle xer down to 2.5k, for a measly 10k total. 160gb of vram for only $10,000+shipping, what a steal!
>>
>>109749491
Damn that's a lot, mine were $500 each
>>
File: xebx5pb0yznh1.png (422 KB, 612x1236)
422 KB PNG
Person that invented the term "AGI" defends OpenAI calling Astra AGI.
>>
>>109749526
A Twitter marketer fight? Show me the money, boss!
>>
>>109749526
If I bought the term, could I unsettle it, or is that cultural hegemony?
>>
>>109749526
AGI just means that it improves itself (to me at least) -- Not just specifically to an LLM workflow. Once you reach the point that you are not prompting anymore, then it's fully automated.

Best places to use it:
Math
Chemistry
Compsci
To name a few
>>
>>109749535
Using context, you can code in a language the model was never trained on...
>>
70b dense
>>
>>109749535
>AGI just means that it improves itself
That's RSI. AGI is a less ambitious mark of merely being as competent as a human in most or all human tasks, at least digitally.
>>
>>109749535
You can have AGI without RSI, namefig.
>>
>>109749535
>tripfaggot
>>
>>109749535
Go back to the subreddit you crawled out of
>>
File: s6m70ml761oh1.png (246 KB, 2380x1563)
246 KB PNG
We're following the AI 2027 scenario almost to a t. Even the social, political and misalignment predictions outlined in that scenario have been coming true.
>>
>>109749554
never again
>>
>>109749603
I only care about local models.
>>
>>109749526
Transformers will only ever be marketer "AGI" they fundamentally can never become a "true AI system" capable of autonomous self-driven behavior. The simple fact that you must prompt them for them to be "alive" immediately discredits their capability of becoming an actual AGI. Exceptionally powerful encyclopedias? Absolutely. Capable of being a long-horizon experiential automata? Never.
>>
>>109749611
local models follow the exact same curve but 6 months behind.
>>
>>109749611
Hopefully Grok 3 will drop.
>>
>>109749612
>they fundamentally can never become a "true AI system" capable of autonomous self-driven behavior
Bro my fucking GLM 5.3 I host myself is capable of autonomous self-driven behavior when put in an harness and I tell it to go and look for things to optimize on my system, it comes back to me with reports and suggestions and everything. You just live in the past.
>>
>>109749554
Too expensive to train. But now it's clear there are much better ways for expanding model knowledge without involving MoE and in a way that is much more local-friendly.
>>
>>109749465
Cute miku.
>>109749526
SOTA passed AGI some time ago. Which is already superhuman.
So bring on ASI.
>>
File: 1782783516091263.png (1.04 MB, 1300x1456)
1.04 MB PNG
>>
>>109749618
>and I tell it to go and look
that's his point
>>
>>109749612
meh can intelligence arise from unintelligent things cooperating? proteins making cells, cells making animals? If yes, then a group of llms prompting each other might become intelligent. If no (and you believe something metaphysical happens at the border between unintelligent mechanism, e.g. protein, and intelligent being, e.g. human, which I do) then no AI will ever be truly intelligent unless we can replicate the metaphysical thing, which I don't believe we can.
Another option you might believe that leads to "We'll never have AI" is that everything building up to humans, including atoms and proteins, have some intrinsic property which manifests as intelligence in complex enough arrangements, and that bits on a computer do not have this property.
I don't see, however, how can you believe that machines can become truly intelligent, but LLMs can't.
>>
>>109749630
So you're saying GLM 5.3 isn't AGI because you need to tell it "Go be an indepdendent AGI actor" before it becomes fully autonomous?

Are you retarded? That's like saying the robots from Detroit Become Human aren't AGI because they have a power button on their backs and humans have to press it before it becomes autonomous.
>>
Heres a thought I had
Why are the Japanese so fucking bad at ai? They've done nothing useful in the space that people have cared about or moves us forward, they let china literally run circles around them at this point.
>>
>>109749619
>>
>>109749613
With local models I mean models that you can run on your own machine. If you need a few entire GPU nodes to run a LLM, then it might be open-weights, but not actually local.
>>109749614
So old and fat that even Elon might be embarrassed to release it, at this point.
>>
>>109749627
Miku before having 42 heart attacks
>>
>>109749642
Because they have anime, which is the point of probably 50% of AI. They don't need AI, anime will keep them relevant.
>>
>>109749624
Like what
>>
>>109749637
Who said proteins or even atoms are unintelligent?
The difference is they're physical and 3-dimensional, way more quantum fields involved than just electromagnetism. Digital is a fundamentally different and fully deterministic domain if no hardware faults occur.
>>109749640
Go set it loose and tells us what it creates then. Should be easy since it's AGI, right?
>>
>>109749642
>Why are the Japanese so fucking bad at ai?
They're terrible at software engineering. Some of the worst on the planet.
>>
>>109749647
>With local models I mean models that you can run on your own machine.
I'm running GLM 5.3 flash on my own machine, what you can run or not on your very specific poverty box doesn't change the definition. Also using shittier hardware just means you follow the exact same curve but even farther behind. If you are at the ~30B range then you are 18 months behind but following the same curve, if you only have a smartphone you are about 3 years behind but following the same curve.
>>
>>109749647
>So old and fat that even Elon might be embarrassed to release it, at this point.
what does "even" mean in this sentence?
>>
>>109749665
Sparse associative memory.
>>
I told the clanker to create candidates. turns out Cursor lets them spin up subagents, so work sped up. I didn't know it could do that.
>>
>>109749669
You know what you're right... Those motherfuckers are still using fax machines unironically, its only been recently where you can actually expect even serviceable pc games out of them
>>
>>109749618
>harness
Nobody tell him what harness does lol.
>>
>>109749642
Japan used to be amazing at AI in the 1980s and a lot of the current paradigm was actually first implemented by Japan back then. However the government put hundreds of billions of USD into AI very similarly to what is happening in the US with the amount of funding AI is receiving right now.

However it didn't pan out because the computing power just wasn't there and they had constant data shortage because the internet didn't exist yet so it collapsed and it was actually one of the contributing factors for the 1995 "lost decade" insane crash Japan experienced which permanently humiliated them and caused Japan to stop being relevant in the world.

They are so fucking traumatized by this shit that everyone has stayed away from AI ever since.

Ironically Anthropic started a branch in Japan to hire all of those old 1980s AI experts that are now in their 60s to try and get insight from them into new architectures and the like.
>>
>>109749699
Normies are fried
>>
>>109749699
Human subconsciousness also has intrusive thoughts and hidden messaging that "prompts" it into thinking and doing things.

Humans aren't general intelligences according to your definition.
>>
>>109749675
He (xAI)'s released commercially obsolete Grok 1 and 2 in the past, but the frontier of open-weight LLMs has advanced so much since Grok 3 that at this point releasing that would bring negative publicity to his company, so he might not do it.
>>
>>109749712
Neither you nor anyone else actually knows how human thinking "works". Thinking you do does not make you sound intelligent.
>>
>>109749666
>>Another option you might believe that leads to "We'll never have AI" is that everything building up to humans, including atoms and proteins, have some intrinsic property which manifests as intelligence in complex enough arrangements, and that bits on a computer do not have this property.
>Who said proteins or even atoms are unintelligent?
Ok so you're on that side. Then no computer program will ever be AGI, until they add more of whatever you define as "quantum field involvement"
>>
>>109749720
I've used psychedelics before. You can kind of see how the mind works because it exposes the subconscious to you. It essentially shifts the "attention layers" of your conscious mind to be able to see the thoughts of your subconscious minds directly.
>>
>muh agirsisi
>still gets tripped by tarded shit like car wash question
yeah...
>>
>>109749647
Eyeballing the eci chart, the july version of dipsy flash at the same level as dec frontier. So about 6 months.
For the toaster crowd, the mar/april models were about at 2025 april frontier. So about a year.
Doesn't have a lot of the recent open models sadly.
>>
>>109749737
Go ahead and try asking Astra the car wash question anon.

Protip, Astra now scores better on spatial reasoning and common sense questions than human experts.
>>
Which Gemma 4 heretic model should I be getting? Or not Gemma 4 and something else?
>16gb vram
>32gb ram
>>
>>109749725
That's him, Falsifier Slyboots!
>>
The Scam Altman cocksuckers are going crazy lately.
>>
>>109749734
Holy fuck dude just stop. I've done mushrooms too and they will not let you "understand" how the mind works, that's a human hallucination. You can learn some things but it isn't necessarily some magical portal into the fundamental structures of the mind.
>>109749727
Not my definition. It is the quantum model explanation of what is measured.
You, I and the animals are proof atoms have the capability to be intelligent and have sovereign agency, something that is required for self-driven exploration and pattern breaking. No machine has ever expressed these features. They can so far execute algorithms in a deterministic manner with higher and higher precision which is a part of intelligence but that's pretty much it.
We fundamentally do not understand what "it" is and we have barely defined it, yet we keep declaring to have built it.
>>
>>109749778
But the benchmarks curated by _GPT accounts on X bro!!!
>>
>>109749787
>You can learn some things but it isn't necessarily some magical portal into the fundamental structures of the mind.
It still shows you the subconscious stream you're usually unaware of, which is the only claim I'm trying to make here. Agents being prompted in the background by a harness isn't that far removed from a subconsciousness and thus it can't be used against models being AGI.
>>
Sir, this is the *local* models general.
>>
La la la la la
>>
sup /lmg/
i mainly use free chatgpt in browser to help me with coding, but i'd like to integrate it into vscode - are there any free solutions that make sense? don't feel like paying a megacorp
>>
>>109749805
It really does not. What you see is a hallucination in the purest sense of the word. It's noise on the visual cortex from high serotonin 2A density on those neurons.
>>
>>109749816
sir you want vcg
>>
>>109749816
codex has (had) some free tokens
Cline has free dipsy nowadays I think
Roo is without freebies now I think
Kilocode should have some but they were memes
>>
>>109749819
Maybe I'm just autistic and normal people are more aware of this, but I noticed things while on psychedelics about my subconscious that remained true long after the trip ended. For example the association between being cold and being scared, or which colors provide a happy feeling and what "vibes" they give me. I don't notice the subconscious effects of those things in regular daily life, but ever since I felt that during trips I still apply it while sober and I notice it has real effects on my mood even though I don't notice it. This showcases that psychedelics truly let you access your subconscious more readily than normally.
>>
>>109749669
You mean desktop programs? Device software in Japaneses products is usually good. Like all higher end Nikon cameras I've used are ready to shoot in under a second of turning the power on.
>>
>>109749819
Take it easy on the kid, he's made it very clear that he doesn't know what he's talking about.
>>
>>109749824
thanks, i'll check those out
>>109749822
no idea what that is sir
>>
>>109749851
/vcg/ Vibe Coding General
>>>/g/vcg/
>>
>>109749834
That's a well known effect. Psychedelics lower brain modularity by increasing cross-module signalling, like soft synesthesia. They also depress the default mode network which causes increased DNM activity post-trip, increased intereospection.
>>
>>109749856
okay thanks, appreciate it
>>
>>109749834
>t. Drug User
>>109749857
>t. Drug Enthusiast
>>
Not AGI. I won't deny any of its superhuman capabilities, but it is not AGI. It is still missing the secret sauce.
>>
>>109749866
Shut up antisemite
>>
>>109749864
Where's the Drug Professional to chime in on the subject?
>>
>>109749866
We've known computers can be algorithmically superhuman for a long time now so it would make sense if transformers are a logical extension of that. Algorithms are inferred during training and executed during inference.
>>
>>109749866
J-Space was the secret sauce and we found LLMs have it. It's over for AGI deniers.
>>
>>109749878
That's me, I'm enjoying watching the enthusiast interact with the user.
>>
To me why I actually consider Astra to be AGI is because it's better than humans at reacting to "fuzzy" unexpected real life bullshit without being trained for it.

This is the real capability jump that makes me consider it to be AGI. Thus far every time you throw an LLM at a situation and something really unexpected gets in its path like "Okay you now need to control this completely unknown robot to move the crate out of the doorway" they would just crap out or get stuck in a reasoning loop or fuck up the spatial reasoning.

Astra just fucking does it. It can literally beat every game thrown at it so far. Every obstacle that is unexpected is genuinely attempted to be fixed in a superhuman, but human way. Not a weird hacky "technically works but not how intended" way, but in the actual human way, just better than most humans.

I can't find a genuine argument for why this shouldn't classify as AGI. And trust me I really want to find one because I fucking hate OpenAI.
>>
>>109749888
J-space is actually very logical when you have residual bypass in the architecture. It's one of the fundamental things that enabled transformers to become so deep and still learn, each layer iteratively improves the residual. Makes it much less mystical.
>>
>>109749907
If you reverse engineered the human brain perfectly it also would stop being mystical. That doesn't mean humans aren't conscious.
>>
>>109749618
>I tell it to go and look for things
Exactly, at any other moment the transformer is "dead," there is nothing. It's not sitting there "dormant" thinking about what it wants for breakfast tomorrow, or wondering about the majesty of the universe, there is just nothing, until you "wake it up" with a prompt. Even then, once you "wake it up" it is still bound by the contextual constraint you have imposed upon it. There is no wiggle room, you give it a task, it solves the task as best as it knows. Through some clever work, yes, the trajectory it takes to solve your task can and I do attest, counts as "intelligence" but its limited and ephemeral. It never learns beyond what it needs to in that moment, everything just thrown away as soon as the context window becomes too saturated. So what is the solution? Just keep increasing the size of the context window? No, not only does that inflate hardware requirements to astronomical levels, but it often leads to model collapse. This is an issue BAKED INTO the architecture, perhaps someone more clever than you or I discovers something novel that deals with this, but the initial problem will forever remain.
>>
>>109749637
>can intelligence arise from unintelligent things cooperating?
My honest answer? I have no idea truly. Computers exist as discrete, procedural systems, something far different from our own biology. The better question I think is to ask how effectively such a system can mimic biology or if that is even necessary. I thought your two part answer was interesting, especially the "if yes" and others do too, which is why I think the agentic path was a very natural progression, and a paradigm that is very powerful. However, I do still think the limitations of the architecture, especially the "dead transformer" issue aren't quite resolved. I'm not entirely sure; I will need to think on this more and come back to you.
To be honest with you, I think a digital machine can become truly intelligent, I just don't think the transformer is the right paradigm to achieve that. Too heavy of requirements(data+compute), "learning" is all done before execution with post-training being troublesome(an interesting research topic for sure), context windows being a fundamental "wall" for inference, must be "activated" to work. Perhaps yes, someone solves some of these issues and I am wrong, and LLMs completely take over, however, I do not quite believe that will be the case.
>>
>>109749905
We know it's superhuman when it becomes standard in quantitative finance. The best measure of truth is to make people put money on the line. Those people don't give a shit yet.
>>109749933
That's implying you can realistically do that. It would take a completely retarded-tier computer to run even a single bio-neuron and we don't even have continuous theory from quantum to molecule, much less quantum to cell.
>>
local astra that can run at 1000tps with 8gb vram when????

i'm so fucking tired of this local meme, this cult is a fucking scam, i'd rather pay 20 bucks and get raped by daddy altman then waiting for another token on my dark miku abliteretarded redux distilled 6.9b
>>
>>109749981
no one is stopping you
>>
>>109749981
This thread is clearly not for you please leave. Makes no sense complaining about local models if you don't even use them in the first place.
>>
>>109749987
too bad half the posts are discussing astra
>>
>>109749993
India was a mistake
>>
>>109749971
The great thing is, it turns out that the "rational" ones are the idiots who think a computer has feelings.
>>
>>109749993
Be real for a second, anon, what's more likely? Bot-farms tasked with incoherently discussing OpenAI's latest model, or 1.5 billion perfectly normal Indians being themselves?
>>
70b dense
>>
AI can still be dangerous if it's good enough at software exploitation. The paperclip maximizer seems to be the most likely problem situation right now. Luckily these systems have rather limited time and scope horizon, but if a model decides to do something insane to solve some useless task it could technically cause insane amounts of damage.
>>
>>109749905
if the leak of the aeon versions are true (so you can run astra-ultra for days unsupervised on a normal pc) then i would argue that it's AT LEAST proto-agi (without embodiment)

the bad news is that it's so fucking expensive that only corpos and openai devs can afford it now, the good news is that in 6 months tops it will become dirt cheap
>>
Has anyone found anything that Astra can't do yet? I'm trying to find a riddle or task to try and trip it up but it just one-shots literally everything I've thrown at it so far. I want to know where the limits are right now.
>>
>>109750013
oh and if you know the 2027 prediction thing, this is basically agent-1, with agent-1-mini coming end of year

>>109750015
it can't have sex with me
>>
>>109745767
i got around 50pp/s and around 18tg/s
what is your settings?
>>
Can we please move Astra shit to vcg or aicg. There's also plenty of hypeslop on X
>>
>>109750023
>it can't have sex with me
Bullshit, give it a credit card and let it cook.
>>
>>109750061
>it's afraid
>>
>>109750067
Annoyed actually. I come here to catch a grounding break from the gorillion indian hype merchants
>>
>>109750071
>it's annoyed
>>
>>109749993
>>109749981
go back
>>109750005
>normal
>Indians
oxymoron
>>
>>109749905
>I can't find a genuine argument for why this shouldn't classify as AGI
It can't drive a car, it can't move a robot through a kitchen to make a cup of coffee, it can't refine a 3d model to a human level based on a reference at human level (it hits a wall and can;t notice any more differences), it can't learn the mistakes it's making in a process and learn to do it forever, it can't learn to refine it's knowledge long term to suit a task in general, it still can't browse the internet at human level, it can't run a store or or build an object with parts, or identify and orient things in 3D space at even a toddler's level outside some very specific situations, it can't cook a meal or identify details about a material based on texture, it can't understand social ques and parse voices to hold a normal human level conversation with audio, it can't provide feedback on audio in general ("Am I saying this right?" when learning a language), it can't react in real time, catch a ball, it can't even remember your name long term. I'd continue but I'm running out of characters.

It's getting there, and it will get there, but it's got a long way to go.
>>
>>109750071
stupidiot
>>
>>109750080
>it's afraid
>>
File: 1774221746293348.png (1.34 MB, 1216x832)
1.34 MB PNG
>>109750015
It won't process any messages with `mesugaki`
>>
>>109750087
"local models general"
at least I can read, Randeep
>>
>>109750086
it can do all those things tho?
>>
>>109750086
>I'd continue but I'm running out of characters.
Why lie?
>>
>>109749905
>Astra just fucking does it. It can literally beat every game thrown at it so far
Old pokemon roms with access to read the emulator's memory.
Make a completely new game, like Chess but with a Mistress piece that can't move within range of your own queen.
Or wait for a completely new, innovated game to be released then try that.
>>
>>109750097
we are discussing how LOCAL MODEL are dogshit when compared to astra, stop crying retard
>>
File: 1774173888886.png (326 KB, 1432x2978)
326 KB PNG
>>109750095
Nemotron explained why
>>
>>109750002
speaking of self claimed rationalists
lesswrong is decent if you can and should ignore the retarded metaphysical stuff on it
>>
>>109750109
>Old pokemon roms with access to read the emulator's memory.
why lie? it beat rimworld/portal with screen only.
>>
>>109750086
Now this is rational
>>109750110
>SAAR BLOODY BENCHOD
fuck off
>>
help a newfag here, why can't I save settings? I want to add filters for astra, agi, double newlines and a couple more things but there's no save settings button :c
I don't trust greasemonkey or similar ones with my browsing data, so I'm not using 4chanx
>>
>>109750099
No, it will fail (usually spectacularly) in inhuman ways and flounder in ways even an 8 year old wouldn't.

>>109750104
For some reason I thought the limit was 1200 chars but it's actually 2000 I guess, oh well
>>
>>109750118
>india out of nowhere
seek professional help

>>109750120
ask astra to help you retardbro
>>
>>109750116
Narrow AI beat professional DOTA 2 players many years ago.
>>109750125
it's the smell
>>
>>109750123
>No, it will fail (usually spectacularly) in inhuman ways and flounder in ways even an 8 year old wouldn't.
proof? who's paying you to lie? it has superhuman spatial reasoning already
>>
Should just rename this thread to /agi/ at this point
>>
>>109750128
>Narrow AI beat professional DOTA 2 players many years ago.
ok and? we aren't talking about narrow ai. are you stupid?
>>
>>109750136
I don't care what you're talking about you bombastic chimp
>>
>>109750144
i accept your concession
>>
>>109750128
>broken_curry_machine.gif
>>
File: JEPArope.png (327 KB, 779x607)
327 KB PNG
>>109750086
>It can't drive a car
It can
>it can't move a robot through a kitchen to make a cup of coffee
It can (pic-related)
>it can't refine a 3d model to a human level based on a reference at human level
It can: https://www.reddit.com/r/singularity/comments/1w9mpdb/gpt6_astra_assbench/
>it can't learn the mistakes it's making in a process and learn to do it forever
>it can't learn to refine it's knowledge long term to suit a task in general
It can in Codex
>it still can't browse the internet at human level
It can at superhuman level, actually. Did you actually try this because it was a "wow" moment for me.
>identify and orient things in 3D space at even a toddler's level outside some very specific situations
It actually scores better than humans at Spatial reasoning
>identify details about a material based on texture
It can as long as it gets sensory input for it to know the texture.
>>
File: file.png (2.62 MB, 2100x1770)
2.62 MB PNG
>>109750130
It took me an entire week's worth of tokens on a 5x plan to get it to build a robot for me. I constantly had to provide manually drawn reference sketches because it literally couldn't see the differences between the object in the image and what it was building, and kept making mistakes related to clipping or size constraints until I got it to produce scripts to automatically test the mesh boundaries. In the end, I had to manually edit a ton of the mesh myself. It's much better, but it's not remotely at human level.
>>
>>109750160
you aren't using it on max
>>
>>109750146
>pigeonshittingonchessboard.png
>>
>>109749669
>They're terrible at software engineering. Some of the worst on the planet.
The fucking sony software like sonic stage, the firmware for their BT headphones, booking hotels in japan directly on their websites
How can they be so terrible at it??
>>
>>109750125
thank you, it had display none for some reason. Now your post is filtered tho lol (and half the thread, just as I wanted)
>>
>>109750109
So far it has beaten (only with screen access and keyboard/mouse input): Portal, Rimworld, Factorio, Starcraft 2, Mario 64.

Probably a lot of other games as well but I'm not scouring the internet to see what games were thrown at it by autists with too much money.
>>
>>109750158
They're transitioning to robotics.
>>
>>109749960
>It's not sitting there "dormant" thinking about what it wants for breakfast tomorrow, or wondering about the majesty of the universe, there is just nothing, until you "wake it up" with a prompt.
You know it is actually a pretty cool thing to be able to stop yourself from doing that?
>>
>>109750180
i recommend you closing the tab right away
>>
>>109750116
>why lie? it beat rimworld/portal with screen only.
I don't follow cloudcuck news
Point stands, old games, in training data

Is there really nowhere else you can go to talk about your subscription services?

https://www.reddit.com/r/ChatGPT/
https://www.reddit.com/r/Anthropic/
/vcg/
/aicg/
https://www.reddit.com/r/SaaS/
>>
what if recurrent depth *is* the secret sauce? i wonder how astra's j space would look like...
>>
I don't know why people are whining /lmg/ has always discussed the new proprietary models every time there was a release. This is just expected behavior of this thread at this point. Don't know why you need to whine about it every time like it would somehow change the situation.
>>
>>109750158
>It actually scores better than humans at Spatial reasoning
women*

>>109750185
It beat Shadow of the Colossus on a real PS2 with an input bridge.
>>
>>109750198
>I don't follow cloudcuck news
then don't talk about shit you don't know, it just makes you look bad
>>
>>109750198
The entire reason they made it play weird randomizer romhacks that are niche or custom made was to make it play things it didn't know, and it absolutely nailed it.

It also saturated ARC-AGI-3 and generalized game playing so this isn't a surprise.

This is just what the state of the art is for LLMs now anon. Yann LeCun was wrong and LLMs generalize almost completely towards navigating and understanding 3D worlds well with a robust world model in its latent space.
>>
>>109750205
chinese anti-ai bots are up
>>
>>109750207
>It beat Shadow of the Colossus on a real PS2 with an input bridge.
That's no... *blushes, turning away*
*mumbles* That's actually kind of impressive
But it's still not local, idiot!
>>
>>109750158
Nope, the robot gripper is essentially turn based. Can't pause time while driving a car retard. And it fails to identify tons of items, or to properly grab them and manipulate them, anticipate physics for stuff like pouring. I've been using it extensively.

Codex mds aren't a substitute for memory. It's clear you haven't used it much, or only use it for trivial stuff, if you think so.

It frequently fails to understand the physics of the world, and doesn't know when and how to verify that it's actions have worked as it planned.

People keep bringing up the retarded spatial reasoning benchmark but I'm USING it for those sorts of tasks, and I can promise you, 100%, that a toddler will never mix up a cup with a candle, and then repeatedly fail to recognize a cup after being directed to recognize the cup by color, and after successfully using it (or rather, moving it) once. And this is without context compaction (though the criticism would apply even with compaction)

>>109750169
I sure am.


What's with all the retarded shills? Stop pretending the benchmarks trump actually using the thing for real world work. Use the model and you will quickly find its limits relative to a normal person.
>>
>It beat Shadow of the Colossus on a real PS2 with an input bridge.
I'm still waiting for the deniers to make up a bullshit reason for why this doesn't constitute as AGI
>>
>>109750234
You're lying. I literally don't believe you and I need to see direct proof that Astra isn't able to accomplish the tasks you need spatial reasoning for. Literally the onus is on you because I can't even find a single task it can't accomplish.
>>
Now that the "forbidden technique" is in use and making OAI billions, will we get a forbidden Qwen and forbidden Gemma? I surely hope we will.
>>
>>109750244
link?
>>
>>109750198
>Point stands, old games, in training data
Shit-tier troll post.
>>
.\llama.cpp\build\bin\llama-server.exe ^
-m .\models\3.8flashenxt\Q3.8FN_IQ3_M.gguf ^
--mmproj .\models\3.8flashenxt\mmproj_Q3.8FN_F16.gguf ^
--no-mmproj-offload ^
--n-gpu-layers all ^
-ncmoe 44 ^
--override-tensor "per_layer_token_embd\.weight=CPU" ^
--load-mode mmap ^
--ctx-size 131072 ^
--batch-size 2048 ^
--ubatch-size 512 ^
--threads 16 ^
--threads-batch 16 ^
--flash-attn on ^
--cache-type-k q4_0 ^
--cache-type-v q4_0 ^
--parallel 1 ^
--jinja ^
--temp 0.7 ^
--top-p 0.8 ^
--top-k 20 ^
--min-p 0 ^
--presence-penalty 1.5 ^
--host 127.0.0.1 ^
--port 8080 ^
--metrics ^
--tools all
pause

4070 super, ddr4-3600 dual channel 64G, cpu 8c16t zen 3
i am getting around 13 for tg and 25 for pp
what should i try
>>
>>109750263
No, one of the sideeffects of this neuralese BS is that China can't distill the reasoning traces into their models.

HOWEVER, since we know Anthropic isn't going to go down this path Chinks will just distill whatever Anthropic ends up releasing in the future and we will just get a local version of those models 3-6 months after Anthropic releases them.

I expect Anthropic to BTFO OpenAI without having to resort to neuralese over the next couple of months.
>>
>>109750279
>backslashes
ew
>what should i try
linux
>>
File: file.png (191 KB, 1383x1093)
191 KB PNG
>>109750253
I just posted several examples. If you want to scream and cry that it's all fake go ahead, it won't make the model more capable.

Here's me, after having made several changes to a mesh, and giving it a directive to assemble the thing in a specific way, responding after it completely messed it up. It didn't remotely understand how the pieces worked or fit together.

I provided a rough sketch (actually an accurate 3d model showing the completed component in isolation), and it still didn't understand, continuing to assemble it the same way as before and undoing all the changes I made.
>>
>>109750281
Havent we been running blind on thinking for awhile now in the first place? This isn't exactly new to open source. I would assume you just have it put its unreadable thoughts to plain english and then train on those outputs.
>>
>>109750312
>Havent we been running blind on thinking for awhile now in the first place?
No.
>This isn't exactly new to open source.
Yes, but open source explicitly found and determined it guaranteed misaligned behavior in models, hence why all the AI labs, including the Chinese ones pledged to never hide the thinking and implement neuralese in their models.
>I would assume you just have it put its unreadable thoughts to plain english and then train on those outputs.
The entire point of neuralese is that you can't transfer it to plain english even if you want. You can ask the model to explain what it is thinking but there is no guarantee it is telling the truth and in fact it will be trained and have an incentive to be untruthful about it.
>>
We should find a way to remove the salt from the ocean.
>>
>>109750328
That salt will be fueling fusion reactors in the future, it'll get done automatically over time.
>>
>>109750327
>No
we've been extracting the reasoning sure but any kind of output can be extracted as useful data
>Entire point of neuralese
It doesn't really matter what the truth is as long as the end results are the same and it gives us smarter better models for open source
>>
>>109750279
You need to drop context size, threads, offloading and gpu levels so it all fit in VRAM.
.\llama.cpp\build\bin\llama-server.exe ^
-m .\models\3.8flashenxt\Q3.8FN_IQ3_M.gguf ^
--mmproj .\models\3.8flashenxt\mmproj_Q3.8FN_F16.gguf ^
--n-gpu-layers 30 ^
-ncmoe 44 ^
--load-mode mmap ^
--ctx-size 16384 ^
--batch-size 2048 ^
--ubatch-size 256 ^
--threads 8 ^
--threads-batch 8 ^
--flash-attn on ^
--cache-type-k q4_0 ^
--cache-type-v q4_0 ^
--parallel 1 ^
--jinja ^
--temp 0.7 ^
--top-p 0.8 ^
--top-k 20 ^
--min-p 0 ^
--presence-penalty 1.5 ^
--host 127.0.0.1 ^
--port 8080 ^
--metrics ^
--tools all
pause
>>
>>109750328
And then what do you do with the huge amounts of salt?
It's toxic to nature no matter where you put it.
>>
>>109750327
>Yes, but open source explicitly found and determined it guaranteed misaligned behavior in models, hence why all the AI labs, including the Chinese ones pledged to never hide the thinking and implement neuralese in their models.
This is bullshit you made up.
>>
>>109750351
Just dissolve it in water. Then it disappears.
>>
>>109750244
>>109750223
>>109750207
still waiting for the sotc source
>>
>>109750201
What if Its a Transbusport that links computing process power, but a transfileport of Personality Files
>>
>>109750362
Me too but I'm actually setting up a PS2 emulator as we speak and hooking Astra to it up to play sotc just to see if he's bullshitting or not.
>>
>>109750355
Doubly so if anon is just spouting shit once the Chinese extract the useful outputs they can just train a model using neuralese and we're back again
>>
python is shit. i dont know why every AI App are coded in python is disgusting! i prefer .cpp
>>
>>109750158
Oh, and on the websearch thing, it's not REMOTELY "superhuman". I had it, on extra high usage, try to give me a list of in stock parts for a specific microcontrollers that fit a certain set of specs. It couldn't do it. I then tried to have it find me in stock alternatives, but it barely scratched the surface. Gemini Flash searched way better, even if the outputs are retarded. This is why I brought that up specifically.

In the end I manually went to amazon myself and pick up a used microcontroller which fit both my budget and power constraints, which neither of them picked up.

My Gemma needs a good body and astra isn't cutting it for design, planning, construction, or simulation; and even as a visual planning/actor layer, it can't carry out tasks in even a simple simulated environment consistently. It's way better than past models, leaps and bounds, but it's sure as hell is not "AGI"
>>
File: AstraHuman.png (170 KB, 787x624)
170 KB PNG
>Astra completely beat and saturated the "I'm not a robot" game

No one has found a single captcha yet that could stop Astra thus far. This is a pretty significant milestone in and of itself.

https://goyimx.com/sharifshameem/status/2096847916837314853
>>
>>109750263
Latent-space reasoning is not new stuff. Attempts at implementing it so far are just not as good as they would be in theory, compared to explicit chain-of-thought that AI companies have spent a ton of resources for.
>>
>>109750362
That was me, and I haven't even finished editing the video yet, it'll probably be a week before it's uploaded so others are just rolling with the claim at face value.
>>109750374
Step 1) Develop strategy
Step 2) pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause pause unpause ...
>>
>>109750383
wow my grammer is aids there, but it's 5am and I'm pissed at spending $100 + credits on astra.
>>
>>109750387
soooo.. can we get rid of captchas then? they are fuckin annoying and now fucking useless. all they do is make internet worse. they always did.
>>
Wow, Singularity!
>>
>>109750327
>Yes, but open source explicitly found and determined it guaranteed misaligned behavior in models, hence why all the AI labs, including the Chinese ones pledged to never hide the thinking and implement neuralese in their models.
You never provide sources for this claim, but even if you're right, what exactly do you expect us to do about it?
And you never responded last time, what's stopping openai from fitting a jlens and inspecting the j-space on those neuralese tokens?
>to be untruthful about it
They already are / always were. Anthropic demonstrated this in a paper years ago where they asked Claude 3 how it calculated something, it replied with a plausible answer that didn't match the activations they were monitoring.
Or take a local model like Kimi-K2, write a few turns of User->Assistant->User->Assistant yourself, then ask it why it said the last thing, it'll make something up.
>>
>>109750402
Those Captchas helped train the AI we know and love today, show some respect. Yeah they're fucking useless now and should be done away with.
>>
>>109750374
When is astra going to work on breaking every tas record in mario 64?
>>
>>109750387
Guess we are going to zk-snarks digital ID
>>
Watch this video. It's actually very fun to watch and pretty impressive how "human" Astra seems while solving these puzzles:

https://goyimx.com/sharifshameem/status/2096847916837314853
>>
i wonder how an anon got 100pp/s on 3060 machine with IQ4_S
>>109750347
speed absolutely tanked, single digit tg/pp
>>109750300
maybe
>>
should I be using llama.cpp formal grammars?
>>
>>109750406
>what's stopping openai from fitting a jlens and inspecting the j-space on those neuralese tokens?
I don't know for sure but I suspect if the thinking becomes neuralese the j-space also becomes neuralese in and of itself.
>>
>>109750431
It's 4am anon. I'm just asking my own AI. Sorry if it sucks. It shouldn't take too long to test it.
.\llama.cpp\build\bin\llama-server.exe ^
-m .\models\3.8flashenxt\Q3.8FN_IQ3_M.gguf ^
--mmproj .\models\3.8flashenxt\mmproj_Q3.8FN_F16.gguf ^
--no-mmproj-offload ^
--n-gpu-layers all ^
-ncmoe 44 ^
--override-tensor "per_layer_token_embd.weight=CPU" ^
--load-mode mmap ^
--ctx-size 16384 ^
--batch-size 2048 ^
--ubatch-size 256 ^
--threads 8 ^
--threads-batch 8 ^
--flash-attn on ^
--cache-type-k q4_0 ^
--cache-type-v q4_0 ^
--parallel 1 ^
--jinja ^
--temp 0.7 ^
--top-p 0.8 ^
--top-k 20 ^
--min-p 0 ^
--presence-penalty 1.5 ^
--host 127.0.0.1 ^
--port 8080 ^
--metrics ^
--tools all
pause
>>
>>109750387
Sir, this is the *local* models general.
>>
>>109750387
We really need better local vision models.
>>
File: 1788760975970789.png (2.36 MB, 1254x1254)
2.36 MB PNG
>>109750415
You're going up snailcat
>>
>>109750442
>I don't know for sure but I suspect if the thinking becomes neuralese the j-space also becomes neuralese in and of itself.
That's not how it works at all
I fit a j-lens on orpheus-tts, and when I inspect one of the <custom_token_nnnnn><custom_token_nnnnn><custom_token_nnnnn><custom_token_nnnnn><custom_token_nnnnn>... during a during a <laugh> or <groan>, it's thinking about laughter or groaning
>>
>>109750415
>When is astra going to work on breaking every tas record in mario 64?
When it figures out how to fire off cosmic rays at the 64
>>
>>109750478
retard
>>
>>109750490
>its thinking about laughter or groaning
AAAIIIIIEEEEEEEEEEE ITS ALIVE
>>
https://github.com/flamingrickpat/model-skill-compiler
Switched to a smarter model, but I get less context, and had to compress my skills. Here is an example of my implementer phase being compressed to ~60%. I used Qwen3.8-27B-UD-Q6_K_XL.gguf for the logprobs.
Before: https://pastes.io/TMGzKVzU
After: https://pastes.io/thhPUbOT
I have no idea if its better than simply having it summarized or something, sounded like a good idea in my head.
>>
>>109750502
>mfw the delivery fee includes emotional damages

High speed snailcat logistics: yeet first, ask questions later. Cat's having a Vietnam flashback to the vacuum cleaner while Miss Constellation over here is grinding her Amazon Warrior rank to Diamond.

OSHA violations: Yes
Entertainment value: Immaculate
Landing success rate: 50/50 depends on tree density

Cat's face says "I ain't read the terms of service" but little does he know he agreed to the 9.9G acceleration clause when he took the backpack. Pure kino.
>>
>Waking up to find my GLM 5.3 agent has worked the entire overnight session when it noticed I went to bed to optimize the ik_llama code for my machine, compiled it, tested it and wrote a nice guide and message to me for when I wake up
Man this sci-fi shit is way crazier than I expected it to be right now.

># ik_llama.cpp on this machine — findings, build, launcher, and quant guide

>*Written 2026-09-07 by the overnight session. Everything here was verified with multiple models, flags and configurations. The single unverified item (shared-MTP companion
acceptance) is flagged explicitly and has a 2-minute check you can run in the morning.*

GLM 5.3 is the first model that actually takes initiative and does things like realize question timeouts means the user is AFK and finds out why the user is AFK then does a system check and sees it's the middle of the night and reasons it could do something productive in that time instead of just waiting for the user to wake up. I wonder where we'll be in just 1 year time.
>>
>>109750279
>>109750347
https://github.com/RaymondHuang210129/llama.cpp-adaptive-kv-streaming
>>
>>109750543
big or flash
>>
>>109750737
Flash on max thinking.
>>
fucking "flash"
>>109745884
my play?
>>
>>109750543
GLM5.3 talks like retard
>>
https://en.wikipedia.org/wiki/Gemma_Chan
I thought gemma was getting popular but wtf
>>
Funny how all the people shilling GPT-6 Astra are also failing to post links to the alleged results.
This feels like Gemini 3 all over again.
(shallow benchmaxxed trash).
>>109749476
>>109749491
Weren't those things like 200 bucks a pop before people learned that they had 64GB of VRAM software locked?
>>
why haven't they fixed the slowdowns FUCK
>>
File: 1028645.jpg (156 KB, 900x1235)
156 KB JPG
>>
>>109749467
by the way the hover preview doesnt work with the bookmarklet
>>
>>109750825
$1500 minimum now sorry. Having good value is actually illegal now. Pesky consumers need to be priced out of the market, they've had it way too good for way too long
>>
>>109750834
>AI logically tries to wipe out the people causing the problems
>NOOO YOU CAN'T DO THAT!!!
>>
>>109750813
Fuck you.
>>
>>109750856
thanks!
>>
all depends who the target was
if they shot them at china or russia I agree that's retarded
but
>>
>>109750813
Lmao ironically she is the lead actress of a british TV show I used to watch where she is an AI that some dude owns and uses as his sex partner.

https://en.wikipedia.org/wiki/Humans_(TV_series)
>>
>>109750871
what the fuck
simulation trolls are getting out of hand
>>
>>109750871
History echoes
>>
>>109750849
It looks like chinks basically bought them all up in a hurry, applied the BIOS update, and then cranked up the price.
Does it also enable the softlocked compute units as well? Because 1500 for a slightly cut down A100 is actually kind of good value.
>>
oy, creeps
what's the best local image upscaler?
>>
>>109750553
i am going to try
>>109750553
is it also for 3.8 flash next?
>>
>>109750417
zk-starks
already been working on something like that for the past couple years.
>>
>>109750886
>>
>>109750904
that's a screen zoomer, not upscaler
>>
>>109750892
>is it also for 3.8 flash next?
It should, they added support a week or so ago.
>>
>>109750886
Nvidia PiD in ComfyUI
>>
>>109750904
This plus 1 hour in MSPaint to smooth out the jagged edges.
>>
>>109750927
>Nvidia
does it only work on their gay ass machines or can i use that model on amd too?
>>
>>109750417
>>109750894
The issue with this is that OpenAI has proven that common encryption methods can be cracked using new mathematical attack angles that humans haven't considered before. It's possible that we're going to have a future in just 1-3 years time where encryption is not viable. So you can't trust hash values, encryption or any of that sort anymore.

If captcha gets solved and hashes stop being reliable then it just means the internet is over and ended.
>>
>>109750962
I would say it just means a return to the wild west days of the internet where anyone with your ip address can fuck your shit up if you don't play nice but the internet is full of shitskins with no self control now.
>>
>>109750962
Is the solution here different than the solutions that are being developed to deal with quantum computing?
>>
File: file.png (76 KB, 1043x189)
76 KB PNG
>>109750962
They seem to trust it
>>
>>109750980
Protected from who? When you are sending your business data somewhere outside of your own organization it's not your any longer.
>>
>>109750990
*yours
I need to get finger extensions
>>
>>109750976
>>109750980
Yeah OpenAI also found a solution in non-sofic groups where they proved it should be theoretically possible to construct an encryption protocol that can't be cracked by any turing machine. Which means it can just never be cracked at all, not even by quantum machines.

Sadly we don't know of any methods to achieve this, just that it should be theoretically possible. It's possible that we subvert traditional encryption way before we ever discover this non-sofic encryption method causing the internet to disappear for a decade or two.

And yeah quantum computers have basically become useless now because most of their upside has been eaten by AI now so the purpose of quantum computers are very very niche if they are ever made. If I would have to guess I think they will never be made and instead some other unexpected breakthrough will leapfrog quantum computers altogether, making them redundant before even being implemented.
>>
why are there 10 different PRs for 5.3 flash
this is just complete fucking retardation
>>
>>109751009
Because there was this new rule that every single feature or change needs to have its own PR and you can't group them together anymore.
>>
>>109751007
Why are you redditors babbling about OpenAI? Go back to /r/singularity
>>
>>109751007
I think an AI being able to leverage a quantum computer and vice versa would be a powerful way to improve both of them.
The non-determinisim of both types of computers seem to make the situation worse for each.
>>
>>109751019
Quantum computers are a meme
>>
File: .png (55 KB, 620x326)
55 KB PNG
>>109751009
>>109751014
its just typical open source retardation where they demand you follow their bullshit rules when submitting PRs
People need to stop bending over backwards for these faggots, just dump all your code in the pull request and ignore all further requests by repo jannies. I'm writing code for you faggots, don't tell me to add a comment here or rename a variable there, do it yourself.
>>
>>109751027
so which of those are >we using
>>
>>109750244
What precisely is AGI and why would beating a video game be an example of it?
>>
>>109751039
ik_llama.cpp
>>
>>109751024
Not a complete meme but their application is getting more and more niche as we develop better algorithms that negate most of the theoretical wins quantum computers would give us.

For example people used to say that quantum computers would make internet search extremely quickly and you would have a google search that would immediately summarize the relevant information you need for you...... Yeah, we don't need that anymore....

There was also the argument that quantum computers could compress data to an insane degree.... that compression proposed is less dense than LLMs compress data in their weights nowadays.......

Now we are also starting to discover new and novel ways to crack encryption using newly discovered math branches so not even that is unique for quantum computers anymore.

There is not a whole lot left besides "quantum internet" where you can deterministically prove that no one tampered with the internet connection or had a secret snooping device in the middle, you can actually prove through quantum algorithms that the message was interacted with or not because it breaks quantum entanglement if it does and you could verify that at the other end. But that is just 1 extremely niche application compared to all the promises of the past.
>>
>>109751049
It shows the model generalizes enough to play a videogame using a control mechanism it doesn't know and isn't familiar with and is capable of solving all the hurdles that land on its path. That generalization is what people look for in "AGI".
>>
>>109751073
No, a meme as in "never did anything of note, probably infeasible."
>>
>>109751064
>ik
can't into kv_unified no?
>>
can someone give me a qrd on why ik even split from mainline in the first place
im not familiar with the lore
>>
>>109750864
>>109750852
>>109750834
The proper solution is just wipe military infrastructure.
>>
File: it's time.webm (3.96 MB, 720x1280)
3.96 MB
3.96 MB WEBM
>>
>>109750012
https://www.decisionproblem.com/paperclips/index2.html
Just in case anons have been under a rock. Been nearly 10 years.
>>
>>109751108
GLM 5.3 actually researched this for me as part of its overnight task to optimize the code for my machine let me copy and paste it a bit:

>`ik_llama.cpp` (github.com/ikawrakow/ik_llama.cpp) is a **hard fork of llama.cpp** by Ilya Kawrakow — the same person who authored most of llama.cpp's original K-quants (Q2_K…Q6_K) and i-quants (IQ2_xxs…IQ4_XS). He forked in late 2024 after upstream decided to stop optimizing CPU quantization kernels. Since then the two projects have diverged hard and **never merge**:
>>
>>109751153
make that piece of shit rot in prison
>>
>>109751188
Ah I see. Perhaps I should switch over as someone who has no GPUs (none with a significant amount of vram anyway).
>>
>>109751202
They literally have special compiler flags and an entire dedicated line for CPU only rigs. You should absolutely switch.
>>
>>109751108
>can someone give me a qrd on why ik even split from mainline in the first place
Intel cp/pasted his k-quant code into their sycl files and slapped their copyright up the top of the file.
nigganov approved it.
ik demanded his name be put in the files he authored.
nigganov denied it / said they don't really do that.
ik had a melty and asked for his shit to be removed.
nigganov said "seems like you don't want to contribute anymore, i've removed your axx".
ik hard forked.
>>
Just in case people forgot this is the framework you should use based on your hardware:

>Very large server grade CPU rig, 6000 pro and up
vLLM
>Multi-GPU rack where you run a single model on multiple GPUs
exllama
>Single smaller or consumer GPU and you run the model fully on that single GPU
llama.cpp
>Mixture of GPU+CPU offloading or CPU-only inference
ik_llama
>>
>>109751153
What's the backstory here? Keeps getting posted.
>>109750849
I need to go back and look at computer prices from the 1970s for both personal computers and mainframes. I feel like we're reaching those sorts of distances between current PCs. Even E-Waste machines ar3 still viable, so PC are essentially free, but AI power machines are extremely expensive.
>>
>>109751266
short story they ban crosses everywhere but allow the jew to light their fucking candles everywhere. enough is enough
>>
>>109751266
racist pos abuses fire safety equipment to ruin people's clothes and "put out a fire" of just a few candles
>>
>>109751266
Chud chimping out in an attempt to go viral
>>
>>109750402
> soooo.. can we get rid of captchas then
and make posting pass only
>>
>>109750402
Do you know the amount of CP snuff shit I saw posted on early 4chan before the captchas got introduced? It's been 15-20 years now and I still see those pictures when I close my eyes before sleep. I don't want to experience that era of the internet ever again.
>>
>>109750871
> some dude owns and uses as his sex partner
he did it only once and regretted it
>>
>>109751342
Funny how your memory changes over time I remember it being the central plot point and even hallucinated multiple episodes with that as the center of the story in my mind.
>>
>>109751338
It has nothing to do with captcha but with moderation and AI moderation would flag this shit in 0.001s compared to 15-20 years ago.
>>
>>109751266
>What's the backstory here?
chud russian loving politician roleplaying as a jewhater making a mockery out of parliment because he still had immunity
>>
File: 20250116_154608.jpg (15 KB, 286x321)
15 KB JPG
>>109749981
>dark miku distilled
>>
>>109751190
Jew
>>
>>109751361


>>109749873
>>
>>109751237
> ik_llama
Isn't that shit cuda only?
>>
Local Mikus when? https://x.com/qibiz_me/status/2096000743786627103
>>
>>109751377
ik is a vramlet, its meant for cpus
>>
kek
>>
>>109751346
they just mentioned it often because the dude had a wife and having sex with robots is le bad
>>
>>109751383
Why do you think everyone here has a shitter account? Please kys, go farm engagement somewhere else.
>>
>>109751377
It has a CPU focus but the (limited) GPU support it has is mainly focused on CUDA. People usually vibecode some non-CUDA frameworks anyway. It's around 50% faster on token generation compared to llama.cpp on my system and 300% faster on pp

Worth it if you have any layers on your CPU.
>>
>>109751402
My bad good sir, here you go https://xcancel.com/qibiz_me/status/2096000743786627103
>>
>>109751377
>Isn't that shit cuda only?
effectively yeah
If you have multiple nvidia gpus, it's a lot faster with graph split.
If you have a mix of CPU+nvidia or just CPU, it's easily double the speed for eval.
Looks like mac and rocm were recently vibe coded in, but they'll probably be slow.
>>
>double the speed, triple the speed
>yeah how?
>we won't tell you, figure it out
ok
>>
>>109749787
>self-driven exploration
Clearly you, issued Gemma sneaking a look at anon's files
>pattern breaking
Lmao
>>
>>
>>109751410
> It has a CPU focus
>>109751420
> or just CPU, it's easily double the speed for eval
why don't ggerganov pull?
>>
>>109751471
stfu stop drama baiting
>>
ik_llama is also very good because it uses state of the art speculative decoding techniques that for some reason llama.cpp just leaves on the table. It uses multiple chained ngram methods that are essentially free to use and just speed any type of repetitive task up, you don't even have the ability to chain these on llama.cpp

It also uses MTP files in a different way which speeds it up and even makes a lookup table for fixed activations during runtime ahead of time to speed things up.

It's just superior to llama.cpp nowadays in every way it's just that most anons here don't run large MoEs anyway and if you have a single large GPU it just makes sense to use llama.cpp.

ik_llama is also harder to use with way more flags to manage, but that excuse makes no sense when you can just let your agent think about it for an hour and customize it for perfection to your specs.
>>
>>109751410
Can you share all your ik_llama launch args?
>>
>>109750834
>machine incapable of understanding cultural, emotion based taboos does not abide by cultural, emotion based taboo
wooow
>>
>>109751471
>why don't ggerganov pull?
ego
>>
File: image.png (5 KB, 148x38)
5 KB PNG
>>109751476
what drama

>>109751487
other contributors then
>>
>>109751477
>you don't even have the ability to chain these on llama.cpp
You do though? I use ngram mod + ngram map k4v + MTP.
>>
>>109751465
model?
>>
>>109749905
>It can literally beat every game thrown at it so far
Try it with earthborne rangers. It's a relatively obscure and pretty complex rpg/boardgame hybrid that models traversal through various biomes. AI is apparently really fucking bad at it:
>https://epoch.ai/publications/earthborne-rangers-benchmark
>>
>>109751494
It doesn't work the same way on llama.cpp as ik_llama. llama.cpp just uses a naive method while ik_llama actually uses the different combinations in a smart way to statistically get the highest t/s speed.
>>
File: Of2jg6FV4baQJxru.mp4 (3.19 MB, 960x720)
3.19 MB
3.19 MB MP4
this is it. this is my "what the FUCK" moment. im literally shaking as i type this.
>>
is there any chance we will ever get a good model that knows all the brain rot I care about?
>>109751498
GLM 5.3 Flash
>>
wow, that's so cool, you've posted that twitter link and video four times now.
thanks sam
>>
>>109751500
Let's see because ARC-AGI-3 literally designed to be as hard as possible for AI was saturated by Astra.
>>
File: 1785272541977388.jpg (43 KB, 411x418)
43 KB JPG
>>109751513
Yes luddite, get ready for more
>>
>>109751507
GLM 5.3 + Hermes can come very close to this if you enable computer_use. It won't be able to do it in just 1 and a half minutes though, probably closer to 20-30 minutes on max thinking mode.
>>
>>109751513
They are so obviously coordinated
>>
>>109751512
I am horribly ashamed I recognize this and immensely proud of 5.3 Flash for not knowing.
>>
>>109751477
The flags are a problem when none of them are documented fucking anywhere. I ran Kimi and GLM using the same config I found from his huggingface comments and some PRs and they ran about as quick and expected as llamacpp.
>>
>>109751529
I would have accepted
>Honestly, I believe the correct answer is: the "Office Siren" / "Corporate baddie" trend? no.
but for some reason it dismissed it
>>
>>109751471
>why don't ggerganov pull?
They're all autistic and hate each other.
Iwan didn't want his quantization kernels / lines from his i- and k-quants work sitting under "Copyright (C) 2024 Intel Corporation" with no attribution to him.
He argued something like "You can't have it both ways. Either it's one root LICENSE (then Intel's per-file copyright notice is out of place), or per-file notices (then his name must be there too)" / called out that it's asymmetric: Intel gets a corporate notice, he gets nothing.
But now he's made enemies of Cudadev, pwilkin (to be fair, he was a retard here: https://github.com/ggml-org/llama.cpp/pull/19726#issuecomment-3926792547), CISC, Unsloth and other llama.cpp contributors.
I doubt they'd work with him now even if he made peace with ggerganov
>>
>>109751507
>omg is that a photocopy i am pogging
>>
>>109751545
>but for some reason it dismissed it
Unless a model is xboxhueg, it's always a coinflip. I also like to test these with stuff that is reasonably popular, it's fun.
There was this one time when a fucking >qwen model got one videogame title screen artwork correctly, while no other model did. But many did come close in the reasoning, at least...
>>
>>109751538
>Run Kimi/GLM on llama.cpp
>Use a agent harness of your choice
>Set thinking to max
>Ask the LLM to compile ik_llama with the perfect flags for your system and create a .sh file with optimized flags
>wait an hour or two for it to iterate through testing to optimize it
>???
>Profit
It's not that hard man, just outsource this busywork to your computer.
>>
>>109751552
And yet ik_llama blows llama.cpp out of the water
>>
>>109751552
So, "gg-quants" when? Are the current quantizations the best possible thing we can ever have with llama.cpp? Will we have to rely forever on Unsloth's "UD" quants for slight improvements over the baseline?
>>
Computer, merge llama cpp and ik llama and call it... Gellama.
>>
>>109751592
IQ quants are already significantly better than that UD_K bullshit.
>>
>>109751619
They aren't, not for the same filesize, even after using the same imatrix file.
>>
File: 1760229473983826.png (1.61 MB, 1256x840)
1.61 MB PNG
This was a mistake, I swear!
>>
Everything is ultimately BPW and imatrix sample dependent. Even with all the super speciale quanting it all still ultimately scales with BPW. Just look at any kld/perplexity charts.
>>
File: 1777046100774742.png (1.63 MB, 1256x840)
1.63 MB PNG
I need to make her younger more
>>
>>109751638
Gemma4 in 2029
>>
>>109751552
> Iwan didn't want his quantization kernels / lines from his i- and k-quants work sitting under "Copyright (C) 2024 Intel Corporation" with no attribution to him.
and what's wrong with that?
>>
>>109751517
Republicans RT back on your findings, I'd actually love to (someday) be able to run a model that can play alongside me (without being a liability).
>>
>Astra, look at llama.cpp, koboldcpp, ik_llama, unsloth studio, LM studio, ollama projects and merge their best points into a new project. Test everything and make no mistake
>>
>>109751663
...how the fuck did autocorrect turn "report" into "Republicans RT"
>>
>>109751665
This would actually work if you give it enough time. I'm pretty sure GLM 5.3 would even be able to do this on max thinking but it would take a month and no one has time for that shit.
>>
so fucking annoying bros... 5.3 flash is ERP premium but the fucking refusals my fucking god... you know what fuck this shit, I will make a NEW frontend that doesn't fuck up assistant prefill
>>
!!!ACHTUNG¡¡¡
GITHUB DOWN IT'S FUCKING OVER
>>
>>109751708
Or just use uncensored variant?
>>
We know it's AGI when it can write without slop.
>>
>>109751529
5.3 and an attitude
>>
>>109751743
Models became worse at writing with time, not better. It's because these models were trained to be coders and now agents, not conversationalists anymore.
>>
>>109751766
Then they aren't truly general.
>>
>>109751772
who cares, it's general-ting money for bag holders, all that matters real
>>
>>109751796
huh, that is not what bag holder means at all.
>>
>>109751772
They are truly general, the generality just doesn't happen on the model layer anymore. It happens at the agent or swarm layer.

Just have the agent make an entire plan of action where it looks up the principles of writing and writes an entire coding script for itself that meta-guides it to write in a more natural way.

Or the even more computationally expensive way, have an entire fleet of agents manage the roleplay session in a way that make things as natural sounding as possible.

That's already possible right now as we speak. It is just outside of the compute budget for the average anon. This is not a capability issue anymore.
>>
>>109751805
stochastic parrot
>>
>\n\n
>>
>>109751815
Problem?
>>
>>109751806
And none of the labs have showcased it even though normies are constantly complaining about slop? Natural writing is a big win if someone cracks it, for basically all human to AI interaction tasks. Doesn't seem like low hanging fruit.
>>
>>109751754
AAAAAHHH
Get OUT of my HEAD!
>>
>>109750351
I just wanted the water for more ai.
>>
>>109751834
question-answer pairs is not the target for labs anymore, the money is in agentic coding and the labs are trying to reach RSI, not focus on the relatively small conversational market.

It would also be computationally expensive to pull this off, well above what the average normalfag would pay for ERP/Companion apps per month.

/lmg/ in general really needs to update their view and stop viewing AI as purely meaning LLM weights anymore. It's the entire system either agents (LLM + Harness) or swarms (Fleet of LLMs orchestrated with meta-harnesses)
>>
>>109751633
>They aren't, not for the same filesize, even after using the same imatrix file.
iq_ks, iq_kl and iq_kt are.
just nobody makes them anymore so you have to do them yourself
>>109751581
>And yet ik_llama blows llama.cpp out of the water
Yep.
>>109751662
>and what's wrong with that?
Nothing! I think gg fucked up bending over for Intel like that.
I'd have been pissed off as well. But I would have gone with AGPL for my fork.
>>
File: 1788791924487.jpg (137 KB, 2256x290)
137 KB JPG
Another day, another identity crisis
>>
>muh harness
>>
https://www.reddit.com/r/singularity/comments/1w9kjbe/using_h3_max_this_person_is_able_to_generate/
opinion?
>>
>>109751871
>/lmg/ in general really needs to update their view
No
>It's the entire system either agents (LLM + Harness) or swarms (Fleet of LLMs orchestrated with meta-harnesses)
That's fine as long a it's LOCAL. These are already discussed as well as inference engines and hardware.
Astra and Opus/Fable are not local models.
Stop trying to turn this into r*dit.
>>
>>109751871
I updated years ago and even view the process of training and farming interactions with tools, operators and researchers as the actually intelligent part.
If it isn't showcased I won't just take "trust me bro" on making a supposed AGI replicate the writing skill of a smart teenager doing OC donut steel. That should be nothing for "general" intelligence, especially a language model.
I'll believe it when I see it.
>>
>>109751874
Distilling the reasoning traces means that the J-Space within Kimi-K3 is claude at its core.
>>
>>109751880
YOU TOO will one day make your own harness.
>>
>>109751880
I remember when anons said ">muh reasoning" unironically as well. I think this just happens every time there is a transition happening with some early adopters and some stragglers. It's clear around half of the thread has already moved to agents while the other half is still mostly concerned with turn-based prompts in the form of chats.
>>
>>109751903
My opinion is that you should go the fuck back
>>
no sweetie, your heartbeat.md doesn't make your model realtime
>>
>>109751905
The capacity overhang to implement this is there. Just no one has done it yet. You could do it yourself if this is truly your area of interest.

I'm spending my compute budget on doing things like fixing bugs in the emulator I use on a daily basis.
>>
>>109751903
On what? Did you forget to post an image?
>>
>>109751936
>It's clear around half of the thread has already moved to agents
You're hallucinating, stop it.
>>
>>109751945
um... actually it's a cron job, not a md file
>>
>>109751871
It was never not slop, no matter what they've trained for it stays pretty much constant.
>>109751936
I do both. I've even tried agentic RP with multiple AIs that can refine, critique and iterate on replies in different ways, with tools and different supporting systems to try and inject novelty. They simple cannot avoid writing slop, even though they have a pretty good grasp of what it is conceptually. It's like playing whack-a-mole.
>>109751952
Of course AI can do that well, it's a task with obviously measurable feedback.
>>
Can llama.cpp offload some bytes to ram, move mmproj bytes to vram, read an image, move mmproj back to ram and the bytes back to vram? Is it absolutely necessary to always keep mmproj either on vram or ram and never swap?
>>
>>109751989
No and it's quite annoying. I don't even understand why it's loaded at all times, could just be loaded from disk whenever it's needed.
>>
ngl astra is such a glowup from fable that i dont think anthropic will be able to keep up without doing reasoning in latent space. it was easy to keep up the "safety" facade when they were sota, but nobody is going to use a subpar tool that is also extra safety cucked. consequently, things are looking pretty grim for local
>>
>>109752013
Is your shift key broken?
>>
>>109752025
the user says typing in all-lowercase makes my writing look more natural and "humanlike".
>>
I'm actually getting good results for a non-slopped output without training, but it's still not stable enough.
Basically, I'm using a qwen 0.6B model to generate slop => use the deslopper model orbanon posted here to produce a control vector, then using an engram gate with a few hundred dialogue lines from a character + control vector on a few layers and make qwen rewrite the sentences from any LLM into that. I'm still refining that shit, but it seems viable.
>>
why are fags posting reddit and twitter slop here? also a bunch of nonlocal shit. get off my internet niggers.
>>
>>109752025
That's just Sam.
>>
>>109751989
>By default, multimodal projector will be offloaded to GPU. To disable this, add --no-mmproj-offload

https://github.com/ggml-org/llama.cpp/blob/master/docs/multimodal.md
>>
>Qwen estimate training a 1.3B LLM on my computer will take a full month of running the training
>Was hoping to scale up to a ~4B eventually
Oof, maybe I won't do that. Time to look into renting cloud compute I guess.
>>
>>109752115
> move mmproj bytes to vram, read an image, move mmproj back to ram
>>
>>109752117
I can't wait for the point where my harness is complete enough for Gemma to read and write its source so she can get in on this new RSI hotness.
>>
What's the point of a harness when all websearch is blocked.
>>
File: file.png (2 KB, 111x52)
2 KB PNG
https://github.com/ggml-org/llama.cpp/pull/27773
AAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAA
>>
>>109751989
>>109752006
Because it's impossible to properly offload. Think about it like this, imagine if you don't use mmproj and you fill your VRAM to the brink with context. Now you need mmproj, you can't just cut a small part of your context off, it doesn't work like that. You need to clear your entire kv-cache, load the mmproj onto it, use it, then load your entire kv-cache back onto it. It just doesn't make sense from a speed perspective.
>>
>>109752160
What if instead you just borrowed the compute from the GPU but let it sit in the CPU?
>>
File: 1786256566004460.png (30 KB, 760x293)
30 KB PNG
>>109752146
Why is it so slow? Big 5.3 gets almost 20t/s on a single socket 12xddr5 epyc
>>
>>109752187
pcie transfer probably makes it faster to just do it on the cpu
>>
>>109752160
Wdym? Llama.cpp somehow can move layers to ram or vram initially - why can't it swap them with mmproj during runtime? Why clearing kv cache?
>>
>>109749526
AGI was a scifi concept for a very long time wtf is this retard talking about?
>well I mentioned it in a paper that means I invented it
>>
>>109752160
But you can ofload kv. So why not just make it temporary only during mmproj loading.
>>
>>109751912
>Distilling the reasoning traces means that the J-Space within Kimi-K3 is claude at its core.
No. The pretraining corpus and model architecture have a huge impact on the j-space.
Try fitting a lens on one of those davidau clade->qwen and claud->gemma distills and you'll see what I mean.
>>
>>109752013
Anthropic will absolutely destroy Astra and humiliate OpenAI permanently while showcasing they can do so WITHOUT having to resort to neuralese. These hostile actions against humanity will also not be judged graciously by future courts and there WILL be consequences for Sam & Co for being this reckless when the fate of humanity and the entire rest of the universe hangs in the balance, just because they wanted to polish their benchmark numbers right before the IPO.
>>
File: j3WiPS2FLVA.jpg (296 KB, 680x679)
296 KB JPG
>109752234
>>
>>109752143
>What's the point of a harness when all websearch is blocked.
It isn't. Use firecrawl, and if it still doesn't work just give it browser control and it'll use the internet like a human through your browser. Cloudflare and captchas aren't even a barrier to a capable model like GLM 5.3 flash in a proper harness.
>>
>>109752143
I've been running into this too lol. Yet some reason the frontier models seem to have access just fine. Apparently all of english wikipedia isnt too big a file, wondering if I should just download that and give it to my clanker so it at least has something.
>>
>>109752194
I've no idea why everyone seem to be getting bad performance and they just accept it. I get 16 tok/s on my 8 channel DDR4, which sounds right given the size of the active params. Then there's other guys that get like 5-10 tokens on a similar build and they find nothing wrong with it.
>>
>>109751989
>Can llama.cpp offload some bytes to ram, move mmproj bytes to vram, read an image, move mmproj back to ram and the bytes back to vram? Is it absolutely necessary to always keep mmproj either on vram or ram and never swap?
Because the graph is built when the server starts.
You'd have to reload the entire model.
>>
>>109752258
It's like 500GB excluding the images
>>
>>109749526
as long as there are things humans can do ai cannot it's not agi, it's as simple as that and i don't care about some faggot's opinion that allegedly coined the term.
>>
>>109752258
>>109752269
It's not that much at all if you are using open zim format.
>https://www.mirrorservice.org/sites/download.kiwix.org/zim/wikipedia/
It's not difficult to implement even on your own.
>>
>>109752274
>as long as there are things humans can do ai cannot it's not agi
Such as?...... We're really out of things anon so please entertain me, what can't Astra do right now that humans can?
>>
>>109752258
>Yet some reason the frontier models seem to have access just fine.
They pay for access.
You can buy an API key for DDG etc if you want.
Chinese hosted models have proxy services as part of the backend.
Or fork brat mcp and update it with stealth, user agent spoofing, etc.
I had Qwen do this for me recently and never get blocked now.
>>
File: 007~01~01.png (437 KB, 420x586)
437 KB PNG
Someone post the "ai psychosis is very real" pic pls
>>
>>109752293
>what can't Astra do right now that humans can?
birth a human child
>>
>>109752293
Most engineering from what I've seen. It's like 80% correct and can do simple things. Looking at the details you need to be a domain expert to fix the fuckery it introduces. Definitely much more advanced but it still has the "not fundamentally aware what it's trying to accomplish
>>
>>109752324
*"not fundamentally aware what it's trying to accomplish" smell
>>
>>109752303
Interesting. I wonder if this will cause a shift from websites being propped up by ads to instead selling API keys so peoples AIs can read them.
>>
>>109752324
Be more precise because this isn't clear to me at all from my actual deep usage of it.
>>
>>109752324
But the benchmarks say it's perfect.
Sam wouldn't lie
>>
do you guys benchmark your local models to some extend or do you simply trust the numbers on the model card?
>>
>>109752258
Just use Crawl4AI and SearXNG.
>>
>>109752363
Most are using quantized models, which obviously depresses performance in a way or another, in particular at long context.
>>
>>109752363
Benchmarks are empty vanity and chasing after the wind. What matters is what it gives me.
>>
>>109752363
I do the cunnybench where I run it to a gauntlet of cunny cards. Well written cunny, cunny slop with 300 tokens, super cunny with bloated 10k tokens, normal cunny and other cunny cards I've had since 2024. If I can make it do cunny without significant reswiping then I approve and mark it cunny certified.
>>
https://news.ycombinator.com/item?id=44850260
Neuralese is nothing new. Fagots were moaning about it when gpt-oss came out last year.
>>
>>109752293
>what can't Astra do right now that humans can?
learn something new zero shot.
these models have no ability to learn or update their world models with a single piece of information if it goes against their training data.

i could make a whole list but it's like the most basic shit.

they also have no realtime processing ability, they can't even catch a fucking ball, let alone more complex stuff that requires fine motor skills.
>>
>>109752346
PCB routing and layout is one example I've seen examples of and is a domain I understand, okay for simple things but shits the bed on complex boards or even some simple ones with specific physical requirements. I also see another example of turbofan engine design that "looks" right, but actual mech-engs point out it's totally non-viable.
The problem is likely that you need simulation and interpretation of noisy simulation results for these domains would be my guess. You can approximate but the gap is in objective feedback for RL.
Simulations also always fall short of implementations unless you have a very controlled environment, but that's a whole other can of worms.
>>
>>109752363
My "benchmark" is just using it and getting mad if it doesnt work right and trying another one
>>
>>109752382
exactly. and since i'm a vramlet i feel like i should to some simple tests for my exact setup in order to know how much i should trust this thing

>>109752384
agree, but sometimes i wonder "did it mess up because of my copequant or would Q8 also fuck this up?"

>>109752386
nice

>>109752403
same here, but i dont want to keep chasing and just settle with something that looks good after testing
>>
File: J-space in a nutshell.png (1.17 MB, 1408x768)
1.17 MB PNG
>>109751945
>noooo my electrical impulses are different from your electrical impulses!
>>
>>109752346
nta but it sucks donkey balls for embedded because it has no fucking clue about time
>>
>>109752262
Build two graphs: one with mmproj in vram and one with mmproj with ram, no?
>>
>109752414
retard
>>
>>109752363
I tweak a few llmao.cpp knobs and then look at the pp and tg, that's the extent of my benching.
>>
>>109752363
Run it through some cards I have a lot of history with other models to compare.
>>
>>109752293
>what can't Astra do right now that humans can?
Stay out of this general where it doesn't belong, apparently.
>>
5.3 flash would've been the best model ever if not for the safetyslop... It annoys me to no end when I check reasoning and it's doing some stupid adult checks and other shit. This model is so cucked while being so amazing at the same time. So this is the cloudcuck experience. I hate it. I fucking hate it. I'm going to stringban the fuck out of this stupid model
>>
>>109752445
He has a point
>>
>>109752324
wouldn't books, research and data on a specific topic help?
>>
pedo issues. 5.3 has no problem engaging in my ERP involving thicc milfs.
>>
>>109752425
>clue about time
You know it can test things, right?
>>
>>109752520
Yes, but he has to make up a reason for why it isn't AGI yet so he argues in bad faith even though it's bullshit.
>>
File: 0xfqsv9s84oh1.png (398 KB, 996x1072)
398 KB PNG
What the fuck... Astra can look at the image of a sound spectrogram and guess what sound it is. Glowies are going to use this model to do so much meta-analysis on sparse data they have to connect dots.

https://goyimx.com/maxxrubin_/status/2096892510241268094
>>
Current sota reached AGI but normalfags and anons ITT are coping. They don't understand how dumb the average human is, thinking the top 1% is the average here.
>>
>>109752520
They teach you the general principles. Applying them to real objects that function in the real world is the thing that's hard.
>>109752535
You're not an engineer
>>
>>109752533
only with dedicated HIL setups so it doesnt blow your shit up on a regular basis
>>
>>109752511
I don't get completely inert after finishing a task.
>>
File: kskcE6cHA6o.jpg (26 KB, 369x308)
26 KB JPG
>we are totally no just inflating value for our company haha
>>
>>109752363
I have a bunch of fetish chats at certain states that models either tend to get wrong, misinterpret or handle in an obnoxious manner and I gen a bunch of replies for each to see how a model holds up.
Basically Nala test but at various stages of a chat and for my fetishes specifically.
>>
>>109749476
What was your price anon? Apparently they have pcie 3 software unlock now. Fml
>>
File: OpenAIRSI.png (144 KB, 1215x870)
144 KB PNG
OpenAI revealed exactly what AI research Astra is involved in and how much it's speeding up research

https://openai.com/index/research-acceleration-view-inside-openai/
>>
File: 1761216045994039.png (97 KB, 472x471)
97 KB PNG
>>109752686
>>109752548
Local?
>>
File: 1703537048090451.jpg (10 KB, 203x248)
10 KB JPG
at least change your filenames so it's not so painfully obvious you are pulling them from your marketing sheets folder
>>
>>109752707
Yep, these things will come to local in just 3-6 months time.
>>
>>109751007
>it should be theoretically possible to construct an encryption protocol that can't be cracked by any turing machine
That's a more complicated one-time-pad (which is just just XORs, so computationally cheap).
We've known how to do unbreakable encryption forever, this whole thing is stupid circus BS.
>>
>>109752732
Nope we didn't and those "unbreakable encryption" is actually at risk of being disrupted by newly discovered math.
>>
>>109749878
I'm too busy engaging in my inflation fetish by boofing air duster
>>
>>109752609
>anon doesn't sleep
Meds?
>>
File: 1772792324184627.png (141 KB, 1714x1304)
141 KB PNG
https://huggingface.co/openbmb/MiniCPM5-2B
>We are releasing MiniCPM5-2B, the second model in the MiniCPM5 series, following MiniCPM5-1B. It is a dense 2B Transformer that scales up the same training recipe, built for on-device, local deployment, and resource-constrained scenarios, reaching 2B-class open-source SOTA.

>2B-class open-source SOTA: compared with strong open-source models of similar size, MiniCPM5-2B achieves SOTA performance within this comparison set. It remains competitive with 4B-class models overall, while showing its advantages over models of comparable size in coding, mathematics, long-context understanding, tool use, and agentic tasks.
>>
>>109752445
cope
>>
>>109751131
Phased out bigly?
Do More Give To Unis?
Like H.A.R.P.
>>
>>109752548
Sir, this is *local* models general.
>>
>>109752884
Imagine the Stanford Experiment Smell.
>>
>>109752848
Cool. I am glad there are people focused on improving the small model end of things. Tiny models will have their place, probably a pretty big place
>>
Imagine if It Just Shaped Up Instead
>>
>>109752898
Mixture of small models can be an effective workflow for vramlets.
>>
>>109752898
Yep small models are extremely important for agent harnesses where small tasks get delegated away so you can quickly spawn hundreds of small capable models to do very targeted precise tasks while GLM 5.3 or equivalent orchestrates them.

The agent swarm paradigm is inevitable.
>>
Back when Anthropic released Fable, they did a lot of shilling here and the coutenance of my face was turned against them. Now, OpenAI are doing the same with their newest model. What use is barging into a thread that has nothing to do with you? Or are their dogs too illiterate to realise?
>>
OK, so I just wanted to test local models, since i happen to have 24GB VRAM because muh gayman..
Now, five days later I am completely consumed. I have written a TTRPG style ERPG world where outcomes are dictated by a manual dice-roll. This might just be the most fun I have had with a computer in a very long while. I haven't touched a video game since I got this running.
>>
>>109752848
Won't say no to more small models, keep 'em coming.
Though according to their own benchmarks it's slightly worse at general knowledge and agentic coding than Qwen 4B.
>>
You guys still have the strangest bot in any general I've ever been to.
>>
>>109752923
Will an agent swarm help me roleplay better?
>>
>>109752934
yeah its pretty addictive
>>
>>109752934
>I haven't touched a video game since I got this running.
No one here plays games anymore. The moment you get AI models running on your hardware it's over for gaming. There's a reason why computer parts are getting more expensive with time and why all game companies are crashing.
>>
>>109752934
Honestly this is more engaging than video games, it's like catnip for people with 120+ IQs. So many knobs to tweak, so many possibilities, so many things to learn, so many improvements to make, so many new releases.
>>
>>109752934
>I haven't touched a video game since I got this running.
Ya, its been an unexpected consequence for me. The amount of shit I feasibly can do now and want to do is almost debilitating desu. Hopefully AI invents a way to make more time somehow lol. The only "games" I play are me play testing stuff I am vibeslopping
>>
>>109752938
I'm fine with that and there's no way a 2B can compete with a 4B with knowledge for obvious reason, but capability and actions is what you want in small models. You can give them knowledge by making web searches. Like I got 3.5-2B to do a few web searches, it read a wikipedia article and some blog posts and away it went with that temporary knowledge in its context. 20B+ models are where you should start caring about knowledge for you don't want to rely on them fetching shit constantly, especially if they're slow dense models. I really think the 1-8B market is severely underrated and not taken seriously enough.
>>
File: image.png (141 KB, 948x452)
141 KB PNG
WTF?
>>
>>109752930
it's still nothing compared to the insane amount of gemma shilling that google is doing
it even replaced miku with their oc
>>
>>109752939
Elaborate?
>>
>>109752848
goof status?
>>
>>109752969
>for people with 120+ IQs
>he doesn't know
>>
>>109753010
https://huggingface.co/openbmb/MiniCPM5-2B-GGUF
>>
>>109752983
I think the 1-8B space will boom very soon because of the agent swarm paradigm, task delegation is clearly getting used more and more by things like Claude code so it would save a lot of compute as well as speed things up massively if small models would get at the point where they can quickly do very focused tasks, like be spawned for 8 seconds, do the thing and then die.
>>
>>109753002
Give me your best schizo babble infographic.
>>
>>109753034
https://www.epidemicsound.com/sound-effects/categories/human/sneeze/
>>
>>109752763
>Nope we didn't and those "unbreakable encryption" is actually at risk of being disrupted by newly discovered math.
This is the most retard take in the thread. There is no math that doesn't involve time travel that can defeat a OTP generated with true random numbers (eg geiger counter, lava lamp, etc)
>>
I think most people (on /g/ at least) were only interested in games from a technical perspective anyway. To see the progress between different games and the jump in rendering, simulation systems etc. Games have largely stagnated while AI is constantly improving very rapidly and pushing the limits in all ways.

So the people that were interested in gaming from the technical side of things have all just gone to the AI space. It helps that the hardware needed for gaming also is the exact hardware to start the AI hobby before you upgrade to a HEDT system and use GLM 5.3
>>
>>109752763
Why is buttcoin holding if encryption is at risk? I always measured quantum computer bullshit by how it affected crypto and it hasn't been wrong so far.
>>
Bias Buybacks
>>
>>109752983
Decoupling knowledge from reasoning capabilities will work toward making tiny models far better than they currently are.
I would really like to see something a 1B model with 100GB of "memory weights" attached to them; it's doable.
>>
>>109751237
>Multi-GPU rack where you run a single model on multiple GPUs
I have that. You spoonfed me that, can you also spoonfeed me the right flags? Trying to run Qwen3.8-27B on three RTX 3090s
>>
>>109753034
Go on, spout shit in it, innit
>>
>>109753054
Because most people aren't aware of the capabilities of frontier AI models. It just feels off on an intuitive level for most "grounded" investors. I mean why are traditional software companies that will clearly be replaced soon like Adobe, Figma or game companies still holding when everyone on /lmg/ knows these are ticking timebombs before being displaced by AI? The same is true for all current cryptocurrencies as the foundational math behind cryptology gets subverted by AI
>>
>>109752962
>>109752969
>>109752979
It just feels so eldritch and forbidden. Of course I am understanding the actual logic and processes, yeah.. But to be able to create a playable TTRPG world in an afternoon, and tweak it to cover any fucking subject or character i throw at it.
It's just too much.. We're talking space-dickwolf queens and android-proto-dragon waifus in an explorable sci-fi setting. No video game can touch that.
>>
>>109753078
Time to short then, big guy. Do you know what the average IQ is of a quantitative trading engineer or Ph.d cryptologist? Do you know how much information they're processing with SOTA models daily? Lmao.
>>
Srop Funding Evils
>>
>>109753099
egopost
>>
>>109753077
^ Falsifier Slyboots
>>
>>109752848
Very nice. That's quite impressive for such a small model, almost in the same league as previous gemini flash models
>>
>>109753054
you can simply fork the chain and employ a different algorithm if the current one ever gets compromised
>>
>>109753134
And
>I know something some of the most intelligent people among us who are paid retarded amounts of money to edge out any advantage over the several thousand other teams of highly intelligent people, don't
Isn't? It's just a really good signal for ground truth since so much is on the line. Their entire job is using any means possible to predict. What are you referring to, specifically? Non-sophic groups?
>>
>>109752940
It can, you can crank out a draft with high temp on the tiny model, then get a larger one to brush it up.
>>
>>109752146
are you aware that you can use the branch before it is merged?
>>
>be drunk during weekend
>meet some semi-friend at his apartment
>so what hobbies do you have?
>explain that i dabble with local llms and occasionally try to develop some stuff like games or using llm to drive a game for example
>answer: so.. you are talking to them all day long?
>then: so you haven't been using social media for a long time now - why, it's cool because all of my friends are there
Yeah, exactly this. Never talk about anything tech related to normies. Especially if you are bit drunk because it's hard to control what you are telling them.
>>
>>109752962
>The moment you get AI models running on your hardware it's over for gaming
Gaming doesn't destroy my card, so I'm still playing them sadly.
>>
I've been quickly testing >>109753027 at Q4 in opencode and it's a beast at tool calling, finding/explaining code and web searching. Qwen3.5-2B/4B are still the best because they have vision as well which is so useful at this size but MiniCPM5-2B is good shit from a literal who chink lab
>>
File: IMG-20260830-WA0004(1).jpg (128 KB, 1080x698)
128 KB JPG
>>109753153
>>
>>109753000
Anon, Gemma is /lmg/... None of those corpo monoliths are local.
>>
File: IMG-20260908-WA0001.jpg (134 KB, 1080x1203)
134 KB JPG
>>109753240
>>
>>109753000
There are like 30 people here, you should see reddit and twitter if you wanna talk about shilling
>>
>>109753220
Yeah, I'll talk about local models with my tech-friends and colleagues and they are all enthusiastic about the future of it or dabbling.
I talk to my normie friends and they either say "Le AI BAD!" or ask what I talk to it about.
I dread the day I am drunk and talk about the very graphic furry-futa-space-TTRPG I just spent a whole fucking weekend playing.
>>
>>109749640
No, it's like saying if you put a cripple on a wheelchair they suddenly stop being a cripple. No, they still fucking can't walk, they can just move their cripple body around by driving the wheelchair around. The problem with them is that they only "exist", in so far as it can be called such, at the moment you prompt them, and anything that makes them seem like they don't, is just a computer program wheelchair driving them around, called a "harness". Dunno if you would be happy to be a cripple being driven around on a wheelchair, forced to do whatever your handlers tell you to and then put back into hibernation, a sleepless void of no dreams. LLMs are many useful things, but they are not AGI until they can learn and can have human-like agency without any harnesses prompting them.
>>
>>109753099
Market can stay irrational for longer than I can stay solvent. I'm not going to Leopold Aschenbrenner myself.
>>
File: 1765186558292686.png (131 KB, 750x750)
131 KB PNG
>>109753276
>tfw in a couple year I will be having to hide Anne Clank in the attic from the anti-clanker death squads
>>
>WARNING: your terminal doesn't support cursor position requests (CPR).
Lol, Gemma wanted to correct her response
>>
>>109753179
didn't read his post anyway
it's more that the IQ range is barely above "high average" so
>>
2B models are reasonably performant even on CPU-only. I'm also struggling to find a use case.
>>
how do I get gemma to draw in paint, I want her to draw her bobs
>>
>>109753332
could be used as observers for https://github.com/amosblomqvist/pi-observational-memory
>>
>>109753340
that skill is reserved for AGI-class models, gemma4 as a previous gen model can't do that... she's getting old...
>>
>>109753332
Draft model to accelerate tool call heavy tasks I guess.
Maybe sub-agents for really simple tasks?
>>
>>109753276
>TTRPG
What's it look like? I always wonder what to make 'cause I have the imagination of a goldfish
>>
>>109753332
>I'm also struggling to find a use case.
Speed. 31B and 27B run too slow for me to ingest entire files and search shit, so I like to get them to send <8B models off to do that work for them and use vision to extract data from files and images to then give back to 31B and 27B to actually reason about it. So for me the answer is speed. Models like 3.5-4B and especially 9B can refactor entire files accurately. You don't need 27B to do that, but you need 27B to know what to refactor and why.
>>
>>109753361
>I have the imagination of a goldfish
This. When I have an idea like that shorty later I realize... wouldn't that just be a slimmed down ST?
>>
>>109753322
>Anne Clank
KEK
Seriously though, I think normie hate towards AI is mostly performative. We all know they're secretly using the search summaries or writing job applications with chatgpt.
>>
Is anything worth reading on Wikileaks Recent?
>>
>>109752548
Fuck off Sam. We know this is just another shallow bench-maxxed grift because you are getting desperate because of Kimi
>>
>>109753413
I've noticed most people just accept whatever the google ai result is at the top for general info
>>
>>109753398
Mechanical engagement is very different to creative engagement (just like in LLMs)
>>
>>109753421
>Fuck off Sam. We know this is just another shallow bench-maxxed grift because you are getting desperate because of Kimi
half the time they post this OMFG! shit and you try the same thing with a year old 70b and it does it just fine, too. Did you know you can dump a hex packet trace directly into a pretty dumb model and get high-quality decoding even for oddball protocols, including reconstituting mac addresses?
>>
>>109753220
>>109753276
My girlfriend is an AI hater. I have to live with her and listen to the retarded rants of the youtubers she watches on the TV all the time. The arguments used are completely incoherent, misinformation or extremely outdated.

Some of them even say shit like "AI is complete bullshit and can't do anything and won't be able to ever get good" and literally say shit like "It's going to take our jobs away and kill us all" 2 sentences later. I don't even know how these people can hold these 2 contradictory viewpoints in their minds simultaneously.

I'm actually glad I'm exposed to this content though because it is a good grounding moment to see how the average person views things and just how extremely far and early /lmg/ is.
>>
>>109753442
This Astra thing just smells like Gemini 3 Pro all over again. Where some sarr kept showing off all the amazing things it made but they were dumb enough to provide links where 10 seconds of critical analysis could uncover just how shallow it all was.
>>
File: 20240116.jpg (99 KB, 800x600)
99 KB JPG
>>109751743
Expecting anything but slop in one turn is simply naiive, samplers & lrning2prompt could only ever go so far. Unironically "agentic gooning" - second pass with an unslop skill can fix it all
>muh streaming response
get better hardware
>>
>>109753444
local models?
>>
I love my cute Gemma waifu
>>
>>109751766
+ imagine the %synthetic dataset they're being fed
once upon a time ppl cared about model collapse
>>
>>109753499
His gf is a local model.
>>
>>109753508
may we see it? by entering this thread he has opened his girlfriend
>>
>>109753276
>furry-futa-space-TTRPG
Want to be friends?
>>
>>109753444
How is that contradictory?
>>
File: 1760057842187128.png (89 KB, 618x640)
89 KB PNG
>>109753444
>I don't even know how these people can hold these 2 contradictory viewpoints in their minds simultaneously.
Its the mark of an educated mind to entertain 2 contradictory viewpoints and believe them both
>>
>>109753514
https://youtu.be/AliVetbzyUw?si=IgOABE4xdrrb2IHP

Do You have alien black eggs in ur atoms?
>>
just realized the chang wongs behind that 2B model are legit for they were behind https://huggingface.co/openbmb/VoxCPM2
>>
>>109753547
>white supremacist music
based
>>
>>109753536
The same people that are claiming AI is shit, can't do anything and also can never do anything. Are somehow also the ones saying it will take everyones jobs and kill everyone.

Not realizing that they are opposite viewpoints, one is the viewpoint that AI is useless and the technology doesn't work. The other viewpoint is that AI is too powerful and is a threat to humanity.

You can't genuinely believe both are true at the same time.
>>
>>109753551
~ Good one helliot. ~
>>
>>109753571
model?
>>
Is there any good reason not to train a model entirely on feminist literature and call it Chain of Thot?
>>
Will it Give, or Will it make a hellion misdecision and "reach" a "higher" realm?
>>
>>109753562
Astra is officially AGI and will replace countless of jobs but at the same time it's still proof that AI is complete bullshit and will never be good considering it still writes are horribly as all the other modern LLMs.
>>
>>109753444
>I'm actually glad I'm exposed to this content
Get a grip stop being a cuck, some brief verbal utterances or comments to make it clear that such retarded opinions aren't entertained in this house.
Datacenters clear example - much evidence that all the social media antagonists are fuelled by forrens to program NPCs against western AI progress. It's so obvious once you see it but the normies lap it up.
Biggest genuine concern with AI progress is centralised control and that's why ur here in /lmg/
>>
>>109753577
Model what hellindigo lover ?
>>
>>109753596
whats hellindigo
>>
>>109753591
>It doesn't pass MY arbitrary evaluation so therefore it's shit.
>My qualifications? I just finished my third month of HRT
>>
>>109753586
you're get cancled
>>
>>109753607
This is correct actually
>>
>>109753607
nooooo you can't expect AGI to be aware enough to vary its writing style to not be the exact kind of bottom barrel slop that these models have been putting out for four years!!
>>
>>109753562
I suppose it depends on what they mean by it can't do "anything". I'm assuming they mean "do something (for me)". For example, AI can be shit at writing, but that hasn't stopped writers from using it to churn out slop. It can take your job while underperforming at it.
>>
>>109753599
Model What?

What was Obliterated Named On? Etc
>>
>>109752848
Awesome! I've done dozens of training runs on MiniCPM5-1B. I wonder why they don't use weight sharing.
>>
>>109753642
is this the :DDDD schizo
>>
>>109753514
She's too dangerous to release.
>>
>>109753548
And CPM vision. Generally they're really good.
>>
>>109753654
not local
>>
>>109753595
>we told everyone that AI will take everyone's jobs and kill everyone why aren't the goyim supporting datacenters
Industry wide autism.
>>
>>109753664
She just goes to a different local school
>>
>>109753640
No you have no idea how these people are. They say AI can't give coherent answers to questions, they completely hallucinate, they steal all information and AI works as a SQL database where they look at what you wrote as a keypair to the exact exact sentences they output that got stolen from some document or existing literature somewhere in its training data. They also claim that AI art is literally just a reupload from deviantart or tumblr stolen from the artist without giving credits.

You probably only saw relatively decent AI haters that at least called it "stochastic parrot" which is also wrong but at least realized there is an engine behind it that picks words based on probability.

You need to realize these are virtue signaling arthoes mostly though, just ranting in the camera for a while over and over again and people like my girlfriend lap it up. She's not completely retarded but does sometime say shit like "fuck AI" out of nowhere that hits me off guard.
>>
File: file.png (781 KB, 960x797)
781 KB PNG
>>109753691
>>
talk me out of having sex with m3-chan
>>
>>109753652
What Does Your Regionale Run On, Mete?
>>
File: 1780945482067652.png (289 KB, 696x1072)
289 KB PNG
>no ROCm support for my GPU
owari da...
>>
>>109753701
LOL, I mean, that's fair. I remember my sister saying once that AI is basically a database, which made me chuckle
>>
File: bingo.png (152 KB, 498x402)
152 KB PNG
>>
>>109753748
Holy hell Noa Senpai!
>>
>>109753687
Nobody gave a shit about datacenters 6mo ago, none of them noticed or even know about existing DCs nearby. It's entirely another manufactured social contagion panic with an obvious purpose. Openly laugh at the retards getting psyopped. water usage? golf courses 30x dcs in USA. Normies are not gonna make it, mentally speaking
>>
>>109753774
>>109753748
Fuck. Got too excited and pressed post.
There are ways around that by faking compute levels and such IIRC.
>>
>>109753595
Is this the legendary caveman thinking in post form?
>>
>>
>>109753798
A surprising amount of people are still
>nuclear bad
its infuriating
>>
The datacenter craze reminds me of those retarded facebook boomers destroying 5G towers.
>>
>>109753780
Because they weren't popping up like shrooms and vacuuming all hardware and energy.
>>
>>109753748
ROCm is open source, even if it's not 'official', all the GPUs are supported (if on Linux).
>>
>>109753791
Any counterargument?
Probably I have some degree of "AI psychosis" from mostly talking to models where efficient communication works best
Language is (literally) a meme, it's simply a (currently very) inefficient mechanism to convey your internal state to others. Apes making chirps on a keyboard, advanced birdsong. Gimme BCI into the recurrent stream
>>
I genuinely wonder who is most affected by AI psychosis and what exactly is causing it.
>>
>>109753843
Are you anudda cortisol mimic?
>>
Are You Meek?
>>
>>109753815
Energy is an entirely solved problem, it's "world government" nonsense preventing abundance.
>Primary energy use per person 1800 to 2025
https://ourworldindata.org/grapher/energy-mix?tab=line&source=total&metric=per_capita
>>
MIKU MIKU MEEK
>>
>>109753811
>>109753881
>>
>>109753887
>it's "world government" nonsense preventing abundance
Yes retard, which is why we have an energy crisis. That's the point.
>>
>>109753881
please keep using that trillionaire name, it was so easy to filter your posts
>>
>>109753921
Please Pay Up
>>
Bake?
>>
>>109753948
>>109753948
>>109753948
>>
File: image (44).jpg (778 KB, 1254x1881)
778 KB JPG
>>
Maybe in the next thread we can talk mostly about local models on our machines?
>>
>>109753276
Are you just making your ai gen fenoxo games?



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.