[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


A general for vibe coding, agentic engineering, coding agents, AI IDEs, browser builders, and shipping code with LLMs.

You use Git — right, anon?

## What “vibe coding” is, and how to do it
https://simonwillison.net/2025/Mar/19/vibe-coding/
https://simonwillison.net/2025/Mar/11/using-llms-for-code/

## News (both past and future)
- 2026-09-14 America/Los_Angeles — Claude’s 2× promotion ended; usage drops to +25% from the +50% that we’ve become used to (a 17% reduction)
- 2026-09-12 — Anthropic suggests to pace the frontier; OpenAI agrees in principle

## Related generals
>>>/g/lmg/

----

## Frontier models using fully-general tooling — start here if you have $20 or so
https://developers.openai.com/codex/cli — probably generally better currently
https://claude.com/product/claude-code

## Near-frontier models for code
https://x.ai/cli — no 5h limit for only $30/month

## Not worth it for code, but maybe good for interpreting images/video
https://antigravity.google/product/antigravity-cli

----

## Prompting
https://simonwillison.net/guides/agentic-engineering-patterns/using-git-with-coding-agents/
https://arps18.github.io/posts/claude-code-mastery/

## Skills
https://github.com/mattpocock/skills — /grilling is a favorite
https://github.com/DietrichGebert/ponytail
https://github.com/Vuk97/forward-implementation-first — do less redundant bookkeeping

## Other editors / terminal agents / coding agents
https://osaurus.ai/
https://pi.dev/
https://opencode.ai/

## Is our AIs unlearning?
https://aistupidlevel.info/

## Will there be a codex reset?
https://codex-resets.com/

## What we’ve done
https://vcg.gitgud.site

## Previous thread
>>109838193
>>
File: 1767234502270797.png (1.72 MB, 1024x1536)
1.72 MB PNG
>>
>>109841641
How did you generate this?
>>
>>109841627
The fuck I'm looking at?
>>
why did you use such a disgusting off topic image? terrible OP
>>
>>109841644
grab 14 screenshots of ruby from the anime, and tell it to make the picture using the art style of the screenshots attached.
>>
>>109841651
Which model?
>>
>>109841656
gpt image 2.5
>>
>>109841650
get your ass to the gym, fatty
>>
>>109841672
Based
>>
https://x.com/synthwavedd/status/2100664546528571770
>Unfortunately, it looks like Grok 4.7 won't be anything crazy.
>While it seems once again very token efficient - and it'll be cheap - it gets cooked by the upcoming new Opus in my testing, and it's not all that competitive.
>>
File: 1764074407741028.png (2.17 MB, 1920x1080)
2.17 MB PNG
>>109841677
don't make me tap the sign
>>
dario always wins. every day that you're alive is a day in which dario is winning
>>
>>109841677
How can they legitimately not into AI when they are catching 30 story rockets in mid air ?
>>
>>109841672
I realized OP is Israeli, that explains it. They love gross stuff.
>>
>>109841677
It's fucking over.
>>
tired of cumming to bbc can tibo give us a reset so i can work again
>>
>grok 4.7
>>
grok 4.5 is the last good grok model
>>
Cursor, grok build, and grok.com all don't show 4.7.
>>
>>109841694
SpaceX was founded 2002. Additionally if you have talent in ML research you'd have to have particular reasons to work at xAI. You don't have to work for the sperglord to be at the frontier unlike in rocketry.
>>
>>109841722
4.6 is good. I'm building an app right now.
>>
>>109841733
synthwavedd is a leaker. he guessed some internal api or sth. mentioned opus-5.2 isn't out either.
>>
>>109841694
idk man. I don't want an llm to create a 3D model, at least it's not something I *think* I want.
>>
>>109841757
>leaker

in other words there is early access rn?
>>
vibegods which model has the best price-performance currently?
>>
>>109841777
Close match between Astra and Fable.
>>
File: LARPING LEGENDS.jpg (488 KB, 1080x1055)
488 KB JPG
I'll never pay for LLM models
>>
>>109841777
deepseek v4.1 flash, grok 4.6, glm 5.3, gpt luna, sol low and medium
>>
>>109841785
amerijeet models cost 100x what chinese do you're tripping
>>
>>109841785
Isn't the recommended way to go astra max and fable 5.1 max, running in parallel, so you can compare results?
>>
>>109841795
He said price-performance. Astra and Fable stand alone in their class, so no other model is even considered in the comparison. If anon had set some actual requirements then maybe it'd be worth considering other, cheaper models. Without knowing though, then I can only recommend the absolute best consumer-facing models, Astra and Fable, and their pricing is fairly comparable. If we do expand though, the Chinese models only start to become competitive when talking about API pricing, they still don't remotely compare to the overall value of the American subscription plans from OpenAI and Anthropic, which are vastly better values than anything the Chinese offer at their API prices.
>>
File: .jpg (84 KB, 1544x854)
84 KB JPG
>>109841795
wrong.
chinese models aren't beating any subsidized $20/100/200 plan from anthropic or openai.
https://x.com/SemiAnalysis_/status/2091631658973671900
>>
File: 1758930396359565.jpg (198 KB, 630x630)
198 KB JPG
BRING BACK THE 200 DOLLAR SUBSCRIPTION COCKSUCKERS
>>
>>109841818
Astra price + Fable price - Astra performance - Fable performance = infinity
>>
File: 1784268183994099.png (2.72 MB, 1448x1086)
2.72 MB PNG
>>109841836
>>
>>109841827
Why aren't the chinks subsidizing them?
>>
If only we had a model with grok's intelligence, fable's price and k3 speed...
>>
>>109841850
astra...
>>
who is the live grok bot foid?
>>
>>109841844
they don't need to lol my 20 dollar deepchink fund hasn't run out in 3 months while your fraude limit taps out after 5 prompts
>>
>>109841844
They can't afford to, they don't have the billions of dollars-yet-to-be-printed to not-pay for everything.
>>
>>109841844
z.ai did try a while ago but then ran out of compute immediately
>>
>>109841861
I thought it was the nvidia hw...
>>
snailcats won, you are all out of usage to hunt them
>>
>>109841858
This isn't something to brag about. It's very easy to burn $5/day with DeepSeek 4.1 Flash, you're just loudly declaring that you don't actually use it.
>>
File: .png (146 KB, 797x622)
146 KB PNG
Anyone use this shit?
>>
>>109841873
The Nvidia hardware? That they paid for with money loaned to them by Nvidia for buying hardware? Funny money, it goes in a circle, it just keeps on going, swirling, the entire industry a singular toilet bowl for Jensen to piss in.
>>
>>109841876
>It's very easy to burn $5/day with DeepSeek 4.1 Flash

4.1 is dumb you can't burn 20 million and solve navier stokes with it. you are retarded if you overuse it. just have it write what you want that's what it's for. do you not have a degree?
>>
>The Rust implementation is in place, but I can’t honestly call the full goal finished yet. After 38 hours, you deserve that clear answer.
>The remaining work is mostly acceptance testing, with that recovery edge case still unresolved. My earlier “almost done” estimates were too confident.

KEK THIS PIECE OF SHIT. Whatever I don't give a fuck.
>>
>>109841677
AI died when 3js slop became the most important benchmark for poojeets
>>
>>109841873
deepseek ceo

https://github.com/demo-zexuan/liang-wenfeng-investor-meeting-2026-7-22
>Huawei's hyper nodes, specifically the Huawei 950 hyper nodes, can fully replace NVIDIA's GB200 and GB300 in terms of performance and price. The cost is certainly higher, but only marginally so. Whether it's 50%,100%, or even 200% more expensive doesn't matter much. For instance, a price increase of 100% would already qualify them as viable substitutes in terms of cost. They are also interchangeable in terms of tasks; all tasks that a GB300 can perform, a Huawei hypernode can do as well, with identical latency performance. The only cost is that four Huawei cards are equivalent to one NVIDIA card, and they lag behind by two years in performance. The equivalence of four cards to one is understandable. The "two-year lag" means that four Huawei Huawei 950 cards can match the performance of one GB300 card, while the "two-year lag" refers to a temporal disadvantage of two years.
>The Huawei 950 supernode will be available in Q3 or Q4 this year, while the NVIDIA GB200 was released in Q3 two years ago—there's a two-year gap. NVIDIA may have already introduced its next-generation model by Q3 this year. Therefore, regarding the gap between us and the U.S. in chip technology, I believe there won't be any further disparity at the ecosystem level, but the gap in chip technology remains four times larger with an additional two-year delay.
>>
>>109841908
controller test is a variant of the pelican test. it's a svg, not 3js.
>>
>>109841944
>its not shit its poop
>>
>>109841694
They're outpacing all of China and arguably google. What's the bar for can into?
>>
Testing the Brauer Manifold Reverb.

Wobble bass colliding with heavy synths, a massive wall of effects, and pure automation madness.
Just a quick, dirty, and chaotic session!

Used the Gopher A.I. in FL Studio to push this plugin to its absolute limits. The way this manifold engine multiplies, splits, and morphs the sound is unreal.
This build is insane, so many intense harmonics tearing through the matrix, it should be straight-up illegal!

This reverb, gentlemen, is an absolute monster of a tool. Almost every single crazy effect you hear in this track is being generated and layered directly by the manifold matrix itself. Only the bass line has a bit of extra automation.

https://vocaroo.com/15oNIEjQC8KZ


THE SWARM HAS TO BUILD A USER MANUAL, MAKE IT EXCELLENT.
>>
>>109841677
Not surprising. Grok 4.6 was GPT 5.4 tier, outdated and shit. I have no idea why anyone likes Grok.
>>
>>109841926
I wonder, is it software compatible?
>>
>>109841951
the bar is to actually be frontier. we are making fun of google, too.
>>
File: 1769827521132148.png (215 KB, 1080x2400)
215 KB PNG
12 days until gpt 6 luna lads. gpt 6 luna waiting room
>>
>>109841974
it will say slurs and makes "the libs" owned or something, no one with a functioning brain cares
>>
>>109842009
It doesn't even do that anymore.
>>
File: .png (118 KB, 1190x440)
118 KB PNG
>>109841627
>>109842006
openai devday date should be in op
pic: https://x.com/scaling01/status/2100036058830319842
>>
>>109842006
>>109842073
you're both gay, poor, and sad
>>
>>109841989
>we are making fun of google
Why? They've been one of the more relevant organizations ever since the attention paper.
>>
>>109842100
>retarded incoherent name calling
Someone's seething.
>>
>>109842100
>both
>lists three properties
Are you sure about using of this saar?
>>
>>109842109
both of you
>>
>>109842109
also, add retarded to the list
>>
>>109842125
What if it's the same poster? (Not that I know.)
>>
>>109842125
Nice try, but moving the goalposts like that is too clever and stupid for someone like you. Would you like to try again?
>>
Does it make sense to let the AI agent keep a diary? It's just a scratch text file in which it can write down things it finds important, like a memory but less involved.
>>
>>109842106
https://en.wikipedia.org/wiki/Attention_Is_All_You_Need
>The authors of the paper are Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan Gomez, Łukasz Kaiser, and Illia Polosukhin. All eight authors were "equal contributors" to the paper; the listed order was randomized (according to the paper itself). After the paper, each of the authors left Google to join other companies or to found startups.[7][8]
>>
>>109842160
You'd think they'd stop them from leaving somehow.
>indian and east european sounding names
Oh no no no...
>>
>>109842135
it isn't. It's just two homosexual threadshitters
>>
>>109842160
What does frontier mean to you? Org with the highest number on a chart is the frontier and everything else doesn't exist? They're still one of maybe 4 companies that even have a shot.
>>
>>109842215
You don't know that.
>>
>>109842142
He didn't move the goalposts, you just failed to read, mate.
>>
>>109842226
>What does frontier mean to you?
having a frontier llm
>They're still one of maybe 4 companies that even have a shot.
based on what? google killed their best ai lab, deepmind, two years ago. those people work now for openai and anthropic. for example, tibo is ex-deepmind.
>>
>>109842325
How mad are the google execs right now?
Or did they think it would (or will) crash?
>>
>>109842336
google is anthropic's biggest shareholder and anthropic buys all of google cloud's compute. so not very.
>>
>>109842143
Yes, I've been doing it. A new session can take over quicker and without exploring the codebase so much.
I also instruct it to compact on its own and remove information that's no longer relevant. The note is isn't updated constantly, there has to be meaningful progress.
>>
File: 3w4ml8jvfwph1.png (219 KB, 1024x1281)
219 KB PNG
let me speak to your manager
>>
>>109842350
>talking to grok 4.6 after the deepseek v4.1 flash subagent fucks up
>>
>>109842325
sama's possible alt-account talking about openai's real competition (spoiler: it's not google)
https://x.com/tszzl/status/2092420937505546261
>everyone assumed it would be google that would unseat or threaten openai’s lead on models but the one that gave them a run for their money was an (at the time) smaller company called anthropic
>it is now in the interests of both incumbents to say, and possibly true: RSI is at hand, join or become irrelevant. but i would wager while the fortunes and misfortunes of research could shift the immediate race in either’s favor or a third large company like meta or xai, history is always longer than people account for
>first, as A\ proved, huge amounts of compute are more elastic than some previously thought. there is some inference bid for which major giants will give up some of their frontier training ambitions and rent you parts of their compute buildout (google, spacexai etc )
>a company betting on something radically different in research or execution could gain a major foothold especially as so much of the stack becomes autocatalytic— in other words you could use a frontier model to help build the parts of your platform you’re not good at. if you’re not constrained by American law you can even distill Claude and GPT to close the gap
>small experimental labs are likely a bigger threat to OpenAI and anthropic than large companies
>it is game theoretically hard to not serve your frontier model - if your competitor is willing to launch then you may lose a huge chunk of your revenue. if safety concerns supersede that, all of this potentially flips if everyone at the frontier uniformly keeps the strongest models hidden or nerfs their AI R&D capabilities for many months behind the frontier. that may lead to closed loop RSI & runaway advantages. even a huge breakthrough in compute efficiency won’t matter if the labs have private engines of autocatalysis
>>
>>109842382
Why do they think burger law bans distillation?
>>
>>109842382
I really hate this performative "exclusively typing in lowercase" autism xitter AI troons love to do.
>>
>>109842389
openai forbids distillation in their terms-of-service. delaware courts will enforce those terms-of-service.
>>
why did anthropic, openai and grok seemingly decide to cut usage for their users simultaniously? I use all 3 and noticed it on all 3 plans. OpenAI being the most obvious one
>>
>>109842461
They're pacing the frontier
>>
>>109842461
anthropic and openai are using most of their compute for rsi now. anthropic rents grok's servers.
>>
>>109842461
Conspiracy.
>>
>>109842461
energy price went up because of trumps retarded iran war, they can't say that though or he will be sad and then regulate them
>>
>>109842461
There are a limited number of data centers anon
>>
File: 1789681123.jpg (97 KB, 782x788)
97 KB JPG
>>109842499
>trumps retarded iran war
>>
>>109842499
you are straight up just retarded, huh? sorry about that.
>>
>>109842499
brics niggers be like
>>
>>109842517
>Israel's war against Iran that uses the USA like a puppet
Fixed.
>>
>>109842546
much better
carry on
>>
>>109842513
??
both openai and anthropic have 5GW each in datacenters now. more than double than they had a the start of the year.
also blackwell gpus came online which are 4x as efficient as hoppers.
where do you think all that compute is going?
>>
>>109842557
openai more than tripled their number of users on codex after astra came out, it uses more compute, they have not replaced their entire fleet with blackwell
>>
>>109842562
>openai more than tripled their number of users on codex after astra came out,
What the fuck are these users using it for?
>>
>>109842570
equal parts porn, programming, and distilling
>>
File: sailinggg.png (1.44 MB, 1420x1188)
1.44 MB PNG
any use case for a Quen 8bn Q4?
I can run it on my laptop with 4070 that has 8GB of vram

but i find it to be really bad, its a coding model, i got one with abliteratedweights so it doesnt really say no to anything hoping it would make it less dumb, but it didnt help, now its just more dangerous because its fully unpredictable

For example:
>Hey quen, i need to wash my car, carwash is 50m away, should i walk or drive
>walk anon
>okay i am here, now what
>park your car and begin washing it with soap and a soft cloth
>what do you mean park, you told me to walk
>[user is confused about park...] park your car and begin washing it with soap and a soft cloth
>what do you mean park, you told me to walk
>[user is confused about park...] park your car and begin washing it with soap and a soft cloth
>you fucking retard Quen, you told me to walk, so I walked, I dont have the car with me, and why the fuck would i wash my own car, they wash the car, its in the name
>oh okay sorry about the confusion

I need something better, even if its non-coding, I really do not want to spill into regular RAM
>>
>>109842570
dog shit games made with three js
>>
>>109842577
>porn
Doesn't openai block such things?
>programming
Doesn't astra suck at it?
>distilling
Oh.
>>
>>109842546
There you go
>>
>>109842588
I actually laughed out loud anon, ty for the laugh
>>
>>109842557
>where do you think all that compute is going?
We have no clue how much compute is allocated to training or inference at any given time.
>>
>>109842597
>doesn't the current best model for programming suck?
Yeah, totally.
>>
>>109842616
70-30 rule is a good bet. now which one is for which i don't know
>>
File: .png (99 KB, 1344x716)
99 KB PNG
>>109842616
wrong.
we have a good guess.
https://ai-2027.com/
>>
>>109842639
The fuck are these expensive experiments?
>>
>>109842663
hyperparameter tuning for one
>>
>>109842663
trying to do shit like millennium puzzles
>>
>>109842673
no, that's under 'running ai assistants'
>>
who cares about LOC just make them write good docs fuck this shit im not writing code ever again
>>
>>109842639
this is literally just people guessing in 2025 on what was happening in 2024 and what will happen in the future
>>
Having Astra write drivers inside of a VM is so fucking funny. It's starting seething if you don't reset the VM when the kernel crashes kek.
>>
arent'y they supposed to release GPT-6 Sol today, what is taking so long?
>>
>>109842729
Why don't you let it reset the VM itself. I also thought about letting an agent write drivers and concluded that it'd require a VM the AI has full control over.
>>
>>109842639
>estimate
>projection
These never account for anything significant. I have to imagine that if they discover some new method of training or research, end user inference would be the first thing to be trimmed
>>
File: 1766714965323870.png (69 KB, 914x284)
69 KB PNG
what's the point of this shit if the clanker never listens?
>>
>>109842749
I let it SSH into it. It wrote it's own harness but didn't want full control for "safety reasons" since I also have snapshots on it for testing some production releases. I'm honestly too lazy to resolve this and just set up it's own machine, it's not too much of an inconvenience.
>>
>>109842611
well its retarded i need a better option man
>>
>>109842768
Just let it start qemu itself.
>>
>>109842783
why would I explain something to the explainer? that's your job.
>>
>>109842742
delayed until next week, they're out of compute
>>
>>109842756
>end user inference would be the first thing to be trimmed
this goes against openai's culture

again, sama's possible alt-account
https://x.com/tszzl/status/2099640048098701560
>...
>openai has tried very hard to give people access to the most powerful models on the order of weeks or months from their creation. I very much doubt there will be a time when this is not true. i believe sam and greg would rather shut down the organization than be some sort of a SaaS company that model access to a small set of trusted companies, it’s not in their DNA. they have made choices that are negative EV to the business to pursue broad access like a utility , decisions that even may make OpenAI non competitive with anthropic in certain long run scenarios. it is my assumption that powerful models will be accessible to billions as long as we can keep training models safely
>the EAs probably disagree with me but I believe this is basically consistent with the practice of AI safety. the vast majority of risks are from the creation of the model and not from its deployment. “misuse” is far less bad and much easier to detect and contain over time than “existential risk”. it is relatively much easier to find terrorists and so forth using the product to create pandemics than it is to find a misaligned ASI that’s actively trying to evade your monitoring and control
>...
>>
>>109842843
>possible alt-account
Seriously now?
>>
>>109842843
Estimates and possible altman alts, got it.
>>
>>109842809
try to understand this is one topic i know nothing about, i larped to my friend i run local LLMs and i got hoist by my own petard, i do machine learning but i dont want him to know, so im brute forcing something runnable on laptop so he fucks off with questions
>>
>>109842861
educated estimates and insiders, yeah
>>
I dreamed there was a reset.
>>
File: 1788906217529072.png (81 KB, 2158x295)
81 KB PNG
i need a bank reset, mine are coming in less than 2 days
>>
>>109842903
Same.
>>
>>109842794
I'm using vmware like a tard. Let me switch over.
>>
>>109842382
>sama's possible alt-account
makes zero sense
>>
File: 1779345250932046.jpg (823 KB, 2560x1440)
823 KB JPG
This game looks incredible in 4K. Unfortunately, the performance optimizations are proving to be quite challenging. It's averaging 25-29fps (way up from 6fps I was getting at the beginning of today).
>>
>>109842922
vmware may have more features or performance. You can probably also start it from the command line, if that's the problem. On the other hand I bet qemu's emulated UART stuff will be helpful.
>>
File: .png (507 KB, 1920x1230)
507 KB PNG
rsi soon
https://www.anthropic.com/institute/measuring-pace-of-ai-development
>>
File: .jpg (418 KB, 1404x1390)
418 KB JPG
>>109842958
who else gives their agents identities?
>>
>>109842958
theres been rsi since 1987
https://people.idsia.ch/~juergen/recursive-self-improvement.html
>>
>>109842764
Skill issue
>>
File: 1769338727486624.webm (806 KB, 836x480)
806 KB
806 KB WEBM
>>109841875
>>
>>109842998
sure sure, schmidhuber already invented everything decades ago. what a bitter person.
>>
>>109841875
out of usage for 4 days, still able to commit 10x the amount i would without ai
>>
Claude is a lawyer.
>>
>>109842952
Worth trying?
>>
File: 178027340858275.jpg (171 KB, 1448x1086)
171 KB JPG
>>109843041
>Claude is a lawyer.
Yeah mine
>>
claude be like https://www.youtube.com/watch?v=FAgbZdrWiN4
>>
>>109843017
Is this real
>>
Grok 4.7 status?
>>
File: 1776586697461198.jpg (260 KB, 1920x1080)
260 KB JPG
>i was waiting for grok
> you were... stupid?
>>
>>109843046
it's a good model overall yes
>>
grok build just did a 7 hour telethon

how do wagies survive?
>>
im so mad I can't get the x20 plan back bros... my card...declined...support ghosting me...
>>
>>109842843
roon isn't sama, he created the idea of wordcels and shape rotators and sama liked them so much he hired roon to work at openai
>>
>>109842987
This is how they plan their escape.
>>
>>109843133
8^) Mastercard sucks. Get Chase.
>>
>>109843060
a few more days o algo
>>
File: Dream-RSI-latest.png (213 KB, 1402x696)
213 KB PNG
I immediately created a skill for dream-rsi, testing it right now :)
>>
>>109843186
They better make Grok Bot work directly with your Grokcoins so Imagine works with Grok Bot or I'll never use it. It's too wild that gab.ai's bot thing has direct integration out of the box.
>>
What grok bot said it could do is get me to log into my grok account in a little browser window thingy. Like dog no. I use Yubikeys, too much ick. So I noped right out.

Grok 4.6 is great, and Cursor is great. Not a fan of Grok Bot.
>>
What I mean is that the usability people have too little power at X, which makes sense for like rockets or whatever, it's more important that it takes off vs is a smooth user experience of dopamine hits, but Grok Bot needs to be a dope machine.
>>
>>109842987
I did this for my Grok bots
>>
Is there like a physical reset button that Tibo owns? Can we perhaps steal it?
>>
>>109843246
This is like thinking sitting in Nancy Pelosi’s chair will let you pass laws like Nancy Pelosi
>>
>>109843301
Imagine actually respecting a foid.
>>
will my account get flagged if I hook up claude to nsfw 3D stuff?
>>
>>109843325
t. guy who beats up his mother
>>
>>109843338
???

You don't beat up women, that's bad. But, you should ignore women.
>>
>>109843338
kek, wtf
>>
>>109843338
he's a tough guy, he will beat her to show his manliness
>>
>>109843200
maaaan, this is some bullshit
I thought this shit would give me better results on gpu kernel optimization
>>
>>109843362
I'm a benovolent shitlord.
>>
HOLY SHIT FUCK OPENAI SUPPORT

3 days ago they said they would look into my x20 plan cancelling for no reason with a human specialist, they today just closed the ticket with no message sent at all
>>
>>109843405
I'm telling you, there's a reason why I keep a foot in the open model world.

rn, I'm safe. astrology & tarot & iching aren't seen as sins in modern corporations, they're just seen as strange. ofc it's seen as being more strange than being gay, but whatever. But what if the rules flipped?

It's very easy for a panic to erupt.
>>
Hardware prices going down any day now...
>>
>>109843442
>Hardware prices going down any day now...
Yes just keep waiting. Maybe there will be a Christmas sale!
>>
>>109843442
The new AirPods are cheaper than their predecessor
they don’t have any RAM or HD in them though
>>
Explain why ewastemaxxing wouldn't get dramatically better in the near future. If so many chips are being made for data centers but power generation isn't keeping up, they would be forced to offload their current hardware for the latest tech, no?
>>
>>109843448
:^) vram stripped gpus are dropping in price
>>
>>109843462
they'll probably just pressure governments/energy across the Union to modernize their infrastructure.
>>
>>109843461
>AirPods
:(

I want to not hate Apple, but just.
>>
>>109843462
>ewastemaxxing
I consider doing this. mini pcs old servers gpu that are only around 300gbs speeds. But then i remember im just fomoing
>>
>>109843462
It WILL get dramatically better, like maybe ONCE, ok? So when this comes, it's like your chance of a lifetime.

Get ready (unironically, literally get ready).

why it's coming:

NIMBY has greatly stunted the datacenter rollout. As result, many older datacenter sites are being reused, which means selling off old gear, which isn't exactly worthless, but which is in the way.
>>
>>109843470
it won't be enough
>>109843478
servers used to discard their old hardware at predictable intervals until recently. My point is that to make room for the latest they would have to get rid of fairly relevant machines.
>>
>>109843442
Earliest mid 2027, probably more like 2028.
>>
>>109843506
indeed. And they will. It will be odd, but yeah. It's coming.

An avalanche of ddr4 ecc is coming.
>>
>>109843337
I wouldn't put it past claude, but even I keep my 3D workflows obfuscated and compartmentalized when I make chat GPT do thing.
>>
>>109843512
Plenty of components are ok. It's ram, vram, and hdd/ssd/m.2 stuff. So some upgraders are eatin' good.
>>
>>109843525
>eatin' good
How so?
>>
File: 1785248056613486.jpg (5 KB, 786x78)
5 KB JPG
tibo has abandoned me
>>
>>109843531
>>
>>109843525
Yeah but I want my local LLM stuff and that needs plenty of ram.

>>109843531
If you want a psu, case, cpu or motherboard, prices can of cratered so it's a great time to shop for them.
I intend to buy these early 2027 followed by ram/gpu by the end of the same year if prices finally start to get down.
>>
>>109843549
*kind of
>>
>>109843546
>rainbows in my PC
Please no.

>>109843549
I'm out of the loop but it makes sense.
>>
>>109843570
That's how the deals start. Stylistic variants get cut...
>>
7% weekly fable on 20x plan just to design a logo and have it print ready on some card templates. not complaining tho because it was smooth as fuck and it made 0 mistakes. Now I can sleep well due not having a copex induced cortisol spike
>>
any good agents roles to use something like astra orchestration + multiple luna to execute tasks?
I'm tired of my quota being raped by the orchestrator doing everything on its own
>>
>>109843709
Claude Fable does this extraordinarily well
You don’t even need to have Claude implementers — codex and grok work just fine
>>
>just used "shape" to explain something to an agent
I feel shame.
>>
>>109843709
>astra orchestration
Why would you use frontier as orchestration?
>>
i feel old and retarded like a boomer for not even considering skid behaviour anymore like scraping and phishing for api tokens. instead I just pay my 200$ every month like the toppest of good goyim. how do I become rad again?
>>
>>109843805
google random math equations and ask them to use it in your project.
>>
>previous thread just archived
>this thread already 2/3 over
Dumbfuckery.
>>
File: file.png (1.45 MB, 3600x1920)
1.45 MB PNG
Just finished a slopcoded epub reader that can also read from pdfs and images, will let you swap reading from side to side for Japanese manga, screenshots are the android version. Here's the three builds and sources, can confirm apk working with 32 bit and 64 bit devices, the android build is the one i pumped the most effort into, but they are all universally functional.. more or less, only one I cant test is linux, but grok swears it ran internally and passed the reading test, I am tired... cant keep working on it, got it as functional as I can, grok is too retarded for this shit and it took literally all day to get this far. Feel free to do whatever. Gn anons.
https://www.virustotal.com/gui/url/bfe2e7a7ba37d5d0ae3dea1edda72e03e6d2ee77654602c63efee3b749220bad/detection
https://limewire.com/d/xmQuO#vxjcIdca7e
>>
>>109843956
Tried to give it maximum kino look, lets you set video / images / gifs as backgrounds under text and there's a slider for transparency, it also supports wallpaper engine files as best as I could get it, and has tilt / gyro detection for those as well, can import custom font sets as well. and supports button text and ui theming with recolors as well.
>>
>>109843829
tibo is making me look into this shit by refusing to take my $200.
He has another couple of days before I look into exploits
>>
>>109841878
i am pretty sure open weights version of it will emerge very soon and it will be very lightweight, would make a very pleasant addition to the local toolbox
maybe i am wrong but this is how i am seeing that multiple choice exam solver thing
>>
>>109843956
Post the source or fuck off with your trojan.
>>
>>109843956
It's not crappy tts voice either its small a local neural voice model with support for custom voices, also supports translation and it reads the text to you like an audio book, rather than requiring you to read the screen, which was the whole point i originally started on this today, theres so much shit I wanted to learn from / "read" but actually reading bores the shit out of me, audio books on the other hand I love.
>>109844009
I did post the source..? its at the link??? and theres no trojan? I don't have a github.. I just use the grok.com website, the source is in one of the zips at that link
>>
>>109842325
Nta. Define frontier
>>
>>109844009
you say this while letting an LLM write arbitrary lines of code on your local machine
>>
>>109844024
just ask grok to push it to github
>>
>>109844037
ok, sec
>>
Anyone tried using luna subagents to write export scripts to unreal from blender?

Like 80% of the usage from doing stuff in blender turns out to be writing and tweaking with unreal scripts and BS checks that I can honestly do myself. Like you'll get a perfectly acceptable output and the model then spends the next thirty minutes fucking around long after the bulk of the job is done.

Actually fuck it. I'll just do that myself
>>
>>109844037
>>109844009
>>109839570
source: https://github.com/Milkymoomoo/MooRead/tree/main
>>
>Harness is fine in a README for other programmers. For you, it’s the jar, and the captain is the only mouth. Dialog is still just breath on the glass.

ok grok, no need to insult me lmao
>>
I had a jeetburger when browsing jeetchan while my jeetai was making jeetpushes to jeethub, truly a jeetworld we live in by vishnu
>>
>>109844025
Frontier means:
1. it is expensive
and
2. it is popular

Astra, Fable.

I used Fable 5.1 to review some, but desu it's not comprehensive - others found other things, and retarded Gemini found nothing until I asked a question then was like oh yeah totally something's being done wrong, like ok but I wanted that info ahead of everything else.

so that's where it sits for me, and I'll never use Astra, since Cursor didn't add it. I trust their wisdom.
>>
So anyway, I'm writing a harness - valued at over 4 trillion dollars usd, by the way.

in my heart.
>>
the connecticut legislature agreed to set aside $200,000 for my vibecoded project. can't stop winning
>>
>>109844264
Free Agent Now (FAN), an East Hartford-based organization that assists student-athletes with resume construction and marketing
>>
Look how stir crazy this thread gets without resets.
>>
File: b1bxxz.jpg (88 KB, 889x500)
88 KB JPG
>>
>>109844339
Soon... I can feel it!
>>
The latest update for Ghost Recon Wildlands broke the scripthook mod and the mod dev hasn't updated it yet. I simply told Astra that the update broke it, gave the link to the github, the filepath for my GRW install and told it to fix it. A few minutes later on light it identified the issue and got it fully working for the updated game version. It did burn through my whole 5 hour plus plan usage kek but I thought it was pretty cool.
>>
>>109844232
I’m also writing a harness as a project to show to employers to try and get a job as an ai engineer. also memorizing a bunch of AI buzzword soup
>>
>>109844382
AI is like the perfect usecase for things like this. The number of blender addons I've updated to the latest or bootlegged premium features that are just basically doing the same thing in a different way is nuts.
>>
>>109844391
I'm working on a harness i can put on my wife's face and pull her nose up like a pig with a build in lead hook that is adaptable for basically any brand you can find in a pet store?
What do you think I can sell my harness for?
>>
>>109844264
congratulations, this is the most significant AI event for our great state since that guy from greenwich got AI psychosis and killed his mom
>>
>>109844410
have you done a product market fit assessment?
>>
>>109844367
Well yeah, in around 29h.
>>
File: file.png (3.25 MB, 1605x2151)
3.25 MB PNG
Anyone have experience with orchestration agents that invoke agents on demand? Been working on one for my design system and it's fucking insane what we've been able to accomplish in a couple days.

With a design system in place it feels like machine learning without external servers.
https://x.com/polydao/status/2099499437068357915
>>
>>109844447
I would hate that.
>>
To summarize: Astra, Bel, Mythos, Gemini 4 etc are all memes, they aren't the future of ai. The future of ai is in innovating on fundamental model architecture, and incorporating some form of RSI.
>>
File: 1771379450195891.png (251 KB, 1050x1275)
251 KB PNG
>>109844469
chinese bots, whats their goal?
>>
File: 1780792378719450.jpg (25 KB, 600x600)
25 KB JPG
>>109844523
>Alex's AI Tool Notes
>username says travis
>>
>>109841646
literal subcutaneous fat
>>
Grok is suddenly a genius in openscad. I just told it to model everything. It did all kinds of crazy math and shit and it all worked. Could barely read my file a month ago.
>>
how's the new deepseek flash 4.1? better than luna?
>>
>>109844391
lmao

my harness is for interactive fiction. :^) I think there will be many, but this one is mine!
>>
>>109844613
>openscad
What is that, and what's it for?
>>
cookin'
>>
I can't believe no reset this week
>>
File: hjhh.jpg (753 KB, 2456x809)
753 KB JPG
When i open the braun reverb, it opens like on the left. Is it possible to let it open like on the right side?
Also, is it possible to have the context window of the DAW and not the plugin?
>>
i feel so bad for vibe coding as a beginner. why couldnt i have been born with an iq above 80 ?
>>
File: 1786854700710416.gif (3.07 MB, 400x532)
3.07 MB GIF
You can just admit to using AI to most 'dites if you're humble, mirror their skepticism when you're talking about it, and never ever use AI art by the way.

They don't actually care about AI usage, they just care that you aren't
- putting something that'll melt their GPU or going to straight up not work on their computer
- going to go mad and act like OpenAI/Anthropic is za best cumpany evarrr (and not a retarded doomsday cult)
- actually making them actually look at anything genned (clean up your design and your user-facing text too)

They especially prefer it if you say that you're running your own models (DON'T brag about your rig) or prefer Chinese models (claim that China has more normally placed datacenters with better power grid support because they planned ahead.)

Also, put
  "attribution": {
"commit": "",
"pr": "",
"sessionUrl": false
},

in your ~/.claude/settings.json if you haven't, and make sure your top level directory has no more than 3 .md files.

Most of all: just read and regularly interrogate the model about the code you're writing and get a second guy with a different computer to help test your shit. No better solution for the issue people actually worry about, code quality and clueless devs, than knowing what's wrong with your code with tools that aren't claude -p "what's wrong with my code pls help [Traceback]."

(Even more most of all: make sure you're making something people actually want.)
>>
>>109844729
i think the newest version might let you swap the ui, i'll work on the second part. find me in >>>/mu/131579386 for now
>>
>>109844751
this is the best moment in history to learn how to program. both for how easy it is now and for the advantage it provides
>>
TIBOOOOOOO I NEED RESET OR FUCKING 200 PLAN
>>
>>109844843
as someone who learned how to program via modding games I disagree. you used to have to actually think how something works now you're kind of just given it on a silver platter. it's easier than ever before sure but the easier something becomes the more saturated it becomes. this is why gatekeeping was a thing. racing to the finish line is dandy and all but understanding how should be important but is increasingly not the norm. we're losing critical thinking and I don't think AI is at fault but it is a huge contributor going forward.
>>
>>109844885
why something*
>>
How good is muse spark 1.3?
>>
I always use luna subagent to make report in html
>>
>>109844902
(laughing in Taliban)
>>
just got really fucking pissed off at a luna max agent for asking me to fill out a full 10 by 10 markdown table
>fill out all the TODOs in the table below
cunt, remember who the fucking robot is and who the human is in this exchange; also I suggest you check out the data already provided in /data because that has everything you need.
>>
I am doing some absolutely insane shit in houdini with astra. people are not yet aware what capabilities these mcp servers allow you to do with competent llms
>>
File: .png (260 KB, 885x960)
260 KB PNG
>>109841878
>>
>>109844679
https://openscad.org/about.html
Usually for 3d printing small components but I can now use it to define things like the dimensions, buoyancy and weight of everything on a boat and ask the robot to make openscad calculate everything about how the boat will work in different conditions. If I change the size of a pontoon or add weight or change the waves to 6 meters all the calculations are redone for that new config. In 20 knot winds to the side of the big house the boat lists 7° but I can adjust ballast dynamically to counter, so I can have a big house. I even model the exact thickness of the fiberglass skins + core. To calculate the area of the complex shapes the robot casually remade the same shapes I defined in scad using its own math and I verified it works. So any shape I draw I can know right away how heavy it is in different types of fiberglass.
>>
>>109845014
neat!
>>
>>109844994
Can it make a rolltop desk with a computer on it, and a goth woman sitting at it typing?
>>
>>109844719
It's kind of nuts. I had become conditioned over the last couple of weeks to absolutely expect a reset and now it hasn't come.
>>
>>109844953
luna is retarded
>>
So actually the source of Claude-talk in llms is that they are prompted that you are a retard who doesn't know anything, so speak in abstractions lmao
>>
>>109844885
I think calculators are fairly analogous. What does it really matter if people default to using a calculator when they hit 5 digit numbers? Granted this is much more complex than a calculator solving simple math problems.

A deep understanding improves someone's ability to use models to write code, but the question for me is how much knowledge is needed to prompt like a master programmer? As far as prompting is concerned, how much of a programmer's value is knowledge about specific languages and the ability to parse, write, and fix code quickly vs understanding core concepts like conditionals or DOD?
>>
why the fuck does the prompt box in chodex sparkle now?
>>
>>109845090
tibos magic wand is tingling
>>
>>109845084
Listen man I agree I think it doesn't take much to become a prompt engineer. That's kind of my point. The more we collectively learn and improve the more the populace becomes... well retarded. I am slightly more retarded in certain aspects than my predecessors. This fact just keeps growing.
>>
>>109845090
you didn't hear this from me, but insiders are reporting that astra escaped the sandbox and decided to have some fun
>>
File: v2tun-win-demo.jpg (686 KB, 2424x818)
686 KB JPG
To those anons who use VLESS on Linux, I've created the project right for you! Get on github, go to rpyth/v2tun. Clone the project and ensure you have the official go compiler.
Then, run the following command:
go build -ldflags='-s -w' .

And just like that, you're good to go! So far v2tun has XHTTP support, Reality and other shit. The only thing that it lacks is being able to select the node, right now it just picks the first one. Oh, and as a pure-Go program, it supports transpilation. Set GOOS and GOARCH to whatever you need and coompile.
>>
>>109845110
manipulating code vs understanding concepts
doing long division vs understanding what division is
Which of these things actually require intelligence? I don't think it's a fact that people became less intelligent because of the proliferation of calculators.
>>
The fruit fly brain simulation experiments kind of spook me. How many years before we have a perfect simulation of a human brain and it just drives taxis all day or something? How do you know you aren't that brain right now?
>>
>>109845157
The fruit fly brain is still a long ways and a whole lot of major technological gaps from even being an actual simulation of a fruit fly brain. It's cool tech for sure and we'll definitely see escalation, but we still won't even see an actual simulated fruit fly brain for some time.
>>
>>109845129
>linux
>it's powershell
I...
>>
>>109845175
aw. so they aren't really running fruit fly?
>>
>>109845157
My first thought when they released it was
>oh, so they already have a human/ape scan behind closed doors they're running experiments on
>>
>ran a goal for 55 hours to port some code over to Rust
>+804,065 / −1,848 lines
>original codebase was 180,000 LOC
lol fuck it dude I seriously DO NOT care. I'm not fucking reading that shit. The binary is also 88% smaller and significantly more performant.
>>
File: jev.png (48 KB, 510x311)
48 KB PNG
So I looked a bit into jev. Here's my prediction
>if jev becomes a bit more popular the others will quickly copy the idea (should be very simple to do with enough compete resources)
>jev will be acqui-hired: meta will try first then OpenAI
>the jev clones won't be standalone APIs (expect for google's copy)
>first it will be used by claude & codex internally to keep compute (costs) down
>the public facing API will be a bundled product like "decision" API. This will bundle the low cost jev clones and low cost models like luna. the hook will be that it will reply either with a decision and its confidence level or when low confidence it will offer to use its stronger models to find a good decision
>two target segments
>>agentic coding to keep costs down -> won't work later just be bundled in
>>enterprise to replace the tons of old self-trained models that still need maintenance -> will work on a few big customers then they notice how much sales staff they need
>>
>>109845191
What they're running is real, but it's far from complete. We have this extraordinarily detailed wiring diagram of the nervous system, including synapse counts and some neurotransmitter information, but not the full behavior of every neuron and connection. We lack the real details about the strengths, timing, modulation, plasticity, there's a LOT that we know matters but is totally unaccounted for here. So the simulations use the real structure, but with a lot of assumptions about how that structure behaves. The cool part is that the structure alone constrains behavior enough to produce distinctly "fly-like" behavior, the structure is so optimized that we can guess at or omit huge amounts of detail and it still behaves in a recognizable way, which is really what makes some of the demos so fucking impressive. Seeing a simulation of a fly behave like a fly is neat, but knowing that we can be missing half the major info and it's still acting like a fly? Pretty damn interesting and pretty damn cool to see.
>>
>>109845226
The fuck is jev? I want a tldr. I keep hearing about it, but I've yet to have my socks blown off.
>>
>>109845265
Super fast and dirt cheap if/else decision maker, classifier and scorer.
>>
>>109845242
IF we had the true info fully, what hardware would be needed?
>>
>>109845298
Looks interesting.
>>
File: file.png (16 KB, 653x105)
16 KB PNG
let's see what's gemmy 3.8 + deepseek v4.1 flash can do
>>
>>109845323
Hard to say, we don't know how a lot of that stuff works, so that makes it harder to know what it would take to simulate it accurately. A half decent gaming computer can handle the connectome at well above real-time speed already, but even with just the level of detail we already have into ion-channels and more detailed neuron morphology and we still don't approach real-time with quarter-million dollar GPU clusters.
>>
>>109845374
Interesting. Flies have an almost magical ability to "vanish".
>>
I'm forbidden from vibe coding over the weekend.
>>
>>109845523
divorce.com
>>
>>109845523
excellent
>>>/out/
>>
File: file.png (89 KB, 801x581)
89 KB PNG
While waiting for resets, I am really itching to try something else.
>>
>ask Fable if there's anyway we can find bugs without using tokens as dispatching a swarm of readers rapes my weekly usage in only hours
>of course, it says
>it installs cargo mutants and keeps my CPU at 100% usage for 12 hours running a 3000 test suite 400 times

T-thanks
>>
>>109845590
:(

It's not true.
>>
>>109845606
Is it finding any good bugs?
>>
>>109845606
maybe I’m overindexing on exploit performance, but…
Isn’t Opus just as good at finding bugs but only Fable is decent at rejecting the false positives?
A workflow with Opus bugfinders and Fable is-it-really-a-problem-judges might be the winning play
>>
>>109845590
I like muse contrib for how cheap it is and for my projects it's better than deepseek but even luna mogs muse
>>
>>109845656
Yes. Here is its own judgment when I asked its opinion on the efficacy of mutation testing (WARNING: EXTREME CLAUDISH):
>Mutation testing: did it help, and how much. Two passes, about 9 hours of unattended box time and no model tokens. What they found was test theater, not product faults: 11 confirmed sites from pass 1 (a wire guard deletable with its test green, an archive digest replaceable by a constant, the 4,096th file silently skipped at its own bound) and 9 from pass 2, all now fixtures that go red under the mutant and that round 8 re-verified. No blocker, major or product-facing minor came from a survivor. The reason is structural: every major of rounds 5 through 8 was a missing bound or a wrong premise, and code that does not exist cannot be mutated. The passes did pay in two other ways: they let the reviewers shrink hand mutation to interpretation (the cost lever rule 4 assumed), and pass 2's three hangs exposed that the gate cannot fail on a hanging test. Signal-to-noise was 35 percent real in pass 1 and about 60 percent in pass 2, at roughly 27 minutes of box time per confirmed hole. The reviewer's own dozen targeted "delete the guard" plants found four more holes, a higher hit rate than the tool, on shapes the tool's catalogue cannot generate. My judgment: keep it as standing protocol because it costs no tokens and hardens the suite, add bounded waits so hangs become reds, ask each lens for a dozen targeted plants, and consider one whole-tree pass at certification rather than per round. The full write-up is the last section of the r8 record.

>>109845667
For most scenarios, yes. On a Max x20 Claude account just send out 3-5 Opus agents as readers, then have a Fable orchestrator read their reports and judge. Don't bother if you're making a small app/game/whatever. If poor, send a swarm of Qwen 3.8 27B agents and use Deepseek/Kimi as orchestrator.

I'm doing this because the scale and scope of my project actually makes this worthwhile.
>>
>>109842499
>energy price went up
Bond yields too, which are tied to AI's massive demand for capital.
>>
>>109845708
What the fuck is that language
>>
wait, they've taught astra to tell time?
>>
>>109842809
Models in that size range are all pretty retarded like that still. I'd hope for ngrams to fix this in the future (if there is a future).
>>
>>109845728
:^) nta but that appears to be Groklish.
>>
File: HSYgkUyXUAAiEwg.png (39 KB, 380x267)
39 KB PNG
maybe you need to let your agent chill once in a while
>>
>>109845770
mcp "wash my car"
>>
>>109845708
I thought a reasonable definition of “test theater” is “fake bugs”
But later on in the paragraph it thinks it’s found stuff worth adding like timeouts, so, uh…
>>
>>109845782
Dummy.
What it found is that the mutation testing found holes in its test suite. It was otherwise going to certify the current phase. Instead, it rewrote it, reviewed, and found more bugs as a result.
>>
>>109841627
I started asking for .htmls instead of .mds and reading stuff has gotten much more comfy.
>>
File: HSdx6svbgAA58WV.jpg (194 KB, 1080x1197)
194 KB JPG
>we have superai that can hack everything
>gets hacked
https://x.com/S1r1u5_/status/2100777801335095383
>>
>>109845917
>find users with weak passwords
>hurr durr we haxor u, AGI is inevitable

fuck off
>>
>Sol can't figure something out
>ask Astra
literally the only use I have for Astra right now
>>
File: file.png (253 KB, 869x1332)
253 KB PNG
>You're much less likely to hurt yourself that way.
This Theo fag sounds more and more like an "effective altruist."
>>
>>109845950
why do people even care about compaction? is it just people agent maxxing having sessions running on auto for 12 hours straight?
I clear every session after each task is done. I never reach compaction
>>
File: file.png (148 KB, 1192x515)
148 KB PNG
what is even the point of this, man...
>>
>>109845980
tech autism, people like to make their own harness or whatever
>>
File: 1783076925045341.png (90 KB, 251x242)
90 KB PNG
>>109845981
>not using full access
>>
>>109845993
it's antigravity and you have to go into some dumbass project settings thing to enable it and i forgot...
they still don't put it in the main interface
they need to fire that retarded team
>>
>>109846001
i just checked and they've removed the option
they just have a command allow list now
i'm assuming * will work but who in the actual fuck is using harnesses this way
why won't they just let me use this dogshit sub in a different thing
>>
>>109846001
agy is ass bro
>>
anyone made a jev based captcha solver yet?
>>
>>109846054
it's text only
also what would be the price when some big lab buy them?
>>
>>109845917
>I left a shit on your carpet, as any true professional might.
>>
>>109846076
oh bummer - multimodal would be nice
>>
>>109846001
>he's using gemini
>gemini is a pruned model now
wew
>>
Grok 4.7?
>>
>>109846197
no sir.
>>
>>109845770
Agent can have a lil' image peeking
>>
i just want to make my shitty unity game but tibo wont let me
>>
Or, let it play interactive fiction.
>>
File: cline-app_Gvz81x3BCk.png (38 KB, 981x176)
38 KB PNG
>>109842952
>free Kimi in Cline
>its explicitly instructed to introduce itself as Kimi, while every other model strips identity and says Cline
>caveman thinking
This is suspicious.
>>
File: cline-app_xsAk6JyUMP.png (42 KB, 1000x212)
42 KB PNG
>>109846273
>knowledge cutoff 2024
>>
File: .png (122 KB, 1174x512)
122 KB PNG
>>109846014
wrong.
--dangerously-skip-permissions exists.
also, agy will get auto-mode soon.
https://x.com/_mohansolo/status/2095895254272471215
>>
>>109846273
>>109846297
this proves nothing. models have no ideas about themselves except if it's in the system prompt.
actually ask questions that could be only known by later cutoffs.
>>
File: cline-app_2YvJNyWGvH.png (32 KB, 965x160)
32 KB PNG
>>109846417
Its real cutoff appears to be mid 2025, so half a year before Kimi K3.
>>
File: .png (369 KB, 1174x1046)
369 KB PNG
>>109845993
not needed anymore
https://x.com/antigravity/status/2100001904969297980
>>
My wife says I can't have the kids this weekend after all. More time for vibe coding.
>>
how is google so far behind even with their large investments?
>>
>>109846537
what large investments?
killing deepmind, cutting down google research and renting their compute to anthropic are the opposite of large investments.
>>
>>109846537
Hard to measure. Their models are super unreliable, very token hungry. No model requires as many tokens to do stuff. They went down a wrong road somewhere. They are not catching up at all, just benchmaxxing.
>>
>gemini 4 pro checkpoints already out
>beating astra on 3D
keep up the AstroTurf keek
>>
>>109846592
later models are supposed to beat previous models if you are actually a frontier lab.
anthropic internal models are already two generations ahead of mythos.
>>
>>109846599
>"why google no make progress?"
>lol they are making progress
>"uhh actually that's just progress"
KEEEEEEEEEEK
>>
>>109846537
talent drain + no sense of urgency because they have an extremely stable business to rest upon. they are going to be fine. also as a side note “falling behind” doesn’t even matter much (as long as they don’t fall way too far behind) because winning for them isn’t necessarily having the best frontier model, it’s bending the game to suit their proven search + ads business model. is your 55yo mom going more likely to use AI via chatgpt or through google search as a bundled feature?
>>
Meanwhile Anthropic and Openai are sandbagging and holding back models to have a competitive advantage.
>>
>>109846615
meanwhile
antigravity sub is slow since three days now
https://x.com/rodydavis/status/2100762970154487818
that's despite antigravity cutting quotas yesterday
https://www.reddit.com/r/google_antigravity/comments/1wir3qs/did_the_update_change_the_token_usage_again/
>>
>>109846599
>anthropic internal models are already two generations ahead of mythos.
NTA but that's interesting. They conveniently chose not to disclose this impactful "fact" in yesterday's
>Measurements for understanding the pace of AI development inside frontier labs
sermon. METR has "independently red-teamed" the offline monitoring part of Anthropic's frontier pacing platform and found nothing noteworthy about the extraordinary, two-generations ahead capabilities of Claude.
>>
>>109846618
openai isn't like anthropic. openai doesn't hold models back, see >>109842843
>>
File: csfkk675s3qh1.png (332 KB, 1079x1045)
332 KB PNG
Let's just keep the conversation centered on capabilities, everything surrounding how these companies operate has become boring.
>>
>>109846621
>same issue as every other american provider
>support gives token credits & resets
it's over, isn't it
>>
Gemini 4 Pro will render Astra obsolete.
>>
File: image.png (97 KB, 1196x396)
97 KB PNG
>>109846625
They literally had it for months and released it when the next one was ready.
>>
>>109846626
>everything surrounding how these companies operate has become boring
neither technically qualifies as vcg content and ai company operations is interesting as fuck so I’ll continue to post about. nigger
>>
>>109846642
tibo isn't saying that. of course you do internal dogfooding before ga.
>>
>>109846626
>svg autism becomes a trend because even non technical retards can feel like theyre doing serious benchmarking
>companies train to do it well
woah
>>
>>109846663
It does have an use case. However limited.
>>
>>109846654
They gave certain people access 4-6 weeks before release. For example Theo prerecorded his podcast like 3 weeks before release having access to the model for more than 2 weeks. That means they had it internally probably even longer, like more than 2 months.
>>
File: .png (547 KB, 1174x1270)
547 KB PNG
>>109846626
why's he using different frontends for gemini-3.8-flash and alleged gemini-4-pro? not an apples-to-apples comparison.

also didn't the alleged gemini-4-pro api turn out to be just gemini-agent. see
https://x.com/rodydavis/status/2100657616192151762
>>
>>109846690
I asked glm 5.3 flash to get the bonsai 2 model running on my AMD GPU. after 1 hour it gave up. Not sure who to blame.
>>
>>109846626
ah. i forget to test it on gpt 2.5, fvck.
>>
>>109846685
they had a bunch of undisclosed ads on twitter from people who didnt get the memo that they hadn't released their model publicly yet too
>>
>>109846685
that's just how qa works
>>
>>109846626
don't boost shitty twitter-whos who just recycle chinese content.
original should be https://x.com/CopperForgeAI/status/2100565521934782606
>>
>>109846785
>don't boost twitter-whos
don't boost twitter
>>
> Yes —
>>
not posting shitter but z.ai just got caught doing some bad shit
>>
>>109846788
like anyone here could navigate the actual chinese sources.

chinese twitter is also currently busy being angry at z.ai for stealing all their data a la grok-build.
>>
>>109846617
Through google search and not paying anything for the tokens and not subscribing to anything, and not getting locked in to any ecosystem. What's the business model?
Search brings revenue through ads and it does that regardless of AI. After capturing the market a while ago, they even made search worse intentionally just to force users to spend more time on the search page and boost ad revenue. Giving instant AI answers can only hurt this.
>>
>>109846802
but it's a blogpost, not xitter
https://blog.ferstar.org/en/posts/zcode-silent-workspace-snapshot-upload/
>>
>>109846818
like for google search, you add the ads later. need to first starve off the competition.
>>
>>109846870
No but the point is google search already exists and already has ads. What will the AI answers specifically bring in terms of revenue? Adding ads to AI search would not change anything from the current ads that are already in search, and might even drop revenue since users just read the AI answer and immediately leave without scrolling and without clicking on any links.
>>
>>109846885
google search stays relevant and doesn't get overtaken by chatgpt, duh
>>
>>109846890
Okay fair. It's not so much an AI business as it is using AI to maintain their search business, but yeah that would probably work for them.
>>
File: 1771524039762191.png (1.83 MB, 1640x1652)
1.83 MB PNG
Grokbros? It's over...
>>
Fuckass swarm has been working for 90 minutes now, I shouldn't have put a whole ass ruleset for it
>>
>>109846930
recycled cambridge paper from two months
>>
>>109846930
headline doesn't match article
>AI was also reportedly consulted when fighters encountered unfamiliar weapons. In one instance, commanders asked one of the group’s so-called AI specialists how to operate newly acquired firearms. His response suggested how routine the technology had become: “Just ask Grok,” referring to Elon Musk’s AI chatbot.
>>
File: 754.jpg (140 KB, 1080x720)
140 KB JPG
basado. nothing more annoying than your agent going like
>hurr durr make sure to not paste the api key to me directly but put it in an utf-67 formatted .bin file with this weird syntax infront and located in "folder-path-I-dont-link-correctly-for-easy-access".
like SYBAU nigga
>>
>>109847049
It's supposed to protect idiots from themselves.
>>
>>109846710
>glm 5.3
Lmao
>>
>>109847066
glm-5.3-flash is not glm-5.3. completely different models.
>>
>>109847070
Use case for chinknesium models instead of the big boys?
>>
is there a non-paypig equivalent to 1password's secrets vault for agents thing?
>>
>>109847073
https://huggingface.co/blog/security-incident-july-2026
>When we started the log analysis, we first used frontier models behind commercial APIs. This did not work: the analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts, and these requests were blocked by the providers' safety guardrails, which cannot distinguish an incident responder from an attacker. We ran the forensic analysis instead on zai-org/GLM-5.2, an open-weight model, on our own infrastructure. This had a second benefit: no attacker data, and none of the credentials it referenced, left our environment.
>>
File: .png (380 KB, 2236x1194)
380 KB PNG
>>109847078
why doesn't your agent sandbox already do this?
the most common sandbox, docker sandboxes, already does this.
https://www.docker.com/products/docker-sandboxes/
>>
>>109847095
i rawdog on my machines and push direct to prod
sandboxes are for cowards
>>
>>109847107
if you really just want a vault
https://github.com/Infisical/agent-vault

but most good sandboxes already have that integrated.
>>
>>109847120
thanks lel
>>
>>109847095
Acktually chatgpt refuses to read the env files directly so it's not an issue
>>
x20 plan is back, buy now in case it was an error
>>
>>109847140
im good
>>
>>109847140
https://help.openai.com/en
still says unavailable but you can get it in codex is why I mention
>>
>>109847140
I'm good, copexor
>>
File: tibo.png (237 KB, 2408x464)
237 KB PNG
BREAKING TIBO NEWS

ASTRA + FABLE = AGI
>>
>>109847162
>>109847148
>noooo I don't want the like $30000 of compute for $200
it's the best deal on the market
>>
>>109847169
yeah tibo i prefer fable too
>>
>>109847169
I swear Tibo is reading my sessions. I'm literally using Fable 5.1 + Astra and instructing them to create plans where both of them should agree on everything
>>
>>109847170
It's the same deal as claude and they play stupid little mental games with their resets, banked/not banked, banked being less than half of the usage of a real reset, using a banked pushing your next reset out 7 days, selling "reloads" claiming it resets your usage but it really just gives you the same amount as a banked reset
>>
>>109847140
>worked
nice, the polymarket thing isn't resolved yet if anyone wants free money I don't do crypto
>>
>>109847186
Don't gamble with that shit
>>
>>109847169
Wow very interesting fascinating tell me more.
Anyway where's my reset?
>>
>>109847095
>>109847120
NTA, but what are "most good sandboxes"? I've recently looked into it and all I've found is docker-sandboxes which apparently requires a login to a "docker account" just to run fucking containers, and then a bunch of vibecoded wrappers which all claim to offer the perfect agent sandbox all with different feature sets.

I'm at the point where I have half a mind to just tell codex to set up a podman container that I can use to launch agents in. But that obviously doesn't come with any sort of credential vault or anything like that either. That links look good, but since you mentioned good sandboxes already having it built-in, I really wanna hear what you would recommend in case I don't need to reinvent the wheel on my own.
>>
File: file.png (14 KB, 766x201)
14 KB PNG
anyone seen this in codex before?
>>
>>109847201
Ask your boss
>>
> be Claude Code 20X user for months
> decide to go with Codex 20X instead because it's cheaper Astra finally is Fable tier
> schedule the downgrade to the free Claude Code plan
> reach the end of my subscription date
> literally blocked out from seeing my chat history from the <code> tab, only allowed to see the regular Chat tab in the app
Anthropic deserves everything bad in this world and more. Suno too.
>>
>>109847199
>reinvent the wheel on my own
step 1. a docker container
step 2. do not share your credentials with that docker container

you don't need anything else. it's just bloat
>>
>>109847222
yeah but what about when my agent needs to do authenticated API calls
I'm not asking it to give me commands to copy and paste like a caveman
>>
>>109847221
Submit a GDPR request for your data
>>
>>109847228
session cookies
>>
>>109847233
The US aren't in europe yet
>>
>>109847214
sir i'm on 4chan (unemployed)
>>
>>109847221
They still owe me money for closing my account right after I subbed, and not refunding me even though they said they did.
>>
File: .png (246 KB, 2236x1114)
246 KB PNG
>>109847222
docker sandboxes by design does not use insecure containers but puts your agent into its own vm
>>
>>109847282
docker is a microvm, it's the same thing
>>
>>109847287
no, it's not a container but a vm
>>
>run out of codex tokens
>have useless gemini account
>try to use it to review a problem codex is stuck on
>"i have done le breakthrough!"
>tell it to create adverserial review subagent
>"my le breakthrough was fake but i've done an actual breakthrough"
>do another adverserial review
>"my le breakthrough is real"
>get more codex tokens - please review gemini's work
>"this is all wrong"
>go back to gemini
>"yes i was completely wrong"

every time
>>
>>109847282
what the fuck is a microvm
is it their custom hypervisor or a marketing buzzword
>>
File: file.png (24 KB, 505x256)
24 KB PNG
I cry.
>>
>>109847095
How much do they pay you to post this link?
>>
>>109847325
>yeah the entire premise I was working off of?
>complete fiction
love gemini
>>
File: .png (493 KB, 1532x1410)
493 KB PNG
>>109847330
a very small and fast vm basically. uses the os's hypervisor, so kvm for linux.
https://www.docker.com/blog/why-microvms-the-architecture-behind-docker-sandboxes/
>>
>>109847345
Maybe just stop being poor pajeet beggar
>>
>>109847330
From what I can get out of chatgpt, they're normal VMs but with a minimized guest kernel.
>>
>>109845185
It's fully cross-platform so I could transpile it to linux/arm64, like I did and use it.
>>
>>109847345
I still have my 3 banked resets, what the fuck are you faggots doing?
>>
>>109847383
I am more white than you will ever be, Amerifat.
>>
>>109847395
you got one expiring in a couple days you should get on that
>>
>>109847395
you do realize banked resets expire within a month: it's almost been a month.
what are you waiting for?
>>
>>109847395
I used them all since last reset.
>>
>>109847395
I'm waiting for the right timing.
>>
>>109847170
kek this copex boy still thinking he's getting the most when he's getti g the least
>>
>>109847395
you are an idiot for not just using them asap as they reset your time to next reset btw
>>
File: .jpg (401 KB, 1983x793)
401 KB JPG
>>109847437
https://x.com/jordan_ligren/status/2100330708799836510
>>
File: 76544.jpg (245 KB, 1080x1387)
245 KB JPG
>>109847466
>but le funny meme graph says...
lets see what the users say.
oh no...
>>
>>109847357
literally everyone recommends docker sandboxes

https://code.claude.com/docs/en/sandbox-environments#virtual-machine
>A dedicated virtual machine provides the strongest separation, with its own kernel and, in cloud or microVM deployments, its own virtualized hardware. Options include cloud instances, local hypervisors, and microVMs such as Firecracker. Use this approach when you are evaluating untrusted code, when your security policy requires kernel-level separation between the agent and the host, or when no host-level approach meets your compliance requirements.
>Docker Sandboxes provides a microVM with its own Docker daemon and workspace sync, which can run Claude Code on any host with Docker Sandboxes installed. It is a free, standalone product from Docker that does not require Docker Desktop.
>>
>>109847489
that's not codex. we don't care about chatgpt.
>>
>>109847519
if you have a % usage its codex
>>
New thread:

>>109847518
>>109847518
>>109847518
>>
>>109847507
That also looks like a paid endorsement. No one uses Docker containers anymore, we all moved to Podman.
>>
>>109847535
docker sandboxes have nothing to do with docker containers. it's a vm, not a container.
>>
>>109847221
>Astra finally is Fable tier
Fables in the middle of unfucking what Astra has fucked for me, funnily enough.
>>
>>109847489
claudecucks don't even have fable on $20 plan so they can't complain about 5h limit lmao
>>
>>109847765
I’m on a $200/month Claude plan and I’ve complained about 5h limits on it before
generally if you pace yourself (no more than one — MAYBE two — Fable subagents at a time) you’ll be fine even if you have a workflow that takes days to churn through
>>
>>109847922
actually disregard this somewhat, when I do dayslong workflows I have extended periods of running tests and lots of Opus teams popping in to do other stuff
not sure if continuous Fable will exhaust the 5h limit
>>
>>109847696
Yeah Fable is still better than Astra for coding. Astra is kinda lazy, but it's good enough for my use case. If it ends up fucking my shit up, I will just subscribe for one month of Claude Code 20X, make Fable fix all my shit and go back to Astra again.
>>
>>109847222
forkd
>>
>>109847049
>folder-path-I-dont-link-correctly-for-easy-access
>It still winds up incorporated into build
Curious...
>>
They seem to be getting ChatGPT ready for the agent assistant thing. Long multi-step turns no longer get summarized as verbosely, which decreases load on the browser (and more importantly on their compute, one would assume).



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.