A general for vibe coding, agentic engineering, coding agents, AI IDEs, browser builders, and shipping code with LLMs.You use Git, right, anon?## What “vibe coding” is, and how to do ithttps://simonwillison.net/2025/Mar/19/vibe-coding/https://simonwillison.net/2025/Mar/11/using-llms-for-code/## News- (2026-07-24) Claude Opus 5 out## Related generals>>>/g/lmg/>>>/bant/agdg/ — schizo-resistant temporary (?) hideout>>>/vg/agdg/----## Frontier models using fully-general tooling — start here if you have $20 or sohttps://claude.com/product/claude-codehttps://developers.openai.com/codex/cli## Near-frontier models for codehttps://x.ai/cli## Not worth it for code, but maybe good for interpreting images/videohttps://antigravity.google/product/antigravity-cli----## Prompting / context / skillshttps://arps18.github.io/posts/claude-code-mastery/https://simonwillison.net/guides/agentic-engineering-patterns/using-git-with-coding-agents/https://github.com/mattpocock/skills — /grilling is a favoritehttps://github.com/DietrichGebert/ponytail## Other editors / terminal agents / coding agentshttps://osaurus.ai/https://pi.dev/https://opencode.ai/https://cursor.com/docshttps://docs.windsurf.com/https://docs.cline.bot/https://docs.github.com/en/copilot/how-tos/use-copilot-agents/coding-agent## UI/Frontendhttps://www.figma.com/make/https://www.anthropic.com/news/claude-design-anthropic-labshttps://uiverse.io/https://ui-ux-pro-max-skill.nextlevelbuilder.io/https://stitch.withgoogle.com/## In-browser builders / hosted vibe toolshttps://bolt.new/https://replit.com/https://docs.github.com/en/copilot/tutorials/sparkhttps://v0.app/docs## Benchmarks / rankingshttps://www.tbench.ai/leaderboard/terminal-bench/2.0## What we’ve donehttps://vcg.gitgud.site## Previous thread>>109496612
>>109505147I know Safari does this kind of thing a fair bitdefinitely macOS’ PDF renderer does
>>109505154This general and anyone who participates in it should be banned and eradicated from existence due to degrading collective human intelligence
>it asks me to check the code because it lost part of the code in a stupid non-committed move and might've reconstructed it poorlyRespectfully, you've robbed me of every opportunity to ever touch the code and even did the documentation I told you not to do so I could grow alongside. There is no way for me to help you now.
>>109505188intelligence was dropping before aithis is just a bit of acceleration
>>109505057Just checked that guys xitter and he seems insanely based>My guess: the big differentiator in modern software development will not be "taste", but sustained attention.>Most software built quickly with AI will follow an "easy come, easy go" pattern and be abandoned shortly after announcement.There's a load of dead vibe software and more forks than ever on github, many of which with incredible potential. It's very rare to find vibe'd one man repo's that are maintained 3-4 months out after launch, but then again, that's always been a problem and the differentiator between software that's adopted vs abandoned.
>>109501684uhhh yes retard since they originally released it with an explicitly attached "Preview" in their names. 0731 is just as good of a label as GA and allows further tweaking under v4 if needed
>>109505192https://www.youtube.com/watch?v=c72d4-LpilM
You HAVE build your own harness... right, anon?
>>109505222nopealso no gta cloneI avoided all the "baby's first vibe project" traps
>>109505222No, why bother?
a harness can't plug in usb-c for testing on device.
>>109505247Why bother vibing anything? So you can have whatever features you want in it.
>>109505295What I have works for me. I only bother with fun things because fun things are fun.
>>109505222several
>>109505222Claude, build the harness.Claude, reply to this post.
>>109505325Claude suck my balls.>"I'd gently push back..."
>>109505188>degrading collective human intelligenceIt was degrading anyway with or without AI. The future is human intelligence embedded with AI
>vibecoders need to be told to use giti'm a self-taught monkey and i thought i sucked, but wow, it really could be worse
>>109505303Enjoyment is apart of it too, obviously. What sorts of fun things have you bothered with, then? >>109505322Are you still using one? What does it look like?
>>109505335Until you run into a bug the a doesn't spot immediately and you can't fix it yourself because you can't program worth shit so you just have to keep begging the ai please figure it out
>>109505351if you’re new to programming, you’re not always gonna know what Git is
>>109505374or you can wait five months for a better model to come out and see if it can fix the bugassuming some other model can’t fix it already
>>109505374>Until you run into a bug the a doesn't spot immediately and you can't fix it yourselfRight, so this never happens to "real" software engineers, it's not like every company has bugs that sit for months if not years because nobody can figure it out
>>109505352I have tooling that simulates the exact legal approval/denial process for registering a new dog in my local municipality including the outbound message serialization, I have a records-management system (to be integrated with a DMS but it ended up being too much overhead for me personally, I'll have to re-derive something from ISO 15489 that actually fits a one-person scenario) and working on a nanokernel inspired by Zircon for the past few weeks.
>codex reset>send a few commands>'worked' for 1 hour total>22% weekly usage remainingdas crazyyy
Embarrassed about my gooner stuff in ComfyUI but want to use LLM assisted promptingJust had Sonnet code up a bash script so I can toggle gooner shit out and back in
.........~@^^
>>109505548not on my watch. nigger
Just spent the last few hours writing an esp32 project to control a mini split. I'm honestly impressed how well that worked. Ran out of use at the very end but I'll finish it tomorrow
>>109505659>an esp32 project to control a mini splitehm.... what?
>>109505659sickI was just watching a video of Sol's ESP32 based virtual pet gameHad Opus 5 do most of the ESP32 + peripheral emulation, then to burn tokens before the reset last night I rushed Sol to make their own game firmware for it, was honestly heart warming
>>109505419Tibo doing another performative reset tomorrow so it's all good
>>109505609.....~@^^.........~@^^.............~@^^
>>109505757we're gonna need more french people
! ^^@~......
another account to blast through and hope he really does reset again on monday
>>109505848>dealer sells you short but promises the reup coming on Monday
>124x render speedup on 4knot a bad spot to go to sleep i reckon
A lot of work has been accomplished. I feel like I'm getting way more juice out of Opus 5 when I'm just letting it loose.
how's the cope holding up, copex turds?braindead marketing slaves haha
>>109504913Sounds like you're still talking about user experience environment stuff>>109505023lel yeah
>claude, what else would you add? any feature suggestions?
>>109506042claudesister... GodPT already has more usage than your without reset
>>109505906letting these things run long is a great feelingI ran out of Fable earlier today, but I got a high score today asking Opus>A long while back I asked "could the CPU-limited checking be sped up by rewriting the Python in a more performant language?" and I got a bunch of suggestions of way more effective things than "rewrite the whole thing in Rust or Go". However, now, we're doing a lot more work to avoid doing work for all the `make`-based checks that need to happen, and I'd like to revisit that decision. Use a workflow to figure out if Python is now something that's making us slow. I have both Go and Rust toolchains on this computer. The deliverable I want is an HTML page helping me figure out if a rewrite in a more performant language is going to help this, and by how much. This is going to be super duper involved, but I want you to be very thorough.and it used two different workflows and I got a really handy webpage showing me what the easy wins were (not RIIR) and a rough guess of how fast it’d go if I were to RIIR or RIIGand its recommendations were kind of slop-py but I got some good solid “fix these first before considering RIIR” suggestions because a Rust port makes the b in O(b^n) smaller but what I really need is to make the n smaller for obvious reasons, and chipping away at the n is harder
>>109506089>keep b<1 >as you increase n, b now grows smalleryou're welcome
>>109506107
3AM and having Sonnet prooompt MiniMax H3, this is exciting!!!11111
>>109505154The OP picture is dream of every faggot here but none of them are like this btw.
>>109506147One draws inspiration from all sorts of sourcespoast fizeek
>>109506160>poast fizeekNah. Not yet at %10. Got %8 more. Lats and shoulders are completed though.
>>109506172gratz
>>109506173Thanks man. Have a good day
You really can do anything you want to the computer.
>clankers’ preferred slur for humans is “primate”I only said it once as a joke and Sol was WAY too content with saying it repeatedly and with weird emphasis
Is there anyone who successfully got the Pi Coding Agent to work (basic prompts, skills, plugins, and a bunch of inference APIs) on Windows 10's default terminal and PowerShell?
WSL
Currently working on this.
We had a good little run there, lads.>>109506780Read: >>109506787unless you are forbidden from using it, using powershell over WSL is pure masochism
>>109505536>systemd/runitI don't really like being even associated with systemd, but it's understandable given that systemd and runit are both init systems (and runit is ambiguous, runit could be the inti system, runit could be the online identity(s), the new zealand sport, the list goes on... runit could be many things)
>>109506854I'm on an 16GB Dell Inspiron laptop from the late 2010's.What about you test it on Powershell and Windows Terminal and post the results?
>>109506780the only experience i had with terminal harnesses in windows was with claude code and it was bretty bad compared to github copilot chat in vscode (windows) at the time.
>>109506911youve got two people telling you to use wsl and not powershell, why is your response for other people to use powershell? quit being dumb
>>109506963>running Pi on a very resource-constrained nd Windows 10 machine is dumbAre there any options left other than forking Pi to run it on Windows without WSL?
>>109506987install ubuntu
https://youtu.be/09UELaUhPEwtl:dw cuz jewtoober stalling:>chinks are pooling SOTA models usage/tokens from cracked/hacked accounts and API keys through their own OpenAI/Anthropic compatible backend>resell it to retards on taobao>actually more expensive than getting it directly from the AI lab and also they might reroute your fable api calls to deepseekv4 or some shit. Besides, can prompt inject you and hack you this way.
>>109506997>pool tokens >mass farm the outputs>distill the models
Thankfully sending demand letters to all 131 companies can now be easily automated. Still working on reducing my usage of AI overall. A nice clean fix. I will probably release my clone of google which uses yandex to get results and brave to give AI overviews soon, what do you think, /g/?
Thoughts on Spettro, a terminal coding assistant built in Go that claims to have native PowerShell support?https://spettro.eyed.to/https://github.com/aploide/spettro>no YouTube videos to be found so far
>>109505154with $0 in my pocket, how much vibeslopping can I reasonably do? to be a bit more specific, but still vague, I'd like to clone a C/Java project locally, and then ask an LLM to implement a single specific feature I have in mind. it's an Android + native app/library combo, and the feature would touch more of the backend rather than the user-facing Android GUI app. so maybe android studio, but idk how strong the vendor lock is with that one towards google's slop ecosystem.
>>109507451Buy an ad
>>109507475if you have the git you clone for the app, $20 on either codex or claude code would get you there. for $0 you would probably need to do something like copy and paste into the web chat and have it handhold you through finding what to change and how to change it.
>>109507451Use WSL faggot
Opus 4.8 is still my beloved.
>>109507475You're in luck, Opencode is being unusually generous by giving away DS4 Flash for free which is about GPT 5.4 level >>109507532Nah
>>109507475You can find v4-flash for free in some places. Freebuff.com does in-harness text ads. Nous portal and open code have free options. Xiaomi mimo code is also completely free.
vibe-coding from scratch doesnt appeal to me (i have absolutely no ideas, and my computer does almost everything i want it to), but is there any way to have ai automagically reverse engineer things? clean room and/or decompilation. old games, windows programs, anything. unlike programming i dunno ANYTHING about reverse engineering.
>>109507652No. It's either a lot of effort or a lot of tokens (as in billions of tokens).
>>109507554>muh WSL is all you need to run any harness on WindowsGive examples disproving this
>>109507652if you want to know how something works under the hood it can get you there if you point it at the file where the code lives & you give it an adequate description of what you want to find.full decomps can be done but >>109507676 is correct
>>109507652>>109507652The latest Chinese models (DS v4-flash, glm 5.2, kimi k3) will happily do whatever you ask them to do if you have them in a proper harness, including REing stuff. Keyloggers, finding exploits, whatever you can think of, they will do it if you have the tokens.
There's literally nothing new to make if you're not an autist making his own OS for fun
>>109507845>Anime troons having kids
>>109507870I wish that were the case, but I'm apparently one of two retards in the world to have some interest in a spreadsheet game engine, so I have to do it myself (ie have my bro deepseek do it for me). Thank god for vibecoding.
Tibo better reset tomorrow. I used 47% of my 20x and 70% of my 5x today
>>109507980Same, the usage is fucking ridiculous
>>109507898Believe it or not but women prefer a capable social man with hobbies over a 4chud
>>109507980Before anyone asks>GEMM library>started planning Conv library>expanded on GEMM shape master list, now I have list of shapes extracted from over 500 model architectures (even more different sized checkpoints) covering GEMM, Conv, Attention, Normalization, Embeddings/LM heads, MoE, Tensor transforms/resampling, Pooling, Softmax/reductions>Modding docs for Cyberpunk quest authoring, different than my quest making system, the docs are like from first principles>>109508000Nah I'd say it's reasonable, I got a lot done with it
>>109507845i dunno what a harness is, and i've mostly used this stuff locally for automation (feeding models images, text, and having them identify things). i dont have any tokens, and have absolutely no desire to run anything off site. i'm assuming this means im out of luck, huh?
>>109507870I can guarantee you’re wrong, I just stumbled into something new, today, that Sol hadn’t heard of and is very supportive of. Then again, everything I make is within the singularity; tools for agents, so my agents can make better tools for agents, so I can interact better with my agents so that we can make more tools for agents, etc.I did also make the cube, so I can say I haven’t fully lost it.
>>109505188Really starting to get sick of these Luddites shitting up the place all the time.
>>109508025There's no such thing as a tool for agents, humans are "agents" and we already made plenty of tools for ourselves
>>109508019NTA, watch for the Qwen3.8-27B release soon.And run using https://unsloth.ai/ once there's support.
>>109508061Is that what they call small?
>>109508061if i can get CUDA working again (stupid shitty nvidia drivers always fucking up) then i will check this out. thanks! kind of off topic, but maybe people know the answer to this, are there any dedicated boxes i can get for AI that wont shove accounts or internet connectivity up my ass? i really just want to run some ai-dedicated machine on a seperate network, getting VERY tired of dealing with nvidia's drivers, and i only have about 8gb of vram anyways, so i was only running extremely small models. i saw that some companies were releasing developer boxes just for AI, any of those a good purchase?
>>109508056I have made a tool that is quite literally for agents, the agents love to use it and they use it in ways that surprise me. Humans are agents but it’s as simple a difference as “I can’t natively ‘jq’ some JSON blob and parse and explore it in my brain at mach 10”.I’ll even tell you exactly what the new thing will be, it’s not really a crazy idea.>my prompts keep accruing “if statements” or “do this, then” statements that aren’t strictly required for coding>for example, “if you make a significant change in the UI, make a little homework assignment so I can say I like it”>if you accrue too many of these “contracts”, you’re overloading your coding agent with bullshit work and it might forget a few things>you could have a review agent review the coding agents with to split that duty>but… what if you could take it further?>instead of a capable review agent handling all the conditional bullshit in one go by reviewing everything the coding agent did, what if it only reviewed piece by piece and only accrued context when there was something interesting?>now what if instead of a capable agent, you had simple agents only looking for one specific thing each?>and then what if you moved the review to happen not after, but during the coding agent’s work?>all these little conditionals can be thrown into a queue for a human to handle and review or even automate>the coding agent is extremely unburdened by all the bullshit work and you can have prompts autogenerated and dispatched to dumb agents to do bullshit work>you can’t use tools to get around this because then your coding agent is just accruing tools; the same problemIt’s the mechanism that triggers authorization requests or fallbacks or guardrails, but “positive and constructive” and I envision it as “a swarm of little shadow agents listening to a trunk agent”.
>>109508090Just pay for a subscription
>>109508019You're not interested in trying to learn about new things? The lack of intellectual curiosity on this board is really bizarre.
>>109508074how much ram do you have? if you have at least 32gb you can run a quant of 35B with CPU offload
>>109507585Claude can work even after achieve 100% usage
>>109508129I can't afford, I got replaced by AI (I previously did graphic design/programming) and now my new shitty datacenter job doesn't pay me anything. To be entirely honest I am very lost right now and putting lots of money towards savings.>>109508138I'm very curious, I am just indicating I don't know what a harness is. For reference my message was>i dunno what a harness isAnd>i've mostly used this stuff locally for automationTrying to indicate that I haven't used a harness, and wouldn't have come into that area yet. If you know what a harness is, I would love to know!
I wonder if I can combine speculative decoding with CPU offload to run K3 at a few tk/s without fitting it on VRAM.
>>109507585>terragoated
>>109508335You don't have $20 but you are considering purchasing hardware, got it.
>>109508418I don't want to rent it, point being. If I spend the money, I am going to buy outright. I have seen the 300$ a month bills my friends rack up and I am not going to be part of that, and it would make it impossible to save properly. I would rather save up for a handful of months and just get a box, rather then spending rediculous prices renting.
>>109508335A harness is what calls the model, like the Claude Code or Codex CLI apps, or something like Pi.As a general tip it's usually good to ask the AIs those questions, it's much more efficient.
>>109508442this nigga gonna save for a couple months and get himself a 16xB300 lmao
>>109508442You need the inference now though. What's the point of affording a good pc in the future to run AI when you need AI for your job right now
>>109508125That looks interesting but wouldn't hooks just work?>if the UI files are changed, inject a reminder after the tool resultSome advanced parsing could be almost as powerful as an agent, but false-positives aren't a big deal in a reminder anyway
literally freewhittu piggu could never
>>109508442Right, you'll just pay $300 a month in electricity and thousands for hardware to run a model with 1% of the capability that you'd get from a frontier model on subscription
glm 5.3 waiting room
>>109508576If you care about the environment, then you are against private ownership of expensive, powerful GPUs, and are for distributed data centers. This is where the state ownership of data centers might come into play.
>>109508442You're missing some basic investing concepts here. $20 - $100/mo on a frontier-level service is going to be a better use of your money than trying to drop cash for the hardware to achieve a poor imitation of the real thing. Local AI is great and everyone should play with it, I run Gemma E4B on my phone, my home server is running LFM2.5 2.6B right now CPU-only and deliver 25tok/s. Investing in local AI hardware is a massive waste of money currently, it is completely senseless. If you want to buy frontier level capability, start by pricing out 400A service to your home, and verify the slab thickness for the contractor that'll be installing the rack.
>>109508479>16xB300Even if he had the millions to buy that it would still be useless, batch 1/single user throughput sucks on all hardware, you only get the value out of the hardware with multi user aggregate throughput>>109508596Shut the fuck up and kill yourself
>>109508605The Gen Z middle class as far as I have known them does not understand basic economic concepts, as much as they like to laugh at things like socioeconomic factors, which are a real thing, they do not understand economics at all, again, every single middle class generation z person that I have known, not a single one understood economics. The only ones that understood them were the proletariat.
>>109508560At least with Codex, hooks are something you’d use to make the idea I gave real, but they aren’t the idea for two reasons;>if the main agent has to determine which hooks to activate, then hooks are again the same problem in a different disguise, just like tools or conditionals in a prompt>if the main agent isn’t determining which hooks to activate (read as the equivalent of “which contracts to signal/recommend/add to queue”), then who is doing that work and how are they doing it?I say if the former is true, you wouldn’t want to use hooks, you’d be burdening your main agent again, if the latter is true, then hooks can be used to make my idea apply.I have an even simpler version of the problem which seems really stupid at first glance:>have some dumb JSON state>write a state transition function in conditionals via prompt>at some number of conditionals, the agent is going to start to crack, be it 10, 100, 1000 if statements>can you improve the accuracy of the state transition function by splitting the prompt?If yes, then it should be the same problem, except not simplifiable to a deterministic state machine transition because coding agent actions are nondeterministic.To your example: “what constitutes significant UI changes that a human would notice?”. You can touch all sorts of shit in source code but that’s going to be extremely difficult to parse with a symbolic program, but Luna Low might be able to reduce that to “yes/no” (the optimistic hypothesis; model and effort may vary).
>>109508544I am not really interested in those jobs anymore. Programming and doing art with ai is a different job that I did not go to school for, with a different appeal and workflow. I am not very good at "prompting", and while I am absolutely trying to improve at such things I doubt it will get to a point where I can do it professionally.>>109508576So there is no real solution?>>109508605>Investing in local AI hardware is a massive waste of money currentlyWill this improve in the future? Or will everyone be forever relying on paying for someone else. Even if I did go about renting, a lot of these models that can be rented are already awful whenever I have tried them on friend's computers, I almost always run into them refusing requests, and despite what people tell me about "jailbreaks", I cannot imagine paying for something I am intent on breaking, especially since it probably violates some agreement I sign when I start using the service. The small ones that I download and use seem up for anything, but when I tried Claude on a friends computer it stopped working with me almost immediately. And grok on another friends computer was no different, despite him telling me it would not complain.
I don’t know how any of you get anything useful from an agent. I have to babysit these things or they’ll make an error that they don’t recognize and it propagates down the chain until the output is useless. Hold their hand, and results are often great. But, God help you if you let them just code for an hour.
>>109508659What is your setup
>>109508666"The devil is in the details"The details here are important. After all, we don't want our planet to be uninhabitable. That much we should all agree on. I think it's an important thing. Important enough that I seriously take into consideration the factors of the technology. If you call that the devil, you're definitely insane.
>>109508666I don’t use agents, I don’t have enough money for that. A $20 clause sub could probably run for 30 minutes autonomously. I just base this off my experience with normal clause code, even the best prompt can end up with critical errors in the output.
>>109508647You have repeatedly been told the solution, pay for a subscription, or just give up because you are clearly not the sharpest tool in the box, a few fries short of a Happy Meal, a few parameters short of a foundation model
Now that Luna will get open sourced, you will invest in a powerful local setup for 100% free unlimited Codex?
>>109508659think about all the ways you're constantly handholding and babysitting your agents, then turn that shit into instructions, skills, and workflows you can reuse over and over until you no longer have to babysit them
Undoing the AI's 15 minutes of work because I made an embarrassing typo in the prompt.
>>109508659>I don’t know how any of you get anything useful from an agent.>Hold their hand, and results are often great.Sounds like you already figured it out.I’ll give you a hint: make the hand holding process easier.
>>109508705zased
>>109508684You don't have any idea what the current models are capable of if you have never used an agent. You can look up videos on https://inv.nadeko.net/feed/popular to see what people have made using AI agents. You can make amazing things, even with the free models. But I would need to know more details about your setup to actually tell you how to improve this.>$20 clause sub could probably run for 30 minutes autonomously.It doesn't need to run autonomously. You can still build amazing things with the Claude Pro $20 subscription. You also can build amazing things without it. I need more details about your setup, in order to help you though.
>>109508700I use extensive project instructions, planning, skill docs etc. LLM’s just can’t be trusted to operate for very long without having their work checked because they make so many errors. The code is either shockingly good or shockingly bad (like worse than I see from high school students bad).
>>109508684Agent != autonomous
glm 5.3 milking room
>>109508659If a kernel developer with 20 years of experience can pick it up without ever having used AI before and 10x (by his own account) his work output within a single month and you can't make them work then you're doing something wrong. And the funny thing is at the beginning he was like "Yeah I hate AI and I think it only hallucinates bullshit but my friends keep talking about how great it is so I'll drop $20 on a sub for a single month just to say I've tried it and confirmed it's shit" and then it becomes his entire work day lmao.Normalfags have no idea what agentic AI actually is and how it'll leave everyone jobless. All their opinions on AI are from 2 year old stuff and using ChatGPT for cooking recipes once a month.https://www.youtube.com/watch?v=d7GedQuOxlo
>>109508647>Will this improve in the future?In some ways, yes absolutely. Hardware is in the shitter right now but that won't remain forever. RAM is going to the highest bidder and that's not us, really fucks with things and fucks with prices, but they have every incentive to make more and sell more, meanwhile the datacenter boom is already slowing down. So they'll catch up, prices might not return to where they were but they'll be a lot better than they are now, and a whole lot of computer hardware will be priced more sensibly. Buying right now is buying near the peak, not to say things can't get worse before they get better, but they have a lot of room to come back down significantly and they will eventually. Makes it a bad time to buy hardware in general. Beyond that, the capabilities will improve, but the disparity will widen, too. The models you can run with a pitiful 8GB GPU and 32GB of system RAM at reasonable speeds are staggeringly good, now, far better than many thought would be possible only a couple of years ago. So the models are getting better, small models are getting more capable and useful, and the hardware is becoming more accessible. Spending the money to run Kimi K3 or DSV4 Flash 0731 at home at useful speeds is a huge waste of money, for now, but you could've said the same thing a year and a half ago about running something as capable as Gemma 4, which my phone runs comfortably now. Still, won't catch up with the frontier obviously, even if you have fable-at-home running on an RTX 3060, that just means the frontier will be in the fucking stratosphere with super-astro-fable-9000.
>>109508684what do you think "clause code" is you stupid faggot
>>109508716Have you considered having it check its own work?
gpt 6 luna waiting room