[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
▼ Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: 1770040199597033.jpg (72 KB, 1024x1024)
72 KB JPG
A general for vibe coding, agentic engineering, coding agents, AI IDEs, browser builders, and shipping code with LLMs.

You use Git — right, anon?

## What “vibe coding” is, and how to do it
https://simonwillison.net/2025/Mar/19/vibe-coding/
https://simonwillison.net/2025/Mar/11/using-llms-for-code/

## News (both past and future)
- 2026-09-29 — OpenAI releases GPT-6.1 Sol
- 2026-09-28 — Anthropic releases Sonnet 5.5
- 2026-09-22 — OpenAI releases GPT-6 Sol and Luna
- 2026-09-22 — Anthropic releases Opus 5.5
- 2026-09-22 — Anthropic increases subscription plans's 5-hour limits by 20%
- 2026-09-14 — Anthropic reduces subscription plans's weekly limits by 17%
- 2026-09-12 — Anthropic suggests to pace the frontier. OpenAI agrees in principle.
- 2026-09-10 — OpenAI pauses new sign-ups for their $200 subscription
- 2026-09-04 — OpenAI releases Astra
- 2026-09-01 — Anthropic releases Fable 5.1

## Related generals
>>>/g/lmg/

----

## Frontier models using fully-general tooling — start here if you have $20 or so
https://claude.com/product/claude-code — probably generally better currently
https://developers.openai.com/codex/cli

## Near-frontier models for code
https://x.ai/cli — no 5h limit for only $30/month

## Not worth it for code, but maybe good for interpreting images/video
https://antigravity.google/product/antigravity-cli

----

## Prompting
https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/overview
https://developers.openai.com/api/docs/guides/latest-model

## Skills
https://github.com/mattpocock/skills — /grill-with-docs is a favorite
https://github.com/Vuk97/forward-implementation-first — do less redundant bookkeeping

## Other editors / terminal agents / coding agents
https://osaurus.ai/
https://pi.dev/
https://opencode.ai/

## Is our AIs unlearning?
https://aistupidlevel.info/

## Will there be a codex reset?
https://codex-resets.com/

## What we’ve done
https://vcg.gitgud.site

## Previous thread
>>109935982
>>
File: 1782503482447673.png (2.06 MB, 1122x1402)
2.06 MB PNG
>OpenAI
>>
File: 1771657071142455.png (147 KB, 861x877)
147 KB PNG
>>109940099
>>109940129
>>109940150
>>109940151
>We've halved your allotted compute on your """20x""" plan so more important people can pay us more for the same product's outputs in less time!
i guess i'm moving to claude
>>
File: claude costanza.png (114 KB, 600x600)
114 KB PNG
>>109940179
>Anthropic isn't going to copy this at the first opportunity
>>
File: 1762039075013900.png (40 KB, 303x303)
40 KB PNG
>>109940197
well they haven't yet and opus 5.5 is sota so i'll enjoy it while it lasts i guess
>>
File: 1765022744875839.jpg (17 KB, 399x400)
17 KB JPG
>>109940179
Will do it too, 6 more days before the end of my $200 plan. Time to waste those banked reset.

When I moved to the $20 on chatgpt, they nerfed it one week after, then I moved to the $200 the next month, and then they nerfed it 3 weeks after. Enough is enough.
>>
>>109940223
inb4 you're the reason they keep nerfing plans and you bring it to claude too
>>
>OpenAI wants to be the new Microsoft Office
yawn
>>
File: 1760574694159914.jpg (235 KB, 1346x1000)
235 KB JPG
>>109940237
Well, if Claude start to nerf shit after next week then....
>>
>>109940179
subscriptions are only good for getting your feet wet. once you do you should spend time replacing subscription reasoning with your own model stack. once you do that you start actually learning how models work and why subscriptions are to be used maybe 20% of the token time. most serious teams shoot for 5% "frontier". so your $20 or $200 should 10x or 20x usage because you are learning to build independently.
>>
>shameless Apple mimicking
pathetic
>>
File: 1769971759571011.jpg (27 KB, 283x323)
27 KB JPG
>>109940257
okay well i'm still going to buy a claude 20x sub because i get more out of that
>>
File: 1769437063141357.png (60 KB, 882x278)
60 KB PNG
HOLY SHIT. THIS CHANGES EVERYTHING (nothing actually changes)
>>
File: 1778498182931122.jpg (7 KB, 250x217)
7 KB JPG
>OMG LE HECKIN' TIBO RESET
yeah my weekly allowance is still halved though
>>
File: file.png (10 KB, 1273x178)
10 KB PNG
Wow must be a lot of resets coming, it's taking awhile to load.
>>
well that was shit
>>
>>109940291
that's fine if this is all you know how to do and don't want to learn anything else. it's more than enough for a lot of people. I push billions of tokens a week so this kinda thing doesn't work for me.
>>
was that even 20 products or did i fall asleep and miss like 16 of them
>>
>>109940378
>6.1 Sol
>dots
>Jev with Vision
there was a 4th?
>>
File: 1790704784.jpg (19 KB, 400x400)
19 KB JPG
>>109940378
anon.... it was you all along
....you are the product
>>
Again the 100 math results were mentioned, but not released.
>>
I use planning mode a lot and generate a phase-oriented outline. That's where I've found value in Claude Code's quickness. I'm okay with the actual coding and implementation part being slow.
I have a couple of laptops with i5-8250U CPUs and around 8 to 16 GB of RAM. What options do I have for running something good locally?

Make no mistakes.
>>
>>109940428
even people with good specs don't have good choices for local models.
>>
>>109940428
>8 to 16 GB of RAM
>running something good locally
If those arent on some lite linux distro forget it. You can get summarizers and document sorters with the small gemma 4 models. but anything else will be super slow or not fit. maybe a small qwen for some coding but if its not the 27b 3.8 its not great either.
I guess depending on how many laptops you could set up a swarm of small models for fun? but you are going to get like 10-15 tk/s max on just ram.
>>
>>109940428
use claude to build automations designed to replace claude
>>
>>109940450
Damn, I’ll keep searching, but for now, $20 for Claude Code is a deal that pays for itself. I doubt it’s going to stay this cheap, though.

>>109940479
At the moment, I use Gemma-4-E4B-it-Q4_K_M and OpenHermes-2.5-Mistral-7B.Q5_K_M with llama.cpp. Integration is my biggest pain point right now. Those models are okay, but I don’t trust them 100%.
>>
My personal benchmark is for every new model release I try to get a model that manages to make a chrome extension which gets rid of pic related (and I don't mean the popup, this is trivial, I mean the 5-15 seconds of fake buffering that they are doing before the video starts)
and so far none of them have managed, including Fable and Astra on Ultra.
AGI has not been achieved
>>
>>109940501
>I’ll keep searching >>109940501
>>109940450
doesn't know how to use models or computing in general
>>
>>109940549
why would they fake the buffering?
unless your extension plugs into the nearest server, you're not gonna be able to fix that
>>
>>109940315
Maybe I can swap from sol medium to sol high
Wait a minute it's 1/5th the price of astra, not 1/5th the price of sol. I thought I was getting a 5x efficiency upgrade.
What a disappointment. Well fuck this I'm going to bed, not waiting around for that trash. Probably not even going to use it since it'll still blow through my usage.
>>
>>109940549
>fake buffering
kek brainworms
>>
>>109940577
anti-adblock measure
not him but I can't even play 720p videos anymore, they buffer every 2 minutes
>>
>>109940542
You can try the qwen 3.5 9b and bonsai 2 but they arent to be really trusted either. its okay if you let them reason but it will take forever at those speeds.
If you are willing to rig you can take a 1070ti which is like $90 and hook it up to run a qwen 35a3b which is a lot better but you will have to set up a egpu and the loading the model times will be horrible and i dont know if it will be actually fast as it loads and unloads experts could drag you to single digit tk/s or prefill.
>>
>>109940583
Oh it's the same api price as gpt 6 sol. I'll switch to it and stay on medium then and give it a shot.
>>
At work so not watching but was the alleged reset banked or normal reset, deciding if I need to use a banked to keep working as I was expecting a normal one today
>>
>>109940577
>>109940588
it's a literal thing. I have gigabit fiber. 4k video streams on netflix load instantly but youtube takes its sweet time buffering 20 seconds before loading a 720p video
see https://community.brave.app/t/how-to-deal-with-fake-buffering-on-youtube/654516
>>
>>109940626
banked
>>
>>109940628
>i have niggabitfiburr
>but i use brave
>why is it not working???
kek wormed
>>
>>109940628
>brave
ngmi
>>
>>109940633
can you even read you turbofaggot
i am using chrome, the brave forum link is just an explanation of the issue
>>
>>109940610
thanks dude I will explore the 1070ti option futher.
>>
>>109940628
mine does it for maybe 2 seconds
still don't see how another extension would fix this - if it's fixable. would probably have to fix ublock
>>
>>109940645
google fingerprinted you as a brave user, uninstall brave
>>
File: 1790573669256158.gif (801 KB, 264x264)
801 KB GIF
>>109940645
>chrome
>>
>>109940645
Doesn't change the fact you put in a brave link
>>
>>109940549
which tools and mcps are you giving the model to work with?
>>
Is there a way to update Claude's context without forcing a response?
For example I want to tell it I changed color from one to another and to remember that in the future.
>>
While local models are the endgame for this AI bubble, there’s no reason not to abuse subsidized tokens as much as possible in the meantime. When the bubble pops, hardware prices will come back down, and you can run the local larger models you want.
>>
File: file.png (5 KB, 162x83)
5 KB PNG
the only good thing that happened today
>>
>>109940702
sounds like you're looking for the pre-prompt? or just have it update memory
>>
>>109940702
auto-memory should do that all by itself
https://claude.dev/blog/the-new-rules-of-context-engineering-for-claude-5-generation-models/#then-memory-in-claudemd-files
>>
File: 1781102074166029.jpg (12 KB, 184x184)
12 KB JPG
>>109940549
There is a trick for that: load comments, if you scroll long enough for the comments and the right side bar, you get the video buffering less and almost insta load.

Try it you will see.
>>
File: shoulder-tap-son.gif (65 KB, 498x292)
65 KB GIF
>>109940718
>When the bubble pops
>>
>>109940718
>here’s no reason not to abuse subsidized tokens as much as possible in the meantime.
yeah using it to make your set ups better.
>the bubble pops.
If. we could still be on track for own nothing and renting compute they will try it.
>>
25% more usage for 250% the cost... I am genuinely speechless.
>>
>>109940749
did you miss the sol price-cuts? sol costs now half as much. opus-class model for the price of sonnet.
>>
>super long days at work this week
>don't have time to vibe until friday
>>
Newfag to vibecoding stuff.
I wanna restart my Claude instance and get back into working while not eating so many tokens. Should I tell it to write a design document based on what it did by now? Anything else worth throwing in?
>>
>>109940762
no, don't restart but keep using your warm cache
>>
>>109940775
cache stays hot across instances dingus
>>
apparently 6.1 sol is monster
>>
>>109940789
more like a gobblin'
>>
>>109940762
don't listen to dumb retard over here >>109940775
>Should I tell it to write a design document based on what it did by now?
yes. Or, if you're project isn't that big, just tell it to update memory. When you boot up the next one, tell it to check memory (the save it under your current user directory)

in my experience it's cheaper to kill them once their task is done to save on context than it is to care about cache (unless you're only using 10-20% of context)
>>
someone let me know if it's over or we are back
>>
>>109940799
>the save it
they save it*
>>
How good is 6.1 Sol?
>>
>>109940808
it so over AI winter is here. expect nothing but gimmicks to get more non coders to sub.
>>
use case for dots? Cheating in video games?
>>
what IDE can I use on a 4gb ram laptop without it crashing intellij-idea is too bloated and makes it froze
>>
>>109940829
>4gb ram
yeah can you even run cargo or cmake with that
>>
>>109940799
It's eating 10% of my usage from just sending a new message so I imagine it's time?
Or is it time for me to stop asking it within Claude website and install the Code malware?
>>
>>109940841
>It's eating 10% of my usage from just sending a new message so I imagine it's time?
it's fuckin time anon
>Or is it time for me to stop asking it within Claude website
it's fuckin time
>>
seems OpenAI is falling behind Anthropic, we need to stop the frontier
>>
>>109940789
It devoured sol 6 pretty fast.
>>
>>109940839
I never tried cmake I am not used to c only java
>>
>>109940819
it's not
>>
>>109940787
restart still leads to cache misses
https://code.claude.com/docs/en/prompt-caching#actions-that-keep-the-cache
>Actions that keep the cache
>These actions either append to the end of the conversation or don’t touch the request at all. Some of them, such as editing CLAUDE.md, keep the cache for the same reason the change doesn’t reach the running session until /clear, /compact, or a restart.
>Editing files in your repository
>Editing CLAUDE.md mid-session
>Changing permission mode
>Changing output style
>Invoking skills and commands
>Running /recap
>Rewinding the conversation
>Spawning a subagent
>>
>>109940857
i don't think you have enough ram to do your tests and builds, get a better rig thirdie
>>
>>109940864
>restart
keke is this a claude thing? i'm used to the way agy works and what i was informed by gemini
>>
er was no reset even announced? We got pointless cloud shit, some weird always on assistant (I don't want this) and my usage is going down by half on astra.
>>
File: 1785700880202539.jpg (123 KB, 903x1080)
123 KB JPG
>dot.com redirects to grok bot
>>
>>109940884
tibo said current $200chuds get extra token credits lol
>>
>>109940884
>>109940893
I have 1 banked reset come in
>>
>>109940893
i think ill let them keep that
>>
is astra worth it?
>>
can someone seriously give me a use case for dots? I can't work out why I would care, is it for middle managers or something
>>
File: .jpg (117 KB, 1461x1076)
117 KB JPG
>>109940799
don't manually manage your context. if you want to save money do as codex does and set a smaller context window

https://x.com/thsottiaux/status/2076543065045795309
>[...]
>The actual reason [for a lower default context] is the what you can see depicted in the chart below, which is the difference in the orange line and the blue line. It is caused by overall cost of cache reads going up with the size of the context being shuffled back and forth between toolcalls. The sweet spot in terms of cost is therefore not necessarily to use the maximum possible context length.
>[...]
>>
>>109940917
competing with muse spark. life organization. not sure if it does nsfw like muse. i'm happy with my spark, it's not a genius but it's a good worker bunny with access to github and a bunch of normalnog apps i don't use.
>>
NIJIKA IS CUTE!!!!
>>
>>109940917
no, I watched their ad for it and it's more "choosing where you can eat!!!!" stuff lol

I don't think they know what to use it for either
>>
>>109940944
>I don't think they know what to use it for either
because it's an afterthought competitive product. they couldn't let zuck capitalize on an empty niche.
>>
>>109940929
it's just called muse.
https://muse.ai/

muse-spark is the name of a model.
https://dev.meta.ai/models/muse-spark
>>
When I get home, gonna try Ponytail
>>
>>109940919
I never compact, I always kill them at 20-25% (unless the task takes them over, then after task is done), and always at full 1m window
some random xitterfag won't convince me of going against what 3-4 years experience with llm's taught me
>>
>>109940950
>it's just called 4chan
>/g/ is the name of a board
kek
>>109940955
>poonytail
>not in OP anymore
kek absolute shit
>>
>>109940956
>tibo
>head of codex
>random xitterfag
first day here?
>>
File: 1776393267055358.jpg (44 KB, 681x614)
44 KB JPG
>>109940917
meta muse competitor
but normgroids already get muse for free. i don't know who is paying for this and i don't know why they thought it was a reasonable tradeoff for mangling the $200 plan.
>>
Opus 5.5 spent his 5 hour usage searching every research paper on the web yet again.

Not that I can complain since I keep asking for extreme scientific accuracy so I will just have to keep going at it slowly I guess.
>>
>>109940970
>normgroids already get muse for free
it has a full ubuntu vm with 99gb free. for free. wtf. too bad it's too late to mine on that bitch.
>>
>>109940965
bad metaphor.
muse agent will soon be powered by watermelon, not muse-spark
>>
>>109940969
he is to me
I don't care about your gay resets
and believing anything they say is a sure sign of retardation
>>
File: file.png (42 KB, 490x550)
42 KB PNG
...hello?
>>
>>109940988
>obsessed with watermelon
what's that beeping noise? anyone know? when the model updates to muse friedchicken i will call it muse friedchicken
>>
>>109940990
tibo literally said in the same tweet

https://x.com/thsottiaux/status/2076543065045795309
>[...] The benefits of higher context lengths are mostly overall speed (as you don't wait for compaction), ability to deal with humongously large input and potentially cost if the system is well tuned and you hit your cache perfectly. [...]
>>
>>109940315
>(nothing actually changes)
the cache cost is actually huge
>>
>>109941003
your language lacks precisions

a model doesn't make a personal agent.
muse-spark is just a model and also used in meta.ai, muse code, ...
>>
>>109940965
everyone who hates Ponytail worships demons.
>>
>>109941004
tibo hits the reset button like a good monkey. don't take his tweets too seriously.
>>
>>109941004
so fucking what? go suck his cock
>the benefits of higher context lengths are mostly overall speed
no it's fucking not
it's that the closer it gets to filling up it's window, the more retarded and unpredictable it gets
>>
>>109941016
>your language lacks precisions
kek maybe it does, that's why i vibecooode. the agenda is to fully merge the ecosystem and put the personal agent in the harness eventually, it seems.
>>
>>109941001
i don't even bother with openai releases in the first 2 days
anthropic manages to get their new models out on day 1
openai always takes at least a day for me to drop them in codex
>>
>>109941030
>go suck his cock
he's kinda cute desu
and he has a LOT of tokens...
>>
>>109941030
>it's that the closer it gets to filling up it's window, the more retarded and unpredictable it gets
this is not true anymore for modern models. neither anthropic nor openai have that problem.
>>
>>109941062
not my experience
it's not as bad as in the past but the less garbage in the context the better the outcomes tend to be
>>
>>109941062
it's might not be as bad as before because the overall baseline has gotten higher
I see no reason why it shouldn't still hold true
>>
>>109941001
The order is usually
>US Pro users
>EU/Asia Pro users
>rest of the world Pro users
>US Plus users
>EU/Asia Plus users
>rest of the world Plus users
>>
>pre-nerf opus 5.5 with good limits
>not-sol sol 6.1 with decent limits due to 50% cache cost drop
yea I'm thinking codeslop time, things probably will get worse pre-ipo rather than better
>>
>>109941022
works for glue code, but absolutely shit for my usecase, which is dsp code, topologies, and rust
>>
>pre-nerf op--
fuck off



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.