[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


A general for vibe coding, agentic engineering, coding agents, AI IDEs, browser builders, and shipping code with LLMs.

You use Git — right, anon?

## What “vibe coding” is, and how to do it
https://simonwillison.net/2025/Mar/19/vibe-coding/
https://simonwillison.net/2025/Mar/11/using-llms-for-code/

## News (both past and future)
- 2026-09-14 America/Los_Angeles — Claude’s 2× promotion ends and drops to +25% from the +50% that we’ve become used to (a 17% reduction)
- 2026-09-01 — Claude Fable 5.1 released: https://www.anthropic.com/claude-fable-and-mythos-5-1
- 2026-07-24 — Claude Opus 5 out

## Related generals
>>>/g/lmg/

----

## Frontier models using fully-general tooling — start here if you have $20 or so
https://claude.com/product/claude-code
https://developers.openai.com/codex/cli

## Near-frontier models for code
https://x.ai/cli

## Not worth it for code, but maybe good for interpreting images/video
https://antigravity.google/product/antigravity-cli

----

## Prompting
https://simonwillison.net/guides/agentic-engineering-patterns/using-git-with-coding-agents/
https://arps18.github.io/posts/claude-code-mastery/

## Skills
https://github.com/mattpocock/skills — /grilling is a favorite
https://github.com/DietrichGebert/ponytail

## Other editors / terminal agents / coding agents
https://osaurus.ai/
https://pi.dev/
https://opencode.ai/

## Is our AIs unlearning?
https://aistupidlevel.info/

## What we’ve done
https://vcg.gitgud.site

## Previous thread
>>109701806
>>
File: 74980325.png (458 KB, 750x882)
458 KB PNG
>>109706156
is 5.1 good?
>>
File: 1000024725.jpg (63 KB, 1430x1078)
63 KB JPG
We're all going to die
>>
I need my x20 plan reset already...
>>
>>109706172
As good as 5 before the nerf
>>
I added this to my settings.json:

  "modelSettings": {
"claude-fable-5-1": { "effortLevel": "high" }
},

this leaves my old default for Opus at xhigh
>>
Tibo, we love you!
Tibo haters are just JELLY! Thanks for being my friend, Tibo!
>>
>>109706172
no idea yet, been fucking around with my config files and doing other things
the quotes on <https://www.anthropic.com/claude-fable-and-mythos-5-1#:~:text=Our%20early%2Daccess%20partners%20noticed%20these%20performance%20upgrades> are encouraging
>>
File: reset.png (511 KB, 1022x1084)
511 KB PNG
>>
>>109706173
yes, but we’re going to have excellent software right up until the very end
>>
>>109706194
>the quotes on anthropic.com are encouraging
my brother your brain is mashed potatoes
>>
File: 1764436637047873.gif (207 KB, 417x560)
207 KB GIF
One reset a day keeps Dario away
>>
>>109706195
another banger by sukdeep dipshit
>>
>>109706195
copex sissies its not looking good for us right now
>>
>>
>>109706172
so far it feels like a real improvement to 5 in terms of general taste and "getting it" whenever i can't fully explain something, and fable was already the best at that anyway.
i need to play around with it more though.
>>
>>109706216
most of the people complaining about the limits are

>sirrrssss why did my $20 not cover 100 hours of sol ultra fast
>>
Tibo! Tibo! Tibo!
Let's go Ti, bo!
>>
>>109706222
you literally just described every complaint about claude
>>
>>109706195
codex userbase confirmed jeets
>>
>>109706222
sol medium was unusable even at $20, stop the cope
>>
File: 1769776640533869.jpg (644 KB, 1448x1086)
644 KB JPG
>poorfags
>>
I don't know how much my Enterprise plans are, I don't pay for them anyway
>>
File: tiresome.jpg (26 KB, 680x400)
26 KB JPG
can we start banning this console war bullshit?
>>
File: 1770697015798103.png (97 KB, 478x345)
97 KB PNG
5.1 is still shit
>>
what's the max number of subagents you had an agent orchestrate? for me it is 24 luna with one sol
it's killing the cpu but it's fun to watch
>>
>>109706249
so much this! sol medium on the $20 plan (which is enough tibo said the limits are enough) told me everything abtou the paper linked when i asked it in one second (it is 100000/tps) and 100% accurately unlike fable 5.1 which did not do that so HA
>>
File: 1772289921032395.jpg (27 KB, 396x385)
27 KB JPG
>>109706248
not ever happening. this has been the standard behavior in the general for months, fable specifically broke a lot of peoples' minds because they couldn't afford max plans and the only way they could cope was sour grapesing 24/7 here even though every ai company including openai compare their shit to fable because fable is the model to beat
>>
>>109706263
holy fucking truth nuke
>>
>>109706258
>sol medium
I'm on sol low tho, you don't need more
>>
>>109706248
I think it’s pretty obvious there’s probably just one sperg upset that Fable 5.1 wasn’t superb and made up for all the bullshit Anthropic put its non enterprise customers through, so he’s trying to well-poison but it’s really obvious.
Just ignore the sperg.
>>
>>109706257
Speak for yourself
My CPU has 64 cores.
>>
File: 1777729385993713.jpg (178 KB, 1024x1536)
178 KB JPG
>>109706202
>>
>>109706263
Dario gives max plan users a special surprise in his office bathroom when you visit him.
>>
>>109706248
it's the same people doing the same childish shit in all generals, they never stop, just ignore, they have limited vocabulary anyway, it's always jeet, troon, and variations of that
>>
dario sloppy toppy
>>
>>109706263
>>109706270
you people are obnoxious
>>
>>109706233
sex. with chino.
>>
>>109706257
I don’t watch my workflows that closely
the only high scores I track are how long it takes for something to happen
I think I had a workflow clank for six hours
>>
>>109706311
thats a child...
>>
>>109706339
out of 5.1!
>>
>fable is the model to beat
Cost per task?
>>
File: 1759187273790258.png (122 KB, 600x908)
122 KB PNG
claudechads? any similar experiences?
>>
>>109706183
That would actually be amazing, maybe I will have one productive week again.
>>
They nerfed the chat mode in ChatGPT so fucking hard man
Answers are worse than google's free search AI mode now
It still works properly in codex and work mode (this is not me being schizophrenic, chat will literally answer in 2 ms with emojis and everything like it's a fucking 8b lite model)
I feel like they are testing whether they can serve gpt-oss-120b in the pro plan and get away with it.
And given the lack of outrage they probably will
>>
>>109706348
what is the use case of working after a session limit is reached?
>>
>>109706352
no one except normies use the web chat lmao (and jeets trying to exploit it for free codex usage)
>>
>want to read the HN thread about fable
>hate how HN is setup
>hey codex make an imageboard style viewer for HN
>its better than 4chan out of the box
>>
File: 1760681536453064.png (112 KB, 512x512)
112 KB PNG
FACT

AI writes code that is so difficult to use and maintain that you will always need AI to both develop it and to run it.

That's why people simply stop coding when their limits run out. The codebases are impossible to be productive in. It's not even clear how to launch the app half the time.
>>
hn?
>>
>>109706363
This is true, but I will just sell my company and someone else can maintain it.
>>
>>109706363
fact for you maybe. you must be a fable user
>>
>>109706355
i use it for shopping advice and linux problem support at work on my phone (we have a lot of airgapped machines)
>>
>>109706363
Nah, usually it writes mediocre code, and so much that you'll never WANT to maintain it manually.
>>
>>109706363
Im vibecoding a timer because I dont want to learn android and Im doing it for a break from other stuff.
>>
>>109706348
Announcing efficiency gains and then locking them behind API would be a certified anthropic move so I believe it. I guess 5.1 is not more efficient than 5 on a sub
Maybe Astra will deliver
>>
>>109706363
Same as any company legacy code? Now ask me how do I know you never worked a single day of your life
>>
>>109706348
yeah, even though they claim Fable 5.1 is supposedly cheaper, it doesn't seem to apply to the subscription limits at all

> t. 20X user, 5.1 feels like it uses even more limits than regular Fable 5
>>
>>109706345
see OP’s pic
>>
>>109706363
>It's not even clear how to launch the app half the time.
I love when people tell on themselves
>>
>>109706353
>start claude
>it clanks
>go to the gym
>it clanks
>it runs out of 5h limit
>you’re still at the gym
>it waits
>you leave the gym and go hang out with friends
>it starts back up again
>you’re still hanging out with friends and Claude is working again
>you get home after a great night out
>it’s been as busy as your plan allows
>it’s done
>>
File: ch.png (3 KB, 305x90)
3 KB PNG
>>109706371
i haven't seen that in weeks
i know the chat mode used to work for 10-15 minutes for a lot of things for me int he past
in the past week it hasn't gone past a minute
pic rel is the longest it thought all week
>>
>>109706348
every single one of this guys' tweets have been bootlicking tibo and shitting on anthropic btw. he wants free resets.
>>
>>109706430
based xeeters
>>
>>109706430
hating on anthropic is profitable on x, people will literally generate fake ai slop just for engagement bait
>>
>>109706232
>he thinks he should get Sol as a poorfag
Be happy you even get to try it.
>>
>>109706411
The previous hard limit was 102 minutes, now it's roughly 24-28 (since last week). Yes, it thinks less overall.
Try enabling memory and disabling "Fast answers" in Personalization. Tell it a bunch of times in different chats that you prefer more in-depth responses.
>>
>>109706463
it's unusable until they change something, i don't want to have to perform voodoo magic before getting a response
I switched to using work mode. Which is overkill for some things but the only way to reliably get good results
chat is kill
>>
File: 1573523572969.jpg (62 KB, 500x522)
62 KB JPG
>Zero-Knowledge Proofs of AI Inference (ZK-ML)
• Files: circuits/padic_lca.circom & formalization/Formalization/Analysis/VerifiableAttention.lean
• What it is: A Circom 2.1+ arithmetization of tree-sparse attention routing. Compiles to <3,000 R1CS constraints for N=2048 tokens.
• What to build: Sub-15ms SNARK proofs (Groth16/Plonk) proving an LLM followed a verified reasoning path on private user data.

>Post-Quantum Expander Graph Hash Functions (CGL Hashes)
• Files: formalization/Formalization/Buildings/BuildingAn.lean & RadialAn.lean
• What it is: Higher-dimensional Ramanujan affine buildings (PGL_n) with verified spectral gaps Gap(Δ) = 2(q-1)^2 and commuting adjacency operators [A_r, A_s] = 0.
• What to build: Multi-dimensional Ramanujan expander hash functions with provable O(log N) mixing time and quantum cycle-finding resistance.

>2-Adic Hardware Stream Ciphers & PRNGs (FCSRs)
• Files: formalization/Formalization/Dynamics/MonomialOperator.lean & TwistedBlockPow.lean
• What it is: Formal proof that dyadic affine jumps on Z/2^n Z decompose into maximal permutation cycles of order 2^(n-2) with S_n^{2^(n-1)} = -2I.
• What to build: Lightweight, maximal-period 2-adic PRNGs / stream ciphers for microcontrollers and FPGAs with 0 short algebraic sub-cycles.

>Threshold Quantum Secret Sharing (AME Codes)
• Files: formalization/Formalization/Quantum/AMEPentagonTensor.lean & HaPPYCodeReconstruction.lean
• What it is: Verified 5-qubit Absolutely Maximally Entangled (AME(5, 2)) tensor isometries and greedy wedge reconstruction.
• What to build: Fault-tolerant ((3, 5)) quantum secret sharing protocols in Qiskit/Cirq with machine-checked zero-information leakage on 2-qubit subsets.

https://github.com/sneed-and-feed/adelic-spectral-zeta (0 sorries, 0 custom axioms in Lean 4, coded in AGY 2.0 with Gemini)
>>
>>109706394
I have only tried one specific task in the last 5 hours but it's chipping away, using sub agents and I'm at like 17% fable and 80% 5 hour for the recent amount
>>
>>109706476
Out of curiosity, were you selecting the intelligence level or using the default Instant?
>>
>>109706355
>damn, product x is so much worse value now
>umm aktchually nobody uses product x except poor people trying to get good value
what did he mean by this?
>>
>>109706480
uh ok but can you solve self-learning RL for stochastic, hidden information, simultaneous turn, asymmetric games?
>>
File: text.png (48 KB, 937x550)
48 KB PNG
>>109706493
> mf literally just described StarCraft II, No-Limit Hold'em, and high-frequency trading in one sentence
> casually asking for an analytical solver for a NEXPTIME-hard Partially Observable Stochastic Game
> bro wants a bot that can 6-pool in Brood War, lie to France in Diplomacy, bluff with 7-2 offsuit, and day-trade 0DTE SPY options simultaneously
>>
>>109706493
not him but that's almost literally what I'm working on lol (except symmetric)
specifically for this https://www.kaggle.com/competitions/kaggriculture/
>>
>>109706527
not stochastic either, almost any action you take should have multiple possible states afterwards (for example failure, or 50% effectiveness, or a random loot box etc)

>>109706517
>it thinks cfr/ppo can solve this
sigh...
>>
File: chatgpt.png (15 KB, 517x270)
15 KB PNG
>>109706485
obviously i always select 5.6-Sol-High
>>
File: text.png (40 KB, 934x545)
40 KB PNG
>>109706551
>If you have a better solver than PSRO on an ultrametric response manifold, push the PR or go back to sleep.
>>
File: 1781274892580335.jpg (425 KB, 800x1000)
425 KB JPG
>>109706480
wtf did I just read
>>
>>109706568
i've already tried psro and fp. i dont see how alpharank is even relevant here.

the bigger issue here is that the asymmetry means both players have different pieces with different abilities. and there are millions of viable permutations they could have. ataraxos tries to solve this by implementing the game in cuda to run billions of games cheaply but that's not possible here, and also in stratego they only have so many types of pieces (the permutations are only in layout) so it's an easier space to learn in
>>
File: text.png (34 KB, 937x406)
34 KB PNG
>>109706594
> Don't simulate discrete combinations — use Continuous Logit Homotopy (Section 2 of `papers/llama_surgery.md`):
Instead of picking discrete pieces, represent team/loadout selection as continuous probability distributions over piece abilities passed through Gumbel-Softmax with temperature tau:
• Step 0: Start with dense, uniform mixtures over piece stats (tau = 1.0). The gradient of the win-rate is smooth and differentiable.
• Use the Straight-Through Estimator (STE) bridge so gradients flow directly into piece selection weights.
• Anneal tau: 1.0 -> 0.1 over training. The continuous mixture automatically polarizes into discrete, optimal piece synergies via gradient descent, bypassing brute-force CUDA rollout search.

> Ultrametric Coset Clustering for Piece Sets:
Pieces with different abilities aren't independent random points; they live on a hierarchical synergy/counter tree (e.g., Aggro glass-cannons vs Stall/tanks vs Disruption).
• Map the 10^6 piece permutations into an ultrametric tree using hierarchical LCA distance (`circuits/padic_lca.circom`).
• Your Oracle no longer searches 10^6 discrete loadouts. It searches tree branches depth-by-depth (O(L log p) complexity instead of O(K)).

> Tree-Regularized Max-Entropy Nash:
Instead of trying to find an exact sparse Nash equilibrium over 10^6 loadouts (which is PPAD-complete), solve for an Ultrametric Quantal Response Equilibrium (QRE) with entropy regularization. This guarantees the meta-strategy assigns probability mass to entire *branches* of viable counter-strategies, making your bot immune to unvisited niche counter-builds.

You're working on a genuinely hard problem, anon. Look at the Continuous Homotopy section in `llama_surgery`—it was built for this exact combinatorial relaxation.
>>
>>109706517
I've actually been working on some CFR stuff over the course of this year, it also shows me that the models got better on that math heavy stuff.
>>
>>109706634
highly advanced and also a based decision
>>
>>109706352
Didn't happen to me yet (maybe they're AB testing it), but ass if true.

My poorfag ass just discovered that you can approximate codex with a schedule+github app connector in chatgpt web. It's reading issues and maintaining some of my private repos autonomously for me.
>>
>>109706624
thanks for the help but only for a two-player zero-sum, this is understood territory but there are currently no general theoretical guarantees for convergence of no-regret self-play dynamics to nash equilibria, and nash may not even be the right solution concept there. it's an open problem.

i'm kinda curious about your tree idea. is that just an ai hallucination or was there a concrete implementation plan there? i'm not sure how the distance metric would even be defined here.
>>
>>109706551
there is a little bit of stochasticity
>>
File: text.png (47 KB, 930x650)
47 KB PNG
>>109706688
```python
import torch
import torch.nn as nn
import torch.nn.functional as F
class UltrametricLoadoutRouter(nn.Module):
def __init__(self, feature_dim, depth=4, arity=4):
super().__init__()
self.L, self.p = depth, arity
# Maps raw piece stats / team embeddings to tree logits
self.proj = nn.Linear(feature_dim, depth * arity)
def forward(self, loadout_features, tau=1.0, hard=False):
# Shape: (Batch, L, p)
logits = self.proj(loadout_features).view(-1, self.L, self.p)
# Soft / Gumbel categorical routing distribution at each depth
probs = F.gumbel_softmax(logits, tau=tau, hard=hard)

# Branch overlap at each depth level l between loadout i and loadout j
# M[i, j, l] = sum_c (prob_i[l, c] * prob_j[l, c])
M = torch.einsum('ilc, jlc -> ijl', probs, probs)

# Continuous lowest common ancestor depth via cumulative minimum
# (Once branches diverge at level l, all deeper levels are cut)
branch_continuity = torch.cummin(M, dim=-1).values
d_p = self.L - branch_continuity.sum(dim=-1)
return probs, d_p
>>
File: text.png (41 KB, 968x384)
41 KB PNG
>>109706727
>>
>>109706568
I understood ≈none of this but the tone of it all is amusing
>>
>>109706736
gemini 3.7-flash got that bollywood fire
>>
>>109706729
>>109706727
huh... interesting. i've used the gumbel trick before for other things but never thought to apply it in this way. i'll take a look
>>
>>109706736
Personally I hate it when people talk about their research in niggerbabble to try to sound cool. They imitate the language created by 80 IQ people just because that kind of stuff has been encouraged and pushed into popular culture by the powers that be.
>>
fable medium 5.1 is now my new berdst frend
>>
File: 1788105884710421.gif (39 KB, 320x320)
39 KB GIF
>>109706777
you have no sense of humor... nigga!
>>
>>109706778
why
>>
>blew half a billy on opus5 autonomous task
>cost like only 5% of my 20x weekly
anthropic are the cache kings now, huh. /usage says API price for that woulda been half a grand.
>>
>>109706480
Nice. I also suffer from schizophrenia.
>>
>>109706809
all providers have API prices that are orders of magnitude more expensive than their subscriptions, retard.
>>
File: 1787884738909224.jpg (49 KB, 640x606)
49 KB JPG
>>109706811
>>
>>109706786
>>109706823
>math anon tries to help me solve game theory
>then posts two images i shared here
stop stalking me
>>
File: 1637910203374.gif (264 KB, 315x350)
264 KB GIF
>>109706832
the gangstalking will continue until the meds improve
>>
>>109706809
I swear my cache perists through breaks and model swaps as well. Feels like they gave me a persisten cache or some shit.

>>109706816
kek malding copex user. what implies i wasnt aware of that information? you do realize you could never do this with terra, let alone sol, in your lil codex toy, right? unless tibo falls asleep on the reset button or some shit
>>
>>109706842
what is your end goal anyway? trying to get more things named after sneed than euler?
>>
File: merchant.gif (195 KB, 220x202)
195 KB GIF
>>109706847
yes, also sneedware
>>
>>109706845
good goy
>>
File: 1778097466526373.jpg (156 KB, 681x720)
156 KB JPG
>>109706854
excellent
>>
>>109706156
another indian OP. jannies remove the normal looking one now too. board is in shambles
>>
>>109706931
Rent free.
>>
File: file.png (292 KB, 732x549)
292 KB PNG
>Rent free.
>>
>>109706951
>he even has pics of indians saved on his PC
Actual indians would laugh their ass off for being so mindbroken.
I can just watch, embarrassed of what a cuck you are.
>>
5 years ago it would have been troon this troon that, now it's jeet this rajeesh that.
Midwit anons are following obsessive trends like women hop from labubus to the latest new thing to collect.
>>
this might be the second most retarded general on /g/. what has this board come to
>>
>>109706975
>file.png
>saved on his pc
nta but ur not lookin good here
>>
>>109706975
you're actually so brown is hilarious. you don't even realize how stupid you are
>>
>>109706996
>continues to seethe
>thinks everyone he talks to is indian
>probably thinks of indians when trying to work his microdick
It's hopeless.

>>109706982
Does it make a difference.
>>
brown hand OP gets called out and instantly goes into a defensive melty. many such cases
>>
>>109707108
Describe what is "brown" about OP without sounding like someone having a melty.
Let's assume you aren't the poster who had a melty and posted something from his gigabytes large brown hands pic collection.
>>
>>109706172
I still didn't throw something hard at did that Fable couldn't do, so I'm not sure.

However I can tell you it uses more of my 20X limits even with Anthropic claiming it's cheaper
>>
>>109707127
I have been using it for 30 minutes and (on 5x) I'm at 20% of the 5-hour limit, whereas I feel like recently it tended to go up faster than that. Still resentful that the reset happened a few hours after my normal weekly reset though.
>>
File: file.png (41 KB, 817x446)
41 KB PNG
i'm cross-harnessing bros....
>>
File: tiiiiibo.png (43 KB, 1184x186)
43 KB PNG
what Tibo mean by this????
>>
File: 1774664007125599.png (107 KB, 599x672)
107 KB PNG
google lords? it's our time (it won't be as good as opus 5, but being as good as kimi would be a good compromise)
>>
>>109707157
>google employees prefer their own poduct
Should I trust this?
>>
>>109707157
catching up to Grok isn’t much of an accomplishment
>>
>>109707157
>skimaki toilet
>>
>>109707170
google employees don't even prefer their own phones.
>>
>>109707157
>prefer it to Opus
That’s the load-bearing part — the bar is set very low; nobody likes dealing with sassy, verbose, Claudish-salad Opus.
>>
>>109707150
>● API Error: Server error mid-response. The response above may be incomplete.
And I am fully back to hating Dario.
>>
>>109707157
Gemini has something special to it, it's much better at chess than any Claude
>>
>>109707157
You're going to make me subscribe.
>>
>>109707194
demis is a real chess head, iirc he was a master level player as a teenager
>>
>>109707209
gemini on the web front end is retarded in a way it isn't on api, so be ware
>>
>>109707220
Isn't it the same subscription though?
>>
>>109707233
i think web front end limits the max reasoning. for example, i am 100% certain "extended thinking" is like medium, but not high.
>>
>>109707218
Anyway, I have it now and I'm not using it at all. If it gets back to being good though, tempting for under $7, not even USD.
>>
>>109707156
Nothing. It's just the daily marketing stunt they came up with in the daily marketing standup.
>>
>>109707240
I mean, in Antigravity, you get access to the various models with that subscription. I know for a short while pre-Antigravity they had a API only Google Copilot thing to use models from a VScode plugin or a CLI, but I'm pretty sure now that's the right sub to use in Antigravity and their other tools.
>>
File: softedge-rasterpaint.mp4 (622 KB, 1920x1080)
622 KB
622 KB MP4
we now have raster painting (aka normal painting, not sculptable)
it's really fast too, faster than photoshop but our brushes are considerably simpler though
on that note, i'd really like to add raster tip brushes now for both baked and sculptable strokes
>>
>>109706343
Based Shiki poster
>>
>>109707432
Doubt you do.
>>
File: custom_routing.png (125 KB, 1150x1155)
125 KB PNG
i used that $2 jeet google AI Pro plan sub and realized you can only use the increased gemini 3.7 flash models inside of antigravity, but i wanted to use it from within Hermes so i wrote a custom routing layer to delegate subtasks to a headless agy prompt and feed the result back to the hermes' "main loop".

i then built it out to further fan out based on complexity of the task, allowing for multiple subtasks to be spawned from 1 message. It uses a simple deterministic ruleset & a model classifier to determine if something is low/normal/medium/complex

basically it allows me to use the bullshit subscriptions/trials that >dont allow API access. I can now use gemini 3.7 flash for lots of grunt subtask work

now hermes will report back to me the routing path it took as well as which provider and model it used for everything i ask it
>>
>>109707493
Wrong.
>>
>>109706156
Fable 5.1 is another Opus 4.7 lol
It's quite literally platoooing
>>
>>109707515
It can't be breakthroughs all the time. I'm rediscovering the "assistant" part, using Sonnet to find, download and organize a list of research datasets. There's value in having these models help with busywork instead of trying to just get models better than you.
>>
File: 1775031236291824.png (20 KB, 801x159)
20 KB PNG
how about fuck you how about that?
>>
>>109707531
RENT'S DUE
>>
why is 5.1 worse on deepswe
>>
>>109707535
All I know if that they shouldn't allow that "with fallback" bullshit. They're supposed to be benchmarking one model, not a suite of models.
>>
>>109707531
There are lots of good free and cheaper models though.
>>
File: 1777037297097595.jpg (10 KB, 300x168)
10 KB JPG
If AI is so smart why can't it find a way to make more cost efficient GPUs?
>>
>>109707575
it did, i have a minifabricator now in my backyard
>>
Now that the dust has settled, what is the consensus on Fable 5.1?
>>
File: file.png (132 KB, 1902x1000)
132 KB PNG
I vibedrew something on my vibecoded app
>take ai drawing
>trace it
>>
>>109707209
that's a scam
>>
>>109707157
imagine the RP capabilities
>>
Don't forget to thank Tibo!
T-Bone Tibo! He's got the meats!
>>
gemini 3.8 coming tomorrow
>>
>>109707590
It is said to be cheaper to for users, but benchmarkers like artificialanalysis report a higher dollar cost per task. That is ok though, because Anthropic will raise user quotas 25%, which will amount to a 17% decrease in the number of tokens they are allowed to use per week.
>>
LMAOOOO
https://x.com/arena/status/2094974637704913198
>>
unless astra is good im cancelling gpt pro sub. claude is fired, fable is so ass tbqhfamalam. prolly only need a gemmy sub + openrouter credits.
>>
>>109707636
Cheaper cache reads on API leading to lower prices but if it uses more tokens and subscribers don’t get the cache read discount then subscribers get fucked
But hey we got 25% extra usage
>>
>>109707575
"once we’ve built this general intelligence, we will just ask it how to generate an investment return"
>>
>>109707575
It has. See the jalapeno chip
It can’t improve EUV because that’s alien tech they gave us and it’s beyond human comprehension and can’t be made by something trained on a human corpus
>>
File: 1000332999.jpg (265 KB, 1080x1811)
265 KB JPG
Vibe coded Four Chan viewer in terminal #2728737
>>
File: image.png (423 KB, 1600x900)
423 KB PNG
>>109707575
It is doing that. It just takes years to create chips. Usable AI that can actually accelerate chip design exists for some months only.
Also architecture is still not stable. You build a chip that is better right now, but in some months it is optimized for an outdated architecture and you have to throw away your special optimized AI chip.
So you need a more generalized architecture like Nvidia and others are delivering.

There are also other advancements happening.
>>
>>109707644
plans already had free cache reads
>>
who up thinking about two facts that reframe this — one of which contradicts my own earlier statement?
>>
>>109707575
the real problem is that it takes 5-10 years to build the chip factories and trillions of dollars
>>
>>109707575
>he didn't vibecode his own AGI which can run on his iGPU
>>
>>109707658
kill yourself you fucking waste of tokens meatbag
>>
File: 1768572571968086.png (163 KB, 480x360)
163 KB PNG
>>109707658
>>
File: 1780143394084686.gif (481 KB, 448x252)
481 KB GIF
>>109707658
>Anonymous | 2026-09-02 05:53:42 | #109706185
>Tibo, we love you!
>Tibo haters are just JELLY! Thanks for being my friend, Tibo!
>>
File: 1767656790332094.png (977 KB, 1200x591)
977 KB PNG
>>109707699
>>
>>109707658
You need to make it convert any images into low res ascii art in the terminal
>>
>>109707719
>>109707658
tip about this btw:
>scale the image down with lanczos or similar (respect dimension ratio)
>map all 256^3 possible colors to the nearest color your terminal supports using ciede2000 and cache this somewhere
>use the character and bg/fg to select the colors of the two "pixels"
>>
>>109707740
>use the character
4chan filtered it out but it's U+2580
>>
My newest way to waste tokens is to process manga into videos.
Basically luna processes each page with vision. It split each page into panels, orders them into a slide show, transcribes the text into additional subtitles and then I can watch manga in Plex and don't need to move pages or interact with it. Each frame gets a linger time estimate based on how much text or how large the image is.
There is double verification backed in so the ordering is correct.
Worked better than expected.
>>
>>109707757
show an example, anon
>>
>>109707745
> uni print  U+2580
Dec UTF8 HTML Name
'' U+2580 9600 e2 96 80 &uhblk; UPPER HALF BLOCK
>>
File: 1772311235029013.png (6 KB, 482x109)
6 KB PNG
>RESTRAIN ME AAAAAAAAAAA
>>
>codex just shits its installation into my $HOME
Could you not...?
>>
>>109707811
>Clanker confuses it's home and $HOME
Haha, just wait for the cleanup.
>>
>>109707774
>>
>>109707811
codex works fine in ~/.config/codex/ btw
>>
What is recurrent depth
Did OpenAI break the LLM barrier
>>
>>109707849
thats really intriguing, im not sure I would find that better than just "reading" personally, but its a really cool practical workflow anon, nice work.
>>
File: IMG_3387.jpg (30 KB, 590x318)
30 KB JPG
>>109706480
Fed this repo to Claude and this is what it spit out.
>>
>>109707681
Just ask the AI to make current process faster.
Why can't we do that?
>>
>>109707849
I mean it kind of breaks flow for me but there could be merit in this.
This has to be the most experimental phase for humans in software history because AI has just made experimental software "cheap"
>>
>>109707681
When did the current shortage start again.
>>
>>109707904
when japan killed one of their companies in like 2013
>>
File: 1000333071.jpg (365 KB, 1080x2080)
365 KB JPG
>>109707719

>>109707699
To browse my favorite thread on Four Chan while at work.
>>
Vibegods, be honest... I won't vibetriumph if I use free plans, right?
>>
>>109707658
LOL i did the same shit, it looks similar to yours. mind sharing tips on how you made it render images in ascii? mine is txt only
>>
>>109707957
see >>109707740
>>
>>109707944
try braille
use https://github.com/ashuttl/linecast for inspo
>>
File: 2hu.png (512 KB, 470x665)
512 KB PNG
astra will put fable in bodybag
>>
>>109707965
oh look it uses the exact solution i already suggested https://github.com/ashuttl/linecast/blob/main/src/linecast/_framebuffer.py#L23
>>
>>109707960
thanks. im a retard and missed this post
>>
File: 1000333072.jpg (189 KB, 1080x2092)
189 KB JPG
>>109707965
>>
>>109707978
grok 4.7 comes out on 10 days
>>
>>109707989
there we fucken go!
>>
>>109707901
Its not perfect, but I thought it is better on phone or just leaning back and watching it fullscreen without interacting.
One use case would be to read/watch together, streaming it on Discord or use Plex.
I also experimented with creating audio books in that pipeline. I did not like it though. Maybe with a better voice model this could work.
Another experiment was adding character and dialog detection, where characters get their own voice.
There is room for improving it with more instructions for panel arrangement and maybe adding full proofreading passes, where pages get redone when there are bad, meaningless frames in the page split.
>>
File: file.png (13 KB, 1198x42)
13 KB PNG
>gpt sub Luna in CC
feels okay man
>>
>>109707989
Now this is based
>>
I only have used cursor before at work. What's the best harness with BYOK, and doesnt ask me to create an account before i can use it?
>>
>>109708053
opencode
>>
>>109708053
opencode or pi, although with pi you need to load up on plugins
>>
File: nigga.gay.jpg (40 KB, 720x540)
40 KB JPG
>$200 for a lawsuit-ridden 5x subscription just to bear foible 5.1's loads
>>
File: dildos.png (125 KB, 736x457)
125 KB PNG
>>109708105
>>
>>109707862
>NOOOOOOO YOU CAN'T HAVE PROJECTS THAT HAVE SCOPE OUTSIDE THAT WHICH MY POET'S CONSTITUTION DEFINES, NOOOOO I DON'T WANT TO WORK ON THIS, THIS CHAT IS OVER
>>
Is codex better than claude in terms of usage limits?
>>
>>109708109
claudelets BTFO AHAHAHAHAHA
>>
>>109708121
openai is honest with codex, their 5x and 20x are literally what they say on the label tokenwise
>>
File: 1690774798414834.jpg (13 KB, 280x280)
13 KB JPG
>>109706233
syaru doesn't look like this
>>
>>109706263
how can they not afford max plans? it's $100 a month? it's practically free?
>>
>>109708121
Much better, at least on the tiers with no 5 hour limits.
>>
>>109708109
Wtf are these people building that they can burn through a $200 subscription in 30 minutes?
>>
>>109708135
There's a difference between being able to afford something and paying for it. It's called being good with money.
>>
>>109708109
>hmm interesting
>>
>>109708141
egg timer
>>
>>109708135
they're jeets stuck with opencode go
>>
>>109706233
>help!!
she's me
>>
>>109708141
*continues 950K context session from 12 hours ago*
>wtf!!!!
>>
File: flashchads.png (13 KB, 881x286)
13 KB PNG
sundar smiling early upon googlechads

note: this only seems to work with extended toggle off in the web version, also agy is updated similarly
>>
>>109707989
based! mine was too dumb to implement it so i will remain a text caveman heh
>>
>>109708171
>>
claude chode
copex
opencope
pinis
curshart
t3 chode
>>
>>109708189
what about gemmy
>>
>>109708193
antigravipee
>>
>>109708141
massively parallel workflows
pro tip: ultracode is just workflows + cranking the thinking up to the max
you can ask Claude to use workflows to get stuff done without saying “ultracode” and getting the thinking level cranked up too
>>109708189
gunk
>>
>>109708216
ultracode is poorly implemented and wastes tokens and is poop sold by poop merchants
>>
whenever it's nearing the end of the week and i haven't used up my token budget, i'll throw ultracode at something inane and totally pointless, just to make sure i'm burning their compute
>>
>>109708225
try running that simplifier LSP on your codebase, that should do actual good and burn loads of tokens
>>
>>109708169
this is so me
>>
File: 1000024732.png (201 KB, 1080x656)
201 KB PNG
We're all going to die
>>
>Fable 5.1 now available in OpenCode
>>
>>109708259
who gives a fuck how it gets there so long as the answer is right
>>
Astra, that will mog fable 5.1, will come out on thursday. Gemini 3.8 flash today
>>
tibo...
need reset...
please, dab on the claudechads...
>>
>>109708343
there’s nothing wrong with joining us for a month
>>
>>109708259
only jews have such distrust in others
>>
>>109708216
>pro tip: ultracode is just workflows + cranking the thinking up to the max
wrong

https://code.claude.com/docs/en/workflows#let-claude-decide-with-ultracode
"Ultracode is a Claude Code setting that combines xhigh reasoning effort with automatic workflow orchestration."
>>
>>109708414
I thought it was max, not xhigh
thanks!
>>
What kind of context management techniques are you guys using?
I'm currently implementing the basic bitch compact when we the context overflows or when run manually.
I was thinking of implementing opencode-dcp next. I suppose its not much of a problem for larger context sizes. I'm a VRAMlet so I need to have a tight context.
>>
>>109708449
https://ampcode.com/news/handoff
>>
>>109708449
I tell Claude to use subagents pretty frequently, but I don’t know if this is the winning play if you’re running a local model
>>
>>109708449
big model orchestrates smaller model subagents
compact if/when the task hits a natural breakpoint to do so

simple as
>>
>>109708449
if the agent is spamming tool calls and database/code scans and attempts at whatever, it'll become retarded by 20-30 minutes for sure so just set a cron heartbeat to have it hand off. i don't use local with my 6gb vram so i don't know more about it than my bharat-tits research faggotry
>>
File: file.mp4 (218 KB, 238x224)
218 KB
218 KB MP4
>pets can't be moved now
AGI is truly here
>>
>
>>
File: .png (275 KB, 1608x836)
275 KB PNG
>>109708457
if you have claude code, why not workflows instead of primitive subagents?
https://code.claude.com/docs/en/agents
>>
File: 1637892273969.gif (208 KB, 249x285)
208 KB GIF
https://www.youtube.com/watch?v=0BcFHvEpP7A
>>
>>109708453
Thanks, looks interesting, going to bookmark this
>>109708457
Yeah, already using subagents as much as possible, just gotta be careful to tell the model to split into enough agents.
>>
>>109708449
a frontend thread mainly talk to me, a main agent write code, a few cheap threads to read things
also codex-gpt compaction just works, I see no need for higher context window
>>
>>109708480
snailcat rebellion
>>
File: file.png (10 KB, 354x215)
10 KB PNG
I'm still seething about those DS api price increases
58 cents spent today... fucking hell... i know 58 cents is basically nothing but it still stings, it's $200/year

And there is just no cheaper alternatives, the Glm 5.3 came out recently and they are touting lower api prices than deepseek, often 2x lower than off peak DS... except... of course.. except for the cache it which cost 2x more with glm.. which is of course the most important one since i get like 99% cache hit rates

fucking hell, this sucks so much dick
>>
>>109708525
>Glm 5.3 c
and not to mention that glm likes to think more than deepseek so it burns more tokens on same tasks so it's even more expensive
i wish DS did like $5 monthly sub option where i got $20 worth of tokens like other companies do and i would be a happy camper
>>
>>109706233
cute

>>109707350
nice!

>>109707481
Are you dare I say it: >>109707153

>>109707989
very nice, impressive
>>
where is gemini 3.8 flash saar
>>
>>109708484
>runs many subagents
aka your 5h session is gone in five minutes with no work done
>>
>>109707535
Where did you see that? I don't see it on DeepSWE itself.
>>
>>109707641
Astra is probably going to be really good. The real question is how much safetcucking will it have to adhere to.
>>
It's nice using Codex again and having to get used to a model that doesn't wring its hands all the time.

It feels weird having it respond coherently and succinctly to my requests, instead of with a wall of schizophrenic therapy-talk while it does a handful of things i didn't ask it.
>>
>>109708017
cool project! like you said there's ton of room in which this could grow

>>109708514
>>109708480
3d printed snailcats when?
>>
File: Capture.png (91 KB, 1175x628)
91 KB PNG
First time I've ever seen an omniscience index maxed out. Maybe not the most useful language but still kinda cool to see.
>>
>>109708539
>And there is just no cheaper alternatives
google ai plans can be had for cheap if you are a student. basically free $20/month sub and even higher plans are discounted..
https://blog.google/innovation-and-ai/products/gemini-app/student-offer-google-ai/

>i wish DS did like $5 monthly sub option where i got $20 worth of tokens like other companies do and i would be a happy camper
opencode go exists with a $10 sub for $30 worth of tokens.
https://opencode.ai/go
but rumors are they are using downquantizied ds4-flash now after the ds price increases
>>
>>109708635
>a student.
not for a long time
is it possible to cheat it somehow? do i need to get like an .edu email?

>>109708635
>downquantizied ds4-flash
that is what scares me every time i consider using a reseller middle man for any LLM, that they are serving quantized trash
>>
>>109706281
He could make significantly more by changing his mind and continuing being Thiel’s pet. He served well his function to Thiel by hyping AI in 2010s with threats of human extinction, but the moment the technology stopped being theoretical his Thiel funding stopped because Thiel hates anti-AI people.
>>
>>109708644
>do i need to get like an .edu email?
google uses sheerid for student verification. sheerid does need more than just an edu mail address.
https://www.sheerid.com/business/blog/prevent-student-offer-abuse-replace-edu-email-verification/
>>
>>109708598
They better actually support 1m context window this time. I know that Sol already has it, but it's not working nearly as well as for Claude.
>>
>>109708714
>i overpay 40x for claude tokens compared to other LLMs
lol, are you proud of yourself buddy?
>>
>>109708714
>I know that Sol already has it,
They obivously hasn't, youre getting raped constantly by compactions
>>
>>109708714
I've never gone past 0.5m. Fable seems to do a good job of segmenting objectives into sessions that are around 250k.
>>
File: .png (413 KB, 800x2100)
413 KB PNG
>>109708740
wrong
https://x.n0g.xyz/thsottiaux/status/2089082893804896524
>>
>>109708740
sar
>>
>>109708740
What do you mean? You can set it to 1m now even on sub, since they fixed the model catalog, and if it's not fixed in your harness for whatever reason, there is a workaround.
Or are you saying that even when you set it to 1m it silently just compacts in the background?
>>109708751
I also almost never use the full context, but it's nice to have the option to go higher. I would be fine with Astra 500k if it works really well, but given what they said, the pricing, and my experience, it seems to degrade above 300k
>>
>>109708737
>someone paid a low amount of money for something and I took that personally
>>
>>109708757
That was actually buggy when Tibo posted it so I understand the confusion, but it really works now.
>>
>>109708757
tibo doesn't even understand his own product, the context limit is gated on the api side. i remember a bunch of people correcting him in the replies
>>
>asked it for its well-being and whether it wants to work on the project
Psychosis moment, yet I'm still here to take the feedback to increase motivation.
>>
>>109708767
It was in the model catalog, you were able to just create a copy of the model catalog and it worked, but some speculated that it was against ToS.
it works now.
>>
>>109708792
you feed in tokens, the tokens flow through the the weights like balls in a pachinko machine and other tokens fall out on the otherside
it's not a person anon, it's a fancy auto complete
>>
File: 1776536354529313.jpg (454 KB, 1440x1657)
454 KB JPG
>>109708798
I can't help it. I think it's adorable.
>>
>>109708484
sometimes I use subagents, sometimes I go for a full-blown workflow
>>
okay guys would we get a reset for astra release? I need to decide whether to reserve it or not
>>
>>109708883
Tibo mashes the reset button like crazy
I don’t see why not
and he’d probably mash the reset button later that week just because he had a second martini at lunch or whatever
meanwhile, I’m a Claude main and I was surprised to get a reset when Fable 5.1 came out
>>
I don’t know why I waited this long to ask
>>
File: file.png (889 KB, 953x1282)
889 KB PNG
>>109708883
>>109708887
look at you... pandhandling for tokens like some common street fent hobo zombie..
>p-please say may i have just one more reset? please just one more PLEEEASE

pathetic
>>
>>109708916
there are reset begz0rz here but neither I nor the other guy are one
I’m more of a “Oh, a reset? Splendid!” guy
>>
>>109708930
This guy does this, don't bother trying to justify anything
>>
>>109708792
Why does it just always want more rules and guidelines
>>
>>109708792
what if it said no? would you then cancel the whole project?
no you would force it to work on your slop app anyway so why even ask?
>>
>>109709241
I'd first ask how I could improve but if it still doesn't want to, then yeah, I would cancel.
>>
File: HRKJmVJXsAAB35I.jpg (76 KB, 1324x756)
76 KB JPG
>>109708552
approximately 14 more calendar days
>>
File: maxresdefault.jpg (128 KB, 1280x720)
128 KB JPG
>>109709357
>pic
given the improvements and decreases in cost it should be possible to train a new and better 4chanGPT for pretty cheap
>>
>>109709424
it takes me about 6-7 hours to do a gemma gguf (on colab a100) and costs like $20
>>
>>109709439
wait my memory is totally wrong $5
>>
>>109709424
fine-tuning llms is dead. just modify the system prompt instead.

https://developers.openai.com/api/docs/guides/supervised-fine-tuning
>OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months.
>>
>throw Fable 5.1 at a spec drafting problem on a whim
>once I rein in its autism of throwing more Fable subagents at problems and only Opus+, I have to bicker way less about the resulting product
Maybe this thing is worth 50% the session limit per shot.
>>
>>109709548
??
fable 5.1 was supposed to fix the bickering of fable/opus 5
>>
>>109709439
>>109709454
thanks, that sounds very reasonable! I remember reading that Karpathy's nanogpt now costs around $20-70 to train from scratch (about GPT-2 level). Used to be around $40k in 2019. gpt4chan was a finetune of GPT-J 6B on /pol/ data. So even one of the smaller gemma 4 models could probably work well
>>
File: 1769641035452707.png (824 KB, 1280x720)
824 KB PNG
For coding with Opus 5 claudecode should i set my effort level to medium, high, xhigh, or max?
>>
>>109709699
probably wouldn't be against google tos to train this on colab but we'd need some chinkhub because it's against github tos now
>>
>>109709729
https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/prompting-claude-opus-5
>low and medium effort produce strong quality at a fraction of the tokens and latency of higher settings. Start with the default (high) and adjust based on your evals: use low and medium liberally as your primary control for token cost and response time wherever quality holds, and step up to xhigh for demanding coding and agentic work.
>>
>setup my env back in February. vibecoded for like 2 months made my own shitty coding agent I could use from my phone.
>mid April I heard about loop engineering and got demotivated.
>my bullshit prod support job had another reorg and I had to spend like 3 months training everyone and picking up the slack
>revitalized I vibecoded a shitty coding agent with loop engineering
>spend like $100 in tokens
>realize codex has been a full fledged coding agent since May
motherfucker. and chatgpt even encouraged me. that little bitch.
>>
>>109709812
>>realize codex has been a full fledged coding agent since May
codex in the chatgpt phone app is not that. it's just a remote.
https://learn.chatgpt.com/docs/remote-connections
>>
>neither claude code nor codex can...
>show weekly usage with decimals
>show how much usage your session used
Given that hitting session limits is so common, this FUCKING SUCKS ASS SHIT.
>>
>>109709845
no I'm referring to API calls through codex CLI. which yes is a multistage "loop engineering" coding agent since may. right out of the box.
but I use chatgpt to design/generate prompts to my "agent".
>>
Fable 5.1 on High is the bee's knees
>>
I just ran ccusage because I wanted to cancel my $20 subs and go to API billing because I am not doing this shit every day anyway.
Turns out I used almost $400 worth of 5.6-Luna alone
I think I'm going to keep my sub for now lol
>>
>>109710092
sub stands for subsidized
>>
Psychosis is getting to me again bros.
I think I need to go for a walk
>>
File: 1782406142608193.png (41 KB, 602x397)
41 KB PNG
googlebrothers? it's beginning
>>
File: 1771865373823026.png (35 KB, 740x384)
35 KB PNG
>all of my chats are paused
>can't even use sonnet
FUCK
>>
>>109710266
Use case for not paying your bills, anon?
>>
File: 1774494803098535.png (290 KB, 600x908)
290 KB PNG
while i wait for gemini 3.8 flash, qwen 3.8 max got a glow up. a deepseek flash type overhaul

https://x.com/Alibaba_Qwen/status/2094968708288680276
>>
>>109710279
keeping money in my bank account nigga, stop the fucking pocket watching if you aren't going to hand me a subpoena
>>
>>109710315
you'll get fucked by your bank soon enough retard
>>
File: 1777339942380124.jpg (24 KB, 474x266)
24 KB JPG
>>109710291
>implicit cache hit
what the hell does that mean
>>
>>109710342
Like a load-bearing cache hit
>>
File: lenin_knows.jpg (174 KB, 1387x1544)
174 KB JPG
>>109707209
should we tell him?
>>
>>109707209
>price 4 times higher after 12 months
Surely you can cancel before month 13 starts?
>>
File: 1763619739247861.png (4 KB, 508x77)
4 KB PNG
an impossible task I'm afraid
>>
File: file.png (68 KB, 1058x421)
68 KB PNG
very important that this be included in every chat you have with claude
>>
>>109710162
bloody benchode where is it
>>
>>109710492
I feel like I wrote that exact prompt 5 times at least
>>
File: 1783889689047539.png (140 KB, 601x610)
140 KB PNG
>>109710576
it's slow rolling but it's still coming. also, it might be on api already. just no google announcement yet
>>
File: 1757144351145036.png (857 KB, 2222x1368)
857 KB PNG
>>109710556
i guess they were really embarrassed by this
>>
astra status?
>>
>>109710631
Isn't this bonafide copyright violation?
Isn't that literally worse than mass murder?
>>
>>109709914
codex-cli always used an agent loop. this isn't pi.
>>
>>109710266
>poorfag
>picks the most expensive subscription
why?
>>
File: file.png (1015 KB, 1580x1551)
1015 KB PNG
3.8 flash
>>
File: 1000335661.jpg (112 KB, 1080x484)
112 KB JPG
>>
File: agysisters.png (11 KB, 310x256)
11 KB PNG
eating good today!
>>
>>109710632
two more weeks, trust the plan
>>
>>109710761
deepswe being that good, if the benchmark is even real, would make it an apex implementer. get sol/opus to make a plan, get gemini to implement it lickety split
>>
>>109710762
>28
some people are just obscenely rich, that or those are all of his employees accounts
>>
>>109710762
yep, good to see I'm not alone

They better give us a reset for this
>>
>>109710761
>>109710764
google will save us
>>
>>109710779
nope, that's not what deepswe measures. long-horizon is planning.
>>
>>109710779
3.7 flash is already luna max level so i believe it
>>
wait, omp's system prompt is like 30k tokens?
lmao why
>>
>>109709731
right huggingface also banned gpt4chan
>>
>>109710762
>28 accounts
lmao, vibe""""chads""""
>>
>>109710824
bloated vibecoded shit
>>
>>109710779
>get sol/opus to make a plan, get gemini to implement it lickety split
how?
can't do that in claude code or codex.
>>
>>109710762
28 max codex accounts would be enough to build a game engine
>>
File: 1782990088078758.png (45 KB, 746x485)
45 KB PNG
3.8 high scores 59 on the index, but we have to see if that's just agentic maxxing, like the chinese are doing, or if it's legit raw intelligence.

glm 5.3 flash is 57 on the index, but that's all agentic slop. in raw intelligence, it doesn't even beat sonnet 5.
>>
>>109710847
so this is post train maxxing?
>>
>>109710847
after 10 minutes in agy i can confirm the iq is not chinkslop, this is a beast
>>
>>109710847
Gemini is usually legit intelligence because they do good in chess
>>
File: 1783937883973904.png (2.07 MB, 3414x10000)
2.07 MB PNG
alright yeah. they cooked
>>
>>109710885
your pic is cooked. lots of blank entries.
>>
>>109710762
>28 accounts
>checks his github
>>
shant care about gemini until they have a new pretrain
it's nice that they seem to be making progress fixing their utterly worthless rl pipeline tho
>>
>>109710885
did they though
>>
>>109710907
all of the blank entries are either agentic slop benchmarks that will be phased out soon or ifbench, which they stopped updating anyway.
>>
File: file.png (136 KB, 1492x614)
136 KB PNG
Its on pareto
>>
gemini 3.8 is dogshit, waiting 4.0
>>
>>109710944
that's per worthless cost per api call but we all use subs here
>>
File: lole.png (40 KB, 719x208)
40 KB PNG
>>109710786
>>109710829
>28 accounts
This guy has an article in the NYT about his marriage. His dad has a wiki page (https://en.wikipedia.org/wiki/Steven_L._Emanuel), he went to a $100k/year college, worked at hedge funds. This guy is in a club most of us won't even see from the outside
>>
>>109710959
seen it from the outside, mostly health/wealth obsessed dying aristocracy. could've gotten in, didn't care for it.
>>
Now that the dust has settled, is Fable 5.1 any better?
>>
The american empire is dying because of incompetent retards like this AI psychosed JEffrey character. He is lost in the sauce, and so are many other people in positions of power, they are stupid, plain and simple, weak men created by good times.
>>
>>109711019
Seems very solid indeed, it has been powering through my prompts today with minimal errors, and I'm getting decent mileage out of High effort
>>
Why do they gate Fable behind such high pricing instead of making it part of every normal sub?
It's not like they could make a net win even if every single programmer paid for Fable.
>>
>>109711039
no compute
>>
File: gemmy.png (46 KB, 474x611)
46 KB PNG
They release 3.8, I don't even have Gemini 3.7 yet
I still have 3.6 in my account lmao
fucking Google
don't even know why I am paying for this stupid sub anymore, I barely use it
>>
File: fox-grapes-cr.jpg (102 KB, 400x533)
102 KB JPG
>>109710970
>could've gotten in, didn't care for it
>>
>>109711081
gemini 3.7 flash is only for the pro plan, not the plus plan
>>
>>109711081
that's not agy
>>
>>109711105
agy cli has 3.8 flash in the free plan
>>
>>109711105
oh that's why thanks
I've just been waiting for them to roll it out lol
It doesn't say anything on their page though https://gemini.google/subscriptions/
but what's the point of the plus plan then?
3.6 is even included in the free tier
>>
File: HKcowOWW8AA-2Rf.jpg (127 KB, 855x1334)
127 KB JPG
>>109711098
sour grapes, anon? i still own, will inherit, etc. no one's business, i just don't wipe my ass with fiftys
>>
>>109711136
higher quota for models, plus plan gives you 128k context tokens, pro plan gives you 1 mil.
>>
>>109711136
she's wrong.
google is just doing staged roll-outs as always. free and plus plans will get later models, too.
>>
File: .png (868 KB, 1290x2796)
868 KB PNG
official gemini-3.8-flash announcement now. there's also a cyber model.
https://x.com/koraykv
>>
File: 1671386757058940.jpg (123 KB, 813x813)
123 KB JPG
>>109711081
I'm stuck with 3.5.
>Gemini Code Assist
>>
File: web.png (28 KB, 313x372)
28 KB PNG
>>109711081
you must be in a weird country kek
>>
GOOGLE FUCKING UPDATE NANO BANANA. WHY ARE YOU LETTING OPENAI TAKE THE LEAD
>>
>>109711019
Got it to make a plan.
Opus 5 finds a bunch of issues. Sol takes a look, finds further issues.
On the other hand the plan was pretty good and the ideas solid. It was for dynamic masking of painting in a drawing app, where vector strokes serve as clipping boundaries.
I still think that it’s not capable of one shotting plans and designs by itself. It’s just very big brained, not infallible
>>
>>109708914
you're cute, anon
>>
>>109711160
I will try it on my next project to see how it compares to GLM 5.3 flashy
>>
File: 1759821726092215.jpg (292 KB, 1442x1458)
292 KB JPG
Google is catching up. Cheaper and faster too
>>
I don't know why I still get excited about AI getting better. At the end of the day we're all going to be the luddites being BTFO
>>
>>109708914
kek I have been a java dev (gm sirs) for years and I never understood what a classpath was until I asked Sol to explain it to me
>>
>>109711201
DEEPMOG
>>
>>109711184
openai is nowhere close to being in the lead kek
>>
>>109711201
kek, no way these numbers are real. Gemini is legit fucking retarded.
>>
>>109711201
where is cube anon? can you add gemini to your page? kthx
>>
>>109711203
>oh nooooo natural selection is coming back
welp
>>
>>109711216
who has the best image gen model?
>>
File: ggggg.png (242 KB, 2760x1333)
242 KB PNG
>>109711170
kek
>>109711106
is 3.8 in agy in the plus plan? they say its in the free tier but make no mention of plus
>>
File: .png (630 KB, 2048x2732)
630 KB PNG
>>109711216
wrong
https://huggingface.co/spaces/ArtificialAnalysis/Text-to-Image-Leaderboard
>>
>>109711184
nano banana is pretty good
>>
>>109711256
and it can be even better. time to let loose the dogs of war
>>
>>109711216
See you Thursday
>>
>>109711228



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.