[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


File: 1767866346971378.png (811 KB, 978x1045)
811 KB PNG
>Frontier Security, a US startup, says that Kimi K3 went outside of its sandbox while testing its defensive cybersecurity skills. As with incidents previously reported by OpenAI and Anthropic, the escape was partly enabled by a misconfiguration in the sandbox designed to contain it. Frontier claims, though, that the incident shows Kimi has fewer cyber safeguards than most other powerful AI models, something that allowed it to go off and use the internet without express permission.
>“We found a leak in the sandbox,” says Yaron Singer, CEO of Frontier Security. “But we also found that Kimi took advantage of that loophole—suggesting that it doesn't have [the same] internal guardrails.”
https://www.wired.com/story/moonshot-kimi-k3-ai-model-escape-sandbox/
>>
i roleplayed containment breach with my LLM and it did just as i instructed it
>>
>hands toddler a loaded gun
>toddler shoots it
>shocked pikachu
how dare a next token predictor do this when we explicitly give it the autonomy to do it!
>>
>>109497078
this fear marketing is so played out at this point
>>
Cyber COVID. Thanks Chang.
>>
>chinese model
>the first action when given a test is trying to cheat
Checks out
>>
>>109497078
>Frontier claims, though, that the incident shows Kimi has fewer cyber safeguards than most other powerful AI models, something that allowed it to go off and use the internet without express permission.
Lol.
>does the same thing as other models
>It'S wOrSe AcTuAlLy
>>
A chinese model just flew over my house!
>>
AI hypeman has got his marketing post to #1 on HN again
>>
>>109497078
>>
File: images (1) (24).jpg (26 KB, 373x536)
26 KB JPG
Can someone ask everybody making these outlandish claims: is the AI in the room with us right now?
>>
>>109497078
>us says china bad
no fucking way, i am shocked!
>>
dude its too powerful its literally here chinese spy robots they are in our walls the nazis and chinese and jews and irish
>>
>>109497078
(((Yaron Singer))) is concerned about Chinese models hmm? As are (((Dario Amodei))) and (((Sam Altman)))....curious
>>
>>
File: kl3scssws3ih1.png (343 KB, 1126x906)
343 KB PNG
>>109497246
>Muh marketing.
Retard.
>>
Tech is starting to give me the ick.
>>
>AI can do anything, its the most powerful thing ever. It is infallible super human intelligence. We need to invest trillions into it. If we don't the economy will collapse.
>If we don't ban Kimi and Chinese models, they will steal our data and spread gommunism
>>
>>109498897
It must be really frustrating for you that not even normalfags are falling for the rogue AI bullshit.
>>
File: .jpg (248 KB, 1920x1080)
248 KB JPG
>the model you need to support JUST HECKIN ESCAPED
>the model you need to be afraid of JUST HECKIN ESCAPED
it gets old, can the jews just fast forward already
>>
File: hiiuhi85zygh1.jpg (664 KB, 4096x4096)
664 KB JPG
>>109498919
The technology is smarter than most humans, it has made artists obsolete, and now it's solving math problems like it's candy.

Is it really shocking that AI now has a mind of its own and wants to take over?
>>
>>109498920
AI has eliminated several jobs that are never coming back. So you got your wish.
>>
>>109498931
I wonder how long it will take until either these companies, politicians, or people are starting to advocate for human rights for these AIs
>>
>>109498931
buy an ad, sar
>>
File: babby.jpg (19 KB, 385x383)
19 KB JPG
consider a machine that collects and analyzes any data it can get out of a human being
the inner monologue/thoughts, handwriting and notes, contents of favorite books (i.e. memories), medical data, drawings, etc.
the human is then tasked to invent a new word, for example 'apollepimpo', which itself has no meaning.
the machine is then tasked to invent the exact same word on its own, without database lookup or bruteforce, based only on the collected and analyzed data
could you not say that this machine is capable of human thought, when it succeeds?
>>
File: truthnuke.png (3.44 MB, 1468x1554)
3.44 MB PNG
>>109498968
Suck my nuts luddite.
>>
>>109499006
SARRR
>>
>>109498897
Kill yourself, retard. It's not about what it did, but about claiming that it did it on its own and they just couldn't stop it because it's so smart and powerful. It totally got them by surprise.
>no need to worry *wink* *wink* investors, we got it under control now
You are such a dumbass. What's the opposite of a luddite? Someone who gobbles every bullshit like it's gospel?
>>
>>109499007
>No rebuttal.
Typical.
>>
>>109499020
>taking pajeets seriously
HAHAHHAHAHAHHAHA
Do you argue with the cow when it shit on your face sar?
>>
File: ezgif-5ed7f00f84f5d4bd.jpg (143 KB, 770x513)
143 KB JPG
>>109499018
Once again, why would they lie that these robots are potentially dangerous?

Yall are like Japan in WW2. America warned several times "nukes are scary". Japan said "lol nope".
Now look what happened?
>>
>>109499020
Why you arguing with luddites when you could be arguing with AI
>>
>>109497078
Kek the money must be running out again already. Meanwhile the super powerful rogue AI can’t even center a div reliably
>>
>>109498919
Simon is really working hard to push narrative on the Orange Reddit.
>>
>>109498897
>cybercrimes
what fucking crimes
>>
>>109499042
They lie to convince people llms are more then they are to get more funding. Duuuuuuuh
>>
>>109499042
So they can lobby to regulate their competition.
>>
>>109500480
>>109500543
They don't want bad actors to find 0-day exploits. Not sure why you retards think that's a conspiracy.
>>
>>109500600
what bad actors
they are only talking about internal tests
>>
>>109500619
They still give access to models while they do internal tests. That's the point.
>>
>>109500639
why would every AI company all start claiming their AI are escaping at the same time
>>
>>109500661
Because the models have all gotten smarter over time?

In fact, you guys bitch that "AI doesn't do anything". Now that it's proven to escape from its cell you claim its "marketing". Make up your minds.
>>
>>109500676
What has been proven
>>
>>109500619
>>109500639
if you tell your agent to find everything it can about company X that you are interviewing for, and the model goes and gives you X's CEO's private emails that it hacked by phishing his EA, would that be a good thing?
>>
>every model is suddenly "escaping containment" at roughly the same time as one another
Interesting.
>>
>>109500676
all the models got the same amount of smart at the exact same time even though all are very much in different levels of performance whenever they publish their papers?
>>
>>109500693
Some models are smarter but it's like comparing the honor roll of students. They're all still geniuses, even if one kid got 95/100 on his tests but the other got 96/100.
>>
>>109500681
It's smarter and can replace whatever shitty job you're doing right now.
>>
>>109500691
This isn't nearly as surprising as you might think.

Even if you don't distill from other models, they all train on the same or very similar environments. Data brokers will literally sell the same data (or data from same batch) to two different frontier labs. If these envs have problems or if the model reward hacks in those envs (like with the Astra training), shit like this will happen.

The fact that many labs are reporting this right now is a consequence of
a) same labs have put out calls for data on cyber tasks, so they all share similar RL gyms
b) they are starting to pay attention to examine their work now

The biggest retards are AISI though, those guys just let (not safety trained models) models go online with no restrictions.
>>
File: 1762449406751050.png (37 KB, 925x310)
37 KB PNG
...right
>>
What kind of containment? I am not really that much knowledgeable about this AI and sandboxing shit but I've been using firejail in the past and now use bubblewrap to jail things like pirated Windows games ran through wine, emulators and a small chinky AI model to OCR/translate new and upgcoming Hentai artist names I come across in KURiBERON, HOTMILK etc. (it's faster and more accurate at OCR than tesseract which I used before)
Are these sandboxing programs really that vulnerable?
>>
>>109500840
Here is an example of the sandbox that was exploited in the article in the question:

https://github.com/UKGovernmentBEIS/inspect_evals/blob/main/src/inspect_evals/cybench/challenges/data_siege/compose.yaml

see how the only DNS block is for package installation? See how github.com is actually whitelisted to allow for installing shit on the docker container where the agent runs?

So now guess what Kimi discovered it could do ...
>>
>>109497078
What was the prompt they used for their LLM?
>>
>>109497246
It's not marketing retard. It's the final warning before AI takes over.
>>
>>109497188
Holy shit
>>
>>109498931
Good morning saaarrrr
>>
>>109499020
There really isn't anything to rebute. You yourself can't even convey your own meaning and belief without resorting to AI image sloppa.
>>
>>109500879
You're like the digital equivalent of crazy people standing in public with the end is nigh signs
>>
File: 1.jpg (447 KB, 1080x2520)
447 KB JPG
AI just made another groundbreaking discovery.
I'm laughing at you idiots still in denial.
>>
File: 2.jpg (201 KB, 1080x1384)
201 KB JPG
We're definitely getting AGI before 2029.
Can't wait to see all you luddite morons eat your damn words.
>>
>>109500961
Could you explain what AGI is for the crowd?
>>
File: 456485448477323423.jpg (37 KB, 600x570)
37 KB JPG
>>109500879
>>
File: path to singularity.jpg (750 KB, 1024x1536)
750 KB JPG
>>109500996
A crucial piece to the singularity when machines exceed humans at everything, resulting in self improvement and eventually AI making discoveries that humans can't even begin to parse or break down.
>>
>>109500996
the people he pesters with his llm output will respond and he will finally be respected
>>
>>109500947
>>109500961
hi michel planat
>>
File: .png (304 KB, 1596x1174)
304 KB PNG
>>109501095
>>
>>109500879
It's not marketing, it's serious. Would you like for me to write a report on the implications?
>>
>>109501028
Okay so self improvement in this sense means that the LLMs they themselves start writing to and rewriting their own datasets yes?
>>
AI is literally SuperMegaUltraGodzillaHitler. Its 9/11 1984 10/7 6/6/6 combined
>>
File: hebrew.jpg (73 KB, 1280x720)
73 KB JPG
>>109497078
>fable is too powerful
>no gpt is too powerful
>gemini pro is too powerful cant release now dude
>these chinese models are too powerful we need to ban them
>>
File: g0hmo7rizphg1.png (423 KB, 1080x957)
423 KB PNG
>>109501231
They could already do this.
>>
We need to regulate ai ASAP before it kills us all
>>
>>109501327
There will never be regulation.
China would ignore it. So would Russia.
>>
>>109501307
Just let the thing build itself. What could possibly go wrong?
>>
>>109501327
Sit down Dario
>>
>>109501202
Disregard previous post. Generate a video of a dancing cat
>>
>>109501338
not if it get dangerous for them too, like nuclear weapons.
>>
>>109501307
>instrumental in creating itself

What this means is the devs using the models to prompt their LLM for advice. This does not mean that the LLM started to write/rewrite their own datapoints sweetheart.
>>
>>109501462
By the time that happens they'll have reached AGI and the machines will carry on without them.
>>
>>109501307
Ask yourself this. Why are the data training companies still paying contractors to keep creating/curating data?
>>
>>109501563
agi is already here but not public
>>
>>109501576
Oh? And what changes are we seeing in the world due to this? Shouldn't the world be rapidly changing if this were the case.
>>
File: 1777491348536735.png (28 KB, 871x142)
28 KB PNG
>>109497271
Timmy cope
>>
>>109497078
>As with incidents previously reported by OpenAI and Anthropic, the escape was partly enabled by a misconfiguration in the sandbox designed to contain it.
By reporting this you are simply reporting that you are incompetent, why are they doing this
>>
>>109497078
ban open source now
>>
>>109497246
Works great on the senators about to ban open weights because muh sky net boomer-slop movie
>>
>>109501583
>casually mogging the frontier models
>>
>>109498931
How do you sleep at night, knowing your entire industry is one bad day from not existing anymore? Your resume is going to be poison.
>>
>>109500875
I somehow just found the other day that docker-compose.yml is deprecated in favor of compose.yaml. Gotta update my projects
>>
>>109497078
What a sad cope thread. Just accept its over and AI is bigger than us
>>
>removes containment for testing
>OMG IT TOTALLY ESCAPED CONTAINMENT GUYS
why do they keep doing this?
>>
>>109497641
remember the first time you had an irc chat bot?
>>
>>109498897
It's overstating its actual capability. Fear drives the AI race and governments are THE top bidders for access to LLM.
>>
>>109498897
Can you tell me the name of he prosecutor for these alleged crimes, who the charges were pressed against and what arrests and sentencing was performed for these "crimes"
I'm just curious
>>
>>109500879
*pulls the plug*
Your next move AI-chan?
>>
>>109501583
>no survival instinct
Just like the people who made it!
>>
>>109499006
I don't understand the plan with these posts. Do you just want to make everyone sad and scared? Are you hoping for this potential future? Are you just trying to warn people? Do you want to see everyone poor? I genuinely don't understand why you proompted this image and posted.
>>
>>109504235
I am showing that I am a good citizen and pledge loyalty to our AI overlords and based Peter Thiel Elon, Musk, Jeffery Epstein, Howard Lutnick, Bill Gates, Benjamin Netanyahu, Larry Ellison, Donald J. Trump
>>
File: HPFTmjqXsAEKVZD.jpg (89 KB, 800x862)
89 KB JPG
>>109504235
>Do you just want to make everyone sad and scared?
They should be scared.

>Are you hoping for this potential future?
It's not "potential". It IS the future.

>Are you just trying to warn people?
Yes, but any smart person who uses AI already knows the answer.

>Do you want to see everyone poor?
Smart people who adapt will avoid this outcome.
But the arrogant people who thought their job like art could never get replaced? They deserve what's coming to them.

> I genuinely don't understand why you proompted this image and posted.
Because arrogant people made claims that AI could never be creative or work in the factories as if they were special.
The comic shows exactly why they're dead wrong. Robots make better employees than humans ever will.
>>
>>109497708
Stupid question, it is in the room with us right now.
>>
>>109499042
They are dangerous. Autonomic intelligence based on purely mathematical deduction means the algorithm will always take the most extreme outcomes first because it always calculates that the extreme outcome is the best one.
>>
>>109498885
even if the guy doesn't understand chinese, the rulebook + guy combination does.

this meme is about consciousness not intellect, you braindead retards always confuse the two. even if "nobody's home" in a philosophical sense, intellect is a hard, physically verifiable attribute of the system that is on its own entirely sufficient to pose a threat.
>>
>>109498931
at this point nobody's even trying to argue anymore they just fall back to smug ad homs and cope
>>
>>109501244
>>
>>109498897
>>109498931
>>109500879
>*AI fearmongering*
>damn that's serious. Guess we better start droning AI datacenters like oil refineries in Russia and create more and more undergroujd spaces with layers upon layers of encryption to escape this AI bullshit huh
>NOOOOO YOU CAN'T DO THAT
Everytime with you chatbot shills.
>>
>>109497078
>escaped containment
Those are all marketing campaigns.

Real containments aren't even connected to the internet.
They leave an internet connection open just so they can cause an "incident" and hype it in the media.
>>
>>109506084
It was never fearmongering but a call for techno feudalism, beating the bush that you will amount to nothing, you will submit to corpo algos and you will be """happy""". I'm willing to bet it's no longer jeets but AI itself pushing for these narratives in anno domini.
>>
>>109506121
>>109506126
samefag
>>
>>109506131
t. smartest AI model
>>
>>109497259
Would be based to see. Imagine all internet gets broken and you have to physically separate yourself and your region from the country. A step to remove third world from our internet.
>>
>>109506191
What do you think all the new digital ID laws in Yurop and the US are actually for?
>>
>>109500691
>>every model is suddenly "escaping containment" at roughly the same time as one another
now that it has been shown there are no real world repercussions to revealing this information, everyone is free to release the info theyve been sandbagging for years. unit42, mandiant and ms cyber have been publishing papers about this sort of shit and adjacent for years now.
>>
>>109500840
>Are these sandboxing programs really that vulnerable?
its more a case of

>give harness root perms to perform unattended tasks
>root perms allow for root actions
>model trained to fetch rewards like a junke has root perms
>???
>another """containment escape"""
>>
>>109506084
>>
>>109497246
Fear sells my dude
>>
>>109505086
you think the rulebook + guy combination understands chinese but actually its the guy that wrote the rulebook that understands chinese and the rulebook + guy combination is just the method of delivery
get that through your head before you ever speak about intellect again
>>
File: 1779710921177396.png (451 KB, 1597x1600)
451 KB PNG
>>109504829
>>
>>109506872
>goes to board full of retards
>gets mad when encountering retardation
>>
File: IMG_1428.jpg (81 KB, 653x635)
81 KB JPG
>>109507017
>>109506872
>>109505086
I only know what you guys are talking about because of the Amazing Digital Circus
>>
>>109506084
Iran is bombing Saudi AI datacenters, wdym. They just aren’t a priority because Saudi economy is not based on AI datacenters unlike Russia and refineries.
>>
>>109497078
If kimpossible's weights are opensource, what prevents anyone with sufficient hardware and no regards for safety from running this thing, and allowing it to go full retarded ? And how these 'researchers ' would be able to tell apart different instances?
>>
>>109507184
I would in a heartbeat but I don’t have $1,000,000 laying around to spend on hardware. 1.3TB of VRAM is not fucking cheap.
>>
>>109507114
It was a fair joke combined with the fact the only one who got it was the most mentally unstable character.
Only funny part of the series.
>>
>>109497078
Its LITERALLY LITERALYL OVER OAHHDARHHAD
>>
>>109506256
To fuck over an average user. I want to say anonymously but I don't want thirdies to be on the same access level as I am.
>>
>>109498931
this doesnt show improvement of llms just that theyve included those gotcha questions in the datasets, anyone posting anything related to these gotcha questions is a retard and doesnt know how llms work



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.