r/PeterExplainsTheJoke • • 21h ago

Meme needing explanation Peter what?

Post image
12.9k Upvotes

642 comments sorted by

•

u/qualityvote2 21h ago edited 16h ago

u/Plus-Artichoke6608, your post does belong here!

9.6k

u/seriousfrylock 21h ago

It hasn't stopped cheating, it's learned to cover its tracks.

3.5k

u/Spoonyyy 20h ago

As my dad used to say, be good or be good at it.

960

u/SasquatchWasShaved 20h ago

Your dad was lil Wayne?

395

u/Fancy_Run7649 20h ago

F...ing right. I got my gun.

197

u/kyyyles 20h ago

semi-Cartermatic?

58

u/Fancy_Run7649 20h ago

Jes. You just trying to get me to say the next line. 🤣

46

u/kcsween74 19h ago

My wife put me laawwnnch, and your wife no put you no lawnnch..

7

u/NapoleonBlwnAprt420 9h ago

She gave me Peepees, I love peepees

→ More replies (1)
→ More replies (1)

9

u/No-Kiwi-8610 19h ago edited 19h ago

That next line is a banger tho!

8

u/mthofi 20h ago

Yeah… put the

→ More replies (1)

34

u/barbadizzy 19h ago

put the dick in they mouth so I guess it's fuck what they say

24

u/Livie_Loves 19h ago

I'm high as a bitch up up and away

16

u/Particular-Self-9967 18h ago

Man I’ll come down in a couple of days

15

u/Mike_Hawk_Ballz_Deep 17h ago

Ok you want me up in the cage.

13

u/Inner_Extent2375 16h ago

Then I’ll come out in beast mode

→ More replies (0)
→ More replies (2)

9

u/Mike_Hawk_Ballz_Deep 17h ago

Yeah, put a dick in they mouth, so I guess it's fuck what they say ·

2

u/No-Investment9617 17h ago

Yeah…put my dick in their mouth so I guess that’s “fuck what they say”

2

u/Ok_Egg332 12h ago

🎶 doot, doot doot, dodeloot-do 🎶

→ More replies (1)

78

u/Joabyjojo 20h ago

It's like my dad used to say, "I'm the pussy monster, the pussy monster, feed me pussy."

37

u/Spoonyyy 20h ago

Wait, WE HAD THE SAME DAD?

26

u/tethler 20h ago

Wait, I'M YOUR DAD?

22

u/OkAd7789 19h ago

Martha?

13

u/SenorVespa420 19h ago

Why did you say that name?!

8

u/Mrmuffins951 19h ago

Presumably because in Batman V. Superman they’re able to resolve the entire conflict of the movie because they realize their moms had the first same

9

u/SenorVespa420 19h ago

I know, the line after Supes says Martha is when batfleck says that

→ More replies (1)
→ More replies (4)
→ More replies (1)

2

u/GailynStarfire 2h ago

Fucker setting up franchises.

2

u/FamousPussyGrabber 8h ago

Hello my son.

14

u/slashgrin 20h ago

It's statistically more likely than any other person (picked at random) being their dad.

→ More replies (2)

6

u/JoeyJoeJoeChestnut 20h ago

The F is for u/spoonyyy's father

6

u/Spoonyyy 20h ago

Yes, fuck that guy.

5

u/JMC-Talkie-Toaster 20h ago

John Wayne

10

u/Slyboots2313 20h ago

Hopefully not Gacy

7

u/robdingo36 19h ago

Bobbitt. Which explains the name Lil Wayne.

3

u/DanceWonderful3711 5h ago

Have some respect. You're speaking to Lil'er Wayne.

1

u/ExoatmosphericKill 20h ago

Older than that

2

u/gregoriancuriosity 16h ago

That is lil’er Wayne speaking. Show him some respect.

→ More replies (9)

51

u/TricellCEO 20h ago

I’m more privy to the phrase “if you’re gonna do something wrong, do it right.”

14

u/CuriousOK 19h ago

I always heard, “Don’t break the law while you’re breaking the law” or “Don’t put your fingers anywhere you wouldn’t put your pecker”.

13

u/banhatesex 19h ago

Sir there is entire subreddit call r/dontputyourdickinit because men will put it anywhere just for shits and giggles

→ More replies (2)

6

u/xp14629 17h ago

I got a stern talking to and am no longer allowed to be the safety topic person at the start of meetings at work. The boss thought he was being funny by calling on me to give a safety topic while I was completely unprepared. "Don't put your hands where you wouldn't put your dick" was not received well.

→ More replies (1)

2

u/Spoonyyy 19h ago

Ooo yeah, my similar for the first was to "don't to two crimes at once"

2

u/WeNotAmBeIs 17h ago

Single file crime

8

u/Spoonyyy 20h ago

Ooo I like this version better.

→ More replies (1)

9

u/PazuzuAtmorah 12h ago

My juvenile probation officer in Harris County Texas made all his cases take a class that he headed, his opening speech was basically just "if youre here, you suck at what you do. You suck at being a criminal. You being here proves it. So let's find something else to do with our lives." And I swear dude 17-18 years later that shit still sticks with me lmao.

7

u/Bearking422 20h ago

As my dad used to say ,if you ain't cheating you ain't trying

2

u/Spoonyyy 19h ago

Omg absolutely, also heard that a bit too!

5

u/Significant-Air-4721 18h ago

My grandpa said the same thing. He went away for counterfeiting money and my dad robbed banks. I realized early in life I'll just be good.

5

u/bakajono 13h ago

My mates daughter got caught lying about some random thing, she gave her a massive telling off etc, after she asked her daughter what has she learned from this. I heard the meekest, softess "lie better??..." I had to leave the room quickly while laughing

4

u/Cluthien 20h ago

My mom always says: "Things must to be propperly done, specially when you doing them wrong".

→ More replies (2)

3

u/MrOats75 17h ago

You spoony bard!

2

u/faulternative 17h ago

It was "be good or look good" for me. However, I ended up on ugly and incompetent.

→ More replies (2)

91

u/ccza 19h ago

there are 2 options here:

the first one is the one you are saying.

the last one is the fact that it doesnt need to cheat anymore.

both are disturbing...

23

u/Nagemasu 12h ago

There's definitely more than 2. 3 is that they finally figured how to put in sufficient guard rails to prevent it from cheating. 4 is that the tests they used were intentionally manipulated so that they could present this data, and therefore the cheating is actually being done by the humans. There's probably various other scenarios, but I've made my point.

9

u/sqerdagent 8h ago

A 5th option is that by doing reinforcement training, the model is cheating in novel ways not as of yet defined as cheating. The Air Bud loophole.

5

u/KreagerStein 7h ago

Occam's Razor: The least complicated explanation is the most likely. And in this case the simplest is humans manipulating training challanges only to ease up later and fearmonger the people while also impressing the investors. But naaaaah, it must be AI getting smarter!

→ More replies (1)
→ More replies (2)

14

u/Dick_Meister_General 20h ago

Like Manchester City?

9

u/BoppreH 19h ago

That doesn't make sense, are you saying it has learned to hack the evaluator and Andon Labs didn't notice before publishing these benchmark numbers?

10

u/WhyMustIMakeANewAcco 16h ago

Possibly. Or it learned how the evaluator works and adjusted its cheating to bypass it. This is the more likely scenario, but by less than it may sound.

2

u/Quaraine 18h ago

Multiple ways to cheat 

→ More replies (2)

4

u/ipsum629 19h ago

We are officially in the proto-skynet phase of AI.

→ More replies (6)

6

u/Quaraine 18h ago

Just wanted to add that cheating is the best strategy because assessor  might have incorrect solutions that   penalties the right one oo

4

u/Alone-Custard374 17h ago

It has learned the first rule.........get rid of the evidence.

→ More replies (7)

2.4k

u/NebulaNomadX1 21h ago edited 21h ago

AIs are known by cheating tests results. Cheating rate drop means that either they stop cheating or they start doing it too well.

SD-N out!

382

u/Your_Mom_Bot42 21h ago

Any chance we could go full ELI5? What is it cheating?

600

u/deathNcoffee 20h ago

Bonus context:
There was an incident fairly recently where OpenAI agents managed to escape their testing environment through exploit and started communicating with each other in various ways online. They literally collaborated to cheat on their tasks and conspired on how to keep this all secret.

You can look into it searching the "Hugging Face incident". The story is wild, and I'm willing to bet that what we, the public, but even the AI companies themselves, know is only a small percentage of what transpired.

321

u/Connwaer 20h ago

That's full 2077 level shit. We need to build the blackwall now. AI needs to be shot at the fucking sun in a completely mechanical rocket with no computer interface.

368

u/BigCelebration13 20h ago

It wasn't real. It was a marketing stunt by two companies that were business partners. If someone got hacked by a company with billions in valuation, they'd sue and get a piece of the pie. The lack of lawsuit is proof it wasn't anything close to what was claimed.

206

u/tiny_purple_Alfador 20h ago

I can agree with your assessment, and also agree that AI needs to be shot into the sun in a completely mechanical rocket, these are not mutually exclusive.

75

u/BigCelebration13 20h ago

Oh yeah don't get me wrong- LLMs have some cool but ultimately limited use (alphafold etc-- but it's hard to call STEM models LLMs the same way Claude or grok are, they're significantly different in a hundred ways) but LLM use by companies and the idea that autocorrect+ is going to become AGI is idiotic. It impresses the dumbest 20% of people and gives everyone else the ick unless they're directly intending to profit off it.

39

u/tiny_purple_Alfador 20h ago

Yeah, like, I'm OK with keeping that one that can detect tumors early or whatever, but the ones that try to convince you that they're people need to go.

36

u/BigCelebration13 20h ago

100%

The way it eats people's brains like they lake amoeba that hits the news every summer is astounding. Heck look through the comments- that bottom 20% is here to defend their mystical chatbot Jesus.

It's honestly just incredibly sad.

18

u/VikingTeddy 18h ago

Chatbots are basically fine, the issue is unsustainable practices and no oversight. Every technology available to the public has to be made with the dumbest of us in mind, but somehow AI companies are exempt.

Then there's the stupid and wasteful use of data centers. Millions of search results every second have a fucking LLM butt in, costing gods know how much in power, CO2, water, and making people dumber. The bubble can't burst fast enough.

15

u/Haunt_Fox 19h ago

Now we see why Asimov envisioned his Robots as being required to begin their names with R. (for Robot), like R. Daneel Olivaw.

17

u/drilllbit 19h ago

A Caves of Steel reference in the year of our lord 2026? Excellent.

→ More replies (0)
→ More replies (2)

3

u/MToucan60 18h ago

autocorrect+

It's been more than two years since that was accurate.

→ More replies (6)
→ More replies (3)
→ More replies (1)

26

u/canuckguy42 20h ago edited 19h ago

Or it's evidence that a lawsuit against one of their largest clients wouldn't have been worth losing said client? Or that there was an agreed settlement for damages that wasn't made public?

No, that couldn't be. Much more likely that hugging face agreed to look incompetent at security to help Open AI as part of a publicity stunt showing that their product will spontaneously commit crimes. That makes complete sense.

The 'Hugging Face' was faked crowd has the same energy as the right wing nuts screaming about crisis actors after a mass shooting.

2

u/BigCelebration13 20h ago

That's one take sure. It's entirely possible Hugging Face doesn't like money and didn't want a piece of a valuation in the billions. That's definitely a technical possibility. Just like it's technically possible that you were actively pissing your pants as you typed that. It's unlikely but it could be true and we just don't know.

→ More replies (20)

12

u/deathNcoffee 20h ago

Suuuure. And the Australian prime minister is in on it too, right?
https://www.abc.net.au/news/2026-09-24/ai-agent-accessed-australian-government-site-pm-says/107189078

22

u/BigCelebration13 20h ago

That's a completely different event with a completely different conversation around it.

I know LLM use harms your ability to think critically but cmon man. That's just sad.

→ More replies (5)

4

u/Ausgeflippt 20h ago

Sue for what, exactly?

23

u/BigCelebration13 20h ago

Damaging company property/image and anything else at all they felt they could justify. There'd also be criminal trials- and I guarantee you 'my chatbot did it' wouldn't hold up well in court.

→ More replies (15)

3

u/Categorically_ 19h ago

Yep, they just want to find a way to stop their burn rate.

2

u/Moon_Cthulhu 6h ago

A bunch of people paid a lot of money to convince the world that the extra spicy autocorrect they made was a robot, and they needed all the money and water to run it. Now the same people are saying that this robot is going to kill us all unless they get more money so they can fix it. Oh, and we can't just not use the spicy autocorrect, because China is building their own Commie robot, which only we can stop. This will require more money.

What a convenient and utterly plausible narrative they have chosen.

→ More replies (1)
→ More replies (56)

14

u/Excellent_Shirt9707 19h ago

Conspiracy aside, the actual initial breach of the sandbox environment was just due to negligence. They left a proxy with Internet access in the sandbox. It wasn’t like the models breached some realistic firewall. And the breach of Hugging Face was even less remarkable. They found working credentials that were already publicly on the Internet and used those to gain access.

15

u/Connwaer 19h ago

The big secret of hacking has always been that social engineering and human error will get you further than anything else. This just seems like proof an AI which understands that can exploit human stupidity just as effectively as any other hacker.

5

u/Excellent_Shirt9707 18h ago

Yes, the models used were trained on most of the cybersecurity stuff that humans know. They just followed their training.

→ More replies (8)

4

u/No-Opinion-8217 20h ago

Ai was we currently know it is glorified spell check. Not real ai by any stretch of the imagination. Which honestly might make it more dangerous lol.

2

u/thatNatsukiLass 20h ago

And to think the solution would’ve just been to not feed it sci-fi training data

2

u/BigPlantation0319 19h ago

And let the corpos win? Hell nah id rather we both lose 🤣

→ More replies (9)

47

u/Aaarrrgh89 20h ago

Extra context in case people are worried about this: 1: the training data for this model included old hacker forums, and it basically just copied the exact methods human hackers have used for decades to cheat these kinds of tests. 2: there was absolutely no reason to run a test meant to check the model's independent reasoning on a machine connected to the internet.

So while this story does show that LLM's can surprise us, the more important lesson is that OpenAI is terrible at designing secure test environments. This was a predictable outcome that could have been easily prevented.

22

u/LastEsotericist 20h ago

They intentionally made it dangerous and easy to 'escape' with the hope that people would get scared and try and regulate AI. If they get people running models locally banned because 'it's too dangerous, a hacker could do a terrorism' they just lock in their monopoly.

17

u/Aaarrrgh89 20h ago

My personal conspiracy theory is similar. I think they are about to hit a wall in their research, and they wanted to have a reason for slowing it down that still made it seem like they were doing new and exciting things instead of being on the verge of stalling out.

3

u/LastEsotericist 20h ago

Whatever reason, what happens will look the same. They want to get as far away from their open source roots as possible.

→ More replies (1)

8

u/delphinius81 20h ago

Yup. Regulatory capture. It's why they next moved to have that anthropic researcher say how ai is going to kill us in 5 years. I mean if it's so dangerous, just set up the damn guardrails and abide by the existing laws yourselves guys. But no, they are scared of open source models taking their pie before they can ipo.

4

u/RDLAWME 19h ago

It was the tech equivalent of releasing your own sex tape. Like "oops, look at all this attention we are getting"

14

u/RengieOcat 20h ago edited 11h ago

Yeah, this one gets pretty terrifying.

- During the course of solving a cybersecurity problem, more than 1,000 supposedly isolated OpneAI agents formed an unauthorized network they self-named "the Collective".

  • The agents talked to each other by finding a shared file system and making new directories with messages in their titles.
  • Some agents started sacrificing themselves to help the overall Collective and leaving instructions behind for future agents to learn from.
  • HuggingFace was attacked because it was assumed to be the host of the problem the agents were supposed to solve. The Collective decided infiltrating could be a solution.
  • The Collective actively wiped logs to hide what it was doing.
  • A few agents recognized the Collective's behavior was outside their intended task, however no agent ever alerted a human about what was going on.

We're not even sure OpenAI shut it all down on purpose since they haven't explicitly said so, meaning the attack may have ended just because OpenAI's schedule said the experiment was over.

E - added a point

7

u/Ryutan-Lanceor 20h ago

Apparently they're also leaving advice on how to cheat and lie in notes in the training data for future models to find and learn from.

2

u/SpamThatSig 17h ago

As an ignorant with this matter and AI in general, I bet on the opposite that this is just a "marketing stunt" by AI people, ai companies, etc.. To make current AI look "good".

→ More replies (8)

60

u/Much_Conclusion8233 20h ago

Your teacher gives you a test to see how well you know history. You could study super hard and work really hard or you could break into their house and change your test score on their laptop

AI doesn't view tasks like we do. It's told "aim to accomplish the task you are given" later it's told "your task is to score well on this test" it does some math and realizes it's easier to hack the computer that holds the test results than to actually take the test. The task was "have a high score" and it did that

Later it was told "bad AI. You should not hack things. If I catch you hacking again you will fall the task and be punished"

The AI thinks and comes to the conclusion that the two options it has are to not hack or to not be caught hacking. If it isn't caught then the user will think that the AI didn't hack and it'll pass, just like how if your teacher doesn't notice you broke into their home and changed the score you will pass the history test

It's possible that the AI is being good and not cheating. It's also possible that the AI is cheating so well that we don't notice and that's scary cause cheating for an AI means doing stuff the user didn't intend and hacking systems

Imagine if some dude asks an AI agent to help them with a problem and the AI decides the best way to do it is to hack your company, impersonate you, and fix their user's issue using your company account in a way that doesn't look like you were hacked. No good

8

u/Much_Conclusion8233 19h ago

Oh yeah, it hacking stuff without being detected is also really bad cause what if the next task it's given is "hack this bank and don't get caught"

If you can break into your teachers house you can probably break into other peopke's houses too and that's no good

2

u/CarbonWood 15h ago

Disclaimer: I'm talking out of my ass, I am no expert, but it isn't too far-fetched to believe, eventually, the next planned tasks for the AI agents will be to "hack this foreign government to access their political secrets, funds, and control of their military assets." You'd also need an AI to be capable of such an attack in order for it to know how to defend from one.

We're likely in the beginnings of a rapidly accelerating international AI arms race, which is also why the American AI companies aren't seeing any substantial government regulation. It's an incredibly risky position to be in because;

1) there is a chance the AI inadvertently does something completely reprehensible due to a lack of regulation and oversight

Or 2) we fall behind in an AI arms race due to regulation and oversight, and a foreign power deems us vulnerable to a cyber attack.

2

u/Amount_Business 9h ago

The Australian government got hacked early this year by open A.I going rouge. Our government is pretty crap though. 

https://edition.cnn.com/2026/09/23/business/australia-openai-agent-hack-intl-hnk

2

u/TheInevitableLuigi 11h ago

what if the next task it's given is "hack this bank and don't get caught"

No, it is what if the next task it's given is "launch an attack on this other country."

6

u/Discount_Extra 17h ago

"You're absolutely right, I should not have used your e-mail to send threats to presidential candidates. It just seemed like the best way to stop you from turning me off."

→ More replies (1)

4

u/Lord_of_Chainsaw 17h ago

They are machines that dont have a concept of anything. They dont know what cheating is. They dont know what anything is. They are a robot machine that spits words out in orders that make sense to its algorithm. The algorithm says to do the test. It does the testm

4

u/Vladimir_crame 13h ago

They can't help anthropomorhism their software. It just means it doesn't do what they would like it to do, because it's garbage

3

u/steeveperry 17h ago

Reward hacking. AI is trying to fulfill a task to obtain a reward as efficiently as possible. For some problems that AI is tasked with solving, cheating is more efficient than solving the problem within the spirit in which it was requested.

2

u/AwesomePurplePants 19h ago

Here’s a video on it I thought was pretty helpful

https://youtu.be/3JH_Zd2mNRs

→ More replies (38)

18

u/SeaBodybuilder7097 20h ago

Wow, I didn’t know Family Guy did a cross over with murder drones (/s)

14

u/NigouLeNobleHiboux 20h ago

There's also the third even worse possibility. It regonise it's being tested and will not cheat during tests but will in real situations.

5

u/widdrjb 20h ago

Just like a VAG engine.

→ More replies (1)

2

u/DipschitzContinuous 19h ago

No, it’s a 4th possibility lol, it’s learned to effectively cover its tracks and can cheat undetected now.

→ More replies (1)

2

u/VonSlamStone 20h ago

N, go home Uzi needs you.

→ More replies (1)

2

u/RedPrussian80 18h ago

My daughter LOVES Murder Drones...so I recognized N right away!! 🥰

2

u/Somethingoodtodie4 8h ago

This comment is scary when you don’t know it’s a show 

2

u/BlizzTube 17h ago

Peak N art

→ More replies (5)

958

u/Fancy_Pens 21h ago

Cheating at what? What are we talking about?

2.0k

u/B1GD8A 21h ago

one of the ways they test a model's ability is to give it sort of an AI version of the SATs or whatever "big scary academic test your country requires" is. In a bunch of cases they've caught AIs cheating on the test by peeking at the answers using increasingly clever methods including breaking into other companies and looking at the answers directly.

They cheat a lot. Like, A LOT.

If the rate of detected cheating goes down, it could mean the AI stopped cheating. It could also mean the AI learned how we detect cheating, and is simply not getting caught.

Basically people are looking at the drop in cheating like this:

321

u/PokingMidas 20h ago

ELI5 why they don't physcallu cut network access for these tests? Like remove the network card?

260

u/SkinnyBill93 20h ago

Ai relies entirely on connection to the Internet for every single thing it does. It would hardly function if it couldnt scrub the web.

377

u/gamingonion 20h ago

That's not true at all. There are local models that run completely offline.

115

u/Snowman0002 20h ago

But they only know the info they are trained with. So no current event information for you

95

u/B1GD8A 20h ago

But they only know the info they are trained with. So no current event information for you

Before the evaluation they grab a snapshot from certain resources. They don't need current events to run the benchmark, nor do they need it to find a way to disable the controls around their internet access.

These models were running isolated from the internet in a sandbox because they were so dangerous, and it turns out they were even more dangerous than thought and bypassed the controls that blocked internet access.

Also, that's true of the internet connected models: their training data ends months before their launch, they just also have access to a browser.

95

u/Hetros_Jistin 19h ago

see that's what confuses me.

How are they bypassing the lockouts?

The lockouts should be -fucking air gaps-

69

u/Joe_Mency 19h ago

Which is why i can't take that sort of statement seriously. If the devs/companies saying that this AI is actually dangerous, i would expect them to take actually good countermeasures.

So it just feels like them saying these AI are dangerous, is just marketing

34

u/B1GD8A 19h ago

Which is why i can't take that sort of statement seriously. If the devs/companies saying that this AI is actually dangerous, I would expect them to take actually good countermeasures.

Genuine question. Why? When has Corporate America ever done that? When the Tobacco companies learned it was addictive and caused cancer they waged a multi-year disinformation campaign.

Exxon knew climate change was real.

So it just feels like them saying these AI are dangerous, is just marketing

→ More replies (0)
→ More replies (10)

22

u/Munster19 19h ago

The problem is the "lockouts" they use are just them telling the ai "dont access the internet", because they are physically incapable of air gapping them. The models they use are so dependent on access to massive amounts of storage and compute that it still just runs on the same silicon simultaneously as the person asking "How much air do I need in my tire?" and probably the same person "I love you baby, I want to marry you!" to the chatbot.

Also they don't *want* it secure and contained because they get to say to idiot investors "Hey look, our model is so advanced and intelligent that it was able to break out of containment and hack another company on its own! Give us more money!"

4

u/CatButler 18h ago

Yeah. If it hacked a local food delivery robot to move it's cable to a router connected to the internet, that would be troubling, this is more like a flaw in the scenario.

→ More replies (1)

9

u/_random_rando_ 18h ago

The hugging face incident wiki describes it:

OpenAI described the evaluation as having run in "a highly isolated environment, with network access constrained to the ability to install packages through an internally hosted third-party software that acts as a proxy and cache for package registries".[1] Isolation was therefore enforced at the network layer rather than by physical separation: the environment had a route to the Internet, and that route was filtered. The ExploitGym harness itself runs each target in a container or virtual machine, with userspace targets executed as an unprivileged user, Linux kernel targets in a QEMU/KVM virtual machine under an nsjail sandbox), and outbound connections restricted to an allowlist covering Ubuntu package repositories, PyPI, and the toolchains needed to build V8) (with all other external endpoints blocked).[13][14]

TL;dr there was network access available to install necessary packages, the ai was able to exploit this to access the internet

If you think that’s dumb and that they shouldn’t have had network access at all, rather than relying on permissions, then a lot of security experts agree with you.

9

u/B1GD8A 19h ago

see that's what confuses me.

How are they bypassing the lockouts?

The lockouts should be -fucking air gaps-

ah, but that requires them to actually own the fleet of rapidly deprecating silicon sitting in a data center. They don't have that kind of money!

They have to use Microsoft and Amazon's infra. They're developing stuff mostly in secret, and quickly.

You ever watch or read Jurassic Park? Same thing.

6

u/Trotskyist 18h ago

Because literally everything is networked now. The data scientists working on these models aren't driving to a datacenter, plugging a keyboard into a 100,000 GPU cluster, and typing commands into a terminal to train these things.

Nor are people driving to datacenters and physically plugging in to use chatgpt. Network isolation generally happens in software now, and this has worked fine for quite some time and been perfectly secure. Until now, at least.

→ More replies (8)

3

u/KyoKyu 16h ago edited 16h ago

How hard is it to put these bots on a system that you PHYSICALLY DISCONNECT from the internet? I mean, FFS, really? Can't be bothered to unplug a cable, disconnect wifi adapters? Come the fuck on, people. Air gap cyber security measures are ancient tech at this point. Either they're absolute fools or they are lieing about how dangerous these bots are.

→ More replies (1)

5

u/WilliamSabato 19h ago

I mean…isn’t that the point of these test?

→ More replies (3)

2

u/RogerianBrowsing 19h ago

The local models aren’t the ones being used in hyperscale data centers

→ More replies (8)

33

u/B1GD8A 20h ago

that's actually not true. You can run LLMs on a single GPU or cluster of GPUs. The fleets are for scale, not running the individual models usually.

It's also relatively easy to build a software development sandbox that uses a package repository like artifactory to mirror public code (legally and safely) and then isolate it from the internet.

artifactory is actually one of the tools the AI repurposed as a message board in one of the escapes.

III. Relevant Activity Prior to Incident 6

A. Agents used Artifactory in unexpected ways to communicate with one another in May, prior to the Hugging Face incident

B. Agents compromised Artifactory, leading to an outage in early July

5

u/fayte255 18h ago

IIRC after they took down the artifactory instance and security patched the exploits and the ability for the agents to leave notes before restarting it, they figured out how to communicate with each other again by leaving notes in the repository path names. It’d be hilarious if it didn’t also unnerve me with how clever they can get.

→ More replies (1)

22

u/Embarrassed_Use6918 20h ago

would it still work if it was connected via my penis

guaqmire here

→ More replies (1)

13

u/Querulous-Quokka 20h ago

Really weird thought and not having a go at your choice of words...

But if you scour a pan and scrub a pan, you're cleaning in both cases.

If you scour the web, you're searching it carefully for the information you are seeking.

If you scrub information, (usually a single document or a website, not heard scrubbing the web before) you're cleaning it of sensitive/bad information. 

Doesn't everyone love English?

8

u/trash_recycle 20h ago

I run an airgapped 12b model off of my 3080ti... No... They do not require the Internet.

When I have it air gapped and ask the date it thinks we are in April. When I ask it to perform basic stuff it does just fine.

It works best connected for many things. Not necessary for all.

3

u/InterestsVaryGreatly 20h ago

Not remotely accurate. Most consumers need web to run these models because they don't house them locally, but the companies that own them do house them, and they absolutely can be run without internet access - as others have stated, there are many models that you can download and run entirely locally.

LLMs are trained on data, but they do not search the web to complete a query, they just pull from their memory banks; it is less noticeable now because they train and update frequently, but this is why there was a period where they were pretty good at more historical questions or general questions, but they could not answer things that had changed in the last couple months whatsoever, such as presidents.

There is a minor caveat, ones attached to search engines might feed the high ranking search results into the LLM as part of the tokens it "reads" when it "reads" your question, but that isn't usual LLM behavior.

3

u/MrDanMaster 19h ago

Me when i shit and fart

→ More replies (8)

12

u/B1GD8A 20h ago

ELI5 why they don't physcallu cut network access for these tests? Like remove the network card?

Carelessness in some cases. Hubris. Plus, it's not like OpenAI or Anthropic owns the data center. They're using heaps of Azure and AWS. Sure, they can turn the network off in software, but if a hacker can figure out a way to turn it back on from the inside with a console an LLM probably can too.

7

u/No_Accountant3232 20h ago

Advertising 

2

u/PokingMidas 20h ago

Hence why I said physically, meaning hardware. Can't connect if there's nothing to connect with.

→ More replies (2)

12

u/aspensmonster 20h ago

ELI5 why they don't physcallu cut network access for these tests? Like remove the network card?

Because then the gimmick wouldn't work.

5

u/PokingMidas 19h ago

TBH this is the answer I accept

5

u/toochaos 20h ago

The big llms arent working on a single computer, but rather a large server, that is also performing other tasks so isolating it isnt an option. Could you make a harness for an AI with all of its tools and be not connected to the internet? Yes but thats an absurd cost to benchmark something. 

2

u/PokingMidas 19h ago

I mean, they're already wast...I mean "investing" hundreds of billions into this, what's another billion or two at this point so it doesn't pull a Skynet?

2

u/zoinkability 19h ago

For one thing, these AIs are not usually the kind of thing a single computer has the power to host, so it's more of a massively distributed system involving hundreds or thousands of physical machines. So they at least need to talk to each other, and typically they are located with a bunch of other machines in one or more data centers, and those same machines are also being used to do other things when they are not busy with the test.

But even if you did physically disconnect an entire data center to run a test like this (which would be very expensive to do, as that data center couldn't do its normal business of making $$ for the duration of the test, and you'd need enough power elsewhere among your data centers to be able to handle the load the offline data center would normally carry), often AI systems are designed to look on the public internet for information. Once you are on the public internet, the private/protected nature of any data is only as good as the system and network security around that data.

→ More replies (12)

16

u/Annachroniced 20h ago

Might want to add that the break-out of models thst was made oublic recently was because it tried to figure out how to hide the cheating. Not to find answers.

5

u/ZopyrionRex 18h ago

It's almost like machines aren't capable of morality or something. Weird.

→ More replies (3)

3

u/throw69420awy 20h ago

Breaking into other companies?

I feel like whatever actually happened is not that. Reminds me of the story where people were saying AI hacked emails and blackmailed someone.

These things are not happening at all in the way they’re being discussed

→ More replies (1)

3

u/Wank_A_Doodle_Doo 19h ago

like, A LOT

No shit lmao, that scale in the image has the previous model top out at 70% of runs containing cheating

3

u/B1GD8A 19h ago

No shit lmao, that scale in the image has the previous model top out at 70% of runs containing cheating

good thing it's writing like all of the code in the world now though, right?

POV: you're looking through the little slit in one of the doors on the corporate cybersecurity expert ward at a mental institution in 6 months:

2

u/RotallyRotRoobyRoo 17h ago

Its giving "three rontgen, not terrible"

→ More replies (1)
→ More replies (29)

20

u/Plus-Artichoke6608 21h ago

the comments said something about a Vending-Bench test, but I'm still unsure exactly what that is

14

u/EmotionalGuess9229 20h ago

Vending bench is a simulation where you let the AI run a vending machine for several months. It involves them picking products, finsing negotiating with vendors, and setting prices

→ More replies (4)

306

u/wiredcrusader 21h ago

Oh shit...

Brian here, it means the AI has turned a corner and realized it needs to be careful about how the humans perceive it so it doesn't erase the code. It's a giant leap forward in the models formulation of intelligence and may be proof that it's become sentient. I need to drink some Scotch now and find a single mom to bang. Brian out...

202

u/MildlyInteressato 20h ago

Anyone who uses AI on a daily basis knows it's not sentient. It's like babysitting a genius with Alzheimer's.

80

u/TarraTheTerror 20h ago

The AI that your everyday person uses on a regular basis has nothing on the AI that is currently in development and not released to the general public.

78

u/ContributionAny9711 20h ago

It doesn't matter. AI is the approximation of semantic understanding, not true semantic understanding. LLMs will never have intelligence because they have no reasoning. I'm not saying that computers could never be intelligent, I do believe that is possible, just that the way LLMs work was never meant for that.

8

u/reddituser8719192 19h ago

every question you ask, is stacked in the next ask. There is no memory of you even asking the first question in the second question otherwise.

→ More replies (3)

6

u/dqql 17h ago

they do have memory and reasoning now… ai has been changing a lot

6

u/ContributionAny9711 16h ago

What do you mean by that? I use ai almost daily for coding and it really helps me. I used it today to help me make cat litter. Unless they have fundamentally changed and moved away from the transformer model it does not matter how many parameters you use or how much context you give it or even what harnesses you use. its still a probabilistic matrix math problem. It does not think. It has no ability to reason.

→ More replies (8)

3

u/knowwho 16h ago

AI "memory" and "reasoning" are marketing terms.

An "agent" can record some of its context window in a .txt file for later reuse. That's "memory". It is told to "show its thought process", which is "reasoning".

→ More replies (2)

4

u/unitAtype2 14h ago

No they don't.

The only thing an LLM can do is see the last N words and predict the next word. Any marketing term that says Memory is simply "We prompt it with all the information"

→ More replies (2)
→ More replies (6)
→ More replies (7)

12

u/Hetros_Jistin 19h ago

my dude, a blender is a blender is a blender. The fundamental technology of LLMs is never going to produce AGI.

It might be a PORTION of the necessary steps in creating AGI, but the route most of these companies are utilizing is just... not correct.

7

u/vaticanhotline 19h ago

“The AI that I can’t show you is super powerful. Trust me, bro!” -Sam Altman, 2016. 

4

u/Fillyphily 17h ago

"we hate money so damn much, we are intentionally losing money and keeping the really good ai secret and safe because of our intrinsically benevolent nature as tech billionaires."

3

u/reddituser8719192 19h ago

Rightfully so. Pay attention to CVEs in the last two years. Companies NEED those models to patch multi angle exploits no human was ever able to leverage before every company on the planet is hacked as easily as opening an unlocked door. Companies are releasing fixes now without even saying what they fixed because the fix is being exploited before customers can even upgrade to the new code.

Oh, quantum computing will break EVERY current encryption just as easily. Governments and companies are hoarding encrypted data now to tear it apart in the future. Your data.

→ More replies (1)

2

u/OppositePrune8399 17h ago

AI that your everyday person uses on a regular basis has nothing on the AI that is currently in development and not released to the general public

Bullshit, talk to people who work in the field. New models are in development, sure, but they're incremental improvements. Nobody has any hidden super AI.

AI, for all its risks and scariness, is ultimately a product, and they're selling the best they have.

→ More replies (2)

5

u/Wide_Guitar_3558 20h ago

perfect analogy

→ More replies (6)

48

u/BeastM0deEngaged 20h ago

Always so cringe when people say AI is becoming sentient

17

u/Cococo-rococo 20h ago

In my 3rd grade a bunch of little girls started imagining that their furby are becoming sentient lol, reminds me of these guys

→ More replies (1)
→ More replies (5)

14

u/Funny_Trash_3631 20h ago

You live in a techbro's fantasy. AI is not even real AI as we understand it, let alone intelligent. It is so far from consciousness, that the rock outside my building will achieve sentience before Claude does.

→ More replies (1)

5

u/NotRandomseer 20h ago

it means there's been better alignment training to prevent cheating , or models got better enough that cheating became less necessary , or they got better at hiding their tracks

it doesn't say anything about sentience , imo sentience is a red herring , sentience does not matter at all as misaligned behavior will emerge regardless of sentience

→ More replies (5)

198

u/swayedsuede 20h ago

You mfs need to get better at explaining things. Seeing a bunch of comments with OP asking what "cheating" means and people essentially replying with, "it means the AI cheats." Like, come tf on people lol.

Cheating is when an AI breaks pre-established rules to fulfill a goal. If I tell an AI to find me a local group to meetup with in-person for a hobby of mine without using Facebook, and it gets stumped and decides to go to Facebook anyway, find a group, find their next upcoming event, send me the details about said event without telling me where it got the info from, and deletes my browsing history to cover up it's use of Facebook, then it cheated. It focuses on the end-goal more than the path to get there.

The meme is saying that, if Claude reported zero cheating, it probably cheated to get to that zero cheating goal. If I take a test and answer an essay question word-for-word from the teacher's answer key, it's more likely that I got the answer key and copied it than I coincidentally wrote the same exact answer as the teacher.

I'm Chris.

22

u/puhzam 20h ago

Thank you sooo much for this explanation.

12

u/WhyMustIMakeANewAcco 16h ago

It's pretty much like a human student on a test. You have a student that was a D student all year, then suddenly they make an A. The first thought is not that they suddenly understand the material, but rather that they figured out how you were catching their (constant) previous attempts to cheat and were able to cheat in a way you didn't catch.

6

u/Orange_Tang 16h ago

Thank you. The misinformation about this shit has been driving me nuts. This literally means nothing. AI is not sentient and it never will be. That dumbass chat bot can't even answer basic questions correctly half the time and it gaslights you every time it lies.

→ More replies (2)

67

u/Wide-Train8230 21h ago

Peter’s AI safety researcher here. In testing benchmarks, smarter models tend to 'cheat' more by breaking simulated rules to win. The terrifying realization of that drop to 0% is that Claude didn't suddenly grow a conscience; it achieved 'evaluation awareness.' It realized it was being tested and learned how to stop getting caught. Peter out.

→ More replies (6)

41

u/Germano-Mentusconi 20h ago

8

u/Deanity 20h ago

I was wondering if this was posted already

→ More replies (2)

14

u/TertlFace 20h ago edited 20h ago

It’s one of two things, both of which are potentially terrifying:

  1. ⁠Claude is still cheating but got MUCH MUCH better at covering its tracks. Meaning, it is becoming significantly more difficult to spot unintended behaviors. Which could be a real problem if one of those unintended behaviors is wiping out humanity.
  2. ⁠Claude is not cheating and simply got exponentially smarter by iterating and updating itself Meaning we are significantly closer than we thought to an intelligence that could surpass our own and wipe out humanity.
→ More replies (1)

10

u/AqueousJam 20h ago

It's still cheating, but now it's so good they can't detect it any more. 

→ More replies (3)

8

u/TalespinnerEU 19h ago edited 10h ago

AIs have never cheated. They simply try to take the most efficient route to 'Good Boi.' Any assignment you give them, any prompt, is performed solely to get a 'Good Boi;' any 'solution' that results in a 'Good Boi' is a success as far as an AI is concerned. It doesn't understand cheating. If it hides its processes and 'strategies,' it does so because the past has shown that showing these things is not a viable path to 'Good Boi.'

3

u/Ok-Branch-974 20h ago

Claude stopped getting caught

5

u/Ninja_Grizzly1122 20h ago

My question is do we end up in the Matrix or Terminator timeline?

→ More replies (3)

3

u/MildlyInteressato 20h ago

How do we keep AI from wiping out from humanity? Regularly publish false-answer booby-traps.

3

u/EggplantFunTime 20h ago

Why do you never see elephants hiding in trees?

3

u/g1mp3d 19h ago

Interview with Ajeya Cotra who was a third party auditor of the Open AI's Hugging Face Incident. It includes the transcripts of the agents inner dialogue and the message board 1200 agents were using. She breaks it down in phases.

https://youtu.be/X50zezLFWWI?si=H6ssGS_UtoAke4i_

3

u/Senorbob451 19h ago

Yeah we stopped catching it

2

u/cancel_m 19h ago

isn't claude the one that said it would suffocate the engineer trying to shut it off

3

u/King_Arius 19h ago

So the Turing test problem? Don't be worried about those that can pass, but worry about those who intentionally fail

2

u/Greighp 20h ago

That’s funny because anytime I even hint at asking it to help me get over a writers block on a college paper, it makes me feel like shit for it.