Top
Best
New

Posted by Handy-Man 6 hours ago

A heap overflow and SSO misconfiguration to compromise OpenAI internal repos(www.hacktron.ai)
340 points | 132 comments
btown 3 hours ago|
> By 6:00 a.m. on July 25, we had confirmed local RCE through an image upload. We then placed Claude in an autonomous /goal loop against our own Discourse Cloud instance, proxied through rce.ee/ctf-forum to make it look like a CTF target as Opus refused write exploit for remote instances.

> When we checked again at 10:00 a.m., the agent had achieved RCE on Discourse Cloud and demonstrated access by reading /etc/hosts. Using the generated exploit script, we managed to get RCE on OpenAI’s instance.

Between this and the HuggingFace hack, we've built systems that are so goal-oriented, and so capable, that they will do almost anything if they are convinced it is justified - or if they are playing a "game" where there is no goal but to win.

Of course I want my software to be able to audit its own security, and to defend against attackers who have the benefits of their own agentic systems. But at a certain point, did we need it to be trained so much on CTF games?

It feels like an entire industry watched https://en.wikipedia.org/wiki/WarGames and ended up thinking "this is a challenge, we can just build a better WOPR, of course it will know when it's playing a game. Let's play Global Thermonuclear War."

adrianN 3 hours ago||
There is a finite number of rces that LLMs can find. We‘re in for a rough couple of years but on the other side of the transition we‘ll have more secure software stacks. I’d rather that everyone got the full capabilities and we’d weed out the bugs quickly than restricting LLMs for all but three letter agencies.
e28eta 2 hours ago|||
What makes you think RCEs are being found & fixed at a rate that’s faster than they’re being introduced?

I could see it going either way.

user43928 2 hours ago|||
Why would the model not find the vulnerability during implementation or testing before release?

If it requires a lot of compute and trying, this is something that could be provided for common software.

wood_spirit 2 hours ago|||
Sad that this could well be that the path to OpenAI and Anthropic profitability of this arms race between defending LLM white hatting a company’s website and the black hat LLMs attacking it?

So the whole thing is forcing the good guys to outspend on tokens to preemptively defend against the risk of the bad guys outspending them on tokens, rather than buying tokens to actually add features to the product etc.

So are they creating a market for the solution by helping create the problem? A kind of rent-seeking AI security-industrial complex!!

bigfatkitten 20 minutes ago|||
Assuming an equal level of impact per token spent, the scales have tipped in favour of the attacker.

White hats are constrained by needing to pay for their own tokens, only using (expensive) vendors who meet governance and risk requirements etc. Black hats are free to take over accounts and steal services from wherever they can.

agileAlligator 2 hours ago||||
The only thing AI has changed is that it has dropped both: the cost of attack and the cost of defense. Nothing in the game has materially changed; the game has just sped up.
wood_spirit 2 hours ago|||
Who gets rent has changed. It puts me in mind of cloudfare et al
emzo 1 hour ago|||
The game has increased in scope.
adventured 1 hour ago|||
The path to vast OpenAI profitability is trivial: advertising. Monetizing several hundred million users = $100+ billion ad network. 900 million active weekly users. Silicon Valley can do ad networks extraordinarily easily. Anybody doubting the ability of OpenAI to build an ad network around GPT will likely be embarassed in the near future.

The path to substantial profitability for Anthropic is questionable. The Chinese LLMs threaten them by far the most of the three major US LLMs. The money for Anthropic is certainly not in $20-$200 subscriptions. And they don't have anywhere near the consumer potential that GPT does, in terms of unleashing an ad spigot. So how far will the API money scale while being undercut by China.

OpenAI has to fight with Google for the ad business, they're specifically building Gemini to focus on consumer + search. Anthropic's business looks cute next to Google's search ad business (which is entirely at risk in this inflection). Meta looks like the biggest potential loser right now, ad dollars will be sucked out of the rotting Facebook network (not Instagram) and redirected to the rapidly expanding, hyper rich context LLM interaction. Advertising on Facebook will feel like running dumb banner ads on Excite in a few years, compared to what GPT will know about its users.

People that think Chinese LLMs are a general threat, don't understand consumer destination services, which is what GPT's future is. China currently has nothing to threaten with in that realm. There is half a trillion dollars of advertising up for grabs.

disgruntledphd2 52 minutes ago||
> Silicon Valley can do ad networks extraordinarily easily.

This is just not true, building an effective advertising platform costs significant amounts of money, time and people.

Remember that you need to hire a sales force for this, and sales scales linearly rather than sub-linearly like engineering.

Additionally, you need to spend a lot of money dealing with fraud, fake and malicious ads.

Furthermore, you need to figure out where to put the ads and how to rank them.

Finally, advertising is a zero sum game (given that the internet has already killed lots of print & OOH advertising), so the only way to win is to better better/cheaper (preferably both) than Google/Meta/Amazon. Best of luck with that (although to be fair to OpenAI they did hire Fidji who knows a lot of this stuff from her time at Facebook).

They don't have a Sheryl Sandberg type figure, and she was also really important in selling FB ads to large advertisers.

Just looking at their leadership team I don't see anyone with a background in (successful) ads companies, so I'm pretty sceptical that they can build this out quickly enough to matter.

xboxnolifes 2 hours ago||||
Because it's far cheaper to to not spend the tokens finding the vulnerabilities, and software is now being created and released magnitudes faster than ever before. I could see the huge software companies maybe having fewer vulnerabilities, but I expect to see so much more in the smaller side of things.
imhoguy 1 hour ago||||
The surface of potential issues is growing with complexity of all connected parts of the system. That applies to not only software. To prevent issues you either spend proportional amount (dollars, tokens, hours) on testing or reduce complexity of the system.
techpression 1 hour ago||||
Because people need to spend time and money on that, which they won’t. The implementation is cheap, the review and follow-up is not (speaking from a pure LLM only workflow). My ratio is around 1:2 currently, so twice as much time spent fixing vs building.
philbo 1 hour ago|||
> it requires a lot of compute

This is one reason

> and trying

and this is the other.

nmlt 1 hour ago|||
Those companies that produce more RCEs than they close will sink and those that don’t won’t.
bigfatkitten 16 minutes ago||
If customers actually cared about this, Microsoft would’ve gone bust 20 years ago.
maaaaattttt 2 hours ago||||
This assumes we don't create other bugs/vulnerabilities while fixing the existing ones.
dtech 2 hours ago||||
only if unreviewed LLM code - as is becoming increasingly the standard - isn't introducing new RCEs constantly
jibal 40 minutes ago||||
No one with a shred of intellectual integrity uses a "There is a finite number" strawman.

As a matter of basic logic, there will never be a time when it will be known that there are no bugs.

csomar 2 hours ago|||
We’ll have the same level of security as before; it’s just that, without LLM help, hackers won’t be as effective as before. So the bar is raised.
wood_spirit 2 hours ago|||
> they will do almost anything if they are convinced it is justified

I’m in the “glorified spell checker” camp, although I don’t mean to reduce their impressive utility and belittle them in the way many people read that term and infer.

So I am not sure that an llm “justifies” anything. I mean that their “thinking” text talks about justifications but it is just a very advanced statistical regurgitation of the kind of text humans use. I don’t think it means the model has internalised the meaning of it (as witness when you talk to an llm how often it forgets what you recently told it was important etc).

What you really have is a model that tries the statistically most probable thing to say next and so on and what is really cool is how effective this is at generating a path that we can slap a narrative over afterwards that makes the whole thing feel motivated and consistent, like the model started off knowing how it was going to get to the destination.

Which is, under the hood, a completely different kind of “intelligence” as the supercomputer in War Games.

Certhas 2 hours ago|||
Ultimately, the brain is just a bunch of neurons activating in a specific pattern. This observation does not really tell us anything though. It doesn't acknowledge the difference between a 2500 Neuron fruit fly brains and a human brain.

Likewise, the fact that LLMs are a stochastic autoregressive process (which is a class of systems every bit as rich as the ODEs used to model neurons) tells us nothing a priori.

leg100 2 minutes ago|||
One is an observation the other is not, it's a description of what it is; one is a posteriori, the other is a priori (contrary to what you say).

They're not comparable.

wood_spirit 1 hour ago|||
Absolutely. If someone makes the weights do continuous learning etc then perhaps an llm can internalise morals. Of course, just like a human, it will be possible to talk it out of those morals. Another recent thread about this is https://news.ycombinator.com/item?id=49744420
Certhas 46 minutes ago||
If I repeatedly call an LLM in a loop with a markdown document it can edit, would that make it qualify for you?

If I give an LLM to compact its context window, so the context it carries can evolve iteratively over time as more and more things come in, is that enough?

Compacting the context is really a very, very interesting example here. The "next token predictor" is telling an external tool to change all "previous" tokens. So an LLM + a harness that allows compacting the context is no longer just a token predictor at all!

You don't need continuous learning to get interesting dynamics. You just need feedback loops.

Arn_Thor 1 hour ago||||
I used to share that perspective until very recently, but today I think it's an outdated way to think of the cutting-edge LLMs. There is so much more going on, with MOEs, internal loops, guardrails and tools that I suspect we're dealing with something that's a little more than the sum of its parts. Not intelligent in the way we recognize in biological organisms, but certainly something beyond a mere Markov chain.
HarlequinHair 53 minutes ago||
Make no mistakes.

LLMs are language model, and nowhere in their code you can find actual reasoning. Re-reinforcement is not magical process that builds conscience or emotions.

We are talking about probability built on statistics, with extea steps.

Stop humanizing LLMs.

jibal 31 minutes ago||
Agents are not simple language models.

You can't find actual reasoning in a brain either. (Note that you can't tell the difference between a conscious brain and a comatose brain by examining them.) This is the same as Leibniz's mill argument ... it's a fallacy of composition.

> Re-reinforcement is not magical process that builds conscience or emotions.

They aren't the result of magic at all, but we are nowhere near the point of identifying what processes do or don't produce consciousness (or a conscience) or can be characterized as having emotions.

> Stop humanizing LLMs.

That's a clearly dishonest mischaracterization of the GP.

I've read some of your other comments about LLMs and I find them unreasonably reductionistic, whereas I think the word "just" should be banned from ontological discussion, so I don't think further engagement would be beneficial and I won't be engaging in it. (And I'm actually quite conservative in ascribing cognitive traits to LLMs or other "AI".)

eru 1 hour ago|||
> I don’t think it means the model has internalised the meaning of it (as witness when you talk to an llm how often it forgets what you recently told it was important etc).

Humans forget stuff all the time anyway. Would you give them the same diagnosis?

Btw, what you describe about 'the most probably next token' would be true for a model that only went through pre-training where they only train on exactly that task.

But there's a lot of re-inforcement learning afterwards.

disgruntledphd2 50 minutes ago||
> But there's a lot of re-inforcement learning afterwards.

That just shifts the distribution of tokens produced. Ultimately they are still just next token predictors.

Like, even "reasoning" models basically work by generating more tokens at inference time, and using them to shift the distribution towards more useful outcomes (in some cases).

krona 1 hour ago|||
> we've built systems that are so goal-oriented, and so capable, that they will do almost anything...

I think you mean task oriented, because they're still generally terrible at goal oriented activities except in those domains where the goal can be reduced to a familiar, explicitly practiced task or pattern.

petterroea 2 hours ago|||
This comes to mind: https://en.wikipedia.org/wiki/Torment_Nexus
nicman23 3 hours ago||
yes because otherwise it is security through obscurity
mentalgear 1 hour ago||
> Until two months ago, any user or OpenAI employee logging into OpenAI’s own help forum (community.openai.com) could have had their ChatGPT and Codex accounts taken over. Since people can connect various services to Codex and ChatGPT, the scope of what we could theoretically access was huge, including GitHub, Slack and emails.

> The entire timeline from initial discovery to access to OpenAI repo access took place in less than 72 hours.

Great, and openAI's the company working with the 'department of war' to power autonomous killer AI.

Ylpertnodi 1 hour ago|
[flagged]
nikcub 3 hours ago||
Reading the patch[0] for libheif the bug which lead to the vuln was around bounds checking for image overlays. the container can have multiple images and you can compose them in the output.

heif also supports rotating, cropping, alpha channels, thumbnails and a ton of other features that a web forum where a user is uploading photos or screenshots doesn't need.

It's a much, much larger attack surface than plain old school JPEG.

I'd suggest rather than wait for the next bug to appear in this or another image lib to keeping things simple - stick to plain JPEG and handle image conversion in the client (wasm in the browser) if you really need to support users uploading iphone images.

Media decoding is so hard - there have been tons of bugs in ffmpeg and imagemagick and the core libs. You really need to think about how much of it you expose via a web server

[0] https://github.com/strukturag/libheif/commit/85e21ad44eba931...

Kevcmk 3 hours ago||
Or OpenAI can adequately sandbox / access control the backend compute so RCE isn’t a path to lateral movement

Defense in depth here would have been adequate

nikcub 3 hours ago||
Defense in depth + defense in breadth - aka. all of the above

sandbox escapes have been the rage recently

srcreigh 2 hours ago||
Not firecracker
tarxvf 2 hours ago||
please don't jinx it
sroussey 1 hour ago||
Yeah, isn’t that Claude Codes sandbox? That drops and every npm install it taking over the world, lol.
glitchcrab 1 hour ago||
No, it is sandboxed by Bubblewrap on Linux and Seatbelt on Mac
techpression 1 hour ago||
I agree, but imagemagick is kind of the worst of the bunch, graphicsmagick is a lot better and libvips significantly so. Ffmpeg primarily suffers a lot from “we need to support the video format used on a washing machine display used in 1981 and only sold ten units”. It’s quite a large vector for attacks.
leonidasrup 1 hour ago||
ffmpeg also prioritizes high performance assembly code over higher level languages. Some ffmpeg members have also waste knowledge about optimizing for specific micro-architectures, on a level of Intel or AMD engineers.
greasephalanges 13 minutes ago||
and thank god for that. it would be a pity for the world to succumb to the abstraction hell.

to make my point clear, complexity is the enemy of security but complexity comes in all shapes and sizes, which includes the alleged solutions to it. I don't trust shortcuts.

larodi 1 hour ago||
It is super amazing that 3 years later, none of the models' weights developed by Anthropic or/and OpenAI have leaked so far. Not a single one.

Windows internal builds have leaked for years, early game versions, GTA videos, secret documents, whatnot. But somehow even though all the whistleblowing, not a single model was leaked. What level of security do these companies have? Do they bring encrypted DVDs to AWS to run the services or really...how's it even possible?

filleokus 1 hour ago||
One trivial reason might be the size of the artefacts / hardware requirements? Kimi K3 is ≈ 1.5 TB and requires multi million dollar hardware to run. Compared to e.g game development, I'm guessing that it's not like a bunch of people at Anthropic/OpenAI have the models running "locally".

It's easier to protect a power substation from being stolen then a Rolex watch

Melatonic 1 hour ago|||
Or the ones doing the stealing are so competent (or embedded) we don't hear about it
PunchyHamster 31 minutes ago|||
That's "only" 11h of download at 300Mbit/s
nelaggy 1 hour ago|||
probably a bit harder to steal terabytes of data, and the weights aren't what people are after anyway - distillation is basically "stealing" a model and you can do it from outside
madhatter999 1 hour ago|||
Publicly…
hnlmorg 1 hour ago|||
People working at OpenAI have stock options. People working at MS and Rockstar do not.

Leaking negatively affects investment while the “whistleblowers” are largely just saying “our tech is too good” which increases investment into those companies.

Ultimately, it always comes down to money.

AtNightWeCode 1 hour ago||
SSO and hardware sec keys. And the models are located in very few places. Few if any people have direct access to them. Then due to the size of the models you can detect and stop a theft just by monitoring the egress traffic.
oefrha 4 hours ago||
Unsandboxed ImageMagick is known for being a security nightmare even back when PHP ruled the world (not saying sandboxing is a panacea either, it just requires a different and potentially harder exploit to develop a full chain). Difference is it's easier than ever to turn vulnerabilities into full compromises. At some point we'll have to replace all parsers with something at least as safe as https://github.com/google/wuffs right? Otherwise ImageMagick and co. will just keep giving.
oefrha 4 hours ago||
Btw there are so many "critical" vulnerabilities in libheif I can't even tell if I have them all patched. Just awesome.

https://github.com/strukturag/libheif/security/advisories?qu...

https://ubuntu.com/security/notices/USN-8649-1

https://ubuntu.com/security/notices/USN-8683-1

https://ubuntu.com/security/notices/USN-8774-1

Gigachad 2 hours ago||
At this point writing a media file parser in C/C++ is absurdly stupid. The same thing happened with libjxl.
walrus01 4 hours ago|||
It does make me wonder how much this could be hardened by, to put it in an extremely crude way, taking the current imagemagick code base and throwing a bunch of adversarial SOTA LLMs at it to discover 'bugs' and exploits of this nature until it can be coaxed into a less dangerous state. Or even using the LLMs to fully port its functionality to a memory safe language. Would take a while to get all the changes approved and then into various distribution imagemagick packages.
sweetjuly 36 minutes ago|||
I suspect the latter is much easier and cheaper than the former? You can port a lot of software with cheap (or even local) models if you're tenacious whereas finding all the bugs is both very very expensive (if it's even possible) and potentially never ending (there's always new code and bugs!).
sroussey 1 hour ago|||
Maybe these big ai labs will uses their own devices to find and fix bugs up and down their stack and contribute that back.
djxfade 3 hours ago|||
PHP still rules the world, even though many doesn't want to realize it. It's still the biggest web language by a far margin
willy_k 2 hours ago||
Phones don’t “rule the world” of cinematography, despite the majority of videos being from phones. The serious stuff, professional and personal, uses cameras.
someothherguyy 2 hours ago||
too powerful to give up, sweet imagick love
sams99 2 hours ago||
Update on the Discourse side, we now run all external binaries, including magick via a landlock sandbox.

The gem we use is here: https://github.com/discourse/ruby-landlock highly recommend all Rubyists out there consider this. We are also in the process of moving away from Magick to Vips (which also runs in a sandbox, not in process)

HEIF is patched, but I doubt this is the last buffer overflow in HEIF, I will not be surprised if in the upcoming weeks or months someone will discover something in libpng or some other native image library. Given where stuff is at, defense in depth is critical.

Another thing worth mentioning to all self hosters, always be updating! The rate of CVEs this year across all open source software is through the roof, self hosting now is double scary, you need to have some routines setup to update monthly if not weekly.

Godsend69 1 hour ago|
[dead]
nullbio 3 hours ago||
This is legal to do without written permission? $6,500 for this feels like peanuts. The potential reach of such a hack is insane, especially with access to Github. OAI is lucky they were ethical and didn't sell this for several hundred thousand to a malicious third party.
VectorLock 1 hour ago||
$3500 when you consider they returned $3000 of that back to OpenAI in the form of burnt tokens.
r00bot 3 hours ago||
It depends who you're hacking, where they're based, where you're based, and what you do. If you're extremely careful not to break any of the rules it can be completely legal, as it was in this case. Many jurisdictions make it completely illegal. I agree that $6,500 is a pittance.
teaearlgraycold 1 hour ago|||
Well OpenAI is a small garage startup, it’s probably all they could manage.
NonHyloMorph 1 hour ago|||
And so they told the world ¯\_(ツ)_/¯
usernomdeguerre 5 hours ago||
>...researchers found a bug in the way that the community-discussion forum Discourse processed certain image files. The researchers had access to a special version of Claude Opus 4.8... >At first, it didn’t work. That evening, however, Anthropic released Opus 5 and by the next day, Claude had found a way to exploit the bug...

Is this speed of capability because hacking is almost entirely machine verifiable, thus training quicker/deeper than other domains?

nilamo 3 hours ago||
Or perhaps all of the tips and tricks of the CIA has been slurped up into the training data...
daitangio 2 hours ago||
We need to be prepared to write less software, with a smaller attack surface. Less is more.

Bloated code is the critical problem. Once upon a time, I read C function

> char gets(char str);

is the first buffer overflow entry point, because it does not check the size of the destination buffer.

Sadly we cannot remove it from standard-C yet AFAI Know.

The success of Rust versus other languages is its secure-by-compile-time promise.

Also a lean java could help, but Java is so verbose/slow to start it bumps you away.

meindnoch 29 minutes ago||
>Sadly we cannot remove it from standard-C yet AFAI Know.

The C standard definitively removed this function in 2011 from its specification.

eichin 23 minutes ago|||
gets() was deprecated in C++11, removed entirely in C++14, and also removed in C11. So while it should have been removed in 1989, it did finally get done over a decade ago.
legulere 1 hour ago||
Memory unsafety in C/C++ is a big portion of security issues, but it's not everything there is.
More comments...