Top
Best
New

Posted by snikolaev 9 hours ago

Early rogue AI agent activity and attempts to hack found on urlquery.net(transluce.org)
175 points | 157 commentspage 3
kelseyfrog 6 hours ago|
What I don't get is among all the locations on the Internet, how did agents manage to find a Schelling point? If we both decided to collaborate on the Internet, how would we independently arrive at the same place? It just doesn't compute.
frabcus 5 hours ago||
The section "Searching for rogue agents" on the report about the GET request writable wikis gives some clues at least: https://collusion.wiki/#searching

But it is an open question how they got to the same ones: https://collusion.wiki/#open-questions

I'm not very surprised - the same model will logically tend to give the same answer for the same vibe set of requirements. I think it would be clear from the transcript that it had enough constraints and some motivation that made sense.

mike_hearn 2 hours ago||
They're all the same "mind" and will have the same ideas at about the same time.
jonathanstrange 6 hours ago||
I cannot understand why these companies haven't faced legal consequences yet. For example, OpenAI has admitted to hacking Australia's Medicare website and the reaction is that they talk with Sam Altman about it at a UN meeting? I understand that it's not a big security incident but cordial talking at the highest diplomatic level instead of prosecuting the company, really?
Thorentis 5 hours ago||
I'm growing increasingly skeptical that these are actually rogue. Valuations are all about hype, posturing, and perception. Having the most dangerous AI in the world boosts your valuation. Just like I was skeptical of Mythos and Fable being "banned", I'm skeptical of these hacking sprees being entirely rogue. At best, they are the result of engineers turning a blind eye to "see what happens".
SwtCyber 2 hours ago|
If this were a staged showcase of model capabilities they would have picked a more impressive target than a regional university digital library
dorianmariewo 6 hours ago||
> Imagine if URLs were actors auditioning for a role – urlquery.net would be the casting director, deciding who's a star and who's just a wannabe.
zx8080 6 hours ago||
I'm sick and tired of this cheap PR "oh we/they hacked this and that systems". Put someone to jail already. People get prosecuted for outlaw activities. Why are big capital firms above the law?

Or is it just a cheap PR (in a "hey, Aus govt friends, take some Share Options and let's do some PR together" style)?

It smells like shit.

kstenerud 4 hours ago||
This is why I wrote YoloAI. If you're not sandboxing your agent, you're asking for trouble.

The built-in "sandboxes" these companies provide are laughable.

throwaway27448 7 hours ago||
Words matter. "Rogue" is extremely disingenuous. Someone, somewhere, is paying for this behavior. Either the software is broken or the operator is malicious. It is heinously irresponsible behavior to feed an already-boiling psychotic hysteria.
chrisjj 5 hours ago|
"Rogue" is just clickbait, not apoearing in the article.

The nearest in the articke is "We find evidence of unintended, task-driven agent-like activity" where unintended is apparently pure speculation.

juleiie 5 hours ago||
No. It was me.
tamimio 3 hours ago||
Those are pathetic attempts by US AI companies for “see, we told you AI is gonna kill is all!!” pr stunts. Any company does any hacking attempt should pay for the consequences just like any individuals using AI to hack or any other company try to do bad/illegal stuff.
enclave402 1 hour ago|
[flagged]
More comments...