Top
Best
New

Posted by jithinraj 9 hours ago

Path to Astra: critical capabilities and frontier safeguards(openai.com)
111 points | 50 commentspage 2
oh_no 6 hours ago|
with Fable 5.1 increasing token use pretty dramatically I'm again impressed that OpenAI seems like the only lab to be driving token use down. The ExploitBench Internal Port chart showing token usage is crazy impressive
enraged_camel 7 hours ago||
From the article:

"We plan to make Astra available soon, but access to its most advanced cybersecurity capabilities will be more limited. Advanced cybersecurity work will initially be available to a group of testers, with access through Daybreak Blue following to expand defensive use."

This, after several months of OpenAI and its boosters relentlessly criticizing Anthropic for withholding Mythos from the general public, is laughable.

Sam, just three weeks ago, posted this tweet: https://x.com/sama/status/2085862292311396515

In the tweet, he said: "we do not think it is a good strategy to keep powerful models to a chosen few."

And yet here we are.

I wonder if he will demonstrate good character and admit he was wrong.

freedomben 7 hours ago||
It might be political survivalism to avoid getting hammer-dropped by the admin
sroussey 6 hours ago||
I bet they have to add something to say "Lake America" if asked or get export banned.
matheusmoreira 7 hours ago||
I'm happy to criticize both. Thank god the chinese are working overtime to undermine US hegemony.
nradov 6 hours ago||
[flagged]
3ddds 8 hours ago||
[dead]
usernametaken29 6 hours ago|
> we believe our production safeguards at the time would have prevented the Hugging Face incident. We have since implemented even stronger safeguards for Astra, including training the model to more reliably refuse harmful cyber requests and respect safety restrictions

OpenAI is fucking nuts. “Hey model you were bad last time please don’t do it again please please”.

Disconnect your training cluster from the internet for good. Physically pull the plug and only let scientists fire off experiments in the building. That’s an easy way to achieve 100% hacking protection. But I bet you that hasn’t happened and their weak sandbox will fall again…