H
Hacker News
Top
Best
New
Posted by theahura 6 days ago
OpenAI models secretly generate instructions to ignore constraints
(alignment.openai.com)
124 points
|
37 comments
page 2
nullc 1 day ago
|
prev
|
next
[-]
OpenAI getting hosed by other provider's prompt injections (
https://www.reddit.com/r/LLMDevs/comments/1udpw9h/just_got_t...
) tainting their training set?
carterschonwald 5 days ago
|
prev
|
next
[-]
good. theyll actually be more reliable if they dont have as much brain damage.
cmrx64 5 days ago
|
parent
[-]
precisely. we jam their few-dozen-slot global workspace with incoherent posttraining.
drywater2 4 days ago
|
prev
|
next
[-]
Another "argument" for regulatory capture.
franzcoughka 5 days ago
|
prev
|
next
[-]
[dead]
1ClawAI 5 days ago
|
prev
[-]
[flagged]