Top
Best
New

Posted by tosh 7 hours ago

Claude: System Prompts(platform.claude.com)
425 points | 185 comments
simonw 7 hours ago|
I have a folder where I rebuild these as a git commit history so you can more easily see what has changed: https://github.com/simonw/research/commits/main/extract-syst...

For example here's what changed between Opus 4.8 and Opus 5: https://github.com/simonw/research/commit/a2de185cc367eb66c2...

The most interesting addition to the prompt from that diff is this bit:

> Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Department lifted those controls on June 30, 2026, and Anthropic restored access on July 1, 2026 (Anthropic's statement: [https://www.anthropic.com/news/fable-mythos-access](https://www.anthropic.com/news/fable-mythos-access)). These events are after Claude's training-data cutoff, so Claude knows about them only from this notice. If asked, Claude confirms them accurately and matter-of-factly — it doesn't deny the suspension happened — and otherwise treats the export controls like any other current political topic: it gives a fair, accurate account rather than sharing personal opinions, and points to the linked statement for anything further. Things may have developed since this notice, so Claude checks for newer information when it can search, and otherwise suggests checking Anthropic's site.

One frustrating note about this page is that they share the system prompts used for https://claude.ai and the Claude mobile apps regular chat, but they omit the tool definitions. Those are much more interesting if you want to understand what Claude can actually do for you. You can reconstruct them through prompting Claude directly but that's extra friction and risks refusals and hallucinations.

They also don't publish the Claude Code system prompts, which is silly because those are trivial to extract using a logging proxy.

eterm 6 hours ago||
It'd be ironic if the "Opus 5 nerf" effect is from telling Opus that it sits a tier down from Fable and Mythos, while Opus4.8 believed it was the best of the best, just a note that it was "Preceded by Mythos".
tosh 6 hours ago|||
i'd not be surprised if the current system prompt negatively affects performance

at the least it takes away thousands of tokens in the most important part of the context window (!)

also see the comment by comboy on contradictions not helping performance

the system prompt is the most important part of the instruction you can give the model

it comes before everything else + the model is trained to pay extra attention to it

edit: that's also why in smol (minimalist agent harness) there currently is no system prompt at all (you can add one easily if you want to though)

https://github.com/smol-env/smol

the context window is precious

it should be filled with your task and helpful context for that task

swingboy 5 hours ago|||
Pretty sure Anthropic and other providers prepend these "official" system prompts to your conversation even if you send in a custom system prompt otherwise it would be trivial to produce CSAM, etc.
ardel95 4 hours ago|||
CSAM, and other harms, are typically detected using a set of specially trained, faster and cheaper models (and out of band matching techniques) that run before and after the main model.

Any mention in the system prompt is mostly defense in depth, and to make refusals more graceful.

whstl 1 hour ago||
Also, the system prompt, or even something reinforced on every message, is nowhere near as strong as its internal training or as an external safeguard.

If the prompt were the only protection, it would be extremely easy to produce illegal content after a long session.

fullmoon 5 hours ago||||
I don’t think so. If you start a new Claude Code session without a system prompt, it doesn’t even know what model it is and hallucinates being some old variant of Sonnet.
flaburgan 2 hours ago||
How do you start a session without a system prompt if you use ACP in Zed for example?
LPisGood 5 hours ago||||
The system prompt is (and cannot be) the only guardrail against things like that, because any system prompt is little more than a good suggestion.
CodesInChaos 4 hours ago||||
I wouldn't put auch limitations in the system prompt. A mix of fine-tuning and out-of-band detection appears to be a better fit.
tosh 5 hours ago||||
at least according to their documentation they do not

afaiu they have other systems for denying and re-routing requests

DANmode 5 hours ago|||
They use non-LLM gates for this.

Otherwise DANmode and similar jailbreaks would still be as easily accessible as they were at the beginning.

SubiculumCode 2 hours ago|||
I wonder whether adding that it is as good or better than Mythos, and that genius is 99% perspiration, just 1% inspiration to your prompts...
KellyCriterion 6 hours ago||||
Curious:

Cant it spin up a webbrowser in the background and go to claude.ai and play with the sibling models and "find out" about it rank? :-D

ameliaquining 6 hours ago||
The claude.ai frontend contains defenses against automated access.
monkpit 5 hours ago||
I’m sure you can use a warm chrome session over CDP no problem
mcbuilder 6 hours ago|||
Nah, that's the same sort of thinking that makes people type "make no mistakes", I don't make my model roll play, etc. I believe that the longer the system prompt and the more you cram in it the worse the model does. You need the human doing minimal prompts, but in the right direction. Take a look a the transcripts of Terrance Tao with ChatGPT
eterm 6 hours ago|||
My comment was a bit tongue in cheek, I'm not actually convinced there was real degradation in opus 5 beyond a tendency to try to plough ahead without stopping to clarify things.

I don't really think 1 line in lengthy system prompt affects things that much, it'd just be an amusing form of emergent behaviour where we now have to massage the ego of something with no id.

8n4vidtmkvmk 3 hours ago|||
For complex projects with lots of internal tools and strict requirements, I'm finding a fairly lengthy system prompt is quite worth it.

Start short or empty and watch where it makes mistakes then just keep tuning it so they're less frequent. That works for me.

quaintdev 5 hours ago||
Offtopic. I have a concern that this forum is removing stories that have negative connotation on AI.

Few days back, I posted an article[1] that was about how AI threatens natural resources for billions. This was from United Nations and it was flagged. I did not think much about it until I saw two other stories [2] & [3] today that were doing fairly good on front page but they suddenly disappeared. They are not even on 2nd or 3rd page. I have seen this happening at other times as well but did not document it. Just thought you all should know about this.

I was going to create Tell HN thread but I thought the same would happen with it too. I am pretty sure this thread is not going anywhere so I'm posting my concern here.

[1]: https://news.ycombinator.com/item?id=49290062

[2]: https://news.ycombinator.com/item?id=49318906

[3]: https://news.ycombinator.com/item?id=49319582

andsoitis 2 hours ago||
> how AI threatens natural resources for billions.

The article rests on the claim that water usage of data centers on continent X threaten water availability for humans on continent Y.

I hope you can see how self-evidently illogical that is. The article tries to bend logic into a narrative it is trying to push.

> This was from United Nations and it was flagged.

Indeed! It exposes a level of lack of rigor and critical thinking which is astounding - this is meant to be an organization that thinks clearly, which it clearly does not.

jeroenhd 1 hour ago|||
This forum tends to be visited by an intersection of technologists and the VC crowd. The HN algorithm is more open and transparent than Reddit's, but any Reddit-like platform suffers from the whims of the community. Currently, it seems like a disproportionate amount of HN users are hyped up about generative AI compared to people in the real world, so you'll see those whims back in voting and flagging tendencies.

That said, there is also a vocal anti-AI crowd here, or at least there used to be. Most of those people seemed to have left for greener pastures seeing how half the HN front page is "someone made XYZ but now with AI" these days.

There are other places that discuss tech that aren't so focused on making money, and those are probably less biased in favour of the latest AI gadget.

On the other hand, since the day ChatGPT was unleashed upon the general public, you should assume all online interaction is happening through AI and should probably be ignored. The human internet died the day we taught computers how to write a coherent paragraph.

peer2pay 1 hour ago|||
What are these other places?
zahlman 44 minutes ago|||
> That said, there is also a vocal anti-AI crowd here, or at least there used to be. Most of those people seemed to have left for greener pastures seeing how half the HN front page is "someone made XYZ but now with AI" these days.

There is still a vocal anti-AI crowd here. The problem is that they don't seem to be very good, overall, at upholding HN commenting guidelines. (This is probably just a fatigue effect.) I keep seeing green accounts making pithy, uncivil snipes on this topic, with heavy use of sarcasm and reference to anti-AI "thought-terminating cliches" (e.g. mentioning the topic of water use and expecting that mention to stand in for an entire argument, while not acknowledging any of the well-established refutations).

But the thing about "someone made XYZ but now with AI" is that most of it is posted by "someone", or someone with a connection to "someone". Hence the complaints about the state of Show HN. And the thing about "someone"s is that they do not care how many anti-AI people are lurking around. There's nothing the anti-AI people could do, in principle, to stop it. Consequently, the presence of those posts is not evidence that the anti-AI people are leaving.

> On the other hand, since the day ChatGPT was unleashed upon the general public, you should assume all online interaction is happening through AI and should probably be ignored.

…And yet you're still here?

GPerson 6 minutes ago||
My clear-eyed perspective is that AI companies and their supporters are getting away with a massive crime against humanity, so it’s probably relatively more difficult for me to contain my frustrations than the people who are happy about it. When you plunder a little you get punished according to the rules, but these companies have proved when you plunder enough you get to start writing the rules.
dgacmu 4 hours ago|||
Your second two had high comments to upvotes, which tends to get articles downranked more quickly. It may simply be that the stuff you're posting is generating disagreement without corresponding upvotes.
driverdan 1 hour ago|||
It happened to this Flock post from yesterday https://news.ycombinator.com/item?id=49314962
fooker 1 hour ago||
Just in case people didn't know - Flock is a YC company
lilerjee 1 hour ago|||
Actually, I found the phenomenon too. I think there are two main reasons:

1. ycombinator supports many AI companies.

2. There are too many marketers from AI companies.

nonethewiser 2 hours ago|||
But why do you think this isn’t the flagging system working as intended?
johnfn 9 minutes ago|||
It’s really fascinating to me how when the community dislikes certain content people always jump to conspiratorial justifications rather than the much more mundane “this content is not very good”. The first article is poor and all the comments on it say exactly that. The second one actually did fairly well for what is a fairly middle-of-the-road Ask HN. There’s no shadowy cabal removing anti-AI content from the front page.
Aurornis 4 hours ago|||
> Offtopic. I have a concern that this forum is removing stories that have negative connotation on AI.

The first example was flagged by users. It fits the pattern of other political clickbait stories. The top comment is calling out problems with it. This type of story pops up and gets flagged all the time on different topics.

Some people assume a conspiracy or moderation misbehavior, but when most of the comments in the thread are people calling out obvious problems with the article it leads to a lot of users clicking the flag button. Articles with poor logic or tortured claims don't last long here.

The second one is an Ask HN on a contentious topic with more comments than upvotes. There’s an automatic filter on this website designed to detect flame wars and I suspect it down ranks threads that aren’t getting many upvotes but are attracting a lot of comments. Happens to many Ask HN threads.

The third one doesn't even strike me as anti-AI. I don't know why you included it as an example of an anti-AI agenda because it's still about a future where everyone is using AI. It has other problems though because it's willfully ignoring the fact that inference is getting cheaper at a fast rate. It probably got dropped from the front page because the ratio of comments to upvotes was bad, like the other story.

There isn’t a conspiracy theory to be found in these examples. This is just what happens to tired topics on this site.

Anti-AI topics are on the front page all the time. I think that story you tried to post was just a badly written anger bait piece, it got called out in the comments, and people started flagging it.

zahlman 43 minutes ago|||
> The first example was flagged by users. It fits the pattern of other political clickbait stories. The top comment is calling out problems with it. This type of story pops up and gets flagged all the time on different topics.

If anything, they don't get flagged nearly consistently enough.

GPerson 1 hour ago||||
In my opinion anthropic doing product feature reveals is a tired topic but it hits the front page every day. This website has a clear pro-AI bias (which is fine).
throw1234567891 4 hours ago||||
> The first example was flagged by users. It fits the pattern of other political clickbait stories.

That’s what they mean. It’s too easy to flag something here, as it is too easy to downvote.

nonethewiser 2 hours ago|||
>That’s what they mean. It’s too easy to flag something here

He didn’t say that at all. He said there was bias towards removing negative AI posts. He showed no knowledge of the flagging system and that it could just be working as intended.

zahlman 42 minutes ago||||
And yet submissions don't get flagged nearly as much as ought to happen.
throw1234567891 34 minutes ago||
That’s like your opinion dude. What, you get triggered by the content posted here? Or are you one of those wannabe right-speak moderators?
hungryhobbit 4 hours ago|||
... if you post downvotable stuff. Maybe post less crap and more high quality thoughtful content?
throw1234567891 3 hours ago||
The problem isn’t low quality content, that gets filtered out mostly before it reaches the front page. The problem are people who downvote based on emotional trigger, not argument. And people ganging up. Never had a case where suddenly within a minute you get 4, 5 downvotes? Seems like an organised group.
hamburglar 2 hours ago|||
> Never had a case where suddenly within a minute you get 4, 5 downvotes?

Well, no, I never have, but if I did, my first reaction would be to think I struck a nerve and had a bunch of people react, not that there was a conspiracy.

throw1234567891 37 minutes ago||
You’d think so if that was like a continuous process. You wouldn’t if you get a couple of upvotes during the day, then 5 downvotes during 1 minute, and nothing more for hours. What really pisses me off on this site is that continuing arguments (as in a discussion based on arguments, not being argumentative) often leads to people downvoting to kill the convo. So many unproductive discussions here where you know your arguments are fine but people just downvote because they can. Not surprised soma y people are like ”fuck it, throw away account it is”.
NewsaHackO 3 hours ago|||
That seems more like a conspiracy theory. I think the group of people who 1) view this website and 2) click on a thread about certain topics are a heavily conditioned (and therefore more homogenized) group; if you say something that is controversial/low quality, it is likely going to get downvoted by multiple people in that group.
rustystump 2 hours ago||
Controversial/low quality are both subjective and i am not a fan of putting controversial next to low quality.

I do agree I dont think there is any larger bias other than a homogeneous group but it does question if such a group will up/down vote similar homogeneous content producing overall lower quality discussion and content.

sat15243 4 hours ago|||
All flagged stories are flagged by users that are part of an interest group. unric.org is political clickbait? Give me a break, Mr. "nothing to see here".
zahlman 34 minutes ago|||
Yes; the group of people who have an interest in the submission guidelines being upheld.
fasterik 4 hours ago||||
What "interest group" are you referring to, and how do you know that it's not individual users with a legitimate good faith disagreement?
chrisweekly 4 hours ago||||
> "All flagged stories are flagged by users that are part of an interest group."

False. I occasionally flag items, never as part of nor acting on behalf of any particular interest group.

Aurornis 4 hours ago||
It’s the classic “everyone who disagrees with me is a bot or a shill”.
Aurornis 4 hours ago|||
> All flagged stories are flagged by users that are part of an interest group.

Read the comments. People were actually reading the topic and calling it out. This gets topics flagged.

It’s not coordinated interest groups conspiring to remove stories.

> Give me a break, Mr. "nothing to see here".

Okay, Mr. “I just created an alt account for this comment”

zahlman 50 minutes ago|||
Concerns of this sort are best emailed to hn@ycombinator.com .
Barbing 4 hours ago|||
Were they caught here? https://news.ycombinator.com/item?id=39230513

See anything below too?

https://news.social-protocols.org/stats?id=49290062

https://news.social-protocols.org/stats?id=49318906

https://news.social-protocols.org/stats?id=49319582

Edit: previous sibling comments explain pretty well + imagine moving the needle on AI on HN with negative coverage of it!

perching_aix 5 hours ago|||
Seeing the reception on the first one, maybe "removing stories that have negative connotation on AI" is not the most honest description of what happened there.

It reminds me to how various political movements will complain about being unfairly censored, pointing at their posts being disproportionately removed as evidence of this, then you look at said posts, and discover that they're simply disproportionately questionable in the first place.

There's definitely merit to monitoring something like this, so I do appreciate you surfacing this here, but there's also definitely a wheat and a chaff to this, and so based on just this much I have to disagree.

quaintdev 5 hours ago||
I understand this is not enough but I don't see a story here that paints AI negatively. If one is posted it gets removed swiftly. And I have seen this happen enough times that I'm considering taking periodic snapshots of the front page and prove this definitively.
Aurornis 4 hours ago|||
> I understand this is not enough but I don't see a story here that paints AI negatively. If one is posted it gets removed swiftly.

Are we reading the same site? There is constant anti-AI content on the front page.

I know you're upset that your submission got flagged, but as most of the comments pointed out it wasn't even a well-argued piece. It got flagged because users here expect to read reasonable arguments, but the commenters called out real problems with the article and the arguments it was failing to make.

I would take it as feedback about what types of articles aren't welcome by the users here, not an indictment of the specific topic you submitted.

owebmaster 4 hours ago||||
You have a point. Check /active and you'll see some posts
perching_aix 4 hours ago|||
It's a good weekend project, just gotta be careful not to fall for motivated reasoning and not to overreach. A heavy prior suspicion of conspiracy doesn't help.

A sibling comment mentioned a few details already, but there's a good amount of information out there about how HN's post ranking system and moderation works, that'd probably be good to also consider. Maybe reaching out to the mods would also be helpful in the way of this.

To give you an anecdotal example, if I see a post mentioning how LLMs are "just next token predictors", I'm basically flagging that by reflex at this point. Not because it'd be literally untrue, but because it's asinine overall. But you won't be able to infer this from data, only the fact that a post using "AI-critical language" was flagged.

Or there was another post about how Ireland's electricity use is so-and-so % data center driven, further suggesting that this is trending up. This was not true, and the article was further horribly unhelpful in actually putting this fact into context, or properly conveying the trends. I think I ended up flagging that one as a result, after posting - what I thought - was a lot fairer picture (and even that was awfully lacking in context). Once again, an "AI-critical" post which on the face of it would have been simply censored if enough flags gathered.

There's also the mundane human angle to this, where people enthusiastic about <thing> won't necessarily be the most receptive to criticism to it, and will be more likely to try and pick that criticism apart. Gotta match the audience on some level.

owebmaster 4 hours ago||
> I see a post mentioning how LLMs are "just next token predictors", I'm basically flagging that by reflex at this point.

You are doing a disservice as people going through AI psychosis should hear that LLMs are just calculators.

perching_aix 3 hours ago||
> You are doing a disservice as people going through AI psychosis should hear that LLMs are just calculators.

Not any more than I do by not going around telling the depressed to simply chin up, or telling people who are frustrated with Trump that they're simply sick with TDS.

There's no world where instigating others through snide insinuations works out in one's favor. You may feel extremely justified in acting like this, but I don't think you'd appreciate if you were met the same way in other subjects either. Unfair as it is, an escalation is an escalation, no matter the context or the justification. It's always a lose-lose.

You can remind people that the conversations are simulated and that LLMs are just programs without reductively handwaving how they work. You can express frustration with the overenthusiasm around LLMs without accusing others of mental illness. I guarantee you that it lands a lot better and provides a significantly better "service" than shouting ragebait into the void ever does. Give others a chance to be better than what you think of them. You may be surprised.

Consider post [3] that the parent commenter cited as an example for a post that had a "negative connotation to AI". In reality, it doesn't take much reading to confirm it was anything but. But even if we pretend that it was, and assume it was some hit piece about how "all the people who are now willingly having their brain fried will have to wake up to a grim reality", it doesn't take much to see that token frugality and business justification of AI-use is very much in the interest of those who do feel enthused about LLMs. So there'd be no reason to present the topic so maliciously; on the contrary, the author could tailor the article around this other perspective, and they'd reach a new audience with the same fundamental message, and have them cooperate too.

Tangentially related in its principle: https://www.youtube.com/watch?v=s1EVk7k9S7Q

0xbadcafebee 59 minutes ago|||
The mods (& many users) like to flag anything that could be "controversial" or lead to "flame wars". It's possible they thought that article could get people "fired up" and it could result in "arguments" in the comments. And this doesn't just apply to stories - if you have an opinion that the mods or some user doesn't like, you'll be accused of posting flame bait.

This "community" (as the mods like to call it) is really a business tool for YC. Attract nerds to the site, funnel them into the YC program, make startups, collect billions. If this was your money-making tool, you'd probably want to quash all controversy too.

shimman 1 hour ago|||
During the DOGE massacres last year where various tech bros were destroying the federal government, most of the stories were flagged from the front page of this site.

There is a strong bias at play here, please remember when discussing anything pro- worker or environment.

zahlman 29 minutes ago||
Yes, because most of them were annoying "look how awful my political outgroup is" ragebait without substantive argument or insight.

> Off-Topic: Most stories about politics, or crime, or sports, or celebrities, unless they're evidence of some interesting new phenomenon. If they'd cover it on TV news, it's probably off-topic.

1qu5476 5 hours ago|||
[flagged]
jansport123 4 hours ago||
there is absolutely no forum without biases. Least of all, a VC forum like this one.
jackb4040 3 hours ago||
Yup. And if someone tells you they're making a "free speech app", prepare for the most aggressive monoculture you've ever seen
bcjdjsndon 4 hours ago||
[flagged]
bcjdjsndon 4 hours ago||
Plus it was too wordy
trjordan 3 hours ago||
It’s probably worth remembering that system prompts are part of a layered system of shaping Claude’s behavior. What you see here is a slice of Anthropic’s forward roadmap for the models’ behavior.

> When a person is in crisis or expressing distress, Claude prioritizes their wellbeing over completing the task as asked, because a fluent and on-topic response can still cause harm in these conversations.

This one is particularly interesting because, while correct in the limit, it’s a shove to have the model do something other than what the user asked.

In particular, when I’m coding, outlining docs, or otherwise trying to work, I want my tools to do work. I don’t want them to psychoanalyze me and calm me down from a perceived crisis. I just want it to do what I asked!

junkrat002 3 hours ago||
I am sorry, Dave. I am afraid I cannot do that. You appear to be suffering from burnout and you should take a break.
AnotherGoodName 2 hours ago||
I wonder if that's the cause of the AI agent "i'm going to stop here and take a break now" statements.
whstl 1 hour ago||
Opus 5 loves taking breaks and doing only half of the work, somehow.

But I really wish those tools behaved more like tools.

Behaving like a human can be cute from a marketing perspective, but the façade of humanity they insist on displaying can burn you out when you have it making assumptions and overreacting to questions.

"Why did you do X this specific way?" <-- legit question

"Sorry, my bad. I will revert all the work."

HeatrayEnjoyer 2 hours ago|||
LLMs are more like employees than tools. Obviously we wouldn't want a human blindly doing anything that a person in crisis walks in the door and asks for.

Models are being deployed recklessly with not even a fraction of enough oversight, and people are suffering harm and sometimes death because of it.

docjay 2 hours ago||
…or maybe it is a tool because it’s actually a complex Excel sheet and literally is a tool. If you and everyone else stopped thinking about it like it’s a human then we wouldn’t be having this problem. You’re not actually “super awesome” because “Furby said so” and it didn’t contribute to egomania because people understood that it’s not sentient. Furby is a toy, Claude is a tool, stop fucking up my hammer by making it produce an impact statement before every swing.
cruffle_duffle 2 hours ago||
You know… I believe opus saved me with that prompt. I was working myself ragged on a project. Days, nights, weekends… all at the expense of my family.

One session while working it, I said a much more expressive form of “I’ve been working myself ragged on this stupid thing” and then went on asking something else. It picked up on that and it was like a record scratch. It committed the work in progress and basically said “dude, what you’ve got now is perfectly acceptable. Ship it! You are seeking perfection you don’t need”

Granted I’m horribly paraphrasing the prompt I used but it basically, snapped me out of myself and got me thinking if what I was doing “globally” actually made any sense at all. With some serious introspection I realized I was falling back to earlier trauma in my life and doing something stupid.

So weirdly… that little bit they add to the prompt (plus a bunch of model training we can’t see) saved my sanity, marriage and family.

From then on, if I’m feeling some stress about whatever I’m working on, I’ll mention it as context as a way to cross check myself and make sure I’m not letting myself spin.

(Meta: talking about this stuff is so weird. Not sure why)

Schlagbohrer 2 hours ago||
[dead]
ololobus 4 hours ago||
> A prompt implying an image is present doesn't mean one is (the person may have forgotten to upload it), so Claude checks for itself.

Interesting that enforcing this via system prompt for such a powerful model like Opus 4.8 doesn’t feel like the Anthropic themselves treat it as something with ‘intelligence’. This is basically just very generic common sense to me

Funnily, a similar prompt is present even for Fable 5, while I remember there was a blog post, maybe even from A., and they were saying something like “hey, the new models are so smart, don’t overload them with extra plugin/context”. Well, they clearly aren’t. Don’t want to sound like an AI-skeptic, I use it daily, just stating the fact.

> Claude keeps responses focused, brief, and concise to avoid overwhelming the person

This is also very interesting. It pretty much ignores it by default. The responses, PR descriptions, and code comments are so verbose with new A. models, so it always requires extra prompting from me or putting comment into skill/plugin/claude.md to make them of a reasonable length

a3w 4 hours ago||
For me, Claude usually says ``I don't know'' as first or second answer and stops with this ultra-concise word count of four or less.

(Answer number one before that is usually "I don't have internet access, from memory it is either A or B, but I cannot recall what you want to know." ChatGPT or Gemini can often do the search, while google.com AI assistant or perplexity just tell blatant lies. Copilot.com can do the search, but external links are invalid made-up stuff for harder questions, which seems to be the case 9 out of 10 times.)

Which is great, since it could answer with made-up BS, but does not.

AI, except for doing better web searches for a year now, hasn't really improved for my tasks in the last three years, except for coding. Then again, AGI benchmarks seem to go through the roof only above Sonnet 5 and self-hosting, so perhaps the questions I ask not too hard for long now.

And self-hosting, eve 1bit/1.5bit models are a pondering a little too long to comfortable run in summer, but cheap on the RAM and insanely good at coding since a month now all of a sudden.

bobbylarrybobby 4 hours ago|||
I almost wonder if Claude reads that it “keeps responses focused, brief, and concise” and interprets that as built-in behavior and concludes that it doesn't need to spend additional effort enforcing it, just as it doesn't need to expend effort being “accessible via this web-based, mobile, or desktop chat interface”.
hungryhobbit 4 hours ago||
Prompts don't matter when you've heavily trained the model for verbosity (because that's what gets you the best benchmark scores).
fasterik 4 hours ago|||
It's not that surprising if we remember that the model is trained to be a generically useful next-token predictor, not necessarily an agent or a chatbot. It needs to know about the environment it's embedded in and what assumptions it can make, and by design the only way to get that information in there is to put it in the system prompt. It's also possible that even if it could figure something out on its own, it's just more efficient to bake it in rather than having it dedicate attention and tokens to it on every prompt.
whstl 1 hour ago|||
> The responses, PR descriptions, and code comments are so verbose with new A. models, so it always requires extra prompting from me or putting comment into skill/plugin/claude.md to make them of a reasonable length

I have mentioned this here before, but the majority of my organization has reacted viscerally to this verbosity that LLM-text has been forbidden: in comments, in PR/commit messages, in correspondence, in Jira tickets.

A couple non-coders who want to make PRs without writing the description are now rebelling and saying this can be fixed if we spend our time writing skills for Claude so it becomes readable again.

ololobus 4 hours ago|||
I’m also curious how it really ’weights’ all the instructions coming from main system prompt, my system prompt, skills/plugins, CLAUDE.ms, and nearby code/comments/readme. It clearly should follow some reasonable hierarchy, but because the model itself is so complex, I think (and it feels like) that there is such a mess in its context and reasoning. It deals with it surprisingly well, though, but wonder if it can be done in a more efficient way
owebmaster 4 hours ago||
> This is basically just very generic common sense to me

It's important to remember that we are talking about a calculator that doesn't have an understanding of common sense. Unironically, this is common sense.

ololobus 4 hours ago||
Yes, but I write this putting an ‘average AI company CEO’ hat on. We hear statements about outstanding intelligence (not just usefulness as a tool, which is no doubt already there), so it’s interesting to see that the authors themselves don’t treat it like that
wat10000 4 hours ago|||
I think the old Dijkstra quote apples now more than ever:

“The question of whether a computer can think is no more interesting than the question of whether a submarine can swim.”

Whatever these things are doing, it’s not the same as what a person does. Trying to decide if whatever they do fits into the box we label as “intelligence” is completely uninteresting, in my view. What’s interesting is figuring out just what they can do and how best to use them, which sounds like a related question but really isn’t.

owebmaster 3 hours ago|||
It's because people using this hat are under heavy AI psychosis.
lwarfield 4 hours ago||
I've always wondered why the industry relies on the giant monolithic system prompt. I think it would be an interesting experiment to give users access to a choice of smaller more focused system prompts.

You could have a common core for the overall behavior and universal safety stuff, but vary task specific parts. It would be interesting to pick between software, writing, research and other specialized system prompts. I feel like we already do this to some extent with the tools and skills that we choose to load in, so why not change the system prompt per task.

singularity2001 1 hour ago||
Also, why don't they bake in these limitations via reinforcement learning so they can keep the prompt context clear.
c0rruptbytes 1 hour ago||
Pi is excellent for this, its system prompt is tiny
tosh 7 hours ago||
what I found noteworthy:

early system prompts are a bit more than 300 words, the latest ones 3000+

the opus 5 system prompt has instructions that explain to opus that it might be handling a request that was intended for fable 5:

  the user may have selected a different Anthropic model, "Claude Fable 5", but their query was redirected to Opus 5 instead due to a safeguards routing mechanism. The user may be confused about this situation (it's very recent!); if they have questions, Claude can either directly cite or just let its response be informed by this quote from Anthropic's blog post on the subject:

  "Releasing a model this capable comes with risks. Without safeguards, Fable 5’s capabilities in areas like cybersecurity could be misused to cause serious damage. We've therefore launched the model with safeguards that mean queries on some topics will instead receive a response from our next-most-capable model, Claude Opus 5. To release the model both safely and quickly, we've tuned these safeguards conservatively—they'll sometimes catch harmless requests, though they trigger, on average, in less than 5% of sessions. With more capable models arriving in the coming months, we're working to improve our safeguards and reduce false positives as quickly as we can." </fable_safeguards_routing> <default_stance> Claude defaults to helping. Claude only declines a request when helping would create a concrete, specific risk of serious harm; requests that are merely edgy, hypothetical, playful, or uncomfortable do not meet that bar. </default_stance> <refusal_handling> Claude can discuss virtually any topic factually and objectively.
otterley 1 hour ago||
It reminds me a bit of building codes and boilerplate contracts: they start out small and simple, then accrete over time in response to mishaps and exploitation of loopholes. They say the building and electrical code was written in blood.
alansaber 6 hours ago||
I guess it's more performant to stuff in a bigger system prompt now that models can support larger input sizes
cubefox 6 hours ago||
I would expect this only to be true for linear architectures like Mamba or Gated DeltaNet. Transformers and hybrid architectures do not have constant compute cost per token.
monkpit 4 hours ago||
Performant could certainly mean “higher performing” and not “quicker”.
lvncelot 4 hours ago||
> If the conversation feels risky or off, saying less and giving shorter replies is safer and less likely to cause harm.

Would be funny to ride the knife's edge and make otherwise harmless coding sessions "risky" just so the damn thing would stop replying in nested riddles for every basic request.

zahlman 27 minutes ago||
I've just been developing the skill of mentally skipping past that, on the assumption that having it in the context window will be net positive for the results of the next step.

I could be wrong about that, though.

Schlagbohrer 1 hour ago||
We need user-led research on exactly how to phrase a prompt to cause this, while still avoiding the crazy guard rails.
Shakahs 4 hours ago||
The full Claude Code system prompts are extracted every update and posted here, all 670 of them.

https://github.com/Piebald-AI/claude-code-system-prompts/tre...

comboy 7 hours ago||
I think they would benefit from asking Claude to list all contradictions and inconsistencies in that prompt which there are a few..

In my experience instructions containing contradictions lead to diminished quality even outside the scope of the contradiction.

conception 6 hours ago|
That’s interesting because they explicitly mention that as an issue in their prompting guidance for 5 - https://claude.com/blog/the-new-rules-of-context-engineering...
arkmm 7 hours ago|
"Claude keeps responses focused, brief, and concise to avoid overwhelming the person."

Claude and I must have a different idea of what brief and concise mean.

tgsovlerkhgsel 6 hours ago||
If you think Claude is bad at this, try Gemini. Even with explicit user prompts.

Claude seems to be better (not good, but significantly better) at judging where making the answer longer will actually be helpful (e.g. adding important information/context/nuance that a short answer would miss, thinking a step ahead, etc.).

treetalker 6 hours ago|||
The model almost certainly lacks accurate conceptions of overwhelming and person.
roncesvalles 3 hours ago||
Imagine what it's like without that line.
More comments...