Top
Best
New

Posted by Labo333 1 day ago

Show HN: The load-bearing vocabulary of Claude(louisabraham.github.io)
515 points | 246 commentspage 7
mohamedkoubaa 13 hours ago|
>unit-tested

Some of this is less to do with Claude vocabulary and more to do with the expectation that Claude justifies it's work. That expectation (probably) came from reinforcement learning.

heisenzombie 14 hours ago||
L
vadansky 18 hours ago||
It's a waste of tokens to do it post-hoc instead of fixing it upstream, but when I see a wall of text I just prompt "TLDR;"
jimmaswell 17 hours ago|
Depends what you mean by "upstream". It would probably be inefficient to force the model to use more human-readable wording in its internal thinking traces, while translating a block of text at the end is a pretty trivial task for an LLM.
nottorp 12 hours ago||
"Fuck" isn't there. Did they only analyze "AI" PRs?
jorl17 13 hours ago||
I'd like to share with HN an almost exact message I got from Fable today, with exactly ONE word changed for privacy reasons.

I love what I can build now, but I sure as hell don't love the headaches this trend has been giving me.

> So the full honest arc on the case we set out to fix: the expiry rules and day note tripled the loose version of the story, the relay fix carried the device’s own guardrail through the pipeline, the fair replay then revealed the last mechanism — ticket-anchoring — which none of the shipped layers reach. Remaining options, in order of my confidence: making the resolved-ticket summaries in the AI’s context carry their day so the expiry rules have something to bite on (small, mechanical, targeted at the observed anchor); and the plan-B second-model check, which structurally catches this class no matter how the model reasons. About $25 of headroom remains. Which way?

Yikes.

(The worst part is that I understand it)

DarmokTanagra 3 hours ago|
And yet you trust this abomination to write working code?
modelhub_dev 2 hours ago||
[dead]
stas4000 15 hours ago|
[flagged]