Top
Best
New

Posted by Labo333 21 hours ago

Show HN: The load-bearing vocabulary of Claude(louisabraham.github.io)
416 points | 193 commentspage 2
avsn 10 hours ago|
Missing a version of “this is where X earns its keep”. Noticed lately that Claude (and other LLMs) really love to use it.
modeless 2 hours ago||
I want to see the PRs in the cluster from early 2025. Did Claude adopt the style of some specific people? Or is it just random noise?
sroussey 14 hours ago||
I'm surprised vacuous is not on the list.

The word selection and way of writing has taken the joy out of using Claude.

nonethewiser 6 hours ago||
Im surprised by "provenance." I see it all the time. Except I've brought this up before and have never had it corroborated. I'm starting to think it's just my Claude.
halfmatthalfcat 13 hours ago|||
Also can't believe vacuous is not there.
user43928 13 hours ago||
I'm also missing the "latch" that "wedged" my test run.
sroussey 13 hours ago||
At least "wedged" is a thing i would say on a spinning out of control test that is stuck. "latch" though... not so much. I really wonder if this is the EU AI Act interfering with everything Claude does these days. As a non-EU citizen, I want a version without the rewriting of with watermarking in text.
user43928 13 hours ago||
I don't think so.

I understand it's not active yet, and when it will be, it should only nudge the chances between choices that are anyway likely and are already randomized today via temperature.

Watermarking is not the reason Claude talks like that.

sroussey 35 minutes ago||
Oh, it is active, and a pre-release version went out with v5 models.

I understand there are other reasons (overdone RLHF, for example) why claude is wedged into this weird way of writing that takes the joy out of conversing with it, but this is one as well.

The meaning of the words it uses can be oh so close, but the popularity of the words are not, and not in that context -- but the use of these other words changes the context ever so slightly, and then it uses other words where those words would work better.

I had something in code that related to people over time periods, and once it switches to a vacuous word choice, i found it starting talking using all ERP terms. I had to google the whole sentence to understand that, individual words were fine they just didn't make any sense to me.

customguy 14 hours ago||
Thanks to the infinite well of human creativity I am able to read "load-bearing" both as the intended affectation (I won't call it meaning) as as well "being full of shit".
Labo333 14 hours ago|
That was the intention behind my title!
elias_junit 4 hours ago||
I think you should recognise machine text not only by a list of words, but also by format. Right now, when it handles any complexity of text so well, the only thing that can differ is the structure. Machine text differs from human text in that it is just well structured. It does not allow non-linear narration. And it very often repeats some known social media patterns.
danpalmer 4 hours ago||
The vocabulary truly is load-bearing, without these words the model is less able to think. Where a human can understand a concept without words, an LLM plainly cannot. This is based both on the technological limitations and based on the evidence we see: as these models get better at working they get worse at communication.
fnordpiglet 3 hours ago|
I think these are more like RL tics caused by over zealous alignment towards specific goals, not all of which are to our benefit. A lot of the language used by Claude now is excusing of responsibility and inducing it to exit loops of work early and sit idle. This, IMO, is a naked attempt to offload load by quieting the models early and escaping from clear work to do. It’s gotten so bad that opus 5 loop escapes even as it claims it’s about to do something. I’ve mandated by engineering teams switch back to opus 4-8. Whatever frontier problem opus 5 excels at is so obscured by its inability to achieve any goal successfully without enormous amounts of hand holding that it feels like regressing to 2024 models.

The florid over exaggeration do certain words in bizarre ways is a reflection of their aggressive alignment towards too many goals, leading to weirdness in both behavior and language. The alignment functionally lobotomized opus-5 for any practical task.

Anthropic had a real gem in 4-6 and managed a near total market capture, which they have since squandered in the fastest burning of developer good will I’ve ever seen. It feels like exceeding the unity licensing implosion but without the single stupid decision.

danpalmer 33 minutes ago||
You may be right, but I would push back on near total market capture. I guess it depends which market you're referring to, but even scoping to just software engineering I don't think they got anywhere near total capture. OpenAI has always had a good foothold there, Gemini, while it may be lagging at the moment, had some great results with earlier models that I'm sure have held some market, and that's not to mention the Chinese market.

I think it would be fairer to say that Anthropic held the mindshare in the Silicon Valley style tech scenes around the world and the companies built on that model, plus a substantial portion of other software engineering. Now it seems that's dwindling quite rapidly.

Sharlin 8 hours ago||
The README of this project is very ironic. https://github.com/louisabraham/load-bearing/blob/main/READM...
d1l 3 hours ago|
This kinda dissonance is uncomfortable. How do people reconcile it?
souvlakee 8 hours ago||
I once asked Claude to replace “byte-identical” with a simpler word, such as “duplicate.” He refused and said “duplicate” does not mean the same thing as “byte-identical” so it should not be changed. He was very nerdy about it, so maybe that is the right way of his evolution.
whywhywhywhy 8 hours ago||
I'm shocked "shape" isn't near the top
Labo333 4 minutes ago|
[delayed]
MrDrDr 14 hours ago|
I find Claude language often hard to process and having to wade through these words can be draining. Embarrassingly, I’ve recently caught myself using them in conversations! Do all models have the their own jargon?
polalavik 11 hours ago||
I swear claude took a detour recently, its written output has been nearly incomprehensible to me. At first i thought i was getting AI-brained and just lost critical thinking but as i dug into response after response its was just the most obtuse language to explain what was going on. Really mentally taxing to wade through it all day.
swader999 4 hours ago||
It's almost unusable.
mike_hearn 10 hours ago||
I haven't noticed GPT 5.6 having any obvious jargon.
More comments...