Top
Best
New

Posted by davidest 4 hours ago

Ask HN: What is one simple thing LLMs are insanely bad at?

I am looking for ideas on what to train a specialized model for!

What is one simple thing you repeatedly ask ChatGPT, Claude, or another model to do that it still somehow messes up?

26 points | 63 commentspage 4
eli 4 hours ago|
I have been working on a personal benchmark suite to test new models and ironically one thing all the models are bad at is writing new benchmark tasks. I guess it’s the different layers of abstraction between the task and how it’s evaluated? Or maybe just a lack of “imagination”

Tasks it writes are typically too easy but also it utterly fails to see how a different model might misunderstand a vague part of the prompt.

flippy_flops 4 hours ago||
humor
veganmosfet 4 hours ago|
+1 We need humor benchmarks!
ipaddr 3 hours ago||
Generating money or profitable ideas
respectattentio 4 hours ago||
science?!! but I'm working to fix that...
newsomix9xl 4 hours ago||
ASCII charts.
rufi 4 hours ago||
very bad at financial calculation
bpodgursky 4 hours ago||
Claude is still not perfect at reading and interpreting noisy graphical data (imagine something like an EKG or chromosomal microarray plot). Still better than an average person but makes mistakes, not sure if this fits your description.
Conol_ai 4 hours ago||
[flagged]
senectus1 4 hours ago|
providing value for the actual cost (not the price we're being charged atm, the actual cost)