Top
Best
New

Posted by robin_reala 2 days ago

Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence (2025)(arxiv.org)
172 points | 104 commentspage 2
lowsong 1 day ago|
In decades to come we view "chatbot" AI as one of the most dangerous inventions of this century. Once regulation catches up and they are outlawed, we'll look back on this period of history with horror.
rramadass 1 day ago||
This is an important paper which everybody should read and then accordingly tune/guard their interactions with AI; especially true when people use AI for personal/psychological support/validation. It will completely distort reality and push people into fantasy land which when mapped to the real world can have disastrous consequences.

Sycophancy is the "stickiness factor" of AI analogous to that of Social Media.

For some background read Jagged Intelligence: The Dangerous Unknowns at the Heart of LLMs - https://news.ycombinator.com/item?id=48577159

Here is an experiment that i did;

Ask AI to build a "Character Profile"(in the broadest sense) of a person based on their available public writings. This requires "commonsense reasoning" (https://en.wikipedia.org/wiki/Commonsense_reasoning), understanding human motivations and behaviour, context, assumptions, societal knowledge etc.

I know "me" and so i asked AI to use my HN comments/submissions as input :-) The sycophancy/flattery/praise was quiet excessive. I am realistic and old enough to not need ego-soothing (a little is fine but a lot makes me suspicious) and so i asked AI whether these phrases were not too over-the-top. It apologized and agreed to drop the fluff. I then asked it to identify job roles (any) for which i might not be a good fit given my character profile. This acts as an external constraint which focuses attention on shortcomings and hence forces the AI to look at the other side of the coin. Now the results were better and more in line with what some Human may deduce from my HN persona (which obviously is not my complete real-life persona) but still wanting in many aspects.

The above is a perfect example of "Jagged Intelligence" exhibited by AI. Excellent in formal symbolic manipulation sciences, regurgitation and simple reasoning but highly deficient in human-like commonsense reasoning.

monocola 1 day ago||
[flagged]
theturtle 2 days ago||
[dead]
davesiknow 2 days ago||
[dead]
alexbelyanin 1 day ago||
[flagged]
jtrn 1 day ago||
[dead]
daun_gee 2 days ago||
[flagged]
LAC-Tech 2 days ago||
I think this goes along way to explain people's very defensive reactions when you are skeptical about Agentic Development.
artnanika 2 days ago||
Those are two different things. Chatgpt being sycophantic while playing the role of a therapist with a teenager is much different from Codex being sycophantic while reviewing user feedback. The former is extremely dangerous for society.
LAC-Tech 2 days ago||
I make no societal judgments; just that the nastiness of the reactions I get suggests something more than technical choice is going on.

I had one prominent "influencer" publicly challenge me to a coding competition over it. It was very weird.

qarl2 2 days ago||
I'm one of the people who gets accused of this. I don't mind skepticism. But every time I turn around, someone's claiming nobody gets any value from these tools and anyone who says otherwise is deluded or running a scam. That gets really old.
LAC-Tech 2 days ago|||
I can believe people get value out of them. I think I do.

But "OH MY GOD EVERYTHING HAS CHANGED THE OLD WAY IS DEAD 10X MORE PRODUCTIVE" - no. If that were true, we'd notice it in the software all around us.

qarl2 2 days ago||
I think we are noticing it. Are you not?

I've shipped three software projects in the last two months. I shipped zero in the preceding year.

Right now I'm building a system to decompile MAME ROM games into idiomatic JavaScript. It completes one game in about 3 days - complete with extensive commenting. I started it about 2 weeks ago.

I'm not sure "EVERYTHING HAS CHANGED" but if you're not seeing dramatic change, I suspect you're not looking.

Planktonne 2 days ago|||
> if you're not seeing dramatic change, I suspect you're not looking

If the people claiming that everything has changed were even a fraction as productive as they think, then it would be visible even to people who weren't looking.

As it is, a lot of people are actively looking but still not seeing it.

qarl2 18 hours ago|||
I'd love for HN to explain why I was flagged here for answering:

> If the people claiming that everything has changed were even a fraction as productive as they think, then it would be visible even to people who weren't looking.

with:

> If you can't see this - you should probably get your eyes checked.

He is clearly suggesting I am delusional about my productivity gains, and I'm the one who gets flagged.

I daresay I'm noticing a bias in the moderation.

qarl2 2 days ago|||
[flagged]
chaps 2 days ago|||
"you should probably get your eyes checked."

Don't be a jerk. A lot of people have had bad experiences with these systems from people basically saying exactly what you're saying. This "you should get your eyes checked" mentality is pervasive with the nerds who confidently say that their analysis is solid when it's... not. It's a big problem in communities that have non-conventional videogame puzzles for example. Someone will come in, announce some novel solution.... leading to a four hour fight about the efficacy of LLMs, only to find a mistake in their code the next day.. never to be seen again.

Mind you, I'm not saying that these systems aren't great at reverse engineering and whatnot. They're spectacularly good at that. But RE is a relatively constrained problem because everything is still right in front of you. For complicated problems where the signal is mixed between noise... less so.

qarl2 2 days ago||
As near as I can tell - your argument boils down to "I've seen other people make mistakes so you must be also."
chaps 2 days ago||
If that's what you got from it, then I hope you have a good life.

Be kinder, friend.

qarl2 2 days ago||
If you can point to a single error in my argument I would love to talk to you about it.

Otherwise, you have a good life as well.

chaps 2 days ago||
My point is just that you're talking past people and being an asshole in the process. Not everything is about winning an argument. Lower your temperature and you'll have more productive discussions that revolve around disagreements.

Cheers.

qarl2 2 days ago||
No. I am not talking past anyone. I am providing cogent argumentation and evidence. And I'm being met with nonsense and platitudes.

And... I didn't start it.

Cheers.

EDIT: Good choice.

Planktonne 2 days ago||||
I'm pleased for you, but these tools have been out for many, many weeks, and there are many, many people claiming incredible productivity gains.

Where's the rest of it?

qarl2 2 days ago||
I just showed you something that would have been considered a miracle one year ago.

And your response was "So what - show me another one."

Planktonne 2 days ago||
I'm not disputing that it's a hard thing to do, but we had decompilers before; "miracle" is rather stretching it.

Even granting that framing though, the claimed increase in productivity isn't a one-off, but asserted for everyone using them; you're claiming mass-produced miracles but trying to depend on a single example.

qarl2 2 days ago|||
> but we had decompilers before

Cite one example of a decompiler that provides English names and comments.

fragmede 2 days ago|||
Snowboard kids 2 https://news.ycombinator.com/item?id=48284494

Silent Hill 1 PC https://youtu.be/0niadQJmYx0

Super Smash Bros/other GameCube and Wii games https://www.reddit.com/r/decomps/comments/1uvttvy/moderngekk...

overgard 2 days ago|||
I'm sure it's useful for that, but we've had pretty good decompilers before AI.
qarl2 2 days ago||
I would love for you to cite a single one.

It must take a binary executable and derive English names for the memory addresses.

overgard 2 days ago||
I used to use dotPeek all the time back in like 2012. There are also things like Ghidra that the NSA made. I dunno, google them, there are a lot. I'm not saying LLMs aren't useful here, just that it's not a huge deal. Also, why are AI people always making tools to take other people's IP? You are kind of fucking with people's intellectual property without permission it seems like. I'm not saying it's the crime of the century or anything but it's distressing how many AI projects come out of "let me repackage something someone else made"
qarl2 2 days ago|||
> there are a lot

There are none.

Ghidra does not provide semantic names.

dotPeek shows you the symbols that were not stripped from the .NET.

Neither of these can do what I asked.

fragmede 2 days ago|||
lol you think he hasn't heard of Ghidra? How stupid do you think he is?
bigstrat2003 2 days ago||||
> if you're not seeing dramatic change, I suspect you're not looking.

Ironically, this is the inverse of the very thing you complained about people doing to you.

qarl2 2 days ago||
Except I'm providing evidence.
xracy 2 days ago||
I'm seeing more examples of you saying you're providing evidence than of you actually providing evidence. So far I've heard a single anecdote about what you did.

That's a pretty far cry from the "evidence" you claim to have dropped on multiple comments.

qarl2 2 days ago||
LOL. Providing a counter example is far from just an anecdote.
vhantz 1 day ago|||
An anecdote is an anecdote is an anecdote
qarl2 1 day ago||
A proof isn't an anecdotal story.

The claim was this is not possible. I show it happening. That's a proof.

You're confused because the example is my own project. But I'm not merely describing it - I am giving you full access to it to confirm or reject my argument.

Not anecdotal. Saying it three times does not make it true.

chaps 1 day ago||
What things have you tried that didn't go well? At this point, that's more interesting.

If you want to have a whack at a hard problem, have a try at the cryptographic puzzle in Noita if you want to see the silly failure modes of these systems. Lots of people have tried to solve it with AI and nothing's budged. If you can solve it, great.

qarl2 1 day ago||
Why are you changing the subject?

What does your new subject have to do with the fact that the previous arguments were all pretty much baloney?

Pointing to a problem that has not been solved is completely unrelated to the miraculous things that have been solved.

And I thought you'd decided (three times now) that I'm not worth talking to?

chaps 1 day ago||
Because I find what you have to say interesting. Sorry about that, I'll try to consider you less interesting.
qarl2 1 day ago||
If you'd like to change the subject I'd be happy to do so.

You're right - there are still many problems that have not been solved by AI. AI has not found a cheap and effective way to turn lead into gold, as an example.

But I sorta think that's a silly metric. There are always going to be unsolved problems for any system. The interesting metric is how many problems have been solved - and how surprising are they.

Like I said above - a decompiler that produces semantic understanding of the machine code it's looking at (this address is score; this address is how many lives are left; this address is...) would have been considered literally impossible before LLMs. Today, I was able to implement such a system in under two weeks.

That's an impressive result. If you disagree, I'd love to discuss it with you.

chaps 1 day ago||
I'm aware of what they're good at. I got claude 4.6 to run linux as a native postgres module inside linux inside postgres inside linux inside postgres on a mac. It was impressive then and it's still impressive. IOW, I really think you're misinterpreting my comments as significantly more inflammatory than I actually mean. And you seem to be taking it personally.

But what I'm interested in is in the edges of their failure modes because those failure modes prop up frequently in pernicious ways. For a similar reason that I want to understand the edges of my own intellectual failure modes by regularly challenging myself. This conversation is an example of that.

It seems like you're just not interested in discussing the edges and failure modes. You call that a "silly metric", I call it, "understanding your tools".

qarl2 1 day ago||
My primary interest in these threads is to debunk the claims that "AI is a hoax".

Granted, that's not exactly the claim these days, the claim's goalposts shift constantly, as these things do.

But if you go back to the top of this conversation - I asked why people aren't seeing the miracles (not cute tricks - useful things that simply were impossible before) and I was largely met with the response "you're delusional if you see miracles."

But the thing is - I am not. So I am going to continue arguing that case every time I see a version of it.

So no, I'm not interested in exploring other issues. Sorry. Yours in particular I find as interesting as discussing whether Microsoft Word can be used to edit video files. Not interesting.

But if you need validation, here it is: there are many problems that AI cannot solve. Of course there are. There always will be.

qarl2 1 day ago|||
HEH. Not sure why you're implying I have changed my position. I said as much above and would have said the same at any time. I don't think it's controversial in the least that there are things AI can't do.

Do you have a weird need to reframe this so you can feel like you won? You shouldn't do that. It's not healthy.

chaps 1 day ago|||
Great! Glad you were able to take off your rose tinted glasses, even for a moment. Peace.
qarl2 1 day ago||
Oh hey - the timer finally kicked off. In case you didn't see my reply - here you go:

HEH. Not sure why you're implying I have changed my position. I said as much above and would have said the same at any time. I don't think it's controversial in the least that there are things AI can't do.

Do you have a weird need to reframe this so you can feel like you won? You shouldn't do that. It's not healthy.

chaps 18 hours ago||
Thanks!
xracy 1 day ago|||
What counter example did you provide? You said you had tried a few projects. You can just claim that without any evidence. I have no reason to believe you're not lying.

You could say "I used AI to cure cancer." This isn't a counter example, because there's no evidence for it.

qarl2 17 hours ago||
Agreed.

Which is why I provided a link to a the source of one such project above. I have several projects in that Github account if you need more.

Perhaps you didn't see it.

LAC-Tech 2 days ago||||
I think we are noticing it. Are you not?

I am not. Congratulations on your side project.

qarl2 2 days ago||
Projects, plural.

And weren't you the one who started this complaining of nastiness?

qarl2 2 days ago||
I'm sorry downvoters - it's true. In his opening statement he complained about nastiness and then became extremely nasty himself, when I had said nothing like that to him. Other than just disagree and provide evidence.
veqq 2 days ago|||
> I think we are noticing it. Are you not?

The software around us seems worse than ever, constantly breaking and significantly worse than 1-3 decades ago.

pc86 2 days ago|||
You think the software around today is worse than software in the late 90s?
qarl2 2 days ago|||
I guess my experience is different than yours - and I'm providing evidence.
sidrag22 2 days ago||||
I've had one conversation that was sorta like this, it was around the time some kid kept releasing videos of him interacting with a very bad TTS gpt model or something and it would give very very bad answers confidently.

Their frame of reference was that, and this was at a time when opus 4.5 was around. Hard to expect a reasonable conversation when one person has such a different view of what the tools are capable of.

chrisjj 2 days ago|||
> But every time I turn around, someone's claiming nobody gets any value from these tools

Really? I find agreement that gullible users get something they value, even if it is worthless.

atleastoptimal 2 days ago|
This is true, but likely happening in talk therapy too with overly sycyphantic therapists
jalev 2 days ago|
Sure, but there is a marked difference between a therapist and an LLM Chatbot: you can't talk to your therapist for +4h a day, at any time convenient to yourself. It's also not a guarantee you're going to hit upon a therapist that is like that, versus a product that is very literally intended to maximise your usage of it.
paimapi 2 days ago|||
a therapist should, if they follow their training, not be a sycophant whatsoever. CBT, for eg, is a multi-step process including Socratic self-dialogue, reframing practices, etc to guide the person towards the development of insight and healthier coping mechanisms
lstodd 2 days ago|||
Any good therapist should train you in the methods of coping with and eventually overcoming your problem. This boils down to private grad+postgrad psychology/psychiatry course, because to cope+overcome you have to understand what's going on and there is no one who can do shit but you yourself.

But that is only ~half of job. The other half consists of gently reminding one that not all is lost yet.

There is no LLM that can do this and I think there never will be.