Top
Best
New

Posted by claytonwramsey 3 days ago

I'm a seeing-eye dog for a computer(claytonwramsey.com)
72 points | 72 comments
dwedge 3 hours ago|
> I used to argue with people on the internet, after about six replies, you realize that you’re speaking to someone incapable of thought

He realises the current woe of things but doesn't realise he's been arguing with bot farms and teams of people hired just for this reason - to sow doom, arguments and engagement.

Around 10 years ago I noticed this happening on trending topics of Twitter - it wasn't that the opponents were stupid because they disagreed, it was that they were simultaneously intelligent and stupid in the way they spoke in a way that I realised I'd never seen in genuine people, making me realise it was probably different people or bots under one account.

If I realised it 10 years ago it was probably happening for at least 15. He wasn't better than these people he was falling into their trap

arjie 2 hours ago||
Haha it’s just a foolish move made by those from the ‘90s. We all internalized a different lesson. The younger generations have learned to brush it off and stop responding. Even if the person is human, when you fail to move toward Aumann Agreement, there is no point in persisting.

Us older ones are still fighting the last war. If only we explain it properly, maybe the bots will understand!

SturgeonsLaw 36 minutes ago|||
I justify it to myself by saying that even if I'm arguing with a bot, I want to have a counterargument posted for the benefit of the people who are reading the comments
dwedge 2 hours ago||||
> If only we explain it properly, maybe the bots will understand!

I'm going to be thinking about this perfect summary of my frustration of the last 20 years all day

blauditore 2 hours ago||||
Arguing on the internet is not about convincing the other side. It's about winning the majority opinion and getting more votes. Almost like politics.
neuroticnews25 2 hours ago||
Obvioulsy not true for all the arguments, only a subset of those not worth having.
skrebbel 2 hours ago|||
I googled Aumann Agreement and it’s all about agreeing about whether things are true/false/likely/unlikely.

I don’t understand your point at all in this context. What’s “Aumann Agreement” in the context of a discussion about values?

arjie 2 hours ago|||
Haha, it’s figurative and requires squinting. “rational actors with common axioms will reach shared conclusions through repeated exchange of information”.

So if you’re not reaching a shared conclusion then you’re not exchanging information, you don’t have shared axioms, or one of you isn’t rational.

If it didn’t land with you, it was just a poor choice of phrase and while I care to attempt to explain it, I don’t care to attempt to defend it.

Values are not different from other positions that people might hold. I believe values and principles are just lossy compressions of positions about specifics because we cannot express them all. When you poke at the edges you find they are fractal. So “what’s your estimate for random variable X?” And “are open borders good?” Are not really different questions so much as questions about different levels of specificity.

My incomplete self-note about this here https://wiki.roshangeorge.dev/w/Principles_Are_Compressed_Im...

neuroticnews25 1 hour ago|||
Have you heard about the free energy principle and free energy model of emotions? In this model, all brain does is minimize the prediction error, so all values and beliefs are bayesian priors by necessity.
skrebbel 1 hour ago|||
Appreciate the writeup, thanks.
bratbag 2 hours ago|||
The outcomes those values achieve.
hliyan 3 hours ago|||
I've now completely stopped engaging with anyone on that platform that I can't reasonably traced back to a real person. If any non-real-name account says anything that remotely feels bad-faith in replies, I block them. Having been on the Internet since the mid 90's when it was the frontier, and everyone helped/trusted everyone else, the policy feels very wrong. But unfortunately Twitter has given no tools or alternatives to deal with the situation.
walrus01 2 hours ago|||
I have found that the only way to deal with this these days is basically to only use my time to discuss things with people in private, invitation only forums. HN is a rare exception for "I am pretty sure you all are not bots (yet)" and yet anonymous.

Everything else that I care about is in private Signal groups or an equivalent of same with anywhere from 15 to 200 people, and adding any new member means they need to be vouched for by an existing one. And at least 55-75% of the people know each other face to face through industry vertical specific events, trade shows, conferences and similar.

I can't even imagine trying to engage with people on something like x/twitter or other social media. Not knowing whether the 'person' you're replying to is somebody's LLM authoried bot farm.

hypfer 2 hours ago|||
> And at least 55-75% of the people know each other face to face through industry vertical specific events, trade shows, conferences and similar.

Tbh, my experience with _that_ has been worse than with a random sample of internet background radiation, as people seem to have no idea how to filter correctly and who to vouch for and why.

_Especially_ with people they know via work, conferences, etc. Spaces that do not filter for "genuinely relevant contributer to a conversation" at all, but encourage faking that.

walrus01 2 hours ago||
Within this context it's specifically people in the ISP and telecom business, most of whom have been in the business 15, 20 or 25 years now. Some of us started out as junior NOC techs, junior sysadmin/Linux/BSD guys in the late 1990s at different dotcom 1.0 startups.

At a certain scale you know who the people are that have super-admin privileges at very specific ASNs and they are well known in the industry. Basically much as somebody in academia in a journalism school on the west coast would know who is the department head of the school of journalism at a major university in Boston or NY and recognize their name.

Now within a Signal group of 150 people adjacent to the neteng teams at large to medium sized US/Canadian ISPs, it's entirely possible that somebody let an uninvited rando in, but we don't discuss anything super sensitive (that's for more 1:1 messages). A person who has been lurking for weeks or months and suddenly starts chatting in an inauthentic way would be found out very quickly.

One of the things that ISP technical operations groups do very well is filter for non-authentic participant in something, because sales persons getting into groups of technical discussion with the angle of selling some new thing is a perpetual problem and is a well known factor to deal with.

Also sales people generally don't do well when a complex operational-related technical matter is put in front of them in a fast paced environment where they can't engage their engineering team to get an accurate reply. Ask anyone that's ever tried to get their direct cellphone number removed from a Cogent IP transit sales person's CRM system....

hypfer 2 hours ago||
Yes, I got that, but what I actually mean is that these criteria do not necessarily lead to good conversations.

I know the kind of space you are referring to. I'm in those too. But my point is that they regularly disappoint me.

Which is usually worse disappointment than with randoms on the internet, because the people should know better but don't.

With these tight knit spaces you get all the social dysfunctions of people liking each other not for merit but because they've been around each other for long, and that's - for me personally - worse than the public alternative, as you suddenly do not argue with logic, but with the tribe and social cohesion.

novok 2 hours ago||||
That feels like the intellectual equivalent of limiting your career opportunities within a small tiny villages vs. a large city. There is so much more in the city.
walrus01 2 hours ago|||
It's not that I don't read a great deal from a wide variety of sources, but that I reserve the small amount of free time I have per day for back-and-forth discussions (like this one) with groups of people I am fairly sure are not sock puppets. The amount of oversight that dang exercises on HN to ensure that frequent and verbose commenters are authentic people should not go without notice.

I have a theory that a good part of the filtering for authentic humans has already been done here, which gives me a higher level of confidence that I'm not metaphorically shouting into the void by conversing with somebody's LLM or social media influence apparatus.

sznio 1 hour ago||||
I have moved back to my small hometown, from a large city.

My mental health improved significantly. I actually feel valuable at work. I feel pride in contributing to my local community. And the paycheck isn't actually worse.

hypfer 2 hours ago|||
Sounds just like gambling, honestly.

Which might be a winning strategy, but also might not be.

ErroneousBosh 2 hours ago|||
> discuss things with people in private, invitation only forums

One of the great things about not having ads on rangerovers.pub is that Google is just not interested in indexing it, because it doesn't make them any money. It shows up here and there in search results, sporadically, especially if you know a username or phrase specific to the site, but other than that if you want to know about it you need to be told about it.

I was never a fan of web forums but they are a nice compromise between really oldschool mailing lists and FAANG content-to-eyeball transformers.

walrus01 2 hours ago||
I have seen a bit of a resurgence in your classic 2005-era type interface php bulletin board in a browser, as people relegate their facebook/instagram/whatever accounts to only keeping up with family stuff posted by boomers who won't use anything else, or things like niche neighborhood based buy & sell groups (or community notice groups in like a town of 12,000).

I guess there's at least 3 or more active and currently developed software packages with open source licenses, the flaskbb that site uses, simplemachinesforum, the classic phpbb, probably some others I've forgotten about... The LAMP stack dependencies for all seem very mundane and very lightweight.

arjie 2 hours ago||||
This is the only way to realistically deal with it but it’s not scalable. A platform is only as useful as its non-bot users. On Twitter even the positive replies are likely bots. Real engagement is limited for most users.

On a site like HN, blocklists are very effective at keeping people off your feed.

hypfer 2 hours ago||||
Same. (I would love to say, but they still get me way too often)

I can also highly encourage people to keep notes and do some basic OSINT. The effective internet is smaller than one might think, so that proves useful time and time again.

dwedge 2 hours ago||
Can you elaborate on the OSINT suggestion? Do you mean on the people you engage with?
hypfer 2 hours ago||
Yes, exactly. Though, not by default.

If something seems _odd_, I suggest looking up _why_ it might be odd. Gathering context, essentially.

Like "Okay, this guy is weird. Aah, okay, LinkedIn says that he works there. Okay _now_ that makes sense".

Based on that, you can then decide how to approach the (previously failing) interaction.

__

HN is a very easy place for that, because, to give somewhat concrete examples, if someone is shilling for something and seems unreachable for common sense, LinkedIn usually tells you that their salary depends on that.

And people play very open here, because this is treated as a business networking event.

gambiting 2 hours ago|||
>>anyone on that platform that I can't reasonably traced back to a real person

Not sure why that makes any difference - I've had plenty of arguments on Facebook with people who are perfectly happy to spew racist and/or conspiratorial bullshit while having their full names, their holiday photos and photos of their children and/or grandchildren attached to their identity. The weird thing is that these people are 99/100 times incapable of actually having an argument, they either start insulting you straight away or say some variation of "if you don't like what I'm saying then leave" or actually majority of time "what does a foreigner like you know about this topic".

The real names policy has yielded zero of the promised effects imho.

dwedge 2 hours ago||
[dead]
sznio 2 hours ago|||
>it wasn't that the opponents were stupid because they disagreed, it was that they were simultaneously intelligent and stupid in the way they spoke in a way that I realised I'd never seen in genuine people, making me realise it was probably different people or bots under one account.

I believed they were all bots until I actually interacted with such a person in real life. It's terrifying.

xboxnolifes 2 hours ago|||
I used to think that the people who could not apply any semblance of logical reasoning online were just bots or trolls, but then I started meeting them in real life.
dwedge 2 hours ago||
Sure, they exist, but online they are simulatenously unable to apply any logic and then surprisingly coherent and intelligent in the next reply.
Fr0styMatt88 2 hours ago|||
I'd love to know if there's a documentary or something that has been done about this, that I could point non-techie people in my life to.
okasaki 52 minutes ago||
This is some schizo thinking. People you interact with online almost certainly aren't bots or paid.
Cthulhu_ 49 minutes ago||
Certainly aren't bots is dubious, although if there's actual interactions it's likely they aren't bots.

Paid, lmao I wish. Nobody gets paid to repeat propaganda, generally.

onion2k 3 hours ago||
One of the first things I tell the junior/mid-level developers I mentor is "You can't debug something just by reading the code." We all have a mental model of how our code works, and it's usually a bit wrong. Bugs are the real world manifestations of those mistakes. When you read the code it's all filtered through your model, and that makes you blind to seeing why something unexpected happened. In order to debug something you have to be able to put the system in the state where the bug happens to see why it occurred.

LLMs generally only debug systems by reading the code with whatever information you give them in a prompt. The image in the article is meta-prompt - the prompt is whatever comes from the vision model the AI happens to use to 'understand' the red circle annotation. That won't work. To successfully debug what's going on it will need much better state information. Has the 'shelf' been explained to is? Is the contrast and lack of shadows in the image messing up the vision model? Why isn't the 'lid' in the image? And so on.

LLMs are clever but they're not magical. Treat them like a naive junior dev. Give them enough data about the state of something to understand it properly.

rcxdude 2 hours ago||
Hmmm, I don't think that's necessarily true. Often times once I have witnessed a bug, I have found it just by reading through the code with the behaviour of the bug in mind.

For LLMs, this is likely to be disproportionately effective as well: especially because they don't really build up a persistent view of the codebase, they're generally re-reading it each session, and they tend to be surprisingly good at predicting the behaviour of code.

(That said, knowing where and how to gather more evidence to make things clearer is a pretty core skill in troubleshooting, so it's generally good advice anyhow)

valzam 1 hour ago||
Also Claude Code is very good at writing small scripts/on-off test cases to confirm bugs, so I wouldn't even say the initial premise is correct.
onion2k 17 minutes ago||
That's AI getting an example of replicating the state that shows the bug which is exactly what I'm talking about. It does that far more than humans do, and it's ace. That's how you should be debugging a system - replicate the issue, understand why it breaks in that given state, and then make a code change to fix it.

Sometimes you can do that mentally and fix the code. Often your fix will be right especially in a relatively simple part of the code. However, equally often you'll fix a different problem (or something that wasn't a problem at all), and the original bug will remain but you'll believe you corrected the issue. This is why you should always replicate a bug to understand it, and why you should always add a test whenever you fix a bug to prove you actually fixed it as well as preventing future regressions.

OtherShrezzing 2 hours ago||
I call this a tautological mental model. You can read the code over and over again, but your second reading will be mostly an echo of the mental model you built up in your first.
hspeiser 4 hours ago||
I completely understand this. I’ve worked on robot hands and 5.6/Fable 5 were practically useless at helping me debug anything visually.

What I have found super useful actually is having models make a interactive 3d viewer in which I use move / highlight / paint (soft body painting directly onto the geometry for issues and different colors mean different failures). This gives a much better way to communicate the physical relationships and positions that are hard to get across in a labeled screenshot.

Its for sure still a lot of manual work so the "seeing-eye dog" description definitely holds. But I have found that after a couple of examples with the extra context the model gets much better at handling the problem and becomes useful.

beklein 3 hours ago||
A bit off topic, but I absolutely love the little robot on the author's main project's landing page (https://rerun.io/). I normally condemn mouse hijacking, but this implementation will be allowed.
hobofan 3 hours ago|
Rerun is pretty dope, but I'm not sure how you came to the conclusion that it's the "author's main project"? There is a whole company backing it, with no affiliation that I could find, apart from the author being an occasional contributor?
beklein 2 hours ago||
They mentioned "... and my visualizer tool comes with an MCP server ...", with a link to the rerun project. I guess it makes more sense that his visualizer tool uses rerun...Sorry for the confusion from my side.
Fr0styMatt88 2 hours ago||
I've found that LLMs are specifically bad at a certain kind of debugging, though I can't quite put my finger on what that is.

"Spot the bug in this code" when the code can be looked at and pattern-matched against bugginess is something they seem really good at.

Some parts of debugging, like "Here is this logfile, what do you think is going on?" are also surprisingly good.

It's that thing kind of in the middle -- I know it when I see it honestly is the best way I can put it into words. An example from recently, I'm receiving some bad data on a network message parser. Immediately I don't know whether it's a my-side or their-side thing, but I know if I try and just vaguely describe the behaviour to the LLM it will start churning tokens.

My current approach to problems like this is -- I need to tell the LLM what it needs to do to give itself the data it needs to solve the problem. My first reaction now isn't "It's not working, there's a bug, it's not doing X". It's "Okay, this isn't quite working properly; I need you to add some debug logging around X, Y and Z so we can figure this out". That tends to avoid spirals and get me out of the situation much more quickly.

The seeing eye dog analogy is pretty apt actually. I would love to see some transcripts from the author if they are able.

Edit to add: I think the 'thing' I'm alluding to might be -- if I have trouble expressing the buggy behaviour clearly in words, then I know it's probably going to be a fair few back-and-forths with the LLM to get something; the harder I find it to concisely describe, the more risk that it'll fall into a pit. Doubly so if I offer up a hypothesis which turns out to be wrong.

creichenbach 2 hours ago||
Those three colored shapes at the bottom look a lot like the EPA logo, a former grocery store chain: https://de.wikipedia.org/wiki/EPA_%28Warenhaus%29?wprov=sfla...
dostick 3 hours ago||
LM still can not see and understand the desktop app UI on a level that is acceptable for testing. All the advances in coding are from web dev and thanks to the nature of html UIs. Try to develop a desktop app and it’s like working with a legally blind person who can see some part of the screen is they squint in a certain way but surely will miss all minor details.
walrus01 2 hours ago||
> I often handwrite the code myself, but I’ve found that LLM coding assistants’ limitless patience ameliorates the drudgiest work of coding.

I've been using a few different "smart" LLM to work on an analysis, parsing, search and correlation tool that ultimately deals with a 5.5GB on disk (with indexes) mariadb database that has its origin as a federal government department's 905,000 row plain text CSV file.

There are a ridiculous number of data entry errors and just plain weird fuckups in the data origin that don't seem they will be ameliorated any time soon, so automating the drudge work of cleaning it up and rectifying it into something usable is a textbook case for this. Very pleased with the results so far.

nannal 2 hours ago||
You could setup a webcam and have a vision llm stalk the breakroom of left over pizza and alert you.
hypfer 3 hours ago|
I mean there's a reason why we're doing MoCap for video games. If computers were good at this, we wouldn't be needing that. But actual motion and all seems to be much more complex than the systems can predict, apparently.

Also.. uh.. isn't this.. good? I thought AI was to steal all our jobs.

___

Beside that, kinda weird self-description.

Isn't the computer executing your commands and you're just filling in where it cannot do that?

Being that dog implies that the computer is in the driver seat.

I mean it's supposed to be a joke I guess, but I read it as one that leaks internal metadata which seems to be incorrectly calibrated.

smugglerFlynn 3 hours ago||

   > Also.. uh.. isn't this.. good?
It is weird if you think about it this way: it is AI that waits for you, its ‘eyes’, to provide a feedback so it can continue working. It literally uses you as its organ.

<!!spoiler ahead!!>There is a TV show called Person of Interest <!!spoiler ahead!!>, where Machine (AI connected to Internet and CCTV networks) has no legs or eyes, so when it needs to go and check something not covered by CCTV feeds, it gives instructions to a real person. In the show it is called an ‘analog interface.’

hypfer 3 hours ago||
> It literally uses you as its organ.

But it isn't. That's my point.

I told the clanker "hey do that", and like the intern/junior it emulates, it eventually says "boss! Help! I can't do this alone".

It is I who is in the driver seat.

smugglerFlynn 2 hours ago||
If intern is making a breakfast, and boss is the one who suddenly runs to the grocery store because eggs are missing, is boss still the one in a driver’s seat?

From the original goal point of view yes, as it was boss who has initiated whole breakfast procedure. But from an execution standpoint it is intern who gives its boss a job of a grocery store run. He could give same job to anyone else, boss as a persona is irrelevant here.

wolfi1 3 hours ago||
>Also.. uh.. isn't this.. good? I thought AI was to steal all our jobs. the problem is, the CEOs still think AI solves their problems ie minimize paid jobs
More comments...