Top
Best
New

Posted by nickthegreek 4 hours ago

What happened to TheNumbers.com(stephenfollows.com)
179 points | 64 commentspage 2
hebleb 2 hours ago|
Good read, I was so curious how this happened a few months ago
dumberquestions 3 hours ago||
Ironic that polymarkets were being advertised as helping society make better predictions.
jambalaya8 2 hours ago||
John Brunner and Alvin Toffler both made stark warnings wrapped in futuristic giddiness about things like it. People like Fuller no doubt thought polling on a large scale was terrific. There were old usenet groups and BBS subs (minus the money aspect) experimenting with the model. I do not believe they ever are or were good in a largescale model (money or not).
mrandish 2 hours ago||
I mean wisdom of crowds, super-forecasters, calibration and pre-registration are useful tools that can result in better predictions, turning it into online gambling is where it went sideways.
hyperhello 3 hours ago||
If the scrapers are going to get it anyway, put the data up as a zip somewhere.
jjgreen 3 hours ago||
It would not make any difference.
codemonkey-zeta 2 hours ago||
Indeed, the article mentions Wikipedia experiencing similar scraping pains, even though they already DO have bulk data available.
HeatrayEnjoyer 2 hours ago||
Who are running these bots? I presume developers at all of the frontier labs know (or at least would know to look for) Wikipedia has bulk APIs for automated access. Unnecessary scraping increases their workload/costs too, so why in 2026 is this still a problem?
esseph 1 hour ago||
Black market and gray market data. All the firms want data. All the other firms want data. The banks want data. The other criminals also want data for their crimes and schemes. Oh insurance companies, and the ATS systems. Everybody wants as much data as they can get and they don't care how they get it.
antisthenes 1 hour ago||
> Read the Docs, a non-profit that hosts documentation for open-source software, who watched a single crawler download 73 terabytes of zipped HTML in one month, costing it over $5,000 in bandwidth

From the article.

Not the same site, but an example of the same issue.

tehjoker 2 hours ago||
The real story here is that prediction markets were banned for a reason and loosening the rules is causing chaos just as was expected. AI plays little role in this story, hacking by humans would also be motivated by financial returns, unless the element is that AI hacking is cheaper and the returns are not so big.
jambalaya8 3 hours ago||
I always liked this site, but reading this and seeing the anger about expecting the site maintainer to do things for you is repulsive. Frankly, if he wanted to pull his site down with no notice that is perfectly within his right. It was/is his site. He doesn't owe anyone a .tar.gz either. His work.
BraveOPotato 3 hours ago|
I agree. I've seen it happen before on a project I use. I decided to take a look at the repo for one of the plugins, and I saw a heinous issue that basically was TELLING (not even asking) the maintainer to fix it.

Deplorable behavior indeed

jambalaya8 2 hours ago||
Like dominoes, as soon as it is accepted in a few places, people think it is acceptable to push to 'share'. It's almost terroristic sometimes, the pressure some maintainers are under.
NetMageSCW 2 hours ago||
What does a cyber attack have to do with AI scraping?
john_strinlai 1 hour ago|
they both have a risk of harm which the site operator was no longer comfortable with.
brcmthrowaway 3 hours ago||
Its just too easy for a technically minded bored person to produce slop that hammers websites. There needs to be a penalty for this.
draw_down 3 hours ago||
Sorry! Just yesterday many of us decided that AI scraping isn't a real problem, and anytime it is blamed, it's a cover for something else.

https://news.ycombinator.com/item?id=49005747

aaron695 1 hour ago|
[dead]