As someone who operates a large non-profit public data driven website, I have some VERY strong feelings about scrapers. We looked into various commercial solutions (Datadome, HUMAN) and based on our traffic estimates from logs we'd be looking at at least 250k/yr for bot mitigation. Anubis is offering a temporary reprieve, but after reading the recent kernel.org article [0] it's increasingly clear that this is a temporary bandaid.
The cheapest solution is to require a login and rate limit by API key. I also have strong feelings about the tragedy of the commons.
How much of the current valuation is due to Google's resources and support, though?
"Sell company now and it's worth more a decade later" is the entire point; acquisitions are an investment and the investment is supposed to appreciate.
It’s impossible to know for sure but if they’d stayed private I think they’d be more not less successful, just because startups can move more nimbly. Although we’d need to define success, we’d be here forever…
It wouldn't have. They had over a billion in debt, which Google casually fixed. Google then very conveniently supplied them with GCP server power. And remember all of this was in like 2014, way before the AI hype. They wouldn't have had access to capital to survive long enough to make it to today.
RTLinux kernel works fine on both platforms, and includes external synchronous scheduling context clock peripheral input on arm64 (the kernel scheduled tasks are synchronized across all RT processors.)
Indeed, people can write shit code in any language.
Trying to avoid a problem is fine, but sometimes people still do silly things for irrational reasons. =3
Ultimately I think artificial scarcity/limited production for products offered online is a futile and erroneous venture. The bots exist because the product is valuable. The product is valuable because it's artificially limited.
There's a niche market for certain player's sneakers. If Nike services them at a price point (e.g. $200 base) then people can keep buying from Nike and will buy the next version.
If instead some sneakerbot takes ALL the consumer surplus (buys for $200, resells for $1000) yes that is the actual demand curve but it hurts NIKE'S business and the consumer.
Nike can make them not worth $1000 and get more revenue by selling more shoes. They are actively shooting themselves in the foot by putting resellers over real customers.
Its fine if you are targeting rich people. It doesn't work if you are targeting (relatively) poor people. You can't have both scarsity and low prices, something has to give.
yes, but people will only pay up to a certain price for a sneaker if they are not collectors and only consumers. and as we now see, a shoe business is not running that well, if it mainly caters to resellers and collectors.
i can remember a few instances about 5 years ago when I was really interested in a nike shoe "drop", but didnt win in the shoe lottery to purchase the show. the result? i just stopped looking at Nikes completely.
Non-collectors don't need what the bots are buying, they can just go to the store and get normal sneakers. I don't see how limited edition shoe scalping affects the normal market at all.
I don't bother buying Nike's because every time I see a pair I like online, it turns out they're "collectibles" and sold out instantly.
Yeah, I can go and buy a pair of black Nikes in any shop today. But even though they make lots of variety, they never seem to make it in enough volume that I can buy it. I'm not interested in being a collector - I just want shoes that aren't black, white or grey.
So I shoes from companies that don't have as much weird hype around them.
Yes, I think this is the main issue. I don't care what policies the AI labs have or enforce, but they need to stop acting like ToS violations are an international crisis demanding intervention instead of a boring civil dispute at most.
Anthropic happily paid billions to settle a lawsuit for pirating books. It's a trivial cost of doing business. If you're lucky you'll get a pittance after the fact by suing them, but a contract doesn't prevent them from doing the thing you don't want them to do and that they are obviously going to do given their past behaviour.
Lucky for us Apple is already alleging something to this effect in their trade secret lawsuit, so you know they'll make sure discovery turns this up if it exists.
ZDR is based on the exact same pinky-promise as training opt-outs. There is no technical barrier to OpenAI, or whoever is running your compute, retaining your prompt after they run inference on their servers. If you don't control the hardware the model is being inferenced on, you don't control your data.
A lot substance is hinged on the exact definition of the word "data" or "user data". In the age of post-truth everyone is claiming that they keep no "user data". Except that after running it once through some transformer program it's no longer "user data", it's something entirely else and these corpos gave ZERO promises regarding such laundered/transformed data at all, ever.
Just a thought experiment: considering training seems to be 'fair use', I wonder if they trained a tiny model to retain key info from your prompts, would mean that this would still constitute fair use, and allow them to legally claim they don't retain your data.
The guarantee on this is a (contractual) “trust me bro”, and a right to try to sue a multi-trillion-dollar company who will absolutely drive you into the ground with legal red tape.
If you are big enough to be able to withstand that, you’re already running (or trying to run) your own/open-weight models.
You dont need an LLM to figure out to make anthrax. Anybody who can figure out how to make a home lab can make all sorts of dangerous stuff pretty easily. Same with college grad from a respectable chemistry program. This all FUD.
If you weren't arguing against the business case, what were you arguing with the C-level types about? I don't think I have ever heard a case for react that wasn't centered around shared codebases mean fewer resources/cost/time to deliver a feature.
On the other hand, the Internet Archive is a non-profit offering a free public resource.
reply