Live data from Hacker News

Google Books (or similar) all book scans – $200k bounty (2025)

software.annas-archive.gl

21–30 of 368 posts

Re: Google Books (or similar) all book scans – $200k bounty (2025)

#21
post #18

The US should just find a way to quietly share literature access with the Russians, rather than letting piracy be promoted and facilitated for US consumers as freedom-fighter "archiving". Between all the piracy, and all the AI training and the purchase/visitor-circumventing AI services, the practice of writing and publishing genuinely good work is being wiped out. We're killing the goose that lays the eggs, for selfi…

[deleted]

Re: Google Books (or similar) all book scans – $200k bounty (2025)

#23
post #9
post #8

Earlier quoted context omitted.

Chinese companies giving away expensive models for free is a symptom of the AI bubble, too. It's not a law of nature that they'll always be able to scrounge up the money for yet another training run.

Shaping the tool that does the thinking is quite valuable when you're in the business of changing how people think - I think we can expect propaganda agencies to be subsidizing model creation forever. This doesn't strike me as a symptom of a bubble - except in so far as the bubble pushes the competitors models forwards and thus they need to invest more to stay competitive.

All the models, have to respect their local laws, and most of all, pressure from users and the employees.

They all carry political weights, because humans behind defend their interests, and are promoting some social values.

https://pastebin.com/hjhvsBFg

This answer from Claude is so biased that it is ridiculous

Re: Google Books (or similar) all book scans – $200k bounty (2025)

#28

One of my hopes is that when the AI bubble bursts, some brave person will sneak out a copy of the last frontier model.

Not worried about that, you will only have to wait 3-6 months and get a Chinese model just as good.

That’s misunderstanding why these models are behind. A large part of why they’re behind is they aren’t able to do the reinforcement learning post-training steps that takes a pre-trained model and turns it into a frontier model like GPT 5 or Opus. Instead they do their best to recreate these models using distillation.

Fundamentally, you can never distill your way to being the teacher, so these approaches will not advance the frontier.

[edit, after thinking about it I think my phrasing is unfair. It's not necessarily that aren't able to do it, but they haven't yet shown that they are willing to do it.]

Re: Google Books (or similar) all book scans – $200k bounty (2025)

#30
post #18

The US should just find a way to quietly share literature access with the Russians, rather than letting piracy be promoted and facilitated for US consumers as freedom-fighter "archiving". Between all the piracy, and all the AI training and the purchase/visitor-circumventing AI services, the practice of writing and publishing genuinely good work is being wiped out. We're killing the goose that lays the eggs, for selfi…

Possibly but this act of governmental self-harm is useful to The People. We live in a world where if your valuation is ~1T you can more or less just do what you like. And the work of The People is stolen from you and launderd.

In such a world, isnt it useful that governments are stupid enough to give adversaries reasons to undermine it? When the government props up a corporate tyranny domestically, and racketeering, should we make a temporary alliance with all its enemies?

(Eg., the provision to AI companies of all corporate secretes and competitive practices via prompts, eventually to be used against their capital interests and their labour interests).

Post reply on HN