Live data from Hacker News

Visualizing All ISBNs

annas-archive.org

91–100 of 146 posts

Re: Visualizing All ISBNs

#91

Earlier quoted context omitted.

They explicitly provide that data for you to do as you wish. They are in a grey area, not you. You can download it no problem.

is there legal precedent for that? already asked LLMs so please don't copy/paste an LLM response.

Depends on your jurisdiction.

Re: Visualizing All ISBNs

#92

Earlier quoted context omitted.

They can't even have a tiny fraction of the world's books. Each edition of the book gets a new ISBN... if a book is released as a paperback, hardback, kindle edition, pdf, and epub then there are supposed to be five ISBNs. The vast, vast majority have only been released as dead-tree versions. They have none of those. The books they scan may have an ISBN, but the scans do not have them. Like all Project Gutenberg book…

Worthless semantics in the context of the mission of the project. What you've described is that the archived content can be mapped to multiple ISBNs. It's clear the only element of concern here is the content itself. The failure to preserve a particular binding or printer's choice of typeface is irrelevant. Failing to recognize this requires an almost malicious level of pedantry

A successful archival of one of those ISBNs will light up; four of those ISBNs remain dark. Yet they have that content archived. It means that lighting up the entire grid is not necessary to achieve their goal.

Indeed a bigger problem is that it’s much harder to know which areas of the grid are never going to light up because the ISBN has not been used.

Re: Visualizing All ISBNs

#93

I see that bounty at the bottom, so tossing away my chances here, but this visualization is just asking to be mapped onto a Hilbert Curve. [0] When you "stripe" the data like this, points that are sorted close together could end up pretty far apart, since a distance in the Y axis skips an entire row of data as you move down, rather than a distance in the X axis which is 1-to-1 with the source data. If you map it onto…

And there's a generalized Hilbert curve, the Gilbert curve, for non powers of two rectangular regions [0] (online demo [1]).

[0] https://github.com/jakubcerveny/gilbert

[1] https://jakubcerveny.github.io/gilbert/demo/

Re: Visualizing All ISBNs

#94

The thing is, ISBNs aren't hierarchical --- they are bought in blocks (or even individually at an exorbitant markup, says the guy who bought one to reprint a single book), so this doesn't show anything really interesting/useful. A visualization using LoC or even Dewey Decimal would be far more useful, esp. if it also linked to public domain and copyright-free repositories/lists, say an interactive and visual version…

One thing it shows is how ISBNs are allocated much faster than they are used, judging by the amount of black pixels.

The image contains 1000*800 pixels at 2500 ISBNs per pixel, so it's visualizing 2e9 ISBNs. ISBN-13 contains 12 digits plus one check digit, so we might have expected the image to be 500 times bigger/denser than the current image. The fact that it's at its current size suggests that only ISBNs with 978 and 979 prefixes are included, and since the bottom half is more sparse, that probably corresponds to the new 979 range.

Re: Visualizing All ISBNs

#96

Earlier quoted context omitted.

I don't think this page, which links to libgen and sci-hub, is that concerned about copyright.

annoying non-answer to my question. i already know all about anna's archive. i'm asking if a person can download these isbns and use them to make data visualizations without fear of breaking a law? https://software.annas-archive.li/AnnaArchivist/annas-archiv...

Seeing as nobody has provided a real answer. The question is, maybe.

Anna's Archive is getting sued currently for scraping vast amounts of essentially public metadata which was being gate-keeped by a single organisation.

Here's the longer and more complicated answer for you:

https://libraries.emory.edu/research/copyright/copyright-dat...

Re: Visualizing All ISBNs

#97

Earlier quoted context omitted.

annoying non-answer to my question. i already know all about anna's archive. i'm asking if a person can download these isbns and use them to make data visualizations without fear of breaking a law? https://software.annas-archive.li/AnnaArchivist/annas-archiv...

Seeing as nobody has provided a real answer. The question is, maybe. Anna's Archive is getting sued currently for scraping vast amounts of essentially public metadata which was being gate-keeped by a single organisation. Here's the longer and more complicated answer for you: https://libraries.emory.edu/research/copyright/copyright-dat...

feist is what comes up when i search around, too. the ISBNs might be poisoned if anna broke terms of service to get the ISBNs

Re: Visualizing All ISBNs

#98

Earlier quoted context omitted.

They can't even have a tiny fraction of the world's books. Each edition of the book gets a new ISBN... if a book is released as a paperback, hardback, kindle edition, pdf, and epub then there are supposed to be five ISBNs. The vast, vast majority have only been released as dead-tree versions. They have none of those. The books they scan may have an ISBN, but the scans do not have them. Like all Project Gutenberg book…

Worthless semantics in the context of the mission of the project. What you've described is that the archived content can be mapped to multiple ISBNs. It's clear the only element of concern here is the content itself. The failure to preserve a particular binding or printer's choice of typeface is irrelevant. Failing to recognize this requires an almost malicious level of pedantry

>Worthless semantics in the context of the mission of the project.

Hardly worthless... often times, the edition of the book matters as much as the title. Steven King wrote two books named The Stand, and one isn't anything like the other. He pulled a Lucas pretty early on.

He's hardly the only author to ever do this. But it's not just authors either. Editors, collectors, translators all make their mark, and give you works that though they might be slightly different to you, the differences actually matter to the rest of us. It's not that you're ignorant that offends me, it's the arrogance about a subject you seem to know so little about that makes it difficult to tolerate.

There is no pedantry here, just a desire to actually preserve books and to organize them.

Re: Visualizing All ISBNs

#99
post #83
post #76

Earlier quoted context omitted.

> The books they scan may have an ISBN, but the scans do not have them. Like all Project Gutenberg books, their books have no ISBNs at all. From a strict point of view, they've released new editions of these books. Are you saying they actively remove ISBN numbers from scans? If I downloaded one of the books, it wouldn't have an ISBN? Why? That seems like a bunch of extra processing per book, makes it harder for users…

> Are you saying they actively remove ISBN numbers from scans? No, he‘s playing the pointless „well, actually a scan of a book is a different thing from the book itself“ game.

No, I'm saying that the ISBN doesn't describe titles, it describes editions, and editions matter.

Re: Visualizing All ISBNs

#100
post #78

It appears that the IP of the server is blocked in the EU. I get this from my ISP (Ziggo, in the Netherlands): Deze website is geblokkeerd Europese sancties De Raad van Europa heeft besloten dat de websites van RT (voorheen Russia Today) en Sputnik News niet meer mogen worden doorgegeven. De website die je probeert te bezoeken, valt onder deze Europese sanctie. VodafoneZiggo is verplicht de sanctie uit te voeren en h…

Running your own recursive resolver has certain advantages…
Post reply on HN