Earlier quoted context omitted.
For that to happen they'd have to be the owners of the Copyright, they're not exactly patent trolls who can use the law creatively. The Newspapers are largely regional (and mostly defunct) papers. Who at the Derby Evening Telegraph is going to waste lawyers fees getting a Library to remove 100year old content?
Patent trolls own the patent.
The British Library puts 1M newspaper pages online for free
51–60 of 60 posts
Re: The British Library puts 1M newspaper pages online for free
#52Earlier quoted context omitted.
From personal experience, I can tell you that there are automated systems searching for such pieces of art. As for newspapers, all it takes is one copyright troll to realize it's happening, and suddenly there's lawsuit settlements everywhere, making him rich.
For that to happen they'd have to be the owners of the Copyright, they're not exactly patent trolls who can use the law creatively. The Newspapers are largely regional (and mostly defunct) papers. Who at the Derby Evening Telegraph is going to waste lawyers fees getting a Library to remove 100year old content?
Re: The British Library puts 1M newspaper pages online for free
#53Trove does this for Australian newspapers https://trove.nla.gov.au/
Like the Carrington Event & Krakatoa, or Alexander Graham Bell's firewall to stop people stealing electricity, or pirate attacks on junks off Hong Kong.
Re: The British Library puts 1M newspaper pages online for free
#54See: https://foreignpolicy.com/2012/07/03/how-did-the-british-pre...
Re: The British Library puts 1M newspaper pages online for free
#55Do anyone know any existing effort on converting these scanned image to text corpus ( probably a new OCR model needed to be developed on these old text ) ? I think it would be more usable if they are in text form in terms of search and research purpose.
Re: The British Library puts 1M newspaper pages online for free
#56Re: The British Library puts 1M newspaper pages online for free
#57Do anyone know any existing effort on converting these scanned image to text corpus ( probably a new OCR model needed to be developed on these old text ) ? I think it would be more usable if they are in text form in terms of search and research purpose.
Re: The British Library puts 1M newspaper pages online for free
#58This is great. Historical newspapers are one of the largest corpora of information that has yet to be adequately brought on line. In the U.S. the Library of Congress has digitized a fair number, but at the state and local level it's really hit or miss. Some states such as California and New York have put quite a bit on line, but many others rely on individual towns and historical societies. Different pay services cov…
> Historical newspapers are one of the largest corpora of information that has yet to be adequately brought on line. Not just information, but works of art as well! A few years ago, I trained an image classifier to help me find Krazy Kat comics in newspaper archives. In the process of doing that, I came across a shocking amount of other comics and artwork. I was honestly surprised to see how many amazing illustration…
Re: The British Library puts 1M newspaper pages online for free
#59Earlier quoted context omitted.
This needs to change, big time. There is almost no cash value to an article one day later, yet we completely impoverish the public domain for its sake. Not only are creators of valuable works is usually pretty distant from direct ownership anyway, there's no possible way for them to profit directly from this work. The only way a spotify-like deal works is because copyright ownership is conglomerated. IMO, public (esp…
> There is almost no cash value to an article one day later, yet we completely impoverish the public domain for its sake. We're not imposing publishing restrictions on past works in order to preserve their cash value. We're imposing publishing restrictions on past works in order to stop them from competing with present works. It keeps the cash value of present works up.
Even if the works are not economically valuable, people's attention is, as is gatekeeping control over archives (allowing one to set narratives and agendas and control context).
The time I spend going through an archive (I immediately hit the 3-article registration wall at the BLNA, so ... little time) is time I'm not spending consuming the present-moment adverts-laden infotainment stream.