Live data from Hacker News

History LLMs: Models trained exclusively on pre-1913 texts

github.com

191–200 of 452 posts

Re: History LLMs: Models trained exclusively on pre-1913 texts

#191
post #46

I wonder if you could query some of the ideas of Frege, Peano, Russell and see if it could through questioning get to some of the ideas of Goedel, Church and Turing - and get it to "vibe code" or more like "vibe math" some program in lambda calculus or something. Playing with the science and technical ideas of the time would be amazing, like where you know some later physicist found some exception to a theory or some…

This is my curiosity too. Would be a great test of how intelligent LLM's actually are. Can they follow a completely logical train of thought inventing something totally outside their learned scope?

You definitely won't get that out of a 4B model tho.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#192
post #184

Earlier quoted context omitted.

Time to create the Torment Nexus, I guess

There's a thriving startup scene in that direction.

Wasn't that the elevator pitch for Palentir?

Still can't believe people buy their stock, given that they are the closest thing to a James Bond villain, just because it goes up.

I mean, they are literally called "the stuff Sauron uses to control his evil forces". It's so on the nose it reads like an anime plot.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#193

Earlier quoted context omitted.

This isn’t science fiction anymore. CIA is using chatbot simulations of world leaders to inform analysts. https://archive.ph/9KxkJ

Zero percent chance this is anything other than laughably bad. The fact that they're trotting it out in front of the press like a double spaced book report only reinforces this theory. It's a transparent attempt by someone at the CIA to be able to say they're using AI in a meeting with their bosses.

Unless the world leaders they're simulating are laughably bad and tend to repeat themselves and hallucinate, like Trump. Who knows, maybe a chatbot trained with all the classified documents he stole and all his twitter and truth social posts wrote his tweet about Ron Reiner, and he's actually sleeping at 3:00 AM instead of sitting on the toilet tweeting in upper case.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#194
post #97

Wait so what does the model think that it is? If it doesn't know computers exist yet, I mean, and you ask it how it works, what does it say?

What would a human say about what he/she is or how he/she works ? Even today, there's so much we don't know about biological life. Same applies here I guess, the LLM happens to be there, nothing else to explain if you ask it.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#196
I'd love for Netflix or other streaming movie and series services to provide chat bots that you could ask questions about characters and plot points up to where you have watched.

Provide it with the closed captions and other timestamped data like scenes and character summaries (all that is currently known but no more) up to the current time, and it won't reveal any spoilers, just fill you in on what you didn't pick up or remember.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#197
Everyone learns that the renaissance was sparked by the translation of Ancient Greek works.

But few know that the Renaissance was written in Latin — and has barely been translated. Less than 3% of I’m working on a project to change that. Research blog at www.SecondRenaissance.ai — we are starting by scanning and translating thousands of books at the Embassy of the Free Mind in Amsterdam, a UNESCO-recognized rare book library.

We want to make ancient texts accessible to people and AI.

If this work resonates with you, please do reach out: Derek@ancientwisdomtrust.org

Re: History LLMs: Models trained exclusively on pre-1913 texts

#198
post #154

Ontologically, this historical model understands the categories of "Man" and "Woman" just as well as a modern model does. The difference lies entirely in the attributes attached to those categories. The sexism is a faithful map of that era's statistical distribution. You could RAG-feed this model the facts of WWII, and it would technically "know" about Hitler. But it wouldn't share the modern sentiment or gravity. In…

I think much of the semantic proximity to evil can be derived straight from the facts? Imagine telling pre-1913 person about the holocaust.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#199
post #57

Earlier quoted context omitted.

This is definitely fascinating - being able to do AI brain surgery, and selectively tuning its knowledge and priors, you'd be able to create awesome and terrifying simulations.

Respectfully, LLMs are nothing like a brain, and I discourage comparisons between the two, because beyond a complete difference in the way they operate, a brain can innovate, and as of this moment, an LLM cannot because it relies on previously available information. LLMs are just seemingly intelligent autocomplete engines, and until they figure a way to stop the hallucinations, they aren't great either. Every piece o…

> LLMs are just seemingly intelligent autocomplete engines

BINGO!

(I just won a stuffed animal prize with my AI Skeptic Thought-Terminating Cliché BINGO Card!)

Sorry. Carry on.

Re: History LLMs: Models trained exclusively on pre-1913 texts

#200

It would be interesting to see how hard it would be to walk these models towards general relativity and quantum mechanics. Einstein’s paper “On the Electrodynamics of Moving Bodies” with special relativity was published in 1905. His work on general relativity was published 10 years later in 1915. The earliest knowledge cuttoff of these models is 1913, in between the relativity papers. The knowledge cutoffs are also r…

> It would be interesting to see how hard it would be to walk these models towards general relativity and quantum mechanics. Definitely. Even more interesting could be seeing them fall into the same trappings of quackery, and come up with things like over the counter lobotomies and colloidal silver. On a totally different note, this could be very valuable for writing period accurate books and screenplays, games, etc…

Accurate-ish, let's not forget their tendency to hallucinate.
Post reply on HN