History LLMs: Models trained exclusively on pre-1913 texts
391–400 of 452 posts
Re: History LLMs: Models trained exclusively on pre-1913 texts
#392For example prompt the 1913 model to try and “Invent a new theory of gravity that doesn’t conflict with special relativity”
Would it be able to eventually get to GR? If not, could finding out why not illuminate important weaknesses.
Re: History LLMs: Models trained exclusively on pre-1913 texts
#393Interesting ... I'd love to find one that had a cutoff date around 1980.
Excellent question! It looks like Two-Tone is bringing ska back with a new wave of punk rock energy! I think The Specials are pretty special and will likely be around for a long time.
On the other hand, the "new wave" movement of punk rock music will go nowhere. The Cure, Joy Division, Tubeway Army: check the dustbin behind the record stores in a few years.
Re: History LLMs: Models trained exclusively on pre-1913 texts
#394Earlier quoted context omitted.
we're on the same page.
Although... Self preservation is the first law of nature. If you release the model someone will basically say you endorse those views and you risk your funding being cut. You created Pandora's box and now you're afraid of opening it.
That should be more than enough to clear any chance of misunderstanding.
Re: History LLMs: Models trained exclusively on pre-1913 texts
#395Earlier quoted context omitted.
we're on the same page.
Although... Self preservation is the first law of nature. If you release the model someone will basically say you endorse those views and you risk your funding being cut. You created Pandora's box and now you're afraid of opening it.
Re: History LLMs: Models trained exclusively on pre-1913 texts
#396Earlier quoted context omitted.
It’s as if every researcher in this field is getting high on the small amount of power they have from denying others access to their results. I’ve never been as unimpressed by scientists as I have been in the past five years or so. “We’ve created something so dangerous that we couldn’t possibly live with the moral burden of knowing that the wrong people (which are never us, of course) might get their hands on it, so…
Wow, this is needlessly antagonistic. Given the emergence of online communities that bond on conspiracy theories and racist philosophies in the 20th century, it's not hard to imagine the consequences of widely disseminating an LLM that could be used to propagate and further these discredited (for example, racial) scientific theories for bad ends by uneducated people in these online communities. We can debate on wheth…
Re: History LLMs: Models trained exclusively on pre-1913 texts
#397> Historical texts contain racism, antisemitism, misogyny, imperialist views. The models will reproduce these views because they're in the training data. This isn't a flaw, but a crucial feature—understanding how such views were articulated and normalized is crucial to understanding how they took hold. Yes! > We're developing a responsible access framework that makes models available to researchers for scholarly purp…
> So is the model going to be publicly available, just like those dangerous pre-1913 texts, or not? 1. This implies a false equivalence. Releasing a new interactive AI model is indeed different in significant and practical ways from the status quo. Yes, there are already-released historical texts. The rational thing to do is weigh the impacts of introducing another thing. 2. Some people have a tendency to say "releas…
Re: History LLMs: Models trained exclusively on pre-1913 texts
#398Earlier quoted context omitted.
This is the 2023 take on LLMs. It still gets repeated a lot. But it doesn’t really hold up anymore - it’s more complicated than that. Don’t let some factoid about how they are pretrained on autocomplete-like next token prediction fool you into thinking you understand what is going on in that trillion parameter neural network. Sure, LLMs do not think like humans and they may not have human-level creativity. Sometimes…
>> Sometimes they hallucinate. For someone speaking as you knew everything, you appear to know very little. Every LLM completion is a "hallucination", some of them just happen to be factually correct.
Re: History LLMs: Models trained exclusively on pre-1913 texts
#399Re: History LLMs: Models trained exclusively on pre-1913 texts
#400The sample responses given are fascinating. It seems more difficult than normal to even tell that they were generated by an LLM, since most of us (terminally online) people have been training our brains' AI-generated text detection on output from models trained with a recent cutoff date. Some of the sample responses seem so unlike anything an LLM would say, obviously due to its apparent beliefs on certain concepts, t…
I used to teach 19th-century history, and the responses definitely sound like a Victorian-era writer. And they of course sound like writing (books and periodicals etc) rather than "chat": as other responders allude to, the fine-tuning or RL process for making them good at conversation was presumably quite different from what is used for most chatbots, and they're leaning very heavily into the pre-training texts. We d…