Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
1–10 of 21 posts
Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#2I also assumed many questions get routed to simpler models or programs to answer correctly, but it almost surprisingly didn't seem that way from the post.
Anyways, great post.
Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#3Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#4Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#5But there is no real way to know how much of this "waiting" any lab is doing, if we can get better estimates this way maybe we can gauge how far the open weights models really are.
Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#6I wonder if this kind of analysis will give us a way to check if the frontier labs are waiting for the right moment to release their models. To me it feels obvious that these companies are not releasing models as soon as they are done doing their post training / testing with any new model. But there is no real way to know how much of this "waiting" any lab is doing, if we can get better estimates this way maybe we ca…
Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#7I wonder if this kind of analysis will give us a way to check if the frontier labs are waiting for the right moment to release their models. To me it feels obvious that these companies are not releasing models as soon as they are done doing their post training / testing with any new model. But there is no real way to know how much of this "waiting" any lab is doing, if we can get better estimates this way maybe we ca…
If you have the new "best" model you may only have a few weeks of time in the market before some other lab releases a new model that beats yours.
This means you should get it out ASAP so you can maximize the time during which your model is "best". Once your model isn't best any more you're going to lose a lot of revenue to the new leader.
Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#8Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#9Re: Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
#10From my own experiments with various models, I suspect that LLMs have distinct/partitioned cutoff dates; for example, historical literature doesn't change (Greek history, Shakespeare, Goethe), general knowledge (updated only in certain areas), technologies (updated regularly), software also remains surprisingly stable - for example, with GIT, a basic command set is sufficient to do 99% of the jobs - new features are…