Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines
A probing methodology estimates frontier models’ knowledge cutoffs, training timelines, data mixtures, and exposure to other models’ outputs. HN discusses how fixed API weights, post-training, distillation, and product-layer updates complicate those inferences.