Links indicate relevance, not agreement. How to use this site →
Methods for inferring hidden details about how large language models like GPT-5 and Claude were trained, including estimating parameter counts, dataset mixtures, and training timelines by probing models with carefully curated questions. Covers the three main stages of frontier model training (pre-training, capability fine-tuning, and post-training) as context for understanding what these probing techniques reveal.