In some applications, AI is indeed already training models smaller than itself. When OpenAI unveiled its GPT-5.6 series in July this year, for example, it explained how Luna, the small model, was post-trained. After Luna completed its initial pre-training, the larger model, Sol, carried out its post-training, adjusting the parameters itself and following the same approach used in its own post-training. The researchers gave only fairly brief instructions. Sol handled everything itself, from choosing the training configuration and allocating GPUs to launching the run and checking that all was well. It's rather like a big model bringing up a little one.