I don't think there's anything like that going on. They just word vomit into a secondary area, and then there is an internal prompt that says "clean this up and summarize for the user".
Less "internal prompt" and more "they are trained to summarize after a </think> token"
Less "internal prompt" and more "they are trained to summarize after a </think> token"