After post-training, the model’s output distribution collapses a ton.
I feel like the recent models though are extremely annoying to use. GPT 3/3.5/4/4o/4.1/4.5 and even 5 was fine. There were patterns (like everything being a list for 3-series) but nothing this annoying.
Ever since they started optimizing thinking and especially coding, these models have gone down the shitter in terms of their “writing style”
