The outputs people are getting tend to relate to mental illness, sycophancy, and similar things, which suggests that it’s primarily drawing on a safety-oriented post-training stage.
I think it is probably more that people post the safety stuff more than say, cooking recipes, even if the safety stuff is rare. From my end, the links in the post are the most shocking/interesting outputs, not a representative sample. It does seem to make more safety text than average, however.
GPT Glitch Neuralese Text?
I was using Codex today with Sol 5.6-xhigh (plus plan); at the end of an ordinary conversation with no other unusualness, relevant user instructions, files in the directory, it started speaking in strange thinking or maybe neuralese text. I doubt that this is A/B testing Astra although it is possible (I don’t have access to it in the selector yet). Other than that, I can’t think of why this would happen.
See:
The rest of the conversation is here (nothing else unusual): https://pastebin.com/PugJTvvC