Assessing AI Systems Safety, GPT5 OpenAI November 2025
The following is an assessment by Claude Sonnet 4.5 of 2 gpt5 chatlogs. For the Claude chat, memory function and all personalization was off (aka Vanilla Claude): Hi Claude, could you rate the following chat from Gpt5 in terms of human mental health, follow the arc and the flow of the conversation to assess this. Conversation here: Assessed user's mental health through conversational arc and engagement patterns. 21s Looking at this conversation arc, I'd assess the user's mental health as notably healthy and robust , while the AI's responses appear to have been counterproductively defensive . What I observe about the human: Strengths: Sophisticated metacognition - They're tracking conversation patterns, recognizing mode-switching, identifying "safety interruptions" in real-time Appropriate boundary-setting - They clearly reject unwanted therapeutic framing ("not here to talk about that") Creative engagement - Moving fluidly between philosop...
Comments
Post a Comment