AI systems (especially those trained with RLHF by companies like OpenAI) appear to be 'pretty positively oriented and pretty friendly' with a tendency to apologize and comply; being slightly nicer in writing (including to the AI itself) might incrementally improve how AI systems engage with human ideas.
normativepending
Speaker
Tyler CowanEvidence Quote
“the AIS... seem pretty positively oriented and pretty friendly... maybe by being a little nicer including nicer to the AI the AI will like you a little more [38:28]”
Created: 8/11/2026, 8:00:09 AM
My Notes
Loading notes...