Hacker News new | ask | show | jobs
by k-thimmaraju 75 days ago
Very interesting approach to showing the relationship between RLHF and AI Psychosis: the idea of taking clinical conversations and prompting the model with it seemed like a grounded start. As I'm also investigating AI Psychosis, this approach seems like something to adopt for my work.
1 comments

I ran the data through our LLM behavioral analysis system at splabs.io, and it alerts Red on the RLHF-optimized output compared to a Yellow on the no-RLHF.

You can check out the analysis here: https://splabs.io/ai-psychosis-and-cognitive-cost