Hacker News new | ask | show | jobs
by Barbing 18 days ago
Do technologists have more respect for the idea you can train a model to be on your side with a constitution than they might’ve at first?

I'm sure the concept seemed just about purely preposterous to many when the models were in their infancy. Now I figure instead it seems mostly preposterous to many.

(Though I guess Anthropic‘s success doesn’t necessarily prove anything about the constitution)

2 comments

I don’t think anyone imagines that it’s an ironclad steering method, but it seems to help, so why not?
Anthropic train it to 'reason morally' based on the constitution's principals.

https://www.anthropic.com/research/teaching-claude-why