Hacker News new | ask | show | jobs
by mike_hearn 22 days ago
They're definitely willing to consider it. Read the parts of the paper where they use J-space interpretations to train more ethical behavior into the model by interrupting it.