|
|
|
|
|
by TZubiri
4 days ago
|
|
Most people seem to think that agreeableness is a personaility thing that vendors can just turn up or down at will. But the usefulness of LLMs comes from following what you say. An LLM that follows your lead when you say "The answer to the collatz conjecture is" is much more useful than one that answers "not known and if you think you know it you are wrong." Reminds me of the tip about working with lawyers, if you ask them whether you can do something, the answer will often be no. However if you ask them how you can do something, they tend to give you more advice on how to do it. |
|
Because it is, more or less ... it is an emergent property of RLHF (reinforcement learning from human feedback), and that feedback follows corporate policy.
> An LLM that follows your lead when you say "The answer to the collatz conjecture is" is much more useful
What lead? Follow it where? How tf am I or the LLM supposed to know what response to that is something you consider far more useful than the truth?
> than one that answers "not known"
I value the truth and that is the truth. Of course I expect an LLM to respond with a lot more detail about the CC, why it's difficult, what progress has been made (e.g., the N for which all values <= N have been shown to satisfy the conjecture and Terence Tao's proof that "almost" all integers satisfy the conjecture), the fact that most mathematicians think the conjecture is true, etc. -- and they do.
> and if you think you know it you are wrong."
Where tf did that come from? The query said nothing about knowing the answer. Don't project being a snarky ah onto the LLMs for no apparent reason.
> Reminds me of the tip about working with lawyers, if you ask them whether you can do something, the answer will often be no.
Irrelevant and inappropriate analogy. Lawyers (among others) want to avoid committing to something that they can be held liable for. LLMs clearly have no such limitations, as they often give wrong advice quite authoritatively.