|
|
|
|
|
by bensyverson
20 hours ago
|
|
I was designing a system that pulls structured data out of an unstructured transcript, and Opus initially suggested a confidence score for each value. When I asked it “do you think the model will generate a meaningful number there?” the answer was basically “ha, you caught me.” What did work was asking the model to emit citations in the form of line numbers from the transcript. It dramatically dropped hallucination. |
|
This also isn't likely to give you good results, models don't have a good sense of their own capabilities.
This is also in the same camp of asking a model why it did something. It'll generate something plausible and likely totally incorrect.