|
|
|
|
|
by jamesrcole
28 days ago
|
|
I would expect a model's result each time to be of a similar quality to the other times. There's something wrong if it does a way better or worse job, at the same problem, sometimes. It's possible, but I haven't heard anyone saying that they do. |
|
People use LLMs to do vulnerability scanning by throwing them repeatedly at a codebase. Depending on the run they return with nothing, with a false positive, with a true vulnerability. These are very different destinations when faced with the same problem, sometimes.
Since GPT2, people have been throwing a ton of crap at the wall just to pick out one nugget that's uncharacteristically more solid than the others. Honestly? It's not just possible—it's core to how they operate. And it always has been.