Y
Hacker News
new
|
ask
|
show
|
jobs
by
prime_ursid
29 days ago
This is a technique called LLM-as-Judge. Well-studied at this point. Here’s a good intro:
https://www.evidentlyai.com/llm-guide/llm-as-a-judge
I recommend reading Hamel.dev posts. Here’s an example:
https://hamel.dev/blog/posts/evals/
1 comments
WickyNilliams
29 days ago
Awesome, thank you! I am starting to do some work building LLM workflows and would like to stand on some giant's shoulders to skip the initial flailing :)
link