Hacker News new | ask | show | jobs
by deno 19 days ago
Unless I'm missing something I think to a large degree you're just comparing system prompts.

If I add "Research the question extensively" to your prompt at the end I get the correct answer from Haiku and Sonnet Med on first try and I've reproduced the original prompt not returning the answer.

Unfortunately every other run now gets your gist in results.

1 comments

It could be variations in system prompt but I'm not sure how "change the user prompt to encourage tool usage" proves that either way
The point of this was supposed to be "model tool use competency" not baseline over-eagerness.