Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Unless I'm missing something I think to a large degree you're just comparing system prompts.

If I add "Research the question extensively" to your prompt at the end I get the correct answer from Haiku and Sonnet Med on first try and I've reproduced the original prompt not returning the answer.

Unfortunately every other run now gets your gist in results.



It could be variations in system prompt but I'm not sure how "change the user prompt to encourage tool usage" proves that either way


The point of this was supposed to be "model tool use competency" not baseline over-eagerness.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: