Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

How can the overall trajectory length be the same across reasoning efforts? I don't see how this is possible even if reasoning is not included in the trajectory length calculation.
 help



I think Tibo was just keeping all else fixed and it’s an illustrative example rather than a perfect real-world trajectory.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: