Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

Specific task based benchmarks don't reflect a lot of day to day agentic use cases in my experience. If you are working on a series of discrete tasks and can clear context after each one and move to the next, you might get that sort of efficiency from Opus low effort. I often find that when working through a real problem, iterating and discovering, context length can creep up, and that is where opus tends to get expensive.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: