Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Their claim is not about the prompt or skill tokens, it's about output tokens - skills can help the model bypass some thinking tokens or avoid reasoning deadends, and that way reduce output token usage. That's what they seem to have found empirically from their testing. (If it's truly 2x-4x, the time savings in waiting for the output is a pretty nice benefit too.)


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: