Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

No, as the sibling comment mentioned, your understanding was correct there.

What’s more, the only technical difference in speeds could be, and likely also is with the HC1 chip, between prefill (prompt processing) and decode (text generation) speeds. I don’t know whether it’s the case with Taalas’ chip, but in the “software-based” LLMs we typically see and use so far, those two stages hit different parts of a computer (processing/compute-bound vs. memory/bandwidth-bound).



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: