Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> Speed is the major limiting factor for high-level automation.

Yes, but the point is the quality of inference is more important than speed. What good is speed if inference is shit?



It's not a tradeoff in this case, this is an optimized megakernel for the same model for better throughput. And no, in most cases accuracy can be sacrificed in favor of throughput or latency (assessing it automatically is the harder part).




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: