Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

It's not like same parameter count models are identical, so that doesn't appear to be an indicator for quality, or even compute requirements?

There seems to be more to producing a better model than brute forcing parameter count after all.



Training and serving large models does require increasingly more compute, though. (The Chinese labs have clearly found some massive optimizations, but my point was that you'd think at some point even those optimizations wouldn't be enough to keep up with exponentially increasing model sizes.)




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: