Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> While its overall performance still trails the most powerful proprietary models, Claude Fable 5 and GPT 5.6 Sol

I'm inclined to believe that, however according to their own benchmarks Kimi K3 actually even beats the other two in many metrics, no?



Benchmarks are meaningless. You can beat any model on any benchmark with the right training.




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: