Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> While its overall performance still trails the most powerful proprietary models, Claude Fable 5 and GPT 5.6 Sol

I'm inclined to believe that, however according to their own benchmarks Kimi K3 actually even beats the other two in many metrics, no?



Benchmarks are meaningless. You can beat any model on any benchmark with the right training.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: