Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

If the Terminal Bench 4.0 scores are to be believed[0] GPT-6.1 is an incredibly efficient model.

Yes, benchmarks aren't real work blah blah, but the delta here is so large compared to Astra, it makes it seem like this is distilled Bel or similar.

[0]https://x.com/thsottiaux/status/2105007628460109953

 help



Consider applying for YC's Winter 2027 batch! Applications are open till November 2.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: