Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

It's just another Qwen 3.6 27B finetune (not even 3.8...), claiming to beat TB-sized models. This is probably the 57th model like that, just this week, claiming the same thing... And they never deliver, to say the least.


Didn’t they deliver something here? Why does it matter what base model they used?


What they all deliver is benchmaxxed models that are awful in actual use.


But it gives the lab good publicity, as seen in the article.


I'm giving out more unpopular opinions today for entertainment and thought :)

Every model is benchmaxxing today. Many of them are awful until tuned for any particular task.


Qwen is a really awesome model to use!


> (not even 3.8...)

3.6 is better for non-coding tasks, noticeably so.


The finetunes will continue until morale increases.




Consider applying for YC's Winter 2027 batch! Applications are open till November 2.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: