Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This is needs ~80GB of fast memory at 4 bits per weight. Faster memory is better, but probably even something like 3090 + 64GB RAM should work (not fast, but maybe even 20-30 t/s? llama.cpp support pending).


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: