Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I've done my own benchmarks where I've hit over 100 TFLOPS on the V100, and that's about 85% of the peak theoretical throughout of them. Granted the matrix size needs to be large enough, but it's definitely doable. Anandtech also showed similar results in their V100 review. I haven't yet seen a comparable SGEMM done on the 2080Ti, so I don't know how it'll compare. I have some coming in at the end of the month though, so I should know soon.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: