Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Dude im hoping we get rid of Nvidia completely. I can run llama.cpp inference on a 7B model on my 24 core cpu intel machine using just CPU and it only uses about 4gb of ram and is not that slow. If we could have massive parallel arm core or even riscv machines without the Cuda issues with proprietary driver hell it would be much more open source. And much less wonkage for the normie user


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: