If you have 500GB of SSD, llama.cpp does disk offloading -> it'll be slow though less than 1 token / s
3 t/s isn't going to be a lot of fun to use.
If you have 500GB of SSD, llama.cpp does disk offloading -> it'll be slow though less than 1 token / s