Hacker News new | past | comments | ask | show | jobs | submit
> Local: Typical scenario is hours just to download the software, the model weights, and then faffing around with CUDA and matching your GPU drivers.

Download LM studio, search models, click download, wait minutes, prompt and have fun

“If you have the prerequisite hardware, then… know which model you want out of thousands of a variants… and your drivers are up to date, then it is fast!”