Hacker News new | past | comments | ask | show | jobs | submit
A 2s LLM call is pretty slow.
Try using Digital Ocean. Minutes spent on inference.