When we consider:
* LLM usage is new for the world
* Models are evolving quickly with high worldwide competition
* Hardware is evolving despite RAM shortages
Is investing a huge sum of money in equipment for local inference a wise use of money? Or are M5 Ultra and equivalently priced local inference hardware future-proof enough to be worth it relative to how the market is evolving? Maybe it’s all a question of what you’d spend otherwise on serverless or dedicated GPU spend…
loading story #49481561
loading story #49483794
loading story #49481977
loading story #49486180
loading story #49483056
loading story #49482101
loading story #49482518
loading story #49482228
loading story #49482103