Small models, real work: running inference locally
When a laptop is enough, and how to tell before you provision a single GPU.
Felipe Goncalves1 min read
When a laptop is enough, and how to tell before you provision a single GPU.
[Draft: the full article goes here.]