/api/ps immediately after the unload call: {"models":[]} prompt: "Reply with a single short sentence: why do developers like small local models?" response: Developers often prefer small local models because they are more efficient, faster, and easier to deploy and train. total-duration: 901ms load-duration: 1ms prompt-eval-count: 44, prompt-eval-duration: 39ms eval-count: 23, eval-duration: 856ms 26.87 tokens/sec (eval-count / eval-duration)