/api/ps immediately after the unload call: {"models":[]}

prompt: "Reply with a single short sentence: why do developers like small local models?"
response: Developers often prefer small local models because they are more efficient, faster, and easier to deploy and train.

total-duration: 901ms
load-duration: 1ms
prompt-eval-count: 44, prompt-eval-duration: 39ms
eval-count: 23, eval-duration: 856ms
26.87 tokens/sec (eval-count / eval-duration)