
MicroPrompt: Running Small Language Models on an Apple Watch
I wanted to see whether my first-generation Apple Watch SE could run a small language model locally. Notes on getting llama.cpp working, reducing the wait for answers, and adding a few tools.