Local AI runs on a 2020 Apple Watch Series 6, beating Raspberry Pi
Better Stack ran the 90M-parameter Falcon H1 LLM on a 2020 Apple Watch Series 6 with 1GB RAM at 15-24 tokens per second. The team adapted llama.cpp to bypass Core ML incompatibility, while a larger 135M-parameter model proved too heavy for the watch.