SpruceChat runs Qwen2.5-0.5B on handheld gaming devices using llama.cpp. no cloud, no wifi needed. the model lives in RAM after first boot and tokens stream in one by one. runs on: Miyoo A30, Miyoo Flip, Trimui Brick, Trimui Smart Pro performance on the A30 (Cortex-A7, quad-core): - model load: ~60s