Latest note
10 Small Language Models on an iPhone: 4K/8K Memory and Agent Accuracy
A controlled comparison of 10 compact GGUF models, including Llama 3.2, SmolLM, and LFM2.5, with iPhone 4K/8K resource data and a 96-case bilingual Agent test.
on-device AILLMmobile inferencebenchmarkagent routing