Benchmarking Pocket-Scale Inference

What happened
Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
Summary assembled by rule from the sources below

Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
Summary assembled by rule from the sources below