Benchmarking Pocket-Scale Inference
Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
Read full article →Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
Read full article →