Get the latest tech news

Benchmarking Pocket-Scale Inference


Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.

None

Get the Android app

Or read this on Hacker News

Read more on:

Photo of Scale Inference

Scale Inference