Get the latest tech news
Benchmarking Pocket-Scale Inference
Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
None
Or read this on Hacker NewsGet the latest tech news
Benchmarking of language model inference on mobile phones. We measure output speed, latency and memory use across local models.
None
Or read this on Hacker News