Artificial Analysis launches benchmarks for small models on mobile phones

ArtificialAnlys · x · 2026-08-25

Artificial Analysis, in partnership with Liquid AI, has launched intelligence and inference benchmarks for small models on mobile phones. The intelligence benchmark uses five evaluations tailored for function calling, knowledge, and reasoning tasks. The inference benchmark measures speed, latency, and memory on real devices. The article highlights that benchmarking small models is difficult due to the clustering of scores on traditional frontier benchmarks and the immaturity of mobile runtimes. Tests use 4-bit or smaller quantized builds.

Related event: Artificial Analysis and Liquid AI Launch On-Device Small Model Benchmarks(8 posts)→

Original post →

More from Infra

Infra channel →