SPIEval: Evaluating LLMs as Mobile Assistants over Scattered Personal Information

Junjie Ye · hf · 2026-08-12

SPIEval benchmarks large language models in mobile assistant scenarios, focusing on their ability to handle scattered personal data. The evaluation reveals significant gaps in current models regarding information retrieval and verification.

Original post →

More from Apps

Apps channel →