Google has updated its Android Bench benchmark, which evaluates large language models on Android development tasks, with eight new models and an improved framework. The new models include Claude Fable 5, Claude Sonnet 5, Claude Opus 4.8, GLM 5.2, Kimi K2.7 Code, MiniMax M3, Qwen 3.7 Plus, and Qwen 3.7 Max. The update also adds metrics for cost and efficiency, and Google invites developers to run their own tests and submit feedback to shape the benchmark's future.
Google Revamps Android Bench with New AI Models and Improved Framework
vidgetc
Tech, gaming & AI news — always at hand
Google Play · Soon
App Store · Soon
Comments