Google Expands Android Bench with New LLMs, but Gemini Still Lags Behind

Google has updated its Android Bench benchmark for evaluating large language models on Android development tasks, adding eight new models such as Claude Fable 5, Claude Sonnet 5, and Qwen 3.7 Max. The benchmark now includes metrics for cost and efficiency, and a new framework designed to be easier for developers to use. Despite the additions, Google's own Gemini models continue to trail behind the new competitors on the leaderboard.

vidgetc Tech, gaming & AI news — always at hand Google Play · Soon App Store · Soon
💬 Discuss

Comments

Next articleWeb Scraper Declares 'Google and Reddit Do Not Own the Internet' After Court Victory
Start typing to search