Google updates Android Bench with new LLMs, but Gemini still lags behind
谷歌更新 Android Bench 并引入新的 LLM,但 Gemini 仍然落后
Author: Ryan Whitwam | Category: AI, Google, AI benchmarks, android | Original: arstechnica.com
Published: Wed, 08 Jul 2026 16:39:48 +0000
Code generation is emerging as one of the most popular applications for large language models (LLMs), but not all agents are equally good at all development tasks. Google created a benchmark earlier this year to evaluate how LLMs perform in Android app development, and Android Bench is getting a big update today. The leaderboard now includes a raft of new models, and Google has adopted a new framework that should be easier to use. Developers are invited to run their own tests and submit feedback that could shape the future of Android Bench.
⋯ 继续阅读请登录会员 ⋯