2026年7月30日 · 星期四
● 每日更新·改变自己
Eurekar·TOP
捕捉真实世界的英语信号
25信息来源
8,218精选文章
167单词卡片
18照片图片
全部6,479口语2,048免费395新闻1,599帖子1,379hackernews658tmz458techmeme457随笔323slashdot310外刊291arstechnica230techcrunch218Cards167simonwillison73动态69bloomberg54sethgodin43图片18youtube7
← 返回

谷歌更新 Android Bench 并引入新的 LLM,但 Gemini 仍然落后

2026-07-09 arstechnica

← 上一篇返回列表下一篇 →

Google updates Android Bench with new LLMs, but Gemini still lags behind

谷歌更新 Android Bench 并引入新的 LLM,但 Gemini 仍然落后

Author: Ryan Whitwam | Category: AI, Google, AI benchmarks, android | Original: arstechnica.com

Published: Wed, 08 Jul 2026 16:39:48 +0000

Code generation is emerging as one of the most popular applications for large language models (LLMs), but not all agents are equally good at all development tasks. Google created a benchmark earlier this year to evaluate how LLMs perform in Android app development, and Android Bench is getting a big update today. The leaderboard now includes a raft of new models, and Google has adopted a new framework that should be easier to use. Developers are invited to run their own tests and submit feedback that could shape the future of Android Bench.

⋯ 继续阅读请登录会员 ⋯

🔒

MEMBERS ONLY

这篇是会员专享内容,你看到的是预览段。

会员每天解锁 6000+ 篇真实英语素材——双语科技、口语、外刊、单卡,不设上限。

年会员 ¥365 —— 一天一块钱,续费一直能用。

了解会员 →

已是会员?点此登录解锁全文。

← 上一篇返回列表下一篇 →