Summary
Google has updated its Android Bench with eight new large language models (LLMs) and a new framework to evaluate their performance in Android app development. While code generation is a popular application for LLMs, Google's own Gemini model still lags behind in this benchmark.