{"id":701955,"date":"2026-06-13T15:52:30","date_gmt":"2026-06-13T15:52:30","guid":{"rendered":"https:\/\/www.newsbeep.com\/us\/701955\/"},"modified":"2026-06-13T15:52:30","modified_gmt":"2026-06-13T15:52:30","slug":"gemini-3-5-flash-costs-3x-price-of-3-1-pro-in-android-coding-test","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/us\/701955\/","title":{"rendered":"Gemini 3.5 Flash costs 3x price of 3.1 Pro in Android coding test"},"content":{"rendered":"<p>\t<img width=\"1600\" height=\"800\" src=\"https:\/\/www.newsbeep.com\/us\/wp-content\/uploads\/2026\/06\/gemini-3-5-flash-2.jpg\" class=\"skip-lazy wp-post-image\" alt=\"\"  decoding=\"async\" fetchpriority=\"high\"\/><\/p>\n<p class=\"wp-block-paragraph\">Google has released another set of benchmark results to determine the best AI models for Android coding, along with how much each model costs per token. Google\u2019s Gemini 3.5 Flash is easily the most resource-intensive in Android development, and it doesn\u2019t even make the top five.<\/p>\n<p class=\"wp-block-paragraph\">As the hype for general chatbots is dying down, companies like Google, OpenAI, and Anthropic are shifting towards agentic models with a strength in coding. Users have begun relying on these models for \u201cvibe coding,\u201d which essentially offloads the bulk of software development to LLMs.<\/p>\n<p class=\"wp-block-paragraph\">Recent models have dramatically improved their Android coding, and Google has kept tabs on which models perform best over the <a href=\"https:\/\/9to5google.com\/2026\/03\/06\/google-says-these-ai-models-are-best-at-coding-android-apps\/\" rel=\"nofollow noopener\" target=\"_blank\">past few months<\/a>. The \u201cAndroid Bench\u201d goes through updates as Google releases its own models, like the recent Gemini 3.5 Flash, and compares them to the competition.<\/p>\n<p class=\"wp-block-paragraph\">The main takeaway is how Google breaks these models down. Each model gets a score out of 100, indicative of the percentage of Android coding cases it can successfully solve across 10 runs. Google lists expected performance and the date the last test was run, with some high performers sticking around since February.<\/p>\n<p>\tAdvertisement &#8211; scroll for more content<\/p>\n<p class=\"wp-block-paragraph\">In the latest edition of Android Bench, the results paint a more expensive picture. Gemini 3.5 Flash ranks 6th in the Android Bench list under models like GPT 5.5 and Gemini 3.1 Pro Preview, which was tested in February.<\/p>\n<p class=\"wp-block-paragraph\">Gemini 3.5 Flash was touted as a <a href=\"https:\/\/9to5google.com\/2026\/05\/19\/gemini-app-google-io-2026\/\" rel=\"nofollow noopener\" target=\"_blank\">cheaper and faster alternative<\/a> to Gemini 3.1 Pro, with an expected performance gap of 6.1%. The new benchmark results say otherwise in regards to Android development, as Gemini 3.5 Flash has a higher latency and 9% gap in performance success.<\/p>\n<p class=\"wp-block-paragraph\">The kicker \u2013 Google\u2019s latest model costs an average of 355.9 tokens at $147.1 for one benchmark run, compared to Gemini 3.1 Pro Preview\u2019s 73.3 tokens used at around a third of that cost.<\/p>\n<p class=\"wp-block-paragraph\">Of course, it\u2019s worth noting that Google lists the preview version of Gemini 3.1 Pro. That being said, the preview model scores higher than a model meant to be faster and more efficient.<\/p>\n<p class=\"wp-block-paragraph\">GPT 5.5 ranks similarly in cost per run, but Gemini 3.5 Flash used up 5.5x more tokens in Android Bench tests. Claude\u2019s previous model, Opus 4.7, ranked 4th at a slightly lower run cost and token usage, sitting right in the middle of the pack. Google has not released benchmark scores for Opus 4.8 or <a href=\"https:\/\/9to5google.com\/2026\/06\/09\/anthropic-claude-mythos-fable-5-model-release\/\" rel=\"nofollow noopener\" target=\"_blank\">Fable 5<\/a>, for that matter.<\/p>\n<p class=\"wp-block-paragraph\">Here are the top ten models ranked by Google in the latest <a href=\"https:\/\/developer.android.com\/bench\" rel=\"nofollow noopener\" target=\"_blank\">Android Bench<\/a> release:<\/p>\n<p>ModelScoreAvg LatencyAvg Total TokensAvg CostGPT 5.57415.764.7$134.2GPT 5.472.421.264.2$91.7Gemini 3.1 Pro Preview72.411.173.3$47.9Claude Opus 4.768.711.690.0$124.3Claude Opus 4.666.69.969.5$84.4Gemini 3.5 Flash63.714.2355.9$147.1GLM 5.159.733.480.2$46.7Kimi K2.658.629.994.3$42.5Claude Sonnet 4.658.48.247.9$40.4DeepSeek V4 Pro55.435.8132.7$13.7Claude Sonnet 4.553.713.194.2$61.0<\/p>\n<p class=\"wp-block-paragraph\">The list includes several open-weight models listed among the well-know closed-weight models like Claude and GPT. The high end of the list has effectively remained unchanged since the last Android Bench, with the exception for GPT 5.3 Codex which has been removed from the list.<\/p>\n<p class=\"wp-block-paragraph\">You can see the <a href=\"https:\/\/developer.android.com\/bench\" rel=\"nofollow noopener\" target=\"_blank\">full rankings on Google\u2019s website<\/a>.<\/p>\n<p class=\"wp-block-paragraph\">Google has regularly updated this list as more models are tested. At its core, it seems like a solid indicator of model performance in Android development. Gemini 3.5 Flash has been a solid improvement for other LLM and agentic tasks, even as Google has shifted cost and <a href=\"https:\/\/9to5google.com\/2026\/05\/28\/gemini-new-usage-limits\/\" rel=\"nofollow noopener\" target=\"_blank\">usage limits<\/a> around. Google\u2019s release numbers can\u2019t be disregarded completely, though Android coding is apparently not Gemini 3.5 Flash\u2019s strong suit.<\/p>\n<p>More on AI:<\/p>\n<p>\t\t<a target=\"_blank\" rel=\"nofollow noopener\" href=\"https:\/\/google.com\/preferences\/source?q=https:\/\/9to5google.com\" aria-label=\"Add 9to5Google as a preferred source on Google\"><br \/>\n\t\t\t<img decoding=\"async\" class=\"google-preferred-source-badge-dark\" src=\"https:\/\/www.newsbeep.com\/us\/wp-content\/uploads\/2026\/01\/1768288030_634_google-preferred-source-badge-dark.png\" alt=\"Add 9to5Google as a preferred source on Google\"\/><br \/>\n\t\t\t<img decoding=\"async\" class=\"google-preferred-source-badge-light\" src=\"https:\/\/www.newsbeep.com\/us\/wp-content\/uploads\/2026\/01\/1768288030_158_google-preferred-source-badge-light.png\" alt=\"Add 9to5Google as a preferred source on Google\"\/><br \/>\n\t\t<\/a><\/p>\n<p class=\"disclaimer-affiliate\">FTC: We use income earning auto affiliate links. <a href=\"https:\/\/9to5mac.com\/about\/#affiliate\" rel=\"nofollow noopener\" target=\"_blank\">More.<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Google has released another set of benchmark results to determine the best AI models for Android coding, along&hellip;\n","protected":false},"author":2,"featured_media":701956,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[45],"tags":[182,181,507,74],"class_list":["post-701955","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/701955","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/comments?post=701955"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/701955\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media\/701956"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media?parent=701955"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/categories?post=701955"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/tags?post=701955"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}