{"id":310414,"date":"2026-02-21T20:54:11","date_gmt":"2026-02-21T20:54:11","guid":{"rendered":"https:\/\/www.newsbeep.com\/ie\/310414\/"},"modified":"2026-02-21T20:54:11","modified_gmt":"2026-02-21T20:54:11","slug":"googles-gemini-3-1-pro-preview-tops-artificial-analysis-intelligence-index-at-less-than-half-the-cost-of-its-rivals","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ie\/310414\/","title":{"rendered":"Google&#8217;s Gemini 3.1 Pro Preview tops Artificial Analysis Intelligence Index at less than half the cost of its rivals"},"content":{"rendered":"<p>Google&#8217;s Gemini 3.1 Pro Preview leads the Artificial Analysis Intelligence Index four points ahead of Anthropic&#8217;s Claude Opus 4.6, at less than half the cost. The model ranks first in six of ten categories, including agent-based coding, knowledge, scientific reasoning, and physics. Its hallucination rate dropped 38 percentage points compared to <a href=\"https:\/\/the-decoder.com\/gemini-3-pro-tops-new-ai-reliability-benchmark-but-hallucination-rates-remain-high\/\" rel=\"nofollow noopener\" target=\"_blank\">Gemini 3 Pro, which struggled in that area<\/a>. The index rolls ten benchmarks into one overall score.<\/p>\n<p><a href=\"https:\/\/www.newsbeep.com\/ie\/wp-content\/uploads\/2026\/02\/benchmark_aa_31_pro-scaled-1.jpg\"><img data-lazyloaded=\"1\" fetchpriority=\"high\" decoding=\"async\" class=\"wp-image-32360 size-full\" src=\"https:\/\/www.newsbeep.com\/ie\/wp-content\/uploads\/2026\/02\/benchmark_aa_31_pro-scaled-1.jpg\" alt=\"Bar chart of the Artificial Analysis Intelligence Index: Gemini 3.1 Pro Preview leads with 57 points, followed by Claude Opus 4.6 at 53, Claude Sonnet 4.6 at 51, GPT-5.2 at 51, and GLM-5 at 50. Other models like Kimi K2.5, Gemini 3 Flash, and Grok 4 follow with lower scores.\" width=\"2560\" height=\"705\"\/><\/a>Gemini 3.1 Pro Preview scored 57 points in the Artificial Analysis Intelligence Index, four points ahead of Claude Opus 4.6, six ahead of GPT-5.2. | Image: Artificial Analysis<\/p>\n<p>Running the full index test with Gemini costs $892, compared to $2,304 for GPT-5.2 and $2,486 for Claude Opus 4.6. Gemini used just 57 million tokens, well under GPT-5.2&#8217;s 130 million. <a href=\"https:\/\/the-decoder.com\/chinese-ai-lab-zhipu-releases-glm-5-under-mit-license-claims-parity-with-top-western-models\/\" rel=\"nofollow noopener\" target=\"_blank\">Open-source models like GLM-5<\/a> come in even cheaper at $547. When it comes to real-world agent tasks, though, Gemini 3.1 Pro still falls behind Claude Sonnet 4.6, Opus 4.6, and GPT-5.2.<\/p>\n<p>As always, benchmarks only go so far. In our own internal fact-checking test, 3.1 Pro does significantly worse than Opus 4.6 or GPT-5.2, verifying only about a quarter of statements in initial tests, even fewer than Gemini 3 Pro, which was already weak here. So find your own benchmarks.<\/p>\n<p>\t\t\t\tAI News Without the Hype \u2013 Curated by Humans<\/p>\n<p>\n\t\t\t\t\tAs a THE DECODER subscriber, you get ad-free reading, our weekly AI newsletter, the exclusive &#8220;AI Radar&#8221; Frontier Report 6\u00d7 per year, access to comments, and our complete archive.\t\t\t\t<\/p>\n<p>\t\t\t\t<a href=\"https:\/\/the-decoder.com\/subscription\/\" class=\"inline-block text-white bg-(--heise-primary) mt-3 hover:bg-blue-800 focus:ring-4 focus:outline-none focus:ring-blue-300 font-medium rounded-sm w-full sm:w-auto  pl-3 pr-3 py-2.5 text-center newsletter-submit-button hover:no-underline\" rel=\"nofollow noopener\" target=\"_blank\"><br \/>\n\t\t\t\t\tSubscribe now\t\t\t\t<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Google&#8217;s Gemini 3.1 Pro Preview leads the Artificial Analysis Intelligence Index four points ahead of Anthropic&#8217;s Claude Opus&hellip;\n","protected":false},"author":2,"featured_media":60862,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[6],"tags":[2210,1711,61,60,80],"class_list":["post-310414","post","type-post","status-publish","format-standard","has-post-thumbnail","category-technology","tag-gemini","tag-google","tag-ie","tag-ireland","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/310414","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/comments?post=310414"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/310414\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media\/60862"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media?parent=310414"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/categories?post=310414"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/tags?post=310414"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}