{"id":479979,"date":"2026-06-06T22:56:14","date_gmt":"2026-06-06T22:56:14","guid":{"rendered":"https:\/\/www.newsbeep.com\/il\/479979\/"},"modified":"2026-06-06T22:56:14","modified_gmt":"2026-06-06T22:56:14","slug":"researchers-put-ai-models-in-charge-of-analyzing-sports-and-they-choked-spectacularly","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/il\/479979\/","title":{"rendered":"Researchers Put AI Models in Charge of Analyzing Sports, and They Choked Spectacularly"},"content":{"rendered":"<p class=\"article-paragraph skip\">Sign up to see the future, today<\/p>\n<p class=\"article-paragraph skip\">Can\u2019t-miss innovations from the bleeding edge of science and tech<\/p>\n<p class=\"pw-incontent-excluded article-paragraph skip\">Good news for sports broadcasters and fans who\u2019d prefer their play-by-plays to have that human-touch: AI doesn\u2019t know ball.<\/p>\n<p class=\"article-paragraph skip\">A new study by researchers at the University of North Carolina at Chapel Hill and Northeastern University found that top AI models are horrible at analyzing professional sports. The <a href=\"https:\/\/arxiv.org\/abs\/2605.31529\" rel=\"noreferrer nofollow noopener\" target=\"_blank\">yet-to-be-peer-reviewed study<\/a> sought to analyze how capable the most popular AI models are in the fields of perception, reasoning, simulation, and agency \u2014 four traits which are difficult to evaluate with existing testing methods.<\/p>\n<p class=\"article-paragraph skip\">To probe how AI performs in these areas, researchers turned to the wide world of sports to create a new kind of AI test. Called strategic video intelligence, or \u201cSVI-bench,\u201d the novel test comprised 35,000 hours of sports footage from basketball, soccer, and hockey, as well as 15 million annotated plays, 15,000 hours of professional analysis, 23,000 post-game reports, and 103,000 statistical records.<\/p>\n<p class=\"article-paragraph skip\">Where AI performed the best was in perception: identifying which player performs which action at a given point in the match. But even there, they struggled badly. The models, which included ChatGPT, Google\u2019s Gemini, and the open-source model Qwen, successfully eyeballed which player was doing what roughly 74 percent of the time \u2014 a rate which would get even a volunteer Little League announcer sacked.<\/p>\n<p class=\"article-paragraph skip\">The AI models did far worse on causal reasoning, or explaining why certain plays went down the way they did, with success rates falling near 40 percent on average. For example, when researchers asked the models to identify what was unusual about a <a href=\"https:\/\/svi-bench.github.io\/#about\" rel=\"noreferrer nofollow noopener\" target=\"_blank\">Cody Martin three-pointer<\/a> \u2014 which bounced off the top of the backboard before landing in the bucket \u2014 ChatGPT replied that it was \u201chis first made three of the game.\u201d<\/p>\n<p class=\"article-paragraph skip\">Simulation, or asking AI to find evidence to predict things like where a player would physically go based on their trajectory, was also dismal. During these tests, the best-performing model was functionally flipping a coin in order to guess a player\u2019s next steps, and performance dropped even further when models were asked to plot out longer motion toward a goal or basket.<\/p>\n<p class=\"article-paragraph skip\">As computer science researcher at Northeastern and study co-author Lorenzo Torresani said in a <a href=\"https:\/\/news.northeastern.edu\/2026\/06\/01\/ai-sports-predictions\/\" rel=\"noreferrer nofollow noopener\" target=\"_blank\">press blurb by the university<\/a>, AI \u201ccannot tell you why things happen, and it cannot tell you what\u2019s gonna happen next.\u201d<\/p>\n<p class=\"article-paragraph skip\">When researchers probed the models\u2019 agency \u2014 basically asking them to make complex post-game analysis of stats and trends, like a human broadcaster would \u2014 its accuracy fell to just 5 percent.<\/p>\n<p class=\"article-paragraph skip\">\u201cA good sportscaster does much more than describe what\u2019s on screen \u2014 they explain why a play worked, anticipate what\u2019s next, and\u2026 decide which moments matter,\u201d Torresani said. \u201cOur study shows AI is already reasonably good at the descriptive part, but collapses on the rest.\u201d<\/p>\n<p class=\"article-paragraph skip\">While sportscasters can definitely breathe a sigh of relief, the study\u2019s findings are also good news for other knowledge workers, at a time when there\u2019s been relentless fear of AI automation turning the job market <a href=\"https:\/\/futurism.com\/artificial-intelligence\/ai-automation-replacement\" rel=\"nofollow noopener\" target=\"_blank\">inside out<\/a>.<\/p>\n<p class=\"article-paragraph skip\">\u201cThe same gap shows up in any job whose value lies not in describing what\u2019s visible, but in understanding why events unfold, anticipating what comes next, deciding what matters, and recommending what to do about it,\u201d Torresani concluded.<\/p>\n<p class=\"article-paragraph skip\">More on AI in sports: <a href=\"https:\/\/futurism.com\/artificial-intelligence\/nfl-new-york-jets-ai\" rel=\"nofollow noopener\" target=\"_blank\">Fans Aghast as New York Jets Say They\u2019re Switching to AI<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Sign up to see the future, today Can\u2019t-miss innovations from the bleeding edge of science and tech Good&hellip;\n","protected":false},"author":2,"featured_media":479980,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[345,343,344,85,46,125],"class_list":["post-479979","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-il","tag-israel","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/479979","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/comments?post=479979"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/479979\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media\/479980"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media?parent=479979"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/categories?post=479979"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/tags?post=479979"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}