{"id":331853,"date":"2025-12-07T02:25:18","date_gmt":"2025-12-07T02:25:18","guid":{"rendered":"https:\/\/www.newsbeep.com\/au\/331853\/"},"modified":"2025-12-07T02:25:18","modified_gmt":"2025-12-07T02:25:18","slug":"google-outlines-miras-and-titans-a-possible-path-toward-continuously-learning-ai","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/au\/331853\/","title":{"rendered":"Google outlines MIRAS and Titans, a possible path toward continuously learning AI"},"content":{"rendered":"<p>                                    <a class=\"article-menu__content__link\" href=\"#summary\"><br \/>\n                        <img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/the-decoder.com\/resources\/icons\/summary.svg\" alt=\"summary\" width=\"27\" height=\"24\" data-no-lazy=\"1\"\/><br \/>\n                        Summary<br \/>\n                    <\/a><\/p>\n<p>A year after publishing its Titans paper, Google has formally detailed the architecture on its research blog, pairing it with a new framework called MIRAS. Both projects target a major frontier in AI: models that keep learning during use and maintain a functional long-term memory instead of remaining static after pretraining.<\/p>\n<p>Google frames the motivation in familiar terms. Traditional Transformers struggle with very long inputs like books, genome sequences, or extended videos because their computational cost grows quadratically with context length. Faster alternatives such as modern RNNs or state-space models scale better but compress the entire context into a single internal state, losing important details. Titans is designed to bridge that gap by combining precise short-term memory through windowed attention with a separate, trainable long-term memory that can update during inference and selectively retain surprising or unexpected information.<\/p>\n<p>The company is also introducing MIRAS, a theoretical framework first described in April in the paper <a target=\"_blank\" rel=\"noopener nofollow\" href=\"https:\/\/arxiv.org\/pdf\/2504.13173\">&#8220;It&#8217;s All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization&#8221;<\/a>. The researchers argue that many of the new sequence models released in recent years &#8211; from Transformer variants to RetNet, Mamba, DeltaNet, and RWKV &#8211; can be viewed as different implementations of the same underlying idea: an internal lookup system that links inputs (keys) to outputs (values).<\/p>\n<p>MIRAS breaks this system into four design questions. What does the lookup structure look like &#8211; a vector, a matrix, or a small or deep network? What internal scoring rule determines what gets stored well? How quickly does new information overwrite old entries? And what update rule governs how those entries change over time? Using this perspective, Google derives new attention-free models like Moneta, Yaad, and Memora that deliberately explore these design spaces and, in tests with extremely long contexts, sometimes outperform Mamba2 and standard Transformers.<\/p>\n<p>Ad<\/p>\n<p>THE DECODER Newsletter<\/p>\n<p>The most important AI news straight to your inbox.<\/p>\n<p>\u2713 Weekly<\/p>\n<p>\u2713 Free<\/p>\n<p>\u2713 Cancel at any time<\/p>\n<p>Titans and MIRAS reflect the limits of today\u2019s dominant Transformer architectures and may represent part of the shift toward what Ilya Sutskever recently described as a new era of AI research. In an interview with Dwarkesh Patel, the former OpenAI chief scientist argued that simply scaling data and compute is hitting diminishing returns, and outlined his startup SSI\u2019s vision of a superintelligence that learns on the job more like a talented teenager than a fully formed AGI dropped out of a training cluster.<\/p>\n<p>Google\u2019s approach differs from Sutskever\u2019s but targets the same gap: moving beyond static, one-time pretrained models toward systems that expand their capabilities over time, whether through explicit memory modules like Titans or through new learning paradigms still waiting to be discovered.<\/p>\n<p>Original article from January 17, 2025<\/p>\n<p>Google researchers have developed a new type of Transformer model that gives language models something similar to long-term memory. The system can handle much longer sequences of information than current models, leading to better performance across various tasks.<\/p>\n<p>The new &#8220;Titans&#8221; architecture takes inspiration from how human memory works. By combining artificial short and long-term memory through attention blocks and memory MLPs, the system can work with long sequences of information.<\/p>\n<p>Recommendation<\/p>\n<p>                                            <a class=\"link-overlay\" href=\"https:\/\/the-decoder.com\/openai-outperforms-humans-and-google-at-the-worlds-top-collegiate-programming-contest\/\" aria-label=\"OpenAI outperforms humans and Google at the world&#039;s top collegiate programming contest\" rel=\"nofollow noopener\" target=\"_blank\"><\/p>\n<p>                                                        \t\t\t<a class=\"post-thumbnail\" href=\"https:\/\/the-decoder.com\/openai-outperforms-humans-and-google-at-the-worlds-top-collegiate-programming-contest\/\" aria-hidden=\"true\" tabindex=\"-1\" rel=\"nofollow noopener\" target=\"_blank\"><\/p>\n<p>\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t<img decoding=\"async\" data-lazyloaded=\"1\" src=\"https:\/\/www.newsbeep.com\/au\/wp-content\/uploads\/2025\/12\/OpenAI-Gold-title-375x250.webp.webp\" loading=\"lazy\" alt=\"OpenAI outperforms humans and Google at the world's top collegiate programming contest\" width=\"375\" height=\"250\"\/><br \/>\n\t\t\t\t\t\t\t<\/a><\/p>\n<p>                \t\t\t<a class=\"post-thumbnail\" href=\"https:\/\/the-decoder.com\/openai-outperforms-humans-and-google-at-the-worlds-top-collegiate-programming-contest\/\" aria-hidden=\"true\" tabindex=\"-1\" rel=\"nofollow noopener\" target=\"_blank\"><\/p>\n<p>\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t\t<img decoding=\"async\" data-lazyloaded=\"1\" src=\"https:\/\/www.newsbeep.com\/au\/wp-content\/uploads\/2025\/12\/OpenAI-Gold-title-375x250.webp.webp\" loading=\"lazy\" alt=\"OpenAI outperforms humans and Google at the world's top collegiate programming contest\" width=\"375\" height=\"250\"\/><br \/>\n\t\t\t\t\t\t\t<\/a><\/p>\n<p>One of the system&#8217;s clever features is how it decides what to remember. Titans uses &#8220;surprise&#8221; as its main metric &#8211; the more unexpected a piece of information is, the more likely it gets stored in long-term memory. The system also knows when to forget things, helping it use memory space efficiently.<\/p>\n<p>The team created three different versions of Titans, each handling long-term memory differently:<br \/>&#8211; Memory as Context (MAC)<br \/>&#8211; Memory as Gate (MAG)<br \/>&#8211; Memory as Layer (MAL)<\/p>\n<p>While each version has its strengths, the MAC variant works especially well with very long sequences.<\/p>\n<p><img data-lazyloaded=\"1\" fetchpriority=\"high\" decoding=\"async\" class=\"wp-image-20586 size-medium\" src=\"https:\/\/www.newsbeep.com\/au\/wp-content\/uploads\/2025\/12\/Titans-Google-MAC-770x357.png\" alt=\"\" width=\"770\" height=\"357\"\/>Image: Google<\/p>\n<p>Share<\/p>\n<p>Recommend our article<\/p>\n<p>        Share<\/p>\n<p>Better performance on long-context tasks<\/p>\n<p>In extensive testing, Titans outperformed traditional models like the classic Transformer and newer hybrid models like Mamba2, particularly when dealing with very long texts. The team says it can handle context windows of more than 2 million tokens more effectively, setting new records for both language modeling and time series prediction with long contexts.<\/p>\n<p>The system also excelled at the &#8220;Needle in the Haystack&#8221; test, where it needs to find specific information in very long texts. Titans achieved over 95% accuracy even with 16,000-token texts. While some models from OpenAI, Anthropic, and Google perform better, they&#8217;re much larger &#8211; Titans&#8217; biggest version has only 760 million parameters.<\/p>\n<p><img loading=\"lazy\" data-lazyloaded=\"1\" decoding=\"async\" class=\"wp-image-20587 size-medium\" src=\"https:\/\/www.newsbeep.com\/au\/wp-content\/uploads\/2025\/12\/Titans-Google-bench-770x305.png\" alt=\"\" width=\"770\" height=\"305\"\/>Titans models also beat significantly larger language models in tasks that require an understanding of larger contexts. | Image: Google<\/p>\n<p>Titans really showed its strength in the <a target=\"_blank\" rel=\"noopener nofollow\" href=\"https:\/\/huggingface.co\/spaces\/RMT-team\/babilong\">BABILong benchmark<\/a>, a challenging test of long-term comprehension where models need to connect facts spread across very long documents. The system outperformed larger models like <a class=\"mixed-keyword\" href=\"https:\/\/the-decoder.com\/open-ai-gpt-4-announcement\/\" rel=\"nofollow noopener\" target=\"_blank\">GPT-4<\/a>, RecurrentGemma-9B, and Llama3.1-70B. It even beat Llama3 with Retrieval Augmented Generation (RAG), though some specialized retrieval models still perform better.<\/p>\n<p>The team expects to make the code publicly available in the near future. While Titans and similar architectures could lead to language models that handle longer contexts and make better inferences, the benefits might extend beyond just text processing. The team&#8217;s early tests with DNA modeling suggest the technology could improve other applications too, including video models &#8211; assuming the promising benchmark results hold up in real-world use.<\/p>\n","protected":false},"excerpt":{"rendered":"Summary A year after publishing its Titans paper, Google has formally detailed the architecture on its research blog,&hellip;\n","protected":false},"author":2,"featured_media":331854,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[256,40304,254,255,64,63,5069,113,105],"class_list":["post-331853","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-ai-training","tag-artificial-intelligence","tag-artificialintelligence","tag-au","tag-australia","tag-generative-ai","tag-google","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/331853","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/comments?post=331853"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/331853\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media\/331854"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media?parent=331853"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/categories?post=331853"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/tags?post=331853"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}