{"id":692012,"date":"2026-06-08T18:08:10","date_gmt":"2026-06-08T18:08:10","guid":{"rendered":"https:\/\/www.newsbeep.com\/us\/692012\/"},"modified":"2026-06-08T18:08:10","modified_gmt":"2026-06-08T18:08:10","slug":"big-tech-is-quietly-admitting-that-if-it-wants-to-sell-people-on-ai-it-better-be-cheap","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/us\/692012\/","title":{"rendered":"Big Tech Is Quietly Admitting That If It Wants to Sell People on AI, It Better Be Cheap"},"content":{"rendered":"<p>The AI mania that\u2019s spread across Silicon Valley like a fever over the past few years is running up against some hard economic realities.<\/p>\n<p>In recent weeks, big tech companies have been forced to admit that spending on tokens\u2014the basic unit of measurement for AI usage\u2014has gotten out of control. Amazon had to shut down an in-house competition to use as many tokens as possible at work, telling employees, \u201cPlease don\u2019t use AI just for the sake of using AI,\u201d <a href=\"https:\/\/www.businessinsider.com\/amazon-ai-leaderboard-tokenmaxxing-2026-5\" rel=\"nofollow noopener\" target=\"_blank\">according to Business Insider<\/a>; Uber has <a href=\"https:\/\/www.bloomberg.com\/news\/articles\/2026-06-02\/uber-caps-usage-of-ai-tools-like-claude-code-to-cut-costs\" rel=\"nofollow noopener\" target=\"_blank\">reportedly<\/a> capped employee spending on tokens to $1,500 per month after the company exhausted its annual AI budget earlier this year. And most tellingly, the companies building the big AI models have also woken up to this sobering reality. At a recent <a href=\"https:\/\/openai.com\/business\/intelligence-at-work\/\" rel=\"nofollow noopener\" target=\"_blank\">event<\/a> hosted by OpenAI, company chief executive Sam Altman admitted that token usage had become \u201ca huge issue\u201d for companies that were promised big productivity gains if they incorporated AI across their organization.<\/p>\n<p>That\u2019s a hard pivot from just a few months ago, where the general vibe across the industry was the more that employees use AI, the better off they\u2014and the companies they work for\u2014will be. So-called \u201ctokenmaxxing\u201d became a meme, and more or less synonymous with \u201cfuture-proofing\u201d: in a day and age when everyone and their neighbors are using AI, those who know how to use AI will have a sharp edge. Not every job will necessarily be replaced by AI (so the thinking goes), but employees who don\u2019t use AI will definitely be replaced by those who do.<\/p>\n<p>But AI has always been expensive, and training and inference costs for new models are only <a href=\"https:\/\/arxiv.org\/abs\/2405.21015\" rel=\"nofollow noopener\" target=\"_blank\">getting higher<\/a>. Meanwhile, the industry\u2019s dedicated push into agents\u2014AI systems that can work with little to no human oversight for extended periods of time\u2014has led to a token usage explosion. One <a href=\"https:\/\/arxiv.org\/abs\/2604.22750\" rel=\"nofollow noopener\" target=\"_blank\">preprint study<\/a> posted in April found that agents use 1,000 times as many tokens as other AI systems.\u00a0<\/p>\n<p>It\u2019s the companies and individual users who have overwhelmingly had to eat those costs. No wonder some developers have resorted to <a href=\"https:\/\/gizmodo.com\/should-you-hijack-a-corporate-ai-chatbot-for-free-tokens-2000767595\" rel=\"nofollow noopener\" target=\"_blank\">pirating free online chatbots<\/a> like Chipotle\u2019s customer service bot, Pepper, to bypass the big companies\u2019 token-hungry models. GitHub <a href=\"https:\/\/github.blog\/news-insights\/company-news\/github-copilot-is-moving-to-usage-based-billing\/\" rel=\"nofollow noopener\" target=\"_blank\">announced<\/a> earlier this week that it was rolling out a new payment model in which users would be charged by the number of tokens they burn. Judging from some of the <a href=\"https:\/\/www.theregister.com\/ai-and-ml\/2026\/06\/02\/github-copilot-users-threaten-exit-as-metered-billing-kicks-in\/5249826\" rel=\"nofollow noopener\" target=\"_blank\">early user feedback<\/a>, it hasn\u2019t been going well. <\/p>\n<p>Big tech desperately needs to find a new way to sell people on the future of AI without the exorbitant token costs. If they don\u2019t, companies and users will just switch to some open model they can use for free.<\/p>\n<p> Close to the edge <\/p>\n<p>Some big tech companies have literally been forced to the edge by the rising costs of AI usage.\u00a0<\/p>\n<p><a href=\"https:\/\/nvidianews.nvidia.com\/news\/nvidia-microsoft-windows-pcs-agents-rtx-spark\" rel=\"nofollow noopener\" target=\"_blank\">Microsoft<\/a> and <a href=\"https:\/\/developers.googleblog.com\/bringing-gemma-4-12b-to-your-laptop-unlocking-local-agentic-workflows-with-google-ai-edge\/\" rel=\"nofollow noopener\" target=\"_blank\">Google<\/a> recently announced new AI products\u2014Gemma 4 12B and the RTX Spark laptop, respectively\u2014that are based on so-called \u201cedge\u201d computing. That\u2019s when a model is powered by computing power from a specific device, rather than by the cloud (i.e., energy-guzzling data centers). Obviously, a model of the magnitude of a Claude Opus 4.8 or a GPT-5 isn\u2019t going to be able to be run directly from your laptop; that\u2019s like trying to provide enough energy for a Falcon 9 rocket launch by plugging a stationary bike into a generator. But the logic behind Microsoft and Google\u2019s new products is that actually, not everyone needs the latest, greatest, most token-hungry models directly in the devices they\u2019re using daily. For most people most of the time, a smaller, leaner model will work just fine. And crucially, it will save everyone some money on tokens.<\/p>\n<p>Make no mistake, Microsoft\u2019s and Google\u2019s investments in edge computing are minuscule compared to what they\u2019re spending on data centers; cloud computing is still very much the backbone of both of their business models. But in their embrace of edge computing, we\u2019re seeing an, at least tacit, acknowledgement that the cost of massive AI models just isn\u2019t worth the squeeze that it\u2019s placing on most consumers.<\/p>\n<p> Water promises <\/p>\n<p>While they\u2019re pushing new edge computing products\u2014promising powerful AI capabilities at a lower cost\u2014Microsoft and Google are also trying to pacify a public that\u2019s become increasingly concerned over data centers\u2019 water demands. (Data centers usually use water to keep GPU clusters from overheating.) On Tuesday, during the opening keynote of Microsoft Build, the company\u2019s annual developer conference, CEO Satya Nadella claimed Microsoft\u2019s new data centers\u2019 annual water usage \u201cis roughly equivalent to what a single restaurant would use.\u201d\u00a0<\/p>\n<p>The following day, Google <a href=\"https:\/\/blog.google\/company-news\/outreach-and-initiatives\/sustainability\/new-water-stewardship-commitments\/\" rel=\"nofollow noopener\" target=\"_blank\">announced<\/a> plans to \u201creplenish more water than we consume\u201d from data center cooling by 2030, along with other \u201cwater stewardship commitments.\u201d For a little extra sprinkle of intended comfort, the press release noted that \u201cU.S. data centers use less than 1% of the water that Americans use on their lawns annually\u201d\u2014though that\u2019s probably more of a damning picture of Americans\u2019 lawn-watering habits than it is an absolution of the water-guzzling sins of the AI industry.<\/p>\n","protected":false},"excerpt":{"rendered":"The AI mania that\u2019s spread across Silicon Valley like a fever over the past few years is running&hellip;\n","protected":false},"author":2,"featured_media":692013,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[27],"tags":[40303,182,28,7258,168,1281],"class_list":["post-692012","post","type-post","status-publish","format-standard","has-post-thumbnail","category-business","tag-agents","tag-ai","tag-business","tag-data-centers","tag-google","tag-microsoft"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/692012","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/comments?post=692012"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/692012\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media\/692013"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media?parent=692012"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/categories?post=692012"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/tags?post=692012"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}