{"id":712071,"date":"2026-07-26T14:07:12","date_gmt":"2026-07-26T14:07:12","guid":{"rendered":"https:\/\/www.newsbeep.com\/uk\/712071\/"},"modified":"2026-07-26T14:07:12","modified_gmt":"2026-07-26T14:07:12","slug":"ai-enthusiast-adds-nvidia-tesla-v100-as-loud-as-a-lawnmower-to-gaming-pc-for-266-32gb-of-vram-rig-can-run-27-billion-parameter-model-at-32-tokens-per-second","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/uk\/712071\/","title":{"rendered":"AI enthusiast adds Nvidia Tesla V100 as loud as a lawnmower to gaming PC for $266 \u2014 32GB of VRAM rig can run 27 billion parameter model at 32 tokens per second"},"content":{"rendered":"<p id=\"elk-894fe6de-86ab-11f1-847e-039b8a00b047\">A computing enthusiast has <a data-analytics-id=\"inline-link\" href=\"https:\/\/blog.tymscar.com\/posts\/v100localllm\/\" target=\"_blank\" data-url=\"https:\/\/blog.tymscar.com\/posts\/v100localllm\/\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\">repurposed<\/a> a very noisy and largely obsolete enterprise GPU (with lots of VRAM) for local LLM inference purposes. They are now enjoying a system that has doubled its total VRAM quota to 32GB for just a $266 (\u00a3200) outlay. That\u2019s a good result, especially in the midst of a <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/the-secret-to-building-a-pc-during-the-rampocalypse-are-bundles-here-are-some-of-the-best-ones-and-why-theyre-so-popular\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/the-secret-to-building-a-pc-during-the-rampocalypse-are-bundles-here-are-some-of-the-best-ones-and-why-theyre-so-popular\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/the-secret-to-building-a-pc-during-the-rampocalypse-are-bundles-here-are-some-of-the-best-ones-and-why-theyre-so-popular\" rel=\"nofollow noopener\">RAMpocalypse<\/a>.<\/p>\n<p>Oscar Molnar explains that a cheap <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/news\/nvidia-tesla-v100s-graphics-card-data-center\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/news\/nvidia-tesla-v100s-graphics-card-data-center\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/news\/nvidia-tesla-v100s-graphics-card-data-center\" rel=\"nofollow noopener\">Tesla V100<\/a> SXM2 with 16GB HBM2 was sourced, as was an SXM2-to-PCIe adapter, and a PWM mod for the loud-as-a-lawnmower cooler, to complete this VRAM expansion for the hefty local LLMs project. Indeed, these GPUs do look cheap right now, as I can see them <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.ebay.com\/sch\/i.html?_nkw=Tesla+V100&amp;mkevt=1&amp;mkcid=1&amp;mkrid=711-53200-19255-0&amp;campid=5337827784&amp;customid=tomshardware-us-4673853830647331051\" target=\"_blank\" data-url=\"https:\/\/www.ebay.com\/sch\/i.html?_nkw=Tesla+V100\" referrerpolicy=\"no-referrer-when-downgrade\" rel=\"sponsored noopener nofollow\" data-hl-processed=\"hawklinks\" data-google-interstitial=\"false\" data-placeholder-url=\"https:\/\/www.ebay.com\/sch\/i.html?_nkw=Tesla+V100&amp;mkevt=1&amp;mkcid=1&amp;mkrid=711-53200-19255-0&amp;campid=5337827784&amp;customid=hawk-custom-tracking\" data-merchant-name=\"eBay\" data-merchant-id=\"1219\" data-merchant-network=\"Ebay\" data-merchant-url=\"ebay.com\" data-mrf-recirculation=\"inline-link\">listed on eBay US for under $140<\/a> each, if you don\u2019t mind buying from China.<\/p>\n<p><a id=\"elk-seasonal\"\/><\/p>\n<p id=\"elk-894fe6de-86ab-11f1-847e-039b8a00b047-2\">As mentioned above, you can\u2019t just get one of these Tesla V100 SXM2 cards with abundant VRAM and plug it into your PC. Molnar says they spent about $66 on an <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/you-can-install-nvidias-fastest-ai-gpu-into-a-pcie-slot-with-an-sxm-to-pcie-adapter-nvidia-h100-sxm-can-fit-into-regular-x16-pcie-slots\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/you-can-install-nvidias-fastest-ai-gpu-into-a-pcie-slot-with-an-sxm-to-pcie-adapter-nvidia-h100-sxm-can-fit-into-regular-x16-pcie-slots\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/you-can-install-nvidias-fastest-ai-gpu-into-a-pcie-slot-with-an-sxm-to-pcie-adapter-nvidia-h100-sxm-can-fit-into-regular-x16-pcie-slots\" rel=\"nofollow noopener\">SXM2-to-PCIe adapter<\/a>, also on eBay.<\/p>\n<p>Latest Videos FromTom&#8217;s Hardware<\/p>\n<p>You might think that was enough. However, the PC and local LLMs enthusiast baulked at the noise of \u201cthe fan from hell,\u201d which came as standard with the Tesla V100 SXM2. That shrieking cooler was measured outputting 82dB of noise. Molnar described it as \u201csomewhere between a garbage disposal and a lawnmower.\u201d This may be the most complicated tweak yet, but basically the existing fan wires just needed rerouting and plugging into the motherboard PWM fan header. You could also simply purchase a \u201c2.54mm male to PH2.0 female jumper cable\u201d for the task. Apparently, the fan only needs to run at 10% to keep the Tesla V100 under 50C at full load.<\/p>\n<p class=\"vanilla-image-block\" style=\"padding-top:93.82%;\">\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/07\/5UhZUv7mNhAoDYNYbE8BJf.jpg\" alt=\"Nvidia Tesla V100\"   loading=\"lazy\" data-new-v2-image=\"true\" data-original-mos=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/07\/5UhZUv7mNhAoDYNYbE8BJf.jpg\" data-pin-media=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/07\/5UhZUv7mNhAoDYNYbE8BJf.jpg\" class=\"rounded-[var(--image--border-radius,0)] inline expandable\"\/><br \/>\n<a href=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/07\/5UhZUv7mNhAoDYNYbE8BJf.jpg\" target=\"_blank\" class=\"expand-button icon-expand-image icon\" data-url=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/07\/5UhZUv7mNhAoDYNYbE8BJf.jpg\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\"><\/p>\n<p>(Image credit: Nvidia)<a id=\"elk-da982952-86ab-11f1-b06f-cfeff2008f57\"\/>27 billion parameter LLM runs at 32 tokens per second<\/p>\n<p id=\"elk-da9829de-86ab-11f1-a60c-55ce38400aec\">With the hardware all now fitted and finessed, Molnar had a 32GB VRAM system at their disposal \u2013 that\u2019s a PC with <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/reviews\/nvidia-geforce-rtx-4080-review\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/reviews\/nvidia-geforce-rtx-4080-review\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/reviews\/nvidia-geforce-rtx-4080-review\" rel=\"nofollow noopener\">RTX 4080<\/a>: 16GB VRAM, <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/features\/nvidia-ada-lovelace-and-geforce-rtx-40-series-everything-we-know\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/features\/nvidia-ada-lovelace-and-geforce-rtx-40-series-everything-we-know\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/features\/nvidia-ada-lovelace-and-geforce-rtx-40-series-everything-we-know\" rel=\"nofollow noopener\">Ada architecture<\/a> and Tesla V100: 16GB VRAM, <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/news\/nvidia-volta-gv100-gpu-ai,35297.html\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/news\/nvidia-volta-gv100-gpu-ai,35297.html\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/news\/nvidia-volta-gv100-gpu-ai,35297.html\" rel=\"nofollow noopener\">Volta architecture<\/a>. They note you can get Tesla V100s with 32GB of VRAM, but they are double the price.<\/p>\n<p>            You may like<\/p>\n<p>Getting the system to make use of this 32GB of total VRAM for LLMs wasn\u2019t tricky, says the DIYer. They used NixOS with a legacy Nvidia driver that overlapped support for both Volta and Ada architectures. Testing a <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/ditching-the-cloud-for-local-ai-how-i-use-two-mini-pcs-to-process-millions-of-tokens-a-day-and-save-money-on-costly-api-fees\" target=\"_blank\" data-url=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/ditching-the-cloud-for-local-ai-how-i-use-two-mini-pcs-to-process-millions-of-tokens-a-day-and-save-money-on-costly-api-fees\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/ditching-the-cloud-for-local-ai-how-i-use-two-mini-pcs-to-process-millions-of-tokens-a-day-and-save-money-on-costly-api-fees\" rel=\"nofollow noopener\">local LLM<\/a>, they got a 27 billion parameter model running at 32 tokens per second, which they say is \u201cfast enough for interactive use\u201d and faster than most cloud API alternatives.<\/p>\n<p><a href=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" id=\"elk-da982a60-86ab-11f1-819a-5b93b501d6c5\" data-url=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" target=\"_blank\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" rel=\"nofollow noopener\"><\/p>\n<p class=\"vanilla-image-block\" style=\"padding-top:31.51%;\">\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/01\/7cUTDmN2PHNRiNBVqbKf56.png\" alt=\"Google Preferred Source\"   loading=\"lazy\" data-new-v2-image=\"true\" data-original-mos=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/01\/7cUTDmN2PHNRiNBVqbKf56.png\" data-pin-media=\"https:\/\/www.newsbeep.com\/uk\/wp-content\/uploads\/2026\/01\/7cUTDmN2PHNRiNBVqbKf56.png\" class=\"rounded-[var(--image--border-radius,0)] pull-left\"\/>\n<\/p>\n<p><\/a><\/p>\n<p id=\"elk-da982ad8-86ab-11f1-a196-39b72dbead7d\">Follow<a data-analytics-id=\"inline-link\" href=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" target=\"_blank\" data-url=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\"> Tom&#8217;s Hardware on Google News<\/a>, or<a data-analytics-id=\"inline-link\" href=\"https:\/\/google.com\/preferences\/source?q=\" target=\"_blank\" data-url=\"https:\/\/google.com\/preferences\/source?q=\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\"> add us as a preferred source<\/a>, to get our latest news, analysis, &amp; reviews in your feeds.<\/p>\n<p class=\"newsletter-form__strapline\">Get Tom&#8217;s Hardware&#8217;s best news and in-depth reviews, straight to your inbox.<\/p>\n","protected":false},"excerpt":{"rendered":"A computing enthusiast has repurposed a very noisy and largely obsolete enterprise GPU (with lots of VRAM) for&hellip;\n","protected":false},"author":2,"featured_media":712072,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[554,733,4308,86,56,54,55],"class_list":["post-712071","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-technology","tag-uk","tag-united-kingdom","tag-unitedkingdom"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/712071","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/comments?post=712071"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/712071\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media\/712072"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media?parent=712071"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/categories?post=712071"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/tags?post=712071"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}