Peer-reviewed io.net research accepted at AIBC 2025 shows idle consumer GPUs (RTX 4090) can deliver near-H100 throughput at roughly half the cost.Peer-reviewed io.net research accepted at AIBC 2025 shows idle consumer GPUs (RTX 4090) can deliver near-H100 throughput at roughly half the cost.

io.net Benchmarks Reveal Cost-Performance “Sweet Spot” for RTX 4090 Clusters

For feedback or concerns regarding this content, please contact us at crypto.news@mexc.com
streams of data gpu nodes digital globe

A peer-reviewed paper accepted to the 6th International Artificial Intelligence and Blockchain Conference (AIBC 2025) argues that idle consumer GPUs, exemplified by Nvidia’s RTX 4090, can meaningfully reduce the cost of running large language model inference when used alongside traditional datacenter hardware.

Titled Idle Consumer GPUs as a Complement to Enterprise Hardware for LLM Inference, the study from io.net is the first to publish open benchmarks of heterogeneous GPU clusters on the project’s decentralized cloud. The analysis compares clusters of consumer cards against datacenter-grade H100 accelerators and finds a clear cost-performance tradeoff that could reshape how organizations design their inference fleets.

According to the paper, clusters built from RTX 4090 GPUs can deliver between 62 and 78 percent of the throughput of H100s while operating at roughly half the cost. For batch workloads or latency-tolerant applications, token costs fall by as much as 75 percent. The researchers underscore that these savings are most compelling when developers can tolerate higher tail latencies or use consumer hardware for overflow and background tasks such as development, batch processing, embeddings generation and large-scale evaluation sweeps.

Aline Almeida, Head of Research at IOG Foundation and Lead Author of the study, said, “Our findings demonstrate that hybrid routing across enterprise and consumer GPUs offers a pragmatic balance between performance, cost and sustainability. Rather than a binary choice, heterogeneous infrastructure allows organizations to optimize for their specific latency and budget requirements while reducing carbon impact.”

Hybrid GPU Fleets

The paper does not shy away from H100s’ strengths: Nvidia’s datacenter cards sustain sub-55 millisecond P99 time-to-first-token performance even at high load, a boundary that keeps them indispensable for real-time, latency-sensitive applications such as production chatbots and interactive agents. Consumer GPU clusters, by contrast, are better suited to traffic that can tolerate extended tail latencies; the authors point to a 200–500 ms P99 window as realistic for many research and dev/test workloads.

Energy and sustainability are also part of the calculus. While H100s remain roughly 3.1 times more energy-efficient per token, the study suggests that harnessing idle consumer GPUs can lower the embodied carbon footprint of compute by prolonging hardware lifetimes and leveraging grids that are rich in renewable generation. In short, a mixed fleet can be both cheaper and greener when deployed strategically.

scalling

Gaurav Sharma, CEO of io.net, said, “This peer-reviewed analysis validates the core thesis behind io.net: that the future of compute will be distributed, heterogeneous, and accessible. By harnessing both datacenter-grade and consumer hardware, we can democratize access to advanced AI infrastructure while making it more sustainable.”

Practical guidance from the paper is aimed squarely at MLOps teams and AI developers. The authors recommend using enterprise GPUs for real-time, low-latency routing while routing development, experimentation and bulk workloads to consumer clusters. They report an operational sweet spot in which four-card RTX 4090 configurations hit the best cost per million tokens, between $0.111 and $0.149, while delivering a substantial portion of H100 performance.

Beyond the benchmarks, the research reinforces io.net’s mission to expand compute by stitching together distributed GPUs into a programmable, on-demand pool. The company positions its stack, combining io.cloud’s programmable infrastructure with io.intelligence’s API toolkit, as a full solution for startups that need training, agent execution and large-scale inference without the capital intensity of buying solely datacenter hardware.

The full benchmarks and methodology are available on io.net’s GitHub repository for those who want to dig into the numbers and reproduce the experiments. The study adds an important, empirically grounded voice to the debate about how to scale LLM deployments affordably and sustainably in the years ahead.

Market Opportunity
IO Logo
IO Price(IO)
$0.1086
$0.1086$0.1086
-2.68%
USD
IO (IO) Live Price Chart
Disclaimer: The articles reposted on this site are sourced from public platforms and are provided for informational purposes only. They do not necessarily reflect the views of MEXC. All rights remain with the original authors. If you believe any content infringes on third-party rights, please contact crypto.news@mexc.com for removal. MEXC makes no guarantees regarding the accuracy, completeness, or timeliness of the content and is not responsible for any actions taken based on the information provided. The content does not constitute financial, legal, or other professional advice, nor should it be considered a recommendation or endorsement by MEXC.

You May Also Like

Is Doge Losing Steam As Traders Choose Pepeto For The Best Crypto Investment?

Is Doge Losing Steam As Traders Choose Pepeto For The Best Crypto Investment?

The post Is Doge Losing Steam As Traders Choose Pepeto For The Best Crypto Investment? appeared on BitcoinEthereumNews.com. Crypto News 17 September 2025 | 17:39 Is dogecoin really fading? As traders hunt the best crypto to buy now and weigh 2025 picks, Dogecoin (DOGE) still owns the meme coin spotlight, yet upside looks capped, today’s Dogecoin price prediction says as much. Attention is shifting to projects that blend culture with real on-chain tools. Buyers searching “best crypto to buy now” want shipped products, audits, and transparent tokenomics. That frames the true matchup: dogecoin vs. Pepeto. Enter Pepeto (PEPETO), an Ethereum-based memecoin with working rails: PepetoSwap, a zero-fee DEX, plus Pepeto Bridge for smooth cross-chain moves. By fusing story with tools people can use now, and speaking directly to crypto presale 2025 demand, Pepeto puts utility, clarity, and distribution in front. In a market where legacy meme coin leaders risk drifting on sentiment, Pepeto’s execution gives it a real seat in the “best crypto to buy now” debate. First, a quick look at why dogecoin may be losing altitude. Dogecoin Price Prediction: Is Doge Really Fading? Remember when dogecoin made crypto feel simple? In 2013, DOGE turned a meme into money and a loose forum into a movement. A decade on, the nonstop momentum has cooled; the backdrop is different, and the market is far more selective. With DOGE circling ~$0.268, the tape reads bearish-to-neutral for the next few weeks: hold the $0.26 shelf on daily closes and expect choppy range-trading toward $0.29–$0.30 where rallies keep stalling; lose $0.26 decisively and momentum often bleeds into $0.245 with risk of a deeper probe toward $0.22–$0.21; reclaim $0.30 on a clean daily close and the downside bias is likely neutralized, opening room for a squeeze into the low-$0.30s. Source: CoinMarketcap / TradingView Beyond the dogecoin price prediction, DOGE still centers on payments and lacks native smart contracts; ZK-proof verification is proposed,…
Share
BitcoinEthereumNews2025/09/18 00:14
House Democrat smacks down Trump's rambling ICE threat: 'This man can't win'

House Democrat smacks down Trump's rambling ICE threat: 'This man can't win'

A House Democrat smacked down President Donald Trump's rambling threat to deploy Immigration and Customs Enforcement agents to airports nationwide.Trump wrote on
Share
Rawstory2026/03/22 07:23
Tether CEO Delivers Rare Bitcoin Price Comment

Tether CEO Delivers Rare Bitcoin Price Comment

Bitcoin price receives rare acknowledgement from Tether CEO Ardoino
Share
Coinstats2025/09/17 23:39