Shenzhen Manufacturer Lists Modded RTX 5090 With 96GB VRAM on Alibaba for $3,888
A Shenzhen manufacturer has done something Nvidia never authorized: it tore apart a GeForce RTX 5090, tripled its memory, and put the result up for sale on Alibaba for less than the price of a decent laptop. On September 11, 2026, Tom’s Hardware reported that Shenzhen Suqiao Intelligent Technology is selling a modified RTX 5090 carrying 96GB of VRAM, up from the card’s stock 32GB, for $3,888. Other Alibaba sellers are listing similar rebuilds closer to $5,900. Neither price is official Nvidia pricing, and neither card is an Nvidia product anymore in any sense the company would recognize.
What Actually Showed Up on Alibaba This Week
The listing itself is straightforward: a GeForce RTX 5090, Nvidia’s current flagship gaming GPU built on the Blackwell architecture, with its factory 32GB of GDDR7 memory replaced by 96GB. Tom’s Hardware framed the math directly in its headline, describing the card as offering “3x more VRAM at 65% the cost of the original” relative to comparable market pricing for high-memory alternatives. Whatever the exact baseline for that percentage, the headline figures are simple enough: 96GB of VRAM for $3,888 versus Nvidia’s official $1,999 Founders Edition MSRP for the unmodified 32GB card, or the $2,500 to $5,000 range the RTX 5090 actually trades at on the open retail market in 2026.
Suqiao is not a hobbyist operation. Reporting describes it as an OEM/ODM manufacturer, meaning it already runs board-level production and assembly work for other brands and is applying that same capability to memory rebuilds. Tom’s Hardware also flagged something worth dwelling on: rumors of leaked Nvidia firmware had reportedly circulated for months before these 96GB cards appeared, suggesting the software side of the mod, getting the GPU to correctly address and report the expanded memory pool, relied on a firmware mechanism that was never meant to leave Nvidia’s internal toolchain.
This isn’t even the most extreme example on the market. A separate Tom’s Hardware report covers an even rarer modded RTX 5090 carrying 128GB of VRAM, priced at roughly $13,000 and described explicitly as a limited-run prototype rather than a production item. That card sits well outside normal consumer or even prosumer budgets, but it confirms the ceiling modders are chasing: as close to datacenter memory capacity as a gaming die can be pushed.
Stock RTX 5090 vs. the Modded Alibaba Cards
Laid side by side, the spec gap explains why buyers are willing to accept a rebuilt card with no factory backing. The table below compares Nvidia’s official RTX 5090, the mainstream 96GB Alibaba mod, the rarer 128GB prototype, and Nvidia’s own China-compliant variant, the RTX 5090D V2, which is a separate and fully legal product built to satisfy export-control performance ceilings rather than a mod.
| Card | VRAM | Price | Status | Architecture |
|---|---|---|---|---|
| RTX 5090 (Founders Edition) | 32GB GDDR7 | $1,999 MSRP (Jan. 30, 2025 launch) | Official Nvidia product | Blackwell |
| RTX 5090 96GB (Suqiao mod) | 96GB | $3,888 on Alibaba | Unofficial third-party rebuild | Blackwell (modified board) |
| RTX 5090 96GB (other Alibaba sellers) | 96GB | ~$5,900 | Unofficial third-party rebuild | Blackwell (modified board) |
| RTX 5090 128GB (prototype) | 128GB | ~$13,000 | Unofficial, limited-run prototype | Blackwell (modified board) |
| RTX 5090D V2 (China) | ~25% less VRAM/bandwidth than global RTX 5090 | $2,299 MSRP in China | Official Nvidia China SKU | Blackwell (export-compliant) |
That last row matters for context. Nvidia already sells a China-specific, deliberately downgraded RTX 5090 variant, the RTX 5090D V2, which trims roughly a quarter of the VRAM and memory bandwidth off the global card to stay under export-control performance thresholds, while keeping the same $2,299 MSRP in China according to Tom’s Hardware’s coverage. In other words, Nvidia’s own compliant answer to Chinese demand for the RTX 5090 is a card with less memory. The Alibaba mod scene is doing the opposite: taking the globally sold, uncapped RTX 5090 and pushing memory well past what Nvidia offers in any version, official or restricted.
This Isn’t New: The RTX 4090 48GB Mod Economy Since 2023
The RTX 5090 story is a sequel. Chinese workshops have been rebuilding RTX 4090 cards since 2023, doubling their stock 24GB of GDDR6X to 48GB and marketing the results under names like “RTX 4090D 48G” (a companion mod bumps the RTX 4080 from 16GB to 32GB). These conversions were never supported by Nvidia and were documented independently by outlets including Tom’s Hardware, TweakTown, ExtremeTech, and HotHardware, along with a steady stream of Reddit and Taobao listings tracked by the r/LocalLLaMA community.
The technique is closer to factory rework than garage tinkering. Technicians desolder the original Ada Lovelace GPU die and memory chips from a donor RTX 4090, then transplant them onto a custom “clamshell” PCB that carries memory packages on both sides of the board instead of one. Twelve additional Micron GDDR6X modules go on alongside the original twelve, doubling the total to 24 chips, and a custom BIOS is flashed to make the card correctly enumerate 48GB instead of 24GB. Tom’s Hardware traced the software trick back to a leaked Nvidia BIOS mechanism for partial memory-controller changes, the same category of leak now suspected in the RTX 5090’s 96GB conversion.
Pricing has followed a consistent pattern across three years and two GPU generations, summarized below using figures reported by Tom’s Hardware, TweakTown, BigGo News, and UK/EU repair shops that now offer the same service as a standing catalog item rather than a one-off mod.
| Source / Market | Mod | Reported Price | Notes |
|---|---|---|---|
| Chinese retail (TweakTown, BigGo News) | RTX 4090 24GB → 48GB | ~$3,400 (with AIO cooler) | Sold openly in China, early 2025 |
| Taobao listings (Reddit-tracked) | RTX 4090 24GB → 48GB / “4090D” variant | ~$3,300 / ~$2,900 | Multiple active sellers as of 2025-2026 |
| Component cost breakdown (Tom’s Hardware) | Upgrade kit only | ~$430–$1,800+ (excludes donor card) | 12 extra GDDR6X modules ~$288 total, plus custom PCB/cooler |
| UK repair shop (Blackstone Repair) | RTX 4090 24GB → 48GB | From £1,000 | Customer sends in own card |
| EU repair shop | RTX 4090 24GB → 48GB | €1,600 (own card) / €3,900 (card included) | 3-month shop warranty |
| Alibaba (Shenzhen Suqiao) | RTX 5090 32GB → 96GB | $3,888 | September 11, 2026 listing |
The RTX 4090 mod economy has had three years to mature into something resembling an actual supply chain, with component sourcing, standardized kits, and international repair shops now offering the same service outside China. The RTX 5090 conversion is following the identical playbook on a newer, more expensive donor card.
Why VRAM, Specifically, Is the Thing Worth Rebuilding a GPU For
Gamers rarely need more than 16GB to 24GB of VRAM even at 4K with ray tracing enabled. AI inference and fine-tuning workloads are a different story: model size, context length, and batch size all scale directly with available memory, and running out of VRAM doesn’t degrade a large language model’s performance gracefully, it simply stops the job from running at all. That’s the specific pain point these mods target. Chinese buyers documented on Reddit have used 48GB RTX 4090 mods to run Llama 3.1 70B locally and to train video generation models at 720p resolution with a batch size of four, workloads that would otherwise require a genuine datacenter card or a cloud GPU rental.
That’s the economic logic driving both the RTX 4090 and RTX 5090 mod markets: a rebuilt gaming card that costs $3,000 to $4,000 delivers memory capacity that would otherwise require equipment priced in the tens of thousands of dollars, with none of the export paperwork.
The Export Control Backdrop Making These Mods Worth the Risk
None of this happens in a vacuum. U.S. export policy on Nvidia’s AI-class chips to China shifted twice in 2026, and both shifts help explain why a modded gaming card looks attractive to Chinese AI teams. In a January 2026 final rule, the Bureau of Industry and Security moved Nvidia’s H200 and AMD’s MI325X from a “presumption of denial” posture to “case-by-case licensing” for exports to China and Macau, according to BIS’s own press release. The change applies only to chips under a specific performance ceiling (total processing performance below 21,000 and DRAM bandwidth below 6,500 GB/s), and approval still requires third-party U.S. testing, customer-screening commitments, and a hard cap limiting China to no more than 50% of the volume sold to American customers.
Nvidia’s Blackwell-generation datacenter chips, the B100, B200, and GB200, remain flatly excluded from that opening. They sit under presumption-of-denial status for China in 2026, a restriction reinforced on May 31, 2026, when BIS issued weekend guidance clarifying that license requirements apply to any company whose ultimate parent is headquartered in China, even if the purchasing entity sits in a subsidiary outside the country, according to reporting from Al Jazeera. That guidance closed a subsidiary workaround that some Chinese-linked firms had reportedly used to acquire Blackwell hardware indirectly. Blackwell chips also depend on TSMC’s 4NP process node, which is separately barred from China-bound production runs under U.S. national-security rules, adding a manufacturing-level barrier on top of the licensing one.
Put plainly: the newest, most capable Nvidia AI silicon is essentially unreachable for Chinese buyers in 2026, and even the mid-tier H200 requires clearing a licensing gauntlet most smaller labs and startups can’t navigate. A rebuilt gaming GPU, bought retail and modified domestically, sidesteps every part of that process. It’s not a datacenter-class part, but for teams doing inference or fine-tuning at a moderate scale, 96GB of memory for under $4,000 is a workable substitute that requires no license, no compliance review, and no wait.
What the Warranty Fine Print Actually Says
Buyers of these cards are trading Nvidia’s backing for the mod shop’s own guarantee, and that trade is total. Nvidia’s official GeForce Graphics Cards Warranty covers manufacturing defects and hardware component failures for three years from the date of purchase, but explicitly excludes damage tied to abuse, misuse, negligence, or misapplication of service by a non-authorized party. Nvidia’s broader service terms for its DGX line go further, stating coverage applies only to unmodified products used exactly as documented, and that opening a case for self-service, let alone desoldering the GPU die, voids support entirely.
Once a card like the RTX 5090 or RTX 4090 has its memory chips physically replaced and its BIOS re-flashed, it is no longer, in Nvidia’s terms, an unmodified product. The three-year factory warranty is gone the moment the rework begins. That’s precisely why the mod shops profiled by Tom’s Hardware, Blackstone Repair, and various EU services offer their own short-term coverage, typically just three months, reflecting the real failure risk baked into a hand-reworked, non-standard board.
Modded Gaming Cards vs. Real Datacenter Silicon
Even at 96GB or 128GB, a modified RTX 5090 is not a substitute for genuine enterprise AI hardware, it’s a budget alternative for teams priced out of the real thing. The gap in both capacity and cost is stark.
| GPU | Memory | Hardware List Price | Cloud Rental (per GPU/hr) |
|---|---|---|---|
| RTX 5090 96GB (modded) | 96GB GDDR7 | $3,888 | Not applicable (retail card) |
| Nvidia H100 | 80GB HBM | ~$31,000; 8-GPU HGX system $250,000–$320,000 | $0.57–$14.90 (avg. ~$3.99–$4.09 on major platforms) |
| Nvidia H200 | 141GB HBM3e | ~$39,999 (retailer list) | $0.93–$13.78 (~$3.22 typical) |
| Nvidia B200 | 192GB HBM3e | ~$50,000–$70,000 (rumored) | $3.35–$16.11 (~$5.25 typical) |
The comparison explains the appeal without overselling it. A single H100 card lists around $31,000, and an eight-GPU HGX H100 system runs $250,000 to $320,000, according to pricing data compiled by GetDeploying and Morphllm. A modded RTX 5090 costs roughly an eighth of a single H100’s list price for a bit more raw memory (96GB versus 80GB), but none of the HBM bandwidth, error correction, NVLink interconnect, or driver-level enterprise support that make H100, H200, and B200 the actual workhorses of large-scale AI training. For a closer look at how the stock RTX 5090 stacks up against its predecessor on pure gaming performance, see our RTX 5090 vs RTX 4090 comparison. What the modded card buys is memory capacity for inference and mid-scale fine-tuning, not a datacenter replacement.
A Bigger Pattern: 2026’s Year of Export-Control Workarounds
The Alibaba GPU mod is one data point in a much larger story that’s played out across 2026: Chinese entities finding creative, sometimes legally contested paths around U.S. chip restrictions. Reporting this year uncovered Megaspeed’s Malaysian subsidiary, Speedmatrix, purchasing an estimated $2 billion in Nvidia chips, and Chinese server maker Inspur routing roughly $5.6 billion in Nvidia hardware purchases through Aivres, an offshore-registered entity. Samsung and SK Hynix have also been drawn into the picture, testing China-made chipmaking tools after four of their fabs lost Verified End User status earlier in 2026, forcing a review of which equipment those fabs can keep using. Meanwhile, Beijing has publicly rejected U.S. claims that Chinese AI labs distilled American models to shortcut their own development, naming six firms in its rebuttal.
None of those stories involve physically modifying consumer GPUs, but they share the same underlying dynamic: export restrictions create a price and access gap, and someone, whether a shell subsidiary, a chip-tool importer, or a Shenzhen OEM with a soldering iron, moves to close it. The RTX 5090 mod is simply the version of that story that plays out on a retail marketplace instead of in a corporate structuring memo.
Market Impact: What This Means for Nvidia’s Product Strategy
For Nvidia, a thriving gray-market mod economy is an awkward signal wrapped around otherwise good news. The fact that Chinese buyers are willing to pay a premium to rebuild an already-expensive consumer card into something resembling a workstation GPU confirms just how much unmet demand exists for mid-tier AI compute inside China, demand Nvidia’s own compliant product, the RTX 5090D V2, was specifically designed not to fully satisfy by trimming its VRAM and bandwidth below the global card. Every modded RTX 5090 sold outside Nvidia’s channel is also a card Nvidia didn’t get standard margin on twice: once at original retail, and again on whatever premium the reseller charges after the rebuild, since the donor card itself still had to be purchased at market price before the mod shop touched it.
There’s also a support-cost angle. Even though Nvidia’s warranty explicitly excludes modified hardware, a card as visibly non-standard as a clamshell-rebuilt RTX 5090 with double-sided memory chips is unlikely to pass unnoticed if it ever reaches an authorized service center, but the sheer volume of RTX 4090 mods documented since 2023 suggests enforcement at the point of sale, rather than after the fact, has not meaningfully slowed the practice.
Competitive Landscape: How This Compares Across the GPU Industry
Nvidia isn’t the only company whose chips are caught in this dynamic, but it is by far the largest target because of its dominant position in AI compute. AMD’s MI325X falls under the same case-by-case H200-tier licensing framework the January 2026 BIS rule established, meaning it faces an identical, if smaller-scale, version of the same access gap. There’s no public reporting of a comparable gray-market VRAM-mod economy built around AMD’s Radeon gaming cards, which likely reflects Nvidia’s overwhelming share of both the gaming GPU market and the AI software ecosystem, since CUDA compatibility matters enormously for anyone trying to repurpose a gaming card for machine learning work, and that ecosystem lock-in points buyers toward GeForce cards specifically.
On the memory side, Micron and Samsung GDDR6X and GDDR7 chips are the components actually being harvested and re-soldered in these mods, meaning the memory makers are, indirectly, supplying both Nvidia’s official product line and the parallel gray-market rebuild industry at the same time, through entirely separate distribution channels.
The Cost-Per-Gigabyte Math Buyers Are Actually Running
Strip away the marketing angle and the decision comes down to simple arithmetic that Chinese AI buyers are clearly already running for themselves.
- Stock RTX 5090: $1,999 / 32GB = $62.47 per GB
- Modded RTX 5090 96GB: $3,888 / 96GB = $40.50 per GB
- Nvidia H100: ~$31,000 / 80GB = $387.50 per GB
- Nvidia H200: ~$39,999 / 141GB = $283.68 per GB
- Nvidia B200: ~$60,000 / 192GB = $312.50 per GB (midpoint of rumored range)
By raw cost-per-gigabyte, the modded card is cheaper than every legitimate option on the list, including the stock RTX 5090 itself. That math ignores bandwidth, reliability, support, and the fact that GDDR7 is not HBM3e, but for a buyer whose bottleneck is simply fitting a large model into memory at all, it’s an easy number to be persuaded by.
Historical Context: Gray Markets Have Always Chased Export Bans
The pattern isn’t unique to GPUs. Whenever a government restricts a technology’s flow to a specific market, that market historically self-organizes to route around the restriction, from Cold War-era electronics smuggling to more recent cases of restricted networking and semiconductor equipment reaching sanctioned buyers through third countries. What’s notable about the RTX 4090 and RTX 5090 mod economy is how visible it is. These aren’t hidden transactions, they’re listed openly on Alibaba, discussed openly on Reddit and Bilibili, and covered openly by mainstream tech press. The openness itself is a signal that mod shops and their customers view the practice as occupying a gray zone rather than a clearly illegal one, since VRAM modification of a legally purchased consumer product sits in a different legal category than smuggling a restricted datacenter chip across a border.
Predictions: Where This Goes From Here
Based on the trajectory from the RTX 4090 mod cycle to the RTX 5090 cycle now underway, a few outcomes look likely over the next 12 to 18 months.
- Mod pricing will compress. The RTX 4090’s 48GB conversion dropped from an early premium toward roughly $2,900 to $3,400 within about two years of appearing; expect the RTX 5090’s 96GB mod to follow a similar downward curve as more Shenzhen shops enter the market and component sourcing scales.
- Nvidia will keep segmenting China-specific SKUs rather than chasing individual mod shops. The RTX 5090D V2’s reduced-spec, same-price approach suggests Nvidia’s strategy is to control what it can, official channel products, rather than police what it can’t, aftermarket rebuilds of cards already sold.
- U.S. export policy will likely see further incremental adjustment rather than a single dramatic reversal, continuing the back-and-forth seen in the January 2026 H200 easing followed by the May 2026 subsidiary-loophole crackdown.
- Expect a 128GB or higher modded tier to become more commercially available, following the same escalation path that took the RTX 4090 mod scene from 48GB toward the rarer 96GB variant over roughly two years.
- Scrutiny of secondhand and gray-market GPU resale will increase globally as regulators in the U.S. and EU pay closer attention to where high-VRAM consumer cards ultimately end up, given their dual-use potential for AI workloads.
What Buyers Outside China Should Take Away From This
For readers outside China who might be tempted by a heavily discounted, high-VRAM RTX 5090 or RTX 4090 showing up on a resale platform or an overseas Alibaba storefront, the practical risk calculus hasn’t changed from what has applied to modded 4090s since 2023: no Nvidia warranty, no guarantee the BIOS mod remains stable under sustained load, and no recourse beyond whatever short-term guarantee the individual mod shop offers. For workloads that genuinely need the memory and can tolerate that risk, particularly hobbyists and small labs running local large language models who would otherwise be priced out of any high-VRAM option entirely, the economics are real. For anyone running production workloads, the gap between a $3,888 gray-market card and a supported, warrantied piece of hardware is the entire point of paying more in the first place. For more on how AI chip pricing and availability are reshaping the broader hardware market, see our ongoing AI chips coverage.
Frequently Asked Questions
What is the China-modified RTX 5090 with 96GB of VRAM?
It’s a standard Nvidia GeForce RTX 5090 that has had its factory 32GB of GDDR7 memory removed and replaced with 96GB through a board-level rework. Shenzhen Suqiao Intelligent Technology began selling the modified cards on Alibaba around September 11, 2026, for $3,888, with other sellers listing similar cards closer to $5,900.
How does the price compare to the official RTX 5090?
Nvidia’s official RTX 5090 launched on January 30, 2025, at a $1,999 Founders Edition MSRP with 32GB of VRAM. The modded 96GB version costs roughly $3,888 to $5,900 depending on the seller, roughly double to triple the original MSRP for three times the memory.
Does modifying an RTX 5090 or RTX 4090 void the Nvidia warranty?
Yes. Nvidia’s official GeForce Graphics Cards Warranty excludes damage or issues caused by abuse, misuse, or servicing by a non-authorized party, and the company’s broader service terms state that support only applies to unmodified products. A card with its GPU die and memory desoldered and replaced is no longer considered unmodified.
Why do buyers want more VRAM than Nvidia officially provides?
AI inference and fine-tuning workloads scale directly with available memory. Larger VRAM pools let a single GPU load bigger language models, longer context windows, or larger training batch sizes. Buyers of 48GB RTX 4090 mods have reported running Llama 3.1 70B locally and training video generation models, workloads that would otherwise require far more expensive datacenter hardware.
What’s the difference between this mod and Nvidia’s own RTX 5090D V2?
The RTX 5090D V2 is an official, Nvidia-sanctioned China-market SKU that reduces VRAM and memory bandwidth by roughly 25% compared to the global RTX 5090, in order to comply with U.S. export-control performance thresholds, while keeping a $2,299 MSRP. The Alibaba mod is the opposite: an unofficial, unsanctioned rebuild of the full global RTX 5090 that increases memory well beyond any version Nvidia sells.
Are these modded GPUs legal to buy and own?
Buying and owning a modified consumer GPU is not the same legal question as exporting a restricted datacenter AI chip. The RTX 5090 and RTX 4090 are retail gaming products with no export restriction on the base card; the mods are aftermarket hardware modifications, which typically fall into a legal gray area around warranty and consumer protection rather than export control law. Buyers should not assume any manufacturer support or recourse.
What is the current status of U.S. export controls on Nvidia AI chips to China in 2026?
As of 2026, Nvidia’s H200 and AMD’s MI325X are eligible for case-by-case export licensing to China and Macau under a January 2026 BIS rule, subject to performance ceilings, third-party testing, and volume caps. Nvidia’s Blackwell-generation datacenter chips (B100, B200, GB200) remain under a presumption of denial and are effectively banned for China, a position reinforced by May 2026 guidance closing a subsidiary-based workaround.
Can these modded cards still be used for gaming?
In principle yes, since the GPU die itself is unchanged, but the added memory provides no gaming benefit since no current game uses anywhere near 32GB of VRAM, let alone 96GB. The mods are built and marketed specifically for AI workloads, not gaming performance.

