Nvidia gaming GPUs modded with 2X VRAM for AI workloads — RTX 4090D 48GB and RTX 4080 Super 32GB go up for rent at Chinese cloud computing provider

GeForce RTX 4090
GeForce RTX 4090 (Image credit: Nvidia)

AI enthusiast 青龍聖者 has discovered two fascinating graphics cards in China:  the GeForce RTX 4090D 48GB and GeForce RTX 4080 Super 32GB. The mysterious SKUs are clearly modified versions of the GeForce RTX 4090D and GeForce RTX 4080 Super, which contend with the best graphics cards.

The GeForce RTX 4090D 48GB and GeForce RTX 4080 Super 32GB are available for rent at AutoDL, a Chinese cloud computing provider that rents servers for AI work. Pricing is a steal. You can rent a single GeForce RTX 4080 Super 32GB for $0.03 hourly. However, the service is currently restricted to China, as you need a Chinese phone number to sign up.

Latest Videos FromTom's Hardware
Zhiye Liu
News Editor, RAM Reviewer & SSD Technician

Zhiye Liu is a news editor, memory reviewer, and SSD tester at Tom’s Hardware. Although he loves everything that’s hardware, he has a soft spot for CPUs, GPUs, and RAM.

  • Squishynidas
    They need to make GPU ram modular like system ram. We keep buying stuff we could easily reuse.
    Reply
  • hotaru251
    Squishynidas said:
    They need to make GPU ram modular like system ram.
    no.
    there is a reason its soldered on now-a-days.
    Reply
  • Notton
    Squishynidas said:
    They need to make GPU ram modular like system ram. We keep buying stuff we could easily reuse.
    okay, how are you going to fit it?
    Reply
  • Makaveli
    hotaru251 said:
    no.
    there is a reason its soldered on now-a-days.
    And also giving customers easy way to increase VRAM and not making them buy a new product is bad for the bottom line of AMD, NV and Intel.
    Reply
  • Evildead_666
    Makaveli said:
    And also giving customers easy way to increase VRAM and not making them buy a new product is bad for the bottom line of AMD, NV and Intel.
    Back in the day, when this was possible, they were almost all proprietary, and were not cheap...
    I remember getting a memory upgrade for my Matrox Mystique...

    Also, the memory type changes every couple of generations.

    I could see it happening on High end Pro cards, as a memory doubler thing though.
    Reply
  • jlake3
    Notton said:
    okay, how are you going to fit it?
    In addition to physical size, how about the trace length and signal integrity problems?

    Also GPU memory controllers don't work like CPU memory controllers, so if it were possible you'd probably end up with something like a single CAMM socket rather than multiple DIMM slots, and because the bus widths and number of memory modules per GPU aren't constant, you'd need different daughterboards for each configuration.
    Reply
  • toaste
    Notton said:
    okay, how are you going to fit it?

    You could do it with something like CAMM on the back side of the board, but only if you dropped the GDDR speed dramatically. Nobody would like the crippling performance drop the GDDR bandwidth reduction would bring with it.

    Makaveli said:
    And also giving customers easy way to increase VRAM and not making them buy a new product is bad for the bottom line of AMD, NV and Intel.
    Yeah, that's not a concern. If it were possible, then Intel or AMD would do it. They're competing for like 2% and 10% of the market, and increased longevity would be massively offset by the piles of money afforded by taking a chunk of Nvidia's pie.

    It's not possible because the GPU relies on increasing memory bandwidth to keep up with increasing shader vertex or texture processing.

    The most extreme GDDR5 overclocking is at around DDR5 7200 or so. So 7.2Gbit/s. It pumps MCLK at a languid 3.6GHz. Pathetic.

    Your GDDR6 gpu is pumping the WCK pin at 8 or 9 GHz to hit 16 or 18Gb/s. GDDR6x uses QAM so it's "only" pushing around 5GHz.

    The PCB routing is kept physically as short as possible, and arranging it so the traces are all EXACTLY the same length and impedance is critical. You simply cannot get a signal through a PCB at that rate with a mechanical connector in the way.
    Reply
  • Lucky_SLS
    This is where tiered memory comes into play. Intel already has this in its Lunar lake. Consider 16 gigs of GDDR7 memory and supplement the rest with CAMM2. This will still have performance penalty, but at least you are no longer limited by memory.
    Reply
  • nocturn9x
    Lucky_SLS said:
    This is where tiered memory comes into play. Intel already has this in its Lunar lake. Consider 16 gigs of GDDR7 memory and supplement the rest with CAMM2. This will still have performance penalty, but at least you are no longer limited by memory.
    How is that even remotely worth it? Lots of memory is useless if you can't access it fast enough
    Reply
  • Lucky_SLS
    Who said this is for gaming or AI? Productivity applications does not benefit the same from fast memory. But memory capacity improves the capability.
    Reply