Nvidia revives Rubin CPX AI chip with major redesign for 2027 production
- Ming-Chi Kuo confirms Nvidia revived Rubin CPX AI chip for Q1 2027 production
- New design features 168GB HBM4 memory and standalone MGX ETL rack architecture
- CPX will handle AI prefill tasks alongside Vera Rubin GPUs in a 1:1 ratio
- Nvidia Q2 revenue surged 106% YoY to $96.22 billion, beating estimates
- Shares rose 1.36% to $220.50 amid strong Vera Rubin production ramp

*this image is generated using AI for illustrative purposes only.
Nvidia Corp (NASDAQ: NVDA) has revived its Rubin CPX AI accelerator program after market speculation suggested it had been dropped from the roadmap. Analyst Ming-Chi Kuo stated on Monday that industry checks confirm the chip giant plans to begin production in the first quarter of 2027.
The revived Rubin CPX features a substantially redesigned architecture aimed at delivering stronger prefill performance. Prefill is the stage of AI inference where a model reads and processes input before generating a response. Kuo noted that the new design includes significant changes to both GPU specifications and rack architecture compared to the earlier CPX iteration.
Technical Specifications and Architecture
The updated Rubin CPX will feature 168GB of HBM4 high-bandwidth memory per GPU. This sits between the 288GB of HBM4 memory found in Nvidia’s standard Rubin GPUs and the 128GB of GDDR7 memory in the previous CPX design.
| Component | Memory Type | Capacity | Notes |
|---|---|---|---|
| Rubin CPX (Revived) | HBM4 | 168GB | Redesigned for prefill |
| Rubin GPU | HBM4 | 288GB | Standard inference |
| CPX (Previous) | GDDR7 | 128GB | Earlier design |
Kuo explained that the revived CPX will move into a standalone MGX ETL rack rather than sharing space with Rubin GPUs. Customers can configure systems with 64, 128, 192, or 256 CPX GPUs. Within each 64-GPU module, eight compute trays house eight CPX GPUs each, alongside a switch tray. NVLink will connect the eight GPUs within each tray, while Ethernet handles communication between trays and rack modules.
Integration with Vera Rubin Systems
Rubin CPX is expected to work alongside Nvidia’s Vera Rubin NVL72 systems rather than replacing them. Kuo said Nvidia recommends a 1:1 ratio of CPX to Rubin GPUs. In this configuration, CPX handles the prefill stage and generates the KV cache before transferring that information to Rubin over Ethernet RDMA for the decode stage.
This announcement follows Nvidia’s GTC 2026 event, where the company appeared to remove CPX from its roadmap, instead highlighting Groq 3 LPUs and LPX racks. Nvidia did not immediately respond to requests for comment.
Financial Context
Nvidia reported $96.22 billion in second-quarter revenue in August, marking a 106% increase from a year earlier. This figure surpassed the Street consensus estimate of $92.18 billion. The company also confirmed that Vera Rubin was ramping into full production, with CoreWeave, Nebius, Microsoft Azure, Google Cloud, and Oracle Cloud among its partners.
Nvidia shares closed at $220.50 on Monday, up 1.36%. In after-hours trading, the shares were down 0.10%. According to Benzinga Edge Rankings, Nvidia ranks in the 98th percentile for growth and maintains positive price-trend ratings across short-, medium-, and long-term time frames.
How might the 1:1 CPX-to-Rubin GPU configuration impact total cost of ownership for hyperscalers compared to using standard Rubin GPUs for both prefill and decode stages?
What are the potential supply chain implications for HBM4 memory manufacturers given the revived CPX's specific 168GB requirement versus the 288GB standard Rubin GPUs?
Could the separation of CPX into standalone MGX ETL racks create integration challenges or latency issues when communicating with Vera Rubin NVL72 systems over Ethernet RDMA?

































