7 Sources
[1]
AMD announces MI350P PCIe AI accelerator card with 144GB of HBM3E -- roughly 40% faster in FP16 and FP8 theoretical compute compared to Nvidia's H200 NVL competitor
AMD now has the fastest AI accelerator card on the market that fits in a traditional PCIe slot. AMD has launched a new member of the MI350-series that comes in a PCIe form factor. The new Instinct MI350P comes with 128 CUs and 144GB of HBM3E memory and is designed to be a drop-in upgrade solution
[2]
AMD Instinct MI350P: PCIe Add-In Card For High Performance Open-Source AI/Compute Review
While there is the AMD Instinct MI400 series coming this year, today AMD announced an interesting and arguably overdue offering for the Instinct MI350 series: the MI350P. The AMD Instinct MI350P is a PCIe add-in-card to add Instinct MI350 compute capabilities to existing PCIe 5.0 air-cooled servers
[3]
AMD puts out new slottable GPU for AI-curious enterprises
AMD hopes to win over enterprise AI customers with a more affordable datacenter GPU that can drop into conventional air-cooled servers. Announced on Thursday, the MI350P is the House of Zen's first PCIe-based Instinct accelerator since the MI210 debuted all the way back in 2022. Until now, AMD's
[4]
AMD Instinct MI350P PCIe Targets Air-Cooled Enterprise AI Servers
AMD has introduced the Instinct MI350P PCIe GPU, a new enterprise accelerator designed for AI inference workloads in existing data center environments. The card uses a dual-slot PCIe format and is intended for standard air-cooled servers, giving organizations a way to add GPU acceleration without
[5]
AMD launches the Instinct MI350P GPU with 144GB of HBM3E and a 600W TBP
TL;DR: AMD introduced the Instinct MI350P, a cost-effective PCIe GPU accelerator for AI workloads, featuring 128 compute units, 144GB HBM3E memory, and up to 4,600 TFLOPS performance. Designed for easy integration in air-cooled servers, it targets enterprises seeking scalable AI infrastructure with
[6]
AMD Launches MI350P, Its First PCIe "Instinct" In Four Years - Packs CDNA 4 GPU With 4.6 PFLOPs AI Compute, 144 GB HBM3E at 600W
AMD has announced its brand new Instinct MI350P PCIe GPU accelerator, which is the first PCIe design in years and is aimed at AI workloads. With the Instinct MI350P PCIe GPU, AMD gives enterprise users an option to expand their AI computing capabilities without having to invest in expensive
[7]
AMD Launches Instinct MI350P PCIe GPUs for Enterprise AI Workloads
Our new AMD Instinct™ MI350 PCIe® cards give your enterprise a third option: Leadership AI performance designed to fit the data center infrastructure you already own. Performance That Drops into Your Existing Racks Designed to help you prepare for the agentic AI era, AMD Instinct MI350P PCIe
Share
Copy Link
AMD unveiled the Instinct MI350P, a PCIe AI accelerator card with 144GB of HBM3E memory designed for air-cooled enterprise servers. The dual-slot card delivers up to 4,600 TFLOPS performance and outpaces Nvidia's H200 NVL by roughly 40% in FP16 and FP8 theoretical compute. With support for up to eight cards per system, AMD targets enterprises seeking scalable AI infrastructure without major platform overhauls.
AMD has introduced the Instinct MI350P, marking the company's first PCIe-based AI accelerator since the MI210 debuted in 2022
1
. The new PCIe AI accelerator card addresses a critical gap in AMD's portfolio by offering enterprises a drop-in upgrade path for existing infrastructure without requiring specialized accelerator platforms3
. This launch positions AMD to compete directly against Nvidia's H200 NVL competitor in the enterprise AI market, particularly for organizations exploring on-premise AI deployments.
Source: Wccftech
The Instinct MI350P features 128 compute units with 8,192 stream processors and 512 Matrix cores, built on the CDNA4 architecture using TSMC's 3nm and 6nm FinFET process . The dual-slot card measures 10.5 inches and integrates 144GB HBM3E memory with 4TB/s of memory bandwidth across an 8192-bit bus
1
. AMD rates the card at a 600W TBP, though power capping allows operation at 450W for more power-constrained environments2
. The passively cooled design relies on chassis fans in air-cooled enterprise AI servers, making it compatible with standard 19-inch rack-mounted configurations3
.
Source: Guru3D
AMD claims the MI350P delivers estimated performance of 2,299 TFLOPS, with peak MXFP4 performance reaching 4,600 TFLOPS—the highest currently available in an enterprise PCIe card
4
. Compared to Nvidia's H200 NVL, the card demonstrates approximately 20% better FP64, 43% better FP16, and 39% better FP8 theoretical compute performance1
. The AI accelerator supports native MXFP6 and MXFP4 lower-precision formats to accelerate Large Language Models (LLMs), along with FP8, MXFP8, INT8, and BF16 precision formats4
. AMD also promotes the card as capable of handling 200 to 250 billion parameter large language models per GPU2
.The Instinct MI350P supports configurations ranging from one to eight cards per system, enabling data centers to scale performance based on workload requirements
1
. AMD targets the card specifically at AI inference workloads, RAG pipelines, and production AI deployments across small, medium, and large enterprise implementations4
. However, the card relies on PCIe 5.0 for chip-to-chip communications at 128GB/s, lacking the high-speed Infinity Fabric interconnects found on AMD's OAM-based MI350X and MI355X accelerators3
. This limitation may affect performance in larger multi-GPU configurations compared to systems with dedicated accelerator interconnects.Source: Phoronix
Related Stories
AMD positions the MI350P within its open enterprise AI software environment, supporting ROCm, Kubernetes GPU Operator, AMD Inference Microservices, and native framework support including PyTorch
4
. The company provides its enterprise AI reference stack to partners without licensing costs, contrasting with proprietary approaches4
. However, widespread adoption remains uncertain given Nvidia's entrenched position with CUDA in the AI market1
. AMD continues developing its ROCm software stack to improve competitiveness, though the company faces an uphill battle against established developer ecosystems.The MI350P launch arrives as Nvidia has not yet announced a PCIe version of its latest B200 Blackwell GPUs with HBM memory, temporarily giving AMD the most advanced AI accelerator in PCIe form factor
1
. Nvidia currently offers its RTX Pro 6000 Blackwell cards to enterprise customers, which sell for $8,000 to $10,000 but deliver significantly lower specifications—the MI350P provides 2.3x higher peak TFLOPS, 2.5x the memory bandwidth, and 50% more VRAM3
. While AMD has not disclosed pricing, competitive positioning against both the H200 NVL and RTX Pro 6000 will prove critical for market penetration. Intel's upcoming Crescent Island AI accelerator with 160GB LPDDR5X memory will add another competitor later this year, though it lacks the high-bandwidth HBM3E memory of AMD's offering2
. The MI350P is now available through AMD partners, though the timing is notable given that AMD's MI400 series is expected to launch within the year2
.Summarized by
Navi
[2]
[3]
11 Jun 2025•Technology

11 Oct 2024•Technology

06 Feb 2025•Technology

1
Technology

2
Policy and Regulation

3
Technology
