7 Sources
[1]
OpenAI's New Open Models Accelerated Locally on NVIDIA GeForce RTX and RTX PRO GPUs
The groundbreaking open-weight models are now available with optimizations for RTX AI PCs. In collaboration with OpenAI, NVIDIA has optimized the company's new open-source gpt-oss models for NVIDIA GPUs, delivering smart, fast inference from the cloud to the PC. These new reasoning models enable
[2]
OpenAI and NVIDIA Propel AI Innovation With New Open Models Optimized for the World's Largest AI Inference Infrastructure
NVIDIA delivers industry-leading gpt-oss-120b performance of 1.5 million tokens per second on a single NVIDIA Blackwell GB200 NVL72 rack-scale system. Two new open-weight AI reasoning models from OpenAI released today bring cutting-edge AI development directly into the hands of developers,
[3]
OpenAI and NVIDIA set global AI benchmark with gpt-oss models
These models mark a major step forward in open AI development, offering state-of-the-art performance, broad flexibility, and efficiency across a wide range of deployment environments. Trained on NVIDIA's H100 GPUs and optimized for deployment across its massive CUDA ecosystem, the models run best
[4]
OpenAI's new open-weight reasoning model can be run locally on an RTX card but you still need a pretty beefy rig to run it
If you like the premise of AI doing, well something, in your rig, but don't much fancy feeding your information back to a data set for future use, a local LLM is likely the answer to your prayers. With OpenAI's latest model, you can do just that, assuming you have the hardware to power
[5]
You can now Deploy gpt-oss-20b Offline on NVIDIA GeForce RTX GPUs with 16GB VRAM
With recent updates from NVIDIA and OpenAI, you can now run sophisticated language models entirely on your own PC -- no cloud account, no monthly fees. All you need is a GeForce RTX card with at least 16 GB of VRAM. This opens the door to powerful AI capabilities directly on your desktop, whether
[6]
NVIDIA's RTX GPUs Deliver Fastest AI Performance On OpenAI's Latest "gpt-oss" Models
NVIDIA & OpenAI have brought the latest gpt-oss family of AI open models to consumers, offering the highest performance on RTX GPUs. NVIDIA's RTX 5090 Delivers 250 Tokens/s Performance on OpenAI's gpt-oss 20b AI Model, PRO GPUs Also Ready For gpt-oss 120b Press Release: Today, NVIDIA announced
[7]
OpenAI GPT-OSS Models Optimized for NVIDIA RTX GPUs
NVIDIA and OpenAI have collaborated to release the gpt-oss family of open-source AI models, optimized for NVIDIA RTX GPUs. These models, gpt-oss-20b and gpt-oss-120b, bring advanced AI capabilities to consumer PCs and workstations, enabling faster and more efficient on-device AI
Share
Copy Link
OpenAI and NVIDIA collaborate to release open-weight AI models, gpt-oss-20b and gpt-oss-120b, optimized for local deployment on NVIDIA GPUs, enabling developers to run advanced AI models offline on personal computers and workstations.
In a groundbreaking move, OpenAI and NVIDIA have joined forces to release two new open-weight AI reasoning models, gpt-oss-20b and gpt-oss-120b. This collaboration marks a significant step forward in democratizing access to advanced AI technologies, allowing developers, enthusiasts, and organizations to run sophisticated language models locally on their own hardware
1
2
.
Source: NVIDIA
The gpt-oss-20b model, designed for broader accessibility, can run on GPUs with at least 16GB of VRAM. It offers performance comparable to OpenAI's o3-mini model on common benchmarks
3
. For more demanding applications, the gpt-oss-120b model achieves near-parity with OpenAI's o4-mini on core reasoning benchmarks and requires an 80GB GPU3
.NVIDIA has optimized these models for their hardware, showcasing impressive performance metrics:
1
.2
.
Source: Guru3D
Users have multiple options for deploying these models locally:
1
.1
.1
5
.The release of these open-weight models under the Apache 2.0 license allows for full commercial and research use, potentially accelerating AI innovation across various sectors
3
. Jensen Huang, founder and CEO of NVIDIA, emphasized the significance of this release:"OpenAI showed the world what could be built on NVIDIA AI -- and now they're advancing innovation in open-source software. The gpt-oss models let developers everywhere build on that state-of-the-art open-source foundation, strengthening U.S. technology leadership in AI -- all on the world's largest AI compute infrastructure."
2
Related Stories

Source: PC Gamer
While the gpt-oss-20b model is accessible to a wider range of users with RTX GPUs featuring at least 16GB of VRAM, the more powerful gpt-oss-120b model requires more substantial hardware. AMD has also announced support for these models, with CEO Lisa Su confirming compatibility with AMD AI CPUs and GPUs
4
.Running these models locally offers several advantages, including enhanced privacy, reduced dependence on cloud services, and the ability to work offline. This makes the technology particularly attractive for sectors like finance, healthcare, and government, where data sensitivity is a primary concern
5
.As AI continues to integrate into various aspects of computing and industry, the release of these open-weight models by OpenAI and NVIDIA represents a significant milestone in making advanced AI capabilities more accessible and customizable for developers and organizations worldwide.
Summarized by
Navi
[3]
06 Aug 2025•Technology

07 Aug 2025•Technology

18 Oct 2024•Technology

1
Science and Research

2
Policy and Regulation

3
Technology