Skip to main content

NVIDIA Platforms Expand Support for DeepSeek, Gemma, and Nemotron Models

25 AUGUST 2026·2 MIN READ·2 SOURCES·Independently corroborated

NVIDIA is accelerating the deployment of community-built open models, including DeepSeek and Google DeepMind's Gemma, alongside its own multimodal Nemotron models.

NVIDIA Platforms Expand Support for DeepSeek, Gemma, and Nemotron Models

Key takeaways · 3

  • 01

    NVIDIA's AI inference platform supports community-built AI models, including Google DeepMind's Gemma and DeepSeek.

  • 02

    NVIDIA offers its own Nemotron open models specifically designed for long-running AI agents.

  • 03

    DeepSeek models can be customized with the NeMo framework or optimized for data centers using TensorRT-LLM.

Open Models and Inference

NVIDIA offers an AI inference platform for exploring and deploying community-built AI models. [2] NVIDIA Nemotron provides high-efficiency, multimodal, open models designed for long-running AI agents. [1] Another featured option is Gemma, which is a family of lightweight, open models created by Google DeepMind. [2]

DeepSeek Integration

The DeepSeek model family uses an open-source mixture-of-experts (MoE) architecture to deliver advanced reasoning capabilities. [2] These models can be optimized for data center deployments utilizing TensorRT-LLM or customized through the NeMo framework. [2] Additionally, the NVIDIA DeepSeek R1 FP4 model is a quantized variant created with the TensorRT Model Optimizer. [2]

What it means

NVIDIA is actively expanding its ecosystem to support diverse model architectures, positioning its proprietary Nemotron agent models alongside prominent community models like Google DeepMind's Gemma and DeepSeek's MoE architectures. By providing targeted optimization tools such as TensorRT-LLM, the NeMo framework, and TensorRT Model Optimizer, the company ensures its infrastructure remains the premier deployment layer for both proprietary and open-source models in data centers. What the sources don't address: How the inference efficiency and performance of NVIDIA's Nemotron models compare directly to optimized DeepSeek deployments on identical enterprise hardware.

The integration of diverse open-source models with specialized hardware optimization pipelines allows enterprises to more effectively deploy advanced architectures. This reduces the friction of adopting specialized models for discrete inference workloads.

Why it matters
Daily session

Turn this story into practical AI skill after launch.

Get the release link for daily sessions built around your role and industry.

Join the waitlist

How this developed

  1. 25 August 2026

    NVIDIA Accelerates DeepSeek, Gemma, and Nemotron Open Models

  2. 25 August 2026

    Event created from source cluster.

Sources

AI fluency, one session a day, built for your work.