Nvidia
Nvidia
@nvidia·2.2M subscribers·2.7K videos

Scaling AI Inference Performance in the Cloud with Nebius

Posted

November 10, 2025

Views

439,336

Likes

476

Engagement

0.11%

Search the Record

Indexed

Every word spoken in this episode is indexed. Type any phrase to jump straight to the moment it was said.

Type any word or phrase that may have been spoken. Click a result to seek the player to that exact moment.

Try a name, a topic, or a quoted line

YouTube Description

as posted by the channel

When it comes to future-proofing AI deployments, you need reliable underlying AI infrastructure that is purpose-built for high performant and scalable inference.

Explore how Nebius built and deployed its high-reliance AI cloud, powered by NVIDIA.

Using managed Kubernetes with auto-scaling, Nebius optimizes its AI cloud to deliver multi-node training and inference of frontier models and AI applications for startups and enterprises. Nebius, an ecosystem partner for Dynamo, is enabling AI inference at scale with NVIDIA infrastructure.

Nebius, an ecosystem partner for Dynamo, is enabling AI inference at scale with NVIDIA infrastructure. Learn how NVIDIA Dynamo and Kubernetes help scale highperformance AI inference in our new Think SMART blog

Explore Nebius and NVIDIA collaboration

Watch Nebius leadership session from GTC

Guests & Subjects Covered

NVIDIA UsingNVIDIA DynamoThink SMARTExplore NebiusWatch Nebius

Sentinel Indexing in Progress

Metadata and chapters are available. Claim extraction for this episode is pending.

All video content is delivered via YouTube embedded players in accordance with the YouTube Terms of Service. Sentinel provides research tools that promote discovery and accountability across political media.

Scaling AI Inference Performance in the Cloud with Nebius · Nvidia · Sentinel