Scaling AI Inference Performance in the Cloud with Nebius
November 10, 2025
439,336
476
0.11%
Search the Record
IndexedEvery word spoken in this episode is indexed. Type any phrase to jump straight to the moment it was said.
Type any word or phrase that may have been spoken. Click a result to seek the player to that exact moment.
Try a name, a topic, or a quoted line
Nvidia Episodes Around November 10, 2025
See what was published immediately before and after this episode.
1:59Now PlayingScaling AI Inference Performance in the Cloud with Nebius
YouTube Description
as posted by the channelWhen it comes to future-proofing AI deployments, you need reliable underlying AI infrastructure that is purpose-built for high performant and scalable inference.
Explore how Nebius built and deployed its high-reliance AI cloud, powered by NVIDIA.
Using managed Kubernetes with auto-scaling, Nebius optimizes its AI cloud to deliver multi-node training and inference of frontier models and AI applications for startups and enterprises. Nebius, an ecosystem partner for Dynamo, is enabling AI inference at scale with NVIDIA infrastructure.
Nebius, an ecosystem partner for Dynamo, is enabling AI inference at scale with NVIDIA infrastructure. Learn how NVIDIA Dynamo and Kubernetes help scale highperformance AI inference in our new Think SMART blog
Explore Nebius and NVIDIA collaboration
Watch Nebius leadership session from GTC
Links & Promotions
Guests & Subjects Covered
Sentinel Indexing in Progress
Metadata and chapters are available. Claim extraction for this episode is pending.
All video content is delivered via YouTube embedded players in accordance with the YouTube Terms of Service. Sentinel provides research tools that promote discovery and accountability across political media.









