Pruning AI Models for Peak Performance - NVIDIA DRIVE Labs Ep. 31
September 19, 2023
10,540
190
1.80%
Search the Record
IndexedEvery word spoken in this episode is indexed. Type any phrase to jump straight to the moment it was said.
Type any word or phrase that may have been spoken. Click a result to seek the player to that exact moment.
Try a name, a topic, or a quoted line
Nvidia Episodes Around September 19, 2023
See what was published immediately before and after this episode.
3:16Now PlayingPruning AI Models for Peak Performance - NVIDIA DRIVE Labs Ep. 31
Chapters
YouTube Description
as posted by the channelCheck out HALP (Hardware-Aware Latency Pruning), a new method designed to adapt convolutional neural networks (CNNs) and #transformer-based architectures for real-time performance. HALP optimizes pre-trained models to maximize compute utilization. In testing with NVIDIA DRIVE Orin™ on the road, it consistently outperformed alternative approaches.
Product page
Watch the full series here
Learn more about DRIVE Labs
Follow us on social:
#NVIDIADRIVE
Guests & Subjects Covered
Sentinel Indexing in Progress
Metadata and chapters are available. Claim extraction for this episode is pending.
All video content is delivered via YouTube embedded players in accordance with the YouTube Terms of Service. Sentinel provides research tools that promote discovery and accountability across political media.









