Nvidia filed 4 patents across 3 categories: AI & Machine Learning (2), Graphics (1), and Audio (1).
The AI & Machine Learning patents cover a hybrid training method for creating virtual characters that move realistically and adapt to new situations, plus a system for generating synthetic videos by separately controlling content appearance and motion. The Graphics patent describes AI-powered multi-frame interpolation technology designed to increase perceived frame rates in real-time with minimal latency. The Audio patent details an AI-driven spatial Audio system that tracks head pose via camera to deliver 3D sound across VR and gaming applications.
A single Graphics patent tackles the challenge of boosting frame rates without sacrificing the split-second responsiveness that competitive gaming and VR demand. The system uses deep neural networks to generate multiple interpolated frames between each rendered frame, going beyond existing techniques that typically produce only one additional frame. This approach aims to deliver smoother visuals while maintaining the ultra-low latency that keeps gameplay feeling immediate and immersive.
Two AI & Machine Learning patents explore different aspects of creating believable virtual motion and content. One addresses video synthesis by separating appearance from movement within a generative adversarial framework, giving creators independent control over what appears on screen and how it moves. The other focuses on character animation through a hybrid reinforcement learning method that merges motion tracking with distribution matching, allowing virtual characters to replicate specific movements accurately while retaining the flexibility to adapt their behavior to new scenarios. Where earlier methods required choosing between precision and adaptability, this approach trains a single model that achieves both qualities simultaneously.
An Audio patent connects visual head tracking to spatial sound generation through an optimized pipeline that reduces the computational overhead typically required for real-time 3D Audio. The system uses camera-based pose estimation to determine head position and orientation, then adjusts Audio output accordingly to maintain spatial accuracy as users move. By streamlining the link between visual tracking and Audio synthesis, the approach aims to make immersive sound practical across VR and gaming applications without demanding excessive processing resources.
All data sourced from USPTO patent filings. Google Patents may take several weeks to index recent publications. If a link is unavailable, search for the patent number at USPTO Patent Public Search.