Patents by Inventor Fitsum Reda

Fitsum Reda has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Publication number: 20260134258
    Abstract: Neural network architectures and machine learning techniques that support tokenization of raw visual input to generate a compact representation in a latent feature space as well as de-tokenization to generate raw visual output. In at least one embodiment, tokenization systems and methods leverages wavelet transforms and causal operations to capture spatial and temporal dependencies in the raw visual input.
    Type: Application
    Filed: March 26, 2025
    Publication date: May 14, 2026
    Inventors: Fitsum Reda, Jinwei Gu, Xian Liu, Songwei Ge, Ting-Chun Wang, Haoxiang Wang, Ming-Yu Liu
  • Publication number: 20260094239
    Abstract: The disclosed method for generating images includes performing, based on one or more inputs, one or more first denoising diffusion operations using a first trained machine learning model to generate a first image at a first resolution; and performing, based on the one or more inputs and the first image, one or more second denoising diffusion operations using a second trained machine learning model to generate a second image at a second resolution.
    Type: Application
    Filed: August 19, 2025
    Publication date: April 2, 2026
    Inventors: Yogesh BALAJI, Ting-Chun WANG, Jiaojiao FAN, Qinsheng ZHANG, Xiaohui ZENG, Maciej BALA, Yin CUI, Yuval ATZMON, Aaron LICATA, Pooya JANNATY, Siddharth GURURANI, Seungjun NAH, Yu ZENG, John LEWIS, Jacob Samuel HUFFMAN, Yunhao GE, Fitsum REDA, Ming-Yu LIU
  • Publication number: 20260094245
    Abstract: The disclosed method for generating images includes performing, based on one or more inputs, one or more first denoising diffusion operations using a first trained machine learning model to generate a first image at a first resolution; and performing, based on the one or more inputs and the first image, one or more second denoising diffusion operations using a second trained machine learning model to generate a second image at a second resolution.
    Type: Application
    Filed: August 19, 2025
    Publication date: April 2, 2026
    Inventors: Yogesh BALAJI, Ting-Chun WANG, Jiaojiao FAN, Qinsheng ZHANG, Xiaohui ZENG, Maciej BALA, Yin CUI, Yuval ATZMON, Aaron LICATA, Pooya JANNATY, Siddharth GURURANI, Seungjun NAH, Yu ZENG, John LEWIS, Jacob Samuel HUFFMAN, Yunhao GE, Fitsum REDA, Ming-Yu LIU
  • Publication number: 20260089325
    Abstract: The disclosed method for quantizing one or more latent embeddings includes receiving one or more latent embeddings, generating, based on the one or more latent embeddings, one or more channel groups, and generating, based on the one or more channel groups, one or more quantized latent embeddings.
    Type: Application
    Filed: August 15, 2025
    Publication date: March 26, 2026
    Inventors: Dawit Mureja ARGAW, Fitsum REDA, Ming-Yu LIU
  • Publication number: 20260087351
    Abstract: The disclosed method for tokenizing video frames includes receiving one or more video frames from one or more I/O devices, generating, using an encoder, one or more latent embeddings based on the one or more video frames, wherein the encoder comprises one or more patchify modules, one or more spatial-temporal Mamba modules, and one or more token pooling modules, and generating, using a quantizer, one or more quantized latent embeddings based on the one or more latent embeddings.
    Type: Application
    Filed: August 15, 2025
    Publication date: March 26, 2026
    Inventors: Dawit Mureja ARGAW, Fitsum REDA, Ming-Yu LIU
  • Patent number: 12568184
    Abstract: Apparatuses, systems, and techniques to generate interpolated video frames. In at least one embodiment, an interpolated video frame is generated based, at least in part, on a first set of pixel data sampled from a first video frame, and a second set of pixel data sampled from a second video frame based, at least in part, on a set of forward pointing motion vectors from the first video frame to the second video frame.
    Type: Grant
    Filed: July 30, 2020
    Date of Patent: March 3, 2026
    Assignee: NVIDIA Corporation
    Inventors: Fitsum Reda, Karan Sapra, Robert Thomas Pottorff, Shiqiu Liu, Andrew Tao, Bryan Christopher Catanzaro
  • Publication number: 20250384588
    Abstract: Video compression systems based on a variational autoencoder, the variational autoencoder including an encoder and a decoder coupled via a latent space embedding component, the encoder configured to transform an input video into a feature maps of the input video at different feature resolution scales, the latent space embedding component configured to transform the feature maps into a latent space parameter distribution, and the decoder configured to sample the latent space parameter distribution to generate a compressed version of the input video.
    Type: Application
    Filed: June 9, 2025
    Publication date: December 18, 2025
    Applicant: NVIDIA Corp.
    Inventors: Fitsum Reda, Dawit Mureja Argaw, Ming-Yu Liu, Qinsheng Zhang
  • Publication number: 20250278883
    Abstract: Methods are provided for selecting pairs of images from which to generate simulated video sequences by interpolating between the pairs of images. A variety of filters are provided by assessing the quality of such image pairs, such that the quality of the simulated video generated by pairs of images selected thereby is improved. These filters can be performed sequentially, with subsequent filters only executed if all preceding filters have ‘passed,’ thereby reducing the computational cost. Some of the filters include generating, for a pair of images, respective embeddings of the images into a representational space and then comparing the embeddings. Some of the filters include generated a test interpolation image between the pair of images and then assessing a similarity between the test image and the pair of images. Some of the filters include determining optical flow between the pair of images.
    Type: Application
    Filed: May 24, 2022
    Publication date: September 4, 2025
    Inventors: Fitsum REDA, Jane KONTKANEN, Claire YAO, Andrew GALLAGHER, Ting LIU
  • Patent number: 12340484
    Abstract: Apparatuses, systems, and techniques for texture synthesis from small input textures in images using convolutional neural networks. In at least one embodiment, one or more convolutional layers are used in conjunction with one or more transposed convolution operations to generate a large textured output image from a small input textured image while preserving global features and texture, according to various novel techniques described herein.
    Type: Grant
    Filed: March 9, 2020
    Date of Patent: June 24, 2025
    Assignee: NVIDIA Corporation
    Inventors: Guilin Liu, Andrew Tao, Bryan Christopher Catanzaro, Ting-Chun Wang, Zhiding Yu, Shiqiu Liu, Fitsum Reda, Karan Sapra, Brandon Rowlett
  • Publication number: 20240265490
    Abstract: Provided is a computer system that includes one or more processors and one or more non-transitory computer-readable media that collectively store a machine-learned image interpolation model. The machine-learned image interpolation model is configured to: extract, for each of multiple different scales, a respective set of feature values from each of a pair of input images; generate, for each of the multiple different scales, a respective flow estimate for each of the pair of input images that indicates a respective flow from the interpolation time to the respective capture time; warp, for each of the multiple different scales, the respective set of feature values for each of the pair of input images according to the respective flow estimate to generate respective warped sets of features; and generate a interpolated image based on the respective warped sets of features for the pair of input images and the multiple different scales.
    Type: Application
    Filed: February 8, 2024
    Publication date: August 8, 2024
    Inventors: Janne Matias Kontkanen, Eric Tabellion, Brian Lee Curless, Fitsum Reda, Deqing Sun, Caroline Rebecca Pantofaru
  • Publication number: 20230186428
    Abstract: Apparatuses, systems, and techniques for texture synthesis from small input textures in images using convolutional neural networks. In at least one embodiment, one or more convolutional layers are used in conjunction with one or more transposed convolution operations to generate a large textured output image from a small input textured image while preserving global features and texture, according to various novel techniques described herein.
    Type: Application
    Filed: February 6, 2023
    Publication date: June 15, 2023
    Inventors: Guilin Liu, Andrew Tao, Bryan Christopher Catanzaro, Ting-Chun Wang, Zhiding Yu, Shiqiu Liu, Fitsum Reda, Karan Sapra, Brandon Rowlett
  • Publication number: 20220038654
    Abstract: Apparatuses, systems, and techniques to generate interpolated video frames. In at least one embodiment, an interpolated video frame is generated based, at least in part, on a first set of pixel data sampled from a first video frame, and a second set of pixel data sampled from a second video frame based, at least in part, on a set of forward pointing motion vectors from the first video frame to the second video frame.
    Type: Application
    Filed: July 30, 2020
    Publication date: February 3, 2022
    Inventors: Fitsum Reda, Karan Sapra, Robert Thomas Pottorff, Shiqiu Liu, Andrew Tao, Bryan Christopher Catanzaro
  • Publication number: 20220038653
    Abstract: Apparatuses, systems, and techniques to generate interpolated video frames. In at least one embodiment, an interpolated video frame is generated based, at least in part, on one of a plurality of possible motions of one or more objects from a first video frame to a second video frame.
    Type: Application
    Filed: July 30, 2020
    Publication date: February 3, 2022
    Inventors: Fitsum Reda, Karan Sapra, Robert Thomas Pottorff, Shiqiu Liu, Andrew Tao, Bryan Christopher Catanzaro
  • Publication number: 20210279841
    Abstract: Apparatuses, systems, and techniques for texture synthesis from small input textures in images using convolutional neural networks. In at least one embodiment, one or more convolutional layers are used in conjunction with one or more transposed convolution operations to generate a large textured output image from a small input textured image while preserving global features and texture, according to various novel techniques described herein.
    Type: Application
    Filed: March 9, 2020
    Publication date: September 9, 2021
    Inventors: Guilin Liu, Andrew Tao, Bryan Christopher Catanzaro, Ting-Chun Wang, Zhiding Yu, Shiqiu Liu, Fitsum Reda, Karan Sapra, Brandon Rowlett
  • Publication number: 20210067735
    Abstract: Apparatuses, systems, and techniques to enhance video. In at least one embodiment, one or more neural networks are used to create, from a first video, a second video having a higher frame rate, higher resolution, or reduced number of missing or corrupt video frames.
    Type: Application
    Filed: September 3, 2019
    Publication date: March 4, 2021
    Inventors: Fitsum Reda, Deqing Sun, Aysegul Dundar, Mohammad Shoeybi, Guilin Liu, Kevin Shih, Andrew Tao, Jan Kautz, Bryan Catanzaro