Patents by Inventor Long Mai
Long Mai has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).
-
Publication number: 20260197473Abstract: Embodiments are disclosed for using a progressive growing variational autoencoder to boost temporal compression. The method may include receiving a request to compress an input video. The method further includes providing the input video to a progressive encoder. The progressive encoder includes a top pipeline and a bottom pipeline. The method further includes generating, by the progressive encoder, a temporally compressed representation of the input video using a first latent space representation determined by the top pipeline and a second latent space representation determined by the bottom pipeline.Type: ApplicationFiled: January 9, 2025Publication date: July 9, 2026Applicant: Adobe Inc.Inventors: Long Mai, Aniruddha Mahapatra, David Bourgin, Feng Liu
-
Publication number: 20260073579Abstract: The present disclosure relates to systems, methods, and non-transitory computer-readable media that leverages a dual-variational autoencoder model. For example, the disclosed systems generate an image embedding from a first frame of a sequence of frames by using a two-dimensional variational autoencoder. Moreover, the disclosed systems generate motion embeddings from motion within a video by using a three-dimensional variational autoencoder. Further, the disclosed systems generate a reconstructed image from the image embedding and a reconstructed video from the motion embeddings and the image embedding. Additionally, the disclosed systems modify parameters of a dual-variational autoencoder model based on a measure of accuracy of the reconstructed image and the reconstructed video.Type: ApplicationFiled: October 29, 2024Publication date: March 12, 2026Inventors: Zhifei Zhang, Jianming Zhang, Feng Liu, Long Mai
-
Publication number: 20250245891Abstract: Embodiments of the present disclosure disclose an image editing method and apparatus, an electronic device, and a storage medium. The method includes: receiving an image to be edited and an editing theme; generating, by using a preset vision-language model, an editing instruction and an editing position corresponding to the editing instruction based on the image to be edited and the editing theme; and editing the image to be edited based on the editing instruction and the editing position, to obtain a target image.Type: ApplicationFiled: January 8, 2025Publication date: July 31, 2025Inventors: Tiancheng SHEN, Jun Hao LIEW, Long MAI, Lu QI, Jiashi FENG
-
Patent number: 12254633Abstract: The present disclosure relates to systems, non-transitory computer-readable media, and methods for training and utilizing scale-diverse segmentation neural networks to analyze digital images at different scales and identify different target objects portrayed in the digital images. For example, in one or more embodiments, the disclosed systems analyze a digital image and corresponding user indicators (e.g., foreground indicators, background indicators, edge indicators, boundary region indicators, and/or voice indicators) at different scales utilizing a scale-diverse segmentation neural network. In particular, the disclosed systems can utilize the scale-diverse segmentation neural network to generate a plurality of semantically meaningful object segmentation outputs. Furthermore, the disclosed systems can provide the plurality of object segmentation outputs for display and selection to improve the efficiency and accuracy of identifying target objects and modifying the digital image.Type: GrantFiled: March 18, 2022Date of Patent: March 18, 2025Assignee: Adobe Inc.Inventors: Scott Cohen, Long Mai, Jun Hao Liew, Brian Price
-
Publication number: 20250078392Abstract: An image generation system is described. The system comprises a neural network model configured to perform a diffusion process to generate a set of multi-view images from a same input prompt. The set of multi-view images have a same subject from different view orientation. The neural network model comprises a self-attention layer configured to relate pixels across the set of multi-view images.Type: ApplicationFiled: August 28, 2023Publication date: March 6, 2025Inventors: Yichun SHI, Peng WANG, Jianglong YE, Long MAI, Xiao YANG, Xiaohui SHEN
-
Patent number: 12147896Abstract: Embodiments of the present invention provide systems, methods, and non-transitory computer storage media for generating an ambient occlusion (AO) map for a 2D image that can be combined with the 2D image to adjust the contrast of the 2D image based on the geometric information in the 2D image. In embodiments, using a trained neural network, an AO map for a 2D image is automatically generated without any predefined 3D scene information. Optimizing the neural network to generate an estimated AO map for a 2D image requires training, testing, and validating the neural network using a synthetic dataset comprised of pairs of images and ground truth AO maps rendered from 3D scenes. By using an estimated AO map to adjust the contrast of a 2D image, the contrast of the image can be adjusted to make the image appear lifelike by modifying the shadows and shading in the image based on the ambient lighting present in the image.Type: GrantFiled: April 6, 2023Date of Patent: November 19, 2024Assignee: Adobe Inc.Inventors: Long Mai, Yannick Hold-Geoffroy, Naoto Inoue, Daichi Ito, Brian Lynn Price
-
Patent number: 11871145Abstract: Embodiments are disclosed for video image interpolation. In some embodiments, video image interpolation includes receiving a pair of input images from a digital video, determining, using a neural network, a plurality of spatially varying kernels each corresponding to a pixel of an output image, convolving a first set of spatially varying kernels with a first input image from the pair of input images and a second set of spatially varying kernels with a second input image from the pair of input images to generate filtered images, and generating the output image by performing kernel normalization on the filtered images.Type: GrantFiled: April 6, 2021Date of Patent: January 9, 2024Assignee: Adobe Inc.Inventors: Simon Niklaus, Oliver Wang, Long Mai
-
Publication number: 20230244940Abstract: Embodiments of the present invention provide systems, methods, and non-transitory computer storage media for generating an ambient occlusion (AO) map for a 2D image that can be combined with the 2D image to adjust the contrast of the 2D image based on the geometric information in the 2D image. In embodiments, using a trained neural network, an AO map for a 2D image is automatically generated without any predefined 3D scene information. Optimizing the neural network to generate an estimated AO map for a 2D image requires training, testing, and validating the neural network using a synthetic dataset comprised of pairs of images and ground truth AO maps rendered from 3D scenes. By using an estimated AO map to adjust the contrast of a 2D image, the contrast of the image can be adjusted to make the image appear lifelike by modifying the shadows and shading in the image based on the ambient lighting present in the image.Type: ApplicationFiled: April 6, 2023Publication date: August 3, 2023Inventors: Long MAI, Yannick Hold-Geoffroy, Naoto Inoue, Daichi Ito, Brian Lynn Price
-
Patent number: 11663467Abstract: Embodiments of the present invention provide systems, methods, and non-transitory computer storage media for generating an ambient occlusion (AO) map for a 2D image that can be combined with the 2D image to adjust the contrast of the 2D image based on the geometric information in the 2D image. In embodiments, using a trained neural network, an AO map for a 2D image is automatically generated without any predefined 3D scene information. Optimizing the neural network to generate an estimated AO map for a 2D image requires training, testing, and validating the neural network using a synthetic dataset comprised of pairs of images and ground truth AO maps rendered from 3D scenes. By using an estimated AO map to adjust the contrast of a 2D image, the contrast of the image can be adjusted to make the image appear lifelike by modifying the shadows and shading in the image based on the ambient lighting present in the image.Type: GrantFiled: November 21, 2019Date of Patent: May 30, 2023Assignee: ADOBE INC.Inventors: Long Mai, Yannick Hold-Geoffroy, Naoto Inoue, Daichi Ito, Brian Lynn Price
-
Patent number: 11645328Abstract: Systems and methods for performing image search are described. An image search method may include generating a feature vector for each of a plurality of stored images using a machine learning model trained using a rotation loss term, receiving a search query comprising a search image with object having an orientation, generating a query feature vector for the search image using the machine learning model, wherein the query feature vector is based at least in part on the orientation, comparing the query feature vector to the feature vector for each of the plurality of stored images, and selecting at least one stored image of the plurality of stored images based on the comparison, wherein the at least one stored image comprises a similar orientation to the orientation of the object in the search image.Type: GrantFiled: March 17, 2020Date of Patent: May 9, 2023Assignee: ADOBE INC.Inventors: Long Mai, Michael Alcorn, Baldo Faieta, Vladimir Kim
-
Patent number: 11468318Abstract: Systems, methods, and computer-readable media for context-aware synthesis for video frame interpolation are provided. A convolutional neural network (ConvNet) may, given two input video or image frames, interpolate a frame temporarily in the middle of the two input frames by combining motion estimation and pixel synthesis into a single step and formulating pixel interpolation as a local convolution over patches in the input images. The ConvNet may estimate a convolution kernel based on a first receptive field patch of a first input image frame and a second receptive field patch of a second input image frame. The ConvNet may then convolve the convolutional kernel over a first pixel patch of the first input image frame and a second pixel patch of the second input image frame to obtain color data of an output pixel of the interpolation frame. Other embodiments may be described and/or claimed.Type: GrantFiled: March 16, 2018Date of Patent: October 11, 2022Assignee: PORTLAND STATE UNIVERSITYInventors: Feng Liu, Simon Niklaus, Long Mai
-
Publication number: 20220321830Abstract: Embodiments are disclosed for video image interpolation. In some embodiments, video image interpolation includes receiving a pair of input images from a digital video, determining, using a neural network, a plurality of spatially varying kernels each corresponding to a pixel of an output image, convolving a first set of spatially varying kernels with a first input image from the pair of input images and a second set of spatially varying kernels with a second input image from the pair of input images to generate filtered images, and generating the output image by performing kernel normalization on the filtered images.Type: ApplicationFiled: April 6, 2021Publication date: October 6, 2022Inventors: Simon NIKLAUS, Oliver WANG, Long MAI
-
Publication number: 20220207745Abstract: The present disclosure relates to systems, non-transitory computer-readable media, and methods for training and utilizing scale-diverse segmentation neural networks to analyze digital images at different scales and identify different target objects portrayed in the digital images. For example, in one or more embodiments, the disclosed systems analyze a digital image and corresponding user indicators (e.g., foreground indicators, background indicators, edge indicators, boundary region indicators, and/or voice indicators) at different scales utilizing a scale-diverse segmentation neural network. In particular, the disclosed systems can utilize the scale-diverse segmentation neural network to generate a plurality of semantically meaningful object segmentation outputs. Furthermore, the disclosed systems can provide the plurality of object segmentation outputs for display and selection to improve the efficiency and accuracy of identifying target objects and modifying the digital image.Type: ApplicationFiled: March 18, 2022Publication date: June 30, 2022Inventors: Scott Cohen, Long Mai, Jun Hao Liew, Brian Price
-
Patent number: 11282208Abstract: The present disclosure relates to systems, non-transitory computer-readable media, and methods for training and utilizing scale-diverse segmentation neural networks to analyze digital images at different scales and identify different target objects portrayed in the digital images. For example, in one or more embodiments, the disclosed systems analyze a digital image and corresponding user indicators (e.g., foreground indicators, background indicators, edge indicators, boundary region indicators, and/or voice indicators) at different scales utilizing a scale-diverse segmentation neural network. In particular, the disclosed systems can utilize the scale-diverse segmentation neural network to generate a plurality of semantically meaningful object segmentation outputs. Furthermore, the disclosed systems can provide the plurality of object segmentation outputs for display and selection to improve the efficiency and accuracy of identifying target objects and modifying the digital image.Type: GrantFiled: December 24, 2018Date of Patent: March 22, 2022Assignee: Adobe Inc.Inventors: Scott Cohen, Long Mai, Jun Hao Liew, Brian Price
-
Publication number: 20210294834Abstract: Systems and methods for performing image search are described. An image search method may include generating a feature vector for each of a plurality of stored images using a machine learning model trained using a rotation loss term, receiving a search query comprising a search image with object having an orientation, generating a query feature vector for the search image using the machine learning model, wherein the query feature vector is based at least in part on the orientation, comparing the query feature vector to the feature vector for each of the plurality of stored images, and selecting at least one stored image of the plurality of stored images based on the comparison, wherein the at least one stored image comprises a similar orientation to the orientation of the object in the search image.Type: ApplicationFiled: March 17, 2020Publication date: September 23, 2021Inventors: Long Mai, Michael Alcorn, Baldo Faieta, Vladimir Kim
-
Publication number: 20210158139Abstract: Embodiments of the present invention provide systems, methods, and non-transitory computer storage media for generating an ambient occlusion (AO) map for a 2D image that can be combined with the 2D image to adjust the contrast of the 2D image based on the geometric information in the 2D image. In embodiments, using a trained neural network, an AO map for a 2D image is automatically generated without any predefined 3D scene information. Optimizing the neural network to generate an estimated AO map for a 2D image requires training, testing, and validating the neural network using a synthetic dataset comprised of pairs of images and ground truth AO maps rendered from 3D scenes. By using an estimated AO map to adjust the contrast of a 2D image, the contrast of the image can be adjusted to make the image appear lifelike by modifying the shadows and shading in the image based on the ambient lighting present in the image.Type: ApplicationFiled: November 21, 2019Publication date: May 27, 2021Inventors: Long MAI, Yannick HOLD-GEOFFROY, Naoto INOUE, Daichi ITO, Brian Lynn PRICE
-
Publication number: 20200202533Abstract: The present disclosure relates to systems, non-transitory computer-readable media, and methods for training and utilizing scale-diverse segmentation neural networks to analyze digital images at different scales and identify different target objects portrayed in the digital images. For example, in one or more embodiments, the disclosed systems analyze a digital image and corresponding user indicators (e.g., foreground indicators, background indicators, edge indicators, boundary region indicators, and/or voice indicators) at different scales utilizing a scale-diverse segmentation neural network. In particular, the disclosed systems can utilize the scale-diverse segmentation neural network to generate a plurality of semantically meaningful object segmentation outputs. Furthermore, the disclosed systems can provide the plurality of object segmentation outputs for display and selection to improve the efficiency and accuracy of identifying target objects and modifying the digital image.Type: ApplicationFiled: December 24, 2018Publication date: June 25, 2020Inventors: Scott Cohen, Long Mai, Jun Hao Liew, Brian Price
-
Publication number: 20200012940Abstract: Systems, methods, and computer-readable media for context-aware synthesis for video frame interpolation are provided. A convolutional neural network (ConvNet) may, given two input video or image frames, interpolate a frame temporarily in the middle of the two input frames by combining motion estimation and pixel synthesis into a single step and formulating pixel interpolation as a local convolution over patches in the input images. The ConvNet may estimate a convolution kernel based on a first receptive field patch of a first input image frame and a second receptive field patch of a second input image frame. The ConvNet may then convolve the convolutional kernel over a first pixel patch of the first input image frame and a second pixel patch of the second input image frame to obtain color data of an output pixel of the interpolation frame. Other embodiments may be described and/or claimed.Type: ApplicationFiled: March 16, 2018Publication date: January 9, 2020Applicant: Portland State UniversityInventors: Feng Liu, Simon Niklaus, Long Mai
-
Patent number: D853529Type: GrantFiled: May 7, 2018Date of Patent: July 9, 2019Inventor: Long Mai