Patents by Inventor Tan Yu
Tan Yu has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).
-
Patent number: 12400418Abstract: An image segmentation method includes: obtaining an image to be segmented containing a target object; performing at least one image feature fusion processing on associated feature points based on the image to be segmented, and extracting global feature information during each image feature fusion processing, wherein the associated feature points are at least two feature points having a location association relation; and determining, based on the global feature information extracted, a segmentation mask for the target object.Type: GrantFiled: December 30, 2022Date of Patent: August 26, 2025Assignee: BEIJING BAIDU NETCOM SCIENCE TECHNOLOGY CO., LTD.Inventors: Jingyi Yang, Tan Yu, Mingming Sun, Ping Li
-
Patent number: 12147497Abstract: Current pretrained vision-language models for cross-modal retrieval tasks in English depend upon on the availability of many annotated image-caption datasets for pretraining to have English text. However, the texts are not necessarily in English. Although machine translation (MT) tools may be used to translate text to English, the performance largely relies on MT's quality and may suffer from high latency problems in real-world applications. Embodiments herein address these problems by learning cross-lingual cross-modal representations for matching images and their relevant captions in multiple languages. Embodiments seamlessly combine cross-lingual pretraining objectives and cross-modal pretraining objectives in a unified framework to learn image and text in a joint embedding space from available English image-caption data, monolingual corpus, and parallel corpus. Embodiments are shown to achieve state-of-the-art performance in retrieval tasks on multimodal multilingual image caption datasets.Type: GrantFiled: April 7, 2022Date of Patent: November 19, 2024Inventors: Hongliang Fei, Tan Yu, Ping Li
-
Patent number: 12033078Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: GrantFiled: August 4, 2023Date of Patent: July 9, 2024Assignee: SNAP INC.Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20230376757Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: ApplicationFiled: August 4, 2023Publication date: November 23, 2023Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Patent number: 11755910Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: GrantFiled: August 1, 2022Date of Patent: September 12, 2023Assignee: SNAP INC.Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Patent number: 11704893Abstract: Aspects of the present disclosure involve a system comprising a storage medium storing a program and method for receiving a video comprising a plurality of video segments; selecting a target action sequence that includes a sequence of action phases; receiving features of each of the video segments; computing, based on the received features, for each of the plurality of video segments, a plurality of action phase confidence scores indicating a likelihood that a given video segment includes a given action phase of the sequence of action phases; identifying a set of consecutive video segments of the plurality of video segments that corresponds to the target action sequence, wherein video segments in the set of consecutive video segments are arranged according to the sequence of action phases; and generating a display of the video that includes the set of consecutive video segments and skips other video segments in the video.Type: GrantFiled: September 2, 2021Date of Patent: July 18, 2023Assignee: Snap Inc.Inventors: Zhou Ren, Yuncheng Li, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20230133218Abstract: An image segmentation method includes: obtaining an image to be segmented containing a target object; performing at least one image feature fusion processing on associated feature points based on the image to be segmented, and extracting global feature information during each image feature fusion processing, wherein the associated feature points are at least two feature points having a location association relation; and determining, based on the global feature information extracted, a segmentation mask for the target object.Type: ApplicationFiled: December 30, 2022Publication date: May 4, 2023Applicant: BEIJING BAIDU NETCOM SCIENCE TECHNOLOGY CO., LTD.Inventors: Jingyi Yang, Tan Yu, Mingming Sun, Ping Li
-
Publication number: 20230034794Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: ApplicationFiled: August 1, 2022Publication date: February 2, 2023Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20220383048Abstract: Current pretrained vision-language models for cross-modal retrieval tasks in English depend upon on the availability of many annotated image-caption datasets for pretraining to have English text. However, the texts are not necessarily in English. Although machine translation (MT) tools may be used to translate text to English, the performance largely relies on MT's quality and may suffer from high latency problems in real-world applications. Embodiments herein address these problems by learning cross-lingual cross-modal representations for matching images and their relevant captions in multiple languages. Embodiments seamlessly combine cross-lingual pretraining objectives and cross-modal pretraining objectives in a unified framework to learn image and text in a joint embedding space from available English image-caption data, monolingual corpus, and parallel corpus. Embodiments are shown to achieve state-of-the-art performance in retrieval tasks on multimodal multilingual image caption datasets.Type: ApplicationFiled: April 7, 2022Publication date: December 1, 2022Applicant: Baidu USA LLCInventors: Hongliang FEI, Tan YU, Ping LI
-
Patent number: 11410439Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: GrantFiled: May 8, 2020Date of Patent: August 9, 2022Assignee: Snap Inc.Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20210407548Abstract: Aspects of the present disclosure involve a system comprising a storage medium storing a program and method for receiving a video comprising a plurality of video segments; selecting a target action sequence that includes a sequence of action phases; receiving features of each of the video segments; computing, based on the received features, for each of the plurality of video segments, a plurality of action phase confidence scores indicating a likelihood that a given video segment includes a given action phase of the sequence of action phases; identifying a set of consecutive video segments of the plurality of video segments that corresponds to the target action sequence, wherein video segments in the set of consecutive video segments are arranged according to the sequence of action phases; and generating a display of the video that includes the set of consecutive video segments and skips other video segments in the video.Type: ApplicationFiled: September 2, 2021Publication date: December 30, 2021Inventors: Zhou Ren, Yuncheng Li, Ning Xu, Enxu Yan, Tan Yu
-
Patent number: 11158351Abstract: Aspects of the present disclosure involve a system comprising a storage medium storing a program and method for receiving a video comprising a plurality of video segments; selecting a target action sequence that includes a sequence of action phases; receiving features of each of the video segments; computing, based on the received features, for each of the plurality of video segments, a plurality of action phase confidence scores indicating a likelihood that a given video segment includes a given action phase of the sequence of action phases; identifying a set of consecutive video segments of the plurality of video segments that corresponds to the target action sequence, wherein video segments in the set of consecutive video segments are arranged according to the sequence of action phases; and generating a display of the video that includes the set of consecutive video segments and skips other video segments in the video.Type: GrantFiled: December 20, 2018Date of Patent: October 26, 2021Assignee: Snap Inc.Inventors: Zhou Ren, Yuncheng Li, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20210209155Abstract: Embodiments of the present disclosure disclose a method and apparatus for retrieving a video, a device and a medium, and relate to the field of data processing technology, and particularly to the field of smart retrieval technology. The method may include: determining, according to a query text and a candidate video, a unified space feature of the query text and a unified space feature of the candidate video based on a conversion relationship between a text semantic space and a video semantic space; determining a similarity between the query text and the candidate video according to the unified space feature of the query text and the unified space feature of the candidate video; and selecting a target video from the candidate video according to the similarity, and using the target video as a query result.Type: ApplicationFiled: September 16, 2020Publication date: July 8, 2021Inventors: Yi YANG, Yi LI, Shujing WANG, Jie LIU, Tan YU, Xiaodong CHEN, Lin LIU, Yanfeng ZHU, Ping LI
-
Patent number: 10988735Abstract: Described herein are tissues containing semiconductor nanomaterials. In some embodiments, the tissues include vascular cells, cardiomyocytes, and/or cardiac fibroblasts. The tissue may be scaffold-free. In some embodiments, the tissue includes an electrically conductive network. The tissue may exhibit synchronized electrical signal propagation within the tissue. In some embodiments, the tissue exhibits increased functional assembly of cardiac cells and/or increased cardiac specific functions compared to a cardiac tissue prepared using a conventional tissue culture method. Methods of preparing and using such tissues are also described herein.Type: GrantFiled: January 15, 2016Date of Patent: April 27, 2021Assignees: Clemson University Research Foundation, MUSC Foundation for Research Development, The University of ChicagoInventors: Ying Mei, Tan Yu, Dylan Richards, Donald R. Menick, Bozhi Tian
-
Publication number: 20200356760Abstract: Systems and methods are disclosed for capturing multiple sequences of views of a three-dimensional object using a plurality of virtual cameras. The systems and methods generate aligned sequences from the multiple sequences based on an arrangement of the plurality of virtual cameras in relation to the three-dimensional object. Using a convolutional network, the systems and methods classify the three-dimensional object based on the aligned sequences and identify the three-dimensional object using the classification.Type: ApplicationFiled: May 8, 2020Publication date: November 12, 2020Inventors: Yuncheng Li, Zhou Ren, Ning Xu, Enxu Yan, Tan Yu
-
Publication number: 20170369847Abstract: Provided herein are tissues containing semiconductor nanomaterials and methods of preparing and using the same.Type: ApplicationFiled: January 15, 2016Publication date: December 28, 2017Inventors: Ying Mei, Tan Yu, Dylan Richards, Donald R. Menick, Bozhi Tian