Patents by Inventor Shengjin Wang

Shengjin Wang has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Patent number: 12080084
    Abstract: A method and a system for detecting a scene text may include extracting a first feature map for a scene image input based on a convolutional neural network, and delivering the first feature map to a sequential deformation module; obtaining sampled feature maps corresponding to sampling positions by performing iterative sampling for the first feature map, obtaining a second feature map by performing a concatenation operation in deep learning according to a channel dimension for the first feature map and the sampled feature maps; obtaining a third feature map by performing a feature aggregation operation for the second feature map in the channel dimension, and delivering the third feature map to the object detection baseline network; and performing text area candidate box extraction for the third feature map and obtaining a text area prediction result as a scene text detection result through regression fitting.
    Type: Grant
    Filed: August 20, 2021
    Date of Patent: September 3, 2024
    Assignees: TSINGHUA UNIVERSITY, HYUNDAI MOTOR COMPANY, KIA CORPORATION
    Inventors: Liangrui Peng, Shanyu Xiao, Ruijie Yan, Gang Yao, Shengjin Wang, Jaesik Min, Jong Ub Suk
  • Patent number: 11881038
    Abstract: A method of multi-directional scene text recognition based on multi-element attention mechanism include: performing normalization processing for a text row/column image output from an external text detection module by a feature extractor, extracting a feature for the normalized image by using a deep convolutional neural network to acquire an initial feature map, adding a 2-dimensional directional positional encoding P to the initial feature map in order to output a multi-channel feature map, converting the multi-channel feature map output from a feature extractor by an encoder into a hidden representation, and converting the hidden representation output from the encoder into a recognized text by a decoder and using the recognized text as the output result.
    Type: Grant
    Filed: October 15, 2021
    Date of Patent: January 23, 2024
    Assignees: TSINGHUA UNIVERSITY, HYUNDAI MOTOR COMPANY, KIA CORPORATION
    Inventors: Liangrui Peng, Ruijie Yan, Shanyu Xiao, Gang Yao, Shengjin Wang, Jaesik Min, Jong Ub Suk
  • Publication number: 20220121871
    Abstract: A method and a system of multi-directional scene text recognition based on multi-element attention mechanism are provided. The method includes: performing normalization processing for a text row/column image I output from an external text detection module by a feature extractor, extracting a feature for the normalized image by using a deep convolutional neural network to acquire an initial feature map F0, and adding a 2-dimensional directional positional encoding P to an initial feature map F0 in order to output a multi-channel feature map F; converting the multi-channel feature map F output from a feature extractor by an encoder into a hidden representation H; and converting the hidden representation H output from the encoder into a recognized text by a decoder and using the recognized text as the output result.
    Type: Application
    Filed: October 15, 2021
    Publication date: April 21, 2022
    Applicants: Tsinghua University, Hyundai Motor Company, Kia Corporation
    Inventors: Liangrui Peng, Ruijie Yan, Shanyu Xiao, Gang Yao, Shengjin Wang, Jaesik Min, Jong Ub Suk
  • Publication number: 20220058420
    Abstract: A method and a system for detecting a scene text may include extracting a first feature map for a scene image input based on a convolutional neural network, and delivering the first feature map to a sequential deformation module; obtaining sampled feature maps corresponding to sampling positions by performing iterative sampling for the first feature map, obtaining a second feature map by performing a concatenation operation in deep learning according to a channel dimension for the first feature map and the sampled feature maps; obtaining a third feature map by performing a feature aggregation operation for the second feature map in the channel dimension, and delivering the third feature map to the object detection baseline network; and performing text area candidate box extraction for the third feature map and obtaining a text area prediction result as a scene text detection result through regression fitting.
    Type: Application
    Filed: August 20, 2021
    Publication date: February 24, 2022
    Inventors: Liangrui PENG, Shanyu XIAO, Ruijie YAN, Gang YAO, Shengjin WANG, Jaesik MIN, Jong Ub SUK
  • Publication number: 20090031218
    Abstract: A document is divided into a plurality of regions so that each page has a meaning and, for each of the regions, there is generated a similarity table indicating a similarity degree and a data amount corresponding to the multiplexed resolution level data. By referencing this similarity table, each page of the electronic document in the region specified by the user is represented by a predetermined similarity (resolution) and at the reading speed specified by the user. Moreover, by referencing the similarity table, each page of the electronic document is represented at the maximum speed while guaranteeing the predetermined similarity.
    Type: Application
    Filed: September 23, 2008
    Publication date: January 29, 2009
    Applicant: NEC CORPORATION
    Inventor: Shengjin Wang
  • Patent number: 7458015
    Abstract: A document is divided into a plurality of regions so that each page has a meaning and, for each of the regions, there is generated a similarity table indicating a similarity degree and a data amount corresponding to the multiplexed resolution level data. By referencing this similarity table, each page of the electronic document in the region specified by the user is represented by a predetermined similarity (resolution) and at the reading speed specified by the user. Moreover, by referencing the similarity table, each page of the electronic document is represented at the maximum speed while guaranteeing the predetermined similarity.
    Type: Grant
    Filed: December 12, 2002
    Date of Patent: November 25, 2008
    Assignee: NEC Corporation
    Inventor: Shengjin Wang
  • Publication number: 20040268221
    Abstract: A document is divided into a plurality of regions so that each page has a meaning and, for each of the regions, there is generated a similarity table indicating a similarity degree and a data amount corresponding to the multiplexed resolution level data. By referencing this similarity table, each page of the electronic document in the region specified by the user is represented by a predetermined similarity (resolution) and at the reading speed specified by the user. Moreover, by referencing the similarity table, each page of the electronic document is represented at the maximum speed while guaranteeing the predetermined similarity.
    Type: Application
    Filed: August 19, 2004
    Publication date: December 30, 2004
    Inventor: Shengjin Wang
  • Patent number: 6597380
    Abstract: In an in-space viewpoint control device, a viewpoint control means comprises a recommend vector setting means, a viewpoint information determination means, a viewing point order calculation means, a viewpoint pass generating means, and a relevant information visualization control means. The recommend vector setting means sets recommend vectors for the object which reflects the intentions of the content designer. The viewpoint information determination means is determines a viewing direction and a viewpoint position in the space for viewing the object on the basis of the recommend vector information. The viewing point order calculation means is responsive when a plurality of recommend vectors are being set, as it determines a rotation order for each viewing point being determined in accordance with the recommend vector. The viewpoint pass generating means calculates an appropriate shifting pass for the viewing point following the calculated rotation order, such that the viewing point is shifted smoothly.
    Type: Grant
    Filed: March 10, 1999
    Date of Patent: July 22, 2003
    Assignee: NEC Corporation
    Inventors: Shengjin Wang, Kazuo Kunieda
  • Patent number: D1099752
    Type: Grant
    Filed: June 23, 2024
    Date of Patent: October 28, 2025
    Inventor: Shengjin Wang