Patents Examined by Yi Yang
  • Patent number: 12725382
    Abstract: A digital human generation method, an electronic device and a storage medium are disclosed. The solution relates to the fields of augmented reality technologies, virtual reality technologies, computer vision technologies, deep learning technologies, or the like, and can be applied to scenarios, such as metaverse, a virtual digital human, or the like. An implementation includes: acquiring a corresponding target object model based on a picture of a to-be-generated digital human; acquiring a corresponding point cloud of a head key feature in the picture from a pre-configured feature library based on the head key feature; and fusing the point cloud of the head key feature in the target object model to obtain a digital human figure.
    Type: Grant
    Filed: June 19, 2024
    Date of Patent: September 1, 2026
    Assignee: BEIJING BAIDU NETCOM SCIENCE TECHNOLOGY CO., LTD.
    Inventors: Lei Wang, Xiaodong Zhang, Shiyan Li
  • Patent number: 12725336
    Abstract: Techniques are described for an autoregressive user input-to-video (e.g., text-to-video or image-to-video) generation. Using these approaches, a video of unlimited length and without temporal inconsistencies is generated based on the user input. In an implementation, the system receives user input, having particular content, to generate a sequence of video frames, having the same content. The system generates an output sequence of video frames, having the same particular content, by iteratively denoising the frames and conditioning the generation based on the user input. The system may additionally integrate initial anchor frame features with the user input when conditioning the generation of the output sequence of frames. The system may additionally condition each denoising iteration of the output sequence of video frames based on the features of the previous sequence of output video frames.
    Type: Grant
    Filed: March 18, 2025
    Date of Patent: September 1, 2026
    Assignee: PICSART, INC.
    Inventors: Levon Khachatryan, Daniil Hayrapetyan, Hayk Poghosyan, Vahram Tadevosyan, Zhangyang Wang, Shant Navasardyan, Roberto Henschel, Humphrey Shi
  • Patent number: 12725226
    Abstract: Provided are an image processing method and apparatus, and an electronic device and a computer-readable storage medium, which relate to the technical field of image processing. The method is applied to a terminal, and the terminal comprises a first photographing apparatus and a second photographing apparatus. The method comprises: collecting a first video image by means of a first photographing apparatus, and displaying the first video image on a screen; and when it is detected that a display object in the first video image meets a pre-set switching condition, switching to collecting a second video image by means of a second photographing apparatus, and displaying the second video image on the screen. By means of the embodiments of the present disclosure, the automatic switching of a photographing apparatus can be realized on the basis of a display object in an image collected by the photographing apparatus.
    Type: Grant
    Filed: August 26, 2021
    Date of Patent: September 1, 2026
    Assignee: Douyin Vision Co., Ltd.
    Inventors: Jinyuan Wu, Yongwen Wu, Haitao Lv
  • Patent number: 12718422
    Abstract: A method for generating a traffic event video includes the following. Map alignment of a set of moving trajectory coordinates corresponding to a moving video with an electronic map is performed, and a set of trajectory map information corresponding to the moving trajectory coordinates is obtained from the electronic map. At least one event map information conforming to an event trajectory model from the set of trajectory map information and a plurality of image frame information of the moving video corresponding to the image frame information of the trajectory map information are obtained, and a plurality of location information of a virtual object is generated according to the event trajectory model. A video segment is extracted from the moving video based on the image frame information, and the virtual object and the video segment are synthesized based on the location information to generate a traffic event video corresponding to the event trajectory model.
    Type: Grant
    Filed: April 20, 2023
    Date of Patent: August 25, 2026
    Assignee: Industrial Technology Research Institute
    Inventors: Hung-Kuo Chu, Cheng-Hua Lin, Sheng-Yao Wang, Chia-Hao Yu
  • Patent number: 12710911
    Abstract: An extended reality (XR) platform may determine one or more parameters for displaying XR content. The XR platform may receive a request to display the XR content from an XR device associated with a user. The XR platform may determine the one or more display parameters based on a context associated with the XR device and/or the user of the XR device.
    Type: Grant
    Filed: March 21, 2023
    Date of Patent: August 18, 2026
    Assignee: Taiwan Semiconductor Manufacturing Company, Ltd.
    Inventors: Tushar Agrawal, Christian Compton, Jeremy R. Fox, Sarbajit K. Rakshit
  • Patent number: 12710821
    Abstract: Methods and/or apparatus provide for virtual sickness reduction by carrying out actions, including: positioning a headband of a head-mounted display to wrap around a user's head; causing a display section, coupled to the headband, to be positioned in front of the user when the head-mounted display is worn by the user, and to display a moving image representing a state viewed from a point of view; and controlling whether or not to rock a rocking section, which is located at a front of the headband, depending on an acceleration status of the point of view in the moving image displayed on the display section.
    Type: Grant
    Filed: May 12, 2023
    Date of Patent: August 18, 2026
    Assignee: SONY INTERACTIVE ENTERTAINMENT INC.
    Inventor: Jun Sakamoto
  • Patent number: 12694671
    Abstract: Various embodiments of the invention pertain to an augmented reality interface for facilitating identification of an arriving vehicle and/or a passenger that improve upon some or all of the above-described deficiencies. According to some embodiments of the invention, a mobile device may be used by a passenger to scan scenery. The mobile device may determine whether and where a requested vehicle is located and display an indicator of the requested vehicle on the mobile device. Similarly, a mobile device may be used by a driver to scan scenery. The mobile device may determine whether and where a passenger is located and display an indicator of the requesting passenger on the mobile device.
    Type: Grant
    Filed: July 7, 2023
    Date of Patent: July 28, 2026
    Assignee: Apple Inc.
    Inventors: Alexander J. O'Connell, Justin M. Strawn, Ryan D. Shelby, Sunny Chan, Tadayasu Sasada, Vincent P. Arroyo
  • Patent number: 12694578
    Abstract: A method, apparatus, non-transitory computer readable medium, and system for data processing include obtaining a text prompt and generating a first intermediate noise state based on the text prompt, retrieving a second intermediate noise state based on the text prompt and the first intermediate noise state, and generating a synthetic image based on the text prompt and the second intermediate noise state.
    Type: Grant
    Filed: April 11, 2024
    Date of Patent: July 28, 2026
    Assignee: ADOBE INC.
    Inventors: Md Mehrab Tanjim, Chen-Yi Lu, Kanak Mahadik, Anup Bandigadi Rao
  • Patent number: 12688642
    Abstract: A technique for performing ray tracing operations is provided. The technique includes, in a first iteration of a ray traversal technique, traversing to an instance node of a bounding volume hierarchy; in a second iteration of the ray traversal technique that is subsequent to the first iteration, transforming a ray based on an instance transform of the instance node to generate a transformed ray; and in the second iteration, performing a ray-box intersection test for box node data of the instance node based on the transformed ray.
    Type: Grant
    Filed: December 14, 2022
    Date of Patent: July 21, 2026
    Assignees: Advanced Micro Devices, Inc., ATI Technologies ULC
    Inventors: David William John Pankratz, David Kirk McAllister, David Ronald Oldcorn, Michael John Livesley, Daniel James Skinner
  • Patent number: 12676127
    Abstract: A control apparatus includes a display having a display area for displaying images, an imager configured to capture an image of a plurality of users around the display area, and a controller configured to raise or lower a position of the display area to a height corresponding to a representative eye height of the plurality of users based on the image captured by the imager.
    Type: Grant
    Filed: April 18, 2024
    Date of Patent: July 7, 2026
    Assignee: TOYOTA JIDOSHA KABUSHIKI KAISHA
    Inventor: Wataru Kaku
  • Patent number: 12669639
    Abstract: An input coupling grating (ICG) for a waveguide-based display comprises an input region that receives light and a plurality of unit cells. Each unit cell includes photonic structures arranged at a pitch. The photonic structures of at least one first unit cell have at least one different structural characteristic than the photonic structures of at least one second unit cell.
    Type: Grant
    Filed: November 22, 2022
    Date of Patent: June 30, 2026
    Assignee: Sony Group Corporation
    Inventors: Kazue Shimizu, Christophe Peroz, Sebastien De Cunsel
  • Patent number: 12657842
    Abstract: An augmented reality method for monitoring an event in space includes acquisition of a plurality of images of space having at least two landmarks by a camera of a portable device. Space being associated with a three-dimensional reference frame and the portable device being associated with a two-dimensional reference frame. A three-dimensional position and orientation of the space in relation to the camera is determined. The instantaneous position, within the reference frame of the space, of a mobile element moving in the space is received. The position of the mobile element in the two-dimensional reference frame is calculated from transformation parameters calculated from the three-dimensional position and orientation of the space in relation to the camera. An overlay at a predetermined distance in relation to the position of the mobile element in the two-dimensional reference frame is displayed on the screen. Also, a portable electronic device implements the method.
    Type: Grant
    Filed: January 6, 2024
    Date of Patent: June 16, 2026
    Assignee: IMMERSIV
    Inventors: Stéphane Guerin, Emmanuelle Roger
  • Patent number: 12646135
    Abstract: A cached cloud rendering system for saving power and rendering time, which reduces motion to photon time. The cached cloud rendering system may utilize a cloud server to render and cache frames requested by a mobile computing device, which distributes the processing workload to a server system and results in increased battery life of the mobile computing device. Distributing the processing workload to the server system provides an additional benefit of reducing workload on the Graphics Processing Unit (GPU) of the mobile computing device.
    Type: Grant
    Filed: April 20, 2022
    Date of Patent: June 2, 2026
    Assignee: Snap Inc.
    Inventors: Edward Lee Kim-Koon, Farid Zare Seisan
  • Patent number: 12639889
    Abstract: A method and an apparatus for processing a 3D scene are presented. Techniques are disclosed for determining light source locations, including tracking a current viewpoint of a camera capturing object(s) in a 3D scene and determining a reference viewpoint relative to the current viewpoint of the camera. According to aspects, a light source location is determined by obtaining a registered map of real cast shadows of the object(s) from an input image captured by the camera, registered with respect to the reference viewpoint. Then, for candidates of light sources, obtaining respective maps of virtual shadows of the object(s) created with respect to the reference viewpoint, and determining the location of the light source based on the candidates of light sources with respective maps of virtual shadows that match the registered map of real cast shadows.
    Type: Grant
    Filed: September 19, 2023
    Date of Patent: May 26, 2026
    Assignee: InterDigital Madison Patent Holdings, SAS
    Inventors: Philippe Robert, Salma Jiddi, Tao Luo
  • Patent number: 12633001
    Abstract: One variation of a method for detecting and visualizing objects within a space includes: retrieving a map of the space annotated with known locations of target anchor objects and regions in the space; accessing a set of images annotated with object types and locations of objects captured by a set of sensor blocks; projecting the set of images onto the map to form a visualization representing objects in the space based on known locations of the set of sensor blocks; isolating a target anchor object at a known location in a region of the visualization; detecting a mutable object at a location in the region; calculating an offset distance between the location and the known location; and, in response to the offset distance exceeding an offset distance threshold, highlighting the mutable object in the visualization as a deviation; and generating a notification to investigate the mutable object in the region.
    Type: Grant
    Filed: September 12, 2023
    Date of Patent: May 19, 2026
    Assignee: VergeSense, Inc.
    Inventors: Kelby Green, Kanav Dhir
  • Patent number: 12626419
    Abstract: To present augmented reality features without localizing a user, a client device receives a request for presenting augmented reality features in a camera view of a computing device of the user. Prior to localizing the user, the client device obtains sensor data indicative of a pose of the user, and determines the pose of the user based on the sensor data with a confidence level that exceeds a confidence threshold which indicates a low accuracy state. Then the client device presents one or more augmented reality features in the camera view in accordance with the determined pose of the user while in the low accuracy state.
    Type: Grant
    Filed: February 6, 2024
    Date of Patent: May 12, 2026
    Assignee: Google LLC
    Inventors: Mohamed Suhail Mohamed Yousuf Sait, Andre Le, Juan David Hincapie, Mirko Ranieri, Marek Gorecki, Wenli Zhao, Tony Shih, Bo Zhang, Alan Sheridan, Matt Seegmiller
  • Patent number: 12620141
    Abstract: Embodiments of this application disclose an image style conversion method performed by an electronic device. The method includes: performing quality enhancement on a first target style image to obtain a second target style image; performing feature extraction on the second target style image to obtain a target style feature; performing migration training on a preset target style conversion model by using a full style conversion model and the target style feature to obtain a target style conversion model; inputting a full style feature, the target style feature, and a to-be-converted image into the target style conversion model, and performing style conversion on the to-be-converted image using the target style conversion model to obtain a target image conforming to a target style.
    Type: Grant
    Filed: March 29, 2023
    Date of Patent: May 5, 2026
    Assignee: TENCENT TECHNOLOGY (SHENZHEN) COMPANY LIMITED
    Inventors: Yun Cao, Xinyi Zhang, Junwei Zhu, Ying Tai, Mu Zhang, Chengjie Wang, Feiyue Huang
  • Patent number: 12608852
    Abstract: The present disclosure relates to systems, non-transitory computer-readable media, and methods for processing multimodal content to generate summaries or responses using a multimodal large language model. In one or more embodiments, the disclosed systems the disclosed systems utilize the multimodal large language model to generate various types of synthesized responses corresponding to multimodal content items that contain data and information within images. For example, in some embodiments, in response to receiving a request to generate a synthesized response corresponding to a multimodal content item, the disclosed systems employ preprocessing pipelines that generate thumbnail images from the multimodal content item and use the thumbnail images to generate a data structure for a prompt for the multimodal large language model.
    Type: Grant
    Filed: December 9, 2024
    Date of Patent: April 21, 2026
    Assignee: Dropbox, Inc.
    Inventors: Dongjie Chen, Dhruvil Gala
  • Patent number: 12586304
    Abstract: An information processing system comprises processing circuitry configured to display a first user interface element on a screen of a virtual space, the first user interface element for starting an application browser; start the application browser in a case that an operation is performed on the first user interface element by a user; display a second user interface element which shares information displayed by the application browser; send a sharing request to a server in a case that an operation is performed on the second user interface element by the user, the sharing request being a request to share information displayed by the application browser with a device used by a different user viewing the screen of the virtual space; and display, as shared information, information displayed by the application browser on a display object disposed in the virtual space.
    Type: Grant
    Filed: June 29, 2023
    Date of Patent: March 24, 2026
    Assignee: GREE HOLDINGS, INC.
    Inventor: Yosuke Kanaya
  • Patent number: 12567129
    Abstract: An image processing method includes an electronic device configured to perform style migration processing on a first image sequence based on a target migration style by using a fused style migration model into which a plurality of single-style migration models is fused in order to obtain a second image sequence. A style of a 1st frame of image to a style of a last frame of image in the second image sequence change in a first style order in styles of output images of the plurality of single-style migration models. The first image sequence may be from a video shot by using the electronic device. The electronic device may save a plurality of frames of images in the second image sequence as a video. The video may present an effect of rapid time lapse during play.
    Type: Grant
    Filed: December 3, 2021
    Date of Patent: March 3, 2026
    Assignee: HUAWEI TECHNOLOGIES CO., LTD.
    Inventors: Wendong Chen, Shuai Chen, Meng Liu