Patents by Inventor Xinwei Li

Xinwei Li has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Patent number: 12689811
    Abstract: A multimedia data processing method and apparatus, a device and a medium, wherein the method includes: receiving text information input by a user; generating multimedia data based on the text information, in response to a processing instruction for the text information, and exhibiting a multimedia edit interface for performing an edition operation on the multimedia data, wherein the multimedia data includes a plurality of multimedia clips, and the multimedia edit interface includes a first edit track, a second edit track and a third edit track, and the first track clip, the second track clip and the third track clip whose timelines are aligned on the edit tracks respectively identify a text clip, a video image clip and a voice clip corresponding thereto.
    Type: Grant
    Filed: September 27, 2023
    Date of Patent: July 21, 2026
    Assignee: Beijing Zitiao Network Technology Co., Ltd.
    Inventor: Xinwei Li
  • Publication number: 20260197489
    Abstract: Methods and systems implement interpolation filters for motion prediction, and more specifically interpolation filters for template matching and decoder-side motion vector refinement (“DMVR”). A VVC and later standard encoder and a VVC and later standard decoder can configure one or more processors of a computing system to signal that at least one of a first filter derived by discrete cosine transform-based interpolation filter, a second filter derived by a cubic-based interpolation filter, and a third filter derived by a cubic-based interpolation filter are applied in a template matching-based mode, applied in DMVR search, or applied in both.
    Type: Application
    Filed: December 30, 2025
    Publication date: July 9, 2026
    Inventors: Zhenyu DAI, Ru-ling Liao, Yan Ye, Xinwei Li, Jie Chen
  • Publication number: 20260181240
    Abstract: A method or device for processing an auxiliary image and a readable storage are provided. The method includes: obtaining image information collected by multiple first image collection devices of an aircraft; splitting the image information into two transmission paths, where the image information of one transmission path is to sense obstacles around the aircraft to enable the aircraft to autonomously avoid obstacles based on the image information, and at least part of the image information of the other transmission path is used as a first image for user viewing and a second image collection direction corresponding to the first image changes with a change in the flight velocity direction; and outputting the first image, when the change in the flight velocity direction is always within the view angle range of the same first image collection device, the first image always comes from the same first image collection device.
    Type: Application
    Filed: February 11, 2026
    Publication date: June 25, 2026
    Applicant: SZ DJI TECHNOLOGY CO., LTD.
    Inventors: Mengfei ZHOU, Qingwen WANG, Shuai CHEN, Xinwei LI, Renzhi DUAN, Junyi PAN, Bowen LI
  • Publication number: 20260181161
    Abstract: The present disclosure provides a method of encoding a video sequence. The method includes receiving a video sequence; encoding the video sequence using an extrapolation filter-based intra prediction (EIP) fusion mode. The encoding the video sequence using the EIP fusion mode includes obtaining a first predictor by an extrapolation filter-based intra prediction (EIP) mode; obtaining a second predictor by an intra prediction mode; generating a fused predictor by fusing the first predictor and the second predictor; and predicting one or more pictures using the fused predictor.
    Type: Application
    Filed: November 17, 2025
    Publication date: June 25, 2026
    Inventors: Jiaxu ZHANG, Zixiang ZHANG, Xinwei LI, Jie CHEN, Ru-ling LIAO, Yan YE
  • Publication number: 20260172592
    Abstract: A VVC and later standard encoder and a VVC and later standard decoder are provided, configuring one or more processors of a computing system to perform motion vector refinement for geometric partitioning for motion prediction, and more specifically application of first-pass and second-pass Decoder-side Motion Vector Refinement (“DMVR”) in GPM motion prediction. The VVC and later standard encoder and decoder implement performing motion prediction upon two partitions of a geometric partitioning mode (GPM)-coded coding block, based on respective motion information of each partition, and applying weighted blending to respective motion predictions of each partition of the GPM-coded coding block based on a blending matrix, wherein the blending matrix depends on a split mode.
    Type: Application
    Filed: November 20, 2025
    Publication date: June 18, 2026
    Inventors: Yuqing YANG, Jiaxu ZHANG, Jie Chen, Xinwei Li, Ru-ling Liao, Yan Ye
  • Publication number: 20260172557
    Abstract: A VVC-standard encoder and a VVC-standard decoder are provided, configuring one or more processors of a computing system to apply additional transform cores for inter coded blocks and apply a decoder side intra mode derivation (“DIMD”) process to inter prediction samples or residual samples to derive an intra prediction angular mode.
    Type: Application
    Filed: November 14, 2025
    Publication date: June 18, 2026
    Inventors: Ru-ling Liao, Yan Ye, Jie Chen, Xinwei Li
  • Publication number: 20260156253
    Abstract: A video decoding method includes: selecting a first intra prediction mode and a second intra prediction mode from a most probable mode (MPM) list; determining a first predictor based on the first intra prediction mode; determining a second predictor based on the second intra prediction mode; blending the first predictor and the second predictor to obtain a blended predictor for intra prediction; and decoding one or more pictures using the blended predictor.
    Type: Application
    Filed: November 3, 2025
    Publication date: June 4, 2026
    Inventors: Xinwei LI, Ru-ling LIAO, Jie CHEN, Yan YE
  • Patent number: 12620416
    Abstract: Embodiments of the present disclosure relate to video editing. A method comprises: displaying a video editing draft which records filming indication information and an image material footage of at least one storyboard shot, wherein an image material footage of a target storyboard shot of the at least one storyboard shot is filmed by a second device; displaying a video editing interface based on the video editing draft; presenting on the editing track an operation indicator, and presenting on a preview player an editing effect produced by applying the editing operation to the image material footage, and generating a target video. In this way, the efficiency of video editing is improved through the cooperation between the first device and the second device.
    Type: Grant
    Filed: November 8, 2023
    Date of Patent: May 5, 2026
    Assignee: Beijing Zitiao Network Technology Co., Ltd.
    Inventor: Xinwei Li
  • Publication number: 20260122274
    Abstract: Disclosed are a transformation method, an encoder, a decoder, and a storage medium. The transformation method applied to an encoder includes: determining a prediction mode parameter of a current block; determining a Matrix-based Intra Prediction (MIP) parameter when the prediction mode parameter indicates that MIP is used for the current block to determine an intra prediction value; determining an intra prediction value of the current block according to the MIP parameter and calculating a residual value between the current block and the intra prediction value; determining a Low-Frequency Non-Separable Transform (LFNST) transform kernel used for the current block according to the MIP parameter when an LFNST is used for the current block, setting an LFNST index, and signalling the LFNST index in a video bitstream; and performing a transform processing on the residual value by using the LFNST transform kernel.
    Type: Application
    Filed: December 27, 2024
    Publication date: April 30, 2026
    Inventors: Junyan HUO, Xinwei LI, Wenhan QIAO, Yanzhuo MA, Shuai WAN, Fuzheng YANG
  • Publication number: 20260113430
    Abstract: A method of decoding a bitstream to output one or more pictures for a video stream includes: decoding a bitstream to construct a merge candidate list including one or more merge candidates; determining whether a first candidate from the merge candidate list is a uni-motion candidate; in response to the first candidate being the uni-motion candidate, determining a bi-motion candidate based on the first candidate and one or more candidate motion vectors; and adding the bi-motion candidate to the merge candidate list.
    Type: Application
    Filed: October 3, 2025
    Publication date: April 23, 2026
    Inventors: Jie CHEN, Ru-ling LIAO, Yan YE, Xinwei LI
  • Publication number: 20260113458
    Abstract: A method of decoding a video bitstream includes: determining a control point motion vector predictor (CPMVP) for an affine coded target block; refining the CPMVP based on a template matching (TM) cost to get a refined CPMVP; deriving the control point motion vector (CPMV) based on the refined CPMVP and a control point motion vector difference (CPMVD); and decoding the affine coded target block based on the CPMV.
    Type: Application
    Filed: September 23, 2025
    Publication date: April 23, 2026
    Inventors: Jie CHEN, Ru-ling LIAO, Xinwei LI, Yan YE
  • Publication number: 20260113478
    Abstract: A method of decoding a bitstream includes: receiving a bitstream; and decoding, using coded information of the bitstream, one or more pictures, by: obtaining a first prediction signal based on a first block vector associated with a target subblock on a boundary of a target block of the one or more pictures; obtaining a second prediction signal based on prediction information from a neighboring subblock of the target subblock; and performing a motion compensation, based on the first prediction signal and the second prediction signal, to predict the target subblock.
    Type: Application
    Filed: September 23, 2025
    Publication date: April 23, 2026
    Inventors: Xinwei LI, Ru-ling LIAO, Jie CHEN, Yan YE
  • Publication number: 20260113461
    Abstract: The present disclosure provides a method of encoding a video sequence. The method includes: receiving a video sequence; encoding the video sequence by determining that an implicit geometric partitioning mode (GPM) is applied to a coding block, a coding block being split into two geometric partitions; and performing an intra and inter prediction on the coding block.
    Type: Application
    Filed: September 19, 2025
    Publication date: April 23, 2026
    Inventors: Ke JIA, Jie CHEN, Xinwei LI, Ru-ling LIAO, Yan YE
  • Publication number: 20260113440
    Abstract: The present disclosure provides a method of encoding a video sequence. The method includes receiving a video sequence; encoding the video sequence by deriving one or more intra modes for intra prediction; performing template matching (TM) tests on a template region using different filters for the one or more intra modes; and selecting one or more intra filters based on a TM cost.
    Type: Application
    Filed: September 16, 2025
    Publication date: April 23, 2026
    Inventors: Zixiang ZHANG, Jie CHEN, Xinwei LI, Ru-ling LIAO, Yan YE
  • Publication number: 20260106995
    Abstract: A method for intra prediction includes: obtaining multiple previously reconstructed neighbouring blocks corresponding to a current processing block; determining prediction modes, that are signalled in a bitstream, corresponding to neighbouring blocks of the multiple previously reconstructed neighbouring blocks, to obtain multiple first prediction modes; if the multiple first prediction modes include at least two directional modes, taking directional modes included in the multiple first prediction modes as first prediction directions; performing, according to a preset operation rule, operation on multiple first prediction directions of the first prediction directions to obtain second prediction directions; obtaining a prediction mode set according to the second prediction directions and the multiple first prediction modes; and performing intra prediction on the current processing block based on the prediction mode set.
    Type: Application
    Filed: November 28, 2025
    Publication date: April 16, 2026
    Inventors: Yanzhuo MA, Junyan HUO, Shuai WAN, Fuzheng YANG, Ze GUO, Xinwei LI
  • Patent number: 12603112
    Abstract: The present disclosure provides a video generation method, an apparatus, a device, a storage medium, and a program product, and the method includes: in response to a first instruction triggered for an input text, generating first video editing data based on the input text, in which the first video editing data includes a first video clip and an audio clip, a first target video clip among the first video clip is a vacant clip; displaying the first video clip and the audio clip on a video editing track of a video editor; in response to triggering a second instruction for the target video clip on the video editor, filling the first target video clip with a target video to obtain second video editing data; generating a first video based on the second video editing data.
    Type: Grant
    Filed: December 20, 2023
    Date of Patent: April 14, 2026
    Assignee: Beijing Zitiao Network Technology Co., Ltd.
    Inventor: Xinwei Li
  • Patent number: 12598290
    Abstract: A video processing method includes: determining whether an inter predictor correction is enabled for a coding block; and when the inter predictor correction is enabled for the coding block, performing the inter predictor correction by: obtaining a plurality of predicted samples from a top boundary and a left boundary of a predicted block corresponding to the coding block; obtaining a plurality of reconstructed samples from top neighboring reconstructed samples and left neighboring reconstructed samples of the coding block; and deriving a corrected predicted block based on the plurality of the predicted samples, the plurality of the reconstructed samples and the predicted block.
    Type: Grant
    Filed: February 14, 2024
    Date of Patent: April 7, 2026
    Assignee: Alibaba (China) Co., Ltd.
    Inventors: Xinwei Li, Jie Chen, Ru-Ling Liao, Yan Ye
  • Publication number: 20260089314
    Abstract: Colour component prediction method is provided, which includes that: first reference sample set corresponding to colour component to be predicted of coding block in video image is acquired; when available sample number in first reference sample set is less than preset number, preset component value is taken as predicted value corresponding to the colour component to be predicted; when available sample number in first reference sample set is not less than preset number, first reference sample set is screened to obtain second reference sample set; when available sample number in second reference sample set is equal to preset number, model parameter is determined through second reference sample set, and prediction model corresponding to colour component to be predicted is obtained based on model parameter, prediction model is used for prediction processing of colour component to be predicted to obtain predicted value corresponding to colour component to be predicted.
    Type: Application
    Filed: November 26, 2025
    Publication date: March 26, 2026
    Inventors: Junyan Huo, Yanzhuo Ma, Shuai Wan, Fuzheng Yang, Xinwei Li, Qihong Ran
  • Patent number: 12586610
    Abstract: Embodiments of the present disclosure relates to a method, apparatus, device, storage medium, and program product for video generation. The method comprises: generating initial multimedia data based on received text data; obtaining a target editing template in response to an editing template obtaining request; applying the editing operation indicated by the target editing template to the initial multimedia data to obtain target multimedia data; and generating target video based on the target multimedia data. Embodiments of the present disclosure generates video by directly applying the editing operation in the obtained editing template to the multimedia data, without the need for users to manually clip the video. This can not only reduce the time cost of video production, but also improve the quality of video production.
    Type: Grant
    Filed: May 9, 2023
    Date of Patent: March 24, 2026
    Assignee: Beijing Zitiao Network Technology Co., Ltd.
    Inventors: Xinwei Li, Jiajin Cao
  • Publication number: 20260032237
    Abstract: An input video or video stream may be obtained or received. The input video or video stream may include a plurality of video frames, and each frame may be divided into a plurality of blocks. A current block of the plurality of blocks may be predicted using a planar mode. Depending on which planar mode is used, different reference samples may be used for predicting a current sample in the current block.
    Type: Application
    Filed: September 30, 2025
    Publication date: January 29, 2026
    Inventors: Xinwei Li, Ru-Ling Liao, Jie Chen, Yan Ye