Patents by Inventor Dian YU
Dian YU has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).
-
Patent number: 12614041Abstract: A method and apparatus comprising computer code configured to cause a processor or processors to receive a text comprising a plurality of sentences, by a machine learning model, extract a nonverbal message from one of the sentences and add an annotation to the text, the annotation indicating the nonverbal message, and output a version of the text including the annotation.Type: GrantFiled: October 27, 2023Date of Patent: April 28, 2026Assignee: TENCENT AMERICA LLCInventors: Dian Yu, Xiaoyang Wang, Haitao Mi, Dong Yu
-
Publication number: 20260057265Abstract: The present disclosure describes various methods, systems, and storage medium for training a language model to obtain an iterative Nash policy optimized (INPO) language model. The method includes initializing a first language model by a reference language model; for N-th iteration with N starting from 1 to M, generating a plurality of responses using the N-th language model for each prompt in a plurality of prompts, and constructing a preference dataset using a preference oracle based on the plurality of responses for each prompt, wherein the preference dataset comprises a winning response and a losing response; training the N-th language model to obtain a (N+1)-th language model by minimizing a value of an INPO function for all preference dataset for the plurality of prompts, wherein the INPO function comprises an expectation term, a regularization term, and a modification term; and outputting the (M+1)-th language model.Type: ApplicationFiled: August 22, 2024Publication date: February 26, 2026Applicant: Tencent America LLCInventors: Dian YU, Linfeng Song, Haitao Mi, Dong Yu
-
Publication number: 20260044675Abstract: A method and apparatus that identifies one or more characters within a text; determines one or more informative sections within the text, the one or more informative sections providing information regarding a gender of the one or more characters within the text; selects a most informative section from the one or more informative sections; extracts unlabeled instances corresponding to the gender of the one or more characters from the most informative section; iteratively trains a multi-task model using unlabeled corpora, the multi-task model performing both speaker identification and gender identification; and labels the gender of the one or more characters based on the extracted unlabeled instances and the multi-task model.Type: ApplicationFiled: October 16, 2025Publication date: February 12, 2026Applicant: TENCENT AMERICA LLCInventors: Dian Yu, Linfeng Song, Dong Yu
-
Patent number: 12468889Abstract: A method and apparatus that identifies one or more characters within a text; determines one or more informative sections within the text, the one or more informative sections providing information regarding a gender of the one or more characters within the text; selects a most informative section from the one or more informative sections; extracts unlabeled instances corresponding to the gender of the one or more characters from the most informative section; iteratively trains a multi-task model using unlabeled corpora, the multi-task model performing both speaker identification and gender identification; and labels the gender of the one or more characters based on the extracted unlabeled instances and the multi-task model.Type: GrantFiled: February 21, 2023Date of Patent: November 11, 2025Assignee: TENCENT AMERICA LLCInventors: Dian Yu, Linfeng Song, Dong Yu
-
Patent number: 12293154Abstract: A method, computer program, and computer system is provided for identifying a speaker in at text based work. Labeled and unlabeled instances corresponding to one or more speakers are extracted. Pseudo-labels are inferred for the extracted unlabeled instances based on the labeled instances. One or more of the unlabeled instances are labeled based on the inferred pseudo-labels.Type: GrantFiled: March 8, 2024Date of Patent: May 6, 2025Assignee: TENCENT AMERICA LLCInventors: Dian Yu, Dong Yu
-
Publication number: 20250139389Abstract: A method and apparatus comprising computer code configured to cause a processor or processors to receive a text comprising a plurality of sentences, by a machine learning model, extract a nonverbal message from one of the sentences and add an annotation to the text, the annotation indicating the nonverbal message, and output a version of the text including the annotation.Type: ApplicationFiled: October 27, 2023Publication date: May 1, 2025Applicant: TENCENT AMERICA LLCInventors: Dian YU, Xiaoyang Wang, Haitao Mi, Dong Yu
-
Patent number: 12248753Abstract: There is included a method and apparatus comprising computer code configured to cause a processor or processors to perform generating one or more aligned inventories, wherein the one or more aligned inventories are generated using one or more word sense inventories, obtaining a word in a context sentence, determining one or more semantic equivalence scores indicating semantic similarity between the word in the context sentence and each of one or more associated glosses in the one or more aligned inventories using a semantic equivalence recognizer model, and predicting a correct sense of the word in the context sentence based on the determined one or more semantic equivalence scores.Type: GrantFiled: October 22, 2021Date of Patent: March 11, 2025Assignee: TENCENT AMERICA LLCInventors: Wenlin Yao, Xiaoman Pan, Lifeng Jin, Jianshu Chen, Dian Yu, Dong Yu
-
Publication number: 20240281608Abstract: A method and apparatus that identifies one or more characters within a text; determines one or more informative sections within the text, the one or more informative sections providing information regarding a gender of the one or more characters within the text; selects a most informative section from the one or more informative sections; extracts unlabeled instances corresponding to the gender of the one or more characters from the most informative section; iteratively trains a multi-task model using unlabeled corpora, the multi-task model performing both speaker identification and gender identification; and labels the gender of the one or more characters based on the extracted unlabeled instances and the multi-task model.Type: ApplicationFiled: February 21, 2023Publication date: August 22, 2024Applicant: Tencent America LLCInventors: Dian YU, Linfeng SONG, Dong YU
-
Publication number: 20240211694Abstract: A method including: receiving an input comprising natural language texts; selecting, via a knowledge selector, one of a plurality of knowledge categories from an external memory based on a context of the input; retrieving one or more helpful knowledge pieces from the selected knowledge category; augmenting the input using the one or more helpful knowledge pieces; feeding the augmented input into a text-to-text model; and generating an output answer based on the text-to-text model.Type: ApplicationFiled: December 27, 2022Publication date: June 27, 2024Applicant: TENCENT AMERICA LLCInventors: Xiaoman PAN, Wenlin YAO, Hongming ZHANG, Dian YU, Dong YU, Jianshu CHEN
-
Publication number: 20240211689Abstract: A method, computer program, and computer system is provided for identifying a speaker in at text based work. Labeled and unlabeled instances corresponding to one or more speakers are extracted. Pseudo-labels are inferred for the extracted unlabeled instances based on the labeled instances. One or more of the unlabeled instances are labeled based on the inferred pseudo-labels.Type: ApplicationFiled: March 8, 2024Publication date: June 27, 2024Applicant: TENCENT AMERICA LLCInventors: Dian YU, Dong YU
-
Patent number: 12001795Abstract: A method, computer program, and computer system is provided for identifying a speaker in at text based work. Labeled and unlabeled instances corresponding to one or more speakers are extracted. Pseudo-labels are inferred for the extracted unlabeled instances based on the labeled instances. One or more of the unlabeled instances are labeled based on the inferred pseudo-labels.Type: GrantFiled: August 11, 2021Date of Patent: June 4, 2024Assignee: TENCENT AMERICA LLCInventors: Dian Yu, Dong Yu
-
Publication number: 20230132090Abstract: There is included a method and apparatus comprising computer code configured to cause a processor or processors to perform generating one or more aligned inventories, wherein the one or more aligned inventories are generated using one or more word sense inventories, obtaining a word in a context sentence, determining one or more semantic equivalence scores indicating semantic similarity between the word in the context sentence and each of one or more associated glosses in the one or more aligned inventories using a semantic equivalence recognizer model, and predicting a correct sense of the word in the context sentence based on the determined one or more semantic equivalence scores.Type: ApplicationFiled: October 22, 2021Publication date: April 27, 2023Applicant: Tencent America LLCInventors: Wenlin Yao, Xiaoman Pan, Lifeng Jin, Jianshu Chen, Dian Yu, Dong Yu
-
Publication number: 20230053148Abstract: A method, computer program, and computer system is provided for identifying a speaker in at text based work. Labeled and unlabeled instances corresponding to one or more speakers are extracted. Pseudo-labels are inferred for the extracted unlabeled instances based on the labeled instances. One or more of the unlabeled instances are labeled based on the inferred pseudo-labels.Type: ApplicationFiled: August 11, 2021Publication date: February 16, 2023Applicant: TENCENT AMERICA LLCInventors: Dian YU, Dong YU