Patents by Inventor Krishnaram Kenthapadi

Krishnaram Kenthapadi has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Patent number: 12670320
    Abstract: A determination is made that an explanatory data set for a common set of predictions generated by a machine learning model for records containing text tokens is to be provided. Respective groups of related tokens are identified from the text attributes of the records, and record-level prediction influence scores are generated for the token groups. An aggregate prediction influence score is generated for at least some of the token groups from the record-level scores, and an explanatory data set based on the aggregate scores is presented.
    Type: Grant
    Filed: March 19, 2024
    Date of Patent: June 30, 2026
    Assignee: Amazon Technologies, Inc.
    Inventors: Cedric Philippe Archambeau, Sanjiv Ranjan Das, Michele Donini, Michaela Hardt, Tyler Stephen Hill, Krishnaram Kenthapadi, Pedro L Larroy, Xinyu Liu, Keerthan Harish Vasist, Pinar Altin Yilmaz, Muhammad Bilal Zafar
  • Publication number: 20260072804
    Abstract: Systems and methods for implementing auditing of large language model-based tools for bias in inferences is disclosed. Individual entries of the dataset of dialogs may be modified to include stereotypical details of particular contexts. These modified records may then be submitted to an automated response generator to produce a set benchmark records. The baseline records and benchmark records may then be analyzed for completeness, accuracy and conciseness with respect to the particular contexts and disparities in precision and recall may be determined using differences in the benchmark and baseline records. The determined disparities may then be used to further train or fine-tune the automated response generator.
    Type: Application
    Filed: August 19, 2025
    Publication date: March 12, 2026
    Inventors: Swetasudha Panda, Naveen Jafer Nizar, Hongyu Cai, Daeja M. Oxendine, Qinlan Shen, Sumana Srivatsa, Krishnaram Kenthapadi
  • Patent number: 12554805
    Abstract: Views may be generated for bias metrics or feature attribution captured in machine learning pipelines. A request to create a view of bias metrics or feature attribution may be received. The bias metrics or feature attribution may have been determined in a machine learning pipeline as part of executing a training job that specified the bias metrics or the feature attribution. A development application may access a data store that stores the bias metrics or the feature attribution determined in the machine learning pipeline. A view based on the bias metrics or feature attribution may be generated and provided.
    Type: Grant
    Filed: November 27, 2020
    Date of Patent: February 17, 2026
    Assignee: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L Larroy
  • Patent number: 12547926
    Abstract: Bias metrics may be captured at different stages for training a machine learning model. A training job may specify bias metrics to capture at multiple different stages of a machine learning pipeline for a feature of a training data set used to train a machine learning model. The training job may be executed and the bias metrics determined at the stages as specified in the training job. The bias metrics for the different stages may be stored.
    Type: Grant
    Filed: November 27, 2020
    Date of Patent: February 10, 2026
    Assignee: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L Larroy
  • Patent number: 12367396
    Abstract: Automatic failure diagnosis and correction may be performed on trained machine learning models. Input data that causes a trained machine learning model may be identified in order to determine different model failures. The model failures may be clustered in order to determine failure scenarios for the trained machine learning model. Examples of the failure scenarios may be generated and truth labels for the example scenarios obtained. The examples and truth labels may then be used to retrain the machine learning model to generate a corrected version of the machine learning model.
    Type: Grant
    Filed: March 29, 2021
    Date of Patent: July 22, 2025
    Assignee: Amazon Technologies, Inc.
    Inventors: Nathalie Rauschmayr, Krishnaram Kenthapadi, Dylan Slack
  • Publication number: 20240311685
    Abstract: In an embodiment, a method includes receiving, via a processor of a first compute device, a representation of a set of inputs and a set of outputs that were generated by inputting the set of inputs into a machine learning (ML) model by a set of compute devices not including the first compute device to generate the set of outputs. The method further includes receiving, via the processor, a request for a machine learning (ML) explanation associated with the ML model and at least one explicand. The method further includes generating, via the processor and without using the ML model, a representation of the ML explanation based on the at least one explicand, the set of inputs, and the set of outputs.
    Type: Application
    Filed: March 16, 2023
    Publication date: September 19, 2024
    Inventors: Kaivalya RAWAL, Amit PAKA, Krishna GADE, Krishnaram KENTHAPADI
  • Patent number: 12080056
    Abstract: Explanation jobs may be performed for computer vision tasks. A request to execute an explanation job for a computer vision machine learning model may be received. The execution job may be performed, including extracting different features from the image, determining the respective relative importance values of the different features on inferences generated by the computer vision machine learning model. The result of the explanation job, including the generated heat maps may be provided.
    Type: Grant
    Filed: November 26, 2021
    Date of Patent: September 3, 2024
    Assignee: Amazon Technologies, Inc.
    Inventors: Ashish Rajendra Rathi, Michele Donini, Tyler Stephen Hill, Krishnaram Kenthapadi, Xinyu Liu, Pinar Altin Yilmaz, Muhammad Bilal Zafar
  • Publication number: 20240232526
    Abstract: A determination is made that an explanatory data set for a common set of predictions generated by a machine learning model for records containing text tokens is to be provided. Respective groups of related tokens are identified from the text attributes of the records, and record-level prediction influence scores are generated for the token groups. An aggregate prediction influence score is generated for at least some of the token groups from the record-level scores, and an explanatory data set based on the aggregate scores is presented.
    Type: Application
    Filed: March 19, 2024
    Publication date: July 11, 2024
    Applicant: Amazon Technologies, Inc.
    Inventors: Cedric Philippe Archambeau, Sanjiv Ranjan Das, Michele Donini, Michaela Hardt, Tyler Stephen Hill, Krishnaram Kenthapadi, Pedro L Larroy, Xinyu Liu, Keerthan Harish Vasist, Pinar Altin Yilmaz, Muhammad Bilal Zafar
  • Patent number: 11977836
    Abstract: A determination is made that an explanatory data set for a common set of predictions generated by a machine learning model for records containing text tokens is to be provided. Respective groups of related tokens are identified from the text attributes of the records, and record-level prediction influence scores are generated for the token groups. An aggregate prediction influence score is generated for at least some of the token groups from the record-level scores, and an explanatory data set based on the aggregate scores is presented.
    Type: Grant
    Filed: November 26, 2021
    Date of Patent: May 7, 2024
    Assignee: Amazon Technologies, Inc.
    Inventors: Cedric Philippe Archambeau, Sanjiv Ranjan Das, Michele Donini, Michaela Hardt, Tyler Stephen Hill, Krishnaram Kenthapadi, Pedro L Larroy, Xinyu Liu, Keerthan Harish Vasist, Pinar Altin Yilmaz, Muhammad Bilal Zafar
  • Patent number: 11841863
    Abstract: An algorithm releases answers to very large numbers of statistical queries, e.g., k-way marginals, subject to differential privacy. The algorithm answers queries on a private dataset using simple perturbation, and then attempts to find a synthetic dataset that most closely matches the noisy answers. The algorithm uses a continuous relaxation of the synthetic dataset domain which makes the projection loss differentiable, and allows the use of efficient machine learning optimization techniques and tooling. Rather than answering all queries up front, the algorithm makes judicious use of a privacy budget by iteratively and adaptively finding queries for which relaxed synthetic data has high error, and then repeating the projection. The algorithm is effective across a range of parameters and datasets, especially when a privacy budget is small or a query class is large.
    Type: Grant
    Filed: September 27, 2022
    Date of Patent: December 12, 2023
    Assignee: Amazon Technologies, Inc.
    Inventors: Sergul Aydore, William Brown, Michael Kearns, Krishnaram Kenthapadi, Luca Melis, Aaron Roth, Amaresh Ankit Siva
  • Patent number: 11836163
    Abstract: A set of clusters from a first set of vector representations (VRs) is identified. A center associated with each cluster from the set of clusters to generate a set of centers is determined. For each VR from the first set of VRs, and to generate a first set of distributions, a distribution of that VR is determined that indicates, for each center from the set of centers, similarity between that VR and that center. For each VR from a second set of VRs, and to generate a second set of distributions, a distribution of that VR is determined that indicates, for each center from the set of centers, similarity between that VR and that center. A set of divergence metrics associated with the first set of VRs and the second set of VRs are computed based on comparing the first set of distributions and the second set of distributions.
    Type: Grant
    Filed: July 25, 2022
    Date of Patent: December 5, 2023
    Assignee: Fiddler Labs, Inc.
    Inventors: Amalendu K. Iyer, Bashir Rastegarpanah, Joshua G. Rubin, Krishnaram Kenthapadi
  • Patent number: 11487765
    Abstract: An algorithm releases answers to very large numbers of statistical queries, e.g., k-way marginals, subject to differential privacy. The algorithm answers queries on a private dataset using simple perturbation, and then attempts to find a synthetic dataset that most closely matches the noisy answers. The algorithm uses a continuous relaxation of the synthetic dataset domain which makes the projection loss differentiable, and allows the use of efficient machine learning optimization techniques and tooling. Rather than answering all queries up front, the algorithm makes judicious use of a privacy budget by iteratively and adaptively finding queries for which relaxed synthetic data has high error, and then repeating the projection. The algorithm is effective across a range of parameters and datasets, especially when a privacy budget is small or a query class is large.
    Type: Grant
    Filed: June 28, 2021
    Date of Patent: November 1, 2022
    Assignee: Amazon Technologies, Inc.
    Inventors: Sergul Aydore, William Brown, Michael Kearns, Krishnaram Kenthapadi, Luca Melis, Aaron Roth, Amaresh Ankit Siva
  • Patent number: 11481659
    Abstract: Hyperparameters for tuning a machine learning system may be optimized for fairness using Bayesian optimization with constraints for accuracy and bias. Hyperparameter optimization may be performed for a received training set and received accuracy and fairness constraints. Respective probabilistic models for accuracy and bias of the machine learning system may be initialized, then hyperparameter optimization may include iteratively identifying respective values for hyperparameters using analysis of the respective models performed using an acquisition function implementing constrained expected improvement on the respective models, training the machine learning system using the identified values to determine measures of accuracy and bias, and updating the respective models using the determined measures.
    Type: Grant
    Filed: June 30, 2020
    Date of Patent: October 25, 2022
    Assignee: Amazon Technologies, Inc.
    Inventors: Valerio Perrone, Michele Donini, Krishnaram Kenthapadi, Cedric Philippe Archambeau
  • Publication number: 20220171991
    Abstract: Views may be generated for bias metrics or feature attribution captured in machine learning pipelines. A request to create a view of bias metrics or feature attribution may be received. The bias metrics or feature attribution may have been determined in a machine learning pipeline as part of executing a training job that specified the bias metrics or the feature attribution. A development application may access a data store that stores the bias metrics or the feature attribution determined in the machine learning pipeline. A view based on the bias metrics or feature attribution may be generated and provided.
    Type: Application
    Filed: November 27, 2020
    Publication date: June 2, 2022
    Applicant: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L Larroy
  • Publication number: 20220172099
    Abstract: Bias metrics may be captured at different stages for training a machine learning model. A training job may specify bias metrics to capture at multiple different stages of a machine learning pipeline for a feature of a training data set used to train a machine learning model. The training job may be executed and the bias metrics determined at the stages as specified in the training job. The bias metrics for the different stages may be stored.
    Type: Application
    Filed: November 27, 2020
    Publication date: June 2, 2022
    Applicant: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L Larroy
  • Publication number: 20220172004
    Abstract: Bias metrics and feature attribution may be monitored for a machine learning model. A request to enable monitoring for bias metrics or feature attribution may be received. Monitoring may be enabled to evaluate respective performance of inferences of a machine learning model according to the enabled bias metrics or feature attribution. If a divergence from reference data is detected, then a notification indicating the divergence may be sent.
    Type: Application
    Filed: November 27, 2020
    Publication date: June 2, 2022
    Applicant: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L. Larroy
  • Publication number: 20220172101
    Abstract: Feature attribution may be captured as part of a machine learning pipeline. A training job may include a request to determine feature attribution as part of a machine learning pipeline that trains a machine learning model from a training data set. A reference data set for determining the feature attribution of the machine learning model may be identified. The feature attribution may be determined based on the reference data set. The feature attribution of the trained machine learning model may be stored.
    Type: Application
    Filed: November 27, 2020
    Publication date: June 2, 2022
    Applicant: Amazon Technologies, Inc.
    Inventors: Sanjiv Das, Michele Donini, Jason Lawrence Gelman, Kevin Haas, Tyler Stephen Hill, Krishnaram Kenthapadi, Pinar Altin Yilmaz, Muhammad Bilal Zafar, Pedro L Larroy
  • Patent number: 11195149
    Abstract: Aspects of the present disclosure relate to cryptography. In particular, example embodiments relate to computing a relationship between private data of a first entity and private data of a second entity, while preserving privacy of the entities and preventing inter-entity data sharing. A server includes a first component to compute an intersection of two datasets, without directly accessing either dataset. The server includes a second component to compute a relationship, such as a regression, between data in the first dataset and data in the second dataset, without directly accessing either dataset.
    Type: Grant
    Filed: May 31, 2016
    Date of Patent: December 7, 2021
    Assignee: Microsoft Technology Licensing, LLC
    Inventors: Krishnaram Kenthapadi, Ryan Wade Sandler
  • Patent number: 11188834
    Abstract: In an example, each of a plurality of members of social networking service is mapped to a weighted skill vector, each weighted skill vector including a list of skills for the member with an associated weight indicating strength of the skill. Members of the social networking service that belong to an industry are aggregated to obtain a weighted matrix of members and skills along with compensation vectors indicating compensation for each of the members in the matrix. The weighted matrix of users and skills and corresponding compensation vectors is used to train a machine learning skill monetary value prediction model to output a predicted monetary value for one or more skills contained in a candidate vector fed to the machine learning skill monetary value prediction model.
    Type: Grant
    Filed: October 31, 2016
    Date of Patent: November 30, 2021
    Assignee: Microsoft Technology Licensing, LLC
    Inventors: Krishnaram Kenthapadi, Stuart MacDonald Ambler
  • Patent number: 11106979
    Abstract: Techniques for implementing a learning semantic representations of sparse entities using unsupervised embeddings are disclosed herein. In some embodiments, a computer system accesses corresponding profile data of users indicating at least one entity of a first facet type associated with the user, and generating a graph data structure comprising nodes and edges based on the accessed profile data, with each node corresponding to a different entity indicated by the accessed profile data, and each edge directly connecting a different pair of nodes and indicating a number of users whose profile data indicates both entities of the pair of nodes. The computer system generating a corresponding embedding vector for the entities based on the graph data structure using an unsupervised machine learning algorithm.
    Type: Grant
    Filed: June 28, 2018
    Date of Patent: August 31, 2021
    Assignee: Microsoft Technology Licensing, LLC
    Inventors: Rohan Ramanath, Gungor Polatkan, Qi Guo, Cagri Ozcaglar, Krishnaram Kenthapadi, Sahin Cem Geyik