Patents by Inventor Tim Hertweck

Tim Hertweck has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Patent number: 12444182
    Abstract: Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for selecting actions to be performed by an agent interacting with an environment to accomplish a goal. In one aspect, a method comprises: obtaining an observation characterizing a state of the environment, processing the observation using an embedding model to generate a lower-dimensional embedding of the observation, determining an auxiliary task reward based on a value of a particular dimension of the embedding, determining an overall reward based at least in part on the auxiliary task reward, and determining an update to values of multiple parameters of an action selection neural network based on the overall reward using a reinforcement learning technique.
    Type: Grant
    Filed: July 27, 2021
    Date of Patent: October 14, 2025
    Assignee: GDM Holding LLC
    Inventors: Markus Wulfmeier, Tim Hertweck, Martin Riedmiller
  • Publication number: 20250196347
    Abstract: This specification describes systems and methods, implemented as computer programs on one or more computers in one or more locations, for controlling an agent to perform multiple different tasks in an environment. The described techniques partition the architecture of a controller into a dispatcher that understands the environment and an executor that understands how to control the agent, with a control channel between them that structures the partitioning. This allows implementations of the controller to generalize better.
    Type: Application
    Filed: December 13, 2024
    Publication date: June 19, 2025
    Inventors: Martin Riedmiller, Roland Hafner, Tim Hertweck
  • Publication number: 20230290133
    Abstract: Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for selecting actions to be performed by an agent interacting with an environment to accomplish a goal. In one aspect, a method comprises: obtaining an observation characterizing a state of the environment, processing the observation using an embedding model to generate a lower-dimensional embedding of the observation, determining an auxiliary task reward based on a value of a particular dimension of the embedding, determining an overall reward based at least in part on the auxiliary task reward, and determining an update to values of multiple parameters of an action selection neural network based on the overall reward using a reinforcement learning technique.
    Type: Application
    Filed: July 27, 2021
    Publication date: September 14, 2023
    Inventors: Markus Wulfmeier, Tim Hertweck, Martin Riedmiller