ALEXANDRIA, Va., July 16 -- United States Patent no. 12,668,273, issued on June 30, was assigned to DENSO International America Inc (Southfield, Mich.) and The Regents of the University of California (Oakland, Calif.).
"Hierarchical planning through goal-conditioned offline reinforcement learning" was invented by Minglei Huang (Novi, Mich.), Wei Zhan (Berkeley, Calif.), Masayoshi Tomizuka (Berkeley, Calif.), Chen Tang (Berkeley, Calif.) and Jinning Li (Berkeley, Calif.).
According to the abstract* released by the U.S. Patent & Trademark Office: "A method and system for controlling a device includes training a low-level policy to form a trained low-level policy and a low-level value function to form a trained goal conditioned value functio...