Steve James
Steve James
Home
Publications
Research Lab
Contact
Composition
Weighted Composition for Entropy-regularised Reinforcement Learning
One avenue for creating a generally intelligent agent is to equip it with the ability to combine its previously learned behaviours to …
Caston Nyabadza
,
Benjamin Rosman
,
Steven James
,
Geraud Nangue Tasse
PDF
Cite
Compositional Instruction Following with Language Models and Reinforcement Learning
Combining reinforcement learning with language grounding is challenging as the agent needs to explore the environment while …
Vanya Cohen
,
Geraud Nangue Tasse
,
Nakul Gopalan
,
Steven James
,
Matthew Gombolay
,
Ray Mooney
,
Benjamin Rosman
PDF
Cite
End-to-End Learning to Follow Language Instructions with Compositional Policies
We develop an end-to-end model for learning to follow language instructions with compositional policies. Our model combines large …
Vanya Cohen
,
Geraud Nangue Tasse
,
Nakul Gopalan
,
Steven James
,
Raymond Mooney
,
Benjamin Rosman
PDF
Cite
Skill Machines: Temporal Logic Composition in Reinforcement Learning
A major challenge in reinforcement learning is specifying tasks in a manner that is both interpretable and verifiable. One common …
Geraud Nangue Tasse
,
Devon Jarvis
,
Steven James
,
Benjamin Rosman
PDF
Cite
Facilitating Safe Sim-to-Real through Simulator Abstraction and Zero-shot Task Composition
Simulators are a fundamental part of training robots to solve complex control and navigation tasks. This is due to the speed and safety …
Tamlin Love
,
Devon Jarvis
,
Geraud Nangue Tasse
,
Branden Ingram
,
Steven James
,
Benjamin Rosman
PDF
Cite
Video
Skill Machines: Temporal Logic Composition in Reinforcement Learning
A major challenge in reinforcement learning is specifying tasks in a manner that is both interpretable and verifiable. One common …
Geraud Nangue Tasse
,
Devon Jarvis
,
Steven James
,
Benjamin Rosman
PDF
Cite
Video
Learning to Follow Language Instructions with Compositional Policies
We propose a framework that learns to execute natural language instructions in an environment consisting of goal-reaching tasks that …
Vanya Cohen
,
Geraud Nangue Tasse
,
Nakul Gopalan
,
Steven James
,
Matthew Gombolay
,
Benjamin Rosman
PDF
Cite
Composing Value Functions in Reinforcement Learning
An important property for lifelong-learning agents is the ability to combine existing skills to solve new unseen tasks. In general, …
Benjamin Van Niekerk
,
Steven James
,
Adam Earle
,
Benjamin Rosman
PDF
Cite
Supplementary Material
Will it Blend? Composing Value Functions in Reinforcement Learning
An important property for lifelong-learning agents is the ability to combine existing skills to solve unseen tasks. In general, …
Benjamin Van Niekerk
,
Steven James
,
Adam Earle
,
Benjamin Rosman
PDF
Cite
Cite
×