Juan Rocamonde

Research Fellow

Juan Rocamonde was a research fellow at FAR.AI. Juan has a master's degree in Mathematics from Cambridge and a bachelor's degree in Physics from University College London. He has previously conducted research at Cambridge, Stanford and CERN.

Publications

Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

Alignment

We show how to use Vision-Language Models as reward models for RL agents. Instead of manually specifying a reward function, we only need to provide text prompts to instruct and provide feedback. We find larger VLMs provide more accurate reward signals, so we expect this method to work even better with future models.

October 18, 2023
Date Range

imitation: Clean Imitation Learning Implementations

Alignment

We describe a software package called "imitation" which provides PyTorch implementations of several imitation and reward learning algorithms, including three inverse reinforcement learning algorithms, three imitation learning algorithms, and a preference comparison algorithm.

September 21, 2022
Date Range

News

VLM-RM: Specifying Rewards with Natural Language

Alignment

We show how to use Vision-Language Models (VLM), and specifically CLIP models, as reward models (RM) for RL agents.

October 18, 2023
Date Range

Research

Our research explores a portfolio of high-potential agendas.

Events

Our events bring together global leaders in AI.

Programs

Our programs build the field of trustworthy and secure AI