Elvis Nava

Publications

Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

Alignment

We show how to use Vision-Language Models as reward models for RL agents. Instead of manually specifying a reward function, we only need to provide text prompts to instruct and provide feedback. We find larger VLMs provide more accurate reward signals, so we expect this method to work even better with future models.

October 18, 2023
Date Range

News

VLM-RM: Specifying Rewards with Natural Language

Alignment

We show how to use Vision-Language Models (VLM), and specifically CLIP models, as reward models (RM) for RL agents.

October 18, 2023
Date Range

Research

Our research explores a portfolio of high-potential agendas.

Events

Our events bring together global leaders in AI.

Programs

Our programs build the field of trustworthy and secure AI