Kyunghyun Cho

Publications

Training Language Models with Language Feedback at Scale

Alignment

We introduce Imitation Learning from Language Feedback (ILF), demonstrate that large language models accurately incorporate natural language feedback and that finetuning with ILF scales well with the dataset size, even outperforming finetuning on human summaries.

March 27, 2023
Date Range

Improving Code Generation by Training with Natural Language Feedback

Alignment

We introduce Imitation Learning from Language Feedback (ILF) to improve code generation, demonstrating that a small amount of natural language feedback during training can lead to significant performance gains on program synthesis benchmarks.

March 27, 2023
Date Range

Training Language Models with Language Feedback

Alignment

We propose a three-step learning algorithm to learn from natural language feedback, which conveys more information per human evaluation than comparisons.

November 16, 2022
Date Range

News

No items found.

Research

Our research explores a portfolio of high-potential agendas.

Events

Our events bring together global leaders in AI.

Programs

Our programs build the field of trustworthy and secure AI