An abstract illustration of swirling shapes, meant to denote a futuristic feeling.

Research - Papers

Explore a selection of our published work on a variety of key research challenges in AI.

AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents

Harsh TrivediTushar KhotMareike HartmannNiranjan Balasubramanian

2024

ACL

Autonomous agents that address day-to-day digital tasks (e.g., ordering groceries for a household), must not only operate multiple apps (e.g., notes, messaging, shopping app) via APIs, but also…

Can Language Models Serve as Text-Based World Simulators?

Ruoyao WangGraham ToddZiang XiaoP. Jansen

2024

ACL

Virtual environments play a key role in benchmarking advances in complex planning and decision-making tasks but are expensive and complicated to build by hand. Can current language models themselves…

Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning

Zhouhang XieBodhisattwa Prasad MajumderMengjie ZhaoJulian McAuley

2024

ACL Findings

We consider the task of building a dialogue system that can motivate users to adopt positive lifestyle changes: Motivational Interviewing. Addressing such a task requires a system that can infer…

Data Contamination Report from the 2024 CONDA Shared Task

Oscar SainzIker Garc'ia-FerreroAlon JacoviJinglin Yang

2024

arXiv

The 1st Workshop on Data Contamination (CONDA 2024) focuses on all relevant aspects of data contamination in natural language processing, where data contamination is understood as situations where…

The Illusion of State in State-Space Models

William MerrillJackson PettyAshish Sabharwal

2024

ICML

State-space models (SSMs) have emerged as a potential alternative architecture for building large language models (LLMs) compared to the previously ubiquitous transformer architecture. One…

Skill Set Optimization: Reinforcing Language Model Behavior via Transferable Skills

Kolby NottinghamBodhisattwa Prasad MajumderBhavana DalviRoy Fox

2024

ICML

Large language models (LLMs) have recently been used for sequential decision making in interactive environments. However, leveraging environment reward signals for continual LLM actor improvement is…

Data-driven Discovery with Large Generative Models

Bodhisattwa Prasad MajumderHarshit SuranaDhruv AgarwalPeter Clark

2024

ICML

With the accumulation of data at an unprecedented rate, its potential to fuel scientific discovery is growing exponentially. This position paper urges the Machine Learning (ML) community to exploit…

Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos

Tarun KalluriBodhisattwa Prasad MajumderManmohan Chandraker

2024

ICML

We introduce LaGTran, a novel framework that utilizes text supervision to guide robust transfer of discriminative knowledge from labeled source to unlabeled target data with domain gaps. While…

Climate sensitivity and relative humidity changes in global storm-resolving model simulations of climate change

T. MerlisKai-Yuan ChengIlai GuendelmanStephan Fueglistaler

2024

Science Advances

The climate simulation frontier of a global storm-resolving model (GSRM; or k-scale model because of its kilometer-scale horizontal resolution) is deployed for climate change simulations. The…

PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Kuo-Hao ZengZichen ZhangKiana EhsaniLuca Weihs

2024

CoRL

We present PoliFormer (Policy Transformer), an RGB-only indoor navigation agent trained end-to-end with reinforcement learning at scale that generalizes to the real-world without adaptation despite…

Previous61-70Next