An abstract illustration of swirling shapes, meant to denote a futuristic feeling.

Research - Papers

Explore a selection of our published work on a variety of key research challenges in AI.

Parsing with Multilingual BERT, a Small Treebank, and a Small Corpus

Ethan C. ChauLucy H. LinNoah A. Smith

2020

Findings of EMNLP

Pretrained multilingual contextual representations have shown great success, but due to the limits of their pretraining data, their benefits do not apply equally to all language varieties. This…

PlotMachines: Outline-Conditioned Generation with Dynamic Plot State Tracking

Hannah RashkinAsli CelikyilmazYejin ChoiJianfeng Gao

2020

EMNLP

We propose the task of outline-conditioned story generation: given an outline as a set of phrases that describe key characters and events to appear in a story, the task is to generate a coherent…

Plug and Play Autoencoders for Conditional Text Generation

Florian MaiNikolaos PappasI. MonteroNoah A. Smith

2020

EMNLP

Text autoencoders are commonly used for conditional generation tasks such as style transfer. We propose methods which are plug and play, where any pretrained autoencoder can be used, and only…

PowerTransformer: Unsupervised Controllable Revision for Biased Language Correction

Xinyao MaMaarten SapHannah RashkinYejin Choi

2020

EMNLP

Unconscious biases continue to be prevalent in modern text and media, calling for algorithms that can assist writers with bias correction. For example, a female character in a story is often…

QADiscourse - Discourse Relations as QA Pairs: Representation, Crowdsourcing and Baselines

Valentina PyatkinAyal KleinReut TsarfatyIdo Dagan

2020

EMNLP

Discourse relations describe how two propositions relate to one another, and identifying them automatically is an integral part of natural language understanding. However, annotating discourse…

RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models

Samuel GehmanSuchin GururanganMaarten SapNoah A. Smith

2020

Findings of EMNLP

Pretrained neural language models (LMs) are prone to generating racist, sexist, or otherwise toxic language which hinders their safe deployment. We investigate the extent to which pretrained LMs can…

SciSight: Combining faceted navigation and research group detection for COVID-19 exploratory scientific search

Tom HopeJason PortenoyKishore VasanJevin D. West

2020

EMNLP • Demo

The COVID-19 pandemic has sparked unprecedented mobilization of scientists, already generating thousands of new papers that join a litany of previous biomedical work in related areas. This deluge of…

SLEDGE-Z: A Zero-Shot Baseline for COVID-19 Literature Search

S. MacAvaneyArman CohanN. Goharian

2020

EMNLP

With worldwide concerns surrounding the Severe Acute Respiratory Syndrome Coronavirus 2 (SARS-CoV-2), there is a rapidly growing body of literature on the virus. Clinicians, researchers, and…

Social Chemistry 101: Learning to Reason about Social Norms and Moral Norms

Maxwell ForbesJena D. HwangVered ShwartzYejin Choi

2020

EMNLP

Social norms---the unspoken commonsense rules about acceptable social behavior---are crucial in understanding the underlying causes and intents of people's actions in narratives. For example,…

The Multilingual Amazon Reviews Corpus

Phillip KeungY. LuGyorgy SzarvasNoah A. Smith

2020

EMNLP

We present the Multilingual Amazon Reviews Corpus (MARC), a large-scale collection of Amazon reviews for multilingual text classification. The corpus contains reviews in English, Japanese, German,…

Previous712-721Next