NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

The Role of Abstract Representations and Observed Preferences in the Ordering of Binomials in Large Language Models

https://doi.org/10.18653/v1/2025.acl-short.55

Houghton, Zachary Nicholas; Sagae, Kenji; Morgan, Emily (January 2025, Association for Computational Linguistics)

Full Text Available
Automatically Exposing Problems with Neural Dialog Models

https://doi.org/10.18653/v1/2021.emnlp-main.37

Yu, Dian; Sagae, Kenji (January 2021, Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing)

Neural dialog models are known to suffer from problems such as generating unsafe and inconsistent responses. Even though these problems are crucial and prevalent, they are mostly manually identified by model designers through interactions. Recently, some research instructs crowdworkers to goad the bots into triggering such problems. However, humans leverage superficial clues such as hate speech, while leaving systematic problems undercover. In this paper, we propose two methods including reinforcement learning to automatically trigger a dialog model into generating problematic responses. We show the effect of our methods in exposing safety and contradiction issues with state-of-the-art dialog models.
more » « less
Full Text Available
Beyond NVD: Cybersecurity meets the Semantic Web.

https://doi.org/10.1145/3498891.3501259

Aranovich, Raúl; Wu, Muting; Yu, Dian; Katsy, Katya; Ahmadnia, Benyamin; Bishop, Matthew; Filkov, Vladimir; Sagae, Kenji (October 2021, NSPW '21: New Security Paradigms Workshop)

Full Text Available
Attribute Alignment: Controlling Text Generation from Pre-trained Language Models

https://doi.org/10.18653/v1/2021.findings-emnlp.194

Yu, Dian; Yu, Zhou; Sagae, Kenji (January 2021, Findings of the Association for Computational Linguistics: EMNLP 2021)

Large language models benefit from training with a large amount of unlabeled text, which gives them increasingly fluent and diverse generation capabilities. However, using these models for text generation that takes into account target attributes, such as sentiment polarity or specific topics, remains a challenge. We propose a simple and flexible method for controlling text generation by aligning disentangled attribute representations. In contrast to recent efforts on training a discriminator to perturb the token level distribution for an attribute, we use the same data to learn an alignment function to guide the pre-trained, non-controlled language model to generate texts with the target attribute without changing the original language model parameters. We evaluate our method on sentiment- and topic-controlled generation, and show large performance gains over previous methods while retaining fluency and diversity.
more » « less
Full Text Available
Language Embeddings for Typology and Cross-lingual Transfer Learning

https://doi.org/10.18653/v1/2021.acl-long.560

Yu, Dian; He, Taiqi; Sagae, Kenji (January 2021, Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers))
null (Ed.)
Cross-lingual language tasks typically require a substantial amount of annotated data or parallel translation data. We explore whether language representations that capture relationships among languages can be learned and subsequently leveraged in cross-lingual tasks without the use of parallel data. We generate dense embeddings for 29 languages using a denoising autoencoder, and evaluate the embeddings using the World Atlas of Language Structures (WALS) and two extrinsic tasks in a zero-shot setting: cross-lingual dependency parsing and cross-lingual natural language inference.
more » « less
Full Text Available

Search for: All records