NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts

Xin, J; Yun, S; Peng, J; Choi, I; Ballard, J L; Chen, T; Long, Q (May 2025, https://doi.org/10.48550/arXiv.2505.19190)

Modality fusion is a cornerstone of multimodal learning, enabling information integration from diverse data sources. However, vanilla fusion methods are limited by (1) inability to account for heterogeneous interactions between modalities and (2) lack of interpretability in uncovering the multimodal interactions inherent in the data. To this end, we propose I2MoE (Interpretable Multimodal Interaction-aware Mixture of Experts), an end-to-end MoE framework designed to enhance modality fusion by explicitly modeling diverse multimodal interactions, as well as providing interpretation on a local and global level. First, I2MoE utilizes different interaction experts with weakly supervised interaction losses to learn multimodal interactions in a data-driven way. Second, I2MoE deploys a reweighting model that assigns importance scores for the output of each interaction expert, which offers sample-level and dataset-level interpretation. Extensive evaluation of medical and general multimodal datasets shows that I2MoE is flexible enough to be combined with different fusion techniques, consistently improves task performance, and provides interpretation across various real-world scenarios.
more » « less
Free, publicly-accessible full text available May 25, 2026
CAT s are Fuzzy PETs : A Corpus and Analysis of Potentially Euphemistic Terms

Gavidia, M.; Lee, P.; Feldman, A.; Peng, J. (January 2022, arXiv preprint arXiv:2205.02728.)

Euphemisms have not received much attention in natural language processing, despite being an important element of polite and figurative language. Euphemisms prove to be a difficult topic, not only because they are subject to language change, but also because humans may not agree on what is a euphemism and what is not. Nevertheless, the first step to tackling the issue is to collect and analyze examples of euphemisms. We present a corpus of potentially euphemistic terms (PETs) along with example texts from the GloWbE corpus. Additionally, we present a subcorpus of texts where these PETs are not being used euphemistically, which may be useful for future applications. We also discuss the results of multiple analyses run on the corpus. Firstly, we find that sentiment analysis on the euphemistic texts supports that PETs generally decrease negative and offensive sentiment. Secondly, we observe cases of disagreement in an annotation task, where humans are asked to label PETs as euphemistic or not in a subset of our corpus text examples. We attribute the disagreement to a variety of potential reasons, including if the PET was a commonly accepted term (CAT).
more » « less
Full Text Available
Is self-supervised learning more robust than supervised learning?

Zhong, Y.; Tang, H.; Chen, J.; Peng, J.; Wang, Y.-X. (January 2022, Proc ICML Workshop on Pre-training)

Full Text Available
You Don’t Say....Linguistic Features in Sarcasm Detection

Ducret M.; Kruse L.; Martinez C.; Feldman A.; Peng, J. (January 2021, CLIC-IT 2021: Seventh Italian Conference on Computational Linguistics Bologna)

We explore linguistic features that contribute to sarcasm detection. The linguistic features that we investigate are a combination of text and word complexity, stylistic and psychological features. We experiment with sarcastic tweets with and without context. The results of our experiments indicate that contextual information is crucial for sarcasm prediction. One important observation is that sarcastic tweets are typically incongruent with their context in terms of sentiment or emotional load.
more » « less
Full Text Available
Distributed code for semantic relations predicts neural similarity during analogical reasoning

Chiang, J. N.; Peng, J.; Lu, H.; Holyoak, K. J.; Monti, M. M. (March 2021, Journal of cognitive neuroscience)
null (Ed.)
The ability to generate and process semantic relations is central to many aspects of human cognition. Theorists have long debated whether such relations are coarsely coded as links in a semantic network or finely coded as distributed patterns over some core set of abstract relations. The form and content of the conceptual and neural representations of semantic relations are yet to be empirically established. Using sequential presentation of verbal analogies, we compared neural activities in making analogy judgments with predictions derived from alternative computational models of relational dissimilarity to adjudicate among rival accounts of how semantic relations are coded and compared in the brain. We found that a frontoparietal network encodes the three relation types included in the design. A computational model based on semantic relations coded as distributed representations over a pool of abstract relations predicted neural activities for individual relations within the left superior parietal cortex and for second-order comparisons of relations within a broader left-lateralized network.
more » « less
Full Text Available
Pixel contrastive-consistent semi-supervised semantic segmentation

https://doi.org/10.1109/ICCV48922.2021.00718

Zhong, Y.; Yuan, B.; Wu, H.; Yuan, Z.; Peng, J.; Wang, Y.-X. (January 2021, International Conference on Computer Vision)

Full Text Available
Linguistic Fingerprints of Internet Censorship: the Case of Sina Weibo

Ng Kei Y.; Feldman A.; Peng, J (January 2020, Thirty-Fourth AAAI Conference on Artificial Intelligence (AAAI-20))

This paper studies how the linguistic components of blogposts collected from Sina Weibo, a Chinese microblogging platform, might affect the blogposts’ likelihood of being censored. Our results go along with King et al. (2013)’s Collective Action Potential (CAP) theory, which states that a blogpost’s potential of causing riot or assembly in real life is the key determinant of it getting censored. Although there is not a definitive measure of this construct, the linguistic features that we identify as discriminatory go along with the CAP theory. We build a classifier that significantly outperforms non-expert humans in predicting whether a blogpost will be censored. The crowdsourcing results suggest that while humans tend to see censored blogposts as more controversial and more likely to trigger action in real life than the uncensored counterparts, they in general cannot make a better guess than our model when it comes to ‘reading the mind’ of the censors in deciding whether a blogpost should be censored. We do not claim that censorship is only determined by the linguistic features. There are many other factors contributing to censorship decisions. The focus of the present paper is on the linguistic form of blogposts. Our work suggests that it is possible to use linguistic properties of social media posts to automatically predict if they are going to be censored.
more » « less
Full Text Available
Neural Network Prediction of Censorable Language

Ng Kei Y; Feldman A; Peng J., and C. (January 2019, Proceedings of the 3rd Workshop on NLP and Computational Social Science (NLP+CSS) held in conjunction with NAACL 2019)

Internet censorship imposes restrictions on what information can be publicized or viewed on the Internet. According to Freedom House’s annual Freedom on the Net report, more than half the world’s Internet users now live in a place where the Internet is censored or restricted. China has built the world’s most extensive and sophisticated online censorship system. In this paper, we describe a new corpus of censored and uncensored social media tweets from a Chinese microblogging website, Sina Weibo, collected by tracking posts that mention ‘sensitive’ topics or authored by ‘sensitive’ users. We use this corpus to build a neural network classifier to predict censorship. Our model performs with a 88.50% accuracy using only linguistic features. We discuss these features in detail and hypothesize that they could potentially be used for censorship circumvention.
more » « less
Full Text Available
Measurement of flavor asymmetry of the light-quark sea in the proton with Drell-Yan dimuon production in $p + p$ and $p + d$ collisions at 120 GeV

https://doi.org/10.1103/PhysRevC.108.035202

Dove, J.; Kerns, B.; Leung, C.; McClellan, R. E.; Miyasaka, S.; Morton, D. H.; Nagai, K.; Prasad, S.; Sanftl, F.; Scott, M. B.; et al (September 2023, Physical Review C)

Full Text Available
High-Statistics Measurement of Collins and Sivers Asymmetries for Transversely Polarized Deuterons

https://doi.org/10.1103/PhysRevLett.133.101903

Alexeev, G D; Alexeev, M G; Alice, C; Amoroso, A; Andrieux, V; Anosov, V; Asatryan, S; Augsten, K; Augustyniak, W; Azevedo, C_D R; et al (September 2024, Physical Review Letters)

New results are presented on a high-statistics measurement of Collins and Sivers asymmetries of charged hadrons produced in deep inelastic scattering of muons on a transversely polarized ${}^{6}{LiD}$ target. The data were taken in 2022 with the COMPASS spectrometer using the 160 GeV muon beam at CERN, statistically balancing the existing data on transversely polarized proton targets. The first results from about two-thirds of the new data have total uncertainties smaller by up to a factor of three compared to the previous deuteron measurements. Using all the COMPASS proton and deuteron results, both the transversity and the Sivers distribution functions of the $u$ and $d$ quark, as well as the tensor charge in the measured $x$ range are extracted. In particular, the accuracy of the $d$ quark results is significantly improved. Published by the American Physical Society2024
more » « less
Full Text Available

« Prev Next »

Search for: All records