NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Knowledge-guided machine learning can improve carbon cycle quantification in agroecosystems

https://doi.org/10.1038/s41467-023-43860-5

Liu, Licheng; Zhou, Wang; Guan, Kaiyu; Peng, Bin; Xu, Shaoming; Tang, Jinyun; Zhu, Qing; Till, Jessica; Jia, Xiaowei; Jiang, Chongya; et al (December 2024, Nature Communications)

Abstract Accurate and cost-effective quantification of the carbon cycle for agroecosystems at decision-relevant scales is critical to mitigating climate change and ensuring sustainable food production. However, conventional process-based or data-driven modeling approaches alone have large prediction uncertainties due to the complex biogeochemical processes to model and the lack of observations to constrain many key state and flux variables. Here we propose a Knowledge-Guided Machine Learning (KGML) framework that addresses the above challenges by integrating knowledge embedded in a process-based model, high-resolution remote sensing observations, and machine learning (ML) techniques. Using the U.S. Corn Belt as a testbed, we demonstrate that KGML can outperform conventional process-based and black-box ML models in quantifying carbon cycle dynamics. Our high-resolution approach quantitatively reveals 86% more spatial detail of soil organic carbon changes than conventional coarse-resolution approaches. Moreover, we outline a protocol for improving KGML via various paths, which can be generalized to develop hybrid models to better predict complex earth system dynamics.
more » « less
Full Text Available
Physics-Guided Fair Graph Sampling for Water Temperature Prediction in River Networks

https://doi.org/10.1609/aaai.v39i27.35025

He, Erhu; Kutscher, Declan; Xie, Yiqun; Zwart, Jacob; Jiang, Zhe; Yao, Huaxiu; Jia, Xiaowei (April 2025, Proceedings of the AAAI Conference on Artificial Intelligence)

This work introduces a novel graph neural networks (GNNs)-based method to predict stream water temperature and reduce model bias across locations of different income and education levels. Traditional physics-based models often have limited accuracy because they are necessarily approximations of reality. Recently, there has been an increasing interest of using GNNs in modeling complex water dynamics in stream networks. Despite their promise in improving the accuracy, GNNs can bring additional model bias through the aggregation process, where node features are updated by aggregating neighboring nodes. The bias can be especially pronounced when nodes with similar sensitive attributes are frequently connected. We introduce a new method that leverages physical knowledge to represent the node influence in GNNs, and then utilizes physics-based influence to refine the selection and weights over the neighbors. The objective is to facilitate equitable treatment over different sensitive groups in the graph aggregation, which helps reduce spatial bias over locations, especially for those in underprivileged groups. The results on the Delaware River Basin demonstrate the effectiveness of the proposed method in preserving equitable performance across locations in different sensitive groups.
more » « less
Full Text Available
Multi-Scale Graph Learning for Anti-Sparse Downscaling

https://doi.org/10.1609/aaai.v39i27.35014

Fan, Yingda; Yu, Runlong; Barclay, Janet R; Appling, Alison P; Sun, Yiming; Xie, Yiqun; Jia, Xiaowei (April 2025, Proceedings of the AAAI Conference on Artificial Intelligence)

Water temperature can vary substantially even across short distances within the same sub-watershed. Accurate prediction of stream water temperature at fine spatial resolutions (i.e., fine scales, ≤ 1 km) enables precise interventions to maintain water quality and protect aquatic habitats. Although spatiotemporal models have made substantial progress in spatially coarse time series modeling, challenges persist in predicting at fine spatial scales due to the lack of data at that scale. To address the problem of insufficient fine-scale data, we propose a Multi-Scale Graph Learning (MSGL) method. This method employs a multi-task learning framework where coarse-scale graph learning, bolstered by larger datasets, simultaneously enhances fine-scale graph learning. Although existing multi-scale or multi-resolution methods integrate data from different spatial scales, they often overlook the spatial correspondences across graph structures at various scales. To address this, our MSGL introduces an additional learning task, cross-scale interpolation learning, which leverages the hydrological connectedness of stream locations across coarse- and fine-scale graphs to establish cross-scale connections, thereby enhancing overall model performance. Furthermore, we have broken free from the mindset that multi-scale learning is limited to synchronous training by proposing an Asynchronous Multi-Scale Graph Learning method (ASYNC-MSGL). Extensive experiments demonstrate the state-of-the-art performance of our method for anti-sparse downscaling of daily stream temperatures in the Delaware River Basin, USA, highlighting its potential utility for water resources monitoring and management.
more » « less
Full Text Available
Physics-Guided Foundation Model for Scientific Discovery: An Application to Aquatic Science

https://doi.org/10.1609/aaai.v39i27.35078

Yu, Runlong; Qiu, Chonghao; Ladwig, Robert; Hanson, Paul; Xie, Yiqun; Jia, Xiaowei (April 2025, Proceedings of the AAAI Conference on Artificial Intelligence)

Physics-guided machine learning (PGML) has become a prevalent approach in studying scientific systems due to its ability to integrate scientific theories for enhancing machine learning (ML) models. However, most PGML approaches are tailored to isolated and relatively simple tasks, which limits their applicability to complex systems involving multiple interacting processes and numerous influencing features. In this paper, we propose a Physics-Guided Foundation Model (PGFM) that combines pre-trained ML models and physics-based models and leverages their complementary strengths to improve the modeling of multiple coupled processes. To effectively conduct pre-training, we construct a simulated environmental system that encompasses a wide range of influencing features and various simulated variables generated by physics-based models. The model is pre-trained in this system to adaptively select important feature interactions guided by multi-task objectives. We then fine-tune the model for each specific task using true observations, while maintaining consistency with established physical theories, such as the principles of mass and energy conservation. We demonstrate the effectiveness of this methodology in modeling water temperature and dissolved oxygen dynamics in real-world lakes. The proposed PGFM is also broadly applicable to a range of scientific fields where physics-based models are being used.
more » « less
Full Text Available
Accounting for spatial variability with geo-aware random forest: A case study for US major crop mapping

https://doi.org/10.1016/j.rse.2024.114585

Xie, Yiqun; Nhu, Anh N; Song, Xiao-Peng; Jia, Xiaowei; Skakun, Sergii; Li, Haijun; Wang, Zhihao (March 2025, Remote Sensing of Environment)

Full Text Available
SolarCube: An Integrative Benchmark Dataset Harnessing Satellite and In-situ Observations for Large-scale Solar Energy Forecasting

Li, Ruohan; Xie, Yiqun; Jia, Xiaowei; Wang, Dongdong; Li, Yanhua; Zhang, Yingxue; Wang, Zhihao; Li, Zhili (December 2024, NeurIPS)

Full Text Available
Adaptive Process-Guided Learning: An Application in Predicting Lake DO Concentrations

https://doi.org/10.1109/ICDM59182.2024.00065

Yu, Runlong; Qiu, Chonghao; Ladwig, Robert; Hanson, Paul C; Xie, Yiqun; Li, Yanhua; Jia, Xiaowei (December 2024, IEEE)

Full Text Available
LITE: Modeling Environmental Ecosystems with Multimodal Large Language Models

Li, Haoran Li; Liu, Junqi; Wang, Zexian; Luo, Shiyuan; Jia, Xiaowei; Yao, Huaxiu (July 2024, COLM)

Full Text Available
Deep Learning for Spatial Data

https://doi.org/10.1201/9781003406969-16

Chen, Weiye; Xie, Yiqun; Jia, Xiaowei; He, Erhu (May 2024, CRC Press)

Full Text Available
Fairness-Aware Deep Learning in Space

https://doi.org/10.1201/9781003406969-18

He, Erhu; Chen, Weiye; Xie, Yiqun; Jia, Xiaowei (May 2024, CRC Press)

Full Text Available

« Prev Next »

Search for: All records