Process-Guided Deep Learning Predictions of Lake Water Temperature

Jordan S. Read, Xiaowei Jia, Jared Willard, Alison P. Appling, Jacob A. Zwart, Samantha K. Oliver, Anuj Karpatne, Gretchen J.A. Hansen, Paul C. Hanson, William Watkins, Michael Steinbach, Vipin Kumar

Research output: Contribution to journalArticlepeer-review

201 Scopus citations

Abstract

The rapid growth of data in water resources has created new opportunities to accelerate knowledge discovery with the use of advanced deep learning tools. Hybrid models that integrate theory with state-of-the art empirical techniques have the potential to improve predictions while remaining true to physical laws. This paper evaluates the Process-Guided Deep Learning (PGDL) hybrid modeling framework with a use-case of predicting depth-specific lake water temperatures. The PGDL model has three primary components: a deep learning model with temporal awareness (long short-term memory recurrence), theory-based feedback (model penalties for violating conversation of energy), and model pretraining to initialize the network with synthetic data (water temperature predictions from a process-based model). In situ water temperatures were used to train the PGDL model, a deep learning (DL) model, and a process-based (PB) model. Model performance was evaluated in various conditions, including when training data were sparse and when predictions were made outside of the range in the training data set. The PGDL model performance (as measured by root-mean-square error (RMSE)) was superior to DL and PB for two detailed study lakes, but only when pretraining data included greater variability than the training period. The PGDL model also performed well when extended to 68 lakes, with a median RMSE of 1.65 °C during the test period (DL: 1.78 °C, PB: 2.03 °C; in a small number of lakes PB or DL models were more accurate). This case-study demonstrates that integrating scientific knowledge into deep learning tools shows promise for improving predictions of many important environmental variables.

Original languageEnglish (US)
Pages (from-to)9173-9190
Number of pages18
JournalWater Resources Research
Volume55
Issue number11
DOIs
StatePublished - Nov 1 2019

Bibliographical note

Funding Information:
See supporting information for data access, extended methods details, and example code. See https://doi.org/10.5066/P9AQPIVD for this study's data release and https://doi.org/10.5281/zenodo.3497495 for the versioned code repository. This research was funded by the Department of the Interior Northeast and North Central Climate Adaptation Science Centers, a Midwest Glacial Lakes Fish Habitat Partnership grant through F&WS, NSF Expedition in Computing Grant 1029711 to the University of Minnesota, a postdoctoral fellowship awarded to J.A.Z. under NSF EAR-PF-1725386, and a seed grant from the Digital Technology Center at the University of Minnesota. Access to computing facilities was provided by the University of Minnesota Supercomputing Institute and USGS Advanced Research Computing, USGS Yeti Supercomputer (https://doi.org/10.5066/F7D798MJ). We thank North Temperate Lakes Long-Term Ecological Research (NSF DEB-1440297) and Global Lake Ecological Observatory Network (NSF 1702991) for modeling discussions and data sharing, and Arka Daw, Randy Hunt, Jeff Sadler, Emily Read, and Mike Fienen for the PGDL discussions and ideas. We thank Luke Winslow, Noah Lottig, Madeline Magee, and along with MN DNR and WI DNR for temperature and bathymetric data, with special thanks to Pete Jacobson, Katie Hein, and Madeline Humphrey for collating thousands of temperature records, and Dave Wolock, the editorial group at WRR, and three anonymous reviewers for input that was used to improve this paper.

Funding Information:
supporting information See for data access, extended methods details, and example code. See https://doi.org/10.5066/P9AQPIVD for this study's data release and https://doi.org/10.5281/zenodo.3497495 for the versioned code repository. This research was funded by the Department of the Interior Northeast and North Central Climate Adaptation Science Centers, a Midwest Glacial Lakes Fish Habitat Partnership grant through F&WS, NSF Expedition in Computing Grant 1029711 to the University of Minnesota, a postdoctoral fellowship awarded to J.A.Z. under NSF EAR‐PF‐1725386, and a seed grant from the Digital Technology Center at the University of Minnesota. Access to computing facilities was provided by the University of Minnesota Supercomputing Institute and USGS Advanced Research Computing, USGS Yeti Supercomputer ( https://doi.org/10.5066/F7D798MJ ). We thank North Temperate Lakes Long‐Term Ecological Research (NSF DEB‐1440297) and Global Lake Ecological Observatory Network (NSF 1702991) for modeling discussions and data sharing, and Arka Daw, Randy Hunt, Jeff Sadler, Emily Read, and Mike Fienen for the PGDL discussions and ideas. We thank Luke Winslow, Noah Lottig, Madeline Magee, and along with MN DNR and WI DNR for temperature and bathymetric data, with special thanks to Pete Jacobson, Katie Hein, and Madeline Humphrey for collating thousands of temperature records, and Dave Wolock, the editorial group at WRR, and three anonymous reviewers for input that was used to improve this paper.

Publisher Copyright:
©2019. The Authors.

Keywords

  • data science
  • deep learning
  • lake modelling
  • process-guided deep learning
  • temperature prediction
  • theory-guided data science

Fingerprint

Dive into the research topics of 'Process-Guided Deep Learning Predictions of Lake Water Temperature'. Together they form a unique fingerprint.

Cite this