H2020Individual fellowship2019–2021

GLOMODAT · Enhancing data fusion, parallelisation for hydrological modelling and estimating sensitivity to spatial parameterization of SWAT to model nitrogen and phosphorus runoff at local and global scale

Horizon 2020 — Marie Skłodowska-Curie Actions

Duration
2019-09-01 → 2021-08-31
EU contribution
€148,583
Participants
1
Scheme
MSCA-IF-EF-RI

Lines connect the coordinator with its partners.

Results in brief

Enhancing data fusion, parallelisation for hydrological modelling and estimating sensitivity to spatialparameterization of SWAT to model nitrogen and phosphorus runoff at local and global scale

When rain falls, the water runs off over the surface to the streams as well as infiltrates into the soils and takes parts of fertilizer with it. A similar effect happens during the snowmelt in spring, where the meltwater carries nutrients from the fields. This is an example of diffuse water pollution (DWP). The main pollutants we are looking at are categorized into nitrogen- (N) and phosphorous-based (P) nutrients. Too many nutrients can have negative effects on the overall health of plant and animal life in our lakes and streams. The main prerequisite for combatting excessive nutrient losses to water bodies is to study how the pollution sources can be traced and reduced. Compared to point pollution, where polluted water is directly pumped into a river, e.g. from wastewater plants or industrial water use, DWP is more difficult to control due to its numerous and dispersed sources, and the difficulties in tracing its pathways. Spatially distributed hydrological models like the Soil and Water Assessment Tool (SWAT) have been successfully used for these analyses. However, there are challenges: First, the data demand for these models is considerable. Even with the recent advances in standardised data access, i.e. discovery, quality assessment and conversions are major challenges. Up to 50% of research time is still spent on data processing. Secondly, this type of modelling is very computationally expensive. Finally, there is a lack of scientific understanding, if and how these models would change their predictions in relation to different input datasets. Previous studies have shown that SWAT results are impacted by different choices of input data, but tested typically only one type of input data, i.e. exchanging the soil or land cover dataset for another one, or testing different elevation models. But there are no comprehensive studies on all spatial input data types concertedly and in unison, and the effect at the field, catchment, national or even global level. In particular, the use of high-resolution data has been neglected, mainly because of the unavailability of very high-resolution data and/or the very high computational requirements. That is also the reason why to our knowledge nutrient runoff has not been modelled at a global scale. If we could automate those analyses, computers would excel at testing various scenarios and analysing and predicting pollution load. At first, we want to enhance and automate data preparation. Subsequently, to improve large scale and high-resolution SWAT modelling we aim to design and test a computation framework that spreads stages of the model computation onto multiple servers, just like nowadays’ cloud-computing. When the technical basis is ready, we test and estimate the effect of different resolution datasets of climate, topographical, soil and land use inputs on SWAT modelling results of flow and nutrient runoff at the local scale in smaller catchments. Then try to scale up to a global application in order to analyse and predict global nitrogen and phosphorus runoff. This could allow us in the future to more easily create SWAT models at any desired level with reasonable input data and understand its reliability. We estimated runoff, effects on soil and water quality and applied an extensive parameter sensitivity analysis in Estonian catchments. The results confirm the general patterns, that agriculture is an important contributor in many places. However, predicting future nutrient pollution has limited applicability. Scenario testing prooved very helpful and gave a very detailed insight on sub-catchment level of sources and pathways of pollution. And finally, although high-resolution spatial maps from satellite, radar, and other sources are available, it does not mean that all data should be used in modelling.

Data: CORDIS, © European Union

Project objective

A growing economy and population in the world is causing landscape changes and an increasing pressure is put on water resources. Diffuse water pollution is considered to be one of the major problems for water quality in many countries. Modelling has been successfully used to simulate water quality in catchments to better understand the underlying landscape processes. The widely used Soil and Water Assessment Tool (SWAT) is a spatially distributed model that can be used to estimate flow and nutrient transport at a variety of scales.In current published studies typically only one or two parameters of precipitation, DEM, land use or soil properties are used in. The proposed project aims to investigate how spatial resolution of core input datasets of all types (precipitation, DEM, land use and soil) impacts SWAT modelling results and estimate the nutrient runoff on a local and global scale.Sensitivity analysis to all of precipitation, DEM, land use and soil will therefore be tested. The limitation to one or two parameters in current published studies is due to the computational demands. Due to the way the SWAT model is programmed using a tightly coupled Message Passing Interface (MPI) approaches the available computing power needs to accessible within specialised High Performance Computing (HPC) clusters of limited size. Thus, either scale or resolution is typically compromised.As for higher resolution or global scale data the computational effort becomes too large for automated calibration, we aim to develop a novel method to automate data processing and balancing computational load transparently between many computers.In order to surpass these limitations we test the MapReduce framework as a novel method for parallelization. This entails new ways of data management, model data partitioning and spreading the model partition computations transparently over multiple computing nodes fostering a loosely coupled distributed computation paradigm.

Original text from CORDIS.

Participants

  • TARTU ULIKOOL · TartuCoordinatorEstonia

Links

Data: CORDIS, © European Union