AI-Driven Phosphorus Monitoring in Lakes
AI-Driven Phosphorus Monitoring in Lakes
College of Engineering, University of Guelph, 50 Stone Road East, Guelph, ON N1G 2W1, Canada;
ydeng09@[Link] (Y.D.); daiwei@[Link] (D.P.)
* Correspondence: syang@[Link] (S.X.Y.); bgharaba@[Link] (B.G.)
Abstract
Accurate estimation of Total Phosphorus, referred to as “Phosphorus, Total” (PPUT; µg/L)
in the sourced monitoring data, is essential for understanding eutrophication dynam-
ics and guiding water-quality management in inland lakes. However, lake-wide PPUT
mapping at high resolution is challenging to achieve using conventional in-situ sampling,
and nearshore gradients are often poorly resolved by medium- or low-resolution satel-
lite sensors. This study exploits multi-generation PlanetScope imagery (Dove Classic,
Dove-R, and SuperDove; 3–5 m, near-daily revisit) to develop a hybrid AI framework
for PPUT retrieval in Lake Simcoe, Ontario, Canada. PlanetScope surface reflectance,
short-term meteorological descriptors (3 to 7-day aggregates of air temperature, wind
speed, precipitation, and sea-level pressure), and in-situ Secchi depth (SSD) were used
to train five ensemble-learning models (HistGradientBoosting, CatBoost, RandomForest,
ExtraTrees, and GradientBoosting) across eight feature-group regimes that progressively
extend from bands-only, to combinations with spectral indices and day-of-year (DOY), and
finally to SSD-inclusive full-feature configurations. The inclusion of SSD led to a strong
and systematic performance gain, with mean R2 increasing from about 0.67 (SSD-free)
to 0.94 (SSD-aware), confirming that vertically integrated optical clarity is the dominant
constraint on PPUT retrieval and cannot be reconstructed from surface reflectance alone. To
enable scalable SSD-free monitoring, a knowledge-distillation strategy was implemented
in which an SSD-aware teacher transfers its learned representation to a student using only
satellite and meteorological inputs. The optimal student model, based on a compact subset
of 40 predictors, achieved R2 = 0.83, RMSE = 9.82 µg/L, and MAE = 5.41 µg/L, retaining
approximately 88% of the teacher’s explanatory power. Application of the student model
to PlanetScope scenes from 2020 to 2025 produces meter-scale PPUT maps; a 26 July 2024
case study shows that >97% of the lake surface remains below 10 µg/L, while rare (<1%)
but coherent hotspots above 20 µg/L align with tributary mouths and narrow channels.
The results demonstrate that combining commercial high-resolution imagery with physics-
informed feature engineering and knowledge transfer enables scalable and operationally
Academic Editor: Yuanrong Zhu relevant monitoring of lake phosphorus dynamics. These high-resolution PPUT maps
Received: 17 December 2025 enable lake managers to identify nearshore nutrient hotspots, tributary plume structures.
Revised: 10 January 2026 In doing so, the proposed framework supports targeted field sampling, early warning for
Accepted: 14 January 2026 eutrophication events, and more robust, lake-wide nutrient budgeting.
Published: 19 January 2026
Copyright: © 2026 by the authors. Keywords: PlanetScope; Lake Simcoe; total phosphorus; knowledge distillation; machine
Licensee MDPI, Basel, Switzerland.
learning; deep learning; remote sensing; water quality retrieval
This article is an open access article
distributed under the terms and
conditions of the Creative Commons
Attribution (CC BY) license.
1. Introduction
Eutrophication of inland lakes remains one of the most widespread environmental
challenges worldwide, driven primarily by excessive phosphorus and nitrogen inputs that
stimulate algal blooms, reduce water transparency, and degrade aquatic ecosystems [1–3].
Accurate and spatially continuous estimation of total phosphorus is essential for under-
standing nutrient dynamics and supporting effective lake management.
However, obtaining lake-wide distributions of Total Phosphorus (PPUT; µg/L)
through conventional in-situ sampling remains challenging due to the sparse spatial and
temporal coverage of monitoring stations [4,5]. Furthermore, nearshore phosphorus con-
centrations are difficult to accurately retrieve from medium- or low-resolution satellite
imagery, where pixel mixing and land-water adjacency effects severely distort spectral
signals [6–8].
Remote sensing provides an effective means to complement in-situ observations by
enabling repetitive, synoptic coverage of large water bodies [9,10]. Many studies have
used satellite optical reflectance and derived indices to estimate water-quality parameters
such as chlorophyll-a, turbidity, suspended sediments and, to some extent, phosphorus or
nitrogen [11–13]. However, phosphorus is a non-optically active constituent whose spectral
expression is indirect, primarily mediated by its relationship with other optically detectable
variables such as phytoplankton and turbidity [14–16]. This indirect linkage complicates
retrieval: spectral signals are weak, and model relationships are often site-specific [17,18].
Another constraint is spatial resolution: many water-quality retrieval studies rely on
sensors such as Sentinel-2 MSI (≈10 m resolution) or Landsat 8 OLI (≈30 m resolution).
These resolutions may fail to resolve narrow bays, tributary plumes or near-shore mix-
ing zones, and mixed-pixel effects can degrade retrieval accuracy in heterogeneous lake
zones [19,20]. In this context, high spatial resolution and high temporal revisit satellite data
become highly desirable for fine-scale inland lake nutrient mapping.
The satellite constellation PlanetScope (Dove Classic, Dove-R, SuperDove) offers dis-
tinct advantages: spatial resolution of ~3–5 m, near-daily revisit frequency, and continuity
across multi-generation sensors, enabling high-resolution and frequent monitoring of
lake surfaces and fine-scale features such as shoreline and tributary plumes [21–23]. In
fact, comparative work has shown that PlanetScope imagery outperforms coarser sensors
(Sentinel-2, Landsat-8) in retrieving specific water-quality parameters in optically complex
inland waters [24,25]. Consequently, PlanetScope is especially suited for whole-lake PPUT
retrieval and capturing near-shore nutrient heterogeneity, offering improved spatial detail
and repeat coverage.
On the modelling front, machine-learning (ML) and deep-learning (DL) approaches
have seen increased adoption in water-quality retrieval thanks to their ability to model non-
linear relationships between spectral, environmental and in-situ data [21,26,27]. Algorithms
such as Random Forest (RF), Light Gradient Boosting Machine (LightGBM), Support Vector
Regression (SVR), and histogram-based gradient boosting (HGBR) have been successfully
applied to estimate parameters such as Chl-a, TSS, and even Total Nitrogen (TN)/Total
Phosphorus (TP) in inland waters [28–30].
More recently, deep learning models including Multilayer Perceptron (MLP), one-
dimensional Convolutional Neural Networks (1D-CNN), and Transformer-based archi-
tectures have shown additional potential for handling high-dimensional inputs, temporal
sequences, and spatial patterns [31–34].
However, the performance of such models often degrades when key auxiliary in-situ
variables, such as Secchi depth (SSD), which reflect optical clarity and vertical attenuation,
are unavailable. These auxiliary variables are typically measured at only a limited number
of monitoring stations and thus cannot support continuous spatial mapping across entire
[Link]
Water 2026, 18, 261 3 of 35
lake surfaces. Developing models that maintain high predictive accuracy without reliance
on in-situ auxiliary features, therefore, remains a significant challenge for whole-lake
nutrient retrieval [35–37].
To address this limitation, knowledge distillation (KD) offers a promising solution. In
KD, a high-capacity “teacher” model trained on full-feature information (including auxiliary
in-situ data) transfers knowledge to a “student” model that uses only readily available
features (e.g., remote sensing + meteorology) and can achieve comparable accuracy [38].
KD has been applied in remote-sensing domains for image classification, segmentation
and multi-task learning [39], although its application to inland water-quality retrieval
remains scarce.
Recent advances in machine learning have substantially improved the retrieval of opti-
cally inactive parameters such as TP from multispectral imagery. Xiong et al. [5] developed
machine-learning algorithms for TP in eutrophic Lake Taihu using Landsat-8, demon-
strating that random forest and other non-linear models can outperform conventional
band-ratio approaches, with typical R2 values around 0.70 for TP estimation. Similarly, Cui
et al. [40] and Qin et al. [41] used Sentinel-2 imagery combined with tree-based models
such as random forest and XGBoost to retrieve TP and other nutrients in large Chinese
lakes, reporting R2 in the range of 0.65–0.80 and highlighting the importance of integrating
spectral indices and hydrometeorological variables.
While these studies confirm the feasibility of satellite-based TP retrieval at 10–30 m spa-
tial resolution, they provide limited insight into nearshore gradients and tributary plumes,
and they do not explicitly address how to leverage auxiliary clarity measurements such
as SSD within a transferable, SSD-free prediction framework. The present study extends
this line of work by exploiting 3 to 5 m PlanetScope imagery and a knowledge-distillation
strategy to transfer information from SSD-aware models to operational SSD-free students.
Although some studies have used machine-learning or deep-learning for TP retrieval
in inland waters [42–46], very few have leveraged high spatial–temporal PlanetScope im-
agery for full-lake PPUT mapping, combined systematic comparison of multiple algorithms
under varying feature sets (remote sensing only; remote sensing and meteorology; remote
sensing, meteorology, and SSD), and then applied a KD framework to enable SSD-free
full-lake prediction. Therefore, this research aims to undertake the following:
1. Utilize PlanetScope multi-generation imagery (Dove Classic, Dove-R, SuperDove)
to achieve high-resolution, high-frequency retrieval of PPUT across the entire lake
surface (including near-shore and tributary zones).
2. Systematically compare multiple machine learning algorithms (HistGBM, CatBoost,
RandomForest, ExtraTrees, and GradientBoosting) under 8 distinct feature settings:
combinations of remote sensing, INDICES, DOY, meteorological variables and SSD.
3. Quantify performance degradation when SSD is removed and implement a KD-based
framework to transfer knowledge from the full-feature teacher model to a reduced-
feature student model, thereby enabling full-lake mapping without SSD.
4. Apply the optimal student model to generate whole-lake PPUT distribution maps for
spring (March–April) and summer (July–August) periods and analyze spatial and
seasonal variability of phosphorus in the study lake.
By integrating PlanetScope’s high spatial–temporal resolution, advanced AI modelling,
and knowledge distillation strategies, this work seeks to advance the feasibility of fine-
scale, lake-wide nutrient retrieval and to support near-real-time eutrophication monitoring
and management.
[Link]
Water 2026, 18, 261 4 of 35
2. Methods
This section describes the general methodological framework developed for high-
resolution retrieval of PPUT from PlanetScope imagery, and its application to Lake Simcoe
as a representative case study. We first outline the overall framework, including data
sources, preprocessing, feature construction, model configurations, and evaluation design.
We then detail how this framework is applied to Lake Simcoe—covering the study area
characteristics, available monitoring networks, PlanetScope acquisition strategy, and the
specific experimental regimes used to train teacher and student models for lake-wide
PPUT mapping.
Table 1. Definition of the eight feature-group regimes used for benchmarking and KD training across
Groups A/B and Stages 1–4.
Group A Group B
Stage 1 x1A = [xSSD , xRS ] x1B = [xRS ]
Stage 2 x2A = [xSSD , xRS , xDOY ] x2B = [xRS , xDOY ]
Stage 3 x3A = [xSSD , xRS , xDOY , xIDX ] x3B = [xRS , xDOY , xIDX ]
Stage 4 x4A = [xSSD , xRS , xDOY , xIDX , xMET ] x4B = [xRS , xDOY , xIDX , xMET ]
[Link]
Water 2026, 18, 261 5 of 35
• Auxiliary in-situ variable (SSD): SSD serves as an auxiliary variable in the teacher
model for KD, reflecting water optical clarity and vertical attenuation. The teacher
model must capture detailed information, but the student model aims to achieve
comparable performance without it.
• Day of the year (DOY): A feature derived from sampling dates to capture seasonal
patterns such as light availability and biological productivity.
Collectively, these predictors describe optical conditions (RS), short-term environ-
mental forcing (MET), seasonal timing (DOY), and an optional in-situ clarity proxy (SSD,
available only to the teacher).
[Link]
Water 2026, 18, 261 6 of 35
where α ∈ [0, 1] balances the contribution of ground-truth labels ytrue and teacher
predictions ŷteacher . For feature selection, permutation-based ∆R2 Candidate features are
computed to quantify their importance; student models are then trained using the top-K
features (with K swept) to enhance compactness and generalization. To optimize α, we
perform a grid search over α = 0.0, 0.1, . . . , 0.9 and validate via station-grouped K-fold
cross-validation to prevent spatial leakage.
[Link]
Water 2026, 18, 261 7 of 35
2
∑i (yi − ŷi )
R2 = 1 − 2
(2)
∑i (yi − yi )
2. Root Mean Square Error (RMSE) quantifying the average magnitude of prediction error:
s
1
n∑
RMSE = (yi − ŷi )2 (3)
i
1
n∑
MAE = |yi − ŷi | (4)
i
In addition, mean absolute percentage error (MAPE) was reported only in the strat-
ified analyses to facilitate relative error comparisons across segments with different
concentration ranges:
100 n yi − ŷi
n i∑
MAPE = (5)
=1 yi + ϵ
where ŷ is the mean of y and ϵ is a small constant added to avoid instability when yi is
close to zero.
Our complete modelling dataset consists of 544 water-quality samples from 11 long-
term monitoring stations across Lake Simcoe. Each sample pairs an in-situ PPUT measure-
ment with a temporally matched PlanetScope scene and the corresponding meteorological
descriptors. To prevent spatial leakage and to evaluate the models on genuinely unseen
locations, we adopted a station-based grouped splitting strategy implemented via scikit-
learn’s GroupShuffleSplit. The grouping variable was the station identifier; all samples
from a given station were assigned exclusively to either the training or the test set. Using
a nominal 80/20 split (test_size = 0.2, random_state = 42), this procedure produced a
training set of 399 samples (73.3%) from 8 stations and a test set of 145 samples (26.7%)
from 3 distinct stations, with no station appearing in both sets. The slight deviation from a
perfect 80/20 ratio reflects the constraint that splits are performed at the station level rather
than at the individual-sample level.
Within the training set, we further employed station-grouped 5-fold cross-validation
(GroupKFold, n_splits = 5) for all hyperparameter tuning and feature-selection experiments,
including (i) the K-sweep over the number of retained features (K = 10, 15, . . ., 60) and (ii) the
grid search over the distillation weight α ∈ {0.0, 0.1, . . ., 0.9}. This nested design separates
(i) the outer train–test split, which measures generalization to new spatial locations (unseen
stations), from (ii) the inner station-grouped CV, which selects model hyperparameters
without leaking information across stations.
Such a station-grouped strategy is critical in inland water-quality modelling, because
neighbouring samples at the same station are strongly spatially autocorrelated; random
sample-wise splits would artificially inflate performance estimates and fail to reflect the
true difficulty of transferring models to new monitoring locations.
[Link]
Water 2026, 18, 261 8 of 35
[Link]
Water 2026, 18, 261 9 of 35
[Link]
Water 2026, 18, 261 10 of 35
model. Records underwent quality control to remove invalid or duplicate samples and
were spatially harmonized using the official station coordinates.
Meteorological variables were collected using the Meteostat API [50] and linked to
each in-situ record by selecting the five nearest available stations. The meteorological
fields included daily air temperature (tavg, tmin, tmax), precipitation, wind speed and
gust, atmospheric pressure, snowfall, and sunshine duration. To capture short-term hy-
drometeorological variability associated with watershed loading and water-column mixing,
3-day and 7-day aggregated metrics were derived. These variables were subsequently
merged with SR-based features and in-situ observations to create the final multi-modal
training dataset.
Figure 1. Lake Simcoe study area showing the LSRCA in-situ water-quality sampling stations
(blue dots) and Environment Canada meteorological stations (red dots) used to construct daily and
short-term (3–7 day) aggregated meteorological predictors for each satellite–in-situ match-up.
[Link]
Water 2026, 18, 261 11 of 35
polygon obtained from the Ontario Land Information Ontario (LIO) Waterbody dataset.
This ensured that subsequent modeling and mapping were restricted strictly to lake pixels,
thereby preventing land reflectance contamination in littoral zones, which are abundant
along the lake’s complex shoreline.
Temporal alignment between PlanetScope SR, meteorological observations, and in-
situ sampling was a central requirement. Because Lake Simcoe undergoes rapid changes
during spring melt, strict temporal-matching criteria (±1 day) were applied to minimize
discrepancies caused by rapidly changing optical and hydrologic conditions. Spatially,
samples located near river mouths—especially those of the Holland River and Black River—
were evaluated to ensure that their surrounding SR values were not affected by adjacency
effects or mixed land-water pixels. Where necessary, edge pixels were removed during
preprocessing to maintain data integrity.
The final prepared dataset retained only high-quality, temporally synchronized, and
spatially validated pixel-station pairs, serving as the foundation for model-training and
distillation stages.
[Link]
Water 2026, 18, 261 12 of 35
then used to guide the SSD-free student model. This approach enabled the student model
to inherit domain knowledge about water clarity even when operating without SSD.
To adapt the framework to the Lake Simcoe setting, feature importance rankings were
explicitly generated for this dataset. A Top-K feature selection analysis revealed that a
compact set of forty features provided the best balance between model complexity and
predictive performance. Using these Lake Simcoe-specific configurations, the final student
model achieved an R2 of 0.83 on held-out stations, demonstrating the robustness of the
KD-enhanced workflow for practical deployment.
4. Results
4.1. Influence of Feature Groups and SSD Availability on Model Performance
The multi-model evaluation across eight feature-group regimes was designed to sys-
tematically examine how different categories of predictors influence the stability, accuracy,
and generalizability of PPUT retrieval. Because phosphorus concentrations are governed
by a mixture of hydrological, biogeochemical, and optical processes, the effectiveness of
machine-learning models depends not only on algorithmic design but on the observability
of physically meaningful variables.
SSD represents a vertically integrated clarity metric that encodes water-column light
attenuation and particulate loads, both of which are tightly linked to phosphorus dynamics
in inland lakes. Table 2 presents a comparison of model-prediction error statistics (R2 ,
RMSE, MAE) across all models and feature groups, revealing a clear structural separation
between SSD-aware and SSD-free regimes.
While the raw numerical contrast shows that mean R2 increased from 0.6741 to 0.9364
when SSD was added, the broader implication is that SSD introduces a form of physical
regularization into the prediction problem. Because SSD encapsulates information about
water-column scattering and absorption processes strongly influenced by suspended sed-
iments, algal biomass, and particulate phosphorus, including SSD provides a constraint
that effectively reduces the model’s solution space. For completeness, we evaluated an
SSD-only model as a station-scale reference. Using SSD as the sole predictor, the SSD-only
models achieved test R2 of approximately 0.80–0.85 (best ≈ 0.846), with RMSE ≈ 13.7–15.8
and MAE ≈ 7.44–7.81. While SSD alone provides a strong predictive signal for PPUT
at monitoring stations, it is not deployable for lake-wide mapping because SSD is not
[Link]
Water 2026, 18, 261 13 of 35
Table 2. Single-model performance (R2 , RMSE, MAE) for SSD-aware and SSD-free regimes across the
five evaluated ensemble learners.
The high consistency of SSD-aware performance across five structurally distinct mod-
els (tree ensembles, gradient boosting, and CatBoost) indicates that the improvement is not
algorithm-dependent but arises from the biophysical relevance of SSD itself. The sharp de-
cline in accuracy in SSD-free scenarios (≈39% relative drop) further highlights the difficulty
of inferring vertical water clarity from surface-only features.
The boxplots in Figure 3 deepen this interpretation by demonstrating that SSD influ-
ences not only mean model accuracy but also the distribution of prediction errors. The
substantial reduction in RMSE (from 0.49 to 0.22 µg/L) and MAE (from 0.33 to 0.16 µg/L)
suggests that SSD reduces both bias and variance in the predictive outputs. This stabiliz-
ing effect stems from SSD’s ecological role: because phosphorus concentrations co-vary
strongly with suspended particulate loads and algal biomass, SSD acts as a proxy for the
processes that modulate phosphorus availability.
The fact that even minimal SSD-aware feature sets (e.g., SSD + Bands) outperform
feature-rich SSD-free models indicates that no combination of spectral indices or meteoro-
logical variables can fully substitute for the depth-integrated information that SSD provides.
SSD constrains predictions across heterogeneous optical conditions—such as nearshore
vs. offshore waters—thereby preventing error inflation in areas where surface reflectance
alone is ambiguous.
The R2 heatmap (Figure 4) demonstrates that SSD’s influence is robust across model
architectures with differing inductive biases. All models achieved R2 ≥ 0.93 under SSD-
aware regimes, demonstrating that SSD reduces the complexity of the prediction task to a
level where algorithmic differences become nearly irrelevant. This convergence suggests
that SSD effectively linearizes or simplifies the multidimensional mapping between observ-
[Link]
Water 2026, 18, 261 14 of 35
able features and PPUT, making the problem well-defined even for models that otherwise
underperform in high-variance conditions (e.g., Random Forest).
Figure 3. Boxplots summarizing model performance under SSD-aware versus SSD-free regimes
across the evaluated models and feature groups.
In SSD-free settings, the wide performance spread (0.55–0.75 R2 ) reflects the inherent
ambiguity of estimating phosphorus without clarity information. Here, models must rely
on indirect proxies, reflectance-based indices or meteorological drivers, whose relationships
to PPUT vary seasonally and spatially. The higher sensitivity to model architecture in this
regime highlights the extent to which SSD provides structural information that the model
would otherwise have to infer.
[Link]
Water 2026, 18, 261 15 of 35
The analyses in Figures 1–3 indicate that SSD is the single most influential variable
for PPUT prediction, not merely because it improves accuracy, but because it encodes
the fundamental optical and particulate processes that shape phosphorus dynamics. SSD
transforms the PPUT retrieval task from an underdetermined surface-reflectance inversion
into a physically grounded, well-constrained prediction problem. The strong cross-model
consistency under SSD-aware regimes and the pronounced instability of SSD-free models
demonstrate that high-fidelity phosphorus estimation in inland lakes requires access to
depth-integrated clarity information (either measured directly or approximated through
advanced techniques such as KD). These insights motivate the hybrid AI framework
developed in this study and lay the foundation for operational phosphorus monitoring at
high spatial and temporal resolution.
[Link]
Water 2026, 18, 261 16 of 35
[Link]
Water 2026, 18, 261 17 of 35
Table 3. Definitions of key predictors and their hypothesized physical relevance to PPUT dynamics
in Lake Simcoe.
[Link]
Water 2026, 18, 261 18 of 35
[Link]
Water 2026, 18, 261 19 of 35
For example, K = 32 and K = 40 both achieve similar validation accuracy (for K = 32,
R2 equals to 0.8306, RMSE equals to 9.853 and for K = 40, R2 equals to 0.8318 and RMSE
equals to 9.818), whereas K = 33 shows a local dip (R2 = 0.73, RMSE = 12.49 µg L−1 )
associated with a less favourable feature subset. The global optimum occurs at K = 40,
which attains the highest cross-validated R2 (0.8318) and the lowest RMSE (9.82 µg L−1 )
among all tested K values. For K > 40, both R2 and RMSE systematically deteriorate (e.g.,
K = 50 yields R2 ≈ 0.71–0.75 and RMSE ≈ 12.9–17.6 µg L−1 ), indicating overfitting and
redundancy when too many features are retained.
We therefore select K = 40 as the final configuration, representing the upper end of the
stable plateau and a sparse yet expressive 40-feature subset for the SSD-free student model.
Figure 8 confirms the feasibility of deploying SSD-free models for operational, lake-wide
nutrient monitoring with reasonable accuracy. For the K = 40 feature subset, we performed
a grid search over α ∈ {0.0, 0.1, . . ., 0.9} using station-grouped 5-fold cross-validation
(Group K-Fold) within the training set. α = 0.2 yielded the highest mean cross-validated R2
(0.676) and the lowest RMSE among all tested values, and was therefore adopted for all
reported student-model results.
To further evaluate model performance across different PPUT concentration ranges,
we conducted a stratified analysis by dividing the dataset into concentration bins (Table 4).
We focused on alternative metrics for evaluating performance within concentration subsets,
while R2 is sensitive to the variance of the dependent variable and can yield mislead-
ing values when applied to restricted-range data, we instead employed RMSE, MAE,
and MAPE (Mean Absolute Percentage Error) to assess model accuracy across different
concentration ranges.
[Link]
Water 2026, 18, 261 20 of 35
Figure 8. Predicted versus measured PPUT for the distilled student model on training and held-out
test subsets.
Table 4. Stratified performance of the student model of Lake Simcoe PPUT prediction.
In the low-concentration range (<25 µg/L, comprising 75.6% of all samples), the model
achieved RMSE values of 2.81 µg/L (training) and 6.29 µg/L (testing), with corresponding
MAPE values of 14.0% and 35.6%, respectively. These modest error magnitudes indicate
acceptable predictive performance in the concentration range most frequently observed in
Lake Simcoe. The high-concentration range (≥50 µg/L) showed RMSE values of 15.38 µg/L
(training) and 14.08 µg/L (testing), with MAPE values of 11.5% and 14.3%, demonstrating
consistent relative accuracy despite the larger absolute errors inherent to higher concentra-
tion values. The middle range (25–50 µg/L) exhibited higher variability in error metrics,
with testing RMSE of 24.19 µg/L and MAPE of 52.2%, likely due to the limited sample size
(23 in training, 7 in testing) and greater uncertainty in this transitional concentration zone.
[Link]
Water 2026, 18, 261 21 of 35
tions forming a broad mesotrophic background, while distinct high-value patches emerge
along the shoreline, in narrow channels, and at tributary confluences.
Figure 9. Lake-wide PPUT prediction map for Lake Simcoe on 26 July 2024 generated by the SSD-free
distilled student model from cloud-/land-masked PlanetScope SR mosaics.
The statistical distribution confirms that most of the lake surface remains in a relatively
low-concentration regime. Median and upper-quartile values are p50 = 7.15 µg/L and
p75 = 8.22 µg/L, with the 90th and 95th percentiles reaching 8.97 and 9.49 µg/L, respectively.
Across all water pixels (n ≈ 7.89 × 107 ), 97.40% fall within the 5–10 µg/L bin, and only
1.88% and 0.19% lie in the 10–20 µg/L and 20–30 µg/L ranges. Pixels exceeding 30 µg/L
account for just 0.53% of the lake area, yet they represent ecologically critical hotspots
where particulate phosphorus loading and/or resuspension are strongly enhanced.
At the basin scale, quadrant-wise averages indicate relatively modest cross-lake con-
trasts: the northeast and northwest quadrants show similar means (7.42 and 7.41 µg/L),
whereas the southeast and southwest quadrants are slightly elevated (7.78 and 7.70 µg/L),
consistent with the influence of significant inflows and shallow embayments in the south-
[Link]
Water 2026, 18, 261 22 of 35
ern basins. Overall, the 26 July snapshot demonstrates that the distilled student model
is capable of generating physically plausible, spatially detailed PPUT fields: the lake
interior is characterized mainly by low to moderate concentrations, while high-PPUT
waters are confined to structurally and hydrologically meaningful nearshore and tributary-
influenced zones.
5. Discussion
The experimental results obtained in this study provide an opportunity not only
to benchmark predictive performance but also to understand the physical, optical, and
algorithmic mechanisms that govern phosphorus retrieval from high-resolution satellite
data. Rather than viewing the models as black boxes, we interpreted their behavior in the
context of Lake Simcoe’s bio-optical regime, the contrasting roles of surface reflectance
and SSD, and the influence of short-term meteorological forcing. By jointly analyzing
multi-model feature groups, permutation-based importances, K-sweep dimensionality
patterns, and teacher-student knowledge-distillation behaviour, we can link quantitative
metrics such as R2 , RMSE, and MAE to underlying processes such as light attenuation,
sediment resuspension, watershed inputs, and stratification dynamics.
The following subsections synthesize these lines of evidence, with a focus on
(i) explaining SSD’s dominant predictive role, (ii) assessing the extent to which SSD-free
monitoring is feasible through distillation, (iii) clarifying how feature engineering and
model class shape performance, and (iv) discussing the implications, limitations, and
broader applicability of the proposed PlanetScope-based AI framework.
[Link]
Water 2026, 18, 261 23 of 35
marginal gains once SSD is included, indicating that the remaining error is not dominated by
model inadequacy but by the inherent observational limits of 4-band surface reflectance. In
this sense, SSD acts as a rate-limiting information source for PPUT prediction, transforming
an underdetermined inversion problem into a well-constrained regression task.
[Link]
Water 2026, 18, 261 24 of 35
influential predictors after the dominant meteorological drivers. This pattern is consistent
with indices amplifying the nonlinear balance between scattering and absorption across
the visible–NIR region, thereby isolating optical signatures of suspended particles, phyto-
plankton biomass, and CDOM that are only weakly expressed in individual raw bands.
The feature-importance hierarchy indicates a complementary division of information
content: meteorological variables encode the short-term forcing that redistributes phospho-
rus (runoff pulses, mixing, and resuspension), while reflectance-based indices capture the
optical manifestation of these processes in the surface layer. This coupling helps explain the
strong performance of the hybrid predictor set and supports the physical interpretability of
the learned relationships.
At the category level, meteorological variables emerged as the strongest contributors to
∆R2 , ahead of remote-sensing indices, with spectral bands and temporal descriptors playing
secondary roles. This ordering is physically consistent: short-term temperature, wind, and
pressure patterns govern stratification, resuspension, and runoff, which in turn mobilize
and redistribute phosphorus. Spectral indices then capture the optical manifestations of
these processes through changes in turbidity, pigment concentration, and water colour.
The relatively small incremental benefit of stand-alone bands and DOY implies that once
physically meaningful meteorology and indices are present, additional raw reflectance or
calendar information adds little unique explanatory power.
The K-sweep analysis provides a complementary perspective on feature-space com-
plexity. Accuracy increased rapidly as K rose from 10 to about 20, reflecting the addition of
genuinely informative features that capture independent aspects of phosphorus dynamics.
Between K = 20 and K = 40, performance gains became more modest, indicating a regime
of diminishing returns where newly added predictors were increasingly correlated with
variants of existing ones.
The optimal configuration at K = 40 corresponds to a point at which the model retains
sufficient feature diversity to approximate the teacher’s manifold while avoiding over-
parameterization. Beyond K > 50, performance degrades, consistent with overfitting in
a setting with modest sample size and many highly collinear features. These patterns
support the use of carefully curated, physically interpretable feature subsets in operational
scenarios, rather than indiscriminately including all available predictors.
[Link]
Water 2026, 18, 261 25 of 35
suggests that the primary bottleneck in PPUT retrieval lies in the observability of key
state variables rather than in the representational capacity of modern machine-learning
models. It also reinforces the idea that improvements in auxiliary data streams—such as
routine SSD sampling or other clarity proxies—may deliver greater benefits than further
algorithmic tuning.
Third, the success of the distilled student model demonstrates how teacher-student
frameworks can encode biophysical relationships—such as depth-dependent attenuation
and particulate scattering—into a compressed feature space. The teacher effectively learns
a high-dimensional representation of optical—biogeochemical structure underpinned by
SSD, while the student approximates this representation using only remote-sensing and
meteorological variables. This process resembles manifold learning, in which the student is
constrained to follow the teacher’s learned decision surface rather than exploring spurious
correlations in the SSD-free feature space.
Finally, the alignment between feature-importance patterns and established limnologi-
cal processes (e.g., the influence of wind-driven resuspension, storm-driven inflows, and
algal growth on phosphorus distribution) suggests that the models are learning mechanistic
relationships rather than arbitrary statistical associations. This enhances confidence in the
scientific interpretability and robustness of the proposed AI framework.
5.5. Comparison with Previous TP Retrieval Studies and Rationale for Backbone Selection
Previous work has made substantial progress in satellite-based retrieval of total phos-
phorus, but most studies face one or more constraints related to spatial resolution, feature
completeness, or model deployment ability. Xiong et al. developed a remote-sensing
algorithm for TP in eutrophic lakes using MODIS FAI-type indices combined with both
conventional and machine-learning models, achieving R2 ≈ 0.60 for a Taihu-specific model
and R2 ≈ 0.64 (RMSE ≈ 0.06 mg·L−1 ) for a generalized multi-lake algorithm [5,53]. This
work is valuable because it systematically compares semi-analytical and data-driven ap-
proaches and explicitly addresses generalization across lakes. However, the coarse 250 m
MODIS resolution limits its ability to resolve nearshore gradients and tributary plumes,
and the feature set is largely restricted to surface reflectance and FAI-type indices without
vertically integrated clarity metrics or short-term meteorological drivers.
Qiao et al. used Landsat-8 imagery for the Miyun Reservoir and conducted a com-
prehensive comparison of twelve machine-learning algorithms, showing that Extra Trees
(ETRs) yielded the best TP retrieval performance with R2 > 0.85 and very low MAE on a
single-reservoir dataset [53]. Their study is a strong benchmark for algorithmic comparison
under medium spatial resolution (30 m), but the feature space is dominated by spectral
bands and simple indices. The model is site-specific and operates at a scale where many
littoral processes remain subpixel, and there is no explicit treatment of auxiliary vertically
integrated indicators such as SSD or of temporal meteorological context.
Several recent studies have adopted gradient-boosting ensembles and Sentinel-class
imagery to improve TP retrievals at regional scales. Wang et al. used Sentinel-3/OLCI
images across the Yangtze–Huaihe lake region and found that an XGBoost-based model
outperformed empirical approaches but still achieved only moderate accuracy (R2 ≈ 0.53,
RMSE ≈ 0.08 mg·L−1 ) when generalized across many lakes [54]. Cui et al. optimized an
XGBoost model for TP retrieval in Taihu Lake using Sentinel-2 and a carefully selected
feature combination, achieving R2 ≈ 0.72 with improved stability relative to using all
variables [40]. Lin et al. further advanced this line of work by developing an interpretable
LightGBM-based model that reconstructs long-term TP dynamics (2005–2024) in Lake
Taihu, emphasizing explainability and driver attribution, yet still within a single large
lake and at 10–30 m resolution [55]. These studies collectively demonstrate that boosted
[Link]
Water 2026, 18, 261 26 of 35
tree ensembles (XGBoost, LightGBM and related methods) are consistently among the
top performers for TP retrieval, and that adding carefully engineered indices improves
robustness. However, they generally operate at medium resolution, rarely integrate in-situ
clarity proxies such as SSD into the retrieval model, and do not address how to deploy a
“high-information” model when such auxiliaries are absent.
Other recent work has begun to incorporate meteorological variables into TP or
nutrient modelling. For example, Li et al. proposed a multimodal framework that combined
satellite data and meteorological forcings to estimate several water-quality parameters,
including TP, obtaining R2 ≈ 0.50 for TP at the regional scale [56], while Qin et al. showed
that air temperature is a key driver for TP and TN variability in northeastern lakes when
combined with Sentinel-2 and machine-learning methods [57]. These studies highlight the
importance of meteorological context but typically treat meteorological features as simple
covariates; they neither integrate vertically integrated optical measures (SSD) nor explore
how such rich feature spaces can be transferred to operational settings where some inputs
are missing.
Against this backdrop, the present study differs from and extends prior work in three
main ways. First, it exploits multi-generation PlanetScope imagery (3–5 m) to explicitly
resolve nearshore and tributary structures that are unresolved by MODIS, OLCI, or even
Sentinel-2 in many lakes. This enables detection and mapping of narrow plume-like
phosphorus hotspots and embayment gradients that previous medium-resolution studies
could only infer indirectly or at coarse scales.
Second, the feature space explicitly integrates a vertically integrated clarity metric
(SSD), short-term meteorological descriptors (3–7 day aggregates of temperature, wind
speed, precipitation, and pressure), and physically informed spectral indices. This design
directly addresses two major gaps identified in the TP retrieval literature: the lack of
water-column integrated optical information and the limited incorporation of short-term
hydrometeorological drivers.
Third, the use of a teacher–student knowledge-distillation framework allows the high-
accuracy, SSD-informed teacher to be “compressed” into an SSD-free student that can
operate at any pixel where only remote-sensing and meteorological variables are available,
thereby resolving the common operational dilemma of sparse in-situ data.
The choice of HistGradientBoosting as the backbone model for both teacher and stu-
dent is also grounded in and consistent with the broader literature. Gradient-boosting
ensembles (including XGBoost, LightGBM, and related variants) have repeatedly emerged
as top performers in TP and nutrient retrieval tasks across lakes and regions, often outper-
forming random forests and support-vector regressors because they can capture complex,
non-linear interactions while remaining relatively robust on tabular feature sets [57–59].
In this study, an initial benchmark across five ensemble learners (HistGBM, CatBoost,
RandomForest, ExtraTrees, and GradientBoosting) and eight feature regimes showed that
HistGradientBoosting systematically offered the best or near-best trade-off between accu-
racy, stability across feature groups, and computational efficiency. On this basis, the hybrid
two-stage model used for knowledge distillation employs a HistGradientBoostingRegressor
as the base learner, augmented by a standardized Ridge regression “linear head” trained on
out-of-fold residuals and teacher predictions. This configuration benefits from the strong
non-linear fitting capacity of boosting while allowing a lightweight linear correction layer
to absorb residual structure and KD signals, providing a good balance between accuracy,
stability, and interpretability for both the SSD-informed teacher and the SSD-free student.
By explicitly positioning the proposed framework relative to these earlier studies, the
contribution of this work can be summarized as follows: it brings high-resolution (3–5 m)
TP mapping into an optically complex lake, leverages both vertically integrated clarity and
[Link]
Water 2026, 18, 261 27 of 35
[Link]
Water 2026, 18, 261 28 of 35
[Link]
Water 2026, 18, 261 29 of 35
Taken together, the analyses in Section 5 show that the success of PPUT retrieval in
Lake Simcoe is less determined by the choice of machine-learning architecture than by
which aspects of the lake system are made observable to the model. SSD emerges as a
pivotal depth-integrated constraint that renders the inversion problem well-posed, while
meteorological drivers and carefully designed spectral indices encode, respectively, the
forcing and the optical expression of phosphorus dynamics. KD then provides a principled
way to transfer this structure into an SSD-free student model, enabling scalable mapping
across space and time without fully sacrificing accuracy. At the same time, the identified
limitations—vertical ambiguity, optical water-type dependence, temporal mismatch, and
uncertainty propagation—highlight the need to view such models as components of an
integrated observing system rather than as stand-alone decision tools.
Thus, combining high-resolution PlanetScope imagery, physics-informed feature de-
sign, and teacher-student learning can bridge long-standing monitoring gaps in inland
waters, while preserving a clear mechanistic link between model outputs and the hydrolog-
ical and biogeochemical processes they are intended to represent.
6. Conclusions
This study demonstrates that integrating multi-generation PlanetScope imagery with a
hybrid machine-learning and knowledge-distillation framework provides a robust pathway
for high-resolution, lake-wide retrieval of total phosphorus (PPUT) in an optically complex
inland lake.
By systematically evaluating five ensemble learners across eight feature-group regimes,
we showed that the dominant constraint on PPUT retrieval is not model architecture but
feature observability—most notably the availability of SSD as a vertically integrated indi-
cator of water clarity. Across all models, including SSD increased the mean R2 from
approximately 0.67 to 0.94 and reduced error metrics by more than half, confirming
that water-column transparency strongly governs the learnable structure of the PPUT
prediction problem.
Building on this insight, we developed a two-stage teacher–student framework in
which a physically informed teacher model is trained with SSD and transfers its represen-
tation to an SSD-free student via knowledge distillation. The distilled student retained
about 88% of the teacher’s accuracy (R2 = 0.83, RMSE = 9.82 µg/L, MAE = 5.41 µg/L)
while relying only on PlanetScope reflectance, derived spectral indices, and short-term
meteorological descriptors. The K-sweep analysis further revealed that an intermediate
subset of 40 features provides an optimal balance between predictive skill and parsimony,
indicating that a compact combination of optical and meteorological drivers can encode
the essential dynamics of phosphorus transport, resuspension, and biological uptake.
Application of the SSD-free student model to PlanetScope SuperDove imagery from
2020 to 2025 produced metre-scale PPUT maps that resolve spatial patterns far beyond the
capability of traditional monitoring and medium-resolution satellites. A representative
case study for 26 July 2024 showed that the vast majority of lake-surface pixels (>97%) fall
within a low-concentration band of 5–10 µg/L, whereas rare (<1%) but spatially coherent
hotspots exceeding 20 µg/L occur near tributary mouths, sheltered embayments, and
narrow channels. These hotspots form contiguous clusters that coincide with known
river-inflow corridors and shallow, wind-exposed shorelines, highlighting the role of
nearshore processes and watershed inputs in structuring phosphorus heterogeneity across
Lake Simcoe.
In general, the findings indicate that the lake-wide nutrient monitoring using labour-
intensive field campaigns can be complemented and enhanced using high-resolution
satellite observations and distilled AI models. The proposed framework is readily transfer-
[Link]
Water 2026, 18, 261 30 of 35
able to other lakes where SSD or similar auxiliary measurements are collected at limited
stations but cannot be mapped wall-to-wall. By enabling near-real-time estimation of PPUT
at 3 m resolution, the approach provides a foundation for adaptive monitoring programs,
early-warning tools for eutrophication, and improved watershed-scale nutrient budgeting.
Overall, the combination of commercial high-resolution imagery, physics-informed feature
engineering and knowledge-transfer algorithms constitutes a robust and generalizable
strategy for inland water-quality retrieval.
Author Contributions: Y.D. contributed to the conceptualization, methodology, and the original
and final writing of the manuscript. S.X.Y. and B.G. both contributed to the conceptualization,
writing—review and editing, supervision of the manuscript, project administration, and funding
acquisition. D.P. contributed to the writing—review and editing. All authors have read and agreed to
the published version of the manuscript.
Funding: This research was funded by the Natural Sciences and Engineering Research Council of
Canada (NSERC) Alliance Grant #401643.
Data Availability Statement: The datasets used in this study are publicly available and can
be downloaded from the internet. PlanetScope imagery used in this study was obtained from
Planet Labs PBC under an academic research licence and is not publicly shareable. Interested
researchers can obtain similar imagery directly from Planet Labs “[Link] (ac-
cessed on 28 August 2025)”, subject to licensing conditions. In-situ water-quality and Secchi depth
data for Lake Simcoe and its Tributaries were provided by the Lake Simcoe Region Conserva-
tion Authority’s Open Data portal “[Link] (accessed on
28 August 2025)” and the Ontario Ministry of the Environment, Conservation and Parks Open Data
Portal “[Link] (accessed on 28 August 2025)”. The lake poly-
gon boundaries were obtained from the Land Information Ontario “[Link]
(accessed on 15 October 2025)” geospatial data portal operated by the Ontario Ministry of Natural
Resources and Forestry. Historical meteorological data were retrieved from the Meteostat project
“[Link] (accessed on 28 August 2025)”. Derived PPUT prediction maps and analysis
scripts are available from the corresponding author on reasonable request.
Acknowledgments: The authors gratefully acknowledge the Lake Simcoe Region Conservation
Authority (LSRCA) and the Ontario Ministry of the Environment, Conservation and Parks (MECP)
for providing in-situ water-quality and Secchi depth observations for Lake Simcoe. Land Information
Ontario (LIO) of the Ontario Ministry of Natural Resources and Forestry is thanked for access to
lake-boundary geospatial data used for lake delineation and masking. The authors also thank the
Meteostat project for providing open, quality-controlled meteorological data and Planet Labs PBC
for supplying PlanetScope imagery under an academic research license. The authors sincerely thank
Andrew Paterson (Ontario Ministry of the Environment, Conservation and Parks; Dorset Environ-
mental Science Centre) for his careful review and constructive comments, which helped improve the
manuscript. The authors also appreciate helpful feedback from colleagues and anonymous reviewers.
Conflicts of Interest: The authors declare that the research was conducted in the absence of any
commercial or financial relationships that could be construed as potential conflicts of interest.
[Link]
Water 2026, 18, 261 31 of 35
tion. For our study, HistGBM is particularly valuable for handling high-dimensional input
spaces (a potential consequence of including MET and SSD features) and large datasets,
enabling faster experimentation across the 8 input regimes. Its ability to model non-linear
relationships through additive trees:
T
F(x) = ∑ ht ( x ) (A1)
t =1
where each ht is a decision tree and T is the number of trees, aligns with our goal of
capturing complex dependencies between input features and the target variable.
1 T
T t∑
F(x) = ht ( x; θt , Dt ) (A2)
=1
where θt are the tree parameters (split points, etc.) learned using a random subset of
features, and Dt is the bootstrap sample for tree T. This inherent randomness decorrelates
the trees, reducing variance and improving generalization. RandomForest is well-suited
for our study because it provides a stable, non-parametric baseline that is less prone to
overfitting compared to single decision trees. Its ability to capture non-linear interactions
and its insensitivity to feature scaling make it ideal for analyzing the impact of adding MET
features across various SSD contexts, as it can robustly model both simple and complex
relationships without strong assumptions about the data distribution.
[Link]
Water 2026, 18, 261 32 of 35
1 T ∼
F(x) = ∑
T t =1
ht ( x; θ t , Dt ) (A3)
∼
where θ t now includes randomly selected split thresholds. For our research, ExtraTrees is
valuable as it provides a different inductive bias compared to RandomForest and gradient
boosting methods. Its faster training time (due to avoiding exhaustive split point search)
and potential to find novel splits that a greedy search might miss make it a complementary
model for evaluating the robustness of MET’s added value across diverse input regimes,
including those with high levels of SSD-aware noise or variability.
where L is the loss function, typically MSE L(y, ŷ) = y − ŷ)2 . Subsequent trees ht are fit
to the negative gradient of the loss with respect to the current prediction. The ensemble is
updated as:
Ft ( x ) = Ft−1 ( x ) + νht ( x ) (A5)
where ν is the learning rate. Unlike HistGBM or CatBoost, this implementation typically
uses exact split finding (evaluating all possible split points for each feature) and does not
include specialized handling for categorical features or advanced regularization beyond
subsampling. This model serves as a crucial baseline in our study because it represents the
traditional gradient boosting approach without the optimizations of HistGBM or CatBoost.
By comparing its performance across input regimes to the more optimized variants, we can
better isolate the specific contributions of DOY, INDICES, MET and SSD features from the
effects of model-specific optimizations. Its tendency to overfit if not properly regularized
also makes it a good test for the discriminative power of the above features- if these features
truly add signal, it should help even a simpler GBR model generalize better across different
SSD conditions.
By training these five diverse models under each of the 8 input regimes, we aim to
systematically assess how the inclusion of DOY, INDICES, MET features, under varying
SSD contexts, impacts predictive performance across different learning paradigms, thus
ensuring the robustness and generalizability of our findings.
[Link]
Water 2026, 18, 261 33 of 35
References
1. Wei, Z.; Yu, Y.; Yi, Y. Analysis of future nitrogen and phosphorus loading in watershed and the risk of lake blooms under the
influence of complex factors: Implications for management. J. Environ. Manag. 2023, 345, 118581. [CrossRef] [PubMed]
2. Paerl, H.W.; Havens, K.E.; Xu, H.; Zhu, G.; McCarthy, M.J. Mitigating eutrophication and toxic cyanobacterial blooms in large
lakes: The evolution of a dual nutrient (N and P) reduction paradigm. Hydrobiologia 2020, 847, 4359–4373. [CrossRef]
3. Ding, S.; Chen, M.; Gong, M.; Fan, X.; Qin, B.; Xu, H. Internal phosphorus loading from sediments causes seasonal nitrogen
limitation for harmful algal blooms in a large shallow lake. Sci. Total Environ. 2018, 636, 139–148.
4. Li, J.; Wang, J.; Wu, Y.; Cui, Y.; Yan, S. Remote sensing monitoring of total nitrogen and total phosphorus concentrations in the
water around Chaohu Lake based on geographical division. Front. Environ. Sci. 2022, 10, 1014155. [CrossRef]
5. Xiong, J.; Lin, C.; Cao, Z.; Hu, M.; Xue, K.; Chen, X.; Ma, R. Development of remote sensing algorithm for total phosphorus
concentration in eutrophic lakes: Conventional or machine learning? Water Res. 2022, 218, 118444. [CrossRef]
6. Wu, Y. Adjacency Effect in Nearshore Aquatic Remote Sensing: Modelling, Correction, and Application; University of Ottawa Research
Repository: Ottawa, ON, Canada, 2025.
7. Zhang, R.; Chu, N.; Yin, K.; Dong, L.; Li, Q.; Liu, H. Satellite-Based Analysis of Nutrient Dynamics in Northern South China Sea
Marine Ranching under the Combined Effects of Climate Warming and Anthropogenic Activities. J. Mar. Sci. Eng. 2025, 13, 1677.
[CrossRef]
8. Guo, H.; Huang, J.J.; Zhu, X.; Tian, S.; Wang, B. Spatiotemporal variation reconstruction of total phosphorus in the Great Lakes
since 2002 using remote sensing and deep neural network. Water Res. 2024, 258, 120450. [CrossRef] [PubMed]
9. Batina, A.; Krtalić, A. Integrating remote sensing methods for monitoring lake water quality: A comprehensive review. Hydrology
2024, 11, 92. [CrossRef]
10. Zhai, M.; Zhou, X.; Tao, Z.; Xie, Y.; Yang, J.; Shao, W. Satellite-ground synchronous in-situ dataset of water optical parameters and
surface temperature for typical lakes in China. Sci. Data 2024, 11, 883.
11. Ashphaq, M.; Srivastava, P.K.; Mitra, D. Preliminary examination of influence of Chlorophyll, Total Suspended Material, and
Turbidity on Satellite Derived-Bathymetry estimation in coastal turbid water. Heliyon 2023, 9, e17681.
12. Liu, J.; Qiu, Z.; Feng, J.; Wong, K.P.; Tsou, J.Y.; Wang, Y. Monitoring Total Suspended Solids and Chlorophyll-a Concentrations
in Turbid Waters: A Case Study of the Pearl River Estuary and Coast Using Machine Learning. Remote Sens. 2023, 15, 5559.
[CrossRef]
13. Theenathayalan, V.; Sathyendranath, S.; Kulk, G. Regional satellite algorithms to estimate chlorophyll-a and total suspended
matter concentrations in Vembanad Lake. Remote Sens. 2022, 14, 6404. [CrossRef]
14. Abualhin, K.; Abushaban, S. Predictive models of non-optically active coastal water quality parameters by remote sensing
imagery. Maejo Int. J. Sci. Technol. Res. 2025, 14, 274944. [CrossRef]
15. Lin, W.; Li, N.; Zhang, Y.; Shi, K.; Guo, H.; Zhang, Y.; Qin, B. Widespread decrease of phosphorus and the potential driving
mechanisms in Taihu Basin’s lakes. Environ. Res. Commun. 2025, 7, 125011. [CrossRef]
16. Chen, L.; Liu, L.; Liu, S.; Shi, Z.; Shi, C. The application of remote sensing technology in inland water quality monitoring and
water environment science: Recent progress and perspectives. Remote Sens. 2025, 17, 667. [CrossRef]
17. Yang, X.; Chen, J.; Lu, X.; Liu, H.; Liu, Y.; Bai, X.; Qian, L. Advances in UAV Remote Sensing for Monitoring Crop Water and
Nutrient Status: Modeling Methods, Influencing Factors, and Challenges. Plants 2025, 14, 2544. [CrossRef]
18. Assunção, A.; Silva, T.F.G.; de Carvalho, L.A.S. Assessing water quality restoration measures in Lake Pampulha (Brazil) through
remote sensing imagery. Environ. Sci. Pollut. Res. 2025, 32, 21277–21291. [CrossRef]
19. Tavares, M.H.; Guimarães, D.; Roussillon, J.; Baute, V. A Framework to Retrieve Water Quality Parameters in Small, Optically
Diverse Freshwater Ecosystems Using Sentinel-2 MSI Imagery. Remote Sens. 2025, 17, 2729. [CrossRef]
20. Karimi, N.; Torabi, O. Remote Sensing-Based Bathymetry Mapping in Shallow Lakes: Comparative Analysis of Sentinel-2 and
Landsat-8 Imagery Integrated with Machine Learning. Cont. Shelf Res. 2025, 272, 105075. [CrossRef]
21. Deng, Y.; Zhang, Y.; Pan, D.; Yang, S.X.; Gharabaghi, B. Review of Recent Advances in Remote Sensing and Machine Learning
Methods for Lake Water Quality Management. Remote Sens. 2024, 16, 4196. [CrossRef]
22. Zhao, Y.; He, X.; Pan, S.; Bai, Y.; Wang, D.; Li, T. Satellite Retrievals of Water Quality for Diverse Inland Waters from Sentinel-2
Images: An Example from Zhejiang Province, China. Ecol. Indic. 2024, 164, 114875.
23. Wasehun, E.T.; Hashemi Beni, L.; Di Vittorio, C.A. UAV and Satellite Remote Sensing for Inland Water Quality Assessments: A
Literature Review. Environ. Monit. Assess. 2024, 196, 12342. [CrossRef] [PubMed]
24. Parida, B.R.; Tiwari, S.; Dwivedi, C.S.; Pandey, A.C. Comparative Assessment of Satellite-Based Models through PlanetScope and
Landsat-8 for Determining Physico-Chemical Water Quality Parameters in Varuna River, India. Appl. Water Sci. 2025, 15, 2367.
25. Kabir, S.; Saranathan, A.M.; Barnes, B. Feasibility of PlanetScope SuperDove Constellation for Water Quality Monitoring of Inland
and Coastal Waters. Front. Remote Sens. 2025, 6, 1624783.
26. Liu, B.; Li, T. A Machine-Learning-Based Framework for Retrieving Water Quality Parameters in Urban Rivers Using UAV
Hyperspectral Images. Remote Sens. 2024, 16, 905.
[Link]
Water 2026, 18, 261 34 of 35
27. Tian, D.; Zhao, X.; Gao, L.; Liang, Z.; Yang, Z.; Zhang, P. Inversion of Water Quality Variables Based on Machine Learning
and Cluster-Analysis Empirical Models Using Multi-Source Remote Sensing Data in Inland Reservoirs. Environ. Pollut. 2024,
339, 123021.
28. Xu, X.; Pan, J.; Zhang, H.; Lin, H. Progress in Remote Sensing of Heavy Metals in Water. Remote Sens. 2024, 16, 3888. [CrossRef]
29. Biswas, P. Development of Chlorophyll-a Soft Sensor Using Machine Learning and IoT. ResearchGate Preprint. Master’s Thesis,
Universiti Malaya, Kuala Lumpur, Malaysia, 2023.
30. Tian, S.; Guo, H.; Xu, W.; Zhu, X.; Wang, B.; Zeng, Q. Remote Sensing Retrieval of Inland Water Quality Parameters Using
Sentinel-2 and Multiple Machine Learning Algorithms. Environ. Sci. Pollut. Res. 2023, 30, 44478–44492. [CrossRef]
31. Li, H.; Wang, N.; Du, Z.; Huang, D.; Shi, M.; Zhong, Z.; Yuan, D. Multi-Parameter Water Quality Inversion in Heterogeneous
Inland Waters Using UAV-Based Hyperspectral Data and Deep Learning Methods. Remote Sens. 2025, 17, 2191.
32. Zheng, D.; Lv, A. MosaicFormer: A Novel Approach to Remote Sensing Spatiotemporal Data Fusion for Lake Water Monitors.
Remote Sens. 2025, 17, 1138.
33. Chen, Y.; Xue, P. Dual-Transformer Deep Learning Framework for Seasonal Forecasting of Great Lakes Water Levels. J. Hydromete-
orol. 2025, 26, e519. [CrossRef]
34. Pan, D.; Deng, Y.; Yang, S.X.; Gharabaghi, B. Recent Advances in Remote Sensing and Artificial Intelligence for River Water
Quality Forecasting: A Review. Environments 2025, 12, 158. [CrossRef]
35. Zhou, Y.; Li, W.; Cao, X.; He, B.; Feng, Q.; Yang, F. Spatial-Temporal Distribution of Labeled Set Bias in Remote Sensing Estimation:
Implication for Supervised Machine Learning in Water Quality Monitoring. Ecol. Indic. 2024, 163, 114313. [CrossRef]
36. Sajib, A.M.; Uddin, M.G.; Rahman, A.; Ahmadian, R.; Olbert, A.I. Remote sensing applications for monitoring optically inactive
water quality indicators: A comprehensive review. Earth-Sci. Rev. 2025, 271, 105259. [CrossRef]
37. Ramtel, P.; Feng, D.; Gardner, J. Toward Large-Scale Riverine Phosphorus Estimation Using Remote Sensing and Machine
Learning. J. Geophys. Res. Biogeosci. 2024, 129, e2024JG008121. [CrossRef]
38. Cristofoli, E. Using Satellite Images and Deep Learning to Detect Water Hidden Under the Vegetation: A Cross-Modal Knowledge
Distillation-Based Method to Reduce Manual Labeling. Master’s Thesis, Luleå University of Technology, Luleå, Sweden, 2024.
39. Zhou, W.; Li, Y.; Huan, J.; Liu, Y. MSTNet-KD: Multilevel Transfer Networks Using Knowledge Distillation for Dense Prediction
of Remote-Sensing Images. IEEE Trans. Geosci. Remote Sens. 2024, 62, 4504612. [CrossRef]
40. Cui, J.; Wu, S.; Dai, J.; Xue, W.; Zhang, Y.; You, J.; Lv, X. Satellite Retrievals of Total Phosphorus in Taihu Lake Using Sentinel-2
Images and an Optimized XGBoost Model. Ecol. Inform. 2025, 82, 1004935. [CrossRef]
41. Qin, H.; Fang, C.; Liu, G.; Song, K.; Li, Z.; Li, S.; Tao, H.; Yan, Z. Temperature Is a Key Factor Affecting Total Phosphorus and Total
Nitrogen Concentrations in Northeastern Lakes Based on Sentinel-2 Images and Machine Learning Methods. Remote Sens. 2025,
17, 267. [CrossRef]
42. Yin, L.; Wang, C.; Wang, X. Research on Stacking Ensemble Learning-Based Remote Sensing Retrieval of Total Phosphorus
Concentration in Poyang Lake and Its Multi-Dimensional Driving Mechanisms. IEEE Trans. Geosci. Remote Sens. 2025, 18,
25770–25780.
43. Leggesse, E.S.; Zimale, F.A.; Sultan, D.; Enku, T.; Tilahun, S.A. Advancing non-optical water quality monitoring in Lake Tana,
Ethiopia: Insights from machine learning and remote sensing techniques. Front. Water 2024, 6, 1432280. [CrossRef]
44. Ngamile, S.; Madonsela, S.; Kganyago, M. Trends in remote sensing of water quality parameters in inland water bodies: A
systematic review. Front. Environ. Sci. 2025, 13, 1549301. [CrossRef]
45. Xiong, J.; Lin, C.; Ma, R.; Cao, Z. Remote Sensing Estimation of Lake Total Phosphorus Concentration Based on MODIS: A Case
Study of Lake Hongze. Remote Sens. 2019, 11, 2068.
46. Oliveira Santos, V.; Guimarães, B.M.D.M.; Neto, I.E.L.; de Souza Filho, F.D.A.; Costa Rocha, P.A.; Thé, J.V.G.; Gharabaghi, B.
Chlorophyll-a Estimation in 149 Tropical Semi-Arid Reservoirs Using Remote Sensing Data and Six Machine Learning Methods.
Remote Sens. 2024, 16, 1870.
47. Ontario Ministry of Natural Resources and Forestry. Waterbody (LIO) Dataset. Available online: [Link]
datasets/29a6e59237bd4fbe8f013b52971dbd25_14 (accessed on 12 August 2025).
48. Santos, V.O.; Rocha, P.A.C.; Thé, J.V.G.; Gharabaghi, B. Evaluation of machine learning methods for forecasting turbidity in river
networks using Sentinel-2 remote sensing data. Ecol. Inform. 2025, 90, 103313.
49. Ontario Ministry of the Environment, Conservation and Parks. Lake Simcoe Water Quality Monitoring Data (Chemistry and Secchi
Depth), 1980–2023; Available via Ontario Open Data Catalogue; Queen’s Printer for Ontario: Toronto, ON, Canada, 2023.
50. Meteostat. Meteostat: Historical Weather and Climate Data. 2025. Available online: [Link] (accessed on
12 August 2025).
51. Downes, J.; Bruce, D.; Miot da Silva, G.; Hesp, P.A. Optimising Satellite-Derived Bathymetry Using Optical Imagery over the
Adelaide Metropolitan Coast. Remote Sens. 2025, 17, 849.
52. Palmer, M.E.; Winter, J.G.; Young, J.D.; Dillon, P.J.; Guildford, S.J. Introduction and summary of research on Lake Simcoe:
Research, monitoring, and restoration of a large lake and its watershed. J. Great Lakes Res. 2011, 37, 1–6. [CrossRef]
[Link]
Water 2026, 18, 261 35 of 35
53. Xiong, J.; Lin, C.; Ma, R.; Wang, X.; Xue, K.; Cao, Z.; Hu, M.; Chen, L. Remote Sensing Observations of Phosphorus in Eutrophic
Lakes: From Concentration to Storage. IEEE Trans. Geosci. Remote Sens. 2025, 63, 4203812.
54. Qiao, Z.; Sun, S.; Jiang, Q.; Xiao, L.; Wang, Y.; Yan, H. Retrieval of Total Phosphorus Concentration in the Surface Water of Miyun
Reservoir Based on Remote Sensing Data and Machine Learning Algorithms. Remote Sens. 2021, 13, 4662. [CrossRef]
55. Wang, X.; Jiang, Y.; Jiang, M.; Cao, Z.; Li, X.; Ma, R.; Xu, L.; Xiong, J. Estimation of Total Phosphorus Concentration in Lakes in the
Yangtze–Huaihe Region Based on Sentinel-3/OLCI Images. Remote Sens. 2023, 15, 4487.
56. Lin, W.; Zhou, Y.; Ren, Z.; Zou, W.; Guo, H.; Li, N.; Zhang, Y.; Elser, J.; Woolway, R.I.; Shi, K.; et al. Interpretable Data-Driven
Modeling of Total Phosphorus Dynamics from 2005 to 2024 in a Large Shallow Lake. Water Res. 2025, 291, 125169. [CrossRef]
[PubMed]
57. Li, N.; Yang, G.; Zhao, X.; Liang, Q.; Sun, Y.; Chen, Z.; Zhao, H. Multimodal Remote Sensing and Meteorological Estimation of
Long-Term Spatial-Scale Non-Point-Source Pollution in Inland Reservoirs. Ecol. Indic. 2025, 159, 113744.
58. Chen, X.; Yang, Y.; Jiang, Y. Gradient Boosting and Remote Sensing Analysis of Spatiotemporal Variations of Total Nitrogen and
Phosphorus in Donghu Lake, Wuhan. Inland Waters 2025, 15, 2441618. [CrossRef]
59. Yan, C.; Fu, X.; Gao, H.; Dong, W.; Liu, Z.; Xu, Z. Enhancing Chlorophyll-a Estimation in Optically Complex Waters Using
ZY-1 02E Hyperspectral Imagery: An Integrated Approach Combining Optical Classification and Multi-Index Blending Models.
Remote Sens. 2025, 17, 3795.
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual
author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to
people or property resulting from any ideas, methods, instructions or products referred to in the content.
[Link]