LimeSoDa: A Dataset Collection for Benchmarking of Machine Learning Regressors in Digital Soil Mapping
Authors:
J. Schmidinger,
S. Vogel,
V. Barkov,
A. -D. Pham,
R. Gebbers,
H. Tavakoli,
J. Correa,
T. R. Tavares,
P. Filippi,
E. J. Jones,
V. Lukas,
E. Boenecke,
J. Ruehlmann,
I. Schroeter,
E. Kramer,
S. Paetzold,
M. Kodaira,
A. M. J. -C. Wadoux,
L. Bragazza,
K. Metzger,
J. Huang,
D. S. M. Valente,
J. L. Safanelli,
E. L. Bottega,
R. S. D. Dalmolin
, et al. (11 additional authors not shown)
Abstract:
Digital soil mapping (DSM) relies on a broad pool of statistical methods, yet determining the optimal method for a given context remains challenging and contentious. Benchmarking studies on multiple datasets are needed to reveal strengths and limitations of commonly used methods. Existing DSM studies usually rely on a single dataset with restricted access, leading to incomplete and potentially mis…
▽ More
Digital soil mapping (DSM) relies on a broad pool of statistical methods, yet determining the optimal method for a given context remains challenging and contentious. Benchmarking studies on multiple datasets are needed to reveal strengths and limitations of commonly used methods. Existing DSM studies usually rely on a single dataset with restricted access, leading to incomplete and potentially misleading conclusions. To address these issues, we introduce an open-access dataset collection called Precision Liming Soil Datasets (LimeSoDa). LimeSoDa consists of 31 field- and farm-scale datasets from various countries. Each dataset has three target soil properties: (1) soil organic matter or soil organic carbon, (2) clay content and (3) pH, alongside a set of features. Features are dataset-specific and were obtained by optical spectroscopy, proximal- and remote soil sensing. All datasets were aligned to a tabular format and are ready-to-use for modeling. We demonstrated the use of LimeSoDa for benchmarking by comparing the predictive performance of four learning algorithms across all datasets. This comparison included multiple linear regression (MLR), support vector regression (SVR), categorical boosting (CatBoost) and random forest (RF). The results showed that although no single algorithm was universally superior, certain algorithms performed better in specific contexts. MLR and SVR performed better on high-dimensional spectral datasets, likely due to better compatibility with principal components. In contrast, CatBoost and RF exhibited considerably better performances when applied to datasets with a moderate number (< 20) of features. These benchmarking results illustrate that the performance of a method is highly context-dependent. LimeSoDa therefore provides an important resource for improving the development and evaluation of statistical methods in DSM.
△ Less
Submitted 20 May, 2025; v1 submitted 27 February, 2025;
originally announced February 2025.
Synthesis of One Atom Thin, Two-Dimensional Gold Films and Their Novel Properties
Authors:
Sudhir Kumar Sharma,
Renu Pasricha,
James Weston,
Florian Stumpf,
Thomas Blanton,
Ramesh Jagannathan
Abstract:
Though significant advances have been made in the field of metal nanostructures, researchers are yet to synthesize one atom thin two-dimensional 2D gold nanostructures. We report, for the first time, a technique to synthesize one atom thin gold films and membrane like porous films on silicon and sapphire. These films were essentially liquid at room temperature. Current-Voltage spectroscopy using a…
▽ More
Though significant advances have been made in the field of metal nanostructures, researchers are yet to synthesize one atom thin two-dimensional 2D gold nanostructures. We report, for the first time, a technique to synthesize one atom thin gold films and membrane like porous films on silicon and sapphire. These films were essentially liquid at room temperature. Current-Voltage spectroscopy using atomic force microscopy revealed a typical Schottky behavior with a very high turn on (knee) voltage at 4.15V indicating that these 2D gold structures are semi-conductors. Nanorings comprising self assembled one atom thin gold structures on sapphire were also observed. Subjecting one nanoring in a group of several nanorings to a point force of 2.25 micro newton for 120 seconds resulted in the creation of a mirror structure, accompanied by every nanoring in the group to simultaneously create its individual mirror structures. These mirror structures had an apparent spatial and temporal correlation to each other. They had a finite life time of a few hours and the process was repeatable to a great degree of precision.
△ Less
Submitted 30 July, 2020;
originally announced July 2020.