arXiv · 2610.06369
Environmental sensor readings in two crop disease image datasets identify the session in which each image was taken
Abstract
Integrating environmental sensor data with leaf imagery is widely reported to boost crop disease classification accuracy. In this work, we reveal that these reported gains are often artifacts of dataset construction: because a single sensor reading is shared across many images collected in a single session (one farm on one date), multimodal networks can predict disease simply by memorizing session identities. Analyzing two widely used Korean datasets, the Crop Disease Diagnosis (CDD) benchmark and an AI Hub pest/disease dataset, we demonstrate that nearly all images share sensor values, with 91.9% of CDD test images having exact sensor duplicates in the training set. Remarkably, an image-free classifier given only timestamps matches or exceeds sensor-driven predictions across all seven evaluated crops, and matches the published macro-F1 of a state-of-the-art CDD fusion model. These results indicate that performance gains on standard random splits cannot be disentangled from session leakage. We propose that multimodal crop studies must evaluate on session-held-out splits and report performance against sensor-free date-time baselines to ensure genuine generalization.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sungwoo Kang. 2026-10-05. Environmental sensor readings in two crop disease image datasets identify the session in which each image was taken. https://arxiv.org/abs/2610.06369
Cite the original work for its findings. Save a collection to share your selection of sources.