Digitální knihovna UPCE přechází na novou verzi. Omluvte prosím případné komplikace. / The UPCE Digital Library is migrating to a new version. We apologize for any inconvenience.

Estimation of atmospheric visibility by deep learning model using multimodal dataset

ČlánekOmezený přístuppeer-reviewedpostprint

Abstrakt

Accurate estimation of atmospheric visibility is essential for numerous safety-critical applications, particularly in the field of transportation. In this study, a deep learning-based approach is investigated using a multimodal input representation that combines RGB images from a fixed-position surveillance camera with tabular meteorological variables collected from a nearby meteorological station. The meteorological input includes temperature, absolute pressure, relative humidity, dew point, wet bulb temperature, average and maximum wind speed, amount of precipitation, solar radiation, and ultraviolet index. Six neural network models for visibility estimation were developed and compared: a multimodal model utilizing both image and tabular meteorological inputs; two ablation models that use only unimodal input (image or meteorological data); a regions-of-interest (ROIs)-based model that extracts features from predefined image subregions; and two ablation models that use only a reduced number of meteorological variables. The multimodal model uses EfficientNetV2M for feature extraction and a set of fully connected neural networks to integrate the two modalities. The ROIs-based model also uses EfficientNetV2M, but only on manually selected reference regions of the scene. Evaluation was performed on a dataset of 1,000 annotated images, with visibility manually determined based on reference points in the scene. The multimodal model achieved a mean squared error of 129,716 m², a mean absolute error of 165.4 m, and an R² score of 0.8861, with 84.46% of predictions falling within a 10% relative error margin. Although the ROIs-based model slightly outperformed the multimodal model in some regression metrics, its accuracy within tolerance thresholds was lower, and its reliance on manual scene annotation limits scalability. In contrast, the ablation models demonstrated lower performance in almost all evaluated criteria. The results show that the proposed multimodal input strategy provides a balanced and practical approach to automated visibility estimation. Compared to conventional unimodal models, this architecture offers improved accuracy, stability, and generalization ability, making it suitable for real-world applications where both visual and environmental data are available.

Rozsah stran

p. 1-14

ISSN

0950-7051

Permanentní identifikátor

Projekt

EH23_021/0008402/Mezisektorová a mezioborová spolupráce ve výzkumu a vývoji komunikačních, informačních a detekčních technologií pro řídicí a zabezpečovací systémy, registrační číslo CZ.02.01.01/00/23_021/0008402

Časopis nebo seriál

Knowledge-Based Systems, volume 331, issue: 1

Vydavatelská verze

https://www.sciencedirect.com/science/article/pii/S095070512501771X

Přístup k e-verzi

Práce není přístupná

Název akce

ISBN

Studijní obor

Studijní program

Signatura tištěné verze

Umístění tištěné verze

Přístup k tištěné verzi

Klíčová slova

Atmospheric visibility, Neural network, Multimodal dataset, Meteorological variables, Deep learning, Viditelnost, neuronová síť, multimodální dataset, meteorologické veličiny, hluboké učení

Endorsement

Review

Supplemented By

Referenced By