SearcharxivSearch

arXiv · 1210.7191

HadISD: a quality-controlled global synoptic report database for selected variables at long-term stations from 1973--2011

Abstract

[Abridged] This paper describes the creation of HadISD: an automatically quality-controlled synoptic resolution dataset of temperature, dewpoint temperature, sea-level pressure, wind speed, wind direction and cloud cover from global weather stations for 1973--2011. The full dataset consists of over 6000 stations, with 3427 long-term stations deemed to have sufficient sampling and quality for climate applications requiring sub-daily resolution. As with other surface datasets, coverage is heavily skewed towards Northern Hemisphere mid-latitudes. The dataset is constructed from a large pre-existing ASCII flatfile data bank that represents over a decade of substantial effort at data retrieval, reformatting and provision. These raw data have had varying levels of quality control applied to them by individual data providers. The work proceeded in several steps: merging stations with multiple reporting identifiers; reformatting to netCDF; quality control; and then filtering to form a final dataset. Particular attention has been paid to maintaining true extreme values where possible within an automated, objective process. Detailed validation has been performed on a subset of global stations and also on UK data using known extreme events to help finalise the QC tests. Further validation was performed on a selection of extreme events world-wide (Hurricane Katrina in 2005, the cold snap in Alaska in 1989 and heat waves in SE Australia in 2009). Although the filtering has removed the poorest station records, no attempt has been made to homogenise the data thus far. Hence non-climatic, time-varying errors may still exist in many of the individual station records and care is needed in inferring long-term trends from these data. A version-control system has been constructed for this dataset to allow for the clear documentation of any updates and corrections in the future.

Explore related subjects

Keep this discovery

BibTeXRIS

Robert J. H. Dunn, Kate M. Willett, Peter W. Thorne, Emma V. Woolley, Imke Durre, Aiguo Dai, David E. Parker, Russ E. Vose. 2012-10-26. HadISD: a quality-controlled global synoptic report database for selected variables at long-term stations from 1973--2011. https://doi.org/10.5194/cp-8-1649-2012

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Windowed Envelope Statistics for Time-Domain Significant Wave Height Estimation From HF Radar

Significant wave height (SWH) retrieval from high-frequency (HF) radar typically relies on a weak second-order Doppler continuum that is sensitive to noise, interference, and spectral leakage. This letter presents a Windowed Envelope Statistics Estimator (WESE) that operates directly on beam-formed time-domain voltages. A second-order term obtained from a Neumann expansion of the rough-surface field equation motivates quadratic compensation of localized radar features. WESE extracts the mean, standard deviation, or variance from overlapping windows of the in-phase, quadrature, or envelope-magnitude sequence, followed by quadratic compensation, rank ordering, least-squares regression, and causal smoothing. Evaluation used 335 synchronized hourly observations from a 13.385 MHz, 12-element WERA system at Argentia, Newfoundland and Labrador. The optimal configuration used quadrature variance, a 16-sample window, 896 retained chronological samples, and 30-h smoothing, achieving an RMSE of 0.152 m and a Pearson correlation of 0.978. This represents RMSE reductions of 32.1% and 18.7% relative to previously reported linear and second-order compensated ordered-statistics models, respectively. The results demonstrate robust time-domain SWH estimation without explicit Doppler-spectrum construction.

physics.ao-ph

KiloDA: Reconstructing kilometer-scale near-surface wind states from sparse station observations

Accurate kilometer-scale near-surface winds are important for understanding atmospheric processes over complex terrain, yet remain difficult to reconstruct from sparse and unevenly distributed observations. Here we introduce KiloDA, a diffusion framework for hourly kilometer-scale wind reconstruction from surface stations. KiloDA learns the statistical distribution and spatial structure of wind fields from historical 3-km Weather Research and Forecasting (WRF) model forecasts. At each reconstruction time, no contemporaneous WRF field is used. Instead, station observations provide the only constraints on the current atmospheric state and guide posterior sampling from the learned prior. In idealized WRF experiments, KiloDA recovers localized wind structures when only 0.24% of grid cells are observed and shows an overall advantage over conventional interpolation across terrain conditions and wind speed regimes. This capability largely transfers to real observations. In a fully withheld region, KiloDA reduces the median wind speed root mean square error (RMSE) by 19% relative to ERA5 reanalysis, using only observations outside the region, with the largest improvements over high-elevation and high-relief terrain. A random station holdout further confirms that this advantage extends across different complex-terrain locations and holdout configurations. These results show that historical model archives can provide useful structural knowledge for reconstructing kilometer-scale wind fields from sparse observations without requiring an accurate model estimate of the current atmospheric state.

physics.ao-ph