The Issue
HadISDH.landT has been compared with other land air surface temperature monitoring products (CRUTEM4, GHCNM3, Berkeley, GISS). Agreement is generally very good although HadISDH.landT shows smaller warming trends overall. During this comparison a handful of gridboxes were highlighted that contained strong disagreement between products. On further investigation, it is clear that inhomogeneities remain in all three datasets (HadISDH, CRUTEM and GHCNM - others not compared for this assessment) at various points in time and space. Analysis of the HadISDH homogenisation at the gridbox level has revealed some oddities - mostly due to intermittent temporal sampling of stations such that the gridbox average suffers drop-in and drop-out of stations. This has resulted in spurious changes in variance and possibly means in some cases. While this does not affect conclusions drawn on large scale averages, it does show that users of data at the individual gridbox level should be careful. It also shows that Climate Data Record creation is far from a 'done deal'. There is so much more to explore and learn and this can greatly enhance our understanding of the climate system and our confidence (reduce uncertainty) in main conclusions.Bad Gridboxes
For HadISDH vs CRUTEM there are 13 gridboxes where trends are in different directions (Figure 1 - pink and red boxes) and for HadISDH vs GHCNM there are 11 (Figure 2 - pink and red boxes). These are listed below with the number of stations from each dataset in each gridbox (H=HadISDH, C=CRUTEM, G=GHCNM). Where 'H/C' is labelled this means there is poor agreement between HadISDH and CRUTEM. Where 'H/G' is labelled this means there is poor agreement between HadISDH and GHCNM. The dataset with the lowest number of stations is highlighted in bold as this suggests greater uncertainty/sensitivity in that dataset. All gridboxes with each dataset only having one station are highlighted in purple.Northern Hemisphere (20N to 70N)
-172.500W, 57.500N H/G, H=1, C=1, G=1
-167.500W, 62.500N H/G, H=2, C=2, H=1
Tropics (20S to 20N)
-157.500W, 17.500N, HG, H=1, C=2, G=1
-57.500W, 7.500N, H/C, H=2, C=7, G=11
157.500E, 7.500N, H/C, H=1, C=3, G=1
-82.500W, -7.500S H/C, H=1, C=1, G=1
177.500E, -12.500 H/G, H=1, C=2, G=2
-67.500W, -17.500S H/C, H=2, C=14, G=5
147.500E, -17.500S H/G, H/C, H=4, C=14, G=10
Southern Hemisphere (70S to 20S)
-132.500W, -22.500S H/C, H=1, C=1, G=1
147.500E, -22.500S H/C, H/G, H=1, C=13, G=13
122.500E, -27.500S H/C, H=1, C=3, G=6
112.500E, -27.500S H/C, H/G, H=2, C=2, G=3
-62.500W, -42.500S H/C, H=1, C=3, G=2
167.500E, -47.500S H/C, H/G, H=1, C=2, G=1
-67.500W, -52.500S, H/C,H/G, H=3, C=7, G=4
112.500E, -67.500S H/C, H/G, H=1, C=1, G=1
37.500E, -67.500, H/G, H=1, C=1, G=1
Gridbox -57.5W, 7.5N Case Study
The time series for the gridbox centred on -57.5W, 7.5N is shown in Figure 3. It is clear from this figure that at the gridbox level there can be considerable disagreement between datasets - CRUTEM shows no trend but GHCNM and HadISDH show significant warming trends. We may expect GHCNM and HadISDH to be quite similar as both use the Pairwise Homogenisation Algorithm (PHA) to make adjustments for inhomogeneities. HadISDH applies extra adjustments through indirect PHA utilising detected changepoints from simultaneous dewpoint depression records. CRUTEM does not apply any homogenisation but many of the stations ingested are pre-homogenised at the National Meteorological Service level. It is quite possible that inhomogeneities remain in all datasets and it looks like CRUTEM contains a large inhomogeneity beginning around 2008. However, it is not clear that the lack of any such inhomogeneity in HadISDH and GHCNM is due to homogenisation because the difference series in the lower panel does not show an obvious negative adjustment at that time. However, the HadISDH time series has generally undergone negative adjustments throughout the record.![]() |
| Figure 3 Monthly mean surface temperature anomaly time series for gridbox -57.5W, 7.5N. Data are shown from HadISDH.landT, CRUTEM4 and GHCNM. Decadal trends (and uncertainty ranges) are shown in the top left. Correlations are shown in the bottom right. The lower panel shows the difference series of HadISDH.landT (adjusted) minus HadISDH.landT (raw). This has been normalised to have a mean of zero of the last five years because no adjustments are applied in the last two years of data. The orange line is a lowess smoothed fit using 2 years of data to apply smoothing. This shows the aggregated impact of station homogenisation at the gridbox level. |
A more detailed look at the HadISDH.landT time series from both the adjusted version and the raw version shows some interesting things. Figures 4 and 5 show the raw and adjusted time series respectively. This reveals that homogenisation has actually increased the variance over two periods in particular relative to the raw version. This seems odd because any adjustments applied during homogenisation are flat (non-varying seasonally) and can only be applied with a maximum frequency of 6 months. I think what we're seeing here is temporal intermittency in the underlying station records having an affect on the gridbox average - where a gridbox average is only ingesting three or fewer stations it is very sensitive to temporal drop-in/drop-out.
![]() |
| Figure 4 Grainy ncview dump of HadISDH.landT (raw) for gridbox -57.5W, 7.5N. These are monthly mean anomalies of surface temperature. |
![]() |
| Figure 5 Grainy ncview dump of HadISDH.landT (adjusted) for gridbox -57.5W, 7.5N. These are monthly mean anomalies of surface temperature. |
HadISDH only has two stations in this gridbox where as CRUTEM has 11 and GHCNM has 7. This means that the HadISDH gridbox average is very sensitive to temporal drop-in/drop-out of any one station. Indeed, a look at the number of stations contributing to that gridbox over time shows exactly that (very intermittent temporal sampling), especially over the periods of high variance in the difference series (e.g., 1973-1982, 1999-2005) - see Figure 6. Obviously this isn't ideal as far as climate monitoring is concerned - we require long-term stability. For future versions of HadISDH we may wish to have some intermittency check - removing isolated months of data.
![]() |
| Figure 6 A rather grainy dumped image from ncview showing the contributing number of observations for each month to gridbox -57.5W, 7.5N for HadISDH.landT. |
![]() |
| Figure 7 Annual mean surface temperature time series for station 812020 (Nickerie) for the raw (red) and homogenised (blue) data and all raw neighbours within the station network (black). Decadal trends and 90% uncertainty ranges are shown for the raw (red) and homogenised (blue) series. |
![]() |
| Figure 8 Annual mean surface temperature time series for station 812250 (Zanderij) for the raw (red) and homogenised (blue) data and all raw neighbours within the station network (black). Decadal trends and 90% uncertainty ranges are shown for the raw (red) and homogenised (blue) series. |
On analysis of the monthly mean series (not shown) these adjustments do not look too outlandish but it is also clear that they are not perfect. The breakdown is listed below for each station. b = both PHA and IDPHA implemented adjustments. i = IDPHA implemented adjustments. p = PHA implemented adjustments.
Station 812020:
Start Month, End Month, Actual adjustment, Cumulative adjustment, Type
464, 504, 0.00, 0.00, b
348, 463, 0.45, 0.45, i
305, 347, 0.05, 0.50, p
1, 304, -0.50, 0.00, b
Station 812250:
Start Month, End Month, Actual adjustment, Cumulative adjustment, Type
462, 504, 0.00, 0.00, p
385, 461, 0.25, 0.25, p
320, 384, 0.50, 0.75, p
258, 319, -0.46, 0.29, i
239, 257, -0.16, 0.13, p
218, 238, 0.37, 0.50, i
1, 217, 0.10, 0.60, b
So, in summary, the main reasons for differences between HadISDH, CRUTEM and GHCNM are station selection, number of contributing stations and their temporal intermittency and also homogenisation. The large inhomogeneity in CRUTEM in 2008 is odd given that there are 11 stations contributing which I would have thought would moderate this to some extent.
Other Gridboxes of Interest
Out of the 39 gridboxes analysed in detail, either because the trend ratios were negative or the correlations were very low, a good number show interesting features. I have added the ones which I think show the most interesting things below:
![]() |
| Figure 9 - as Figure 1 |
The HadISDH gridbox time series in Figure 9 has been adjusted upwards relative to the raw time series resulting in a drastically different trend compared to CRUTEM. Similar magnitude jumps are not so apparent in CRUTEM except for the 2006 to 2011 period. This suggests CRUTEM has been adjusted or inhomogeneities were not present to begin with. Also, averaging over the 14 CRUTEM stations reduces sensitivity to inhomogeneities compared to having only 2 stations to average over in HadISDH.
Trends are very different.
Trends are very different.
![]() |
| Figure 10 - as for Figure 1. |
There is a clear and large negative adjustment applied to HadISDH between 1973 to 1980 in Figure 10. For this gridbox all datasets only have one station and it appears to be the same station. Month-to-month variability is very similar suggesting these are the same station and version of the station. So, it ooks like GHCNM has also applied an adjustment at the beginning of the series - one that is larger than for HadiSDH. It looks like CRUTEM has not had any adjustments in this case.
Trends are very different.
Trends are very different.
![]() |
| Figure 11 - as for Figure 1. |
Trends are very different.
![]() |
| Figure 12 - as for Figure 1. |
![]() |
| Figure 13 - as for Figure 1. |
![]() |
| Figure 14 - as for Figure 1. |
![]() |
| Figure 15 - as for Figure 1. |
![]() |
| Figure 16 - as for Figure 1. |
![]() |
| Figure 17 - as for Figure 1. |
![]() |
| Figure 18 - as for Figure 1. |
![]() |
| Figure 19 - as Figure 1. |
![]() |
| Figure 20 - as for Figure 1. |
![]() |
| Figure 21 - as for Figure 1. |
Conclusions
- The vast majority of gridboxes agree very well.There are fewer than 10 gridboxes with correlations lower than 0.6 for each dataset pair.
- For a handful of gridboxes large uncertainties remain at the gridbox level between products.
- The main reasons for these uncertainties are:
- different stations/versions of stations/quantity of stations making up the gridbox average
- temporal intermittency in stations and sensitivity to this in gridboxes with very few contributing stations
- homogenisation methods
- Its is likely that the homogenisation does not always do a good job but it does appear improve the majority of time series.
- Inhomogeneities are apparent in CRUTEM and GHCNM





















No comments:
Post a Comment
Note: only a member of this blog may post a comment.