Aggregating signals of Earth system dynamics across space, time, models, and variables
Abstract. Model Intercomparison Projects (MIPs) provide standardised computer simulations of the Earth system, offering unique opportunities to systematically detect and assess features and dynamics, such as abrupt shifts, across diverse models and variables. Recent advances combine time-series analysis with spatiotemporal clustering to identify dynamically connected regions within individual datasets. Yet, extending this notion of connectivity across the model and variable dimensions of MIP output remains an open challenge. Here, we present a conceptual workflow that addresses this by introducing two aggregation strategies for "detect-then-cluster" pipelines: "Detect-Cluster-Aggregate-Cluster" (DCAC) and "Detect-Aggregate-Cluster" (DAC), enabling systematic synthesis of spatiotemporal signals across multiple datasets. These aggregation algorithms are evaluated and tuned using a customisable Analytic Hierarchy Process (AHP) framework, which allows users to encode prior knowledge about dataset reliability. In anticipation of output from the Tipping Points Modelling Intercomparison Project (TIPMIP) and other MIPs within the Coupled Model Intercomparison Project (CMIP), we implement the proposed aggregation methods using the "Tipping and Other Abrupt Events Detector" (TOAD) package. To demonstrate feasibility, we apply the methods to CMIP6 simulations of Amazon rainforest dynamics, detecting and clustering abrupt vegetation shifts first across multiple variables, where a shared signal indicates a coherent ecosystem response, and then across multiple models, where a shared signal reflects model alignment. Our case study reveals that this aggregation helps distinguish such shared behaviour from dynamics that are specific to individual variables or models, patterns that typically remain obscured when datasets are analysed in isolation. These results illustrate that conclusions about abrupt dynamics depend critically on how information is synthesised across time, space, models, and variables. While showcased here in the context of tipping points, the proposed aggregation framework provides a structured and transferable foundation for multimodel and multivariate risk assessments of diverse Earth-system processes within MIPs.
The paper by De Maeyer et al “Aggregating signals of Earth system dynamics across space, time, models, and variables” proposes several methods of combining complex climate data for studying tipping points, with an example of Amazon rainforest.
I find the paper structure strange and unsuitable for this journal: the metrics and relevant plots and tables are placed in the appendix; this is typical for Nature submissions, and similar journals, but in fact not easy to follow, and better to change into regular structure: Intro – Method – Data – Results.
I find the title of the paper misleading: the paper is not only about aggregation, and space and time are included in the variables that are mentioned along those. Furthermore, aggregation may be not the best term here. I would call the approach assemblage: aggregation usually leads to a single statistics, speaking mathematically - which is what is shown in section 2.1 but is more narrow than what the title states. In Fig.1, aggregation is mentioned as an alternative of clustering, and this contradicts the paper title.
The meaning of “aggregation” (assemblage) should be explained much earlier than it currently is in page 4. Similarly, clusters are mentioned early but explained much later, and this complicates reading of the paper. When clustering is mentioned, it is necessary to explain whether it is spatial or temporal; later there are “spatiotemporal” clusters – this requires clarification. Clustering is discussed in general terms for many pages, and only in page 14 the DBSCAN technique is mentioned.
As I understand it, by “detection time series” (time series analysis of dynamics of interest) the authors mean, in particular, early warning signal indicators – why is it not mentioned explicitly?
In line 200, relative importance vector is introduced but a reference is not provided.
In section 4.2, the processed data represents 150 annual mean datapoints. Considering the context of “abrupt shifts”, annual resolution may be too crude. Also, it would be useful to include an actual vegetation map of the current state of the Amazon rainforest. This would be interesting to compare with the three clusters in Fig.5.
FURTHER COMMENTS
> why sorting is mentioned in Eq.1? Median is a standard metric
> acronyms DAC and DCAC are defined multiple times
> the caption of Fig.2 should mention what data are analysed
> in line 135, what values are meant as “high counts”?
> Fig.3, panel (c) – are frequencies normalised? From the x-range between 0 and 100, this is unlikely. Then why y-axis has label “cluster frequency”? The caption mentions “counts” but the values are below 1 – what processing was applied to integer counts?
> In Eqs.7,8, only three datasets are used – isn’t it a formula for an arbitrary number of datasets and should be expressed by a general sum?
> Tables 1 and D1 can be merged, and the order of their lines can be rearranged according to the decreasing weights
MINOR COMMENTS
> In the 4th affiliation, System should start with a capital letter. The same in line 21 of the main text.
> in line 184, word “priority” can be omitted (a priory is sufficient)
> in Table 1, 4th line from bottom, CO2 should be with 2 as index
> line 381, clarify the meaning of “a clear, and consistently abrupt” – if this was meant as an Oxford comma, that is incorrect.
> in the caption of Fig.6, panel (c), explain that the colours of the curves correspond to the colours of the clusters.
> line 493, 564, 568 start with a dot