Version Notes
Version 3 of the Thrive by Five Index (2021) data was uploaded on 2 April 2023. Version 3 of the data is the Thrive by Five Index data merged with the ECD Baseline Audit data, as in version 1. Changes were made to Version 1 to produce Version 2 of the data, but that version was superceded by Version 3 before it could be published.
Changes in version 3 are the inclusion of a weight variable and the corrections to an item, domain and total ELOM score.
An error was discovered in the digital tablet scoring of Emergent Numeracy and Maths (ENM) item 10 (addition trial), where the programme recorded correct answers as incorrect and some incorrect answers as correct. Due to this programming error, ENM total scores were incorrectly calculated and have now been corrected. The effect of the correction has resulted in a slight increase in the score for the ENM total scores, as well as a slight adjustment of the ELOM 4&5 Total Score. The error in the tablet form has been corrected and no further errors have been detected. It is important to note that only children whose scores were close to one of the ELOM standard score bands would be likely to be misclassified. Data files in the public domain dataset hosted by DataFirst have been corrected.
Version 4 has corrections to the weights in the data the additional variable 'quintile_natemis' which is described below (from page 38 of the Thrive by Five Technical Report).
Description of the 'quintile_natemis' Variable
"The quintile rank system is a valuable proxy for the wealth status of the children assessed and was used for stratification during the sampling process, as well as for disaggregation in the analysis. There are three approaches that can be used in assigning a quintile status to an ELP:
Using the quintile status of the primary school that was used when constructing the sample (quintile_original);
Using the quintile status of the closest primary school in the school sample used to identify clusters (quintile_sample); or
Using the quintile status of the closest school in the DBE 2021 Masterlist data (quintile_natemis).
The research team decided to use the quintile status of the primary schools that was used to construct the sample (quintile_original) for both the construction of the weights, as well as for disaggregation.For the construction of weights, the quintile_original variable is most appropriate because it determined the probability of an ELP having been sampled. For disaggregation, the variable quintile_original was deemed the most conservative choice of classification to use, since it will not introduce any additional measurement error that cannot be accounted for (for example, the closest school being in a typical quintile 5 area, while the ELP is in a neighbouring Q3 area)."