SEAM technical appendix.pdf

Technical Report

The SEAM benchmarks and items were identified from the literature on social-emotional development of young children raised in mainly Western cultures. Certain concepts emerged as essential to the mental health competence of young children (Squires & Bricker, 2007). These benchmarks were reviewed and revised based on feedback from family members and experts in related fields.

Research Questions

  1. What is the item functioning for the Infant and Toddler Intervals?
  2. What is the reliability of the Infant, Toddler, and Preschool Intervals, including internal consistency, test–retest, and interrater reliability?
  3. What is the validity of the Infant, Toddler, and Preschool Intervals, specifically content and congruent validity?

Data Collection

Data were gathered from various caregivers across the United States using both pencil-and-paper and online formats, collecting a total of 2,201 SEAMs, with 1,850 collected online and 351 via paper.

Sample Demographics

Psychometric Properties

Data Analysis Techniques

Data analysis included item response theory (IRT) modeling to examine item functionality and classical test analyses for validity and reliability studies. The item fit statistics generated during these analyses indicated most selected models functioned well across different assessment modes.

Statistical Results

Age Interval Benchmark Infit mean MNSQ(SD) Outfit mean MNSQ(SD)
Infant 1.0 0.98(0.19) 0.95(0.17)
Toddler 1.0 1.00(0.22) 1.00(0.23)
Preschool 1.0 0.99(0.14) 0.92(0.18)

Conclusions

The SEAM system shows promise based on initial results demonstrating valid, reliable, and useful assessments of social-emotional development. Future research is necessary to confirm these findings on a larger scale.