Accurate predictions of life outcomes can inform social theory and policy. Yet, social science predictions often perform poorly. One common explanation is small sample sizes of social datasets. We examine the predictability of having a child within three years under conditions that are close to the best currently possible. We use full-population Dutch registers and LISS survey data, linkable to the registers, in the Predicting Fertility data challenge. LISS data includes a wider range of theoretically relevant variables, including subjective measures such as fertility intentions. Register data provides vast, high-resolution coverage of life-course trajectories but lacks subjective measures. This unique framework enables us to leverage the strengths of these datasets to assess the current predictability. Over 150 people participated in the data challenge and submitted over 70 models, both traditional machine learning and cutting-edge approaches. Predictive performance remained modest for both datasets: even the best models fell below the upper limit of predictability caused by randomness in conception and fetal survival. Survey-based predictions performed slightly better. Combining
📖 افتح في inklap 🔗 DOI 📮 اطلب بحثاً