21c7e211d012b9b3b259ab8db1420f15a0efacbb
Implements only what reviewer 1 and reviewer 2 independently and concordantly required, at the build sources rather than in generated files: - NSDUH 2022/2023/2024 YUSUITHK: strip the codebook editorial note about the 2022 section move from question_text. It was never shown to respondents. The removed text is kept verbatim in notes. - NSDUH 2021 YUSUITHK: drop the stray space before the final question mark. - Solomon Islands items: review_status provisional -> needs_source. No country questionnaire, codebook or fact sheet ships under Dataset/; the only citation is the generic GSHS data dictionary. Keyed off the missing instrument, not hardcoded per country. Also downgrades two design claims that A-20260920-067 verified as unsupported: Jamaica and Wallis and Futuna report a census with a 100% school response rate but carried the same Taylor-linearization plan as the two-stage cluster samples. design_status becomes census_variance_treatment_unresolved and two conflict records are added. variance_method itself is untouched - that is a Stage 4 statistical decision needing its own evidence. Deliberately not done: the instruction sentence missing from the 2021 stem is left absent rather than written in, since that would be composing official question text; and the 84 adjudications remain human work. Only 21 of 84 rows are concordant, so 63 need adjudication. 84 item_version_id unchanged, response_dictionary byte-identical, both signed reviewer CSVs unchanged. Gate 1 audit passes all seven machine deliverables and stays blocked solely on two_actual_human_reviewers. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Description
Reproducible research project for language-conditioned psychometric harmonization across surveys.
2.5 MiB
Languages
Python
86.8%
JavaScript
8.5%
R
3.2%
PowerShell
0.7%
Cuda
0.5%
Other
0.3%