Appendix C — Part III Data Appendix: Press Corpus and Documentary Sources

This appendix is reserved for the data documentation supporting Part III’s process-tracing analysis: the Folha de São Paulo press corpus, legislative records, and other documentary sources. It is designed to travel with Part III if that chapter is developed into a standalone article, in parallel with the role Section B.3’s appendix plays for Part II (see Section A.1’s index). Detailed content — a corpus-level codebook (fields, coding categories, search criteria) and a sourcing log for legislative and official documents — is forthcoming.

C.1 Press Corpus (Folha de São Paulo)

Folha de São Paulo press corpus (1994–2024). A systematically collected corpus of newspaper articles from the Folha de São Paulo digital archive, gathered using a custom R scraper pipeline (4-DA-Code/Scraper/). The corpus covers tertiary education policy coverage and was used for qualitative analysis in Part III. A full codebook (fields, coding categories, search criteria, and any inter-coder reliability checks) will be added here.

C.2 Legislative and Documentary Sources

To be completed: a sourcing log for legislative records and official documents cited in Part III (law number, date, official gazette reference), following the same transparency standard as the press corpus.