Text Corpus of Speeches in German State Parliaments

**StateParl** comprises parliamentary speeches by members of parliament and government representatives from all 16 German state parliaments. The dataset contains over **16 million speech paragraphs** from plenary session transcripts covering **2000–2025**. It provides a coherent, machine-readable corpus with resolved speaker identities, grouped speeches, and rich metadata, enabling systematic computational analysis across disciplines. 

Scroll to Top