<?xml version='1.0' encoding='UTF-8'?><metadata xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:dcterms="http://purl.org/dc/terms/" xmlns="http://dublincore.org/documents/dcmi-terms/"><dcterms:title>Raw data and analysis code for the study “Project 2025 as a technocratic blueprint: A corpus-based linguistic analysis of conservative governance discourse”</dcterms:title><dcterms:identifier>https://doi.org/10.60507/FK2/BK14F8</dcterms:identifier><dcterms:creator>Schilling, Julia</dcterms:creator><dcterms:creator>Fuchs, Robert</dcterms:creator><dcterms:publisher>bonndata</dcterms:publisher><dcterms:issued>2025-11-11</dcterms:issued><dcterms:modified>2025-11-11T12:47:28Z</dcterms:modified><dcterms:description>This repository contains all data and scripts used for the study “Project 2025 as a Technocratic Blueprint: A Corpus-Based Linguistic Analysis of Conservative Governance Discourse” (Schilling &amp; Fuchs, 2025).
The study investigates the language of the Heritage Foundation’s Project 2025, a 900-page conservative policy blueprint, using methods from Corpus-Assisted Discourse Studies (CADS), Political Discourse Analysis (PDA), and psycholinguistic text analysis. The corpus includes Project 2025 and Democratic and Republican Party platforms (2016–2024).

The dataset includes:
Raw data (/data/raw_data/): full texts of Project 2025 and party platforms in CSV format.

Processed data (/data/raw_data/Project2025_lemmaPOS.csv): tokenized, lemmatized, and POS-tagged text.

Keyness results (/data/keyness/): unigram and bigram keyness calculations (log-likelihood, log-ratio).

Collocation results (/data/collocations/): top 10 adjective, noun, and verb collocates per node.

LIWC results (/data/liwc/): LIWC-22 category scores for each corpus.

Analysis scripts (/code/): R Markdown file (analysis_project2025.Rmd) and two Python scripts for collocation and keyness analysis.

All files are in UTF-8 plain-text format. The dataset contains no personal, sensitive, or proprietary data and derives entirely from publicly accessible political documents.</dcterms:description><dcterms:subject>Arts and Humanities</dcterms:subject><dcterms:language>English</dcterms:language><dcterms:IsSupplementTo>Schilling, Julia &amp; Fuchs, Robert (2025). Project 2025 as a Technocratic Blueprint: A Corpus-Based Linguistic Analysis of Conservative Governance Discourse. Submitted manuscript, PLOS ONE.</dcterms:IsSupplementTo><dcterms:date>2025-11-11</dcterms:date><dcterms:contributor>Schilling, Julia</dcterms:contributor><dcterms:dateSubmitted>2025-11-05</dcterms:dateSubmitted><dcterms:license>CC BY 4.0</dcterms:license></metadata>