Études fondées sur les communautés Reddit

Stakeholder perspectives on generative AI in education: A dataset of Reddit posts and comments (2022-2025).

Data Brief . 2026;66 :112871

Résumé

This data article describes a dataset capturing public discourse on generative AI (GenAI) in education, collected from Reddit between 1 September 2022 and 31 October 2025. The dataset comprises 984 unique posts and 8346 associated comments drawn from 320 education-related subreddits, grouped into three stakeholder categories: students, educators, and parents. The posts and comments focus on how GenAI tools, such as ChatGPT, are used, experienced, and debated in relation to teaching, learning, assessment, and broader educational practices. Data were retrieved via the Reddit API and processed through a privacy-preserving pipeline that removed personal identifiers in accordance with platform policies. Each entry is accompanied by rich metadata, including a unique identifier, content type (post or comment), stakeholder group, subreddit, creation timestamp, and engagement metrics. Spanning the period before and after the introduction of ChatGPT, the dataset enables temporal analyses of discourse volume and content evolution. This openly available dataset provides a valuable resource for research in education, human-AI interaction, social media analytics, and natural language processing, and supports reproducibility in studies of public engagement with GenAI technologies.

Tous les articles