Skip to content

Analysis Code and Reproducibility Materials for "Measuring Cognitive Surrender: Early-Warning Signals in AI-Mediated Education"

Sep 2026 · Zenodo (CERN European Organization for Nuclear Research)

Abstract

This repository contains the analysis code, derived results, figures, tables, and reproducibility materials supporting the manuscript “Cognitive Surrender in AI-Mediated Education: Behavioral Reliance, Prospective Signals, and Governance for Human Cognitive Agency.” The study examines reliance-oriented behavioral patterns in student-Large Language Model (LLM) interactions using the publicly available StudyChat dataset. It analyzes 16,851 student-LLM interactions from 203 users across 2,214 chat sessions and investigates whether observable interaction traces can characterize patterns of AI-mediated reliance. The repository includes the analysis pipeline used to construct and evaluate the prototype Reliance Index (RI), an exploratory behavioral composite designed to summarize interaction patterns associated with output seeking, limited contextualization, escalation from hints toward direct solutions, repeated prompting, code/data conversion, and limited observable evidence of prior independent effort. The RI is intended as a research instrument, not a diagnostic or psychometric measure of cognitive surrender. The archived analyses include cleaned intent distributions, RI distributions across users and StudyChat topics, comparisons between the RI and an independently constructed Struggle Index, robustness and sensitivity analyses, and a temporally separated prospective analysis examining whether features observed during the first 20% of each user’s interactions contain information associated with reliance-oriented behavior during the subsequent observation period. RI robustness was evaluated using equal weighting, Monte Carlo perturbation of component weights, and leave-one-component-out specifications. Under equal weighting, the RI showed a Spearman rank correlation of 0.953 with the baseline specification. Across 1,000 random ±20% weight perturbations, the median Spearman correlation was approximately 0.995, with a 5th-95th percentile range of approximately 0.980-0.999. Leave-one-component-out analyses indicated that the RI was more sensitive to removal of some behavioral components, particularly low-context prompting and write-code requests, than to moderate changes in numerical weighting. The prospective analysis used temporally non-overlapping predictor and outcome periods. Logistic Regression achieved an AUC of 0.584 and Random Forest an AUC of 0.598 for identifying users in the highest quartile of subsequent-period RI. Feature-set ablation showed that interaction-volume features alone performed approximately at chance, while reliance-oriented and combined behavioral features provided modest prospective discrimination. These results are interpreted as preliminary evidence of temporal structure in interaction behavior rather than as support for individual-level prediction or automated intervention. The repository is organized into directories containing the final analysis code, derived result tables, analysis summaries, reproducibility outputs, and manuscript figures. The included README provides instructions for reproducing the analyses. The original StudyChat dataset is not redistributed in this archive. Users should obtain the dataset from its original repository and place the downloaded data file in the location specified in the README before running the analysis pipeline. Use of the original dataset remains subject to its applicable licensing and usage conditions. This archive accompanies the manuscript submitted for scholarly publication and is intended to support transparency, reproducibility, and independent verification of the reported analyses.

View source

Similar papers

#computer vision Open access Jun 2016

Software Development in Startup Companies: The Greenfield Startup Model

The results are packaged in the Greenfield Startup Model (GSM), which explains the priority of startups to release the product as quickly as possible, and the need to shorten time-to-market, by speeding up the development through low-precision engineering activities.

Carmine Giardino, Nicolò Paternoster, M. Unterkalmsteiner et al. · 178 citations · ⚡14
#computer vision Open access Oct 2016

Software Startups - A Research Agenda

Software startup companies develop innovative, software-intensive products within limited timeframes and with few resources, searching for sustainable and scalable business models.

M. Unterkalmsteiner, P. Abrahamsson, Xiaofeng Wang et al. · 157 citations · ⚡17
#machine learning Review Open access Oct 2016

“Failures” to be celebrated: an analysis of major pivots of software startups

This study conducts a case survey study based on the secondary data of the major pivots happened in 49 software startups, and demonstrates that customer need pivot is the most common among all pivot types.

Sohaib Shahid Bajwa, Xiaofeng Wang, Anh Nguyen-Duc et al. · 127 citations · ⚡15
#computer vision Review Open access May 2015

A survey study on major technical barriers affecting the decision to adopt cloud services

The comparison of adopter and non-adopter sample reveals three potential adoption inhibitor, security, data privacy, and portability, which underlines the importance of the technical and security perspectives for research investigating the adoption of technology.

Nattakarn Phaphoom, Xiaofeng Wang, S. Samuel et al. · 111 citations · ⚡8
#computer vision Conference Open access Dec 2013

Affordable and Energy-Efficient Cloud Computing Clusters: The Bolzano Raspberry Pi Cloud Cluster Experiment

The ongoing work building a Raspberry Pi cluster consisting of 300 nodes is presented, with potential use cases being an inexpensive and green test bed for cloud computing research and a robust and mobile data center for operating in adverse environments.

P. Abrahamsson, S. Helmer, Nattakarn Phaphoom et al. · 110 citations · ⚡7
#computer vision Book Open access Mar 2017

On the Unhappiness of Software Developers

The results indicate that software developers are a slightly happy population, but the need for limiting the unhappiness of developers remains, and 219 factors representing causes of unhappiness while developing software are identified.

D. Graziotin, Fabian Fagerholm, Xiaofeng Wang et al. · 84 citations · ⚡6

Related blog posts

MIT News · Artificial Intelligence Sep 14, 2026

New method enables AI for safety-critical situations

The “HardFlow” algorithm could help generative AI models produce high-quality outputs that obey strict requirements when “pretty close” doesn’t cut it.

GPT-Lab Sep 10, 2026

Responsible AI Must Consider Its Afterlife

AI may appear weightless, but every model depends on physical infrastructure. To understand responsible AI, we need to look beyond algorithms and consider the entire lifecycle of the hardware behind them. The post Responsible AI Must Consider Its Afterlife appeared first on GPT-Lab.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.