Core Facility for Neuroscience of Self-Regulation, Heidelberg University
Heidelberg, Germany
NCT Number: NCT06511102
The goal of this randomized controlled lab experiment is to examine if using generated artificial intelligence (AI) technology will affect people's academic performance and cognitive abilities in the context of analytical writing among college students. The main questions it aims to answer are:
1. Does using the technology affect students' writing performance? 2. Does using the technology affect students' cognitive effort during the writing process?
Participants will be randomly assigned to either a control group, which is writing without AI assistance, or an experimental group, which is writing with the assistance of ChatGPT. Researchers will compare the two groups to see if ChatGPT affects students' writing performance and cognitive effort.
For each participant, the lab experiment will last for no more than 1.5 hours. An eye-tracker will monitor the participant's gaze activities and pupil size. A functional near-infrared spectroscopy (fNIRS) will monitor the participant's brain activities in the frontal lobe. During the experiment, participants will be asked to:
1. Read learning materials on analytical writing techniques. 2. Based on the previously provided materials, complete an analytical writing assignment that will take approximately 30 minutes either with or without the aid of ChatGPT. 3. Answer survey questions about their experience with the writing assignment, attitudes on using ChatGPT, and demographic backgrounds.
Looking for future studies?
Notify Me18 year–35 year
All sexes
Interventional
Not applicable
Heidelberg, Germany
Healthy volunteers accepted: Yes
Only the study team can determine whether someone qualifies for participation.
Inclusion criteria
The computer interface used for the essay writing task follows a split-screen design. The writing instructions and text input field are administered on a survey platform, placed on the left half of the screen. ChatGPT is placed on the right half of the screen for technology assistance.
Time frame: 1.5 hours
The essay writing task is derived from the Analytical Writing section in the Graduate Record Examinations (GRE), which is a worldwide and standardized computer-based exam developed by the Educational Testing Service (ETS). The participants' essays will be scored on a scale from 0 to 6 by an automatic and validated third-party scoring tool that is also developed by ETS.
Time frame: 1.5 hours
Cognitive effort is quantified by monitoring changes in pupil size. To achieve this, pupil diameters are recorded throughout the writing task using a near-infrared eye tracker, specifically the Tobii Pro Fusion model. At the start of the experiment, individual baseline pupil diameters are measured during a 30-second relaxation task.
Time frame: 1.5 hours
This is a one-item scale:
Using the same grading rubric from before, what score do you think your essay should get (0 being the lowest and 6 being the highest)?
The score ranges from 0 to 6. A higher score indicates higher self-perceived writing performance. The variable is treated as a continuous variable.
Time frame: 1.5 hours
This is a one-item Likert scale adapted from the National Aeronautics and Space Administration-task load index (NASA-TLX; Hart, 2006; Hart & Staveland, 1988):
On a scale of 1 to 7, rate how hard you have to work to accomplish your level of performance.
The Likert score ranges from 1 to 7 (1 being "very low" and 7 being "very high"). A higher score indicates higher self-perceived cognitive effort. The variable is treated as a continuous variable.
References:
Time frame: 1.5 hours
Cognitive Effort is quantified by monitoring changes in the cortical hemodynamic activity in the frontal lobe. To achieve this, the brain activity is recorded throughout the writing task using a functional near-infrared spectroscopy (fNIRS), specifically the NIRSport2 model.
Time frame: 1.5 hours
This is a one-item Likert sub-scale adapted from the Primary Appraisal Secondary Appraisal scale (PASA; Gaab, 2009; Pollak et al., 2020):
On a scale of 1 to 7, how much would you agree or disagree with the following statement on perceived stress: The analytical writing assignment was stressful to me.
The Likert score ranges from 1 to 7 (1 being "strongly disagree" and 7 being "strongly agree"). A higher score indicates higher self-perceived stress. The variable is treated as a continuous variable.
References:
Time frame: 1.5 hours
This is a one-item Likert sub-scale adapted from the Primary Appraisal Secondary Appraisal scale (PASA; Gaab, 2009; Pollak et al., 2020):
On a scale of 1 to 7, how much would you agree or disagree with the following statement on perceived challenge: I find the analytical writing assignment a challenge.
The Likert score ranges from 1 to 7 (1 being "strongly disagree" and 7 being "strongly agree"). A higher score indicates higher self-perceived challenge. The variable is treated as a continuous variable.
References:
Time frame: 1.5 hours
This is a sixteen-item Likert scale that measures three dimensions of writing self-efficacy: ideation, convention and self-regulation (Bruning et al., 2013). The Likert score ranges from 1 to 7 (1 being "strongly disagree" and 7 being "strongly agree"). A higher score indicates higher self-efficacy. The three dimensions will be treated separately, each as a continuous variable.
Reference:
Time frame: 1.5 hours
This is a four-item Likert scale adapted from the situational interest scale (Hulleman et al., 2010). This scale measures participants' situational interest in analytical writing:
On a scale of 1 to 7, how much would you agree or disagree with the following statements on your interest in the analytical writing assignment that you just completed?
The Likert score ranges from 1 to 7 (1 being "strongly disagree" and 7 being "strongly agree"). A higher score indicates higher situational interest. The variable is treated as a continuous variable.
Reference:
Time frame: 1.5 hours
This is a two-item Likert scale that measures participants' behavioral intention in using ChatGPT in the future for essay writing tasks (Albayati, 2024):
On a scale of 1 to 7, how much would you agree or disagree with the following statements on using ChatGPT in essay writing assignments?
The Likert score ranges from 1 to 7 (1 being "strongly disagree" and 7 being "strongly agree"). A higher score indicates higher behavioral intention in using ChatGPT. The variable is treated as a continuous variable.
Reference:
University Hospital Heidelberg
Other
A Randomized Controlled Trial on the Impact of Using Generative Artificial Intelligence on Analytical Writing Performance and Cognitive Abilities
OpenTrials presents study information sourced from ClinicalTrials.gov. The official registry record should be consulted for the latest information.
View the official ClinicalTrials.gov record (opens in a new tab)This listing is for discovery and informational purposes only. It is not medical advice, does not guarantee that a study is recruiting, and does not determine eligibility. Contact the study team and a qualified healthcare professional when considering participation.
Published trials that share one or more normalized conditions with this study.
NCT06107075
Behavior, Cognitive Change
Newcastle upon Tyne, United Kingdom
View Trial DetailsNCT06168526
Behavior, Cognitive Change
Castellon, Castellón, Spain
View Trial DetailsNCT06312956
Apnea, Behavior
Piancavallo, VCO, Italy
View Trial DetailsNCT06721468
Behavior, Cognitive Change
Moscow, Idaho, United States
View Trial Details