<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMC</journal-id>
<journal-id journal-id-type="nlm-ta">CMC</journal-id>
<journal-id journal-id-type="publisher-id">CMC</journal-id>
<journal-title-group>
<journal-title>Computers, Materials &#x0026; Continua</journal-title>
</journal-title-group>
<issn pub-type="epub">1546-2226</issn>
<issn pub-type="ppub">1546-2218</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">79379</article-id>
<article-id pub-id-type="doi">10.32604/cmc.2026.079379</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Codenote: Leveraging AI-Driven Personality Grouping to Foster Students&#x2019; Coding Self-Efficacy</article-title>
<alt-title alt-title-type="left-running-head">Codenote: Leveraging AI-Driven Personality Grouping to Foster Students&#x2019; Coding Self-Efficacy</alt-title>
<alt-title alt-title-type="right-running-head">Codenote: Leveraging AI-Driven Personality Grouping to Foster Students&#x2019; Coding Self-Efficacy</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Lin</surname><given-names>Jia-Rou</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Tseng</surname><given-names>Chun-Hsiung</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><xref rid="cor1" ref-type="corresp">&#x002A;</xref><email>lendle@saturn.yzu.edu.tw</email></contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Lin</surname><given-names>Hao-Chiang Koong</given-names></name><xref ref-type="aff" rid="aff-2">2</xref></contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Huang</surname><given-names>Andrew Chih-Wei</given-names></name><xref ref-type="aff" rid="aff-3">3</xref></contrib>
<aff id="aff-1"><label>1</label><institution>Department of Electrical Engineering, Yuan Ze University</institution>, <addr-line>Taoyuan</addr-line>, <country>Taiwan</country></aff>
<aff id="aff-2"><label>2</label><institution>Department of Information and Learning Technology, National University of Tainan</institution>, <addr-line>Tainan</addr-line>, <country>Taiwan</country></aff>
<aff id="aff-3"><label>3</label><institution>Department of Psychology, Fo Guang University</institution>, <addr-line>Yilan</addr-line>, <country>Taiwan</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Chun-Hsiung Tseng. Email: <email>lendle@saturn.yzu.edu.tw</email></corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic">
<year>2026</year>
</pub-date>
<pub-date date-type="pub" publication-format="electronic">
<day>8</day><month>5</month><year>2026</year>
</pub-date>
<volume>88</volume>
<issue>1</issue>
<elocation-id>31</elocation-id>
<history>
<date date-type="received">
<day>20</day>
<month>01</month>
<year>2026</year>
</date>
<date date-type="accepted">
<day>12</day>
<month>03</month>
<year>2026</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2026 The Authors. Published by Tech Science Press.</copyright-statement>
<copyright-year>2026</copyright-year>
<copyright-holder>The Authors</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMC_79379.pdf"></self-uri>
<abstract>
<p>Effective pair programming relies heavily on optimal partner compatibility, a requirement that is often difficult to scale manually in software engineering education. This study presents the empirical validation of Codenote, an AI-driven Integrated Development Environment (IDE) designed to automate personality-aware group formation. By integrating a behavioral analysis mechanism, Codenote infers student personality traits from coding patterns to construct complementary pairs, thereby facilitating intelligent collaborative learning. To validate the system&#x2019;s effectiveness, a controlled experiment was conducted to assess the impact of this AI-mediated pairing strategy on students&#x2019; self-efficacy across adaptive, innovative, and persuasive domains. Results indicate that students paired via Codenote&#x2019;s complementary algorithm demonstrated significantly higher adaptive self-efficacy compared to the control group. This suggests that the system&#x2019;s ability to expose learners to diverse problem-solving perspectives effectively enhances their confidence in managing complex and dynamic programming tasks. While no significant improvements were observed in innovative or persuasive self-efficacy, these findings identify specific directions for future system iterations, such as integrating automated scaffolding for creative exploration and leadership. Overall, this study demonstrates the viability of Codenote as an intelligent tool for scaling personalized instruction and highlights the crucial role of AI-driven pairing strategies in fostering psychological readiness for collaborative problem solving.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Codenote</kwd>
<kwd>pair programming</kwd>
<kwd>personality-aware grouping</kwd>
<kwd>self-efficacy</kwd>
</kwd-group>
<funding-group>
<award-group id="awg1">
<funding-source>National Science and Technology Council (NSTC)</funding-source>
<award-id>113-2410-H-155-023</award-id>
<award-id>114-2410-H-155-012-MY2</award-id>
</award-group>
</funding-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>In recent years, programming education has become an essential component of modern curricula due to the increasing importance of computational thinking and problem-solving skills. Programming languages are not only tools for human computer interaction but also form the foundation of systematic and logical reasoning. In Taiwan, educational reforms such as the 12-Year Curriculum Guidelines explicitly emphasize computational thinking and problem-solving abilities to prepare students for real-world challenges [<xref ref-type="bibr" rid="ref-1">1</xref>]. Integrating computational thinking into curriculum design has been shown to enhance students&#x2019; problem-solving capabilities and deepen their understanding of programming concepts [<xref ref-type="bibr" rid="ref-2">2</xref>]. Despite these advancements, programming remains a highly abstract and challenging subject for most students. Effective learning requires the simultaneous mastery of syntax memorization, comprehension of example code, completion of practice tasks, and interpretation of immediate feedback from development environments [<xref ref-type="bibr" rid="ref-3">3</xref>]. In large-class settings, students who encounter difficulties without timely support often fall behind, creating a negative cycle of frustration, disengagement, and reduced learning outcomes. To address these challenges, instructional strategies that promote hands-on, collaborative learning have received increasing attention. To address the growing demand for adaptive learning environments, we developed Codenote, which is a personality-aware Integrated Development Environment (IDE) designed with the functionality of AI-powered pair programming.</p>
<p>Pair programming has emerged as a prominent approach for supporting collaborative learning in programming. Originally developed as an agile software development practice, pair programming involves two programmers working together on the same computer to improve code quality and reduce errors. In educational contexts, studies have shown that pair programming reduces students&#x2019; frustration and cognitive load while also alleviating instructors&#x2019; teaching burdens, because students no longer rely solely on instructors as their source of guidance [<xref ref-type="bibr" rid="ref-4">4</xref>,<xref ref-type="bibr" rid="ref-5">5</xref>]. However, challenges such as inappropriate pairing, unequal participation, task difficulty mismatches, and classroom constraints may limit its effectiveness [<xref ref-type="bibr" rid="ref-6">6</xref>]. Research also suggests that pairing students based on complementary personality traits can enhance learning outcomes, particularly in remote collaboration contexts [<xref ref-type="bibr" rid="ref-7">7</xref>].</p>
<p>A major limitation in implementing personality-informed strategies lies in the reliance on traditional self-report questionnaires, which are often time-consuming and prone to low-quality or inattentive responses. To address this limitation, the present study proposes inferring personality traits from students&#x2019; operational behaviors and note-taking activities. Prior studies have suggested that machine learning approaches can be applied to analyze note-taking patterns in order to infer individual personality traits [<xref ref-type="bibr" rid="ref-8">8</xref>]. In addition, Meidenbauer et al. and Buker et al. examined mouse movement patterns and keystroke dynamics as behavioral indicators of personality. Their findings indicate that features such as mouse pauses, clicks, and typing dynamics are significantly associated with traits including Conscientiousness and Agreeableness [<xref ref-type="bibr" rid="ref-9">9</xref>]. Given that programming practice inherently involves frequent note-taking and code annotation, these observable behaviors provide a practical and unobtrusive source of data for inferring students&#x2019; personality traits.</p>
<p>Based on the aforementioned research insights, this study develops a personality trait based pair programming framework supported by a pair programming assistant system to implement a one-to-many pairing model in classroom settings. Students are grouped according to inferred personality traits and engage in collaborative learning with the support of the assistant software. This design aims to enhance the scalability of instructional practices while preserving the advantages of pair programming in promoting learning interaction and conceptual understanding. The primary objective of this study is to examine the impact of a personality-informed, assistant-supported pair programming model on students&#x2019; programming competence, thereby providing empirical evidence for the development of more adaptive and effective programming learning environments. In addition, this study adopts self-efficacy as a core validation indicator to investigate whether personality-based pair programming can effectively enhance programming learning outcomes. Prior research has indicated a significant positive correlation between self-efficacy and academic achievement in programming courses. Mart&#x00ED;nez-Mej&#x00ED;a and Rodr&#x00ED;guez-Villanueva further found that self-efficacy and learning motivation have positive effects on academic performance and contribute to the development of problem-solving skills and creativity among programming learners [<xref ref-type="bibr" rid="ref-10">10</xref>]. To empirically validate these effects within an intelligent environment, the present study implements the proposed framework using the Codenote system.</p>
<p>The architectural framework and core functional design of the Codenote system were initially conceptualized and presented in our preliminary work at TCSE 2025<xref ref-type="fn" rid="fn1"><sup>1</sup></xref><fn id="fn1"><label>1</label><p><ext-link ext-link-type="uri" xlink:href="https://tcse2025.seat.org.tw">https://tcse2025.seat.org.tw</ext-link></p></fn>. As detailed in that study, traditional Integrated Development Environments (IDEs) often lack mechanisms to capture the psychological nuances of learners&#x2019; behaviors. To address this, the original Codenote prototype was developed to demonstrate how students&#x2019; operational behaviors&#x2014;such as note-taking frequency and code editing patterns&#x2014;could be utilized to infer personality traits without invasive questionnaires. While the TCSE 2025 study focused on the technical feasibility of the behavior-to-personality inference algorithm, the present study extends this foundation by deploying the system in a live educational setting. Specifically, we upgraded the system to include an active intervention module: the AI-powered pair programming engine. This evolution marks the transition of Codenote from a passive assessment tool to an intelligent intervention system, aiming to empirically validate its impact on student learning outcomes.</p>
</sec>
<sec id="s2">
<label>2</label>
<title>Theoretical Background</title>
<sec id="s2_1">
<label>2.1</label>
<title>Personality Traits and Their Relationship with Learning</title>
<p>The relationship between personality traits and academic achievement has long attracted scholarly attention. Prior research indicates that higher levels of conscientiousness are associated with effective time management, diligence, and attention to detail, all of which contribute to improved academic outcomes [<xref ref-type="bibr" rid="ref-11">11</xref>]. Additionally, openness to experience, characterized by a willingness to engage with new ideas and experiences, has been shown to be positively related to academic performance, particularly during the early stages of education [<xref ref-type="bibr" rid="ref-12">12</xref>,<xref ref-type="bibr" rid="ref-13">13</xref>].</p>
<p>Based on this body of research, incorporating personality traits as parameters within digital learning systems offers a promising approach for adaptive instructional design. By enabling systems to automatically adjust instructional strategies and pairing learners with complementary personality profiles for pair programming activities, it becomes possible to leverage individual differences to achieve more effective learning outcomes. Building on this premise, the present study employs Codenote to operationalize these theoretical insights. Unlike traditional approaches that rely on static assessments, Codenote utilizes real-time behavioral data to dynamically infer these critical traits. This capability enables the system to implement an automated, scalable mechanism for forming complementary groups, specifically designed to foster the psychological readiness required to enhance students&#x2019; coding self-efficacy.</p>
</sec>
<sec id="s2_2">
<label>2.2</label>
<title>Pair Programming</title>
<p>Pair programming was originally developed as an agile software development practice in the software industry, in which two programmers work collaboratively on the same computer to improve design quality and reduce software defects. Ainsworth and Th argue that the effectiveness of pair programming in software development can be attributed to clear role differentiation and the facilitation of discussion between partners, which help reduce cognitive load and thereby alleviate the burden on working memory [<xref ref-type="bibr" rid="ref-14">14</xref>].</p>
<p>However, the effects of pair programming are not uniformly positive. In recent years, pair programming has been increasingly adopted in programming courses, yet its instructional effectiveness remains a subject of debate. McDowell et al. conducted a large-scale experiment involving 404 students at the University of California and found that, in a CS1 course, students who engaged in randomly assigned pair programming were more likely to persist until the final examination compared with those who did not use pair programming (90.8% vs. 80.4%). Nevertheless, no significant differences were observed in final course grades between the two groups [<xref ref-type="bibr" rid="ref-15">15</xref>]. McChesney conducted a three-year longitudinal study and reported that, on average, pair programming had a positive effect on students&#x2019;learning outcomes; however, the results were inconsistent across different years. Furthermore, qualitative findings revealed that some students perceived pair programming as imposing considerable learning pressure and reported relatively low productivity during the process [<xref ref-type="bibr" rid="ref-16">16</xref>]. Notably, 29% of the participants in this study still considered discussion during pair programming to be an important component of the learning experience.</p>
<p>Further evidence suggests that algorithmic team formation strategies play a critical role in determining the effectiveness of pair programming. In the context of group-aware learning analytics, Poonam and Yasser reported that pairing students with different personality traits through remote collaboration resulted in the most favorable learning outcomes [<xref ref-type="bibr" rid="ref-7">7</xref>]. Similarly, Demir and Seferoglu validated this finding and further examined grouping strategies based on multiple learner characteristics. Their results indicated that pairing students with similar learning styles and levels of interpersonal closeness led to more enjoyable pair programming experiences, whereas pairing students with differing levels of interpersonal closeness and self-regulation resulted in better learning performance [<xref ref-type="bibr" rid="ref-17">17</xref>]. These findings suggest that, in both industrial and educational contexts, pair programming does not inherently guarantee improved work or learning outcomes, and that careful consideration of partner characteristics is essential.</p>
<p>In summary, the existing literature indicates that pair programming based on random assignment or solely on ability measures, such as programming skills or prior academic performance, does not yield consistently positive outcomes. However, several common principles have emerged that may enhance the effectiveness of pair programming. These include pairing learners with complementary characteristics, assigning distinct roles within the pair, and encouraging active discussion between partners to reduce cognitive load. Collectively, these principles provide important theoretical and practical foundations for the design of more effective pair programming instructional models. The Codenote system proposed in this study directly translates these principles into practice through its AI-driven pairing engine. By ensuring that students are matched based on the complementary traits identified in the literature rather than random assignment, the system seeks to maximize the benefits of collaborative interaction. This study aims to provide empirical evidence on how such structured, technology-mediated pairing strategies specifically influence students&#x2019; self-efficacy in a real-world educational setting.</p>
</sec>
<sec id="s2_3">
<label>2.3</label>
<title>Self-Efficacy and Programming</title>
<p>Self-efficacy is widely regarded as a critical psychological factor reflecting students&#x2019; ability to perform effectively within the domain of programming. Kovari and Katona characterize programming self-efficacy not merely as a skill metric, but as a composite construct that combines independent problem-solving abilities, self-confidence, and the sustained motivation required for coding tasks [<xref ref-type="bibr" rid="ref-18">18</xref>]. Students who demonstrate higher levels of this trait tend to comprehend abstract program concepts more effectively. Empirical research supports this link; for instance, Avcu and Ayverdi discovered that among gifted students, a significant positive correlation exists between programming self-efficacy and computational thinking skills [<xref ref-type="bibr" rid="ref-19">19</xref>]. To mitigate these challenges, Schultz and Blaszczyk concluded that the use of process guides can reduce negative emotional experiences, thereby increasing self-efficacy and lowering dropout rates [<xref ref-type="bibr" rid="ref-20">20</xref>]. Additionally, El Khoury et al. highlighted that robust self-efficacy is essential for overcoming initial hurdles and achieving long-term success in programming courses [<xref ref-type="bibr" rid="ref-21">21</xref>].</p>
<p>In the context of this study, self-efficacy is adopted as a primary metric for validating the Codenote system. By leveraging AI-driven pairing to create supportive, complementary partnerships, the system aims to reduce the &#x201C;negative emotional experiences&#x201D; and &#x201C;initial challenges&#x201D; highlighted in the literature, thereby fostering a learning environment conducive to the development of robust programming self-efficacy. Specifically, students&#x2019; self-efficacy was assessed using the self-efficacy scale developed by Ng and Lucianetti [<xref ref-type="bibr" rid="ref-22">22</xref>], a widely recognized and highly cited instrument in the field. The scale further divides self-efficacy into innovative self-efficacy, persuasion self-efficacy, and adaptive self-efficacy. According to Ng and Lucianetti, the scale measures participants&#x2019; innovative behaviors. Innovative self-efficacy is the self-view that one has the ability to produce novel ideas. Persuasion self-efficacy is the extent to which employees are confident in their ability to convince others to accept and adopt new ideas. Adaptive self-efficacy refers to an individual&#x2019;s perceived ability to handle change in the workplace.</p>
</sec>
</sec>
<sec id="s3">
<label>3</label>
<title>Methodology</title>
<p>This section details the research design, system implementation, and experimental procedures conducted to evaluate the pedagogical impact of the Codenote system. We first describe the cloud-based system architecture developed to facilitate real-time classroom interaction (<xref ref-type="sec" rid="s3_1">Section 3.1</xref>). Subsequently, we detail the AI-driven classification models and feature engineering processes used to enable the heterogeneous pairing mechanism (<xref ref-type="sec" rid="s3_2">Section 3.2</xref>). Finally, the experimental setup, including the participant demographics, learning activities, and evaluation metrics, is presented to outline how the intervention was empirically validated in a live educational setting.</p>
<sec id="s3_1">
<label>3.1</label>
<title>System Architecture and Implementation</title>
<p>To facilitate scalable classroom deployment and robust real-time interaction, Codenote is architected as a cloud-based Web service adhering to a Service-Oriented Architecture (SOA). The system is structurally divided into two primary layers: a server-side backend handling core logic and data persistence, and a client-side frontend managing user interaction and feature extensions. The system architecture of Codenote is shown in <xref ref-type="fig" rid="fig-1">Fig. 1</xref>.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>System architecture of Codenote.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-1.tif"/>
</fig>
<p>The backend, positioned at the upper layer of the architecture, is implemented using the SpringBoot framework. It comprises four distinct micro-services arranged to handle specific functional domains. From left to right, these components are: the <italic>Code/Note Management Service</italic>, the <italic>Personality Assessment Module</italic>, the <italic>Pair Programming Management Service</italic>, and the <italic>Code Quality Assessment Module</italic>. Underpinning these services is a centralized Database, which maintains connectivity with all aforementioned modules to ensure data consistency and accessibility across the platform.</p>
<p>The frontend, located at the lower layer, is built upon modern JavaScript modules and centers on user interaction. Its core components include the <italic>Code/Note Editors</italic> (powered by the Monaco Editor) and the <italic>Editing Behavior Collector</italic>. Surrounding these are specialized plugins designed for educational interventions: the <italic>Learning Material Reader Plugin</italic>, the <italic>Pair Programming Plugin</italic>, and the <italic>Code Quality Assessment Plugin</italic>. A dedicated Plugin Interface operates at the foundation, connecting these three plugins directly to the core editing environment, allowing for dynamic feature injection.</p>
<p>The architecture establishes precise communication channels between frontend components and backend services to facilitate data exchange:
<list list-type="bullet">
<list-item>
<p>Content Management: Both the <italic>Code/Note Editors</italic> and the <italic>Learning Material Reader Plugin</italic> interface directly with the backend <italic>Code/Note Management Service</italic> to handle file operations and learning material retrieval.</p></list-item>
<list-item>
<p>Behavior Analysis: The <italic>Editing Behavior Collector</italic> transmits user operational data to the backend <italic>Personality Assessment Module</italic> for personality trait analysis.</p></list-item>
<list-item>
<p>Collaboration: The frontend <italic>Pair Programming Plugin</italic> maintains a direct connection with the backend <italic>Pair Programming Management Service</italic> to synchronize coding sessions.</p></list-item>
<list-item>
<p>Evaluation: The <italic>Code Quality Assessment Plugin</italic> communicates directly with the backend <italic>Code Quality Assessment Module</italic> to facilitate instant feedback on code performance.</p></list-item>
</list></p>
<p>To clarify how the Codenote system integrates into the instructional design, <xref ref-type="fig" rid="fig-2">Fig. 2</xref> below illustrates the end-to-end operational workflow.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Experiment flow.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-2.tif"/>
</fig>
<p>The process is divided into two phases:
<list list-type="simple">
<list-item><label>1.</label><p>Phase 1: Individual Data Collection (Baseline): All students initially engage in individual coding tasks. The system captures their operational behaviors and note-taking patterns (Input) to train the personality inference model, which outputs a specific personality profile for each learner.</p></list-item>
<list-item><label>2.</label><p>Phase 2: Intervention (Control vs. Experimental):
<list list-type="simple">
<list-item><label>(a)</label><p>Control Group: Students continue to execute programming tasks individually using the standard Codenote editor.</p></list-item>
<list-item><label>(b)</label><p>Experimental Group: The system&#x2019;s pairing algorithm utilizes the inferred personality profiles to generate a Pairing List based on complementarity. Students then enter the Collaborative Workspace, where they assume Driver/Navigator roles supported by the system.</p></list-item>
</list></p></list-item>
</list></p>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Personality-Based Grouping Mechanism</title>
<p>Previous studies have indicated that using a single personality trait is insufficient to fully explain the effectiveness of pair programming on learning outcomes [<xref ref-type="bibr" rid="ref-23">23</xref>,<xref ref-type="bibr" rid="ref-24">24</xref>]. The present study adopts a personality-based grouping approach, considering multiple traits as key parameters for pair programming assignments. As prior research has demonstrated a significant relationship between programming behaviors and personality traits [<xref ref-type="bibr" rid="ref-25">25</xref>], this study employs students&#x2019; actual programming and operational behaviors as the basis for personality trait classification.</p>
<p>To operationalize this behavioral analysis, the system employs a specialized data structure named <italic>CodeDiffBean</italic> to capture granular user interactions. The <italic>CodeDiffBean</italic> records two primary attributes: <italic>diffString</italic>, which captures the content of user inputs, and <italic>diffSource</italic>, which identifies whether the input originated from the source code editor or the note-taking panel. An <italic>Editing Behavior Collector</italic> module runs in the background, monitoring the frequency of context switching between coding and note-taking, as well as the volume of typing activity.</p>
<p>The personality trait classification model employed in this study was pre-trained and rigorously validated during a preliminary pilot phase of this research project. The training dataset for this initial phase consisted of 45 undergraduate students at Yuan Ze University participating in a frontend web programming course. Ground-truth labels for personality traits were established using the Five Factor Personality Traits questionnaire, administered prior to the data collection. The model utilizes a two-layer Random Forest architecture to process students&#x2019; note-taking content and system operation behaviors (e.g., switching frequency between code and note interfaces). While the foundational validation of the AI grouping mechanism was established during the aforementioned pilot phase, the selection of the two-layer Random Forest architecture was driven by the specific nature of the behavioral dataset. First, given the relatively small training sample (N &#x003D; 45), an ensemble method like Random Forest was chosen over a single decision tree to mitigate the risk of overfitting through bagging. Second, human behavioral data, such as the variance in coding transitions and keyword frequencies, often exhibits complex, non-linear interactions. Random Forest natively captures these non-linearities more effectively than simpler linear models (e.g., Logistic Regression or Support Vector Machines) without requiring strict assumptions regarding data distribution. Finally, the two-layer hierarchical architecture was implemented to optimize multi-class classification; by decomposing the task into sequential binary decisions, the system more stably isolated distinct personality profiles within the constrained dataset. The predictive models were implemented using the randomForest package in R. To ensure strict reproducibility, a fixed random seed was applied (set. seed(5)). For the first-layer model, the number of trees (ntree) was set to 750, and the number of variables sampled at each split (mtry) was 1. For the second-layer model, ntree was set to 500 with an mtry of 3. Given the relatively small size of the pre-training dataset (N &#x003D; 45), we deliberately avoided aggressive hyperparameter tuning (such as deep grid search or cross-validation matrices) to prevent over-optimizing to the sample. Instead, ntree and mtry were determined empirically based on the stabilization of the Out-of-Bag (OOB) error. Other architectural parameters including maximum tree depth, feature subsampling (standard bootstrap with replacement), class weighting, and minimum samples per leaf (default &#x003D; 1) were kept at their robust library defaults. This conservative configuration strategy was explicitly chosen to minimize variance and control the risk of overfitting.</p>
<p>To construct the predictive models, the raw behavioral log data which contained no missing values after the initial extraction phase was engineered into three distinct feature categories. First, semantic features included the frequencies of specific note-taking keywords (e.g., questions or other annotations), which were standardized using Z-score scaling. Second, temporal features were constructed by dividing the total coding session into five equal time segments. To account for varying session lengths, these segments were normalized as ratios of the total timestamp. Third, behavioral variance metrics were calculated to capture the fluctuation in students&#x2019; editor-switching behaviors. To ensure model interpretability and avoid the &#x201C;black-box&#x201D; nature of automated dimensionality reduction, feature selection was conducted using the Boruta algorithm. As a wrapper method built around Random Forest, Boruta rigorously isolates all statistically relevant variables by comparing original feature importance against randomized shadow features. Guided by these results, the model formulas were optimized to prevent overfitting. The first-layer classification utilized a targeted subset of critical semantic features confirmed as highly relevant, while the second-layer model incorporated a broader combination of temporal and variance features. This transparent, Boruta-driven feature engineering process ensured that the AI mechanism relied on meaningful, pedagogical behavioral patterns rather than noise.</p>
<p>Due to the constrained size of the pre-training dataset (N &#x003D; 45), traditional techniques such as k-fold cross-validation or hold-out validation would significantly reduce the available training data, potentially destabilizing the model. Therefore, performance was evaluated using Out-of-Bag (OOB) estimation. OOB error is widely recognized as an unbiased and highly reliable estimate of true test set error in Random Forests, particularly suited for small sample sizes. The first layer classifies students based on note-taking features to identify high conscientiousness and low openness traits. This layer achieved an Out-of-Bag (OOB) error rate of 24.32% with a significance level of <italic>p</italic> &#x003D; 0.0185. The second layer further classifies the remaining students using vectors derived from both notes and behavioral actions, achieving a lower OOB error rate of 10.71%. To provide a comprehensive evaluation beyond the raw OOB error rate, we calculated standard classification metrics (Accuracy, Precision, Recall, and Macro F1-score) based on the OOB confusion matrices.</p>
<p>The first-layer classification model achieved an overall accuracy of 75.68%, with a Precision of 0.769, Recall of 0.625, and a balanced Macro F1-score of 0.745. The second-layer model, distinguishing between the remaining two profiles, demonstrated even higher robustness, achieving an accuracy of 85.71%. Notably, it yielded a Precision of 0.778 and a Recall of 1.00 for the target group, culminating in a Macro F1-score of 0.854. These metrics confirm that the models maintain a strong balance between precision and recall, ensuring the reliability of the automated personality grouping mechanism prior to the educational intervention. This two-stage approach allows Codenote to infer personality traits from non-invasive behavioral data with established empirical validity, enabling the automated grouping strategy used in the current experimental intervention. <xref ref-type="fig" rid="fig-3">Fig. 3</xref> shows the confusion matrices:</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>The confusion matrix.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-3.tif"/>
</fig>
<p>Specifically, we utilized the Codenote system to collect data on students&#x2019;note-taking and operational behaviors in a front-end web design course, and subsequently classified students according to personality traits such as conscientiousness and openness to experience, which have been shown to be closely associated with effective learning strategies. In the present study, personality trait assessment was conducted prior to the instructional intervention and remained fixed throughout the experimental period; subsequently; students with different personality traits were paired together for programming activities.</p>
<p>To ensure the classification thresholds were ecologically valid for the specific student demographic, we utilized a cumulative dataset collected from 165 undergraduate students at Yuan Ze University between 2021 and the onset of this experiment. Ground-truth personality traits were measured using the International Personality Item Pool (IPIP) scale. The thresholds for determining &#x2018;High&#x2019; vs. &#x2018;Low&#x2019; levels for each trait were established using a median-split method based on this local historical dataset. Consequently, the pairing algorithm defined a student as &#x2018;High Conscientiousness&#x2019; if their inferred score exceeded the historical median of this local population, ensuring that the grouping criteria were calibrated to the specific characteristics of the student cohort.</p>
<p>While calculating the theoretical end-to-end exact-match accuracy of the cascade pipeline yields approximately 64.87% (0.7568 &#x00D7; 0.8571), this metric must be interpreted within the pedagogical context of the system. The fundamental objective of the Codenote system is not to serve as a perfect psychometric diagnostic tool, but to automate heterogeneous pairing, i.e., grouping students with divergent traits to foster collaborative problem-solving. In practical terms, achieving this pedagogical goal does not require flawless fine-grained classification. The robust binary separation executed by the first-layer model isolates the most distinct personality profile from the rest, while the second layer ensures behavioral contrast among the remaining students. Therefore, even if minor misclassifications occur due to algorithmic error propagation, the baseline complementary difference required for paired students is mathematically and practically preserved. By functioning as an unobtrusive automated proxy for traditional questionnaires (thereby eliminating survey fatigue), the AI mechanism successfully generated the &#x201C;productive friction&#x201D; necessary for learning. The significant improvements in students&#x2019; self-efficacy observed in this study empirically validate that this algorithmic accuracy baseline is highly sufficient for real-world educational interventions.</p>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Runtime Code Analysis and Feedback Mechanism</title>
<p>Beyond personality assessment, Codenote incorporates a real-time Code Quality Assessment Module to support student autonomy during pair programming. Unlike static code analysis tools, our system implements a dynamic runtime evaluation. When a student edits a project (HTML/JavaScript), the system instantiates an invisible iframe within the browser context. This isolated environment executes the student&#x2019;s code in the background, intercepting runtime errors and console outputs. The system then compares these runtime events against pre-configured validation scripts defined by the instructor. This mechanism allows the system to simulate a &#x201C;developer console&#x201D; experience, identifying logic errors that syntax highlighters miss, and providing immediate, context-aware feedback to the pair, thereby reducing the cognitive load on the navigator.</p>
<p>To ensure experimental validity and isolate the pairing strategy as the sole independent variable, this runtime analysis functionality was uniformly activated across all conditions. Specifically, the feature was enabled for the &#x2018;Driver&#x2019; (code writer) in the pair programming groups, providing them with immediate diagnostic feedback during collaboration. Simultaneously, students in the control group, who worked individually, also had full access to this runtime analysis feature. This design ensures that any observed differences in learning outcomes are attributable to the personality-based pairing intervention rather than disparities in tool functionality or debugging support.</p>
</sec>
<sec id="s3_4">
<label>3.4</label>
<title>Experimental Procedure</title>
<p>The experimental design of this study consisted of two stages. In the first stage, a pretest was administered to assess students&#x2019; programming competence and self-efficacy levels. Specifically, self-efficacy was measured using the multi-dimensional scale developed by Ng and Lucianetti [<xref ref-type="bibr" rid="ref-22">22</xref>]. In this study, these dimensions were operationalized to assess students&#x2019; confidence in generating novel coding solutions (innovative), influencing peer decisions (persuasive), and coping with debugging challenges (adaptive) within the classroom context. In the second stage, students were assigned to either an experimental or a control group, ensuring no significant initial differences in self-efficacy existed between them. Participants in the experimental group engaged in programming practice using the pair programming assistant component of the Codenote system. They were paired based on personality trait groups, with each pair comprising a navigator and a driver. Based on pretest results, higher-competence students were designated as navigators, while those with lower competence acted as drivers; these roles remained constant throughout the intervention. Conversely, students in the control group practiced with the same system, but without the pair programming function. Finally, posttest self-efficacy levels were compared to evaluate the effects of the proposed approach. This design ensured that any observed differences could be attributed to the pairing strategy rather than system usage.</p>
<p>To facilitate smooth and efficient pair programming, this study developed a dedicated pair programming module integrated into the Codenote system<xref ref-type="fn" rid="fn2"><sup>2</sup></xref><fn id="fn2"><label>2</label><p><ext-link ext-link-type="uri" xlink:href="https://gitlab.com/zoe841228/codenote">https://gitlab.com/zoe841228/codenote</ext-link></p></fn>. As implemented in our previous work, the system is built as a Web service using the SpringBoot framework, featuring a backend <italic>Personality Assessment Module</italic> that utilizes a Random Forest model to process student behaviors. Before class, teachers can pre-assign student groups within the system. By integrating this AI-driven personality trait assessment model, the system allows teachers to review the distribution of students&#x2019; personality traits and to form groups accordingly. The user interface for personality assessment is shown in <xref ref-type="fig" rid="fig-4">Fig. 4</xref>.</p>
<fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Codenote personality trait assessment interface.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-4.tif"/>
</fig>
<p>During class, once students log into the system, pairing is automatically performed based on these predefined group assignments. Within each pair, one student is designated as the code writer (driver), while the other serves as the mentor (navigator). The mentor can observe their partner&#x2019;s code in real time, and both participants are able to communicate synchronously through the system&#x2019;s built-in messaging interface. Through this role structure and communication mechanism, the system is designed to promote sustained interaction and timely feedback, thereby supporting effective collaborative learning. The system interface for paired programming is shown in <xref ref-type="fig" rid="fig-5">Fig. 5</xref>.</p>
<fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>Pair programming interface.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-5.tif"/>
</fig>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Results</title>
<p>During the semester, a total of 30 students from Yuan Ze University (Taoyuan, Taiwan) were recruited to participate in the experiment and were randomly assigned to either a control group or an experimental group, with 15 students in each group. Students in the control group used the same programming editor for learning activities but did not participate in pair programming. In contrast, students in the experimental group used the same programming editor with the pair programming function enabled throughout the learning activities.</p>
<p>During the implementation of pair programming, students in the experimental group were paired according to the results generated by the personality trait analysis module. The grouping process was based on students&#x2019; levels of Conscientiousness and Openness to Experience, two dimensions of the Big Five personality traits, with high and low trait levels used to determine whether students&#x2019; personality profiles were complementary. Based on this classification, individuals with complementary personality traits were paired together. As the total number of participants was an odd number, one pair programming group included a teaching assistant in order to maintain consistency in the experimental design.</p>
<p>Students&#x2019; self-efficacy was assessed using pretest and posttest measures, employing the self-efficacy scale developed by Ng and Lucianetti [<xref ref-type="bibr" rid="ref-22">22</xref>], which comprises three dimensions: adaptive self-efficacy, innovative self-efficacy, and persuasive self-efficacy. <xref ref-type="table" rid="table-1">Table 1</xref> shows the descriptive results.</p>
<table-wrap id="table-1">
<label>Table 1</label>
<caption>
<title>Descriptive statistics.</title>
</caption>
<table>
<colgroup>
<col align="center" width="38mm"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
</colgroup>
<thead>
<tr>
<th></th>
<th></th>
<th>Valid</th>
<th>Missing</th>
<th>Mean</th>
<th>Std. Deviation</th>
<th>Minimum</th>
<th>Maximum</th>
</tr>
</thead>
<tbody>
<tr>
<td align="center" rowspan="2">Innovative Self-Efficacy (Pre-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>13.47</td>
<td>3.25</td>
<td>7.00</td>
<td>21.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>12.80</td>
<td>2.81</td>
<td>8.00</td>
<td>18.00</td>
</tr>
<tr>
<td align="center" rowspan="2">Persuasive Self-Efficacy (Pre-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>22.73</td>
<td>5.48</td>
<td>15.00</td>
<td>35.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>21.27</td>
<td>5.35</td>
<td>12.00</td>
<td>31.00</td>
</tr>
<tr>
<td align="center" rowspan="2">Adaptive Self-Efficacy (Pre-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>20.73</td>
<td>5.50</td>
<td>14.00</td>
<td>32.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>20.27</td>
<td>6.27</td>
<td>10.00</td>
<td>30.00</td>
</tr>
<tr>
<td align="center" rowspan="2">Innovative Self-Efficacy (Post-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>12.73</td>
<td>2.94</td>
<td>6.00</td>
<td>18.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>13.47</td>
<td>3.11</td>
<td>6.00</td>
<td>18.00</td>
</tr>
<tr>
<td align="center" rowspan="2">Persuasive Self-Efficacy (Post-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>21.33</td>
<td>5.53</td>
<td>12.00</td>
<td>34.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>21.20</td>
<td>4.86</td>
<td>13.00</td>
<td>30.00</td>
</tr>
<tr>
<td align="center" rowspan="2">Adaptive Self-Efficacy (Post-test)</td>
<td>Control</td>
<td>15</td>
<td>0</td>
<td>20.07</td>
<td>4.80</td>
<td>11.00</td>
<td>29.00</td>
</tr>
<tr>
<td>Experimental</td>
<td>15</td>
<td>0</td>
<td>30.33</td>
<td>7.08</td>
<td>14.00</td>
<td>42.00</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The analysis results indicated that there was no significant difference between the control group and the experimental group in adaptive self-efficacy at the pretest stage. However, after the completion of the experiment, a significant difference in adaptive self-efficacy was observed between the two groups, as illustrated in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>.</p>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Adaptive self-efficacy: pre-test and post-test comparison between groups.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_79379-fig-6.tif"/>
</fig>
<p>Prior to conducting the inferential analysis, the assumptions of normality and homogeneity of variance were examined. A Test of Normality using the Shapiro&#x2013;Wilk method was performed for both groups, and the results showed that all <italic>p</italic>-values were greater than 0.05, indicating that the data for each group followed a normal distribution, as shown in <xref ref-type="table" rid="table-2">Table 2</xref>.</p>
<table-wrap id="table-2">
<label>Table 2</label>
<caption>
<title>Test of normality (Shapiro-Wilk).</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
</colgroup>
<thead>
<tr>
<th></th>
<th></th>
<th>W</th>
<th><italic>p</italic></th>
</tr>
</thead>
<tbody>
<tr>
<td>Innovative Self-Efficacy</td>
<td>Control</td>
<td>0.954</td>
<td>0.592</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Experimental</td>
<td>0.954</td>
<td>0.591</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy</td>
<td>Control</td>
<td>0.931</td>
<td>0.284</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Experimental</td>
<td>0.951</td>
<td>0.547</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy</td>
<td>Control</td>
<td>0.912</td>
<td>0.143</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Experimental</td>
<td>0.953</td>
<td>0.566</td>
</tr>
<tr>
<td>Innovative Self-Efficacy</td>
<td>Control</td>
<td>0.864</td>
<td>0.027</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Experimental</td>
<td>0.942</td>
<td>0.406</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy</td>
<td>Control</td>
<td>0.933</td>
<td>0.306</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Experimental</td>
<td>0.944</td>
<td>0.435</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy</td>
<td>Control</td>
<td>0.918</td>
<td>0.179</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Experimental</td>
<td>0.960</td>
<td>0.699</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In addition, a Test of Equality of Variances using Levene&#x2019;s test was conducted to assess the homogeneity of variances between the two groups. The results indicated that all <italic>p</italic>-values were also greater than 0.05, suggesting that the assumption of equal variances was satisfied, as presented in <xref ref-type="table" rid="table-3">Table 3</xref>.</p>
<table-wrap id="table-3">
<label>Table 3</label>
<caption>
<title>Test of equality of variances (Levene&#x2019;s).</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th></th>
<th>F</th>
<th>df<sub>1</sub></th>
<th>Df<sub>2</sub></th>
<th><italic>p</italic></th>
</tr>
</thead>
<tbody>
<tr>
<td>Innovative Self-Efficacy (Pre-test)</td>
<td>0.026</td>
<td>1</td>
<td>28</td>
<td>0.873</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy (Pre-test)</td>
<td>0.079</td>
<td>1</td>
<td>28</td>
<td>0.781</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy (Pre-test)</td>
<td>0.266</td>
<td>1</td>
<td>28</td>
<td>0.610</td>
</tr>
<tr>
<td>Innovative Self-Efficacy (Post-test)</td>
<td>0.207</td>
<td>1</td>
<td>28</td>
<td>0.652</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy (Post-test)</td>
<td>0.105</td>
<td>1</td>
<td>28</td>
<td>0.748</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy (Post-test)</td>
<td>2.921</td>
<td>1</td>
<td>28</td>
<td>0.099</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Based on these results, the data met the assumptions required for parametric testing, and an Independent Samples <italic>t</italic>-test was subsequently applied. The statistical results are reported in <xref ref-type="table" rid="table-4">Table 4</xref>. The findings reveal that personality trait&#x2013;based pair programming produced a significant positive effect on the adaptive dimension of self-efficacy. To address the potential inflation of Type I error due to multiple comparisons (three dimensions of self-efficacy), a Bonferroni correction was considered (<italic>&#x03B1;</italic> &#x003D; 0.05/3 &#x003D; 0.016). The results of the Independent Samples <italic>t</italic>-test (<xref ref-type="table" rid="table-4">Table 4</xref>) indicate that the difference in Adaptive Self-Efficacy remained statistically significant even under this stricter threshold (<italic>p</italic> &#x003C; 0.001). Furthermore, given the small sample size (N &#x003D; 30), Hedges&#x2019; g was reported as the measure of effect size to provide a less biased estimate than Cohen&#x2019;s d. In contrast, no statistically significant differences were found between the experimental and control groups with respect to innovative self-efficacy or persuasive self-efficacy.</p>
<table-wrap id="table-4">
<label>Table 4</label>
<caption>
<title>Independent samples <italic>t</italic>-test.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center" width="20mm"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th></th>
<th>Test</th>
<th>Statistic</th>
<th>df</th>
<th><italic>p</italic></th>
<th>Effect Size (Hedges&#x2019; g)</th>
<th>SE Effect Size</th>
</tr>
</thead>
<tbody>
<tr>
<td>Innovative Self-Efficacy</td>
<td>Welch</td>
<td>0.601</td>
<td>27.426</td>
<td>0.553</td>
<td>0.214</td>
<td>0.357</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Mann-Whitney</td>
<td>127.000</td>
<td></td>
<td>0.557</td>
<td>0.129</td>
<td>0.211</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy</td>
<td>Welch</td>
<td>0.741</td>
<td>27.983</td>
<td>0.465</td>
<td>0.263</td>
<td>0.358</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Mann-Whitney</td>
<td>128.500</td>
<td></td>
<td>0.519</td>
<td>0.142</td>
<td>0.211</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy</td>
<td>Welch</td>
<td>0.217</td>
<td>27.525</td>
<td>0.830</td>
<td>0.077</td>
<td>0.356</td>
</tr>
<tr>
<td>(Pre-test)</td>
<td>Mann-Whitney</td>
<td>111.000</td>
<td></td>
<td>0.967</td>
<td>&#x2212;0.013</td>
<td>0.211</td>
</tr>
<tr>
<td>Innovative Self-Efficacy</td>
<td>Welch</td>
<td>&#x2212;0.663</td>
<td>27.907</td>
<td>0.513</td>
<td>&#x2212;0.236</td>
<td>0.358</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Mann-Whitney</td>
<td>91.500</td>
<td></td>
<td>0.381</td>
<td>&#x2212;0.187</td>
<td>0.211</td>
</tr>
<tr>
<td>Persuasive Self-Efficacy</td>
<td>Welch</td>
<td>0.070</td>
<td>27.549</td>
<td>0.945</td>
<td>0.025</td>
<td>0.355</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Mann-Whitney</td>
<td>116.000</td>
<td></td>
<td>0.901</td>
<td>0.031</td>
<td>0.211</td>
</tr>
<tr>
<td>Adaptive Self-Efficacy</td>
<td>Welch</td>
<td>&#x2212;4.649</td>
<td>24.637</td>
<td>&#x003C;0.001<bold>&#x002A;</bold></td>
<td>&#x2212;1.652</td>
<td>0.461</td>
</tr>
<tr>
<td>(Post-test)</td>
<td>Mann-Whitney</td>
<td>27.000</td>
<td></td>
<td>&#x003C;0.001<bold>&#x002A;</bold></td>
<td>&#x2212;0.760</td>
<td>0.211</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="table-4fn1" fn-type="other">
<p>Note: &#x002A;indicates statistical significance at the <italic>p</italic> &#x003C; 0.05 level.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>Effect sizes were calculated using Hedges&#x2019; g to correct for small sample bias. Notably, the effect size for adaptive self-efficacy (g &#x003D; 1.65) indicates a substantial practical impact of the intervention.</p>
</sec>
<sec id="s5">
<label>5</label>
<title>Discussion</title>
<sec id="s5_1">
<label>5.1</label>
<title>The Role of Personality Complementarity in Pair Programming</title>
<p>The study examined the effects of Codenote-facilitated pair programming based on personality trait complementarity on students&#x2019; self-efficacy, with particular attention to adaptive, innovative, and persuasive dimensions. Previous research has shown that pair programming can be beneficial in both industrial and educational settings; however, its effectiveness is highly dependent on pairing strategies and the structure of collaboration rather than on the mere implementation of collaborative programming. Empirical studies have consistently indicated that grouping strategies grounded in learner characteristics, such as personality traits, learning styles, interpersonal closeness, and self-regulation, yield more favorable outcomes than random or ability-based pairing. Specifically, Poonam and Yasser reported that pairing students with different personality traits in a remote collaboration context resulted in the most favorable learning outcomes, suggesting that heterogeneity in personality traits may support more effective learning processes [<xref ref-type="bibr" rid="ref-7">7</xref>]. Similarly, Demir and Seferoglu examined grouping strategies based on multiple learner characteristics and found that pairing students with similar learning styles and levels of interpersonal closeness led to more enjoyable pair programming experiences, whereas pairing students with differing levels of interpersonal closeness and self-regulation was associated with better learning performance [<xref ref-type="bibr" rid="ref-17">17</xref>]. Taken together, these findings indicate that different pairing criteria may lead to different types of outcomes, and that complementary learner characteristics are particularly relevant when the instructional focus is on learning performance rather than on subjective experience.</p>
<p>Building on this body of research, the present study adopted Codenote&#x2019;s AI-driven pairing mechanism that grouped students with complementary traits and examined its effects on multiple dimensions of self-efficacy. Results indicated no significant difference in adaptive self-efficacy between the groups at the pretest stage, suggesting they were comparable prior to the intervention. Following the implementation of Codenote-supported pair programming, the experimental group demonstrated significantly higher adaptive self-efficacy than the control group. This finding aligns with earlier research emphasizing that the effectiveness of pair programming stems from the quality of interaction&#x2014;including discussion, role differentiation, and mutual monitoring&#x2014;rather than from collaboration alone. A plausible explanation is that personality-based pairing exposed students to diverse perspectives and coping strategies. As pair programming requires learners to adjust strategies, negotiate solutions, and respond to errors in real time, interacting with partners who approach problems differently may foster flexibility and resilience. Consequently, this enhances students&#x2019; perceived ability to manage novel or uncertain situations, an interpretation consistent with qualitative findings that highlight interpersonal interaction as a central component of effective pair programming.</p>
</sec>
<sec id="s5_2">
<label>5.2</label>
<title>Dimension-Specific Effects on Self-Efficacy</title>
<p>In contrast, no significant effects were observed for innovative self-efficacy or persuasive self-efficacy. This suggests that the AI-based pairing algorithm alone may be insufficient to influence these dimensions. Prior research has indicated that innovative self-efficacy is closely associated with opportunities for creative exploration and open-ended problem solving, while persuasive self-efficacy is more strongly related to leadership roles, communication demands, and explicit decision-making authority. Although system-mediated collaborative programming was implemented in the present study, the instructional design and system features may not have explicitly emphasized these elements, thereby limiting potential gains in these dimensions of self-efficacy. Overall, the findings reinforce the view that pair programming does not inherently produce uniformly positive effects across all psychological or learning outcomes. Instead, its effectiveness appears to be dimension-specific and contingent upon instructional design and tool capability decisions. The significant improvement observed in adaptive self-efficacy suggests that Codenote&#x2019;s intelligent pairing strategy is particularly effective in supporting students&#x2019; confidence in coping with complex and dynamic programming tasks. By demonstrating a selective effect on adaptive self-efficacy, this study contributes to the literature by clarifying the conditions under which AI-powered collaborative tools may enhance learners&#x2019; psychological readiness rather than assuming broad benefits across all aspects of self-efficacy.</p>
<p>While significant improvements were not observed in innovative or persuasive domains, the isolated gain in adaptive self-efficacy represents a critical pedagogical milestone in computer science education. The field of programming is characterized by rapid technological evolution, where languages, frameworks, and methodologies are in a constant state of flux. Consequently, the ability to adapt to change is arguably the most vital attribute for long-term success. As recent research on career development highlights, individuals equipped with strong career adaptability are significantly more likely to engage in proactive career behaviors, enabling them to successfully navigate uncertainty and rapid technological shifts [<xref ref-type="bibr" rid="ref-26">26</xref>]. In the context of this study, the significant increase in adaptive self-efficacy indicates that personality-based pair programming successfully cultivated students&#x2019; psychological readiness to navigate uncertainty. This suggests that students are not merely learning to code, but are developing the &#x2018;change agility&#x2019; required to proactively learn new skills and drive technological innovation in their future workplaces. Therefore, even in the absence of immediate gains in innovation or persuasion, the enhancement of adaptive capabilities provides a foundational scaffold for students&#x2019; persistent engagement and lifelong learning in engineering disciplines.</p>
</sec>
<sec id="s5_3">
<label>5.3</label>
<title>Pedagogical Implications in the Era of Generative AI</title>
<p>In light of the rapid proliferation of Generative AI (GenAI) in programming education, it is pertinent to distinguish the impact of human peer pairing from AI assistance. While GenAI tools can effectively scaffold syntax acquisition and debugging, the significant improvement in adaptive self-efficacy observed in this study likely stems from the socio-cognitive demands of human partnership. Specifically, the need to negotiate conflicting ideas and navigate interpersonal dynamics. Current AI assistants, which typically function as passive information retrievers, may not replicate this &#x2018;productive friction&#x2019; that fosters psychological resilience.</p>
<p>However, the core principle validated in this study offers a crucial design blueprint for next-generation AI tutors. Rather than providing generic responses, future AI agents could be engineered with specific &#x2018;personas&#x2019; (e.g., a highly conscientious &#x2018;Navigator&#x2019; Agent) to complement the learner&#x2019;s detected traits. By simulating the heterogeneous pairing logic validated in this study, AI-mediated environments could potentially recreate the psychological benefits of human collaboration, moving from simple code generation to personalized socio-emotional scaffolding.</p>
<p>Compared with recent related literature, this focus on &#x201C;adaptive resilience&#x201D; offers a distinct pedagogical value. For instance, Yilmaz and Karaoglan Yilmaz [<xref ref-type="bibr" rid="ref-27">27</xref>] demonstrated that Generative AI tools (e.g., ChatGPT) significantly boost students&#x2019; programming self-efficacy by assisting with task completion and debugging. Similarly, a recent systematic review by Massaty et al. [<xref ref-type="bibr" rid="ref-28">28</xref>] concluded that AI-driven tools generally enhance self-efficacy by providing &#x201C;tailored feedback and support,&#x201D; which directly reinforces students&#x2019; confidence in their academic abilities. While these approaches effectively build confidence through task success (i.e., &#x201C;I can solve this because AI helped me&#x201D;), our personality-based pairing strategy fosters confidence through adaptability. By enhancing adaptive self-efficacy, our system does not merely make programming easier; it equips students with the change agility required to face uncertainty. This resonates with the broader goal of self-efficacy identified in Massaty et al.&#x2019;s review [<xref ref-type="bibr" rid="ref-28">28</xref>] but achieves it via a socio-cognitive mechanism rather than purely technical assistance.</p>
</sec>
<sec id="s5_4">
<label>5.4</label>
<title>Limitations and Future Work</title>
<p>While this study offers promising evidence for the benefits of personality-aware tools, several contextual aspects are worth noting. This pilot implementation involved 30 participants, providing a focused foundation that future research can extend across larger and more diverse student populations. By evaluating Codenote as an integrated instructional environment, this work establishes a baseline for further exploring the synergistic interactions between its various intelligent components. Additionally, the system&#x2019;s innovative use of non-invasive behavioral markers for personality inference opens the door for future studies to incorporate traditional assessments as a means of providing multi-dimensional validation. The stable role assignment used in this intervention helped ensure experimental consistency, though future iterations involving role rotation could provide broader insights into diverse self-efficacy gains. These considerations serve as valuable starting points for the ongoing optimization of AI-supported collaborative learning.</p>
<p>A limitation of the current experimental design is the absence of a &#x2018;randomly paired&#x2019; control group. Due to the constrained class size (N &#x003D; 30), dividing participants into three conditions would have compromised statistical power. Consequently, the study compared personality-based pairing against individual work. While this design leaves open the possibility that the observed benefits stem from collaboration itself rather than the specific matching strategy, prior literature suggests that pair programming does not inherently guarantee superior outcomes. Studies such as McChesney [<xref ref-type="bibr" rid="ref-16">16</xref>] and Demir and Seferoglu [<xref ref-type="bibr" rid="ref-17">17</xref>] have noted that without strategic grouping, pair programming can lead to frustration or unequal participation. Furthermore, the specific gain in adaptive self-efficacy in the absence of significant changes in persuasive or innovative dimensions suggests that the effect is not a generic &#x2018;collaboration bonus&#x2019;. Instead, it points to the specific influence of working with a complementary partner, which requires students to actively adapt to differing perspectives, thereby isolating the mechanism of personality complementarity to a plausible degree. To move beyond this plausible isolation and establish rigorous causality, we explicitly commit to deploying a multi-arm experimental design in future large-scale studies, directly comparing AI-driven heterogeneous pairing against a randomly paired control group. Furthermore, the current evaluation relies primarily on self-reported measures of self-efficacy to validate the psychological impact of the intervention. While fostering these affective states is a critical precursor to learning, we recognize the necessity of objective technical validation. In our future research trajectory, we explicitly commit to integrating objective performance indicators such as code completion rates, algorithmic efficiency, and compilation success frequencies. Combining these objective metrics with subjective psychological assessments will provide a comprehensive evaluation of the Codenote system&#x2019;s true pedagogical efficacy.</p>
</sec>
</sec>
<sec id="s6">
<label>6</label>
<title>Conclusions</title>
<p>The present study validated the effectiveness of the Codenote system, specifically its AI-powered pair programming functionality, on students&#x2019; self-efficacy across adaptive, innovative, and persuasive dimensions. The findings indicate that Codenote&#x2019;s automated pairing mechanism, which groups students with complementary personality traits, significantly enhances adaptive self-efficacy, suggesting that exposure to diverse perspectives and problem-solving approaches strengthens learners&#x2019; confidence in managing complex and dynamic programming tasks. No significant improvements were observed in innovative or persuasive self-efficacy, implying that these dimensions may require additional system-embedded scaffolds such as opportunities for creative exploration, leadership roles, or structured reflection activities. Overall, the results highlight that the effectiveness of AI-supported pair programming is dimension-specific and contingent upon the quality of interaction and intelligent pairing strategies rather than collaboration alone. These findings contribute to the literature by clarifying the conditions under which AI-driven tools can enhance students&#x2019; psychological readiness and adaptive capabilities. Future research should explore long-term effects, the integration of additional automated scaffolds, and the interaction patterns within pairs to further understand the mechanisms underlying the development of self-efficacy in AI-mediated environments.</p>
</sec>
<sec id="s7">
<label>7</label>
<title>Future Work</title>
<p>Based on the findings and limitations of the current study, several strategic directions for future research and system development have been identified.
<list list-type="simple">
<list-item><label>1.</label><p>Enhanced Instructional Scaffolding for Creative and Leadership Domains: While the current personality-based pairing successfully improved adaptive self-efficacy, the lack of significant gains in innovative and persuasive dimensions suggests the need for more explicit pedagogical interventions. Future iterations of Codenote will integrate dynamic role rotation mechanisms. Rather than static assignments, the system could algorithmically prompt role swaps (Driver/Navigator) at critical project milestones. Furthermore, to bolster persuasive self-efficacy, the system could introduce structured &#x201C;conflict resolution&#x201D; scenarios where pairs must negotiate architectural decisions, supported by automated prompts that encourage the quieter partner to lead the discussion.</p></list-item>
<list-item><label>2.</label><p>Integration of Generative AI for Intelligent Feedback: Building upon the Runtime Code Analysis module detailed in <xref ref-type="sec" rid="s3_3">Section 3.3</xref>, future work aims to transition from rule-based validation to Generative AI-driven mentorship. By embedding Large Language Models (LLMs) into the AI Inspector, the system could move beyond error detection to provide conversational, Socratic-style guidance. This would allow Codenote not only to identify what is wrong but to explain why, customizing the complexity of the explanation based on the learner&#x2019;s detected proficiency level, thereby providing a more personalized scaffolding experience.</p></list-item>
<list-item><label>3.</label><p>Multimodal Analysis of Collaboration Dynamics: To deepen our understanding of how complementary personality traits influence collaboration, future research should leverage the data collected by the Pair Programming Plugin. Specifically, applying Natural Language Processing (NLP) techniques to analyze the chat logs between partners could reveal patterns in communication styles&#x2014;such as negotiation frequency, sentiment, and turn-taking behavior. Correlating these linguistic markers with personality profiles would provide granular insights into the mechanisms underlying effective peer modeling, moving beyond outcome-based metrics to process-oriented analysis.</p></list-item>
<list-item><label>4.</label><p>Longitudinal Scalability and Transferability: Finally, longitudinal studies are required to determine the durability of the observed self-efficacy gains. It remains to be seen whether the confidence built through personality-aligned pairing transfers to individual coding tasks or more complex software engineering challenges, such as backend development or system architecture design. Extending the experimental scope to include diverse student populations across different institutions will also be crucial in establishing the generalizability of the behavior-based pairing algorithm.</p></list-item>
</list></p>
</sec>
</body>
<back>
<ack>
<p>Not applicable.</p>
</ack>
<sec>
<title>Funding Statement</title>
<p>This research was partially supported by the National Science and Technology Council (NSTC), Taiwan, under Grant Nos.: 113-2410-H-155-023 and 114-2410-H-155-012-MY2. The grants were received by author Chun-Hsiung Tseng. The sponsor&#x2019;s website is available at <ext-link ext-link-type="uri" xlink:href="https://www.nstc.gov.tw/">https://www.nstc.gov.tw/</ext-link>.</p>
</sec>
<sec>
<title>Author Contributions</title>
<p>The authors confirm contribution to the paper as follows: study conception and design: Chun-Hsiung Tseng, Hao-Chiang Koong Lin, Andrew Chih-Wei Huang; system implementation: Jia-Rou Lin, Chun-Hsiung Tseng; data collection and experiment execution: Jia-Rou Lin; analysis and interpretation of results: Chun-Hsiung Tseng, Hao-Chiang Koong Lin, Andrew Chih-Wei Huang; draft manuscript preparation: Jia-Rou Lin, Chun-Hsiung Tseng. All authors reviewed and approved the final version of the manuscript.</p>
</sec>
<sec sec-type="data-availability">
<title>Availability of Data and Materials</title>
<p>The data that support the findings of this study are available from the corresponding author, Chun-Hsiung Tseng, upon reasonable request.</p>
</sec>
<sec>
<title>Ethics Approval</title>
<p>The study protocol was reviewed and approved by the Human Research Ethics Governance &#x0026; Ethical Review Committee of National Cheng Kung University (Case No.: 112-193). Please note that as Yuan Ze University does not maintain an internal Institutional Review Board (IRB), the ethical review process was officially delegated to and approved by the aforementioned independent committee, which is a standard practice for human subject research in Taiwan. Prior to the experiment, all participants provided written informed consent. They were explicitly informed that their operational data and note-taking behaviors would be analyzed for research purposes only, and that their participation was voluntary. To ensure privacy and confidentiality, all personal identifiers were removed and replaced with pseudo-anonymized IDs during the data analysis process. The behavior-based personality inference was strictly used for grouping purposes within the instructional activity and did not influence students&#x2019; course grades or academic evaluations.</p>
</sec>
<sec sec-type="COI-statement">
<title>Conflicts of Interest</title>
<p>The authors declare no conflicts of interest.</p>
</sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Huang</surname> <given-names>KH</given-names></string-name></person-group>. <article-title>Computer education reform in Taiwan</article-title>. <source>Nat Sci Educ A Compr Sch</source>. <year>2024</year>;<volume>30</volume>(<issue>1</issue>):<fpage>14</fpage>&#x2013;<lpage>23</lpage>. doi:<pub-id pub-id-type="doi">10.48127/gu/24.30.14</pub-id>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Chen</surname> <given-names>JM</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>TT</given-names></string-name>, <string-name><surname>Sandnes</surname> <given-names>FE</given-names></string-name></person-group>. <chapter-title>Exploration of computational thinking based on bebras performance in webduino programming by high school students</chapter-title>. In: <person-group person-group-type="editor"><string-name><surname>Wu</surname> <given-names>TT</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>YM</given-names></string-name>, <string-name><surname>Shadiev</surname> <given-names>R</given-names></string-name>, <string-name><surname>Lin</surname> <given-names>L</given-names></string-name>, <string-name><surname>Star&#x010D;i&#x010D;</surname> <given-names>AI</given-names></string-name></person-group>, editors. <source>Innovative technologies and learning</source>. <publisher-loc>Cham, Switzerland</publisher-loc>: <publisher-name>Springer International Publishing</publisher-name>; <year>2018</year>. p. <fpage>443</fpage>&#x2013;<lpage>52</lpage>. doi:<pub-id pub-id-type="doi">10.1007/978-3-319-99737-7_47</pub-id>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Dwan</surname> <given-names>F</given-names></string-name>, <string-name><surname>Oliveira</surname> <given-names>E</given-names></string-name>, <string-name><surname>Fernandes</surname> <given-names>D</given-names></string-name></person-group>. <article-title>Predi&#x00E7;&#x00E3;o de zona de aprendizagem de alunos de introdu&#x00E7;&#x00E3;o &#x00E0; programa&#x00E7;&#x00E3;o em ambientes de corre&#x00E7;&#x00E3;o autom&#x00E1;tica de C&#x00F3;digo</article-title>. <source>Anais Do XXVIII Simp&#x00F3;sio Brasileiro De Inform&#x00E1;tica Na Educ SBIE 2017</source>. <year>2017</year>;<volume>1</volume>:<fpage>1507</fpage>. doi:<pub-id pub-id-type="doi">10.5753/cbie.sbie.2017.1507</pub-id>.</mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Williams</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wiebe</surname> <given-names>E</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>K</given-names></string-name>, <string-name><surname>Ferzli</surname> <given-names>M</given-names></string-name>, <string-name><surname>Miller</surname> <given-names>C</given-names></string-name></person-group>. <article-title>In support of pair programming in the introductory computer science course</article-title>. <source>Comput Sci Educ</source>. <year>2002</year>;<volume>12</volume>(<issue>3</issue>):<fpage>197</fpage>&#x2013;<lpage>212</lpage>. doi:<pub-id pub-id-type="doi">10.1076/csed.12.3.197.8618</pub-id>; <pub-id pub-id-type="pmid">37995447</pub-id></mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>VanDeGrift</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Coupling pair programming and writing: learning about students&#x2019; perceptions and processes</article-title>. In: <conf-name>Proceedings of the 35th SIGCSE Technical Symposium on Computer Science Education; 2004 Mar 3&#x2013;7; Norfol, VA, USA</conf-name>. <publisher-loc>New York, NY, USA</publisher-loc>: <publisher-name>ACM</publisher-name>. p. <fpage>2</fpage>&#x2013;<lpage>6</lpage>. doi:<pub-id pub-id-type="doi">10.1145/971300.971306</pub-id>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Chigona</surname> <given-names>W</given-names></string-name>, <string-name><surname>Pollock</surname> <given-names>M</given-names></string-name></person-group>. <article-title>Pair programming for information systems students new to programming: students&#x2019; experiences and teachers&#x2019; challenges</article-title>. In: <conf-name>Proceedings of the PICMET&#x2019;08&#x2014;2008 Portland International Conference on Management of Engineering &#x0026; Technology; 2008 Jul 27&#x2013;31; Cape Town, South Africa</conf-name>. <publisher-loc>New York, NY, USA</publisher-loc>: <publisher-name>IEEE</publisher-name>. p. <fpage>1587</fpage>&#x2013;<lpage>94</lpage>. doi:<pub-id pub-id-type="doi">10.1109/PICMET.2008.4599777</pub-id>.</mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Poonam</surname> <given-names>R</given-names></string-name>, <string-name><surname>Yasser</surname> <given-names>CM</given-names></string-name></person-group>. <article-title>An experimental study to investigate personality traits on pair programming efficiency in extreme programming</article-title>. In: <conf-name>Proceedings of the 2018 5th International Conference on Industrial Engineering and Applications (ICIEA); 2018 Apr 26&#x2013;28; Singapore</conf-name>. <publisher-loc>New York, NY, USA</publisher-loc>: <publisher-name>IEEE</publisher-name>; <year>2018</year>. p. <fpage>95</fpage>&#x2013;<lpage>9</lpage>. doi:<pub-id pub-id-type="doi">10.1109/IEA.2018.8387077</pub-id>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Mairesse</surname> <given-names>F</given-names></string-name>, <string-name><surname>Walker</surname> <given-names>MA</given-names></string-name>, <string-name><surname>Mehl</surname> <given-names>MR</given-names></string-name>, <string-name><surname>Moore</surname> <given-names>RK</given-names></string-name></person-group>. <article-title>Using linguistic cues for the automatic recognition of personality in conversation and text</article-title>. <source>J Artif Intell Res</source>. <year>2007</year>;<volume>30</volume>:<fpage>457</fpage>&#x2013;<lpage>500</lpage>. doi:<pub-id pub-id-type="doi">10.1613/jair.2349</pub-id>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Meidenbauer</surname> <given-names>KL</given-names></string-name>, <string-name><surname>Niu</surname> <given-names>T</given-names></string-name>, <string-name><surname>Choe</surname> <given-names>KW</given-names></string-name>, <string-name><surname>Stier</surname> <given-names>AJ</given-names></string-name>, <string-name><surname>Berman</surname> <given-names>MG</given-names></string-name></person-group>. <article-title>Mouse movements reflect personality traits and task attentiveness in online experiments</article-title>. <source>J Pers</source>. <year>2023</year>;<volume>91</volume>(<issue>2</issue>):<fpage>413</fpage>&#x2013;<lpage>25</lpage>. doi:<pub-id pub-id-type="doi">10.1111/jopy.12736</pub-id>; <pub-id pub-id-type="pmid">35591790</pub-id></mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Mart&#x00ED;nez Mej&#x00ED;a</surname> <given-names>RD</given-names></string-name>, <string-name><surname>Rodr&#x00ED;guez Villanueva</surname> <given-names>BP</given-names></string-name></person-group>. <article-title>Key factors influencing initial learning in computer programming</article-title>. <source>Rev Cient&#x00ED;fica Sist E Inform&#x00E1;tica</source>. <year>2024</year>;<volume>4</volume>(<issue>2</issue>):<fpage>3</fpage>.</mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Stajkovic</surname> <given-names>AD</given-names></string-name>, <string-name><surname>Bandura</surname> <given-names>A</given-names></string-name>, <string-name><surname>Locke</surname> <given-names>EA</given-names></string-name>, <string-name><surname>Lee</surname> <given-names>D</given-names></string-name>, <string-name><surname>Sergent</surname> <given-names>K</given-names></string-name></person-group>. <article-title>Test of three conceptual models of influence of the big five personality traits and self-efficacy on academic performance: a meta-analytic path-analysis</article-title>. <source>Pers Individ Differ</source>. <year>2018</year>;<volume>120</volume>:<fpage>238</fpage>&#x2013;<lpage>45</lpage>. doi:<pub-id pub-id-type="doi">10.1016/j.paid.2017.08.014</pub-id>.</mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Verbree</surname> <given-names>AR</given-names></string-name>, <string-name><surname>Maas</surname> <given-names>L</given-names></string-name>, <string-name><surname>Hornstra</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wijngaards-de Meij</surname> <given-names>L</given-names></string-name></person-group>. <article-title>Personality predicts academic achievement in higher education: differences by academic field of study?</article-title> <source>Learn Individ Differ</source>. <year>2021</year>;<volume>92</volume>:<fpage>102081</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.lindif.2021.102081</pub-id>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Brandt</surname> <given-names>ND</given-names></string-name>, <string-name><surname>Lechner</surname> <given-names>CM</given-names></string-name>, <string-name><surname>Tetzner</surname> <given-names>J</given-names></string-name>, <string-name><surname>Rammstedt</surname> <given-names>B</given-names></string-name></person-group>. <article-title>Personality, cognitive ability, and academic performance: differential associations across school subjects and school tracks</article-title>. <source>J Pers</source>. <year>2020</year>;<volume>88</volume>(<issue>2</issue>):<fpage>249</fpage>&#x2013;<lpage>65</lpage>. doi:<pub-id pub-id-type="doi">10.1111/jopy.12482</pub-id>; <pub-id pub-id-type="pmid">31009081</pub-id></mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ainsworth</surname> <given-names>S</given-names></string-name>, <string-name><surname>Th Loizou</surname> <given-names>A</given-names></string-name></person-group>. <article-title>The effects of self-explaining when learning with text or diagrams</article-title>. <source>Cogn Sci</source>. <year>2003</year>;<volume>27</volume>(<issue>4</issue>):<fpage>669</fpage>&#x2013;<lpage>81</lpage>. doi:<pub-id pub-id-type="doi">10.1016/S0364-0213(03)00033-8</pub-id>.</mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>McDowell</surname> <given-names>C</given-names></string-name>, <string-name><surname>Werner</surname> <given-names>L</given-names></string-name>, <string-name><surname>Bullock</surname> <given-names>HE</given-names></string-name>, <string-name><surname>Fernald</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Pair programming improves student retention, confidence, and program quality</article-title>. <source>Commun ACM</source>. <year>2006</year>;<volume>49</volume>(<issue>8</issue>):<fpage>90</fpage>&#x2013;<lpage>5</lpage>. doi:<pub-id pub-id-type="doi">10.1145/1145287.1145293</pub-id>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>McChesney</surname> <given-names>I</given-names></string-name></person-group>. <article-title>Three years of student pair programming: action research insights and outcomes</article-title>. In: <conf-name>Proceedings of the 47th ACM Technical Symposium on Computing Science Education; 2016 Mar 2&#x2013;5; Memphis, TN, USA</conf-name>. <publisher-loc>New York, NY, USA</publisher-loc>: <publisher-name>ACM</publisher-name>; <year>2016</year>. p. <fpage>84</fpage>&#x2013;<lpage>9</lpage>. doi:<pub-id pub-id-type="doi">10.1145/2839509.2844565</pub-id>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Demir</surname> <given-names>&#x00D6;</given-names></string-name>, <string-name><surname>Seferoglu</surname> <given-names>SS</given-names></string-name></person-group>. <article-title>The effect of determining pair programming groups according to various individual difference variables on group compatibility, flow, and coding performance</article-title>. <source>J Educ Comput Res</source>. <year>2021</year>;<volume>59</volume>(<issue>1</issue>):<fpage>41</fpage>&#x2013;<lpage>70</lpage>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Kovari</surname> <given-names>A</given-names></string-name>, <string-name><surname>Katona</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Effect of software development course on programming self-efficacy</article-title>. <source>Educ Inf Technol</source>. <year>2023</year>;<volume>28</volume>(<issue>9</issue>):<fpage>10937</fpage>&#x2013;<lpage>63</lpage>. doi:<pub-id pub-id-type="doi">10.1007/s10639-023-11617-8</pub-id>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Avcu</surname> <given-names>YE</given-names></string-name>, <string-name><surname>Ayverdi</surname> <given-names>L</given-names></string-name></person-group>. <article-title>Examination of the computer programming self-efficacy&#x2019;s prediction towards the computational thinking skills of the gifted and talented students</article-title>. <source>Int J Educ Methodol</source>. <year>2020</year>;<volume>6</volume>(<issue>2</issue>):<fpage>259</fpage>&#x2013;<lpage>70</lpage>. doi:<pub-id pub-id-type="doi">10.12973/ijem.6.2.259</pub-id>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Schultz</surname> <given-names>O</given-names></string-name>, <string-name><surname>Blaszczyk</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Introduction of process in embedded programming supporting students&#x2019; self-efficacy&#x2014;case study</article-title>. In: <conf-name>Towards a new future in engineering education, new scenarios that european alliances of tech universities open up</conf-name>. <publisher-loc>Barcelona, Spain</publisher-loc>: <publisher-name>Universitat Polit&#x00E8;cnica de Catalunya</publisher-name>; <year>2022</year>. p. <fpage>1610</fpage>&#x2013;<lpage>8</lpage>. doi:<pub-id pub-id-type="doi">10.5821/conference-9788412322262.1200</pub-id>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>El Khoury</surname> <given-names>J</given-names></string-name>, <string-name><surname>Safa</surname> <given-names>N</given-names></string-name>, <string-name><surname>Khoury</surname> <given-names>J</given-names></string-name>, <string-name><surname>Nasrallah</surname> <given-names>R</given-names></string-name></person-group>. <article-title>Measuring the relationship between self-efficacy beliefs and performance attainments of first year engineering students in the programming course</article-title>. <source>TechRxiv</source>. <year>2023</year>. doi:<pub-id pub-id-type="doi">10.36227/techrxiv.24639594.v1</pub-id>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ng</surname> <given-names>TWH</given-names></string-name>, <string-name><surname>Lucianetti</surname> <given-names>L</given-names></string-name></person-group>. <article-title>Within-individual increases in innovative behavior and creative, persuasion, and change self-efficacy over time: a social-cognitive theory perspective</article-title>. <source>J Appl Psychol</source>. <year>2016</year>;<volume>101</volume>(<issue>1</issue>):<fpage>14</fpage>&#x2013;<lpage>34</lpage>. doi:<pub-id pub-id-type="doi">10.1037/apl0000029</pub-id>; <pub-id pub-id-type="pmid">26052714</pub-id></mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Chao</surname> <given-names>J</given-names></string-name>, <string-name><surname>Atli</surname> <given-names>G</given-names></string-name></person-group>. <article-title>Critical personality traits in successful pair programming</article-title>. In: <conf-name>Proceedings of the AGILE 2006 (AGILE&#x2019;06); 2006 Jul 23&#x2013;28; Minneapolis, MN, USA</conf-name>. <publisher-loc>New York, NY, USA</publisher-loc>: <publisher-name>IEEE</publisher-name>; <year>2006</year>. p. <fpage>5</fpage>&#x2013;<lpage>93</lpage>. doi:<pub-id pub-id-type="doi">10.1109/AGILE.2006.20</pub-id>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Hannay</surname> <given-names>JE</given-names></string-name>, <string-name><surname>Arisholm</surname> <given-names>E</given-names></string-name>, <string-name><surname>Engvik</surname> <given-names>H</given-names></string-name>, <string-name><surname>Sjoberg</surname> <given-names>DIK</given-names></string-name></person-group>. <article-title>Effects of personality on pair programming</article-title>. <source>IEEE Trans Softw Eng</source>. <year>2010</year>;<volume>36</volume>(<issue>1</issue>):<fpage>61</fpage>&#x2013;<lpage>80</lpage>. doi:<pub-id pub-id-type="doi">10.1109/TSE.2009.41</pub-id>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Karimi</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Baraani-Dastjerdi</surname> <given-names>A</given-names></string-name>, <string-name><surname>Ghasem-Aghaee</surname> <given-names>N</given-names></string-name>, <string-name><surname>Wagner</surname> <given-names>S</given-names></string-name></person-group>. <article-title>Influence of personality on programming styles an empirical study</article-title>. <source>J Inf Technol Res</source>. <year>2015</year>;<volume>8</volume>(<issue>4</issue>):<fpage>38</fpage>&#x2013;<lpage>56</lpage>. doi:<pub-id pub-id-type="doi">10.4018/jitr.2015100103</pub-id>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Haenggli</surname> <given-names>M</given-names></string-name>, <string-name><surname>Hirschi</surname> <given-names>A</given-names></string-name></person-group>. <article-title>Career adaptability and career success in the context of a broader career resources framework</article-title>. <source>J Vocat Behav</source>. <year>2020</year>;<volume>119</volume>(<issue>1</issue>):<fpage>103414</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.jvb.2020.103414</pub-id>.</mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yilmaz</surname> <given-names>R</given-names></string-name>, <string-name><surname>Karaoglan Yilmaz</surname> <given-names>FG</given-names></string-name></person-group>. <article-title>The effect of generative artificial intelligence (AI)-based tool use on students&#x0027; computational thinking skills, programming self-efficacy and motivation</article-title>. <source>Comput Educ Artif Intell</source>. <year>2023</year>;<volume>4</volume>:<fpage>100147</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.caeai.2023.100147</pub-id>.</mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Massaty</surname> <given-names>MH</given-names></string-name>, <string-name><surname>Fahrurozi</surname> <given-names>SK</given-names></string-name>, <string-name><surname>Budiyanto</surname> <given-names>CW</given-names></string-name></person-group>. <article-title>The role of AI in fostering computational thinking and self-efficacy in educational settings: a systematic review</article-title>. <source>Indones J Inform Educ</source>. <year>2024</year>;<volume>8</volume>(<issue>1</issue>):<fpage>49</fpage>. doi:<pub-id pub-id-type="doi">10.20961/ijie.v8i1.89596</pub-id>.</mixed-citation></ref>
</ref-list>
</back></article>