AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.3 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese

How to Conduct Large-Scale Assessment of Creative Thinking?——Exploring a “Human-in-the-Loop” Human-Machine Collaborative Scoring Mode Supported by Large Language Model

Lu ZHANG1,2Jing LU1Xiao-Qing GU3
Shanghai Academy of Educational Sciences, Shanghai, China 200032
Department of Education, Wenzhou University, Wenzhou, Zhejiang, China 325000
Faculty of Education, East China Normal University, Shanghai, China 200062
Show Author Information

Abstract

Creative thinking, as a key element in the cultivation of top-notch innovative talents, its scientific assessment is an important direction for the reform of educational evaluation. The PISA 2022 Creative Thinking Assessment Framework provides a mature reference for large-scale standardized evaluation. However, manual scoring has problems such as difficult standard unification, low efficiency and poor quality guarantee, which urgently needs technological empowerment. Based on this, the paper constructed a “human-in-the-loop” human-machine collaborative scoring mode supported by large language model, and explained the implementation mechanisms of applying this mode for creative thinking assessment from three aspects of theoretical basis, key technologies and practical paths. Subsequently, taking the Shanghai student creative thinking assessment project as a case study, this paper verified the effectiveness of this mode through comparative analysis of relevant data from both manual and human-machine collaborative scoring methods. It was found that both scoring methods were reliable, while the human-machine collaborative scoring showed higher factor loading values and discrimination values compared to manual scoring, and reduced the workload per human rate by approximately 66.7% in the human-machine collaborative scoring. The adaptability of human-computer collaborative assessment was influenced by the degree of task structuring, and expression-type questions required deeper human intervention. The research in this paper provided a human-machine collaborative pathway and empirical basis that can balance quality and efficiency for large-scale standardized assessment of creative thinking, and can also offer a reference for the standardized promotion of more extensive and higher-order thinking assessment enabled by artificial intelligence.

CLC number: G40-057 Document code: A Article ID: 1009-8097(2026)04-0083-09

References

【1】
【1】
 
 
Modern Educational Technology
Pages 83-91

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
ZHANG L, LU J, GU X-Q. How to Conduct Large-Scale Assessment of Creative Thinking?——Exploring a “Human-in-the-Loop” Human-Machine Collaborative Scoring Mode Supported by Large Language Model. Modern Educational Technology, 2026, 36(4): 83-91. https://doi.org/10.3969/j.issn.1009-8097.2026.04.009

5

Views

0

Downloads

0

Crossref

Received: 20 July 2025
Published: 01 April 2026
© The journal of Modern Educational Technology