Journal
Assessing Writing
Published
2026-04-01
DOI
10.1016/j.asw.2026.101044
CompPile
Open Access
Closed
Topics
Export

Citation context

Cited by in this index (0)

No articles in this index cite this work.

References (65) · 5 in this index

  1. Evaluating the quality of AI feedback: A comparative study of AI and human essay grading
    Innovations in Education and Teaching International  ↗
  2. Validity and Reliability of Automated Essay Scoring
    Handbook of automated essay evaluation: current applications and new directions
  3. Transforming education: A comprehensive review of generative artificial intelligence in e…
    Sustainability  ↗
  4. Exploring ChatGPT as a writing assessment tool
    Innovations in Education and Teaching International  ↗
  5. ChatGPT as an automated essay scoring tool in the writing classrooms: How it compares wit…
    Education and Information Technologies  ↗
Show all 65 →
  1. Application of an Automated Essay Scoring engine to English writing assessment using Many…
    Language Testing  ↗
  2. The TOFEL validity argument
    Building a Validity Argument for the Test of English as a Foreign Language™
  3. Validity arguments for diagnostic assessment using automated writing evaluation
    Language Testing  ↗
  4. Statistical power analysis for the behavioral sciences
  5. Evaluating quadratic weighted kappa as the standard performance metric for automated essa…
    16th International Conference on Educational Data Mining. Germany
  6. Practical considerations for using AI models in automated scoring of writing
    Application of artificial intelligence to assessment
  7. Assessing second-language academic writing: AI vs. human raters
    Journal of Educational Technology & Online Learning  ↗
  8. Has artificial intelligence rendered language teaching obsolete?
    Modern Language Journal  ↗
  9. Validity Arguments for Automated Essay Scoring of Young Students’ Writing Traits
    Language Assessment Quarterly  ↗
  10. Assessing Writing
  11. Automated language essay scoring systems: a literature review
    PeerJ Computer Science  ↗
  12. VADER: A Parsimonious rule-based model for sentiment analysis of social media text
  13. When AI meets source use: Exploring ChatGPT’s potential in L2 summary writing assessment
    System  ↗
  14. Automated essay scoring: a survey of the state of the art
    Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence (IJCAI-19)
  15. Automated Essay Scoring With GPT-4 for a Local Placement Test: Investigating Prompting St…
    TESOL Quarterly  ↗
  16. ChatGPT for writing evaluation: Examining the accuracy and reliability of AI-generated sc…
    Exploring Artificial Intelligence in Applied Linguistics
  17. The potential advantages of using an LLM-based chatbot for automated writing evaluation f…
    Language Learning & Technology
  18. A Guideline of Selecting and Reporting Intraclass Correlation Coefficients for Reliabilit…
    Journal of Chiropractic Medicine  ↗
  19. The measurement of observer agreement for categorical data
    Biometrics  ↗
  20. Applying large language models and chain-of-thought for automatic scoring
    Computers and Education: Artificial Intelligence
  21. Evaluating the role of ChatGPT in enhancing EFL writing assessments in classroom settings…
    Humanities & Social Science Communications  ↗
  22. Leveraging ChatGPT for Second Language Writing Feedback and Assessment
    International Journal of Computer-Assisted Language Learning and Teaching  ↗
  23. Enhancing GPT-based automated essay scoring: the impact of fine-tuning and linguistic com…
    Computer Assisted Language Learning
  24. Assessing Writing
  25. Sentiment analysis methods, applications, and challenges: A systematic literature review
    Journal of King Saudade University - Computer and Information Sciences
  26. Exploring the potential of using an AI language model for automated essay scoring
    Research Methods in Applied Linguistics  ↗
  27. Testing the viability of ChatGPT as a companion in L2 writing accuracy assessment
    Research Methods in Applied Linguistics  ↗
  28. Automated evaluation of written discourse coherence using GPT-4
    Proceedings of the 18th Workshop on Innovative Use of NLP for Building Educational Applications
  29. OpenAI. (2023a). GPT-3.5. 〈https://platform.openai.com/docs/models/gpt-3-5〉.
  30. OpenAI. (2023b). Fine-tuning GPT models. 〈https://platform.openai.com/docs/guides/fine-tuning〉.
  31. An Empirical Study of the Non-Determinism of ChatGPT in Code Generation
    ACM Transactions on Software Engineering and Methodology  ↗
  32. Large language models and automated essay scoring of English language learner writing: In…
    Computers and Education: Artificial Intelligence
  33. PRISMA 2020 explanation and elaboration: Updated guidance and exemplars for reporting sys…
    The BMJ
  34. Can ChatGPT reliably and accurately apply a rubric to L2 writing assessments? The devil i…
    Journal of Technology and Chinese Language Teaching
  35. An automated essay scoring systems: A systematic literature review
    Artificial Intelligence Review  ↗
  36. Assessing Writing
  37. Examining the consistency of instructor versus large language model ratings on summary co…
    Language Testing  ↗
  38. Teacher or CHATGPT: The issue of accuracy and consistency in L2 assessment
    Teaching English with Technology
  39. Shermis, M., & Wilson, J. (2024). Introduction to automated essay evaluation. In M. Shermis & J. Wilson (Eds.…
     ↗
  40. Exploratory study on the potential of ChatGPT as a rater of second language writing
    Education and Information Technologies  ↗
  41. Comparing the quality of human and ChatGPT feedback of students’ writing
    Learning and Instruction  ↗
  42. Stureborg, R., Alikaniotis, D., & Suhara, Y. (2024). Large language models are inconsistent and biased evalua…
  43. Assessing Writing
  44. A comparative study of the human, automated scoring model, and GPT-4 ratings of young EFL…
    Language Testing  ↗
  45. Is generative AI ready to replace human raters in scoring EFL writing? Comparison of huma…
    Educational Technology & Society
  46. Artificial intelligence as an automated essay scoring tool: A focus on ChatGPT
    International Journal of Assessment Tools in Education  ↗
  47. A new interpretation of the weighted kappa coefficients
    Psychometrika  ↗
  48. Assessing Writing
  49. Comparison of traditional machine learning and neural network approaches for automated sc…
    Language Testing  ↗
  50. The use of assistive technologies including generative AI by test takers in language asse…
    Language Assessment Quarterly  ↗
  51. Fairness
    The Routledge handbook of language testing
  52. Effectiveness of large language models in automated evaluation of argumentative essays: f…
    Computer Assisted Language Learning  ↗
  53. A Framework for Evaluation and Use of Automated Scoring
    Educational Measurement: Issues and Practice  ↗
  54. Validity and the automated scoring of performance tests
    The Routledge handbook of language testing
  55. Advancing Language Assessment with AI and ML – Leaning into AI is Inevitable, but Can The…
    Language Assessment Quarterly  ↗
  56. Rating short L2 essays on the CEFR scale with GPT-4
    Proceedings of the 18th Workshop on Innovative Use of NLP for Building Educational Applications
  57. The Reliability of using ChatGPT in Rating EFL Writings
    Shanlax International Journal of Education  ↗
  58. An application of many-facet Rasch measurement to evaluate automated essay scoring: A cas…
    Research Methods in Applied Linguistics
  59. Exploring potential biases in GPTGPT-4o’s ratings of English language learners’ essays
    Language Testing  ↗
  60. Utilizing large language models for EFL essay grading: An examination of reliability and …
    British Journal of Educational Technology  ↗