Journal
Assessing Writing
Published
2026-10-01
DOI
10.1016/j.asw.2026.101107
CompPile
Open Access
Closed
Topics
Export

Citation context

Cited by in this index (1)

  1. Assessing Writing

References (70) · 6 in this index

  1. Standards for educational and psychological testing
  2. Automatic item generation unleashed: An evaluation of a large-scale deployment of item models
    International Conference on Artificial Intelligence in Education
  3. Automated essay scoring with e-rater® V. 2
    The Journal of Technology, Learning and Assessment
  4. Barber, J., & Wolfe, E. (2024). Use of data augmentation in automated essay scoring training data. Paper pres…
  5. The criticality of implementing principled design when using AI technologies in test deve…
    Language Assessment Quarterly  ↗
Show all 70 →
  1. Implementing a contributory scoring approach for the GRE® Analytical Writing section: A c…
    ETS Research Report Series  ↗
  2. Comparison of human and machine scoring of essays: Differences by gender, ethnicity, and …
    Applied Measurement in Education  ↗
  3. Bulut, O., Beiting-Parrish, M., Casabianca, J.M., Slater, S.C., Jiao, H., Song, D., Ormerod, C., Fabiyi, D.G.…
  4. Burstein, J., & LaFlair, G.T. (2024). Where assessment validation and responsible AI meet. arXiv. https://arx…
     ↗
  5. Argument-based validation in testing and assessment
  6. An empirical survey off data augmentation for limited data learning in NLP
    Transactions of the Association for Computational Linguistics  ↗
  7. Validity: An integrated approach to test score meaning and use
  8. Improving automated evaluation of student text responses using GPT-3.5 for text data augm…
    Artificial Intelligence in Education. AIED 2023
  9. Deane, P. (2011). Writing assessment and cognition (Research Report No. RR-11-14). Princeton, NJ: ETS.
     ↗
  10. Fang, L., Lee, G., & Zhai, X. (2023). Using GPT-4 to augment unbalanced data for automatic scoring. arXiv pre…
  11. Validity arguments for AI-based automated scores: Essay scoring as an illustration
    Journal of Educational Measurement  ↗
  12. Firoozi, T., Bulut, O., & Gorgun, G. (2024, April). Text augmentation for enhancing the accuracy of automated…
  13. A workflow for minimizing errors in template-based automated item-generation development
    Educational Measurement: Issues and Practice  ↗
  14. Bias and fairness in large language models: A survey
    Computational Linguistics  ↗
  15. Developing and validating test items
  16. Transforming assessment: The impacts and implications of large language models and genera…
    Educational Measurement: Issues and Practice  ↗
  17. Exploring the effectiveness of large-scale automated writing evaluation implementation on…
    Journal of Computer Assisted Learning  ↗
  18. Exploring the effect of human error when using expert judgments to train an automated sco…
    Educational Measurement: Issues and Practice  ↗
  19. International Test Commission and Association of Test Publishers (2025). Guidelines for technology-based asse…
  20. Validation
    Educational measurement
  21. Validating the interpretations and uses of test scores
    Journal of Educational Measurement  ↗
  22. Explicating validity
    Assessment in Education: Principles, Policy & Practice
  23. Kolen, M. (2011). Comparability issues associated with assessment for the common core state standards. Paper …
  24. Simulating text understanding for educational applications with latent semantic analysis:…
    Interactive Learning Environments  ↗
  25. Validity, fairness, and technology-based assessment
    Advancing natural language processing in educational assessment
  26. Validity and validation
    Educational measurement
  27. Lee, G., Fang, L., & Zhai, X. (2024, April). Improving machine scoring performance with unbalanced training d…
  28. Comparative analysis of psychometric frameworks and properties of scores from autogenerat…
    Educational Measurement: Issues and Practice  ↗
  29. Mansour, W.A., Albatarni, S., Eltanbouly, S., & Elsayed, T. (2024, May). Can large language models automatica…
     ↗
  30. Best practices for constructed-response scoring
    ETS Research Report Series  ↗
  31. Exploring the potential of using an AI language model for automated essay scoring
    Research Methods in Applied Linguistics  ↗
  32. Automated scoring of constructed response items in math assessment using large language models
    International Journal of Artificial Intelligence in Education  ↗
  33. National Council of Teachers of English. (2013, April). NCTE position statement on machine scoring. http://ww…
  34. Open AI (2026). ChatGPT 5.4 [Large language model]. OpenAI. https://platform.openai.com.
  35. OpenAI. (2022). GPT3.5 [Large language model]. OpenAI. https://platform.openai.com.
  36. Ormerod, C.M., & Kwako, A. (2024). Automated text scoring in the age of generative AI for the GPU-poor. arXiv…
     ↗
  37. Automated essay evaluation at scale: Hybrid automated scoring/hand scoring in the summati…
    The Routledge International Handbook of Automated Essay Evaluation
  38. Palermo, C., He, Y., Justice, D., & Katula, P. (2025, October). Improving automated scoring accuracy through …
  39. Journal of Writing Research
  40. EssayGAN: Essay data augmentation based on generative adversarial networks for automated …
    Applied Sciences
  41. Assessing Writing
  42. Comparing the validity of automated and human scoring of essays
    Journal of Educational Computing Research  ↗
  43. Do LLMs write like humans? Variation in grammatical and rhetorical styles
    Proceedings of the National Academy of Sciences  ↗
  44. Applications of automated essay evaluation in West Virginia
    Handbook of automated essay evaluation: Current application and new directions
  45. Theory into practice: Reflections on the handbook
    Handbook of automated scoring
  46. Rupp, A.A., & Lorié, W. (2023). Ready or not: AI is changing assessment and accountability. Center for Assess…
  47. Decoding the future: A guide to choosing AI for test development
  48. A review of automatic item generation techniques leveraging large language models
    International Journal of Assessment Tools in Education  ↗
  49. Synonym-based essay generation and augmentation for robust automatic essay scoring
    Intelligent Data Engineering and Automated Learning
  50. Assessing Writing
  51. Wang, Y., Wu, X., Huang, J., Liu, L., Zhai, X., & Liu, N. (2026). BRIDGE the gap: mitigating bias amplificati…
     ↗
  52. Assessing Writing
  53. Establishing fairness when developing educational tests
    Fairness in educational and psychological testing: examining theoretical, research, practice and policy implications of 2014 Standards
  54. A framework for evaluation and use of automated scoring
    Educational Measurement: Issues and Practice  ↗
  55. Assessing Writing
  56. Elementary English learners’ engagement with automated feedback
    Learning and Instruction  ↗
  57. Wolfe, E.W., & Barber, J.O. (2026). Calibrating generative AI to produce realistic essays for data augmentati…
  58. Public perception and communication around automated essay scoring
    Handbook of automated scoring: Theory into practice
  59. On the consistency of automatic scoring with large language models
    Educational and Psychological Measurement  ↗
  60. Contrasting automated and human scoring of essays
    R & D Connections
  61. Journal of Writing Research
  62. Differential feature functioning in automated essay scoring
    Test fairness in the new generation of large-scale assessment
  63. Character-level convolutional networks for text classification
    Advances in Neural Information Processing Systems
  64. AI-generated essays: characteristics and implications on automated scoring and academic i…
    Educational Measurement: Issues and Practice
  65. Test design and development
    Fairness in educational assessment and measurement