Inicio
todoELE
  • Inicio
  • Materiales
    • πŸ“‹ Actividades
    • πŸ“ Conjugación
    • πŸ“Š Corpus
    • πŸ“” Diccionarios
    • βœ… Evaluación
    • βš™οΈ Gramática
    • πŸ“— Manuales
    • ✍️ Ortografía
    • πŸ“… Programación
    • πŸ—£οΈ Pronunciación
    • πŸ“ Recursos
    • πŸ”€ Vocabulario
    • πŸ’» Herramientas digitales
  • Formación
    • πŸ“š Bibliografía
    • πŸ‘₯ Congresos
    • πŸŽ“ Cursos
    • 🏫 Centros
    • 🏒 Organizaciones
    • πŸ“° Revistas
    • 🌍 Atlas de ELE
  • Trabajo
    • πŸ’Ό Ofertas de trabajo
    • ℹ️ Trabajo - Recursos
  • En la red
    • 🌐 Sitios ELE
    • πŸ“° Agregador
    • πŸ“§ Formespa
  • IA
    • ✨ Nuevos contenidos
    • πŸ“š Bibliografía IA
    • 🧰 Herramientas IA
    • πŸ’¬ Prompts
    • πŸ§ͺ Experiencias IA
    • 🌐 Sitios web IA
    • πŸ“° Actualidad IA
  • Comunidad
    • πŸ“° Actualidad ELE
    • 😊 Anécdotas ELE
    • πŸ“ Blog
    • πŸ“ŒTablón de anuncios
  • Buscar

Ruta de navegación

  • Inicio
  • Bibliografia
  • Can AI provide useful holistic essay scoring?

Sección IA: Inteligencia artificial Bibliografía

Can AI provide useful holistic essay scoring?

Tamara P. Tate
Jacob Steiss
Drew Bailey
Steve Graham
Youngsun Moon
Daniel Ritchie
Waverly Tseng
Mark Warschauer
2024
Computers & Education: Artificial Intelligence
7
https://www.sciencedirect.com/science/a…
artículo
estudio empírico
inteligencia artificial
expresión escrita
evaluación
ChatGPT
grandes modelos de lenguaje
educación primaria y secundaria
herramientas
tecnología educativa
IA y enseñanza-aprendizaje de lenguas
IA y evaluación
chatGPT
estudio empírico

Texto completo

Researchers have sought for decades to automate holistic essay scoring. Over the years, these programs have improved significantly. However, accuracy requires significant amounts of training on human-scored texts—reducing the expediency and usefulness of such programs for routine uses by teachers across the nation on non-standardized prompts. This study analyzes the output of multiple versions of ChatGPT scoring of secondary student essays from three extant corpora and compares it to quality human ratings. We find that the current iteration of ChatGPT scoring is not statistically significantly different from human scoring; substantial agreement with humans is achievable and may be sufficient for low-stakes, formative assessment purposes. However, as large language models evolve additional research will be needed to continue to assess their aptitude for this task as well as determine whether their proximity to human scoring can be improved through prompting or training.

Texto completo en abierto (CC BY-NC 4.0).
  • Inicie sesión para enviar comentarios

Enviar publicación

Contenidos relacionados

  • Investigating the affordances of OpenAI's large language model in developing listening assessments
  • Do teachers spot AI? Evaluating the detectability of AI-generated texts among student essays
  • Can ChatGPT score ESL writing? A correlation analysis between teacher and GenAI scores
  • Comparing expert tutor evaluation of reflective essays with marking by generative artificial intelligence (AI) tool
  • Teacher feedback and ChatGPT feedback on Chinese university EFL learners’ English essay revision: A mixed-methods study
  • Synergizing collaborative writing and AI feedback: An investigation into enhancing L2 writing proficiency in wiki-based environments
Sobre Todoele Índice Publica Contacto: todoele@gmail.com
Política de privacidad Créditos