Inicio
todoELE
  • Inicio
  • Materiales
    • πŸ“‹ Actividades
    • πŸ“ Conjugación
    • πŸ“Š Corpus
    • πŸ“” Diccionarios
    • βœ… Evaluación
    • βš™οΈ Gramática
    • πŸ“— Manuales
    • ✍️ Ortografía
    • πŸ“… Programación
    • πŸ—£οΈ Pronunciación
    • πŸ“ Recursos
    • πŸ”€ Vocabulario
    • πŸ’» Herramientas digitales
  • Formación
    • πŸ“š Bibliografía
    • πŸ‘₯ Congresos
    • πŸŽ“ Cursos
    • 🏫 Centros
    • 🏒 Organizaciones
    • πŸ“° Revistas
    • 🌍 Atlas de ELE
  • Trabajo
    • πŸ’Ό Ofertas de trabajo
    • ℹ️ Trabajo - Recursos
  • En la red
    • 🌐 Sitios ELE
    • πŸ“° Agregador
    • πŸ“§ Formespa
  • IA
    • ✨ Nuevos contenidos
    • πŸ“š Bibliografía IA
    • 🧰 Herramientas IA
    • πŸ’¬ Prompts
    • πŸ§ͺ Experiencias IA
    • 🌐 Sitios web IA
    • πŸ“° Actualidad IA
  • Comunidad
    • πŸ“° Actualidad ELE
    • 😊 Anécdotas ELE
    • πŸ“ Blog
    • πŸ“ŒTablón de anuncios
  • Buscar

Ruta de navegación

  • Inicio
  • Bibliografia
  • Investigating the affordances of OpenAI's large language model in developing listening assessments

Sección IA: Inteligencia artificial Bibliografía

Investigating the affordances of OpenAI's large language model in developing listening assessments

Vahid Aryadoust
Azrifah Zakaria
Yichen Jia
2024
Computers & Education: Artificial Intelligence
6
https://www.sciencedirect.com/science/a…
artículo
estudio empírico
inteligencia artificial
comprensión oral
evaluación
creación de materiales
ChatGPT
grandes modelos de lenguaje
enseñanza/aprendizaje de lenguas
herramientas
tecnología educativa
IA y enseñanza-aprendizaje de lenguas
IA y evaluación
IA y creación de materiales
prompts
chatGPT
estudio empírico

Texto completo

To address the complexity and high costs of developing listening tests for test-takers of varying proficiency levels, this study investigates the capabilities of an OpenAI's large language model, ChatGPT 4, in developing listening assessments. Employing prompt engineering and fine-tuning of prompts, the study specifically focuses on creating listening scripts and test items using ChatGPT 4 for test-takers across a spectrum of proficiency levels (academic, low, intermediate, and advanced). For comparability, the 24 topics of these scripts were selected from topics found in academic listening tests. We conducted two types of analyses to evaluate the quality of the output. First, we performed linguistic analyses of the scripts using Coh-Metrix and Text Inspector to determine if the scripts varied linguistically as required by the prompts. Second, we analyzed topic variation and the degree of overlap in the test items. Results indicated that while ChatGPT 4 reliably produced scripts with significant textual variations, the test items generated were often long and exhibited semantic overlaps among options. This effect was also influenced by the topic. We discuss the ethical complexities that arise from the use of generative artificial intelligence (AI), and how generative AI (GenAI) can potentially benefit practitioners and researchers in language assessment, while recognizing its limitations.

Texto completo en abierto (CC BY-NC-ND 4.0).
  • Inicie sesión para enviar comentarios

Enviar publicación

Contenidos relacionados

  • How to train your dragon: Evaluating prompting and fine-tuning for GPT-based item generation in L2 listening assessment
  • Improving EFL students’ cultural awareness: Reframing moral dilemmatic stories with ChatGPT
  • Evaluating the psychometric properties of ChatGPT-generated questions
  • Analysis of LLMs for educational question classification and generation
  • Reinventing assessments with ChatGPT and other online tools: Opportunities for GenAI-empowered assessment practices
  • Evaluating the potential of ChatGPT-reformulated essays as written feedback in L2 writing
Sobre Todoele Índice Publica Contacto: todoele@gmail.com
Política de privacidad Créditos