Computer Science > Artificial Intelligence

arXiv:2407.20360 (cs)

[Submitted on 29 Jul 2024]

Title:Evaluating Large Language Models for automatic analysis of teacher simulations

Authors:David de-Fitero-Dominguez, Mariano Albaladejo-González, Antonio Garcia-Cabot, Eva Garcia-Lopez, Antonio Moreno-Cediel, Erin Barno, Justin Reich

View PDF HTML (experimental)

Abstract:Digital Simulations (DS) provide safe environments where users interact with an agent through conversational prompts, providing engaging learning experiences that can be used to train teacher candidates in realistic classroom scenarios. These simulations usually include open-ended questions, allowing teacher candidates to express their thoughts but complicating an automatic response analysis. To address this issue, we have evaluated Large Language Models (LLMs) to identify characteristics (user behaviors) in the responses of DS for teacher education. We evaluated the performance of DeBERTaV3 and Llama 3, combined with zero-shot, few-shot, and fine-tuning. Our experiments discovered a significant variation in the LLMs' performance depending on the characteristic to identify. Additionally, we noted that DeBERTaV3 significantly reduced its performance when it had to identify new characteristics. In contrast, Llama 3 performed better than DeBERTaV3 in detecting new characteristics and showing more stable performance. Therefore, in DS where teacher educators need to introduce new characteristics because they change depending on the simulation or the educational objectives, it is more recommended to use Llama 3. These results can guide other researchers in introducing LLMs to provide the highly demanded automatic evaluations in DS.

Subjects:	Artificial Intelligence (cs.AI)
Cite as:	arXiv:2407.20360 [cs.AI]
	(or arXiv:2407.20360v1 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2407.20360

Submission history

From: Mariano Albaladejo-González [view email]
[v1] Mon, 29 Jul 2024 18:19:17 UTC (116 KB)

Computer Science > Artificial Intelligence

Title:Evaluating Large Language Models for automatic analysis of teacher simulations

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Evaluating Large Language Models for automatic analysis of teacher simulations

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators