Dalarna University's logo and link to the university's website

du.sePublications
Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • chicago-author-date
  • chicago-note-bibliography
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Artificial Intelligence Agents as Point of Contact Mediators in Student & Teacher Communication
Dalarna University, School of Information and Engineering.
2025 (English)Independent thesis Advanced level (degree of Master (Two Years)), 10 credits / 15 HE creditsStudent thesis
Abstract [en]

Teachers in higher education often face issue with enormous emails from student regarding academic queries. However, many of these questions are repetitive and can be found in course handbook which is already uploaded on the website. This redundant communication load not only consumes valuable time of teachers also sometimes student may not get the responses on time. This thesis explores the use of Artificial Intelligence (AI) agents particularly Large Language Models (LLMs) as a mediator between student-teacher communication by answering to the student questions by solely based on course manual.

This study assesses two modern LLMs, OpenAI’s GPT-4o-mini (OpenAI, 2024) and Meta’s LLAMA 3.2 (Di Palma et al., 2024). A balanced selection of both answerable and unanswerable questions was included in each of the five course handbooks from a recognized academic field. To highlight different models’ behaviours, final accuracy comparisons concentrated on Prompt 1(loose) and Prompt 3(strict), although a three-stage prompt design (White et al., 2023) was evaluated. A Likert-scale evaluation (1-3) was used to validate the model’s outputs, which were then categorized into (TP, TN, FN, FP) and confirmed by a combination of manual and LLM assisted review.

The results revealed that LLaMA 3.2 achieved highest accuracy with strict prompt, with the lowest average errors rates (FN and FP) and the highest average correct responses (TP+TN) while GPT-4o-mini remained prone errors. This emphasise how important prompt clarity and instruction tuning (Raina, Liusie, & Gales, 2024) are. Although, encouraging the study’s scope is constrained by its manual review and handbook structure, indicating the necessity of more extensive validation and automation in future work.

Place, publisher, year, edition, pages
2025.
Keywords [en]
Artificial Intelligence, Large Language Models, Prompt Engineering, Handbook- based Evaluation, Instruction-tuned LLMs, GPT-4o-mini, LLaMA 3.2, Likert Scale, Confusion Matrix, Fallback Handling
National Category
Computer and Information Sciences
Identifiers
URN: urn:nbn:se:du-51111OAI: oai:DiVA.org:du-51111DiVA, id: diva2:1991263
Subject / course
Data Analytics
Available from: 2025-08-22 Created: 2025-08-22 Last updated: 2025-10-09

Open Access in DiVA

fulltext(678 kB)149 downloads
File information
File name FULLTEXT01.pdfFile size 678 kBChecksum SHA-512
d2c19134a31ace174c0af10615ca920a034b896a42b690e44afe35593dca99dc45b643f7ed96a246910e3ac0c95955632d880374ccba0bc18b67c73fd4ff4ed7
Type fulltextMimetype application/pdf

By organisation
School of Information and Engineering
Computer and Information Sciences

Search outside of DiVA

GoogleGoogle Scholar
Total: 152 downloads
The number of downloads is the sum of all downloads of full texts. It may include eg previous versions that are now no longer available

urn-nbn

Altmetric score

urn-nbn
Total: 158 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • chicago-author-date
  • chicago-note-bibliography
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf