arrow
Return

Utilizing Multimodal Large Language Models for Video Analysis of Posture in Studying Collaborative Learning: A Case Study

delete2025-03-19
delete1
delete
OA
AI
R
Ridwan Whitehead *
A
Andy Nguyen
S
Sanna Järvelä
DOI:10.18608/jla.2025.8595delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Incorporating non-verbal data streams is essential to understanding the dynamics of interaction within collaborative learning environments in which a variety of verbal and non-verbal modes of communication intersect. However, the complexity of non-verbal data - especially gathered in the wild from collaborative learning contexts - demands efficient and effective analysis. Methodological advancements are necessary to handle this complexity, enabling researchers to derive meaningful insights from these data streams. The advancement of Generative Artificial Intelligence (GenAI) has significantly broadened its accessibility, making it available to a diverse array of users and demonstrating its utility in aiding data analytics. However, the application of GenAI in multimodal learning analytics, particularly within the context of feature extraction for studying collaborative learning interactions, remains unexplored. This study aims to explore how multimodal large language models (MLLMs) can be utilized as part of the multimodal learning analytics (MMLA) process, focusing on the extraction of postural behaviour. The study focuses on an illustrative case study involving 52 pre-service teachers engaged in a physics-based collaborative learning task, demonstrating how MLLMs can be used for feature extraction. The integration of GenAI techniques in learning research promises aAnew horizon in understanding and enhancing collaborative learning interactions.
Keywords:
(MLLMs)
collaborative learning

Journal

Journal of Learning Analytics cover
Journal of Learning Analytics
IF:
3.6
Papers:
47
Citations:
1.1K

Organization

U
Univ Oulu
Scholars:
532
Papers: 290
Citations: 90