Purpose

Online learning communities on social media platforms can support peer learning, but educators often lack theoretically grounded and measurable approaches for monitoring how participation and discourse evolve across a semester. This study proposes an extended Community of Inquiry (CoI) evaluation framework that integrates Social, Teaching, and Cognitive Presence with a fourth behavioural dimension, Student Presence.

Design/methodology/approach

A sequential exploratory mixed-method design was adopted. Qualitative analysis of prior literature and semester-long observations of two large first-year engineering course Facebook groups (each enrolling 800–1000 students) informed an indicator-based coding scheme, applied quantitatively over Weeks 1–13. Predictive modelling used a persistence baseline, a multi-output Random Forest, and a multilayer perceptron under time-aware evaluation protocols.

Findings

Social Presence was enquiry-driven and peaked in Weeks 3–4; Teaching Presence was frontloaded and primarily reactive; Cognitive Presence was shallow, dominated by remembering and analysing. Student participation was consumption-oriented, with observers consistently outnumbering posters. Random Forest achieved consistent poster prediction (R2 ˜ 0.48–0.49), while observers and non-members remained difficult to forecast due to structural interdependence. Permutation importance identified remembering and evaluating as the most influential cognitive predictors.

Research limitations/implications

The dataset comprises 13 weekly observations from a single platform and institution, limiting generalisability. Future work should collect multi-cohort data, introduce lagged predictors, and explore individual-level modelling.

Practical implications

The framework provides instructors with an early-warning system for low poster activity, enabling timely, evidence-based interventions to support online peer learning communities.

Originality/value

This study makes three contributions: a multi-dimensional coding scheme grounded in the extended CoI framework; a data-driven analytics pipeline enabling descriptive monitoring and predictive modelling of participation roles; and an integrated evaluation framework that combines theory-grounded indicator coding with transparent machine learning to produce actionable insights from social media learning data.

Digital technology has fundamentally transformed how students learn and collaborate. Online learning communities are now integral to higher education, offering flexible environments that extend classroom interaction beyond physical boundaries (Bruggeman et al., 2021). Within these environments, peer interaction is not incidental — it is foundational. Research consistently shows that both instructor–learner and peer-to-peer exchanges are essential for building knowledge, sustaining engagement, and fostering a sense of belonging (Freeman and Jarvie-Eggart, 2019; Garrison, 2022; Lai et al., 2019).

As online education expands, understanding what makes peer interactions effective has become a pressing concern. However, evaluating the quality and impact of these interactions in digital settings remains an open challenge (Babik et al., 2024). Most existing approaches assess isolated dimensions, such as cognitive gains, while privileging outcomes over the processes through which learning unfolds (Yu and Schunn, 2023). Peer learning is inherently dynamic. It evolves in response to shifting group dynamics, course milestones, and individual participation patterns — none of which static, outcome-focused measures adequately capture (Chen et al., 2024; Moon et al., 2024; Zhang et al., 2025). A unified, multi-dimensional framework that evaluates both the process and progress of peer learning over time is therefore lacking.

To address this gap, this study investigates effective methods for assessing peer learning in online learning environments. A mixed-methods investigation was conducted at an Australian university, where student and instructor interactions were analysed across a full semester. The study context comprised two compulsory first-year engineering courses at The University of Queensland, using Facebook-based course discussion groups as the online platform. This setting provided a rich, naturalistic dataset spanning thirteen weeks, enabling analysis of both the processes and outcomes of student participation.

The proposed solution is a comprehensive evaluation framework with two integrated components. The first is a theory-grounded coding scheme informed by the Community of Inquiry (CoI) model (Garrison et al., 1999) and extended through the inclusion of student presence (Dokhanchi et al., 2018). It operationalises four dimensions of peer learning discourse — student, cognitive, teaching, and social presences — into measurable weekly indicators. The second is a data-driven analytics pipeline that couples with the coding scheme to support descriptive monitoring of engagement patterns and exploratory prediction of student participation roles over time. Together, these components equip instructors and researchers with a practical tool for tracking how online peer learning communities develop and identifying which discourse patterns are associated with active, meaningful participation.

The contributions of this study are as follows:

  1. A multi-dimensional coding scheme grounded in the CoI framework and extended with student presence, operationalising four dimensions of peer learning discourse into measurable weekly indicators for systematic evaluation.

  2. A data-driven analytics pipeline coupled with the coding scheme to enable descriptive monitoring of weekly engagement patterns and predictive modelling of student participation roles over time.

  3. A practical, integrated evaluation framework that empowers instructors and researchers to analyse learner–learner and learner–instructor interactions, inform instructional design, and identify opportunities for improving online learning communities.

The remainder of this paper is organised as follows. Section 2 reviews related work. Section 3 details the methodology. Section 4 presents the experiments and results. Section 5 discusses limitations and future directions. Section 6 concludes the paper.

Researchers have previously highlighted the importance of interaction in online learning environments (Freeman and Jarvie-Eggart, 2019; Lai et al., 2019). Collaboration emerging from learner–learner and learner–instructor interactions is the key contributor to successful learning (Palloff and Pratt, 1999), as such interactions help learners test ideas, receive feedback, and refine their understanding (Wagner, 1997). While interactions positively influence students' satisfaction and engagement (Liu et al., 2017), they must be structured and cohesive to enable higher levels of critical thinking and knowledge construction (Garrison and Cleveland-Innes, 2005).

Multiple scholars have attempted to categorise types of interaction. For example, Wagner (1997) focused on learners' achievement, whereas Hare et al. (1994) distinguished between task-driven and socio-emotional interactions. Task-driven interactions emphasise completing assigned work and tend to be instructor-oriented, while socio-emotional interaction focuses on relationships between learners (Rovai, 2002) and incorporates individual factors such as personality and emotions (Delahunty et al., 2014).

There are four types of interactions in online learning: learner–learner, learner–instructor (Moore, 1989, 1993), learner–content (Liu and Kaye, 2016), and learner–interface (Hillman et al., 1994). Table 1 summarises these four widely recognised types of interactions in online learning environments.

Table 1

Types of interaction in online learning environments

Interaction typeDefinitionCitation
Learner–learnerInteraction between one learner and other learners, with or without an instructor presentMoore (1989) 
Learner–instructorInteraction between the learner and an expert who prepared or delivers course materialMoore (1989) 
Learner–contentCognitive or perceptual contact between students and study materials that results in the acquisition of meaningLiu and Kaye (2016) 
Learner–interfaceThe process of manipulating tools to accomplish a learning taskHillman et al. (1994) 

Given the focus of this study on peer learning, the review now turns to the two dimensions most relevant to peer learning: learner–learner and learner–instructor interactions.

Learner–learner interactions in online learning communities play a critical role in fostering a sense of community (Xie and Yen, 2011), supporting cognitive development (Lin et al., 2017), and strengthening interpersonal skills (Lee, 2002). Although active contributors positively shape collaborative learning, a high proportion of passive participants or “lurkers” can negatively impact group interaction dynamics.

Several frameworks have been proposed to analyse learner–learner online interactions. Bravo et al. (2008) introduced a three-stage model consisting of observation, abstraction, and intervention, where interaction data are first observed and collected then analysed, and finally used to design improvements to collaboration. Similarly, Ke and Xie (2009) developed a framework based on Cercone (2008) deep learning model incorporating social interaction, knowledge construction, and regulation of learning. Vuopala et al. (2016) created a coding scheme distinguishing between task-related and group-related interactions to assess collaborative learning processes.

These frameworks highlight the multidimensional nature of learner–learner interaction, reinforcing the need for evaluation models that capture both social and cognitive aspects of peer learning.

Instructor involvement is essential in online peer learning communities because peer interaction alone does not always lead to productive outcomes (Kanuka and Garrison, 2004; Xie et al., 2018; Zhu, 2006). Instructors structure discussions, guide inquiry, and sustain engagement through feedback (Shute, 2008). Majeski and Stover (2007) show that instructor facilitation shapes the depth of discourse, moving students from surface-level exchange toward higher-order reflection. This study therefore analyses instructor and student interactions in parallel rather than treating them as separate phenomena.

This study adopts the extended Community of Inquiry (CoI) framework, which provides a multidimensional lens for analysing interactions in online learning environments. Garrison et al. (1999) originally identified three interrelated presences—cognitive, social, and teaching—as overlapping processes that shape the quality of the learning experience and determine student engagement in online education. Cognitive presence reflects the extent to which learners develop meaningful understanding through inquiry and reflection. Social presence refers to the learners' ability to present themselves authentically and build connections within the learning community. Teaching presence encompasses the design, facilitation, and instructional direction provided by the instructor. Later extensions to the model introduced Student Presence, which emphasises learners' self-regulation, agency, and active participation in managing their learning (Dokhanchi et al., 2018). A study conducted by Dokhanchi et al. (2018) found that student presence, cognitive presence, teaching presence, and social presence affect peer learning in online learning environments. These four dimensions collectively provide a holistic framework for evaluating the complexity of interactions within online peer learning environments.

2.3.1 Cognitive presence

Garrison et al. (1999) categorised students' cognitive engagements as triggering events, exploration, integration, and resolution. Building on this work, Zhu (2006) examined student cognitive engagement processes in online learning communities, identifying processes such as seeking, interpreting, analysing, summarising information, and making decisions. More recently, Fiock (Fiock (2020) provided a comprehensive review of strategies to enhance cognitive presence in online learning environments, aligning instructional design with the Community of Inquiry framework to support deeper engagement and critical thinking.

2.3.2 Social presence

To measure social presence in online learning communities, Rourke et al. (1999) coded students' messages into affective, interactive, and cohesive categories. Swan and Shih (2005) expanded these categories by adding indicators such as expressing values and social expression. Similarly, Vuopala et al. (2016) analysed students' social interactions by including expressing cohesion, decreasing tension, and accompanying indicators in their coding scheme. Borup et al. (2012) used emotional expressions, open communication, and cohesion categories to analyse social presence. According to Borup et al. (2012), the presence of group cohesiveness indicators in online learning communities indicates the existence of a sense of community among members.

2.3.3 Teaching presence

Teaching presence has been conceptualised in various ways. Berge (1995) coded instructor posts into managerial, social, pedagogical, and technical indicators, while Anderson et al. (2001) organised teachers' presence around instructional design and organisation, facilitating discourse, and direct instruction. More recently, Stenbom (Stenbom (2018) reviewed empirical research on teaching presence in online environments and proposed a refined model highlighting the dynamic and context-dependent nature of instructional roles. Teaching presence is particularly relevant in peer learning contexts, as instructors play a crucial role in structuring discussions, guiding inquiry, and sustaining student engagement.

2.3.4 Student presence

Beyond the original three dimensions of the CoI framework, scholars have increasingly recognised the importance of a fourth dimension, often referred to as Student Presence. Early work by Shea and Bidjerano (2009) introduced the idea of Learner Presence as an extension of the model, emphasising the role of self-regulation and learner agency in online education. Their subsequent study (Shea and Bidjerano, 2012) further positioned Learning Presence as a moderating factor within the CoI framework and provided empirical evidence of its significance. Building on this perspective, Redmond (2014) highlighted the centrality of learner agency and self-directed effort in shaping meaningful engagement. More recently, studies identified Student Presence as the missing presence in the CoI framework, formally proposing it as a distinct and essential dimension (Ng et al., 2021). Similarly, Dokhanchi et al. (2018) adopted Student Presence as the fourth presence, framing it as a reflection of students' active participation and responsibility within online learning communities.

Figure 1 shows the theoretical relationships among the four presences, and situates the extended CoI dimensions within the analytical pipeline of the proposed framework. End-to-end pipeline of the proposed framework is shown in Figure 2. The qualitative phase produces the coding scheme and indicator definitions. The quantitative phase aggregates coded indicators into weekly feature vectors, trains three competing models, and evaluates them under time-aware protocols before extracting interpretability results.

Figure 1
A Venn diagram illustrating the relationships between student presence, cognitive presence, social presence, and teaching presence in online learning environments.A Venn diagram with four overlapping circles representing student presence, cognitive presence, social presence, and teaching presence in online learning environments. The student presence circle includes student participation. The cognitive presence circle includes cognitive engagement, educational content, direct learning, and regulate learning. The social presence circle includes sense of community, learner-learner interaction, and peer learning. The teaching presence circle includes educator activity, set up regulations, and seek help. The intersection of all four circles is labeled educational experience. Each section of the Venn diagram highlights the interactions and overlaps between different types of presence, emphasizing the interconnected nature of these elements in creating a comprehensive educational experience.

Extended CoI framework in online learning environments with student presence

Figure 1
A Venn diagram illustrating the relationships between student presence, cognitive presence, social presence, and teaching presence in online learning environments.A Venn diagram with four overlapping circles representing student presence, cognitive presence, social presence, and teaching presence in online learning environments. The student presence circle includes student participation. The cognitive presence circle includes cognitive engagement, educational content, direct learning, and regulate learning. The social presence circle includes sense of community, learner-learner interaction, and peer learning. The teaching presence circle includes educator activity, set up regulations, and seek help. The intersection of all four circles is labeled educational experience. Each section of the Venn diagram highlights the interactions and overlaps between different types of presence, emphasizing the interconnected nature of these elements in creating a comprehensive educational experience.

Extended CoI framework in online learning environments with student presence

Close modal
Figure 2
A diagram of a data processing framework.The diagram illustrates a data processing framework starting with qualitative data sources such as Facebook group posts and comments, and student membership counts. These sources are encoded using indicator coding and then aggregated weekly. The aggregated data is used as input for three models: Persistence Baseline, Random Forest, and MLP. The outputs of these models are evaluated using time-aware evaluation metrics such as MAE, RMSE, and R squared. The results are further analyzed using permutation importance and sensitivity analysis.

End-to-end pipeline of the proposed framework

Figure 2
A diagram of a data processing framework.The diagram illustrates a data processing framework starting with qualitative data sources such as Facebook group posts and comments, and student membership counts. These sources are encoded using indicator coding and then aggregated weekly. The aggregated data is used as input for three models: Persistence Baseline, Random Forest, and MLP. The outputs of these models are evaluated using time-aware evaluation metrics such as MAE, RMSE, and R squared. The results are further analyzed using permutation importance and sensitivity analysis.

End-to-end pipeline of the proposed framework

Close modal

To address the research question and meet the study objectives, a mixed-method research design was adopted (Creswell and Creswell, 2017). This approach was selected to capture both the interpretive depth of peer-learning discourse and the measurable breadth of behavioural participation in online learning environments. Importantly, the choice of mixed methods aligns with the multi-dimensional nature of the proposed evaluation framework, which integrates Social Presence, Teaching Presence, Cognitive Presence, and Student Presence. The study followed a sequential exploratory strategy in which qualitative insights were used to construct the coding scheme and quantitative analysis was then used to apply, validate, and summarise the resulting indicators at scale.

In the qualitative phase, an extensive review of relevant literature was conducted alongside systematic observation of student and instructor interactions within the online learning communities across the semester. This stage provided an in-depth understanding of how peer learning unfolded over time and guided the iterative development of coding categories and indicators grounded in the extended CoI framework. In particular, qualitative analysis supported the refinement of indicator definitions, the clarification of boundary cases, and the alignment of the coding scheme with the communicative practices typical of social media environments. Building on this foundation, the quantitative strand operationalised the developed evaluation scheme to measure weekly interaction patterns. This enabled the research team to quantify participation levels, characterise the distribution of the four presences, and examine how these dimensions manifested and co-evolved across the semester as part of the peer-learning process.

The study was conducted at the University of Queensland using two compulsory first-year engineering courses. These large, team-based courses, each enrolling approximately 800–1000 students annually, were selected because they rely heavily on collaboration and peer interaction, making them well-suited for investigating online peer learning at scale. Course Facebook groups served as the primary online discussion platforms, offering an authentic, student-driven setting for observing peer support, instructor facilitation, and knowledge construction behaviours. Ethical approval for the study was obtained from the University of Queensland (Approval No. 2018000264), and permission to collect and analyse data was granted by the respective course coordinators.

The proposed framework integrates (1) a theory-grounded coding scheme based on the extended Community of Inquiry (CoI) model (Garrison et al., 1999) and (2) a data-driven modelling and validation layer to quantify, predict, and interpret student participation dynamics in online peer learning communities. The coding scheme was developed through an iterative process informed by the original CoI framework and its later extension incorporating student presence. The categories and indicators were designed to reflect the key factors that affect online peer learning—namely student, cognitive, teaching, and social presences (Dokhanchi et al., 2018)—through: (1) reviewing existing coding schemes in the literature and (2) analysing student and instructor interactions in course Facebook groups across Weeks 1–13. Table 2 summarises the parameters used to guide the development of the framework.

Table 2

Parameters in online learning communities that were used to develop the framework

DimensionParameter
Student presenceLevel of student participation in discussions
Cognitive presenceStudents' course-related posts that include cognitive presence indicators
Social presenceStudents' course and non-course related posts that include social presence indicators
Teaching presenceCourse coordinators' and tutors' course and non-course related posts that include teaching presence indicators

Beyond descriptive analytics, the proposed framework contributes a modelling component that learns relationships between weekly aggregates of Social/Teaching/Cognitive indicators (inputs) and Student Presence participation roles (outputs). Specifically, the modelling layer aims to predict weekly Student Presence role counts (Poster, Observer, Non-Member) from weekly aggregates of Social, Teaching, and Cognitive presence indicators. This predictive component serves two purposes: (1) evaluation of whether the coded indicators contain sufficient signal to forecast participation behaviour and (2) insight generation by identifying which indicators are most strongly associated with student role changes over time.

To ensure methodological rigour for time-ordered data, the evaluation uses (1) a time-based holdout split (train on Weeks 1–10; test on Weeks 11–13) and (2) a walk-forward (rolling-origin) backtest that repeatedly trains on early weeks and predicts the next unseen week. These protocols reduce temporal leakage and provide a more credible estimate of performance under realistic deployment settings.

Table 3 summarises how each presence dimension was extended in this study relative to its original formulation.

Table 3

Theoretical contributions of each presence dimension relative to original CoI formulations

DimensionOriginal formulationExtension in this study
Cognitive presenceTriggering event, exploration, integration, resolution (Garrison et al., 1999)Reframed using Bloom's revised taxonomy (Anderson and Bloom, 2001): six cognitive processes from remembering through creating
Social presenceAffective, interactive, cohesive categories (Rourke et al., 1999)Extended with conversational expression (enquiry, statement, sharing information) and affective expression (paralanguage, self-disclosure)
Teaching presenceInstructional design, facilitating discourse, direct instruction (Anderson et al., 2001)Applied to instructor posts in an informal social media context; distinguishes reactive from proactive instructor behaviour
Student presenceLearner Presence emphasising self-regulation (Shea and Bidjerano, 2009)Operationalised as observable weekly role counts: Poster, Observer, Non-Member (Dokhanchi et al., 2018); used directly as prediction targets

Let t ∈ {1, …, T} index academic weeks, where T = 13. Each week contains a set of discourse artefacts (posts and comments) denoted Dt={dt,1,,dt,nt}, where nt is the count of artefacts in week t. A coding function maps each artefact to a K-dimensional binary indicator vector:

(1)

Where ϕk(dt,i) = 1 if indicator k is present in artefact dt,i and 0 otherwise. A single artefact can receive multiple indicators (multi-label coding). Weekly counts are obtained by summing across artefacts:

(2)

The weekly feature vector concatenates all indicator counts:

(3)

In this study, xt combines indicators from Social, Teaching, and Cognitive Presence, yielding K = 24 predictors.

Student Presence is represented by three weekly role counts:

(4)

Where yt(P), yt(O), and yt(N) denote Poster, Observer, and Non-Member counts respectively. The full dataset is then the set of paired weekly observations:

(5)

Because the three targets sum to total enrolment in each week (modulo rounding due to joining and leaving), they are structurally coupled. Specifically:

(6)

Where Nenrol denotes total enrolment. This linear dependence means that an accurate prediction of any two targets almost determines the third, but it also means that errors in one target propagate to the others. This structural constraint is an important consideration when interpreting multi-output model performance.

The modelling layer learns a function fθ(⋅) that maps presence indicators to weekly participation roles:

(7)

Where θ denotes model parameters. Fitting is expressed as empirical risk minimisation:

(8)

Three approaches were used.

3.5.1 Persistence baseline

The simplest reasonable baseline predicts next week's counts using the current week's values:

(9)

A model that cannot outperform this baseline provides no practical forecasting value.

3.5.2 Random Forest (multi-output regression)

Random Forest averages predictions from an ensemble of M decision trees:

(10)

Random Forest is well suited to small tabular datasets because it captures non-linear feature interactions without requiring large training samples, and it produces interpretable importance scores as a by-product of training.

3.5.3 Multilayer perceptron (MLP)

A deep multi-output regressor applies a composition of affine transformations and nonlinearities:

(11)

Where σ(⋅) is a nonlinear activation function (ReLU in this implementation) and θ={W,b}=13 are trainable parameters. The MLP was included to test whether a higher-capacity model improves performance relative to Random Forest when the training set is small.

To prevent temporal leakage, two evaluation strategies were applied.

3.6.1 Time-based holdout split

The model was trained on Weeks 1–10 and tested on Weeks 11–13. This mirrors the situation where an instructor uses early-semester signals to forecast late-semester participation.

3.6.2 Walk-forward (rolling-origin) backtest

Given a minimum training size t0, the model is retrained on weeks {1, …, t − 1} and tested on week t for each t = t0 + 1, …, T:

(12)

Walk-forward evaluation is more conservative than a single holdout split because each prediction is made strictly from past data, and the model must generalise across multiple time points rather than just three.

Figure 3 illustrates both evaluation protocols.

Figure 3
A diagram illustrating two evaluation protocols for model training and testing.Panel A shows a time-based holdout split. It consists of a series of 13 blocks labeled from 1 to 13. The first 10 blocks are colored blue and labeled as Train (W1-10), while the last 3 blocks are colored red and labeled as Test (W11-13). Panel B shows a walk-forward backtest. It consists of a grid of blocks arranged in a diagonal pattern. The blocks are labeled from 1 to 13. The blue blocks represent the training periods, and the red blocks represent the testing periods. The process starts with a training period of six weeks, followed by a testing period of one week. This process repeats, expanding the training period by one week each time and testing on the next single week.

Evaluation protocols. Panel (a) shows the time-based holdout split: the model trains on Weeks 1–10 and is tested on Weeks 11–13. Panel (b) shows the walk-forward backtest: the model is retrained at each step on an expanding window and tested on the next single week, starting from a minimum training window of six weeks

Figure 3
A diagram illustrating two evaluation protocols for model training and testing.Panel A shows a time-based holdout split. It consists of a series of 13 blocks labeled from 1 to 13. The first 10 blocks are colored blue and labeled as Train (W1-10), while the last 3 blocks are colored red and labeled as Test (W11-13). Panel B shows a walk-forward backtest. It consists of a grid of blocks arranged in a diagonal pattern. The blocks are labeled from 1 to 13. The blue blocks represent the training periods, and the red blocks represent the testing periods. The process starts with a training period of six weeks, followed by a testing period of one week. This process repeats, expanding the training period by one week each time and testing on the next single week.

Evaluation protocols. Panel (a) shows the time-based holdout split: the model trains on Weeks 1–10 and is tested on Weeks 11–13. Panel (b) shows the walk-forward backtest: the model is retrained at each step on an expanding window and tested on the next single week, starting from a minimum training window of six weeks

Close modal

For each target q ∈ {P, O, N}, mean absolute error (MAE) and root mean squared error (RMSE) measure average and worst-case prediction error:

(13)

The coefficient of determination R2 expresses explained variance relative to a mean baseline:

(14)

Where y¯(q) is the mean of the true values on the test interval. A negative R2 indicates that the model performs worse than simply predicting the mean, which is itself a practically important finding in small-sample settings.

To identify influential indicators, permutation importance is computed by measuring the degradation in predictive score after randomly permuting feature k:

(15)

Where s(⋅) is the model score (e.g. multi-output R2), and Xtestπ(k) denotes the test matrix with the k-th feature permuted.

Additionally, a one-dimensional what-if sensitivity analysis visualises the effect of varying feature k while holding all other features fixed at a reference vector xref (e.g. training mean):

(16)

Where ek is the k-th canonical basis vector. This yields an interpretable curve relating feature values to predicted participation roles.

3.9.1 Student presence

Student presence reflects the degree to which learners actively participate in the online community. Participation levels were observed and coded weekly across the semester, resulting in three categories: posters, observers, and non-members (Table 4). In addition to descriptive coding, these categories were operationalised as weekly target variables in the modelling layer, enabling the quantitative analysis of how Social/Teaching/Cognitive presence indicators relate to participation dynamics.

Table 4

Student presence indicators

Student typeDefinition
PosterA student who posts at least one message during the academic week
ObserverA student who has joined the community but does not post or comment. Students who like posts are included in this category
Non-memberA student who is enrolled in the course but has not yet joined the community

Posters are active participants who create posts in online learning communities. Observers or lurkers do not actively participate in discussions; however, they learn by observing the discussions. They are passive participants and may become active participants at any time during the semester. Non-members in online communities can join the community anytime and become observers or posters during the semester.

3.9.2 Cognitive presence

In this research, the cognitive presence section of the coding scheme was developed based on observation of students' course-related posts adopting Bloom's taxonomy framework (Anderson and Bloom, 2001). Indicators reflect six cognitive processes: remembering, understanding, applying, analysing, evaluating, and creating. Table 5 presents the categories, definitions, and representative examples drawn from course Facebook group discussions.

Table 5

Cognitive presence indicators

CategoryIndicatorStudent comment
RememberingRecalling prior information“Okay thank you, it's just that in the MEA guide it says to list them all so I just was a little confused.”
UnderstandingConfirming knowledge acquisition“Thanks. It really helps! I can generate a simulation now!”
ApplyingCarrying out a procedure“For calibration. Simply set the Arduino to output a PWM signal of duty cycle 50%.”
AnalysingBreaking knowledge into components to support understanding“I was taking into consideration the dry weather with decreasing rainfall throughout the year where it's very cost effective.”
EvaluatingMaking a judgment based on available information“Here is what we've got for the simulation result. The changes in colour seem alright in the image, but the value of the mass friction is quite close to 0.002 which is not that ideal.”
CreatingCombining information components to produce new knowledge“Considering the resolution of the camera is 1920x1080 pixels and knowing the distance it is looking at (say approximately 10 cm), you can calculate what the minimum strain you will be able to measure using the OSS.”

In the modelling layer, cognitive presence indicators form a subset of the predictor vector xt and quantify the cognitive depth of weekly discourse. This enables empirical testing of whether higher-order cognitive activity is associated with changes in participation roles, and supports interpretability analyses that identify which cognitive processes are most influential.

3.9.3 Social presence

Social presence captures how learners project themselves socially and emotionally within the online community. Students' course and non-course related posts in online discussions were observed. Table 6 presents the coding scheme that was developed in this research to analyse social presence in online learning communities. Students' posts and comments are categorised into group cohesiveness, conversational expression, and affective expression. Each category is divided further into different indicators.

Table 6

Social presence indicators

CategoryIndicatorDefinitionStudent comment
Group cohesivenessGroup ReferenceReferring to community members as “we”“Will we have a PIR for ENGG1200?”
TaggingAcknowledging other students in conversation“John those numbers have come directly from the project C PowerPoint slides this week”
Conversational expressionSharing InformationA post that informs other students“For PVC we got a 0.23 K loss for 3 L/s and 313K”
StatementAn expression of views or ideas“Everyone talks about pulling an all-nighter. No posts after 12 a.m.”
EnquiryA question directed at other students“Has anyone know whereabouts of a grey little pencil case that I left in the advanced engineering building today”
Affective expressionEmotional IconsUse of icons to convey feelings:)
Emotional WordsUse of words to convey feelings“You make me so happy. See you today”
ParalanguageText forms used to convey feeling“Whoa!!! Can’t wait”
Self-DisclosureExpressing a personal feeling“I am offended by this accusation”

In the modelling layer, social presence indicators represent interactional and affective signals and constitute a subset of xt. This supports quantitative evaluation of whether socio-emotional projection and group cohesion co-vary with participation roles (Poster/Observer/Non-Member), complementing the qualitative interpretation of discourse.

3.9.4 Teaching presence

Teaching presence was examined through course coordinators' and tutors' posts in online discussions. Indicators were organized into three categories: design and plan, facilitating discourse, and directing discourse. These categories capture instructors' roles in structuring learning activities, encouraging engagement, monitoring discussions, and providing clarification or guidance. Table 7 sets out the coding scheme developed to investigate teaching presence in online learning communities.

Table 7

Teaching presence indicators

CategoriesIndicatorDefinitionInstructor comment
Design and planSetting GoalsPost that indicates regulations for participating in online discussions“By signing up to this Facebook group you agree to use it responsibly.”
Defining Discussion TopicsPosts that direct discussions“I'm giving an overview of the PIR in Wednesday's lecture. What would you particularly like to know?”
Facilitating discourseIcebreakingPost that warms up the conversation between participants“Yup, a good video! Don't forget to bring your popcorn and 3D glasses for the ultimate movie experience!”
EncouragingPost that encourages students to participate in the discussion“The first speakers for 2015! Well done on addressing 600 people”
MonitoringComments that show the online discussions are being monitored“Woop! Rule breakers identified! Request instructions on appropriate punishment”
AcknowledgingTagging students in the discussions“Seriously Alex, we will make sure that he doesn't mark your logbook.”
Directing discourseClarifyingPost that clarifies students' enquiries or statements“Absolutely! Logbooks are a fantastic way to keep all your thoughts in a safe collected space!”
Giving DirectionsPost that directs students to other resources“Full scale - it's in Doc 2 and my LCA slides.”
Asking QuestionsPost that requires students' responses“Which one are you talking about?”
Providing InformationPost that provides information“Google EAIT cover sheets, and all will be revealed”

Teaching presence indicators are included in xt as contextual signals reflecting instructional design and facilitation. Their inclusion enables the modelling layer to assess whether instructor behaviours (e.g. clarifying, giving directions, and encouraging) relate to transitions between participation roles over time.

Based on the weekly interaction analysis (Table 8) and the predictive modelling outputs, this section synthesises (1) descriptive insights about Social, Teaching, Cognitive, and Student Presence dynamics over Weeks 1–13 and (2) quantitative evidence from exploratory analysis, forecasting experiments, and interpretability methods. Together, these results provide a multi-perspective understanding of online peer learning behaviour in the course Facebook group.

Table 8

Weekly interaction data: indicator counts per Week (W1–W13) for all four presence dimensions

DimIndicatorW1W2W3W4W5W6W7W8W9W10W11W12W13
SocialGroup Ref2318242312661016641618
Tagging33313618111391211111736
Sharing Info10121931873311002
Statement30391139191310677657
Enquiry3041776921159172220132460
Emot. Icons201526279768984617
Emot. Words45109512112135
Paralanguage3622276313853413
Self-Disclosure35810321131102
TeachingSetting Goals2000000000000
Defining Topics3000000000000
Icebreaking35324638719913746345
Encouraging1218211111793543519
Monitoring0132001010000
Acknowledging221731323101444428
Clarifying12924112614225110
Giving Dir46135223321212
Asking Ques107138163211216
Providing Info3455807613221618201815428
CognitiveRemembering412141143361134139
Understanding431614241251256
Applying0164101011120
Analysing615372964412187584
Evaluating031084225120435
Creating0154221220141
StudentPoster40448092374138353943333863
Observer360416440450514513520523519516528526507
Non-Member590530470448439436432432432431429426420

The merged dataset contains 13 weekly observations and 27 variables: 24 predictors (Social, Teaching, and Cognitive Presence indicators) and three targets (Poster, Observer, Non-Member). The small sample size has two practical consequences. First, any individual performance metric is sensitive to a single anomalous week, so MAE, RMSE, and R2 must be interpreted cautiously and in combination. Second, complex models with many parameters (such as the MLP) are highly prone to overfitting. Because the dataset is small and time-ordered, model estimates and generalisation performance are highly sensitive to overfitting and to the chosen evaluation protocol. Therefore, the analysis combines descriptive statistics (Table 8), time-series visualisation, time-aware splits, walk-forward validation, and interpretability tools.

4.2.1 Social presence

Social presence recorded robust interaction over the semester (Table 8), particularly through Enquiry, which peaks in Weeks 3 and 4. This pattern suggests alignment with demanding course content and/or assessment pressure points that triggered clarification-seeking and peer support. Sustained Statement and Tagging activity indicates an active conversational environment that supported informal collaboration. Emotional Icons and Emotional Words appear at moderate levels, while Self-Disclosure remains low, which is typical for formal academic contexts. Overall, the social presence evidence suggests that peer support for academic problem solving was prioritised over emotional bonding.

The dominance of Enquiry over Emotional Words and Self-Disclosure suggests that students used the Facebook group primarily for academic problem solving rather than social bonding. This is an important distinction for instructors: high Social Presence in an engineering course Facebook group does not necessarily imply community warmth; it may simply reflect high information-seeking activity around deadlines.

4.2.2 Teaching presence

Teaching Presence was highest in early weeks and concentrated in two indicators: Providing Information and Icebreaking. Setting Goals and Defining Discussion Topics occurred only in Week 1, and Monitoring was almost entirely absent. This pattern points to a reactive instructor role: instructors respond to student questions, but they do not proactively shape the discourse beyond establishing initial community norms.

A proactive Teaching Presence, in which instructors set discussion topics, prompt students with higher-order questions, and monitor discourse quality over time, may be a realistic lever for shifting Cognitive Presence from lower-order to higher-order processes. The data here do not allow a causal test of that hypothesis, but the pattern creates a testable implication for future studies.

4.2.3 Cognitive presence

Cognitive Presence was the least frequent dimension, and the discourse that was coded remained concentrated at the lower end of Bloom's taxonomy. Remembering and analysing together account for the majority of coded cognitive activity. Evaluating, Creating, and Applying each appeared infrequently, with weekly counts rarely exceeding five.

This finding is consistent with the informal nature of a social media discussion platform, where structured prompts that would support higher-order thinking (e.g. reflective tasks, case analyses) are absent. It also aligns with prior work showing that deeper cognitive engagement in online environments requires sustained instructional scaffolding (Fiock, 2020; Garrison and Cleveland-Innes, 2005).

4.2.4 Student presence

Observer counts dominated throughout the semester, ranging from 360 in Week 1 to a peak of 528 in Week 11. Poster counts were modest, peaking at 92 in Week 4 and then declining to a trough of 33 in Week 11. Non-Member counts declined monotonically from 590 in Week 1–420 in Week 13, indicating steady uptake of group membership. However, membership growth did not translate into posting: of the 170 students who joined between Weeks 1 and 13, virtually all became Observers rather than Posters.

The Week 4 Poster spike coincides with a peak in Enquiry (69), Analysing (29), Tagging (36), and Providing Information (76), suggesting that assessment pressure simultaneously raised the volume of discourse and the proportion of students who posted at least once. The Week 13 Poster spike (63) shows a similar but smaller pattern.

Figure 4 presents a schematic summary of presence dynamics across the semester, and Figure 5 shows the distribution of Cognitive Presence across Bloom's taxonomy levels.

Figure 4
A bar graph showing weekly totals of social, teaching, and cognitive presence indicator counts.The bar graph compares weekly totals of social, teaching, and cognitive presence indicator counts. It features three vertical, grouped bars for each week, representing Social Presence in blue, Teaching Presence in brown, and Cognitive Presence in green. The x-axis is labeled 'Week' and ranges from 1 to 13. The y-axis is labeled 'Total indicator count' and ranges from 0 to 400. Activity peaks in Week 3, dips in Weeks 5 to 10, and partially recovers in Week 13. Cognitive Presence consistently contributes the smallest share of total discourse. All values are approximated.

Weekly totals of social, teaching, and cognitive presence indicator counts. Activity peaks in Weeks 3–4, dips in Weeks 5–10, and partially recovers in Week 13. Cognitive Presence consistently contributes the smallest share of total discourse

Figure 4
A bar graph showing weekly totals of social, teaching, and cognitive presence indicator counts.The bar graph compares weekly totals of social, teaching, and cognitive presence indicator counts. It features three vertical, grouped bars for each week, representing Social Presence in blue, Teaching Presence in brown, and Cognitive Presence in green. The x-axis is labeled 'Week' and ranges from 1 to 13. The y-axis is labeled 'Total indicator count' and ranges from 0 to 400. Activity peaks in Week 3, dips in Weeks 5 to 10, and partially recovers in Week 13. Cognitive Presence consistently contributes the smallest share of total discourse. All values are approximated.

Weekly totals of social, teaching, and cognitive presence indicator counts. Activity peaks in Weeks 3–4, dips in Weeks 5–10, and partially recovers in Week 13. Cognitive Presence consistently contributes the smallest share of total discourse

Close modal
Figure 5
A bar graph showing cognitive presence by Bloom's taxonomy levels.The bar graph compares cognitive presence across different levels of Bloom's revised taxonomy over Weeks 1 to 13. It features six horizontal bars representing Remembering, Understanding, Analysing, Applying, Evaluating, and Creating. The x-axis indicates the total count, ranging from 0 to 200. The y-axis lists the taxonomy levels. Remembering has a count of 97, Understanding 65, Analysing 155, Applying 18, Evaluating 58, and Creating 25. The bars are colored green, with higher-order processes (Evaluating and Creating) appearing at lower rates. All values are approximated.

Distribution of cognitive presence coded indicators across Bloom's revised taxonomy levels (cumulative over Weeks 1–13). Analysing and Remembering dominate. Higher-order processes (Evaluating and Creating) appear at substantially lower rates, indicating that discourse rarely reached the level of knowledge synthesis or critical evaluation

Figure 5
A bar graph showing cognitive presence by Bloom's taxonomy levels.The bar graph compares cognitive presence across different levels of Bloom's revised taxonomy over Weeks 1 to 13. It features six horizontal bars representing Remembering, Understanding, Analysing, Applying, Evaluating, and Creating. The x-axis indicates the total count, ranging from 0 to 200. The y-axis lists the taxonomy levels. Remembering has a count of 97, Understanding 65, Analysing 155, Applying 18, Evaluating 58, and Creating 25. The bars are colored green, with higher-order processes (Evaluating and Creating) appearing at lower rates. All values are approximated.

Distribution of cognitive presence coded indicators across Bloom's revised taxonomy levels (cumulative over Weeks 1–13). Analysing and Remembering dominate. Higher-order processes (Evaluating and Creating) appear at substantially lower rates, indicating that discourse rarely reached the level of knowledge synthesis or critical evaluation

Close modal

4.3.1 Temporal Trends

Figures 6–7 visualise weekly dynamics of (1) the targets and (2) the presence indicators. Abrupt changes can reduce the effectiveness of simple baselines and hinder model generalisation, while smoother patterns can be exploited more reliably by machine learning models.

Figure 6
A line graph showing student presence targets over weeks for three categories: Poster, Observer, and Non-Member.A line graph titled Student Presence over Weeks (Targets) with the horizontal axis labeled Week and the vertical axis labeled Value. The graph includes three lines representing different categories: Poster in blue, Observer in orange, and Non-Member in green. The Poster line starts at a value of around 50 and fluctuates slightly, peaking at around 100 in week 4 before declining and stabilizing around 50. The Observer line starts at a value of around 350, increases sharply to around 500 by week 3, and then stabilizes around that value for the remaining weeks. The Non-Member line starts at a value of around 600, decreases steadily to around 450 by week 4, and then remains relatively stable around that value for the remaining weeks.

Student presence targets (poster, observer, non-member) over Weeks 1–13

Figure 6
A line graph showing student presence targets over weeks for three categories: Poster, Observer, and Non-Member.A line graph titled Student Presence over Weeks (Targets) with the horizontal axis labeled Week and the vertical axis labeled Value. The graph includes three lines representing different categories: Poster in blue, Observer in orange, and Non-Member in green. The Poster line starts at a value of around 50 and fluctuates slightly, peaking at around 100 in week 4 before declining and stabilizing around 50. The Observer line starts at a value of around 350, increases sharply to around 500 by week 3, and then stabilizes around that value for the remaining weeks. The Non-Member line starts at a value of around 600, decreases steadily to around 450 by week 4, and then remains relatively stable around that value for the remaining weeks.

Student presence targets (poster, observer, non-member) over Weeks 1–13

Close modal
Figure 7
A line graph showing presence-related input features over weeks.A line graph titled 'All Presence-related Features over Weeks (Inputs)' displays the values of various presence-related features over a period of 13 weeks. The x-axis represents the weeks, ranging from 0 to 13, and the y-axis represents the value, ranging from 0 to 80. The graph includes multiple lines, each representing a different feature such as Group Reference, Tagging, Sharing Information, Statement, Enquiry, Emotional Icons, Emotional Words, Paralanguage, Self-Disclosure, Setting Goals, Defining Discussion Topics, Icebreaking, Encouraging, Monitoring, Acknowledging, Clarifying, Giving Directions, Asking Questions, Remembering, Understanding, Applying, Analysing, Evaluating, and Creating. Each line shows the trend of the respective feature's value over the weeks. Notable trends include a peak in the values of several features around week 4, followed by a decline. All values are approximated.

Presence-related input features over Weeks 1–13

Figure 7
A line graph showing presence-related input features over weeks.A line graph titled 'All Presence-related Features over Weeks (Inputs)' displays the values of various presence-related features over a period of 13 weeks. The x-axis represents the weeks, ranging from 0 to 13, and the y-axis represents the value, ranging from 0 to 80. The graph includes multiple lines, each representing a different feature such as Group Reference, Tagging, Sharing Information, Statement, Enquiry, Emotional Icons, Emotional Words, Paralanguage, Self-Disclosure, Setting Goals, Defining Discussion Topics, Icebreaking, Encouraging, Monitoring, Acknowledging, Clarifying, Giving Directions, Asking Questions, Remembering, Understanding, Applying, Analysing, Evaluating, and Creating. Each line shows the trend of the respective feature's value over the weeks. Notable trends include a peak in the values of several features around week 4, followed by a decline. All values are approximated.

Presence-related input features over Weeks 1–13

Close modal

4.3.2 Target distributions

Figure 8 presents the empirical distribution of the targets. Differences in spread and skewness can explain discrepancies in predictive difficulty; highly variable targets are harder to forecast in small-sample settings.

Figure 8
Three histograms showing the distribution of student presence targets for different roles.Panel A: A histogram showing the distribution of student presence targets for the Poster role. The horizontal axis represents the presence targets ranging from 40 to 90, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 40 to 50. Panel B: A histogram showing the distribution of student presence targets for the Observer role. The horizontal axis represents the presence targets ranging from 400 to 500, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 500. Panel C: A histogram showing the distribution of student presence targets for the Non-Member role. The horizontal axis represents the presence targets ranging from 450 to 550, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 450.

Histograms of the student presence targets

Figure 8
Three histograms showing the distribution of student presence targets for different roles.Panel A: A histogram showing the distribution of student presence targets for the Poster role. The horizontal axis represents the presence targets ranging from 40 to 90, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 40 to 50. Panel B: A histogram showing the distribution of student presence targets for the Observer role. The horizontal axis represents the presence targets ranging from 400 to 500, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 500. Panel C: A histogram showing the distribution of student presence targets for the Non-Member role. The horizontal axis represents the presence targets ranging from 450 to 550, and the vertical axis represents the frequency of students. There are five vertical bars, with the highest frequency around 450.

Histograms of the student presence targets

Close modal

4.3.3 Correlation structure

Figure 9 shows the correlation matrix over all variables, highlighting associations between presence indicators and student roles, predictor multicollinearity, and dependencies among targets.

Figure 9
A heat map showing the correlation between various predictors and targets.A heat map titled Correlation Heatmap (All Variables) displays the correlation between different predictors and targets. The heat map uses a grid layout with both axes labeled with the same set of variables, including Group Reference, Tagging, Sharing Information, Statement, Enquiry, Emotional Icons, Emotional Words, Paralanguage, Self-Disclosure, Setting Goals, Defining Discussion Topics, Icebreaking, Encouraging, Monitoring, Acknowledging, Clarifying, Giving Directions, Asking Questions, Remembering, Understanding, Applying, Analysing, Evaluating, Creating, Poster, Observer, and Non-Member. The color scale ranges from dark blue to bright yellow, indicating the magnitude of correlation values from -0.75 to 1.00. Darker colors represent lower correlation values, while brighter colors indicate higher correlation values. Notable patterns include clusters of high correlation values along the diagonal axis, suggesting strong correlations between the same variables.

Correlation heatmap of all predictors and targets

Figure 9
A heat map showing the correlation between various predictors and targets.A heat map titled Correlation Heatmap (All Variables) displays the correlation between different predictors and targets. The heat map uses a grid layout with both axes labeled with the same set of variables, including Group Reference, Tagging, Sharing Information, Statement, Enquiry, Emotional Icons, Emotional Words, Paralanguage, Self-Disclosure, Setting Goals, Defining Discussion Topics, Icebreaking, Encouraging, Monitoring, Acknowledging, Clarifying, Giving Directions, Asking Questions, Remembering, Understanding, Applying, Analysing, Evaluating, Creating, Poster, Observer, and Non-Member. The color scale ranges from dark blue to bright yellow, indicating the magnitude of correlation values from -0.75 to 1.00. Darker colors represent lower correlation values, while brighter colors indicate higher correlation values. Notable patterns include clusters of high correlation values along the diagonal axis, suggesting strong correlations between the same variables.

Correlation heatmap of all predictors and targets

Close modal
  1. Top associations with targets (association only; not causal):

    • Poster: Paralanguage (0.973), Enquiry (0.926), Understanding (0.915), Emotional Icons (0.873), Emotional Words (0.872), Self-Disclosure (0.862), Tagging (0.829), Sharing Information (0.826), Acknowledging (0.811), Analysing (0.765).

    • Observer: Non-Member (0.937), Asking Questions (0.798), Acknowledging (0.782), Statement (0.767), Emotional Icons (0.738), Group Reference (0.727), Setting Goals (0.717), Defining Discussion Topics (0.717), Icebreaking (0.691), Clarifying (0.665).

    • Non-Member: Observer (0.937), Setting Goals (0.816), Defining Discussion Topics (0.816), Statement (0.678), Asking Questions (0.605), Acknowledging (0.534), Group Reference (0.532), Emotional Icons (0.464), Icebreaking (0.458), Clarifying (0.444).

  2. Interpretation: Poster correlates strongly with socio-emotional signals (paralanguage, emotional cues, self-disclosure), suggesting that active posters exhibit richer social presence. Observer and Non-Member are strongly coupled (r ≈ 0.94), indicating structural interdependence and complicating independent prediction.

4.4.1 Evaluation protocol

A time-based split was adopted to reflect realistic forecasting conditions and prevent information leakage: models were trained on Weeks 1–10 and evaluated on unseen future weeks (Weeks 11–13). This setup mirrors how an instructor would predict upcoming participation from earlier signals. To strengthen reliability, a walk-forward (rolling-origin) backtest was also performed, repeatedly retraining on expanding histories and testing week-by-week.

4.4.2 Persistence baseline

The persistence baseline produced a negative R2 for all three targets on the holdout set (Table 9). Poster showed the largest error (MAE = 13.33), reflecting the sharp jump in Week 13 (from 38 to 63) that the baseline could not anticipate. The negative R2 values confirm that end-of-semester behaviour is not a simple continuation of mid-semester behaviour.

Table 9

Holdout performance (Weeks 11–13): persistence baseline

TargetMAERMSER2
Poster13.3315.81−0.45
Observer11.0013.03−0.89
Non-member3.674.04−0.17
OVERALL9.3312.06−0.58

4.4.3 Random Forest.

  1. Holdout evaluation (Weeks 11–13): To assess out-of-sample generalisation in a realistic forecasting setting, the RandomForest model was trained on the first ten weeks (Weeks 1–10) and evaluated on the final three weeks (Weeks 11–13), ensuring that no future information leaked into the training stage. Table 10 reports the resulting error metrics (MAE and RMSE) together with R2 for each student-presence target and for the overall multi-output prediction. In addition to the tabulated metrics, Figures 10–11 provide complementary diagnostic views. Figure 10 plots the actual and predicted trajectories across the holdout weeks, allowing visual inspection of whether the model captures the direction and magnitude of changes near the end of the semester. Figure 11 displays residual errors in time order, which helps identify systematic bias (consistent over- or under-estimation) and potential regime changes that may occur in the final weeks (e.g. assessment-related shifts). Together, these results quantify predictive accuracy while also revealing whether errors are stable or concentrated in specific weeks or targets, thereby supporting a more reliable interpretation of model performance under time-ordered educational data.

  2. Walk-forward backtest: To obtain a more robust estimate of time-series generalisation, a walk-forward (rolling-origin) backtest was performed. In this setting, the model is repeatedly retrained on an expanding window of past weeks and evaluated on the next unseen week, better reflecting deployment conditions. Table 11 summarises aggregate performance across all tested weeks, while Figure 12 visualises predicted versus observed values over time (see Table 12).

  3. Interpretation: RandomForest improves Poster prediction consistently, achieving positive R2 values that indicate meaningful predictive signal in the presence of indicators for active participation. In contrast, Observer and Non-Member remain difficult to model, likely due to their strong interdependence (changes in one often imply changes in the other), temporal non-stationarity across weeks, and the limited sample size, which amplifies variance and sensitivity to outliers.

Table 10

Holdout performance (Weeks 11–13): RandomForest (multi-output regression)

TargetMAERMSER2
Poster8.119.360.49
Observer19.3920.94−3.90
Non-Member21.0027.10−51.46
OVERALL16.1720.50−3.57
Figure 10
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13.The image contains three line graphs comparing actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13. The first graph shows the Poster values, with the actual values represented by a blue line and the predicted values by an orange line. The second graph displays the Observer values, again with actual values in blue and predicted values in orange. The third graph illustrates the Non-Member values, following the same color scheme. Each graph spans from week 11 to week 13, with the x-axis representing the weeks and the y-axis representing the respective values. The trends show varying degrees of alignment between actual and predicted values across the different categories. All values are approximated.

Holdout (Weeks 11–13): actual vs. predicted for RandomForest

Figure 10
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13.The image contains three line graphs comparing actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13. The first graph shows the Poster values, with the actual values represented by a blue line and the predicted values by an orange line. The second graph displays the Observer values, again with actual values in blue and predicted values in orange. The third graph illustrates the Non-Member values, following the same color scheme. Each graph spans from week 11 to week 13, with the x-axis representing the weeks and the y-axis representing the respective values. The trends show varying degrees of alignment between actual and predicted values across the different categories. All values are approximated.

Holdout (Weeks 11–13): actual vs. predicted for RandomForest

Close modal
Figure 11
Three scatter plots showing residuals for different groups in a RandomForest model.Three scatter plots depict holdout residuals for a RandomForest model. Each plot represents residuals for a different group: Poster, Observer, and Non-Member. The x-axis for all plots is labeled Test sample index (time order) and ranges from 0.00 to 2.00. The y-axis for each plot is labeled with the respective group's residuals: Res(Poster), Res(Observer), and Res(Non-Member). The top plot shows residuals for the Poster group, with data points at approximately (0.00, -5), (1.00, 0), and (2.00, 15). The middle plot shows residuals for the Observer group, with data points at approximately (0.00, 10), (1.00, 15), and (2.00, 30). The bottom plot shows residuals for the Non-Member group, with data points at approximately (0.00, -10), (1.00, -10), and (2.00, -40).

Holdout residuals (RandomForest)

Figure 11
Three scatter plots showing residuals for different groups in a RandomForest model.Three scatter plots depict holdout residuals for a RandomForest model. Each plot represents residuals for a different group: Poster, Observer, and Non-Member. The x-axis for all plots is labeled Test sample index (time order) and ranges from 0.00 to 2.00. The y-axis for each plot is labeled with the respective group's residuals: Res(Poster), Res(Observer), and Res(Non-Member). The top plot shows residuals for the Poster group, with data points at approximately (0.00, -5), (1.00, 0), and (2.00, 15). The middle plot shows residuals for the Observer group, with data points at approximately (0.00, 10), (1.00, 15), and (2.00, 30). The bottom plot shows residuals for the Non-Member group, with data points at approximately (0.00, -10), (1.00, -10), and (2.00, -40).

Holdout residuals (RandomForest)

Close modal
Table 11

Walk-forward backtest: RandomForest performance

TargetMAERMSER2
Poster5.166.730.48
Observer20.0123.13−11.70
Non-Member20.5525.61−36.98
OVERALL15.2420.30−7.44
Figure 12
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over 13 weeks.The image contains three line graphs comparing actual and predicted values for Poster, Observer, and Non-Member over 13 weeks. The first graph shows the Poster values, with the actual values represented by a blue line and the predicted values by an orange line. The second graph displays the Observer values, with the actual values in blue and the predicted values in orange. The third graph illustrates the Non-Member values, again with actual values in blue and predicted values in orange. Each graph has the weeks on the x-axis and the respective values on the y-axis. The trends, peaks, and variations between actual and predicted values are clearly visible in each graph. All values are approximated.

Walk-forward: actual vs. predicted for RandomForest

Figure 12
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over 13 weeks.The image contains three line graphs comparing actual and predicted values for Poster, Observer, and Non-Member over 13 weeks. The first graph shows the Poster values, with the actual values represented by a blue line and the predicted values by an orange line. The second graph displays the Observer values, with the actual values in blue and the predicted values in orange. The third graph illustrates the Non-Member values, again with actual values in blue and predicted values in orange. Each graph has the weeks on the x-axis and the respective values on the y-axis. The trends, peaks, and variations between actual and predicted values are clearly visible in each graph. All values are approximated.

Walk-forward: actual vs. predicted for RandomForest

Close modal

4.4.4 MLP

To examine whether a nonlinear deep model can capture more complex relationships among presence indicators, a multi-output MLP was trained on Weeks 1–10 and evaluated on the holdout period (Weeks 11–13). Table 12 reports performance metrics, while Figures 13–15 illustrate optimisation behaviour (training/validation loss) and residual errors, highlighting generalisation limitations under scarce data. The MLP performed poorly across all targets and evaluation settings (Table 12). The Observer R2 of −751.39 and Non-Member R2 of −2728.34 indicate catastrophic overfitting: the model fitted Week 1–10 patterns so tightly that it predicted values far outside the plausible range on the holdout. With only 10 training observations, a three-layer MLP with hundreds of parameters has far more capacity than the data can support. This result confirms that deep models are not suitable for datasets of this size, regardless of their theoretical expressiveness.

Table 12

Holdout performance (Weeks 11–13): MLP (multi-output regression)

TargetMAERMSER2
Poster25.4530.33−4.34
Observer206.70259.58−751.39
Non-member178.70195.48−2728.34
OVERALL136.95188.42−385.22
Figure 13
A line graph titled MLP Training History MSE with two lines representing train loss and validation loss.A line graph titled MLP Training History MSE displays the loss values over epochs. The x axis represents the number of epochs ranging from 0 to 250. The y axis represents the loss values ranging from 0 to 160000. The blue line represents the training loss, which starts at around 150000 and gradually decreases to approximately 10000. The orange line represents the validation loss, which also starts at around 150000 and decreases more smoothly to approximately 20000. Both lines show a downward trend, indicating a reduction in loss over time. All values are approximated.

MLP training history (loss curves)

Figure 13
A line graph titled MLP Training History MSE with two lines representing train loss and validation loss.A line graph titled MLP Training History MSE displays the loss values over epochs. The x axis represents the number of epochs ranging from 0 to 250. The y axis represents the loss values ranging from 0 to 160000. The blue line represents the training loss, which starts at around 150000 and gradually decreases to approximately 10000. The orange line represents the validation loss, which also starts at around 150000 and decreases more smoothly to approximately 20000. Both lines show a downward trend, indicating a reduction in loss over time. All values are approximated.

MLP training history (loss curves)

Close modal
Figure 14
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13.Three line graphs depict the comparison of actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13. Each graph has a horizontal axis labeled Week ranging from 11.00 to 13.00 and a vertical axis with different units. The first graph shows Poster values on the vertical axis ranging from 0 to 60. The blue line represents Actual Poster values, while the orange line with crosses represents Predicted Poster values. The second graph shows Observer values on the vertical axis ranging from 0 to 500. The blue line represents Actual Observer values, while the orange line with crosses represents Predicted Observer values. The third graph shows Non-Member values on the vertical axis ranging from 0 to 500. The blue line represents Actual Non-Member values, while the orange line with crosses represents Predicted Non-Member values. In all graphs, the actual values remain relatively stable, while the predicted values show a decreasing trend over the weeks.

Holdout (Weeks 11–13): actual vs. predicted for MLP

Figure 14
Three line graphs compare actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13.Three line graphs depict the comparison of actual and predicted values for Poster, Observer, and Non-Member over weeks 11 to 13. Each graph has a horizontal axis labeled Week ranging from 11.00 to 13.00 and a vertical axis with different units. The first graph shows Poster values on the vertical axis ranging from 0 to 60. The blue line represents Actual Poster values, while the orange line with crosses represents Predicted Poster values. The second graph shows Observer values on the vertical axis ranging from 0 to 500. The blue line represents Actual Observer values, while the orange line with crosses represents Predicted Observer values. The third graph shows Non-Member values on the vertical axis ranging from 0 to 500. The blue line represents Actual Non-Member values, while the orange line with crosses represents Predicted Non-Member values. In all graphs, the actual values remain relatively stable, while the predicted values show a decreasing trend over the weeks.

Holdout (Weeks 11–13): actual vs. predicted for MLP

Close modal
Figure 15
Three scatter plots showing residuals for different models.Three scatter plots display residuals for different models. The x-axis represents the test sample index in time order, ranging from 0 to 2. The y-axis represents the residuals for each model. The top plot shows residuals for the Posterior model, with values ranging from approximately -20 to 40. The middle plot shows residuals for the Observer model, with values ranging from approximately -100 to 400. The bottom plot shows residuals for the Non-Member model, with values ranging from approximately -100 to 300. Each plot has data points at specific intervals, indicating the residuals at those points. The residuals vary significantly across the different models.

Holdout residuals (MLP)

Figure 15
Three scatter plots showing residuals for different models.Three scatter plots display residuals for different models. The x-axis represents the test sample index in time order, ranging from 0 to 2. The y-axis represents the residuals for each model. The top plot shows residuals for the Posterior model, with values ranging from approximately -20 to 40. The middle plot shows residuals for the Observer model, with values ranging from approximately -100 to 400. The bottom plot shows residuals for the Non-Member model, with values ranging from approximately -100 to 300. Each plot has data points at specific intervals, indicating the residuals at those points. The residuals vary significantly across the different models.

Holdout residuals (MLP)

Close modal

Figure 16 provides a direct visual comparison of R2 values for Poster across the three models and both evaluation protocols.

Figure 16
A bar graph comparing poster R squared values for three models under holdout and walk-forward evaluation.A bar graph compares poster R squared values for three models under holdout and walk-forward evaluation. The horizontal axis lists three models: Persistence, RandomForest, and MLP. The vertical axis represents R squared values ranging from -6 to 0.49. The graph includes two sets of bars for each model, one in blue representing holdout evaluation and one in red representing walk-forward evaluation. For Persistence, the blue bar is around -0.5 and the red bar is around -0.4. For RandomForest, the blue bar is around 0.45 and the red bar is around 0.4. For MLP, the blue bar is around -5.5 and the red bar is around -5.5. A dashed line at R squared equals 0 indicates the baseline. Positive R squared values above the dashed line show that the model outperforms the mean baseline. Only RandomForest achieves this consistently. MLP and Persistence fail on both protocols.

Comparison of poster R2 values for the three models under holdout and walk-forward evaluation. positive R2 values (above the dashed line) indicate that the model outperforms the mean baseline. Only Random Forest achieves this consistently. MLP and persistence fail on both protocols

Figure 16
A bar graph comparing poster R squared values for three models under holdout and walk-forward evaluation.A bar graph compares poster R squared values for three models under holdout and walk-forward evaluation. The horizontal axis lists three models: Persistence, RandomForest, and MLP. The vertical axis represents R squared values ranging from -6 to 0.49. The graph includes two sets of bars for each model, one in blue representing holdout evaluation and one in red representing walk-forward evaluation. For Persistence, the blue bar is around -0.5 and the red bar is around -0.4. For RandomForest, the blue bar is around 0.45 and the red bar is around 0.4. For MLP, the blue bar is around -5.5 and the red bar is around -5.5. A dashed line at R squared equals 0 indicates the baseline. Positive R squared values above the dashed line show that the model outperforms the mean baseline. Only RandomForest achieves this consistently. MLP and Persistence fail on both protocols.

Comparison of poster R2 values for the three models under holdout and walk-forward evaluation. positive R2 values (above the dashed line) indicate that the model outperforms the mean baseline. Only Random Forest achieves this consistently. MLP and persistence fail on both protocols

Close modal

4.5.1 Feature importance

To better understand which indicators contribute most to prediction, feature relevance was examined using two complementary approaches. Impurity-based importance (Figure 17) provides a global view derived from the average reduction in split impurity across trees, but it can be biased under multicollinearity. Permutation importance (Figure 18) evaluates test-time influence by measuring performance degradation when a feature is randomly shuffled.

Figure 17
A bar graph showing RandomForest feature importance.The bar graph displays the importance of various features in a RandomForest model. The x-axis lists features such as Tagging, Encouraging, Statement, Asking Questions, Sharing Information, Giving Directions, Setting Goals, Clarifying, Emotional Icons, Defining Discussion Topics, Group Reference, Icebreaking, Inquiry, Creating, Acknowledging, Evaluating, Self-Disclosure, Emotional Words, Paralanguage, and Remembering. The y-axis represents the importance values ranging from 0.00 to 0.25. Tagging has the highest importance, followed by Encouraging and Statement. The bars are vertical and colored blue. All values are approximated.

RandomForest feature importance (impurity-based; global)

Figure 17
A bar graph showing RandomForest feature importance.The bar graph displays the importance of various features in a RandomForest model. The x-axis lists features such as Tagging, Encouraging, Statement, Asking Questions, Sharing Information, Giving Directions, Setting Goals, Clarifying, Emotional Icons, Defining Discussion Topics, Group Reference, Icebreaking, Inquiry, Creating, Acknowledging, Evaluating, Self-Disclosure, Emotional Words, Paralanguage, and Remembering. The y-axis represents the importance values ranging from 0.00 to 0.25. Tagging has the highest importance, followed by Encouraging and Statement. The bars are vertical and colored blue. All values are approximated.

RandomForest feature importance (impurity-based; global)

Close modal
Figure 18
A bar graph showing permutation importance on holdout set.A bar graph compares the importance of different categories on the holdout set. The horizontal axis lists categories such as Remembering, Evaluating, Acknowledging, Setting Goals, Monitoring, Sharing Information, Defining Discussion Topics, Giving Directions, Analysing, Statement, Self-Disclosure, Paralanguage, Understanding, Tagging, Emotional Words, Creating, Applying, Asking Questions, Group Reference, and Enquiry. The vertical axis measures importance with values ranging from -0.8 to 0.4. The bars are vertical and represent the mean decrease in score for each category. Remembering has the highest importance at approximately 0.35, followed by Evaluating at around 0.2. Acknowledging is near zero. The rest of the categories have negative importance values, with Enquiry and Group Reference being the lowest at approximately -0.7. The color scheme is a single blue color for all bars.

RandomForest permutation importance on the holdout set

Figure 18
A bar graph showing permutation importance on holdout set.A bar graph compares the importance of different categories on the holdout set. The horizontal axis lists categories such as Remembering, Evaluating, Acknowledging, Setting Goals, Monitoring, Sharing Information, Defining Discussion Topics, Giving Directions, Analysing, Statement, Self-Disclosure, Paralanguage, Understanding, Tagging, Emotional Words, Creating, Applying, Asking Questions, Group Reference, and Enquiry. The vertical axis measures importance with values ranging from -0.8 to 0.4. The bars are vertical and represent the mean decrease in score for each category. Remembering has the highest importance at approximately 0.35, followed by Evaluating at around 0.2. Acknowledging is near zero. The rest of the categories have negative importance values, with Enquiry and Group Reference being the lowest at approximately -0.7. The color scheme is a single blue color for all bars.

RandomForest permutation importance on the holdout set

Close modal

Acknowledging a Teaching Presence indicator appears next. Instructor posts that tag individual students may stimulate those students and their peers to become more active, although the correlational nature of this finding prevents a causal interpretation.

4.5.2 Sensitivity analysis

The sensitivity curves for Remembering show a non-linear relationship with Poster predictions: as Remembering counts increase from the 10th to the 90th percentile, predicted Poster counts rise and then plateau (Figures 19–20). A similar pattern appears for Evaluating. These curves suggest that moderate levels of cognitive discourse are associated with increases in posting, but gains taper off at high counts. This may reflect a ceiling effect: a class discussion can only sustain a limited number of active contributors regardless of how rich the cognitive content becomes.

Figure 19
A line graph titled What if Sensitivity R F: vary Remembering.The line graph titled What if Sensitivity R F: vary Remembering features three data lines representing Poster, Observer, and Non Member. The x axis is labeled Remembering and ranges from 2 to 12. The y axis is labeled Predicted Student Presence and ranges from 0 to 500. The Observer line remains constant at 500, the Non Member line remains constant at 450, and the Poster line remains constant at 50. All values are approximated.

Sensitivity analysis: varying the top feature (remembering) while holding others fixed

Figure 19
A line graph titled What if Sensitivity R F: vary Remembering.The line graph titled What if Sensitivity R F: vary Remembering features three data lines representing Poster, Observer, and Non Member. The x axis is labeled Remembering and ranges from 2 to 12. The y axis is labeled Predicted Student Presence and ranges from 0 to 500. The Observer line remains constant at 500, the Non Member line remains constant at 450, and the Poster line remains constant at 50. All values are approximated.

Sensitivity analysis: varying the top feature (remembering) while holding others fixed

Close modal
Figure 20
A line graph showing the predicted student presence based on different evaluating roles.A line graph titled What-if Sensitivity (RF): vary Evaluating. The horizontal axis is labeled Evaluating and ranges from 0 to 10. The vertical axis is labeled Predicted Student Presence and ranges from 0 to 500. The graph includes three lines representing different roles: Poster, Observer, and Non-Member. The Poster line is blue and remains constant at around 50 predicted student presence. The Observer line is orange and remains constant at around 500 predicted student presence. The Non-Member line is green and remains constant at around 450 predicted student presence.

Sensitivity analysis: varying the second top feature (evaluating) while holding others fixed

Figure 20
A line graph showing the predicted student presence based on different evaluating roles.A line graph titled What-if Sensitivity (RF): vary Evaluating. The horizontal axis is labeled Evaluating and ranges from 0 to 10. The vertical axis is labeled Predicted Student Presence and ranges from 0 to 500. The graph includes three lines representing different roles: Poster, Observer, and Non-Member. The Poster line is blue and remains constant at around 50 predicted student presence. The Observer line is orange and remains constant at around 500 predicted student presence. The Non-Member line is green and remains constant at around 450 predicted student presence.

Sensitivity analysis: varying the second top feature (evaluating) while holding others fixed

Close modal

Five findings emerge from the combined descriptive and predictive analysis.

  1. Social Presence is enquiry-driven and assessment-sensitive, peaking when deadlines approach and declining in low-pressure weeks.

  2. Teaching Presence is reactive and frontloaded: instructors provide information and break the ice in early weeks, but do not proactively structure discourse across the semester.

  3. Cognitive Presence is shallow, concentrated in remembering and analysing, with limited evidence of evaluating or creating.

  4. Student participation is consumption-heavy: the Observer role dominates throughout, and posting spikes selectively near assessment events rather than growing steadily.

  5. Random Forest predicts Poster behaviour with positive R2 on both evaluation protocols, but Observer and Non-Member remain difficult due to structural coupling and slow within-semester change.

These findings converge on a practical implication: the presence indicators contain enough signal to predict active posting one to three weeks ahead with moderate accuracy, but they do not reliably predict passive participation levels. For instructors, this means the framework is most useful as an early-warning system for low Poster activity, not as a complete monitoring tool for all forms of participation.

The study has several limitations that bound the generalisability and precision of its findings. The dataset contains only 13 weekly observations. At this scale, each R2 value is estimated from three test points (holdout) or seven to eight points (walk-forward), making the metrics sensitive to individual anomalous weeks. Standard errors and confidence intervals on R2 would require bootstrapping procedures that are not feasible with 13 total observations.

The data come from a single platform (Facebook), a single institution (The University of Queensland), and a single discipline (engineering). Facebook groups have specific structural and normative characteristics (e.g. visible likes, informal register, public posts by default) that may not generalise to other platforms such as discussion boards, Slack workspaces, or Microsoft Teams. Coding was performed by a single team, and inter-rater reliability statistics would strengthen confidence in the indicator definitions.

The weekly aggregation level obscures within-week dynamics. A student who posts on Monday of a high-Enquiry week and a student who posts on Friday contribute equally to the weekly Poster count, even though the causal mechanism linking the discourse to their posting decisions may differ.

Future work should address these limitations through several avenues. Collecting data from multiple cohorts and courses would provide the sample sizes needed for more stable estimates and cross-context comparisons. Introducing lagged variables (e.g. last week's Poster count as a predictor) would allow the model to exploit temporal autocorrelation explicitly. Modelling at the individual student level, where the target is a binary “did this student post this week” indicator, would enable a richer analysis and support targeted intervention design. Sequence models (e.g. LSTM or temporal convolutional networks) could be explored once larger longitudinal datasets are available.

This study proposed and tested an extended Community of Inquiry framework for monitoring and predicting online peer learning participation in a social media context. The framework integrates four dimensions: Social Presence, Teaching Presence, Cognitive Presence, and Student Presence, and couples a theory-grounded coding scheme with a data-driven analytics pipeline.

The coding scheme operationalises 24 indicators across three presence dimensions and defines three Student Presence roles (Poster, Observer, Non-Member) as observable weekly targets. Applied to two first-year engineering course Facebook groups over 13 weeks, the scheme produced a structured dataset that is small but theoretically rich. Descriptive analysis showed that Social Presence was enquiry-driven, Teaching Presence was reactive, Cognitive Presence was shallow, and student participation was consumption-oriented.

On the predictive side, Random Forest achieved a consistent Poster R2 of approximately 0.48–0.49 on both holdout and walk-forward evaluation, outperforming the persistence baseline and confirming that the coded indicators carry predictive signal for active participation. Observer and Non-Member targets remained difficult to forecast, primarily because their structural interdependence and slow temporal dynamics leave little residual signal for a model to exploit. The MLP failed entirely under data scarcity, reinforcing the importance of selecting model capacity proportional to dataset size.

Permutation importance identified Remembering and Evaluating as the most influential predictors of posting behaviour. Although this finding is exploratory given the small holdout, it offers a plausible pedagogical interpretation: weeks in which students engage cognitively with course content, even at lower cognitive levels, tend to be weeks of higher active participation.

The framework demonstrates that theory-grounded indicator coding can be combined with transparent machine learning to produce actionable and interpretable insights from social media learning data. With appropriate extensions (more cohorts, individual-level modelling, lagged features), the approach has potential to inform real-time monitoring systems that give instructors timely, evidence-based signals about participation dynamics.

Anderson
,
L.W.
and
Bloom
,
B.S.
(
2001
),
A Taxonomy for Learning, Teaching, and Assessing: A Revision of Bloom's Taxonomy of Educational Objectives
,
Longman
.
Anderson
,
T.
,
Rourke
,
L.
,
Garrison
,
D.R.
and
Archer
,
W.
(
2001
), “
Assessing teaching presence in a computer conferencing context
”,
Journal of the Asynchronous Learning Network
, Vol. 
5
No. 
2
, pp. 
1
-
17
, doi: .
Babik
,
D.
,
Gehringer
,
E.
,
Kidd
,
J.
,
Sunday
,
K.
,
Tinapple
,
D.
and
Gilbert
,
S.
(
2024
), “
A systematic review of educational online peer-review and assessment systems: charting the landscape
”,
Educational Technology Research and Development
, Vol. 
72
No. 
3
, pp. 
1653
-
1689
, doi: .
Berge
,
Z.L.
(
1995
), “
Facilitating computer conferencing: recommendations from the field
”,
Educational Technology
, Vol. 
35
No. 
1
, pp. 
22
-
30
.
Borup
,
J.
,
West
,
R.E.
and
Graham
,
C.R.
(
2012
), “
Improving online social presence through asynchronous video
”,
The Internet and Higher Education
, Vol. 
15
No. 
3
, pp. 
195
-
203
, doi: .
Bravo
,
C.
,
Redondo
,
M.A.
,
Verdejo
,
M.F.
and
Ortega
,
M.
(
2008
), “
A framework for process–solution analysis in collaborative learning environments
”,
International Journal of Human-Computer Studies
, Vol. 
66
No. 
11
, pp. 
812
-
832
, doi: .
Bruggeman
,
B.
,
Tondeur
,
J.
,
Struyven
,
K.
,
Pynoo
,
B.
,
Garone
,
A.
and
Vanslambrouck
,
S.
(
2021
), “
Experts speaking: crucial teacher attributes for implementing blended learning in higher education
”,
The Internet and Higher Education
, Vol. 
48
, 100772, doi: .
Cercone
,
K.
(
2008
), “
Characteristics of adult learners with implications for online learning design
”,
AACE Journal
, Vol. 
16
No. 
2
, pp. 
137
-
159
.
Chen
,
F.
,
Li
,
S.
,
Lin
,
L.
and
Huang
,
X.
(
2024
), “
Identifying temporal changes in student engagement in social annotation during online collaborative reading
”,
Education and Information Technologies
, Vol. 
29
No. 
13
, pp. 
16101
-
16124
, doi: .
Creswell
,
J.W.
and
Creswell
,
J.D.
(
2017
),
Research Design: Qualitative, Quantitative, and Mixed Methods Approaches
,
Sage Publications
.
Delahunty
,
J.
,
Verenikina
,
I.
and
Jones
,
P.
(
2014
), “
Socio-emotional connections: identity, belonging and learning in online interactions. A literature review
”,
Technology, Pedagogy and Education
, Vol. 
23
No. 
2
, pp. 
243
-
265
, doi: .
Dokhanchi
,
M.
,
Kavanagh
,
L.
and
Reidsema
,
C.
(
2018
), “
Factors that influence peer learning in social media enhanced engineering courses
”,
Proceedings of the 29th Annual Conference of the Australasian Association for Engineering Education
,
Hamilton, New Zealand
.
Fiock
,
H.
(
2020
), “
Designing a community of inquiry in online courses
”,
The International Review of Research in Open and Distributed Learning
, Vol. 
21
No. 
1
, pp. 
135
-
153
, doi: .
Freeman
,
M.T.M.
and
Jarvie-Eggart
,
M.E.
(
2019
), “
Best practices in promoting faculty-student interaction in online STEM courses
”,
126th Annual Conference and Exposition
,
Tampa, Florida
.
Garrison
,
R.
(
2022
), “
Shared metacognition in a community of inquiry
”,
Online Learning
, Vol. 
26
No. 
1
, pp. 
6
-
18
, doi: .
Garrison
,
D.R.
and
Cleveland-Innes
,
M.
(
2005
), “
Facilitating cognitive presence in online learning: interaction is not enough
”,
The American Journal of Distance Education
, Vol. 
19
No. 
3
, pp. 
133
-
148
, doi: .
Garrison
,
D.R.
,
Anderson
,
T.
and
Archer
,
W.
(
1999
), “
Critical inquiry in a text-based environment: computer conferencing in higher education
”,
The Internet and Higher Education
, Vol. 
2
Nos
2-3
, pp. 
87
-
105
, doi: .
Hare
,
A.P.
,
Blumberg
,
H.H.
,
Davies
,
M.F.
and
Kent
,
M.V.
(
1994
),
Small Group Research: A Handbook
,
Ablex Publishing
.
Hillman
,
D.C.
,
Willis
,
D.J.
and
Gunawardena
,
C.N.
(
1994
), “
Learner–interface interaction in distance education: an extension of contemporary models and strategies for practitioners
”,
American Journal of Distance Education
, Vol. 
8
No. 
2
, pp. 
30
-
42
, doi: .
Kanuka
,
H.
and
Garrison
,
D.R.
(
2004
), “
Cognitive presence in online learning
”,
Journal of Computing in Higher Education
, Vol. 
15
No. 
2
, pp. 
21
-
39
, doi: .
Ke
,
F.
and
Xie
,
K.
(
2009
), “
Toward deep learning for adult students in online courses
”,
The Internet and Higher Education
, Vol. 
12
Nos
3-4
, pp. 
136
-
145
, doi: .
Lai
,
C.H.
,
Lin
,
H.W.
,
Lin
,
R.M.
and
Tho
,
P.D.
(
2019
), “
Effect of peer interaction among online learning community on learning engagement and achievement
”,
International Journal of Distance Education Technologies (IJDET)
, Vol. 
17
No. 
1
, pp. 
66
-
77
, doi: .
Lee
,
L.
(
2002
), “
Enhancing learners' communication skills through synchronous electronic interaction and task-based instruction
”,
Foreign Language Annals
, Vol. 
35
No. 
1
, pp. 
16
-
24
, doi: .
Lin
,
C.H.
,
Zheng
,
B.
and
Zhang
,
Y.
(
2017
), “
Interactions and learning outcomes in online language courses
”,
British Journal of Educational Technology
, Vol. 
48
No. 
3
, pp. 
730
-
748
, doi: .
Liu
,
J.C.
and
Kaye
,
E.R.
(
2016
), “Preparing online learning readiness with learner-content interaction: design for scaffolding self-regulated learning”, in
Handbook of Research on Strategic Management of Interaction, Presence, and Participation in Online Courses
,
IGI Global
, pp. 
216
-
243
.
Liu
,
W.
,
Xie
,
H.
and
Johnson
,
C.
(
2017
), “
The influencing factors and effects on comprehension of e-reading
”,
2nd Information Technology, Networking, Electronic and Automation Control Conference (ITNEC)
,
Chengdu, China
.
Majeski
,
R.
and
Stover
,
M.
(
2007
), “
Theoretically based pedagogical strategies leading to deep learning in asynchronous online gerontology courses
”,
Educational Gerontology
, Vol. 
33
No. 
3
, pp. 
171
-
185
, doi: .
Moon
,
J.
,
McNeill
,
L.
,
Edmonds
,
C.T.
,
Banihashem
,
S.K.
and
Noroozi
,
O.
(
2024
), “
Using learning analytics to explore peer learning patterns in asynchronous gamified environments
”,
International Journal of Educational Technology in Higher Education
, Vol. 
21
No. 
1
, p.
45
, doi: .
Moore
,
M.
(
1989
), “
Three types of interaction
”,
The American Journal of Distance Education
, Vol. 
3
No. 
2
, pp. 
1
-
7
, doi: .
Moore
,
M.
(
1993
), “Three types of interaction”, in
Distance Education: New Perspectives
, pp. 
19
-
24
.
Ng
,
B.J.M.
,
Han
,
J.Y.
,
Kim
,
Y.
,
Togo
,
K.A.
,
Chew
,
J.Y.
,
Lam
,
Y.
and
Fung
,
F.M.
(
2021
), “
Supporting social and learning presence in the revised community of inquiry framework for hybrid learning
”,
Journal of Chemical Education
, Vol. 
99
No. 
2
, pp. 
708
-
714
, doi: .
Palloff
,
R.M.
and
Pratt
,
K.
(
1999
),
Building Learning Communities in Cyberspace
,
Jossey-Bass
,
San Francisco
.
Redmond
,
P.
(
2014
), “
Reflection as an indicator of cognitive presence
”,
E-Learning and Digital Media
, Vol. 
11
No. 
1
, pp. 
46
-
58
, doi: .
Rourke
,
L.
,
Anderson
,
T.
and
Garrison
,
D.R.
(
1999
), “
Assessing social presence in asynchronous text-based computer conferencing
”,
Journal of Distance Education
, Vol. 
14
No. 
2
, pp. 
50
-
71
.
Rovai
,
A.P.
(
2002
), “
Building sense of community at a distance
”,
The International Review of Research in Open and Distributed Learning
, Vol. 
3
No. 
1
, pp. 
1
-
16
, doi: .
Shea
,
P.
and
Bidjerano
,
T.
(
2009
), “
Community of inquiry as a theoretical framework to foster ‘epistemic engagement’ and ‘cognitive presence’ in online education
”,
Computers and Education
, Vol. 
52
No. 
3
, pp. 
543
-
553
, doi: .
Shea
,
P.
and
Bidjerano
,
T.
(
2012
), “
Learning presence as a moderator in the community of inquiry model
”,
Computers and Education
, Vol. 
59
No. 
2
, pp. 
316
-
326
, doi: .
Shute
,
V.J.
(
2008
), “
Focus on formative feedback
”,
Review of Educational Research
, Vol. 
78
No. 
1
, pp. 
153
-
189
, doi: .
Stenbom
,
S.
(
2018
), “
A systematic review of the community of inquiry survey
”,
The Internet and Higher Education
, Vol. 
39
, pp. 
22
-
32
, doi: .
Swan
,
K.
and
Shih
,
L.F.
(
2005
), “
On the nature and development of social presence in online course discussions
”,
Journal of Asynchronous Learning Networks
, Vol. 
9
No. 
3
, pp. 
115
-
136
, doi: .
Vuopala
,
E.
,
Hyvönen
,
P.
and
Järvelä
,
S.
(
2016
), “
Interaction forms in successful collaborative learning in virtual learning environments
”,
Active Learning in Higher Education
, Vol. 
17
No. 
1
, pp. 
25
-
38
, doi: .
Wagner
,
E.D.
(
1997
), “
Interactivity: from agents to outcomes
”,
New Directions for Teaching and Learning
, Vol. 
1997
No. 
71
, pp. 
19
-
26
, doi: .
Xie
,
V.D.
and
Yen
,
L.L.
(
2011
), “
Relationship between students' motivation and their participation in asynchronous online discussions
”,
MERLOT Journal of Online Learning and Teaching
, Vol. 
7
No. 
1
, pp. 
17
-
29
.
Xie
,
H.
,
Liu
,
W.
and
Bhairma
,
J.
(
2018
), “
Analysis of synchronous and asynchronous e-learning environments
”,
3rd Joint International Information Technology, Mechanical and Electronic Engineering Conference (JIMEC)
,
Chongqing, China
.
Yu
,
Q.
and
Schunn
,
C.D.
(
2023
), “
Understanding the what and when of peer feedback benefits for performance and transfer
”,
Computers in Human Behavior
, Vol. 
147
, 107857, doi: .
Zhang
,
E.
,
Zhang
,
Z.
,
Liu
,
H.
,
Han
,
S.
and
Xue
,
Z.
(
2025
), “
Exploring peer facilitation and critical thinking in asynchronous online discussions: a lag sequential analysis approach
”,
British Journal of Educational Technology
, pp. 
2453
-
2477
, doi: .
Zhu
,
E.
(
2006
), “
Interaction and cognitive engagement: an analysis of four asynchronous online discussions
”,
Instructional Science
, Vol. 
34
No. 
6
, pp. 
451
-
480
, doi: .
Published by Emerald Publishing Limited. This article is published under the Creative Commons Attribution (CC BY 4.0) licence. Anyone may reproduce, distribute, translate and create derivative works of this article (for both commercial and non-commercial purposes), subject to full attribution to the original publication and authors. The full terms of this licence may be seen at Link to the terms of the CC BY 4.0 licence.

or Create an Account

Close Modal
Close Modal