Robust Educational Dialogue Act Classifiers with Low-Resource and Imbalanced Datasets

Lin, Jionghao; Tan, Wei; Nguyen, Ngoc Dang; Lang, David; Du, Lan; Buntine, Wray; Beare, Richard; Chen, Guanliang; Gašević, Dragan

doi:10.1007/978-3-031-36272-9_10

Jionghao Lin^12,13,
Wei Tan¹²,
Ngoc Dang Nguyen¹²,
David Lang¹⁴,
Lan Du¹²,
Wray Buntine^12,15,
Richard Beare^12,16,
Guanliang Chen¹² &
…
Dragan Gašević¹²

Part of the book series: Lecture Notes in Computer Science ((LNAI,volume 13916))

Included in the following conference series:

International Conference on Artificial Intelligence in Education

3496 Accesses
1 Citations

Abstract

Dialogue acts (DAs) can represent conversational actions of tutors or students that take place during tutoring dialogues. Automating the identification of DAs in tutoring dialogues is significant to the design of dialogue-based intelligent tutoring systems. Many prior studies employ machine learning models to classify DAs in tutoring dialogues and invest much effort to optimize the classification accuracy by using limited amounts of training data (i.e., low-resource data scenario). However, beyond the classification accuracy, the robustness of the classifier is also important, which can reflect the capability of the classifier on learning the patterns from different class distributions. We note that many prior studies on classifying educational DAs employ cross entropy (CE) loss to optimize DA classifiers on low-resource data with imbalanced DA distribution. The DA classifiers in these studies tend to prioritize accuracy on the majority class at the expense of the minority class which might not be robust to the data with imbalanced ratios of different DA classes. To optimize the robustness of classifiers on imbalanced class distributions, we propose to optimize the performance of the DA classifier by maximizing the area under the ROC curve (AUC) score (i.e., AUC maximization). Through extensive experiments, our study provides evidence that (i) by maximizing AUC in the training process, the DA classifier achieves significant performance improvement compared to the CE approach under low-resource data, and (ii) AUC maximization approaches can improve the robustness of the DA classifier under different class imbalance ratios.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Log in via an institution

Subscribe and save

Springer+ Basic

EUR 32.99 /Month

Get 10 units per month
Download Article/Chapter or Ebook
1 Unit = 1 Article or 1 Chapter
Cancel anytime

Subscribe now

Buy Now

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 99.00; Price excludes VAT (USA)

Softcover Book: USD 129.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

Notes

1.
FP, Positive Feedback “Well done!”, same abbreviation from [25].
2.
Our study has 31 dialogue acts as the classes to be classified.

References

Al-Luhaybi, M., Yousefi, L., Swift, S., Counsell, S., Tucker, A.: Predicting academic performance: a bootstrap** approach for learning dynamic Bayesian networks. In: Isotani, S., Millán, E., Ogan, A., Hastings, P., McLaren, B., Luckin, R. (eds.) AIED 2019. LNCS (LNAI), vol. 11625, pp. 26–36. Springer, Cham (2019). https://doi.org/10.1007/978-3-030-23204-7_3
Chapter Google Scholar
Boyer, K., Ha, E.Y., Phillips, R., Wallis, M., Vouk, M., Lester, J.: Dialogue act modeling in a complex task-oriented domain. In: Proceedings of the SIGDIAL 2010 Conference, pp. 297–305 (2010)
Google Scholar
Cavalcanti, A.P., et al.: How good is my feedback? A content analysis of written feedback. In: Proceedings of the LAK, LAK 2020, pp. 428–437. ACM, New York (2020)
Google Scholar
Devlin, J., Chang, M.W., Lee, K., Toutanova, K.: BERT: pre-training of deep bidirectional transformers for language understanding. In: Proceedings of NAACL-HLT, Minneapolis, Minnesota, pp. 4171–4186. Association for Computational Linguistics (2019)
Google Scholar
D’Mello, S., Olney, A., Person, N.: Mining collaborative patterns in tutorial dialogues. J. Educ. Data Min. 2(1), 1–37 (2010)
Google Scholar
Du Boulay, B., Luckin, R.: Modelling human teaching tactics and strategies for tutoring systems: 14 years on. Int. J. Artif. Intell. Educ. 26(1), 393–404 (2016)
Article Google Scholar
Ezen-Can, A., Boyer, K.E.: Understanding student language: an unsupervised dialogue act classification approach. J. Educ. Data Min. 7(1), 51–78 (2015)
Google Scholar
Ezen-Can, A., Grafsgaard, J.F., Lester, J.C., Boyer, K.E.: Classifying student dialogue acts with multimodal learning analytics. In: Proceedings of the Fifth LAK, pp. 280–289 (2015)
Google Scholar
Lin, J., et al.: Is it a good move? Mining effective tutoring strategies from human–human tutorial dialogues. Futur. Gener. Comput. Syst. 127, 194–207 (2022)
Article Google Scholar
Lin, J., et al.: Enhancing educational dialogue act classification with discourse context and sample informativeness. IEEE TLT (under reviewing process)
Google Scholar
Min, W., et al.: Predicting dialogue acts for intelligent virtual agents with multimodal student interaction data. Int. Educ. Data Min. Soc. (2016)
Google Scholar
Nguyen, N.D., Tan, W., Buntine, W., Beare, R., Chen, C., Du, L.: AUC maximization for low-resource named entity recognition. In: Proceedings of the AAAI Conference on Artificial Intelligence (2023)
Google Scholar
Nye, B.D., Graesser, A.C., Hu, X.: Autotutor and family: a review of 17 years of natural language tutoring. Int. J. Artif. Intell. Educ. 24(4), 427–469 (2014)
Article Google Scholar
Nye, B.D., Morrison, D.M., Samei, B.: Automated session-quality assessment for human tutoring based on expert ratings of tutoring success. Int. Educ. Data Min. Soc. (2015)
Google Scholar
Raković, M., et al.: Towards the automated evaluation of legal casenote essays. In: Rodrigo, M.M., Matsuda, N., Cristea, A.I., Dimitrova, V. (eds.) AIED 2022. LNCS, vol. 13355, pp. 167–179. Springer, Cham (2022). https://doi.org/10.1007/978-3-031-11644-5_14
Chapter Google Scholar
Rus, V., Maharjan, N., Banjade, R.: Dialogue act classification in human-to-human tutorial dialogues. In: Innovations in Smart Learning. LNET, pp. 183–186. Springer, Singapore (2017). https://doi.org/10.1007/978-981-10-2419-1_25
Chapter Google Scholar
Rus, V., et al.: An analysis of human tutors’ actions in tutorial dialogues. In: The Thirtieth International Flairs Conference (2017)
Google Scholar
Samei, B., Li, H., Keshtkar, F., Rus, V., Graesser, A.C.: Context-based speech act classification in intelligent tutoring systems. In: Trausan-Matu, S., Boyer, K.E., Crosby, M., Panourgia, K. (eds.) ITS 2014. LNCS, vol. 8474, pp. 236–241. Springer, Cham (2014). https://doi.org/10.1007/978-3-319-07221-0_28
Chapter Google Scholar
Samei, B., Rus, V., Nye, B., Morrison, D.M.: Hierarchical dialogue act classification in online tutoring sessions. In: EDM, pp. 600–601 (2015)
Google Scholar
Sha, L., et al.: Is the latest the greatest? A comparative study of automatic approaches for classifying educational forum posts. IEEE Trans. Learn. Technol. 1–14 (2022)
Google Scholar
Tan, W., Du, L., Buntine, W.: Diversity enhanced active learning with strictly proper scoring rules. In: Advances in Neural Information Processing Systems, vol. 34, pp. 10906–10918 (2021)
Google Scholar
Tan, W., et al.: Does informativeness matter? Active learning for educational dialogue act classification. In: Wang, N., et al. (eds.) AIED 2023. LNAI, vol. 13916, pp. 176–188. Springer, Cham (2023)
Google Scholar
Taori, R., Dave, A., Shankar, V., Carlini, N., Recht, B., Schmidt, L.: Measuring robustness to natural distribution shifts in image classification. In: Advances in Neural Information Processing Systems, vol. 33, pp. 18583–18599 (2020)
Google Scholar
Vail, A.K., Grafsgaard, J.F., Boyer, K.E., Wiebe, E.N., Lester, J.C.: Predicting learning from student affective response to tutor questions. In: Micarelli, A., Stamper, J., Panourgia, K. (eds.) ITS 2016. LNCS, vol. 9684, pp. 154–164. Springer, Cham (2016). https://doi.org/10.1007/978-3-319-39583-8_15
Chapter Google Scholar
Vail, A.K., Boyer, K.E.: Identifying effective moves in tutoring: on the refinement of dialogue act annotation schemes. In: Trausan-Matu, S., Boyer, K.E., Crosby, M., Panourgia, K. (eds.) ITS 2014. LNCS, vol. 8474, pp. 199–209. Springer, Cham (2014). https://doi.org/10.1007/978-3-319-07221-0_24
Chapter Google Scholar
VanLehn, K., Graesser, A.C., Jackson, G.T., Jordan, P., Olney, A., Rosé, C.P.: When are tutorial dialogues more effective than reading? Cogn. Sci. 31(1), 3–62 (2007)
Article Google Scholar
Yang, T., Ying, Y.: AUC maximization in the era of big data and AI: a survey. ACM Comput. Surv. 55(8) (2022). https://doi.org/10.1145/3554729
Ying, Y., Wen, L., Lyu, S.: Stochastic online AUC maximization. In: Advances in Neural Information Processing Systems, vol. 29 (2016)
Google Scholar
Yuan, Z., Yan, Y., Sonka, M., Yang, T.: Large-scale robust deep AUC maximization: a new surrogate loss and empirical studies on medical image classification. In: 2021 IEEE/CVF ICCV, Los Alamitos, CA, USA, pp. 3020–3029. IEEE Computer Society (2021)
Google Scholar
Yuan, Z., Guo, Z., Chawla, N., Yang, T.: Compositional training for end-to-end deep AUC maximization. In: International Conference on Learning Representations (2022). https://openreview.net/forum?id=gPvB4pdu_Z
Zhao, L., et al.: METS: multimodal learning analytics of embodied teamwork learning. In: LAK23: 13th International Learning Analytics and Knowledge Conference, pp. 186–196 (2023)
Google Scholar

Download references

Author information

Authors and Affiliations

Monash University, Clayton, Australia
Jionghao Lin, Wei Tan, Ngoc Dang Nguyen, Lan Du, Wray Buntine, Richard Beare, Guanliang Chen & Dragan Gašević
Carnegie Mellon University, Pittsburgh, USA
Jionghao Lin
Stanford University, Stanford, USA
David Lang
VinUniversity, Hanoi, Vietnam
Wray Buntine
Murdoch Children’s Research Institute, Melbourne, Australia
Richard Beare

Authors

Jionghao Lin
View author publications
You can also search for this author in PubMed Google Scholar
Wei Tan
View author publications
You can also search for this author in PubMed Google Scholar
Ngoc Dang Nguyen
View author publications
You can also search for this author in PubMed Google Scholar
David Lang
View author publications
You can also search for this author in PubMed Google Scholar
Lan Du
View author publications
You can also search for this author in PubMed Google Scholar
Wray Buntine
View author publications
You can also search for this author in PubMed Google Scholar
Richard Beare
View author publications
You can also search for this author in PubMed Google Scholar
Guanliang Chen
View author publications
You can also search for this author in PubMed Google Scholar
Dragan Gašević
View author publications
You can also search for this author in PubMed Google Scholar

Corresponding authors

Correspondence to Wei Tan or Ngoc Dang Nguyen .

Editor information

Editors and Affiliations

University of Southern California, Los Angeles, CA, USA
Ning Wang
University of British Columbia, Vancouver, BC, Canada
Genaro Rebolledo-Mendez
North Carolina State University, Raleigh, NC, USA
Noboru Matsuda
Despacho 3.01, UNED-Grupo de Investigación aDeNu, Madrid, Spain
Olga C. Santos
University of Leeds, Leeds, UK
Vania Dimitrova

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Lin, J. et al. (2023). Robust Educational Dialogue Act Classifiers with Low-Resource and Imbalanced Datasets. In: Wang, N., Rebolledo-Mendez, G., Matsuda, N., Santos, O.C., Dimitrova, V. (eds) Artificial Intelligence in Education. AIED 2023. Lecture Notes in Computer Science(), vol 13916. Springer, Cham. https://doi.org/10.1007/978-3-031-36272-9_10

Download citation

DOI: https://doi.org/10.1007/978-3-031-36272-9_10
Published: 26 June 2023
Publisher Name: Springer, Cham
Print ISBN: 978-3-031-36271-2
Online ISBN: 978-3-031-36272-9
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics

Robust Educational Dialogue Act Classifiers with Low-Resource and Imbalanced Datasets