Beyond Hostility: Detecting Subtle and Overt Forms of Online Conflict with Multi-Objective Learning

Authors

DOI:

https://doi.org/10.54195/irrj.25366

Keywords:

Subtle and Overt Hostility, Online Conflict Detection, Social Media Behaviour Analysis, Knowledge Distillation, Transformer-Based NLP

Abstract

Social networks have become a dominant and influential aspect of modern society, with numerous users engaging in observing, creating, and distributing content. The growth of content has led to user conflicts that include bullying, aggression, harassment, and threats. Consequently, recent research has been aimed at identifying and addressing these openly hostile forms of social conflict. However, in current studies, the less overtly hostile yet equally damaging types of conflict, including teasing, criticism, and sarcasm, have been overlooked. We introduce a comprehensive multi-class conflict dataset and develop a robust multi-objective classification model to capture the full spectrum of conflict; from subtle tensions to open hostility, significantly advancing conflict detection capabilities. This innovative approach leverages class-based reward functions to improve model performance and is implemented by fine-tuning pre-trained transformer language models within a decision transformer framework. We also propose and evaluate multiple knowledge distillation strategies that compress large LLM teachers into efficient student models. These methods result in statistically significant increases in classification performances, leading to further evaluation of performance-cost trade-off. Our experiments on three datasets demonstrate superior recall, precision, F1-score, and accuracy compared to traditional state-of-the-art deep learning classifiers. Furthermore, we analyse class ambiguity and its impact on model performance as well as conducting thematic analysis on model misclassifications.

Downloads

Download data is not yet available.

References

Elias Aboujaoude, Matthew W Savage, Vladan Starcevic, and Wael O Salame. Cyberbullying: Review of an old problem gone viral. Journal of adolescent health, 57(1):10–18, 2015.

Karmanya Aggarwal, Pakhi Bamdev, Debanjan Mahata, Rajiv Ratn Shah, Ponnurangam Kumaraguru, et al. Trawling for trolling: A dataset. arXiv:2008.00525, 2020.

Waseem Akram and Rekesh Kumar. A study on positive and negative effects of social media on society. International journal of computer sciences and engineering, 5(10):351–354, 2017.

Davey Alba and Kurt Wagner. Elon Musk cuts more Twitter staff overseeing content moderation, Jan 2023. URL https://www.bloomberg.com/news/articles/2023-01-07/elon-musk-cuts-more-twitter-staff-overseeing-content-moderation.

Mohsan Ali, Mehdi Hassan, Kashif Kifayat, Jin Young Kim, Saqib Hakak, and Muhammad Khurram Khan. Social media content classification and community detection using deep learning and graph analytics. Technological Forecasting and Social Change, 188:122252, 2023.

Fatimah Alkomah and Xiaogang Ma. A literature review of textual hate speech detection methods and datasets. Information, 13(6):273, 2022.

Brooke Auxier and Monica Anderson. Social media use in 2021. Pew Research Center, 1:1–4, 2021.

Salvador V Balkus and Donghui Yan. Improving short text classification with augmented data using GPT-3. Natural Language Engineering, 30(5):943–972, 2024.

Pietro Barbiero, Giovanni Squillero, and Alberto Tonda. Modeling generalization in machine learning: A methodological and computational study. arXiv:2006.15680, 2020.

Chloe Berryman, Christopher J Ferguson, and Charles Negy. Social media use and mental health among young adults. Psychiatric quarterly, 89:307–314, 2018.

Federico Bianchi, Stefanie Hills, Patricia Rossini, Dirk Hovy, Rebekah Tromble, and Nava Tintarev. “It’s not just hate”: A multi-dimensional perspective on detecting harmful speech online. In Proceedings of the 2022 conference on empirical methods in natural language processing, pages 8093–8099, 2022.

Layla Boroon, Babak Abedin, and Eila Erfani. The dark side of using online social networks: a review of individuals’ negative experiences. Journal of Global Information Management (JGIM), 29(6):1–21, 2021.

Mondher Bouazizi and Tomoaki Ohtsuki. Multi-class sentiment analysis on Twitter: Classification performance and challenges. Big Data Mining and Analytics, 2(3):181–194, 2019.

Virginia Braun and Victoria Clarke. Using thematic analysis in psychology. Qualitative research in psychology, 3(2):77–101, 2006.

Virginia Braun and Victoria Clarke. Can I use TA? Should I use TA? Should I not use TA? Comparing reflexive thematic analysis and other pattern-based qualitative analytic approaches. Counselling and psychotherapy research, 21(1):37–47, 2021.

Jan Breitsohl, Holger Roschk, and Christina Feyertag. Consumer brand bullying behaviour in online communities of service firms. Service Business Development: Band 2. Methoden–Erlösmodelle–Marketinginstrumente, pages 289–312, 2018.

Thomas Brewster. Musk’s X fired 80% of engineers working on trust and safety, australian government says, Aug 2024. URL https://www.forbes.com/sites/thomasbrewster/2024/01/10/elon-musk-fired-80-per-cent-of-twitter-x-engineers-working-on-trust-and-safety/.

Stefano Brogi. Online brand communities: a literature review. Procedia-Social and Behavioral Sciences, 109:385–389, 2014.

Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. Language models are few-shot learners. Advances in neural information processing systems, 33:1877–1901, 2020.

Taina Bucher, Anne Helmond, et al. The affordances of social media platforms. The SAGE handbook of social media, 1:233–253, 2018.

Matt Carlson. Embedded links, embedded meanings: Social media commentary and news sharing as mundane media criticism. Journalism studies, 17(7):915–924, 2016.

Tommaso Caselli, Valerio Basile, Jelena Mitrović, and Michael Granitzer. HateBERT: Retraining BERT for abusive language detection in English. In Proceedings of the 5th Workshop on Online Abuse and Harms (WOAH 2021), pages 17–25, 2021.

Tanmoy Chakraborty and Sarah Masud. Nipping in the bud: detection, diffusion and mitigation of hate speech on social media. ACM SIGWEB Newsletter (Winter):1–9, 2022.

Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Misha Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch. Decision transformer: Reinforcement learning via sequence modeling. Advances in neural information processing systems, 34:15084–

, 2021.

Justin Cheng, Cristian Danescu-Niculescu-Mizil, and Jure Leskovec. Antisocial behavior in online discussion communities. In Proceedings of the international aaai conference on web and social media, volume 9, pages 61–70, 2015.

Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al. Scaling instruction-finetuned language models. Journal of Machine Learning Research, 25(70):1–53, 2024.

Snehil Dahiya, Shalini Sharma, Dhruv Sahnan, Vasu Goel, Emilie Chouzenoux, Vı́ctor Elvira, Angshul Majumdar, Anil Bandhakavi, and Tanmoy Chakraborty. Would your tweet invoke hate on the fly? Forecasting hate intensity of reply threads on Twitter. In Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining, pages 2732–2742, 2021.

Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. Automated hate speech detection and the problem of offensive language. In Proceedings of the international AAAI conference on web and social media, volume 11, pages 512–515, 2017.

Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, volume 1 (long and short papers), pages 4171–4186, 2019.

Ashwin Geet d’Sa, Irina Illina, and Dominique Fohr. Classification of hate speech using deep neural networks. Revue d’Information Scientifique & Technique, 25(01), 2020.

Clare Duffy. Meta gets rid of fact checkers and says it will reduce ‘censorship’, CNN business, Jan 2025. URL https://edition.cnn.com/2025/01/07/tech/meta-censorship-moderation/index.html.

Aleksandra Edwards and Jose Camacho-Collados. Language models for text classification: Is in-context learning enough? In Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), pages 10058–10072, 2024.

F Elsafoury. Cyberbullying datasets. Mendeley. URL https://data.mendeley.com/datasets/jf4pzyvnpj/1, 2020.

Hyung Chung et al. Hugging face, Google FLAN-T5 base, 2023. URL https://huggingface.co/google/flan-t5-base.

Paula Fortuna and Sérgio Nunes. A survey on automatic detection of hate speech in text. ACM Computing Surveys (CSUR), 51(4):1–30, 2018.

Paula Fortuna, José Ferreira, Luiz Pires, Guilherme Routar, and Sérgio Nunes. Merging datasets for aggressive text identification. In Proceedings of the First Workshop on Trolling, Aggression and Cyberbullying (TRAC-2018), pages 128–139, 2018.

Paula Fortuna, Juan Soler, and Leo Wanner. Toxic, hateful, offensive or abusive? What are we really classifying? An empirical analysis of hate speech datasets. In Proceedings of the 12th language resources and evaluation conference, pages 6786–6794, 2020.

Paula Fortuna, Juan Soler-Company, and Leo Wanner. How well do hate speech, toxicity, abusive and offensive language classification models generalize across datasets? Information Processing & Management, 58(3):102524, 2021.

Antigoni Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. Large scale crowdsourcing and characterization of Twitter abusive behavior. In Twelfth International AAAI Conference on Web and Social Media, 2018.

Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. Realtoxicityprompts: Evaluating neural toxic degeneration in language models. In Findings of the association for computational linguistics: EMNLP 2020, pages 3356–3369, 2020.

Nicolas Gontier, Pau Rodriguez, Issam Laradji, David Vazquez, and Christopher Pal. Language decision transformers with exponential tilt for interactive text environments. arXiv:2302.05507, 2023.

Reginald H Gonzales. Social media as a channel and its implications on cyber bullying. In DLSU Research Congress, pages 1–7, 2014.

Shloak Gupta, S Bolden, Jay Kachhadia, A Korsunska, and J Stromer-Galley. Polibert: Classifying political social media messages with bert. In Social, cultural and behavioral modeling (SBP-BRIMS 2020) conference. Washington, DC, 2020.

Georgios Hadjiharalambous, Kacper Beisert, and Joemon M Jose. End-to-end hierarchical approach for emotion detection in short texts. In Responsible Data Science: Select Proceedings of ICDSE 2021, pages 1–12. Springer, 2022.

Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. Distilling the knowledge in a neural network. arXiv:1503.02531, 2015.

Sepp Hochreiter and Jürgen Schmidhuber. Long short-term memory. Neural computation, 9(8):1735–1780, 1997.

Michael Janner, Qiyang Li, and Sergey Levine. Offline reinforcement learning as one big sequence modeling problem. Advances in neural information processing systems, 34:1273–1286, 2021.

Shagun Jhaver, Larry Chan, and Amy Bruckman. The view from the other side: The border between controversial speech and harassment on kotaku in action. First Monday, 23(2):4, 2018.

Xiaoqi Jiao, Yichun Yin, Lifeng Shang, Xin Jiang, Xiao Chen, Linlin Li, Fang Wang, and Qun Liu. Tinybert: Distilling BERT for natural language understanding. In Findings of the association for computational linguistics: EMNLP 2020, pages 4163–4174, 2020.

Google Jigsaw. URL https://www.perspectiveapi.com/#/home.

Dacher Keltner, Lisa Capps, Ann M Kring, Randall C Young, and Erin A Heerey. Just teasing: a conceptual analysis and empirical review. Psychological bulletin, 127(2):229, 2001.

Muhammad US Khan, Assad Abbas, Attiqa Rehman, and Raheel Nawaz. Hateclassify: A service framework for hate speech identification on social media. IEEE Internet Computing, 25(1):40–49, 2020.

Mikhail Khodak, Nikunj Saunshi, and Kiran Vodrahalli. A large self-annotated corpus for sarcasm. In proceedings of the eleventh international conference on language resources and evaluation (LREC 2018), 2018.

Haesoo Kim, HaeEun Kim, Juho Kim, and Jeong-woo Jang. When does it become harassment? An investigation of online criticism and calling out in Twitter. Proceedings of the ACM on Human-Computer Interaction, 6(CSCW2):1–32, 2022.

Taehyeon Kim, Jaehoon Oh, Nak Yil Kim, Sangwook Cho, and Se-Young Yun. Comparing kullback-leibler divergence and mean squared error loss in knowledge distillation. In Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, pages 2628–2635. International Joint Conferences on Artificial Intelligence Organization, 2021.

Robin M Kowalski. “I was only kidding!”: Victims’ and perpetrators’ perceptions of teasing. Personality and Social Psychology Bulletin, 26(2):231–241, 2000.

Robin M Kowalski, Gary W Giumetti, Amber N Schroeder, and Micah R Lattanner. Bullying in the digital age: a critical review and meta-analysis of cyberbullying research among youth. Psychological bulletin, 140(4):1073, 2014.

Robert V. Kozinets. The field behind the screen: Using netnography for marketing research in online communities. Journal of Marketing Research, 39(1):61–72, 2002. doi: 10.1509/jmkr.39.1.61.18935.

Robert V Kozinets. Netnography: redefined. Sage, 2015.

Mateusz Lango and Jerzy Stefanowski. What makes multi-class imbalanced problems difficult? an experimental study. Expert Systems with Applications, 199:116962, 2022.

Loı̈c Lannelongue, Jason Grealey, and Michael Inouye. Green algorithms: quantifying the carbon footprint of computation. Advanced science, 8(12):2100707, 2021.

Deborah Roth Ledley, Eric A Storch, Meredith E Coles, Richard G Heimberg, Jason Moser, and Erica A Bravata. The relationship between childhood teasing and later interpersonal functioning. Journal of Psychopathology and Behavioral Assessment, 28:33–40, 2006.

Kuang-Huei Lee, Ofir Nachum, Mengjiao Sherry Yang, Lisa Lee, Daniel Freeman, Sergio Guadarrama, Ian Fischer, Winnie Xu, Eric Jang, Henryk Michalewski, et al. Multi-game decision transformers. Advances in Neural Information Processing Systems, 35:27921–

, 2022.

Jiacheng Li, Yujie Wang, and Julian McAuley. Time interval aware self-attention for sequential recommendation. In Proceedings of the 13th international conference on web search and data mining, pages 322–330, 2020.

Courtney Subramanian Liv McMahon, Zoe Kleinman. Meta to replace “biased” fact-checkers with moderation by users, Jan 2025. URL https://www.bbc.co.uk/news/articles/cly74mpy8klo.

Shayne Longpre, Le Hou, Tu Vu, Albert Webson, Hyung Won Chung, Yi Tay, Denny Zhou, Quoc V Le, Barret Zoph, Jason Wei, et al. The FLAN collection: Designing data and methods for effective instruction tuning. In International Conference on Machine Learning, pages 22631–22648. PMLR, 2023.

Ilya Loshchilov and Frank Hutter. Decoupled weight decay regularization. arXiv:1711.05101, 2017.

Ariadna Matamoros-Fernández and Johan Farkas. Racism, hate speech, and social media: A systematic review and critique. Television & New Media, 22(2):205–224, 2021.

Meta and Kaplan, Joel. More speech and fewer mistakes. https://about.fb.com/news/2025/01/meta-more-speech-fewer-mistakes/, January 2025. Newsroom Post.

Shervin Minaee, Nal Kalchbrenner, Erik Cambria, Narjes Nikzad, Meysam Chenaghlu, and Jianfeng Gao. Deep learning–based text classification: a comprehensive review. ACM computing surveys (CSUR), 54(3):1–40, 2021.

Ioannis Mollas, Zoe Chrysopoulou, Stamatis Karlos, and Grigorios Tsoumakas. Ethos: a multi-label hate speech detection dataset. Complex & Intelligent Systems, 8(6):4663–4678, 2022.

Raymond T Mutanga, Nalindren Naicker, and Oludayo O Olugbara. Hate speech detection in Twitter using transformer methods. International Journal of Advanced Computer Science and Applications, 11(9), 2020.

Usman Naseem, Imran Razzak, and Ibrahim A Hameed. Deep context-aware embedding for abusive and hate speech detection on Twitter. Aust. J. Intell. Inf. Process. Syst., 15 (3):69–76, 2019.

Esteban Ortiz-Ospina and Max Roser. The rise of social media. Our world in data, 2023.

Keiron O’Shea and Ryan Nash. An introduction to convolutional neural networks. arXiv:1511.08458, 2015.

Fabio Poletto, Valerio Basile, Manuela Sanguinetti, Cristina Bosco, and Viviana Patti. Resources and benchmark corpora for hate speech detection: a systematic review. Language Resources and Evaluation, 55:477–523, 2021.

Jing Qian, Mai ElSherief, Elizabeth Belding, and William Yang Wang. Leveraging intra-user and inter-user representation learning for automated hate speech detection. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 2 (Short Papers), pages 118–123, 2018.

Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. Language models are unsupervised multitask learners. OpenAI blog, 1(8):9, 2019.

Nils Reimers and Iryna Gurevych. all-mpnet-base-v2. https://huggingface.co/sentence-transformers/all-mpnet-base-v2, 2021. Accessed: 2025-11-14.

Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio. Fitnets: Hints for thin deep nets, 2015. URL https://arxiv.org/abs/1412.6550.

Dhruv Sahnan, Snehil Dahiya, Vasu Goel, Anil Bandhakavi, and Tanmoy Chakraborty. Better prevent than react: Deep stratified learning to predict hate intensity of Twitter reply chains. In 2021 IEEE International Conference on Data Mining (ICDM), pages 549–558. IEEE, 2021.

Joni Salminen, Maximilian Hopf, Shammur A Chowdhury, Soon-gyo Jung, Hind Almerekhi, and Bernard J Jansen. Developing an online hate classifier for multiple social media platforms. Human-centric Computing and Information Sciences, 10:1–34, 2020.

Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter. arXiv:1910.01108, 2019.

Vishwam Sankaran, Nov 2022. URL https://www.independent.co.uk/independentpremium/world/elon-musk-twitter-layoffs-moderators-b2224818.html.

Shabnoor Siddiqui, Tajinder Singh, et al. Social media its impact with positive and negative aspects. International journal of computer applications technology and research, 5(2):71–75, 2016.

Ge Song, Yunming Ye, Xiaolin Du, Xiaohui Huang, and Shifu Bie. Short text classification: a survey. Journal of multimedia, 9(5), 2014.

Fei Sun, Jun Liu, Jian Wu, Changhua Pei, Xiao Lin, Wenwu Ou, and Peng Jiang. BERT4rec: Sequential recommendation with bidirectional encoder representations from transformer. In Proceedings of the 28th ACM international conference on information and knowledge management, pages 1441–1450, 2019a.

Siqi Sun, Yu Cheng, Zhe Gan, and Jingjing Liu. Patient knowledge distillation for bert model compression. In Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP), pages 4323–4332, 2019b.

Stuart A. Thompson and Kate Conger. Meet the next fact-checker, debunker and moderator: You, Jan 2025. URL https://www.nytimes.com/2025/01/07/technology/meta-facebook-content-moderation.html.

Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. Llama: Open and efficient foundation language models. arXiv:2302.13971, 2023.

Iulia Turc, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. Well-read students learn better: On the importance of pre-training compact models, 2019.

Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin. Attention is all you need. Advances in neural information processing systems, 30, 2017.

Bertie Vidgen, Tristan Thrush, Zeerak Talat, and Douwe Kiela. Learning from the worst: Dynamically generated datasets to improve online hate detection. In Proceedings of the 59th annual meeting of the Association for Computational Linguistics and the 11th international joint conference on natural language processing (volume 1: long papers), pages 1667–1682, 2021b.

Eric Wallace, Yizhong Wang, Sujian Li, Sameer Singh, and Matt Gardner. Do NLP models know numbers? Probing numeracy in embeddings. In Proceedings of the Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP), pages 5307–5315, 2019.

Jie Wang, Alexandros Karatzoglou, Ioannis Arapakis, Xin Xin, Xuri Ge, and Joemon M. Jose. Sparks of surprise: Multi-objective recommendations with hierarchical decision transformers for diversity, novelty, and serendipity. In Proceedings of the 33rd ACM

International Conference on Information and Knowledge Management, CIKM ’24, page 2358–2368, 2024. Association for Computing Machinery. doi: 10.1145/3627673.3679533.

Jie Wang, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. Jose, and Xuri Ge. Beyond accuracy: Decision transformers for reward-driven multi-objective recommendations. IEEE Transactions on Knowledge and Data Engineering, 37(9):5004–5016, 2025. doi: 10.1109/TKDE.2025.3582506.

Qiong Wang, Ruilin Tu, Yihe Jiang, Wei Hu, and Xiao Luo. Teasing and internet harassment among adolescents: The mediating role of envy and the moderating role of the Zhong-Yong thinking style. International Journal of Environmental Research and Public Health, 19(9):5501, 2022.

Tianyi Wang, Ke Lu, Kam Pui Chow, and Qing Zhu. Covid-19 sensing: negative sentiment analysis on social media in china via bert model. Ieee Access, 8:138162–138169, 2020a.

Wenhui Wang, Furu Wei, Li Dong, Hangbo Bao, Nan Yang, and Ming Zhou. Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers. Advances in neural information processing systems, 33:5776–5788, 2020b.

Xia Wang, Chunling Yu, and Yujie Wei. Social media peer communication and impacts on purchase intentions: A consumer socialization framework. Journal of interactive marketing, 26(4):198–208, 2012.

Oliver Warke, Joemon M Jose, Jan Breitsohl, and Jie Wang. Capturing the spectrum of social media conflict: A novel multi-objective classification model. In Proceedings of the ACM SIGIR International Conference on Theory of Information Retrieval, pages 215–225, 2024.

Janis Wolak, Kimberly J Mitchell, and David Finkelhor. Does online harassment constitute bullying? an exploration of online harassment by known peers and online-only contacts. Journal of adolescent health, 41(6):S51–S58, 2007.

Ellery Wulczyn, Nithum Thain, and Lucas Dixon. Wikipedia Talk Labels: Personal Attacks. 2 2017a. doi: 10.6084/m9.figshare.4054689.v6. URL https://figshare.com/articles/dataset/Wikipedia_Talk_Labels_Personal_Attacks/4054689.

Ellery Wulczyn, Nithum Thain, and Lucas Dixon. Ex machina: Personal attacks seen at scale. In Proceedings of the 26th international conference on world wide web, pages 1391–1399, 2017b.

Lan Xia. Effects of companies’ responses to consumer criticism in social media. International Journal of Electronic Commerce, 17(4):73–100, 2013.

Chuanpeng Yang, Yao Zhu, Wang Lu, Yidong Wang, Qian Chen, Chenlong Gao, Bingjie Yan, and Yiqiang Chen. Survey on knowledge distillation for large language models: methods, evaluation, and application. ACM Transactions on Intelligent Systems and Technology, 2024.

Sergey Zagoruyko and Nikos Komodakis. Paying more attention to attention: Improving the performance of convolutional neural networks via attention transfer. arXiv:1612.03928, 2016.

Mark Zuckerberg. More speech and fewer mistakes. Facebook Video, January 2025. URL https://www.facebook.com/watch/?v=1525382954801931.

Downloads

Published

2026-06-17

Issue

Section

Articles

How to Cite

Warke, O., Jose, J., & Breitsohl, J. (2026). Beyond Hostility: Detecting Subtle and Overt Forms of Online Conflict with Multi-Objective Learning. Information Retrieval Research, 2(1), 113-158. https://doi.org/10.54195/irrj.25366