Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Evaluation of Prompting Strategies for Cyberbullying Detection Using Various Large Language Models

View through CrossRef
Sentiment analysis detects toxic language for safer online spaces and helps businesses refine strategies through customer feedback analysis [1, 2]. Advancements in Large Language Models (LLMs) and prompt engineering have introduced novel approaches to sentiment analysis, cyberbullying detection, and toxicity classification. However, several challenges persist, particularly in handling text ambiguity, sarcasm, multilingual contexts, and nuanced emotional comprehension, which limit the ability to achieve accurate and human-aligned results. This study uses the CYBY23 dataset, which contains 112 human-annotated threads. To balance the dataset, synthetic threads were generated using ChatGPT, resulting in a final dataset of 148 threads evenly distributed across two labels: 0 (bullying with no aggression) and 1 (bullying with aggression). Three publicly available LLMs—Deepseek-r1-distillllama-70b (Deepseek), Qwen-2.5-32b (Qwen) and llama3-70b-8192 (Llama)—were systematically evaluated using zero-shot, one-shot, and few-shot prompting strategies, with all models accessed via Groq Cloud APIs. The model outputs were assessed using recall, precision, F1 scores, and accuracy to measure performance in different prompting techniques (PT). In this report, Qwen achieved the highest overall accuracy at 82.43% in few-shot 2, while Llama matched that accuracy in one-shot 2, demonstrating solid performance in few-shot tasks as well. Deepseek showed high variability, thriving with contextual enhancements in zero-shot 2 but struggling in one-shot and fluctuating in few-shot settings. one-shot prompting proved most effective for Llama, while few-shot methods worked best for both Qwen and Llama.
Title: Evaluation of Prompting Strategies for Cyberbullying Detection Using Various Large Language Models
Description:
Sentiment analysis detects toxic language for safer online spaces and helps businesses refine strategies through customer feedback analysis [1, 2].
Advancements in Large Language Models (LLMs) and prompt engineering have introduced novel approaches to sentiment analysis, cyberbullying detection, and toxicity classification.
However, several challenges persist, particularly in handling text ambiguity, sarcasm, multilingual contexts, and nuanced emotional comprehension, which limit the ability to achieve accurate and human-aligned results.
This study uses the CYBY23 dataset, which contains 112 human-annotated threads.
To balance the dataset, synthetic threads were generated using ChatGPT, resulting in a final dataset of 148 threads evenly distributed across two labels: 0 (bullying with no aggression) and 1 (bullying with aggression).
Three publicly available LLMs—Deepseek-r1-distillllama-70b (Deepseek), Qwen-2.
5-32b (Qwen) and llama3-70b-8192 (Llama)—were systematically evaluated using zero-shot, one-shot, and few-shot prompting strategies, with all models accessed via Groq Cloud APIs.
The model outputs were assessed using recall, precision, F1 scores, and accuracy to measure performance in different prompting techniques (PT).
In this report, Qwen achieved the highest overall accuracy at 82.
43% in few-shot 2, while Llama matched that accuracy in one-shot 2, demonstrating solid performance in few-shot tasks as well.
Deepseek showed high variability, thriving with contextual enhancements in zero-shot 2 but struggling in one-shot and fluctuating in few-shot settings.
one-shot prompting proved most effective for Llama, while few-shot methods worked best for both Qwen and Llama.

Related Results

Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
<p><em><span style="font-size: 11.0pt; font-family: 'Times New Roman',serif; mso-fareast-font-family: 'Times New Roman'; mso-ansi-language: EN-US; mso-fareast-langua...
Pengaruh Parental Attachment terhadap Perilaku Cyberbullying pada Remaja di Jawa Barat
Pengaruh Parental Attachment terhadap Perilaku Cyberbullying pada Remaja di Jawa Barat
Abstract. Bullying behavior is now familiar, but along with the development of internet technology, bullying behavior that initially occurred directly has now turned into cyberbull...
Pengaruh Kepribadian (Five Factor Personality) terhadap Perilaku Cyberbullying pada Pengguna Media Sosial
Pengaruh Kepribadian (Five Factor Personality) terhadap Perilaku Cyberbullying pada Pengguna Media Sosial
Abstract. The rapid development of technology provides many conveniences for its users, such as the existence of social media that makes it easier for individuals to interact, uplo...
Tinjauan Kriminologis Mengenai Cyberbullying di Kota Makassar
Tinjauan Kriminologis Mengenai Cyberbullying di Kota Makassar
This thesis, entitled Criminological Review of Cyberbullying in Makassar, is motivated because the author has seen many reports in the mass media about cyberbullying outside Makass...
Pengaruh Loneliness terhadap Perilaku Cyberbullying pada Mahasiswa Jawa Barat
Pengaruh Loneliness terhadap Perilaku Cyberbullying pada Mahasiswa Jawa Barat
Abstract. Technology continues to develop makes humans easier to do their activities, especially with the emergence of the internet which makes humans have other lives in cyberspac...
School Staff's Perceptions and Attitudes towards Cyberbullying
School Staff's Perceptions and Attitudes towards Cyberbullying
<p>Parallel with the spread of technology use, cyberbullying has become a serious problem in schools, particularly those in developed countries where most young people have r...
Pengaruh Self-Esteem terhadap Cyberbullying Victimization pada Remaja di Jawa Barat
Pengaruh Self-Esteem terhadap Cyberbullying Victimization pada Remaja di Jawa Barat
Abstract. Bullying is a problem that occurs a lot. Now with the development of technology and the internet, bullying does not only occur physically but also virtually or called cyb...
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
The actual use of classroom language is principally limited to the classroom environment. As far as foreign language learning is concerned, the classroom often turns out to be the ...

Back to Top