Javascript must be enabled to continue!
Evaluation of Prompting Strategies for Cyberbullying Detection Using Various Large Language Models
View through CrossRef
Sentiment analysis detects toxic language for safer online spaces and helps businesses refine
strategies through customer feedback analysis [1, 2]. Advancements in Large Language
Models (LLMs) and prompt engineering have introduced novel approaches to sentiment
analysis, cyberbullying detection, and toxicity classification. However, several challenges
persist, particularly in handling text ambiguity, sarcasm, multilingual contexts, and nuanced
emotional comprehension, which limit the ability to achieve accurate and human-aligned
results. This study uses the CYBY23 dataset, which contains 112 human-annotated threads.
To balance the dataset, synthetic threads were generated using ChatGPT, resulting in a final
dataset of 148 threads evenly distributed across two labels: 0 (bullying with no aggression) and 1 (bullying with aggression). Three publicly available LLMs—Deepseek-r1-distillllama-70b (Deepseek), Qwen-2.5-32b (Qwen) and llama3-70b-8192 (Llama)—were systematically evaluated using zero-shot, one-shot, and few-shot prompting strategies, with all models accessed via Groq Cloud APIs. The model outputs were assessed using recall, precision,
F1 scores, and accuracy to measure performance in different prompting techniques (PT). In
this report, Qwen achieved the highest overall accuracy at 82.43% in few-shot 2, while Llama
matched that accuracy in one-shot 2, demonstrating solid performance in few-shot tasks as
well. Deepseek showed high variability, thriving with contextual enhancements in zero-shot
2 but struggling in one-shot and fluctuating in few-shot settings. one-shot prompting proved
most effective for Llama, while few-shot methods worked best for both Qwen and Llama.
Advances in Artificial Intelligence and Machine Learning
Title: Evaluation of Prompting Strategies for Cyberbullying Detection Using Various Large Language Models
Description:
Sentiment analysis detects toxic language for safer online spaces and helps businesses refine
strategies through customer feedback analysis [1, 2].
Advancements in Large Language
Models (LLMs) and prompt engineering have introduced novel approaches to sentiment
analysis, cyberbullying detection, and toxicity classification.
However, several challenges
persist, particularly in handling text ambiguity, sarcasm, multilingual contexts, and nuanced
emotional comprehension, which limit the ability to achieve accurate and human-aligned
results.
This study uses the CYBY23 dataset, which contains 112 human-annotated threads.
To balance the dataset, synthetic threads were generated using ChatGPT, resulting in a final
dataset of 148 threads evenly distributed across two labels: 0 (bullying with no aggression) and 1 (bullying with aggression).
Three publicly available LLMs—Deepseek-r1-distillllama-70b (Deepseek), Qwen-2.
5-32b (Qwen) and llama3-70b-8192 (Llama)—were systematically evaluated using zero-shot, one-shot, and few-shot prompting strategies, with all models accessed via Groq Cloud APIs.
The model outputs were assessed using recall, precision,
F1 scores, and accuracy to measure performance in different prompting techniques (PT).
In
this report, Qwen achieved the highest overall accuracy at 82.
43% in few-shot 2, while Llama
matched that accuracy in one-shot 2, demonstrating solid performance in few-shot tasks as
well.
Deepseek showed high variability, thriving with contextual enhancements in zero-shot
2 but struggling in one-shot and fluctuating in few-shot settings.
one-shot prompting proved
most effective for Llama, while few-shot methods worked best for both Qwen and Llama.
Related Results
Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
<p><em><span style="font-size: 11.0pt; font-family: 'Times New Roman',serif; mso-fareast-font-family: 'Times New Roman'; mso-ansi-language: EN-US; mso-fareast-langua...
Pengaruh Parental Attachment terhadap Perilaku Cyberbullying pada Remaja di Jawa Barat
Pengaruh Parental Attachment terhadap Perilaku Cyberbullying pada Remaja di Jawa Barat
Abstract. Bullying behavior is now familiar, but along with the development of internet technology, bullying behavior that initially occurred directly has now turned into cyberbull...
Pengaruh Kepribadian (Five Factor Personality) terhadap Perilaku Cyberbullying pada Pengguna Media Sosial
Pengaruh Kepribadian (Five Factor Personality) terhadap Perilaku Cyberbullying pada Pengguna Media Sosial
Abstract. The rapid development of technology provides many conveniences for its users, such as the existence of social media that makes it easier for individuals to interact, uplo...
Tinjauan Kriminologis Mengenai Cyberbullying di Kota Makassar
Tinjauan Kriminologis Mengenai Cyberbullying di Kota Makassar
This thesis, entitled Criminological Review of Cyberbullying in Makassar, is motivated because the author has seen many reports in the mass media about cyberbullying outside Makass...
Pengaruh Loneliness terhadap Perilaku Cyberbullying pada Mahasiswa Jawa Barat
Pengaruh Loneliness terhadap Perilaku Cyberbullying pada Mahasiswa Jawa Barat
Abstract. Technology continues to develop makes humans easier to do their activities, especially with the emergence of the internet which makes humans have other lives in cyberspac...
School Staff's Perceptions and Attitudes towards Cyberbullying
School Staff's Perceptions and Attitudes towards Cyberbullying
<p>Parallel with the spread of technology use, cyberbullying has become a serious problem in schools, particularly those in developed countries where most young people have r...
Pengaruh Self-Esteem terhadap Cyberbullying Victimization pada Remaja di Jawa Barat
Pengaruh Self-Esteem terhadap Cyberbullying Victimization pada Remaja di Jawa Barat
Abstract. Bullying is a problem that occurs a lot. Now with the development of technology and the internet, bullying does not only occur physically but also virtually or called cyb...
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
The actual use of classroom language is principally limited to the classroom environment. As far as foreign language learning is concerned, the classroom often turns out to be the ...

