Paloma Piot
IRLab, UDC
A Coruña, Spain
I am a PhD researcher in Natural Language Processing at the IRLab, University of A Coruña, Spain. My research focuses on how large language models (LLMs) perceive, assess, and sometimes reproduce hate speech and harmful content in social media. I am particularly interested in the intersection of AI and society, examining how language technologies can reflect social biases and how they can be developed to be more responsible, transparent, and trustworthy.
My work centers on hate speech and abusive language detection, multilingual and low-resource NLP, and the evaluation of LLMs across different languages, cultures, and communities. I investigate the social and sociolinguistic dimensions of online discourse, with an emphasis on fairness, explainability, and robustness in language technologies. I am also interested in data-centric approaches, synthetic data generation, and the development of datasets and evaluation frameworks that help improve the detection and mitigation of online harm.
I am committed to open and reproducible research, and I maintain open-source implementations of my work and related projects on GitHub. You can also follow my academic and professional updates on LinkedIn and X.
Feel free to reach out via email at <paloma.piot[at]udc.es>.
Selected Publications
- MetaHate: A Dataset for Unifying Efforts on Hate Speech DetectionIn Proceedings of the International AAAI Conference on Web and Social Media, 2024
- Decoding hate: Exploring language models’ reactions to hate speechIn Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), 2025
- Towards Efficient and Explainable Hate Speech Detection via Model DistillationIn European Conference on Information Retrieval, 2025
- Personalisation or prejudice? addressing geographic bias in hate speech detection using debias tuning in large language modelsIn Proceedings of the International AAAI Conference on Web and Social Media, 2026
- WATCHED: A Web AI Agent Tool for Combating Hate speech by Expanding DataSoftwareX, 2025