Cardiff University | Prifysgol Caerdydd ORCA
Online Research @ Cardiff 
WelshClear Cookie - decide language by browser settings

Browse by All Cardiff Authors

Number of items: 12.

Shafiei, Mohammadamin, Saffari, Hamidreza, Pilehvar, Mohammad Taher and Raganato, Alessandro 2026. TruthTrap: A bilingual benchmark for evaluating factually correct yet misleading information in question answering. Presented at: EACL, Rabat, Morocco, 24-29 March 2026. Published in: Demberg, Vera, Inui, Kentaro and Marquez, Lluis eds. Findings of the Association for Computational Linguistics: EACL 2026. Association for Computational Linguistics, pp. 2966-2987. 10.18653/v1/2026.findings-eacl.155

Nobakhtian, Melika, Yaghoobzadeh, Yadollah and Pilehvar, Mohammad T. 2025. Evaluating cultural knowledge and reasoning in LLMs through Persian allusions. Presented at: EMNLP 2025, Suzhou, China, 4 - 9 November 2025. Published in: Christodoulopoulos, Christos, Chakraborty, Tanmoy, Rose, Carolyn and Peng, Violet eds. Findings of the Association for Computational Linguistics. Association for Computational Linguistics, pp. 25725-25737. 10.18653/v1/2025.findings-emnlp.1403

Zakizadeh, Mahdi and Pilehvar, Mohammad 2025. Blind men and the elephant: diverse perspectives on gender stereotypes in benchmark datasets. Presented at: 2025 Conference on Empirical Methods in Natural Language Processing, Suzhou, China, 4-9 November 2025. Published in: Christodoulopoulos, Christos, Chakraborty, Tanmoy, Rose, Carolyn and Peng, Violet eds. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, pp. 22838-22851. 10.18653/v1/2025.emnlp-main.1162
file

Marcuzzo, Matteo, Zangari, Alessandro, Albarelli, Andrea, Camacho-Collados, Jose ORCID: https://orcid.org/0000-0003-1618-7239 and Pilehvar, Mohammad Taher 2025. Morables: A benchmark for assessing abstract moral reasoning in LLMs with fables. Presented at: 2025 Conference on Empirical Methods in Natural Language Processing, Suzhou, China, 4-9 November 2025. Published in: Christodoulopoulos, Christos, Chakraborty, Tanmoy, Rose, Carolyn and Peng, Violet eds. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, pp. 27715-27739. 10.18653/v1/2025.emnlp-main.1411
file

Abbasi, Sina, Modarres, Mohammad Reza and Pilehvar, Mohammad Taher 2025. NormXLogit: The head-on-top never lies. Presented at: 2025 Conference on Empirical Methods in Natural Language Processing, Suzhou, China, 4-9 November 2025. Published in: Christodoulopoulos, Christos, Chakraborty, Tanmoy, Rose, Carolyn and Peng, Violet eds. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, pp. 34914-34935. 10.18653/v1/2025.emnlp-main.1769
file

Zangari, Alessandro, Marcuzzo, Matteo, Albarelli, Andrea, Pilehvar, Mohammad Taher and Camacho-Collados, Jose ORCID: https://orcid.org/0000-0003-1618-7239 2025. Pun unintended: LLMs and the illusion of humor understanding. Presented at: 2025 Conference on Empirical Methods in Natural Language Processing, Suzhou, China, 4-9 November 2025. Published in: Christodoulopoulos, Christos, Chakraborty, Tanmoy, Rose, Carolyn and Peng, Violet eds. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, pp. 27924-27959. 10.18653/v1/2025.emnlp-main.1419
file

Mehri, Faridoun, Baghshah, Mahdieh Soleymani and Pilehvar, Mohammad T. 2025. LibraGrad: Balancing gradient flow for universally better vision transformer attributions. Presented at: 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA, 10-17 June 2025. IEEE Computer Society Conference on Computer Vision and Pattern Recognition. IEEE, pp. 67-78. 10.1109/cvpr52734.2025.00016
file

Delbari, Zahra and Pilehvar, Mohammad Taher 2025. Beyond accuracy: revisiting out-of-distribution generalization in NLI models. Presented at: 29th Conference on Computational Natural LanguageLearning, Vienna, Austria, 31 July - 1 August 2025. Published in: Boleda, Gemma and Roth, Michael eds. Proceedings of the 29th Conference on Computational Natural Language Learning. Association for Computational Linguistics, pp. 557-570. 10.18653/v1/2025.conll-1.36
file

Moosavi Monazzah, Erfan, Rahimzadeh, Vahid, Yaghoobzadeh, Yadollah, Shakery, Azadeh and Pilehvar, Mohammad Taher 2025. PerCul: A story-driven cultural evaluation of LLMs in Persian. Presented at: Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), Albuquerque, New Mexico, USA, 29 April - 4 May 2025. Published in: Chiruzzo, Luis, Ritter, Alan and Wang, Lu eds. Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). Albuquerque, New Mexico: Association for Computational Linguistics, pp. 12670-12687. 10.18653/v1/2025.naacl-long.631
file

Mehri, Faridoun, Fayyaz, Mohsen, Baghshah, Mahdieh Soleymani and Pilehvar, Mohammad Taher 2024. SkipPLUS: Skip the first few layers to better explain vision transformers. Presented at: 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Seattle, WA, USA, 17-18 June 2024. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). pp. 204-215. 10.1109/cvprw63382.2024.00025
file

Delbari, Zahra, Moosavi, Nafise Sadat and Pilehvar, Mohammad Taher 2024. Spanning the spectrum of hatred detection: a Persian multi-label hate speech dataset with annotator rationales. Presented at: Thirty-Eighth AAAI Conference on Artificial Intelligence, Vancouver, Canada, 20-27 February 2024. Published in: Woolridge, M., Dy, J. and Natarajan, S. eds. Proceedings of the AAAI Conference on Artificial Intelligence. , vol.38 (16) Washington, DC, USA: Association for the Advancement of Artificial Intelligence, pp. 17889-17897. 10.1609/aaai.v38i16.29743
file

Modarressi, Ali, Fayyaz, Mohsen, Aghazadeh, Ehsan, Yaghoobzadeh, Yadollah and Pilehvar, Mohammad Taher 2023. DecompX: Explaining transformers decisions by propagating token decomposition. Presented at: 61st Annual Meeting of the Association for Computational Linguistics, Toronto, Canada, 9-14 July 2023. Published in: Rogers, Anna, Boyd-Graber, Jordan and Okazaki, Naoaki eds. Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, 2649–2664. 10.18653/v1/2023.acl-long.149
file

This list was generated on Sun Oct 4 05:01:00 2026 BST.