Update cookies preferences
 logo Idiap Research Institute        
Project AlignAI
Name: AlignAI

Publications of AlignAI sorted by recency
Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets, Lei Hsiung, Tianyu Pang, Yung-Chen Tang, Linyue Song, Tsung-Yi Ho, Pin-Yu Chen and Yaoqing Yang, in: Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 2026
[URL]
Politics of Questions in News: A Mixed-Methods Study of Interrogative Stances as Markers of Voice and Power, Victor Bros, Matilde Barbini, Patrick Gerard and Daniel Gatica-Perez, in: Vol. 20 (2026): Proceedings of the Twentieth International AAAI Conference on Web and Social Media, 2026
attachment