| 2026 | EACL | Do Psychometric Tests Work for Large Language Models? Evaluation of Tests on Sexism, Racism, and Morality. | Jana Jung, Marlene Lutz, Indira Sen, Markus Strohmaier |
| 2026 | EACL | Too Open for Opinion? Embracing Open-Endedness in Large Language Models for Social Simulation. | Bolei Ma, Yong Cao, Indira Sen, Anna-Carolina Haensch, Frauke Kreuter, Barbara Plank, Daniel Hershcovich |
| 2026 | EACL | Neural network embeddings recover value dimensions from psychometric survey items on par with human data. | Max Pellert, Clemens Lechner, Indira Sen, Markus Strohmaier |
| 2025 | ACL | Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness. | Shayan Alipour, Indira Sen, Mattia Samory, Tanushree Mitra |
| 2025 | ACL | Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models. | Mats Faulborn, Indira Sen, Max Pellert, Andreas Spitz, David Garca |
| 2025 | ACL | Missing the Margins: A Systematic Literature Review on the Demographic Representativeness of LLMs. | Indira Sen, Marlene Lutz, Elisa Rogers, David Garca, Markus Strohmaier |
| 2025 | EMNLP | The Prompt Makes the Person(a): A Systematic Evaluation of Sociodemographic Persona Prompting for Large Language Models. | Marlene Lutz, Indira Sen, Georg Ahnert, Elisa Rogers, Markus Strohmaier |
| 2025 | NAACL | Tell Me What You Know About Sexism: Expert-LLM Interaction Strategies and Co-Created Definitions for Zero-Shot Sexism Detection. | Myrthe Reuver, Indira Sen, Matteo Melis, Gabriella Lapesa |
| 2024 | ACL | An Open Multilingual System for Scoring Readability of Wikipedia. | Mykola Trokhymovych, Indira Sen, Martin Gerlach |
| 2023 | EMNLP | People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection. | Indira Sen, Dennis Assenmacher, Mattia Samory, Isabelle Augenstein, Wil M. P. van der Aalst, Claudia Wagner |
| 2022 | NAACL | Counterfactually Augmented Data and Unintended Bias: The Case of Sexism and Hate Speech Detection. | Indira Sen, Mattia Samory, Claudia Wagner, Isabelle Augenstein |
| 2021 | EMNLP | How Does Counterfactually Augmented Data Impact Models for Social Computing Constructs? | Indira Sen, Mattia Samory, Fabian Flck, Claudia Wagner, Isabelle Augenstein |
| 2021 | ICWSM | "Call me sexist, but..." : Revisiting Sexism Detection Using Psychological Scales and Adversarial Samples. | Mattia Samory, Indira Sen, Julian Kohne, Fabian Flck, Claudia Wagner |
| 2020 | CSCW | (Mis)Measuring People's Attitudes from Social Media. | Indira Sen |
| 2020 | EMNLP | On the Reliability and Validity of Detecting Approval of Political Actors in Tweets. | Indira Sen, Fabian Flck, Claudia Wagner |
| 2018 | ACL | Language Identification and Named Entity Recognition in Hinglish Code Mixed Tweets. | Kushagra Singh, Indira Sen, Ponnurangam Kumaraguru |