publications
publications by categories in reversed chronological order.
2026
- A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and ModelsJun 2026Preprint AI text detection
- Leveraging Comparable Toxicity Lexicons in Prompt Instructions for Multilingual Text DetoxificationIn Proceedings of the 19th Workshop on Building and Using Comparable Corpora (BUCC), May 2026toxicity
- Misaligned by Reward: Socially Undesirable Preferences in LLMsMay 2026Preprint alignmentreward modelssocial alignmentbenchmark
2025
- “I understand your perspective”: LLM Persuasion through the Lens of Communicative Action TheoryIn Findings of the Association for Computational Linguistics: ACL 2025, Jul 2025persuasionsycophancy
- AI Argues Differently: Distinct Argumentative and Linguistic Patterns of LLMs in Persuasive ContextsIn Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, Nov 2025persuasionlinguistic convergenceargumentation
2024
- Please note that I’m just an AI: Analysis of Behavior Patterns of LLMs in (Non-)offensive Speech IdentificationIn Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, Nov 2024offensive languageerroneous generationstereotypes
2023
- HNC: Leveraging Hard Negative Captions towards Models with Fine-Grained Visual-Linguistic Comprehension CapabilitiesIn Proceedings of the 27th Conference on Computational Natural Language Learning (CoNLL), Dec 2023vision & languageimage–text matchingmisalignmentcounterfactuals