Oana Balalau

prof_pic.jpg

Contact

Inria Saclay, Bâtiment Alan Turing, 1 rue Honoré d’Estienne d’Orvet

Palaiseau, France

I am a researcher (ISFP) at Inria in the CEDAR team and part-time assistant professor at École Polytechnique. I had the priviledge to work, learn and make friends in many nice places: I was a postdoctoral researcher at Max Planck Institute for Informatics in the team of Gerhard Weikum, I did my PhD at Télécom Paris and my advisor was Mauro Sozio, and I also graduated from the University of Bucharest.


My research focuses on natural language processing methods that can improve the quality of public debate. By developing open-source algorithms, my team and I build tools designed to assist citizens and journalists in filtering, interpreting, and assimilating complex data. Ultimately, my work aims to make both public information and NLP tools more reliable and accountable.

Currently, my research spans the following main directions:

  • Fact-Checking & Misleading Argumentation: We build systems that assist journalists in matching public claims with reliable evidence. In collaboration with the Le vrai du faux team at Radio France, we built StatCheck, a pipeline that verifies statistical claims against sovereign databases like INSEE and Eurostat. Furthermore, using argumentation mining, we detect informal fallacies and propaganda in political discussions, demonstrating how deceptive rhetoric actively drives user engagement online.

  • AI Reliability (Hallucinations & Bias): To combat LLM hallucinations, we develop tools that evaluate factual integrity. This includes FactSpotter, which assesses the factual faithfulness of graph-to-text generation, and FDSpotter, which ensures complex discourse relations between atomic facts are preserved across generated texts. Furthermore, through our associated team, MediumAI, we are investigating new methodologies to identify and mitigate inherent biases in language models.

  • Climate NLP: My recent work identifies critical methodological gaps in the evaluation of climate NLP models and automated greenwashing detection. We are currently investigating how greenwashing manifests across diverse corporate communications and developing more robust NLP techniques to expose it.



I accept reviewing for open-access conferences and journals.

Alumni: Tom Calamai (2022 - 2026), co-advised with Fabian Suchanek.
Kun Zhang (2021 - 2025), co-advised with Ioana Manolescu. Now a postdoc at CNRS.