Publications



38 documents

Articles dans une revue

  • Clémentine Bleuze, Karën Fort, Vincent P Martin, Aurélie Névéol. Large Language Models for Mental Health Prediction: Scoping Review of Bias and Clinical Utility Documentation in 2019-2024. JMIR AI, 2026, 5, pp.e88082-e88082. ⟨10.2196/88082⟩. ⟨hal-05717650⟩
  • Vincent P. Martin, Charlotte Menard, Jean-Arthur Micoulaud-Franchi, Jacopo Giansanto Bodini, Adelaïde de Lastic, et al.. French validation of the short version of the ASA questionnaire: An exploration of the socio-technical imaginaries and cosmology of Socially Interactive Agent evaluation. L'Encéphale, In press. ⟨hal-05664209⟩
  • Clément Morand, Aurélie Névéol, Rosy Tsopra, Anne-Isabelle Tropeano, Sophie de Chambine, et al.. Prospectively Evaluating the Environmental Impacts of Digital Health Applications : A Case Study and Recommendations. Journal of the American Medical Informatics Association, 2026, ⟨10.1093/jamia/ocag091⟩. ⟨hal-05628404⟩
  • Fanny Ducel, Aurélie Névéol, Karën Fort. "You'll be a nurse, my son!" Automatically Assessing Gender Biases in Autoregressive Language Models in French and Italian. Language Resources and Evaluation, 2024, pp.1495-1523. ⟨10.1007/s10579-024-09780-6⟩. ⟨hal-04803403⟩

Communications dans un congrès

  • Ayoub Hammal, Pierre Zweigenbaum, Caio Corro. On the Rejection Criterion for Proxy-based Test-time Alignment. ACL 2026 - 64th Annual Meeting of the Association for Computational Linguistics, Jul 2026, San Diego, United States. pp.547-554, ⟨10.18653/v1/2026.acl-short.46⟩. ⟨hal-05689863⟩
  • Clémentine Bleuze, Karën Fort, Vincent P. Martin, Aurélie Névéol. Grands modèles de langue pour prédire la santé mentale : une revue exploratoire de la documentation des biais et de l'utilité clinique. TALN 2026 - 33e Conférence sur le Traitement Automatique des Langues Naturelles, Jun 2026, Nantes, France. ⟨hal-05670826⟩
  • Fanny Ducel, Lucie Digoin-Caparros, Ibrahim Al Kotob, Shayan Ahmed Shariff, Binesh Arakkal Remesh, et al.. Les benchmarks sont une source de biais des LLM : MMLU, CommonSenseQA et MGSM au microscope. TALN 2026 - 33e Conférence sur le Traitement Automatique des Langues Naturelles, Jun 2026, Nantes, France. ⟨hal-05618509⟩
  • Clémentine Bleuze, Bruno Guillaume, Aurélie Névéol, Karën Fort. Omniprésents et anthropomorphisés : analyse lexico-syntaxique des discours sur les LLM. TALN 2026 - 33e Conférence sur le Traitement Automatique des Langues Naturelles, Jun 2026, Nantes, France. ⟨hal-05670834⟩
  • Fanny Ducel, Sayaka Sato, Lucie Escasain, Aurelie Neveol, Karën Fort, et al.. L'écriture inclusive dans les invites des modèles de langue : une stratégie gagnant-gagnant. 21e Conférence en Recherche d'Information et Applications (CORIA) 19e Rencontres Jeunes Chercheurs en RI (RJCRI) 33e Conférence sur le Traitement Automatique des Langues Naturelles (TALN) 28e Rencontre des Étudiants Chercheurs en Informatique pour le Traitement Automatique des Langues (RÉCITAL), Jun 2026, Nantes, France. pp.334-349. ⟨hal-05708359⟩
  • Clément Morand, Aurélie Névéol, Anne-Laure Ligozat. The Rising Unsustainability of AI Graphics Cards Production. LIMITS 2026: 12th Workshop on Computing within Limits, Jun 2026, Online, France. ⟨hal-05666542⟩
  • Clémentine Bleuze, Fanny Ducel, Maxime Amblard, Karën Fort. COCOA: Creation and Exploratory Investigation of a Corpus of Claims from NLP Articles. LREC 2026 - 15th biennial Language Resources and Evaluation Conference, ELRA Language Resources Association, May 2026, Palma de Mallorca, Spain. ⟨10.63317/38hiuxwcq4bc⟩. ⟨hal-05547842⟩
  • Fanny Ducel, Aurélie Névéol, Vidit Khazanchi, Loïc Leclere, Arthur Pedrini, et al.. Code-switching as a Bias Indicator in LLMs: "The consequences are not the same para nosotros". LREC 2026 - 15th biennial Language Resources and Evaluation Conference, May 2026, Palma De Mallorca, Spain. ⟨10.63317/2mq6kqjk9bng⟩. ⟨hal-05529786⟩
  • Hugo Thomas, Caio Corro, Guillaume Gravier, Pascale Sébillot. A Study on Building Efficient Zero-Shot Relation Extraction Models. The Fifteenth Language Resources and Evaluation Conference (LREC 2026), May 2026, Palma, France. pp.6056-6067, ⟨10.63317/5ndtgoxeph5j⟩. ⟨hal-05628350⟩
  • Ayoub Hammal, Pierre Zweigenbaum, Caio Corro. KAD: A Framework for Proxy-based Test-time Alignment with Knapsack Approximation Deferral. EACL 2026 - 19th Conference of the European Chapter of the Association for Computational Linguistics, Mar 2026, Rabat, Morocco. pp.3854-3872, ⟨10.18653/v1/2026.eacl-long.179⟩. ⟨hal-05571208⟩
  • Clémentine Bleuze, Karën Fort, Vincent P. Martin, Aurélie Névéol. Grands modèles de langue pour la détection de pathologies psychiatriques : promesses, réalité, et enjeux. Journée d'étude "LLM@hopital", ATALA, Mar 2026, Paris, France. ⟨hal-05532823⟩
  • Karën Fort. Les enjeux éthiques de l’IA vus depuis le traitement automatique des langues. Journée de lancement du projet Insight, Dec 2025, Nancy, France. ⟨hal-05397834⟩
  • Alberto Muñoz-Ortiz, David Vilares, Caio Corro, Carlos Gómez-Rodríguez. Nested Named Entity Recognition as Single-Pass Sequence Labeling. EMNLP 2025 - Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, Nov 2025, Suzhou, China. pp.9993-10002, ⟨10.18653/v1/2025.findings-emnlp.530⟩. ⟨hal-05457961⟩
  • Karën Fort. Large Language Models: the challenge of evaluation. CLiC-it 2025 - Eleventh Conference on Computational Linguistics, Sep 2025, Cagliari, Italy. ⟨hal-05397849⟩
  • Caio Corro, Mathieu Lacroix, Joseph Le Roux. Bregman Conditional Random Fields: Sequence Labeling with Parallelizable Inference Algorithms. ACL 2025 - 63rd Annual Meeting of the Association for Computational Linguistics, Association for Computational Linguistics, Jul 2025, Vienne, Austria. pp.29557-29574, ⟨10.18653/v1/2025.acl-long.1430⟩. ⟨hal-05360479⟩
  • Fanny Ducel, Karën Fort, Aurélie Névéol. La linguistique appliquée pour une IA plus éthique. NéALA 2025 - Colloque sur Naturel et Artificiel en Linguistique Appliquée : une époque de paradoxes, Jul 2025, Nancy, France. ⟨hal-05457534⟩
  • Clémentine Bleuze, Fanny Ducel, Maxime Amblard, Karën Fort. « De nos jours, ce sont les résultats qui comptent » : création et étude diachronique d'un corpus de revendications issues d'articles de TAL. TALN 2025 - 32ème Conférence sur le Traitement Automatique des Langues Naturelles, Jul 2025, Marseille, France. ⟨hal-05282966⟩
  • Fanny Ducel, Nicolas Hiebel, Olivier Ferret, Karën Fort, Aurélie Névéol. « Les femmes ne font pas de crise cardiaque ! » Étude des biais de genre dans les cas cliniques synthétiques en français. TALN 2025 - 32ème Conférence sur le Traitement Automatique des Langues Naturelles, Jul 2025, Marseille, France. pp.1. ⟨hal-05282965⟩
  • Fanny Ducel, Jeffrey André, Aurélie Névéol, Karën Fort. Introducing MascuLead: the First Gender Bias Leaderboard. EALM 2025 - Ethic and Alignment of (Large) Language Models, Jun 2025, Marseille, France. pp.12-19. ⟨hal-05282981⟩
  • Fanny Ducel, Nicolas Hiebel, Olivier Ferret, Karën Fort, Aurélie Névéol. "Women do not have heart attacks!" Gender Biases in Automatically Generated Clinical Cases in French. NAACL 2025 - Annual Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics, Apr 2025, Albuquerque, United States. pp.7145-7159, ⟨10.18653/v1/2025.findings-naacl.398⟩. ⟨hal-04938811⟩
  • Ayoub Hammal, Benno Uthayasooriyar, Caio Corro. Few-Shot Domain Adaptation for Named-Entity Recognition via Joint Constrained k-Means and Subspace Selection. COLING 2025 - 31st International Conference on Computational Linguistics, Jan 2025, Abu Dhabi, United Arab Emirates. pp.9902-9916, ⟨10.48550/arxiv.2412.00426⟩. ⟨hal-04877776⟩
  • Elie Dina, Rania Ayachi Kibech, Miguel Couceiro. EDAR: A pipeline for Emotion and Dialogue Act Recognition. The 31st International Conference on Computational Linguistics (COLING 2025), Jan 2025, Abu Dhabi, United Arab Emirates. ⟨hal-04878084⟩
  • Fanny Ducel, Aurélie Névéol, Karën Fort. Desiderata for Actionable Bias Research. New Perspectives on Bias and Discrimination in Language Technology, Nov 2024, Amsterdam, Netherlands. ⟨hal-04755691⟩
  • Mohammed Rameez Rameez Qureshi, Naïm Es-Sebbani, Luis Galárraga, Yvette Graham, Miguel Couceiro, et al.. REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning. ECAI 2024 - 27th European Conference on Artificial Intelligence, Oct 2024, Santiago de Compostela, Spain. pp.4027-4034, ⟨10.3233/FAIA240970⟩. ⟨hal-04872144⟩
  • Ajinkya Kulkarni, Atharva Kulkarni, Isabel Trancoso, Miguel Couceiro. Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems. Interspeech 2024, Sep 2024, Kos / Greece, Greece. ⟨hal-04610235⟩
  • Fanny Ducel, Aurélie Névéol, Karën Fort. Évaluation automatique des biais de genre dans des modèles de langue auto-régressifs. 31ème Conférence sur le Traitement Automatique des Langues Naturelles (TALN 2024), Jul 2024, Toulouse, France. pp.64-84. ⟨hal-04621134⟩
  • Clément Morand, Anne-Laure Ligozat, Aurélie Névéol. MLCA: a tool for Machine Learning Life Cycle Assessment. 2024 10th International Conference on ICT for Sustainability (ICT4S), Jun 2024, Stockholm, Sweden. pp.227-238, ⟨10.1109/ICT4S64576.2024.00031⟩. ⟨hal-04643414⟩
  • Guilherme Dean Pelegrina, Miguel Couceiro, Leonardo Tomazeli Duarte. A preprocessing Shapley value-based approach to detect relevant and disparity prone features in machine learning. ACM Conference on Fairness, Accountability, and Transparency (ACM FAccT 2024), Fabro Steibel, Meg Young, Ricardo Baeza-Yates, Jun 2024, Rio de Janeiro (BR), Brazil. pp.279-289, ⟨10.1145/3630106.3658905⟩. ⟨hal-04568343⟩
  • Ajinkya Kulkarni, Anna Tokareva, Mohammed Rameez Qureshi, Miguel Couceiro. The Balancing Act: Unmasking and Alleviating ASR Biases in Portuguese. EACL 2024 LT-EDI WorkShop, Mar 2024, St. Julians, Malta. ⟨hal-04436147⟩

Proceedings/Recueil des communications

  • Romain Michelucci, Bastien Rance, Adrien Coulet. Actes de la journée d’étude sur l’utilisation des LLM à l’hôpital. Journée llm@hopital.fr, Mar 2026, Paris, France. 2026, Journée d'étude ATALA. ⟨hal-05556971⟩

Pré-publications, Documents de travail

  • Ayoub Hammal, Pierre Zweigenbaum, Caio Corro. Suffix-Constrained Greedy Search Algorithms for Causal Language Models. 2026. ⟨hal-05734800⟩
  • Lucas Ondel Yang, Tina Raissi, Martin Kocour, Pablo Riera, Caio Corro. Fast and General Automatic Differentiation for Finite-State Methods. 2026. ⟨hal-05500843⟩
  • Clément Morand, Anne-Laure Ligozat, Aurélie Névéol. The Environmental Impacts of Machine Learning Training Keep Rising Evidencing Rebound Effect. 2025. ⟨hal-04839926v5⟩
  • Clément Morand, Anne-Laure Ligozat, Aurélie Névéol. Characterizing Goals and Impacts of Digitalization: The Case of Promises in French Healthcare Policies. 2025. ⟨hal-05066176⟩