Pourquoi publions-nous ?
Les publications scientifiques issues du laboratoire de recherche témoignent de l'engagement de F.initiatives dans une démarche de R&D rigoureuse, documentée et ouverte sur l'écosystème académique. Elles contribuent à consolider les fondements scientifiques des outils développés en interne et à valoriser les travaux menés autour du traitement automatique du langage naturel, de l'extraction de connaissances et de l'analyse de corpus complexes.
Cette activité de publication répond à plusieurs objectifs. Elle permet d'abord de confronter les approches développées par F.initiatives à l'état de l'art scientifique international, en soumettant les méthodes et résultats à l'évaluation de communautés spécialisées. Elle contribue ensuite à renforcer la crédibilité scientifique des solutions mises à disposition des équipes et des clients. Elle favorise enfin la capitalisation des connaissances produites par le laboratoire et leur diffusion auprès de communautés de recherche pertinentes.
Les travaux publiés portent notamment sur des problématiques liées à :
- La reconnaissance d'entités nommées
- L'extraction de concepts scientifiques
- La structuration de taxonomies
- La recherche d'information
- La détection de contenus générés par intelligence artificielle
Ces recherches alimentent directement les outils opérationnels développés par F.initiatives, en particulier ceux destinés à l'exploration de la littérature scientifique et à l'amélioration de la documentation R&D.
Liste des publications scientifiques
2025
Multi-view Text Data Stream Clustering
Ben-Fares, M., Grozavu, N., Rastin, P., & Holat, P.
Applied Intelligence and Informatics: 4th International Conference, AII 2024, London, UK, 18–20 décembre 2024, Revised Selected Papers, p. 156 — Springer Nature (octobre 2025)
Revue scientifiqueOnline Clustering of Massive Text Data Streams
Ben-Fares, M., Rastin, P., Grozavu, N., & Holat, P.
SN Computer Science, 6(5), 413 (2025)
2024
An autoregressive text-to-graph framework for joint entity and relation extraction
Zaratiana, U., Tomeh, N., Holat, P., & Charnois, T.
Proceedings of the AAAI Conference on Artificial Intelligence, vol. 38, n°17, p. 19477-19487 (mars 2024)
Conférence / workshopFi group at SemEval-2024 task 8: A syntactically motivated architecture for multilingual machine-generated text detection
Ben-Fares, M., Zaratiana, U., Hernandez, S., & Holat, P.
Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024), p. 1166-1171 (juin 2024)
Conférence / workshopGLiNER: Generalist model for named entity recognition using bidirectional transformer
Zaratiana, U., Tomeh, N., Holat, P., & Charnois, T.
Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), p. 5364-5376 (juin 2024)
Conférence / workshopMulti-view Text Data Stream Clustering via Late Fusion
Ben-Fares, M., Grozavu, N., Rastin, P., & Holat, P.
International Conference on Applied Intelligence and Informatics, p. 156-172 — Cham : Springer Nature Switzerland (décembre 2024)
Prépublication arXivEnrico: Enriched representation and globally constrained inference for entity and relation extraction
Zaratiana, U., Tomeh, N., Dauxais, Y., Holat, P., & Charnois, T.
arXiv preprint arXiv:2404.12493 (2024)
Prépublication arXivGraphER: A Structure-aware Text-to-Graph Model for Entity and Relation Extraction
Zaratiana, U., Tomeh, N., El Khbir, N., Holat, P., & Charnois, T.
arXiv preprint arXiv:2404.12491 (2024)
2023
Filtered Semi-Markov CRF
Zaratiana, U., Tomeh, N., El Khbir, N., Holat, P., & Charnois, T.
Findings of the Association for Computational Linguistics: EMNLP 2023, p. 222-235 (décembre 2023)
Conférence / workshopSélection globale de segments pour la reconnaissance d'entités nommées
Zaratiana, U., El Khbir, N., Holat, P., Tomeh, N., & Charnois, T.
Actes de CORIA-TALN 2023 — 30e Conférence sur le Traitement Automatique des Langues Naturelles (TALN), volume 4, p. 169-177 (2023)
2022
Towards automation of topic taxonomy construction
Dauxais, Y., Zaratiana, U., Laneuville, M., Hernandez, S. D., Holat, P., & Grosman, C.
International Symposium on Intelligent Data Analysis, p. 26-38 — Cham : Springer International Publishing (avril 2022)
Conférence / workshopGNNer: Reducing overlapping in span-based NER using graph neural networks
Zaratiana, U., Tomeh, N., Holat, P., & Charnois, T.
Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics: Student Research Workshop, p. 97-103 (mai 2022)
Conférence / workshopHigh dimensional data stream clustering using topological representation learning
Ben-Fares, M., Grozavu, N., Rastin, P., & Holat, P.
2022 IEEE Symposium Series on Computational Intelligence (SSCI), p. 1415-1422 — IEEE (décembre 2022)
Conférence / workshopNamed entity recognition as structured span prediction
Zaratiana, U., Tomeh, N., Holat, P., & Charnois, T.
Proceedings of the Workshop on Unimodal and Multimodal Induction of Linguistic Structures (UM-IoS), p. 1-10 (décembre 2022)
Conférence / workshopGlobal span selection for named entity recognition
Zaratiana, U., El Khbir, N., Holat, P., Tomeh, N., & Charnois, T.
Proceedings of the Workshop on Unimodal and Multimodal Induction of Linguistic Structures (UM-IoS), p. 11-17 (décembre 2022)
Conférence / workshopMethod for Joining Information and Adapting Content from Gamified Systems and Serious Games in Organizations
Guinebert, M., Fabiani, J., Cherdieu, M., Holat, P., & Grosman, C.
CSEDU (2), p. 338-350 (2022)
Prépublication arXivHierarchical transformer model for scientific named entity recognition
Zaratiana, U., Holat, P., Tomeh, N., & Charnois, T.
arXiv preprint arXiv:2203.14710 (2022)
Prépublication arXivDyREx: Dynamic query representation for extractive question answering
Zaratiana, U., Khbir, N. E., Núñez, D., Holat, P., Tomeh, N., & Charnois, T.
arXiv preprint arXiv:2210.15048 (2022)
Autres travaux
Extractive Question Answering with Soft Structural biases
Zaratiana, U., Tomeh, N., Holat, P., & Charnois, T.
DyReF: Extractive Question Answering with Dynamic Query Representation for Free
Zaratiana, U., El Khbir, N., Holat, P., Tomeh, N., & Charnois, T.
Pour aller plus loin
Laboratoire de recherche
Le socle scientifique et technologique à l'origine de ces travaux, entre recherche appliquée et gouvernance responsable de l'IA.
Sécurité et certifications
Notre système de management intégré et nos garanties en matière de sécurité de l'information.
NeoPhi
La plateforme d'exploration et de structuration de la littérature scientifique issue de ces recherches.