France

Florian Boudin

Expert
@boudinfl

Associate professor at the University of Nantes, working on Natural Language Processing and Information Retrieval.

pke. Python Keyphrase Extraction module

1.6k

ake-datasets. Large, curated set of benchmark datasets for evaluating automatic keyphrase extraction algorithms.

148

takahe. takahe is a multi-sentence compression module

54

sume. Sume is an implementation of the concept-based ILP model for summarization.

37

kea. A tokenizer for French

14

taln-archives. TALN Archives is a digital archive of French research articles in Natural Language Processing

13

centrality_measures_ijcnlp13. Centrality Measures for Graph-Based Keyphrase Extraction

13

ir-using-kg. Keyphrase Generation for Scientific Document Retrieval

11

acm-cr. ACM-CR: A Manually Annotated Test Collection for Citation Recommendation

10

hulth-2003-pre. Preprocessed Inspec keyphrase extraction benchmark dataset

8

semeval-2010-pre. Preprocessed SemEval-2010 benchmark dataset for keyphrase extraction

7

duc-2001-pre. Preprocessed DUC 2001 keyphrase extraction benchmark dataset

6

krapivin-2009-pre. Preprocessed Krapivin keyphrase extraction benchmark dataset

5

redefining-absent-keyphrases. Code and dataset for the paper "Redefining Absent Keyphrases and their Effect on Retrieval Effectiveness"

5

marujo-2012-pre. Preprocessed Marujo keyphrase extraction benchmark dataset

4

lina-msc. LINA-msc is a dataset for evaluating Multi-sentence Compression in French.

3

kepy. kepy is a keyphrase extraction module in Python

2

CLIREC. CLinical Information Retrieval Evaluation Collection

2

cross-language_IR. Un cours de deux heures sur la recherche d'information cross-lingue

2

silk. silk: Unsupervised Domain Adaptation for Keyphrase Generation using Citation Contexts

2

pke-benchmarking. Jupyter Notebook

1
21
Apply