resources

  • MUCH: A Multilingual Claim Hallucination Benchmark
  • Geo²CLEF collection: GeoCLEF collection annotated automatically with geographical information
  • ACL-RelAcS corpus: corpus designed for semantic RELation ACquiSition (extraction and classification) in the scientific domain
  • SemEval 2018 Task 7 dataset: data used in the organization of SemEval 2018 Task 7: Semantic Relation Extraction and Classification in Scientific Papers
  • CS-KG: Computer Science Knowledge Graph, a large-scale automatically generated knowledge graph describing 67M statements from 14.5M articles about 24M entities (e.g., tasks, methods, materials, metrics) linked by 219 semantic relations
  • GloVe word vectors for the Genoese language, based on the Ligurian Wikipedia (too few data, no guarantees for good results… but as we say in Genoa, “l’é mêgio o pöco che o ninte”)