Abstract
This paper presents our system for document-level semantic textual similarity (STS) evaluation at SemEval-2022 Task 8: 'Multilingual News Article Similarity". The semantic information used is obtained by using different semantic models ranging from the extraction of key terms and named entities to the document classification and obtaining similarity from automatic summarization of documents. All these semantic information's are then used as features to feed a supervised system in order to evaluate the degree of similarity of a pair of documents. We obtained a Pearson correlation score of 0.706 compared to the best score of 0.818 from teams that participated in this task. Our source code can be found at GitHub.
Cite
CITATION STYLE
Dufour, S., Kandi, M. M., Boutamine, K., Gosset, C., Billami, M. B., Bortolaso, C., & Miloudi, Y. (2022). BL.Research at SemEval-2022 Task 8: Using various Semantic Information to evaluate document-level Semantic Textual Similarity. In SemEval 2022 - 16th International Workshop on Semantic Evaluation, Proceedings of the Workshop (pp. 1221–1228). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.semeval-1.173
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.