MELINDA: A Multimodal Dataset for Biomedical Experiment Method Classification

N/ACitations
Citations of this article
33Readers
Mendeley users who have this article in their library.
Get full text

Abstract

We introduce a new dataset, MELINDA, for Multimodal biomEdicaL experImeNt methoD clAssification. The dataset is collected in a fully automated distant supervision manner, where the labels are obtained from an existing curated database, and the actual contents are extracted from papers associated with each of the records in the database. We benchmark various state-of-the-art NLP and computer vision models, including unimodal models which only take either caption texts or images as inputs, and multimodal models. Extensive experiments and analysis show that multimodal models, despite outperforming unimodal ones, still need improvements especially on a less-supervised way of grounding visual concepts with languages, and better transferability to low resource domains. We release our dataset and the benchmarks to facilitate future research in multimodal learning, especially to motivate targeted improvements for applications in scientific domains.

Cite

CITATION STYLE

APA

Wu, T. L., Singh, S., Paul, S., Burns, G., & Peng, N. (2021). MELINDA: A Multimodal Dataset for Biomedical Experiment Method Classification. In 35th AAAI Conference on Artificial Intelligence, AAAI 2021 (Vol. 16, pp. 14076–14084). Association for the Advancement of Artificial Intelligence. https://doi.org/10.1609/aaai.v35i16.17657

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free