Dataset Card for Scientific Figures, Captions, and Context A novel vision language dataset of scientific figures taken directly from research papers. We scraped approximately ~150k papers, with about ~690k figures total. We extracted each figure's caption and label from the paper. In addition, we searched through each paper to find references of each figure and included the surrounding text as 'context' for this figure. All figures were taken from arXiv research papers. Figure 5: Comparisons between our multifidelity learning paradigm and single low fidelity (all GPT 3.5) annotation on four domain specific tasks given the same total 1000 annotation budget. Note that the samples for all GPT 3.5 are drawn based on the uncertainty score. Figure 3: Problem representation visualization by T SNE. Our model with A&D improves the problem rep resentation learning, which groups analogical problems close and separates non analogical problems. Usage The merged.json file is a mapping between the figure's filename as stored in the repository and its caption, label, and context. To use, you must extract the parts located under dataset/figures/ and keep the raw images in the same directory so that…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy