Papers
arxiv:2607.26848

ICDAR 2026 Competition on Information Extraction from Atomic Layer Deposition/Etching (ALD/E) Scientific Figures

Published on Jul 29
Β· Submitted by
Jennifer D'Souza
on Aug 4
Authors:
,
,

Abstract

Scientific figure comprehension and reasoning using multimodal AI requires integrating visual perception with domain-specific reasoning to extract meaningful knowledge, often not presented in the text of a research publication. The Sci-ImageMiner benchmark dataset, accompanied by a community-driven competition, raises the bar over prior scientific competitions by curating a comprehensive, expert-annotated dataset across four end-to-end complementary tasks. The competition attracted 68 active participants and 1,263 public/private submissions from 9th January 2026 to 8th April 2026. Our results show that state-of-the-art multimodal models perform well on classification and summarization tasks but struggle with data extraction and scientific reasoning, particularly in visual question-answering. These findings reveal key limitations and highlight challenges and opportunities for improving domain-aware multimodal AI systems. Overall, the Sci-ImageMiner benchmark and competition establish a rigorous platform for advancing research in scientific figure comprehension and reasoning and demonstrate the potential of state-of-the-art approaches for a challenging and complex research area.

Community

Paper submitter

πŸ“Š ALD/E-ImageMiner is an expert-annotated multimodal benchmark for understanding scientific figures from atomic layer deposition and atomic layer etching (ALD/E), covering both experimental and simulation studies. It contains 1,951 figures from 205 research papers.

Evaluate your models and submit results to the four CodaBench leaderboards:

πŸ€— Access the complete ALD/E-ImageMiner benchmark dataset on Hugging Face

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2607.26848
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 0

No model linking this paper

Cite arxiv.org/abs/2607.26848 in a model README.md to link it from this page.

Datasets citing this paper 2

Spaces citing this paper 0

No Space linking this paper

Cite arxiv.org/abs/2607.26848 in a Space README.md to link it from this page.

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.