# A Benchmark of PDF Information Extraction Tools Using a Multi-task and Multi-domain Evaluation Framework for Academic Documents

**Type:** Papers  
**Canonical URL:** https://scholariq.org/papers/a-benchmark-of-pdf-information-extraction-tools-using-a-multi-task-and-multi/

## Facts

| Field | Value |
| --- | --- |
| Author Names | Norman Meuschke,Apurva Jagdale,Timo Spinde,Jelena Mitrović,Béla Gipp |
| Citations | 23 |
| DOI | 10.1007/978-3-031-28032-0_31 |
| Fields | Computer Science |
| Open Access | true |
| OA Status | green |
| OA URL | https://arxiv.org/pdf/2303.09957 |
| OpenAlex ID | https://openalex.org/W4323780671 |
| Type | conference-paper |
| Year | 2023 |

## Paper authors

- [Timo Spinde](https://scholariq.org/researchers/timo-spinde/)

## Paper journal

- [Lecture notes in computer science](https://scholariq.org/journals/lecture-notes-in-computer-science/)

## Paper primary topic

- [Web Data Mining and Analysis](https://scholariq.org/topics/web-data-mining-and-analysis/)

## Paper topics

- [Web Data Mining and Analysis](https://scholariq.org/topics/web-data-mining-and-analysis/)
- [Handwritten Text Recognition Techniques](https://scholariq.org/topics/handwritten-text-recognition-techniques/)

---
Source: ScholarIQ — public research metadata, principally OpenAlex. See https://scholariq.org/sources/ for provenance and https://scholariq.org/methodology/ for what these figures mean.
