# LaTr: Layout-Aware Transformer for Scene-Text VQA

**Type:** Papers  
**Canonical URL:** https://scholariq.org/papers/latr-layout-aware-transformer-for-scene-text-vqa/

## Facts

| Field | Value |
| --- | --- |
| Author Names | Ali Furkan Biten,Ron Litman,Yusheng Xie,Srikar Appalaraju,R. Manmatha |
| Citations | 86 |
| DOI | 10.1109/cvpr52688.2022.01605 |
| Fields | Computer Science |
| Open Access | false |
| OA Status | closed |
| OpenAlex ID | https://openalex.org/W4312263373 |
| Type | conference-paper |
| Year | 2022 |

## Paper authors

- [Yusheng Xie](https://scholariq.org/researchers/yusheng-xie/)

## Paper primary topic

- [Multimodal Machine Learning Applications](https://scholariq.org/topics/multimodal-machine-learning-applications/)

## Paper topics

- [Multimodal Machine Learning Applications](https://scholariq.org/topics/multimodal-machine-learning-applications/)
- [Advanced Image and Video Retrieval Techniques](https://scholariq.org/topics/advanced-image-and-video-retrieval-techniques/)
- [Domain Adaptation and Few-Shot Learning](https://scholariq.org/topics/domain-adaptation-and-few-shot-learning/)

---
Source: ScholarIQ — public research metadata, principally OpenAlex. See https://scholariq.org/sources/ for provenance and https://scholariq.org/methodology/ for what these figures mean.
