TextVQA
TextVQA is a dataset to benchmark visual reasoning based on text in images. TextVQA requires models to read and reason about text in images to answer questions about them.
CC-BY-4.0
10K – 100K
Visual Question Answering
English
No discussions yet. Start the first one.
New Discussion