TextVQA

TextVQA

TextVQA is a dataset to benchmark visual reasoning based on text in images. TextVQA requires models to read and reason about text in images to answer questions about them.

CC-BY-4.0
10K – 100K
Visual Question Answering
English
by @AIOZAI
219

Last updated: 8 months ago


Sign in to see dataset files

orCreate an account