LayoutLM for Visual Question Answering This is a fine tuned version of the multi modal LayoutLM model for the task of question answering on documents. It has been fine tuned using both the SQuAD2.0 and DocVQA datasets. Getting started with the model To run these examples, you must have PIL, pytesseract, and PyTorch installed in addition to transformers. NOTE : This model and pipeline was recently landed in transformers via PR 18407 and PR 18414, so you'll need to use a recent version of transformers, for example: About us This model was created by the team at Impira.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy