Document Models (Pretrained) Various pretrained models for analyzing documents. These need to be fine-tuned for a task naver-clova-ix/donut-base Image-to-Text • Updated Aug 13, 2022 • 25.7k • 148 google/pix2struct-base Image-to-Text • Updated Dec 24, 2023 • 8.85k • 60 google/pix2struct-large Image-to-Text • Updated Sep 6, 2023 • 3.61k • 27 microsoft/layoutlmv3-base Updated 24 days ago • 6.41M • 268
Document Models (Fine-tuned) naver-clova-ix/donut-base-finetuned-cord-v2 Image-to-Text • Updated Aug 13, 2022 • 20.4k • 65 google/pix2struct-docvqa-base Visual Question Answering • Updated Dec 24, 2023 • 206k • 33 google/pix2struct-docvqa-large Visual Question Answering • Updated May 19, 2023 • 1.74k • 29 google/pix2struct-screen2words-base Visual Question Answering • Updated May 19, 2023 • 144 • 20
google/pix2struct-screen2words-base Visual Question Answering • Updated May 19, 2023 • 144 • 20