LATTE: Improving Latex Recognition for Tables and Formulae with Iterative Refinement Paper • 2409.14201 • Published Sep 21, 2024 • 3
PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling Paper • 2410.05970 • Published Oct 8, 2024 • 2
Small Language Models can Outperform Humans in Short Creative Writing: A Study Comparing SLMs with Humans and LLMs Paper • 2409.11547 • Published Sep 17, 2024 • 1
PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering Paper • 2404.12720 • Published Apr 19, 2024 • 2
PdfTable: A Unified Toolkit for Deep Learning-Based Table Extraction Paper • 2409.05125 • Published Sep 8, 2024 • 1