Fetching the paper…
Reading the bibliography…
We introduce Docling, an easy-to-use, self-contained, MIT-licensed, open-source toolkit for document conversion, that can parse several types of popular document formats into a unified, richly structured representation.
HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Wolf, T.; Debut, L.; Sanh, V.; Chaumond, J.; Delangue, C.; Moi, A.; Cistac, P.; Rault, T.; Louf, R.; Funtowicz, M.; Davison, J.; Shleifer, S.; von Platen, P.; Ma, C.; Jernite, Y.; Plu, J.; Xu, C.; Scao, T. L.; Gugger, S.; Drame, M.; Lhoest, Q.; and Rush, A. M. 2020 · 1910
Earlier work this paper cites.
Image-based table recognition: data, model, and evaluation
Zhong, X. 2020 · 1911
Earlier work this paper cites.
openpyxl: A Python library to read/write Excel 2010 xlsx/xlsm files
Eric Gazoni, C. C. 2010–2024 · 2010
Earlier work this paper cites.
python-docx: Create and update Microsoft Word .docx files with Python
Canny, S.; and contributors. 2013–2024a · 2013
Earlier work this paper cites.
python-pptx: Python library for creating and updating PowerPoint (.pptx) files
Canny, S.; and contributors. 2013–2024b · 2013
Earlier work this paper cites.
Robust PDF Document Conversion using Recurrent Neural Networks
Livathinos, N.; Berrospi, C.; Lysak, M.; Kuropiatnyk, V.; Nassar, A.; Carvalho, A.; Dolfi, M.; Auer, C.; Dinkla, K.; and Staar, P. 2021 · 2021
Earlier work this paper cites.
Delivering Document Conversion as a Cloud Service with High Throughput and Responsiveness
Auer, C.; Dolfi, M.; Carvalho, A.; Ramis, C. B.; and Staar, P. W. 2022 · 2022
Earlier work this paper cites.
LangChain
Chase, H. 2022 · 2022
Earlier work this paper cites.
LlamaIndex
Liu, J. 2022 · 2022
Earlier work this paper cites.
Tableformer: Table structure understanding with transformers
Nassar, A.; Livathinos, N.; Lysak, M.; and Staar, P. 2022 · 2022
Earlier work this paper cites.
DocLayNet: a large human-annotated dataset for document-layout segmentation
Pfitzmann, B.; Auer, C.; Dolfi, M.; Nassar, A. S.; and Staar, P. 2022 · 2022
Cited alongside, same era.
Optimized Table Tokenization for Table Structure Recognition
Lysak, M.; Nassar, A.; Livathinos, N.; Auer, C.; and Staar, P. 2023 · 2023
Cited alongside, same era.
CCpdf: Building a High Quality Corpus for Visually Rich Documents from Web Crawl Data
Turski, M.; Stanisławek, T.; Kaczmarek, K.; Dyda, P.; and Graliński, F. 2023 · 2023
Cited alongside, same era.
DETRs Beat YOLOs on Real-time Object Detection
Zhao, Y.; Lv, W.; Xu, S.; Wei, J.; Wang, G.; Dang, Q.; Liu, Y.; and Chen, J. 2023 · 2023
Cited alongside, same era.
EasyOCR: Ready-to-use OCR with 80+ supported languages
2024 · 2024
Cited alongside, same era.
OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations
Ouyang, L.; Qu, Y.; Zhou, H.; Zhu, J.; Zhang, R.; Lin, Q.; Wang, B.; Zhao, Z.; Jiang, M.; Zhao, X.; Shi, J.; Wu, F.; Chu, P.; Liu, M.; Li, Z.; Xu, C.; Zhang, B.; Shi, B.; Tu, Z.; and He, C. 2024 · 2024
Later among the works it cites.
Marker: Convert PDF to Markdown Quickly with High Accuracy
Paruchuri, V. 2024 · 2024
Later among the works it cites.
pypdf: A Pure-Python PDF Library
pypdf Maintainers. 2024 · 2024
Later among the works it cites.
PyPDFium2: Python bindings for PDFium
PyPDFium Team. 2024 · 2024
Later among the works it cites.
Beautiful Soup: A Python library for parsing HTML and XML
Richardson, L. 2004–2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
PyTorch 2: Faster Machine Learning Through Dynamic Python Bytecode Transformation and Graph Compilation
Ansel, J.; Yang, E.; He, H.; et al. 2024 · 2024
Cited alongside, same era.
QPDF: A Content-Preserving PDF Document Transformer
Berkenbilt, J. 2024 · 2024
Cited alongside, same era.
Bee Agent Framework
IBM Research. 2024 · 2024
Cited alongside, same era.
Marko: A markdown parser with high extensibility
Ming, F. 2019–2024 · 2024
Cited alongside, same era.
Sudalairaj, S.; Bhandwaldar, A.; Pareja, A.; Xu, K.; Cox, D. D.; and Srivastava, A. 2024 · 2024
Later among the works it cites.
Unstructured.io: Open-Source Pre-Processing Tools for Unstructured Data
Unstructured.io Team. 2024 · 2024
Later among the works it cites.
MinerU: An Open-Source Solution for Precise Document Content Extraction
Wang, B.; Xu, C.; Zhao, X.; Ouyang, L.; Wu, F.; Zhao, Z.; Xu, R.; Liu, K.; Qu, Y.; Shang, F.; Zhang, B.; Wei, L.; Sui, Z.; Li, W.; Shi, B.; Qiao, Y.; Lin, D.; and He, C. 2024 · 2024
Later among the works it cites.
Data-Prep-Kit: getting your data ready for LLM application development
Wood, D.; Lublinsky, B.; Roytman, A.; Singh, S.; Adam, C.; Adebayo, A.; An, S.; Chang, Y. C.; Dang, X.-H.; Desai, N.; Dolfi, M.; Emami-Gohari, H.; Eres, R.; Goto, T.; Joshi, D.; Koyfman, Y.; Nassar, M.; Patel, H.; Selvam, P.; Shah, Y.; Surendran, S.; Tsuzuku, D.; Zerfos, P.; and Daijavad, S. 2024 · 2024
Later among the works it cites.