Struggling with PDF scanned nested tabels to html/md/json conversion

작성자

카테고리:

← 피드로
r/programming · /u/Kakarot_DB · 2026-06-18 개발(SW)

For the past few days, I've been trying to parse a PDF (scanned and text based) which has the same contents. PDF has nested tables Tables start at one page and end at another Currently I have been using (docling)[https://docling-project.github.io] to help me out with…

원문 보기 ↗

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다