Linked List
login
sign up
What's so hard about PDF text extraction?
filingdb.com
· first added by
@afreshcup
saved by
1 person
discussions
·
4
What's so hard about PDF text extraction?
406 pts · 235 comments · Sep 2020
406 pts · 235 comments · Sep 2020
▲
0
0 people saved or upvoted this
What's so hard about PDF text extraction?
11 pts · 15 comments · Oct 2022
11 pts · 15 comments · Oct 2022
▲
0
0 people saved or upvoted this
What's so hard about PDF text extraction?
733 pts · 342 comments · Mar 2020
733 pts · 342 comments · Mar 2020
▲
0
0 people saved or upvoted this
see all 4 →
What's so hard about PDF text extraction?
3 pts · 1 comment · Jan 2026
3 pts · 1 comment · Jan 2026
▲
0
0 people saved or upvoted this
from the discussion
·
5
GitHub - camelot-dev/camelot: A Python library to extract tabular data from PDFs
github.com
github.com
▲
0
0 people saved or upvoted this
GitHub - coolwanglu/pdf2htmlEX: Convert PDF to HTML without losing text or format.
github.com
github.com
▲
0
0 people saved or upvoted this
The sad state of PDF-Accessibility of LaTex Documents
umij.wordpress.com
umij.wordpress.com
▲
0
0 people saved or upvoted this
Peter Selinger: Creating high-quality PDF/A documents using LaTeX
mathstat.dal.ca
mathstat.dal.ca
▲
0
0 people saved or upvoted this
Why GOV.UK content should be published in HTML and not PDF
gds.blog.gov.uk
gds.blog.gov.uk
▲
0
0 people saved or upvoted this
related HN threads
·
3
The sad state of PDF-Accessibility of LaTex Documents (2016)
79 pts · 74 comments · Sep 2020
79 pts · 74 comments · Sep 2020
▲
0
0 people saved or upvoted this
PDF processing and analysis with open-source tools (2021)
186 pts · 40 comments · Oct 2022
186 pts · 40 comments · Oct 2022
▲
0
0 people saved or upvoted this
Examples to compare OCR services: Amazon vs. Google vs. Microsoft
262 pts · 65 comments · Jul 2019
262 pts · 65 comments · Jul 2019
▲
0
0 people saved or upvoted this
Feed
Explore
Sign In