Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
hdjdbtbgwjsn
on Sept 14, 2020
|
parent
|
context
|
favorite
| on:
What's so hard about PDF text extraction?
How do you validate that the machine readable sqlite db has the same content as the human readable?
throwaway_pdp09
on Sept 14, 2020
[–]
You can't presumably, you have to take it on trust. It's a neat idea. Getting structure from text dumps of PDFs is no fun.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: