Awesome - I'm a dabbler, but any thoughts on best engines for PDF tables? I've got tons of PDFs with similar tables embedded deep in them, but all formatted slightly differently. Seems like it should be easy....but nope!
PDF -> Markdown looks like a pretty great use case
Just added box detection support -- maybe I'll start from here https://github.com/junhoyeo/BetterOCR#-box-detection