Right now if you have a PDF in a folder that's corrupt, not in an expected format (e.g. you save a HTML file as a .PDF), or has some other issue, the parser fails. Ideally it would skip files it cannot parse or load and maybe output a note or error log somewhere.
Not putting this here expecting you will fix it, but I want to start logging all the pain points so I can come back to them at some point with a PR.
Right now if you have a PDF in a folder that's corrupt, not in an expected format (e.g. you save a HTML file as a .PDF), or has some other issue, the parser fails. Ideally it would skip files it cannot parse or load and maybe output a note or error log somewhere.
Not putting this here expecting you will fix it, but I want to start logging all the pain points so I can come back to them at some point with a PR.