forked from Archives/langchain
23231d65a9
Different PDF libraries have different strengths and weaknesses. PyMuPDF does a good job at extracting the most amount of content from the doc, regardless of the source quality, extremely fast (especially compared to Unstructured). https://pymupdf.readthedocs.io/en/latest/index.html |
||
---|---|---|
.. | ||
examples | ||
how_to_guides.rst | ||
key_concepts.md |