You cannot select more than 25 topics Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
langchain/libs/community/tests/examples
Erick Friis 3a2eb6e12b
infra: add print rule to ruff (#16221)
Added noqa for existing prints. Can slowly remove / will prevent more
being intro'd
4 months ago
..
README.org community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
README.rst community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
brandfetch-brandfetch-2.0.0-resolved.json community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
default-encoding.py community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
docusaurus-sitemap.xml community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
duplicate-chars.pdf community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
example-utf8.html community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
example.html community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
example.json community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
example.mht community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
facebook_chat.json community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
factbook.xml community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
fake-email-attachment.eml community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
fake.odt community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
fake.vsdx community[minor]: New documents loader for visio files (with extension .vsdx) (#16171) 5 months ago
hello.msg community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
hello.pdf community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
hello_world.js community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
hello_world.py infra: add print rule to ruff (#16221) 4 months ago
layout-parser-paper.pdf community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
multi-page-forms-sample-2-page.pdf community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
non-utf8-encoding.py community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
sample_rss_feeds.opml community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
sitemap.xml community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
slack_export.zip community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
stanley-cups.csv community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
stanley-cups.tsv community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
stanley-cups.xlsx community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago
whatsapp_chat.txt community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463) 6 months ago

README.rst

Example Docs
------------

The sample docs directory contains the following files:

-  ``example-10k.html`` - A 10-K SEC filing in HTML format
-  ``layout-parser-paper.pdf`` - A PDF copy of the layout parser paper
-  ``factbook.xml``/``factbook.xsl`` - Example XML/XLS files that you
   can use to test stylesheets

These documents can be used to test out the parsers in the library. In
addition, here are instructions for pulling in some sample docs that are
too big to store in the repo.

XBRL 10-K
^^^^^^^^^

You can get an example 10-K in inline XBRL format using the following
``curl``. Note, you need to have the user agent set in the header or the
SEC site will reject your request.

.. code:: bash

   curl -O \
     -A '${organization} ${email}'
     https://www.sec.gov/Archives/edgar/data/311094/000117184321001344/0001171843-21-001344.txt

You can parse this document using the HTML parser.