Corrected README

pull/1/head
Yuri Baburov 14 years ago
parent dada82099b
commit f925e3ef05

@ -4,20 +4,12 @@ This is a python port of a ruby port of arc90's readability project
http://lab.arc90.com/experiments/readability/ http://lab.arc90.com/experiments/readability/
In few words,
Given a html document, it pulls out the main body text and cleans it up. Given a html document, it pulls out the main body text and cleans it up.
It also can clean up title based on latest readability.js code.
Ruby port by starrhorne and iterationlabs Based on:
Python port by gfxmonk - Latest readability.js ( https://github.com/MHordecki/readability-redux/blob/master/readability/readability.js )
- Ruby port by starrhorne and iterationlabs
This port uses BeautifulSoup for the HTML parsing. That means it can be - Python port by gfxmonk ( https://github.com/gfxmonk/python-readability , based on BeautifulSoup )
a little slow, but will work on Google App Engine (unlike libxml-based - Decruft effort to move to lxml ( http://www.minvolai.com/blog/decruft-arc90s-readability-in-python/ )
libraries)
**note**: I don't currently have any plans for using or improving this
library, and it's far from perfect (slow, and almost certainly buggy).
So if you do something cool with it or have a better tool that does
the same job, please let me know and I can link to it from here.
If you're looking for alternatives, here's the list so far:
- http://www.minvolai.com/blog/decruft-arc90s-readability-in-python/

Loading…
Cancel
Save