Projects
Loading projects...
Fast, flexible, and lean implementation of core jQuery designed specifically for the server. |
Pushed 2 days ago 169 contributors Created 15 years ago |
30.4k |
Extract the Readable Content from an HTML Document | Pushed 18 days ago 99 contributors Created 11 years ago | 11.4k |
Extract the main content from web pages. | Pushed 4 days ago 37 contributors Created a year ago | 8.58k |
The next web scraper. See through the <html> noise | Pushed 6 years ago 43 contributors Created 11 years ago | 5.9k |
Extract meaningful content from the chaos of a web page | Pushed 3 years ago 58 contributors Created 10 years ago | 5.79k |
A command-line tool to grab web pages as beautifully formatted PDFs | Pushed a year ago 18 contributors Created 8 years ago | 4.66k |
A Node.js scraper for humans. | Pushed 20 days ago 21 contributors Created 10 years ago | 4.07k |
A library to easily scrape metadata from an article on the web using Open Graph, JSON+LD, regular HTML metadata, and series of fallbacks. | Pushed 7 days ago 42 contributors Created 10 years ago | 2.72k |
Extract main article, main image and meta data from URL | Pushed 3 months ago 17 contributors Created 11 years ago | 1.9k |
Download website to local directory (including all css, images, js, etc.) | Pushed 6 days ago 20 contributors Created 12 years ago | 1.74k |
A super simple site crawler and broken link checker | Pushed 3 days ago 30 contributors Created 7 years ago | 1.24k |
Metadata scraper with support for oEmbed, Twitter Cards and Open Graph Protocol for Node.js | Pushed 2 years ago 26 contributors Created 9 years ago | 499 |