diff --git a/README.md b/README.md index edad28bb..91ced2f0 100644 --- a/README.md +++ b/README.md @@ -396,8 +396,8 @@ _Libraries to automate web scraping and extract web content._ - [stagehand](https://github.com/browserbase/stagehand) - A fast and token-efficient browser automation SDK to extract data and perform self-healing actions on web pages. - [jev-ultrafast](https://github.com/browser-use/jev-ultrafast) - A fast browser agent that picks actions from an indexed table of page elements, using a small LLM only to type text. - Content Extraction + - [markdownify](https://github.com/matthewwithanm/python-markdownify) - Convert HTML to Markdown, with customizable tag handling. - [feedparser](https://github.com/kurtmckee/feedparser) - Universal feed parser. - - [html2text](https://github.com/Alir3z4/html2text) - Convert HTML to Markdown-formatted text. - [trafilatura](https://github.com/adbar/trafilatura) - A tool for gathering text and metadata from the web, with built-in content filtering. ### Email