I built Danubia, a web scraping and content extraction API, hosted in the EU.<p>Why: I used to build crawler infrastructure for a search engine company. I wanted to apply my knowledge to building a clean text extraction API, in part because was it fun to do (why else do we progra
A scraping API for developers. Send a URL and get clean markdown back, on demand or in bulk.
Danubia | Turn any web page into clean text danubia Docs Language English Français Dashboard Hi 👋 Here's something I've been working on lately: Danubia, a clean text extraction API. You give it a URL, it sends back the page's main content as clean markdown, on demand or in batch for less. I used to be an engineer on the crawl side of a large search engine. Now I keep watching scraping projects reinvent the same hard, fragile parts of a crawler, and I keep thinking to myself: "Hey, I think I know a better way" So I started applying what I know about running crawlers in
A scraping API for developers. Send a URL and get clean markdown back, on demand or in bulk.
Danubia | Turn any web page into clean text danubia Docs Language English Français Dashboard Hi 👋 Here's something I've been working on lately: Danubia, a clean text extraction API. You give it a URL, it sends back the page's main content as clean markdown, on demand or in batch for less. I used to be an engineer on the crawl side of a large search engine. Now I keep watching scraping projects reinvent the same hard, fragile parts of a crawler, and I keep thinking to myself: "Hey, I think I know a better way" So I started applying what I know about running crawlers in