
LibriVox
Open Library
Standard Ebooks
Archive.org
Audible
Project Gutenberg
LearnOutLoud.com
Libro.fm
Simple Scraper
Octoparse
Scraper API
Diggernaut
Agenty
Apify
eScraper
Crawlbase
Simple scraper is the easiest way to scrape the web โ turn any website into an API in seconds and use ready-made scraping recipes to scrape popular sites with ease.
LibriVox
Simple ScraperBased on our record, LibriVox seems to be a lot more popular than Simple Scraper. While we know about 236 links to LibriVox, we've tracked only 22 mentions of Simple Scraper. We are tracking product recommendations and mentions on various public social media platforms and blogs. They can help you identify which product is more popular and what people think of it.
Then share the result on https://librivox.org and other platforms. - Source: Hacker News / 12 days ago
i've been building this over the last two weeks, it first started as a reader for standard ebooks (https://standardebooks.org/), with clean typography, progress tracking, and nice deep linking (try clicking on a paragraph to create a mark and copy it to your clipboard) but I think I really unlocked something when I discovered the amazing recordings on https://librivox.org, I've been spending the last few days... - Source: Hacker News / 15 days ago
Used to be one could sort of get that with the Project Librivox: https://librivox.org/ e-book app Gutebooks (in addition to their audio app), but it seems to have been deprecated (I'm no longer able to connect to the server on my copy (which I only got 'cause there was an in-app purchase to fund Project Librivox). FWIW, Barnes & Noble has been plundering the public domain using a book composition/keying house in... - Source: Hacker News / 3 months ago
There's a really awesome site called Librivox [1], where volunteers narrate books that are in the public domain. Those recordings are also in the public domain as well (this is just part of the Librivox thing). The quality of those recordings (both the narration, and the actual recording quality) varies quite a bit and most of them aren't at a quality I'd expect people to pay for and thus aren't useable for me.... - Source: Hacker News / 8 months ago
This course has been a _huge_ help to me in my current project (a G-code previewer and programmatic 3D modeling system written for (Open)PythonSCAD). I would recommend pairing it with: https://ocw.mit.edu/courses/6-001-structure-and-interpretation-of-computer-programs-spring-2005/ and some additional online resources which I've found very helpful: - https://mathcs.clarku.edu/~djoyce/java/elements/elements.html -... - Source: Hacker News / about 1 year ago
Data extraction: https://simplescraper.io A project that I launched on HN that became a business. Simplescraper rode the no-code wave of a few years back ('instant structured data without parsing html'). Now working on increasing the surface area for AI agents: MCP support, screenshots API, and (experimentally) x402^ ^ https://simplescraper.io/blog/x402-payment-protocol/. - Source: Hacker News / 5 months ago
1. Clicking the box programmatically โ possible but inconsistent 2. Outsourcing the task to one of the many CAPTCHA-solving services (2Captcha etc) โ better 3. Using a pool of reliable IP addresses so you don't encounter checkboxes or turnstiles โ best I run a web scraping startup (https://simplescraper.io) and this is usually the approach. It has become more difficult, and I think a lot of the AI crawlers are... - Source: Hacker News / over 1 year ago
Making my data extraction Saas (https://simplescraper.io) more LLM friendly. Markdown extraction, improved Google search, workflows - search for this terms, visit the first N links, summarize etc. Big demand for (or rather, expectation of) this lately. - Source: Hacker News / almost 2 years ago
Things are much easier for one-person startups these daysโit's a gift. I remember building a todo app as my first SaaS project, and choosing something called Stormpath for authentication. It subsequently shut down, forcing me to do a last-minute migration from a hostel in Japan using Nitrous Cloud IDE (which also shut down). Just pain upon pain.[1] Now, you can just pick a full-stack cloud service and run with it.... - Source: Hacker News / about 2 years ago
Simplescraper โ Trigger your webhook after each operation. The free plan includes 100 cloud scrape credits. - Source: dev.to / over 2 years ago
Open Library - The ultimate goal of the Open Library is to make all the published works of humankind available to...
Octoparse - Octoparse provides easy web scraping for anyone. Our advanced web crawler, allows users to turn web pages into structured spreadsheets within clicks.
Standard Ebooks - Online library of downloadable e-books that focuses on quality and modern standards in typography.
Scraper API - Scale Data Collection with a Simple API.
Archive.org - Internet Archive is a non-profit digital library offering free universal access to books, movies...
Diggernaut - Web scraping is just became easy. Extract any website content and turn it into datasets. No programming skills required.