Skip to main content
Layer 1
Logo KB Lab
Hoofdnavigatie
Datasets
Tools
Tutorials
News and events
Blogs
About us
Affiliated researchers
Team
Contact
Secondary menu
NL
Open Menu
zoeken
Digital-preservation-22
Extracting text from EPUB files in Python
Johan van der Knijff published a brief introduction to extracting unformatted text from EPUB files.
Dutch Novels 1800-2000
Dataset that contains a corpus of 1346 novels from DBNL.
Keyword generator
A command-line tool to extract significant keywords from a collection of sample texts.