bigscience-workshop / bigscience-workshop/lam
Add dataset: old_book_illustrations
- Dominant language
- No language data
- Stars
- 91
- Forks
- 8
- PR merge metrics
- No merged PRs in 30d
Description
### A URL for this dataset
https://www.oldbookillustrations.com/
### Dataset description
The Old Book Illustrations website contains a dataset of illustrations scanned from old books. Each illustration page also contains infos about the illustrator, the illustration and the book it's taken from as well as a title, a description, and a few keywords. As of today, the website contains 3150 images.
I already wrote a script to scrap all the content since the api does not give access to all the information (for instance the image is not is the best resolution).
Is it a dataset that is relevant for this project?
About the license, the website reads:
* Text content (descriptions, translations, etc.) is published under a Creative Commons [Attribution-NonCommercial-ShareAlike 4.0 International License](http://creativecommons.org/licenses/by-nc-sa/4.0/).
* Although we do our best to offer only Illustrations that are considered public domain in most countries, copyright laws vary from one jurisdiction to another, and you agree that you are solely responsible for abiding by all laws and regulations that may be applicable to using the Illustrations.
More info on the [term of use page](https://www.oldbookillustrations.com/terms-of-use/).
### Dataset modality
Image
### Dataset licence
Creative Commons Public Domain Dedication and Certification
### Other licence
_No response_
### How can you access this data
Other
### Confirm the dataset has an open licence
- [X] To the best of my knowledge, this dataset is accessible via an open licence
### Contact details for data custodian
_No response_
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.