遇见数据集

Books to Scrape

收藏
Zenodo2020-11-08 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

This dataset in CSV format contains all books from the web http://books.toscrape.com which has been got using web scraping method in November 2020. The CSV file has 12 columns called each of them like: title, image, rating, description, category, UPC, producttype, priceextax, priceincltax, tax, availability, numberreviews. The project was born as a practice for a subject of the Master of Science (MSc) of Data Science at the Universitat Oberta de Catalunya (UOC).

本数据集采用CSV格式,收录了通过网络爬虫方法于2020年11月从网站http://books.toscrape.com抓取的全部图书信息。该CSV文件共包含12个字段,分别为:标题(title)、图片(image)、评分(rating)、图书简介(description)、分类(category)、通用产品代码(UPC)、产品类型(producttype)、税前价格(priceextax)、税后价格(priceincltax)、税费(tax)、库存状态(availability)以及评论数(numberreviews)。本项目作为加泰罗尼亚开放大学(Universitat Oberta de Catalunya, UOC)数据科学理学硕士(Master of Science, MSc)课程的实训作业开发而成。

提供机构:
Zenodo
创建时间:
2020-11-08
二维码
社区交流群
二维码
科研交流群
商业服务