遇见数据集

Sample Project: Exploratory Data Analysis with TMDB Data

收藏
Databricks2024-05-09 收录
官方服务:

资源简介:

**Exploring the TMDB Movies Dataset on Databricks** Welcome to our comprehensive guide on Exploratory Data Analysis (EDA) on Databricks. Using a rich movies dataset, you'll journey through exploring and profiling the data, to cleaning and querying it, and ultimately discovering insights about the world of cinema, complemented by compelling visualizations. By the end of this notebook, you'll not only be more familiar with our platform's capabilities, but you'll also gain a hands-on understanding of the fundamental steps in EDA. In this sample project, you will: 1. Read in a sample dataset 2. Profile the data to get a general idea of our dataset's contents 3. Clean up the data and save it as a new table 4. Query the cleaned data to gain insights about movies (with visualizations!) 5. Take the exploration further on your own with a challenge The data set you will use contains information about 10,000 movies collected from The Movie Database (TMDB), including user ratings and revenue. Source: [https://www.kaggle.com/datasets/akshaypawar7/millions-of-movies](https://www.kaggle.com/datasets/akshaypawar7/millions-of-movies)

提供机构:
Databricks
搜集汇总
数据集介绍
Sample Project: Exploratory Data Analysis with TMDB Data 数据集图片
背景与挑战
背景概述
该数据集为TMDB电影数据探索性分析示例项目,提供了10,000部电影的信息,包括用户评分和收入等。项目旨在引导用户通过数据读取、分析、清理和查询等步骤,在Databricks平台上进行EDA并生成可视化结果。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务