Loan-Dataset
收藏资源简介:
这是一个包含20,000条个人和金融数据的合成数据集,旨在促进风险评分的预测模型开发。它有两个主要用途:1. 风险评分回归:预测与每个个体贷款违约或金融不稳定可能性相关的连续风险评分;2. 二元分类:确定贷款批准的二元结果,表明申请人是否可能被批准或拒绝贷款。数据集包括人口统计信息、信用历史、就业状况、收入水平、现有债务和其他相关金融指标等多种特征,为复杂的数据驱动分析和决策提供了全面的基础。
This is a synthetic dataset containing 20,000 entries of personal and financial data, designed to facilitate the development of predictive models for risk scoring. It has two primary applications: 1. Risk score regression: predicting continuous risk scores associated with the likelihood of loan default or financial instability for each individual; 2. Binary classification: determining the binary outcome of loan approval, indicating whether an applicant is likely to be approved or rejected for a loan. The dataset includes various features such as demographic information, credit history, employment status, income level, existing debts and other relevant financial metrics, providing a comprehensive foundation for complex data-driven analysis and decision-making.
Loan-Dataset 概述
数据集基本信息
- 记录数量: 20,000条
- 数据类型: 合成数据
数据集用途
- 风险评分回归: 预测与个人贷款违约或财务不稳定可能性相关的连续风险评分。
- 二元分类: 确定贷款审批的二元结果,预测申请人是否可能获得批准或被拒绝。
数据集特征
- 特征类型:
- 人口统计信息
- 信用历史
- 就业状况
- 收入水平
- 现有债务
- 其他相关财务指标
数据集目标
- 为风险评估和贷款审批建模提供全面的数据驱动分析和决策基础。




