遇见数据集

lighthousehq/eligibility_clean_with_resume_20240717

收藏
Hugging Face2024-07-17 更新2024-07-22 收录
官方服务:

资源简介:

该数据集包含多个与个人信息相关的特征字段,如作者身份强度、作者身份摘要、奖项强度、奖项摘要等。数据集主要用于分析或研究与个人职业或学术背景相关的信息。训练集包含153个样本,总大小为988965字节。

This dataset contains multiple feature fields related to personal information, such as authorship strength, authorship summary, award strength, award summary, etc. The dataset is primarily used for analyzing or researching information related to personal career or academic background. The training set includes 153 samples with a total size of 988965 bytes.

提供机构:
lighthousehq
原始信息汇总

数据集概述

数据特征

  • authorship_strength: 字符串类型
  • authorship_summary: 字符串类型
  • award_strength: 字符串类型
  • award_summary: 字符串类型
  • createdate: 字符串类型
  • critical_role_strength: 字符串类型
  • critical_role_summary: 字符串类型
  • cv: 字符串类型
  • email: 字符串类型
  • firstname: 字符串类型
  • google_scholar_url: 字符串类型
  • high_remuneration_strength: 字符串类型
  • high_remuneration_summary: 字符串类型
  • hs_object_id: 字符串类型
  • judging_strength: 字符串类型
  • judging_summary: 字符串类型
  • lastmodifieddate: 字符串类型
  • lastname: 字符串类型
  • membership_strength: 字符串类型
  • membership_summary: 字符串类型
  • message: 字符串类型
  • optional_scenarios: 字符串类型
  • original_contributions_strength: 字符串类型
  • original_contributions_summary: 字符串类型
  • published_materials_strength: 字符串类型
  • published_materials_summary: 字符串类型
  • cv_text: 字符串类型
  • index_level_0: 整数类型

数据分割

  • train: 包含153个样本,占用988965字节

数据集大小

  • 下载大小: 516120字节
  • 数据集大小: 988965字节

配置

  • config_name: default
    • data_files:
      • split: train
      • path: data/train-*
搜集汇总
数据集介绍
lighthousehq/eligibility_clean_with_resume_20240717 数据集图片
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务