遇见数据集

vannha204/15.12.2025

收藏
Hugging Face2025-12-16 更新2025-12-20 收录
官方服务:

资源简介:

该数据集未提供直接的中文描述,但从其包含的特征(video_id、audio和transcript)可以推断,这可能是一个包含视频或音频记录及对应转录文本的数据集。数据集包含一个名为train的分割,共有19,135个样本,总大小约为17.16 GB。音频数据的采样率为16,000 Hz。

The dataset does not provide a direct description, but based on the features listed (video_id, audio, and transcript), it can be inferred that this is likely a dataset containing video or audio recordings with corresponding transcripts. The dataset includes a single split named train with 19,135 examples and a total size of approximately 17.16 GB. The audio data has a sampling rate of 16,000 Hz.

提供机构:
vannha204
二维码
社区交流群
二维码
科研交流群
商业服务