遇见数据集

lowry02/hs-real-case

收藏
Hugging Face2026-05-17 更新2026-05-31 收录
官方服务:

资源简介:

该数据集包含句子在模型中的隐藏状态表示,具体特征包括基准行索引、句子文本、层索引、词元位置、词元ID、词元字符串以及隐藏状态向量。数据分片为hidden_states,总大小约为850MB,包含196,704个示例,适用于自然语言处理中的模型分析或特征提取任务。

This dataset contains hidden state representations of sentences in a model, with features including benchmark row index, sentence text, layer index, token position, token ID, token string, and hidden state vectors. The split is named hidden_states, with a total size of approximately 850MB and 196,704 examples, suitable for model analysis or feature extraction tasks in natural language processing.

提供机构:
lowry02
二维码
社区交流群
二维码
科研交流群
商业服务