相关数据集
RIVER: A Real-Time Interaction Benchmark for Video LLMs
该项目介绍了RIVER Bench,旨在通过流式视频感知评估视频大型语言模型的实时交互能力,具有记忆、实时感知和主动响应等新任务。根据参考事件、问题和答案的频率和时间,我们进一步将在线交互任务分为四个不同的子类。对于Retro-Memory,线索来自过去;对于live-Perception,线索来自现在——两者都要求立即响应。对于Pro-Response任务,视频大型语言模型需要等待相应的线索出现
github2026-03-05 更新300
Everything_Instruct_Multilingual
Everything Instruct (Multilingual Edition) 旨在为开源AI的大语言模型带来提升,它是一个大型的Alpaca指令格式数据集,涵盖广泛的主题。该数据集包含7,800,783行数据,最大长度为78,451个token,支持包括英语、俄语、中文、韩语、乌尔都语、拉丁语、阿拉伯语、德语、西班牙语、法语、印地语、意大利语、日语、荷兰语和葡萄牙语在内的多种语言。数据内容
OpenCSG2024-07-19 更新100
LLM Merging Methods in the weight space
This table analyzes weight-space LLM merging approaches that operate directly on model parameters. It characterizes their architectural constraints, training requirements, computational cost, experime
DataCite Commons2026-03-02 更新50
NLPLog.json
The dataset developed in the paper titled 'Adapting Large Language Models to Log Analysis with Interpretable Domain Knowledge' is designed to transform information found in logs into interpretable kno
Figshare2025-05-24 更新150
amogh-sinha/Llama-2-7B-Chat-GGML
--- dataset_info: features: - name: text dtype: string splits: - name: train num_bytes: 39206 num_examples: 1 download_size: 16971 dataset_size: 39206 configs: - config_name: d
Hugging Face2023-08-08 更新80



