遇见数据集

ByteDance-Seed/cudaLLM-data

收藏
Hugging Face2025-08-03 更新2025-08-09 收录
官方服务:

资源简介:

CudaLLM数据集是一个包含PyTorch操作测试用例的高质量数据集,旨在对大型语言模型在生成优化CUDA内核方面的能力进行基准测试和评估。该数据集提供了一组问题(标准的PyTorch nn.Module实现)和解决方案(使用自定义CUDA内核的性能优化版本)。这个数据集对于高性能计算、代码生成和编译器优化的AI研究非常有价值。

CudaLLM Dataset is a high-quality dataset of PyTorch operator test cases designed to benchmark and evaluate the capabilities of large language models in generating optimized CUDA kernels. The dataset provides pairs of problems (standard PyTorch nn.Module implementations) and solutions (performance-optimized versions using custom CUDA kernels). This dataset is a valuable resource for AI research in high-performance computing, code generation, and compiler optimization.

提供机构:
ByteDance-Seed
搜集汇总
数据集介绍
ByteDance-Seed/cudaLLM-data 数据集图片
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务