遇见数据集

跨域资源调度测试日志数据集

收藏
官方服务:

资源简介:

跨域资源调度测试日志数据集采集自分布式服务器的自动化日志系统(如Nginx),记录客户端HTTP请求的完整交互信息,每条日志包含IP地址、精确时间戳、请求方法、资源路径、协议版本、状态码、响应大小、引用来源及用户代理等10个核心字段,空值以“-”标记,体量可达百万至千万级。数据生成于真实业务场景(如微服务跨域API调用、CDN资源调度),通过时序分析、状态码分布、高频IP行为等特征,支撑资源调度优化、安全威胁识别(如DDoS攻击检测)及系统性能评估,为网络流量分析、运维决策与算法研究提供多维度基准数据,适用于网络工程、信息安全及大数据领域的高效挖掘与应用建模。

This cross-domain resource scheduling test log dataset is collected from automated logging systems of distributed servers (e.g., Nginx), which records complete interaction details of client HTTP requests. Each log entry contains 10 core fields including IP address, exact timestamp, request method, resource path, protocol version, status code, response size, referrer, and user agent. Null values are marked with "-", and the dataset scale ranges from millions to tens of millions of entries. The data is generated in real-world business scenarios such as cross-domain API calls of microservices and CDN resource scheduling. Leveraging features including time-series analysis, status code distribution, and high-frequency IP behaviors, the dataset supports resource scheduling optimization, security threat identification (e.g., DDoS attack detection), and system performance evaluation. It provides multi-dimensional benchmark data for network traffic analysis, operation and maintenance decision-making, and algorithm research, and is applicable to efficient data mining and application modeling in the fields of network engineering, information security, and big data.

提供机构:
上海交通大学
搜集汇总
数据集介绍
跨域资源调度测试日志数据集 数据集图片
背景与挑战
背景概述
该数据集采集自分布式服务器自动化日志系统,记录了客户端HTTP请求的完整交互信息,包含IP地址、时间戳等10个核心字段,体量可达百万至千万级。数据生成于真实业务场景,如微服务跨域API调用和CDN资源调度,可用于资源调度优化、安全威胁识别和系统性能评估,为网络工程、信息安全等领域提供多维度基准数据。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务