BinSimDB
收藏资源简介:
BinSimDB是由中央俄克拉荷马大学构建的一个细粒度二进制代码相似性分析基准数据集,包含4,426,258对等效的汇编代码片段。数据集主要用于解决二进制代码在不同优化级别或平台下的相似性比较问题,特别是在基本块级别的比较。数据集的创建过程包括使用BMerge和BPair算法来处理不同优化级别或平台导致的二进制代码片段差异。BinSimDB的应用领域广泛,包括漏洞发现、恶意软件分析和代码重用检测等安全相关应用。
BinSimDB is a fine-grained benchmark dataset for binary code similarity analysis developed by the University of Central Oklahoma. It contains 4,426,258 pairs of equivalent assembly code fragments. This dataset is primarily designed to address the problem of similarity comparison of binary codes under different optimization levels or platforms, especially comparisons at the basic block level. The construction process of BinSimDB utilizes the BMerge and BPair algorithms to handle the discrepancies in binary code fragments caused by varying optimization levels or platforms. BinSimDB has a wide range of application scenarios, including security-related applications such as vulnerability discovery, malware analysis, and code reuse detection.
BinSimDB 数据集概述
数据集名称
BinSimDB
数据集描述
BinSimDB 是一个用于细粒度二进制代码相似性分析的基准数据集。
相关研究
- 论文标题: "BinSimDB: Benchmark Dataset Construction for Fine-Grained Binary Code Similarity Analysis"
- 作者: Fei Zuo, Cody Tompkins, Qiang Zeng, Lannan Luo, Yung Ryn Choe, Junghwan Rhee
- 会议: 20th EAI International Conference on Security and Privacy in Communication Networks (SecureComm 2024)
- 日期: October 28, 2024
数据集下载链接

- 1BinSimDB: Benchmark Dataset Construction for Fine-Grained Binary Code Similarity Analysis中央俄克拉荷马大学 · 2024年



