Multi-Pattern Multi-Model Benchmark Runs on Live Network Configuration Tasks
收藏官方服务:
资源简介:
This dataset is a export of 350 benchmark run artifacts from a comparative study of LLM agent orchestration patterns on live network configuration tasks (FRRouting under Containerlab). Five tasks span RIP, OSPF, and BGP protocols at graded difficulty; five orchestration patterns (single-agent, sequential, debate, group chat, concurrent) are evaluated with four language models. The bundle is intended for **reproducibility, independent re-analysis, and citation** alongside the companion manuscript and related publications.
提供机构:
Zenodo创建时间:
2026-04-10



