AgentRE-Bench:AgentRE-Bench is an agentic benchmark that evaluates state-of-the-art models on long-horizon reverse engineering tasks, measuring their ability to analyze binaries, use tooling effectively, and reason over multi-step execution artifacts
skill-eval-harness:Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
reproduce-cgo2017-paper:Artifact Evaluation Reproduction for "Software Prefetching for Indirect Memory Accesses", CGO 2017, using CK.
ISSTA21-JIT-DP:This repo illustrates how to evaluate the artifacts in the paper Deep Just-in-Time Defect Prediction: How Far Are We? published in ISSTA'21.
allo-pldi24-artifact:Artifact evaluation of PLDI'24 paper "Allo: A Programming Model for Composable Accelerator Design"
// repository documentation
Was this content helpful?
★ 0(0 ratings)
Recent Feedback
Download README
Do you want to download the README.md file for micro24-fusemax-artifact?