AttackonCTF: Defending Hardware Security Competition Benchmarks in the Age of LLMs (opens in new tab)

Covered by Semiconductor Engineering

Hardware security competitions such as HackTheSilicon serve as benchmarking platforms for evaluating vulnerability detection methods and for training humans and AI. However, our study reveals that LLMs threaten their validity. Instead of genuine security reasoning, detectors exploit a diff-style syntactic comparison, achieving an 83% detection rate, undermining fair evaluation. To mitigate this, we propose the first LLM-oriented, semantics-prese...

Read the original article

Sign in to keep reading the full article.

Sign Up Log In

Covered in 1 article

Semiconductor Engineering·

Covered in 1 article

Chip Industry Week In Review