1 comment

[ 3.0 ms ] story [ 11.8 ms ] thread
My team and I have worked on an agentic benchmark to see how well agents can perform at reverse engineering threats. The results are pretty interesting, we would love the communities feedback.