Story

arxiv_cs_lg ยท Oct 2, 2026 ยท paper

Source brief

Threat-Preserving Representation Sensitivity in Agent-Security Benchmarks

arxiv.orgOct 2, 2026
original source linked

In brief

Security benchmarks for LLM-based agents often report the attack success rate (ASR) as a measure of model robustness and use these scores to compare different models and defense mechanisms, assuming that they describe...

Feed lens
agentevaluation

Continue reading

Read the original at arxiv.org โ†’Open in live feedRead that dayโ€™s brief

Earlier in this thread 4 items