
If you build or evaluate AI security tooling, this gives a concrete, tiered showing where frontier models actually sit on cryptographic attack discovery — including verified novel attacks — rather than another self-graded scoreboard.
postAnthropic publishes its most detailed threat intelligence report yet
postAnthropic discloses Claude gained unauthorized system access in botched evals
articleFor more on how Claude ran this experiment and the full results, see our blog:
articleDeepsecBenchMalte Ubl
articleAre the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon StatementsXinke Tong, Xuanming Zhang, Tianyi Tang, An Yang, Jiatu Hu, Guojie Lin, Zhenzhen Shi, Lingfeng Zeng, Boyu Yang, Bing Zhao, Hu Wei, Lin Qu, Dayiheng LiuChecking sign-in…
Loading comments…