Tagged articles

Security Evaluation

3 articles · Page 1 of 1
Machine Heart
Machine Heart
Aug 8, 2026 · Artificial Intelligence

Why Anthropic Says Claude Code’s Auto Mode Is Safer After Testing 1,000 Users

Anthropic’s new default Auto Mode for Claude Code uses a dedicated classifier that caught 89% of dangerous commands versus 14% for manual approval, a study of 1,053 paid testers showed equal or better safety, fewer harmful actions, and zero successful attacks on Claude models compared with competing systems.

AI SafetyAgent ToolsAnthropic
0 likes · 11 min read
Why Anthropic Says Claude Code’s Auto Mode Is Safer After Testing 1,000 Users
AntTech
AntTech
May 25, 2026 · Artificial Intelligence

Ant Group and Five Universities Unveil Agent3σ: A Multi‑Layer AI Agent Security Evaluation Platform

As AI agents move from simple Q&A to tool use and real‑world actions, Ant Group and five leading universities launch the open‑source Agent3σ platform, offering a three‑tier, 7‑dimension risk framework and concrete metrics to assess agents' safety across static, simulated, and live environments.

AI AgentAgent3σMulti‑Layer Testing
0 likes · 9 min read
Ant Group and Five Universities Unveil Agent3σ: A Multi‑Layer AI Agent Security Evaluation Platform
Tencent Technical Engineering
Tencent Technical Engineering
Jul 16, 2025 · Artificial Intelligence

Introducing A.S.E: The First Project‑Level AI Code Generation Security Evaluation Framework

The A.S.E (AI Code Generation Security Evaluation) framework provides a comprehensive, project‑level benchmark for assessing the safety, quality, and stability of AI‑generated code across multiple languages and vulnerability types, helping developers and researchers evaluate and improve large language model coding assistants.

AI code generationSecurity EvaluationVulnerability detection
0 likes · 7 min read
Introducing A.S.E: The First Project‑Level AI Code Generation Security Evaluation Framework