Tagged articles

AgentDoG

2 articles · Page 1 of 1
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Aug 11, 2026 · Artificial Intelligence

How DoGNAVY Ranked #3 Globally in AI Security Using a Single Open‑Source Model

DoGNAVY achieved a 90.84% verification rate and placed third on the CyberGym AI‑security leaderboard by leveraging the open‑source GLM‑5.2 model within a multi‑agent workflow that combines reachability analysis, dynamic testing, independent review, and a strict sandbox environment.

AI securityAgentDoGCyberGym benchmark
0 likes · 14 min read
How DoGNAVY Ranked #3 Globally in AI Security Using a Single Open‑Source Model
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jun 7, 2026 · Artificial Intelligence

AgentDoG 1.5: A Lightweight, Extensible Framework for Trajectory‑Level Agent Safety

AgentDoG 1.5 expands AI‑agent safety from final replies to complete execution trajectories, introducing the ATBench family for fine‑grained evaluation, a taxonomy‑guided DataEngine for high‑quality data generation, and demonstrating substantial safety gains in both SFT/RL training and online guardrail deployment with lightweight models.

AI safetyATBenchAgentDoG
0 likes · 14 min read
AgentDoG 1.5: A Lightweight, Extensible Framework for Trajectory‑Level Agent Safety