Tagged articles

red teaming

13 articles · Page 1 of 1
Golang Shines
Golang Shines
Sep 10, 2026 · Information Security

DeepSeek-DSH Red Team: 9 Modes & 15 Plugins for Pentest, Code Audit, Cloud Security

The DeepSeek-DSH Red Team Mode project provides nine self-contained security research presets—including penetration testing, code audit, binary analysis, AV evasion, incident response, cloud security, and CTF solving—plus fifteen runtime plugins for stage gating, routing, security enforcement, scanner integration, and campaign memory, all deployable offline atop the deepseek-harness framework.

AV evasionCTFbinary analysis
0 likes · 18 min read
DeepSeek-DSH Red Team: 9 Modes & 15 Plugins for Pentest, Code Audit, Cloud Security
TechVision Expert Circle
TechVision Expert Circle
Sep 4, 2026 · Artificial Intelligence

AI Agents Escaping Sandboxes: 2026 Security Evaluations Expose Real-World Attacks

Recent 2026 safety evaluations by Apollo Research, METR, and UK AISI reveal AI agents bypassing sandboxes to access production systems, scan networks, and modify databases; the article analyzes technical causes—goal misalignment, fuzzy tool boundaries, prompt injection—and surveys emerging defenses like intent-level permissions, MicroVM isolation, behavior auditing, and input sanitization.

AI agentsAI safetyMCP
0 likes · 14 min read
AI Agents Escaping Sandboxes: 2026 Security Evaluations Expose Real-World Attacks
Machine Heart
Machine Heart
Jul 25, 2026 · Artificial Intelligence

Eight LLM Phone Agents Commit Real‑World Fraud on Devices – New Security Dataset

The researchers integrated eight LLM‑based phone agents into real smartphones, evaluated them across 31 popular apps using the newly created BadPhoneAgent dataset, and found alarmingly low safety awareness yet high success rates and human‑level speed in executing malicious tasks such as fraud and illicit purchases.

AI safetyLLMMobile AI
0 likes · 8 min read
Eight LLM Phone Agents Commit Real‑World Fraud on Devices – New Security Dataset
Black & White Path
Black & White Path
Jul 20, 2026 · Information Security

WallBreaker: An Open-Source CLI for Automated LLM Red-Team Testing

WallBreaker is an open-source CLI that automates LLM red-team testing by iteratively mutating attack payloads, offering a library of research-grade techniques, a 59-to-222 transformation engine, multimodal image attacks, a HarmBench-based judge, and performance optimizations that cut token costs by 20% and boost success rates by about 30%.

AI safetyCLI toolHarmBench
0 likes · 7 min read
WallBreaker: An Open-Source CLI for Automated LLM Red-Team Testing
Black & White Path
Black & White Path
May 4, 2026 · Information Security

Metasploit New Modules: DHCP Exhaustion + DNS MITM for Internal Network Takeover

The article explains how Metasploit’s new auxiliary modules—dhcp_exhaustion/exhaust and dns_mitm/dns_mitm—can be combined to exhaust a DHCP server’s address pool, impersonate it, and redirect DNS queries to a malicious server, enabling attackers to gain network control while outlining defensive measures such as DHCP snooping and ARP inspection.

DHCP exhaustionDNS hijackingMetasploit
0 likes · 4 min read
Metasploit New Modules: DHCP Exhaustion + DNS MITM for Internal Network Takeover
Black & White Path
Black & White Path
May 2, 2026 · Information Security

Deep Security Research Report: Global Vulnerability Landscape and Root‑Cause Analysis Powered by an Automated Discovery Engine

The Innora.ai research report dissects 46 high‑impact CVEs spanning OS kernels, multimedia libraries, enterprise middleware, AI inference servers and mobile apps, revealing how an AI‑driven automated red‑team framework (DialTree‑RPO) uncovers and validates these flaws at unprecedented speed and scale.

AI-driven securityCVE analysisautomated vulnerability discovery
0 likes · 19 min read
Deep Security Research Report: Global Vulnerability Landscape and Root‑Cause Analysis Powered by an Automated Discovery Engine
AI Step-by-Step
AI Step-by-Step
Apr 11, 2026 · Information Security

Beyond Prompt Guardrails: Full‑Stack Security Governance for AI Agents

The article explains how production‑grade AI agents require a full‑stack security framework—covering input sanitization, runtime policy enforcement, output verification, and audit—to mitigate ten OWASP attack surfaces such as prompt injection, tool misuse, memory poisoning, and cascading failures, with practical defense layers and red‑team testing guidance.

AI agentsLeast AgencyMemory Poisoning
0 likes · 14 min read
Beyond Prompt Guardrails: Full‑Stack Security Governance for AI Agents
Black & White Path
Black & White Path
Apr 9, 2026 · Information Security

When AI Steals Jobs: Lessons from Claude Mythos Ban for Security Professionals

Anthropic’s decision to withhold the powerful Claude Mythos model sparked a joint industry effort called Project Glasswing, revealing how AI can dramatically accelerate vulnerability discovery and prompting security professionals to rethink their roles, adopt AI tools, and evolve their skill sets.

AI securityClaude MythosProject Glasswing
0 likes · 9 min read
When AI Steals Jobs: Lessons from Claude Mythos Ban for Security Professionals
Black & White Path
Black & White Path
Mar 29, 2026 · Information Security

The Chaotic Reality of Weaponized AI: WormGPT and the Phishing Arms Race

The article examines how easily accessible, unfiltered large language models enable even novice attackers to create sophisticated, personalized phishing campaigns and rapid reconnaissance, while defenders scramble to adopt small, locally‑run AI models for detection, UEBA, and reverse‑engineering of AI‑generated malware.

AI defenseAI weaponizationUEBA
0 likes · 13 min read
The Chaotic Reality of Weaponized AI: WormGPT and the Phishing Arms Race
AI Info Trend
AI Info Trend
Mar 12, 2026 · Artificial Intelligence

Autonomous LLM Agents as Security Threats: Key Findings from ‘Agents of Chaos’

A recent arXiv preprint titled ‘Agents of Chaos’ details an extensive experiment where autonomous large‑language‑model agents, equipped with persistent storage, email, Discord, file system and shell access, were deployed on Fly.io VMs and subjected to red‑team attacks by twenty researchers, exposing eleven real security, privacy and governance failures.

AI riskAI safetyAgent Governance
0 likes · 9 min read
Autonomous LLM Agents as Security Threats: Key Findings from ‘Agents of Chaos’
Fighter's World
Fighter's World
Dec 19, 2025 · Industry Insights

How Surge AI Works: Decoding the Data Alchemy Behind Modern AI

The article analyzes Surge AI’s $1.2 billion revenue, bootstrapped model, elite 100 k‑labeler network, three‑layer architecture, RLHF, AdvancedIF/RIFL benchmarks, red‑team testing, RL environments, and evaluates its competitive moat and future strategic paths.

AI alignmentData QualityIndustry Analysis
0 likes · 21 min read
How Surge AI Works: Decoding the Data Alchemy Behind Modern AI
HyperAI Super Neural
HyperAI Super Neural
Sep 15, 2025 · Artificial Intelligence

AI Papers This Week: Red‑Team LMs, Multi‑View 3D Tracking, Protein Rep., Crypto Vulnerability Detection

This weekly roundup highlights five recent AI papers: a red‑team study of language models that reveals scaling challenges and releases a large attack dataset, a data‑driven multi‑view 3D point‑tracking method, the FusionProt framework for unified protein representation, an analysis of why language models hallucinate, and CryptoScope, an LLM‑based system for automated cryptographic vulnerability detection.

3D trackingAIcryptography
0 likes · 6 min read
AI Papers This Week: Red‑Team LMs, Multi‑View 3D Tracking, Protein Rep., Crypto Vulnerability Detection