SendTech Times
Analysis
SUPPLY CHECK:

Vijil DART Tests Enterprise AI Agents With Adaptive Red Teaming

Newsroom brief

Help Net Security reported that Vijil released DART, an adaptive red-teaming system that uses adversarial agents to test enterprise AI agents across tools, memory and multi-turn behavior.

Verified against source materialEdited by SendTech Times Cybersecurity DeskSource: Help Net Security
Vijil DART Tests Enterprise AI Agents With Adaptive Red Teaming
Image source: Help Net Security / Vijil article image

Vijil has moved AI-agent red teaming into an adaptive test pattern, Help Net Security reported, with a new DART system that uses adversarial agents to probe security flaws and policy violations across enterprise agents.

The product, formally Diamond Adaptive Red Teaming for Agents, is designed to test more than a model prompt or chat interface.

Its attacks run across tool use, memory and multi-turn behavior, then adjust tactics as the target agent responds.

Those are the same functions that can turn a chatbot-style risk into an operational security problem when an agent has enterprise permissions.

That makes the release a security-control story rather than a simple product update: Vijil is positioning DART as a way to stress agents inside the same operating conditions where they may make decisions or handle sensitive data.

The pressure behind that approach is scale.

Gartner estimates that a typical global Fortune 500 company will use more than 150,000 agents by 2028, compared with fewer than 15 in 2025.

Manual testing would struggle to keep pace with that increase, while attackers can also deploy agents for sustained, multi-turn campaigns against enterprise AI systems.

Static Prompts Are The Baseline DART Tries To Move Past

Existing red-teaming tools check agent output against known attack patterns and static prompt sets.

That approach can miss new strategies, does not fully exercise agent behavior across tools and memory, and often enters the development process late through outside consulting engagements.

When serious findings arrive days before launch, developers may have little time to change code, policies or deployment plans.

DART changes the control path by running waves of attacks that learn across turns, episodes and engagements.

The system evaluates the target response, changes its approach and retries for as many rounds as the user specifies.

It can begin testing without a predefined vulnerability list, which is the main distinction from fixed-script prompt libraries.

Benchmark Result Gives The Release Its Evidence Point

Vijil cites a DecodingTrust-Agent benchmark evaluation in which DART achieved an attack success rate 1.5 times that of its closest competitor.

The system surpassed that competitor in nine of twelve enterprise agent tasks.

The evaluated tasks covered CRM, software coding, support, healthcare, research workflows and travel scenarios.

Those figures should be read as vendor-provided benchmark evidence, not an independent security certification.

Still, they define the claim Vijil is making: adaptive red teaming can uncover more agent weaknesses than static prompt testing when enterprise tasks require multi-step behavior.

Vin Sharma, Vijil’s CEO, framed the issue around agents that can reach confidential data and take consequential action.

He presented DART as a way to reduce long testing cycles and large engagement costs when teams compare it with consultant-led red teaming, broad benchmarks or prototype toolkits.

DART is part of Vijil Diamond inside the broader Vijil platform.

After DART identifies weaknesses, other modules perform root-cause analysis, apply policy-driven guardrails and suggest code changes.

The surrounding platform also includes Discover for finding and fingerprinting agents, Dome for production policy compliance and Darwin for continuous improvement as users, models and attack methods change.

Share this article
inXf

Related articles

More
Cybersecurity M&A Shows Banks Preparing For Machine-Identity Risk
Cybersecurity

Cybersecurity M&A Shows Banks Preparing For Machine-Identity Risk

PYMNTS reported that cybersecurity acquisitions are concentrating on AI security, machine identities, browsers, industrial systems and fraud signals as the enterprise attack surface expands.

Caliptra Hardware Trust Work Shifts From Standard To Deployment
Cybersecurity

Caliptra Hardware Trust Work Shifts From Standard To Deployment

A Semiconductor Engineering article says Caliptra can align hardware trust for data-center devices, but production systems still need lifecycle controls, attestation links, cryptographic agility and SoC-wide security orchestration.

TONTOU CPU Attack Tests Spectre Defenses On Linux Systems
Cybersecurity

TONTOU CPU Attack Tests Spectre Defenses On Linux Systems

Researchers showed a Time-of-Neutralization to Time-of-Use technique that can repollute branch prediction state after Spectre v2 mitigations and leak Linux kernel data in lab tests.

Anthropic Program Pairs Claude With Infrastructure Security Teams
AI

Anthropic Program Pairs Claude With Infrastructure Security Teams

Anthropic is pairing Claude models, its engineers and outside cybersecurity firms to scan critical infrastructure and open-source software for vulnerabilities, with an opt-in service for maintainers.

Atlassian Warns Data Centre Admins To Patch Critical File Access Flaw
Cybersecurity

Atlassian Warns Data Centre Admins To Patch Critical File Access Flaw

Atlassian is urging Data Centre customers to patch CVE-2026-21589, a critical flaw that can let unauthenticated attackers read specific web-root files.

Google Freezes OSS Bug Bounty Reports After AI Submission Flood
Cybersecurity

Google Freezes OSS Bug Bounty Reports After AI Submission Flood

Google has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.

Markey Bill Would Shift AI Hack Reviews To Federal Board
Cybersecurity

Markey Bill Would Shift AI Hack Reviews To Federal Board

Sen. Ed Markey’s bill would create a Cybersecurity and AI Board of Investigations for AI agent-led hacks, with subpoena authority and a mandate covering federal systems and critical infrastructure.

PLDT Pitches Integrated Cloud, Security and Connectivity Strategy for Enterprise Clients
Telco & Connectivity

PLDT Pitches Integrated Cloud, Security and Connectivity Strategy for Enterprise Clients

PLDT is positioning connectivity, wireless, cloud, cybersecurity and digital services as one enterprise relationship as customers modernize around AI, resilience and procurement priorities.

Keep Reading

More Stories

Latest
UK Fibre Deals Face Split Tests As Nexfibre And BT Reviews DivergePoliticsOct 11, 2026UK Fibre Deals Face Split Tests As Nexfibre And BT Reviews DivergeCapacity Media reported that UK regulators are applying competition and public-interest tests to separate broadband consolidation deals involving nexfibre, Netomnia, BT and TalkTalk.Enveda Raises $311M To Push AI-Discovered Medicines Deeper Into TrialsAIOct 11, 2026Enveda Raises $311M To Push AI-Discovered Medicines Deeper Into TrialsFrontier Enterprise reported that Enveda closed a $311 million round led by Catalio Capital Management after two early clinical readouts for PRISM-discovered medicines.Indian Startup Funding Falls To $122.9 Million Despite More DealsFintech & Digital PaymentsOct 11, 2026Indian Startup Funding Falls To $122.9 Million Despite More DealsIndian startups raised $122.9 million across 25 deals in the first week of October, led by DailyObjects, Lumio, Beyond Appliances and StockGro, while IPO and acquisition moves kept the pipeline active.Oxide Raises $445M to Scale Rack-Level Cloud HardwareChips & SemiconductorsOct 11, 2026Oxide Raises $445M to Scale Rack-Level Cloud HardwareOxide Computer raised $445 million in Series D funding led by Eclipse Capital as demand for its pre-integrated data centre racks exceeds supply and the company prepares GPU-capable hardware upgrades.IBM and Red Hat Backport Fixes for 400-Plus Open Source BugsCybersecurityOct 11, 2026IBM and Red Hat Backport Fixes for 400-Plus Open Source BugsIBM and Red Hat say Lightwell has remediated more than 400 previously unknown vulnerabilities in Java libraries, while the new Clearinghouse gives customers a way to submit dependencies for priority review and fixes.Upscale AI Pairs Nvidia Spectrum-X With Its Own SkyHammer FabricChips & SemiconductorsOct 11, 2026Upscale AI Pairs Nvidia Spectrum-X With Its Own SkyHammer FabricUpscale AI is building SkyHammer as a scale-up fabric for AI clusters while using Nvidia Spectrum-X for scale-out switches, a strategy that tests whether Ethernet-based designs can challenge proprietary accelerator domains.Anthropic Opens Free AI Vulnerability Scanner For Open SourceCybersecurityOct 11, 2026Anthropic Opens Free AI Vulnerability Scanner For Open SourceAnthropic is offering open-source projects free AI security scans, with model-generated reports that may speed vulnerability checks but arrive without human triage.Vatar Raises $500,000 After Lagos Life Browser Game Surges to 4.3 Million UsersAIOct 11, 2026Vatar Raises $500,000 After Lagos Life Browser Game Surges to 4.3 Million UsersVatar Inc. has raised a $500,000 angel round at a $10 million valuation after Lagos Life reached 4.3 million registered users, giving the young Nigerian browser-game company capital for product, marketing and hiring.ABC Shareholders Seek Board Seats After South Africa Market SanctionsPoliticsOct 10, 2026ABC Shareholders Seek Board Seats After South Africa Market SanctionsShareholders holding about 76% of Africa Bitcoin Corporation want a meeting to appoint two non-executive directors after South Africa's FSCA sanctioned three former Altvest executives.Morocco King Defends Spain Partnership After Ceuta Migrant RushPoliticsOct 10, 2026Morocco King Defends Spain Partnership After Ceuta Migrant RushKing Mohammed VI said Morocco’s partnership with Spain remains a sovereign choice after more than 70,000 migrants crossed into Ceuta, while promising partners a strategic vision for co-development and stability.Atlassian AMP Targets AI Code Attribution Across Enterprise WorkflowsAIOct 10, 2026Atlassian AMP Targets AI Code Attribution Across Enterprise WorkflowsAtlassian’s Agentic Multiplayer Protocol links agent identity, code attribution, Rovo Work oversight and EU-hosted inference controls to help enterprises track mixed human and AI software work.Unpatched AhsayCBS Flaws Used to Deploy Webshells and Crypto MinersCybersecurityOct 10, 2026Unpatched AhsayCBS Flaws Used to Deploy Webshells and Crypto MinersThreat actors are chaining two AhsayCBS vulnerabilities to bypass authentication, execute commands, install webshells and hide XMRig mining activity on backup management servers.