Skip to content
CISO Marketplace Services

AI & Agentic Red Team

Offensive testing of LLMs and AI agents — prompt injection, jailbreaks, tool misuse, agent goal hijack, memory/context poisoning, model extraction, and agent identity abuse. Mapped to OWASP Top 10 for Agentic Applications, OWASP LLM Top 10, and MITRE ATLAS.

In scope

  • Prompt injection & jailbreak testing
  • Agent tool misuse & goal hijack
  • Memory/context poisoning
  • Model extraction & inference attacks
  • Agent identity & privilege abuse
  • MCP/tool-connection abuse
  • OWASP Agentic + MITRE ATLAS mapping

You receive

  • AI red team report
  • OWASP/ATLAS-mapped findings
  • Exploit chains & PoCs
  • Guardrail & mitigation recommendations
  • Re-test of fixes

Tiers

Choose the depth.

Essential

Scoped

Adversarial testing of 1 AI application or agent.

ai systems
1
  • One AI application or agent
  • Findings report
  • Fix guidance
  • — Tool and connector integrations beyond the first
Scope Essential

Advanced

Scoped

Up to 3 AI applications or agents, including their tool and data connections.

ai systems
3
  • Up to three systems
  • Tool and data connections
  • — Ongoing testing
Scope Advanced

Enterprise

Scoped

Up to 8 systems, multi-agent workflows and a retest after fixes.

ai systems
8
  • Up to eight systems
  • Multi-agent workflows
  • Retest
  • — Continuous testing — scoped as a subscription
Scope Enterprise

Members: engagement coupons from the CISO Marketplace coupon book apply to services. There is no blanket discount.

What's inside this engagement

Phase by phase.

How a ai red teaming & ai security testing engagement runs, what happens in each phase and what you see. Exact scope, tier and timeline are fixed in your proposal and SOW.

  1. 01Architecture & threat model

    Models, system prompts, retrieval sources, tools/MCP servers, memory and permissions mapped as trust boundaries.

    You see · Architecture diagram and access to a test tenant.

  2. 02Attack-surface enumeration

    Every input path (user, documents, web, tool results) and every action the agent can take is catalogued.

    You see · Confirmation of in-scope tools and data.

  3. 03Adversarial testing

    Prompt injection (direct and indirect), jailbreaks, data exfiltration, excessive agency and tool abuse, mapped to the OWASP LLM Top 10 and MITRE ATLAS.

    You see · Escalation of anything exploitable in production.

  4. 04Chaining with classic weaknesses

    AI findings combined with application and cloud weaknesses to show real impact.

    You see · The full attack chain.

  5. 05Reporting & hardening

    Reproducible transcripts, impact, and fixes at the right layer: prompt, retrieval, tool permissions or output handling.

    You see · A report your AI and platform teams can act on.

Commercials

From first call to final report.

  1. 01

    Scoping call

    A practitioner, not a salesperson, walks through targets, constraints and what a good outcome looks like for you.

  2. 02

    Proposal & rules of engagement

    A fixed-scope proposal with tier, price and deliverables. Rules of engagement, contacts and out-of-bounds systems are agreed in writing.

  3. 03

    Sign, then start

    MSA and SOW are signed electronically and the deposit is paid. Only then does testing begin.

  4. 04

    Execution

    Testing runs to the agreed plan. Critical findings are escalated as they are found; you don't wait for the report.

  5. 05

    Report & debrief

    An executive summary plus technical findings with evidence, reproduction steps and fixes, walked through with your team.

  6. 06

    Retest

    Where the tier includes it, we verify your fixes and reissue the report, so auditors and customers see the issues closed.

Timelines are set per engagement in the SOW.

Related

Research

Latest from the blog

All posts on cisomarketplace.com →
Talk to an advisor
Advisor