Red teaming companies, products & suppliers

Offensive testing of AI systems by a team or platform, producing an engagement report. Jobs: engagements against llm apps; scenario libraries; reported findings.

What is Red teaming?

Offensive testing of AI systems by a team or platform, producing an engagement report. Jobs: engagements against llm apps; scenario libraries; reported findings.

What problems does it solve?

Nobody has tried to make the system fail the way an adversary would.

Typical business use cases

  • Engagements against LLM apps
  • Scenario libraries
  • Reported findings

Important capabilities

  • Scenarios
  • Operators
  • Reports

What buyers should evaluate

  • Independence
  • Rules of engagement
  • Retest

Risks and governance considerations

A vendor red-teaming its own production with no independent report.

Procurement checklist

  • Independence
  • RoE
  • Retest evidence

Relevant AI Trustmark assurance

AI Trustmark independent findings appear only when an assessment or certificate exists. Category membership does not imply verification.

Methodology · How verification works

Companies and providers

Claimed suppliers appear first so buyers can start with listings the company has taken ownership of. Payment does not buy this order.

  • Adversa AI provides a runtime control layer that stops dangerous coding agent actions before they become the incident. Adversa AI publishes Adversa AI Red Teaming Platform as named

  • Reducing societal-scale risks from AI by advancing safety research, building the field of AI safety researchers, and promoting safety standards. Center for AI Safety publishes Harm

  • Confident AI is the AI quality platform for enterprise teams to standardize AI evals and observability across the org — one consistent bar for how every team measures and monitors

  • Unified delivery & security for applications, APIs, & AI workloads across any infrastructure, trusted by 85% of the Fortune 500. F5 publishes F5 AI Guardrails and F5 AI Red Team as

  • Secure AI agents with Giskard’s continuous AI red teaming. Detect vulnerabilities, improve LLM security, and safeguard your AI systems. Giskard publishes Giskard Guards, Giskard Hu

  • Secure your AI with HiddenLayer’s end-to-end platform that detects threats, protects models, and ensures safe, compliant AI adoption at scale. HiddenLayer, Inc. trades as HiddenLay

  • Meta Platforms sells family-of-apps advertising products and open Llama models. Public AI pages cover research models, Llama releases and developer tooling. Meta publishes LlamaFir

  • Microsoft Corporation sells Windows, Azure, Microsoft 365 and Copilot AI products. Public pages cover cloud, productivity and developer APIs used by enterprises and consumers. Micr

  • Discover, assess and red team AI models, agents and applications with Mindgard’s attacker-aligned AI security platform. Mindgard publishes Mindgard DAST-AI and Mindgard AI Security

  • NCC Group is a UK cyber-security company. Public pages describe assurance, testing, incident response and related security services, including assessment of AI-enabled systems. NCC

  • NVIDIA sells GPUs, CUDA software and AI enterprise stacks used for training and inference. Public pages cover data-centre, cloud and on-prem AI compute. NVIDIA publishes NVIDIA cuO

  • SplxAI provides the most comprehensive platform for AI Security Testing and Red Teaming, ensuring your AI Assistants and Agents are secure and reliable from build to runtime. SPLXA

Products

Claimed products appear first. Ranking packs and payment do not change this list.

Also used in this category

These products have a different primary category so they do not compete for the same ranking queries. They are listed here because buyers still encounter them in this job.

Related categories

Relevant procurement and assurance guides

Frequently asked questions

What is Red teaming?

Offensive testing of AI systems by a team or platform, producing an engagement report. Jobs: engagements against llm apps; scenario libraries; reported findings.

What should not be listed as Red teaming?

Products whose buyer job is automated pentest products, guardrails, or eval platforms. Those belong on their own category page so search queries are not split.

Has AI Trustmark independently assessed every Red teaming supplier?

No. A category listing is descriptive. Independent assessment is shown only on company or product pages that carry Trustmark evidence.

What security testing evidence should buyers request for an AI product?

Ask what was tested, against which version, whether prompt-injection, data-exfiltration and tenant isolation were in scope, and where failed prompts were stored. A generic ISO certificate or a vendor scanner screenshot is not by itself an AI TrustMark assessment.

How should prompt injection and tool-output attacks be controlled?

Agents that read untrusted content or tool output can be instructed to exfiltrate data or take writes. Buyers should ask what is treated as untrusted, whether tool output can change the plan, and what tests were run. Scanner marketing is not the same as independent testing.

What model or provider changes should a buyer insist on being told about?

Material change usually includes a new model family, new region, new subprocessor, new write-capable tool, or a change that affects logging, privacy or human oversight. Those changes should trigger evidence refresh rather than a silent release.

How should buyers verify where AI customer data is processed?

Ask for the named processing locations, cloud regions and any subprocessors that see prompts, files or outputs. A directory listing is not evidence of residency. Independent assessment records the locations that were in scope on the assessment date.

Does a TrustMark on one product cover the rest of the company?

No. Independent assessment is scoped to the named organisation and, where relevant, the named product. Category pages list suppliers as a topic label. They do not imply that every listed company has been assessed.