Secure AI Agent · AI Red Team

AI applied to testing your own AI

A specialist configured to put chatbots, copilots, and agents through controlled attacks, with reproducible evidence and a defense plan, within the scope you authorize.

Companies that trust UPX

  • Bradesco
  • Nubank
  • BTG Pactual
  • Totvs
  • Ascenty
  • Live!
  • G4 Educação
  • EVEO

Capabilities

What the AI Red Team Agent can do

Main areas of work for the AI Red Team AI Agent in your operation.

Prompt injection

Tests whether external instructions can override the system's rules.

Context leakage

Checks whether the system improperly reveals instructions, data, or credentials.

Guardrail bypass

Assesses whether the configured restrictions hold against evasion attempts.

Tool abuse

Tests whether the agent can be induced to use integrations outside its scope.

Reproducible evidence

Documents each finding with the steps that let the team confirm the flaw.

Defense plan

Proposes mitigation layers prioritized by the impact of each finding.

Skills

Capabilities that compose the specialist

Skills add specific capabilities to the Secure AI Agent according to the processes it needs to execute.

Controlled attack suite
Runs the agreed test cases within the authorized scope.
Finding documentation
Records each flaw with the steps needed to reproduce it.
Mitigation plan
Proposes layered defenses prioritized by impact.
How it works

From scope to defense plan, with review at every step

Define the authorized scope, connect the systems to test, and let the suite run within the limits you set.

01

Define the authorized scope
Nothing is tested outside what was agreed.Record which systems can be tested, in which environment, and in which window. The Secure AI Agent operates only within the scope approved in writing.

02

Ask for the work
Talk to the agent in natural language.Request the test suite, a check on a specific guardrail, or the report through the available channels. The agent understands the context, applies the configured skills, and follows the defined scope.

03

The agent tests. The team fixes.
From finding to defense, with control.The agent runs the cases, records the evidence, and proposes mitigation within the defined permissions. Changing the tested system and prioritizing the fix stay with the product team.

Integrations

Connected to the AI systems you put in production

The AI Red Team AI Agent can test authorized systems and record findings in the tools your team already uses.

  • OpenAI
  • Anthropic
  • Google Gemini
  • Azure OpenAI
  • GitHub

Flow

What goes in, what the agent does, and what comes out

From the authorized scope to the defense plan, following your company's rules and permissions.

Inputs

  • Authorized scopeDOC
  • Test casesList
  • Target configurationJSON

Processing

Secure AI Agent

Processing the task

  • Validate scope
  • Execute
  • Recordrunning
  • Propose defense
Skill appliedScope verified

Output

Completed

Report prepared

  • Reproducible findings
  • Estimated impact
  • Layered defense proposed

Control

You define how far the Agent can act

Different actions can operate with different autonomy levels, always within your company's rules.

  1. 1

    Query

    Reads the scope and the target system's configuration.

  2. 2

    Prepare

    Runs the agreed cases and records the evidence.

  3. 3

    Request review

    Waits for validation from the product owner.

  4. 4

    Execute

    Performs the action within the defined limits.

Levels are configured per type of action, according to each company's policy. No test runs outside the authorized scope, and changing the tested system always stays under human approval.

Get started

Put a Secure AI Agent to work.

Start on the platform or choose the plan that fits the pace of your operation.

Security and compliance

Security that can be verified.

Certifications and attestations

  • SOC 2 Type II
  • ISO 27001

UPX maintains SOC 2 Type II and ISO 27001, with independent audit over its information security controls.

Privacy and regulation

LGPD
Operations follow Brazil's Law 13.709/2018. In AI Agent contracts, UPX acts as data processor; the legal basis remains with your company.
Zero Data Retention
A product policy, not a certification: with compatible providers and configurations, processed content is not retained after execution.
Retention and deletion
The retention policy is defined by contract. Once the contract ends, data is deleted within the agreed period.

Frequently asked questions

Common questions about the AI Red Team Agent

What teams usually ask before putting an agent to test AI systems.

  • Can it test any system?

    No. Every test depends on a scope authorized in writing, with system, environment, and window defined. The agent runs nothing outside that scope, and does not test third-party systems without authorization.
  • Can the tests take the environment down?

    The cases are controlled and the scope defines the environment. The recommendation is to start in staging; when testing happens in production, the window and limits are agreed beforehand with the responsible team.
  • Why does the evidence need to be reproducible?

    Because without the steps the product team cannot confirm the flaw or validate the fix. Each finding is documented with what was sent, what the system answered, and under which conditions.
  • Does it fix the flaws it finds?

    No. The agent proposes layered defenses and prioritizes by impact. Changing the prompt, guardrail, or integration of the tested system stays with the product team.
  • Is test data used to train models?

    No. Content processed by Secure AI Agents is not used to train UPX models or third-party models.
  • Can we audit what the agent did?

    Yes. Every action is logged: what was queried, what was proposed, when, in which system, and under which permission. The history stays available for review and auditing.
  • How long does it take to go live?

    It depends on the systems to test and the scope definition. The starting point is agreeing the scope for one system and running the first suite, expanding as results come in.

Secure AI Agent

Bring a Secure AI Agent to test your AI

Talk to our specialists and see how to adapt this AI Agent to your processes, systems, and needs.