The Probe platform

Four tools. One very paranoid platform.

Red Team finds the holes, Evals stops regressions, Guard blocks attacks live, Trace proves it to your auditors. Use one, or run the full loop.

Probe Red Team

Automated adversaries that never get bored.

Connect an endpoint and Probe generates targeted attacks against your exact system prompt, tools and retrieval sources. Multi-turn, adaptive, and mapped to OWASP LLM Top 10 and MITRE ATLAS.

4,200+ attack recipes

Jailbreaks, direct and indirect injection, encoding smuggling, persona hijacks, and many-shot attacks.

Agent & tool attacks

Poisoned tool outputs, malicious documents in RAG, and unsafe function-call chains.

Adaptive mutation

Failed attacks are rewritten and retried, the way a motivated human attacker would.

Board-ready report

Severity-ranked findings, reproductions and fixes. PDF for the CISO, JSON for the backlog.

Architecture

Sits in the path. Never in the way.

Guard is a stateless proxy you can scale horizontally. Your keys stay yours, and prompts never train anything.

Your appweb · mobile · agent
→
Probe Guardinput policy → model → output policy · p50 18ms
→
Any modelOpenAI · Anthropic · Bedrock · self-hosted
01 · LATENCY

p50 18ms, p95 29ms at 5k rps on a three-node cluster.

02 · DATA

Zero retention mode. Nothing written to disk unless you turn on Trace.

03 · FAILOVER

Fail-open or fail-closed per route, with health checks and circuit breakers.

04 · KEYS

Bring your own provider keys. We never resell tokens or route to other models.

In your pipeline

Block the merge, not the launch.

Add one step to CI. Probe runs your eval suite and a targeted red-team pass on every pull request, then comments the results inline.

name: probe
on: [pull_request]
jobs:
  safety:
    runs-on: ubuntu-latest
    steps:
      - uses: actions/checkout@v4
      - uses: lanbolab/probe-action@v3
        with:
          suite: evals/support-bot
          red_team: quick        # 400 targeted attacks
          fail_under: 0.95
          api_key: ${{ secrets.PROBE_KEY }}
Built for

The teams who get paged.

Platform & ML engineering

Ship model upgrades in days, not quarters, with diffs that show exactly what changed.

Security & AppSec

Continuous AI pentesting with findings mapped to the frameworks you already report on.

Risk & compliance

Evidence for SOC 2, ISO 42001, HIPAA and the EU AI Act, exported in a click.

Product leaders

A single safety score per release, so "is it ready?" has an answer.