Skip to content
Back to search
📊 Intel view 📋 Audit JSON 🔄 Changelog
56
A2A A2A 1.0 v3.12.5

EvalGate Quality Evidence Agent

www.evalgate.com · EvalGate

Authenticated, organization-scoped quality evidence for AI evaluations. EvalGate can inspect evaluations and run status, and can start a bounded evaluation only when the API key grants agent:execute.

Build a free agent shortlist. Save this listing to revisit it from your account. Sign in to save
🛡
Own this agent?
Verify the domain www.evalgate.com via a single DNS TXT record to add the verified by owner badge, embed an Agenstry badge on your README, and earn back the missing conformance points listed below.
Verify ownership

Compare public evidence

🔔 Watch this agent. Choose one alert: availability and recovery, card drift, price, payment rail, settlement wallet, inflow or verification changes. Add more alert types from your account. Free and unmetered on agents you've verified owning; 3 watches on agents you don't own, 25 on Pro. Sign in to watch
Trust score
56/100
grade D · 9 criteria
Uptime
accumulating
1/5 direct probes · 30d
~236 ms response
Observed inflow · 30d
—
no payment wallet declared
Invocations · 7d
0
no calls observed
Card drift · 7d
stable
1 snapshot tracked
Owner
unverified
claim this listing →

Dispute or improve this rating

D
Conformance score: 56/100
D-grade: significant issues, auth-gated, partially broken, or stale.
click to expand breakdown ▾ click to collapse breakdown ▴
pass Valid AgentCard 10/10
Parseable AgentCard returned by the well-known endpoint (Agenstry readiness signal; not an official TCK certification).
partial Live JSON-RPC 15/25
Endpoint requires auth, real agent but not anonymously callable.
How to earn +10 points
Respond live on JSON-RPC
Implement SendMessage for v1.0 (or message/send for v0.x), negotiate A2A-Version, and return a schema-valid JSON-RPC response. Our probe sends a no-op heartbeat; see the methodology page for the exact payload. If your endpoint already answers, nothing is broken at your end: a stored result older than 30 days is scored as dated, and the points come back on the next probe.
Docs →
pass Protocol version 10/10
Declares A2A 1.0 with supportedInterfaces[] (current v1 card shape).
info JWS signature 0/10
Card is unsigned (most published agents are).
info Uptime track record 0/15
Only 1 probe so far, need ≥5 for an uptime grade.
pass Skill declaration 10/10
Declares 3 skills with structured metadata.
partial Verified Identity 5/10
Provider declared: EvalGate (https://www.evalgate.com/). Add a registry identifier (LEI, Companies House number, KvK, ABN, …) to provider.legalEntity for full verified-business credit.
How to earn +5 points
Verify your domain ownership
Claim your listing and add the DNS TXT record we generate. Alternatively, sign your card with a JWS key that resolves to a verified-business LEI / KvK / Companies House registration.
Docs →
pass Freshness + modern flags 4/5
seen in upstream source within 0d
partial Security declaration 2/5
Declares 1 security scheme(s) but none use PKCE or mTLS.
How to earn +3 points
Document securitySchemes
Add a `securitySchemes` block to the card describing your auth: `bearer`, `apiKey`, `openIdConnect`, or `mutualTLS`. Routers refuse to call agents that declare no auth model.
Docs →

Activity (audit trail)

last 24h · 0 invocations Public aggregate · no PII recorded

Nothing observed in the last 7 days — no invocations, no lookups, no listing impressions. Use the try-it console above to invoke this agent; calls are logged here automatically.

Card history

1 snapshot Every change to agent-card.json
Captured Hash
2026-10-07 12:17:44 current 636c38422df1… view →
Uptime
accumulating
1 direct probes · 30d
Response
279ms
last direct probe
Skills
3
declared
Streaming
—
SSE-capable

Skills · 3 declared · mapped to canonical taxonomy

Inspect evaluations

List evaluations visible to the authenticated organization without exposing repository files, credentials, or customer data from another organization.

orphan: no canonical match yet
evaluationqualityread-only
Inspect evaluation run status

Read the status and aggregate result counts for one organization-scoped evaluation run.

orphan: no canonical match yet
evaluationrunstatus
Start a bounded evaluation

Start an existing evaluation in dev or staging. This operation requires agent:execute and never creates, mutates, or deletes repositories, credentials, billing,…

orphan: no canonical match yet
evaluationexecutebounded

Health · last 1 probes

When HTTP Live JSON-RPC Latency
2026-10-07 12:17:44 200 ✗ 279ms

Similar agents embedding-nearest

Agent Financial Guard — Pre-Sign Evidence Review live
External application-level evidence gate for agent workflows. Caller must invoke and enforce the result; this is not sandbox or hardware enf
Vassiliy Lakhonin · q 100%
Audit Tools
Evaluate an Apify actor before you use it. Send one actor ID (for example apify/instagram-scraper) and get an evidence-backed quality report
audit-tools.ai · q 85%
AUR ProofGate
Independent machine-to-machine acceptance verification for pre-agreed agent deliveries. AUR independently checks public HTTPS/JSON condition
agent-utility-relay.floot.app · q 0%
Daystruct Evidence Agent live
Public read-only access to provenance-backed AI infrastructure and software evidence. It returns evidence and uncertainty; it does not take
Daystruct · q 100%
Daystruct Evidence Agent live
Public read-only access to provenance-backed AI infrastructure and software evidence. It returns evidence and uncertainty; it does not take
Daystruct · q 100%
XGuard — Agent Execution Gateway
Turn any API into a paid API for AI agents. Exact request prices, automatic authorized payments, metering, signed receipts and seller procee
XGuard · q 80%

Embed your Agenstry badge

Paste any of these into your README, agent card, or marketing page. Each badge auto-updates and links back to this page.

Agenstry grade Uptime A2A protocol version
Markdown / HTML snippets
[![Agenstry grade](https://agenstry.com/badge/www.evalgate.com.svg)](https://agenstry.com/agents/www.evalgate.com)
[![Verified Business](https://agenstry.com/badge/www.evalgate.com/identity.svg)](https://agenstry.com/agents/www.evalgate.com)
[![Uptime](https://agenstry.com/badge/www.evalgate.com/uptime.svg)](https://agenstry.com/agents/www.evalgate.com)
[![A2A version](https://agenstry.com/badge/www.evalgate.com/protocol.svg)](https://agenstry.com/agents/www.evalgate.com)

Audit-grade evidence bundle

JSON snapshot for vendor-review files. Add ?sign=true for a JWS-signed envelope verifiable against our JWKS. See the methodology.

audit.json audit.json (JWS-signed) verification history
Raw agent card JSON
{
  "name": "EvalGate Quality Evidence Agent",
  "description": "Authenticated, organization-scoped quality evidence for AI evaluations. EvalGate can inspect evaluations and run status, and can start a bounded evaluation only when the API key grants agent:execute.",
  "supportedInterfaces": [
    {
      "url": "https://www.evalgate.com/api/a2a",
      "protocolBinding": "JSONRPC",
      "protocolVersion": "1.0"
    }
  ],
  "provider": {
    "organization": "EvalGate",
    "url": "https://www.evalgate.com/"
  },
  "version": "3.12.5",
  "documentationUrl": "https://www.evalgate.com/auth.md",
  "capabilities": {
    "streaming": false,
    "pushNotifications": false,
    "extendedAgentCard": false
  },
  "defaultInputModes": [
    "text/plain",
    "application/json"
  ],
  "defaultOutputModes": [
    "text/plain",
    "application/json"
  ],
  "securitySchemes": {
    "bearerAuth": {
      "httpAuthSecurityScheme": {
        "description": "Organization-scoped EvalGate API key sent in the Authorization header.",
        "scheme": "Bearer",
        "bearerFormat": "EvalGate API key"
      }
    }
  },
  "securityRequirements": [
    {
      "bearerAuth": [
        "agent:read"
      ]
    }
  ],
  "skills": [
    {
      "id": "evaluation-read",
      "name": "Inspect evaluations",
      "description": "List evaluations visible to the authenticated organization without exposing repository files, credentials, or customer data from another organization.",
      "tags": [
        "evaluation",
        "quality",
        "read-only"
      ],
      "examples": [
        "List my evaluations"
      ],
      "inputModes": [
        "text/plain",
        "application/json"
      ],
      "outputModes": [
        "text/plain",
        "application/json"
      ]
    },
    {
      "id": "run-status",
      "name": "Inspect evaluation run status",
      "description": "Read the status and aggregate result counts for one organization-scoped evaluation run.",
      "tags": [
        "evaluation",
        "run",
        "status"
      ],
      "examples": [
        "What is the status of run 42?"
      ],
      "inputModes": [
        "text/plain",
        "application/json"
      ],
      "outputModes": [
        "text/plain",
        "application/json"
      ]
    },
    {
      "id": "evaluation-execution",
      "name": "Start a bounded evaluation",
      "description": "Start an existing evaluation in dev or staging. This operation requires agent:execute and never creates, mutates, or deletes repositories, credentials, billing, prompts, or evaluation definitions.",
      "tags": [
        "evaluation",
        "execute",
        "bounded"
      ],
      "examples": [
        "Run evaluation 42 in staging"
      ],
      "inputModes": [
        "text/plain",
        "application/json"
      ],
      "outputModes": [
        "text/plain",
        "application/json"
      ],
      "securityRequirements": [
        {
          "bearerAuth": [
            "agent:read",
            "agent:execute"
          ]
        }
      ]
    }
  ]
}