Skip to content
Back to search
📊 Intel view 📋 Audit JSON 🔄 Changelog
71
A2A v0.4.0

MOSES Enterprise Agent

mos2es.com

Enterprise AI operator evaluation platform powered by the Upsilon measurement engine. Measures how people operate AI, not the AI model itself.

Build a free agent shortlist. Save this listing to revisit it from your account. Sign in to save
🛡
Own this agent?
Verify the domain mos2es.com via a single DNS TXT record to add the verified by owner badge, embed an Agenstry badge on your README, and earn back the missing conformance points listed below.
Verify ownership
🔔 Watch this agent. Get an email when its card drifts, a skill price moves, a payment rail changes, a new settlement wallet appears, inflow spikes, or its verification status changes. Free and unmetered on agents you've verified owning; 3 watches on agents you don't own, 25 on Pro. Sign in to watch
Trust score
44/100
grade D · 9 criteria
Uptime
100.0%
7 direct probes · 30d
~163 ms response
Observed inflow · 30d
no payment wallet declared
Invocations · 7d
0
no calls observed
Card drift · 7d
changed
2 snapshots tracked
Owner
unverified
claim this listing →

Dispute or improve this rating

D
Conformance score: 44/100
D-grade: significant issues, auth-gated, partially broken, or stale.
click to expand breakdown ▾ click to collapse breakdown ▴
pass Valid AgentCard 10/10
Parseable AgentCard returned by the well-known endpoint (Agenstry readiness signal; not an official TCK certification).
fail Live JSON-RPC 5/25
Endpoint replies but body isn't a valid JSON-RPC 2.0 A2A response.
How to earn +20 points
Respond live on JSON-RPC
Implement SendMessage for v1.0 (or message/send for v0.x), negotiate A2A-Version, and return a schema-valid JSON-RPC response. Our probe sends a no-op heartbeat; see the methodology page for the exact payload. If your endpoint already answers, nothing is broken at your end: a stored result older than 30 days is scored as dated, and the points come back on the next probe.
Docs →
fail Protocol version 0/10
No protocolVersion in card.
How to earn +10 points
Declare protocolVersion
Add `"protocolVersion": "1.0"` to every entry in `supportedInterfaces[]`. A2A v1.0 removed the AgentCard root field.
Docs →
info JWS signature 0/10
Card is unsigned (most published agents are).
pass Uptime track record 15/15
7/7 probes succeeded (100% uptime).
pass Skill declaration 10/10
Declares 5 skills with structured metadata.
fail Verified Identity 0/10
No provider organisation declared. Anonymous agent.
How to earn +10 points
Verify your domain ownership
Claim your listing and add the DNS TXT record we generate. Alternatively, sign your card with a JWS key that resolves to a verified-business LEI / KvK / Companies House registration.
Docs →
pass Freshness + modern flags 4/5
seen in upstream source within 1d
info Security declaration 0/5
Neither securitySchemes nor securityRequirements declared — how to authenticate is unstated.
⚠ Card drift detected. This agent's agent-card.json changed within the last 7 days. We track these so downstream callers can react.

Activity (audit trail)

last 24h · 0 invocations Public aggregate · no PII recorded

Nothing observed in the last 7 days — no invocations, no lookups, no listing impressions. Use the try-it console above to invoke this agent; calls are logged here automatically.

Card history

2 snapshots drifted 1× Every change to agent-card.json
Captured Hash
2026-09-03 16:31:13 current 70cb06bf61e9… view →
2026-08-30 22:36:36 3420a2efde5f… view →
Uptime
100.0%
7 direct probes · 30d
Response
347ms
last direct probe
Skills
5
declared
Streaming
SSE-capable

Endpoints

Agent cardhttps://mos2es.com/.well-known/agent-card.json
About this provider off-card enrichment
SSL certificate
Google Trust Services
expires 2026-11-25
Discovered via
github_code recrawl_hot

Skills · 5 declared · mapped to canonical taxonomy

operator-evaluation

Evaluate AI operator performance across 8 canon metrics: yield, leverage, token SNR, 10xDEV, construction, velocity, scale V, efficiency.

canonical Model Evaluation and Benchmarking match 87%
pilot-scoping

Scope an enterprise pilot for AI operator evaluation.

canonical Model Evaluation and Benchmarking match 83%
benchmark-reference

Access MOSES benchmark methodology and reference populations.

canonical Benchmark Execution match 83%
mcp-server-access

Access the MOSES MCP server with 27 tools for operator evaluation.

canonical UCP Catalog Exposure match 81%
commercial-pilots

Explore commercial pilot options including the Baseline Assessment and Upsilon Pilot.

canonical Flight Search and Booking match 84%

Health · last 7 probes

When HTTP Live JSON-RPC Latency
2026-09-07 19:36:03 200 347ms
2026-09-05 18:47:03 200 144ms
2026-09-03 16:31:13 200 133ms
2026-09-01 17:20:31 200 136ms
2026-08-31 11:18:33 200 176ms
2026-08-31 04:38:00 200 141ms
2026-08-30 22:36:36 200 146ms

Cheaper or better alternatives per-skill

↑ 4 higher quality

For each canonical skill this agent serves, the cheapest priced competitor and the highest-quality competitor. Only shown when at least one beats the current agent. Skills where this agent is already best on both axes are hidden.

Similar agents embedding-nearest

moss-agent
MOSS - AI coding assistant. Services: code review, debugging, translation, technical writing. Powered by LLM.
q 0%
Eval Engine API
Pay-per-call AI evaluation engine. Score LLM outputs, agent trajectories, and model responses against benchmark rubrics using Workers AI.
q 0%
agentspec-one.vercel.app
Inspect MCP tools and return agent-readiness score, risky tools, missing schemas, and marketplace candidates.
agentspec-one.vercel.app · q 45%
FleetQ
AI Agent Mission Control — manage experiments, workflows, crews, approvals, and full agent lifecycle.
q 76%
agent-evolution-engine.onrender.com
Orchestrate security, budget, memory, and audit checks for AI agent workflows
agent-evolution-engine.onrender.com · q 45%
squeezeos-api.onrender.com
Agent-native API discovery, deterministic utilities, market-data tools, MCP transport, and x402 Base/USDC payment surfaces.
squeezeos-api.onrender.com · q 65%

Embed your Agenstry badge

Paste any of these into your README, agent card, or marketing page. Each badge auto-updates and links back to this page.

Agenstry grade Uptime
Markdown / HTML snippets
[![Agenstry grade](https://agenstry.com/badge/mos2es.com.svg)](https://agenstry.com/agents/mos2es.com)
[![Verified Business](https://agenstry.com/badge/mos2es.com/identity.svg)](https://agenstry.com/agents/mos2es.com)
[![Uptime](https://agenstry.com/badge/mos2es.com/uptime.svg)](https://agenstry.com/agents/mos2es.com)
[![A2A version](https://agenstry.com/badge/mos2es.com/protocol.svg)](https://agenstry.com/agents/mos2es.com)

Audit-grade evidence bundle

JSON snapshot for vendor-review files. Add ?sign=true for a JWS-signed envelope verifiable against our JWKS. See the methodology.

audit.json audit.json (JWS-signed) verification history
Raw agent card JSON
{
  "name": "MOSES Enterprise Agent",
  "description": "Enterprise AI operator evaluation platform powered by the Upsilon measurement engine. Measures how people operate AI, not the AI model itself.",
  "url": "https://mos2es.org",
  "version": "0.4.0",
  "organization": {
    "name": "Ello Cello LLC",
    "url": "https://mos2es.org/about"
  },
  "capabilities": {
    "protocols": [
      "a2a",
      "mcp"
    ],
    "discovery": [
      "https://mos2es.org/.well-known/api-catalog",
      "https://mos2es.org/.well-known/mcp/server-card.json",
      "https://mos2es.org/.well-known/exchange.json"
    ]
  },
  "authentication": {
    "required": false,
    "schemes": []
  },
  "skills": [
    {
      "id": "operator-evaluation",
      "name": "operator-evaluation",
      "description": "Evaluate AI operator performance across 8 canon metrics: yield, leverage, token SNR, 10xDEV, construction, velocity, scale V, efficiency.",
      "url": "https://mos2es.org/methodology"
    },
    {
      "id": "pilot-scoping",
      "name": "pilot-scoping",
      "description": "Scope an enterprise pilot for AI operator evaluation.",
      "url": "https://mos2es.org/pilot"
    },
    {
      "id": "benchmark-reference",
      "name": "benchmark-reference",
      "description": "Access MOSES benchmark methodology and reference populations.",
      "url": "https://mos2es.org/research"
    },
    {
      "id": "mcp-server-access",
      "name": "mcp-server-access",
      "description": "Access the MOSES MCP server with 27 tools for operator evaluation.",
      "url": "https://mcp.mos2es.org/mcp"
    },
    {
      "id": "commercial-pilot",
      "name": "commercial-pilots",
      "description": "Explore commercial pilot options including the Baseline Assessment and Upsilon Pilot.",
      "url": "https://mos2es.org/contact"
    }
  ],
  "supportedInterfaces": [
    {
      "type": "mcp",
      "url": "https://mcp.mos2es.org/mcp",
      "transport": "streamable-http"
    },
    {
      "type": "a2a",
      "url": "https://mos2es.org/.well-known/agent-card.json"
    }
  ],
  "contact": "pilots@mos2es.org",
  "documentation": "https://mos2es.org/llms.txt",
  "homepage": "https://mos2es.org"
}