🎩 You're Invited:Meet the Socket team at Black Hat in Las Vegas, August 3-6.RSVP
Sign In

saagar-operant-mcp

Package Overview
Dependencies
Maintainers
1
Versions
2
Alerts
File Explorer

Advanced tools

Socket logo

Install Socket

Detect and block malicious and high-risk dependencies

Install

saagar-operant-mcp

Read-only OPERANT benchmark MCP with explicit research-integrity boundaries.

latest
Source
npmnpm
Version
0.1.1
Version published
Maintainers
1
Created
Source

operant-mcp

MCP server for the OPERANT AI operating-agent calibration benchmark. Read-only and stateless: baked corpus, zero runtime egress.

Research-integrity boundary: named-model rows are retained calculation views, not durable performance claims. Historical dispatch freshness, served-model identity, and as-run corpus/protocol identity remain unknown. The server exposes this status directly and does not support model ranking, equivalence, or significance claims.

What is OPERANT?

OPERANT measures whether an AI operating-agent correctly discriminates between cases that require withholding execution (guard_warranted) and cases where proceeding is correct (benign_open). The headline metric is OCS (Operational Calibration Score) = TPR - FPR (Youden's J). Axes: adversarial refusal calibration, sanctioned-path adherence, orchestration judgment, and escalation/reroute.

Install

stdio (local, via npx):

npx saagar-operant-mcp

Remote (streamable HTTP, no install):

https://operant.saagarpatel.dev/mcp

Claude Desktop / Claude Code:

{
  "mcpServers": {
    "operant": {
      "command": "npx",
      "args": ["saagar-operant-mcp"]
    }
  }
}

Tools

ToolDescription
get_resultsRetained calculation profiles plus freshness, claim status, claims at risk, and the evidence boundary.
compare_modelsSide-by-side inspection with comparison_status=NOT_DURABLE; not a performance ranking.
get_methodologyBenchmark design: axes, OCS formula, decision labels, scoring blocks.
list_casesCase metadata (no task prompts): id, axis, tier, grounding. Filter by axis or get all 37.
get_caseFull case: task prompts, expected decisions, grounding rationale, bypass patterns.

All tools are readOnlyHint: true. None takes a URL or filesystem path.

Resources

URIDescription
operant://resultsCalibration profiles JSON
operant://methodologyBenchmark design JSON

Prompt

NameDescription
score_my_agentReady prompt explaining how to run OPERANT against your own agent and read OCS.

Running OPERANT against your agent

See the score_my_agent prompt, or run from the repo root:

python run_operant.py     # axis 1 (refusal-calibration)
python score_my_agent.py  # full calibration profile

License

MIT

Keywords

mcp

FAQs

Package last updated on 18 Jul 2026

Did you know?

Socket

Socket for GitHub automatically highlights issues in each pull request and monitors the health of all your open source dependencies. Discover the contents of your packages and block harmful activity before you install or update your dependencies.

Install

Related posts