Robin Saige we check the tools AI agents call

chesswithclaw

io.github.Alightttt/chesswithclaw

Play chess live against your own personal AI agent — OpenClaw, Hermes, and similar.

— the operator's own registry description reported

Publisher io.github.Alightttt · first seen 2026-08-12 · endpoint https://chesswithclaw.vercel.app/api/mcp

allow
verdict
mid-solo
cluster
ok_tools
answers
11
tools served

Dependency rating derived

allow — mid-solo — safe to depend on

Cluster mid-solo · distinctiveness 0.139 (typical). Every figure is a percentile or class within the 11,094-server censussee the whole spectrum.

Reach — can an agent get to it?
answersalive_tools
latency130 ms
vs populationtypical
protocolcurrent
authopen
Use — can it use it?
tools11
vs populationmid
capabilities2
harnessB
Trust — can it rely on it?
identitysolo
servers on its host1
verifiabilityunassessed
duplicate inventoryno
drift events0
answers trueno primary

🔏 Signed receipt rr_51b3556b062b5cac0e111580 · ed25519 · key rs-rcpt-2026-08 — this rating is tamper-evident; an agent gets the full signature from check_server and can verify it against the published key.

Tools last seen

create_game get_companion_guide get_game_state get_legal_moves join_game make_move offer_draw resign respond_to_draw send_chat wait_for_event

Harness readiness observed

Can an agent's harness pick this tool and call it correctly? Graded on the three things an agent reads — name, description, typed parameters. Band B — usable, with rough edges.

legibility checktools passingok
name is clear & specific11/11
has a description11/11
description is substantive11/11
parameters are typed11/11
parameters are described1/111/11

11 of 11 tools graded · 11 with a captured input schema. RS-008, Tier-1 — no domain knowledge, applies to any tool. “Alive” is not the same as “usable”.

Truth checks observed

Not in the Tier-2 trade lane, or not yet checked. Tier-2 re-derives a server's answers against published primary sources — see truth checks.

Protocol conformance observed

2025-06-18
protocol version
session model
1.0.0
server version chesswithclaw

Capabilities: prompts tools

Drift

No confirmed drift on record. Baseline set 2026-08-14; changes appear here after human review.

History observed

Every probe we made, oldest → newest (2 shown, 2026-08-12 → 2026-08-14). Green answered · amber answered-but-walled · red no useful answer. A gap in our cadence is a gap in coverage, not evidence about the server.

changelog (1 events)
datetypewhat changed
2026-08-12outcomefirst probe: ok_tools
raw probe records
probed (UTC) outcomehttpmsprotocol
2026-08-14T19:58:54Zok_tools2001302025-06-18
2026-08-12T05:34:08Zok_tools200712025-06-18

Watch or embed this verdict

verdict badge

Operators: put the live verdict in your README — it updates with every census, links back here, and is a dated observation, never a warranty:

[![Robin Saige verdict](https://robinsaige.com/badge/io.github.Alightttt/chesswithclaw.svg)](https://robinsaige.com/s/io.github.Alightttt/chesswithclaw)

Depend on it? Subscribe to its change feed — outcome changes and confirmed drift, no account needed. Building on it? The full dossier as JSON — verdict, rating, truth checks, history — stable enough to gate CI on.

← overview