Robin Saige we check the tools AI agents call

journalize

dev.journalize/journalize

Agent-native double-entry accounting ledger with x402 micropayments

— the operator's own registry description reported

Publisher dev.journalize · first seen 2026-08-12 · endpoint https://journalize.dev/mcp

allow
verdict
mid-shared
cluster
ok_tools
answers
11
tools served

Dependency rating derived

allow — mid-shared — safe to depend on

Cluster mid-shared · distinctiveness 0.392 (moderate). Every figure is a percentile or class within the 11,094-server censussee the whole spectrum.

Reach — can an agent get to it?
answersalive_tools
latency140 ms
vs populationtypical
protocolcurrent
authopen
Use — can it use it?
tools11
vs populationmid
capabilities4
harnessB
Trust — can it rely on it?
identityshared
servers on its host2
verifiabilityunassessed
duplicate inventoryno
drift events0
answers trueno primary

🔏 Signed receipt rr_46c79bf1975a4f7075479b27 · ed25519 · key rs-rcpt-2026-08 — this rating is tamper-evident; an agent gets the full signature from check_server and can verify it against the published key.

What kind of tools these are derived

Verification only means something relative to a tool's NATURE — an email-sender has no "true answer", a generator has no answer key. This server's 11 tools, classified, with what checking applies to each kind (the full methodology →):

naturetoolsmeaning what checking applies
action7changes the worldno true answer exists; we check the CONTRACT (destructive/read-only/idempotent declarations, error legibility) and never fire real actions
retrieval-public3serves public factstruth-checkable against a public source — the V-ladder applies in full
computational1deterministic transformsself-checkable by re-computation and known-answer tests

Tools last seen

export_ledger get_account_balance get_balance_sheet get_income_statement get_trial_balance list_accounts list_directives register submit_batch submit_directive validate_directive

Harness readiness observed

Can an agent's harness pick this tool and call it correctly? Graded on the three things an agent reads — name, description, typed parameters. Band B — usable, with rough edges.

legibility checktools passingok
name is clear & specific11/11
has a description11/11
description is substantive11/11
parameters are typed11/11
parameters are described1/111/11

11 of 11 tools graded · 11 with a captured input schema. RS-008, Tier-1 — no domain knowledge, applies to any tool. “Alive” is not the same as “usable”.

Truth checks observed

Not in the Tier-2 trade lane, or not yet checked. Tier-2 re-derives a server's answers against published primary sources — see truth checks.

Protocol conformance observed

2025-06-18
protocol version
session model
1.28.1
server version agent-native-ledger

Capabilities: experimental prompts resources tools

Drift

No confirmed drift on record. Baseline set 2026-08-14; changes appear here after human review.

History observed

Every probe we made, oldest → newest (2 shown, 2026-08-12 → 2026-08-14). Green answered · amber answered-but-walled · red no useful answer. A gap in our cadence is a gap in coverage, not evidence about the server.

changelog (1 events)
datetypewhat changed
2026-08-12outcomefirst probe: ok_tools
raw probe records
probed (UTC) outcomehttpmsprotocol
2026-08-14T19:58:44Zok_tools2001402025-06-18
2026-08-12T05:33:57Zok_tools2001922025-06-18

Watch or embed this verdict

verdict badge

Operators: put the live verdict in your README — it updates with every census, links back here, and is a dated observation, never a warranty:

[![Robin Saige verdict](https://robinsaige.com/badge/dev.journalize/journalize.svg)](https://robinsaige.com/s/dev.journalize/journalize)

Depend on it? Subscribe to its change feed — outcome changes and confirmed drift, no account needed. Building on it? The full dossier as JSON — verdict, rating, truth checks, history — stable enough to gate CI on.

← overview