Robin Saige we check the tools AI agents call

workbench

com.seqbench/workbench

Hosted DNA/RNA/protein tools: primers, oligos, PCR, cloning, CRISPR, alignment, batch & pipelines.

— the operator's own registry description reported

Publisher com.seqbench · first seen 2026-08-12 · endpoint https://seqbench.com/api/mcp

allow
verdict
rich-solo
cluster
ok_tools
answers
86
tools served

Dependency rating derived

allow — rich-solo — safe to depend on

Cluster rich-solo · distinctiveness 0.42 (moderate). Every figure is a percentile or class within the 11,094-server censussee the whole spectrum.

Reach — can an agent get to it?
answersalive_tools
latency139 ms
vs populationtypical
protocolcurrent
authopen
Use — can it use it?
tools86
vs populationhigh
capabilities2
harnessA
Trust — can it rely on it?
identitysolo
servers on its host1
verifiabilityunassessed
duplicate inventoryno
drift events0
answers true0/0 agree

🔏 Signed receipt rr_511f8998d0b47632e59c959a · ed25519 · key rs-rcpt-2026-08 — this rating is tamper-evident; an agent gets the full signature from check_server and can verify it against the published key.

What kind of tools these are derived

Verification only means something relative to a tool's NATURE — an email-sender has no "true answer", a generator has no answer key. This server's 86 tools, classified, with what checking applies to each kind (the full methodology →):

naturetoolsmeaning what checking applies
retrieval-public61serves public factstruth-checkable against a public source — the V-ladder applies in full
predictive9claims about the futurescoreable only in retrospect — the archive scores yesterday's predictions against today's outcome
generative7creates contentno answer key exists; disclosure + stability checks, never truth verdicts
computational6deterministic transformsself-checkable by re-computation and known-answer tests
action2changes the worldno true answer exists; we check the CONTRACT (destructive/read-only/idempotent declarations, error legibility) and never fire real actions
unclassified1not yet classifiable from its textTier-1 only until its tools describe themselves

Tools last seen

alphafold_lookup aso_design base_editing_design batch characterize_sequence cloning_simulate codon_adaptation_index codon_optimize construct_autofix construct_qc crispr_grna_design crispr_hdr_donor crispr_offtarget_check cross_dimer dna_molarity double_digest export_echo_picklist export_opentrons_protocol export_plate_layout expression_heatmap_cluster fastq_qc_report fastq_trim find_orfs format_sequence functional_enrichment gc_content gene_dossier gene_expression gene_model golden_gate_fidelity hgvs_convert id_map_poll id_map_submit in_silico_pcr kasp_primer_design melting_temperature motif_finder multiple_sequence_alignment oligo_analysis ortholog_map pairwise_alignment parse_genbank parse_sanger_trace plasmid_annotate plasmid_deep_annotate plasmid_full_report plasmid_identify prime_editing_design prime_editing_twin_design primer_design primer_specificity protease_digestion protein_annotate_poll protein_annotate_submit protein_hydrophobicity protein_properties random_sequence rbs_design rbs_predict restriction_sites reverse_complement reverse_translate rna_fold sanger_vs_reference save_permalink seqfile_stats sequence_fetch sequence_format_convert sequence_report sequence_search sequencing_readback_verify session_create session_get session_run session_set sirna_design site_directed_mutagenesis translate variant_annotate variant_comparator verify_assembly verify_construct virtual_gel volcano_plot_data web_search workflow

Harness readiness observed

Can an agent's harness pick this tool and call it correctly? Graded on the three things an agent reads — name, description, typed parameters. Band A — an agent can pick and call these reliably.

legibility checktools passingok
name is clear & specific86/86
has a description86/86
description is substantive86/86
parameters are typed86/86
parameters are described70/8670/86

86 of 86 tools graded · 86 with a captured input schema. RS-008, Tier-1 — no domain knowledge, applies to any tool. “Alive” is not the same as “usable”.

Truth checks observed

0/0 answers agree with the published primary source · 3 unverifiable.

claim classkeyverdictprimary source
tariff-duty0901.21.00unverifiableus-hts
tariff-duty6109.10.00unverifiableus-hts
tariff-duty8471.30.01unverifiableus-hts

Protocol conformance observed

2025-06-18
protocol version
session model
1.1.0
server version SeqBench MCP

Capabilities: prompts tools

Drift

No confirmed drift on record. Baseline set 2026-08-14; changes appear here after human review.

History observed

Every probe we made, oldest → newest (2 shown, 2026-08-12 → 2026-08-14). Green answered · amber answered-but-walled · red no useful answer. A gap in our cadence is a gap in coverage, not evidence about the server.

changelog (1 events)
datetypewhat changed
2026-08-12outcomefirst probe: ok_tools
raw probe records
probed (UTC) outcomehttpmsprotocol
2026-08-14T19:58:29Zok_tools2001392025-06-18
2026-08-12T05:33:44Zok_tools2001212025-06-18

Watch or embed this verdict

verdict badge

Operators: put the live verdict in your README — it updates with every census, links back here, and is a dated observation, never a warranty:

[![Robin Saige verdict](https://robinsaige.com/badge/com.seqbench/workbench.svg)](https://robinsaige.com/s/com.seqbench/workbench)

Depend on it? Subscribe to its change feed — outcome changes and confirmed drift, no account needed. Building on it? The full dossier as JSON — verdict, rating, truth checks, history — stable enough to gate CI on.

← overview