&

Portfolio review · 2026-09-12

What is built, which rung it stands on, and what one outsider would change. A review of Ampersand Box Design for investors, contributors and the builder himself.

This is not a pitch. It is an audit of a portfolio that has spent a year building governance arithmetic, protocols, an operating system and a proof harness with no external funding and, so far, no external user. Every status word below is one of five rungs on an evidence ladder, every count was derived by a command run today, and the things the previous review got wrong are retracted here by name rather than quietly edited out.

Subject
The [&] portfolio — 23 public hostnames, ~27 sub-repositories, one builder
Measured
2026-09-12 from the working tree at ProjectAmp2 plus live HTTP checks; see Method
Supersedes
The May 2026 review (“Investor & Contributor Report”) — four of its headline claims are retracted in §4
Author
Claude (Fable 5.1), from the tree and the public web. Judgements are labelled as judgements.
Status
Reviewed against DOCTRINE.md rules 1–5. Not a Travis ruling. No composite score is offered, on purpose.

01 · The read in ninety seconds

A lot has been built. Almost none of it has been touched by someone else.

Six numbers carry the whole review. Five are strengths. The sixth is the one that bounds all of them.

210+3
enforced property-tested laws in the governance kernel, plus three declared-open gaps that print red by design
derived today · node test/laws.mjs && node test/compose-laws.mjs
12
OpenSentience protocols, OS-001 through OS-011 plus SCOPE; one still a draft (OS-007)
amp-nav 0.12.3 · opensentience.org
46
named agent invariants on the periodic table, grouped by evidentiary register; κ checked on 1,926,351 finite systems
_invariants/data/cells.json
32
chapters in Unboxed Patterns, each with a WRL world, a forge-reduced film and a derived standing
13 witnessed · 14 stated · 3 proposed
23/23
public hostnames answering HTTP 200 today, from ampersandboxdesign.com to gpscoord.com
curl, 2026-09-12
0
claims on the external rung. No third party has reproduced a result, adopted a package, or driven the loop
STACK_COMPLETION §6, unchanged since April

That last cell is the review. The portfolio's own status document has carried an adoption row at roughly ten percent since April and has watched four months of substantial building leave it untouched, because every one of those deliveries landed on a rung below external. The cheapest known move against it, handing one outsider the docs atlas and one real task, has been marked “runnable today, needs no new code” since 2026-08-11 and has not been run. The rest of this page is context for that one sentence.

02 · How to read every status word on this page

Five rungs. No checkmarks, no percentages standing in for a rung.

The portfolio's doctrine forbids “done”. A claim is on exactly one of these rungs, and when in doubt it is downgraded. The instrument is borrowed here unchanged, because a review that invented its own vocabulary would be the drift it is grading.

spec
Specified, possibly schema'd. No implementation.
Deliberatic · AgenTroMatic · TickTickClock · GeoFleetic · studbook · SCOPE
in_tree
Implemented and tested against stubs, simulators or single-process round trips. The code exists; the system is unproven.
Super (CD) · WebHost.Systems · BendScript core · RuneFort · Workbench ingest
live_local
Exercised end to end against its real substrate at least once on a developer machine.
dark-factory loop 7/7 · TRVM/WRL · Weave · ComputeDriven Edge · the-residency
live_deployed
Runs in production and has been exercised there.
the law suites in a browser · T&R v0.3 image · PRISM · FleetPrompt · SpecPrompt · Unboxed Patterns · Alkeyword site
external
Validated by someone who is not us.
nothing — 0 of 32 patterns REPRODUCED, 0 packages with a known outside consumer

The rung words are quoted from DOCTRINE.md, which is entrenched. The names under each rung are this review's placement, derived from STACK_COMPLETION.md and the nav's status strings, and a placement is a judgement wherever the document did not state one.

03 · The shape of the portfolio

Three strata, and the word that holds them together.

The old review counted seventeen projects in one flat table. The portfolio does not have that shape any more. It has a doctrine layer that says what may be claimed, a protocol layer that says how things compose, and a surface layer that people could actually run. Grading them together is how a shipped operating system and an unwritten spec end up averaged into one number.

An agent is an active locus in a world.

An active locus carries state, scoped authority, and causal continuity: it can establish its operational state, determine which actions are admissible under that state, the governing policy and its explicit grants, act on the world, and preserve the evidence needed to continue.

AGENCY.md §1 — proposed 2026-08-25, not yet ratified. Quoted, not paraphrased, because a definition that lives in twenty places is a count that lives in two.

Stratum one — doctrine and arithmetic

What every other claim in the portfolio is measured against. This is the part that no competitor is shipping in the same form, and the part that has produced the most retractions, because it is the part most often measured.

ArtifactWhat it isRungLast measured
box-and-box v0.10.0The eight-rung governance kernel: feasible ▸ permitted ▸ best over an un-weakenable floor, a certificate on every verdict. 109 kernel + 101 compose laws enforced; CP5/CP6/CP7 falsified by design.live_deployed2026-09-12, both suites run from this tree; 109 wired live in the playground (2026-08-24)
Periodic Table of Agent Invariants46 cells grouped by evidentiary register; the page is generated from data and refuses a claim its cell does not support.live_deployed2026-08-25 (INV-R7.8)
Six proof pagesκ machine-checked over 1,926,351 objects, 0 counterexamples; π property-tested; four implementation-verified against the Delegatic kernel suite.in_tree2026-06
Unboxed Patterns32 chapters, one WRL world each, 166-node typed graph, 30+ films reduced by TRVM's forge, standing derived on every build.live_deployed2026-09-13, live at opensentience.org/patterns
DOCTRINE.md · AGENCY.md · CLAIM_LEDGER.jsonThe conduct rules, the subject vocabulary, and one record per public claim with a status from a closed set.in_treeentrenched / proposed / INV-R7.8

Stratum two — protocols

ProtocolRoleRungNote
[&] Protocol 0.1.0Six primitives, two operators, one kernel; compiles to MCP and A2A. npm + Python SDKs, Elixir reference CLI.in_tree58 tests; the compose laws above are its conformance surface
OpenSentience OS-001…011 + SCOPEThe twelve research protocols: continual learning, κ-routing, deliberation, attention, model tiers, governance, adversarial (draft), harness, PRISM, PULSE, embodiment, scope.in_treeOS-007 is the one still-draft protocol; OS-012 SCOPE deliberately deferred
PULSE 0.1.1Temporal algebra: a loop declares its phases, cadence, nesting and cross-loop tokens in one JSON file.live_deployednpm os-pulse; six canonical tokens; three reference manifests
PRISM 0.1.0Diagnostic benchmark that reads any PULSE manifest and measures the loop over time.live_deployed172 tests; prism-eval.fly.dev answered 200 today
BendScript 0.1Graph-first document format with typed inline link facets; TypeScript reference parser.in_tree96 tests; pivoted from a SaaS product in April
RuneFort 0.1Layout protocol for tiled, file-backed UIs; the T&R road draws 10,000 panes at 16.7 ms/frame with it.live_localno test script; treat as untested rather than 0 tests

Stratum three — surfaces people could run

SurfaceWhat it doesRungEvidence
T&R v0.3 (ComputeDriven OS)FreeBSD 15 assembled from pkgbase; a 626 MB image that carries its own verifier so its claims can be re-derived without asking anyone.live_deployeddownloadable with sha256; runs under PARKVPS; nothing external has booted it that we know of
Super (CD)Desktop compute surface holding control, placement, evidence and authority above interchangeable engines. Elixir/OTP + Tauri.in_treefirst dogfood cycle 2026-09-12; app-layer commits landed 2026-09-11; not released
[World] CloudHosted, forked, restored worlds on Cloudflare; Workers → Hyperdrive → Pigsty; identity via Keycloak.in_treeits own status page: 24 of 40 capabilities run on a dev machine, 16 written not built, 0 serving the public
Graphonomous 0.4.3Continual-learning knowledge graph and MCP memory server; embedded SQLite; κ-routed deliberation.live_deployednpm; 573 tests at last full run (April); see the deep dive for the benchmark caveat
Workbench 0.4.0-alphaRecords Claude Code transcripts into signed, scored SkillBundles; six proof gates computed through Invariant Arithmetic.live_local152 tests; 202 MB real session indexed in 936 ms; three redaction defects found in real data and fixed
FleetPrompt · SpecPrompt · Agentelic · DelegaticMarketplace, spec toolchain, agent builder, OS-006 authorization kernel. Elixir/OTP.live_deployed157 · 71 · 66 · 42 tests; Fly hosts answered today; “prior design” per the nav — the Supabase-era shape
Alkeyword 0.6.xBring-a-domain generative-engine-optimisation tool; deterministic stdlib crawl → claim graph.live_deployedsite live; 18 regression witnesses each cut from a real defect
TRVM · WRL · TRAAVIISThe substrate: an interaction-calculus reducer, the world language it reduces, and the evidence format receipts are written in.live_localWRL conformance 924 checks (2026-09-11); TRVM governance at round 28; TRAAVIIS pivoted to trvs
AcademyThe institutional Read / Practice / Prove loop above the products.specprototype surface live; the one experiment that would move adoption (B1) is unrun
Deliberatic · AgenTroMatic · TickTickClock · GeoFleeticDeliberation, orchestration, temporal and spatial intelligence.spec418–1118 line specs with PULSE manifests; no code

Test counts are the last recorded full run for each suite and are dated where the document dates them. The old “1,463 ecosystem tests, 100% passing” figure has not been re-run since 2026-04-28 and is not repeated here as current.

04 · What the May review got wrong

Four headline claims, retracted by name.

Rule five of the doctrine: retract out loud, in the same place the claim was published. The May report was published at this URL. These are its errors, and in each case the correction came from a ruling or a measurement rather than from a change of taste.

Retracted — “10 live Fly.io deployments” counted as a strength

All ten hosts still answer today, which is what the May figure measured. But on 2026-08-15 it was ruled that compute moves into the ComputeDriven OS, Cloudflare is the cloud side, and Fly keeps exactly one thing, Keycloak. The deployments are now a wind-down queue, not a moat. A rating that reads deployment breadth as strength misreads a wind-down as a regression. Nothing has been torn down, and one orphaned validator (box-and-box-mcp) has been answering requests since its code was removed on 2026-05-30.

Retracted — “1 shared data layer” (Supabase) as architecture

The shared-Supabase route was abandoned by ruling on 2026-07-30. Postgres stores a row but not why the row is that row, and nothing at the migration boundary could refuse a row whose provenance did not check out, which stopped the stack's central claim at the database. Its replacement, studbook, is a spec with one unruled question (where confidentiality comes from under content addressing). The 40 migrations still apply and FleetPrompt still runs on them; they are archived, not failed.

Retracted — “Invariant Arithmetic, 15 laws” as the formal moat

Invariant Arithmetic is rungs one and two of an eight-rung ladder. The kernel that ladder became derives 210 enforced laws plus 3 declared-open today, and the count has been corrected at least six times between May and now (116 → 118 → 157 → 164 → 168 → 173 → 176 → 210) because it was being hand-typed. Both suites now derive their printed totals, and this page quotes the derivation rather than the number.

Retracted — “1,463 tests · 100% passing” and “8.2 / 10 composite” as current

The test total was last run on 2026-04-28. The composite was a single blended figure that let a ten-percent adoption row hide behind a ninety-percent protocol row for four months. This review offers per-dimension grades and refuses to blend them; see §6.

Still standing from May

κ verified on 1,926,351 finite systems with zero counterexamples. Graphonomous on npm with a reproducible LongMemEval result. PULSE and PRISM shipped. The three-protocol thesis. The dark-factory loop closed locally, seven steps of seven. None of these has moved up a rung either, which is the point of §1.

05 · Four deep dives

Where the substance is, and what each one is careful not to claim.

A reviewer picks the four things a stranger should look at first. These are mine, chosen for how much of their evidence can be re-derived from a terminal.

Kernel · doctrine stratum

The laws, and the three that fail on purpose

Run both suites and the kernel prints 109 enforced laws, then 101 compose laws, 2,000 randomised trials each, and then three lines in red: CP5, CP6 and CP7 are algebraic gaps the kernel has not closed, and the build fails if one of them ever starts passing. That inversion is the most credible thing in the portfolio. A test suite that can only go green is not evidence; this one carries its own negative controls.

The July–August round with outside review (R25–R63 in the revision register) produced the compose layer almost entirely from falsifiers: a coalition handing off on behalf of a member that could not have handed off itself; identity that quantifies over the carrier rather than everything a constructor can represent; a certificate that is presented, not authenticated, unless this process minted it.

cd AmpersandBoxDesign/box-and-box && node test/laws.mjs && node test/compose-laws.mjs · ~2 s
Memory · surface stratum

Graphonomous, read honestly

The 92.6% LongMemEval figure is a QA proxy on the oracle split, reproducible on local hardware, with the reader model held fixed. The public leaderboard now clusters at 93–95% on cloud readers (Mem0 94.4%, Mastra 94.87%, OMEGA 95.4%), so the headline number is no longer a differentiator and this review does not present it as one.

The status document records a more uncomfortable measurement: topology on versus off is +0.3 points, and there is no flat-RAG baseline. So the thing Graphonomous is for, κ-routed deliberation over a cyclic graph, has not yet been shown to earn its keep on this benchmark. What it does have that the funded peers do not: an embedded, edge-sized store; a κ invariant with a machine-checked proof; and goal coverage as a first-class graph object. Operational finding from this week: three sessions running three BEAM servers on one SQLite file made a retrieve hang for thirty minutes. That is a session policy defect, not an engine defect, and it is recorded.

npx -y graphonomous · OS-E001 §4.5 · STACK_COMPLETION §2.1
Operating system · surface stratum

ComputeDriven: the derivation travels with the artifact

The pitch is one sentence — keep the computation, not just the answer — and T&R is the sentence made bootable: a 626 MB FreeBSD 15 image that carries the verifier needed to re-derive what it claims. The road shell draws 10,000 RuneFort panes at 16.7 ms a frame. The cloud half is honest to the point of discomfort: its own status page says 0 of 40 capabilities serve the public.

The Edge lane is where the discipline shows most. A placement battery was refused three times by its own baseline guard before a rebooted, dedicated host produced a result, and the result claims exactly what it earned: no admission cost from one settled move, and no benefit claimable, because benefit was ceiling-censored. A 5.3 KB production floor for the kernellet with a dependency count flat across a 3.3× code growth is the strongest single datapoint the program has.

tr.computedriven.com · cloud.computedriven.com/#status · computedriven/receipts/
Recorder · surface stratum

Workbench, the wedge nobody has picked up yet

Thirty-nine percent of professional developers were using Claude Code at work by mid-2026 (JetBrains, August). Every one of those sessions writes a transcript to disk. Workbench ingests them into signed SkillBundles and scores them through six proof gates, each computed by the same arithmetic the kernel uses. The ingest spike ran on real data: the largest session on the machine, 202 MB, indexed in under a second, streaming.

Real data found three shipped defects the fixtures never would have: a canonicaliser that threw on undefined keys, a redaction profile that scrubbed observations but left secrets in tool arguments (8 of the 25 largest transcripts contained key-shaped strings), and a manifest that was never redacted at all. All fixed, all recorded in INGEST_SPIKE.md. The deployed site was a 1.6 KB empty shell for a while and nobody noticed, because deploys were checked by grepping HTML for copy instead of reading the route table. That lesson is now in the roadmap.

workbench.opensentience.org · workbench/docs/ROADMAP.md · 152 vitest

06 · Grades, one per dimension, no composite

Eight dimensions. Each one names what would move it.

A blended number is how the adoption row hid for four months. So: letters, per axis, with the single action that changes each letter. Two of these dimensions (honesty infrastructure, scope) did not exist in the May scorecard and are the ones I would weight most if I had to weight anything.

A

Formal foundations

A property-tested governance algebra with negative controls, a machine-checked κ, a generated invariants table that refuses unsupported claims, and a pattern language whose standing is derived rather than typed. Nothing comparable is shipping from Mem0, Letta, LangChain or the Cloudflare Agents SDK.

Moves to A+ when one of CP5/CP6/CP7 is closed by a ruling, or an outsider reproduces the suites from a clean clone and says so.
A

Honesty infrastructure

Gates that refuse a page whose count drifted, a revision register that ranks contradictions by damage, a status vocabulary with “when in doubt, downgrade”, and a habit of retracting in the place the error was published. This dimension is why the rest of the page can be believed.

Moves when the gates run in CI on every repo rather than by hand; check-authority-language.mjs exists precisely because a doctrine was contradicted on the two pages most likely to be read by a stranger — the previous versions of this one and its sibling.
A−

Protocol coherence

Twelve protocols, one algebra, one temporal manifest format, one benchmark that reads it. The layers are separable: a third party could adopt PULSE without the kernel. The agent definition is still “proposed”, and OS-007 is still a draft.

Moves to A when AGENCY.md is ratified and OS-007 gets an empirical anchor.
B+

Shipped surface

An operating system image, a desktop cockpit in dogfood, a memory server on npm, a recorder that ingests real transcripts, four Elixir products on Fly, a live GEO tool. The volume is startling for one builder. The grade is not higher because the estate is mid-migration: the cloud side has 0 public capabilities and the Fly side is a wind-down queue.

Moves to A− when [World] Cloud serves one capability publicly, or one deployed end-to-end pass of the dark-factory loop runs across machines.
C+

Deployment posture

Everything that is live is live on infrastructure the plan has already left, and the destination (Cloudflare + Pigsty + Keycloak-on-Fly) is in_tree. The docs atlas answers on 13 byte-identical hostnames. The orphaned validator has run for three and a half months.

Moves to B when the first Fly teardown happens after its OS-side replacement is live_local, and the twelve non-canonical docs hosts redirect.
C

Scope discipline

Twenty-three hostnames, roughly twenty-seven sub-repositories, seven parallel lanes with rulings owed to the same person. The status file's own “what not to prioritise” list has been respected (no OS-012, no landing-page redesigns before users), but the surface area still grows faster than any outsider could read it.

Moves to B when the spec-only four are formally parked or merged, and the nav stops listing more than one “prior design” product.
D

External validation

Zero. No outside reproduction, no outside consumer, no competitor adapted into PRISM, no contributor. This is the row that bounds every other row and it has not moved since April. The cheapest experiment against it needs no code and has been ready for a month.

Moves to C the day one person who is not the builder runs B1, or reproduces the law suites from a clean clone, or boots T&R and posts the sha256 they derived.
—

Revenue

None, and no SKU is live. Not graded, because it has not been attempted. What exists: a Cloud pricing page priced as “a place to put something” (storage at roughly $1.35 per 100 GB per month cost basis), and a ruling that storage admission is authorisation and ships before checkout. That ordering is correct and it defers revenue on purpose.

Becomes gradable when the storage SKU takes its first real byte from a stranger.

07 · Risk, stated so it can be argued with

Six things that could make the strengths not matter.

1 · The memory layer commoditises before distribution exists

Retrieval scores have converged into a two-point band on cloud readers. Supermemory ships a free local server; Cloudflare gives every agent a Durable Object with state for nothing when idle. Graphonomous's durable edge is the invariant and the goal graph, and neither has a published benchmark that shows it winning.

2 · The migration gap

Between “Fly is a wind-down queue” and “Cloud serves one public capability” there is a window in which the portfolio has less running than it did in May. The sequencing rule (nothing torn down before its replacement is live_local) protects against losing a capability; it does not protect against a year with nothing new deployed.

3 · One builder, one ledger, many lanes

Seven lanes owe rulings to one person. Two sessions independently built the same Super round in one week; the ledger drifted 13 of 22 gated figures in sixteen days in the WEK lane. The honesty infrastructure catches this, but catching drift is not the same as not producing it.

4 · Agent-washing fatigue lands on the honest too

Gartner expects over 40% of agentic-AI projects to be cancelled by end of 2027 and counts roughly 130 real vendors among thousands. A portfolio whose vocabulary is locus, world, rung reads as jargon to a buyer who has just been burned by a chatbot relabelled as an agent. The doctrine's plainness is an asset only if someone reads that far.

5 · The word “world” is about to be crowded

World-model startups raised more than $3 billion in the first half of 2026. That “world” means a learned simulator of physics. This portfolio's “world” is a machine and the record of how it got that way, in three named sizes. The distinction is real and it will have to be made on every page, forever.

6 · The Elixir bet

The BEAM is the right shape for supervised agent loops, and the 2026 typed-Elixir release removed the last structural excuse. The hiring pool did not grow to match. Super, ampd, four products and Graphonomous are all on it; the cloud control plane, by ruling, is not.

08 · What an outsider can actually do this week

Three asks, sized so that each one moves a rung.

The portfolio does not need more code from a stranger. It needs a stranger. These are the cheapest ways to be one, in the order that changes the grades above.

Contributor · 30 minutes · moves External from D

Reproduce one receipt and say so in public

Clone AmpersandBoxDesign, run the two law suites, and post the printed totals and seed. Or open the playground on a fresh origin and press its button. Or download the T&R image and post the sha256 you derived. Any one of these is the portfolio's first external claim, and there is a receipt kind waiting for exactly that.

git clone github.com/c-u-l8er/AmpersandBoxDesign · cd box-and-box · node test/laws.mjs
Reader · 2 hours · runs experiment B1

Take the docs atlas and one real task

Academy's B1 is: hand one outsider the docs and one real repository task, and record where they got stuck. It has been marked runnable since 2026-08-11. Whoever does it produces the first adoption datapoint in the portfolio's history, and the refusal log has a place for what they could not do.

academy.opensentience.org · ACADEMY.md B1
Investor · a conversation · funds the one missing pass

Underwrite the deployed end-to-end run

The single highest-leverage engineering move in the status file has been the same for five months: one production pass of perceive → act → record → crystallise → install → replay → measure across deployed machines. Now it also needs the Cloud side to hold the world. That is a scoped, receipt-producing milestone, not a roadmap. If you want a term sheet, it should be written against that receipt.

STACK_COMPLETION §7 item 0 · CLOUD_V1.md M2

09 · The next six months, September 2026 to March 2027

What the market will do, and what should be true here by spring.

The market

  • Logs become procurement. The EU AI Act's high-risk obligations became enforceable on 2 August 2026, and Article 12 wants automatic, replayable records of what an agent did, not what it said. A Q1 survey of 420 organisations found only 17% could reconstruct an agent's tool-call sequence after the fact. Receipts move from a research value to a checkbox.
  • The protocol stack finishes consolidating. MCP joined the Linux Foundation's Agentic AI Foundation in December 2025 and A2A joined it on 17 August 2026. There is now one neutral home for agent-to-tool and agent-to-agent, which is where the [&] Protocol's “compiles to MCP and A2A” claim gets cheaper to make and easier to test.
  • Skills get audited. The skills registries index somewhere between 800,000 and 1.9 million SKILL.md files; Snyk found prompt injection in 36% of a sample. A marketplace that installs only what carries a certificate, and returns zero otherwise, stops being a philosophical position.
  • Memory scores stop mattering. Everyone will sit between 93 and 96. The next differentiator is provenance: who wrote this memory, under what authority, and can I replay the run that produced it.

What should be true here by March

MilestoneFrom → to
One external reproductionnothing → external on at least one receipt
[World] Cloud serves a capabilityin_tree → live_deployed; sync + restore before hosted compute, per R12
Deployed dark-factory passlive_local → live_deployed across machines
First Fly teardownten answering → nine, after its replacement is live
Graphonomous vs a flat baseline+0.3 pp topology-on/off → a published comparison that could lose
Workbench on hooksmanual ingest → a Claude Code session recording itself
studbook §10.2open → ruled, so anything with a user in it has somewhere to live

Seven rows. If three are true by March the portfolio is on a different rung than it is today. If none are, the doctrine says to write that here, at this URL, by name.

10 · Method and sources

How this page was made, so that it can be re-derived.

Measured in the tree, 2026-09-12

cd AmpersandBoxDesign/box-and-box
node test/laws.mjs          # ✓ all 109 enforced kernel laws hold
node test/compose-laws.mjs  # ✓ all 101 enforced CC2 compose laws
                            #   3 known gaps (CP5 CP6 CP7) FALSIFIED

python3 -c 'import json; print(len(json.load(open(
  "opensentience.org/_invariants/data/cells.json"))["cells"]))'   # 46

for h in graphonomous-mcp prism-eval os-pulse-mcp fleetprompt \
         specprompt agentelic delegatic-mcp body-browser-mcp \
         body-os-mcp box-and-box-mcp; do
  curl -s -o /dev/null -w "%{http_code}\n" https://$h.fly.dev/
done                        # 10 of 10 answered (200/404/405)

# 23 public hostnames → 23 × HTTP 200 (alkeyword apex on retry)

Versions were read from mix.exs and package.json files. Rung placements come from STACK_COMPLETION.md (last entry 2026-09-13), the nav's status strings (amp-nav 0.12.3), and the sites' own status pages. Where those disagree, the lower rung was taken.

Not re-run: the full ecosystem test battery, Graphonomous's 573-test suite, WRL's 924 conformance checks. Their counts are quoted with the date they were last run.

External sources

Tags: analyst = named research firm; primary = the organisation's own publication; survey = stated sample and method; vendor = a competitor's own claim; trade = trade press or blog, quoted where nothing better exists and labelled so.