Review kit

Evaluate one bounded workflow with the right artifact.

Choose the review packet that matches the question: local execution, scoped verification, reproducibility, scientific-AI record quality, or benchmark credibility.

Review questionDoes this evidence materially change your confidence in the computation?
Review packet matrix

Start from the artifact, not a broad product claim.

The strongest first review is narrow enough to replay and important enough to reveal whether OperatorWorks changes the quality of technical judgment.

Review pathBest forPrimary artifactWhat the reviewer should confirm
Workbench local reviewResearchers, labs, technical evaluatorsNotebook, managed environment, derivation trace, export pathThe workflow installs, launches, executes, preserves order, exposes assumptions, and produces a reviewable output locally.
Verify workflow reviewQuantum, scientific, and high-assurance R&DScoped claim, assumptions, outcome, certificate-style record, replay contextThe result is pass, fail, inconclusive, or unsupported for reasons the reviewer can inspect.
Evidence Bundle reviewResearch handoff, enterprise review, procurementManifest, inputs, trace, outputs, environment, warnings, hashes, notesThe artifact is complete enough to understand, identify, and replay without reconstructing the workflow from screenshots.
Scientific-AI record reviewAI-for-science teams and technical-data buyersRecord schema, provenance, derivation, canonical result, difficulty, uniqueness and verification metadataThe record is useful, non-duplicative, traceable, and better suited to evaluation or training than a plausible but unverified example.
Benchmark reviewTechnical buyers, independent reviewers, partnersTask definition, baseline, environment, expected output, measured result, limitationsThe comparison is fair, reproducible, relevant to a real workflow, and adequate for the precise claim being made.
Reviewer scorecard

Six questions separate evidence from theater.

A strong packet should let a skeptical technical reviewer answer these questions directly.

01 · Correctness

Is the mathematical claim established?

Inspect the identity, rule, canonical form, negative cases, and supported operator family.

02 · Assumptions

Are the conditions explicit?

Check conventions, ordering rules, domains, index conditions, and any facts required by the transformation.

03 · Replay

Can the artifact be reproduced?

Use the recorded environment, versions, inputs, deterministic steps, and replay guidance.

04 · Evidence

Is the trail complete?

Look for manifests, hashes, warnings, outputs, certificates, and links between the claim and its evidence.

05 · Boundary

Are unsupported cases visible?

Confirm that inconclusive and unsupported states remain distinct from verified results.

06 · Claims

Does the language match the proof?

Ensure the packet does not imply production, customer, cloud, API, or superiority status beyond the evidence.

Evidence Bundle anatomy

One packet connects the claim to its context.

The exact schema may evolve, but the review standard remains: preserve enough information for a qualified reviewer to understand what happened and what remains unproven.

01

Identity — workflow ID, product version, run context, and stable artifact identifiers.

02

Scientific context — input expression, assumptions, operator families, conventions, and claim.

03

Computation — transformation trace, outputs, warnings, and explicit verification outcome.

04

Reproducibility — environment, dependency context, hashes, replay guidance, and reviewer notes.

What to send

  • A redacted description of one narrow workflow
  • The reviewer audience and artifact type you need
  • The failure mode: order, assumptions, equivalence, provenance, or replay
  • The cost of a silent error or weak handoff
  • The boundary of information you are authorized to share

What to keep out

  • Unredacted proprietary or confidential equations
  • Customer data, credentials, regulated records, or export-controlled content
  • Requests that assume a hosted service or broad deployed operation exists
  • Claims that depend on undisclosed benchmarks or unreviewed artifacts
  • Broad “prove everything” requests that cannot be scoped or replayed
Controlled review

Bring one narrow workflow and one clear decision.

Request a review path
Review posture: review packets are selected, bounded artifacts. They are not a substitute for independent technical judgment and do not imply broad production readiness, public benchmark leadership, hosted availability, customer adoption, or support outside the stated workflow boundary.