Research · Measurement plan

How we will measure x4me2 review recovery

The homepage assumption sandbox uses invented values. This page states what we intend to count before data collection begins, so later results cannot quietly change the definition of success. It is a public measurement plan, not an independently archived preregistration.

Published 10 August 2026 · Updated 22 August 2026

Status: the prototype is live; measurement has not started. The current phase is qualitative feedback on whether people want to recover reviews, what feels risky, and which controls must stay with the author. Nothing on this page reports a result.

The domain

x4me2 helps people recover reviews they already wrote on platforms, prove authorship, and host them outside the original platform. Payments from AI systems are planned, not live. It is a League project, so this will be an internal product measurement rather than an independent evaluation. Raw counts, including failures, will be published here for others to inspect.

The unit

The main unit is one completed review recovery. A session counts as complete only when all four conditions hold:

The counters

Attempts
Recovery sessions started, including abandoned and failed ones.
Completions
Sessions reaching the live-under-control state defined above.
Completion rate
Completed recoveries divided by all started sessions.
AI use per attempt
Metered model usage across every automated call in the recovery process, averaged across all attempts.
Checking cost per completion
Compute cost plus human review time. The hourly rate for human review will be stated with the results.
Total cost per completion
AI use, authorship checks, human review, and operating cost, divided by completed recoveries.
Manual comparison
The time and cost required to reconstruct an equivalent, verifiable record without x4me2. The comparison procedure and hourly rate must be published before collection begins.

Still to decide before measurement begins

These decisions will be added as a dated amendment before the first measured batch. Until then, the plan is incomplete.

Community feedback

The open question right now is simpler: would recovering a review be useful, what would make the process feel unsafe, and which decisions must remain with the author? Reactions and objections are welcome through the contact form. When measurement begins, batches will include failures and abandoned sessions.

Results that would weaken the claim

Any of these outcomes will be published with the same prominence as a favorable result.

Publication commitments