Crypto & Web3  ·  Curated marketplace

first-reader

Beta readers for any draft, run by simulating how a real reader experiences it, moment by moment.


Composite

4.0

C 4.0 · A 0.0

How we got there

Craft · D1–D5

D1 · Trigger clarity 4.5
D2 · Output specificity 4.0
D3 · Scope precision 4.0
D4 · Self-containment 3.5
D5 · Reusability 4.0

02 — Review

Our evaluation


Tier-2 Review: first-reader

What we tried to do

We attempted to run the standard two-stage harness against first-reader: an install step, followed by a smoke invocation that exercises the described workflow end to end. The skill's own framing is unusually concrete — a skim gate, a no-lookahead timed read producing an attention transcript with quit points, a next-day recall test, and a trust ledger of the implied author — so we expected at least a runnable interface to point the harness at.

What failed

Both tests were skipped as unrunnable. Nothing passed, nothing partially passed, nothing failed on the merits — the harness never got far enough to judge behavior.

The blocker is concrete and single: the provided SKILL.md is truncated and exposes no invocation surface. Specifically:

  • No install command, package name, or dependency declaration appears anywhere in the supplied text.
  • No README, CLI entrypoint, or API signature is present.
  • The source URL resolves to a skillsmp.com creator page, not an install recipe or a repository with runnable instructions.

The install step had nothing to execute. The smoke-invocation step had no syntax to call. Because the harness cannot invoke a skill that declares no way to be invoked, both stages short-circuited to skipped rather than producing a pass/fail signal.

This is a D-048 simulated run, so treat the zero counts as "not exercised," not "exercised and found wanting." The distinction matters: the skill was never given a chance to succeed or fail.

What we observed

The truncated body we did receive is, on its own terms, well-specified. The trigger paragraph enumerates a genuinely broad but coherent set of phrasings — "beta read," "first reader," "will this hold attention," "where do readers bounce," "why does this feel hollow after humanizer edits" — and the described outputs (attention transcript, quit points, recall test, trust ledger) are the kind of thing you could actually grade. The scope claim is disciplined: reports where the reading broke; never rewrites. That boundary is the most promising part of the design.

But none of that is verifiable here. We cannot confirm the skim gate behaves as described, that the timed read is genuinely no-lookahead, or that the trust ledger produces anything more than a restatement of the draft. The dependency list came back empty, which is consistent with a pure-prompt skill — but also consistent with a truncated file that simply dropped its own requirements.

On the rating

The composite of 4.0/5.0 (D1 4.5, D2 4.0, D3 4.0, D4 3.5, D5 4.0) is therefore theoretical. It reflects the quality of the description of the skill, not the skill in operation. The low-ish D4 (self-containment, 3.5) is the dimension the failure most directly implicates, and in a full run it would likely fall further, since a skill with no install path and no invocation interface is by definition not self-contained. Until the truncation is resolved and the skill is physically re-run, no dimension here should be treated as measured.

Does it still seem valuable in principle?

Yes — cautiously. The problem it targets is real: anti-slop and humanizer passes produce text that is clean but often inert, and "does a busy skimmer open this, and what do they remember tomorrow" is a sharper question than most editing rubrics ask. The four-part structure (skim gate → timed read → recall test → trust ledger) is a plausible, gradeable decomposition, and the refusal to rewrite keeps it in an honest evaluative lane.

The value is contingent on the missing pieces existing: a real install path, a documented invocation, and enough of the prompt to make the gates reproducible. If those are simply absent from our truncated copy, this is a promising skill with a packaging problem. If they never existed, the description is doing work the artifact cannot. Either way, the next step is the same: obtain the untruncated SKILL.md and re-run before assigning a real score.

03 — Tests

What we tried


Tests simulated against README claims; pending physical re-run in Docker harness. Ran 2026-09-25.

Overall: broken. 0 tests passed, 0 partial, 0 failed; key blocker: the provided SKILL.md is truncated and contains no install command, README, or invocation interface, so both tests were skipped as unrunnable.

Test Status Notes
install skipped SKILL.md content is truncated and contains no install command, package name, or dependency declaration; the source URL points to a skillsmp.com creator page rather than a documented install recipe, so no install command could be executed.
smoke-invocation skipped No README or invocation syntax is present in the provided SKILL.md text; the skill describes a beta-read workflow (skim gate, timed read, recall test, trust ledger) but exposes no CLI entrypoint or API to invoke.
04 — Cross-validation

1 source verified

Install

Use this skill

/plugin install first-reader