You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
{{ message }}
Repository navigation
Commit 30c84c1
Browse filesBrowse the repository at this point in the historyBrowse files
Referee Mode v2: configurable lenses, reviewer stances, model tiers, cross-vendor
ClaudeR 0.8.0 (R-side only; bridge unchanged at 0.10.0).
Mode-collapse hardening for the referee lenses. Same-model parallel
reviewers share blind spots; lens mandates decorrelate attention but
not priors. New controls:
- referee_prompt(lenses, reviewers_per_lens, model, cross_vendor):
run any lens subset; 2 reviewers per lens = adversarial prosecutor +
verifier stances, 3 adds a backwards reader; model takes one tier for
all subagents ("haiku" quick / "opus" submission-grade) or a named
per-lens vector, passed through to the host CLI subagent dispatch
(Claude Code Task tool model parameter); cross_vendor = TRUE
dispatches logic/methods reviewers to another vendor via codex/agy/
qwen one-shots -- the strongest decorrelation available
- Anti-collapse rules now mandatory in the protocol: reviewer prompts
written from scratch per lens and stance (template-with-one-word-
swapped forbidden), consistency lens reads back-to-front, findings
carry n_independent corroboration counts through dedup and into the
report, and the run configuration is reported honestly
- reviewer_zero_prompt(referee = TRUE) uses the configured defaults
CI: referee v2 configuration test. R CMD check: Status OK.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Copy file name to clipboardExpand all lines: README.md
+12Lines changed: 12 additions & 0 deletions
Display the source diff
Display the rich diff
Original file line number
Diff line number
Diff line change
@@ -44,6 +44,7 @@ claudeAddin()
44
44
<details>
45
45
<summary><b>Recent Updates</b> (click to expand)</summary>
46
46
47
+
-**Referee Mode v2: configurable reviewers (R 0.8.0).**`referee_prompt()` now takes `lenses` (run any subset), `reviewers_per_lens` (2 = adversarial prosecutor + verifier pairs, 3 adds a backwards reader), `model` (a tier for all subagents, e.g. `"haiku"` for quick passes or `"opus"` for submission-grade, or a named per-lens vector), and `cross_vendor = TRUE` to dispatch logic/methods reviewers to a different model vendor via codex/agy/qwen one-shots. Anti-collapse rules are now mandatory in the protocol: reviewer prompts must be written from scratch per lens and stance, the consistency lens reads back-to-front, and every finding carries a corroboration count across independent reviewers.
47
48
-**Referee Mode (R 0.7.0 / clauder-mcp 0.10.0).** The substantive manuscript review that paid services charge ~$50 a pass for, running free on the subscription you already have, and delivered where it belongs: as Word comments in your manuscript. `reviewer_zero_prompt(referee = TRUE)` (or standalone `referee_prompt()`) runs five content-only review lenses — argument logic, methods, internal consistency, evidence presentation, framing — as parallel subagents where the CLI supports them. Every finding must anchor to a verbatim quote and survive an independent verification pass before it lands in the document, severities are kept honest, and a clean report on a sound paper is a valid outcome. Alongside it, the new `check_cross_references` tool deterministically catches dangling references ("see Table 4" when Table 4 no longer exists) and tables or figures the text never mentions.
48
49
49
50
-**Value-first auditing (R 0.6.0 / clauder-mcp 0.9.0).** Built from field feedback after a full manuscript+supplement audit. `read_file` now transparently extracts `.docx`/`.pdf` (previously returned raw bytes), and the extractor preserves structure: headings marked, table cells emitted row-wise with separators (they were previously dropped entirely). New `reconcile_values` tool: enumerates every number in a manuscript and reconciles each against the corpus your code produced, respecting displayed precision (5038.5 matches 5038.46), commas, percents, scientific notation, and `< .001` thresholds; the per-value `values_registry` makes numeric completeness a construction, not a diligence hope. Reviewer Zero now sets audit-clean print options (no more tibble 3-sig-fig false alarms), gates on the value sweep, and takes final verdicts from clean-room runs via `probe_scripts(capture_output = TRUE)`. CrossRef lookups retry with backoff instead of silently truncating the reference check on 429s.
0 commit comments