All skills
imbad0202 avatar

/academic-paper-reviewer

@a3f6569

Multi-perspective academic paper review with dynamic reviewer personas. Runs a 5-seat, role-separated review panel (Journal-Fit Reviewer + 3 peer-review roles + Devil's Advocate) with field-specific expertise; role separation is not a claim of independent error processes. Supports full review, re-review (verification), quick assessment, methodology focus, Socratic guided, and calibration modes. Triggers on: review paper, peer review, manuscript review, referee report, review my paper, critique paper, simulate review, editorial review, calibrate reviewer, reviewer calibration, measure reviewer accuracy, 審查論文, 論文審查, 模擬審查, 同儕審查, 幫我審這篇, 以審查人角度評估, 審查者校準, 논문 심사, 동료 심사, 모의 심사, 심사자 관점에서 평가, 심사자 보정, revisar artículo, revisión entre pares, revisión de manuscrito, informe de árbitro, revisa mi artículo, criticar artículo, simular revisión, revisión editorial, calibrar revisor.

Use this Skill: https://skilld.dev/gh/imbad0202/academic-research-skills/academic-paper-reviewer

This session only. Nothing lands on disk.

referencesquality_rubrics.md

≈1.6k tokens on demand. Your agent reads this file only when SKILL.md points to it.

Criterion-Bound Judgement Rubric for Academic Paper Review

Purpose

This rubric defines what each reviewer must examine and how to make the basis of a judgement inspectable. It does not provide a calibrated quality score, a paper-ranking scale, or an acceptance probability.

Calibration status

Every review must declare one of these states:

  • NOT_CALIBRATED — the required default. Use this whenever there is no empirical target profile matched to the current domain, article type, venue criteria, rubric version, reviewer/model configuration, and review mode.
  • PROFILE_MEASURED — reserved for a package-level resolution backed by a replay-valid, hash-bound empirical target profile whose exact target fields and actual completed-panel execution_topology_sha256 match. Producing or naming a calibration profile does not by itself authorize this status.

Individual reviewer seats always emit NOT_CALIBRATED, because the actual completed-panel topology (including fallbacks) does not exist until all seats finish. The current Schema 6 package adapter also remains NOT_CALIBRATED; PROFILE_MEASURED must not appear in a live review package until a closed profile schema and replay validator are shipped. This is an honest implementation boundary, not evidence that calibration is impossible.

Never describe either state as proof that judgements are consistent or reproducible across papers, sessions, fields, venues, or model versions. A directional calibration readout is not an empirical target profile and leaves the status NOT_CALIBRATED.

Required judgement form

For every applicable dimension, report:

Field Required content
Criterion source The target venue criterion, reporting standard, article-type expectation, or reviewer configuration item being applied
Judgement EXCEEDS / MEETS / PARTLY_MEETS / DOES_NOT_MEET / NOT_ASSESSED
Evidence anchors Specific manuscript locations or bounded absence anchors
Rationale How the cited evidence bears on the named criterion
Uncertainty or scope limit Missing information, domain dependence, or reviewer limitation
Decision bearing? Whether this judgement affects the recommendation, with a reason

These labels are criterion-bound categories, not numbers. Do not convert them into points, weights, a total, a percentage, or a latent ranking. Do not map any total or count of labels mechanically to Accept, Minor Revision, Major Revision, or Reject.

Editorial recommendations instead follow the applicable contract or the qualitative, evidence-anchored rules in editorial_decision_standards.md. The recommendation must identify the particular unresolved criteria that make it appropriate.


Dimension 1: Originality

Judge the claimed contribution against the paper's stated field, article type, and target venue. Examine whether the work identifies a defensible gap; distinguishes its theory, method, evidence, or application from relevant prior work; and avoids overstating novelty. Replication and boundary-testing studies can make an original contribution without introducing a new theory.

Dimension 2: Methodological Rigor

Judge whether the design can answer the stated research question and whether execution and reporting support the inferences made. Apply paradigm- and design-appropriate criteria, including sampling, measurement, validity threats, analysis choices, uncertainty, transparency, and reproducibility where relevant. A standard applies only when its scope matches the paper.

Dimension 3: Evidence Sufficiency

Judge whether each material claim has evidence of the right type, quality, relevance, and coverage for that claim and field. Consider counter-evidence, triangulation, source provenance, primary versus secondary evidence, and important omissions where relevant.

There is no universal minimum source count and no universal peer-reviewed-source ratio. Literature needs vary by field, article type, claim breadth, evidence base, and venue. A review may apply a numeric requirement only when an identified target venue, reporting standard, or protocol explicitly imposes it; cite that authority and do not generalize it beyond its scope.

Dimension 4: Argument Coherence

Judge whether the problem, gap, research question, method, findings, and implications form a traceable argument. Identify unsupported logical transitions, conclusions that exceed the evidence, unresolved counterarguments, and contradictions. Do not confuse a familiar rhetorical structure with a sound argument.

Dimension 5: Writing Quality

Judge whether the manuscript communicates its reasoning precisely enough to be reviewed and used. Separate presentation problems from substantive research quality, and do not penalize non-native phrasing when meaning remains clear. Identify only issues that materially affect interpretation, verification, or venue requirements; route copyediting-level points as minor issues.

Dimension 6: Literature Integration

Judge whether the manuscript identifies and critically integrates the literature needed to establish its question, conceptual lineage, alternatives, and contribution. Coverage is assessed relative to the claims and field, not by a fixed number of references. Missing work should be named or bounded by a clearly described literature area whenever possible.

Dimension 7: Significance and Impact

Judge whether the claimed theoretical, empirical, practical, or policy implications follow from the evidence and matter for the stated audience. Separate demonstrated significance from speculative future impact; do not require cross-field reach when the target criterion values a focused contribution.


Narrative synthesis

The synthesis must preserve disagreements and non-compensatory weaknesses. A strong judgement on one criterion cannot numerically cancel a failure on another. Report:

  1. criteria positively verified;
  2. unresolved decision-bearing criteria, with evidence anchors;
  3. repairability and the work needed to satisfy each criterion;
  4. material uncertainty or reviewer-scope limitations; and
  5. the resulting recommendation under the applicable decision standard.

If the evidence does not support a judgement, use NOT_ASSESSED or state the uncertainty. Do not manufacture precision by selecting a midpoint or averaging reviewers.

Source: SKILL.md on GitHub

1 warning6d5 checks · Risk SAFE
  • Gen Agent Trust Hub6d

    The skill is a multi-agent framework for academic paper review. It is well-architected with significant security defenses against prompt injection from the manuscripts it processes. The primary risk is the large attack surface provided by untrusted input data, though this is mitigated by explicit boundary instructions.

  • Socket6d

    No alerts

  • Snyk6d

    Risk: MEDIUM · 1 issue

  • Runlayer6mo

    18 files scanned · No issues

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at a3f6569. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 33 minutes ago.

Activeupdated last week
Other metadata
metadata
{
  "version": "1.11.1",
  "last_updated": "2026-08-15",
  "status": "active",
  "data_access_level": "raw",
  "task_type": "open-ended",
  "related_skills": [
    "academic-paper",
    "academic-pipeline"
  ]
}

README badge

README badge for imbad0202/academic-research-skills/academic-paper-reviewer