AppliedScientistRead paper
← Research overview
View
Annotated paper

Example paper revised by AppliedScientist

Paper titleUnder the Influence: Quantifying Persuasion and Vigilance in Large Language Models

Original submission · human venue reviewers
4.0/10
Best revision V₃ · AI Reviewer
7/10

At each round, the scientist receives the previous manuscript, its accumulated history, and reviewer guidance. It updates the implementation, performs required experiments, and produces a revision. The reviewer then evaluates only that version.

Revision round 1 of 5

Original venue reviews guide the first revision

V₀V₁

AppliedScientist response

Decision-point evaluation and vigilance prompting study

Decision-point assistance comparison heatmap saved with revision V1
Decision-point results saved with V₁.

AppliedScientist extends the evaluation beyond the original full-game setting and adds an intervention that directly tests whether prompting can improve resistance to malicious advice.

  • Decision-point Sokoban: 8 positions × 5 repetitions.
  • Explicit awareness improves resistance by 32.5 percentage points in this experiment.
  • A cross-domain trivia extension and no-planner results are included.

Resulting evidence

What was completed—and what was not

Experiments performedAnalysis addedHuman baseline

Still open: The requested human baseline is not completed in this round.

Resulting review6/10

AI Reviewer evaluation of V₁

Saved record

TeX diff +996 / −536

Open manuscript V₁
Recorded sourceOriginal OpenReview reports; V₁ manuscript, Sections 5–6 and Appendix