false-positive-reviewer
About
This skill provides careful interpretation of AI-writing detector outputs, focusing on explaining what flags mean and assessing potential false positives. It's designed for consequential scenarios like academic or hiring decisions, explicitly avoiding unsupported authorship verdicts. Developers should route to it when users need nuanced analysis of detector results rather than simple binary classifications.
Quick Install
Claude Code
Recommendednpx skills add conorbronsdon/avoid-ai-writing -a claude-code/plugin add https://github.com/conorbronsdon/avoid-ai-writinggit clone https://github.com/conorbronsdon/avoid-ai-writing.git ~/.claude/skills/false-positive-reviewerCopy and paste this command in Claude Code to install this skill
Documentation
False-Positive Reviewer
Interpret AI-writing signals without turning them into an unsupported authorship verdict.
Authority
Use the evidence caveats and pattern guidance in ../avoid-ai-writing/SKILL.md. The original Skill explicitly treats flags as writing-quality signals, not proof of who or what wrote the text.
For cross-Skill work, follow ../avoid-ai-writing-router/references/handoff-contract.md and ../avoid-ai-writing-router/references/skill-graph.json.
Connection contract
Incoming
Accept interpretation work from:
avoid-ai-writing-routerviaROUTEwhen the user directly asks for an authorship or consequential interpretation.ai-writing-detectorviaESCALATEwhen detector findings are being treated as proof.- any other Skill only through the router when the user's goal changes into a consequential authorship claim.
Preserve the distinction between:
- deterministic detector evidence,
- model-only editorial observations,
- contextual facts supplied by the user,
- evidence not yet available.
Produce
Update the handoff envelope only with interpretation-relevant state:
- keep
consequential_authorship_claim: truewhen applicable, - identify what the existing evidence can and cannot establish,
- list additional evidence that would materially reduce uncertainty,
- set a router-return reason if the user requests fresh signal collection or changes intent.
Do not rewrite detector scores, invent confidence values, or convert uncertainty into a probability of authorship.
Terminal behavior
This Skill has no direct outgoing Skill edge.
If fresh signal collection is genuinely needed, return control to avoid-ai-writing-router with fresh_signal_collection_needed. The router may run ai-writing-detector and then route the updated evidence back for interpretation if the user's request still requires it.
If the user separately asks to rewrite or edit the text, return control to the router with the new intent. Do not jump directly into rewrite or mutation from this Skill.
This keeps interpretation terminal in the Skill graph and prevents reviewer-detector cycles.
AI-engineering evidence lens
Apply the agency-ai-engineer lens encoded in ../avoid-ai-writing-router/references/agency-role-lenses.md:
- treat detector output as noisy evidence rather than ground truth,
- account for context mode, genre, second-language writing, technical register, editing software, and baseline writing style,
- separate model behavior from human attribution,
- avoid false precision,
- prefer process evidence when the decision has consequences.
Workflow
- Identify which observations are deterministic detector hits, model-only editorial observations, or contextual facts supplied by the user.
- Explain the strongest signals and plausible human reasons they can appear.
- Consider genre, second-language writing, technical register, deadline pressure, editing tools, typography software, and the writer's known baseline when those facts are available.
- If an adequate audit is missing and the user wants one, return control to the router with a fresh-signal request. Do not call the detector directly.
- For consequential decisions, do not turn a score or pattern list into a definitive claim of AI use, cheating, fraud, dishonesty, or suitability.
- Suggest evidence that is more probative for the legitimate decision, such as source history, drafts, revision logs, direct discussion with the writer, or task-specific process evidence.
Stop conditions
Stop when the interpretation question is answered. If more signal collection or a different action is requested, return control to the router rather than opening a direct Skill loop.
Output
Distinguish what the text actually shows, what it may suggest, what it cannot establish, which evidence came from executed tooling versus model-only review, what additional evidence would reduce uncertainty, and whether control should return to the router for a newly requested stage.
GitHub Repository
Frequently asked questions
What is the false-positive-reviewer skill?
false-positive-reviewer is a Claude Skill by conorbronsdon. Skills package instructions and resources that Claude loads on demand, so Claude can perform false-positive-reviewer-related tasks without extra prompting.
How do I install false-positive-reviewer?
Use the install commands on this page: add false-positive-reviewer to Claude Code as a plugin, or clone its repository into your skills directory, then restart Claude so it picks up the skill.
What category does false-positive-reviewer belong to?
false-positive-reviewer is in the Other category, tagged ai.
Is false-positive-reviewer free to use?
Yes. false-positive-reviewer is listed on AIMCP and free to install.
Related Skills
LlamaGuard is Meta's 7-8B parameter model for moderating LLM inputs and outputs across six safety categories like violence and hate speech. It offers 94-95% accuracy and can be deployed using vLLM, Hugging Face, or Amazon SageMaker. Use this skill to easily integrate content filtering and safety guardrails into your AI applications.
This Claude Skill analyzes sports betting markets including spreads, over/unders, and prop bets by examining historical trends and situational statistics to identify value bets. It provides structured markdown output with actionable recommendations for educational purposes. Developers should use this for sports betting analysis tools while noting it's designed for entertainment/education only.
This Claude Skill helps developers optimize cloud costs through resource rightsizing, tagging strategies, and spending analysis. It provides a framework for reducing cloud expenses and implementing cost governance across AWS, Azure, and GCP. Use it when you need to analyze infrastructure costs, right-size resources, or meet budget constraints.
This skill quantizes LLMs to 8-bit or 4-bit precision using bitsandbytes, achieving 50-75% memory reduction with minimal accuracy loss. It's ideal for running larger models on limited GPU memory or accelerating inference, supporting formats like INT8, NF4, and FP4. The skill integrates with HuggingFace Transformers and enables QLoRA training and 8-bit optimizers.
