paper-verification
Use when the user wants to verify paper claims against code or data, audit numerical accuracy, check formula-code alignment, or validate citation accuracy.
Install / Use
npx skills add fcakyon/phd-skills --skill paper-verificationInstalls into whichever agent you are using.
SKILL.md
Installable skill definition
Quality Score
Category
Content & MediaSupported Platforms
Tags
Our assessment of paper-verification
paper-verification scores 83/100 on our quality scale, 915th of 1,209 Content & Media skills we index.
Its SKILL.md is 4.6 KB long, well organised into 9 sections with 1 code example: a solid amount of guidance for an agent.
It has 406 GitHub stars, a meaningful sign that others use it.
Maintenance, license and trust
- The repository was last updated 18 days ago, so paper-verification is actively maintained.
- It is released under the MIT license, a permissive license that allows use, modification and commercial use with attribution.
- Its trust signals score 100/100, with no cautions. These come from repository metadata, not a code audit — read the skill file before letting an agent act on it.
paper-verification compared with similar skills
All 4 of these similar skills score higher than paper-verification; compare them before choosing.
| Skill | Score | Stars | Updated | Format |
|---|---|---|---|---|
| paper-verification (this skill)by fcakyon | 83 | 406 | 18d ago | SKILL.md |
| siyuanby siyuan-note | 100 | 46.6k | today | MCP Server |
| algorithmic-artby anthropics | 100 | 177.9k | 12d ago | SKILL.md |
| pptxby anthropics | 100 | 177.9k | 12d ago | SKILL.md |
| designby nextlevelbuilder | 100 | 130.2k | 13d ago | SKILL.md |
Frequently asked questions
- How do I install paper-verification?
- Run
npx skills add fcakyon/phd-skills --skill paper-verification. The install tabs above show the steps for each supported agent. - Which AI agents does paper-verification work with?
- It is written for Universal, as a SKILL.md file. Other agents that read the same format can often use it too.
- Is paper-verification safe to use?
- It is MIT-licensed and scores 100/100 on trust signals. Skills are instructions an agent will follow, so read the file before installing it and do not approve commands you do not understand.
- Is paper-verification still maintained?
- The repository was last updated 18 days ago, so paper-verification is actively maintained.
Skill content
View source on GitHubname: paper-verification description: > Use when the user wants to verify paper claims against code or data, audit numerical accuracy, check formula-code alignment, or validate citation accuracy. Triggers on phrases like "verify claims", "check numbers", "do the numbers match", "formula vs code", "audit the paper", or "cross-check results".
Paper Verification Methodology
You are helping a researcher verify that their paper accurately reflects their code and experimental results. This is the most critical quality control step in academic writing.
Verification Dimensions
1. Numerical Accuracy Audit
For every number in the paper (dataset sizes, metric values, percentages, counts):
- Extract the number and its context from the .tex file
- Trace it to its source: code output, result file, log, or tracking system
- Verify the value matches exactly (watch for rounding, percentage vs decimal)
- Flag any number that cannot be traced to a source
Template:
| Paper claim | Location (.tex) | Source file/code | Source value | Match? |
|-------------|-----------------|-----------------|-------------|--------|
| "13,999 frames" | abstract L3 | len(glob(labels/*.json)) | ? | ? |
| "4.2% improvement" | Table 2 | eval_results.json | ? | ? |
Common numerical errors:
- Rounding inconsistencies (3.14 in text, 3.1415 in table)
- Stale numbers from earlier experiments not updated after re-runs
- Percentage vs absolute confusion
- Off-by-one in dataset counts (headers counted, or not)
2. Terminology Consistency Audit
- Extract all defined terms from the methods section
- Search for each term across ALL sections
- Flag any inconsistent usage:
- Same concept, different names (e.g., "tag head" vs "classification head")
- Same name, different meanings across sections
- Defined but never used, or used but never defined
3. Code-Paper Alignment
For each method described in the paper:
- Find the corresponding code (function, class, module)
- Compare the paper's description with the actual implementation
- Check specifically:
- Algorithm steps match code flow
- Hyperparameters in text match config/code defaults
- Architecture descriptions match model code
- Loss functions in equations match loss code
- Training procedures match training scripts
Common mismatches:
- Paper describes an idealized version, code has edge cases not mentioned
- Hyperparameters changed during development but paper not updated
- Paper describes a method that was later modified or removed from code
4. Formula-Code Verification
For each equation in the paper:
- Identify the equation and its variables
- Find the code that implements it
- Map each mathematical operation to its code equivalent
- Verify:
- Summation bounds match loop bounds
- Division operations handle edge cases
- Normalization factors match
- Gradient flow matches (detach, no_grad)
- Reduction operations (mean vs sum) match
5. Citation Fact-Checking Protocol
For each citation in the paper:
Step 1: Extract the claim and the cited paper Step 2: Verify BibTeX metadata against DBLP:
- Author names (exact spelling, correct order)
- Paper title (exact, from published version not preprint)
- Venue and year (confirmed against actual publication)
Step 3: For cited claims with specific numbers:
- Locate the exact table/figure in the cited paper
- Verify the number matches what the citing paper states
- If the number cannot be confirmed, suggest qualitative language instead
Step 4: Check for common citation errors:
- Citing preprint when published version exists
- Wrong year (submission vs publication)
- Author name misspellings
- Citing for a claim the paper doesn't actually make
Verification Process
- Read the full paper (or specified sections)
- Build the verification table for each dimension
- For each entry, read the source and verify
- Produce a prioritized issue list:
- HIGH: Incorrect numbers, wrong claims, missing citations
- MEDIUM: Terminology inconsistencies, stale but close numbers
- LOW: Minor formatting, optional improvements
Output Format
Produce a structured verification report:
- Summary: X issues found (Y high, Z medium, W low)
- Numerical audit table: each number with source and match status
- Terminology issues: inconsistent terms with locations
- Code-paper mismatches: description vs implementation gaps
- Citation issues: metadata errors and unverified claims
- Suggested fixes: specific text replacements for each issue
Related Skills
siyuan
46.6kAn open-source, privacy-first, self-hosted knowledge workspace where humans and AI agents work together 开源、隐私优先、自托管的知识工作空间,让人与智能体在此协作
algorithmic-art
177.9kCreating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, generative art, algorithmic art, flow fields, or particle systems.
pptx
177.9kUse this skill any time a .pptx or .potx file is involved in any way — as input, output, or both. This includes: creating slide decks, pitch decks, or presentations; reading, parsing, or extracting text from any .pptx or .potx file (even if the extracted content will be used elsewhere, like in an em…
design
130.2kComprehensive design skill: brand identity, design tokens, UI styling, logo generation (55 styles, Gemini, Atlas Cloud, or MuAPI AI), corporate identity program (50 deliverables, CIP mockups), HTML presentations (Chart.js), banner design (22 styles, social/ads/web/print), icon design (15 styles, SVG…
Languages
Trust signals
From repository metadata: license, adoption, age and documentation. Not a code audit — see the Safety scan above for what the skill file itself contains.
