SkillAgentSearch skills...

paper-orchestra

Orchestrate the full PaperOrchestra (Song et al., 2026, arXiv:2604.05018) five-agent pipeline to turn unstructured research materials (idea, experimental log, LaTeX template, conference guidelines, optional figures) into a submission-ready LaTeX manuscript and compiled PDF.

Install / Use

npx skills add Ar9av/PaperOrchestra --skill paper-orchestra

Installs into whichever agent you are using.

About this skill
๐Ÿ“„

SKILL.md

Installable skill definition

Quality Score

92/100

Category

Automation

Supported Platforms

Universal

Our assessment of paper-orchestra

paper-orchestra scores 92/100 on our quality scale, 958th of 2,855 Automation skills we index (top 34%).

Its SKILL.md is 14 KB long, well organised into 22 sections with 7 code examples: a thorough specification that gives an agent plenty to work with.

It has 664 GitHub stars, a meaningful sign that others use it.

Substance
30/30
Structure
20/20
Description
15/15
Adoption
12/20
Freshness
15/15

Maintenance, license and trust

  • The repository was last updated 12 days ago, so paper-orchestra is actively maintained.
  • No license is declared. By default that means all rights are reserved: you can read it, but reusing or redistributing it is not clearly permitted. Ask the author before building on it commercially.
  • Its trust signals score 88/100, with 1 caution from licensing, adoption, age or documentation. These come from repository metadata, not a code audit โ€” read the skill file before letting an agent act on it.

Safety scan

No issues found

Our scan of the whole file found no instruction hijacking, hidden characters, credential access, data exfiltration or destructive commands.

Automated pattern scan on 2026-10-04. It catches known dangerous patterns, not every risk โ€” read a skill before letting an agent act on it.

paper-orchestra compared with similar skills

All 4 of these similar skills score higher than paper-orchestra; compare them before choosing.

SkillScoreStarsUpdatedFormat
paper-orchestra (this skill)by Ar9av9266412d agoSKILL.md
Agent-Reachby Panniantong10090.1k18d agoCLAUDE.md
Scraplingby D4Vinci10085.6k1d agoMCP Server
rufloby ruvnet10073.8ktodayMCP Server
algorithmic-artby anthropics100177.9k11d agoSKILL.md

Frequently asked questions

How do I install paper-orchestra?
Run npx skills add Ar9av/PaperOrchestra --skill paper-orchestra. The install tabs above show the steps for each supported agent.
Which AI agents does paper-orchestra work with?
It is written for Universal, as a SKILL.md file. Other agents that read the same format can often use it too.
Is paper-orchestra safe to use?
Our scan of the whole file found no instruction hijacking, hidden characters, credential access, data exfiltration or destructive commands. It declares no license and scores 88/100 on trust signals. Skills are instructions an agent will follow, so read the file before installing it and do not approve commands you do not understand.
Is paper-orchestra still maintained?
The repository was last updated 12 days ago, so paper-orchestra is actively maintained.

name: paper-orchestra description: Orchestrate the full PaperOrchestra (Song et al., 2026, arXiv:2604.05018) five-agent pipeline to turn unstructured research materials (idea, experimental log, LaTeX template, conference guidelines, optional figures) into a submission-ready LaTeX manuscript and compiled PDF. TRIGGER when the user asks to "write a paper from my experiments", "turn this idea and these results into a paper", "generate a conference submission", "run paper-orchestra on X", or otherwise wants the end-to-end paper-writing pipeline. Coordinates the outline-agent, plotting-agent, literature-review-agent, section-writing-agent, and content-refinement-agent skills. data_access_level: raw

paper-orchestra (Orchestrator)

Top-level driver for the PaperOrchestra pipeline. Read this document and follow the steps below. The detailed prompts and rules live in each sub-skill's SKILL.md and references/ directories โ€” you (the host agent) will load them as you go.

Source paper: Song et al., PaperOrchestra: A Multi-Agent Framework for Automated AI Research Paper Writing, arXiv:2604.05018, 2026. https://arxiv.org/pdf/2604.05018

What this skill produces

A complete submission package P = (paper.tex, paper.pdf) written into workspace/final/, plus a full audit trail under workspace/ (outline, figures, refs, drafts, refinement worklog, provenance snapshot).

Inputs (the (I, E, T, G, F) tuple from the paper)

The workspace MUST contain:

| File | Symbol | Required | Description | |---|---|---|---| | workspace/inputs/idea.md | I | yes | Idea Summary (Sparse or Dense variant โ€” see references/io-contract.md) | | workspace/inputs/experimental_log.md | E | yes | Experimental Log: setup, raw numeric data, qualitative observations | | workspace/inputs/template.tex | T | yes | LaTeX template for the target conference (with \section{...} commands) | | workspace/inputs/conference_guidelines.md | G | yes | Formatting rules, page limit, mandatory sections | | workspace/inputs/figures/ | F | no | Optional pre-existing figures. If empty, the plotting agent generates everything. |

scripts/init_workspace.py will scaffold this layout. scripts/validate_inputs.py will check it before the pipeline runs.

Pipeline (read references/pipeline.md for the full diagram)

Step 1: Outline           โ”€โ”€โ–ถ  outline.json                       (1 LLM call)
Step 2: Plotting     โ”€โ”
                      โ”œโ”€โ”€โ–ถ  figures/*.png + captions.json         (~20-30 calls)
Step 3: Lit Review   โ”€โ”˜                                           (~20-30 calls)
                          intro_relwork.tex + refs.bib

Step 4: Section Writing  โ”€โ”€โ–ถ  drafts/paper.tex                    (1 LLM call)
Step 5: Content Refine   โ”€โ”€โ–ถ  final/paper.tex + final/paper.pdf   (~5-7 calls, ~3 iters)

Step 2 and Step 3 are independent and MUST run in parallel when your host supports parallel sub-agents. If not, run Step 3 first (it has the longer wall time due to Semantic Scholar rate limits) and Step 2 second.

Critical pre-instruction (read once, apply always)

Before any LLM call that writes paper content (outline, intro/related work, section writing, refinement), you MUST prepend the Anti-Leakage Prompt at references/anti-leakage-prompt.md to your system prompt. This is verbatim from Appendix D.4 of the paper and prevents pre-training-data leakage. The paper applies it uniformly across all baselines for fair comparison; we apply it for fidelity and to keep generated papers grounded in the user's inputs.

Step-by-step execution

0. Pre-flight Checks

Before running the pipeline, perform the following quality gates in order:

# 1. Scaffold the workspace
python skills/paper-orchestra/scripts/init_workspace.py --out workspace/
# user drops their inputs into workspace/inputs/

# 2. Validate required files are present and well-formed
python skills/paper-orchestra/scripts/validate_inputs.py --workspace workspace/

# 3. Check input density โ€” idea and experimental log must meet minimum thresholds
python skills/paper-orchestra/scripts/check_idea_density.py \
    --idea workspace/inputs/idea.md \
    --log workspace/inputs/experimental_log.md

# 4. Cross-validate consistency between idea and experimental log
python skills/paper-orchestra/scripts/validate_consistency.py \
    --idea workspace/inputs/idea.md \
    --log workspace/inputs/experimental_log.md

If validate_inputs.py or check_idea_density.py fail (exit code 1 or 2), stop and tell the user what's missing or below threshold โ€” do not proceed until fixed.

validate_consistency.py produces warnings only (exit code 1 = WARN, non-blocking); report warnings to the user but continue.

Before failing on missing inputs, check whether aggregation can supply them:

| Inputs state | Action | |---|---| | idea.md and experimental_log.md both present and non-empty | Continue to Step 1. | | Either is missing/empty, and the user mentioned a directory | Load and run agent-research-aggregator with that directory as --search-roots, then re-validate. | | Either is missing/empty, no directory mentioned | Ask the user: "Your workspace is missing idea.md / experimental_log.md. Do you have a folder with research notes or agent history I can aggregate from? If so, tell me the path โ€” or drop the files manually into workspace/inputs/." |

If validation still fails after aggregation (e.g. template.tex or conference_guidelines.md are missing), stop and tell the user exactly which files remain outstanding.

Also probe the TeX installation (once per workspace, result cached):

python skills/paper-orchestra/scripts/check_tex_packages.py \
    --out workspace/tex_profile.json

The Section Writing Agent reads tex_profile.json to decide which LaTeX patterns to use (e.g., Figure~\ref{} vs \cref{}, whether to include \usepackage{microtype}, etc.). This eliminates compile-time package failures that previously required iterative manual edits.

1. Outline (Step 1 โ€” 1 LLM call)

Load skills/outline-agent/SKILL.md and follow it. Output: workspace/outline.json. Validate with python skills/outline-agent/scripts/validate_outline.py workspace/outline.json. Halt the pipeline if validation fails โ€” every downstream agent depends on the schema.

2 โˆฅ 3. Plotting and Literature Review (in parallel)

Parse outline.json. Extract:

  • outline.plotting_plan โ†’ drives Step 2
  • outline.intro_related_work_plan โ†’ drives Step 3

If your host supports parallel sub-agents (Claude Code's Agent tool with multiple concurrent calls; Cursor's parallel agents; Antigravity's worker pool), spawn two concurrent sub-tasks:

  • Sub-task A: load skills/plotting-agent/SKILL.md, execute the plotting plan, produce workspace/figures/<figure_id>.png for every entry, plus workspace/figures/captions.json.
  • Sub-task B: load skills/literature-review-agent/SKILL.md, execute the research strategy, produce workspace/drafts/intro_relwork.tex and workspace/refs.bib.

If your host does not support parallel sub-agents, run Sub-task B first (it has slower wall-clock due to Semantic Scholar QPS limits) then Sub-task A. The artifacts are independent, so order doesn't affect correctness.

3.5. Outline Reconciliation (after Step 3 completes, before Step 4)

Once Step 3 (Literature Review) has produced citation_pool.json and cross_verification_report.json, run the reconciliation step.

Load references/outline-reconciliation.md and follow its prompt. Output: workspace/outline_reconciled.json.

Validate and diff:

python skills/outline-agent/scripts/validate_outline.py workspace/outline_reconciled.json
python skills/paper-orchestra/scripts/diff_outlines.py \
    --original   workspace/outline.json \
    --reconciled workspace/outline_reconciled.json \
    --summary    workspace/reconciliation_summary.md

If validation fails, fall back to outline.json for Step 4 and warn the user. Show the user the reconciliation_summary.md (even if no changes โ€” it confirms the outline matched the actual literature).

Skip conditions: citation pool empty, Step 3 failed, or Step 2 is still running and the host cannot issue another call concurrently. See references/outline-reconciliation.md for full skip conditions.

4. Section Writing (Step 4 โ€” ONE single multimodal LLM call)

Load skills/section-writing-agent/SKILL.md and follow it. This is one single call in the paper (App. B: "Section Writing Agent (1 call)") โ€” do not split it per section. The agent receives:

  • outline_reconciled.json (use this if it exists; fall back to outline.json)
  • idea.md, experimental_log.md
  • intro_relwork.tex (already-filled from Step 3 โ€” preserve verbatim)
  • refs.bib (the citation map)
  • conference_guidelines.md
  • research_brief.md (if it exists โ€” read ยง1โ€“ยง3 for accumulated pipeline context)
  • The actual figure image files from workspace/figures/ (multimodal input)

Output: workspace/drafts/paper.tex (a complete LaTeX document).

Then run the deterministic gates:

python skills/section-writing-agent/scripts/orphan_cite_gate.py workspace/drafts/paper.tex workspace/refs.bib
python skills/section-writing-agent/scripts/latex_sanity.py workspace/drafts/paper.tex
python skills/paper-orchestra/scripts/anti_leakage_check.py workspace/drafts/paper.tex
python skills/paper-orchestra/scripts/claim_evidence_gate.py \
    --paper workspace/drafts/paper.tex \
    --log   workspace/inputs/experimental_log.md \
    --out   workspace/claim_evidence_report.json

claim_evidence_gate.py is a WARN gate (exit 1 = warnings, not a hard stop). Report the count of unsupported claims to the user. The content-refinement agent will address them in Step 5.

If any gate fails, the host agent must fix the issue (re-prompting the writing step with the gate's error report) before proceeding.

5. Content Refinement (Step 5 โ€” ~3 iterations, ~5-7 calls)

Load skills/content-refinement-agent/SKILL.md and follow it. The skill implements the loop with strict halt rules from halt-rules.md. Maintain workspace/refinement/worklog.json and snapshot each iteration into workspace/refinement/iter<N>/.

Halt conditions (any one triggers the loop to stop and accept the current best snapshot):

  1. Iteration count reaches the cap (default 3, see halt-rules.md).
  2. Overall score from the simulated reviewer decreases vs the previous iteration โ†’ revert to previous snapshot, halt.
  3. Overall score ties but at least one sub-axis decreases while none gain compensatingly (negative net sub-axis change) โ†’ revert, halt.
  4. Reviewer issues no new actionable weaknesses.

The accepted snapshot is copied to workspace/final/paper.tex.

6. Compile and finalize

cd workspace/final && latexmk -pdf paper.tex

Then write workspace/provenance.json capturing input file hashes, outline hash, refs hash, figure hashes, and final tex/pdf hashes (helper: scripts/snapshot.py in the orchestrator scripts dir if you want a one-shot; otherwise the host agent computes hashes inline).

Report to the user: the path to workspace/final/paper.pdf, a brief summary of which sections were drafted, citation count, refinement iterations completed, and any gates that failed mid-pipeline.

Workspace layout

See references/io-contract.md. Summary:

workspace/
โ”œโ”€โ”€ inputs/                          # user-provided
โ”‚   โ”œโ”€โ”€ idea.md
โ”‚   โ”œโ”€โ”€ experimental_log.md
โ”‚   โ”œโ”€โ”€ template.tex
โ”‚   โ”œโ”€โ”€ conference_guidelines.md
โ”‚   โ””โ”€โ”€ figures/                     # optional pre-existing figures
โ”œโ”€โ”€ outline.json                     # Step 1 output
โ”œโ”€โ”€ figures/                         # Step 2 output
โ”‚   โ”œโ”€โ”€ <figure_id>.png
โ”‚   โ””โ”€โ”€ captions.json
โ”œโ”€โ”€ refs.bib                         # Step 3 output
โ”œโ”€โ”€ drafts/                          # Step 3 + Step 4 output
โ”‚   โ”œโ”€โ”€ intro_relwork.tex
โ”‚   โ””โ”€โ”€ paper.tex
โ”œโ”€โ”€ refi

Truncated for display โ€” read the full file on GitHub.

Related Skills

View on GitHub
GitHub Stars664
CategoryAutomation
Updated12d ago
Forks92

Languages

Python

Trust signals

88/100

From repository metadata: license, adoption, age and documentation. Not a code audit โ€” see the Safety scan above for what the skill file itself contains.

1 medium
paper-orchestra โ€” Universal Skill: Install & Safety Check | SkillAgent