Sep 16, 2026

Write an arXiv-Ready Review Paper with Gated Citations Using latex-arxiv-SKILL

An agentic harness that drives Claude Code or Codex through a gated, issue-driven pipeline for ML/AI review papers — every citation web-checked before it reaches ref.bib.

#tutorial#research#python

A machine-learning review paper is two jobs stapled together: reading enough literature to have an opinion, and keeping a LaTeX project honest citation by citation. The latex-arxiv-SKILL harness automates the discipline around the second job — it has around 400 stars on GitHub and runs in both Claude Code and OpenAI Codex.

Why This Skill Matters

The repo packages itself as an "agentic harness for writing machine-learning and AI review papers in LaTeX". You give your coding agent a topic, and it drives a fixed pipeline: a research snapshot (10 to 20 papers, no prose), an IEEEtran project scaffold with a draft plan and candidate titles, a human approval gate, then an issue-by-issue writing loop where every \cite{} is web-checked against a live source before it enters ref.bib.

The guardrails are the point. The agent cannot write a single paragraph into main.tex until you approve the plan and an issues CSV exists. Every section becomes a tracked issue with target citations and acceptance criteria, and nothing is marked DONE until the criteria are met. Delivery requires a clean pdflatex and bibtex build with zero undefined-citation warnings.

The bundle follows the repo's own portable Agent Skills standard, so the same files run in Codex and Claude Code. The deterministic parts — scaffolding, plan and issue generation, arXiv discovery (with a local SQLite cache), validation, compilation — are Python scripts, and the issue-driven workflow is inspired by the same author's appautomaton/agent-designer.

Installation

The README documents no install script or package command. Clone the repository and work from it with your agent:

git clone https://github.com/appautomaton/latex-arxiv-SKILL.git
cd latex-arxiv-SKILL

The repo root carries an AGENTS.md and the four skill bundles under .codex/skills/ (arxiv-paper-writer, latex-rhythm-refiner, collaborating-with-claude, collaborating-with-gemini), with a .claude symlink alongside. The requirements section lists a working LaTeX environment (pdflatex and bibtex, or latexmk), Python 3.8+, an agent runtime with skills enabled, and web search and browsing for citation verification. The author tested on macOS with GPT-5.2 (Extra High).

Real Workflow: Generate a Review Paper in Two Prompts

The README's quickstart produced the example paper in example/v0-single-SKILL/ with two prompts. Start the paper:

write a review article for arxiv that is about SOTA generative image models

The agent does an initial literature pass, drafts a section framework, proposes candidate titles, and writes a plan/<timestamp>-<slug>.md file containing clarification questions. Open that file and answer the questions to steer scope, title, and coverage.

Then delegate the decisions and proceed:

I will let you choose the best title and the topics and inclusion of material that you see the best fit

Even with the plan questions ignored, the harness makes best-effort choices and produces a complete, compiling LaTeX project — main.tex, ref.bib, and main.pdf. The repo ships two generated examples: a generative-image-models review with 55 verified citations, and a video-world-simulators (3D/4D) review with 81 verified citations from a multi-skill run using the SQLite arXiv registry.

Real Workflow: Audit Citations on an Existing LaTeX Project

You do not have to start from zero. Point the harness at an existing LaTeX project and it runs a citation-validation pass that audits and repairs ref.bib without re-scaffolding anything. Claims without evidence become TODOs rather than fabricated references — the FAQ puts this under guardrails: any citation that cannot be verified against a live source does not get added.

Tips

  • After the plan is written, actually answer the clarification questions in plan/ — that is the one manual lever the gate gives you before writing starts.
  • The issues/*.csv file is the single source of truth for progress; the agent splits or inserts issues as scope grows instead of doing untracked work.
  • latex-rhythm-refiner post-processes prose to vary sentence and paragraph rhythm while preserving every citation — run it before the final compile pass.
  • Review and survey articles are the sweet spot out of the box; original or experimental papers work with tailoring, per the FAQ, by shaping the plan and inputs to your goal.

When Not to Use This

You need a working LaTeX environment and Python before anything runs — this is not a hosted service. If your target is a short blog post or an internal memo, the gated pipeline is overhead without payoff. And if you want the agent to freewrite a draft you will fact-check yourself afterwards, the hard gates will fight you the whole way.


See the leaderboard for more skills.