๐Ÿ“
AutomationPython

Scoreable File First

by tangxiangru

Scoreable File First is an Automation skill for Claude Code, published by tangxiangru in AutoR.

804 stars25 forkson tangxiangru/AutoRAdded 2026/08/22Repository updated 2026/08/22
agentaiai-scientistauto-researchclaudeclaude-codecliharnessllmopenaipaperscience
Install in seconds
Install Scoreable File First
Copy Scoreable File First into your Claude Code skills folder. Run the command in your terminal, or review the source on GitHub before installing.
terminal
npx degit https://github.com/tangxiangru/AutoR/tree/main/src/skills/a-scoreable-file-in-the-first-hour ~/.claude/skills/a-scoreable-file-in-the-first-hour

Requires Node.js. Downloads this skill only โ€” not the rest of the repository โ€” into your Claude Code skills folder.

Without Node.js

git clone https://github.com/tangxiangru/AutoR.git

Clones the whole repository, then copy the skillโ€™s own directory into your skills folder yourself.

In this catalog

Source file
src/skills/a-scoreable-file-in-the-first-hour/SKILL.md in tangxiangru/AutoR
Installs to
~/.claude/skills/a-scoreable-file-in-the-first-hour
Collection
One of 25 skills cataloged from this repository
Category
Automation โ€” 1523 skills

What Scoreable File First does

Scoreable File First tells you to create a valid predictions file in the first hour and keep improving it in place. Use it when a run can be scored only if a submission file exists and stays valid.

Scoreable File First is cataloged under Automation on DirSkills. Scoreable File First comes from a repository tagged agent, ai, ai-scientist, auto-research and claude.

Documentation

README

Write a scoreable file before you write anything else

A predictions file that does not exist scores nothing. Not a low score โ€” no score, and on a benchmark that reports valid submission rate as a headline metric beside the score, a missing file costs you on two axes at once.

So the first version is not a milestone to work toward. It is a thing to get out of the way in the first hour, from whatever you can compute immediately, and then improve in place for the rest of the run.

Why this is not the obvious advice

The instinct is that a trivial submission is embarrassing and that a real one is close, so it is better to wait. Two measurements say otherwise.

This is the opening of the README. Read the full README on GitHub.

Frequently asked about Scoreable File First

  • What else does tangxiangru publish alongside Scoreable File First?

    Scoreable File First is one of 25 skills that DirSkills catalogs from tangxiangru/AutoR, the repository it ships in. Its siblings there include A Deliverable Is Not an Instruction, A Value You Did Not Measure Still Has A Source and Answer The Why, Not Only The What. Each one is a separate skill with its own page in this directory, installs the same way Scoreable File First does, and is maintained by tangxiangru in that same repository. The rest of the collection is listed on the tangxiangru/AutoR page.

  • How does Scoreable File First compare to other Automation skills?

    Scoreable File First ranks #988 by stars among the 1523 Automation skills in this catalog. The most-starred ones next to it are Autonomous Loops, Autonomous Agent Harness and Automation Audit Ops. DirSkills ranks by the star count of the repository each skill ships in, so that order reflects how popular those repositories are rather than any review of Scoreable File First against them. Open each page to compare what they document and how they install.

More from tangxiangru/AutoR

Scoreable File First is one of 25 skills cataloged on DirSkills from tangxiangru/AutoR.

See all 25 skills โ†’
๐Ÿงญ
1w ago

A Deliverable Is Not an Instruction

A Deliverable Is Not an Instruction helps separate research deliverables from harness instructions when building a study plan and checking coverage. Use it to keep `report_plan.json` focused on findings and to mark genuinely unreachable items honestly.
AI Engineering
80425
๐Ÿ“
1w ago

A Value You Did Not Measure Still Has A Source

A Value You Did Not Measure Still Has A Source explains how to handle deliverables this run cannot measure or produce, using attributed published values instead of fabricating or omitting them. It covers where to place cited values and how to compare them to your own results.
Writing
80425
๐Ÿ“
1w ago

Answer The Why, Not Only The What

Answer The Why, Not Only The What helps you write results and discussion sections that explain why an effect happens, not just whether it happened. Use it when a reviewer or task asks for mechanism claims, competing explanations, or limits on what the data support.
Writing
80425
๐Ÿงช
1w ago

Assume This Stage Is The Last One You Get

Assume This Stage Is The Last One You Get advises planning each early stage of a timed research run so it leaves a valid model, score, and rerunnable script if the later stages never happen. It is used when wall-clock limits make deferred work unlikely to execute.
AI Engineering
80425
๐Ÿ”ญ
1w ago

Astronomy Caption Specification

Astronomy Caption Specification turns figure captions into a plotting specification for reproducing astronomy figures when the rendered source is unavailable. Use it to keep panel order, series, colors, references, normalization, and error bars aligned with the paper.
Writing
80425
๐Ÿ”ญ
1w ago

Astronomy Error Budget Audit Trail

Astronomy Error Budget Audit Trail helps you document uncertainty propagation, fit bookkeeping, and residual diagnostics before quoting a result. Use it when a measurement depends on calibration chains, covariance, or a model fit.
Writing
80425