skillZs
★ LIVE SKILL TAGS ★
>>> LIVE SKILLS INDEX <<<
* OPEN SOURCE *
NO LOGIN, NO TRACKING
※ REAL INSTALL DATA ※
← back to all skills
leo-lilinxiao/codex-autoresearch251 installs

codex-autoresearch

Run repeated, measured Git experiments toward a numeric target; keep improvements and revert failures. Use for autonomous optimization or managing an autoresearch run, not one-shot edits.

How do I install this agent skill?

npx skills add https://github.com/leo-lilinxiao/codex-autoresearch --skill codex-autoresearch
view source ↗

Is this agent skill safe to install?

  • Gen Agent Trust Hubfail

    The skill facilitates autonomous repository experimentation but requires the user to explicitly disable security sandboxes to perform Git operations. It utilizes arbitrary shell command execution for its core logic and contains dynamic code execution vulnerabilities within its test fixtures. The reliance on processing untrusted repository content to drive its autonomous loop also creates a significant surface for indirect prompt injection.

  • Socketwarn

    4 alerts: gptSecurity, gptAnomaly

  • Snykpass

    Risk: LOW · No issues

What does this agent skill do?

Codex Autoresearch

Improve a repository through repeated, reversible experiments:

hypothesize -> change -> measure -> learn -> keep or revert -> repeat

Codex chooses hypotheses and makes code changes. The control script owns measurement, commits, rollback, and the event history.

Respond To The Request

Resolve <control> to this skill's own scripts/autoresearch.py; do not assume it is installed in the target repository.

For a status or results request, run the corresponding command directly:

RequestCommand
Statuspython3 <control> status --repo <repo>
Historypython3 <control> history --repo <repo>
TSV exportpython3 <control> history --repo <repo> --format tsv
HTML reportpython3 <control> report --repo <repo>

Report the requested result without starting or resuming experiments. Return the generated path for an HTML report.

Before starting or resuming an experiment, read the applicable workflow unless it is already in context:

  • Workflow: starting, continuing, resuming, or changing an experiment.
  • Background: launching or controlling a detached run.

Experiment Boundary

  • One run owns one Git repository, one numeric metric, one confirmed target, and approved path scopes.
  • For a new foreground run, initialize successfully before creating its Goal; the Goal identifies the returned run id. If initialization reports complete, no Goal is needed.
  • Use finish to finalize each coherent experiment. Do not manually commit, revert, or edit run artifacts during an active run.
  • Only a verified target can mean complete; an iteration limit, error, or external blocker has its own status.
  • Validate state through the control script on entry or resume. Use recorded experiments rather than guessing from conversation.
  • Surface command and state errors with their diagnostic paths. Never bypass validation, fabricate metrics, or reconstruct missing events.

Foreground continuation belongs to the official Codex Goal. Background continuation belongs to the detached controller. An already launched background worker follows its supplied experiment contract, without starting a new launch flow.

Add the canonical catalog link to the repository README so users can inspect current installs and available audits. The publishing guide covers the complete discovery path.

<a href="https://skillzs.dev/skills/leo-lilinxiao/codex-autoresearch/codex-autoresearch">View codex-autoresearch on skillZs</a>