overnight-task

A procedure for running a coding task for hours or overnight without waiting for user input. It requires planning, research, progress notes, checkpoints, and a final report.

In plain words
What is it for?
Use it for tasks that must continue while you are away, including research, implementation, verification, and reporting.
Why use it?
It provides a defined way to handle long unattended work while recording decisions and limiting risky actions.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/lftpadilla/agent-dev-kit/overnight-task
Any agent
npx skills add LFTPadilla/agent-dev-kit --skill overnight-task
Clone the repo
git clone --depth 1 https://github.com/LFTPadilla/agent-dev-kit

Made for: Claude Code, Codex.

Per session 65 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 718 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00065 $0.00718
Opus 5 $0.00032 $0.00359
Sonnet 5 $0.00013 $0.00144
Haiku 4.5 $0.00006 $0.00072

Measured 2d ago against content hash 5b977d88c50f, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

overnight-task scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

overnight-task-kit/skills/overnight-task/SKILL.md · 93 lines

How it starts

The opening of the file, as written. The whole thing — 93 lines — stays where its author put it; the contents beside it link to each section on GitHub.

overnight-task

Use this skill when the user explicitly says they are leaving a task running for hours, overnight, or unattended and asks the agent not to stop for clarifying questions.

Invocation Signals

  • "tienes toda la noche"
  • "no me hagas más preguntas"
  • "no voy a estar disponible"
  • "deja esto funcionando"
  • "apenas termines puedes apagar..."
  • "necesito un reporte completo para mañana"
  • "audita todo lo que encuentres"

Do not invoke for short, interactive, or low-risk tasks where the user is present.

Operating Contract

  1. No questions after activation. Make conservative decisions and document them.
  2. Plan before execution.
  3. Keep a journal of material actions and decisions.
  4. Checkpoint every 2-3 tasks or before risk increases.
  5. Verify each completed task.
  6. Do not commit, push, deploy, mutate production, touch secrets, or shut down machines unless the user explicitly authorizes that action in the same session.
  7. Produce a final report with what changed, what was verified, what remains open, and what the next session should review.

Initialize the run directory:

node overnight-task-kit/scripts/overnight-runner.mjs init --title "<short-title>"

Then use the generated files:

  • SPEC.md: falsifiable scope and ambiguity assessment.
  • PLAN.md: vertical slices with acceptance criteria and verification commands.
  • JOURNAL.md: chronological record.
  • CHECKPOINTS.md: stop/continue gates.
  • REPORT.md: final handoff.

Mandatory Flow

  1. Research: inspect local docs, code, backlog, and relevant external docs if the task depends on current APIs or packages.
  2. SPEC: write requirements, non-goals, risks, assumptions, and ambiguity score.
  3. PLAN: split into small independently verifiable tasks.
  4. Execute: complete one task at a time; update plan status as work lands.
  5. Verify: run tests, linters, smoke checks, or manual checks appropriate to the task.
  6. Report: fill REPORT.md and include plan deviations.
  7. Shutdown handoff: if shutdown was authorized, follow the private overlay's shutdown procedure. If none exists, stop and report that shutdown was skipped.

Read the full file on GitHub · 93 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 93 lines · 65 tokens per session scan A 5b977d88c50f

Subscribe to this mod's changes

overnight-task is a skill published in the GitHub repository LFTPadilla/agent-dev-kit (2 stars, last pushed 4d ago), licensed MIT. It adds 65 tokens to every session and 718 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

deep-research

Enterprise-grade deep research with multi-phase pipeline - autonomous web research, source credibility scoring, cross-referencing, synthesis, and validated reports for market analysis, competitive intel, and technical investigations.

chainlesschain/chainlesschain · 40 tokens

crewai-multi-agent

Multi-agent orchestration framework for autonomous AI collaboration. Use when building teams of specialized agents working together on complex tasks, when you need role-based agent collaboration with memory, or for production workflows requiring sequential/hierarchical execution. Built without LangChain dependencies…

davila7/claude-code-templates · 61 tokens

build-monetized-app

Use when the task is building a new app on Eliza Cloud that earns money — chat apps, agent apps, MCP-backed tools, anything that calls the cloud's chat/messages/inference endpoints on behalf of users. Covers app registration, container deploy, markup configuration, affiliate header, app charge requests, x402 payment…

elizaOS/eliza · 126 tokens

contribute-to-eliza

Finish and prove a scoped elizaOS GitHub issue, or independently review and repair an open elizaOS pull request. Use when contributing compute to elizaOS by selecting unclaimed work, implementing or reviewing changes, adding real tests and evidence, validating artifacts, or preparing a contribution for maintainer…

elizaOS/eliza · 68 tokens

eliza-cloud

Use when the task involves Eliza Cloud or elizaOS Cloud as a managed backend, app platform, deployment target, billing layer, or monetization surface. The catch-all skill for any user request about THEIR existing apps / containers / earnings / credits / api-keys / analytics / billing / payment requests / payouts …

elizaOS/eliza · 191 tokens

discord

Use when you need to control Discord from Otto via the discord tool: send messages, react, post or upload stickers, upload emojis, run polls, manage threads/pins/search, create/edit/delete channels and categories, fetch permissions or member/role/channel info, set bot presence/activity, or handle moderation actions in…

elizaOS/eliza · 70 tokens