calibration skills

42 tagged calibration, measured the same way as everything else here.

Browse within: Evaluation 12agent-framework 12agent-memory 12benchmark 12bayesian-inference 5data-assimilation 5derivative-free-optimization 5ensemble-kalman-filter 5ensemble-kalman-inversion 5

base-show

01

CliMA/EnsembleKalmanProcesses.jl

Skill Claude CodeCodex

Add concise Base.show and Base.summary methods to Julia types whose default REPL representation is unhelpful or overwhelming. Use this skill whenever the user mentions that a type prints badly in the REPL, asks to improve how an object is displayed or printed, wants a custom show, summary, or repr for a Julia type, or…

118 4d ago A 170 tokens original Apache-2.0

docstrings

02

CliMA/EnsembleKalmanProcesses.jl

Skill Claude CodeCodex

Add or normalise Julia docstrings on public symbols (exported types, functions, and constants) so the package's public API is fully self-documenting and the Documenter.jl docs build passes its checkdocs check. After writing docstrings, also updates docs/src/API/ pages so every exported symbol appears exactly once…

118 4d ago A 167 tokens original Apache-2.0

CliMA/EnsembleKalmanProcesses.jl

Skill Claude CodeCodex

Rewrite vague, delayed, or low-context Julia error messages into structured, actionable diagnostics. Invoke this skill whenever the user mentions: error message, improve errors, rewrite @assert, ArgumentError, DimensionMismatch, DomainError, vague error, error rewrite, Julia exception, diagnostic, throw, validation…

118 4d ago A 171 tokens original Apache-2.0

fable-thinking

04

mrgoonie/fable-thinking

Skill Claude CodeCodex

Reasoning protocol distilled from Claude Fable 5. Makes any model reason like Fable — evidence-grounded claims, multi-hypothesis diagnosis, concrete simulation, adversarial self-review, calibrated outcome-first delivery. Its never-skipped Floor check catches simple-looking trick questions models answer confidently…

42 1mo ago A 115 tokens original MIT

Avyayalaya/agent-council

Skill Claude CodeCodex

Use when you want a per-claim evidence-tier audit on a text artifact before it ships — assign T1-T6 tiers to every load-bearing claim, surface calibration mismatches (high confidence on weak evidence, or honesty-theater under-claiming), and flag P11 (citation-as-decoration), P17 (pile-of-anecdotes-as-evidence), P54…

10 1mo ago A 129 tokens original MIT

audit-post

06

DanceNitra/agora

Skill Claude CodeCodex

The end-to-end procedure for auditing one already-published post to a scientific-organization standard. It chains the adversarial skills and — critically — re-runs the auditor on the CORRECTED post to confirm it now passes clean before committing. We are a scientific organization; we do not ship missteps. Run EVERY…

3 today A 0 tokens original MIT

dungeon-os-threejs

07

DanceNitra/agora

Skill Claude CodeCodex

Build and extend the Dungeon OS 3D world (the live Three.js renderer in agora-game-server/static/index.html). Use whenever the task touches the dungeon scene, agent avatars in 3D, tiles/walls/props, lighting, the isometric camera, post-processing/bloom, sprites/labels, or "the game looks flat/cheap/laggy/should look…

3 today A 110 tokens original MIT

seo

08

DanceNitra/agora

Skill Claude CodeCodex

SEO + AI-search (AEO/GEO/LLMO) optimizer for the Agora storefront (dancenitra.github.io/agora) — a bilingual EN/SK static research blog on GitHub Pages. Synthesized from 37 SEO YouTube transcripts (vault 04 Resources/raw/YouTube Transcripts - SEO Ranking/) into an actionable pre-publish checklist + site-wide technical…

3 today A 0 tokens original MIT

wick-automate

09

agoradynamics/wick

Skill Claude CodeCodex

Detect when a task is being repeated often enough to be worth automating, then propose the right automation — a PROGRAM (deterministic script) or a SKILL (reusable judgment procedure). Analyzes the current session, the memory/ record, or an observation log; classifies program-vs-skill; estimates payoff; drafts the…

3 6d ago A 106 tokens

wick-catalog

10

agoradynamics/wick

Skill Claude CodeCodex

Extract structured fields from a source — paper, web page, API doc, or UI screenshot — and save a queryable record to memory/catalog/. Domain-agnostic. Use to build a personal index of things you'll want to find again, compare side-by-side, or reference in later decisions. Complements wick-research (which finds) and…

3 6d ago A 87 tokens

agoradynamics/wick

Skill Claude CodeCodex

Produce a one-paragraph narrative summary of a CHANGELOG.md version entry — suitable for release announcements, blog posts, or social copy. Preserves technical specifics (file paths, command names, counts) rather than vaguifying them. Defaults to the latest version; optionally takes a version like "v1.0.0". Refuses to…

3 6d ago A 83 tokens

swatplus-builder

12

AI-Hydro/Skills

Skill Claude CodeCodex

Use this skill when building, running, calibrating, validating, or diagnosing SWAT+ hydrological models. Triggers include SWAT+, watershed modeling, streamflow simulation, hydrograph calibration, USGS discharge matching, routing failures, and outlet selection debugging.

2 2mo ago A 57 tokens

log

13

allenc84/sapience

Skill Claude CodeCodex

Judgment ledger — log a prediction, review or resolve pending assessments, and generate calibration or bias maps from your track record.

0 1mo ago A 24 tokens original MIT

falsify

14

263311487-ux/falsify

Skill Claude CodeCodexCursor

The scientific thinking protocol for AI agents. Use when facing complex, ambiguous, or high-stakes questions where guessing is costly — technical design decisions, architecture choices, debugging theories, data claims, security judgments, or any answer the agent is tempted to give confidently without evidence.…

0 2d ago A 130 tokens original MIT