hallucination skills

23 tagged hallucination, measured the same way as everything else here.

Browse within: frontend-development 12frontend-web 12

anti-lie

01

lc198707/anti-lie

Skill Claude CodeCodex

Use when audited OpenClaw conversations or outgoing Feishu/Slack/channel messages contain concrete business numbers such as revenue, percentages, stock prices, contract values, costs, market share, or funding and need evidence checks plus red/yellow/green audit stamps after sends.

89 3mo ago A 57 tokens

clarity-gate

02

frmoretto/clarity-gate

Skill Claude CodeCodex

Pre-ingestion verification for epistemic quality in RAG systems. Ensures documents are properly qualified before entering knowledge bases. Produces CGD (Clarity-Gated Documents) and validates SOT (Source of Truth) files.

33 6mo ago A 51 tokens

thePM001/AEP-agent-element-protocol

Skill Claude CodeCodex

Use when creating a new AepCaw security policy, including agent sandboxes, CI pipelines, development environments, HTTP service gateways, or Postgres-family database access policies.

5 5d ago A 41 tokens original Apache-2.0

aep-caw-policy-edit

04

thePM001/AEP-agent-element-protocol

Skill Claude CodeCodex

Use when adding, removing, or updating rules in an existing AepCaw policy, modifying security permissions, HTTP service declarations, Postgres-family database rules, resource limits, or policy YAML files.

5 5d ago A 46 tokens copy · 94% Apache-2.0

aep

05

thePM001/AEP-agent-element-protocol

Skill Claude CodeCodex

Use this skill whenever working with AEP (Agent Element Protocol) 2.8, dynAEP (main AEP event runtime), Base Node, Composer Lite, CCA / setup agent, component registry, CAW, UCB, Path A/B connect, dynAEP-TA, dynAEP-TA-P or any AEP governance feature. Triggers include 'AEP', 'dynAEP', 'dynAEP-TA', 'dynAEP-TA-P'…

5 5d ago A 411 tokens original Apache-2.0

rag-evaluation

06

karthikrshet/aiskills

Skill Claude CodeCodex

Use this skill to evaluate the quality of a RAG pipeline on faithfulness, answer relevancy, context precision, context recall, and hallucination rate. Activates after a RAG system is implemented or when retrieval quality is in question. Produces a structured evaluation report with measurable results.

3 12d ago A 63 tokens

belief-assessor

07

hqzzdsda/belief-state-runtime

Skill Claude CodeCodex

LLM-driven epistemic reasoning engine. Evaluates claims against evidence, outputs calibrated confidence and structured belief state (VERIFIED/CONTESTED/UNCERTAIN). v2 adds 4-way constraint system, parameterized configuration, and formula-based confidence intervals. Use when the agent needs to assess whether…

2 2mo ago A 77 tokens

hqzzdsda/belief-state-runtime

Skill Claude CodeCodex

LLM-driven epistemic reasoning engine. Evaluates claims against evidence, outputs calibrated confidence and structured belief state (VERIFIED/CONTESTED/UNCERTAIN). Use when the agent needs to assess whether information is trustworthy, detect contradictions in evidence, or quantify uncertainty.

2 2mo ago A 58 tokens

fact-check

09

Nlai741533/EFC-Plugin

Skill Claude CodeCodex

Systematically fact-check AI-generated research reports and data-heavy documents. Trigger when the user asks to verify, fact-check, validate, or audit a report or document — especially those containing financial figures, market data, or claims sourced from web searches. Not for code review or general editing.

2 3mo ago B 60 tokens original MIT

tooloftruth

10

adigoel07/tooloftruth

Skill Claude CodeCodex

Use when verifying that a tool or skill was actually used, checking tool installation, or generating verification receipts. Triggers on 'verify', 'truth', 'proof', 'did it actually use', 'check tool', 'tooloftruth', '/truth'. Also triggers when the agent is about to claim a tool was used — run verification first.

0 15d ago A 73 tokens original MIT