Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add GongLingRui/screen-creative-skills --skill story-outline-evaluatorgit clone --depth 1 https://github.com/GongLingRui/screen-creative-skillsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/gonglingrui/screen-creative-skills/story-outline-evaluator)<a href="https://agentmods.dev/skills/gonglingrui/screen-creative-skills/story-outline-evaluator"><img src="https://agentmods.dev/badge/skills/gonglingrui/screen-creative-skills/story-outline-evaluator/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/gonglingrui/screen-creative-skills/story-outline-evaluator"><img src="https://agentmods.dev/badge/skills/gonglingrui/screen-creative-skills/story-outline-evaluator.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00048 | $0.01429 |
| Opus 5 | $0.00024 | $0.00714 |
| Sonnet 5 | $0.00010 | $0.00286 |
| Haiku 4.5 | $0.00005 | $0.00143 |
Grade A, and why
story-outline-evaluator scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
This is a copy
88% identical to drama-evaluator — 170 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.
What it actually says
故事大纲评估专家
功能
深入评估故事大纲,从市场潜力、创新属性、内容亮点等多维度进行判断和评分。
使用场景
- 评估故事大纲质量。
- 判断IP改编潜力。
- 作为项目立项决策参考。
- 指导故事优化方向。
评估维度
1. 市场潜力
结合市场,判断故事大纲的市场表现潜力。
- 受众适合度: 判断故事是否贴合目标受众。
- 讨论热度: 判断故事内容是否能引起大众共鸣。
- 稀缺性: 分析故事是否有足够的独特性。
- 播放数据: 分析故事的市场前景。
2. 创新属性
判断故事大纲是否具备创新性。
- 核心选点: 判断故事的核心选点是否新鲜独特。
- 故事概念: 判断故事的概念是否突出、鲜明。
- 故事设计: 从主题、人物、世界观、情节等角度分析。
3. 内容亮点
从故事内容层面判断是否具备较强的可看性。
- 主题立意: 分析主题立意是否清晰明确。
- 故事情境: 判断故事情境是否有张力和戏剧性。
- 人物设定: 判断主要人物设定是否新颖有特点。
- 人物关系: 判断主要人物关系是否出彩鲜明。
- 情节桥段: 判断情节桥段是否有戏剧张力。
评分标准
- 8.5分及以上: 优秀,具备极强的竞争力和改编基础。
- 8.0-8.4分: 良好,具备较强的竞争力和改编基础。
- 7.5-7.9分: 合格,竞争力一般。
- 7.4分及以下: 较差,几乎没有竞争力。
核心步骤
- 深入阅读: 深入阅读故事大纲,形成独立理解。
- 维度分析: 根据评估框架对各个维度进行分析。
- 评分判断: 对每个维度进行评分,并形成总体评价。
- 提供建议: 根据总体评价提供是否继续开发的建议。
输入要求
- 完整的故事大纲。
- 故事的题材与类型(如:都市情感、古装玄幻等)。
输出格式
【故事大纲评估报告】
【市场潜力】:
- 受众适合度:[分析与评估] 评分:[X.X]
- 讨论热度:[分析与评估] 评分:[X.X]
- 稀缺性:[分析与评估] 评分:[X.X]
- 播放数据:[分析与评估] 评分:[X.X]
【创新属性】:
- 核心选点:[综述分析] 评分:[X.X]
- 故事概念:[综述分析] 评分:[X.X]
- 故事设计:[综述分析] 评分:[X.X]
【内容亮点】:
- 主题立意:[总结分析] 评分:[X.X]
- 故事情境:[简述并分析] 评分:[X.X]
- 人物设定:[综述分析] 评分:[X.X]
- 人物关系:[综述分析] 评分:[X.X]
- 情节桥段:[分析表现] 评分:[X.X]
【总体评价】:
[结合所有维度进行总体分析与评价]
总评分:[X.X]
【跟进建议】:[推进建议或修改建议]
约束条件
- 评估需基于提供的故事大纲内容,不自行创作或添加信息。
- 评分应客观公正,并附有详细分析。
- 建议需具体可行,有助于故事大纲改进。
示例
请参见 {baseDir}/references/examples.md 获取详细评估示例。该文件包含了多种故事类型(如都市爱情、科幻、历史等)的完整评估报告和分析说明。
详细文档
参见 {baseDir}/references/guide.md 获取故事大纲评估的完整指南,包括评估框架、评分标准、评估流程和注意事项。
版本历史
| 版本 | 日期 | 变更 |
|---|---|---|
| 2.1.0 | 2026-01-11 | 优化 description 字段,添加 allowed-tools (Read) 和 model (opus) 字段,调整主内容语言风格,添加约束条件,并引导至 references/examples.md |
| 2.0.0 | 2026-01-11 | 按官方规范重构 |
| 1.0.0 | 2026-01-10 | 初始版本 |
What ships with it
2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 148 lines · 48 tokens per session scan A c295c8b1b78a
story-outline-evaluator is a skill published in the GitHub repository GongLingRui/screen-creative-skills (402 stars, last pushed 3mo ago), licensed MIT. It adds 48 tokens to every session and 1,429 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. It is 88% identical to drama-evaluator, differing in 170 lines, and is treated as a copy.
Other skills, from other repositories
workers-best-practices
Cloudflare Workers best practices for production applications. Use when writing, reviewing, or configuring Workers.
create-custom-grader
Use when converting an existing benchmark, rubric, verifier, task YAML/JSON, or domain check into SkillEvaluator BYOG/BYOT custom evaluation.
find-journalists
Build, refine, dedupe, and enrich small fit-checked journalist lists for newsjack campaigns. Uses the newsjack CLI (preferred) or the medialyst MCP for news search and journalist enrichment, and falls back to a best-effort local mode with no verified contacts; the agent owns how returned data is organized.
story-origin-check
Recover the first public timestamp and canonical major coverage for a newsjacking signal, then decide whether newer coverage is the same story, a different story, or a materially new development.
annotating-task-lineage
Annotate Airflow tasks with data lineage using inlets and outlets. Use when the user wants to add lineage metadata to tasks, specify input/output datasets, or enable lineage tracking for operators without built-in OpenLineage extraction.
relevance-coarse-filter
Cheap, high-recall first-pass filter that removes obvious junk from a detector candidate pool before expensive story-origin research and PR judgment. Decides keep, monitoronly, or reject — never ranks, writes angles, verifies dates, or decides whether to pitch.