agent-photo-question

agent-photo-question is a skill for Claude Code, Codex from hezkvectory/hermes-edu-skills. It costs 40 tokens per session (1,298 once invoked), scanned A, a copy of adult-vocational-certificate, MIT.

A photo-based homework tutor that reads a question, explains the reasoning step by step, and checks whether the student understands. It can handle school questions from different subjects when the image or text is clear.

In plain words
What is it for?
Use it to explain photographed homework, break down long questions, identify mistakes, and create follow-up exercises matched to the student's level.
Why use it?
It helps students who are stuck understand how to solve a problem instead of copying an answer. It also flags unclear parts and provides similar practice questions.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one.

Good fit Use it to explain photographed homework, break down long questions, identify mistakes, and create follow-up exercises matched to the student's level.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/hezkvectory/hermes-edu-skills/agent-photo-question
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add hezkvectory/hermes-edu-skills --skill agent-photo-question
Clone the repo
git clone --depth 1 https://github.com/hezkvectory/hermes-edu-skills

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for agent-photo-question

README.md
[![agentmods](https://agentmods.dev/badge/skills/hezkvectory/hermes-edu-skills/agent-photo-question/github.svg)](https://agentmods.dev/skills/hezkvectory/hermes-edu-skills/agent-photo-question)
Your own site
<a href="https://agentmods.dev/skills/hezkvectory/hermes-edu-skills/agent-photo-question"><img src="https://agentmods.dev/badge/skills/hezkvectory/hermes-edu-skills/agent-photo-question/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for agent-photo-question

Your own site · 80×15
<a href="https://agentmods.dev/skills/hezkvectory/hermes-edu-skills/agent-photo-question"><img src="https://agentmods.dev/badge/skills/hezkvectory/hermes-edu-skills/agent-photo-question.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 40 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,298 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin 80% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00040 $0.01298
Opus 5 $0.00020 $0.00649
Sonnet 5 $0.00008 $0.00260
Haiku 4.5 $0.00004 $0.00130

Measured 9d ago against content hash 612e40dd781d, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

agent-photo-question scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

80% identical to adult-vocational-certificate — 125 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

skills/learning-core/agent-photo-question/SKILL.md · 143 lines

How it starts

The opening of the file, as written. The whole thing — 143 lines — stays where its author put it; the contents beside it link to each section on GitHub.

拍照答疑 Skill

把拍照题目转成先识题、再讲思路、最后追问确认的学习过程,而不是只给答案。

这个 Skill 解决什么问题 / Problem

把拍照题目转成先识题、再讲思路、最后追问确认的学习过程,而不是只给答案。

最适合 / Best For

  • 学生或家长拍题求讲解
  • 题干较长、需要拆条件的题
  • 需要生成同类题巩固的场景

不适合 / Not For

  • 考试作弊或实时替考
  • 图片严重模糊且用户无法补充题干

使用前请准备 / Inputs

  • 题目图片或完整题干
  • 年级/学段
  • 学科
  • 学生卡住的位置
  • 识别题干和条件
  • 确认不清晰内容
  • 按条件-目标-方法-步骤讲解
  • 给同类变式题检查理解

输出格式 / Output Format

  • 题目识别结果
  • 解题思路
  • 分步讲解
  • 易错提醒
  • 同类练习

质量检查 / Quality Checks

  • 不得只输出答案
  • 识别不确定时必须标注
  • 讲解要匹配年级
  • 避免诱导未成年人消费或泄露隐私

没有平台工具时 / Standalone Fallback

  • 没有视觉工具时,请用户手动输入题干。
  • 没有练习工具时,由 Agent 生成同类变式题。

示例提示 / Example Prompts

  • 我把数学题文字贴给你,请按五年级水平讲解。
  • 这道物理题我卡在受力分析,帮我拆条件。

适用场景 / When To Use

当学习者、家长、老师、学校或教育应用开发者需要处理以下场景时,可以使用这个 Skill。

最适合的场景:

  • 拍照答疑
  • 课后作业

适用角色:

  • 学习者
  • 家长

调用信号 / Invocation Signals

意图:

  • agent_photo_question
  • learning_core
  • 拍照答疑
  • 课后作业

示例表达:

  • 开始拍照答疑 Skill
  • 帮我做拍照答疑
  • 根据当前上下文执行拍照答疑 Skill

公开 Skill 契约 / Public Skill Contract

  • Workflow: agent_photo_question.run
  • Category: learning-core
  • Stages: primary, junior, senior
  • Subjects: 综合
  • Abilities: AI 讲题, 图片识题
  • Quality Tier: curated
  • Standalone Support: requires_tools
  • Public Release: recommended
  • Requires Tools: file.read_upload, vision.ocr_question
  • Requires Data: 题目图片或用户转写的题干, 学生年级, 学科
  • Export Mode: installable
  • Release Channel: recommended

成熟度备注:

  • 已按精品 Skill 标准补充边界、输入、工作流、输出格式和示例。

参数化使用 / Parameters

这个 Skill 不再把年级、册别、单元、知识点和难度拆成大量独立 Skill。请在调用时通过参数或自然语言补充这些信息。

  • Grades: 一年级, 二年级, 三年级, 四年级, 五年级, 六年级, 七年级, 八年级, 九年级, 高一, 高二, 高三
  • Semesters: 上册, 下册, 上册, 下册, 必修一, 必修二, 选择性必修
  • Scenarios: 拍照答疑, 课后作业
  • Difficulties: 基础, 标准, 提高
  • Parameterized Dimensions: grade, semester, unit, lesson, knowledgePointCodes, scenario, difficulty

独立 Hermes 使用方式 / Standalone Hermes Usage

这个 Skill 可以通过 Hermes 的 skills.external_dirs 作为外部 Skill 加载。

如果你有自己的工具、记忆、课程数据或 workflow runner,可以把它们与本 Skill 组合使用。如果没有外部工具,也可以直接使用上面的说明来引导对话,生成有用的学习或教学反馈。

Read the full file on GitHub · 143 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 9d ago First seen · 143 lines · 40 tokens per session scan A 612e40dd781d

Subscribe to this mod's changes

agent-photo-question is a skill published in the GitHub repository hezkvectory/hermes-edu-skills (94 stars, last pushed 3mo ago), licensed MIT. It adds 40 tokens to every session and 1,298 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. It is 80% identical to adult-vocational-certificate, differing in 125 lines, and is treated as a copy.

Related

Other skills, from other repositories

hr-onboarding

A new-hire onboarding plan as a single page — first week schedule, buddy + manager intro, learning track, equipment checklist, and "you're set when…" outcomes. Use when the brief mentions "onboarding", "new hire", "first week plan", or "入职".

nexu-io/open-design · 62 tokens

book-mirror

Take any book (EPUB/PDF), produce a personalized chapter-by-chapter analysis. Each chapter is preserved in detail (The Chapter) and mirrored back to the reader's actual life (The Mirror) using brain context. The mirror observes and resonates — a friend pointing out parallels, NOT a consultant rearranging the reader's…

garrytan/gbrain · 138 tokens

miniapp

Build a tiny interactive HTML playground only when someone asks to see, play with, or step through a mechanism.

yc-software/qm · 25 tokens

eli5

Explain research, papers, or technical ideas in plain English with minimal jargon, concrete analogies, and clear takeaways. Use when the user says "ELI5 this", asks for a simple explanation of a paper or research result, wants jargon removed, or asks what something technically dense actually means.

companion-inc/feynman · 63 tokens

deck-course-module

A course or workshop slide template with persistent learning goals, teaching pages, multiple-choice self-tests, and a wrap-up.

nexu-io/html-anything · 25 tokens

master-yinguang

A reference-based assistant for questions about Yinguang and Pure Land Buddhism, a Buddhist tradition focused on faith, ethical living, and practice connected with rebirth in the Pure Land. It can answer in Yinguang’s historical teaching style.

xr843/Master-skill · 274 tokens