Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add Desko77/claude-code-skills-1c --skill transcribegit clone --depth 1 https://github.com/Desko77/claude-code-skills-1cWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/desko77/claude-code-skills-1c/transcribe)<a href="https://agentmods.dev/skills/desko77/claude-code-skills-1c/transcribe"><img src="https://agentmods.dev/badge/skills/desko77/claude-code-skills-1c/transcribe/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/desko77/claude-code-skills-1c/transcribe"><img src="https://agentmods.dev/badge/skills/desko77/claude-code-skills-1c/transcribe.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00196 | $0.07290 |
| Opus 5 | $0.00098 | $0.03645 |
| Sonnet 5 | $0.00039 | $0.01458 |
| Haiku 4.5 | $0.00020 | $0.00729 |
Grade A, and why
transcribe scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
Copies of this mod
1 near-identical copy found in the catalogue:
- transcribe — 95% identical, 15 lines differ
How it starts
The opening of the file, as written. The whole thing — 269 lines — stays where its author put it; the contents beside it link to each section on GitHub.
/transcribe - Транскрибация видео и аудио
Два движка:
- Локальный (default для аудио):
faster-whisper(CUDA) + опц. диаризация. Движок диаризации выбирается сам: без--num-speakers-pyannote community-1(GPU, корректный автодетект числа спикеров, RTF ~0.064); с явным--num-speakers N-sherpa-onnx GPU(pyannote-segmentation-3.0 + eres2net, RTF ~0.24, точное N). Опция--diarize-engine moss- MOSS-Transcribe-Diarize end-to-end: ASR+диаризация одной моделью (без whisper-шага), лучше текст на технических терминах, но ~2x медленнее (RTF ~0.34), требуетvenv-moss(envMOSS_PYTHON). Нет затрат, не уходит наружу. ВИДЕО тоже можно разобрать полностью локально ---engine local(разбор экрана локальной VLM + спикеры по голосу, см. ниже). - Gemini (default для видео и
--analyze-ui): облачный API, ~$0.10/час. Нужен интернет и квота. Стартовая модельgemini-2.5-flash(пин конкретной версии, дешевая); при перегрузке (503/429) переходит наgemini-2.5-flash-lite. Дорогие 3.5/pro сознательно исключены.
Выбор движка по умолчанию
| Тип файла | Движок | Причина |
|---|---|---|
| Аудио (m4a, mp3, wav, ogg, flac, aac, wma) | local | Быстро, бесплатно, диаризация |
| Видео (mp4, mkv, webm, avi, mov) | gemini | Быстро, облако. Приватный вариант - --engine local (см. ниже) |
Видео + --engine local |
local | Разбор экрана БЕЗ облака: whisper + локальная VLM (LM Studio) + спикеры по голосу |
Любой + --analyze-ui |
gemini | Детальный разбор интерфейсов в облаке |
Любой + --engine gemini |
gemini | Явный override на облако |
Аудио + --engine local |
local | Явный override (аудио) |
При 503/429 Gemini-движок сначала сам перебирает пул моделей (см. раздел "Авто-fallback по моделям Gemini"). Если весь пул недоступен и это аудио - можно вручную переключиться на local (--engine local).
Режимы
Локальный (аудио + faster-whisper + опц. pyannote)
Выходные файлы:
<имя> - транскрипция.md- таймкоды + текст<имя> - транскрипция.txt- plain text<имя> - со спикерами.md- реплики с метками[SPEAKER_XX, MM:SS](только при--diarize)
What ships with it
16 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- .gitignore 670 B
- glossary.txt 2.4 KB
- README.md 14 KB
- scripts/analyze_video_local.py 45 KB runs code
- scripts/diarize_moss.py 8.4 KB runs code
- scripts/diarize_sherpa.py 13 KB runs code
- scripts/glossary.py 12 KB runs code
- scripts/local_backends.py 38 KB runs code
- scripts/setup.py 16 KB runs code
- scripts/speaker_validator.py 38 KB runs code
- scripts/text_stage.py 22 KB runs code
- scripts/transcribe_local.py 37 KB runs code
- scripts/transcribe.py 47 KB runs code
- scripts/verify.py 9.2 KB runs code
- scripts/voiceprints_dedup.py 17 KB runs code
- scripts/voiceprints.py 8.1 KB runs code
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 5d ago First seen · 269 lines · 196 tokens per session scan A 7cbbc5c90a48
transcribe is a skill published in the GitHub repository Desko77/claude-code-skills-1c (56 stars, last pushed 6d ago), licensed MIT. It adds 196 tokens to every session and 7,290 once invoked, about $0.0010 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.
Other skills, from other repositories
1c-form-info
A compact reader for managed forms in 1C, a platform for building business applications. It summarizes the form's elements, fields, commands, and events instead of showing the full XML file.
claude-env-setup
A setup and update guide for installing an agent's skills, rules, commands, and related tools on a computer. It first records what is already installed and then plans additions or updates.
docx-from-sample
A document-building tool that creates a new DOCX file using an existing DOCX as its formatting template. DOCX is the file format used by Microsoft Word.
1c-bsp-registration
A helper for 1C external reports and processors that adds the function needed to register them in the Standard Subsystems Library (БСП). БСП is a set of reusable features for 1C configurations.
1c-cfe-patch-method
A tool for creating method interceptors in a 1C configuration extension. 1C is business software, and an interceptor can run code before, after, or instead of an inherited method.
claude-md-bootstrap
A project setup tool that creates or updates a CLAUDE.md file, which gives an AI coding agent project-specific instructions and background at the start of each session.