documentdb-full-text-search

documentdb-full-text-search is a skill for Claude Code, Codex from Azure/documentdb-agent-kit. It costs 135 tokens per session (743 once invoked), scanned A, original, MIT.

A guide to keyword search in Azure DocumentDB, where text is matched and ranked using search indexes and the `$search` database operation. It covers exact phrases, approximate spelling, prefixes, and custom text processing.

In plain words
What is it for?
Use it to create and query full-text indexes, search one field at a time, rank matches, support phrase or fuzzy searches, and match prefixes or identifier paths.
Why use it?
It prevents using search syntax or index commands from other MongoDB guides that do not apply to DocumentDB's full-text search path.

Skill for Claude CodeCodex

Part of the documentdb plugin — 17 skills, 1 agent shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/azure/documentdb-agent-kit/full-text-search
Any agent
npx skills add Azure/documentdb-agent-kit --skill full-text-search
Clone the repo
git clone --depth 1 https://github.com/Azure/documentdb-agent-kit

Made for: Claude Code, Codex.

Or install documentdb, the plugin that ships this one along with the rest of its 17 skills, 1 agent.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for documentdb-full-text-search

README.md
[![agentmods](https://agentmods.dev/badge/skills/azure/documentdb-agent-kit/full-text-search.svg)](https://agentmods.dev/skills/azure/documentdb-agent-kit/full-text-search)
Your own site
<a href="https://agentmods.dev/skills/azure/documentdb-agent-kit/full-text-search"><img src="https://agentmods.dev/badge/skills/azure/documentdb-agent-kit/full-text-search.svg" alt="Measured on agentmods" height="20"></a>
Per session 135 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 743 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00135 $0.00743
Opus 5 $0.00068 $0.00371
Sonnet 5 $0.00027 $0.00149
Haiku 4.5 $0.00014 $0.00074

Measured 3d ago against content hash 26bb606a2ba9, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

documentdb-full-text-search scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/full-text-search/SKILL.md · 29 lines

How it starts

The opening of the file, as written. The whole thing — 29 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Full-Text Search — Azure DocumentDB ($search + createSearchIndexes)

Azure DocumentDB's full-text search is driven by search indexes built with the createSearchIndexes database command and queried through the $search aggregation stage. Scoring is BM25, exposed via $meta: "searchScore". The community $text operator and { field: "text" } index type are not the DocumentDB search path.

Key syntax points that differ from many blog posts and older docs:

  • Index command is createSearchIndexes (not createIndexes) — each index has a name and a definition.mappings.fields block; dynamic: false is the safe default.
  • Custom analyzers live inside definition.analyzers and are referenced per-field via analyzer / searchAnalyzer.
  • $search targets an index by name via index: "<name>" — the engine does not auto-pick when multiple exist.
  • No count field inside $search — use a downstream { $limit: N } stage.
  • No compound operator yet — query one field at a time and merge in the application (see fts-multifield-index).

Rules

  • fts-create-search-index — Create a search index via runCommand({ createSearchIndexes }); use definition.mappings with dynamic: false.
  • fts-basic-search$search + text operator for BM25 keyword search; target index, project searchScore, cap with $limit.
  • fts-fuzzy-search — Add fuzzy: { maxEdits: 1 } to tolerate typos; keep maxEdits small.
  • fts-phrase-searchphrase operator with slop for ordered-proximity matching.
  • fts-custom-analyzers — Keyword tokenizer + lowerCase + asciiFolding + edgeGram for case-insensitive prefix matching on IDs, SKUs, part numbers. Index-time vs search-time analyzer pair.
  • fts-path-hierarchypathHierarchy tokenizer for hierarchical identifiers (BN-747-ENG-2024.05, dotted, slash paths).
  • fts-multifield-index — One search index mapping multiple fields; fan-out-and-merge in the app while $search compound is unavailable.
  • fts-hybrid-search — Combine BM25 and vector search (RRF) on the same collection.

Read the full file on GitHub · 29 lines

Files

What ships with it

8 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 3d ago First seen · 29 lines · 135 tokens per session scan A 26bb606a2ba9

Subscribe to this mod's changes

documentdb-full-text-search is a skill published in the GitHub repository Azure/documentdb-agent-kit (5 stars, last pushed 1mo ago), licensed MIT. It adds 135 tokens to every session and 743 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

schema-exploration

Lists tables, describes columns and data types, identifies foreign key relationships, and maps entity relationships in a database. Use when the user asks about database schema, table structure, column types, what tables exist, ERD, foreign keys, or how entities relate.

langchain-ai/deepagents · 57 tokens

agent-platform-rag-engine-management

Manage and query Agent Platform RAG Engine Corpora and retrieve grounded contexts using the Google GenAI SDK. Use when listing RAG corpora or files, inspecting a corpus, retrieving contexts, or generating content grounded in a RAG corpus. Do not use for standard database queries (use SQL/Spanner skills), Google…

google/skills · 85 tokens

moderator-page-migration

Port a moderator page from the main Next.js app (src/pages/moderator/) into apps/moderator. Use when asked to migrate, move or cut over a /moderator/ page to the spoke, or to port its tRPC procedures and Prisma services to SvelteKit loads/actions and Kysely.

civitai/civitai · 71 tokens

dsql

Build with Aurora DSQL — manage schemas, execute queries, handle migrations, diagnose query plans, diagnose cluster performance, load data, and develop applications with a serverless, distributed SQL database. Covers IAM auth, multi-tenant patterns, MySQL-to-DSQL and PostgreSQL-to-DSQL schema conversion, FK…

awslabs/agent-plugins · 227 tokens

sql-translate

Translate SQL queries between database dialects (Snowflake, BigQuery, PostgreSQL, MySQL, etc.).

AltimateAI/altimate-code · 26 tokens

cwicr-material-procurement

Generate material procurement lists from CWICR data. Calculate quantities with waste factors, group by supplier categories, and create purchase orders.

datadrivenconstruction/DDC_Skills_for_AI_Agents_in_Construction · 34 tokens