13,736 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
An image-to-text bridge that lets the text-only DeepSeek model work with pictures. It uses codex exec to send images to gpt-5.4-mini, which returns structured descriptions in Chinese for the conversation.
Converts vague, messy, or under-specified requests into clearer, efficient AI-ready instructions while preserving user intent. Use before solving complex, ambiguous, or high-risk tasks.
★not rated 1 2mo agoA
tokens not measured
originalMIT
This skill should be used when users want to train or fine-tune language models using TRL (Transformer Reinforcement Learning) on Hugging Face Jobs infrastructure. Covers SFT, DPO, GRPO and reward modeling training methods, plus GGUF conversion for local deployment. Includes guidance on the TRL Jobs package, UV…
Monitor ML models in production for data drift, concept drift, and performance degradation. Outputs drift detection pipelines, alerting thresholds, retraining triggers, and monitoring dashboards.
Expert guide for building applications with Qdrant Edge — the embedded, offline-capable vector search engine for edge devices (robots, kiosks, mobile phones, IoT, home assistants). Use this skill whenever the user mentions Qdrant Edge, qdrant-edge-py, EdgeShard, on-device vector search, offline vector search, embedded…
Ballast: pull a quantized knowledge corpus (or build your own) and ground any local model — OpenAI proxy, MCP server, model profiling, and a three-arm grounding benchmark. Runs locally from the openballast Python package.
Natively ingest an ArchiveBox archive into the epistemic-graph knowledge graph via the archivebox-api MCP server — push snapshots as typed :Snapshot nodes (with :Tag + :hasTag links and per-snapshot :Document page-text) and archive results as :ArchiveResult nodes, best-effort, with the Wire-First ingest tools. Use…
Enhanced MCP server for Google Gemini 3 with Image Generation, Batch API integration (50% cost, async processing), advanced file handling, and conversation management. Features Gemini 3 Pro (default) and Gemini 3 Pro Image models with state-of-the-art rea. Runs locally from the @mintmcqueen/gemini-mcp npm package.
★not rated 1 4mo agoA
tokens not measured
originalMIT
MCP server for image recognition using vision APIs (Anthropic, OpenAI, Cloudflare Workers AI). Runs locally from the mcp-image-recognition Python package.
★not rated 1 9mo agoA
tokens not measured
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: