Pranjay-kumar

8 mods across 1 repository, 2 stars between them.

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Trigger when the user wants to collect, structure, evaluate, crawl, extract, refresh, or build reusable data acquisition pipelines from websites, APIs, portals, files, or rendered apps. Use for dataset design, source classification, feasibility, endpoint discovery, authorized/owned-session scraping plans, Patchright…

2 2mo ago A 126 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use for Patchright/Playwright-based public or authorized browser probing: warm-session cookie/storage generation, browser network capture, JSON/API route discovery from page loads, rendered DOM fallback, screenshots, tiny DOM samples, and user-owned storage-state workflows. Do not use for CAPTCHA solving, credential…

2 2mo ago A 72 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Shared core for the data acquisition skill tree. Use when a data acquisition task needs source access classification, output contracts, compliance boundaries, feasibility scorecards, probing standards, pipeline quality standards, or shared references used by sibling data-acquisition skills. Do not use alone for…

2 2mo ago A 65 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use when the user needs to decide what data to collect before scraping or API work: DatasetNeed, DatasetSpec, entity grain, required vs nice-to-have fields, freshness, history, coverage targets, join keys, exclusions, and uselessness criteria. Use for vague business goals, all data requests, and scope control before…

2 2mo ago A 72 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use for discovering and reverse-engineering data sources: official APIs, XHR/fetch, GraphQL, persisted queries, Algolia, Shopify, Salesforce Commerce Cloud, sitemaps, feeds, embedded JSON, hydration state, page-data routes, pagination limits, headers, params, and endpoint templates.

2 2mo ago A 66 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use when the user wants to know whether a dataset/source is worth pursuing, compare routes, score feasibility, identify trapdoors, classify Green/Yellow/Red, or decide whether to stop, sample, narrow, license, use owned-session access, or build a pipeline.

2 2mo ago A 61 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use when the user wants a production-grade scraping/API/browser pipeline design or implementation plan: pipeline.yaml, schemas, raw/staged/normalized outputs, dedupe, incremental refresh, checkpoints, retries, rate-limit strategy, quality gates, observability, run reports, and recovery.

2 2mo ago A 61 tokens original MIT

Pranjay-kumar/universal-data-acquisition-pipeline-skill

Skill Claude CodeCodex

Use when packaging real data acquisition results for publication: probe-backed case studies, README summaries, evidence tables, sample rows, feasibility reports, and publishability checks. Do not publish hypothetical case studies, owned-session outputs, cookies, credentials, private data, or non-public authorized…

2 2mo ago A 62 tokens original MIT