Azure/gpt-rag-ingestion

The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich, multimodal retrieval.

About the project

GPT-RAG Data Ingestion is a service that processes documents such as PDFs, images, spreadsheets, transcripts, and SharePoint files so they can be searched through Azure AI Search. It prepares data with format-specific chunking and text or image embeddings for multimodal retrieval in agent-based applications.

189Stars on the repository
19Mods indexed here, across every type
yesterdayLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Copilot instructions for Azure/gpt-rag-ingestion, covering repository development and release instructions, branching, versioning, changelog lifecycle and release safety.

not rated 189 changed today A 968 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 208 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 250 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 225 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 249 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 203 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file GitHub Copilot ✓ vendor

Instructions for Azure/gpt-rag-ingestion, a project described as: The GPT-RAG Data Ingestion service automates processing of diverse documents—PDFs, images, spreadsheets, transcripts, and SharePoint—readying them for Azure AI Search. It applies smart chunking, generates text and image embeddings, and enables rich…

not rated 189 yesterday A 219 tokens original MIT

Azure/gpt-rag-ingestion

Instructions file CodexOpenCode ✓ vendor

AGENTS.md instructions for Azure/gpt-rag-ingestion, covering gpt-rag ingestion engineering-agent contract, priority, what this repository is, repository boundaries and data, security, and configuration.

not rated 189 yesterday A 1,304 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: