This is an open-source implementation of the ITU P.808 standard for "Subjective evaluation of speech quality with a crowdsourcing approach" (see https://www.itu.int/rec/T-REC-P.808/en). It uses Amazon Mechanical Turk as the crowdsourcing platform. It includes implementations for Absolute Category Rating (ACR), Degradation Category Rating (DCR), and Comparison Category Rating (CCR).
P.808 Toolkit is software for running crowdsourced subjective speech-quality assessments using methods defined by ITU-T recommendations. Researchers and speech-system developers use it with platforms such as Amazon Mechanical Turk and Prolific, or with their own remote worker panels, to evaluate speech and noise-suppression systems.
These files are microsoft/P.808's own configuration. They tell Claude Code, GitHub Copilot, Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.
.github/copilot-instructions.md A 781 tok .github/instructions/language.instructions.md A 480 tok AGENTS.md A 282 tok CLAUDE.md A 169 tok .github/agents/analyze-results.agent.md A 52 tok .github/agents/create-study.agent.md A 79 tok