FluidInference/FluidAudio

Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.

About the project

FluidAudio is a Swift SDK that runs local audio AI models on Apple devices for speech recognition, text-to-speech, voice activity detection, and identifying who is speaking. It is intended for developers building iOS and macOS apps with low-latency audio processing. The catalogue add-ons provide workflows for using the SDK.

Latest release v0.15.6 · 19 Aug 2026

These files are FluidInference/FluidAudio's own configuration. They tell Claude Code, GitHub Copilot, Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.

2,735Stars on the repository
7Files it configures its agents with
3,646Tokens loaded in every session
4Agents configured

Instructions

Settings

Hooks

Agents