Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
FluidAudio is a Swift SDK that runs local audio AI models on Apple devices for speech recognition, text-to-speech, voice activity detection, and identifying who is speaking. It is intended for developers building iOS and macOS apps with low-latency audio processing. The catalogue add-ons provide workflows for using the SDK.
Latest release v0.15.6 · 19 Aug 2026
These files are FluidInference/FluidAudio's own configuration. They tell Claude Code, GitHub Copilot, Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.
.github/copilot-instructions.md A 552 tok AGENTS.md A 551 tok CLAUDE.md A 2,543 tok .claude/settings.json A — .claude/settings.json A — .claude/agents/apple-neural-performance-expert.md A — .claude/agents/code-search-ast-grep.md A —