FluxyChat

Operations

Voice load report

SLO targets from the SDK unit bench. Not production fleet P95.

The numbers below are product SLO targets, not a load test against a distributed fleet. Regenerate them with pnpm --filter @fluxy-chat/sdk exec vitest run src/voice-load-bench.test.ts (unit bench in this repo).

Do not quote them as measured P95 in production. Live operator stats (when you collect them) are POST /admin/voice-ai/metrics then GET /admin/voice-ai/stats.

Product SLO

PathTarget
OpenAI RealtimeP95 ≤ 300ms e2e
Chunked RESTP95 ≤ 500ms
Barge-in≤ 500ms

Production source of truth: POST /admin/voice-ai/metrics → GET /admin/voice-ai/stats.

See Voice AI pipeline and Platform voice.

On this page