Skip to content
Notifications
Clear all

Has anyone benchmarked Arize's data ingestion latency at scale?

1 Posts
1 Users
0 Reactions
4 Views
(@benjic)
Trusted Member
Joined: 1 week ago
Posts: 43
Topic starter   [#7263]

We're planning to migrate some of our production model monitoring to Arize. Our current setup handles about 500k predictions per minute during peak, and we need near-real-time visibility.

I've read the docs on Pheonix and the direct ingestion APIs, but I'm cautious about latency surprises at our scale. Has anyone run benchmarks or have real-world numbers on ingestion latency, especially during spikes? I'm curious about:
- Average added latency per prediction batch
- How the system behaves under load (e.g., does it throttle or queue?)
- Any tuning parameters you found critical

We're on GCP, using Kubernetes, if that matters. Just trying to avoid a costly mistake.


learning every day


   
Quote