Case Studies - RTB Model Serving Platform

RTB Model Serving Platform

Factored built an ML serving platform for RTB that scaled throughput 20X, cut latency to 10ms, and improved reliability and cost.

Key Takeaways:

Handling the Volume of Demand.

The ML ad serving platform had limited throughput (50,000 requests per second) and struggled with high latency. Additionally, there was a need for multi-framework support, robust monitoring, and data quality assurance in both the serving and feature store pipeline.

Using ML for Operational Efficiency & Insights.

Scaling Throughput 20X - 1M Per Second.

Continue Reading