Expedia Upgrades Ranking Models to Keras 3: 30% Faster Training, 70% Lower Latency
fchollet · x · 2026-08-12
Expedia recently migrated its ranking models to a state-of-the-art Keras 3 setup. The upgrade resulted in a 30% increase in training speed and a massive 70% reduction in inference latency.
Keras creator François Chollet highlighted the writeup as a great source of technical insights for serving large-scale ranking and recommendation models at ultra-low latency.
Related event: Expedia Migrates to Keras 3, Slashing Inference Latency by 70%(2 posts)→
More from Infra
- Compute as Collateral? Silicon Data Raises $30.5M Series A — dinabass · 2026-08-12
- Deep Optimization of Automatic1111 for Apple Silicon: 40% Render Time Reduction — Time-Conversation528 · 2026-08-12
- Designing a 42U On-Prem AI Pod: 32 GPUs with Plug-and-Play Infrastructure — dee_hw · 2026-08-12
- YMTC-Backed Fund Invests in China's Alternative Chipmaking Route — pstAsiatech · 2026-08-12
- Should AI Data Centers Disguise Themselves as Victorian Buildings? — david_stillwell · 2026-08-12
- Best Quantization for Sub-2-bit? Beyond QTIP, What Papers to Read? — Aggravating-Push-207 · 2026-08-12