Niels Claeys shares how his team built a data platform processing up to 1.5 million core hours monthly. He explains the specific optimizations they discovered through production experience, from scheduler changes to achieving 97% spot instance usage without reliability issues.
You will learn:
- How to achieve 97% spot instance adoption through strategic instance type diversification, region selection, and Spark-specific techniques
- Node pool design principles that balance Kubernetes overhead with workload efficiency
- Platform-specific gotchas like AWS cross-AZ data transfer costs that can spike bills unexpectedly
Watch (or listen to) it here: https://ku.bz/hGRfkzDJW
🌟 This episode is brought to you by Testkube—the ultimate Continuous Testing Platform for Cloud Native applications. Scale fast, test continuously, and ship confidently https://ku.bz/lnxYK3s0L
With @Birthmarkb "Almost 40" Farrell
Post #1532
236
Forwarded from KubeFM