We are seeking a world-class Data Platform Engineer to take complete operational and architectural ownership of one of the largest Apache Druid deployments in the industry. Our environment operates at extreme velocity and petabyte scale, processing real-time streaming data via Kafka across ~2,000 parallel ingestion tasks. If you have pushed Druid, Zookeeper, and the Linux kernel to their absolute breaking points—and know exactly how to tune them to stay up—this role is built for you. In this role, you will bridge the gap between deep infrastructure engineering, Unix Linux systems administration, and distributed big data platforms. You will design, scale, and automate massive Middle Manager/Indexer fleets, optimize deep storage, eliminate metadata DB bottlenecks, and implement bulletproof query isolation. Utilizing Ansible and Infrastructure-as-Code methodologies, you will ensure our clusters are highly available, self-healing, and meticulously monitored. If you are a platform engineer who thrives on solving complex concurrency, JVM Garbage Collection, and high-throughput streaming challenges at scale, join us to redefine what is possible with real-time analytics.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed