The CloudOps Engineer - Software role sits within the eShepherd R&D team, reporting to the Regional Lead for North America, with day-to-day collaboration with the Software Team Lead for Cloud and Apps and engineers in Melbourne and Hamilton. eShepherd is a virtual fencing solution for cattle, utilizing solar-powered neckbands, a phone app, and a platform that enables farmers to manage livestock remotely. The company has grown from a startup to a scale-up within Gallagher, a long-standing farm fencing company. This role is crucial for evolving the platform's architecture to handle significant growth in device count, regional expansion, and new product development. The position involves investigating performance under load, identifying potential breaking points, and recommending solutions to the R&D team. Key areas of focus include query performance, data modeling at scale, ingest paths, caching, connection handling, storage strategy, regional topology, and failover. The role also encompasses improving monitoring, incident response, and capacity planning to ensure smooth growth. As part of the DevOps function, the engineer will extend it into the North American region, managing pipelines, infrastructure as code, environments, automation, and deployment tooling in collaboration with the global team. This includes setting up regional environments, deploying releases, and ensuring automation supports multi-region operations. The role contributes to the platform's next major extension, involving design work for new products and scaling. The CloudOps Engineer will serve as the R&D presence in North America, acting as the on-the-ground engineer during Melbourne's off-hours for incident response and providing timely answers to North American customer success and operations teams. The job emphasizes a cycle of design, test, deploy, and iterate, prioritizing measurement and weekly improvements over lengthy development cycles. The ideal candidate is results-oriented, data-driven, comfortable working independently across time zones, and possesses strong written communication skills. An understanding of the physical constraints of devices on farms (e.g., battery life, cellular coverage) is important for designing robust systems. The role requires commercial experience running production cloud infrastructure on AWS at scale, deep database competence with MySQL, a track record of improving system performance and reliability, experience with monitoring and incident response, and proficiency in infrastructure as code and CI/CD practices. Familiarity with IoT architectures and protocols is also beneficial.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed