Core Infrastructure Engineer

OracleNashville, TN

About The Position

This role focuses on implementing and optimizing components within existing distributed systems under guidance. The engineer will apply basic scalability requirements, conduct performance/load testing, and configure resiliency features like retries, circuit breakers, and timeouts to handle network variability. Key responsibilities include building telemetry, alerts, and runbook-driven procedures, delivering scoped features and fault-injection tests, and assisting with basic data replication and synchronization. The position also involves participating in on-call rotations, using automation/IaC scripts for troubleshooting, and adhering to change, security, and compliance procedures while escalating complex issues to senior engineers.

Requirements

  • 5 years of experience in software development OR Bachelor's of Technology (B.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 1 year of experience in software development OR Bachelor's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 1 year of experience in software development
  • Demonstrated ability in or knowledge of troubleshooting, including diagnosing and resolving issues across various technical domains.
  • Demonstrated proficiency writing maintainable, effective code in high-level programming languages.
  • 1 year of academic or professional experience with cloud platforms (e.g., AWS, Azure, Google, Oracle Cloud).
  • 2 years of experience working in testing and automation at the system level.
  • 1 year of experience working with delivering and operating large-scale distributed systems.

Nice To Haves

  • 6 years of experience in software development OR Bachelor's of Technology (B.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 2 years of experience in software development OR Bachelor's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 2 years of experience in software development OR Master's of Technology (M.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field OR Master's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field.
  • 1 year of experience with cloud platforms (e.g., AWS, Azure, Google, Oracle Cloud).

Responsibilities

  • Assist in the implementation of components of distributed systems that support horizontal and vertical scaling under the guidance of senior engineers.
  • Optimize code segments and/or systems for large-scale data processing with oversight from senior engineers.
  • Implement scalability requirements for assigned components.
  • Learn about the use of data plane platforms for large-scale data retrieval, storage, and processing.
  • Execute performance and load testing, with guidance.
  • Collaborate with the team to build fault-tolerant components capable of withstanding in-service updates by learning about redundancy, replication, and automatic failover mechanisms.
  • Learn about recovery oriented computing principles and assist in applying them to component designs.
  • Configure and test retry mechanisms, circuit breakers, and timeouts to help handle network unreliability, with guidance.
  • Implement testing and alarming configurations to detect issues/failures.
  • Support efforts to recover from failures by drafting and executing runbooks and operational procedures, under guidance.
  • Help build dashboards, telemetry systems, and alerting mechanisms to monitor component health.
  • Implement functional requirements and testing for assigned features within an existing system.
  • Implement test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness, under guidance.
  • Help implement basic data replication and synchronization techniques to maintain data integrity and availability.
  • Assist in diagnosing and debugging issues in system components to support ongoing operation, under supervision.
  • Follow protocols to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues.
  • Run basic automation scripts and tooling to troubleshoot operational issues.
  • Participate in operational support rotations, assisting in incident responses and root cause investigations.
  • Assist in implementing basic security measures to protect data and applications in multi-tenant environments, including encryption and access controls.
  • Assist in the execution of remediation plans to address identified security gaps, under supervision.
  • Support the creation and updating of documentation to ensure cloud infrastructure is in compliance with relevant industry standards and regulations.
  • Assist in maintaining basic automation scripts and tools (e.g., Infrastructure as Code (IaC)).
  • Adhere to change management plans for patching, updating, and rolling back applications, under guidance.
  • Track timelines with minimal supervision, ensuring work is completed in a timely manner and is in alignment with project requirements.
  • Prioritize and adjust work as resources or timelines change, with some guidance.
  • Collaborate within the team to better understand expectations and achieve shared objectives.
  • Leverage a foundational understanding of business, stakeholder, and/or customer needs to build partnerships with limited guidance.
  • Actively listen and ask questions to enhance collaboration.
  • Build a basic understanding of business, stakeholder, and/or customer needs with guidance.
  • Identify and address issues, escalating problems to senior staff as needed in accordance with standard procedures.
  • Compile and review data and/or information from multiple sources to troubleshoot standard and non-standard errors.
  • Seek opportunities to gain knowledge and learn new skills and/or tools aligned with industry trends and best practices.
  • Utilize feedback and training to improve skills.
  • Participate in a culture of continuous learning and knowledge sharing.
  • Implement updates to processes, protocols, and workflows to increase efficiency and effectiveness as directed, with some guidance.
  • Contribute to ideation for future process improvements.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service