Reliability & Production Engineer

mthreeNew York, NY
$85,000 - $110,000Onsite

About The Position

We are looking for an experienced Reliability & Production Engineer (RPE) to join a dynamic technology team for one of our clients, experienced in supporting mission-critical equities trading platforms and alternative trading systems. This role offers direct exposure to electronic trading, market structure, production support, and platform reliability in a fast-paced trading environment. We would like to talk to you if you: Are interested in supporting distributed systems and high-availability trading platforms. Like to work in a fast-moving environment and are not afraid to improve processes and systems. Enjoy new technological challenges and solving complex production issues. Are keen to learn about electronic trading, market structure, and production engineering while helping users and stakeholders solve critical business challenges. Believe a team working well together is smarter than the single smartest person on that team. Aspire to grow as a person and as a teammate. Have grit, drive, and a deep sense of ownership. Thrive in a highly collaborative environment partnering with traders, developers, quantitative teams, and business stakeholders. As a Reliability & Production Engineer, you'll help support and maintain mission-critical trading applications and alternative trading systems, ensuring platform stability, operational excellence, incident resolution, and continuous improvement across the production environment. About mthree Since 2010, mthree has been helping clients solve their business and technological challenges. We are a technology and business consultancy with a global workforce delivering significant business and IT projects in some of the largest financial services organizations worldwide. Core Services Consulting and Advisory Managed Services Alumni Graduate Program Alumni Pro Program We have a global presence and are experts in delivering exceptional quality to our client base, providing consulting services across Risk, Regulation & Compliance, Vendor Products, Application Support, Application Development, Cyber & Information Security, Data Science, and DevOps. Our Expert program offers experienced professionals access to top roles in technology, finance, aviation, and insurance. Join us to work on innovative technology projects and critical platforms while gaining valuable experience within prestigious global organizations.

Requirements

  • 2+ years of experience supporting production applications, distributed systems, trading platforms, or technology operations environments.
  • Bachelor's degree in Computer Science, Engineering, Finance, Mathematics, or a related discipline.
  • Strong Linux/Unix administration and troubleshooting skills.
  • Experience supporting distributed real-time applications in a production environment.
  • Proficiency in Python or another scripting language.
  • SQL and database query experience.
  • Familiarity with market data systems and the FIX Protocol.
  • Strong analytical, troubleshooting, and problem-solving abilities.
  • Excellent verbal and written communication skills.
  • Ability to perform effectively in a fast-paced, high-pressure production support environment.
  • Applicants must be currently authorized to work in the United States on a full-time basis. The Company will not sponsor applicants for work visas.

Nice To Haves

  • Experience supporting electronic trading systems, exchange connectivity, or brokerage platforms.
  • Understanding of U.S. Equities market structure and the order lifecycle.
  • Familiarity with alternative trading systems (ATS), dark pools, crossing engines, smart order routing, or algorithmic trading platforms.
  • Experience with monitoring tools such as Splunk, Grafana, Prometheus, or similar technologies.
  • Knowledge of Agile software development practices.

Responsibilities

  • Monitor real-time order flow, execution activity, client connectivity, and overall platform health.
  • Support highly available trading and execution platforms in a production environment.
  • Investigate and resolve production issues including order rejects, missing executions, connectivity issues, market data discrepancies, and trade reporting issues.
  • Serve as a primary escalation point during production incidents impacting business operations and users.
  • Proactively monitor applications, market data feeds, infrastructure, and system performance.
  • Lead incident management efforts from detection through resolution and post-incident review.
  • Perform root cause analysis and implement preventative measures to improve platform resiliency.
  • Create and maintain operational runbooks, procedures, dashboards, and support documentation.
  • Participate in disaster recovery, failover, capacity planning, and system testing exercises.
  • Build automation and tooling to improve observability, diagnostics, and operational efficiency.
  • Develop solutions using Python, SQL, Linux/Unix tools, and monitoring technologies.
  • Leverage analytics to identify trends in platform behavior and system performance.
  • Work with development teams to enhance platform stability, application design, and supportability.
  • Collaborate with trading, development, quantitative, compliance, and infrastructure teams to resolve issues and improve platform performance.
  • Communicate effectively during production incidents and stakeholder escalations.
  • Participate in reviews and operational discussions to ensure best practices are adopted across teams.
  • Assist with trading inquiries, investigations, regulatory reviews, and audit requests.
  • Support reporting requirements and controls related to electronic trading operations.

Benefits

  • competitive compensation
  • comprehensive benefits package
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service