Software Engineer - Forensics

EverpureSanta Clara, CA
$149,000 - $224,000Onsite

About The Position

Drive the technical resolution and stability of Everpure's highest-stakes cloud storage deployments for top-tier hyperscale partners. As a Senior Software Engineer on the Hyperscale Product Escalations team (known internally as a Forensics Engineer), you will serve as the technical bridge between field telemetry and core engineering, diagnosing high-complexity distributed systems failures and creating automation that anticipates issues before they impact customer workloads. You will directly influence customer trust, enable multi-million dollar account success, and help architect the future of hyperscale storage solutions built for AI-driven infrastructure.

Requirements

  • Systems & Storage Expertise: Deep hands-on experience in Linux systems engineering, platforms, or low-level firmware, along with a strong grasp of storage technologies (such as SSDs, NVMe, NAND, or distributed storage architectures).
  • Advanced Debugging & Software Engineering: Proven mastery of troubleshooting complex distributed systems using languages like Python, Go, or C++, paired with the ability to build automated tools for log analysis, telemetry, and system diagnostics.
  • Technical Communication & Problem-Solving: Ability to deconstruct intricate distributed systems failure modes into clear technical action plans, facilitating seamless collaboration with cross-functional SMEs and external engineering leadership.

Responsibilities

  • Lead Root-Cause Investigations: Analyze complex system behavior, memory dumps, field telemetry, and low-level code across platform, firmware, and software layers to isolate and resolve critical failures in large-scale distributed systems.
  • Build System Health Automation: Design and deploy automated diagnostics, predictive health-monitoring services, and AI-assisted workflow tools that reduce manual triage time and prevent recurring failure patterns across growing global fleets.
  • Direct Hyperscaler Engineering Collaboration: Partner directly with technical teams at hyperscale customer organizations to solve deep integration challenges, ensuring platform reliability and unlocking major expansion opportunities.
  • Engineer Fleet-Wide Reliability Improvements: File, prioritize, and drive long-term engineering fixes from escalation findings back into core product roadmaps to continuously elevate system stability.

Benefits

  • flexible time off
  • wellness resources
  • company-sponsored team events
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service