Storage Architect

SURATech, LLC (Jefferson Lab)Newport News, VA
$118,400 - $186,500Hybrid

About The Position

At Jefferson Lab, you’ll champion cutting-edge science and operational excellence while shaping the future of discovery. Join us and make your mark – where excellence meets purpose, and great minds truly matter. The good-faith pay range for this role is $118,400 - $186,500 per year. Actual compensation may vary and may be above the posted range based on factors such as a candidate's skills, experience, education, certifications, and work location. Jefferson Lab, in collaboration with Lawrence Berkeley National Laboratory are building the High-Performance Data Facility (HPDF), a landmark DOE ASCR investment in scientific data infrastructure. Spanning from the Virginia coast to the California Bay Area, HPDF will deliver seamless, integrated data services across both institutions, purpose-built for the demands of modern data science and research. Unlike traditional facilities where data is an afterthought, HPDF is designed to be a data-first facility that will treat scientific data as a first-class asset and provides the capabilities researchers need to extract its full scientific value. This is a greenfield storage architecture role. You are the principal authority on storage architecture for the High-Performance Data Facility (HPDF), leading the design of a large-scale, multi-tiered, distributed storage system purpose-built for DOE scientific data lifecycle. You research and evaluate the full landscape of commercial vendors, cloud providers, and open-source distributed storage solutions, develop performance and cost models, and work closely with peer architects to ensure coherent, well-grounded technology selections as the project advances toward its design milestones. When the facility is complete, you lead the HPDF storage operations team, owning storage infrastructure across its distributed locations.

Requirements

  • Experience with hybrid cloud architectures for HPC workloads including cloud bursting and cross-site data movement over a WAN.
  • 10 or more years HPC, scientific computing, large-scale distributed systems, or equivalent.
  • 4 or more years designing and deploying high-performance parallel or distributed file systems at petabyte scale or beyond.
  • Experience designing multi-tiered, highly available, and resilient storage architectures, including data staging, archival, and immutability strategies across distributed locations.
  • Hands-on experience designing and deploying parallel or distributed file systems at petabyte scale or beyond in an HPC or scientific computing environment.
  • Experience developing storage performance models, capacity planning frameworks, and total cost of ownership analyses at facility scale.
  • Bachelor's Degree Computer Science, Information Systems, or other relevant information technology.

Nice To Haves

  • Experience at a DOE national laboratory, research university HPC center, or equivalent scientific computing environment.
  • Master's Degree Computer Science, Information Systems, or other relevant information technology.
  • Familiarity with scientific data workflows including detector data ingest, MPI-IO access patterns, and HSM.
  • Familiarity with data integrity, encryption, and compliance requirements including FAIR data principles, data retention policies, and relevant federal data management standards.

Responsibilities

  • Lead the technical design of a multi-tiered, widely distributed storage system incorporating reliability, immutability, data integrity, and high-throughput performance across distributed locations.
  • Evaluate storage technologies: HPC platforms (VAST, Weka, Spectrum Scale), open-source file systems (Lustre, DAOS, Ceph), and cloud. Engage vendors, assess solutions, and develop analyses to support architecture and acquisition decisions.
  • Develop and validate models to predict storage performance, capacity, and total cost of ownership at scale. Define KPPs and benchmarking strategies to govern storage system selection and validation.
  • Define initial SLOs and SLIs for storage performance and availability. Architect and implement monitoring and alerting solutions supporting design validation and future facility operations.
  • Define the operational framework and staffing plans for the future storage team, which will be responsible for the lifecycle management of the entire storage infrastructure.
  • Lead small team

Benefits

  • Medical, Dental, and Vision Care Plans
  • Flexible Spending Accounts
  • Paid Time-off and Leave Programs (Paid Parental, vacation, holidays, and sick leave)
  • 401(k) Plan – 9% Lab Contribution; 100% vested
  • Flexible Work Arrangements (Remote & Alternate Work Schedules available)
  • Tuition Assistance, Training and Professional Development Programs
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service