Senior Storage Engineer, File & Block

CoreWeaveSunnyvale, CA
$165,000 - $242,000

About The Position

CoreWeave is seeking a Senior Storage Engineer, File & Block to join their Storage team. This role will be instrumental in building and operating the file and block storage services that power CoreWeave's high-performance, multi-tenant infrastructure, crucial for AI training, inference workloads, and internal stateful services. The engineer will focus on enhancing storage capabilities to not only keep pace with GPU growth but to actively drive it, taking ownership of the data path and control plane for both file and block storage. This involves running these services on CoreWeave's own hardware fleet and developing customer-facing features that support large-scale deals. Collaboration with compute, platform, and infrastructure teams will be key to ensuring storage is fast, reliable, and easily consumable at scale.

Requirements

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
  • 6–10 years of experience building storage systems, distributed systems, or infrastructure services.
  • Strong hands-on experience with distributed or networked file systems and/or block storage in production (e.g., NFS/POSIX file semantics at scale, and/or block volume services with durability, snapshots, replication).
  • Depth in the storage data path, including IO performance, caching, small-file and metadata-heavy workloads, crash consistency, and resilience under load.
  • Experience with multi-tenant services or storage control planes, including provisioning, tenancy/quota, and data-protection features (snapshots, encryption, key management).
  • Proficiency in a systems/back-end language, with Go strongly preferred (C or Rust a plus).
  • Experience building on Kubernetes, including controllers/operators, CRDs, and CSI (dynamic provisioning, resize, snapshots, RWO volumes).
  • Familiarity with distributed databases (e.g., CockroachDB), workflow orchestration (e.g., Temporal), and gRPC/Protobuf service design.
  • Experience with distributed or parallel storage stacks such as Ceph/RBD, Lustre, GPFS/Spectrum Scale, BeeGFS, WEKA, VAST, or DAOS is ideal.
  • Familiarity with storage observability tools and telemetry pipelines (e.g., ClickHouse, Prometheus, Grafana).
  • Bonus: exposure to high-performance data-path technologies (RDMA, GPUDirect Storage, RoCE, InfiniBand, SPDK).
  • Strong debugging and problem-solving skills in distributed, high-performance environments.
  • Clear communicator, able to work collaboratively across teams and share technical insights effectively.

Nice To Haves

  • Experience with distributed or parallel storage stacks such as Ceph/RBD, Lustre, GPFS/Spectrum Scale, BeeGFS, WEKA, VAST, or DAOS.
  • Exposure to high-performance data-path technologies (RDMA, GPUDirect Storage, RoCE, InfiniBand, SPDK).

Responsibilities

  • Design, build, and operate highly scalable, multi-tenant file and block storage services, covering both the data path (NFS for file; durable, high-performance block volumes) and the control plane for tenancy, provisioning, and data protection.
  • Manage the file system data path, focusing on NFS hot-path reliability, IO resilience, and performance, particularly for small-file/IOPS-heavy workloads and throughput for GPUs.
  • Develop a native block storage service offering durable, high-performance block volumes for customer workloads and internal stateful services.
  • Deliver file and block storage as first-class Kubernetes citizens, including designing and implementing a backend-neutral CSI driver with dynamic provisioning, resize, snapshots, and RWO block volumes, backed by a public provisioning SLO.
  • Build the control plane for storage, encompassing multi-tenant isolation, quota, QoS, snapshots, key management (BYOK), audit, and lifecycle policy to meet the needs of regulated and enterprise customers.
  • Enhance the reliability, durability, and observability of the storage stack, partnering with operations to monitor, analyze, and optimize performance, latency, and resilience using telemetry, metrics, and dashboards.
  • Collaborate with platform, product, and infrastructure teams to ensure seamless storage integration across the stack, including tiering cold data to object storage.
  • Share knowledge and mentor other engineers on best practices for building distributed, high-performance systems.
  • Work with technologies such as RDMA, GPU Direct Storage, RoCE, InfiniBand, SPDK, and distributed filesystems to optimize storage performance and efficiency.
  • Participate in efforts to improve the reliability, durability, and observability of the storage stack.
  • Collaborate with operations teams to monitor, analyze, and optimize storage systems using telemetry, metrics, and dashboards to improve performance, latency, and resilience.
  • Work cross-functionally with platform, product, and infrastructure teams to deliver seamless storage capabilities across the stack.
  • Share knowledge and mentor other engineers on best practices in building distributed, high-performance systems.

Benefits

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave
  • Company-paid Life Insurance
  • Voluntary supplemental life insurance
  • Short and long-term disability insurance
  • Flexible Spending Account
  • Health Savings Account
  • Tuition Reimbursement
  • Ability to Participate in Employee Stock Purchase Program (ESPP)
  • Mental Wellness Benefits through Spring Health
  • Family-Forming support provided by Carrot
  • Paid Parental Leave
  • Flexible, full-service childcare support with Kinside
  • 401(k) with a generous employer match
  • Flexible PTO
  • Catered lunch each day in our office and data center locations
  • A casual work environment
  • A work culture focused on innovative disruption
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service