Cloudflare logo
Cloudflare·Verified

Senior Infrastructure Engineer, Storage Platform - Cloudflare

Hybrid remoteFull-timeSenior$185K - $254KUnited StatesUnited KingdomAustin+2 more#goBonus

About the Role

As a Senior Infrastructure Engineer on the Storage Infrastructure team within Emerging Technologies & Incubation (ETI), you will build and operate a shared storage platform for products like R2, Workers KV, and Durable Objects. You will manage underlying storage hardware, distributed databases, and object storage clusters, focusing on fleet lifecycle automation, capacity, and production operations.

Responsibilities

  • Automation: Design and build operator tooling for provisioning, configuring, upgrading, and decommissioning storage hardware and clusters.
  • Resilience: Engineer globally distributed storage fleets to tolerate hardware failures, network disruptions, and capacity pressure.
  • Observability: Develop alerting and safety controls to ensure auditable production changes and fleet health.
  • Incident Response: Diagnose complex issues across Linux, storage, and networking; participate in on-call rotations.
  • AI Integration: Utilize AI tools to accelerate debugging, root-cause analysis, and toil reduction while maintaining sound engineering judgment.
  • Hardware Validation: Characterize storage hardware performance under various conditions (normal, failure, rebuilds).
  • Collaboration: Partner with cross-functional teams including R2, Workers KV, and Capacity Planning to turn requirements into infrastructure capabilities.

Requirements

  • Experience designing and operating large-scale infrastructure platforms or production fleets.
  • Strong Linux systems knowledge (compute, storage, and networking).
  • Experience operating distributed systems in production (observability, reliability, change management).
  • Proficiency in at least one programming language (Go, Rust, or Python).
  • Excellent communication skills for cross-team collaboration.

Bonus Points

  • Experience with distributed databases, object storage, or storage fundamentals (SSD behavior, replication).
  • Experience with bare-metal provisioning or hardware lifecycle management.
  • Familiarity with infrastructure-as-code (Terraform, SaltStack) or observability (Prometheus, Grafana).
  • Knowledge of capacity modeling or cross-datacenter networking.

Benefits

Bonus

Open to

Austin · United States · Seattle · London +1

Sign in to track applications and earn points.

More roles at Cloudflare

Similar remote roles