Staff Infrastructure Engineer, AWS & Terraform
Overview
You will architect and own the cloud platform that every engineer at Headway deploys on, making deploys boring and scaling automatic. You will lead the shift toward per-service deploy isolation, redesign the ECS and EKS footprint, and build a self-serve Terraform platform with guardrails. You will set the technical direction for compute, networking, and deployment, and raise the infrastructure bar org-wide. This role offers staff-level influence in a Series D company serving over 75,000 providers and 1 million patients, with the mandate to make every change safe as AI accelerates code velocity.
What You'll Do7
- 1Redesign deployment architecture to isolate blast radius, ensuring a mistake in one service cannot block others, with per-service deploy isolation and functional-area slices.
- 2Own the ECS and EKS container footprint, evaluate broader EKS adoption for AI workloads, and design the next iteration of inter-service network connectivity.
- 3Build capacity for spiky workloads by computing floors ahead of demand, self-deriving from data, and catching drift early.
- 4Build the Terraform self-serve platform with guardrails so engineering teams own standard infrastructure changes, and Reliability Engineering reviews only non-standard ones.
- 5Stand up per-team cost attribution across AWS, Datadog, and LLM spend, making infrastructure costs visible and attributable.
- 6Own the Python runtime and dependency health: garbage collection, event loop contention, and runtime limits, leading framework and package upgrades.
- 7Make deploys safe and self-serve for other teams, not just your own, through architecture reviews, runbooks, and paved-road tooling.
Requirements7
- 18+ years in platform, infrastructure, or SRE roles at companies running significant production traffic.
- 2Deep AWS expertise and production ownership of compute and networking at scale: ECS, EKS, RDS, networking, IAM.
- 3Strong infrastructure-as-code experience, particularly Terraform, including designing self-serve platforms for other engineering teams.
- 4Hands-on autoscaling and capacity engineering, and container orchestration with ECS and/or EKS.
- 5Track record making deploys safe and self-serve for other teams, not just your own.
- 6Staff-level influence: drive decisions across team boundaries and raise the infrastructure bar org-wide without requiring management authority.
- 7Nice to have: FinOps and cloud cost optimization experience, Kubernetes and EKS depth, observability tooling at scale (Datadog), healthcare experience, event-driven systems.
Salary Insight
$265 - $331k per year
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.
Cambia Health Solutions
VerifiedPlatform Infrastructure Engineer (AWS, Terraform)
You will own the design and implementation of secure, scalable cloud foundations and shared services across AWS and Azure in a high-uptime, high-security environment. You will lead modernization of platform operations through Infrastructure as Code, CI/CD, and observability standards, influencing leadership decisions and mentoring engineers. This role combines hands-on delivery with coaching, championing responsible automation including AI-accelerated workflows. You will drive enterprise best practices and reference architectures for hybrid integrations, reducing toil and improving reliability.

Zensar Technologies Inc.
VerifiedCloud Platform Engineer, AWS & Kubernetes
You will own the design and delivery of cloud infrastructure at scale, powering enterprise workloads on AWS and Kubernetes. You join a collaborative team of 12 engineers, partnering with product and security to ship resilient platforms. This role blends deep infrastructure work with automation, making you the go-to for Terraform and CI/CD pipelines. You will drive reliability and cost efficiency from day one, shaping the platform roadmap.
Tetrix
VerifiedSenior Infra/DevOps Engineer, AWS & CI/CD
Own infrastructure for a AWS-native platform serving institutional investors in private markets. You will design scalable systems for high-volume document ingestion and AI-powered data pipelines at scale. Work with a small team of engineers, reporting to the CTO, and collaborate directly with product and clients. This role stands out by combining infrastructure leadership with client-facing work and strategic input on product direction.
NinjaOne
VerifiedStaff DevOps Engineer - CI/CD & AWS
Own the design, operation, and continuous improvement of deployment pipelines at scale, ensuring code moves from commit to production safely and predictably. As part of the Deployment Engineering team, you will architect progressive delivery strategies, automate release orchestration, and mentor engineers. Collaborate closely with Development, QA, Security, and IT to keep releases secure and resilient. This role offers hands-on technical leadership with strategic impact, driving down lead time while building guardrails for confident shipping.
Bright Vision Technologies
VerifiedStaff Software Engineer Remote | Cloud & Distributed Systems
You will own the technical vision for cross-cutting initiatives, influencing architecture across multiple teams in a 100% remote U.S. environment. You'll work with AWS, Kubernetes, and Terraform to build scalable cloud-native systems, collaborating closely with product, design, and engineering leaders. Your mandate is to raise the engineering bar through code review, design review, and mentorship of senior engineers. This role stands out for its direct impact on multi-year platform modernization and the freedom to shape technical strategy without the noise of office politics.

Apex Systems
VerifiedDevOps Engineer - AWS, Kubernetes, Terraform
Own the AWS and Kubernetes (EKS) environments at a Denver-based company, with a focus on security-first infrastructure. You will build and maintain IaC pipelines with Terraform, automating everything from library publishing to application deployments and data promotions. As part of a tight DevOps team, you will collaborate with software engineers to productionize services and ensure high availability. This role stands out because it demands a strong software development mindset, not just ops experience.