LambdaVerified Source

IT Systems Engineer - Internal Platforms & SRE

206K–275K
Partially · San Francisco, California
Posted July 30, 2026
payroll

Overview

Design and deliver software to improve availability scalability reliability and efficiency of internal IT systems. Solve problems related to mission critical services and build automation to prevent problem recurrence. Influence architecture standards for large scale distributed systems. Engage in capacity planning and performance analysis. Be an excellent communicator producing documentation. You have a keen interest in system design and experience with AWS GCP Azure. Think carefully about edge cases failure modes and specific implementations.

What You'll Do11

  • 1Design Build Write and deliver software to improve availability scalability reliability and efficiency of internal IT systems
  • 2Solve problems relating to mission critical services and build automation to prevent problem recurrence
  • 3Influence architecture standards for large scale distributed systems
  • 4Engage in service capacity planning and demand forecasting software performance analysis and system tuning
  • 5Be an excellent communicator producing documentation and related artifacts
  • 6Have solid programming skills in Python Go etc.
  • 7Have an urge to collaborate and communicate asynchronously combined with a desire to record and document issues and solutions
  • 8Have an enthusiastic go-for-it attitude when seeing something broken you cannot help but fix
  • 9Have an urge for delivering quickly and effectively iterating fast
  • 10Practical experience implementing and managing paging alerting and on-call scheduling flows
  • 11Positive attitude combined with a desire to learn and collaborate

Requirements8

  • 15+ years building ETL pipelines with Spark and Airflow
  • 2Experience with Chef Ansible Terraform GitHub Actions
  • 3Knowledge of AWS GCP Azure
  • 4Configuration management systems and toolchains
  • 5Urge to deliver quickly and effectively
  • 6Interest in ML/AI workloads and compute
  • 7Practical experience with paging alerting and on-call scheduling flows
  • 8Positive attitude and collaborative mindset

Salary Insight

$206 - $275k per year

Location

Typepartially
LocationSan Francisco, California

Required Skills

PythonGoAWSGCPAzureChefAnsibleTerraform
Share:

Similar open positions

Explore active roles that match your skills and interests.

Uipath

23h agoDenver, Coloradopayroll

Senior Software Engineer SRE

Design and engineer SRE platform systems using AI while leading cross‑team initiatives. You will identify gaps across teams, design solutions, build, ship, and adopt them, and drive measurable improvements in reliability scalability and performance. This role focuses on ownership accountability and fostering a culture of continuous iteration without relying on generic statements.

160K–210K
pythonawsreact+2 more
Xoriant Corporation

Xoriant Corporation

1d agoSan Jose, Californiapayroll

Senior/Staff SRE, AI/ML Platform Infrastructure

Own the reliability of a large-scale AI/ML platform serving millions of requests daily. Kubernetes and Docker are your primary tools. Join a team of 8 SREs supporting 50+ microservices on AWS and GCP. This role focuses on incident command, automation, and platform improvements.

Competitive salary
Production on-callIncident commandBlameless postmortem+15 more

Empire Abrasive Equipment Co.

21h agoPhiladelphia, Pennsylvaniapayroll

Systems Engineer Empire Abrasive Equipment Co.

Design and maintain complex systems to ensure optimal performance and reliability. Lead development of system architectures and collaborate across teams. Own system upgrades and provide technical support. Drive improvements in system efficiency and scalability. Stand out by delivering measurable impact in a dynamic environment.

Competitive salary
Operating SystemsWindowsLinux+8 more
BCforward

BCforward

15h agoMinneapolis, Minnesotacontract

Technology & Information Architectures - Performance & Reliability Engineer

Own performance and reliability engineering initiatives scaling across hybrid cloud environments. Deliver high availability solutions leveraging Python AWS React Kubernetes and Spark. Drive observability and automation using industry best practices. Impact multiple global teams.

70K–74K
pythonawsreact+2 more
Cambia Health Solutions

Cambia Health Solutions

20h agoPortland, Oregonpayroll

Platform Infrastructure Engineer (AWS, Terraform)

You will own the design and implementation of secure, scalable cloud foundations and shared services across AWS and Azure in a high-uptime, high-security environment. You will lead modernization of platform operations through Infrastructure as Code, CI/CD, and observability standards, influencing leadership decisions and mentoring engineers. This role combines hands-on delivery with coaching, championing responsible automation including AI-accelerated workflows. You will drive enterprise best practices and reference architectures for hybrid integrations, reducing toil and improving reliability.

Competitive salary
AWSTerraformAnsible+8 more

Amazon.com Services LLC

18h agoAustin, Texaspayroll

System Development Engineer II, Manufacturing Systems

You will architect, build, and administer commercial-off-the-shelf CAD, simulation, and product software for Amazon's Robotics and Industrial Systems, ensuring low-latency operations across global manufacturing sites. Working within the Manufacturing Systems Engineering team, you will establish a departmental IT team from the ground up, managing software vendor relationships and contract negotiations. You will build AWS-based monitoring and integrations, set up identity management (AAA) with Amazon standard tooling, and drive high availability for critical applications. This role combines deep technical work with strategic ownership, directly impacting production efficiency at scale.

129K–175K
PythonRubyGolang+7 more