IT Systems Engineer - Internal Platforms & SRE
Overview
Design and deliver software to improve availability scalability reliability and efficiency of internal IT systems. Solve problems related to mission critical services and build automation to prevent problem recurrence. Influence architecture standards for large scale distributed systems. Engage in capacity planning and performance analysis. Be an excellent communicator producing documentation. You have a keen interest in system design and experience with AWS GCP Azure. Think carefully about edge cases failure modes and specific implementations.
What You'll Do11
- 1Design Build Write and deliver software to improve availability scalability reliability and efficiency of internal IT systems
- 2Solve problems relating to mission critical services and build automation to prevent problem recurrence
- 3Influence architecture standards for large scale distributed systems
- 4Engage in service capacity planning and demand forecasting software performance analysis and system tuning
- 5Be an excellent communicator producing documentation and related artifacts
- 6Have solid programming skills in Python Go etc.
- 7Have an urge to collaborate and communicate asynchronously combined with a desire to record and document issues and solutions
- 8Have an enthusiastic go-for-it attitude when seeing something broken you cannot help but fix
- 9Have an urge for delivering quickly and effectively iterating fast
- 10Practical experience implementing and managing paging alerting and on-call scheduling flows
- 11Positive attitude combined with a desire to learn and collaborate
Requirements8
- 15+ years building ETL pipelines with Spark and Airflow
- 2Experience with Chef Ansible Terraform GitHub Actions
- 3Knowledge of AWS GCP Azure
- 4Configuration management systems and toolchains
- 5Urge to deliver quickly and effectively
- 6Interest in ML/AI workloads and compute
- 7Practical experience with paging alerting and on-call scheduling flows
- 8Positive attitude and collaborative mindset
Salary Insight
$206 - $275k per year
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.
Uipath
VerifiedSenior Software Engineer SRE
Design and engineer SRE platform systems using AI while leading cross‑team initiatives. You will identify gaps across teams, design solutions, build, ship, and adopt them, and drive measurable improvements in reliability scalability and performance. This role focuses on ownership accountability and fostering a culture of continuous iteration without relying on generic statements.

Xoriant Corporation
VerifiedSenior/Staff SRE, AI/ML Platform Infrastructure
Own the reliability of a large-scale AI/ML platform serving millions of requests daily. Kubernetes and Docker are your primary tools. Join a team of 8 SREs supporting 50+ microservices on AWS and GCP. This role focuses on incident command, automation, and platform improvements.
Empire Abrasive Equipment Co.
VerifiedSystems Engineer Empire Abrasive Equipment Co.
Design and maintain complex systems to ensure optimal performance and reliability. Lead development of system architectures and collaborate across teams. Own system upgrades and provide technical support. Drive improvements in system efficiency and scalability. Stand out by delivering measurable impact in a dynamic environment.

BCforward
VerifiedTechnology & Information Architectures - Performance & Reliability Engineer
Own performance and reliability engineering initiatives scaling across hybrid cloud environments. Deliver high availability solutions leveraging Python AWS React Kubernetes and Spark. Drive observability and automation using industry best practices. Impact multiple global teams.
Cambia Health Solutions
VerifiedPlatform Infrastructure Engineer (AWS, Terraform)
You will own the design and implementation of secure, scalable cloud foundations and shared services across AWS and Azure in a high-uptime, high-security environment. You will lead modernization of platform operations through Infrastructure as Code, CI/CD, and observability standards, influencing leadership decisions and mentoring engineers. This role combines hands-on delivery with coaching, championing responsible automation including AI-accelerated workflows. You will drive enterprise best practices and reference architectures for hybrid integrations, reducing toil and improving reliability.
Amazon.com Services LLC
VerifiedSystem Development Engineer II, Manufacturing Systems
You will architect, build, and administer commercial-off-the-shelf CAD, simulation, and product software for Amazon's Robotics and Industrial Systems, ensuring low-latency operations across global manufacturing sites. Working within the Manufacturing Systems Engineering team, you will establish a departmental IT team from the ground up, managing software vendor relationships and contract negotiations. You will build AWS-based monitoring and integrations, set up identity management (AAA) with Amazon standard tooling, and drive high availability for critical applications. This role combines deep technical work with strategic ownership, directly impacting production efficiency at scale.