Senior Platform Engineer Cloud Infrastructure Go
Overview
We seek a senior engineer to design and operate cloud-native platforms using Go. This role involves building Kubernetes clusters managing networking workload isolation and multi-region topologies implementing service mesh features like mTLS and traffic management optimizing containerized workloads developing Go services and middleware writing unit and integration tests leading incident response defining SLOs and building alerting systems managing infrastructure as code and CI/CD pipelines planning cloud migrations and ensuring observability through metrics and tracing collaborating with product and security teams.
What You'll Do10
- 1Design and operate production Kubernetes clusters including networking workload isolation and multi-region topologies
- 2Work with Kubernetes internals resource quota management scheduling and cluster behaviour network policy enforcement and custom controllers or operators
- 3Implement service mesh capabilities mTLS between services service account level authentication and authorization and traffic management across internal and external request paths
- 4Optimize containerized workloads for performance cost and resource efficiency
- 5Write refactor and maintain production Go services controllers and middleware extending platform functionality read and contribute to complex codebases including open source projects
- 6Build HTTP REST and gRPC interfaces lead incident response for platform level issues investigate root causes and author postmortems define and drive SLOs and build actionable alerting systems
- 7Own infrastructure as code authoring reusable modules and maintaining them as platform evolves
- 8Build and improve CI/CD and GitOps delivery workflows balance developer velocity against reliability security and compliance plan and execute cloud migration initiatives maintain reliability during workload moves
- 9Build and maintain metrics dashboards alerting policies and distributed tracing ensure observability instrument services for diagnosable failures
- 10Partner with product security and infrastructure teams collaborate on design reviews and mentor engineers to raise engineering standards
Requirements22
- 16+ years professional experience in software platform infrastructure or SRE including production distributed systems
- 2Production experience writing Go with primary language daily
- 3Experience designing systems from ambiguous starting points and delivering to production
- 4Migration of cloud services planning and executing production workload moves between providers
- 5Degree in Computer Science Engineering or equivalent practical experience
- 6Strong understanding of Kubernetes internals networking CNI NetworkPolicy resource management and cluster behaviour under load
- 7Hands-on experience with service mesh Istio Envoy Linkerd or similar and mTLS workload identity
- 8Solid Linux fundamentals including cgroups and resource management
- 9Infrastructure as code at scale Terraform or equivalent
- 10Production experience with major cloud platforms GCP AWS or Azure multicloud a strong plus
- 11Observability tooling Prometheus Grafana OpenTelemetry PromQL and database expertise PostgreSQL or managed services
- 12Debugging and performance profiling skills
- 13Preferred Go testing frameworks Ginkgo Gomega experience building Kubernetes controllers operators or API server extensions
- 14Additional strength in Python
- 15Identity and access management SSO Keycloak OIDC SAML or secret management with Vault or cloud equivalents
- 16High availability and disaster recovery across regions or providers
- 17Monorepo experience and build systems such as Bazel
- 18Ability to read Java
- 19Compliance frameworks SOC 2 GDPR
- 20Interest in LLM-backed systems cloud-native environments
- 21Alibaba Cloud experience highly preferred
- 22Large-scale Data Platform technologies Apache Spark and Apache Flink
Salary Insight
Salary not disclosed in listing
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.
PAC Panasonic Avionics Corporation
VerifiedMember of Technical Staff III – Platform as a Service
Design and operate platform software and infrastructure enabling applications on Kubernetes-based edge platforms. Engineer across software development cloud-native infrastructure Kubernetes networking security automation and distributed systems. Collaborate with engineers across platform infrastructure DevOps Security and application teams to deliver reliable scalable capabilities from cloud to edge environments.
Xora Innovation
VerifiedSenior Platform Engineer Cloud Xora Portfolio Company
Elemynt seeks a Senior Platform Engineer to architect and maintain the cloud infrastructure and core platform services that power our AI-driven materials platform. You will own networking identity access Kubernetes and infrastructure-as-code ensuring reproducibility and security across deployments. This hands-on role involves building scalable services observability pipelines and deployment processes while standing out in a fast-paced environment.

Apex Systems
VerifiedCloud DevOps Engineer - SaaS IGA Platform
Own the cloud infrastructure and CI/CD pipelines for a SaaS Identity Governance and Administration (IGA) platform in San Francisco. Design and maintain automation that keeps the platform reliable at scale, collaborating with engineering and security teams. Work with AWS, Kubernetes, Terraform, and GitLab to ship changes daily with confidence. This role puts you at the core of platform reliability and security in a cloud-native environment.

Xoriant Corporation
VerifiedSenior/Staff SRE, AI/ML Platform Infrastructure
Own the reliability of a large-scale AI/ML platform serving millions of requests daily. Kubernetes and Docker are your primary tools. Join a team of 8 SREs supporting 50+ microservices on AWS and GCP. This role focuses on incident command, automation, and platform improvements.
Clinically AI
VerifiedDevOps Engineer, Cloud Infrastructure GCP Kubernetes
Own the cloud infrastructure, deployment systems, and reliability of a healthcare AI platform serving behavioral health organizations. You will design, build, and evolve GCP environments using Terraform, Kubernetes, and Helm, supporting real-time AI processing and customer-facing applications. Collaborate with Backend, AI, and Product teams to scale infrastructure securely and efficiently. This role offers significant autonomy and the chance to shape infrastructure strategy as the company grows.
Early Warning Services, LLC
VerifiedSenior Platform Engineer - Security Solutions
Lead the technical strategy and roadmap for the Platform Engineering team while driving security initiatives. Own the design and implementation of scalable infrastructure solutions. Shape the future of platform engineering through innovative architecture and cross-functional collaboration. Stand out as a technology leader who influences organizational direction.