Serko Ltd
Serko LtdVerified Source
Remote

Principal Engineer AI Platform Operations

168K–230K
Remote · Washington, District of Columbia
Posted August 12, 2026
payroll

Overview

Serko Ltd seeks a Principal Engineer to architect the AI Platform & Operations. You will define the long-term technical roadmap for AI products, establish engineering benchmarks, and optimize GPU compute efficiency. This role drives platform stability and reliability while mentoring senior engineers and championing internal developer platforms.

What You'll Do11

  • 1Define the long-term technical roadmap for our AI platform covering model serving and feature stores
  • 2Establish engineering benchmarks for deployment versioning and A/B testing
  • 3Lead strategies for GPU compute efficiency and cost optimization during LLM inference
  • 4Design sophisticated monitoring and alerting systems for AI workloads
  • 5Drive platform stability through architecture reviews and evaluation of emerging cloud services
  • 6Mentor senior engineers and lead architecture discussions
  • 7Champion reliability by partnering with application teams to meet product needs
  • 8Operate internal platforms treating engineers as primary customers
  • 9Optimize ML inference through quantization batching and latency reduction
  • 10Implement CI/CD pipelines for machine learning models
  • 11Scale self-serve developer tools for rapid deployment and monitoring

Requirements10

  • 15+ years building ETL pipelines with Spark and Airflow
  • 2Expert-level knowledge of model serving systems such as Triton vLLM and Ray Serve
  • 3Deep experience with Kubernetes and cloud ecosystems AWS GCP Azure
  • 4Proven use of MLflow Weights & Biases or Kubeflow
  • 5High proficiency in Python and containerization technologies Docker Helm
  • 6Hands-on experience operating LLM inference at scale
  • 7Track record of building internal platforms that improve engineering velocity
  • 8Strong background in systems thinking Docker image creation Helm chart deployment
  • 9Understanding of observability solutions for AI workloads
  • 10Experience with CI/CD practices for machine learning pipelines

Salary Insight

$168 - $230k per year

Location

Typeremote
LocationWashington, District of Columbia
This is a remote position

Required Skills

pythonsparkairflowtritonvllm
Share:

Similar open positions

Explore active roles that match your skills and interests.

Unisoft Technology Inc

Unisoft Technology Inc

19h agoWashington, District of Columbiacontract

Sr Lead AI Data Engineer

Lead design and delivery of AI infrastructure to drive scalable machine learning solutions across enterprise platforms. Own development of end-to-end ML pipelines and foster collaboration between data science and engineering teams. This role differs by focusing on cross-functional leadership and production-grade MLOps implementation.

Competitive salary
PythonPyTorchTensorFlow+4 more
VDart, Inc.

VDart, Inc.

16h agoAtlanta, Georgiapayroll

Senior AI DevOps Engineer AI Ops Platform Engineering

Design and implement scalable AI-driven CI/CD pipelines using Python AWS React Kubernetes Spark Airflow. Own automated workflows that accelerate software delivery while integrating generative AI and model context protocol solutions.

114K–119K
pythonawsreact+2 more
TetraScience

TetraScience

20d agoRemotepayroll

Lead Software Platform Engineer MLOps TetraScience

We seek a Lead Software Platform Engineer at the intersection of distributed systems and MLOps to own and scale AI and data infrastructure for customers. This role involves architecting cloud-based services and MLOps platforms enabling production-grade AI workflows for pharmaceutical clients while ensuring security and compliance. The ideal candidate thrives in regulated environments and drives technical strategy for multi-tenant AI products.

200K–270K
TypeScriptPythonAWS+6 more
Lorven Technologies, Inc.

Lorven Technologies, Inc.

8d agoColumbus, Ohiopayroll

AI Engineer Lorven Technologies Columbus Ohio Onsite Payroll

We seek an experienced AI Engineer to design and deploy AI/ML solutions at scale. This role involves leading technical initiatives that deliver measurable business impact across large datasets. The ideal candidate thrives in a collaborative environment while driving innovation through cutting-edge technologies.

125K–146K
PythonTensorFlowLLM+1 more
Bain & Co.

Bain & Co.

14h agoHouston, Texaspayroll

Senior AI/ML Engineer, LLMOps & RAG Systems

You build production inference, serving, and LLMOps infrastructure for Python-based ML systems, taking models from prototype to governed deployment. Bain's Private Equity Group Innovation team creates proprietary data and software products, and your work supports over 1,000 professionals across the investment lifecycle. You collaborate with Data Scientists, Data Engineers, and the Agent/AI squad to deliver reliable, observable ML systems. This hands-on role sets engineering standards and mentors mid-level engineers, with daily use of MLflow, Kubernetes, and Databricks.

141K–169K
PythonMLflowLLMOps+12 more
TechniPros, LLC

TechniPros, LLC

18h agoCharlotte, North Carolinacontract

Senior Generative AI Platform Engineer

Lead design and delivery of scalable enterprise AI platforms using modern LLMs AI orchestration frameworks cloud-native technologies and Model Context Protocol (MCP). Drive platform ownership and ship production-grade GenAI applications. This role differs by focusing on MCP integration and production scalability across multiple locations.

Competitive salary
pythonawsreact+2 more