LavendoVerified Source
Remote

AI Field Engineer - Infrastructure Scaling

176K–224K
Remote · San Francisco, California
Posted July 24, 2026
payroll

Overview

Own end-to-end AI infrastructure scaling for enterprise and AI-native clients. Lead discovery and production deployments from initial contact to live customer environments. Partner with VP Engineering and CTO-level leaders to drive technical wins. Work remotely with occasional travel to US hubs.

What You'll Do6

  • 1Run full pre-sales field cycle including discovery POC scoping model evaluations and final model selection
  • 2Ship production code deploying open-model LLMs using vLLM SGLang and TensorRT-LLM
  • 3Act as true partner to sales shaping deal strategy beyond advisory support
  • 4Build relationships across engineering and executive teams to move deals forward
  • 5Deploy and fine-tune LLMs while maintaining hyperscaler infrastructure on AWS Azure or GCP
  • 6Collaborate with AI/ML engineers to debug inference trade-offs and resolve production issues

Requirements6

  • 13+ years client-facing AI/ML role with hands-on experience in LLM inference and fine-tuning
  • 2Strong Python skills and familiarity with GPU infrastructure and Kubernetes
  • 3Hyperscaler experience in AWS Azure or GCP with production workloads
  • 4Background building AI features natively at AI-native companies or SaaS firms
  • 5Full-time availability with willingness to travel to enterprise customers across US
  • 6Visa sponsorship available for appropriate candidates

Salary Insight

$176 - $224k per year

Location

Typeremote
LocationSan Francisco, California
This is a remote position

Required Skills

PythonKubernetesGPU infrastructureOpen‑model LLM inferenceLLM fine‑tuningvLLMSGLangTensorRT‑LLM
Share:

Similar open positions

Explore active roles that match your skills and interests.

Wise Skulls Corp.

Wise Skulls Corp.

14h agoAustin, Texaspayroll

AI Engineer (LLM Agents & Data Engineering)

Lead design and delivery of AI solutions that scale across multiple platforms. Own the end-to-end pipeline from concept to production while driving innovation in large language models. This role shapes how our systems learn and adapt.

Competitive salary
PythonLLMsPrompt Engineering+8 more
Advent Global Solutions, Inc.

Advent Global Solutions, Inc.

20h agoDallas, Texascontract

Sr. AI Engineer, LLM & Generative AI

You own the design, build, and deployment of production-grade AI/ML solutions using Python, SQL, and modern AI frameworks. You join a team of engineers and data scientists, shipping LLM and Generative AI features that impact core business processes. You collaborate with product and infrastructure teams in a contract-to-hire role, 3 days onsite in Irving, TX. Your work directly improves model accuracy and system reliability from day one.

60K–65K
PythonSQLAPI+5 more
RELX Inc. Company

RELX Inc. Company

22h agoRaleigh, North Carolinapayroll

Senior Machine Learning Engineer III

Lead the implementation and scaling of AI systems for legal products. Partner with Data Scientists to turn validated models into reliable high-performance customer-facing systems. Own system architecture infrastructure and productionization of ML/LLM solutions. Based in Raleigh NC hybrid fully remote.

118K–220K
PythonRustGo+6 more
Intel

Intel

22h agoSan Jose, Californiapayroll

AI Infrastructure Engineer Intel

Performance‑obsessed AI Infrastructure Engineer at Intel in San Jose, California. You will drive inference performance and redefine peak performance on Intel’s next‑generation GPU architectures.

170K–315K
C++PythonGPU Computing+8 more
Lorven Technologies, Inc.

Lorven Technologies, Inc.

8d agoColumbus, Ohiopayroll

AI Engineer Lorven Technologies Columbus Ohio Onsite Payroll

We seek an experienced AI Engineer to design and deploy AI/ML solutions at scale. This role involves leading technical initiatives that deliver measurable business impact across large datasets. The ideal candidate thrives in a collaborative environment while driving innovation through cutting-edge technologies.

125K–146K
PythonTensorFlowLLM+1 more
Technogen, Inc.

Technogen, Inc.

10d agoCharlotte, North Carolinapayroll

Senior LLM Inference & GPU Systems Engineer

Technogen, Inc. seeks a Senior On-Premise LLM Inference & GPU Systems Engineer to architect and optimize high-performance inference platforms in Charlotte, NC. You will own the deployment and tuning of large language models on on-premise GPU clusters, ensuring low-latency, high-throughput serving for enterprise workloads. Collaborate with data scientists and infrastructure teams to build scalable ML pipelines. This role offers direct impact on production AI systems within a Woman-Owned Small Business with 15+ years of IT services.

687K
nvidia tritonkubernetespython+2 more