PRIMUS Global Services Inc.
PRIMUS Global Services Inc.Verified Source

AI Engineer, Python FastAPI & LLM Systems

70K–75K
Onsite · San Jose, California
Posted August 13, 2026
contract

Overview

Own the design and deployment of an AI-powered support bot for Hypercare query resolution, processing thousands of enterprise queries daily. You will build RAG pipelines, Agentic AI Workflows, and integrate Large Language Models into production. Work with Python and FastAPI in an onsite Sunnyvale, CA team. This role stands out for its focus on agentic systems and knowledge assist, requiring deep technical ownership from day one.

What You'll Do8

  • 1Build the core RAG pipeline using Python and FastAPI to retrieve and synthesize answers from enterprise knowledge bases.
  • 2Design Agentic AI Workflows that break down complex queries into sub-tasks, coordinate tool calls, and produce reliable resolutions.
  • 3Ship LLM-based features such as prompt chaining, context windowing, and model routing to optimize response accuracy and latency.
  • 4Integrate the support bot with internal data sources and APIs to enable real-time Hypercare query handling.
  • 5Debug and tune LLM performance, managing token budgets and logging traces to minimize hallucination.
  • 6Drive testing and evaluation of retrieval quality, using RAG metrics to refine chunking and embedding strategies.
  • 7Scale the FastAPI service to handle high concurrency, implementing caching and async patterns.
  • 8Collaborate with product and data teams to define knowledge schemas and ingestion workflows for the assistant.

Requirements8

  • 15+ years building AI applications with Python and FastAPI in production environments.
  • 23+ years hands-on with LLM frameworks like LangChain, LlamaIndex, or Haystack for RAG systems.
  • 32+ years designing Agentic AI Workflows, including tool orchestration and autonomous decision-making.
  • 4Experience integrating enterprise knowledge bases (e.g., Elasticsearch, Pinecone, Weaviate) with vector search.
  • 5Proven track record shipping end-to-end AI products from prototype to production, including monitoring and observability.
  • 6Strong understanding of prompt engineering, fine-tuning, and model evaluation for support chatbots.
  • 7Bachelors or Masters in CS or related field, with a focus on AI/ML.
  • 8Onsite availability in Sunnyvale, CA for the length of the contract.

Salary Insight

$70 - $75k per year

Location

Typeonsite
LocationSan Jose, California

Required Skills

pythonfastapiragllmagentic-ai
Share:

Similar open positions

Explore active roles that match your skills and interests.

Cynet Systems

Cynet Systems

1d agoRemotecontract

Senior AI Engineer, LLM & RAG Systems

You will own the design and deployment of AI-powered backend systems for a client in Woodland Hills, CA, operating at enterprise scale. You will build on a foundation of Python expertise, integrating LLMs, RAG architectures, and vector databases to deliver production-grade applications. You will collaborate with a small team of engineers and data scientists to solve complex problems. This contract role offers the chance to shape the AI infrastructure from the ground up, with direct impact on business outcomes.

50K–55K
PythonLarge Language Models (LLMs)Prompt Engineering+4 more
QUANTUM TECHNOLOGIES LLC

QUANTUM TECHNOLOGIES LLC

4h agoBoston, Massachusettscontract

AI/ML Engineer, LLM & Agentic Systems

You will architect and ship production-grade LLM applications and RAG pipelines in a 10-month contract for a Boston-area client. You will design multi-agent frameworks with LangChain or AutoGen, optimize retrieval with Pinecone or Weaviate, and deploy models on AWS or Azure. You will work with a tight team of engineers, integrating your agents with Kubernetes and Docker. This role demands 80% coding and 20% strategy, with direct ownership of the AI roadmap.

60K–70K
pythonlangchainkubernetes+2 more
Esvee Technologies Inc

Esvee Technologies Inc

11h agoBaltimore, Marylandcontract

AI Engineer

You will own the design and deployment of production AI systems, specifically focusing on NLP and LLMs within a regulated environment. You will join a team of engineers and data scientists, building on a successful track record of delivering AI solutions. Your work will directly impact the firm's compliance and operational efficiency, requiring close collaboration with product and legal teams.

Competitive salary
nlpllmrag+2 more

TechniPros, LLC

11h agoNew York, New Yorkcontract

Lead Agentic AI Engineer, LLM & LangGraph

Build enterprise-grade AI agents that reason, plan, and execute autonomously. You'll architect scalable applications with LangGraph, RAG, and MCP, integrating cloud AI platforms like AWS Bedrock and Azure AI. You'll lead a team of engineers, shaping the technical roadmap from day one. This role demands deep expertise in LLM orchestration and production deployment.

Competitive salary
pythonlanggraphaws+2 more
Saim Technologies

Saim Technologies

11h agoSan Jose, Californiapayroll

Lead AI Engineer, Agentic AI Systems

You will own the architecture and delivery of production-grade Agentic AI solutions at scale. You will lead a team of AI engineers, defining technical strategy and driving end-to-end system development. This role stands out for its focus on deploying autonomous agents in real-world enterprise environments, requiring deep expertise in Python, LangChain, and LLM orchestration.

125K–146K
pythonlangchainaws+2 more
Xoriant Corporation

Xoriant Corporation

11h agoSan Jose, Californiacontract

Junior Software Engineer, Agentic AI & Python

Build production applications powered by agentic AI as a Junior Software Engineer in San Jose. You'll own feature development from design to deployment, working primarily in Python to integrate LLM capabilities into real software systems. Join a collaborative team of engineers and AI specialists, shipping code that directly impacts enterprise clients. This contract role offers hands-on experience with cutting-edge agentic AI, distinguishing it from typical junior positions.

50K–70K
pythonfastapiflask+2 more