Senior LLMOps Engineer

Steampunk
  • McLean, Virginia
  • $145,000–$185,000 Per Year
20 days ago

Job Description

Overview:

We are looking for an experienced Senior LLMOps Engineer to design, implement, and maintain production-grade large-language-model (LLM) pipelines, deployment architectures, and monitoring systems across enterprise environments. The Senior LLMOps Engineer will play a critical role in operationalizing generative AI capabilities, ensuring that LLM-based applications are scalable, secure, reliable, and compliant with emerging AI risk and governance frameworks. This role spans the spectrum of model deployment, orchestration, evaluation, and optimization. 

Contributions:
  • Architect and maintain scalable LLM and RAG pipelines, including model hosting, inference optimization, retrieval layers, and context management frameworks. 
  • Lead the design and implementation of secure GenAI infrastructure across cloud environments, ensuring reliability, performance, and cost efficiency. 
  • Build and manage automated evaluation systems that assess LLM output quality, safety, latency, and adherence to AI governance requirements. 
  • Develop CI/CD workflows tailored for LLM- and GenAI-based applications, including dataset versioning, model lineage, and automated testing of prompt and model behaviors. 
  • Collaborate with AI Product Engineers and Data Scientists to productionize LLM-based prototypes into enterprise-grade, maintainable systems. 
  • Integrate vector databases, model gateways, content filters, and guardrail frameworks into end-to-end LLM solutions. 
  • Implement observability and monitoring solutions that track performance metrics, hallucination rates, cost profiles, and user interaction patterns. 
  • Lead troubleshooting and root-cause analysis for issues related to LLM deployment, inference performance, or pipeline reliability. 
  • Stay current with emerging LLM architectures, inference optimizations, fine-tuning techniques, and relevant MLSecOps patterns. 
  • Ensure compliance with data privacy, ethical AI, and AI-governance frameworks throughout pipeline design and operations. 
  • Mentor junior engineers and contribute to Steampunk’s AI engineering best practices, tooling, and reusable infrastructure patterns. 
  • You will contribute to the growth of our AI & Data Exploitation Practice! 

 

Qualifications:
  • Ability to hold a position of public trust with the U.S. government. 
  • Bachelor’s, Master’s, or Ph.D. in Computer Science, Machine Learning, Data Engineering, or a related field. 
  • 5+ years of experience in software engineering, data engineering, MLOps, or cloud engineering, with 2+ years focusing specifically on LLM or GenAI operations. 
  • Strong experience deploying models using frameworks such as Hugging Face Transformers, vLLM, TensorRT-LLM, or similar. 
  • Proficiency in Python and operational tooling such as FastAPI, PyTorch, LangChain, LlamaIndex, and vector databases (FAISS, Milvus, Pinecone, or similar). 
  • Advanced knowledge of cloud platforms (AWS, Azure, GCP) including model hosting, distributed compute, and secure networking patterns. 
  • Hands-on experience building CI/CD pipelines, automated testing frameworks, and environment provisioning for AI/ML workloads. 
  • Experience with Docker, Kubernetes, and infrastructure-as-code (Terraform, CloudFormation). 
  • Familiarity with MLSecOps, AI governance, model hardening, prompt injection defenses, and content safety monitoring. 
  • Strong understanding of logging, observability, and performance profiling for high-throughput LLM inference systems. 
  • Excellent written and verbal communication skills, with the ability to explain trade-offs and architectural decisions to technical and non-technical stakeholders. 
  • Demonstrated ability to balance long-term platform thinking with hands-on operations and rapid problem solving. 
  • Experience working in agile teams and using modern project management tools. 

Preferred:

  • Experience building and maintaining classical ML pipelines, including feature engineering, model training, and automated retraining workflows.
  • Familiarity with ML experiment tracking and model versioning tools such as MLflow or Weights & Biases.
  • Familiarity with batch and streaming data pipeline orchestration (Airflow or similar) supporting model training workflows.
  • Experience supporting the full ML lifecycle, from data ingestion through model deployment, for classification, regression, or recommendation systems.

 

About steampunk:

Steampunk relies on several factors to determine salary, including but not limited to geographic location, contractual requirements, education, knowledge, skills, competencies, and experience. The projected compensation range for this position is $145,000 to $185,000.  The estimate displayed represents a typical annual salary range for this position. Annual salary is just one aspect of Steampunk’s total compensation package for employees. Learn more about additional Steampunk benefits here. 

 

Identity Statement

As part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

 

Steampunk is a Change Agent in the Federal contracting industry, bringing new thinking to clients in the Homeland, Federal Civilian, Health and DoD sectors.  Through our Human-Centered delivery methodology, we are fundamentally changing the expectations our Federal clients have for true shared accountability in solving their toughest mission challenges.  As an employee owned company, we focus on investing in our employees to enable them to do the greatest work of their careers – and rewarding them for outstanding contributions to our growth. If you want to learn more about our story, visit http://www.steampunk.com.

Numbers & Facts

LocationMcLean, Virginia
Salary$145,000–$185,000 Per Year

Skills

  • Amazon Web Services (AWS)unmatched
  • Architectural Servicesunmatched
  • Artificial Intelligence (AI)unmatched
  • Best Practicesunmatched
  • Cloud Computingunmatched
  • Communication Skillsunmatched
  • Computer Networksunmatched
  • Computer Scienceunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Contract Requirementsunmatched
  • Data Managementunmatched
  • Data Modelingunmatched
  • Database Designunmatched
  • Dockerunmatched
  • Federal Contractsunmatched
  • GCP (Good Clinical Practices)unmatched
  • Governmentunmatched
  • High Throughputunmatched
  • Identify Issuesunmatched
  • Injectionsunmatched
  • Machine Learningunmatched
  • Machine Toolunmatched
  • Maintain Complianceunmatched
  • Mentoringunmatched
  • Microsoft Windows Azureunmatched
  • Modeling Languagesunmatched
  • Performance Analysisunmatched
  • Performance Metricsunmatched
  • Position of Public Trustunmatched
  • Presentation/Verbal Skillsunmatched
  • Problem Solving Skillsunmatched
  • Project Management Softwareunmatched
  • Python Programming/Scripting Languageunmatched
  • Riskunmatched
  • Root Cause Analysisunmatched
  • Safety/Work Safetyunmatched
  • Software Engineeringunmatched
  • Systems Administration/Managementunmatched
  • Systems Analysisunmatched
  • Team Playerunmatched
  • Test Automationunmatched
  • Test Harnessunmatched
  • Training Data Setsunmatched
  • United States Department of Defense (DoD)unmatched
  • User Interface/Experience (UI/UX)unmatched
  • Writing Skillsunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder