Staff AI/ML Engineer (Large Language Model)

Caci Inc
King of Prussia, PA, United States
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$108,400.0 - $227,500.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Artificial Neural Networks Computer Vision C Sharp (Programming Language) C++ (Programming Language) Nvidia CUDA Linux
+28 more
Python (Programming Language) Machine Learning NumPy Software Deployment Software Engineering Virtualization Technology Reinforcement Learning Rust (Programming Language) ReactJS Large Language Models Prompt Engineering Gitlab Git Pandas Containerization Scikit Learn Kubernetes Information Technology HuggingFace Rancher Bitbucket Machine Learning Operations Functional Programming GPT Software Version Control Software Library Docker Golang

Job description

Lead and mentor a multidisciplined team consisting of developers and researchers to implement machine learning algorithms to solve a broad set of challenges for our various customers

  • Lead and mentor a multidisciplinary team delivering advanced AI/ML solutions
  • Apply LLMs to complex domain-specific problems and operational workflows
  • Adapt and fine-tune foundation models for specialized use cases
  • Design and implement retrieval-augmented generation (RAG) systems and semantic search architectures
  • Build production-grade LLM applications and agentic systems
  • Deploy scalable AI solutions across cloud, on-prem, and hybrid environments
  • Analyze large, multi-modal datasets to extract meaningful features and actionable insights
  • Translate emerging research into applied, mission-relevant capabilities
  • Communicate technical strategy, status, and risks to internal and external leadership

Requirements

Required:

  • Active TS/SCI U.S. Government Security Clearance
  • B.S. in machine learning, computer science, mathematics, or related fields
  • 8+ years of experience, preferably in software development or as a data scientist with 2+ years of building LLM applications using some of the following:
  • Fine-tuning foundational models
  • Steering Techniques (e.g Sparse auto encoders, representation tuning)
  • Building adapters to use foundational models (e.g. PEFT, llama factory)
  • Prompt engineering techniques / Inference time techniques (e.g. chain of thought, tree of thoughts, etc.)
  • Using Retrieval Augmented Generation techniques to populate and query vector databases (e.g. Weaviate, pinecone, pgvector)
  • Using LLM Frameworks (e.g. LangChain, DSPy, Microsoft Agent Framework)
  • Using AI APIs ( e.g AWS Bedrock, OpenAI)
  • Using LLM deployment frameworks (eg llama.cpp, vllm, tgi)
  • Developing UIs with ReAct
  • Experience leading an interdisciplinary team of researchers and software developers and working with a program manager to define project scope and schedule to ensure we meet project milestones as defined by our customers
  • Experience with Python and data science / machine learning libraries (e.g. NumPy, Pandas, Polars, scikit-learn, etc.)
  • Experience contributing on a team using version control (e.g. git, GitLab, Bitbucket)

Desired:

  • M.S. or PhD in machine learning, computer science, mathematics, or related fields
  • Experience leading an interdisciplinary team of researchers and software developers
  • Experience with any of the following:
  • Large Language Models and experience identifying ways to incorporate them into new domains and applications
  • Applying Transformer-based architectures to domains in other areas outside of Natural Language Processing (NLP) such as computer vision
  • Natural Language Processing algorithms such as BERT
  • Reinforcement learning and familiarity with Gymnasium Gym, OpenEnv, TorchRL, RLlib, and Stable Baselines
  • Applying clustering algorithms and/or deep neural networks to real life problems
  • Implementing tracking and pattern-of-life algorithms
  • Experience with GenAI Ops techniques (e.g. LLM-as-a-judge) and frameworks (e.g. LangFuse, MLFlow, Arize Phoenix)
  • Experience with Machine Learning libraries and frameworks such as HuggingFace and LangChain
  • Experience with Linux
  • Experience with CUDA and Python libraries such as CuPy, Numba, CuSignal, CuDF, etc.
  • Familiarity with using AWS cloud computing resources such as EC2, S3, Lambda, etc.
  • Experience with any of the following additional languages: Java, C++, Rust, Go, and/or C#
  • Experience in application deployment, virtualization, and containerization (e.g. Podman, Docker, Kubernetes, Rancher)
  • Experience shaping and writing proposals
  • Adjudicated Counter Intelligence or Full Scope Polygraph

Benefits & conditions

There are a host of factors that can influence final salary including, but not limited to, geographic location, Federal Government contract labor categories and contract wage rates, relevant prior work experience, specific skills and competencies, education, and certifications. Our employees value the flexibility at CACI that allows them to balance quality work and their personal lives. We offer competitive compensation, benefits and learning and development opportunities. Our broad and competitive mix of benefits options is designed to support and protect employees and their families. At CACI, you will receive comprehensive benefits such as; healthcare, wellness, financial, retirement, family support, continuing education, and time off benefits.

The proposed salary range for this position is: $108,400 - 227,500 USD

About the company

At CACI, we place character and innovation at the center of everything we do. As a valued team member, you’ll be part of a high-performing group dedicated to our customer’s missions and driven by a higher purpose - to ensure the safety of our nation.

An environment of trust.

CACI values the unique contributions that every employee brings to our company and our customers - every day. You’ll have the autonomy to take the time you need through a unique flexible time off benefit and have access to robust learning resources to make your ambitions a reality.

A focus on continuous growth.

Together, we will advance our nation’s most critical missions, build on our lengthy track record of business success, and find opportunities to break new ground - in your career and in our legacy.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Applying large language models to infrastructure tasks

Alfonso Sandoval Rosas Alfonso Sandoval Rosas · Europe 2026 Virtual

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

4:20 min

Combating human workforce shortages with specialized language models

Markus Hacker Markus Hacker +3 · WWC 2024

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · WWC 2025

Videos

See all

Related articles

See all