Senior Data Engineer / AI ML Engineer with Python, AI/ML & LLMs

Ampcus Inc
Reston, VA, United States
2 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$77,600.0 - $176,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon Web Services Microsoft Azure Cloud Computing Cloud Storage Cyber Security Databases Information Engineering Data Infrastructure Extract Transform Load (ETL)
+32 more
Data Systems Database Schema Python (Programming Language) NoSQL NumPy Software Tools Tensorflow SQL Databases Unstructured Data Google Cloud Data Ingestion Pytorch Delivery Pipeline Large Language Models Backend Git Pandas Containerization Pyspark Scikit Learn Kubernetes Information Technology Apache Flink HuggingFace Apache Kafka Machine Learning Operations Api Design Stream Processing Artificial Intelligence Markup Language (AIML) Software Version Control Data Pipelines Docker

Job description

  • We are seeking a highly skilled and motivated Data Engineer to join our growing AI/ML team. This role is ideal for someone passionate about building scalable data pipelines, enabling machine learning workflows, and integrating cutting-edge Large Language Models (LLMs) into production systems.
  • You will work closely with data scientists, ML engineers, and software developers to design and implement robust data infrastructure that powers intelligent applications., * Design, build, and maintain scalable and efficient ETL/ELT pipelines using Python and modern data engineering tools.
  • Collaborate with AI/ML teams to support model training, evaluation, and deployment workflows.
  • Develop and optimize data schemas, storage solutions, and APIs for structured and unstructured data.
  • Integrate and fine-tune LLMs (e.g., OpenAI, Hugging Face Transformers) for various business use cases.
  • Ensure data quality, governance, and compliance across all data systems.
  • Monitor and troubleshoot data workflows and model performance in production.
  • Automate data ingestion from diverse sources including APIs, databases, and cloud storage.
  • Contribute to the development of internal tools and libraries for ML experimentation and deployment.

Preferred Skills:

  • Experience with MLOps tools (MLflow, Airflow, Kubeflow).
  • Understanding of data privacy and security best practices.
  • Exposure to vector databases (e.g., Pinecone, FAISS, Weaviate).
  • Experience with real-time data processing (Kafka, Flink)., Cybersecurity Data Scientist The Opportunity: As a Cybersecurity Data Scientist, you will operate as a hands-on technical contributor and applied research leader responsible fo…
  • 1 month ago +

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Data Engineering, or related field.
  • 3 years of experience in data engineering or backend development.
  • Strong proficiency in Python and libraries such as Pandas, NumPy, PySpark, etc.
  • Experience with AI/ML frameworks (e.g., TensorFlow, PyTorch, Scikit-learn).
  • Hands-on experience with LLMs and NLP tools (e.g., LangChain, Hugging Face, OpenAI API).
  • Proficiency in SQL and working with relational and NoSQL databases.
  • Familiarity with cloud platforms (AWS, GCP, Azure) and containerization (Docker, Kubernetes).
  • Knowledge of CI/CD pipelines and version control (Git).

About the company

Ampcus Inc. is a certified global provider of a broad range of Technology and Business consulting services. We are in search of a highly motivated candidate to join our talented Team.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all