AI/ML Engineer

CCS, LLC
Plano, TX, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Amazon S3 Cloud Engineering Continuous Integration Python (Programming Language) Data Ingestion Large Language Models Prompt Engineering Cloudformation Git Flow
+7 more
HuggingFace Api Design Terraform Data Pipelines Serverless Computing Docker Microservices

Requirements

Bachelor’s Degree 6+ years cloud architecture experience 3+ years building production GenAI/LLM systems on AWS. Strong Python and AWS expertise, including Lambda, ECS/EKS, S3, SageMaker, Docker and Kubernetes. Production experience with vector databases and designing ingestion + embedding pipelines for both batch and streaming workloads. Hands-on with prompt design, evaluation, LLM orchestration, and RAG implementation patterns. Experience deploying and operating model- serving or MCP - like server infrastructure (selfhosted or managed). Proficient with IaC and delivery tooling, including Terraform/CloudFormation, GitOps, and CI pipelines. Experience with model-serving infrastructure, such as Amazon SageMaker, NVIDIA Triton, Ray Serve, or similar platforms. Hands-on experience with GenAI libraries and frameworks, including LangChain, LlamaIndex, Hugging Face, and OpenAI APIs. Deep operational expertise with vector databases, such as Pinecone, Milvus, Weaviate, or Qdrant. AWS Solutions Architect, AWS DevOps Engineer, or equivalent industry certifications. Responsibilities Cloud Architecture & Infrastructure, Design scalable, secure AWS architectures LLM & GenAI Platforms, Lead integration of API-based and self-hosted LLMs, implement RAG solutions Prompting & Evaluation, Develop prompt engineering strategies, reusable templates, and evaluation frameworks Vector Databases & Retrieval Pipelines, Implement and maintain vector stores (OpenSearch, Pinecone, Milvus, Qdrant) Data Ingestion & Processing Pipelines Microservices & Serverless Systems Python Development & AI Tooling Security, Governance & Cross-Functional Leadership

Benefits & conditions

Bonus based on performance Dental insurance Health insurance Vision insurance

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.wayup.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

5:02 min

Mapping Git flow branches to application tester segments

Majid Hajian · LIVE

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all