AI/ML Platform Engineer

SURGE INFOTECH LLC
Alexandria, VA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Distributed Systems Amazon DynamoDB Python (Programming Language) Machine Learning Natural Language Processing Tensorflow
+21 more
Azure Machine Learning Systems Integration Cloud Platform System Pytorch Large Language Models Prompt Engineering Deep Learning Keras Fastapi Build Management Kubernetes Information Technology AWS Fargate Build Tools Machine Learning Operations Functional Programming Api Gateway Software Coding Terraform Docker Microservices

Job description

  • Build and deploy productiongrade ML/AI pipelines and services
  • Develop LLMpowered and NLPdriven applications
  • Write, optimize, and maintain highquality Pythonbased ML code
  • Implement scalable infrastructure using Terraform, AWS, and Kubernetes
  • Build FastAPIbased inference services and cloud APIs
  • Collaborate with crossfunctional engineering teams to deliver highimpact systems
  • Troubleshoot, optimize, and own systems endtoend as a handson engineer

Preferred Skills

  • Experience with distributed systems and microservices
  • Strong understanding of ML model lifecycle, deployment patterns, and operational monitoring

Requirements

Do you have experience in System deployment?, We are seeking a hands-on Senior AI/ML Platform Engineer with 10+ years of IT experience and a strong track record of building, deploying, and operationalizing AI/ML systems. The ideal candidate is a doer who excels in implementing scalable, production-grade AI/ML solutions across cloud environments., * 10+ years of IT/engineering experience

  • 3+ years of handson AI/ML development experience
  • 4+ years working directly with AWS services (Lambda, EC2, S3, DynamoDB, IoT Core, API Gateway, Fargate/ECS)
  • Proven experience deploying ML systems into production environments
  • Strong coding skills and ability to build systems endtoend, * Deep Learning frameworks: TensorFlow, PyTorch, Keras
  • LLMs, prompt engineering, NLP pipelines
  • Python and Java as primary languages; strong engineering fundamentals
  • FastAPI and microservices for ML inference
  • InfrastructureasCode (Terraform)
  • Kubernetes and Docker for scalable ML workloads
  • Distributed/cloud systems design with AWS
  • Edgetocloud system integration experience
  • Handson build experience (not just design/architecture)

About the company

We are an IT Consulting Company specializing in transformative solutions. We provide comprehensive services including Enterprise Architecture, Application Modernization, and Talent Management. With a focus on innovation and efficiency, we empower businesses to navigate the complexities of the digital landscape. Our dedicated team of experts collaborates closely with clients to optimize their IT infrastructure, streamline applications, and cultivate a high-performing workforce. Through cutting-edge technologies and strategic insights, we drive sustainable growth and enhance operational excellence for organizations of all sizes.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

1:18 min

Converting existing Keras models to TensorFlow format

Håkan Silfvernagel · LIVE

1:42 min

Introduction to the fast API web framework

Sebastián Ramírez · WWC 2022

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all