Senior Technical Architect for AI Model Training

Johns Hopkins Applied Physics Laboratory
Washington, DC, United States
8 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$100,000.0 - $245,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Systems Engineering Automation of Tests Microsoft Azure Cloud Computing Software Debugging DevOps Disaster Recovery Distributed Systems Identity and Access Management Machine Learning
+14 more
Reliability Engineering Trusted Systems Reinforcement Learning Google Cloud Cloud Platform System Autoscaling Multi-Cloud Build Management Containerization Kubernetes Infrastructure Automation Frameworks Restful APIs Terraform Programming Languages

Job description

Apply senior platform engineering and production operations expertise to design realistic cloud infrastructure challenges that train and evaluate next-generation AI systems. You will create Reinforcement Learning environments that test an AI model’’s ability to design, deploy, secure, scale, troubleshoot, and recover production-grade cloud systems. No prior AI experience is required, the role prioritizes hands-on production ownership and domain expertise. Key Responsibilities

  • Create realistic cloud infrastructure tasks that cover distributed systems, networking, security, scalability, and reliability.
  • Design and build reproducible, containerized environments, including valid golden reference solutions and intentionally defective variants for evaluation.
  • Define measurable requirements across infrastructure configuration, deployed topology, and runtime behavior.
  • Develop deterministic integration, load, security, failure-injection, deployment, and recovery tests to validate model behavior.
  • Build Reinforcement Learning environments that exercise IAM, queues, durable storage, observability, rolling deployments, and disaster recovery scenarios.
  • Debug environments, document technical decisions, and review and improve tasks created by other experts., + $105,000-245,000 per year Description Are you eager to architect and deliver robust, scalable, and secure software systems that enable real operational decisions for warfighters? Are you excited to dire…

  • 2 days ago, + $105,000-245,000 per year Description Are you eager to architect and deliver robust, scalable, and secure systems that enable real operational decisions for warfighters? Are you excited to directly impa…

  • 2 days ago +

Requirements

  • Senior-level experience in technical architecture, cloud infrastructure, platform engineering, DevOps, systems engineering, or SRE, including personal ownership of a production platform.
  • Strong knowledge of distributed systems, scalable APIs, queues, autoscaling, durable storage, and partial-failure scenarios.
  • Practical experience with IAM, private networking, least-privilege access, and service-to-service security.
  • Experience with observability, measurable SLOs, rolling deployments, rollback strategies, and disaster recovery.
  • Ability to write infrastructure automation or testing tools and to debug containerized environments using a relevant programming language.

Preferred Qualifications

  • Experience with Terraform or OpenTofu.
  • Experience with AWS, Azure, GCP, Kubernetes, or multi-cloud infrastructure.
  • Experience building internal developer platforms, edge infrastructure, or shared platform services.
  • Familiarity with chaos engineering, fault injection, local cloud emulators, or resilience testing.
  • Experience creating technical evaluations, automated grading systems, or AI training/evaluation environments is helpful but not required.

Benefits & conditions

  • Output-based compensation, experts are paid per task that meets project specifications.
  • Time required to complete work varies by expert experience and workflow, minimum submission requirements apply.
  • Experts must submit a minimum number of tasks per week.
  • Assignments and volume depend on project availability and may vary over time.

Compensation

  • Pay range: 60 to 130 hourly.
  • Compensation is paid per completed task that meets the project specifications, rather than a fixed hourly guarantee.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · WWC 2022

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all