Technical Architects

Mag Llc.
New York, NY, United States
10 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$124,800.0 - $270,400.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Systems Engineering Automation of Tests Microsoft Azure Cloud Computing Cloud Engineering System Configuration Software Debugging DevOps Disaster Recovery
+15 more
Distributed Systems Identity and Access Management Network Security Reliability Engineering Requirements Management Software Engineering Software Requirements Analysis Data Logging Google Cloud Autoscaling Multi-Cloud Kubernetes Infrastructure Automation Frameworks Restful APIs Terraform

Job description

Cloud Architecture & Distributed Systems

  • Design realistic infrastructure scenarios spanning distributed systems, networking, security, scalability, and reliability
  • Evaluate architectures involving scalable APIs, queues, durable storage, autoscaling, and service coordination
  • Model partial failures, degraded services, and realistic production constraints
  • Assess trade-offs across performance, availability, security, and operational complexity
  • Define clear system requirements and measurable success criteria

Infrastructure Environments & Validation

  • Build reproducible, containerised technical environments
  • Create reference implementations and intentionally defective variants
  • Develop deterministic integration, load, security, failure-injection, deployment, and recovery tests
  • Validate infrastructure configuration, topology, and runtime behaviour
  • Support repeatable provisioning, execution, and teardown workflows

Security, Reliability & Operations

  • Design scenarios involving IAM, least privilege, private networking, and service-to-service security
  • Incorporate logging, metrics, tracing, SLOs, and operational telemetry
  • Evaluate rolling deployments, rollback strategies, disaster recovery, and resilience
  • Develop realistic troubleshooting and incident scenarios
  • Write automation or testing tools supporting environment setup and validation

AI Evaluation & Technical Review

  • Create reinforcement-learning environments for multi-step infrastructure reasoning
  • Build golden solutions and defective or adversarial variants
  • Review peer-created tasks for ambiguity, unrealistic assumptions, or validation gaps
  • Improve benchmark difficulty, reproducibility, and grading reliability
  • Maintain strong engineering standards across project deliverables

Requirements

We are sharing a specialised consulting opportunity for experienced Senior Technical Architects with strong expertise in cloud infrastructure, distributed systems, platform engineering, DevOps, SRE, production architecture, and resilience engineering to contribute to an advanced AI training and cloud-infrastructure evaluation project.

Selected professionals will design realistic cloud-infrastructure tasks and reinforcement-learning environments that test an AI system’s ability to design, deploy, secure, scale, troubleshoot, and recover production-grade systems. The work requires substantial hands-on production ownership, strong systems judgement, and the ability to create reproducible environments, deterministic tests, and rigorous reference solutions. No prior experience in AI is required., * Senior-level experience in technical architecture, cloud infrastructure, platform engineering, DevOps, systems engineering, or SRE

  • Demonstrated ownership of production infrastructure or a production platform
  • Strong knowledge of distributed systems, scalable APIs, queues, autoscaling, and durable storage
  • Practical IAM, private-networking, and service-security experience
  • Strong observability, SLO, deployment, rollback, and disaster-recovery experience
  • Ability to write infrastructure automation or testing tools
  • Strong debugging skills in containerised environments
  • Experience with Terraform or OpenTofu is advantageous
  • Experience with AWS, Azure, GCP, Kubernetes, or multi-cloud environments is valuable
  • Background in internal developer platforms, edge infrastructure, chaos engineering, fault injection, or resilience testing is beneficial
  • No prior AI-training or model-evaluation experience is required, Amazon Web Services (AWS), Application Programming Interface (API), Artificial Intelligence (AI), Autoscaling, Benchmarking, Cloud Architecture, Cloud Computing, Consulting, Debugging Skills, DevOps, Disaster Recovery, Distributed Computing, GCP (Good Clinical Practices), Identify Issues, Industry/Trade Analysis, Injections, Metrics, Microsoft Windows Azure, Network Security, Onboarding, Production Systems, Project Evaluation, Requirements Management, Software Engineering, Systems Engineering, Systems Scalability, Technical Consulting, Telemetry, Test Automation, Test Plan/Schedule, Test Tools, Testing, Topology, Writing Skills

Benefits & conditions

Engagement Details

  • Independent contractor engagement
  • Fully remote
  • Displayed compensation range: $60-$130/hour
  • Actual compensation structure is output-based, with payment made per task that meets project specifications
  • Task completion time may vary depending on experience and workflow
  • Minimum weekly submission requirements apply; the source does not specify the exact number of required tasks
  • Work will involve cloud architecture, distributed systems, networking, IAM, observability, resilience testing, deployment workflows, and AI evaluation environments
  • Roles are typically filled within approximately 48 hours
  • Selected experts are expected to begin initial tasks within approximately 24-48 hours after onboarding
  • Project scope, technical environments, benchmark requirements, and evaluation standards may evolve depending on project needs
  • Work must be completed without using confidential or proprietary information belonging to any employer, client, cloud provider, software organisation, infrastructure environment, or other third party

About the Platform

This opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Transitioning from traditional software development to artificial intelligence consulting

Patrick Schnell Patrick Schnell · Coffee With Developers

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:15 min

Deploying local container pods to Kubernetes clusters

Stevan Le Meur Stevan Le Meur · World Congress 2024

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

3:27 min

Defining DevOps through its historical origins and foundational texts

Sonal Patil · LIVE

Videos

See all

Related articles

See all