> Markdown version of [/jobs/ext/3077924-cloud-engineer](https://www.wearedevelopers.com/jobs/ext/3077924-cloud-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloud Engineer - **Company:** Princeton University - **Location:** Princeton, NJ, United States (Remote available) - **Experience:** Expert - **Salary:** $75,000.0 - $110,000.0 - **Contract:** Temporary contract - **Skills:** Application Programming Interfaces (APIs), Software System Penetration Testing, Audit Trail, Microsoft Azure, Big Data, Cloud Computing, Cloud Engineering, Cloud Storage, Continuous Delivery, Information Engineering, Data Governance, Data Infrastructure, Data Security, DevOps, Github, Identity and Access Management, Image Management, Python (Programming Language), Key Management, Network Segmentation, Node.Js, Role-Based Access Control, Azure Active Directory, Tensorflow, Prometheus, Runbook, Service Pack, Software Engineering, TypeScript, Management of Software Versions, Data Processing, Software Modules, ReactJS, Grafana, Backend, Containerization, Kubernetes, Information Technology, HuggingFace, Bicep, Machine Learning Operations, Front End Software Development, Software Coding, Terraform, Data Pipelines, Dynatrace, Devsecops, Docker, Databricks, Vulnerability Analysis - **Published:** September 25, 2026 - **Apply:** https://www.jofdav.com/jobs/59864040-cloud-engineer ## About the Role This is a role for a senior, self-directing engineer who is equally comfortable designing architecture and writing code, and who takes satisfaction in building systems that are reliable, secure, and well-understood by the people who depend on them. The right candidate brings deep cloud expertise alongside strong software engineering fundamentals - someone who can own infrastructure end to end and contribute meaningfully to application development. This position is classified at the Senior Engineer level, corresponding to 5-8 years of relevant experience. The individual will, * 5-8 years of experience in cloud engineering, DevOps, or a software engineering role with significant infrastructure ownership. * Bachelor's degree in Computer Science, Engineering, or a related field or equivalent work experience. * Strong proficiency in Python; experience with at least one additional language (TypeScript/Node.js, Go, or equivalent). * Deep hands-on experience with Azure cloud services, including compute, networking, storage, identity, and managed services; familiarity with Azure CAF landing zones, subscription governance, and resource management at scale. * Proficiency with Terraform, including module development, remote state management, and PR-based workflow; Bicep familiarity a plus. * Experience designing and implementing CI/CD pipelines, preferably using GitHub Actions. * Production experience with Kubernetes / AKS - deploys, scaling, health management, upgrades, and cluster operations. * Solid Docker and container image management skills; experience building and maintaining containerized services in production. * Azure networking and security fundamentals, including Private Link, network rules, NSGs, and RBAC; comfort managing secrets hygiene across environments. * Azure cost management and FinOps awareness: budget alerts, cost/cluster policies, anomaly detection and response. * Comfort operating in and improving existing codebases with limited live handoff - able to orient independently, read unfamiliar infrastructure, and contribute quickly without extensive documentation. * Experience with Databricks or equivalent large-scale data processing platforms. * Solid understanding of data security principles, IAM patterns, and compliance frameworks (SOC 2, HIPAA, ISO 27001, or equivalent). * Experience operating a self-hosted observability stack - specifically Grafana, Loki, and Prometheus - including patching, upgrades, and dashboard maintenance; equivalent stack experience considered. * Ability to work independently on complex, ambiguous problems and communicate technical decisions clearly to non-technical stakeholders. * Strong written communication skills; comfortable producing architecture documentation, runbooks, and technical specifications. Preferred * Experience supporting research computing or academic data infrastructure environments. * Familiarity with ML infrastructure tooling (MLflow, Hugging Face Hub, model serving frameworks). * Experience with IRB-compliant research data environments or sensitive data handling at scale. * Frontend development experience (React or equivalent) - useful on a small cross-functional team. * Relevant certifications: Azure Administrator (AZ-104), Azure Solutions Architect (AZ-305), Azure DevOps Engineer (AZ-400), or equivalent., A combination of relevant work experience and education equivalent to 5-8 years of hands-on cloud engineering or software engineering experience, with a demonstrable record of owning and delivering complex infrastructure and software projects. A bachelor's degree in Computer Science, Engineering, or a related field is preferred but not required - equivalent professional experience will be considered. ## Description * Plan and execute work independently, applying sound judgment in the evaluation, selection, and adaptation of technical approaches across infrastructure, security, and software development. * Design and implement solutions with broad ownership - from initial architecture through deployment and ongoing operations - with supervisory input primarily at the level of objectives and critical decisions rather than day-to-day methods. * Devise new approaches to novel problems, drawing on extensive knowledge across cloud infrastructure, DevSecOps, and software engineering disciplines. * Serve as a technical resource for the broader team, contributing to architectural decisions and engineering standards., Cloud Infrastructure * Design, deploy, and maintain cloud infrastructure on Azure, with responsibility for performance, cost-effectiveness, and reliability across research and production environments. * Architect and manage Databricks workspaces, including compute cluster configuration, access controls, and cost optimization for large-scale data processing workflows. * Manage Azure networking, storage, identity (Azure AD / Entra ID), and resource governance across multiple environments. * Implement infrastructure-as-code using Terraform and/or Bicep; maintain version-controlled, reproducible infrastructure definitions including modules, remote state management, and PR-based workflow. * Deploy, operate, and maintain AKS clusters running containerized workloads - including containerized data crawlers - managing deploys, scaling, health monitoring, patching, and upgrades. * Administer Azure Blob Storage, including lifecycle policies, redundancy configuration, and access tier management. * Manage Azure networking and security, including Private Link, network rules, RBAC, and secrets hygiene across environments. * Own Azure cost management: budget alerts, cost/cluster policies, anomaly detection and response, and FinOps practices to keep infrastructure spend predictable and efficient. Software Development & DevOps * Design, build, and maintain backend services, APIs, and data pipelines using Python and/or TypeScript/Node.js. * Develop and maintain CI/CD pipelines using GitHub Actions, ensuring reliable and automated delivery of infrastructure and application changes. * Build and maintain internal tooling that improves the experience and efficiency of the research and operations teams. * Contribute to frontend integrations where needed; comfortable working across the stack on a small team. Data Engineering & ML Infrastructure * Develop and support data pipelines for ingesting, transforming, and serving large-scale behavioral and social media datasets to researchers. * Implement and maintain infrastructure for machine learning workflows, including model serving, experiment tracking, and compute resource management. * Support integration with ML frameworks and tools (e.g., MLflow, Hugging Face, or equivalent) within the managed environment. Security & Compliance * Implement and maintain security controls across all systems, including encryption at rest and in transit, identity and access management, network segmentation, and secrets management. * Design and operate environments meeting IRB, data governance, and institutional compliance requirements; ensure adherence to standards equivalent to SOC 2, HIPAA, or ISO 27001 as applicable. * Conduct regular security reviews, vulnerability assessments, and penetration test coordination; manage remediation tracking. * Implement audit logging, access controls, and data handling procedures for sensitive research data in compliance with IRB protocols and data use agreements. Observability & Operations * Operate, patch, and upgrade the self-hosted observability stack - Grafana (dashboards), Loki (log aggregation), and Prometheus (metrics) - including security patching and version upgrades; implement and maintain alerting, distributed tracing, and platform-wide monitoring. * Own incident response, root cause analysis, and operational reliability for production systems. * Develop and maintain runbooks, architecture documentation, and operational procedures. ## Related Videos - [Back(end) to the Future: Embracing the continuous Evolution of Infrastructure and Code](https://www.wearedevelopers.com/videos/440-back-end-to-the-future-embracing-the-continuous-evolution-of-infrastructure-and-code) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)