> Markdown version of [/jobs/ext/482610-principal-cloud-engineer](https://www.wearedevelopers.com/jobs/ext/482610-principal-cloud-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Cloud Engineer - **Company:** Fmr LLC - **Location:** Hudson, NH, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Computer-Aided Design, Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Computing Platforms, Microsoft Azure, Cloud Engineering, Continuous Integration, Software Debugging, Distributed Systems, Identity and Access Management, Python (Programming Language), Machine Learning, Open Source Technology, Performance Tuning, Prometheus, Software Engineering, Datadog, System Availability, Grafana, Multi-Cloud, Cloudformation, Kubernetes, Information Technology, Terraform - **Published:** June 8, 2026 - **Apply:** https://find.jobs/jobs-near-me/principal-cloud-engineer-observability-hudson-new-hampshire/2808207742-2/ ## About the Role We are seeking a highly motivated Principal Cloud Engineer to join our Observability Platform team within Fidelity Architecture and Engineering. In this role, you will help design, build, and operate scalable, cloud-native observability solutions that support our most critical digital services., * 10+ years of software engineering or cloud engineering experience * Deep expertise in AWS cloud stack, especially: + EKS (Kubernetes on AWS) + Core services (IAM, EC2, networking, storage, etc.) * Strong experience working with Kubernetes and cloud-native ecosystems * Hands-on experience with observability tools/platforms (e.g., Prometheus, Grafana, Datadog, OpenTelemetry, etc.), including hosting and operational ownership * Proficiency in Python and/or Go for platform and tooling development * Strong understanding of CI/CD practices and tools * Experience working with Infrastructure as Code (Terraform, CloudFormation, etc.) * Familiarity with the Azure cloud stack and hybrid/multi-cloud environments * Working knowledge of OpenTelemetry (OTel) concepts and implementation Bonus Skills * Experience building or operating large-scale internal platforms * Exposure to eBPF-based observability or advanced profiling solutions * Experience integrating observability across multi-region / multi-cloud environments * Experience or interest in applying AI/ML techniques to observability (e.g., anomaly detection, predictive insights, intelligent alerting, or AIOps) * Active participation in or contributions to open-source projects * Strong background in performance optimization and distributed systems ## Description You will work in a collaborative, transparent, and innovation-driven environment where engineering excellence, continuous learning, and open-source contribution are core to how we operate. This is a high-impact role where your expertise will influence platform architecture, engineering practices, and the developer experience across the organization. What You'll Do * Lead the design and implementation of cloud-native observability platforms across AWS and Azure environments * Develop and operate highly scalable systems on AWS (EKS, core services) with strong focus on reliability, performance, and automation * Own the end-to-end lifecycle of observability tooling, including hosting, maintenance, scaling, and optimization * Drive adoption of OpenTelemetry (OTel) standards for metrics, traces, logs, and profiling * Build and enhance platform capabilities using Python and/or Go * Architect and optimize CI/CD pipelines enabling rapid, secure, and reliable deployments * Collaborate with cross-functional teams to improve system visibility, debugging capabilities, and performance insights * Define and promote best practices for cloud engineering, observability, and platform reliability * Mentor engineers and provide technical leadership across squads ## Related Videos - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)