> Markdown version of [/jobs/ext/2716785-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2716785-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Engineer - **Company:** Bayesian Health, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Computing Platforms, Automation of Tests, Backup Devices, Cyber Security, Relational Databases, DevOps, Disaster Recovery, Failover, Fault Tolerance, Github, PostgreSQL, MySQL, Performance Tuning, Reliability Engineering, Datadog, Circleci, Delivery Pipeline, Grafana, Kubernetes, Terraform - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/infrastructure-engineer-bayesian-health-inc-8137862 ## About the Role * 5+ years of experience building and operating production cloud infrastructure on AWS as a DevOps, Infrastructure, Site Reliability Engineer, or similar role. * Proficient with Kubernetes, preferably with EKS, including cluster bootstrapping and day-2 ops. * Strong operational knowledge of relational databases such as PostgreSQL/MySQL (backups, failover, performance tuning). * Deep expertise in Terraform (or equivalent IaC) and an eye for building clean and scalable modules. * Familiarity with observability tools, particularly the Datadog * Experience building infrastructure with sensitive data that contains PHI/PII. * Knowledge of CI/CD pipelines, preferably with CircleCI * Excellent communication skills and a proven ability to collaborate with cross-functional teams (e.g., engineering, data science) to translate requirements into robust technical solutions. * Experience handling ambiguity and uncertainty in a startup., * Experience with using AI agents to optimize infrastructure management or DevOps workflows. * Experience with disaster recovery or business continuity plans. * Experience with multi account, multi cluster topologies. * Experience building systems in healthcare, life sciences, or similarly regulated industries * Chaos engineering or game-day facilitation. * Experience implementing and maintaining a GitOps framework. ## Description As an Infrastructure Engineer, you will build and maintain the networking and infrastructure for the Bayesian platform and develop CI/CD pipelines to enable other team members such as software engineers, data scientists, etc. to accelerate their development. This role is crucial to drive expansion of our clinical AI/ML module offerings, health system enterprise-wide implementations, and revenue growth., * Design cost-optimized, fault-tolerant infrastructure for scale: Propose enhancement to our infrastructure design to enable us to expand our client base and deploy new products on our platform while managing cloud costs and ensuring reliability. * Streamline development and deployment: Define a branching and promotion strategy that allow us to comply with the regulatory change control process. Build and maintain CI/CD pipelines using GitHub for automated testing and deployment. * Establish and evangelize infrastructure best practices: Create infrastructure guidelines and templates such as Terraform modules, and educate team members in leveraging them. * Infrastructure support and maintenance: Continuous monitoring of system performance and reliability, and apply software upgrades accordingly. Collaborate with other team members in troubleshooting infrastructure issues and optimize performance. * Secure infrastructure: Partner with SecOps engineer to implement security best practices complying with HIPAA, HITRUST, FDA, and client requirements. * AI Ops Platform Architecture: Architect and build a secure, internal AI Ops platform to safely host and manage AI/ML agents for infrastructure and DevOps optimization. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Coding for Good: Achieving social change with an app](https://www.wearedevelopers.com/videos/1645-coding-for-good-achieving-social-change-with-an-app) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence)