> Markdown version of [/jobs/ext/1137299-devops-engineer-gcp](https://www.wearedevelopers.com/jobs/ext/1137299-devops-engineer-gcp). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps Engineer (GCP) - **Company:** Knotch, Inc. - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $90,000.0 - $120,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Software as a Service, Cloud Computing, Configuration Management, Continuous Integration, Data as a Services, DevOps, Digital Architecture, Distributed Systems, Github, Monitoring of Systems, Identity and Access Management, Platform as a Service (PAAS), Performance Tuning, Reliability Engineering, Prometheus, Workflow Management Systems, Data Logging, Google Cloud, Cloud Platform System, Snowflake, Grafana, Reliability of Systems, Containerization, Kubernetes, Infrastructure Automation Frameworks, Deployment Automation, Terraform, Data Pipelines, Docker - **Published:** July 2, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=dc2a9cbd0edf34f9 ## About the Role You have a minimum 5+ years of experience in DevOps, Site Reliability Engineering, or Infrastructure Engineering roles within SaaS, PaaS, or cloud-native environments, with at least 3 years of exposure to GCP cloud environments., * Prior experience in growth-stage and/or startup environement scaling from $10M to $20M+ ARR with a lean team. * Strong experience with Google Cloud Provider (GCP) is required, including IAM, networking, and data services. * Hands-on experience with Infrastructure as Code tools such as Terraform. * Experience building and maintaining CI/CD pipelines (GitHub Actions, ArgoCD, or similar). * Solid experience with Kubernetes, Docker, and containerized environments. * Familiarity with deployment tools such as Helm. * Experience with monitoring and observability tools like Prometheus and Grafana. * Strong understanding of system reliability, scalability, and performance optimization. * Ability to work across multiple systems and priorities in a dynamic environment. * Strong documentation and communication skills, with attention to clarity and detail., * Supplementary experience supporting AI/ML or data-intensive workloads in production environments. * Familiarity with workflow orchestration or data pipeline tools. * Experience with cost optimization strategies for cloud infrastructure. * Exposure to security frameworks and compliance best practices. * Experience working with distributed or globally deployed systems. How to be Successful * Infrastructure ownership mindset: You take responsibility for system reliability, performance, and scalability - not just deployments. * Strong DevOps fundamentals: You understand CI/CD, IaC, observability, and containerization deeply and apply best practices consistently. * Systems thinking: You think holistically about how services interact, scale, and fail - and design accordingly. * Collaboration-first approach: You work closely with engineers across disciplines to enable velocity and reliability. * Pragmatic decision-making: You balance speed, cost, and reliability without over-engineering. * Operational excellence: You prioritize monitoring, alerting, and incident response as core parts of system design * Adaptability: You thrive in fast-moving environments and can context switch effectively across priorities. * Continuous improvement mindset: You proactively identify gaps and improve systems, processes, and tooling over time. ## Description As we evolve into an AI-native platform powered by agentic systems and large-scale data pipelines, the reliability, scalability, and observability of our infrastructure becomes mission-critical. We're not just deploying services - we're operating complex, production-grade AI systems that enterprise clients depend on every day. As a DevOps Engineer, you'll take part in building and scaling the foundation that powers everything we ship - from our core platform to our AI agents. You'll work across infrastructure, CI/CD, observability, and security to ensure our systems are fast, resilient, and cost-efficient. This role goes beyond "keeping the lights on". You'll help define how Knotch operates as an AI-first company, shaping infrastructure strategy, enabling developer velocity, and ensuring our systems scale alongside rapid product innovation. If you want to own infrastructure in a high-impact environment, work closely with engineering teams across the stack, and directly influence how production systems are built and operated, this is *that* role. If you've made it this far, please kindly input this code with your application: DEVOPS-ENG-2026. Responsibilities * Design, build, and maintain scalable, secure, and highly available infrastructure across pre-production and production environments. * Develop and manage CI/CD pipelines to enable fast, reliable, and repeatable deployments across multiple environments. * Own infrastructure as code (IaC) practices using tools like Terraform to ensure consistency and reproducibility. * Manage environment lifecycle (development, staging, production), including promotion workflows and configuration management. * Partner closely with Engineering, Data, and AI teams to support system performance, reliability, and scalability. * Implement and maintain monitoring, logging, and alerting systems to ensure high visibility into system health and performance. * Optimize infrastructure for cost, performance, and reliability, especially for compute- and data-intensive AI workloads. * Support Kubernetes-based deployments and container orchestration for distributed systems. * Contribute to security best practices across infrastructure, including IAM, networking, and application-level protections. * Create dashboards and reporting systems to provide visibility into system performance, uptime, and operational metrics. * Document architecture, operational processes, and infrastructure decisions to support knowledge sharing and onboarding. * Act as a DevOps/SRE partner across teams, helping troubleshoot issues and improve system reliability., We offer a unique opportunity to build and scale infrastructure that powers a truly AI-native platform. Our stack includes modern tools like Kubernetes, Terraform, AWS, Snowflake, and cutting-edge AI systems, all supported by a team actively building at the intersection of data, AI, and enterprise SaaS. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Why Attend a Developer Event in 2026?](https://www.wearedevelopers.com/magazine/688-why-attend-a-developer-event-in-2026) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)