> Markdown version of [/jobs/ext/2862568-platform-engineer](https://www.wearedevelopers.com/jobs/ext/2862568-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Platform Engineer - **Company:** Perry Weather - **Location:** Dallas, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Bash Shell, Databases, Continuous Integration, Domain Name System (DNS), Github, Python (Programming Language), Key Management, Policy as Code, Scripting, Performance Testing, Delivery Pipeline, Large Language Models, Kubernetes, Api Gateway, Terraform, Code Restructuring - **Published:** September 12, 2026 - **Apply:** https://startup.jobs/platform-engineer-perry-weather-9991757 ## About the Role * 4+ years in platform, infrastructure, DevOps, or backend engineering, with real ownership of production infrastructure * Terraform you've written and maintained. You've reviewed a plan and caught something you didn't want applied, and you know how state drift happens * Kubernetes in production. You can take a failing workload, trace it through configuration, networking, and resource limits, and explain what went wrong to the engineer who owns the service * CI/CD you've built and debugged, ideally GitHub Actions. You've fixed a pipeline engineers had stopped trusting, and you can separate a failing test from failing infrastructure * Working depth in a major cloud, including managed databases, networking, identity, and secrets. We're primarily on Azure, and AWS or GCP experience transfers * Scripting in Python, Go, or Bash good enough to automate a manual process end to end and leave it maintainable for someone else * Observability work you've done yourself: instrumentation you added, a dashboard other people used, and an alert you tuned because it fired too often * On-call experience for production systems, including at least one incident you drove to resolution and wrote up afterward * Product instincts about internal tooling. You've built something for other engineers, watched them use it, and changed it based on what you saw * Fluency with AI-native engineering tools and agentic workflows (e.g., terminal-native coding agents, LLM-assisted code refactoring and generation) to multiply technical output and speed up development cycles * Least-privilege credential design for automated systems: CI service accounts, scoped tokens, short-lived credentials, audit trails. Agents are the newest consumers of that work, and the principles carry over * A view on how automated changes should be reviewed. More code and configuration now arrive generated, which puts weight on the checks that run before a merge or an apply, and you should have opinions about what those checks need to cover A strong plus is experience with policy-as-code, cloud cost management or FinOps practices, API gateway operations, DNS and CDN configuration, Helm chart authoring, managed time-series or high-volume databases, load and performance testing, running AI or ML workloads on shared infrastructure, and exposure to fleets of connected hardware. ## Description We're looking for a Platform Engineer to help build and run the infrastructure the rest of engineering depends on. You'll partner closely with our Lead Platform Engineer, designing and delivering the work together, and own real surface area from the start: infrastructure as code, our Kubernetes environments, delivery pipelines, observability, and the guardrails that keep all of it safe to change. Engineering is your customer. Success in this role looks like fast service setup, uneventful deploys, and engineers who can diagnose a failing workload without asking for help. Engineering here also works with AI coding agents daily, which adds to the platform's job: checks that catch generated infrastructure errors before they apply, scoped credentials and audit trails for automated actors, and cost reporting that accounts for agent usage. You'll use agents in your own work and build these controls for everyone else. What You'll Own * Infrastructure as code. Write and maintain the Terraform that defines our cloud footprint, review changes for blast radius, and keep our modules usable by engineers outside the platform team. * Kubernetes and workloads. Run our clusters and the Helm-deployed services on them, including resource tuning, scaling behavior, and the failure modes that only show up under load. * Delivery pipelines. Build and maintain CI/CD in GitHub Actions so that builds, tests, and deploys are fast and dependable enough that engineers rely on them. * Observability and reliability. Improve the instrumentation, dashboards, and alerting that tell us how the platform is behaving, and make alerts specific enough that people act on them. * Cost and capacity. Track where our cloud spend goes, find the waste, and add controls that catch expensive changes before they ship. * Guardrails. Secrets management, least-privilege access, dependency and image scanning, and policy checks that catch mistakes during review. * AI and automated actors. Coding agents and automation are regular consumers of our platform. You'll give them scoped identities and audit trails, add the checks that catch generated infrastructure mistakes before they apply, and track what their usage costs alongside the rest of our spend. * On-call and incident support. Share the on-call rotation, respond when the platform degrades, and make the change that prevents a repeat. ## Related Videos - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [Platform Engineering vs. DevOps Why not both?](https://www.wearedevelopers.com/videos/885-platform-engineering-vs-devops-why-not-both) - [Implementing Feature Environments with AWS and Terraform](https://www.wearedevelopers.com/videos/531-implementing-feature-environments-with-aws-and-terraform) - [Empowering Developer Innovation - Balancing Speed, Security, and Scale](https://www.wearedevelopers.com/videos/1689-empowering-developer-innovation-balancing-speed-security-and-scale) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)