Senior SRE / Platform Engineer - AI Platform
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+12 more
Job description
I’m currently working with a leading organisation that is looking for an experienced Senior SRE / Platform Engineer to join their team on an initial contract basis.
This is an excellent opportunity for someone with a strong SRE, DevOps or Platform Engineering background who has experience supporting modern software platforms and, ideally, working with AI/LLM platforms such as OpenAI or Anthropic.
The role
You’ll be responsible for ensuring the organisation’s platforms are reliable, scalable, automated, secure and cost-effective.
The role sits across SRE, DevOps, software engineering and cloud infrastructure, with a strong focus on automation, CI/CD, observability and supporting AI-enabled platforms., * Supporting and maintaining production platforms and services in a BAU/SRE environment
- Building and managing CI/CD pipelines and deployment orchestration
- Implementing infrastructure and automation using Terraform
- Working with APIs and API gateways, including technologies such as Kong
- Supporting and administering AI platforms, including OpenAI, Anthropic or similar services
- Managing and monitoring API/token usage
- Implementing observability and monitoring using tools such as Datadog
- Working with message queuing and distributed systems
- Supporting software deployment and understanding application architecture
- Reviewing code and maintaining software repositories
- Implementing and maintaining zone-based architecture and networking
- Improving platform reliability, scalability and performance
- Managing infrastructure capacity and optimising cloud/platform costs
- Working closely with Software Engineering, DevOps, Security and Architecture teams
Requirements
We’re looking for someone with a strong background in SRE, Platform Engineering, DevOps or DevSecOps, with experience across:
- SRE / Site Reliability Engineering
- DevOps / DevSecOps practices
- Terraform / Infrastructure as Code
- API configuration and API gateways
- Cloud infrastructure and networking
- Automation and orchestration
- Observability tooling - ideally Datadog
- Software engineering/deployment principles
- Repository and code management
- Infrastructure and application architecture
- Cost, capacity and performance optimisation
Experience with OpenAI, Anthropic, LLMs or other AI platforms would be highly advantageous.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Navigating the AI Shift
Where To Find Software Engineering Jobs
Find a Developer Job: 12 Best Job Sites For Developers
Dev Digest 121 - AI goes offline