Senior Site Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+6 more
Job description
As a Senior Platform Engineer, you are a champion for DevOps and SRE culture and industry best practice within Megaport. You will work alongside talented team members in multiple timezones ensuring that systems are secure, maintainable and available. External to the team you will be engaging with stakeholders in requirements analysis and demonstrations. Technically you will be very hands on. Continually evolving your skills through a mix of peer reviews and research. Ultimately your obsession is customer success and ensuring company goals are met., * Improving production reliability and system resilience within an SRE scoped team
- Championing high standards of work and industry best practices
- Communicating with teams and stakeholders at all stages
- Bringing fresh ideas to the table and encouraging others
- Diving into complex technical problems with a can-do attitude
- Working across numerous technologies in a fast-changing industry
- Participating in on-call rotation, incident response, and blameless post-incident reviews
- Writing code, handling alerts, improving solutions, and supporting others
- Playing a crucial role in the success of your company and team, Please see Part 2 of our Privacy Policy to see what information Megaport collects from job applicants, why, and how we store and use it. Note that you’re entitled to know what personal data of yours Megaport holds, to request updates, rectification, and in some circumstances restriction or deletion thereof if you object (you being entitled to withdraw your consent to our holding your information at any time). Please see Part 5 of our Privacy Policy for more details on this and how to contact Megaport’s data protection officer if you have any further privacy-related questions. Candidates who meet the selection criteria will be invited to attend an interview. Strictly no Recruitment Agencies.
Requirements
- 5+ years administering Linux systems and related infrastructure in production environments
- A collaborative SRE mindset, with familiarity around SLIs/SLOs/SLAs, error budgets, blast radius, and blameless postmortems
- A focus on automation, reducing toil, and preventing problem recurrence
- A track record of writing runbooks that work for the broader team, not just yourself
- Strong Kubernetes and broader ecosystem fundamentals
- Cloud infrastructure experience; AWS strongly preferred and bare-metal is a bonus
- Strong tool development - Bash, plus either Python or Go preferred, or similar
- Infrastructure-as-code tooling experience - Terraform preferred
- CI/CD and version control, GitHub preferred
- Database experience - one of Postgres, Cassandra, or ClickHouse preferred
- Experience operating a production observability stack (metrics, logs, traces), with an eye for signal over noise
- Comfortable working on live production infrastructure, with strong troubleshooting instincts and ownership of incident response
- A history of continual professional development
- A self-directed style suited to an async, globally distributed team, and comfortable picking up adjacent work when the situation calls for it
Benefits & conditions
- Flexible working environment - a remote-first culture with coworking options available.
- Generous leave plans - including 4 weeks of paid annual leave, parental leave, birthday leave, and a purchased annual leave program.
- Health and wellness support - through a wellness allowance and employee wellbeing initiatives.
- Comprehensive learning support - generous study and training allowance plus 5 days of paid study leave
- Creative, modern workspaces - designed to inspire when you’re not working remotely
- Motivated, inclusive team - work alongside industry experts and fresh talent
- Recognition programs - celebrate achievements with our Legend and Kudos awards
About the company
We’re not your typical tech company - and we don’t want to be. Megaport is the global leader in Network as a Service (NaaS), and has transformed the way businesses connect to the cloud, data centers, and each other. We’re publicly listed on the Australian Stock Exchange and partnered with the biggest names in tech like Amazon, Microsoft, Google, Oracle, IBM, and more. Headquartered in Brisbane with a crew of over 600 people spread across Asia-Pacific, Europe, and the Americas, our employees enjoy an environment that is collaborative, supportive, and (actually) fun. Our Team Culture We’re a team of problem solvers, pixel pushers, code slingers, and cloud fanatics. Culture is more than a poster on the wall here - collaboration beats hierarchy, curiosity fuels our growth, and everyone’s voice matters. We take our work seriously, but not ourselves. We work across time zones to execute on our global vision, trust each other to get things done, and never compromise our values for commercial gain. Most importantly, we place our customers at the center of everything we do.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fully Remote Software Engineer Jobs
Is Software Engineering Over-Saturated?
Find a Developer Job: 12 Best Job Sites For Developers
Dev Digest 120 - Apple and peers