Software Engineer (Site Reliability
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+30 more
Job description
Job Summary: As a Senior Software Engineer-Site Reliability on the Digital Fullfilment team, you’ll deliver complex code solutions. You’ll support the build and deployment pipeline and when necessary, diagnose / solve production support or on-call issues. You’ll contribute to overall system design, architecture, security, scalability, reliability, application performance and provide end-to-end support., * Develops and maintains tooling used for environment monitoring and task automation
- Engages in and improves whole lifecycle of services, including inception and design, deployment, operation, and refinement
- Analyzes and establishes efficient configurations for software and servers, DB connections / indexes, drivers, etc.
- Collaborates with development teams to design service architectures, software platforms and frameworks, capacity planning, release plans and launch reviews
- Monitors internal and vendor service level objectives (SLOs) and agreements (SLAs); identifies / resolves SLO / SLA gaps
- Serves as technical subject matter expert (SME) for cross-functional engineering Teams; assists with / troubleshoots systems-related issues and maintenance
The responsibilities and essential functions outlined above describe the general nature and level of work assigned to this position. This is not an exhaustive list of all duties, responsibilities, and skills required. Duties and responsibilities may be modified at any time based on business needs. Employees may be required to perform other job-related tasks as requested by their supervisor, subject to reasonable accommodations.
Requirements
- 5+ years experience designing, analyzing, developing, or troubleshooting distributed systems
- 3+ years of SRE experience managing Google Kubernetes Engine (preferred), K8s, or AWS environments
- 2+ years of Java (Spring) programming experience preferred
- 3+ years of using Terraform to maintain cloud infrastructure
- 3+ years of CI Pipeline experience with either Gitlab Pipelines, or GitHub Actions
- Experience with tools such as Gitlab, JIRA, Slack, Confluence and Intellij is preferred
- Experience with microservices architecture patterns
- Experience working with PostgreSQL, Kubernetes, Docker, Linux, Google Cloud Platform, Terraform, and APIs using REST and GraphQL
- Experience working with monitoring and visualization tools such as Datadog, Grafana, or New Relic
- Strong proficiency with scripting languages such as Python, Ruby, Groovy, Bash
- Proven track record of researching, understanding, and effectively applying Scalability and High Availability principles
Knowledge/Skills/Abilities:
- Advanced knowledge in system and data architecture, data modeling, and design and capable of architecting and designing at the application or service level using well-accepted design patterns -
- Able to review platform designs for strength of engineering solutions, namely performance, sustainability, and iterative development potential. -
- Comprehensive knowledge of Computer Science fundamentals: data structures, algorithms, design patterns, system architecture and design patterns -
- Advanced understanding of development methodologies and processes -
- High degree of personal accountability to self and team for continued growth -
- Adjust - Leverages Agile metrics to improve team performance and deliverables. Evaluates and adjusts resources, self, and team as necessary. -
- Collaborate - Ability to work on tasks which span multiple domains, requiring cross-team collaboration, which have a high impact on your project. -
- Agility - Embraces risk, change, and helps team manage ambiguity within the team’s scope of work. -
- Able to drive progress without having a complete picture and can articulate potential tradeoffs and prioritize when faced with ambiguity. -
- Connect - Delivers clear, concise, effective messages across different levels; can tailor communication based on intended audience. -
- Growth Mindset - Fosters a culture of mentoring and coaching across multiple technical teams and other stakeholders. -
- Relate - Fosters a culture within their team where people are encouraged to share their opinions and contribute to discussions in a respectful manner, approach disagreement non-defensively with inquisitiveness, and use contradictory opinions as a basis for constructive, productive conversations. -
Education:
- A Computer Science degree or comparable formal training, certification, or work experience -involving software / systems engineering
Physical Demands & Working Conditions:
- Travel by car or plane with overnight stays
- Work extended hours; sit for extended periods
- Work rotating and on-call schedules, as needed
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Highest Paying Tech Companies for Developers
Best Countries for Software Engineers
Software Engineer Salary London