Software Engineer (Site Reliability

Heb Grocery Company, LP
Austin, TX, United States
19 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Shift work
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Agile Methodology Amazon Web Services Application Performance Management Computing Platforms Systems Engineering Confluence JIRA Bash Shell Cloud Computing Data Architecture
+30 more
Data Structures Data Visualization Software Design Patterns Linux Distributed Systems Github Groovy Iterative and Incremental Development IntelliJ IDEA Python (Programming Language) PostgreSQL Reliability Engineering Ruby Service-Oriented Architecture Systems Architecture Datadog Scripting Google Cloud System Availability Grafana Gitlab Build Management Kubernetes Information Technology Slack Graphql Terraform New Relic (SaaS) Docker Microservices

Job description

Job Summary: As a Senior Software Engineer-Site Reliability on the Digital Fullfilment team, you’ll deliver complex code solutions. You’ll support the build and deployment pipeline and when necessary, diagnose / solve production support or on-call issues. You’ll contribute to overall system design, architecture, security, scalability, reliability, application performance and provide end-to-end support., * Develops and maintains tooling used for environment monitoring and task automation

  • Engages in and improves whole lifecycle of services, including inception and design, deployment, operation, and refinement
  • Analyzes and establishes efficient configurations for software and servers, DB connections / indexes, drivers, etc.
  • Collaborates with development teams to design service architectures, software platforms and frameworks, capacity planning, release plans and launch reviews
  • Monitors internal and vendor service level objectives (SLOs) and agreements (SLAs); identifies / resolves SLO / SLA gaps
  • Serves as technical subject matter expert (SME) for cross-functional engineering Teams; assists with / troubleshoots systems-related issues and maintenance

The responsibilities and essential functions outlined above describe the general nature and level of work assigned to this position. This is not an exhaustive list of all duties, responsibilities, and skills required. Duties and responsibilities may be modified at any time based on business needs. Employees may be required to perform other job-related tasks as requested by their supervisor, subject to reasonable accommodations.

Requirements

  • 5+ years experience designing, analyzing, developing, or troubleshooting distributed systems
  • 3+ years of SRE experience managing Google Kubernetes Engine (preferred), K8s, or AWS environments
  • 2+ years of Java (Spring) programming experience preferred
  • 3+ years of using Terraform to maintain cloud infrastructure
  • 3+ years of CI Pipeline experience with either Gitlab Pipelines, or GitHub Actions
  • Experience with tools such as Gitlab, JIRA, Slack, Confluence and Intellij is preferred
  • Experience with microservices architecture patterns
  • Experience working with PostgreSQL, Kubernetes, Docker, Linux, Google Cloud Platform, Terraform, and APIs using REST and GraphQL
  • Experience working with monitoring and visualization tools such as Datadog, Grafana, or New Relic
  • Strong proficiency with scripting languages such as Python, Ruby, Groovy, Bash
  • Proven track record of researching, understanding, and effectively applying Scalability and High Availability principles

Knowledge/Skills/Abilities:

  • Advanced knowledge in system and data architecture, data modeling, and design and capable of architecting and designing at the application or service level using well-accepted design patterns -
  • Able to review platform designs for strength of engineering solutions, namely performance, sustainability, and iterative development potential. -
  • Comprehensive knowledge of Computer Science fundamentals: data structures, algorithms, design patterns, system architecture and design patterns -
  • Advanced understanding of development methodologies and processes -
  • High degree of personal accountability to self and team for continued growth -
  • Adjust - Leverages Agile metrics to improve team performance and deliverables. Evaluates and adjusts resources, self, and team as necessary. -
  • Collaborate - Ability to work on tasks which span multiple domains, requiring cross-team collaboration, which have a high impact on your project. -
  • Agility - Embraces risk, change, and helps team manage ambiguity within the team’s scope of work. -
  • Able to drive progress without having a complete picture and can articulate potential tradeoffs and prioritize when faced with ambiguity. -
  • Connect - Delivers clear, concise, effective messages across different levels; can tailor communication based on intended audience. -
  • Growth Mindset - Fosters a culture of mentoring and coaching across multiple technical teams and other stakeholders. -
  • Relate - Fosters a culture within their team where people are encouraged to share their opinions and contribute to discussions in a respectful manner, approach disagreement non-defensively with inquisitiveness, and use contradictory opinions as a basis for constructive, productive conversations. -

Education:

  • A Computer Science degree or comparable formal training, certification, or work experience -involving software / systems engineering

Physical Demands & Working Conditions:

  • Travel by car or plane with overnight stays
  • Work extended hours; sit for extended periods
  • Work rotating and on-call schedules, as needed

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

Videos

See all

Related articles

See all