Server Lifecycle Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
- Execute server decommissioning workflows including wiping, powering off, and coordinating with Infrastructure Operations for de-racking as appropriate
- Manage server repurposing by coordinating with Sales, Customer Success, and Solutions Engineering to assign servers to new workloads and customers
- Perform firmware updates on repurposed and transitioning servers in coordination with assignment changes
- Coordinate and execute server migrations including OS upgrades and platform transitions
- Manage inactive server inventory, identifying servers that need lifecycle action and routing hardware issues to the Hardware/Infrastructure team
- Handle deeper lifecycle management tasks beyond initial server onboarding, including capacity reclamation and fleet optimization
- Develop and refine runbooks, standard operating procedures, and documentation for all lifecycle workflows
- Maintain accurate records of decommissioning, repurposing, and migration activities including validation and handoff documentation
- Collaborate across teams to ensure clean handoffs and clear ownership throughout the server lifecycle
- Respond to lifecycle-related incidents and contribute to post-incident reviews when decommissioning or migration work impacts production
Requirements
- 3+ years of hands-on Linux system administration experience in a production environment
-
Strong working knowledge of bare-metal server hardware, including BIOS, BMC/IPMI, and common server troubleshooting workflows
-
Familiarity with server decommissioning processes including secure disk wiping and power management procedures
-
Comfortable with shell scripting or at least one general-purpose scripting language such as Python for automation of lifecycle tasks
-
Excellent organizational and coordination skills, with experience working across Sales, Customer Success, and Infrastructure Operations teams
-
Demonstrated experience operating in a 24/7 production environment with clear written communication
-
Strong diagnostic instincts with the ability to distinguish hardware faults from platform faults and route issues appropriately
- Prior experience at a cloud provider, hosting company, or large-scale infrastructure operator is a strong plus
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fully Remote Software Engineer Jobs
Why Upskilling And Reskilling is Important For Developers
Now is the time for industrialized software development
Is Software Engineering Over-Saturated?