Cloud Infrastructure Operations Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+16 more
Job description
We are looking for a Junior Cloud Infrastructure Operations Engineer to support a highly critical production environment. In this role, you will serve as an initial point of contact for cloud infrastructure issues reported by customers and application teams.
The ideal candidate will have hands-on exposure to AWS and Azure cloud infrastructure, Red Hat Linux administration, troubleshooting, and scripting/automation. You will work closely with senior infrastructure engineers, following established procedures to troubleshoot issues, gather information, and escalate complex problems when needed.
This is an excellent opportunity for someone who is technically hands-on, eager to learn, and looking to develop their cloud infrastructure and operations skills in a production environment.
Responsibilities
- Respond to incoming cloud infrastructure issues and perform initial troubleshooting and triage.
- Collect and analyze system logs, network traces, and filesystem information.
- Use Red Hat Linux skills to perform basic LVM and filesystem operations, analyze syslog, and troubleshoot OS-level processes and services.
- Work with application owners and customers to gather information needed to troubleshoot reported issues.
- Help determine whether issues are related to the application, operating system, or cloud infrastructure.
- Open and manage support cases with cloud providers such as AWS and Azure when required.
- Gather logs, evidence, and troubleshooting information and prepare documentation for escalation to senior L3 infrastructure engineers.
- Use AI assistants and AIOps tools to assist with log analysis, troubleshooting, and evidence gathering.
- Monitor AWS and Azure backup status and assist with VM/image-based restore operations, including AWS AMI and Azure image restores, following established procedures.
- Develop and maintain basic automation scripts using Bash, Python, and Ansible for data collection and routine troubleshooting tasks.
- Follow established troubleshooting, escalation, and operational procedures to maintain system availability and stability.
- Provide support during off-hours, nights, or weekends when critical infrastructure issues require immediate attention.
Requirements
Work Authorization: Candidates must be authorized to work in the U.S. without sponsorship., Hands-on exposure to AWS and Azure core services, including:
- Compute: EC2, Virtual Machines, AMI, Managed Images
- Networking: VPC/VNet, Subnets, Security Groups, NSGs, VPN, Direct Connect, ExpressRoute, Load Balancers
- Storage: S3, EBS, EFS, Azure Blob Storage, Managed Disks, Vault
- Security: IAM Policies, RBAC, Secret Keys
- Backup & Restore: Backup configuration, monitoring, and restore operations
Linux & OS-Level
Red Hat Linux administration experience is required, including:
- Basic LVM and filesystem administration
- Filesystem creation and expansion
- Managing system services and daemons
- OS-level logging and syslog analysis
- Basic troubleshooting of CPU, memory, and I/O issues
Scripting & Automation
Basic to intermediate experience with:
- Python
- Bash
- Ansible
AI for Infrastructure Support
- Familiarity with AI tools or AI assistants for troubleshooting and log analysis
- Interest in using AI/AIOps tools to improve infrastructure support and troubleshooting, * Hands-on Mindset: Willingness to troubleshoot issues and learn through real-world production support.
- Problem Solver: Ability to investigate issues, gather relevant information, and identify when additional assistance is needed.
- Team Player: Comfortable working with application teams, customers, and senior infrastructure engineers.
- Continuous Learner: Eager to learn new cloud technologies, automation techniques, and infrastructure best practices.
- Communication: Ability to clearly document troubleshooting steps, findings, and escalations., * Experience: 1 3 years of IT experience with hands-on exposure to cloud infrastructure administration and operational support.
- Cloud: Exposure to AWS and Azure environments.
- Linux: Hands-on Red Hat Linux administration experience.
- Scripting: Experience with Python, Bash, or Ansible.
- Education: Bachelor’s degree in Computer Science, Information Technology, or equivalent experience.
About the company
Techwish is seeking a Junior Cloud Infrastructure Operations Engineer for a contract opportunity with one of it’s clients in Princeton, NJ.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
7 Cloud Computing Trends Coming in 2025 for Developers
Is Software Engineering Over-Saturated?
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Highest Paying Tech Companies for Developers