Cloud and AI Ops Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+24 more
Job description
Experteer Overview As part of HPE Networking, you design and deliver cloud services for SASE, building end-to-end solutions for mission-critical apps. You work onsite in San Jose, collaborating with cross-functional teams to enable scalable, observable cloud infrastructure. You apply AI-driven approaches to detect and troubleshoot issues, raising reliability and performance. This role blends cloud engineering, SRE practices, and AI-enabled ops to accelerate secure access services for customers. Compensation / Benefits * Deploy cloud infrastructure using Docker containerization * Automate cloud orchestration with Kubernetes, Python, and Jenkins * Utilize NoSQL databases (MongoDB, DynamoDB) for cloud-native apps * Develop and deploy Lambda functions using AWS-SAM * Monitor and generate alerts with Grafana, Influx, or Kibana * Apply Infrastructure-as-Code with Terraform to deploy cloud services * Integrate logs, metrics, alerts, and telemetry into AIOps platforms * Provide SRE support and monitoring for HPE Networking SASE products * Design AI-driven strategies to detect, address, and auto-troubleshoot issues Tasks * 3+ years Python programming * 3+ years developing cloud-native apps, Kubernetes, and containerized environments * Experience with automation/CI-CD tools: Terraform, Ansible, Jenkins, and/or Git * Experience with monitoring tools: Grafana, Datadog, Prometheus, Observe, or Splunk * Public cloud experience (AWS, Azure, GCP, etc) * Familiarity with data analysis and basic machine learning concepts * Knowledge of NoSQL databases (MongoDB, DynamoDB) preferred * Networking knowledge (routing, TCP/IP, BGP, OSPF/ISIS, NetFlow, SNMP, ITE techniques) * Excellent written and verbal communication; growth mindset Key requirements * Health & wellbeing programs * Personal & professional development programs * Unconditional inclusion and flexible work arrangements * Culture of bold moves and collaboration
Requirements
and monitoring for HPE Networking SASE products * Design AI-driven strategies to detect, address, and auto-troubleshoot issues Tasks * 3+ years Python programming * 3+ years developing cloud-native apps, Kubernetes, and containerized environments * Experience with automation/CI-CD tools: Terraform, Ansible, Jenkins, and/or Git * Experience with monitoring tools: Grafana, Datadog, Prometheus, Observe, or Splunk * Public cloud experience (AWS, Azure, GCP, etc) * Familiarity with data analysis and basic machine learning concepts * Knowledge of NoSQL databases (MongoDB, DynamoDB) preferred * Networking knowledge (routing, TCP/IP, BGP, OSPF/ISIS, NetFlow, SNMP, ITE techniques) * Excellent written and verbal communication; growth mindset Key requirements * Health & wellbeing programs * Personal & professional development programs * Unconditional inclusion and flexible work arrangements * Culture of bold moves and collaboration
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on us.experteer.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
Highest Paying Tech Companies for Developers