AIOps (Artificial Intelligence For IT Operations)

CareerCircle
Wilmington, DE, United States
13 days ago
Apply on www.careercircle.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$145,000.0 - $235,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Systems Engineering Audit Trail Automation of Tests Microsoft Azure Bash Shell Bioinformatics Cloud Computing Configuration Management Code Coverage
+85 more
Code Generation Code Review Cyber Security Computational Biology Continuous Integration Information Engineering Extract Transform Load (ETL) Data Structures Data Visualization Data Warehousing Software Design Patterns Linux Disaster Recovery Failover Github R (Programming Language) Monitoring of Systems Network Topologies Information Technology Operations Intrusion Detection and Prevention Intrusion Detection Systems JSON Python (Programming Language) Linux System Administration Machine Learning Network Configuration and Change Management NetFlow Network Architecture Networking Basics Network Protocols Operational Databases Parsing Windows PowerShell Systems Development Life Cycle Red Hat Enterprise Linux Ansible Salesforce.Com Shell Script Security Information and Event Management Software Deployment SQL Databases Traffic Analysis Virtualization Technology Workflow Management Systems YAML Software Organization Network Routing Snort (Software) Scripting Cloud Platform System Data Ingestion GitHub Copilot Delivery Pipeline Prompt Engineering Mitre Att&ck Mttr Multi-Cloud Parallel Computation Technical Debt Cyber Threat Analysis HybridCloud Firewalls (Computer Science) Infrastructure as Code (IaC) Gitlab SC Clearance Data Lakes Pyspark Kubernetes Infrastructure Automation Frameworks Information Technology Deployment Automation Hashicorp Data Lakehouse Virtual Agents CIS Benchmarks Puppet Restful APIs Terraform Cyber Warfare Webhooks Software Version Control Data Pipelines Docker Servicenow Amazon Redshift

Job description

Hybrid Cloud Computing Intelligent Automation Environment Management Artificial Intelligence Software Design Patterns Configuration Management Bash (Scripting Language) Self Service Technologies Infrastructure Automation Verbal Communication Skills Employee Assistance Programs Infrastructure as Code (IaC) Python (Programming Language) Continuous Improvement Process Systems Development Life Cycle Key Performance Indicators (KPIs) Information Technology Operations Red Hat Certified Engineer (RHCE) Cloud Financial Management (FinOps) Artificial Intelligence Infrastructure Network Configuration And Change Management AIOps (Artificial Intelligence For IT Operations), The IT Operations Automation & AIOps Engineer is a hands-on technical role responsible for designing, building, and operating the automation and intelligent operations frameworks that power the organization’s AI-first IT environment. Reporting to the VP of Intelligent Automation and IT Operations, this engineer is a core builder of the greenfield IT operational environment - translating manual processes into automated, repeatable, and auditable workflows, and evolving those workflows toward agentic operations where AI systems take autonomous action with appropriate human oversight. This role spans infrastructure provisioning, configuration management, CI/CD pipeline integration, AI-assisted SDLC tooling, and the orchestration of AI agents across IT operational domains. The destination is not just automation - it is agentic IT operations: a model where AI agents handle routine operational events end-to-end while humans focus on strategy, governance, and edge-case resolution., * Infrastructure Automation & Provisioning

  • Design and maintain Infrastructure as Code (IaC) templates using Ansible to provision and manage cloud and on-premises resources consistently and repeatably.
  • Develop and maintain Terraform playbooks for configuration management, application deployment, patch automation, and compliance enforcement.
  • Build automated provisioning workflows for compute, storage, networking, and end-user environments.
  • Integrate automation pipelines with ITSM platforms (ServiceNow or equivalent) to enable self-service IT capabilities.
  • CI/CD & AI-Native SDLC for IT Operations
  • Implement and maintain CI/CD pipelines for infrastructure code
  • Apply software development best practices - version control, peer review, automated testing - to all infrastructure automation code.
  • Integrate and operationalize AI-assisted coding tools (GitHub Copilot, Amazon Q, or equivalent) into the IT operations SDLC, enabling AI-augmented code generation, review, and remediation for automation scripts and runbooks.
  • Develop automated compliance-as-code checks to validate infrastructure against security baselines and policy requirements.
  • Agentic Operations & AIOps Platform
  • Operate the AIOps platform, configuring AI-driven anomaly detection, alert correlation, and automated triage workflows.
  • Build and maintain the agentic workflow orchestration layer: design multi-step, AI-agent-executed operational workflows using orchestration frameworks, enabling IT agents to autonomously handle routine incidents and operational tasks.
  • Define human-in-the-loop thresholds for agentic operations: specify which actions agents may take autonomously, which require human approval, and how agent actions are logged, audited, and reviewed.
  • Integrate automation platforms with monitoring and observability tools to enable closed-loop remediation - from alert detection through root cause analysis to ticket resolution, without human intervention for defined incident types.
  • Identify and prioritize manual IT operational processes suitable for agentic automation, developing business cases and ROI models for each initiative.
  • Cloud & Hybrid Environment Management
  • Automate provisioning and lifecycle management of workloads across AWS, Azure, and/or GCP environments.
  • Implement FinOps practices including automated tagging, resource scheduling, and rightsizing recommendations to control cloud spend.
  • Maintain automation tooling for hybrid cloud environments, ensuring consistent policies across on-premises and cloud-hosted systems.
  • Contribute to disaster recovery automation, including scripted failover and failback procedures with validated RTO/RPO targets.
  • Track and report automation and agentic operations KPIs including tickets automated, MTTR reduction, agent action accuracy, and hours saved per sprint cycle.
  • Documentation & Continuous Improvement
  • Create and maintain comprehensive documentation for all automation frameworks, agentic workflows, runbooks, and operational procedures.
  • Conduct regular reviews of existing automations to identify optimization opportunities, address technical debt, and evaluate candidates for elevation from script-based automation to agentic orchestration.
  • Mentor junior team members and IT operations staff on automation tooling, AIOps platforms, agentic workflow design, and AI-assisted coding practices.
  • Track and report automation KPIs including tickets automated, MTTR reduction, agent action success rate, and hours saved per sprint cycle., At V2X, we are deeply committed to both equal employment opportunity, including protection for Veterans and individuals with disabilities, and fostering an inclusive and diverse workplace. We ensure all individuals are treated with fairness, respect, and dignity, recognizing the strength that comes from a workforce rich in diverse experiences, perspectives, and skills. This commitment, aligned with our core Vision and Values of Integrity, Respect, and Responsibility, allows us to leverage differences, encourage innovation, and expand our success in the global marketplace, ultimately enabling us to best serve our clients. Related Jobs Senior Translational Data & AI Engineer Actalent

Wilmington, DE*Remote

CI/CD Writing Tooling Biology PySpark Research Genomics Visionary AI Agents Pipelines Automation Innovation Biomarkers Data Lakes Proteomics Code Review Scalability Data Quality Observability Data Modeling Code Coverage Reconciliation Failure Causes Data Ingestion Data Pipelines Bioinformatics Data Lakehouse GitHub Copilot Version Control Amazon Redshift Computer Science Data Warehousing Data Engineering Inventory Staging Data Visualization Scientific Studies Industry Standards Biological Studies Workflow Management Amazon Web Services Workflow Automation Parallel Processing Operational Databases Computational Biology Artificial Intelligence R (Programming Language) SQL (Programming Language) Engineering Design Process Extract Transform Load (ETL) Data Warehouse Architectures Python (Programming Language) Active Directory Application Mode Application Programming Interface (API) +0

Salesforce Developer AI & Security Infrastructure Integration Engineer Leidos

Alexandria, VA*Remote

JSON YAML Linux Triage Ansible Parsing NetFlow Firewall Equities Webhooks Terraform Pipelines Leadership Automation Market Data RESTful API Shell Script Cyber Defense Change Control Cyber Security Virtualization Data Structures Network Routing Ancient History Machine Learning Network Topology Threat Detection Secret Clearance Traffic Analysis Network Protocols Elevation Drawings Prompt Engineering Workflow Management Linux Administration Network Infrastructure MITRE ATT&CK Framework Artificial Intelligence Infrastructure Security Bash (Scripting Language) Cyber Threat Intelligence IAT Level II Certification Cyber Kill Chain Framework CSSP Infrastructure Support Python (Programming Language) Snort (Intrusion Detection System) Puppet (Configuration Management Tool) Application Programming Interface (API) Security Information And Event Management (SIEM) Top Secret-Sensitive Compartmented Information (TS/SCI Clearance) +0

Google Cybersecurity IT Operations Automation & AI Ops Engineer V2X

Remote

CI/CD Triage Gitlab Github Writing Ansible Tooling Auditing Failover Scripting Terraform AI Agents HashiCorp Pipelines Scheduling Operations Automation Governance Innovation Kubernetes ServiceNow Peer Review Multi-Cloud Observability Cloud Hosting Technical Debt GitHub Copilot Professionalism Version Control Test Automation Microsoft Azure Problem Solving Code Generation Computer Science Secret Clearance Docker (Software) Disaster Recovery Autonomous System Anomaly Detection Windows PowerShell Service Industries Security Clearance Technical Training Workflow Management Root Cause Analysis Systems Engineering Amazon Web Services Time Off Management Software Development Lifecycle Management IT Service Management System Administration Information Technology Application Deployment Hybrid Cloud Computing Intelligent Automation Environment Management Artificial Intelligence Software Design Patterns Configuration Management Bash (Scripting Language) Self Service Technologies Infrastructure Automation Verbal Communication Skills Employee Assistance Programs Infrastructure as Code (IaC) Python (Programming Language) Continuous Improvement Process Systems Development Life Cycle Key Performance Indicators (KPIs) Information Technology Operations Red Hat Certified Engineer (RHCE) Cloud Financial Management (FinOps) Artificial Intelligence Infrastructure Network Configuration And Change Management AIOps (Artificial Intelligence For IT Operations) +0 Login | JoinContact

Requirements

Triage Gitlab Github Writing Ansible Tooling Auditing Failover Scripting Terraform AI Agents HashiCorp Pipelines Scheduling Operations Automation Governance Innovation Kubernetes ServiceNow Peer Review Multi-Cloud Observability Cloud Hosting Technical Debt GitHub Copilot Professionalism Version Control Test Automation Microsoft Azure Problem Solving Code Generation Computer Science Secret Clearance Docker (Software) Disaster Recovery Autonomous System Anomaly Detection Windows PowerShell Service Industries Security Clearance Technical Training Workflow Management Root Cause Analysis Systems Engineering Amazon Web Services Time Off Management Software Development Lifecycle Management IT Service Management System Administration, + Bachelor’s degree in Information Technology, Computer Science, Systems Engineering, or related field OR an equivalent combination of education and experience from which comparable knowledge and job skills can be obtained. (One year related experience may be substituted for one year of education, if degree is required) ., + A minimum of five (5) years of experience in IT operations, systems administration, or a related technical discipline.

  • A minimum of three (3) years of hands-on experience writing and maintaining automation using Ansible and/or Terraform.
  • Demonstrated experience with scripting languages: Python and/or Bash required; PowerShell a plus.
  • Experience with CI/CD pipelines and version control systems (GitLab, GitHub, or equivalent).
  • Experience with AI-assisted coding tools (GitHub Copilot, Amazon Q, or equivalent) in a production IT operations context preferred.
  • Familiarity with ITSM platforms such as ServiceNow and integration of automation into ticketing workflows.
  • Experience operating in cloud environments (AWS, Azure, or GCP); multi-cloud experience a plus.
  • Experience in the government contractor/services industry preferred.
  • U.S. Citizenship required; ability to obtain and maintain a security clearance.
  • Other Requirements:
  • Ability to travel to project and customer locations as needed.
  • Certifications:
  • Industry certifications preferred: HashiCorp Terraform Associate, Red Hat Certified Engineer (RHCE), AWS/Azure/GCP Associate-level or equivalent.
  • Ability to obtain and maintain CMMC Level 2 certification.
  • Preferred Requirements:
  • Master’s degree or equivalent advanced technical training preferred.
  • Active Secret security clearance desired.
  • Experience with container orchestration (Kubernetes, Docker) and related automation tooling.
  • Familiarity with AI agent governance patterns: human-in-the-loop design, agent audit logging, and permission scoping for autonomous systems.
  • Experience designing or operating agentic IT operations platforms.
  • Skills:
  • Strong proficiency in Ansible, Terraform, Python, and Bash scripting.
  • Hands-on experience with AIOps platforms and closed-loop remediation design.
  • Familiarity with agentic workflow orchestration frameworks and human-in-the-loop design patterns.
  • Working knowledge of AI-assisted SDLC tools and their application to IT operations code.
  • Working knowledge of networking fundamentals sufficient to automate network configuration tasks.
  • Familiarity with security baselines, CIS benchmarks, and compliance-as-code practices.
  • Strong problem-solving skills and a bias toward automation-first and agentic-operations thinking.
  • Clear written and verbal communication skills, including the ability to document technical processes for non-technical audiences.

Benefits & conditions

  • The successful candidate’s starting pay will be based on, but not limited to, their job related skills, experience, qualifications, work location, and market conditions.
  • The following salary range is intended to display the value of the company’s base pay compensation and may be modified at the discretion of the company.
  • USD $ 145,000 -235,000
  • Provided salary range minimum and maximum values correspond to variances between regional/geographic locations across the United States.
  • Please speak with a recruiter for additional information.
  • Employee benefits include the following:
  • Healthcare coverage
  • Life insurance, AD&D, and disability benefits
  • Retirement plan
  • Wellness programs
  • Paid time off, including holidays
  • Learning and Development resources
  • Employee assistance resources
  • Pay and benefits are subject to change at any time and may be modified at the discretion of the company, consistent with the terms of any applicable compensation or benefit plans.

About the company

Working across the globe, V2X builds smart solutions designed to integrate physical and digital infrastructure from base to battlefield. We bring 120 years of successful mission support to improve security, streamline logistics, and enhance readiness. Aligned around a shared purpose, our $4.5B company and 16,000 people work alongside our clients, here and abroad, to tackle their most complex challenges with integrity, respect, responsibility, and professionalism.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careercircle.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:19 min

Applying code assistant capabilities to infrastructure and cloud operations

Ryan J Salva · Coffee With Developers

1:35 min

Centralizing configuration logic with native YAML block references

Matthieu Vincent Matthieu Vincent · Europe 2026 Virtual

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · World Congress 2021

2:00 min

Introduction to YAML syntax and basic formatting

Chris Ayers · LIVE

2:03 min

Distinguishing type definition constructs from data validation routines

Clemens Vasters Clemens Vasters · World Congress 2025

Videos

See all

Related articles

See all