Cloud Platform Senior Consultant - Remote
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+54 more
Job description
Arity is part of Allstate Corporation, which means we share the same innovative drive that keeps us a step ahead of our customers’ evolving needs. We collect and analyze enormous amounts of data to provide cutting-edge solutions for companies invested in transportation.
The Team Our engineers are fueled by a passion to impact the future of mobility. As part of an Agile team, they have the freedom to innovate and the opportunity to see projects through from start to finish. Using a variety of languages and a modern technology stack, our engineers advance sensor technology, enterprise engineering, and platform development. Our team collaborates across a global engineering organization while maintaining trust, transparency, and empathy for the end user.
The Operational Data Management (ODM) team within Engineering owns the reliability, performance, and scalability of Arity’s data infrastructure. We operate mission-critical database and streaming platforms-including PostgreSQL, Redis/Valkey, Amazon Redshift, Google BigQuery, Amazon MSK (Kafka), and Google Pub/Sub-as well as analytics layers such as Starburst Galaxy and AWS Athena. We partner closely with application development teams to tune, troubleshoot, and optimize applications that depend on these technologies, ensuring the data platforms powering Arity’s mobility insights remain highly available and performant at scale., Arity is seeking a Cloud Platform Senior Consultant to join ODM. This is a fully remote position. You will design, build, deploy, and operate cloud-native data infrastructure across AWS and Google Cloud Platform, with hands-on work across databases, data streaming, and distributed systems., * Build, operate, and optimize data streaming infrastructure using Amazon MSK (Kafka) or Google Pub/Sub to support Real Time and batch data pipelines.
- Design, deploy, and manage highly available database and caching platforms-including PostgreSQL, Redis, Valkey, Amazon Redshift, and Google BigQuery-across multi-cloud environments.
- Develop and maintain infrastructure-as-code, CI/CD pipelines, and cloud automation using Terraform, Python and industry-standard tooling to enable repeatable, secure deployments.
- Implement monitoring, alerting, and observability for data platform services to proactively detect and resolve issues before they impact customers.
- Partner with application development teams to troubleshoot, tune, and optimize application performance, query patterns, and data access layers backed by team-managed platforms.
- Administer and optimize analytics and query engines including Starburst Galaxy and AWS Athena to deliver performant, cost-effective access to large-scale datasets.
- Participate in incident response, root cause analysis, and post-incident reviews for production database and streaming systems; drive remediation and preventive improvements.
- Share on-call rotation to provide support for mission-critical data infrastructure.
- Review application source code, identify performance or reliability issues, and collaborate on targeted fixes or optimization guidance with development teams.
- Evaluate emerging tools and automation approaches-including AI-assisted workflows-to improve operational efficiency and developer experience.
- Contribute to capacity planning, disaster recovery, security hardening, and cost optimization initiatives across the data platform estate., The candidate(s) offered this position will be required to submit to a background investigation.
Joining our team isn’t just a job - it’s an opportunity. One that takes your skills and pushes them to the next level. One that encourages you to challenge the status quo. One where you can shape the future of protection while supporting causes that mean the most to you. Joining our team means being part of something bigger - a winning team making a meaningful impact.
Allstate generally does not sponsor individuals for employment-based visas for this position.
Effective July 1, 2014, under Indiana House Enrolled Act (HEA) 1242, it is against public policy of the State of Indiana and a discriminatory practice for an employer to discriminate against a prospective employee on the basis of status as a veteran by refusing to employ an applicant on the basis that they are a veteran of the armed forces of the United States, a member of the Indiana National Guard or a member of a reserve component.
Requirements
You will help ensure the platforms that ingest, store, and serve billions of miles of driving data remain resilient, observable, and cost-efficient-directly enabling Arity’s products and the customers who rely on them. The ideal candidate combines solid cloud platform skills with strong database and streaming fundamentals, production-grade Python experience, and a collaborative, ownership-minded approach to production support and performance improvement with application teams. A willingness to learn new technologies quickly and deliver practical solutions in production is highly valued., * 3-5 years of software engineering or infrastructure experience, with at least 2 years in SRE, DevOps, or platform engineering operating production systems at scale.
- Hands-on experience designing, deploying, and managing cloud infrastructure on AWS and/or Google Cloud Platform, including networking, identity, and security fundamentals.
- Production experience operating data streaming platforms; hands-on work with Apache Kafka (including Amazon MSK or Confluent Kafka) and a solid understanding of partitions, consumer groups, delivery semantics, and backpressure.
- Production experience with relational and NoSQL databases; PostgreSQL required, plus familiarity with distributed data stores.
- Strong Python and Shell Scripting for automation, custom tooling, and operational solutions that go beyond basic scripts.
- Strong experience with infrastructure-as-code (eg, Terraform), CI/CD (eg, Jenkins, Git), Ansible, and container orchestration (eg, Kubernetes) in production environments.
- Experience implementing and automating monitoring, logging, and alerting for distributed systems (eg, Prometheus, Grafana, CloudWatch, Datadog, or equivalent), including automated runbooks where applicable.
- Proven ability to contribute to root cause analysis for production incidents spanning infrastructure, databases, streaming pipelines, and application code layers.
- Strong problem-solving, communication, and documentation skills with a track record of ownership in on-call and incident management environments.
- Working understanding of distributed systems principles including high availability, fault tolerance, consistency models, and disaster recovery.
Desired Skills
- Production experience with Apache Cassandra, including cluster operations, performance tuning, and troubleshooting in distributed environments.
- Knowledge of DynamoDB, ElastiCache, Apache NiFi, or self-managed Apache Flink, including checkpoint management, flow design, and streaming job troubleshooting.
- Advanced experience with Apache Kafka, Google Pub/Sub, and operating streaming workloads across both AWS and GCP.
- Experience administering or optimizing Starburst Galaxy, Trino, or AWS Athena for large-scale analytics workloads.
- Experience building AI agents, Model Context Protocol (MCP) servers, or LLM-based tooling to automate DevOps, observability, or operational workflows.
- Familiarity with Java/Spring Boot application troubleshooting, JVM diagnostics (heap dumps, GC tuning, thread dumps, connection pool analysis), or reading Golang production code for performance analysis.
- Experience with data pipeline orchestration (eg, Apache Airflow, dbt), event-driven architectures, or enterprise PaaS platforms (eg, Cloud Foundry).
- AWS or Google Cloud professional-level certifications, performance benchmarking, query plan analysis, database capacity planning, APM/distributed tracing, or open-source contributions in database, streaming, or infrastructure projects., Amazon CloudWatch, Apache Airflow, Apache Kafka, AWS DynamoDB, Cloud Foundry, Cloud Infrastructure, Cloud Native, Cloud Platform, Datadog, Data Infrastructure, Data Pipelines, Data Query, Java (Programming Language), Kubernetes, NoSQL Databases, PostgreSQL, Software Development, Terraform (Software), Allstate provides a comprehensive technology setup, including a laptop, monitors, headset, keyboard, and mouse. Employees eligible to work from home also receive a monthly connectivity reimbursement to help offset Internet costs.
When working from home, you must have a dedicated, private workspace free from distractions, along with appropriate desk and seating. Reliable Internet is required, with minimum speeds of 50 MB download and 5 MB upload.
Benefits & conditions
Compensation offered for this role is 90,700.00 - 153,925.00 annually and is based on experience and qualifications.
About the company
At Allstate, great things happen when our people work together to protect families and their belongings from life’s uncertainties. And for more than 90 years, our innovative drive has kept us a step ahead of our customers’ evolving needs. From advocating for seat belts, air bags and graduated driving laws, to being an industry leader in pricing sophistication, telematics, and more recently, device and identity protection.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Best Paying Remote Jobs
7 Cloud Computing Trends Coming in 2025 for Developers
How Much FAANG Companies Actually Pay Software Engineers in 2025
Fully Remote Software Engineer Jobs