> Markdown version of [/jobs/ext/513066-confluent-kafka-flink-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/513066-confluent-kafka-flink-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Confluent Kafka & Flink Infrastructure Engineer - **Company:** State Farm Insurance - **Location:** Bloomington, IL, United States - **Salary:** $95,800.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Continuous Integration, Digital Architecture, Disaster Recovery, Distributed Systems, Identity and Access Management, Performance Tuning, Runbook, Data Streaming, System Availability, Infrastructure as Code (IaC), Amazon Virtual Private Cloud (VPC), Backend, Kubernetes, Information Technology, Apache Flink, Apache Kafka, Cloudwatch, Dynatrace, Confluent - **Published:** June 11, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=d1f758446f4ca855 ## About the Role Do you have experience in System tuning?, * Hands-on production experience administering Apache Kafka or Confluent Platform (cluster administration, security, topic management, performance tuning). * Experience operating Apache Flink (or similar streaming frameworks) at the infrastructure level (checkpointing and state management). * Strong IaC experience and AWS familiarity (e.g., MSK/EC2/EKS, IAM, VPC, CloudWatch, S3). * Proven experience with HA/DR and reliability practices (SLA/SLO ownership, incident response, root-cause analysis). * Knowledge of security and networking patterns for distributed systems. * Experience integrating infrastructure changes into CI/CD with safe deployment and rollback. * Nice to Have * Experience with Confluent Cloud, container orchestration (e.g., EKS/ROSA), OpenTelemetry/distributed tracing, or GitOps. * Exposure to infrastructure considerations for agentic/AI-driven workloads. ## Description Help build and operate the Event Hub platform that enables event-driven capabilities across the enterprise. In this role, you will own the underlying Confluent Kafka and Apache Flink infrastructure, ensuring it is production-grade, highly available, scalable, and resilient, with robust disaster recovery., * Provision, configure, and manage Confluent Kafka (brokers, topics/partitions, ACLs, schema registry) and enforce platform governance standards (retention, quotas, security). * Operate Apache Flink infrastructure (job/task managers, checkpointing, state backend) and support runtime availability and performance. * Design and maintain high availability (AZ/region) and disaster recovery plans, including failover testing and runbooks. * Use Infrastructure as Code (IaC) and automation to deliver reproducible, auditable infrastructure and support infrastructure CI/CD. * Own observability and security for Kafka/Flink (monitoring/alerting/runbooks, encryption, IAM/VPC/network isolation, certificate and access management). * Partner with software teams to align standards, support onboarding, and document architecture and operational procedures. ## Related Videos - [Let's Get Started With Apache Kafka® for Python Developers](https://www.wearedevelopers.com/videos/565-let-s-get-started-with-apache-kafka-for-python-developers) - [Kafka Streams Microservices](https://www.wearedevelopers.com/videos/168-kafka-streams-microservices) - [Technical Documentation - How Can I Write Them Better and Why Should I Care?](https://www.wearedevelopers.com/videos/681-technical-documentation-how-can-i-write-them-better-and-why-should-i-care) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) - [What If We've Been Scaling Stream Processing Wrong All Along?](https://www.wearedevelopers.com/videos/100139-what-if-we-ve-been-scaling-stream-processing-wrong-all-along) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)