Kafka/Spark Developer

CGI Technologies and Solutions, Inc.
Pittsburgh, PA, United States
18 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$70,800.0 - $156,700.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Big Data Cloud Storage Cloudera Impala Information Engineering Data Transformation Data Warehousing Relational Databases DevOps Distributed Computing Environment Apache Hadoop
+26 more
Hadoop Distributed File System Apache Hive Python (Programming Language) PostgreSQL Neo4j NoSQL Object-Oriented Software Development Oracle (Applications) Performance Tuning Scrum Methodology Software Engineering SQL Databases Data Streaming Systems Integration Unstructured Data Freeform SQL Enterprise Software Applications Apache Spark Event Driven Architecture Pyspark Real Time Data Apache Kafka Spark Streaming Data Management Stream Processing Data Pipelines

Job description

CGI is looking for mid level Kafka and Spark Software Developers to join our Applications Development and Maintenance team, supporting our client which is a large US Bank, working in an advanced technology environment., As Kafka Spark Software developers, you will be responsible for developing and maintaining scalable big data solutions using Hadoop, Spark, Kafka, and Impala to support enterprise data processing and analytics initiatives.

. Design, build, and optimize batch and real-time data pipelines for ingesting, processing, transforming, and delivering large volumes of structured and unstructured data.

. Develop Spark applications using PySpark, Scala, or Java for data transformation, aggregation, cleansing, and analytical processing.

. Build and maintain Kafka producers, consumers, topics, and streaming workflows to enable reliable real-time data ingestion and event-driven architectures.

. Design and implement logical and physical data models to support data warehousing, reporting, analytics, and business intelligence requirements.

. Monitor, troubleshoot, and tune Kafka and Spark streaming jobs to improve performance, scalability, and operational reliability.

. Optimize Hadoop ecosystem components, Spark jobs, Kafka configurations, and Impala queries to improve system performance and resource utilization.

. Collaborate with architects, data engineers, DevOps teams, and business stakeholders to design and implement modern streaming and event-driven data platforms.

. Analyzing user requirements, and defines technical project scope and assumptions for assigned tasks.

. Creating technical designs for new systems, and/or modifications to existing systems.

. Translating detailed requirements into functional system designs.

. Prioritizing work, meeting deadline and also establishing and maintaining effective working relationships with clients, project team members, supervisors, and employees from other departments.

. Partner with business leaders, enterprise architects, and product owners to identify new graph-based use cases, evaluate emerging technologies, and align Neo4j initiatives with digital transformation goals.

Requirements

At least 5+ years of experience in Big Data development, data engineering, or distributed data processing environments.

. Strong hands-on experience with Apache Kafka, topic configuration, producer/consumer development, Kafka Connect, and Schema Registry.

. Extensive experience developing real-time data processing applications using Apache Spark Streaming and/or Spark Structured Streaming.

. Proficiency in Java, Scala, or Python (PySpark) with strong object-oriented programming and software development skills.

. Proficiency in writing and optimizing complex SQL queries using Impala, Hive, or similar distributed query engines.

. Hands-on experience with Hadoop ecosystem components including HDFS, Hive

. Experience integrating Kafka and Spark with relational databases, NoSQL databases, cloud storage platforms, and enterprise applications.

. Strong analytical, troubleshooting, and performance tuning skills in distributed streaming environments.

. Excellent communication, collaboration, and stakeholder management skills, with the ability to work effectively in Agile/Scrum teams.

. Experience working in Agile development environments with strong collaboration, technical leadership, problem-solving, and stakeholder communication skills., + Agile

  • Apache Kafka

  • Apache Spark

  • Hadoop Ecosystem (HDFS)

  • Impala

  • Oracle

  • Oracle RBDMS Audit

  • Postgre SQL

  • Python

  • SQL

Benefits & conditions

CGI is required by law in some jurisdictions to include a reasonable estimate of the compensation range for this role. The determination of this range includes various factors not limited to skill set, level, experience, relevant training, and licensure and certifications. To support the ability to reward for merit-based performance, CGI typically does not hire individuals at or near the top of the range for their role. Compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range for this role in the U.S. is $70,800.00 - $156,700.00.

CGI’s benefits are offered to eligible professionals on their first day of employment to include:

. Competitive compensation

. Comprehensive insurance options

. Matching contributions through the 401(k) plan and the share purchase plan

. Paid time off for vacation, holidays, and sick time

. Paid parental leave

.Learning opportunities and tuition assistance

About the company

Life at CGI is rooted in ownership, teamwork, respect and belonging. Here, you’ll reach your full potential because…

You are invited to be an owner from day 1 as we work together to bring our Dream to life. That’s why we call ourselves CGI Partners rather than employees. We benefit from our collective success and actively shape our company’s strategy and direction.

Your work creates value. You’ll develop innovative solutions and build relationships with teammates and clients while accessing global capabilities to scale your ideas, embrace new opportunities, and benefit from expansive industry and technology expertise.

You’ll shape your career by joining a company built to grow and last. You’ll be supported by leaders who care about your health and well-being and provide you with opportunities to deepen your skills and broaden your horizons.

Come join our team-one of the largest IT and business consulting services firms in the world.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:24 min

Comparing Neo4j and GraphQL conceptual models

William Lyon · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:09 min

Evaluating mature stream processing frameworks for production systems

Soroosh Khodami Soroosh Khodami · WWC 2024

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all