Big Data Engineer II

IntraEdge, Inc.
United States
19 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Computer-Aided Design Amazon Web Services Software Applications Application Performance Management Computer Programming Couchbase Servers Data Visualization Data Warehousing Github Apache Hadoop MapReduce Apache HBase
+16 more
Apache Hive Python (Programming Language) Unix Shell MongoDB NoSQL Scrum Methodology Software Engineering SQL Databases Tableau (Software) Apache Spark Pyspark Apache Kafka Stream Processing Looker Analytics Data Pipelines Microservices

Job description

Strong HIVE, SPARK, SQL, UNIX

  • Develop and design software applications, translating user needs into system architecture. Assess and validate application performance and integration of component systems and provide process flow diagrams. Test the engineering resilience of software and automation tools.
  • Be part of an enthusiastic, high performing technology team developing solutions to drive engagement and loyalty within our existing cardmember base and attract new customers to the Amex brand.
  • The position will also play a critical role partnering with other development teams, testing and quality, and production support, to meet implementation dates and allow smooth transition throughout the development life-cycle.
  • The successful candidate will be focused on building and executing against a strategy and roadmap focused on moving from monolithic, tightly coupled, batch-based legacy platforms to a loosely coupled, event-driven, microservices-based architecture to meet our long-term business goals.

Requirements

  • 5+ years of software development experience and leading teams of engineers and scrum teams
  • 3+ years of hands-on experience of working with Map-Reduce, Hive, Spark (core, SQL and PySpark)
  • Hands-on experience on writing and understanding complex SQL(Hive/PySpark-dataframes), optimizing joins while processing huge amount of data
  • Experience in UNIX shell scripting

Additional Good to have requirements:

  • Solid Datawarehousing concepts
  • Knowledge of Financial reporting ecosystem will be a plus
  • Experience with Data Visualization tools like Tableau, SiSense, Looker
  • Expert on Distributed ecosystem
  • Hands-on experience with programming using Python/Scala
  • Expert on Hadoop and Spark Architecture and its working principle
  • Ability to design and develop optimized Data pipelines for batch and real time data processing
  • Should have experience in analysis, design, development, testing, and implementation of system applications
  • Demonstrated ability to develop and document technical and functional specifications and analyze software and system processing flows
  • Aptitude for learning and applying programming concepts.
  • Ability to effectively communicate with internal and external business partners. Preferred Qualifications:
  • Knowledge of cloud platforms like GCP/AWS, building Microservices and scalable solutions, will be preferred
  • 2+ years of experience in designing and building solutions using Kafka streams or queues
  • Experience with GitHub and leveraging CI/CD pipelines
  • Experience with NoSQL i.e., HBase, Couchbase, MongoDB

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.intraedge.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all