Spark Data Engineer with Apache Iceberg / Trino

Infosys
Austin, TX, United States
6 days ago
Apply on us.experteer.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Business Analytics Applications Apache HTTP Server Big Data Code Review Computer Programming Data Warehousing Database Design Performance Tuning Systems Development Life Cycle Standard Sql Tableau (Software)
+8 more
Strategies of Testing Software Repository Scripting Apache Spark Pyspark Information Technology Data Analytics Data Pipelines

Job description

Experteer Overview As Technology Consultant 2, you will translate requirements into actionable development tasks and guide the technical team through design decisions. You’ll contribute across the SDLC, from elicitation to deployment, while resolving production issues and documenting knowledge for future projects. You will collaborate with cross-functional teams to build robust data analytics, reporting, and data warehousing solutions. This role offers the opportunity to shape scalable BI and analytics platforms in a dynamic, AI-enabled environment. You will work with cutting-edge tools to deliver high-quality, value-driven software in a global IT services setting. Compensation / Benefits * Document assigned parts of business requirements per guidance * Facilitate design discussions and record design decisions * Code, integrate features, and maintain application stability * Conduct code reviews and maintain code repositories * Implement test strategies, analyze results, and coordinate fixes * Develop user training, documentation, and support frameworks * Resolve production issues and propose preventive strategies * Maintain records of code, tests, and support activities Tasks * Experience with Spark/PySpark, Apache Iceberg and Trino * Complex SQL coding with performance tuning expertise * Python scripting and programming skills * Gen-AI skills (Claude Code) and client-context solution development * BI, Reporting and Data Warehousing experience * End-to-end reporting/dashboard design and data pipelines * Stakeholder collaboration to translate requirements into technical design * Tableau dashboard development and database design * Experience with large data volumes and leading offshore data engineers Key requirements * Medical/Dental/Vision/Life Insurance * Long-term/Short-term Disability * Health and Dependent Care Reimbursement Accounts * Insurance (Accident, Critical Illness, Hospital Indemnity, Legal) * 401(k) plan and contributions * Paid holidays and Paid Time Off

Requirements

_ fixes * Develop user training, documentation, and support frameworks * Resolve production issues and propose preventive strategies * Maintain records of code, tests, and support activities Tasks * Experience with Spark/PySpark, Apache Iceberg and Trino * Complex SQL coding with performance tuning expertise * Python scripting and programming skills * Gen-AI skills (Claude Code) and client-context solution development * BI, Reporting and Data Warehousing experience * End-to-end reporting/dashboard design and data pipelines * Stakeholder collaboration to translate requirements into technical design * Tableau dashboard development and database design * Experience with large data volumes and leading offshore data engineers Key requirements * Medical/Dental/Vision/Life Insurance * Long-term/Short-term Disability * Health and Dependent Care Reimbursement Accounts * Insurance (Accident, Critical Illness, Hospital Indemnity, Legal) * 401(k) plan and contributions * Paid holidays and Paid Time Off

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all