Data Engineer

Calsoft Private Limited
Irvine, CA, United States
7 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$125,000.0 - $180,000.0
Working hours
Regular working hours
Job source

Tech stack

Clean Code Principles Airflow Amazon Web Services Amazon S3 Unit Testing Cloud Computing Code Review Databases Continuous Integration Data as a Services Data Validation Information Engineering
+33 more
Data Infrastructure Extract Transform Load (ETL) Data Warehousing Distributed Computing Environment Fault Tolerance Monitoring of Systems Identity and Access Management Python (Programming Language) PostgreSQL Machine Learning Metadata MySQL Object-Oriented Software Development Performance Tuning Query Optimization Cloud Services SQL Databases Data Streaming Data Processing Sql Optimization System Availability Database Optimization Apache Spark Git Cloudformation Data Lakes Pyspark Deployment Automation Apache Kafka Cloudwatch Terraform Data Pipelines Docker

Job description

Data Engineering & Platform Development

  • Design, develop, and maintain scalable batch and streaming data pipelines on AWS.

  • Build robust ETL/ELT frameworks for ingesting data from multiple internal and external sources.

  • Develop reusable data services that support analytics, reporting, machine learning, and operational workloads.

  • Design efficient data models optimized for both analytical and operational use cases.

  • Ensure data quality, consistency, lineage, and governance across the data platform.

Cloud & Infrastructure

  • Build and optimize cloud-native data infrastructure using AWS services.

  • Design highly available, fault-tolerant, and scalable data processing architectures.

  • Automate deployment, monitoring, and operational workflows using Infrastructure as Code and CI/CD practices.

  • Optimize storage, compute utilization, and overall cloud costs.

Performance & Reliability

  • Improve pipeline performance, scalability, and reliability.

  • Monitor production workloads and proactively resolve bottlenecks.

  • Optimize SQL queries, data partitioning, indexing strategies, and processing performance.

  • Troubleshoot production issues and implement preventive improvements.

Collaboration

  • Work closely with Data Scientists, Software Engineers, Product Owners, and Architects to understand business requirements.

  • Support downstream analytics, dashboards, and reporting solutions.

  • Participate in architecture discussions, code reviews, and technical design sessions.

  • Mentor junior engineers and promote engineering best practices.

Requirements

  • Strong Python programming skills.

  • Experience building production-grade data engineering solutions.

  • Knowledge of object-oriented design, testing, and clean coding practices.

AWS

Hands-on experience with several of the following:

  • S3

  • Glue

  • Lambda

  • EMR

  • Athena

  • Redshift

  • RDS

  • Step Functions

  • EventBridge

  • CloudWatch

  • IAM

  • ECS/EKS (preferred)

Data Engineering

  • ETL/ELT pipeline development

  • Data warehousing concepts

  • Batch and streaming data processing

  • Data modeling

  • Data validation and quality frameworks

  • Metadata and lineage concepts

Databases

  • Advanced SQL

  • PostgreSQL

  • MySQL

  • Redshift

  • Performance tuning and query optimization

Development Practices

  • Git

  • CI/CD pipelines

  • Unit testing

  • Docker

  • Agile/Scrum development

Preferred Skills

  • Apache Spark or PySpark

  • Kafka or Kinesis

  • Airflow or AWS Managed Workflows

  • Infrastructure as Code (Terraform or CloudFormation)

  • Data Lake architecture

  • Lakehouse concepts

  • Experience supporting Machine Learning data pipelines

  • Experience with observability and monitoring tools

Desired Experience

  • 6-10 years of experience in Data Engineering.

  • Experience building cloud-native data platforms on AWS.

  • Experience handling large-scale structured and semi-structured datasets.

  • Strong understanding of distributed data processing and performance optimization.

  • Experience working in enterprise production environments with high availability requirements.

Benefits & conditions

Base Annual Compensation: $125,000 - $180,000

The salary range listed represents the Company’s good-faith estimate of the base compensation reasonably expected to be offered for this position at the time of posting. The actual salary offered to the selected candidate may vary based on job-related factors, including work location, market conditions, relevant experience, education and training, technical skills, demonstrated abilities, internal equity, and the results of role-specific technical evaluations conducted during the selection process. This position may also be eligible for benefits and other forms of compensation, as applicable. Compensation decisions are made in accordance with applicable federal, state, and local laws.

Benefits we offer:

  • Comprehensive group health insurance coverage subject to plan terms.
  • 401(k) retirement savings plan participation, in accordance with applicable plan provisions.
  • Flexible work environment
  • Holidays and paid time off program, subject to company policy.

Calsoft core values:

  • Customer Focus
  • Integrity
  • Innovation
  • Ownership
  • Excellence
  • Team Spirit

About the company

Calsoft is an engineering-led digital product partner for global ISVs and tech-driven enterprises. Since 1998, we’ve helped customers build platforms, modernize systems, and scale with confidence, across cloud, AI, and connected ecosystems. With engineering roots in Silicon Valley and delivery strength in India, we operate at the intersection of product development, intelligent automation, and secure digital transformation.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all