Azure Data Engineer

Tata Consultancy Services Limited
Bellevue, WA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$64,000.0 - $104,000.0
Working hours
Regular working hours
Job source

Tech stack

Query Performance Microsoft Azure Databases Data Cleansing Information Engineering Data Integration Extract Transform Load (ETL) Data Transformation Data Warehousing Monitoring of Systems Python (Programming Language) Microsoft SQL Server
+14 more
SQL Azure Performance Tuning Power BI Kusto Query Language SQL Stored Procedures SQL Databases Parquet Microsoft Power Automate Azure Data Factory System Availability Apache Spark Pyspark Data Management Data Pipelines

Job description

Synapse; SQL Server; Python; Azure Storage; ADF; Cosmos; Kusto; Digital : PySpark; Fabric; Data Engineering; Data Warehouse Roles & Responsibilities: Unified Analytics Platform Work across Fabric workloads: COSMSO, Kusto, Data Factory (ETL), Synapse Data Engineering (Spark/SQL), Synapse Data Warehousing (SQL), and OneLake for storage. Data Loading & Architecture Design data loading patterns, Lakehouse architectures, and orchestration processes for enterprise-scale analytics. Data Preparation & Enrichment Prepare and enrich data for analysis, ensuring data quality and consistency checks. Implement semantic models for BI and self-service analytics. Security & Compliance Secure and manage analytics assets, enforce governance policies, and monitor compliance and Managed Identity. Performance & Optimization Monitor and optimize Data pipelines and workloads for cost efficiency and high availability. Collaboration & Stakeholder Engagement, Offshore team collaboration Work closely with architects, analysts, and business teams to translate requirements into scalable solutions. Enable Power BI integration for visualization and reporting. Data Integration & Orchestration Build and manage data pipelines using COSMOS, Azure Data Factory (ADF), Synapse, Fabric pipelines for batch and incremental ingestion. Automate ETL workflows and implement parameterized pipelines for scalability. Data Transformation & Modeling Develop ETL code / pipelines to clean, transform, and convert raw files into optimized formats (Parquet/Delta). Create views, stored procedures, and dimensional models for reporting and analytics. Performance Optimization Monitor Dynamic Management Views (DMVs) in Synapse to analyze system health and reduce bottlenecks. Implement workload management strategies for query performance and resource governance. Security & Governance Apply role-based authentication, encryption/decryption logic, and compliance controls. Automate governance tasks to reduce manual intervention. Reliability & Monitoring Build failure monitoring systems integrated with Logic Apps for alerting and proactive issue resolution. Strong in Azure Data Services ADF, Azure SQL database, Synapse Ability to understand vast amounts of data, identify and fix data issues.

Requirements

Do you have experience in Spark implementation?, Strong in database and DW concepts, Excellent technical & analytical skills with strong business acumen Strong communication skills Must have Microsoft Azure Cloud and services experience.

Benefits & conditions

Pulled from the full job description

  • Pet insurance
  • Health insurance
  • Vision insurance
  • Dental insurance, Discretionary Annual Incentive. Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans. Family Support: Maternal & Parental Leaves. Insurance Options: Auto & Home Insurance, Identity Theft Protection. Convenience & Professional Growth: Co mmuter Benefits & Certification & Training Reimbursement. Time Off: Vacation, Time Off, Sick Leave & Holidays.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

57 sec

Overview of the Kusto big data analytics platform

Prof Smoke Prof Smoke · WWC 2025

Videos

See all

Related articles

See all