Database Architect

SHYN I.T BUSINESS SOLUTIONS PRIVATE LIMITED
United States
16 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$2,080.0
Working hours
Regular working hours
Job source

Tech stack

Query Performance Business Analytics Applications Information Systems Databases Data Architecture Data Validation Information Engineering Data Integrity Extract Transform Load (ETL) Data Systems Database Development Monitoring of Systems
+18 more
Issue Tracking Systems Python (Programming Language) Performance Tuning Standard Sql Software Deployment SQL Databases Technical Data Management Systems Data Processing Data Ingestion Backend Data Layers Data Lakes Pyspark Information Technology Data Lineage Tools for Reporting Data Pipelines Databricks

Job description

  • We are seeking an experienced Data Architect (Databricks / Backend Data Engineer) to support a mission-critical U.S. Air Force program. The selected candidate will be responsible for designing, developing, and maintaining scalable data architectures and production-grade data pipelines within the ADVANA environment. This role focuses on building and optimizing Databricks-based data solutions using the Medallion Architecture (Bronze, Silver, Gold) to deliver reliable, high-quality data for enterprise analytics and reporting.
  • The ideal candidate will have strong expertise in Databricks, Delta Lake, PySpark, SQL, Python, data pipeline development, and data validation, along with experience working in secure DoD environments.
  • Key Responsibilities
  • Data Architecture & Engineering
  • Design, develop, and maintain scalable Databricks Delta Lake data architecture.
  • Build and optimize Bronze, Silver, and Gold data pipelines using Medallion Architecture.
  • Develop curated datasets for downstream analytics and reporting platforms.
  • Design efficient Delta Table schemas and optimize storage and query performance.
  • Implement business rules and transformation logic within the Databricks data layer.
  • Data Pipeline Development
  • Develop and maintain ETL/ELT pipelines using PySpark, SQL, and Databricks.
  • Build automated ingestion pipelines from multiple authoritative data sources.
  • Implement incremental loads, upserts, scheduling, monitoring, and automated workflows.
  • Manage production deployments, pipeline monitoring, troubleshooting, and performance tuning.
  • Optimize data processing for scalability, reliability, and cost efficiency.
  • Data Quality & Validation
  • Perform data validation, reconciliation, row count verification, checksum analysis, and performance benchmarking.
  • Investigate and resolve discrepancies between source systems and transformed datasets.
  • Maintain metadata, data lineage, schema documentation, and transformation documentation.
  • Ensure data integrity, consistency, and compliance with organizational standards.
  • Platform Administration
  • Manage Databricks workspaces, clusters, job orchestration, and access control.
  • Support platform operations, environment configuration, and production maintenance.
  • Coordinate platform requests, issue tracking, and operational support activities.
  • Monitor system performance and recommend optimization strategies.
  • Collaboration
  • Partner with data architects, visualization developers, and business stakeholders.
  • Support integration with reporting and analytics platforms.
  • Translate business requirements into scalable technical data solutions.
  • Participate in technical reviews, architecture discussions, and continuous improvement initiatives.

Requirements

  • Active U.S. Secret Security Clearance (or higher).
  • Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field (or equivalent professional experience).
  • 5+ years of experience in Data Engineering, Data Architecture, or Backend Data Development.
  • Strong hands-on experience with:
  • Databricks
  • Delta Lake / Delta Tables
  • PySpark
  • SQL
  • Python
  • Experience designing and implementing Medallion Architecture (Bronze, Silver, Gold)., * Bachelor’s (Required)

Experience:

  • Database Architect: 6 years (Required)
  • Medallion Architecture (Bronze, Silver, Gold).: 3 years (Required)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

2:40 min

Evolving through early developer roles and technology stacks

Liam Hurrel +1 · WWC 2021

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

4:01 min

Managing application isolation via pluggable database models

Wei Hu Wei Hu · WWC 2022

Videos

See all

Related articles

See all