Solution Architect

HCLTech
London, UK
10 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Airflow Audit Trail Automation of Tests Batch Processing BigQuery Cloud Database Cloud Storage Code Review Continuous Integration Data Validation Data Governance
+34 more
Extract Transform Load (ETL) Data Systems Data Warehousing Dimensional Modeling Data Flow Control JSON Python (Programming Language) Metadata Meta-Data Management Query Optimization Release Management Cloudera Software Engineering SQL Databases Unstructured Data Parquet Data Processing Google Cloud Real Time Systems Apache Spark Git Data Lakes Kubernetes Infrastructure Automation Frameworks Information Technology Avro Apache Kafka Restful APIs Terraform Stream Processing Software Version Control Data Pipelines Apache Beam Docker

Job description

  • Design, develop, test, deploy, and maintain scalable ETL/ELT pipelines for structured and unstructured data.
  • Build batch and real-time processing solutions using BigQuery, Dataflow, Cloud Storage, Pub/Sub, and related GCP services.
  • Develop data-processing components and integrations using Java or Python, SQL, REST APIs, Kafka, and containerized services on GKE.
  • Implement reusable ingestion and transformation frameworks supporting full, incremental, and event-driven loads.
  • Design data models and optimize BigQuery tables through partitioning, clustering, efficient SQL, and cost-aware processing patterns.
  • Produce high-level and low-level technical designs, document assumptions, and contribute to architecture reviews.
  • Implement data-quality checks, schema validation, reconciliation, audit logging, lineage, monitoring, alerting, and error-handling mechanisms.
  • Troubleshoot complex production issues, perform root-cause analysis, and deliver sustainable corrective actions.
  • Apply security, privacy, access-control, reliability, and operational-support standards across data solutions.
  • Use source control, automated testing, CI/CD, and infrastructure-as-code practices to promote repeatable deployments.
  • Participate in code reviews, improve engineering standards, and mentor less-experienced engineers.
  • Collaborate with architects, analysts, application teams, platform teams, and business stakeholders to deliver reliable data products.

Requirements

  • Strong hands-on experience with GCP data services, particularly BigQuery, Dataflow, Cloud Storage, and Pub/Sub.
  • Proficiency in SQL, including complex transformations, query tuning, data validation, and analytical processing.
  • Proficiency in Java or Python for data pipelines, automation, API integration, and production support.
  • Good understanding of ETL/ELT, data warehousing, data lakes, dimensional modelling, batch processing, and stream processing.
  • Experience with Apache Beam, Kafka, REST APIs, Docker, Kubernetes, or GKE.
  • Experience with data formats such as JSON, CSV, Avro, and Parquet.
  • Working knowledge of Git, automated testing, CI/CD pipelines, observability, and release management.
  • Ability to design solutions for scalability, resilience, security, performance, operability, and cost efficiency.

Preferred Skills

  • Experience with Cloud Composer or Apache Airflow, Dataproc or Spark, Dataform or dbt, and metadata-driven pipeline frameworks.
  • Exposure to Terraform or another infrastructure-as-code tool.
  • Knowledge of Dataplex, data governance, metadata management, lineage, and access-control practices.
  • Experience modernizing legacy or on-premises data workloads to GCP.
  • Google Cloud Professional Data Engineer certification or an equivalent cloud data certification., Degree in Computer Science, Software Engineering or a related discipline, or equivalent practical experience. Experience delivering software solutions in agile teams. Knowledge of modern development frameworks, tools and engineering practices.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:03 min

Distinguishing type definition constructs from data validation routines

Clemens Vasters Clemens Vasters · World Congress 2025

1:52 min

Customizing block storage tiers and formats

Ricardo Sueiras Sueiras · LIVE

Videos

See all

Related articles

See all