> Markdown version of [/jobs/ext/641475-remote-senior-software-engineer-python-and-data-ecosystem](https://www.wearedevelopers.com/jobs/ext/641475-remote-senior-software-engineer-python-and-data-ecosystem). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Remote Senior Software Engineer - Python and Data Ecosystem - **Company:** ClickHouse - **Location:** Norwich, UK (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Query Performance, Application Programming Interfaces (APIs), Artificial Intelligence, Airflow, Cloud Computing, Databases, Information Engineering, Data Transformation, Dataspaces, Data Systems, Java Virtual Machine (JVM), Python (Programming Language), Machine Learning, Multiprocessing, Online Analytical Processing, NumPy, Open Source Technology, Query Optimization, Search Technologies, Software Engineering, SQL Databases, Retrieval-Augmented Generation, Large Language Models, Pandas, Kubernetes, Machine Learning Operations, Vertica, Api Design - **Published:** June 25, 2026 - **Apply:** https://find.jobs/jobs-near-me/remote-senior-software-engineer-python-and-data-ecosystem-norwich-norfolk/2835132644-2/ ## About the Role * 7+ years of software development experience, ideally with hands-on time as a Data Engineer, Data Scientist, or ML Engineer * Deep, proven experience designing, building, and maintaining production-grade Python connectors, SDKs, or integrations for at least one major platform (orchestration, BI, MLOps, or data transformation) * Solid experience with the Python data ecosystem: Pandas, NumPy, Pydantic, and related libraries * Prior contributions to, or deep practical experience with, popular data orchestration tools (Airflow, Dagster, or Prefect) * Hands-on experience with AI/ML in data engineering contexts: embedding generation, vector search, feature pipelines, or LLM-powered tooling in production, not just experimentation * Strong understanding of database fundamentals: SQL, data modeling, query optimization, and familiarity with OLAP/analytical databases * Solid experience with concurrent Python: threading, multiprocessing, and async patterns * Outstanding written and verbal communication skills; comfortable collaborating across engineering functions and with open-source communities Bonus points for: * Experience deploying AI/ML models in production, including inference APIs and vector databases * Prior experience as a Data Engineer or Data Scientist in a product-facing or platform role * Familiarity with ClickHouse or similar high-performance OLAP platforms * Familiarity with the JVM ecosystem ## Description As a Senior Software Engineer specializing in Python and the Data Ecosystem, you'll be a core contributor owning and evolving critical parts of ClickHouse's data engineering ecosystem. This role sits at the intersection of high-performance database engineering and developer experience. You'll craft tools that enable Data Engineers and Data Scientists to harness ClickHouse's speed and scale in the frameworks they already use. We're looking for someone who has lived the Data Engineer or Data Scientist experience firsthand. The data practitioner's world is shifting rapidly: databases are no longer just query targets, but they're becoming active participants in AI-powered workflows, serving as vector stores for RAG pipelines, backends for LLM-powered agents, and real-time feature stores for ML inference. You understand these workflows not from the outside, but because you've operated within them. You don't just build integrations, you bring product-level insight into what we should build and why. You'll own the full lifecycle of key Python integrations, driving architecture, performance, and feature direction across: * Orchestration Platforms: Apache Airflow, Dagster, Prefect * Transformation Tools: dbt, SQLMesh * AI & LLM Ecosystem: LangChain, LlamaIndex, n8n, and broader AI tooling: embedding pipelines, retrieval-augmented generation with ClickHouse as a vector store, ML feature stores, and LLM-powered data applications ClickHouse's columnar architecture and query performance make it exceptionally well-positioned in this new landscape. Your job is to make that potential real: building the robust, production-ready connectors that make ClickHouse the natural choice when data practitioners design their next-generation AI and data systems., * Own and evolve ClickHouse's Python connector and SDK ecosystem, raising the bar on performance, reliability, and API design * Build and maintain integrations with orchestration platforms (Airflow, Dagster, Prefect) and transformation tools (dbt) to enterprise-grade quality standards * Drive the AI/LLM integration strategy: designing connectors and patterns that make ClickHouse a natural fit in RAG architectures, ML feature pipelines, and LLM-powered data applications * Engage actively with the open-source community: triage issues, support contributors, advocate for users, and shape the roadmap based on real-world feedback * Collaborate with Product, Cloud, and other engineering teams to align integration work with broader platform priorities * Bring a practitioner's perspective to roadmap decisions, grounding prioritization in genuine Data Engineer and Data Scientist workflows ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Python Data Visualization @ Deepnote (w/ PyViz overview)](https://www.wearedevelopers.com/videos/113-python-data-visualization-deepnote-w-pyviz-overview) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)