> Markdown version of [/jobs/ext/2720339-software-data-engineer](https://www.wearedevelopers.com/jobs/ext/2720339-software-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software / Data Engineer - **Company:** Sleuth Insights, Inc. - **Location:** Los Angeles, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Applications Architecture, Big Data, Software Quality, Data as a Services, Information Engineering, Extract Transform Load (ETL), Data Stores, Data Systems, Relational Databases, Integrated Development Environments, Python (Programming Language), PostgreSQL, Software Tools, Cloud Services, Software Engineering, Large Language Models, Generative AI, Information Technology, Docker - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-software-data-engineer-sleuth-insights-8311546 ## About the Role We are seeking an experienced data engineer who has built enterprise-grade, cloud-native data infrastructure and products. In this role, you'll support our forward-deployed engineering efforts to build and ship data assets as part of pharma projects involving competitive intelligence and complex analytical tasks. You will be the primary bridge between these projects and the core engineering efforts to transform those assets into standard Sleuth platform AI-powered features. You will need to have knowledge and working experience with complex datasets in the pharma and biotech industry, a broad knowledge of data engineering tools and technologies, and the ability to swing between short/mid-term deliverables with customers and the mid/long-term sleuth technology roadmap., * Experience: 10 years of professional software and/or data engineering experience in the industry with 2+ years in small companies or startup environments. * Languages and tools: proficiency in Python, Docker, orchestration framework (e.g., Airflow), * Data stores: relational databases (specifically PostgreSQL). Experience with graph and vector databases is a big plus. * Generative AI: Knowledge of LLMs and GenAI frameworks (LangChain, LangSmith, LangGraph, AutoGen, MCP, etc.). Hands-on experience is a big plus. * Infrastructure & Ops: experience in leveraging cloud environments and data services for development and operation, strong in infrastructure-as-code. * Security & Compliance: experience in secure software development and familiar with SOC2 or regulated software development environments. * Education: BS or MS in computer science, engineering, math, biology, or a related scientific field. Additional hands-on certificates are great to have. * Working style: collaborative and organized with a strong sense of end-to-end ownership. * Domain Knowledge: familiarity with biopharma, biotech, or life sciences environments and experience with the industry datasets (e.g., diseases, drugs, clinical trials, etc.) ## Description * Design, develop, and operate AI-powered data solutions (ETL pipelines, entity extractions, and analytics) to deliver client projects. * Investigate new tools and technologies and develop proof of concepts to accelerate the delivery of data solutions. * Identify patterns that could be turned into platform capabilities and product features, document and analyze requirements, and deliver technical proposals. * Collaborate closely with the rest of the engineering team to ship platform capabilities and product features. * Develop applications and pipelines leveraging LLMs and modern agentic frameworks and tooling. * Leverage large-scale datasets for advanced analytics. * Contribute to engineering best practices across security, compliance, and software quality. * Build monitoring and observability into the infrastructure and all data components. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)