> Markdown version of [/jobs/ext/2222173-senior-data-science-engineer](https://www.wearedevelopers.com/jobs/ext/2222173-senior-data-science-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Science Engineer - **Company:** Legitscript Llc - **Location:** Portland, OR, United States (Remote available) - **Experience:** Expert - **Salary:** $160,000.0 - $175,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Airflow, Big Data, Continuous Integration, Data Validation, Information Engineering, Extract Transform Load (ETL), Data Warehousing, DevOps, Fraud Prevention and Detection, Python (Programming Language), Machine Learning, Performance Tuning, Cloud Services, Azure Machine Learning, Workflow Management Systems, Sql Optimization, Large Language Models, Snowflake, Prompt Engineering, Apache Spark, Generative AI, Git, Pyspark, Git Flow, Machine Learning Operations, Terraform, Docker, Databricks - **Published:** August 25, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/pf1v12r3od ## About the Role * 5-8+ years spanning data engineering and data science/ML, with a demonstrated track record of shipping models to production * Strong Python proficiency; experience with Spark/PySpark for large-scale data processing * Advanced SQL for complex transformation, analysis, and data modeling * Hands-on experience with cloud data platforms such as Databricks or Snowflake * Experience with ETL/ELT frameworks - dbt, Lakeflow Declarative Pipelines, Databricks Autoloader, Informatica, or similar * Familiarity with ML experiment tracking tools such as MLflow or Weights & Biases * DevOps fluency: Git-based development, branching strategies, CI/CD, IaC (DABs/Terraform), and Docker * Experience with orchestration tools such as Databricks Workflows or Apache Airflow, * Hands-on experience with LLMs and Generative AI techniques in a production context (prompt engineering, RAG architectures, fine-tuning, or evaluation frameworks) * Experience building or operating ML platforms, feature stores, or model registries * Prior work in risk, compliance, fraud detection, or other high-stakes ML domains ## Description * Research, prototype, and develop ML and LLM-based models to solve complex business problems, with a current focus on risk detection and prioritization * Wrap models into production-ready APIs and integrate them into our core product * Ensure model outputs are interpretable - translating predictions into actionable reason codes for end users * Partner directly with operational teams to gather feedback, refine features, and improve model relevance over time _Data Engineering _ * Design, build, and maintain scalable pipelines to ingest data from disparate sources into our data warehouse/lake * Implement robust data validation, quality checks, and transformation workflows across raw, curated, and serving layers * Build and maintain curated datasets optimized for both analytics and model training use cases _MLOps & Production Ownership _ * Implement and maintain CI/CD pipelines for both data workflows and ML model deployment across environments * Monitor pipeline latency, data drift, and model performance in production; design alerting and retraining triggers * Own the business outcomes of your models - define success metrics, track ROI, and iterate based on real-world efficacy * Manage infrastructure as code and containerized deployments to ensure reproducible, environment-consistent releases ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)