> Markdown version of [/jobs/ext/1319432-data-scientist-ll-digital-intelligence](https://www.wearedevelopers.com/jobs/ext/1319432-data-scientist-ll-digital-intelligence). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist ll - Digital Intelligence - **Company:** Socure Inc. - **Location:** Carson City, NV, United States - **Experience:** Expert - **Salary:** $140,000.0 - $170,000.0 - **Contract:** Permanent contract - **Skills:** Data Analysis, Big Data, Biometrics, Code Review, Encodings, Cyber Security, Database Queries, Distributed Computing Environment, Fraud Prevention and Detection, Monitoring of Systems, Virtual Private Networks (VPN), Python (Programming Language), Machine Learning, NumPy, Tensorflow, Signal Processing, Supervised Learning, Feature Engineering, Pytorch, Apache Spark, Model Validation, Pandas, Pyspark, Spoofing, Scikit Learn, Information Technology, Xgboost, Machine Learning Operations, Categorical Data, Unsupervised Learning, Databricks - **Published:** July 17, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=e83c437377562639 ## About the Role * Bachelor's, Master's, or Ph.D. in Computer Science, Machine Learning, Statistics, Mathematics, Data Science, or a related quantitative field, or equivalent practical experience. * 5+ years of experience in data science, applied machine learning, statistical modeling, analytics engineering, or a related technical role. * Experience building, evaluating, and improving machine learning models, features, analytical pipelines, or risk signals. * Strong SQL skills and experience working with large-scale, complex datasets. * Strong proficiency in Python and experience with data science libraries such as pandas, NumPy, scikit-learn, XGBoost, TensorFlow, PyTorch, or similar. * Experience with distributed data processing tools such as Spark, PySpark, Databricks, or equivalent frameworks. * Solid understanding of supervised learning, unsupervised learning, feature engineering, model evaluation, statistical validation, and experiment analysis. * Ability to work with noisy data, imperfect labels, missing values, instrumentation gaps, and changing data distributions. * Strong analytical judgment across data quality, feature design, model selection, explainability, and business impact. * Experience collaborating with engineering, product, analytics, or risk teams to move data science work toward production or operational use. * Clear communication skills, including the ability to explain technical work, assumptions, tradeoffs, and results to non-specialist stakeholders. * Ability to operate independently on defined problem areas while seeking guidance appropriately on ambiguous or high-risk decisions., * Background in fraud detection, identity verification, trust and safety, anomaly detection, cybersecurity, risk modeling, or another adversarial data domain. * Experience with device intelligence, browser/mobile fingerprinting, behavioral biometrics, network intelligence, VPN/proxy detection, or telemetry signal processing. * Experience developing features from high-cardinality categorical data using techniques such as aggregation, frequency encoding, target encoding, embeddings, graph features, or representation learning. * Familiarity with production ML workflows, model monitoring, feature monitoring, or batch and near-real-time decisioning systems. * Experience with dashboarding, model explainability, feature documentation, or customer-impact analysis. * Interest in adversarial behavior, fraud patterns, telemetry quality, and applied ML systems that operate in real-world production environments. ## Description This is a hands-on role for a data scientist who can independently deliver well-scoped projects, work with complex and noisy data, and partner with engineering, product, and risk teams to improve fraud detection, identity confidence, and customer outcomes. You will deepen your expertise in Digital Intelligence while contributing to models and signals used in real-world production decisions., * Develop machine learning features, models, and analytical methods for device, network, browser, mobile, session, and behavioral intelligence. * Work on scoped fraud and identity risk problems where data quality, labels, telemetry coverage, and product tradeoffs need careful analysis. * Build features from large-scale, high-cardinality, sparse, noisy, and platform-dependent telemetry. * Analyze signal patterns such as spoofing, emulator behavior, automation, proxy/VPN usage, low-entropy fingerprints, telemetry gaps, and device or session fragmentation. * Design and execute validation analyses, including train/test splits, holdout checks, leakage review, drift assessment, customer impact analysis, and feature stability review. * Use supervised, unsupervised, statistical, and heuristic approaches to identify durable fraud and identity risk signals. * Investigate imperfect labels, delayed outcomes, instrumentation gaps, and changing fraud patterns to distinguish useful signal from data artifacts. * Partner with senior data scientists, engineering, product, risk, and platform teams to clarify requirements, prepare data, implement features, and support production rollout. * Contribute to model documentation, feature definitions, explainability materials, dashboards, and production-readiness reviews. * Communicate methods, assumptions, findings, limitations, and recommendations clearly to technical and cross-functional stakeholders. * Support junior data scientists and analysts through code review, analytical feedback, and sharing effective modeling and validation practices. ## Related Videos - [TikTok's Privacy Innovation](https://www.wearedevelopers.com/videos/1036-tiktok-s-privacy-innovation) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Explainable machine learning explained](https://www.wearedevelopers.com/videos/589-explainable-machine-learning-explained) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [How to start an AI project for a good cause and boost your career](https://www.wearedevelopers.com/magazine/15-how-to-start-an-ai-project-for-a-good-cause-and-boost-your-career) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How machine learning can help us tell fact from fiction](https://www.wearedevelopers.com/magazine/509-how-machine-learning-can-help-us-tell-fact-from-fiction)