Data Practitioner/Data Lead-In

Info Dinamica Inc
New York, NY, United States
8 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Artificial Intelligence Amazon Web Services Amazon S3 Data Analysis CA Workload Automation Ae Cloud Computing Code Review Continuous Integration Data as a Services Data Architecture
+47 more
Information Engineering Data Governance Data Integrity Data Security Data Systems Linux DevOps Distributed Data Store Fraud Prevention and Detection Revision Control Systems Apache Hadoop Python (Programming Language) Linux System Administration Machine Learning Meta-Data Management Microsoft SQL Server Neo4j Azure Data Lake Shell Script SQL Databases Toolchain Enterprise Data Management Enterprise Software Applications Cloud Platform System Feature Engineering Data Ingestion Azure Data Factory Pytorch Snowflake Apache Spark Deep Learning Git Microsoft Fabric Kubernetes Data Lineage AWS Glue Data Analytics Real Time Data Apache Kafka Data Management Machine Learning Operations Restful APIs Azure Synapse Analytics Data Pipelines Databricks Microservices Control M

Job description

We are seeking a highly experienced Senior Data Practitioner / Data Lead to drive enterprise-wide data management initiatives across data acquisition, provisioning, governance, and analytics. This role will act as the lead for the data function, partnering closely with business stakeholders in New York and global technology teams in India to deliver scalable, secure, and high-quality data solutions.

The ideal candidate will bring strong expertise in Financial Crime, Fraud Analytics, AML, and Banking data ecosystems, with a proven track record of building and governing real-time data platforms while enabling business insights through modern data engineering practices and AI-driven automation

Day-to-Day Job Duties:

  • Lead end-to-end data management initiatives across data acquisition, data provisioning, data governance, data quality, metadata management, and analytics enablement.
  • Design, develop, and maintain scalable data platforms that support Fraud Technology, Financial Crime, AML, and Non-Financial Risk business functions.
  • Build and manage high-volume batch and real-time data pipelines using modern distributed data technologies.
  • Develop data ingestion, transformation, enrichment, and provisioning frameworks to support operational, analytical, and machine learning workloads.
  • Enable real-time data persistence and delivery to fraud investigation, surveillance, and risk management applications.
  • Design and maintain enterprise data lake environments, ensuring data is curated, governed, secure, and analytics ready.
  • Partner with Data Scientists and AI teams to support feature engineering, model training, model inference, and deployment of fraud detection and risk scoring solutions.
  • Utilize PyTorch to support machine learning workflows, fraud analytics models, anomaly detection frameworks, and AI-enabled automation initiatives.
  • Build and optimize data pipelines that feed AI/ML models and support model monitoring, retraining, and operationalization.
  • Implement enterprise data governance standards covering data quality, cataloging, lineage, metadata management, security, privacy, and regulatory compliance.
  • Develop automated frameworks and dashboards for monitoring Data Quality KPIs, data integrity, and platform performance.
  • Collaborate with business stakeholders in New York and global technology teams in India to deliver strategic data solutions.
  • Drive automation of data engineering and data management processes using AI, machine learning, and advanced analytics techniques.
  • Participate in architecture reviews, code reviews, Agile ceremonies, production support, and platform modernization initiatives.
  • Mentor junior engineers and promote adoption of modern data engineering, AI, and MLOps best practices.

Requirements

  • Minimum 8+ years of experience in Data Engineering, Data Management, Data Architecture, or Data Governance.
  • Minimum 5+ years of experience leading enterprise-scale data initiatives and delivering complex data solutions.
  • Strong experience within Banking, Financial Services, Capital Markets, or Financial Institutions.
  • Proven experience supporting Fraud Analytics, Financial Crime, AML, KYC, Risk Management, or Regulatory Reporting functions.
  • Strong expertise in data architecture, data modeling, and enterprise data management practices.
  • Extensive hands-on experience developing enterprise data solutions using Python and SQL.
  • Strong experience with machine learning and AI-enabled data platforms.
  • Hands-on experience with PyTorch for machine learning model development, feature engineering, training, inference, and deployment support.
  • Experience building and maintaining data pipelines that support fraud detection, predictive analytics, anomaly detection, and AI-driven decision-making.
  • Strong understanding of MLOps concepts, model lifecycle management, and integration of machine learning models within enterprise applications.
  • Experience designing and supporting real-time and near real-time data architecture.

Strong hands-on experience with:

  • Apache Spark
  • Hadoop Ecosystem
  • SQL Server
  • Distributed Data Platforms
  • Enterprise Data Lakes
  • Experience implementing enterprise data governance practices including:
  • Data Quality
  • Data Cataloging
  • Metadata Management
  • Data Lineage
  • Data Security & Compliance
  • Experience working in Linux environments and developing Python/Shell scripts.
  • Experience with Agile development methodologies, CI/CD pipelines, source control systems, and DevOps practices.
  • Strong analytical, problem-solving, communication, and stakeholder management skills.
  • Ability to work effectively across geographically distributed teams and business functions.

Required Technology Experience:

  • Python
  • PyTorch
  • Apache Spark
  • Hadoop
  • SQL Server
  • SingleStore (Distributed SQL Database)
  • Kubernetes
  • Python-based Microservices
  • REST APIs
  • Graph APIs
  • Enterprise Data Lakes
  • Data Governance Frameworks
  • Data Quality Platforms
  • Linux
  • Shell Scripting
  • Autosys, Control-M, or equivalent scheduling tools
  • Git, CI/CD, and DevOps Toolchains

Preferred Qualifications:

Experience leveraging PyTorch for fraud detection, graph analytics, anomaly detection, predictive risk scoring, and AI-driven surveillance solutions.

Experience developing and deploying deep learning models in production environments.

Experience with MLOps platforms and machine learning deployment frameworks.

Experience with Graph-based analytics using Neo4j.

Experience processing real-time business events using Kafka or similar streaming platforms.

Experience building near real-time feature engineering pipelines for machine learning inference.

Strong knowledge of Azure data platforms including:

  • Azure Databricks
  • Azure Data Factory
  • Azure Synapse Analytics
  • Azure Data Lake Storage
  • Experience with AWS data services including:
  • Amazon S3
  • AWS Athena
  • AWS Glue

Experience with Snowflake Data Cloud.

Experience implementing AI and Generative AI solutions to automate data engineering and operational processes.

Experience supporting enterprise Fraud Prevention, Fraud Investigation, Financial Crime Monitoring, or Risk Analytics programs.

Knowledge of modern data architecture patterns including Data Mesh and Data Fabric.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:24 min

Comparing Neo4j and GraphQL conceptual models

William Lyon · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all