> Markdown version of [/jobs/ext/2719738-corporate-vice-president-data-scientist-model-validation-and-ai-governance-organization](https://www.wearedevelopers.com/jobs/ext/2719738-corporate-vice-president-data-scientist-model-validation-and-ai-governance-organization). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Corporate Vice President - Data Scientist - Model Validation and AI Governance Organization - **Company:** New York Life Insurance Company - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $147,500.0 - $211,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, Cyber Security, Data Files, Fraud Prevention and Detection, Monitoring of Systems, Systems Analysis, Internet Security, Python (Programming Language), Machine Learning, Open Source Technology, Regression Testing, SQL Databases, Systems Architecture, Privacy Controls, Data Logging, Scripting, Large Language Models, Multi-Agent Systems, Prompt Engineering, Model Validation, Generative AI, Information Technology, Data Management, Virtual Agents - **Published:** September 4, 2026 - **Apply:** https://www.careerbuilder.com/job-details/corporate-vice-president-data-scientist-model-validation-and-ai-governance-new-york-ny--ae6ad8f9-974c-46e9-9fcd-00de931fea75 ## About the Role * Advanced degree in Statistics, Computer Science, Data Science, Mathematics, Economics, Engineering, or a related quantitative discipline, with strong knowledge of statistics and 7+ years of experience in model validation, model governance, or model risk management for predictive and AI/ML solutions within regulated environments. * Deep hands-on knowledge of traditional statistical and machine learning approaches, combined with demonstrated experience validating generative and agentic AI systems, including multi-step workflows, tool use, orchestration and planning, guardrails, autonomy controls, human oversight, and agent-specific failure modes. * Demonstrated experience designing, building, or independently reviewing AI evaluation frameworks, including dataset curation, benchmark design, rubric-based and LLM-as-a-judge evaluation, human review, regression testing, retrieval-quality assessment, and adversarial or red-team testing. * Practical experience with AI/ML monitoring and observability, including tracing, logging, telemetry, dashboards, and alerting, as well as experience validating or overseeing third-party, vendor-hosted, foundation-model, or other limited-transparency AI solutions. * Proficiency in Python and SQL, with experience using agent orchestration, evaluation, observability, or related AI tooling and the technical depth to independently replicate results, conduct diagnostic analysis, and build challenger approaches when appropriate. * Strong communication, leadership, and analytical skills, including the ability to interpret technical and regulatory requirements, translate them into practical controls, lead complex validations end-to-end, mentor colleagues, and present and defend technical findings with senior leaders and cross-functional governance bodies. Preferred Skills * Experience evaluating Generative AI solutions, including prompting strategies, retrieval-augmented generation (RAG), retrieval quality, model adaptation or fine-tuning trade-offs, and AI-specific privacy, security, interpretability, and stress-testing considerations. * Knowledge of agentic AI risks and controls, including prompt injection, excessive agency, tool and permission scoping, action reversibility, guardrail effectiveness, and hands-on exposure to commercial or open-source AI evaluation, observability, or guardrail technologies. * Familiarity with model and AI risk frameworks and emerging regulatory expectations, including SR 11-7, the NIST AI Risk Management Framework, the NAIC Model Bulletin on AI, and the EU AI Act, as well as experience collaborating with Legal, Compliance, Cybersecurity, and Third-Party Risk Management functions. * Applied understanding of AI use cases within financial services or insurance, such as underwriting support, fraud detection, marketing and sales enablement, agent productivity, and customer service, including the strengths, limitations, and evolving failure modes associated with generative and agentic AI., Analysis Skills, Artificial Intelligence (AI), Benchmarking, Best Practices, Communication Skills, Computer Science, Cross-Functional, Customer Support/Service, Data Analysis, Data Management, Data Science, Data Sets, Due Diligence, Economics, Failure Analysis, Financial Services, Injections, Insurance, Internet Security, Interpret Regulations, Leadership, Leading Edge Technology, Legal, LinkedIn, Machine Learning, Machine Tool, Marketing, Mathematics, Memory Hardware, Mentoring, Metrics, Model Validation, Newsroom, Open Source, Privacy Controls, Problem Solving Skills, Product Control, Python Programming/Scripting Language, Regression Testing, Regulations, Regulatory Compliance, Regulatory Requirements, Reporting Dashboards, Risk, Risk Management, Risk Management Framework (RMF), Risk Modeling, SQL (Structured Query Language), Sales, Statistics, Stress Testing, System Architecture, Systems Analysis, Team Player, Technical Presentation, Telemetry, Test Harness, U.S. National Institute of Standards and Technology (NIST), Underwriting, Use Cases, Vendor/Supplier Evaluation, Volunteer Experience ## Description The Corporate Vice President, Data Scientist - Model Validation and AI Governance will play a key leadership role in strengthening New York Life's approach to model validation and responsible AI governance across predictive, generative, and agentic AI solutions. Working closely with Model Risk Management and partners across Artificial Intelligence & Data, Technology, Risk, Legal, Compliance, Cybersecurity, and Third-Party Risk Management, this role will translate model risk requirements into rigorous, practical validation approaches and controls. This role will lead complex validation engagements end-to-end for internally developed and third-party solutions, independently challenging model methodologies, evaluation approaches, controls, and monitoring strategies. A particular focus will be establishing robust approaches for evaluating agentic and multi-step AI systems, including tool use, orchestration, autonomy, guardrails, human oversight, observability, and emerging failure modes. The successful candidate will combine deep quantitative and AI expertise with strong risk judgment and communication skills. They will establish validation standards and reusable practices, mentor junior colleagues, and communicate technical findings, limitations, and conditions of use in clear terms that enable senior stakeholders and governance forums to make informed decisions. What Youll Do * Lead independent validation and effective challenge across predictive, machine learning, generative AI, and agentic AI solutions, assessing model design, data and feature pipelines, methodologies, evaluation metrics, assumptions, limitations, and business impact. For agentic systems, evaluate architecture, planning and orchestration, tool use and permissions, retrieval and prompt design, memory and state, autonomy boundaries, guardrails, human-in-the-loop controls, and escalation and fallback mechanisms. * Design and advance rigorous AI evaluation practices by developing and challenging evaluation frameworks, benchmark and golden datasets, rubric-based and LLM-as-a-judge scoring, human review protocols, regression suites, offline and online testing, and adversarial or red-team evaluations. Assess statistical rigor, coverage, reproducibility, and the strength of validation evidence for high-impact use cases. * Establish monitoring, observability, and third-party validation approaches that address model drift, bias and fairness, stability, hallucinations, business outcomes, and agentic failure modes. Evaluate vendor and foundation-model solutions through independent testing and due diligence, and partner with teams to establish tracing, logging, telemetry, dashboards, alerts, compensating controls, and audit-ready evidence. * Translate risk and regulatory expectations into practical controls by interpreting technical standards, regulatory guidance, and internal procedures and converting requirements into validation methodologies, checklists, playbooks, standard operating procedures, monitoring expectations, and evidence requirements. Present validation findings, limitations, and conditions of use to senior stakeholders and governance committees. * Set standards and strengthen validation capabilities across teams by developing reusable templates, guidelines, test harnesses, and best practices for predictive, generative, and agentic AI; mentoring and reviewing the work of junior colleagues; partnering across data science, engineering, product, and control functions; and remaining current on evolving modeling techniques, AI research, evaluation methodologies, observability technologies, and governance practices. ## Related Videos - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Introduction to TXT](https://www.wearedevelopers.com/videos/30-introduction-to-txt) - [Crypto-secure Data Management with In-Database Blockchain](https://www.wearedevelopers.com/videos/632-crypto-secure-data-management-with-in-database-blockchain) - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Intermediate Bitcoin Script](https://www.wearedevelopers.com/videos/25-intermediate-bitcoin-script) ## Related Articles - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to start an AI project for a good cause and boost your career](https://www.wearedevelopers.com/magazine/15-how-to-start-an-ai-project-for-a-good-cause-and-boost-your-career)