Asst Dir-Data Scientist
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
- Evaluate and validate large language models for production-grade analytical and decision-support systems
- Design and implement evaluation frameworks for assessing LLM performance in credit analytics and decision-support contexts
- Develop metrics and benchmarks to measure model robustness, reliability, consistency, and output quality
- Analyze model behavior across diverse inputs, identifying failure modes, edge cases, and areas for improvement
- Collaborate with model development and deployment teams to integrate validation processes into the model lifecycle
- Conduct systematic assessments of model stability over time and across updates
- Evaluate model outputs for bias, fairness, and economic relevance to credit risk applications
- Develop and maintain documentation for evaluation methodologies, findings, and recommendations
- Contribute to the advancement of best practices for LLM evaluation within the Credit COE
About the Team Our Credit Center of Excellence (COE) team is responsible for maintaining and enhancing our industry-leading credit analytics and predictive modelling capabilities. We work closely with various teams including product management, commercial strategy, and go-to-market leaders to ensure the delivery of high-quality credit risk assessments and solutions. By joining our team, you will be part of exciting work in credit analytics with a global team spread across all US time zones, GMT, and GMT+1.
Requirements
- Ph.D. in Computer Science, Machine Learning, Natural Language Processing, Statistics, or a related quantitative field; or Master’s degree with 2-3 years of experience in machine learning evaluation or a related area
- Strong foundations in statistical methods, experimental design, and hypothesis testing
- Experience evaluating machine learning or NLP models, including designing experiments and interpreting results
- Familiarity with LLM evaluation benchmarks and methodologies
- Strong programming skills in Python or R
- Excellent communication skills in English (both written and verbal)
Preferred:
- Experience evaluating LLMs or generative AI systems
- Experience with production machine learning systems
- Exposure to cloud platforms such as AWS, GCP, or Azure
- Publications or demonstrated work in model evaluation, benchmarking, or related areas, * Ph.D. in Computer Science, Machine Learning, Natural Language Processing, Statistics, or a related quantitative field; or Master’s degree with 2-3 years of experience in machine learning evaluation or a related area
Benefits & conditions
For US-based roles only: the anticipated hiring base salary range for this position is $116,500.00 - $169,000.00, depending on factors such as experience, education, level, skills, and location. This range is based on a full-time position. In addition to base salary, this role is eligible for incentive compensation. Moody’s also offers a competitive benefits package, including not but limited to medical, dental, vision, parental leave, paid time off, a 401(k) plan with employee and company contribution opportunities, life, disability, and accident insurance, a discounted employee stock purchase plan, and tuition reimbursement.
About the company
hackajob is collaborating with Moody’s Corporation to connect them with exceptional professionals for this role.
At Moody’s, we unite the brightest minds to turn today’s risks into tomorrow’s opportunities. We do this by striving to create an inclusive environment where everyone feels welcome to be who they are-with the freedom to exchange ideas, think innovatively, and listen to each other and customers in meaningful ways. Moody’s is transforming how the world sees risk. As a global leader in ratings and integrated risk assessment, we’re advancing AI to move from insight to action-enabling intelligence that not only understands complexity but responds to it. We decode risk to unlock opportunity, helping our clients navigate uncertainty with clarity, speed, and confidence.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production
How to Become an AI Engineer
MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production
MLOps And AI Driven Development