Researcher - Foundations of Generative AI

Microsoft
Redmond, United States of America
1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 235K

Job location

New York, United States of America

Tech stack

Artificial Intelligence
Software Debugging
Github
Python
Machine Learning
Open Source Technology
TensorFlow
Software Engineering
PyTorch
Large Language Models
Information Technology
HuggingFace

Job description

  • Apply research and engineering skills to develop, prototype, and evaluate cutting-edge research ideas.
  • Work closely with other researchers and engineers to rapidly prototype and test new research ideas, driving a high-impact agenda and publishing results where appropriate.
  • Collaborate hands-on with other researchers, engineers, and internal and external product groups to deliver high-impact solutions to real-world problems.
  • Embody our culture and values.

Requirements

  • Doctorate (or currently pursuing) in Computer Science or relevant field
  • OR equivalent experience., * Doctorate in Computer Science or relevant field AND 2+ years related research experience
  • OR equivalent experience.
  • Research program demonstrated by public artifacts like models, tools, code in the AI space or publications at the following conferences: NeurIPS, ICML, ICLR, ACL, NAACL, CVPR, COLT, ECCV, ICCV, EMNLP.
  • 2+ years of academic or industry experience in developing, applying, and/or implementing algorithms for machine learning/statistics, using common ML engineering programming languages and platforms such as Python, Python numerical libraries, PyTorch, TensorFlow and/or HuggingFace.
  • Experience publishing academic papers as a lead author or essential contributor in a top AI conference or journal.
  • Deep understanding of frontier model architectures, especially transformers and state space models
  • Hands-on experience building and working with Large Language Models (LLMs) or multimodal models (VLMs, VLAs), including pre-training, fine-tuning, and inference
  • 2+ years of industry or academic experience with building, debugging and optimizing large-scale ML training pipelines.
  • Demonstrated software engineering excellence building and deploying prototypes, applications, or open-source (OSS) technologies. Providing a link to a GitHub profile and/or code samples on your CV/resume, is highly encouraged.
  • Ability to work independently and ramp-up quickly on complex projects or unfamiliar code
  • Ability to collaborate, communicate effectively, and work as part of a multi-disciplinary team
  • Keen interest in real-world applications and impact.

About the company

Microsoft is a global technology company headquartered in Redmond, Washington. Our mission is to empower every person and every organization on the planet to achieve more. We develop, license, and support a wide range of software products, services, and devices that help individuals and businesses realize their full potential.

Our flagship products include the Microsoft 365 productivity cloud, Windows operating system, Azure cloud platform, and Dynamics 365 business applications. We are also a leader in areas such as artificial intelligence, cybersecurity, developer tools, and gaming through Xbox and Game Pass.

With operations in more than 190 countries and over 220,000 employees worldwide, Microsoft is committed to responsible innovation, inclusive economic growth, and sustainability. We work closely with governments, industries, and communities to ensure that technology serves the public good and helps address some of the world’s most pressing challenges.

As we celebrate our 50th anniversary in 2025, we continue to look forward—investing in AI, cloud, and quantum computing to shape the future of work, education, and society at large scale.

Apply for this position