AI Security Engineer

Microsoft
New York, NY, United States
6 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$102,100.0 - $202,200.0
Working hours
Regular working hours

Tech stack

Microsoft Access Application Programming Interfaces (APIs) Artificial Intelligence Data Analysis Software System Penetration Testing Automation of Tests Microsoft Online Services Cloud Computing Cyber Security Computer Programming Home Automation Internet Security
+13 more
Python (Programming Language) Key Management Microsoft Software Language Modeling Systems Development Life Cycle Software Engineering Scripting Large Language Models Software Security Model Validation Generative AI Information Technology Vulnerability Analysis

Job description

MRT Strategic AI is seeking an AI Security Engineer focused on Cyber Capability Evaluation. This is a hands-on technical role for implementing and executing evaluations of frontier AI models and agentic systems in realistic, controlled cyber environments. The engineer will work at the intersection of offensive security, AI agents, model evaluation, and experimental engineering to build and maintain evaluation tasks, run experiments, analyze model behavior, and determine whether observed results represent meaningful cyber capability, an evaluation artifact, or a limitation in the test environment.

The role requires the ability to work independently within established technical direction, solve engineering and security problems, and contribute improvements to evaluation methods, cyber ranges, scoring, telemetry, and containment. The right candidate combines offensive-security experience with experimental discipline, programming skills, and an interest in emerging AI capabilities.

Responsibilities

  • Design, implement, and operate offensive cyber-capability evaluations for frontier, preview, production, and open-weight AI models.
  • Build and maintain realistic evaluation tasks covering vulnerability analysis, exploit development, application and system exploitation, attack-path reasoning, post-exploitation activities, and multi-step offensive workflows.
  • Execute controlled experiments using AI models; define success criteria and baselines; collect reliable telemetry; and document results.
  • Analyze model trajectories and investigate unexpected behavior to determine whether it reflects genuine capability, task leakage, environmental flaws, scoring errors, or other evaluation artifacts.
  • Develop and maintain cyber ranges, vulnerable applications, exploit-development targets, and evaluation harnesses, and apply security controls for sandboxing, secrets management, telemetry, access, and containment when testing autonomous AI systems.
  • Partner with red-team operators, researchers, engineers, and Responsible AI stakeholders to translate evaluation findings into security insights, and communicate results through reports, documentation, and briefings.

Requirements

  • Masters Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 1+ year(s) experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
  • OR Bachelors Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 2+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
  • OR equivalent experience.

Other Requirements:

Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings:

Microsoft Cloud Background Check:

  • This position will be required to pass the Microsoft background and Microsoft Cloud background check upon hire/transfer and every two years thereafter.
  • This role will require access to information that is controlled for export under export control regulations, potentially under the U.S. International Traffic in Arms Regulations or Export Administration Regulations, the EU Dual Use Regulation, and/or other export control regulations. As a condition of employment, the successful candidate will be required to provide proof of citizenship, U.S. permanent residency, or other protected status (e.g., under 8 U.S.C. § 1324b(a)(3)) for assessment of eligibility to access the export-controlled information. To meet this legal requirement, and as a condition of employment, the successful candidate’s citizenship will be verified with a valid passport. Lawful permanent residents, refugees, and asylees may verify status using other documents, where applicable.
  • This position requires verification of citizenship due to citizenship-based legal restrictions. Specifically, this position supports United States federal, state, and/or local government agency customers and is subject to certain citizenship-based restrictions where required or permitted by applicable law. To meet this legal requirement, and as a condition of employment, the successful candidate’s citizenship will be verified with a valid passport., * Doctorate in Statistics, Mathematics, Computer Science, Computer Security, or related field OR Masters Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 3+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
  • OR Bachelors Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 5+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
  • OR equivalent experience.
  • Programming ability, particularly in Python, with experience building automation, scripts, or security tooling.
  • Familiarity with large language models, generative AI systems, coding models, AI agents, or model APIs.
  • Ability to follow experimental methodology, record results accurately, and document technical findings clearly.
  • Exposure to security labs, cyber ranges, capture-the-flag environments, or vulnerable applications.
  • Experience contributing to cyber-capability evaluations or model-evaluation work for AI or agentic systems.
  • Coursework, internship, competition, or project experience in exploit development, vulnerability research, penetration testing, application security, or red teaming.
  • Experience working with coding agents, tool-using models, or autonomous agent frameworks.
  • Experience building capture-the-flag challenges, vulnerable applications, or exploit-development targets.
  • Familiarity with PyRIT, adversarial-testing tools, or comparable model-evaluation frameworks.
  • Familiarity with containers, sandboxed execution, or cloud-based test infrastructure.
  • Hands-on cybersecurity experience with systems, networks, applications, vulnerability analysis, penetration testing, red teaming, or exploitation in authorized environments.
  • Experience working with large language models, generative AI systems, coding models, AI agents, model APIs, or model-evaluation frameworks.
  • Experience designing controlled experiments, defining success criteria, analyzing results, and documenting technical findings.
  • Familiarity with common evaluation risks, including contamination, task saturation, unreliable scoring, weak baselines, environmental leakage, and limited reproducibility.

MSSecurity, Analysis Skills, Application Programming Interface (API), Applications Security, Artificial Intelligence (AI), Artificial Intelligence (AI) Agents, Background Investigation, Cloud Computing, Comparative Analysis, Computer Programming, Computer Science, Computer Security, Documentation, Experiment Design, Home Automation, Import/Export, Internet Security, Legal, Local Government, Machine Tool, Mathematics, Microsoft Product Family, Microsoft Windows System Internals/Programming, Modeling Languages, Penetration Testing, Problem Solving Skills, Python Programming/Scripting Language, Record Keeping, Regulations, Regulatory Requirements, Scripting (Scripting Languages), Security Analysis, Software Development Lifecycle (SDLC), Statistics, Technical Strategy, Technical Writing, Telemetry, Test Tools, Threat Modeling, United States Citizen

Benefits & conditions

Everyone works differently and is motivated by different things. We also understand that there’s more to you than your job. That’s why we offer competitive pay and a wide assortment of benefits– to help you make the most of life at work and away from it.

About the company

The Cloud & AI organization accelerates Microsofts mission to secure digital technology platforms, devices, clouds, and AI systems across customers heterogeneous environments and Microsofts internal estate. Our culture is centered on a growth mindset, inspiring excellence, and helping teams and leaders bring their best every day.

Microsoft Red Team (MRT) emulates real-world advanced persistent threats against Microsoft, external customers, and frontier AI systems. As AI systems become increasingly capable at reasoning, coding, tool use, exploitation, and autonomous execution, understanding when and how those systems materially increase offensive cyber capability is an important component of Microsofts AI security mission., Make your mark on the world’s most used technologies. Develop the next hit mobile application. Pioneer a startup that could be the next big thing. At Microsoft, you choose your path.

Headquartered in Redmond, Washington, Microsoft is a top innovator in both the consumer and enterprise technology industry. Just a few of the many things our products do are unleash creativity, connect businesses, and make learning more fun. But our continued success is based on one thing: our employees. We hire amazing, talented people and give them the opportunities-and the tools-to succeed.

WHY MICROSOFT? As a Microsoft employee, you’re surrounded by a diverse group of the smartest people in your field. This fosters new ideas, better business results, and creates a dynamic work environment. In the office, you’re constantly challenged and supported by your colleagues. Every day holds something new and exciting.

We also offer unparalleled depth and breadth of career opportunities. As an industry leader in multiple fields, working for Microsoft means being able to do whatever you feel passionate about-and being able to make an impact in that field. From day one, we give our employees significant responsibility. This means that you’ll know that you directly contributed to something that has a positive impact on people worldwide. Whether you choose to work in management, dive deep into the newest technology, or explore multiple professions, you’ll find everything you need at Microsoft to drive your career-and to make a difference.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

1:48 min

Automating exploratory data analysis within training pipelines

Dora Petrella · World Congress 2023

2:37 min

Tracing the evolution from early AI to generative AI

Mike Mike · World Congress 2025

6:10 min

Tech headline trivia on artificial intelligence and cybersecurity

Andrew MacLean Andrew MacLean +2 · LIVE

1:53 min

Evaluating traditional scripting languages for modern development tasks

Jens Knipper Jens Knipper · Europe 2026 Virtual

1:20 min

Utilizing industry threat models for AI security

Balázs Kiss · World Congress 2023

Videos

See all

Related articles

See all