AI Security Engineer

Posted Yesterday
Hiring Remotely in United States
Remote
102K-261K Annually
Junior
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
The Role
Designs and operates controlled cyber-capability evaluations for frontier AI models and agentic systems. Builds vulnerable applications, cyber ranges, exploit-development targets, and evaluation harnesses; runs experiments; collects telemetry; analyzes model behavior and evaluation artifacts; and documents security findings. The role also develops sandboxing, containment, access, and secrets-management controls while collaborating with red-team operators, researchers, engineers, and Responsible AI stakeholders.
Summary Generated by Built In
Overview

The Cloud & AI organization accelerates Microsoft's mission to secure digital technology platforms, devices, clouds, and AI systems across customers' heterogeneous environments and Microsoft's internal estate. Our culture is centered on a growth mindset, inspiring excellence, and helping teams and leaders bring their best every day.

Microsoft Red Team (MRT) emulates real-world advanced persistent threats against Microsoft, external customers, and frontier AI systems. As AI systems become increasingly capable at reasoning, coding, tool use, exploitation, and autonomous execution, understanding when and how those systems materially increase offensive cyber capability is an important component of Microsoft's AI security mission.

MRT Strategic AI is seeking an AI Security Engineer focused on Cyber Capability Evaluation. This is a hands-on technical role for implementing and executing evaluations of frontier AI models and agentic systems in realistic, controlled cyber environments. The engineer will work at the intersection of offensive security, AI agents, model evaluation, and experimental engineering to build and maintain evaluation tasks, run experiments, analyze model behavior, and determine whether observed results represent meaningful cyber capability, an evaluation artifact, or a limitation in the test environment.

The role requires the ability to work independently within established technical direction, solve engineering and security problems, and contribute improvements to evaluation methods, cyber ranges, scoring, telemetry, and containment. The right candidate combines offensive-security experience with experimental discipline, programming skills, and an interest in emerging AI capabilities.


Responsibilities
  • Design, implement, and operate offensive cyber-capability evaluations for frontier, preview, production, and open-weight AI models.
  • Build and maintain realistic evaluation tasks covering vulnerability analysis, exploit development, application and system exploitation, attack-path reasoning, post-exploitation activities, and multi-step offensive workflows.
  • Execute controlled experiments using AI models; define success criteria and baselines; collect reliable telemetry; and document results.
  • Analyze model trajectories and investigate unexpected behavior to determine whether it reflects genuine capability, task leakage, environmental flaws, scoring errors, or other evaluation artifacts.
  • Develop and maintain cyber ranges, vulnerable applications, exploit-development targets, and evaluation harnesses, and apply security controls for sandboxing, secrets management, telemetry, access, and containment when testing autonomous AI systems.
  • Partner with red-team operators, researchers, engineers, and Responsible AI stakeholders to translate evaluation findings into security insights, and communicate results through reports, documentation, and briefings.

Qualifications

Minimum Qualifications:

  • Master's Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 1+ year(s) experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
    • OR Bachelor's Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 2+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
    • OR equivalent experience.

Other Requirements:

Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings:

Microsoft Cloud Background Check:

  • This position will be required to pass the Microsoft background and Microsoft Cloud background check upon hire/transfer and every two years thereafter.
  • This role will require access to information that is controlled for export under export control regulations, potentially under the U.S. International Traffic in Arms Regulations or Export Administration Regulations, the EU Dual Use Regulation, and/or other export control regulations. As a condition of employment, the successful candidate will be required to provide proof of citizenship, U.S. permanent residency, or other protected status (e.g., under 8 U.S.C. § 1324b(a)(3)) for assessment of eligibility to access the export-controlled information.  To meet this legal requirement, and as a condition of employment, the successful candidate’s citizenship will be verified with a valid passport. Lawful permanent residents, refugees, and asylees may verify status using other documents, where applicable. 
  • This position requires verification of citizenship due to citizenship-based legal restrictions. Specifically, this position supports United States federal, state, and/or local government agency customers and is subject to certain citizenship-based restrictions where required or permitted by applicable law. To meet this legal requirement, and as a condition of employment, the successful candidate’s citizenship will be verified with a valid passport. 

Preferred Qualifications:

  • Doctorate in Statistics, Mathematics, Computer Science, Computer Security, or related field OR Master's Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 3+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
    • OR Bachelor's Degree in Statistics, Mathematics, Computer Science, Computer Security, or related field AND 5+ years experience in software development lifecycle, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, and/or anomaly detection.
    • OR equivalent experience.
  • Programming ability, particularly in Python, with experience building automation, scripts, or security tooling.
  • Familiarity with large language models, generative AI systems, coding models, AI agents, or model APIs.
  • Ability to follow experimental methodology, record results accurately, and document technical findings clearly.
  • Exposure to security labs, cyber ranges, capture-the-flag environments, or vulnerable applications.
  • Experience contributing to cyber-capability evaluations or model-evaluation work for AI or agentic systems.
  • Coursework, internship, competition, or project experience in exploit development, vulnerability research, penetration testing, application security, or red teaming.
  • Experience working with coding agents, tool-using models, or autonomous agent frameworks.
  • Experience building capture-the-flag challenges, vulnerable applications, or exploit-development targets.
  • Familiarity with PyRIT, adversarial-testing tools, or comparable model-evaluation frameworks.
  • Familiarity with containers, sandboxed execution, or cloud-based test infrastructure.
  • Hands-on cybersecurity experience with systems, networks, applications, vulnerability analysis, penetration testing, red teaming, or exploitation in authorized environments.
  • Experience working with large language models, generative AI systems, coding models, AI agents, model APIs, or model-evaluation frameworks.
  • Experience designing controlled experiments, defining success criteria, analyzing results, and documenting technical findings.
  • Familiarity with common evaluation risks, including contamination, task saturation, unreliable scoring, weak baselines, environmental leakage, and limited reproducibility.

#MSSecurity


Security Research IC3 - The typical base pay range for this role across the U.S. is USD $102,100 - $202,200 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $133,800 - $219,200 per year.

Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay

Security Research IC4 - The typical base pay range for this role across the U.S. is USD $119,800 - $234,700 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $160,200 - $261,000 per year.

Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:
https://careers.microsoft.com/us/en/us-corporate-pay


This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.



Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.

Skills Required

  • Master's degree in Statistics, Mathematics, Computer Science, Computer Security, or a related field and 1+ year of experience in software development, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, or anomaly detection
  • Bachelor's degree in Statistics, Mathematics, Computer Science, Computer Security, or a related field and 2+ years of experience in software development, large-scale computing, threat analysis or modeling, cybersecurity, vulnerability research, or anomaly detection
  • Equivalent experience may substitute for the stated education and experience combinations
  • Ability to meet Microsoft, customer, and/or government security screening requirements
  • Eligibility to access export-controlled information through proof of citizenship, U.S. permanent residency, or other protected status
  • Doctorate in Statistics, Mathematics, Computer Science, Computer Security, or a related field
  • Master's degree in a relevant field and 3+ years of related experience
  • Bachelor's degree in a relevant field and 5+ years of related experience
  • Programming ability, particularly in Python, including automation, scripting, or security tooling
  • Familiarity with large language models, generative AI systems, coding models, AI agents, or model APIs
  • Ability to follow experimental methodology, record results accurately, and document technical findings clearly
  • Exposure to security labs, cyber ranges, capture-the-flag environments, or vulnerable applications
  • Experience contributing to cyber-capability evaluations or AI and agentic-system model evaluation
  • Coursework, internship, competition, or project experience in exploit development, vulnerability research, penetration testing, application security, or red teaming
  • Experience working with coding agents, tool-using models, or autonomous agent frameworks
  • Experience building capture-the-flag challenges, vulnerable applications, or exploit-development targets
  • Familiarity with PyRIT, adversarial-testing tools, or comparable model-evaluation frameworks
  • Familiarity with containers, sandboxed execution, or cloud-based test infrastructure
  • Hands-on cybersecurity experience with systems, networks, applications, vulnerability analysis, penetration testing, red teaming, or authorized exploitation
  • Experience designing controlled experiments, defining success criteria, analyzing results, and documenting technical findings
  • Familiarity with evaluation risks including contamination, task saturation, unreliable scoring, weak baselines, environmental leakage, and limited reproducibility

Microsoft Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Microsoft and has not been reviewed or approved by Microsoft.

  • Fair & Transparent Compensation Pay is presented as broadly competitive overall, with clear role/level/location variation and an emphasis on using posted ranges and band information for apples-to-apples comparisons.
  • Retirement Support Retirement benefits are described as a standout, highlighted by a strong 401(k) match structure and immediate vesting, plus additional plan features for tax-advantaged saving.
  • Parental & Family Support Family-oriented benefits are portrayed as a meaningful strength, with substantial paid parental leave and added supports like back-up care and adoption/surrogacy assistance.

Microsoft Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Redmond, WA
206,870 Employees
Year Founded: 1975

What We Do

At Microsoft, our mission is to empower every person and every organization on the planet to achieve more. Our mission is grounded in both the world in which we live and the future we strive to create. Today, we live in a mobile-first, cloud-first world, and the transformation we are driving across our businesses is designed to enable Microsoft and our customers to thrive in this world.

Similar Jobs

Affirm Logo Affirm

Security Engineer

Big Data • Fintech • Mobile • Payments • Financial Services
Easy Apply
Remote
United States
2200 Employees
204K-290K Annually

DigitalOcean Logo DigitalOcean

Principal Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
In-Office or Remote
Seattle, WA, USA
1400 Employees
230K-288K Annually

DigitalOcean Logo DigitalOcean

Principal Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
In-Office or Remote
San Francisco, CA, USA
1400 Employees
230K-288K Annually

DigitalOcean Logo DigitalOcean

Principal Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
In-Office or Remote
Denver, CO, USA
1400 Employees
230K-288K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account