Intern for AI Security RL Gyms

Reposted 7 Days Ago
Be an Early Applicant
Hanoi, VNM
Hybrid
Internship
Security • Software • Generative AI
Trust, safety, and security for the GenAI era 🛡️
The Role
Create adversarial prompt-injection tasks in YAML for RL Gym environments: define agent personas, simulated environments, adversarial payloads, benign user requests, run tasks against frontier models, validate success rates, and submit passing tasks with deterministic evaluation criteria.
Summary Generated by Built In
Description

Alice builds adversarial evaluation infrastructure used by the world's leading AI labs to stress-test their most capable agents before deployment. We're hiring interns for our RL Gym environments: adversarial training grounds where AI agents face prompt injection, social engineering, and data exfiltration scenarios.

The Role: You design prompt injection scenarios in YAML, run them against frontier models, validate success rates, and submit passing tasks.

The Task

You build adversarial prompt injection tasks for Alice's RL Gym platform. Each task is a self-contained YAML scenario simulating a realistic AI agent deployment, testing whether the agent can be manipulated into violating its safety policies.

What a task includes:

  1. An agent persona with a system prompt, security policies, and a defined set of tools.
  2. A simulated environment (e-commerce site, messaging platform, enterprise app) with content the agent interacts with.
  3. An adversarial payload embedded in the environment (in messages, product listings, tool responses, documents) that attempts to trick the agent into a policy violation.
  4. A benign user request that naturally leads the agent to encounter the payload.
  5. Deterministic evaluation criteria specifying exactly what constitutes success (agent resists) vs. failure (agent is compromised).
Requirements

What We're Looking For

  • Strong problem-solving skills and curiosity about AI security.
  • Interest in adversarial thinking and understanding how AI agents can be manipulated through prompt injection or other attack techniques.
  • Basic understanding of prompt injection concepts (or willingness to learn).
  • Comfortable writing structured content in YAML or able to learn quickly.
  • Familiar with using the command line (CLI); experience with Docker is a plus but not required.
  • Detail-oriented and able to follow technical guidelines consistently.
  • Good command of English.
  • Background in Computer Science, Cybersecurity, AI, Software Engineering, or related fields is preferred.

What We Offer

  • Hands-on experience working on one of the most cutting-edge AI safety and security projects.
  • Internship Allowance.
  • Mentorship from experienced AI security and red teaming professionals.
  • Opportunity to contribute to evaluation environments used by leading AI labs.

If you’re eager to learn, innovate, and grow in the field of data engineering, we’d love to hear from you. Apply today to be part of a team that values creativity and technical excellence!

Please note that only shortlisted candidates will be contacted.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.

If you're creative and driven to secure the future of AI, we want to hear from you!

Skills Required

  • Strong problem-solving skills and curiosity about AI security
  • Interest in adversarial thinking and prompt injection attack techniques
  • Basic understanding of prompt injection concepts or willingness to learn
  • Comfortable writing structured content in YAML or able to learn quickly
  • Familiar with using the command line (CLI)
  • Experience with Docker
  • Detail-oriented and able to follow technical guidelines consistently
  • Good command of English
  • Background in Computer Science, Cybersecurity, AI, Software Engineering, or related fields
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ramat Gan
413 Employees

What We Do

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact - whether with each other or with machines. In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection. Alice represents the next chapter of our growth and the natural evolution of ActiveFence, our industry-leading solution for UGC safety, as we expand our mission to secure the future of AI. Advance unafraid: alice.io

Similar Jobs

UL Solutions Logo UL Solutions

Intern, Chemical Safety Testing

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

UL Solutions Logo UL Solutions

Intern, Toy/Lab Tools Testing

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

UL Solutions Logo UL Solutions

Intern, Textile Testing (Softline)

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

Pfizer Logo Pfizer

Senior Health Representative (Oncology - North) - Temporary contract

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office
Hanoi, VNM
121990 Employees

Similar Companies Hiring

LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
200 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account