GenAI Safety Analyst

Reposted 15 Hours Ago
Be an Early Applicant
Hiring Remotely in New York, NY, USA
In-Office or Remote
80K-87K Annually
Mid level
Security • Software • Generative AI
Trust, safety, and security for the GenAI era 🛡️
The Role
Analyze content infringements and write adversarial prompts to find model vulnerabilities across LLMs, text-to-image/video, and agents. Manage datasets and projects end-to-end, investigate evasion tactics, collaborate with engineering, product, and policy teams, and promote knowledge sharing to improve model safety.
Summary Generated by Built In
Description

Alice is seeking a driven, detail-focused professional to become a vital part of our team as a GenAI Safety Analyst. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.

Your tasks will involve writing adversarial prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs.

Responsibilities:

  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.
  • Managing projects end-to-end, from initial planning and oversight through quality assurance to final delivery.
  • Handling extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.
  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.
  • Working alongside diverse teams, engineering, product, policy, to tackle new challenges and craft forward-thinking strategies and resolutions.
  • Promoting a culture of knowledge exchange and continual learning within the team.
Requirements

Must have:

  • Background in AI Safety and/or Responsible AI and/or Trust and Safety 
  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.
  • Command of English at a near-native level.
  • Attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently.

Additional Wants:

  • Experience with various model types (Text-to-Text, Text-to-Image) is desirable.
  • Prior experience with OSINT (Open Source Intelligence) will be considered an asset.
  • A self-starter attitude, with the energy to excel in a fast-moving and variable environment.

The salary range for this role is $80K - $87K - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact—whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Skills Required

  • Background in AI Safety and/or Responsible AI and/or Trust and Safety
  • Familiarity with recent Generative AI models and agents
  • Near-native command of English
  • Strong attention to detail, organizational skills, and ability to manage multiple tasks
  • Experience with various model types (Text-to-Text, Text-to-Image)
  • Prior experience with OSINT (Open Source Intelligence)
  • Self-starter attitude and ability to work in fast-moving environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ramat Gan
413 Employees

What We Do

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact - whether with each other or with machines. In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection. Alice represents the next chapter of our growth and the natural evolution of ActiveFence, our industry-leading solution for UGC safety, as we expand our mission to secure the future of AI. Advance unafraid: alice.io

Similar Jobs

Remote or Hybrid
US
15100 Employees
135K-188K Annually

Wipfli Logo Wipfli

Architect

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
117K-158K Annually

Bounteous Logo Bounteous

Engagement Lead - Enterprise Data Warehouse

Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Remote
United States
5000 Employees

Liberty Mutual Insurance Logo Liberty Mutual Insurance

Inside Sales Representative

Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Remote or Hybrid
11 Locations
40000 Employees
45K-85K Annually

Similar Companies Hiring

Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
200 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account