Senior GenAI Safety Researcher

Reposted Yesterday
Be an Early Applicant
Hiring Remotely in USA
Remote
105K-115K Annually
Senior level
Security • Software • Generative AI
Trust, safety, and security for the GenAI era 🛡️
The Role
Lead AI safety role developing adversarial prompt strategies, owning end-to-end projects, mentoring junior analysts, managing multilingual abuse datasets, researching circumvention tactics, and partnering with engineering, product, and policy teams to secure generative AI models.
Summary Generated by Built In
Description

Alice is seeking a driven, detail-focused Senior Generative AI Researcher to take on a leading role within our US team. In this position, you will operate at the cutting edge of AI Safety and Trust & Safety, analyzing potential vulnerabilities and content safety risks across the newest wave of Generative AI tools.

As a Senior Researcher, you won't just run tests, you will design robust testing methodologies, act as a core content expert to support the Program Lead, and actively expand the team’s internal knowledge base. You will partner closely with cross-functional teams and external stakeholders to secure models across multiple modalities, including LLMs, Text-to-Image, Text-to-Video, and AI Agents.

Key Responsibilities

Methodology & Strategy

  • Architect rigorous, scalable testing methodologies and red-teaming frameworks to evaluate foundational models, multimodal systems, and AI agents.
  • Develop sophisticated prompt strategies across diverse risk domains (e.g., Hate Speech, Misinformation, IP & Copyright infringement, Child Safety) to expose complex model vulnerabilities.
  • Conduct ongoing research into emerging jailbreak tactics, prompt injection techniques, and novel circumvention strategies used against foundational safety measures.

Subject-Matter Expertise

  • Serve as a trusted content and domain expert, providing deep technical and policy insight to support the Program Lead in scoping projects, assessing risks, and driving strategy.
  • Lead efforts to continuously document, synthesize, and expand Alice’s internal AI Safety knowledge base, standardizing best practices, taxonomies, and research findings across the team.
  • Mentor junior analysts, foster a culture of continual learning, and elevate the team’s analytical standards.

Operational Excellence

  • Own engagement lifecycles from initial planning and methodology design through execution, quality assurance (QA), and final delivery.
  • Oversee complex, multi-language datasets across multiple areas of abuse, ensuring the highest precision, accuracy, and output quality.
  • Partner effectively with engineering, product, policy, and client-facing teams to communicate research findings and inform mitigation strategies.
Requirements

Must-Have

  • 5+ years of experience in AI Safety, Responsible AI, Trust & Safety, or aligned research domains.
  • Proven expertise in research design and building qualitative or quantitative evaluation methodologies for GenAI.
  • Strong domain expertise in content risks (e.g., toxicity, copyright, misinformation, safety policy violations).
  • Track record of project ownership, leading deliverables end-to-end with high attention to detail in fast-paced, variable environments.
  • Deep familiarity with modern Generative AI architectures, prompt engineering, red-teaming, and AI agents.
  • Strong communication skills to act as a core subject-matter contact for program leads, internal teams, and clients.

Nice-to-Have

  • Proven track record of published research in academia, industry whitepapers, or a research institute.
  • Hands-on experience evaluating multimodal systems (Text-to-Image, Text-to-Video, Audio).
  • Experience mentoring, leading, or QAing the work of junior analysts and researchers.

The salary range for this role in the US is $105K - $115K - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.

If you're creative and driven to secure the future of AI, we want to hear from you!

Skills Required

  • 5+ years of experience in AI Safety, Responsible AI, or Trust and Safety.
  • Proven track record of managing projects end-to-end from planning through QA to delivery.
  • Experience leading projects with multiple stakeholders in fast-paced, variable environments.
  • Experience working directly with clients, including attending client meetings and managing relationships.
  • Familiarity with recent Generative AI models and agents (LLMs, text-to-image, text-to-video); direct technical coding experience not required.
  • Strong attention to detail, organizational capabilities, and ability to manage multiple concurrent tasks.
  • Proven track record of research in academia or at a research institute.
  • Experience with various model types (Text-to-Text, Text-to-Image).
  • Prior experience with OSINT (Open Source Intelligence).
  • Experience mentoring or leading junior team members.
  • Self-starter attitude, able to excel in fast-moving, variable environments.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ramat Gan
413 Employees

What We Do

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact - whether with each other or with machines. In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection. Alice represents the next chapter of our growth and the natural evolution of ActiveFence, our industry-leading solution for UGC safety, as we expand our mission to secure the future of AI. Advance unafraid: alice.io

Similar Jobs

Alice (Formerly ActiveFence) Logo Alice (Formerly ActiveFence)

Senior GenAI Safety Researcher

Security • Software • Generative AI
Remote
United States
413 Employees
105K-115K Annually

Gratia Health, Inc. Logo Gratia Health, Inc.

Business Development Representative

Healthtech • HR Tech • Software • Analytics
In-Office or Remote
Austin, TX, USA
65K-85K Annually

Rula Logo Rula

Senior Analytics Engineer

Healthtech • Social Impact • Software • Telehealth
Remote
United States
620 Employees
152K-186K Annually

AlertMedia Logo AlertMedia

Social Media Contractor

Artificial Intelligence • Cloud • Information Technology • Security • Social Impact • Software
Easy Apply
Remote or Hybrid
2 Locations
450 Employees
25-25 Hourly

Similar Companies Hiring

Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
200 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account