United States
2 months ago

Job Overview

Job Type
Full Time
Pay
USD 80,000 – 87,000 (annual)

Job description

Salary:
USD 80,000 - 87,000 per year
Location:
United States
Work arrangement:
On-site

Role Summary

Description

Alice is seeking a driven, detail-focused professional to become a vital part of our team as a GenAI Safety Analyst. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.

Your tasks will involve writing adversarial prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs.

Additional Wants

Experience with various model types (Text-to-Text, Text-to-Image) is desirable.

Prior experience with OSINT (Open Source Intelligence) will be considered an asset.

A self-starter attitude, with the energy to excel in a fast-moving and variable environment.

The salary range for this role is $80K - $87K OTE - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact—whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Responsibilities

  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.
  • Managing projects end-to-end, from initial planning and oversight through quality assurance to final delivery.
  • Handling extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.
  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.
  • Working alongside diverse teams, engineering, product, policy, to tackle new challenges and craft forward-thinking strategies and resolutions.
  • Promoting a culture of knowledge exchange and continual learning within the team.

Requirements

  • Background in AI Safety and/or Responsible AI and/or Trust and Safety
  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.
  • Command of English at a near-native level.
  • Attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently.
Role:
GenAI Safety Analyst
Job Type:
Full Time

Company profile

ActiveFence

alice.io

Alice (ActiveFence) is an Israeli-American AI safety and security company, formerly known as ActiveFence. It helps enterprises and foundation model teams secure AI apps and agents with red teaming, adversarial intelligence and runtime guardrails built on real-world adversarial data. The company also helps technology companies track disinformation, hate speech, fraud and child abuse across user-generated content platforms and generative AI systems. Clients include NVIDIA, Amazon, TikTok and Cohere.

Headquarters
Tel Aviv, Israel
Founded
2018
Founders
Noam Schwartz, Alon Porat, Eyal Dykan, Iftach Orr

More jobs at ActiveFence

Similar Intelligence jobs at other companies