ActiveFence

GenAI Safety Tech Lead

ActiveFence$105K — $115K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience in AI Safety, Responsible AI, or Trust and Safety.
  • Proven ability to manage projects from start to finish, including planning, execution, and quality assurance.
  • Experience leading projects with multiple stakeholders in dynamic environments.
  • Strong client interaction skills with experience in relationship management.
  • Familiarity with Generative AI models; technical expertise not mandatory.
  • Exceptional attention to detail and organizational skills.

Responsibilities

  • Develop prompt strategies to uncover vulnerabilities in AI models.
  • Oversee projects from planning to delivery, ensuring accountability for results.
  • Mentor junior analysts to foster knowledge sharing and learning.
  • Manage large datasets with precision across multiple languages.
  • Investigate new methods for bypassing safety measures in foundational models.
  • Collaborate with engineering, product, and policy teams on innovative strategies.

Benefits

  • Opportunity to work on cutting-edge Generative AI technology.
  • Collaborative environment with experts across various fields.
  • Mentoring opportunities to encourage team development.
  • Exposure to diverse challenges in AI safety.
  • Potential for leadership role in a fast-paced setting.
Full Job Description
Description

Alice is seeking a driven, detail-focused professional to take on a leading role in our team as a GenAI Safety Tech Lead. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools, while owning the end-to-end delivery of engagements from planning through execution and quality assurance. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.

Your tasks will involve writing adversarial prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs, bringing a proven track record of managing complex projects from the ground up.

Responsibilities:

  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.
  • Owning projects end-to-end - from initial planning and execution through quality assurance to final delivery - with accountability for outcomes.
  • Mentoring junior analysts and promoting a culture of knowledge exchange and continual learning within the team.
  • Managing extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.
  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.
  • Partnering with cross-functional teams - engineering, product, policy - to tackle new challenges and craft forward-thinking strategies and resolutions.

Requirements

Must have:

  • 5+ years of experience in AI Safety and/or Responsible AI and/or Trust and Safety.
  • Proven track record of managing projects end-to-end, from planning and execution through quality assurance to final delivery.
  • Experience leading projects with multiple stakeholders in fast-paced, variable environments.
  • Experience working directly with clients, including attending client meetings and helping manage those relationships day to day.
  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.
  • Strong attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently.

Nice to Have:

  • Proven track record of research in academia or at a research institute.
  • Experience with various model types (Text-to-Text, Text-to-Image) is desirable.
  • Prior experience with OSINT (Open Source Intelligence) will be considered an asset.
  • Experience mentoring or leading junior team members.
  • A self-starter attitude, with the energy to excel in a fast-moving and variable environment.


The salary range for this role is $105K - $115K OTE - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About ActiveFence

ActiveFence is an Israeli cybersecurity company that specializes in detecting and mitigating online threats. The company's platform uses advanced AI and machine learning algorithms to monitor online activity and identify potential threats to businesses and organizations. ActiveFence's technology is used by a variety of clients, including financial institutions, social media companies, and government agencies. The company has raised significant funding from investors and is rapidly expanding its customer base.
Learn more about ActiveFence
Size
50 employees
Industry
Founded
2018

Similar Jobs

More Jobs at ActiveFence

  • ActiveFence
    GenAI Safety Team Lead
    $127K — $140K *
    New York, NY 10025 (New York County)
    Information Technology
    In-Person
  • ActiveFence
    GenAI Safety Tech Lead
    $105K — $115K *
    New York, NY 10025 (New York County)
    Information Technology
    In-Person

More Information Technology Jobs

Find similar GenAI Safety Tech Lead jobs: