Alice (Formerly ActiveFence) Logo

Alice (Formerly ActiveFence)

GenAI Safety Analyst

Posted One Month Ago
Remote
Hiring Remotely in United States
80K-87K Annually
Entry level
Remote
Hiring Remotely in United States
80K-87K Annually
Entry level
Analyze generative AI content infringements and develop adversarial prompts to identify model vulnerabilities across hate speech, misinformation, intellectual property, and other abuse areas. Manage multilingual datasets, investigate safety circumvention tactics, oversee projects and quality assurance, and collaborate with engineering, product, and policy teams to improve AI safety strategies.
The summary above was generated by AI
Description

Alice is seeking a driven, detail-focused professional to become a vital part of our team as a GenAI Safety Analyst. In this role, you'll dive into the cutting-edge of technology, meticulously analyzing various content infringements to secure the new wave of Generative AI tools. Your duties will include collaborating with experts in diverse fields such as Hate Speech, Misinformation, Intellectual Property and Copyright, among others.

Your tasks will involve writing adversarial prompts to identify weaknesses in various AI models, including Large Language Models (LLMs), Text-to-Image, Text-to-Video, AI Agents and beyond. You'll also oversee data management to guarantee the highest quality of outputs.

Responsibilities:

  • Developing adversarial and risky prompt strategies across several areas of abuse to expose potential vulnerabilities in models.
  • Managing projects end-to-end, from initial planning and oversight through quality assurance to final delivery.
  • Handling extensive datasets across multiple languages and areas of abuse, ensuring precision and meticulous attention to detail.
  • Ongoing investigation into new tactics for circumventing foundational models' safety measures.
  • Working alongside diverse teams, engineering, product, policy, to tackle new challenges and craft forward-thinking strategies and resolutions.
  • Promoting a culture of knowledge exchange and continual learning within the team.
Requirements

Must have:

  • Background in AI Safety and/or Responsible AI and/or Trust and Safety 
  • Familiarity with recent Generative AI models and agents is essential, though direct technical experience is not a prerequisite.
  • Command of English at a near-native level.
  • Attention to detail, organizational capabilities, and the capacity to juggle numerous tasks concurrently.

Additional Wants:

  • Experience with various model types (Text-to-Text, Text-to-Image) is desirable.
  • Prior experience with OSINT (Open Source Intelligence) will be considered an asset.
  • A self-starter attitude, with the energy to excel in a fast-moving and variable environment.

The salary range for this role in the US is $80K - $87K - Range may vary based on experience. Salary at the time of offer will be commensurate with experience.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.

If you're creative and driven to secure the future of AI, we want to hear from you!

Similar Jobs

One Month Ago
Remote
USA
80K-87K Annually
Mid level
80K-87K Annually
Mid level
Security • Software • Generative AI
Analyze content infringements and write adversarial prompts to find model vulnerabilities across LLMs, text-to-image/video, and agents. Manage datasets and projects end-to-end, investigate evasion tactics, collaborate with engineering, product, and policy teams, and promote knowledge sharing to improve model safety.
Top Skills: Ai AgentsGenerative AiLarge Language Models (Llms)OsintText-To-ImageText-To-Video
20 Minutes Ago
Remote or Hybrid
United States
211K-285K Annually
Senior level
211K-285K Annually
Senior level
Artificial Intelligence • Cloud • Information Technology • Sales • Security • Software • Cybersecurity
Leads Rapid7’s global technical sales organization across Sales Engineering Specialists, Forward Engineering, Labs, and the Center of Excellence. Develops technical sales strategy, organizational roadmaps, succession frameworks, and engineering capabilities while partnering with sales leadership. Oversees integrations, demo infrastructure, enterprise platform architecture, ROI modeling, resource planning, and cross-functional execution. Drives organizational scaling, executive alignment, talent development, and measurable contributions to revenue and product innovation.
Top Skills: APIsCloud ComputingEnterprise Security PlatformsSecurity Orchestration Automation And Response (Soar)
20 Minutes Ago
Remote or Hybrid
United States
89K-121K Annually
Mid level
89K-121K Annually
Mid level
Artificial Intelligence • Cloud • Information Technology • Sales • Security • Software • Cybersecurity
Supports customers using Rapid7’s Vector Command Red Team service through attack surface analysis, reconnaissance, OSINT, reporting, customer communications, and request prioritization. The role coordinates with Red Team, MDR, and MVM teams; conducts manual network and service reconnaissance; develops scripts to analyze attack surface data; monitors customer exposures; and performs entry-level external penetration testing. It also includes customer onboarding, monthly reporting, update calls, and translating technical security concepts for nontechnical stakeholders.
Top Skills: Ieee 802.11Internet Protocol SuiteLinuxOsintPenetration Testing ToolsPowershellPythonUnixVector Command PlatformWindows

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account