Empresa: ActiveFence
Description
Alice is seeking a driven, detail-focused Senior Generative AI Researcher to take on a leading role within our US team. In this position, you will operate at the cutting edge of AI Safety and Trust & Safety, analyzing potential vulnerabilities and content safety risks across the newest wave of Generative AI tools.
As a Senior Researcher, you won't just run tests, you will design robust testing methodologies, act as a core content expert to support the Program Lead, and actively expand the team’s internal knowledge base. You will partner closely with cross-functional teams and external stakeholders to secure models across multiple modalities, including LLMs, Text-to-Image, Text-to-Video, and AI Agents.
Key Responsibilities
Methodology & Strategy
- Architect rigorous, scalable testing methodologies and red-teaming frameworks to evaluate foundational models, multimodal systems, and AI agents.
- Develop sophisticated prompt strategies across diverse risk domains (e.g., Hate Speech, Misinformation, IP & Copyright infringement, Child Safety) to expose complex model vulnerabilities.
- Conduct ongoing research into emerging jailbreak tactics, prompt injection techniques, and novel circumvention strategies used against foundational safety measures.
Subject-Matter Expertise
- Serve as a trusted content and domain expert, providing deep technical and policy insight to support the Program Lead in scoping projects, assessing risks, and driving strategy.
- Lead efforts to continuously document, synthesize, and expand Alice’s internal AI Safety knowledge base, standardizing best practices, taxonomies, and research findings across the team.
- Mentor junior analysts, foster a culture of continual learning, and elevate the team’s analytical standards.
Operational Excellence