Cyber Digital Trust & Online Safety Manager
Indexed description
Recruiting for this role ends on 12/31/3026.
Work you'll do
As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:
- Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
- Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
- Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
- Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
- Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
- Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.
- Ability to work independently and collaborate as part of a team
- Effective written and verbal communication skills
- Meticulous attention to detail and quality of work product
- Ability to build and sustain professional relationships
- Ability to lead projects or workstreams
- Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
- Strong interpersonal skills and professional demeanor
- Ability to meet deadlines
- Ability to mentor and provide clear guidance to others
Qualifications
Required:
- Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
- Demonstrated hands-on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
- Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
- Limited immigration sponsorship may be available.
- Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
- Experience designing and operationalizing trust and safety testing programs for large-scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
- Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements
You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.
#CyberDTP27
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search