The AI Safety Research & Testing Professional will play a key technical role in advancing AI safety research and evaluation capabilities. The role involves designing and conducting structured, code-based evaluations of publicly accessible AI models and AI-enabled products, and developing Python-based evaluation pipelines to support these activities.
The Professional will conduct hands-on technical testing, adversarial evaluation, failure analysis, and comparative benchmarking, while monitoring emerging AI safety and security research. The role will translate empirical findings, published frameworks, and established standards into practical guidance, readiness assessments, and recommendations for industry members and partners.
This is a hands-on, technically focused research and testing role covering model APIs, open-source evaluation frameworks, evaluation datasets, automated testing, RAG, tool use, agents, and other AI-enabled application components.
Research & Landscape Analysis
Technical Testing & Evaluation
Evaluation Methodology & Failure Analysis
Automation, Regression & Mitigation Testing
Advisory, Training & Knowledge Sharing