Research Engineer: AI Safety Benchmarks & Evaluation
REMOTE
il y a 14 heures
White Circle is hiring a research engineer to own and maintain our internal benchmark suite for single and multi-turn content guardrails and agentic safety. You will work with the product team to ensure evals cover core functionality of flagship models and to adapt benchmarks to new features and data.
You will also study realistic agent behaviours in the wild and publish insights as part of a rigorous research program.
#J-18808-Ljbffr
Entreprise
White Circle
Plateforme de publication
WHATJOBS
Offres pouvant vous intéresser
PARIS, 75
il y a 20 jours
PARIS, 75
il y a 1 jour
REMOTE
il y a 1 jour
PARIS, 75
il y a 1 jour