White Circle, an AI Safety company, seeks a Research Engineer to own our internal benchmarks for single/multi-turn content guardrails and agent safety. You will build evals for flagship models and extend benchmarks as product data evolves, collaborating with the product team and research groups.
You will craft synthetic data, reproduce published benchmarks, and maintain production-grade code with clean abstractions. Office in Paris with hybrid setup; relocation support may apply.
#J-18808-Ljbffr