LLM Evaluation post training
Employer not named by the sourceRemote
Frontier is not the employer and does not collect applications.
About this role
Machine Learning (ML), AI Model Development, AI Research, Large Language Models (LLMs), LLM Fine-Tuning · Senior ML Engineer / Advisor / Technical co-founder: Model Post-Training & Alignment
About Bentham Research Grounding Machine Intelligence in the Humanities
Bentham is an applied research lab working at the intersection of AI and the humanities. We build doctorate-authored, peer-reviewed evaluation instruments, training environments, and datasets for frontier AI labs, enterprises, and governments. Our core focus spans ethics, moral reasoning, philosophy, political theory, law, theology, and history. We believe expanding model capabilities requires anchoring machine judgment in human wisdom through bottom-up, scholar-led workflows paired with AI-in-the-loop validation.
The Role We are seeking an ML Research Engineer or Technical Advisor with hands-on experience in post-training models at a major frontier AI lab. You will bridge our team of PhD humanities scholars and technical alignment workflows, moving Bentham from core methodology pressure-testing into active project execution. You will help design, build, and validate the pipelines that translate complex humanities rubrics into high-yield evaluation benchmarks and post-training datasets.
Key Responsibilities
Technical Archi