Newcode.ai
Head of Evaluations (Legal AI Benchmarking)
New York, New York, United States
Sponsorship not specifiedDetected 7 days ago
PythonMachine LearningNumPyStatisticsComplianceContract ManagementResearch
About the role
- Note: We believe in being transparent about what it's like to work at Newcode.
Responsibilities
- Design Legal Benchmarks for: Contract Drafting, Information Extraction, Legal Research, and Contract Review Build, source and maintain relevant datasets Audit AI Output: Review and score complex AI-generated legal text, contract analyses, and statutory interpretations for accuracy and precision and lay out a strategy.
- Collaborate with Engineering: Partner directly with Engineering to translate legal errors into actionable technical feedback for model fine-tuning.
Requirements
- Establish clear criteria for grading model performance, specifically focusing on logical reasoning, citation accuracy, and the model's ability to safely abstain from answering.
Skills
- Proven ability to break down complex statutory frameworks and case law into structured, logical data points.
- Tech-Savviness: python, panda, numpy, jupiter notebooks and similar statistical models
- That means not every process, playbook, or framework is already in place, and priorities can shift quickly.
- The people who thrive here are comfortable with ambiguity, take ownership, and don't wait for perfect direction.
- Keep reading to learn more.
Benefits
- PhD or Masters in statistics, mathematics, machine learning or equivalent Analytical Skills: Proven ability to break down complex statutory frameworks and case law into structured, logical data points.
Apply directly at Newcode.ai →Create a free account for alerts like thisView Newcode.ai immigration profile
This listing is sourced directly from Newcode.ai's careers page and normalized into a canonical job model.