Newcode.ai

Newcode.ai

Head of Evaluations (Legal AI Benchmarking)

New York, New York, United States

Sponsorship not specifiedDetected 7 days ago
PythonMachine LearningNumPyStatisticsComplianceContract ManagementResearch

About the role

  • Note: We believe in being transparent about what it's like to work at Newcode.

Responsibilities

  • Design Legal Benchmarks for: Contract Drafting, Information Extraction, Legal Research, and Contract Review Build, source and maintain relevant datasets Audit AI Output: Review and score complex AI-generated legal text, contract analyses, and statutory interpretations for accuracy and precision and lay out a strategy.
  • Collaborate with Engineering: Partner directly with Engineering to translate legal errors into actionable technical feedback for model fine-tuning.

Requirements

  • Establish clear criteria for grading model performance, specifically focusing on logical reasoning, citation accuracy, and the model's ability to safely abstain from answering.

Skills

  • Proven ability to break down complex statutory frameworks and case law into structured, logical data points.
  • Tech-Savviness: python, panda, numpy, jupiter notebooks and similar statistical models
  • That means not every process, playbook, or framework is already in place, and priorities can shift quickly.
  • The people who thrive here are comfortable with ambiguity, take ownership, and don't wait for perfect direction.
  • Keep reading to learn more.

Benefits

  • PhD or Masters in statistics, mathematics, machine learning or equivalent Analytical Skills: Proven ability to break down complex statutory frameworks and case law into structured, logical data points.

This listing is sourced directly from Newcode.ai's careers page and normalized into a canonical job model.