Principal AI Data & Evaluation Strategist
You will serve as a consultative expert and thought partner alongside program managers, providing strategic guidance during project discovery, solution design, data planning, evaluation-framework creation, quality-rubric development, and pilot execution. You'll work directly with stakeholders, engineering teams, researchers, and program leadership to define project requirements before execution begins. This is not an operational-delivery role; you'll bring subject-matter expertise, challenge assumptions, identify risks early, and ensure projects are designed for quality, scalability, and measurable outcomes. It is a remote, flexible engagement contributing to frontier AI work.
Responsibilities
- Lead project discovery and translate business goals into AI data requirements
- Define project success metrics and identify risks and failure modes early
- Design training and evaluation datasets and define taxonomy structures
- Recommend data-sourcing methodologies and establish ground-truth standards
- Identify coverage gaps and edge cases
- Develop human-evaluation frameworks and quality-measurement methodologies
- Establish acceptance criteria and design benchmarking approaches
- Build calibration mechanisms
- Challenge assumptions that may impact project quality
- Recommend industry best practices and guide teams through ambiguity
- Support executive and client reviews
Requirements
- 10+ years in AI, data science, machine learning, research, or AI operations
- Experience designing large-scale AI data programs
- Experience with LLM evaluation, human feedback systems, AI benchmarking, and dataset development
- Strong written English and the ability to communicate complex ideas clearly
- Comfortable working independently in a remote environment
- Based in the United States