An expert-reasoning data platform for a frontier AI research lab
Client: an applied AI research lab, United States. Name withheld under NDA.
- 4data product lines launched
- 12 wkfrom brief to first paid delivery
- 3×contributor throughput after workflow redesign
Situation
The lab had a strong thesis: models trained on expert reasoning improve where models trained on outputs plateau. What it lacked was the machinery to capture that reasoning from professionals at scale, grade it consistently and package it for customers building foundation models.
What we built
A contributor platform with domain-specific task templates for supervised fine-tuning pairs, rubric-graded reinforcement learning prompts, tool-based agent environments and recorded computer-use trajectories. A two-tier review pipeline with automated consistency checks, and a delivery layer that exported datasets in each customer's schema.
Outcome
Four product lines went live in twelve weeks. Reviewer agreement rose above the customer's acceptance threshold in the second sprint, and contributor throughput roughly tripled once the task UI stopped fighting the experts. The platform now carries the lab's enterprise deliveries.
