We are looking for a data scientist with 8+ years experience (or 3+ with relevant PHD). In this role you will:
++ Design and build the systems that generate longitudinally coherent synthetic environments for agent training and evaluation, including persona modeling, task generators, and verifiable ground truth.
++ Build and maintain synthesis models that generate realistic replacement values at very large scale, preserving format, statistical distribution, and semantic consistency so de-identified data stays useful downstream.
++ Train and improve the NER models behind our entity detection, driving accuracy and recall across free text, structured fields, and mixed enterprise data at scale.
++ Build evaluation infrastructure that grades agent outcomes, not just traces, and produces real discrimination between frontier models on real tasks.
++ Fine-tune and evaluate open-weight models on Tonic-generated data, and turn benchmark results into product and research direction.
++ Expand coverage into new domains, languages, and entity types, and handle the long tail of formats and edge cases that real customer data throws off.
++ Own model evaluation across the board: precision and recall on detection, utility preservation on synthesis, and outcome-level grading for agents.
++ Optimize inference so models run efficiently on large volumes of sensitive data inside customer environments.
++ Partner directly with frontier labs and enterprise ML team to turn hard data problems into shipped model improvements.
++ Set technical direction for a small, senior team and raise the bar on rigor, reproducibility, and shipping.
Apply here: https://jobs.ashbyhq.com/TonicAI/048a114d-fb5f-46ef-b0ff-b62... but also shoot me an email at adam + (company domain name).