Audit
Find out why your data is hurting the model.
Two weeks, fixed price
Open to consulting. Replies in a day.
Collection, cleaning, evaluation, and the operations around them.
How I help
Find out why your data is hurting the model.
Two weeks, fixed price
A pipeline end to end, collection through evaluation.
Six to twelve weeks
A second pair of eyes on a team already building.
From four hours a week
Budget capped before the first request.
Streaming, so a large corpus never has to fit in memory.
A cost ceiling on every run.
Every sample rated. 97.6% human approved.
Grouped by source page first, so no sample can cross the evaluation boundary.
JSONL, Parquet, or straight to the Hub.

Selected work
Research
831 samples at 97.6% approved, with splits that cannot leak by construction.DataForge: A Streaming, Leak-Aware Pipeline for Synthetic LLM Fine-Tuning Datasetsdoi:10.5281/zenodo.22906072The paper and four open questions

Process

About
I started by fine-tuning models and kept hitting the same wall: the model was fine, the data was not. Every interesting problem turned out to be upstream.
More about meThirty minutes, no preparation needed.
Book a call, opens in a new tab on Google CalendarOpens in a new tab on calendar.app.google, a Google page.
hello@iantoo.space