Jacob Weiss

AI researcher

I work on post-training for AI agents. An agent in production fails in the same few ways every day, and a prompt edit is a guess. While catches the failure in traffic, trains the model on it, and proves the gain on a held-out set before anything ships. I co-founded it and I write most of the code.

As Head of AI at Silvia, I built its evals and judges, post-trained its specialist models, and wrote the character spec that shapes how it talks about money. People ask it hard questions about their taxes, mortgages and portfolios, and a wrong answer costs them. That is the pressure the method came out of.

I came to this from mathematics. I have three master’s degrees, from Johns Hopkins and Georgia Tech, and I spent the early agent years contributing to open-source frameworks. The part that was missing was the measurement: whether an agent actually got better, and by how much. That is the problem I work on now.

Everything below was learned by running the experiment. Each result has a recipe you can rerun, a held-out set, and an interval.

Research

Code

Education