A student research case study that built a leakage-aware benchmark to test whether a person's survey history could be compressed into model weights well enough to predict unseen answers.
What it shows: A good experiment should be able to stop an attractive idea. The useful result was not a digital twin; it was a benchmark, a failure diagnostic, and a better next decision.