0:38 · 971 kB · MP4 Download the video
Walkthrough 1
Inside the detector
Open the weights explorer, read the token ids that push each detector towards human or machine, switch between domains and learners, and see how little the two domains agree.
- 1Each detector is one learnt weight per token id: 5,000 weights and an intercept
- 2Logistic regression, domain 1: the token ids that push hardest towards human
- 3…and the token ids that push hardest towards machine
- 4Switch to domain 2: a different set of token ids leads on both sides
- 5The specimen sheet: estimator, chosen C, sparsity and the weight distribution
- 6SGD on domain 1: the same data, another learner, other leading tokens
- 7Cross-examination: the two domains' weights barely correlate
Transcript
Silent screen recording of inside the detector, captioned step by step. A highlighted circle shows the pointer.
- 0:00Each detector is one learnt weight per token id: 5,000 weights and an intercept
- 0:03Logistic regression, domain 1: the token ids that push hardest towards human
- 0:07…and the token ids that push hardest towards machine
- 0:11Switch to domain 2: a different set of token ids leads on both sides
- 0:16The specimen sheet: estimator, chosen C, sparsity and the weight distribution
- 0:22SGD on domain 1: the same data, another learner, other leading tokens
- 0:28Cross-examination: the two domains' weights barely correlate