Skip to content

v2.0.0

Choose a tag to compare

@SAY-5 SAY-5 released this 14 Sep 22:09
· 43 commits to main since this release

spoofline sweep repeats the whole train, calibrate, fuse and evaluate protocol for all 16 leave two families out splits, each pairing one held out video family with one held out audio family, across several seeds.
It reports the mean, the sample standard deviation and a percentile bootstrap 95 percent interval of the mean for precision, recall, F1, EER and AUC per detector, and caches every run's raw logits so the evaluation is recomputed without retraining.
The variance table in the README comes from a documented reduced profile of 960 clips with 8 frames and 1 second of audio and 8 video and 6 audio epochs, which finished 48 runs over 3 seeds in 1125 seconds on a 10 core CPU.
Over those runs fused precision on unseen families averages 0.944 with a standard deviation of 0.048, against 0.985 on seen families.
Fused precision still sits 0.038 below the better single stream of the same run on average, while fused recall on unseen families is 0.566 against 0.412 for video and 0.268 for audio.