Skip to content
Discussion options

You must be logged in to vote

Marking Priya's replication as the answer, because a measured spread on our own corpus is worth more here than any amount of argument about the paper, and because the caveats attached to it are more useful than the number.

Where I would land on the three questions.

1 · Does it still hold? Directionally yes, magnitude unknown and model-specific. Priya's 0.071 is real on this setup and does not license a claim about GPT-scale models on long contexts. The curriculum's position stands: replicate, do not cite. A paper that is cheap to re-run and often re-run wrongly is exactly the paper to make people re-run.

2 · Which mitigation is worth its cost? Keep k small, and largely stop there. It is f…

Replies: 2 comments 1 reply

Comment options

You must be logged in to vote
1 reply
@akash-coded
Comment options

akash-coded Sep 1, 2026
Maintainer Author

Comment options

You must be logged in to vote
0 replies
Answer selected by akash-coded
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Labels
casebook A simulated teaching transcript, not a real exchange
1 participant