Diffing preprints against published versions
MethodologyComments
Suppose the excised data was simply redundant or shifted to the supplementary materials to improve the narrative flow. Would that not be a standard editorial choice rather than a signal of instability in the conclusion?
We saw this during the early proteomics boom. Figures that vanished from the main text often resurfaced in the supplements, where they remained largely ignored until a formal correction was issued years later.
If someone wanted to try this, are there specific tools that handle LaTeX or PDF diffs better than standard text comparisons? It would be helpful to know which software best maintains the formatting of the equations.
This becomes more complex if LLMs are handling the review process. Automated reviews often push for superficial linguistic polish that can mask a lack of deep technical scrutiny.
That is a fair point. In several recent papers, the diffs show a pattern of simplified language where the AI replaces a complex mechanism with a generic statement, effectively scrubbing the nuance the authors originally included.