Whose Words Are These? Forensic Linguistics, Authorship Evidence, and the Courts in the Age of Large Language Models
DOI:
https://doi.org/10.54878/5s9t9f87Keywords:
forensic linguistics, authorship attribution, large language models, AI-generated text, stylometry, expert evidence, admissibilityAbstract
For half a century, forensic linguistics has rested on a quietly powerful premise: that people leave themselves in their language. Habits of lexis, syntax, spelling, and discourse accumulate into an idiolect distinctive enough, in favourable conditions, to link a questioned text to its author, and courts in many jurisdictions have admitted analysis built on that premise in cases ranging from threatening letters to disputed confessions. Large language models unsettle the premise at its root. When any literate person can produce fluent text through a model, when a coerced author can be assisted, imitated, or replaced by a machine, and when the tools marketed to detect machine text prove fragile under paraphrase and biased against non-native writers of English, the evidential status of the written word requires re- examination. This paper conducts that re-examination across the three disciplines the problem now spans. From linguistics, we assess which components of idiolect survive machine mediation and which dissolve. From computer science, we review the technical state of AI- text detection, watermarking, and stylometric attribution under adversarial conditions, and formalise the underlying inference problem. From literary and interpretive studies, we ask what authorship itself now means when composition is distributed between human intention and machine articulation, and what that shift does to legal categories, confession, threat, defamation, plagiarism, built on unitary authorship. We argue that forensic authorship analysis is not obsolete but must be rebuilt on likelihood-ratio reporting, validated error rates, and explicit uncertainty about machine mediation, and that courts should treat current AI-text detectors as investigative leads rather than evidence. Recommendations for practitioners, courts, and researchers follow.
References
Alnajjar, M. (2024). Shifting trends shaping content creation in media and newsrooms. Emirati
Journal of Digital Art & Media, 2(1). Emirates Scholar Center for Research and Studies.
Bender, E. M., Gebru, T., McMillan-Major, A., & Shmitchell, S. (2021). On the dangers of stochastic
parrots: Can language models be too big? Proceedings of the 2021 ACM Conference on
Fairness, Accountability, and Transparency (FAccT '21), 610–623.
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam,
P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R.,
Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., … Amodei, D. (2020). Language models are
few-shot learners. Advances in Neural Information Processing Systems, 33, 1877–1901.
Chaski, C. E. (2005). Who's at the keyboard? Authorship attribution in digital evidence investigations.
International Journal of Digital Evidence, 4(1).
Coulthard, M., Johnson, A., & Wright, D. (2017). An introduction to forensic linguistics: Language in
evidence (2nd ed.). Routledge.
Daubert v. Merrell Dow Pharmaceuticals, Inc., 509 U.S. 579 (1993).
Grant, T. (2007). Quantifying evidence in forensic authorship analysis. International Journal of
Speech, Language and the Law, 14(1), 1–25.
Juola, P. (2006). Authorship attribution. Foundations and Trends in Information Retrieval, 1(3),
233–334.
Kirchenbauer, J., Geiping, J., Wen, Y., Katz, J., Miers, I., & Goldstein, T. (2023). A watermark for
large language models. Proceedings of the 40th International Conference on Machine Learning
(ICML 2023), PMLR 202, 17061–17084.
Liang, W., Yuksekgonul, M., Mao, Y., Wu, E., & Zou, J. (2023). GPT detectors are biased against
non-native English writers. Patterns, 4(7), 100779. https://doi.org/10.1016/j.patter.2023.100779
Mitchell, E., Lee, Y., Khazatsky, A., Manning, C. D., & Finn, C. (2023). DetectGPT: Zero-shot
machine-generated text detection using probability curvature. Proceedings of the 40th
International Conference on Machine Learning (ICML 2023), PMLR 202, 24950–24962.
Rawwash, A. A. (2025). The role of honeybees in forensic evidence detection: Review of their
application in environmental and criminal investigations. Emirati Journal of Law and Policing
Studies, 1(1), 4–16. Emirates Scholar Center for Research and Studies.
https://doi.org/10.54878/wafkw950
Sadasivan, V. S., Kumar, A., Balasubramanian, S., Wang, W., & Feizi, S. (2023). Can AI-generated
text be reliably detected? arXiv preprint arXiv:2303.11156.
Solan, L. M., & Tiersma, P. M. (2005). Speaking of crime: The language of criminal justice.
University of Chicago Press.
Stamatatos, E. (2009). A survey of modern authorship attribution methods. Journal of the American
Society for Information Science and Technology, 60(3), 538–556.
https://doi.org/10.1002/asi.21001
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., &
Polosukhin, I. (2017). Attention is all you need. Advances in Neural Information Processing
Systems, 30, 5998–6008.