Comparing LLM-generated and human-authored news text using formal syntactic theory
Olga Zamaraeva, Dan Flickinger, Francis Bond, Carlos Gómez-Rodríguez
Abstract
This study provides the first comprehensive comparison of New York Times-style text generated by six large language models against real, human-authored NYT writing. The comparison is based on a formal syntactic theory. We use Head-driven Phrase Structure Grammar (HPSG) to analyze the grammatical structure of the texts. We then investigate and illustrate the differences in the distributions of HPSG grammar types, revealing systematic distinctions between human and LLM-generated writing. These findings contribute to a deeper understanding of the syntactic behavior of LLMs as well as humans, within the NYT genre.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cc5fa530-d84d-489d-b7b9-9d07a7241170Cited by top-tier papers2
- More Aligned, Less Diverse? Analyzing the Grammar and Lexicon of Two Generations of LLMsAdrián Gude, Roi Santos-Rios, Francis Bond, Dan Flickinger et al.ACL 2026
- Learn-to-learn on Arbitrary Textual Conditioning: A Hypernetwork-Driven Meta-gated LLMLuo Ji, Qi Qin, Ningyuan Xi, Teng Chen et al.ICML 2026
Builds on1
Related papers
- Linguistic and Embedding-Based Profiling of Texts Generated by Humans and Large Language ModelsSergio E. Zanotto, Segun AroyehunEMNLP 2025 · 3 citations
- Threads of Subtlety: Detecting Machine-Generated Texts Through Discourse MotifsZae Myung Kim, Kwang Hee Lee, Preston Zhu, Vipul Raheja et al.ACL 2024 · 1 citation
- Beyond the Final Actor: Modeling the Dual Roles of Creator and Editor for Fine-Grained LLM-Generated Text DetectionYang Li, Qiang Sheng, Zhengjia Wang, Yehan Yang et al.ACL 2026
- Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language ModelsYanlin Wang, Tianyue Jiang, Mingwei Liu, Jiachi Chen et al.FSE 2025 · 6 citations
- Beyond Checkmate: Exploring the Creative Choke Points for AI Generated TextsNafis Irtiza Tripto, Saranya Venkatraman, Mahjabin Nahar, Dongwon LeeEMNLP 2025 · 1 citation
