Research
Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era
arXiv:2604.08568v1 Announce Type: cross Abstract: The evolution of writing assistance tools from machine translation to large language models (LLMs) has changed how researchers write. This study inves
arXiv:2604.08568v1 Announce Type: cross Abstract: The evolution of writing assistance tools from machine translation to large language models (LLMs) has changed how researchers write. This study investigates whether this shift is homogenizing research papers by analyzing native language identification (NLI) trends in ACL Anthology papers across three eras: pre-neural network (NN), pre-LLM, and post-LLM. We construct a labeled dataset using a semi-automated framework and fine-tune a classifier to detect linguistic fingerprints of author backgrounds. Our analysis shows a consistent decline in NLI performance over time. Interestingly, the post-LLM era reveals anomalies: while Chinese and French show unexpected resistance or divergent trends, Japanese and Korean exhibit sharper-than-expected declines.
Related
- Daily and Weekly Periodicity in Large Language Model Performance and Its Implications for Research
- The Human Condition as Reflected in Contemporary Large Language Models
- Overstating Attitudes, Ignoring Networks: LLM Biases in Simulating Misinformation Susceptibility
- LLMs Underperform Graph-Based Parsers on Supervised Relation Extraction for Complex Graphs
Source: arXiv cs.AI | 2026-04-13