Skip to main content
PLOS One logoLink to PLOS One
. 2026 Apr 24;21(4):e0346776. doi: 10.1371/journal.pone.0346776

A multi-dimensional analysis of native and non-native academic research articles in twelve disciplines

Jiaqi Deng 1,2, Ghayth Kamel Shaker Al-Shaibani 3,*
Editor: Xiaoming Tian4
PMCID: PMC13108751  PMID: 42030257

Abstract

Multi-dimensional analysis (MDA) approach has been widely adopted to compare linguistic and register variations between native and non-native researchers’ academic writing in numerous studies. However, only a few MDA studies have identified specific linguistic features for academic writing to develop an MDA framework suitable for this genre. To fill in this gap, this study identified 62 academic English linguistic features which were tagged with MDA tagger and counted by PatCount software programs to develop a novel MDA model for distinguishing native English and Chinese researchers’ research article writing. This yielded four dimensions: academic involvement and interaction vs. information density; interactive argumentation vs. static description; impersonal evaluation vs. personal opinion; and explicit elaborating style vs. simplified reporting style. It was found that native English researchers are more academically and interactionally involved; they are more interactive and personal, and they use an explicit elaborating style than Chinese researchers do. In disciplines of hard or pure sciences, native researchers are more strategic to freely express their author stance, but Chinese researchers tend to be conservative as they follow pre-set disciplinary conventions. These findings suggest that Chinese researchers should exhibit their authorial stance, interact with the readers with confidence and employ more interactive devices to make their writing coherent and explicit. Meanwhile, Chinese conciseness displays clarity and efficiency for native researchers prone to verbosity; Chinese impersonality offers objectivity where natives risk bias. This study also proposed a new method to tag linguistic features in MDA research so that researchers can utilize this method to develop novel MDA models based on their research objectives.

1. Introduction

Over the past 10 years, a considerable amount of research [16] on disciplinary variation between native English researchers and non-native researchers has been conducted. Understanding these differences helps non-native researchers meet the standards of English-speaking academic writing communities and enhances their writing skills to publish in reputable international journals [7]. Biber’s (1988) [8] multi-dimensional analysis (MDA) model, which is favored by many researchers, stands out for its indispensable role in analyzing linguistic variations since it involves both macroscopic and microscopic perspectives. The macroscopic perspective refers to an MDA approach that identifies textual dimensions after analyzing the co-occurring patterns of linguistic features in selected research articles, whereas the microscopic perspective refers to interpreting these dimensions in functional terms by observing a large quantity of linguistic features in a substantial number of texts simultaneously.

The MDA model is based on the assumption that strong co-occurrence patterns of linguistic features mark underlying functional dimensions. Linguistic features do not randomly co-occur in texts. If certain features consistently co-occur, it is reasonable to look for an underlying functional influence that encourages their use [8]. Our MDA model has been constructed through identifying, tagging, counting linguistic features, then conducting a factor analysis which was widely used in many studies to distinguish language variation. Additionally, when comparing the differences in academic writing, only few MDA studies have identified linguistic features in academic writing by only adopting Biber’s (1988) tagger [8] which was originally used to distinguish spoken and written linguistic features. Therefore, this study extends the previous studies in identifying linguistic features in academic writing, tagging and extracting these features through code writing and PatCount software, and eventually developing a novel MDA model. Furthermore, few previous studies have compared the linguistic variation in different disciplines between native English researchers and Chinese researchers via MDA model. In this study, a total of 2400 English research articles from 12 disciplines written by native English and Chinese researchers were collected to answer the following three research questions:

  • (i)

    What are the dimensions that may identify the lexical and grammatical differences between native English and Chinese researchers’ academic writing in 12 disciplines?

  • (ii)

    What are the differences in the academic writing of native English and Chinese researchers across the 12 disciplines along these dimensions?

  • (iii)

    To what extent do language background (native English researchers vs Chinese researchers) and discipline influence the dimension scores of native English and Chinese researchers’ academic writing?

2. Literature review

2.1 Discipline-based studies on research articles

Research articles are very common for scholars to communicate their findings to the academic community, adhering to specific disciplinary conventions and expectations of the target academic discipline in which they are written [9]. Writing differs across disciplines meanwhile different academic disciplines shape writing practices [10,11]. Exploring how academic writing varies across disciplines helps non-native writers know how to adapt to these conventions [12]. While reviewing the previous research, two main trends were found. One was to compare the differences in linguistic features such as lexical choices, syntactic complexity, hedging and stance, audience awareness and rhetorical structures among different academic disciplines [13,14], and the other is conducted in subfields within one discipline. For instance, how various subfields use language differently within the field of applied linguistics [15], how sociolinguistics and language documentation differ in approaching linguistic variation within linguistics [16]; and metadiscourse choice variation within the field of English for academic purposes [17].

Three main perspectives were considered in these investigations: the macro perspective, the micro perspective, and the multi-dimensional perspective. A comparison of linguistic variation between native and non-native speakers from a macro perspective has usually been conducted via functional analysis of various genres, namely, move analysis [18]; a micro perspective has focused mostly on single lexical features, such as critical stances and evaluations [19], abstracts [20], the use of first personal pronouns [21], and hedges [22]. In addition, the MDA model has attracted the researchers’ attention because it combines the advantages of the above-mentioned perspectives and has been applied to compare linguistic variation both macroscopically and microscopically. Studies on MDA are reviewed in detail in the following section.

2.2 MDA development and application

The origins of the MDA research can be dated back to the 1970s when the importance of the co-occurrence of linguistic features in comparative studies of registers attracted researchers’ attention. Carroll (1960) provided a methodological basis for MDA by employing a statistical analysis of linguistic co-occurring patterns and performing all linguistic analyses manually to conduct a study on vectors of prose style. After that, Douglas Biber improved MDA in 1984 in his PhD dissertation. Biber selected 23 categories of 481 texts covering the spoken and written registers from the Lancaster-Oslo-Bergen (LOB) Corpus of British English and the London-Lund Corpus (LLC) of Spoken English in the UK, and he identified 67 linguistic features. Through statistical methods as factor analysis, Biber performed a multi-dimensional comparative analysis of spoken and written English of the co-occurrence of these linguistic features to identify seven dimensions [8]. The MDA represents a groundbreaking method for linguistic variation. It allows researchers to handle large corpora to identify patterns and dimensions of variation across different texts and registers by analyzing the co-occurring patterns of linguistic features via computational techniques. This approach is instrumental in understanding how language varies systematically according to various situational contexts.

The early MDA model was used to investigate the variation between oral and written English [8], and was widely used for other different registers [23,24]. After that, studies applied this approach to examine register variation in other languages, such as in Somali [25]. MDA subsequently experienced diachronic language variation research [26,27], and dialect variation research [28,29].

Two main paradigms have been employed in many studies comparing linguistic variations using MDA. One used Biber’s (1988) [8] framework [30,31]. The other one developed a novel MDA suitable for one’s research domain [3234].

For these two paradigms, the one adopting Biber’s (1988) [8] framework conducted the study under Biber’s original framework with those linguistic features distinguishing spoken and written registers even in specialized disciplines. This approach would be less effective in analyzing variation in more restricted discourse domains. Thus, the other paradigm was commonly agreed to be more suitable especially for studies on a specialized discipline. However, using this paradigm requires selecting linguistic features for a specific discipline, tagging those features, counting the frequencies, and performing new factor analysis to identify new dimensions. It is tough for Applied Linguistics researchers with limited computer skills to tag those linguistic features related to their research, thus most studies use Biber’s (1988) tag [8] and Biber’s (2006) tag [35]. Biber’s (1988) tag [8] is not freely accessible online; however, Biber’s (1988) tag [8] was completely replicated by Nini [36] and can be accessed through multi-dimensional analysis tagger (MAT) programming for free. Biber’s (2006) tag [35] added additional semantic categories to Biber’s (1988) tag [8] which is not freely available either, but may be accessed by contacting Douglas Biber’s Corpus Linguistics Research Program at Northern Arizona University [37]. This study solved the difficulty of tagging specific linguistic features by using Nini’s MAT and PatCount computer programming through the writing of computer codes and developed a novel MDA model related to academic writing in articles.

3. Methods

This section presents our research design in terms of data collection and data analysis in the sub-sections below resulting in developing a novel MDA model.

3.1 Data collection

This research adopts Biber’s (1988) MDA development procedure: (1) the use of computer-based text corpora; (2) the use of computer programs to count the frequency of certain linguistic features in a large number of texts, enabling analysis of the distribution of many linguistic features across many texts; (3) the use of factor analysis to determine co-occurrence relationships among the linguistic features; (4) the use of microscopic analysis to interpret the functional parameters underlying the quantitatively identified co-occurrence patterns; and (5) a comparison of texts with respect to dimensions on the basis of computed factor scores [8].

To collect the research articles written by native English researchers and Chinese researchers in various disciplines, the disciplines to be studied in this research should be determined first. The criteria for selecting the disciplines followed the Chinese National System of Level One Disciplines for Degree Education, namely, 13 disciplines as agriculture, art, economics, history, law, literature, management science, medicine, natural science, education, philosophy, engineering and military. To make the native English researchers’ corpus and Chinese researchers’ corpus comparable, we selected the same disciplines and the same number of research articles for both corpora. When we checked native English researchers’ disciplines in the Web of Science (WoS), there was no military discipline. Therefore, the number of final disciplines involved in this study was 12 when the military discipline was excluded. After the 12 disciplines were selected, we adopted stratified random sampling to select the research articles (RAs) in these 12 disciplines [38]. First, we used Journal Citation Reports (JCR) of the Web of Science (WoS) to look for the most influential ten international Science Citation Index (SCI), Social Science Citation Index (SSCI) and Science Citation Index Expanded (SCIE) journals in each discipline to select RAs based on the highest impact factor. Five of those ten journals in each discipline were then randomly selected. To select native English researchers’ RAs, we chose the top international journals with the highest impact factor because they could be set as a benchmark. Twenty full-text articles from each journal for the past five years (2019–2023) were selected and included as data. The sample size for native English researchers’ RAs is 1200. For native researchers’ identification, it was not feasible for a study of this scale to directly contact all authors to confirm their L1 identity, thus we followed a widely adopted method in Applied Linguistics and English for Academic Purposes (EAP) research using a multi-tiered screening method of first author’s affiliated institution identification and name identification [33,3942]. First, the nation of the institution affiliated with which the paper is published was determined. Only institutional affiliation belongs to nations in the inner circle of English-speaking countries [43], specifically the United States, the United Kingdom, Canada, Australia, and New Zealand were included. Articles with which the nationality of the institution cannot be determined or do not belong to these five English-speaking nations were discarded. Then we determined whether the first author’s nationality was matched with that of their publishing institution, by using a name origin database [44]. If the ethnic origin of the first author’s surname was consistent with the nationality of the publishing institution, the author was regarded as originating from that nation. Articles that did not meet this criterion were excluded from the final corpus.

For the collection of Chinese researchers’ RAs, the RAs were all chosen randomly from academic English journals published in mainland China whose authors were Chinese and were affiliated with an institution in mainland China to reflect the Chinese researchers’ average English writing level. All the English academic journals published in China were chosen according to Bao and Zhang [45]. The number of journals and articles was the same as that in the native researchers’ corpus to ensure that the two corpora are balanced. The sample size was 1200 as well. These two corpora were named as the Native Researchers Corpus (NRC) and the Chinese Researchers Corpus (CRC). Both corpora comprised the same academic disciplines to avoid potential disciplinary biases in the selection process. Both corpora were collected from June 2024 to August 2024 and the total sample size of the research articles collected in this research is 2400. This sample size is sufficient for performing this research because in a factor analysis, the database should include five times as many texts as linguistic features to be analyzed [46]. In our study, 62 linguistic features were ultimately identified and included; thus, we have a sufficient 2400 articles as a sample for this study. The composition of the corpora including the total number of words for each discipline is shown in Table 1.

Table 1. Composition of corpora.

Discipline English native researchers Chinese researchers
No. of articles Tokens No. of articles Tokens
Agriculture 100 824200 100 557500
Art 100 516800 100 733600
Economics 100 1077500 100 675000
Engineering 100 729600 100 236400
History 100 1331400 100 793300
Law 100 718900 100 933800
Literature 100 800700 100 1010100
Management science 100
1015100

100

693000
Medicine 100 297200 100 362000
Natural
science

100

423000

100

287200
Education 100 1109700 100 776200
Philosophy 100 278700 100 745100
Total 1200 9122800 1200 7803200

Most of the collected RAs are in PDF format. ABBYY FineReader, Adobe Acrobat Pro and other software were used to convert PDF format files into TXT format files. Two rounds of manual verification were carried out on the converted text to clean up the problems such as garbled code, formatting and spelling errors in the conversion process. Moreover, all the examples, directly quoted paragraphs, interview excerpts in the RAs and other languages that were not written by the author(s) were deleted. The quoted languages in the RAs do not represent the real language level of Chinese and English native researchers have been removed because retaining them may affect the accuracy of the results. Data such as charts and chart captions, which cannot reflect linguistic characteristics in the text, were deleted as well.

3.2 Data analysis

Textual analysis is a method of transforming textual data into quantifiable forms, revealing patterns and trends through classification, coding, and statistical analysis of texts [47]. Therefore, this study used textual analysis to identify and tag the linguistic features suitable for academic writing in the collected research articles, quantify the frequency of co-occurring linguistic features, and interpret these co-occurring dimensions after conducting a factor analysis.

3.2.1 Identifying, tagging and counting the linguistic features.

We determined that all linguistic features from Biber’s (1988) tag [8] should be included as part of the linguistic features for this study because Biber’s (1988) tag [8] is for spoken and written linguistic features and some Chinese researchers’ academic writing is marked with spoken language (i.e., informal). Additionally, we included academic linguistic features for this study after reviewing the literature on academic writing research. We selected attitude markers [9,48], identity markers [9,49], boosters [9,50], reporting verbs [51,52], noun-noun phrases [53,54] and transition signals [55] which were the most investigated academic linguistic features in previous research. All the above-mentioned linguistic features for academic writing together with Biber’s (1988) tag [8] constituted 73 new linguistic features considered for this study (see S5 Appendix).

After the linguistic features were identified, tagging these linguistic features was a vital and challenging step in developing a novel MDA. The linguistic features adopted from Biber [8] could be tagged through Biber’s (1988) tag [8], but Biber’s (1988) tag [8] is not accessible online. However, Nini developed a multi-dimensional analysis tagger (MAT) [36] by accurately replicating Biber’s (1988) tag [8], and this program can be accessed online and downloaded. Therefore, for the linguistic features involving Biber’s (1988) tag [8], we used Nini’s MAT [36] for tagging. For those selected academic linguistic features, we first used the Stanford Part of Speech (PoS) tagger to tag all the words in the RAs, and then used PatCount software to extract and count the linguistic features selected in this study. The Standford PoS tagger is a software program that assigns parts of speech, such as nouns, verbs, and adjectives, to each word (and other token) of each text. When each word of every text was tagged by PoS, any linguistic feature can be extracted via the pattern code through PatCount. PatCount is a software program developed by the Language Engineering Laboratory of China Foreign Language Education Research Center at Beijing Foreign Studies University [56] and can be downloaded online and used for free of charge. It functions by using pattern matching technology in natural language processing to automatically analyze a corpus. By using pattern matching software through regular expressions, it is convenient to count the frequency of multiple linguistic features in a corpus. In this study, the codes of regular expressions of the selected academic linguistic features were written and input into PatCount, and then the frequency of each linguistic feature was counted through the written pattern files. For instance, absolute, a linguistic feature of boosters, was tagged by the Stanford PoS tagger as absolute_JJ and its regular expression was written as \sabsolute_JJ. When the code \sabsolute_JJ was input into the PatCount pattern file, the frequency of “absolute” in all the texts was counted by PatCount. Another example is “noun-noun phrase”. Its regular expression was written as \S + _NN\s\S + _NN\s. When this pattern file was put into the PatCount software, the frequency count of the noun-noun phrase structure in all the texts was automatically performed by PatCount. In this way, the frequencies of all the academic linguistic features in all the texts can be counted.

3.2.2 Conducting a factor analysis.

Finally, the frequencies of linguistic features obtained from MAT and PatCount were normalized to rates per 1000 words and were put in an Excel file for the next step of factor analysis using SPSS v.25 (see S1 Table). Several pilot factor analyses were performed and only factors with communities over 0.25 and factor loadings higher than 0.30 were retained for final factor analysis. Hence, a total of 62 linguistic features were ultimately retained (see S4 Table). The final factor analysis was run via Principal Axis Factoring with Promax rotation to allow some correlation between factors. A four-factor solution was found to be optimal. The eigenvalues for the four-factor solution accounted for approximately 60.419% of the cumulative variance explained (see S4 Table). Dimension scores for each discipline (see S3 Table) were calculated using z-scores [8] (see S2 Table) to quantify the extent to which a text utilizes the patterns of variation identified by each factor. The mean dimension scores of 100 texts in each discipline of each corpus constituted the dimension score for each discipline of the CRC and NRC for comparison (see Table 2).

Table 2. Mean scores of each discipline on four dimensions in NRC and CRC.
Discipline Dimension NRC CRC
Mean SD Mean SD
1. Agriculture D1 8.0 17.3 −1.2 8.3
D2 3.4 16.3 −0.1 9.3
D3 −1.4 16.8 −0.1 9.5
D4 2.7 13.7 −0.1 9.6
2. Art D1 4.4 13.5 −1.8 8.9
D2 2.2 13.0 −0.1 8.6
D3 −1.0 13.4 0.7 8.4
D4 2.4 11.8 1.0 9.2
3. Economics D1 4.3 12.3 −2.4 8.5
D2 2.6 12.1 0.3 9.3
D3 −0.2 12.2 −0.9 8.2
D4 2.1 11.0 −0.6 8.1
4. Engineering D1 −1.1 8.4 −3.9 4.4
D2 2.6 10.0 −2.8 5.5
D3 0.0 9.2 0.3 6.7
D4 2.1 8.6 −2.1 9.8
5. History D1 2.1 11.4 −1.2 8.0
D2 0.3 11.4 −1.6 8.7
D3 −2.1 11.7 0.7 9.0
D4 0.5 11.0 −0.4 9.3
6. Law D1 3.6 12.6 −1.5 7.7
D2 2.0 12.2 −1.0 8.7
D3 −0.4 12.7 0.6 8.5
D4 1.8 11.7 −0.4 9.7
7. Literature
D1 1.9 10.6 −1.6 7.0
D2 0.4 10.8 −0.3 8.1
D3 −1.23 11.0 0.5 8.4
D4 0.5 10.3 −0.4 9.1
8. Management science D1 0.2 9.4 −2.0 7.8
D2 −0.1 9.8 −1.2 8.7
D3 −1.2 9.7 1.2 8.6
D4 0.3 10.4 −1.5 8.8
9. Medicine
D1 1.4 9.7 −2.8 7.9
D2 1.3 9.9 −1.1 8.1
D3 0.8 10.9 1.4 8.1
D4 −0.6 9.8 −1.8 8.7
10. Natural science
D1 2.0 9.9 −3.2 8.1
D2 0.2 9.7 −0.4 7.5
D3 0.9 9.5 0.7 8.5
D4 0.6 9.7 −1.5 9.0
11. Education D1 0.3 9.4 −1.9 7.4
D2 −1.3 8.8 −1.9 8.8
D3 −0.2 9.1 −0.5 7.8
D4 −0.5 9.8 −2.7 8.5
12. Philosophy D1 0.0 8.3 −3.5 5.5
D2 −1.2 9.7 −2.3 7.4
D3 −0.3 8.7 1.7 7.6
D4 1.1 9.4 −2.6 9.6

ANOVA tests were employed to examine whether there were significant differences between the NRC and CRC along each dimension. With two independent variables (i.e., language background and discipline), this study examined the effects of language background (native English researchers and Chinese researchers) and 12 disciplines (agriculture, art, economics, history, literature, law, medicine, natural science, engineering, education, philosophy and management science) on the dimension score of the newly developed MDA. Due to the non-normal distribution of the data, a nonparametric approach, an Aligned Rank Transform (ART ANOVA) was employed.

3.2.3 Reliability and validity.

To ensure the validity of this research, we controlled the bias in collecting and analyzing the qualitative data by using the method of triangulation (of investigators); that is, we assigned two research assistants to collect, clean and analyze the data. We trained these research assistants first and ensured that they were completely clear about the criteria of selecting journals and research articles and how to clean the textual data and use the software to analyze the data.

After counting the frequencies of the linguistic features via the computer programs, we obtained an Excel data file (See S1 Table) and then we used the Kaiser-Meyer-Olkin measure of sampling adequacy (KMO) to test the validity of these data before we used SPSS to perform the factor analysis. The KMO value was  .974, indicating that the data were suitable for factor analysis (See S4 Table).

4. Results and discussion

4.1 Results of factor analysis

The KMO of the data in this study yielded a score of  .974, indicating that correlation patterns were noticeable in the data [57]. Barlett’s Test for Sphericity (Approximate Chi-Square = 198364.098, df = 2628, p = .000) was significant, indicating that adequate correlations existed in the correlation matrix and that factor analysis was suitable for the data in this study [58]. After conducting factor analysis, four dimensions were extracted and interpreted in the following sections.

4.2 Interpretation of the dimensions

Four dimensions were extracted through factor analysis. Following Biber’s [8] way of interpreting the dimensions, the detailed analysis and discussion of each dimension is presented in the following sub-sections.

4.2.1 Dimension one: Academic involvement and interaction versus information density.

Dimension one, labelled as Academic Involvement and Interaction versus Information Density, is composed of sixteen positive features and three negative features (see Table 3).

Table 3. Linguistic features on Dimension 1.
Positive linguistic features with loadings
Total adverbs
(.953)
Infinitives TO
(.948)
Emphatics
(.913)
Reporting verbs
(.905)
Pronoun IT
(.900)
Present tense
(.892)
Attributive adjectives
(.886)
Perfect aspect
(.874)
Attitude markers
(.870)
Public verbs
(.844)
Subordinator that deletion
(.844)
WH- clause
(.826)
Boosters
(.813)
Identity markers
(.797)
First person pronoun
(.793)
Necessity modals (.726)
Negative linguistic features with loadings
Total other nouns (−.934) Total prepositional phrases (−.789) Nominalizations (−.551)

On this dimension, positive linguistic features (e.g., first-person pronouns, reporting verbs, public verbs, attitude markers, necessity modals, emphatics, and boosters) are mostly associated with elaborate and involved language production. First-person pronouns are markers of ego-involvement in a text [8]. They are used to bring author and readers into the discourse, creating a sense of interaction. Attitude markers, emphatics and boosters are devices for self-mentioning and expressing one’s attitude in language use, a textual signal by which the writers personally involve themselves in their discourse to interact with the readers. These features create a picture of involving the author’s interaction with the readers to highlight the author’s stance. On the other hand, the features loaded negatively on Dimension 1 are all noun-phrase structures (nouns, nominalizations and prepositional phrases), closely aligned with Biber’s [8,35] and Gray’s [59] findings in which they associated these features with an informational purpose. This utilization of structures is meant to pack high amounts of information into academic nominals [8]. Overall, for Dimension 1, positive features exhibit a more academically involved and interactional style, whereas negative features reveal a more information density writing style.

Fig 1 shows the dimension scores across the 12 disciplines and language background on Dimension 1. It can be seen that native English researchers’ academic writing is more involved and interactional than that of Chinese researchers. ART ANOVA tests were first used to examine the overall main effect of language background, discipline and the interaction between them on Dimension 1. The results reveal that the main effect of language background on Dimension 1 is statistically significant (F = 75.940, p < 0.001, partial eta-squared = 0.031), suggesting that Dimension 1 can differentiate research articles written by native English researchers and Chinese researchers. Given that the overall main effect of language background on Dimension 1 was significant, we further conducted Dunn’s post-hoc pairwise comparisons and applied the Benjamini-Hochberg (BH) method to adjust the p-values for controlling the false discovery rate to confirm on what disciplines researchers with different language backgrounds differed from each other. Dunn’s test suggests that the difference between native English researchers’ academic writing and Chinese researchers’ is significant for the disciplines of agriculture (p = 0.004), art (p = 0.008), economics (p = 0.002), law (p = 0.034), medicine (p = 0.016), natural science (p = 0.004) and philosophy (p = 0.029), but insignificant for the disciplines of history, literature, engineering, education and management science, as shown in Fig 1.

Fig 1. Dimension 1 scores across disciplines and language background.

Fig 1

Moreover, ART ANOVA tests reveal that the overall main effect of discipline (F = 2.42, p = 0.006, partial eta-squared = 0.011) is statistically significant, but the interaction effect between language background and discipline (F = 1.190, p = .0.292, partial eta-squared = 0.006) is insignificant, indicating that disciplines influence Dimension 1 scores but the interaction between language background and discipline has no effect on Dimension 1 scores. Dunn’s post-hoc pairwise comparisons with BH method was then done to identify the disciplinary differences within the same language background because the overall main effect of discipline on Dimension 1 scores is significant. Dunn’s test shows that the differences between disciplines within Chinese researchers’ academic writing are insignificant on Dimension 1, but in native researchers’ academic writing, the differences between the disciplines of agriculture and education (p = 0.032), agriculture and engineering (p = 0.004), agriculture and management science (p = 0.018), agriculture and philosophy (p = 0.027), art and engineering (p = 0.029), economics and engineering (p = 0.013), economics and management science (p = 0.050) are significant, as shown in Fig 2.

Fig 2. Disciplinary differences within the same language background on Dimension 1.

Fig 2

Two examples were excerpted from the agriculture discipline from native English researchers corpus and engineering discipline from Chinese researchers corpus because the dimension scores of these two disciplines are the highest on the positive and negative ends, respectively. In the two examples, positive features are marked in boldface and negative features are underlined (this procedure is applied through all 8 typical examples). In Example (1), the researcher heavily relies on the use of positive features, such as first-person pronouns (e.g., we, our), reporting verbs (e.g., require, propose), adverbs (e.g., directly, instead, wholly, surprisingly, conceptually), attributive adjectives (previous, multiple, simple, significant) and present tense to highly stress and recommend their novel and effective research approach to the reader. Choosing these linguistic features reveals that the researcher would like to explicitly convey their stance to persuade readers about their innovation reported in their research. There are very few negative linguistic features, such as nominalization and preposition, and only some necessary nouns are found in Example (1).

(1) As far as we know, our approach is the first to generate bounding box proposals directly from edges. Unlike all previous approaches we do not use segmentations or super pixels, nor do we require learning a scoring function from multiple cues. Instead, we propose to score candidate boxes based on the number of contours wholly enclosed by a bounding box. Surprisingly, this conceptually simple approach out-competes previous methods by a significant margin.

(NRC_AGRI27)

In contrast, Example (2) shows the researchers’ preference for negative linguistic features on Dimension 1. Using a large number of nouns, nominalizations and prepositions mark the article with dense information. Compared with Example 1, Example 2 avoids involving and interacting with the readers to conceal their authorial identity and opinion by using rare positive features on Dimension 1, but it is featured with conveying information objectively.

(2) Nick and Satao compared the aerodynamic characteristics of two models—short and optimized—to determine the best pod shape for drag reduction. The pressure drag of the optimized model was 24% less than that of the short model.

(CRC_ENGR01)

Dimension 1 in this study reflected a distinction that has been consistently identified in MD analyses, i.e., it was more involved language production and high-density informational language. In the academic writing register, native English researchers were more likely to involve themselves and interact with readers to clearly and confidently express their opinions than Chinese researchers were. This finding aligned with those of previous studies. Native English writers preferred to project themselves in explicit interactional academic discourse, construct authorial presence and show their engagement with their readers in arguments [60]. While EFL learners were overwhelmingly reluctant to establish an authorial identity in their writing, Chinese EFL learners presented a passive, test-oriented writer identity due to influences from learning and testing cultures [61]. Similar findings were reported by Liu and Zhang [62], who found that the frequency of interactional metadiscourse utilized by Chinese master students in academic writing was less than that used by international journal authors. This was perhaps because that Chinese researchers, as EFL learners, were taught to eliminate explicit agency in their academic writing [63] to establish objectivity. Chinese students were taught that academic writing was a rigorous writing so that the authors should avoid involving or engaging the readers in the discourse and not claim their authorial presence. Therefore, Chinese researchers preferred to disguise their interpretative statements for objectivity, whereas Anglo-American rhetoric tended to project a more reader-oriented attitude characterized by more reader guidance and an explicit authorial presence [64]. While native researchers excel at authorial engagement and elaboration, Chinese researchers demonstrate strengths in information density and conciseness, efficiently packing content with minimal redundancy, a style valued in fast-paced, technical fields as engineering. This aligns with English for Academic Purposes (EAP) research showing L2 writers’ precision can model concise argumentation for L1 peers. Thus, both groups offer mutual lessons of natives in interactivity and Chinese in streamlined reporting.

This distinction on Dimension 1 was especially significant on the disciplines of agriculture, art, economics, law, medicine, natural science and philosophy, but insignificant in the disciplines of education, engineering, history, literature and management science, meaning that native researchers prefer a more involved and interactional writing style while Chinese researchers tend to be informational in the disciplines of agriculture, art, economics, law, medicine, natural science and philosophy. These disciplines belong to hard sciences, except art, law and philosophy based on the existing body of disciplinary classification theories [6567]. Although art, law and philosophy disciplines do not belong to hard sciences, they are regarded as pure sciences compared with applied sciences. Pure sciences are concerned with the development of knowledge for its own sake and testing theoretical formulations [66], rather than with any form of practical outcome [68]. Therefore, native English and Chinese researchers show significant differences in writing in hard or pure sciences with the characteristics of having highly internationalized research paradigms, rigorous reporting standards and a relatively unified terminological system. Interestingly, it was found that even for hard sciences that emphasize objectivity, logical rigor and highly conventionalized formats, native researchers exhibit more flexibility by appropriately incorporating authorial identity and engagement with the readers, rather than being informational and objective as it should be. This finding demonstrated native researchers strategically declared their author stance within the disciplinary norms due to their greater writing fluency and more liberal expressions in writing. Whereas, Chinese researchers appeared to be conservative in writing in hard and pure disciplines, by strictly adhering to hard sciences’ disciplinary conventions, focusing on objective facts and rigorous academic norms. In Hyland’s study [69], he pointed out that the conventions of the hard sciences stress the demonstration of objectivity, precision, and logical reasoning, minimizing the researcher’s role to highlight the phenomena under study, while in the soft disciplines, the interpretation of personal experiences is central and writers are more able to draw on their own authority to persuade readers. Our finding aligns with Hyland’s [69] observation in disciplinary variation in his meta-discourse theory, but extends his theory by finding out that the disciplinary variation between hard sciences and soft sciences was also influenced by second language writing challenges and language proficiency (i.e., competence). Native researchers’ writing style in hard science is not completely informative and objective as expected; on the contrary, they also showed an involved and interactional writing style. However, Chinese researchers, as L2 writers, often adhere rigidly to impersonal constructions in hard science. This rigidity may stem from high cognitive load [70] during composition in English, leaving limited capacity for rhetorical flexibility. Conversely, native researchers’ linguistic proficiency affords them the flexibility to strategically follow these conventions. As Hyland [71] suggested, expert writers employ self-mention not to violate objectivity, but to skillfully construct their authorial credibility and responsibility within the disciplinary framework. Thus, in hard or pure sciences, native researchers are more strategic to freely express their author stance, but Chinese researchers tend to follow the generally pre-set disciplinary conventions.

In the disciplines of education, engineering, history, literature and management science, which are generally regarded as soft disciplines except engineering, the difference between native researchers being involved and interactional and Chinese researchers being informational is insignificant. Soft disciplines and engineering discipline falling under applied sciences, are characterized by rhetorical and personalized nature emphasizing personal interpretation more where the individual voice and argumentative skill of the researcher are paramount [68]. Therefore, academic writing in soft disciplines permits a diversity of rhetorical styles and modes of expression. This intrinsic variability may obscure differences induced by various language backgrounds, making statistical differences insignificant.

Fig 2 shows the disciplinary differences within the same language background. It can be seen that within Chinese researchers the variations among disciplines are insignificant, but within native researchers, the differences between agriculture and education, agriculture and engineering, agriculture and management science, agriculture and philosophy, art and engineering, economics and engineering, economics and management science were significant. All disciplines except engineering within native researchers’ language background located on the positive end of Dimension 1 (dimension score is more than 0), meaning that these 11 disciplines exhibit a generally involved and interactional writing style (Fig 2). Statistically significant differences observed between agriculture and education, agriculture and management science, agriculture and philosophy, economics and management science were only in the extent to which they all exhibit the features of being involved and interactional. Notably, among all disciplines, only engineering stood out on the negative end of Dimension 1 (dimension score is below 0), meaning that native researchers preferred a writing style of being informational in engineering discipline. Engineering discipline was significantly different from the soft sciences as art and economics and other hard applied sciences as agriculture, as shown in  Fig 2. This phenomenon can be explained by the subtle distinctions among disciplines. Engineering is part of applied sciences that relies heavily on mathematical equations to design or create new devices or structures. Its emphasis on precision, predictability, and safety necessitates a highly anonymized, manual-like writing style to convey technical information, more than personal interpretation and rhetorical expression [68]. In contrast, the research topics of agriculture science are often complex living systems, economics deals with the production, distribution, and consumption of goods and services, and art is grounded in personal aesthetic interpretation. These three disciplines, while also providing information, must argue for and interpret the contextual and uncertain nature of their findings [72]. Consequently, their writing requires authorial intervention and interactive elements, thereby exhibiting involved and interactional writing style. This finding powerfully demonstrates that the influence of disciplines on writing style is deeply embedded within each discipline, beyond a simple hard-soft binary.

4.2.2 Dimension two: Interactive argumentation versus static description.

Dimension two is made up of thirteen positive features and six negative features (see Table 4), labelled as Interactive Argumentation versus Static Description.

Table 4. Linguistic features on Dimension 2.
Positive linguistic features with loadings
Transitional signals (.975) Present participial clauses (.905) Conjuncts
(.878)
Phrasal coordination (.877)
Other adverbial subordinators (.867) Past participial clauses (.862) Split auxiliaries (.853)
Present participial WHIZ deletion relatives (.826)
Concessive adverbial subordinators (.820)
Pied-piping relative clauses (.815)
Past participial WHIZ deletion relatives (.808)
Causative adverbial subordinators (.715)
Conditional adverbial subordinators (.662)
Negative linguistic features with loadings
Existential there (−.861) Be as main verb (−.826) Demonstrative pronouns (−.814) Third person pronouns (−.790)
Past tense (−.783) Predicative adjectives (−.779)

Linguistic features such as transitional signals, conjuncts, participles, adverbial clauses and relative clauses co-occurred on positive end of Dimension 2, providing an interactive function of managing information flow to establish writers’ interpretations. Conjuncts explicitly mark logical relations between clauses; and participles are used for integration or structural elaboration [8]. Adverbial clauses and relative clauses serve the function of cohesion by connecting clauses in a good logic. Therefore, the positive features for Dimension 2 help create this interactive argumentation discourse, showing that in the academic writing register, the author is aware of the logical relations between clauses and paragraphs to form argumentative discourse.

Fig 3 shows that native English researchers’ academic writing is more interactive and argumentative than that of Chinese researchers. ART ANOVA tests suggest that the main effect of language background on Dimension 2 is statistically significant (F = 12.685, p = 0.000, partial eta-squared = 0.005). Since the overall main effect of language background on Dimension 1 was significant, Dunn’s post-hoc pairwise comparisons were conducted to confirm what disciplines the researchers with different language backgrounds differed from each other. Dunn’s test shows that the difference between native English researchers’ academic writing and Chinese researchers’ is statistically significant only in the discipline of engineering (p = 0.010). ART ANOVA tests also reveal that the main effect of both discipline (F = 1.581, p = 0.098, partial eta-squared = 0.007) and the interaction effect between language background and discipline (F = 1.024, p = 0.422, partial eta-squared = 0.005) on Dimension 2 scores are not statistically significant.

Fig 3. Dimension 2 scores across disciplines and language background.

Fig 3

Examples (3) and (4) were chosen from the native researchers’ agriculture discipline and the Chinese researchers’ engineering discipline, respectively representing texts with the highest positive dimension scores and the highest negative dimension scores on Dimension 2.

In Example (3), the researchers argue an event using interactive devices through positive features on Dimension 2 (i.e., conjuncts, phrasal coordination, present participial WHIZ deletion relatives, past participial WHIZ deletion relatives, adverbial subordinators, and concessive adverbial subordinators). By using conjuncts and participial structures, cohesion is achieved. Researchers present a clear relationship between clauses and sentences. These positive features provide a picture of interactive argumentation.

(3) Although the amyloid load reached a plateau early after symptom onset, astrocytosis and microgliosis increased linearly throughout the disease course, thus integrating both lesions as a marker of disease severity. Moreover, glial responses correlated positively with tangle burden, whereas astrocytosis correlated negatively with cortical thickness. However, neither correlated with amyloid load.

(NRC_AGRI79)

On the contrary, Example (4) exhibits the function of negative linguistic features, creating a static description discourse. The BE verb, existential there structure and predicative adjective are markers of the static, informational style that are common in writing, but they are considered non-complex constructions with a reduced informational load, thus being informal as a spoken style [8]. Unlike active verbs, BE and existential there are usually used to describe a static state. In Example (4), simple sentence construction of “Subject + be” is used three times in a paragraph to describe a simple and basic fact about democracy.

(4) Democracy is not a simple form of government and judgments about the nature of different governments that claim to be democratic should not be made in a simplistic manner. Certainly, degrees of democracy are possible and a crucial criterion is the proportion of citizens voting correctly at any particular time.

(CRC_ENGR38)

In summary, Dimension 2 differentiated native English researchers’ interactive argumentation writing style from Chinese researchers’ static descriptive writing style. This was explainable because previous studies have shown that native English researchers tended to present a greater number of logical markers than L2 writers do [73,74]. To achieve logical elaboration, a good command and proficiency in the English language are needed. Chinese researchers, as EFL learners, usually presented a less causal content than the native speaker’s writing [75] and inadequate use of logical connectors [76]. Native researchers exhibited an explicit logical connection of the ideas in a linear way, but Chinese researchers did not.

On the other hand, Chinese researchers presented a more static descriptive writing style. EFL learners tended to use is and are frequently, and usually added a be verb into a simple subject structure to express meaning [77]. Chinese researchers’ English vocabulary was not rich enough, and they cannot organize and combine various language bundles skillfully; thus, their sentence structures were mostly simple repetitions in collocation, using less complex sentence patterns to enrich and refine meaning expression. The same held for existential there, Chinese university students’ excessive use of existential sentences in their academic writing was reported by Zhang et al. [78] Dimension 2 differentiated native researchers’ interactive argumentation from Chinese researchers’ static descriptive style. This reflects proficiency differences that natives chain ideas linearly with complex connectors, while Chinese EFL writers favor direct A is B structures due to vocabulary and syntax constraints. Yet this L2 simplicity confers strengths in clarity and conciseness, avoiding native overelaboration that may obscure meaning. Such straightforwardness enhances readability in disciplines with heavy jargons.

In the discipline of engineering, this difference was statistically significant. Writing in engineering emphasizes preciseness and objectivity to convey the information, usually complex profession-related terms and technical information. L2 Chinese researchers, likely constrained by their limited language proficiency and the cognitive load of writing in a second language, tend to produce texts with simple sentence structures in such a discipline requiring professional information and demanding logical rigor. In contrast, native researchers, proficient in their discourse community, can strategically employ interactive devices to guide readers through reasoning, not limited to static description, in a hard discipline.

4.2.3 Dimension three: Impersonal evaluation versus personal opinion.

Interpreted as an Impersonal Evaluation versus Personal Opinion, Dimension 3 is characterized by eight positive features and six negative features (see Table 5). The use of passives provides a sense of objective detachment in expository prose. This sense of objectivity is part of scientific culture, and is often expected in scientific writing [35]. Another two features (downtowners and hedges) generally reflect the uncertainty of the language. Positive features show the author’s preference to be impersonal in presenting an objective fact without expressing their subjective opinion, whereas negative features indicate personal and persuasive discourse.

Table 5. Linguistic features on Dimension 3.
Positive linguistic features with loadings
Agentless passives (.865) By-passives (.805) Noun-noun phrase (.792) Hedges (.762)
Demonstratives (.744)
Possibility modals (.744) Downtoners (.704)
Analytic negation (.617)
Negative linguistic features with loadings
Suasive verbs (−.974) That relative clauses on object position (−.947) Amplifiers (−.940) That adjective complements (−.932)
That verb complements (−.887)
Predicative modals (−.860)

Fig 4 shows the distribution of the mean dimension scores across the 12 disciplines and language background along Dimension 3. Chinese researchers’ academic writing is generally more impersonal than that of the native English researchers in all disciplines except economics, natural science and education, whereas native English researchers’ writing is all at the negative end of Dimension 3 except for medicine, natural science and engineering. ART ANOVA tests reveal that the main effect of language background on Dimension 3 is statistically significant (F = 10.330, p = 0.001, partial eta-squared = 0.004), suggesting that Dimension 3 could differentiate research articles written by native English and Chinese researchers. But Dunn’s post-hoc pairwise comparisons with BH method was then done to find no statistically significant difference for any discipline. Meanwhile, the main effect of discipline (F = 1.010, p = 0.435, partial eta-squared = 0.005) and the interaction effects between language background and discipline (F = 0.730, p = 0.711, partial eta-squared = 0.003) on Dimension 3 are not statistically significant.

Fig 4. Dimension 3 scores across disciplines and language background.

Fig 4

Example (5) was selected from the Chinese researcher corpus in the philosophy discipline because it had the highest positive dimension scores. Researchers are showing their impersonal tone when presenting human research ethics by using passive structures, demonstratives and noun-noun phrases.

(5) Human Research Ethics aims to ensure that research is conducted to the highest ethical standard and that human participants in research are protected.

(CRC_PHIL05)

In contrast, negative features on Dimension 3 suggest a personal opinion expressing style in writing. Example (6) was selected from the native researchers’ corpus in the discipline of history because it had the highest negative dimension score. Suasive verbs imply intentions to make some change in the future (e.g., command, stipulate) [8] and amplifiers have the effect of boosting the force of the verb [79]. In Example (6), suasive verbs, that verb complements and that relative clauses in the object position are used when the researchers are making conclusions for their research. This reveals that the researchers are stating their personal opinions and trying to persuade the readers about their findings.

(6) We insist that the prefixal agreement and the theme suffix are fundamentally different kinds of morphemes. Therefore, we propose that the special character of inverse contexts arises from the fact that the EA never agrees with the core probe, and this failure is what must be repaired.

(NRC_HIST95)

Dimension 3 distinguished the native English researchers’ and the Chinese researchers’ writing in that the Chinese researchers generally adopted a more impersonal style, whereas the native researchers tended to make personal argumentation in their research. This finding aligned with previous research in which the frequent use of simple passive verbs was found in L2 learners’ academic writing [80]. Chinese impersonality offers strengths in objectivity and humility, aligning with disciplinary conventions that prioritize collective knowledge over individual voice. This restrained writing style models with cultural sensitivity for natives who risk perceived bias through overt self-positioning.

4.2.4 Dimension four: Explicit elaborating style versus simplified reporting style.

Six positive features and four negative features load on Dimension 4 (see Table 6) which is labelled as Explicit Elaborating Style versus Simplified Reporting Style.

Table 6. Linguistic features on Dimension 4.
Positive linguistic features with loadings
Type-token ratio (.947) Independent clause coordination (.932) Sentence relatives (.872)
Word length (.840) That relative clauses on subject position (.769) Split infinitives (.628)
Negative linguistic features with loadings
WH relative clauses on subject position (−.900) Contractions (−.850) Private verbs (−.788)
Gerunds (−.704)

Positive features on Dimension 4 create a picture of explicit elaboration. Explicitness has been measured by features such as word length and the type/token ratio which is the ratio of the number of different words to the total number of words [8]. They are features revealing lexical specificity. Complex words and a careful selection of vocabulary result in a high type/token ratio [8], indicating good syntactic complexity. Together with relative clause structures, these features reflecting authors’ proficiency in the English language serve the function of an explicit explanation of the authors’ argumentation.

Fig 5 shows the distribution of the mean scores of Dimension 4 across disciplines and language background. The native English researchers use a more explicit elaborating style than the Chinese researchers do in academic writing. ART ANOVA tests reveal that the main effect of language background (F = 14.210, p = 0.000, partial eta-squared = 0.006) on Dimension 4 are statistically significant. Then Dunn’s post-hoc pairwise comparisons with BH method was conducted, this difference was not statistically significant for any discipline. Meanwhile, the main effect of discipline (F = 1.380, p = 0.177, partial eta-squared = 0.006) and the interaction effect between language background and discipline on Dimension 4 are insignificant (F = 0.560, p = 0.863, partial eta-squared = 0.003).

Fig 5. Dimension 4 scores across disciplines and language background.

Fig 5

Examples (7) and (8) were chosen from disciplines with the highest positive and negative dimension scores, respectively on Dimension 4. In Example (7), positive features on Dimension 4 as split infinitives, that relative clauses in the subject position and independent clause coordination are used to help explicitly illustrate the benefits of sustainable farming. (The type/token ratio is computed by counting the number of different lexical items, so it is not marked in Example 7).

(7) Sustainable farming practices, that aim to significantly improve soil health and biodiversity, not only help to reduce environmental impact but also strive to increase crop yields.

(NRC_AGRI09)

In contrast, the cooccurrence of negative features such as private verbs, gerunds, contractions, and relative clauses on Dimension 4 suggests that researchers prefer simple structures to report their research more colloquially. Private verbs [62] describe or refer to mental states (e.g., know, learn, think) and non-observable intellectual acts that are private, such as emotive acts (feel, hope), mental acts (realize, understand), and cognitive acts (believe, conclude, forget, recognize) [81]. In Example (8), the researchers focus more on reporting their personal opinions when using private verbs.

(8) Moreover, the authors hope to improve the prediction results by tuning the hyperparameters and designing more sophisticated features using deep learning models.

(CRC_EDU46)

In summary, native English researchers preferred a more explicit elaborating style because native English researchers had much better language proficiency than Chinese researchers did. They certainly surpassed Chinese researchers in terms of syntactic complexity, lexical diversity and explicit elaboration. Chinese researchers were not skilled English language users and most of them could use only simple structures to report their research in their writing. Dimension 4 contrasted native researchers’ explicit elaborating style with Chinese researchers’ simplified reporting style. Natives excelled in elaboration due to superior proficiency. Chinese simplicity provides strengths in efficiency and readability, prioritizing research content over linguistic display. This reader-centered writing style models concise impact for natives prone to overelaboration.

5. Limitations

While this study offers novel insights into L1/L2 academic writing patterns, two methodological limitations warrant consideration. First, author classification proxies have inherent edge cases. Second, disciplinary sampling lacks sub-disciplinary nuance explained further below.

5.1 Author classification limitations

Although our double-safeguard method, i.e., institutional restriction to US/UK/Canada/Australia/NZ with surname origin matching, virtually eliminates L2 misclassification as L1, rare edge cases persist (e.g., naturalized citizens with fully Westernized names). These proxy limitations, while standard in large-scale EAP research, suggest caution in future validation through direct author verification.

5.2 Disciplinary sampling limitations

Following China’s Level One Disciplines system, we sampled broadly across fields without subdiscipline stratification. This captures major disciplinary tendencies but misses intra-discipline variation, potentially masking specific field patterns. Representativeness is thus sufficient for broad discipline-level effects rather than fine-grained subfield differences. Future studies should build stratified corpora to enhance generalizability across specialized academic subdomains.

6. Conclusions

While MDA has effectively captured linguistic variation across general registers and L1/L2 learner corpora, few studies have developed genre-specific MDA frameworks for academic research articles by identifying linguistic features for certain disciplines [35,82]. Existing comparisons in academic writing using MDA method typically focused on L1/L2 variations rather than disciplinary nuances. This study addresses these gaps by: (1) selecting 62 linguistic features for academic writing; (2) using PatCount software for automated tagging; and (3) deriving four novel dimensions that reveal nuanced L1/L2 differences beyond the hard-soft science binary. These contributions provide both a new analytical model and pedagogical tools absent in the previous MDA research.

The current study analyzed the co-occurrence of patterns of the selected 62 linguistic features for academic writing which resulted in constructing a novel multidimensional analysis model with four dimensions distinguishing native English researchers’ and Chinese researchers’ academic writing differences across twelve disciplines. They are (1) academic involvement and interaction versus information density; (2) interactive argumentation versus static description; (3) impersonal evaluation versus personal opinion; and (4) explicit elaborating style versus simplified reporting style.

ART ANOVA confirmed significant language background effects across all dimensions (p < .001), with discipline effects limited to Dimension 1 and no interaction. Natives showed greater academic involvement, interaction, and elaboration while Chinese researchers favored a more informational, static, descriptive, impersonal and simplified reporting style in their writing.

Native researchers were more academically involved and interactional primarily in agriculture, art, economics, law, medicine, natural science, and philosophy (Dunn’s post-hoc, BH-adjusted, p < .05), while Chinese researchers tended to be more informational in these seven disciplines, indicating that in rigorous disciplines as hard or pure sciences, native researchers are more strategic to freely express their author stance, but Chinese researchers tend to be conservative and follow the generally pre-set disciplinary conventions. While these differences were insignificant in soft or applied fields because writing in soft disciplines permits a diversity of rhetorical styles and ways of expression. This intrinsic variability may obscure differences induced by various language backgrounds. Notably, engineering discipline revealed intra-native variation, i.e., natives tended to be more informational than in soft sciences (art & economics) or other hard fields (agriculture), and only in engineering discipline Chinese researchers showed a more static, descriptive and simplified reporting style than native English researchers did, challenging simple hard-soft binaries.

These findings proved that researchers’ preference for linguistic features in different disciplines in their academic writing shaped their unique writing styles, and this preference was not simply attributed to hard-soft sciences binary difference, but also related to researchers’ language proficiency and subtle distinctions within disciplines under same hard or soft sciences. This suggests that Chinese researchers should exhibit their authorial presence, interact with readers more, and employ more interactive devices to make their writing coherent and explicit.

In addition to the above-mentioned key findings from different language backgrounds and disciplines, this research also contributed to enriching the existing MDA research by providing a newly complementary perspective. Unlike previous studies stressing the differences, this research innovatively emphasized the mutual strengths from these explored differences. This balanced comparison reveals that mutual strengths of Chinese conciseness (Dimensions 1 & 4) model clarity and efficiency for natives prone to verbosity; native involvement and interaction (Dimensions 1 & 4) guides Chinese toward engagement. Chinese impersonality (Dimension 3) offers objectivity where natives risk bias; natives’ logic (Dimension 2) complements L2 simplicity.

This study contributed pedagogically and methodologically as well. Pedagogically, the newly developed model (Fig 6) enables targeted EAP instruction by discipline nuances. Educators should tailor pedagogical approaches to address distinct writing styles between native English and Chinese researchers across various disciplines. By interpreting discipline-specific writing styles on the four dimensions of this novel MDA model, educators can increase students’ discipline awareness. Being aware of disciplinary variation and native researchers’ writing style empowers learners to engage in quality academic practices. Methodologically, PatCount automation with Regular Expression code writing overcomes traditional MDA tagging barriers, scalable for genre-specific models.

Fig 6. Newly developed MDA model for academic writing in research articles.

Fig 6

Limitations in this study include proxy L1/L2 classification and broad disciplinary categories without subdiscipline stratification. Future research should test proficiency subgroups and finer-grained fields.

This framework advances MDA toward academic genres, equipping non-native researchers with discipline-sensitive tools while recognizing cross-cultural writing assets.

Supporting information

S1 Table. Frequencies of all linguistic features in 2400 articles for factor analysis.

(XLSX)

pone.0346776.s001.xlsx (1,003.5KB, xlsx)
S2 Table. Z-score of each article on four factors after factor analysis.

(XLSX)

pone.0346776.s002.xlsx (129KB, xlsx)
S3 Table. Mean dimension score of each discipline on four factors.

(XLSX)

pone.0346776.s003.xlsx (9.9KB, xlsx)
S4 Table. Results of factor analysis (KMO, cumulative variance explained and pattern matrix).

(DOC)

pone.0346776.s004.doc (7.2MB, doc)
S5 Appendix. Linguistic features selected in this study.

(PDF)

pone.0346776.s005.pdf (60.6KB, pdf)
S6 Appendix. List of abbreviations.

(PDF)

pone.0346776.s006.pdf (59KB, pdf)

Acknowledgments

We are grateful to Yang Yang and Jie Liu for their assistance with data collection and data cleaning, to Mei Feng for her insightful comments on the draft research, and to Wanmeng Xiao for her guidance in the data analysis. We also thank the reviewers and editors for providing constructive feedback on the manuscript to improve it.

Data Availability

All relevant data are within the paper and its Supporting Information files.

Funding Statement

Southwest Medical University, Luzhou, Sichuan, China Award Number: 2022YB024 | Recipient: Jiaqi Deng, MA The funder had no role in the study design, data collection and analysis, decision to publish, or preparation of the manuscript.

References

  • 1.Biber D, Reppen R, Staples S, Egbert J. Exploring the longitudinal development of grammatical complexity in the disciplinary writing of L2-English university students. Int J Learn Corpus Res. 2020;6(1):38–71. [Google Scholar]
  • 2.Omidian T, Siyanova-Chanturia A, Biber D. A new multidimensional model of writing for research publication: an analysis of disciplinarity, intra-textual variation, and L1 versus LX expert writing. J Engl Acad Purp. 2021;53:101020. [Google Scholar]
  • 3.Li F, Han Y. Student feedback literacy in L2 disciplinary writing: Insights from international graduate students at a UK university. Assess Eval High Educ. 2022;47(2):198–212. [Google Scholar]
  • 4.Dontcheva-Navratilova O. Self-mention in L2 (Czech) learner academic discourse: Realisations, functions and distribution across master’s theses. J Engl Acad Purp. 2023;64:101272. [Google Scholar]
  • 5.Shen C, Guo J, Shi P, Qu S, Tian J. A corpus-based comparison of syntactic complexity in academic writing of L1 and L2 English students across years and disciplines. PLoS One. 2023;18(10):e0292688. doi: 10.1371/journal.pone.0292688 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 6.Gray B, Nuttall C. Disciplinary discourses and second language research. In: The Routledge handbook of second language acquisition and discourse. Routledge; 2024. p. 242–54. doi: 10.4324/9781003177579-21 [DOI] [Google Scholar]
  • 7.Shamsi AF, Osam UV. Challenges and support in article publication: perspectives of non-native English speaking doctoral students in a “publish or no degree” context. Sage Open. 2022;12(2):21582440221095021. doi: 10.1177/21582440221095021 [DOI] [Google Scholar]
  • 8.Biber D. Variation across speech and writing. Cambridge University Press; 1988. [Google Scholar]
  • 9.Hyland K. Disciplinary interactions: metadiscourse in L2 postgraduate writing. J Second Lang Writ. 2004;13(2):133–51. doi: 10.1016/j.jslw.2004.02.001 [DOI] [Google Scholar]
  • 10.Boginskaya OA. Cross-disciplinary variation in metadiscourse: a corpus-based analysis of Russian-authored research article abstracts. Train Lang Cult. 2022;6(3). [Google Scholar]
  • 11.Zhang W, Cheung YL. The different ways to write publishable research articles: using cluster analysis to uncover patterns of appraisal in discussions across disciplines. J Engl Acad Purp. 2023;63:101231. [Google Scholar]
  • 12.Seyri H, Rezaei S. Disciplinary and cross-cultural variation of stance and engagement markers in soft and hard sciences research articles by native English and Iranian academic writers: a corpus-based analysis. Interdiscip Stud Engl Lang Teach. 2023;1(1):81–99. [Google Scholar]
  • 13.Ho SS, Looi J, Leung YW, Goh TJ. Public engagement by researchers of different disciplines in Singapore: a qualitative comparison of macro- and meso-level concerns. Public Underst Sci. 2020;29(2):211–29. doi: 10.1177/0963662519888761 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 14.Dong Y, Wang J, Jiang F. Epistemic positioning by science students and experts: a divide by applied and pure disciplines. Appl Linguist Rev. 2024;15(3):927–54. [Google Scholar]
  • 15.Alghazo S, Al Salem MN, Alrashdan I, Rabab’ah G. Grammatical devices of stance in written academic English. Heliyon. 2021;7(11). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 16.Meyerhoff M. Unnatural bedfellows? The sociolinguistic analysis of variation and language documentation. J R Soc N Z. 2019;49(2):229–41. [Google Scholar]
  • 17.Hyland K, Jiang FK. Metadiscourse choices in EAP: an intra-journal study of JEAP. J Engl Acad Purp. 2022;60:101165. [Google Scholar]
  • 18.Stoller FL, Robinson MS. Chemistry journal articles: an interdisciplinary approach to move analysis with pedagogical aims. Engl Specif Purp. 2013;32(1):45–57. [Google Scholar]
  • 19.Wu J, Pan F. Stance construction via that-clauses in telecommunications research articles: a comparison of L1 and L2 expert writers. Text Talk. 2023;44(3):387–410. doi: 10.1515/text-2021-0170 [DOI] [Google Scholar]
  • 20.Viera RT. Lexical richness of abstracts in scientific papers in anglophone and non-anglophone journals. 3L. 2022;28(2):224–39. [Google Scholar]
  • 21.Contemori C. Changing comprehenders’ pronoun interpretations: immediate and cumulative priming at the discourse level in L2 and native speakers of English. Second Lang Res. 2019;37(4):573–86. doi: 10.1177/0267658319886644 [DOI] [Google Scholar]
  • 22.Mur-Dueñas P. There may be differences: analysing the use of hedges in English and Spanish research articles. Lingua. 2021;260:103131. doi: 10.1016/j.lingua.2021.103131 [DOI] [Google Scholar]
  • 23.Cao Y, Xiao R. A multi-dimensional contrastive study of English abstracts by native and non-native writers. Corpora. 2013;8(2):209–34. doi: 10.3366/cor.2013.0041 [DOI] [Google Scholar]
  • 24.Gray B. Linguistic variation in research articles: when discipline tells only part of the story. John Benjamins; 2015. [Google Scholar]
  • 25.Biber D, Hared M. Dimensions of register variation in Somali. Lang Var Change. 1992;4(1):41–75. [Google Scholar]
  • 26.Biber D, Finegan E. Diachronic relations among speech-based and written registers in English. In: Variation in English. Routledge; 2001. 264 p. [Google Scholar]
  • 27.Westin I, Geisler C. A multi-dimensional study of diachronic variation in British newspaper editorials. ICAME J. 2002;26(1):133–52. [Google Scholar]
  • 28.Friginal E. Outsourced call centers and English in the Philippines. World Engl. 2007;26(3):331–45. [Google Scholar]
  • 29.Biber D, Conrad S. Register, genre, and style. Cambridge University Press; 2019. [Google Scholar]
  • 30.Kruger H, Van Rooy B. Constrained language: a multidimensional analysis of translated English and a non-native indigenised variety of English. Engl World-Wide. 2016;37(1):26–57. [Google Scholar]
  • 31.Mu C. A multidimensional contrastive analysis of linguistic features between international and local biology journal English research articles. Scientometrics. 2021;126(9):7901–16. doi: 10.1007/s11192-021-04102-x [DOI] [Google Scholar]
  • 32.Hardy JA, Römer U. Revealing disciplinary variation in student writing: a multi-dimensional analysis of the Michigan Corpus of Upper-level Student Papers (MICUSP). Corpora. 2013;8(2):183–207. doi: 10.3366/cor.2013.0040 [DOI] [Google Scholar]
  • 33.Pan F. A multidimensional analysis of L1–L2 differences across three advanced levels. S Afr Linguist Appl Lang Stud. 2018;36(2):117–31. [Google Scholar]
  • 34.Jin B. A multi-dimensional analysis of research article discussion sections in an engineering discipline: corpus explorations and scientists’ perceptions. Sage Open. 2021;11(4):21582440211050401. doi: 10.1177/21582440211050401 [DOI] [Google Scholar]
  • 35.Biber D. University language: a corpus-based study of spoken and written registers. John Benjamins Publishing Company; 2006. [Google Scholar]
  • 36.Nini A. The multi-dimensional analysis tagger. In: Multi-dimensional analysis: research methods and current issues. Bloomsbury; 2019. p. 67–94. [Google Scholar]
  • 37.Friginal E, Weigle S. Exploring multiple profiles of L2 writing using multi-dimensional analysis. J Second Lang Writ. 2014;26:80–95. doi: 10.1016/j.jslw.2014.09.007 [DOI] [Google Scholar]
  • 38.Creswell JW. A concise introduction to mixed methods research. SAGE Publications; 2021. [Google Scholar]
  • 39.Abdi J, Farrokhi F. Investigating the projection of authorial identity through first person pronouns in L1 and L2 English research articles. Int J Lang Lit. 2015;3(1):156–68. [Google Scholar]
  • 40.Yağız O, Demir C. A comparative study of boosting in academic texts: a contrastive rhetoric. Int J Engl Linguist. 2015;5(4). [Google Scholar]
  • 41.Kareema MIF, Hakmal MHM. A genre analysis of abstracts written by Sri Lankan and western academics on social science discipline. KALAM Int Res J. 2023;16(1):319–30. [Google Scholar]
  • 42.Ridley R, Wu Z, Zhang J, Huang S, Dai X. Addressing linguistic bias through a contrastive analysis of academic writing in the NLP domain. Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing; 2023. p. 16765–79. doi: 10.18653/v1/2023.emnlp-main.1042 [DOI] [Google Scholar]
  • 43.Kachru BB. Standards, codification, and sociolinguistic realism: the English language in the outer circle. In: English in the world: teaching and learning the language and the literature. Cambridge University Press; 1985. p. 1–30. [Google Scholar]
  • 44.FamilySearch International. FamilySearch - free family history and genealogy records; 1999 [cited 2025 Oct 7]. Available from: https://www.familysearch.org/
  • 45.Bao F, Zhang Y. Zhongguo 548zhong yingwen xueshu qikan de xueke fenlei yu minglu [Subject Classification and Catalogue of 548 Chinese English-Language Academic Journals]. Bianji xuebao [Editorial Board]. 2018;30(06):574–85. [Google Scholar]
  • 46.Gorsuch RL. Factor analysis. 2nd ed. Hillsdale: Lawrence Erlbaum Associates; 1983. [Google Scholar]
  • 47.Weber RP. Basic content analysis. Sage; 1990. [Google Scholar]
  • 48.Wu B, Paltridge B. Stance expressions in academic writing: a corpus-based comparison of Chinese students’ MA dissertations and PhD theses. Lingua. 2021;253:103071. doi: 10.1016/j.lingua.2021.103071 [DOI] [Google Scholar]
  • 49.Li Z. Authorial presence in research article abstracts: a diachronic investigation of the use of first person pronouns. J Engl Acad Purp. 2021;51:100977. [Google Scholar]
  • 50.Bal Gezegin B, Bas M. Metadiscourse in academic writing: a comparison of research articles and book reviews. Eurasian J Appl Linguist. 2020;6(1):45–62. doi: 10.32601/ejal.710204 [DOI] [Google Scholar]
  • 51.Jarkovská M, Kučírková L. Citation practices in EFL academic writing: the use of reporting verbs in Master’s thesis literature reviews. Indones J Appl Linguist. 2020;10(2):571–9. [Google Scholar]
  • 52.Barghamadi M. Reporting verbs in the humanities and medical sciences research articles. Lang Teach Res Q. 2021;22:17–32. [Google Scholar]
  • 53.Badley GF. Post-academic writing: human writing for human readers. Qual Inq. 2019;25(2):180–91. [Google Scholar]
  • 54.Heron M, Gravett K, Yakovchuk N. Publishing and flourishing: writing for desire in higher education. High Educ Res Dev. 2021;40(3):538–51. [Google Scholar]
  • 55.Maamuujav U, Olson CB, Chung H. Syntactic and lexical features of adolescent L2 students’ academic writing. J Second Lang Writ. 2021;53:100822. doi: 10.1016/j.jslw.2021.100822 [DOI] [Google Scholar]
  • 56.Liang MC, Xiong WX. Wenben fenxi fongju PatCount zai waiyu jiaoxue yu yanjiu zhong de [Applications of PatCount in Foreign language teaching and research]. Waiyu dianhua jiaoxue [Technol Enhanc Foreign Lang Educ]. 2008;123:71–6. [Google Scholar]
  • 57.Kaiser HF. An index of factorial simplicity. Psychometrika. 1974;39(1):31–6. doi: 10.1007/bf02291575 [DOI] [Google Scholar]
  • 58.Tabachnick BG, Fidell LS, Ullman JB. Using multivariate statistics. Pearson; 2007. [Google Scholar]
  • 59.Gray B. More than discipline: uncovering multi-dimensional patterns of variation in academic research articles. Corpora. 2013;8(2):153–81. doi: 10.3366/cor.2013.0039 [DOI] [Google Scholar]
  • 60.Pahor T, Smodiš M, Pisanski Peterlin A. Reshaping authorial presence in translations of research article abstracts. ELOPE. 2021;18(1):169–86. doi: 10.4312/elope.18.1.169-186 [DOI] [Google Scholar]
  • 61.Yu S, Jiang L. L2 university students’ motivational self system in English writing: a sociocultural inquiry. Appl Linguist Rev. 2023;14(3):553–78. [Google Scholar]
  • 62.Liu G, Zhang J. Interactional metadiscourse and author identity construction in academic theses. J Lang Teach Res. 2022;13(6):1313–23. [Google Scholar]
  • 63.Hyland K. Disciplinary identities. Cambridge: Cambridge University Press; 2012. [Google Scholar]
  • 64.Hyland K, Jiang FK. In this paper we suggest: changing patterns of disciplinary metadiscourse. Engl Specif Purp. 2018;51:18–30. [Google Scholar]
  • 65.Becher T. The disciplinary shaping of the profession. In: The academic profession. University of California Press; 2023. p. 271–303. doi: 10.2307/jj.8306011.11 [DOI] [Google Scholar]
  • 66.Biglan A. The characteristics of subject matter in different academic areas. J Appl Psychol. 1973;57(3):195–203. [Google Scholar]
  • 67.Kolb DA. Learning styles and disciplinary differences. Mod Am. 1981:232–5. [Google Scholar]
  • 68.Becher T, Trowler P. Academic tribes and territories. McGraw-Hill Education (UK); 2001. [Google Scholar]
  • 69.Hyland K. Persuasion, interaction and the construction of knowledge: representing self and others in research writing. Int J Engl Stud. 2008;8(2):1–23. [Google Scholar]
  • 70.Schoonen R, Gelderen AV, Glopper KD, Hulstijn J, Simis A, Snellings P, et al. First language and second language writing: the role of linguistic knowledge, speed of processing, and metacognitive knowledge. Lang Learn. 2003;53(1):165–202. [Google Scholar]
  • 71.Hyland K. Authority and invisibility: authorial identity in academic writing. J Pragmat. 2002;34(8):1091–112. [Google Scholar]
  • 72.Hyland K. Stance and engagement: a model of interaction in academic discourse. Discourse Stud. 2005;7(2):173–92. doi: 10.1177/1461445605050365 [DOI] [Google Scholar]
  • 73.He Z. Cohesion in academic writing: a comparison of essays in English written by L1 and L2 university students. TPLS. 2020;10(7):761. doi: 10.17507/tpls.1007.06 [DOI] [Google Scholar]
  • 74.Zhao J. Native speaker advantage in academic writing? Conjunctive realizations in EAP writing by four groups of writers. Ampersand. 2017;4:47–57. doi: 10.1016/j.amper.2017.07.001 [DOI] [Google Scholar]
  • 75.Xu T, Cui F, Zhang B. Challenges in the use of logical connectors for EFL learners: frequency, variety, suitability, and improper use with pedagogical implications. Int J Appl Linguist Engl Lit. 2022;11(3):41. [Google Scholar]
  • 76.Yang S. The functions of the nontarget be in the written interlanguage of Chinese learners of English. Lang Acquis. 2014;21(3):279–303. [Google Scholar]
  • 77.Zhang M, Lan G, Yang K. Grammatical complexity: insights from English for academic purposes teachers. J Second Lang Writ. 2023;60:100974. doi: 10.1016/j.jslw.2023.100974 [DOI] [Google Scholar]
  • 78.Quirk R, Greenbaum S, Leech G, Svartvik J. A comprehensive grammar of the English language. London: Longman; 1985. [Google Scholar]
  • 79.Mahdun M, Chan MY, Thai Yap N, Eng Wong B, Mohd Kasim Z. Overpassivisation in L2 acquisition: an examination of L1 Malay ESL tertiary students’ passivisation of intransitive verbs in English. JSSH. 2023;31(3):995–1013. doi: 10.47836/pjssh.31.3.05 [DOI] [Google Scholar]
  • 80.Hinkel E. Adverbial markers and tone in L1 and L2 students’ writing. J Pragmat. 2003;35(7):1049–68. [Google Scholar]
  • 81.Hyland K. Writing in the disciplines: research evidence for specificity. Taiwan Int ESP J. 2009;1(1):5–22. [Google Scholar]
  • 82.Biber D, Egbert J. Register variation online. Cambridge University Press; 2018. [Google Scholar]

Decision Letter 0

Fasih Ahmed

28 Jan 2025

Dear Dr. Al-Shaibani,

Thank you for submitting your manuscript to PLOS ONE. After careful consideration, we feel that it has merit but does not fully meet PLOS ONE’s publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.

Please submit your revised manuscript by Mar 13 2025 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

  • A rebuttal letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols....

We look forward to receiving your revised manuscript.

Kind regards,

Fasih Ahmed

Academic Editor

PLOS ONE

https://journals.plos.org/plosone/s/file?id=wjVg/PLOSOne_formatting_sample_main_body.pdf and https://journals.plos.org/plosone/s/file?id=ba62/PLOSOne_formatting_sample_title_authors_affiliations.pdf.

Additional Editor Comments:

Dear Author,

The review regarding your article are complete now. The reviewers have suggest the changes. You are required to address all the issues, raised by the worthy reivewers, and comply recommendations.

The detail of the reviewers comments is as follows.

[Note: HTML markup is below. Please do not edit.]

Reviewers' comments:

Reviewer's Responses to Questions

Comments to the Author

1. Is the manuscript technically sound, and do the data support the conclusions?

Reviewer #1: Partly

Reviewer #2: Yes

**********

2. Has the statistical analysis been performed appropriately and rigorously? -->?>

Reviewer #1: Yes

Reviewer #2: I Don't Know

**********

3. Have the authors made all data underlying the findings in their manuscript fully available??>

The PLOS Data policy requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.-->

Reviewer #1: No

Reviewer #2: No

**********

4. Is the manuscript presented in an intelligible fashion and written in standard English??>

Reviewer #1: Yes

Reviewer #2: Yes

**********

Reviewer #1: General Comments

The manuscript presents an interesting and valuable study examining the differences between native and non-native English-speaking academic researchers across multiple disciplines. This is a relevant and timely topic given the increasing global nature of academic publishing. The authors aim to explore the impact of disciplinary variation on the academic writing practices of native and non-native English researchers and how this affects their ability to publish in international journals.

Overall, the paper addresses an important gap in the literature, but there are several areas that need improvement to enhance the quality and clarity of the research.

Specific Comments

1. Clarity and Transparency of Methodology:

While the research design is reasonable, the manuscript lacks detailed transparency regarding the methodology used. Specifically:

Sample Selection: The authors should provide more details on how the sample was selected. Was it randomly selected from various disciplines, or were specific criteria used to select the participants? How were "native" and "non-native" researchers defined, and how large were the sample sizes for each group? Clearer explanation of these factors is necessary to ensure that the study's results are generalizable.

Control and Bias Considerations: There is no clear mention of how potential confounding variables, such as the researchers' academic backgrounds, proficiency in English, or cultural differences, were controlled. The authors should consider discussing how they accounted for these potential biases to strengthen the study's validity.

Data Collection Methods: The paper does not specify what kinds of data were collected—was the analysis based on qualitative data (e.g., textual analysis) or quantitative data (e.g., citation counts, journal acceptance rates)? Additionally, the statistical methods used to analyze the data are not outlined. A clear description of the statistical tests performed is essential to assess the rigor of the analysis.

2. Statistical Analysis:

There is no mention of the statistical methods used to test hypotheses or analyze the data. This omission raises concerns about whether the statistical analysis was performed rigorously. I recommend that the authors provide more detailed information on the types of statistical tests used and whether the sample sizes were sufficiently large to ensure meaningful results. Additionally, the manuscript should include more information on how the data were processed and analyzed.

3. Data Availability:

The manuscript does not mention whether the underlying data supporting the findings will be made publicly available. To enhance the transparency and reproducibility of the study, the authors should make the data available, either in a public repository or as supplementary material. This would allow other researchers to verify the results and build upon this work.

4. Conclusion and Data Support:

While the conclusions drawn in the manuscript are logical, they need to be more firmly supported by data. The authors mention that non-native researchers face challenges in meeting academic writing standards and publishing in high-impact journals, but the manuscript does not provide specific quantitative or qualitative data to back up these claims. A clearer connection between the data and the conclusions would make the study's findings more robust.

5. Language and Presentation:

The manuscript is written in standard academic English, and the ideas are generally presented clearly. However, some sentences are overly complex and could be simplified for better readability. I recommend that the authors revise some of the more convoluted passages to improve the overall flow of the paper.

Suggestions for Improvement:

Enhance Methodological Transparency: The authors should provide more detailed information on the sample selection process, including how participants were chosen and the size of each group (native and non-native researchers). More discussion is needed on how biases and confounding factors were controlled.

Clarify Statistical Methods: The manuscript should include detailed information about the statistical analysis, including the types of tests performed and whether the sample sizes were appropriate. Transparency in data processing and statistical methods will strengthen the validity of the findings.

Provide Data Access: The authors should make the underlying data available for public access, either through a repository or as supplementary material. This will improve the study’s transparency and allow others to verify the results.

Strengthen the Connection Between Data and Conclusions: The authors should ensure that their conclusions are firmly supported by specific data points. This could involve providing more concrete examples or statistical evidence to support their claims regarding the challenges faced by non-native researchers.

Revise for Readability: While the manuscript is generally written well, some sentences could be streamlined for clarity. I recommend that the authors revise sections where the language is overly complicated to improve the overall readability.

Ethical Considerations:

Research Ethics: There are no concerns regarding research ethics in this manuscript. The research appears to be conducted in accordance with ethical standards, and there is no indication of any issues with plagiarism or improper citation practices.

Publication Ethics: The manuscript does not appear to have been previously published or under review elsewhere, based on the information provided. However, I recommend the authors ensure that the paper is not submitted to multiple journals simultaneously and that proper citations and acknowledgements are included.

Final Recommendation:

While the manuscript addresses an important and timely topic, there are several areas that need improvement, particularly in terms of methodological transparency, statistical analysis, and data availability. I recommend major revisions to address these concerns before the manuscript can be considered for publication.

Reviewer #2: Dear authors,

The article is very interesting and informative, the title matches the content , and the sections and subsections are appropriate and well-written, yet the following points are suggested to enrich the work:

1. Yes/no research questions are not recommended, it is better to re-phrase the third question to be 'To what extent language backgrounds ...?'.

2. I suggest having both sections 1 and 2 under one section titled 'Literature Review'.

3. Number 5 in 4.1 should be written as a word.

4. The ampersand '&' should not be used within the text, only used within the in-text citation.

5. Consistency is required in some places such as having POS or PoS, having MDA or the full words, etc.

6. Is the example in "... and the regular expression of which was written as \sabsolute_JJ, ..." correctly written?

7. Figure 1 is to be placed in its correct position within the text so readers can make use of it. The same applies to the other figures.

8. I wonder why the researchers do not use 'interactional' and 'informational' instead of positive and negative with regard to Table 3 and what comes after! The word negative gives the impression of having a very low value. To remove this vagueness, it is better to make it clear that both positive and negative are used to help in drawing the line graphs rather than underestimation.

9. Some pedagogical implications can be added to the conclusion section.

10. Editing is needed to fix some issues.

Best of luck.

**********

what does this mean?). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.

If you choose “no”, your identity will remain anonymous but your review may still be made public.

Do you want your identity to be public for this peer review? For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our Privacy Policy..-->

Reviewer #1: No

Reviewer #2: Yes:Nawal Fadhil AbbasNawal Fadhil AbbasNawal Fadhil AbbasNawal Fadhil Abbas

**********

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

While revising your submission, please upload your figure files to the Preflight Analysis and Conversion Engine (PACE) digital diagnostic tool, https://pacev2.apexcovantage.com/. PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at figures@plos.org. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.

PLoS One. 2026 Apr 24;21(4):e0346776. doi: 10.1371/journal.pone.0346776.r002

Author response to Decision Letter 1


14 Mar 2025

Reviewer 1:

1. The authors should provide more details on how the sample was selected. Was it randomly selected from various disciplines, or were specific criteria used to select the participants? How were "native" and "non-native" researchers defined, and how large were the sample sizes for each group? Clearer explanation of these factors is necessary to ensure that the study's results are generalizable.

Response:

We sincerely thank the reviewer for this insightful comment. We have revised this part in the manuscript and highlighted the revisions in red. This research does not involve any human participants; and it uses the method of textual analysis involving academic research articles written by native English researchers and Chinese researchers. The textual analysis of linguistic differences of these research articles is the main research focus. Therefore, the sample selected in this research is research articles in various disciplines.

As for the selection criteria, firstly, the criterion for selecting the disciplines is based on the Chinese National System of Level One Disciplines for Degree Education, namely, agriculture, art, economics, history, law, literature, management science, medicine, natural science, education, philosophy, engineering and military. To make the native English researchers’ corpus and Chinese researchers’ corpus comparable, we selected the same disciplines and same number of the research articles to build the corpora. When we checked native English researchers’ disciplines in the Web of Science (WoS), there was no military discipline. Therefore, the number of final disciplines involved in this study was 12 by excluding military discipline.

Secondly, after deciding the disciplines, for selecting native English researchers’ research articles (RAs), we used Journal Citation Reports (JCR) of the WoS to look for the most influential 10 international Science Citation Index (SCI), Social Science Citation Index (SSCI) and Science Citation Index Expanded (SCIE) journals in each discipline according to the journals’ highest impact factor since the native English researchers’ RAs were set as the norm and then randomly selecting 5 journals from these 10 in each discipline to select RAs (Table 1). In each journal, 20 RAs from 2019-2023 were collected. Thus, the sample size for native English researchers corpus is 1200 articles. As for selecting the research articles for Chinese researchers, simple random sampling was conducted; that is, academic English journals published in China by referring to Bao and Zhang (2018) were randomly selected because it would reflect the Chinese researchers’ average English writing level in this way. Therefore, the sample size for Chinese researchers corpus is 1200 articles as well. In total, the sample size for the corpora is 2400 research articles.

Pertaining to defining the native researchers, we followed Pan (2018)’s approach “the standard for native researchers is basically the authors’ names should be considered native to English-speaking countries as the UK, the USA, Canada, Australia and New Zealand and were affiliated with an institution in a country where English is spoken as the first language” (Pan 2018, 119).

2. There is no clear mention of how potential confounding variables, such as the researchers' academic backgrounds, proficiency in English, or cultural differences, were controlled. The authors should consider discussing how they accounted for these potential biases to strengthen the study's validity.

Response

We really appreciate this valuable suggestion. However, we did not consider the variables as researchers' academic backgrounds, proficiency in English, or cultural differences when comparing the native English and non-native researchers’ writing differences because the main focus of this research is to explore how different the academic writing written by Chinese researchers with an average writing level from that of high-quality writing level of native English researchers is. We would like to show the Chinese researchers what the writing style a well-written and high-quality native English research article is, and what difference of theirs is from those well-written ones. So we didn’t control the bias of choosing the same level of proficiency in English and academic research background to make comparisons between native English and Chinese researchers. Instead, we selected the native English research articles from the top journals with highest journal impact factor in each discipline in order to set these native English research articles as a benchmark, but for Chinese researchers’ English research articles we selected the articles from journals with varying impact factors and with average writing level. Furthermore, this is possible because this was done in Cao and Xiao (2013)’s study as well. To control the bias in this research, we used the method of triangulation (of investigators), that is, we assigned three research assistants to collect, clean and analyze the data. We trained these assistants first, and made them completely clear about the rule of selecting journals and research articles, how to clean the texts and use the software to analyze the data. For the quantitative data, we used KMO to test the validity of the data. We added section 3.2.3 to explain the validity and reliability in this study and marked them in red.

3. Data Collection Methods: The paper does not specify what kinds of data were collected—was the analysis based on qualitative data (e.g., textual analysis) or quantitative data (e.g., citation counts, journal acceptance rates)? Additionally, the statistical methods used to analyze the data are not outlined. A clear description of the statistical tests performed is essential to assess the rigor of the analysis.

Response

Thank you very much for this valuable comment. We have modified section 3.1 to better present the statistical methods used to analyze the data, and marked in red.

The data collected in this research are published research articles written by the native English and Chinese researchers. To analyze the linguistic features of these research articles, the analysis was based on the qualitative data. We used textual analysis to analyze the articles. The textual analysis was conducted by first extracting the linguistic features suitable for academic writing in the collected research articles, quantifying the frequency of cooccurred linguistic features, then revealing and explaining these cooccurred dimensions after factor analysis. This data collection and data analysis approach was adopted from Biber (1988). Biber pioneered this multi-dimensional analysis approach and then numerous studies followed this approach. The whole data analysis procedure was added in section 3.

4. There is no mention of the statistical methods used to test hypotheses or analyze the data. This omission raises concerns about whether the statistical analysis was performed rigorously. I recommend that the authors provide more detailed information on the types of statistical tests used and whether the sample sizes were sufficiently large to ensure meaningful results. Additionally, the manuscript should include more information on how the data were processed and analyzed.

Response

Thank you very much for this valuable feedback. We have revised this section to clarify the statistical methods and how the data were processed and analyzed. All the revisions were marked in red. As for the sample size, we have 2400 articles as a sample which is sufficient for doing this research because in a factor analysis, the data base should include five times as many texts as linguistic features to be analyzed (Gorsuch 1983, p.332). In our study, 62 linguistic features were identified and included, and 310 (62*5) research articles as a sample are enough for this study, while we have sufficient 2400 articles as a sample for this study.

5. The manuscript does not mention whether the underlying data supporting the findings will be made publicly available. To enhance the transparency and reproducibility of the study, the authors should make the data available, either in a public repository or as supplementary material. This would allow other researchers to verify the results and build upon this work.

Response

Thank you very much for highlighting this issue. We have uploaded all the underlying data supporting our findings as the supplementary material file.

6. While the conclusions drawn in the manuscript are logical, they need to be more firmly supported by data. The authors mention that non-native researchers face challenges in meeting academic writing standards and publishing in high-impact journals, but the manuscript does not provide specific quantitative or qualitative data to back up these claims. A clearer connection between the data and the conclusions would make the study's findings more robust.

Response

We are grateful for this valuable feedback. We have revised the conclusion section to clarify this point and marked it in red.

7. The manuscript is written in standard academic English, and the ideas are generally presented clearly. However, some sentences are overly complex and could be simplified for better readability. I recommend that the authors revise some of the more convoluted passages to improve the overall flow of the paper.

Response

Thank you very much for your suggestion. Our manuscript has been edited by proficient English writer

8. There are no concerns regarding research ethics in this manuscript. The research appears to be conducted in accordance with ethical standards, and there is no indication of any issues with plagiarism or improper citation practices.

Response

Thank you very much for your great feedback. We have already obtained the approval of the ethical application for this study from UCSI university in Malaysia as attached in the supplementary file.

9. The manuscript does not appear to have been previously published or under review elsewhere, based on the information provided. However, I recommend the authors ensure that the paper is not submitted to multiple journals simultaneously and that proper citations and acknowledgements are included.

Response

Thank you very much for your great feedback. Our manuscript has been submitted to PLOS ONE only. For acknowledgements, PLOS ONE required authors should not acknowledge editors and reviewers in the acknowledge section, only to those contribute to the research but not listed as the co-authors.

Reviewer 2:

1. Yes/no research questions are not recommended, it is better to re-phrase the third question to be 'To what extent language backgrounds ...?'

Response

Thank you very much for your good suggestion. We revised third research question and marked it in blue.

2. I suggest having both sections 1 and 2 under one section titled 'Literature Review'.

Response

Thank you very much for your good feedback.

I think you meant section 2 and 3, so we already revised section 2 and 3 under one section 2 titled 'Literature Review' based on your feedback and marked them in blue.

3. Number 5 in 4.1 should be written as a word.

Response

Thank you very much for your good suggestion. I think you meant “the most influential 5 international SCI…” under section 3.1. We revised it and marked it in blue.

4. The ampersand '&' should not be used within the text, only used within the in-text citation.

Response

Thank you very much for your good comment. We changed all the '&' into ‘and’ within the text and marked them in blue.

5. Consistency is required in some places such as having POS or PoS, having MDA or the full words, etc.

Response

Thank you very much for pointing out this issue. We corrected all the POS and MDA and marked them in blue.

6. Is the example in "... and the regular expression of which was written as \sabsolute_JJ, ..." correctly written?

Response

Thank you very much for your good feedback. We checked again the example of \sabsolute_JJ, and the like, all of them were all correct. The format written in this way is a regular expression in computer codes.

7. Figure 1 is to be placed in its correct position within the text so readers can make use of it. The same applies to the other figures.

Response

Thank you very much for your good suggestion. We have followed the journal’s guidelines. The journal guidelines suggest all figures must be submitted separately and only figure captions can be put in where figures should be.

8. I wonder why the researchers do not use 'interactional' and 'informational' instead of positive and negative with regard to Table 3 and what comes after! The word negative gives the impression of having a very low value. To remove this vagueness, it is better to make it clear that both positive and negative are used to help in drawing the line graphs rather than underestimation.

Response

Thank you very much for your feedback. To explain this point, words as 'interactional' and 'informational' are our unique interpretation for each dimension based on the linguistic features extracted after factor analysis; however, positive and negative are used to describe factor loadings. In factor analysis, a loading represents the correlation coefficient (or standardized regression coefficient) between a variable and a factor. The values of loadings typically fall within the range of [-1 and 1]. When a loading is positive, it indicates a positive correlation between the variable and the factor. That is, as the factor score increases, the value of the variable also tends to increase. Conversely, a negative loading implies a negative correlation. In this case, as the factor score rises, the value of the variable tends to decrease. So negative does not imply having a very low value, and it only shows its correlation with this factor. It is positive and negative loading that distinguished various dimensions and reflected the difference between native English researchers’ and Chinese researchers’ academic writing.

9. Some pedagogical implications can be added to the conclusion section.

Response

Thank you very much for your good suggestion. We have revised the conclusion section and marked the revisions in red because the first reviewer also requested this revision

10. Editing is needed to fix some issues.

Response

Thank you very much for your good suggestion. Our manuscript has been edited by proficient English writer.

Attachment

Submitted filename: Resonse to Reviewers.docx

pone.0346776.s008.docx (31.6KB, docx)

Decision Letter 1

Fasih Ahmed

19 Jun 2025

Dear Dr. Al-Shaibani,

Thank you for submitting your manuscript to PLOS ONE. After careful consideration, we feel that it has merit but does not fully meet PLOS ONE’s publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.

Please submit your revised manuscript by Aug 03 2025 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

  • A rebuttal letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols....

We look forward to receiving your revised manuscript.

Kind regards,

Fasih Ahmed

Academic Editor

PLOS ONE

Additional Editor Comments:

Dear Author,

Keeping in view the comments of the reviewers, I recommend you to follow the recommendations made by the reviewers to enrich the quality of the aricle. The reviewers' recommendations mainly belong to the methodogy, results and dicussions.

Regards,

Fasih

[Note: HTML markup is below. Please do not edit.]

Reviewers' comments:

Reviewer's Responses to Questions

Comments to the Author

Reviewer #3: (No Response)

Reviewer #4: (No Response)

**********

2. Is the manuscript technically sound, and do the data support the conclusions??>

Reviewer #3: Yes

Reviewer #4: Yes

**********

3. Has the statistical analysis been performed appropriately and rigorously? -->?>

Reviewer #3: Yes

Reviewer #4: (No Response)

**********

4. Have the authors made all data underlying the findings in their manuscript fully available??>

The PLOS Data policy requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.-->

Reviewer #3: Yes

Reviewer #4: Yes

**********

5. Is the manuscript presented in an intelligible fashion and written in standard English??>

Reviewer #3: Yes

Reviewer #4: Yes

**********

Reviewer #3: 1. Typo in Line 226: "Twentyfull-text articles in each"

2. The study shows the differences between native authors and Chinese authors. However, one suggestion by thestudy is : "This finding indicates Chinese researchers should exhibit their authorial stance and interact with the readers with confidence and employ more interactive devices to make their writing coherent and explicit." Why? Can the way native authors promote the influence of articles? Any evidences?

Reviewer #4: As some of my comments overlap with those provided by other reviewers, I will restrict mine to the following points:

• Line 52: The verb tense and overall language use should be reviewed to ensure consistency with the formal tone expected in academic writing.

• Lines 73, 74, 127, 188, 372, 864: The readability of these sections could be improved. Revising sentence structure and ensuring clarity would enhance the overall flow and accessibility of the text.

• Abstract Discrepancy: There is an inconsistency between the two versions of the abstract. Notably, the second version omits reference to Nini’s MAT, which is included in the first. The abstracts should be aligned to maintain consistency and accurately represent the study’s scope.

• Page 66 (Line 226): Spelling errors, such as “pattens,” should be corrected. A careful proofreading of the manuscript is recommended to eliminate such typographical issues.

• Research Question 3: The response to Research Question 3 appears underdeveloped. Its current focus on language of publication is overly restrictive and does not reflect the potential complexity of the issue. A more comprehensive exploration is needed.

• Line 316: The inclusion of the software download link in the body of the paper is unnecessary.

• Figures and Layout: Figures are not effectively positioned within the text, which disrupts the logical flow and readability. They should be placed adjacent to the relevant discussions to support the argument and maintain textual coherence.

• Clarity of Contribution: The contribution of the paper, particularly the proposed novel MDA model, would benefit from a visual representation. A schematic or diagrammatic illustration would make the model’s structure and innovative aspects more accessible to readers.

• Strength of Argumentation: The argument that “language background and discipline cannot only account for the variation on dimension scores alone, but also influence the dimension scores altogether” is currently unconvincing. The concept of "language background" is applied and analyzed in a rather narrow way. It is recommended that the authors revisit and refine this argument, potentially in relation to a revised and more robust treatment of Research Question 3, to enhance the theoretical and empirical soundness of their claims.

**********

what does this mean?). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.

If you choose “no”, your identity will remain anonymous but your review may still be made public.

Do you want your identity to be public for this peer review? For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our Privacy Policy..-->

Reviewer #3: No

Reviewer #4: Yes:Dima FarhatDima FarhatDima FarhatDima Farhat

**********

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

While revising your submission, please upload your figure files to the Preflight Analysis and Conversion Engine (PACE) digital diagnostic tool, https://pacev2.apexcovantage.com/. PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at figures@plos.org. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.

PLoS One. 2026 Apr 24;21(4):e0346776. doi: 10.1371/journal.pone.0346776.r004

Author response to Decision Letter 2


31 Jul 2025

Dear Editor Dr. Ahmed,

We would like to sincerely thank you and the reviewers for your valuable time and effort in evaluating our manuscript, A multi-dimensional analysis of native and non-native academic research articles in twelve disciplines, and for providing us with constructive feedback. We appreciate the thoughtful comments and suggestions, which have significantly helped us improve the quality of our research.

We have carefully addressed all the comments and concerns raised by the reviewers, and we have made the necessary revisions to the manuscript accordingly. Below, we provide a point-by-point response to each of the reviewers’ comments, detailing the changes we have made. We hope that these revisions meet your expectations and further strengthen the manuscript.

We hope that the revised manuscript is now suitable for publication in PLOS ONE.

Sincerely,

The authors

Comments from reviewer #3

1. Typo in Line 226: "Twentyfull-text articles in each…"

Response:

We sincerely thank the reviewer for this careful comment. We have revised this sentence in Section 3.1, line 226, page 13.

2. The study shows the differences between native authors and Chinese authors. However, one suggestion by the study is: "This finding indicates Chinese researchers should exhibit their authorial stance and interact with the readers with confidence and employ more interactive devices to make their writing coherent and explicit." Why? Can the way native authors promote the influence of articles? Any evidence?

Response:

We really appreciate this valuable comment. The reasons for this suggestion can be found in Section 4.2. By tagging, counting and analyzing the linguistic features of the collected research articles, we found that, for example, on Dimension One, native English researchers showed a heavy reliance on the use of linguistic features, such as first person pronouns (e.g. we, our), reporting verbs (e.g. require, propose), adverbs (e.g. directly, instead, wholly, surprisingly, conceptually), attributive adjectives (previous, multiple, simple, significant) and present tense. These linguistic features create an image of involving the author’s interaction with the readers to highlight the author’s stance explicitly (detailed explanation can be found in Section 4.2.1). In contrast, Chinese researchers prefer to use many nouns, nominalizations and prepositions marking the article with dense information, indicating that Chinese researchers avoid involving and interacting with the readers to conceal their authorial identity and opinion, but tend to convey information objectively. Other findings and explanations on Dimension Two, Three and Four can be found from section 4.2.2 through section 4.4.4. Therefore, we identified, tagged and counted linguistic features frequencies using Stanford Part of Speech (PoS) tagger and PatCount software. Then we conducted a factor analysis and textual analysis to obtain these findings in our research (pp.17-23).

Native English authors indirectly influence non-native authors. Furthermore, our data on native English researchers are all empirical data collected from research articles published in top international journals and authored by native English researchers. They can be set as a benchmark of high-quality writing (Lines 223-226). Therefore, our findings related to the writing style of native English researchers are not aimed to promote their style, rather it is based on the empirical data we collected from native English researchers’ high-quality research articles which show us the way how most influential articles were written by native English researchers. In other words, we need a benchmark for Chinese researchers to learn from, so that the Chinese researchers can imitate the native English researchers’ writing style to improve their writing.

Comments from reviewer #4

1. Line 52: The verb tense and overall language use should be reviewed to ensure consistency with the formal tone expected in academic writing.

Response:

Thank you very much for your good suggestion. For this and all following comments on language use, we have already sent our manuscript to Luzhou Jiayi Language Service Company for language editing by professional language editors, and we submitted the language editing certificate for you to check.

2. Lines 73, 74, 127, 188, 372, 864: The readability of these sections could be improved. Revising sentence structure and ensuring clarity would enhance the overall flow and accessibility of the text.

Response:

The response is in number one above.

3. Abstract Discrepancy: There is an inconsistency between the two versions of the abstract. Notably, the second version omits reference to Nini’s MAT, which is included in the first. The abstracts should be aligned to maintain consistency and accurately represent the study’s scope

Response:

Thank you very much for your comment. The first two reviewers requested a revision and thus we revised the whole manuscript including the abstract based on the first-round comments given by the first-round reviewers.

4. Page 66 (Line 226): Spelling errors, such as “pattens,” should be corrected. A careful proofreading of the manuscript is recommended to eliminate such typographical issues.

Response:

Thank you very much for your comment. There is no page 66, but we checked Line 66 and this spelling error was on line 66. We have corrected the spelling errors in the whole manuscript. Besides, we sent our manuscript to Luzhou Jiayi Language Service Company for language editing by professional language editors and submitted the language editing certificate.

5. Research Question 3: The response to Research Question 3 appears underdeveloped. Its current focus on language of publication is overly restrictive and does not reflect the potential complexity of the issue. A more comprehensive exploration is needed.

Response:

Thank you very much for your comment. In fact, the focus of RQ 3 (and even the whole manuscript) is not on the language of publication, but rather the language of writing style differences between native English and Chinese researchers’ academic writing. In other words, there are differences in their academic writing style as each group has a distinctive pattern as reported in our study. We focused on such differences by employing a multi-dimensional model (line 790-794, pp.47). Hence, the answer of RQ3 was explained as how researchers with different language backgrounds (native English researchers and Chinese researchers) write differently on each dimension of our developed model. Through calculating the dimension score of native English researchers’ writing and Chinese researchers’ writing in 12 disciplines on each dimension (Line 423-444, Line 536-560, Line 643-664, Line 719-733), we made a comparison. After that, we used ANOVA to test whether the language background (native English and Chinese) and discipline can influence the dimension score or not. The detailed explanation for RQ3 can be found from Section 4.2.1through 4.2.4. For example, ANOVA tests showed that the main effect of language backgrounds on Dimension 1 was statistically significant, suggesting that Dimension One can differentiate research articles written by native English researchers and Chinese researchers. By doing this, RQ3 (how language backgrounds and disciplines influence the dimension scores of native English and Chinese researchers’ academic writing) was answered.

6. Line 316: The inclusion of the software download link in the body of the paper is unnecessary.

Response:

Thank you very much for your suggestion. We have deleted the link.

7. Figures and Layout: Figures are not effectively positioned within the text, which disrupts the logical flow and readability. They should be placed adjacent to the relevant discussions to support the argument and maintain textual coherence.

Response:

Thank you very much for your comment. However, we are following the journal’s requirement for formatting the submission. Figures must be submitted separately and only figure captions can be put in where figures should be. Kindly take note we are not allowed to breach the journal guidelines. All the figures are placed after the references list based on the submission system.

8. Clarity of Contribution: The contribution of the paper, particularly the proposed novel MDA model, would benefit from a visual representation. A schematic or diagrammatic illustration would make the model’s structure and innovative aspects more accessible to readers.

Response:

Thank you very much for your suggestion. We followed your suggestion, and now there is a diagram as Figure 5 added in the conclusion section in the manuscript.

9. Strength of Argumentation: The argument that “language background and discipline cannot only account for the variation on dimension scores alone, but also influence the dimension scores altogether” is currently unconvincing. The concept of "language background" is applied and analyzed in a rather narrow way. It is recommended that the authors revisit and refine this argument, potentially in relation to a revised and more robust treatment of Research Question 3, to enhance the theoretical and empirical soundness of their claims.

Response:

Thank you very much for your comment. In our study, language background refers to native Chinese language background and native English language background. The whole manuscript only focused on the comparison of language difference between native English researchers and Chinese researchers in academic writing style. Therefore, our argument “language background would influence the dimension score” was obtained from ANOVA test, meaning that native English researchers wrote differently from Chinese researchers, and the difference is significant, i.e. judging from the linguistic features of the research articles, research articles written by native English researchers and Chinese researchers can be differentiated. Anyway, since you advise us to revise this argument, we deleted and revised some parts of this argument from page 49-50.

Attachment

Submitted filename: Resonse_to_Reviewers_auresp_2.docx

pone.0346776.s009.docx (30.5KB, docx)

Decision Letter 2

Fasih Ahmed

11 Sep 2025

Dear Dr. Al-Shaibani,

Thank you for submitting your manuscript to PLOS ONE. After careful consideration, we feel that it has merit but does not fully meet PLOS ONE’s publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.

Please submit your revised manuscript by Oct 26 2025 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

  • A rebuttal letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols....

We look forward to receiving your revised manuscript.

Kind regards,

Fasih Ahmed

Academic Editor

PLOS ONE

Journal Requirements:

If the reviewer comments include a recommendation to cite specific previously published works, please review and evaluate these publications to determine whether they are relevant and should be cited. There is no requirement to cite these works unless the editor has indicated otherwise.

[Note: HTML markup is below. Please do not edit.]

Reviewers' comments:

Reviewer's Responses to Questions

Comments to the Author

Reviewer #3: All comments have been addressed

Reviewer #5: All comments have been addressed

**********

2. Is the manuscript technically sound, and do the data support the conclusions??>

Reviewer #3: Yes

Reviewer #5: Partly

**********

3. Has the statistical analysis been performed appropriately and rigorously? -->?>

Reviewer #3: Yes

Reviewer #5: Yes

**********

4. Have the authors made all data underlying the findings in their manuscript fully available??>

The PLOS Data policy requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.-->

Reviewer #3: Yes

Reviewer #5: Yes

**********

5. Is the manuscript presented in an intelligible fashion and written in standard English??>

Reviewer #3: Yes

Reviewer #5: No

**********

Reviewer #3: All my comments have been properly addressed. I have no further questions now. The article can be accepted.

Reviewer #5: This study aimed to compare research articles written by native and non-native English speakers across twelve disciplines. However, I find it unclear what new insights this work contributes to the existing body of literature. For instance, the method of determining an author’s first language is problematic. Without directly contacting authors, it is virtually impossible to verify whether a given article was written by an L1 English speaker. Relying on names and institutional affiliations as proxies for “native” identity risks serious misclassification; for example, a Chinese author may have lived and studied in an English-speaking country for an extended period before returning to China, while many L2 scholars work at English-medium institutions.

In addition, the study’s corpus design raises concerns. Corpus-based research typically faces challenges in controlling the number and distribution of research articles, and the criteria here are not sufficiently justified. With respect to Research Question 3, although ANOVA tests are reported, the discussion is underdeveloped. The implications of significant interactions are only briefly noted and not adequately theorized in terms of disciplinary writing practices and cultures.

Given these fundamental concerns regarding the validity of author classification, corpus construction, and theoretical interpretation, I recommend rejection of this paper.

**********

what does this mean?). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.). If published, this will include your full peer review and any attached files.

If you choose “no”, your identity will remain anonymous but your review may still be made public.

Do you want your identity to be public for this peer review? For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our For information about this choice, including consent withdrawal, please see our Privacy Policy..-->

Reviewer #3: No

Reviewer #5: No

**********

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

While revising your submission, please upload your figure files to the Preflight Analysis and Conversion Engine (PACE) digital diagnostic tool, https://pacev2.apexcovantage.com/. PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at . PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at figures@plos.org. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.. Please note that Supporting Information files do not need this step.

PLoS One. 2026 Apr 24;21(4):e0346776. doi: 10.1371/journal.pone.0346776.r006

Author response to Decision Letter 3


29 Nov 2025

Comments from reviewer #5

1. I find it unclear what new insights this work contributes to the existing body of literature.

Response

Thank you very much for this valuable feedback. We reanalyzed our data by using a nonparametric approach, an Aligned Rank Transform (ART ANOVA) instead of previous two-way ANOVA because the data fall under non-normal distribution. This yielded new findings on the disciplinary variations, i.e., even for hard sciences that emphasize objectivity and logical rigor, native English researchers exhibit more flexibility by appropriately incorporating authorial identity and engagement with the readers, rather than being informational and objective as it should be. This finding demonstrated native English researchers strategically declared their authorial stance within the disciplinary norms due to their greater writing fluency and more flexible expressions in writing, but the Chinese researchers’ “being informational in writing” tend to be conservative as they follow pre-set disciplinary conventions. While in soft disciplines, the difference is insignificant. Academic writing in soft disciplines permits a diversity of rhetorical styles and modes of expressions. This intrinsic variability may obscure differences induced by various language backgrounds, making statistical differences insignificant.

Another contribution in this research is that we developed a novel MDA model for academic writing in research articles genre and introduced a method of tagging by using PatCount which identifies and codes regular expressions to tag specific linguistic features. This is another contribution because tagging specific linguistic features has been a stumbling block for many corpus linguistics researchers to develop customized MDA model based on their research purpose.

The detailed revised discussion and conclusion sections can be found on pages 28-30, 34-42, 46-47, 53, 57-61.

2. For instance, the method of determining an author’s first language is problematic. Without directly contacting authors, it is virtually impossible to verify whether a given article was written by an L1 English speaker. Relying on names and institutional affiliations as proxies for “native” identity risks serious misclassification; for example, a Chinese author may have lived and studied in an English speaking country for an extended period before returning to China, while many L2 scholars work at English-medium institutions.

Response

Thank you for raising this important point. We fully agree with the reviewer.

We adopted this method for the following reasons: First, directly contacting all authors to confirm their L1 identity was not feasible for a study of this scale with 2400 articles. Second, the use of the first author’s name and institutional affiliation as proxy indicators is a widely adopted method in Applied Linguistics and English for Academic Purposes (EAP) research for conducting large-scale analyses. The following studies adopted the author’s name and the nation of institution to determine native researchers. For example, Abdi and Farrokhi (2015) compared the use and functions of first-person pronouns in L1 and L2 research articles of Applied Linguistics (AL), Mechanical Engineering (ME), and Medicine (MED) whereby “Articles were judged to be L1 or L2 considering the authors’ names and affiliations” (p.158). In another research, “verification about author nativeness was not ensured by contacting them. Authors’ status of nationality was presumed based on their names or nationalities (Yagizi & Demir, 2015, p.15). In Pan (2018)’s research, “L1 research articles were selected from prestigious international journals (measured by impact factors) whose authors had a first and last name that can be considered native to English-speaking countries and were affiliated with an institution in a country where English is spoken as the first language” (p. 119). Kareema and Hakmal (2023) collected abstracts written by Sri Lankan scholars and Western scholars related to the English field randomly stating that “Both journals are available online and the articles were all checked in terms of the author's nationality” (p. 323). In Cao and Xiao’s (2013) research published on Corpora, they built the corpora of English abstracts written by native English and native Chinese writers from twelve academic disciplines. However, they only considered the most prestigious journals in each discipline as reflected by their impact factors to select native English speakers’ abstracts, and also to select journals with varying impact factors for the Chinese speakers’ abstracts.

We recognize that this classification cannot fully capture the complexity of scholarly identity, but it is widely used in our field. Meanwhile, in our study, we used a double-safeguard method to maximize the chances of the included researchers in native researchers corpus are native researchers. This double-safeguard way was explained in detail in the manuscript on pages 14-15 as well. First, the nation of the institution affiliated with which the paper is published was determined. Only institutional affiliation belongs to nations in the inner circle of English-speaking countries, specifically the United States, United Kingdom, Canada, Australia, and New Zealand were included. Papers with which the nationality of the institution cannot be determined or not belong to these five English-speaking nations were discarded. Then we determined whether the first author’s nationality was matched with that of their publishing institution, by using a name origin database. If the ethnic origin of the first author’s surname was consistent with the nationality of the publishing institution, the author was regarded as originating from that nation. Papers that did not meet this criterion were excluded from the final corpus. Referring to the reviewer’s comment here “many L2 scholars work at English-medium institutions”, we first checked whether the English-medium institutions are in the United States, United Kingdom, Canada, Australia, or New Zealand, if not, they are excluded from this research; if yes, we then checked the L2 scholar’s surname. We know that L2 scholars have unique surnames different from native English speakers, for example, Chinese L2, Japanese L2, German L2, Russia L2, Arabic L2, Portuguese L2 and so on. Only when the nationality of the surname of the first author belongs to is the same as where the institution is located, we regarded the research as done by native researchers. In this way, we may have excluded research articles written by native researchers, but we insured that the researchers included in our research are maximally native.

The reviewer also mentioned here “a Chinese author may have lived and studied in an English speaking country for an extended period before returning to China”, yes, we did not set the standard to select Chinese scholars because in this study we did not consider Chinese scholar’s English proficiency as a variable, we explained this on page 15, we collected all English levels’ Chinese scholars in order to reflect the Chinese researchers’ average English writing level.

Cited sources list

Abdi, J., & Farrokhi, F. (2015). Investigating the projection of authorial identity through first person pronouns in L1 and L2 English research articles. International Journal of Language and Literature, 3(1), 156-168. (Scopus Q1)

Cao, Y., & Xiao, R. (2013). A multi-dimensional contrastive study of English abstracts by native and non-native writers. Corpora, 8(2), 209-234. (ESCI Q3)

Kareema, M. I. F., & Hakmal, M. H. M. (2023). A Genre Analysis of Abstracts Written by Sri Lankan and Western Academics on Social Science Discipline. KALAM International Research Journal, 16(1), 319-330.

Pan, F. (2018). A multidimensional analysis of L1–L2 differences across three advanced levels. Southern African Linguistics and Applied Language Studies, 36(2), 117-131. (SSCI Q4)

Yagiz, O., & Demir, C. (2015). A comparative study of boosting in academic texts: A contrastive rhetoric. International Journal of English Linguistics, 5(4), 12.

3. In addition, the study’s corpus design raises concerns. Corpus-based research typically faces challenges in controlling the number and distribution of research articles, and the criteria here are not sufficiently justified.

Response

Thank you very much for this valuable comment. We noticed this and justified it in our manuscript as explained on page 16. First, for the number of research articles selected for the corpus design, we referred to Gorsuch (1983) whereby in a factor analysis, the database should include five times as many texts as linguistic features to be analyzed (p.332). In our study, 62 linguistic features were ultimately identified and included. Thus, 310 research articles are enough as far as the number of research articles for this study’s corpus design is concerned. In our study, we have collected 2, 400 articles as a sample for our corpus design which is sufficient for corpus-based research (Gorsuch, 1983).

As for the distribution of research articles, we aim to distinguish the disciplinary variation among the selected research articles, so we determined 12 disciplines for the corpus. The criteria for selecting the disciplines followed the Chinese National System of Level One Disciplines for Degree Education, and the selection and distribution of the research articles were fully explained on page 13 through page 16.

4. With respect to Research Question 3, although ANOVA tests are reported, the discussion is underdeveloped. The implications of significant interactions are only briefly noted and not adequately theorized in terms of disciplinary writing practices and cultures.

Response

Thank you very much for this valuable feedback. To revise the discussion section for RQ3 with theoretical and practical implications, we re-analyzed our data. Due to the non-normal distribution of the data, a nonparametric approach, an Aligned Rank Transform (ART ANOVA) was conducted to replace the previous Two-way ANOVA. Then we reanalyzed and revised the results and re-wrote the discussion section. This was done to yield deeper and insightful implications in the discussion, especially on Dimension 1, and conclusion sections. Most pages of extensive discussion of Dimension 1 were provided because only on Dimension 1, the main effect of discipline on the dimension score was statistically significant based on our new ART ANOVA test. Then a focused and deeper discussion on the disciplinary differences was done according to the reviewers’ comments. The revisions made can be found on pages 28-30, 34-42, 46-47, 53, 57-61.

Attachment

Submitted filename: Resonse_to_Reviewers_auresp_3.docx

pone.0346776.s010.docx (34.4KB, docx)

Decision Letter 3

Xiaoming Tian

22 Feb 2026

Dear Dr. Al-Shaibani,

Thank you for submitting your manuscript to PLOS ONE. After careful consideration, we feel that it has merit but does not fully meet PLOS ONE’s publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.

Please submit your revised manuscript by Apr 08 2026 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

  • A letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at . Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols....

We look forward to receiving your revised manuscript.

Kind regards,

Xiaoming Tian, Ph.D.

Academic Editor

PLOS One

Journal Requirements:

If the reviewer comments include a recommendation to cite specific previously published works, please review and evaluate these publications to determine whether they are relevant and should be cited. There is no requirement to cite these works unless the editor has indicated otherwise.

Additional Editor Comments (if provided):

The author has provided detailed and thoughtful responses to the reviewer’s comments, but there are several areas that could benefit from further clarification. First, the novelty of the study’s contributions could be more explicitly linked to gaps in the existing literature to highlight its unique insights. Second, the issue of misclassification regarding scholars with high English proficiency has been addressed, but further elaboration on how this issue was dealt with could make the methodology more robust. Additionally, while the corpus design has been justified, more clarity is needed regarding the representativeness of the sample, particularly in terms of the diversity within disciplines. Expanding on how the corpus accounts for variations within each discipline would help strengthen the generalizability of the findings.

Moreover, a more balanced approach to comparing the writing styles of native and non-native researchers would improve the objectivity and comprehensiveness of the study. While the author highlights areas where Chinese researchers can improve, such as authorial stance and interaction with readers, it would be important to also consider the strengths of native English writers. For example, Chinese scholars often excel at creating concise and information-dense texts, which could offer valuable lessons for English-native researchers. By presenting a more reciprocal view, where both groups can learn from each other, this study would offer a more comprehensive and culturally inclusive perspective on academic writing practices.

[Note: HTML markup is below. Please do not edit.]

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

To ensure your figures meet our technical requirements, please review our figure guidelines: https://journals.plos.org/plosone/s/figures

You may also use PLOS’s free figure tool, NAAS, to help you prepare publication quality figures: https://journals.plos.org/plosone/s/figures#loc-tools-for-figure-preparation.

NAAS will assess whether your figures meet our technical requirements by comparing each figure against our figure specifications.

PLoS One. 2026 Apr 24;21(4):e0346776. doi: 10.1371/journal.pone.0346776.r008

Author response to Decision Letter 4


22 Mar 2026

1. First, the novelty of the study’s contributions could be more explicitly linked to gaps in the existing literature to highlight its unique insights.

Response

Thank you very much for this valuable feedback. We have rewritten the Conclusion (page 58-63) and added a new opening paragraph to the Conclusion that explicitly positions our contributions against MDA limitations: lack of academic-genre frameworks and effective method of tagging linguistic features for specific domains, and insufficient disciplinary nuances. We now state: “This study addresses these gaps by: (1) selecting 62 academic-specific linguistic features; (2) using PatCount for automated tagging; and (3) deriving four novel dimensions for academic writing across 12 disciplines”.

2. Second, the issue of misclassification regarding scholars with high English proficiency has been addressed, but further elaboration on how this issue was dealt with could make the methodology more robust.

Response

Thank you for your comment. To clarify our classification approach, which builds directly on our prior response and established field practices, in the revised manuscript (Methods section, pages 14 &15), we have elaborated our double-safeguard method: (1) restricting to institutions in core English-speaking countries (US, UK, Canada, Australia, New Zealand), excluding non-core or undetermined affiliations; and (2) verifying first-author surname origin matches the institution’s country via name database, specifically filtering out L2-associated surnames (e.g., Chinese or Japanese). This addresses Round 3 reviewer’s concerns about L2 scholars at English-medium institutions (by country restriction first) and extended immersion (acknowledged as a proxy limit). For Chinese authors, we intentionally sampled across proficiency levels (as noted on page 15) to capture average writing patterns because proficiency is not a variable. Therefore, no changes to the results are needed as sensitivity to misclassification is minimal given the sample size.

To elaborate further, this double-safeguard proxy method virtually eliminates misclassification of L2 authors as L1 by requiring both institutional location in core English-speaking countries and surname origin matching that country. While rare edge cases exist (e.g., naturalized citizens who legally changed their ethnic surnames to Western names), such instances are exceptional in academic publishing and unlikely to affect group-level patterns in a corpus of 2,400 articles. However, we have added a Limitations section (pages 57 & 58) to explain this classification.

3. Additionally, while the corpus design has been justified, more clarity is needed regarding the representativeness of the sample, particularly in terms of the diversity within disciplines. Expanding on how the corpus accounts for variations within each discipline would help strengthen the generalizability of the findings.

Response

We thank you for drawing attention to the issue of disciplinary diversity. In the revised manuscript, we clarified that the corpus was constructed based on the Chinese National System of Level One Disciplines for Degree Education which includes 12 broad disciplines (agriculture, art, economics, history, law, literature, management science, medicine, natural science, education, philosophy, and engineering). Within each discipline, we sampled research articles without further subdiscipline classification, as our focus was on broad disciplinary patterns rather than highly specialized subfields. We now explicitly acknowledge this as a limitation in Section 4 (pages 58), noting that our findings should be interpreted as generalizing average tendencies at the level of major disciplines, and we suggest that future research can build larger, subdiscipline-stratified corpora to examine finer-grained variation.

4. Moreover, a more balanced approach to comparing the writing styles of native and non-native researchers would improve the objectivity and comprehensiveness of the study. While the author highlights areas where Chinese researchers can improve, such as authorial stance and interaction with readers, it would be important to also consider the strengths of native English writers. For example, Chinese scholars often excel at creating concise and information-dense texts, which could offer valuable lessons for English-native researchers. By presenting a more reciprocal view, where both groups can learn from each other, this study would offer a more comprehensive and culturally inclusive perspective on academic writing practices.

Response

Thank you very much for such an insightful suggestion. We revised our whole manuscript and offered a more reciprocal view. We considered the strengths of Chinese researchers for the four dimensions and revised our Dimension 1 on pages 34 & 35, Dimension 2 on page 47, Dimension 3 on page 52, Dimension 4 on page 57. We also revised our abstract and our conclusion on page 61. The revised texts are below:

Dimension 1 (pages 34 & 35): While native researchers excel at authorial engagement and elaboration, Chinese researchers demonstrate strengths in information density and conciseness, efficiently packing content with minimal redundancy, a style valued in fast-paced, technical fields as engineering. This aligns with EAP research showing L2 writers’ precision can model concise argumentation for L1 peers. Thus, both groups offer mutual lessons of natives in interactivity and Chinese in streamlined reporting.

Dimension 2 (page 47): Dimension 2 differentiated native researchers’ interactive argumentation from Chinese researchers’ static descriptive style. This reflects proficiency differences that natives chain ideas linearly with complex connectors, while Chinese EFL writers favor direct A is B structures due to vocabulary and syntax constraints. Yet this L2 simplicity confers strengths in clarity and conciseness, avoiding native overelaboration that can obscure meaning. Such straightforwardness enhances readability in disciplines with heavy jargons.

Dimension 3 (page 52): Chinese impersonality offers strengths in objectivity and humility, aligning with disciplinary conventions that prioritize collective knowledge over individual voice. This restrained writing style models cultural sensitivity for natives, who risk perceived bias through overt self-positioning.

Dimension 4 (page 57): Dimension 4 contrasted native researchers’ explicit elaborating style with Chinese researchers’ simplified reporting style. Natives excelled in elaboration due to superior proficiency. Chinese simplicity provides strengths in efficiency and readability, prioritizing research content over linguistic display. This reader-centered writing style models concise impact for natives prone to overelaboration.

Conclusion: In addition to the above-mentioned key findings from different language backgrounds and disciplines, this research also contributed to enriching the existing MDA research by providing a newly complementary perspective. Unlike previous studies stressing the differences, this research innovatively emphasized the mutual strengths from these explored differences. This balanced comparison reveals mutual strengths of Chinese conciseness (Dimensions 1 & 4) model clarity and efficiency for natives prone to verbosity; native involvement and interaction (Dimensions 1 & 4) guides Chinese toward engagement. Chinese impersonality (Dimension 3) offers objectivity where natives risk bias; native logic (Dimension 2) complements L2 simplicity.

Attachment

Submitted filename: Resonse to Reviewers & Editor.docx

pone.0346776.s011.docx (27.8KB, docx)

Decision Letter 4

Xiaoming Tian

24 Mar 2026

A multi-dimensional analysis of native and non-native academic research articles in twelve disciplines

PONE-D-24-57048R4

Dear Dr. Al-Shaibani,

We’re pleased to inform you that your manuscript has been judged scientifically suitable for publication and will be formally accepted for publication once it meets all outstanding technical requirements.

Within one week, you’ll receive an e-mail detailing the required amendments. When these have been addressed, you’ll receive a formal acceptance letter and your manuscript will be scheduled for publication.

An invoice will be generated when your article is formally accepted. Please note, if your institution has a publishing partnership with PLOS and your article meets the relevant criteria, all or part of your publication costs will be covered. Please make sure your user information is up-to-date by logging into Editorial Manager at Editorial Manager® and clicking the ‘Update My Information' link at the top of the page. For questions related to billing, please contact  and clicking the ‘Update My Information' link at the top of the page. For questions related to billing, please contact  and clicking the ‘Update My Information' link at the top of the page. For questions related to billing, please contact  and clicking the ‘Update My Information' link at the top of the page. For questions related to billing, please contact billing support....

If your institution or institutions have a press office, please notify them about your upcoming paper to help maximize its impact. If they’ll be preparing press materials, please inform our press team as soon as possible -- no later than 48 hours after receiving the formal acceptance. Your manuscript will remain under strict press embargo until 2 pm Eastern Time on the date of publication. For more information, please contact onepress@plos.org.

Kind regards,

Xiaoming Tian, Ph.D.

Academic Editor

PLOS One

Additional Editor Comments (optional):

Reviewers' comments:

Acceptance letter

Xiaoming Tian

PONE-D-24-57048R4

PLOS One

Dear Dr. Al-Shaibani,

I'm pleased to inform you that your manuscript has been deemed suitable for publication in PLOS One. Congratulations! Your manuscript is now being handed over to our production team.

At this stage, our production department will prepare your paper for publication. This includes ensuring the following:

* All references, tables, and figures are properly cited

* All relevant supporting information is included in the manuscript submission,

* There are no issues that prevent the paper from being properly typeset

You will receive further instructions from the production team, including instructions on how to review your proof when it is ready. Please keep in mind that we are working through a large volume of accepted articles, so please give us a few days to review your paper and let you know the next and final steps.

Lastly, if your institution or institutions have a press office, please let them know about your upcoming paper now to help maximize its impact. If they'll be preparing press materials, please inform our press team within the next 48 hours. Your manuscript will remain under strict press embargo until 2 pm Eastern Time on the date of publication. For more information, please contact onepress@plos.org.

You will receive an invoice from PLOS for your publication fee after your manuscript has reached the completed accept phase. If you receive an email requesting payment before acceptance or for any other service, this may be a phishing scheme. Learn how to identify phishing emails and protect your accounts at https://explore.plos.org/phishing.

If we can help with anything else, please email us at customercare@plos.org.

Thank you for submitting your work to PLOS ONE and supporting open access.

Kind regards,

PLOS ONE Editorial Office Staff

on behalf of

Dr. Xiaoming Tian

Academic Editor

PLOS One

Associated Data

    This section collects any data citations, data availability statements, or supplementary materials included in this article.

    Supplementary Materials

    S1 Table. Frequencies of all linguistic features in 2400 articles for factor analysis.

    (XLSX)

    pone.0346776.s001.xlsx (1,003.5KB, xlsx)
    S2 Table. Z-score of each article on four factors after factor analysis.

    (XLSX)

    pone.0346776.s002.xlsx (129KB, xlsx)
    S3 Table. Mean dimension score of each discipline on four factors.

    (XLSX)

    pone.0346776.s003.xlsx (9.9KB, xlsx)
    S4 Table. Results of factor analysis (KMO, cumulative variance explained and pattern matrix).

    (DOC)

    pone.0346776.s004.doc (7.2MB, doc)
    S5 Appendix. Linguistic features selected in this study.

    (PDF)

    pone.0346776.s005.pdf (60.6KB, pdf)
    S6 Appendix. List of abbreviations.

    (PDF)

    pone.0346776.s006.pdf (59KB, pdf)
    Attachment

    Submitted filename: Resonse to Reviewers.docx

    pone.0346776.s008.docx (31.6KB, docx)
    Attachment

    Submitted filename: Resonse_to_Reviewers_auresp_2.docx

    pone.0346776.s009.docx (30.5KB, docx)
    Attachment

    Submitted filename: Resonse_to_Reviewers_auresp_3.docx

    pone.0346776.s010.docx (34.4KB, docx)
    Attachment

    Submitted filename: Resonse to Reviewers & Editor.docx

    pone.0346776.s011.docx (27.8KB, docx)

    Data Availability Statement

    All relevant data are within the paper and its Supporting Information files.


    Articles from PLOS One are provided here courtesy of PLOS

    RESOURCES