Skip to main content
PLOS One logoLink to PLOS One
. 2023 May 18;18(5):e0286067. doi: 10.1371/journal.pone.0286067

Causal implicatures from correlational statements

Samuel J Gershman 1,*, Tomer D Ullman 1,*
Editor: Micah B Goldwater2
PMCID: PMC10194916  PMID: 37200364

Abstract

Correlation does not imply causation, but this does not necessarily stop people from drawing causal inferences from correlational statements. We show that people do in fact infer causality from statements of association, under minimal conditions. In Study 1, participants interpreted statements of the form “X is associated with Y” to imply that Y causes X. In Studies 2 and 3, participants interpreted statements of the form “X is associated with an increased risk of Y” to imply that X causes Y. Thus, even the most orthodox correlational language can give rise to causal inferences.

Introduction

Modern scientists are carefully trained to avoid conflating causation and correlation when describing research results. A correlation between two variables may reflect the causal effect of one variable on the other, or the causal effect of another variable on both. For this reason, observational studies report findings using seemingly non-causal stock phrases: variables are “associated” or “linked” with one another. However, it is possible that readers of these correlational statements may interpret them as causal.

Much work has demonstrated that people draw pragmatic inferences from ambiguous or incomplete linguistic utterances [13]. In particular, studies of “implicit causality” in language understanding have shown that people draw inferences about ambiguous causal roles from verbs [46]. For example, people infer that “she” refers to the daughter in the following sentence: “The mother punished her daughter because she admitted her guilt” [4]. In contrast, people infer that “she” refers to the mother in the following sentence: “The mother punished her daughter because she discovered her guilt.” The verbs “admitted” and “discovered” induce different causal role assignments. Similar results have been reported for sentence completion tasks, where the referent of the completed sentence picks out different nouns depending on the verb.

One influential view of implicit causality is that it reflects inferences about event structure [5], rather than reflecting linguistic structure (though see [6]). Supporting this view is evidence that causal role assignments are sensitive to general knowledge and social context, such as the gender and typicality of the agents/patients in the sentence. More broadly, the literature on implicit causality suggests that language is a rich source of causal information.

For our purposes, an important limitation of the implicit causality concept is that it takes as given some causal background knowledge (e.g., that mothers punish daughters when they admit guilt) and asks how people use this knowledge to make inferences about linguistic referents. Our goal in this paper is to flip this around: what happens when people know the referents but not the causal background knowledge? This situation is commonly encountered when people are reading newspaper headlines about scientific discoveries: if scientists report that eating ice cream increases the risk for cancer, a natural question is whether ice cream causes cancer. Careful scientists and journalists can expunge causal language from correlation studies, but can they expunge causal representations from the mental models of readers? To answer this question, we undertook a series of studies that assess what kinds of causal inferences people draw from correlational statements.

Results

We conducted three studies to examine the inference of causality from association statements. The results of Studies 1 and 2 are summarized in Fig 1, and the results of Study 3 are shown in Fig 2. In Study 1, participants were presented with statements of the form “X is associated with Y” and asked to judge whether X caused Y, or Y caused X. In Study 2, participants were presented with statements of the form “X is associated with an increased probability of Y”, and asked to make the similar causal judgments. In both studies, we ran versions with nonsense names designed to sound similar to medical terminology (e.g., “Themaglin” or “Pneuben”) or arbitrary letter symbols (see Methods for details).

Fig 1. Summary of results for Study 1 (left) and Study 2 (right).

Fig 1

Participants were given statements such as ‘Themaglin is associated with Pneuben’ (Study 1) or ‘Denoden is associated with an increased probability of Flembers’ (Study 2). A response of 1 indicates participants interpreted the statement to mean the first variable causes the second, 0 that they took it to mean the second variable causes the first, and 0.5 that they answered at random. The figure shows that without context, simple association is taken to imply the second variable causes the first. With minimal context, the association statement is taken to strongly imply the first variable caused the second. Error bars show standard error estimates for a proportion, p(1-p)/N, where p = M/N, M is the number ‘Yes’ responses, and N is the total number of responses.

Fig 2. Summary of results for Study 3.

Fig 2

Participants were given the same statements as in Study 2A, with the addition of a ‘Neither’ option to report that no causal direction was preferred. The y-axis now indicates the proportion of participants who chose each response. Error bars show standard error estimates for a proportion. The dotted line indicated expected random response levels.

We analyzed the data from Study 1 as follows. If participants chose the first variable as causing the second variable after being presented with the sentence “[X] is associated with [Y]”, this was coded as 1, and if they chose the second variable as causing the first, this was coded as 0. All responses within a study were averaged together (288 responses total in the ‘nonsense names’ condition, 291 responses total in the ‘symbols’ condition). If people were responding randomly, we would expect the average value to be 0.5. Note that a strategy such as ‘just pick the first answer’ is negated by the randomization of the variable order in both questions and answers. However, the mean response for both studies was significantly below chance, by a two-sided proportion z-test using Holm-Bonferroni correction for multiple comparisons (‘nonsense names’ condition: M = 0.23, Z = −11.14, SE = 0.025, p < 10−28; ‘symbols’ condition: M = 0.38, SE = 0.029, Z = −4.29, p < 10−4).

These results suggest that when given a simple association statement between two variables, participants inferred that the second variable causes the first. Put plainly, when presented with a sentence such as ‘Themaglin is associated with Pneuben’, or ‘X is associated with Y’, participants took this to imply ‘Pneuben causes Themaglin’, and ‘Y causes X’.

The analysis of data from Study 2 followed that of Study 1. If participants chose the first variable as causing the second variable after being presented with the sentence “[X] is associated with [RELATIONSHIP] [Y]”, this was coded as 1, otherwise the response was coded as 0. All responses within each relationship and within each study were averaged together. Again, if people were responding randomly, we would expect the average value to be 0.5.

We found that the addition of context drives all responses to be significantly above chance, by a two-sided proportion z-test using Holm-Bonferroni correction for multiple comparisons (‘nonsense names’ condition: Mrisk increase = 0.87, SE = 0.03, Z = 10.58, p < 10−25, Mrisk decrease = 0.94, SE = 0.02, Z = 17.91, p < 10−71, Mprobability increase = 0.88, SE = 0.03, Z = 11.25, p < 10−28, Mprobability decrease = 0.89, SE = 0.03, Z = 12.00, p < 10−32; ‘symbols’ condition: Mrisk increase = 0.84, SE = 0.04, Z = 9.14, p < 10−19, Mrisk decrease = 0.87, SE = 0.03, Z = 10.96, p < 10−27, Mprobability increase = 0.86, SE = 0.04, Z = 10.30, p < 10−24, Mprobability decrease = 0.82, SE = 0.04, Z = 8.16, p < 10−15). By contrast, the simple association statements (which do not include any context) were now not significantly different from chance. Although it is not entirely clear why this aspect of the data did not replicate Study 1, we suspect that it may be related to the context of more causally salient statements (risk increase/decrease, probability increase/decrease) which may have attuned participants to scrutinize association statements more carefully.

One concern with Studies 1 and 2 is that participants were forced to choose one causal direction. We addressed this in Study 3 by giving participants the option to report that neither causal direction was preferred. The results indicate that even with the option of reporting no preferred causal direction, people still typically took the association statement to imply that the first variable caused the second (Fig 2). For the pure association statements, the proportion between ‘X → Y’ and ‘Y → X’ was roughly twice that of Study 2, which we take to indicate that in Study 2 some participants were using the ‘Y → X’ statement to stand in for ‘neither are directly related’.

All proportions were different from one another within the different context questions. That is, we ran 3 two-sided proportion z-tests within each context (risk increase, risk decrease, probability increase, probability decrease, and simple association), comparing ‘X → Y’ to ‘Y → X’, ‘X → Y’ to ‘Neither’, and ‘Y → X’ to ‘Neither, for a total of 15 comparisons, and using the Holm-Bonferroni correction for multiple comparisons (note that the pre-registration has this as 7 questions and so 21 comparisons, due to an typographic error on the part of the experimenters in considering the number of questions asked).

All proportions were also significantly different from the chance response of 33%, except for ‘X → Y’ in the Association context, and ‘Neither’ in the Probability Decrease context. Given the overall pattern of responses, we take these latter two to be coincidental.

Discussion

These results indicate that when given an association statement between two variables with minimal context that indicates a change in the relationship for the second variable, participants inferred a causal relation, such that the first variable causes the second. In other words, when presented with a sentence such as ‘Themaglin is associated with an increased risk of Pneuben’, or ‘X is associated with an increased risk of Y’, participants took this to strongly imply ‘Themaglin causes Pneuben’, and ‘X causes Y’.

While we take the data to show that people draw causal implicatures from correlational statements, a possible concern is that participants were given a forced-choice between two causal relationships, without the option of declaring uncertainty, or rejecting both relationships. But this forced choice was a deliberate feature: because many people are highly drilled in the mantra that correlation does not imply causation, it is likely that when appropriately cued, they will activate the mantra. However, our hypothesis here is that there are implicit causal expectations about linguistic structures that are activated by correlational statements, and these expectations can be made explicit when people are forced to choose between different causal interpretations. If people truly are committed to non-causal interpretations of non-causal language, then participants should just choose arbitrarily between the different causal interpretations available to them. The fact that we found a large systematic preference argues in favor of our causal implicature hypothesis.

To more directly address the concern that forced choices between causal interpretations might yield artifactual results, we conducted a third study in which participants could choose a ‘Neither’ option. Strikingly, the results were essentially the same: participants still showed a strong preference for a particular causal direction in all of the experimental conditions apart from the association condition. Thus, it is unlikely that the response format produced a systematic bias.

One potential concern about our findings is that the association condition produced different results in Studies 1 and 2, deviating from random in the former but not in the latter. We speculated that this might be due to some kind of context effect. Study 3 sheds some additional light on this discrepancy, finding that most people endorse a non-causal interpretation of associative statements, although a large minority (47%) still favor a causal interpretation. Within that large minority, we find again a preference for XY rather than YX. We tentatively propose that many or even most of the people endorsing YX in Study 2 were doing so as a signal of protest against being asked to assign a causal direction.

Our results cannot unambiguously rule out several alternative hypotheses that hinge on different interpretations of the information we provided to participants. First, it is possible that participants did not distinguish between probabilistic and causal interpretations of the statements. For example, the statement that X increases the risk of Y might also imply that Y increases the risk of X, but the increase for Y given X might still be larger than the increase for X given Y. In this case, there is an asymmetry in the risk pattern, even if there is no causal relationship between X and Y. If participants align their causal preferences with the directional asymmetry, then they might show preferences deviating from random, which we have treated as a signature of causal implicature. In essence, this alternative hypothesis posits either that participants do not fully understand what causality is, or that risk directionality is used as a heuristic for causal implicature. We think it is somewhat unlikely that participants simply do not understand what we are asking when we elicit causal interpretations, given the sophisticated causal reasoning abilities demonstrated in many other studies.

A second alternative hypothesis is that participants interpreted statements of the form “X is associated with increased risk of Y” as “Doing X is associated with increased risk of Y” which would seem to imply a kind of causal intervention. Thus, on this hypothesis apparent causal preferences arise from vagueness in the statements. While we cannot rule out this possibility, we would argue that the vagueness hypothesis is really another version of causal implicature, where participants make a pragmatic interpretation that the speaker intends to communicate a causal relationship.

We have argued that the alignment and vagueness hypotheses may or may not be consistent with a causal implicature hypothesis. More work is required to answer this question.

In summary, certain correlational statements are associated with an increased probability of causal implicature. To be clear, we are not implying that these correlational statements cause causal implicature, but rather that they are correlated with causal implicature. In other words, correlation does not imply causation, but it does sometimes “imply” causation.

Methods

The studies were approved by the Harvard Institutional Review Board, protocol no. 19-1861.

Participants were recruited online [7] via the Prolific platform (https://www.prolific.co). Participants were restricted to those located in the USA, and having completed at least 100 prior studies on Prolific, with an acceptance rate of at least 90%. We recruited 100 participants in each study (total of N = 400), following a pilot indicating the expected effect size is such that this number of participants had a power >90%. Participants in any study described here (meaning, each sub-condition) were prohibited from participating in any other study described here.

At the end of each study, participants were asked “Please describe, in a few words, what you were asked to do in this experiment?”. We excluded participants who gave nonsensical or non-sequitur answers, such as ‘Do the causes’ or ‘Opinions’.

The analyses, experiments, exclusion criteria, and sample sizes were pre-registered, as available here: https://aspredicted.org/blind.php?x=JHK_JR1 and here https://aspredicted.org/blind.php?x=BP6_Z8D.

All data are available at https://osf.io/u34ex/?view_only=1ffdf6729af149d6b14f7e32692c0334.

Study 1

To examine the basic implicature of the phrasing ‘X is associated with Y’, without any further context, we designed a simple study in which participants read statements about the existence of an association between nonsense terms such as Themalgin and Pneuben (‘nonsense names’ condition), or abstract symbols like X and Y (‘symbols’ condition).

Participants

We recruited 200 participants, and randomly assigned them evenly to the two conditions (‘nonsense names’ and ‘symbols’). Written/verbal informed consent was obtained from all participants for inclusion in the study. After excluding 4 participants, the mean age of participants in the ‘nonsense names’ condition was 35.4, and 57 identified as female. After excluding 3 participants in the ‘symbols’ condition, the mean age of the remaining participants was 34.7, and 66 identified as Female.

Stimuli

Participants were informed that they will be asked a few simple questions about the possible relationship between different things, and that the questions are unrelated to one another. Participants then saw 3 questions in succession, each phrased as

Suppose you read the following piece of information:

“[X] is associated with [Y]”

Which of the following is more likely?

Participants were then given a forced choice between the statements “[X] causes [Y]” and “[Y] causes [X]”.

In the ‘nonsense names’ condition, [X] and [Y] were replaced with the following nonsense terms: Themaglin, Rebosen, Denoden, Flembers, Agoriv, and Ceflar. In the ‘symbols’ condition, [X] and [Y] were replaced with the following symbols: X, Y, P, Z, G, and D. The ordering of the nonsense term pairs or symbol pairs within each question, the ordering of the forced choice answers to each question, and the question order of the three questions was randomized.

Study 2

In Study 2, we provided additional context of the sort that is often found in the scientific literature, as well as the popular press. In particular, we examined the causal implicature of phrases such as ‘Ceflar is associated with a lower risk of Agoriv’, ‘Pneuben is associated with higher probability of Efogen’, and so on. The design was similar to Study 1: Participants were given statements about the existence of an association and minimal context about the relationship (risk, probability, increase, decrease) between nonsense terms (‘nonsense names’ condition), or abstract symbols like X and Y (‘symbols’ condition).

Participants

We recruited 200 participants in total, randomly and evenly distributing them to the two conditions (‘nonsense names’ and ‘symbols’). Written/verbal informed consent was obtained from all participants for inclusion in the study. After excluding 3 participants, the mean age of participants in the ‘nonsense names’ condition was 35.0, and 55 identified as female. After excluding 6 participants in the ‘symbols’ condition, the mean age of the remaining participants was 34.6, and 70 identified as Female.

Stimuli

Participants were informed that they would be asked a few simple questions about the possible relationship between different things, and that the questions were unrelated to one another. Participants then saw 5 questions in succession, each phrased as:

Suppose you read the following piece of information:

“[X] is associated with [RELATIONSHIP] [Y]”

Which of the following is more likely?

Participants were then given a forced choice between the statements “[X] causes [CHANGE] [Y]” and “[Y] causes [CHANGE] [X]”.

In the ‘nonsense names’ condition, [X] and [Y] were replaced with the following nonsense terms: Themaglin, Rebosen, Denoden, Flembers, Agoriv, Ceflar, Pneuben, Efogen, Turilin, and Laurem. In the ‘symbols’ condition, [X] and [Y] were replaced with the following symbols: T, R, P, E, A, C, X, Y, D, and F. The ordering of the nonsense term pairs or symbol pairs within each question, the ordering of the forced choice answers to each question, and the question order of the five questions was randomized. The [RELATIONSHIP] variable was replaced with: Higher probability, higher risk, lower probability, lower risk, or was left empty. When the [RELATIONSHIP] variable was changed to ‘higher’ the [CHANGE] relationship was replaced with: ‘an increase in’. When the [RELATIONSHIP] variable was changed to ‘lower’, the [CHANGE] relationship was replaced with: ‘a decrease in’. When the [RELATIONSHIP] variable was left empty, the [CHANGE] relationship was left empty, recreating the structure of Study 1.

Study 3

In Study 3, we gave participants a tertiary response rather than a binary response, allowing them to express that neither of the two entities were causally related. More specifically, the study replicated Study 2, condition A (‘nonsense names’), with an additional response option: ‘A third factor is causally related to [X] and [Y], they are not directly related’, where X and Y were the same nonsense terms used in Study 2.

Participants

We recruited 100 participants, matching the sample sizes of the different conditions in Studies 1 and 3. Written/verbal informed consent was obtained from all participants for inclusion in the study. Participants were US-based No participants were excluded. The mean age of participants was 34.8, 63 identified as female, and 37 identified as male.

Stimuli

Study 3 following the logic and stimuli of Study 2A. Participants were informed that they would be asked a few simple questions about the possible relationship between different things, and that the questions were unrelated to one another. Same as Study 2, the ordering of the nonsense term pairs within each question, the ordering of the forced choice answers to each question, and the question order of the five questions was randomized.

Data quality assurance

We appreciate the growing concern about using AI tools to answer online surveys, and we do not think there is an agreed-on safeguard yet against the use of latest tools like ChatGPT-4 to answer online catch questions. We would note that the first two studies were run in early 2022, when the use of large language models was not widespread, and their performance when used was rather limited. Our catch question was meant to weed out low-effort, automatic responses such as ‘have a good day or ‘relationships’. To the degree that we did not weed out all auto-replies, we believe that they only introduced a bit of noise, and given that our studies report differential effects this extra noise on top of the main effect is not a primary concern. The third study was run in late 2022, around the advent of the (now legacy) version of ChatGPT, and before adjusting for how well (or not) such tools can pass various tasks. We cannot guarantee our last question weeds out malicious users of the latest automatic tools, and moving forward we plan to adjust our data validation. However, we believe that the timing of the study was such that it is unlikely a large amount of malicious users on Prolific (to the degree such a group existed) had time to mass adopt ChatGPT into their workflow and rack up hundreds of completed and approved studies in time for our study. We also note that the marginal financial gain of using these tools is likely outweighed by the cost of per-token expenses.

Acknowledgments

We are grateful to Michael Franke for helpful discussions.

Data Availability

All data are available at https://osf.io/u34ex/?view_only=1ffdf6729af149d6b14f7e32692c0334.

Funding Statement

This work was supported by the Center for Brains, Minds and Machines (CBMM), funded by NSF STC award CCF1231216. There was no additional external funding received for this study. The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.

References

  • 1. Grice HP. Logic and conversation. In: Speech Acts. Brill; 1975. p. 41–58. [Google Scholar]
  • 2. Clark HH. Using Language. Cambridge University Press; 1996. [Google Scholar]
  • 3. Goodman ND, Frank MC. Pragmatic language interpretation as probabilistic inference. Trends in cognitive sciences. 2016;20(11):818–829. doi: 10.1016/j.tics.2016.08.005 [DOI] [PubMed] [Google Scholar]
  • 4. Garvey C, Caramazza A. Implicit causality in verbs. Linguistic Inquiry. 1974;5:459–464. [Google Scholar]
  • 5. Pickering MJ, Majid A. What are implicit causality and consequentiality? Language and Cognitive Processes. 2007;22:780–788. doi: 10.1080/01690960601119876 [DOI] [Google Scholar]
  • 6. Hartshorne JK. What is implicit causality? Language, Cognition and Neuroscience. 2014;29:804–824. doi: 10.1080/01690965.2013.796396 [DOI] [Google Scholar]
  • 7. Peer E, Brandimarte L, Samat S, Acquisti A. Beyond the Turk: Alternative platforms for crowdsourcing behavioral research. Journal of Experimental Social Psychology. 2017;70:153–163. doi: 10.1016/j.jesp.2017.01.006 [DOI] [Google Scholar]

Decision Letter 0

Micah B Goldwater

19 Sep 2022

PONE-D-22-19317Causal implicatures from correlational statementsPLOS ONE

Dear Dr. Gershman,

Thank you for submitting your manuscript to PLOS ONE. After careful consideration, we feel that it has merit but does not fully meet PLOS ONE’s publication criteria as it currently stands. Therefore, we invite you to submit a revised version of the manuscript that addresses the points raised during the review process.Reviewer 1 was more positive than Reviewer 2. R1 carefully spells out multiple possible interpretations of the results and suggests ways for the manuscript to account for these interpretations. R2 does not think the paper does an adequate job of engaging with the existing literature on this topic. The paper as it stands is quite short, so more adequately addressing the prior research could still be done in a "brief report" length paper. 

Please submit your revised manuscript by Nov 03 2022 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

Please include the following items when submitting your revised manuscript:

  • A rebuttal letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols.

We look forward to receiving your revised manuscript.

Kind regards,

Micah B. Goldwater, Ph.D

Academic Editor

PLOS ONE

Journal Requirements:

When submitting your revision, we need you to address these additional requirements.

1. Please ensure that your manuscript meets PLOS ONE's style requirements, including those for file naming. The PLOS ONE style templates can be found at 

https://journals.plos.org/plosone/s/file?id=wjVg/PLOSOne_formatting_sample_main_body.pdf and 

https://journals.plos.org/plosone/s/file?id=ba62/PLOSOne_formatting_sample_title_authors_affiliations.pdf

2. Please provide additional details regarding participant consent. In the ethics statement in the Methods and online submission information, please ensure that you have specified (1) whether consent was informed and (2) what type you obtained (for instance, written or verbal, and if verbal, how it was documented and witnessed). If your study included minors, state whether you obtained consent from parents or guardians. If the need for consent was waived by the ethics committee, please include this information.

If you are reporting a retrospective study of medical records or archived samples, please ensure that you have discussed whether all data were fully anonymized before you accessed them and/or whether the IRB or ethics committee waived the requirement for informed consent. If patients provided informed written consent to have data from their medical records used in research, please include this information.

3. Please update your submission to use the PLOS LaTeX template. The template and more information on our requirements for LaTeX submissions can be found at http://journals.plos.org/plosone/s/latex.

4. Thank you for stating in your Funding Statement: 

"This work was supported by the Center for Brains, Minds and Machines (CBMM), funded by NSF STC award CCF1231216."

Please provide an amended statement that declares *all* the funding or sources of support (whether external or internal to your organization) received during this study, as detailed online in our guide for authors at http://journals.plos.org/plosone/s/submit-now.  Please also include the statement “There was no additional external funding received for this study.” in your updated Funding Statement. 

Please include your amended Funding Statement within your cover letter. We will change the online submission form on your behalf.

5. Please include your full ethics statement in the ‘Methods’ section of your manuscript file. In your statement, please include the full name of the IRB or ethics committee who approved or waived your study, as well as whether or not you obtained informed written or verbal consent. If consent was waived for your study, please include this information in your statement as well.

[Note: HTML markup is below. Please do not edit.]

Reviewers' comments:

Reviewer's Responses to Questions

Comments to the Author

1. Is the manuscript technically sound, and do the data support the conclusions?

The manuscript must describe a technically sound piece of scientific research with data that supports the conclusions. Experiments must have been conducted rigorously, with appropriate controls, replication, and sample sizes. The conclusions must be drawn appropriately based on the data presented.

Reviewer #1: Partly

Reviewer #2: Partly

**********

2. Has the statistical analysis been performed appropriately and rigorously?

Reviewer #1: Yes

Reviewer #2: Yes

**********

3. Have the authors made all data underlying the findings in their manuscript fully available?

The PLOS Data policy requires authors to make all data underlying the findings described in their manuscript fully available without restriction, with rare exception (please refer to the Data Availability Statement in the manuscript PDF file). The data should be provided as part of the manuscript or its supporting information, or deposited to a public repository. For example, in addition to summary statistics, the data points behind means, medians and variance measures should be available. If there are restrictions on publicly sharing data—e.g. participant privacy or use of data from a third party—those must be specified.

Reviewer #1: Yes

Reviewer #2: Yes

**********

4. Is the manuscript presented in an intelligible fashion and written in standard English?

PLOS ONE does not copyedit accepted manuscripts, so the language in submitted articles must be clear, correct, and unambiguous. Any typographical or grammatical errors should be corrected at revision, so please note any specific errors here.

Reviewer #1: Yes

Reviewer #2: No

**********

5. Review Comments to the Author

Please use the space provided to explain your answers to the questions above. You may also include additional comments for the author, including concerns about dual publication, research ethics, or publication ethics. (Please upload your review as an attachment if it exceeds 20,000 characters)

Reviewer #1: This paper examines how people interpret statements of association (e.g. "X is associated with increased risk of Y"). Participants were asked to respond to statements like this in a forced-choice task, choosing whether it is more likely that "X causes Y" or "Y causes X." The authors find that participants are biased to interpret statements of association as implying specific causal directionality. They conclude that this is due to causal implicatures conveyed by the association statements, which the authors call the "causal implicature hypothesis."

Overall, the paper is nicely and clearly written. The methods appear appropriately sound and the results are very clear with respect to the statistical hypotheses tested. I also commend the open materials and data though I confess I have not examined the data myself.

I have a few points to raise for revisions:

First, a question about one condition in Study 2 and its relation to Study 1. In Figure 1, the "association" condition in S2 appears to show chance responding. From the methods, I understand this condition was essentially the same as Study 1, which showed non-chance responding. Am I correct in interpreting this as a failure to replicate S1's results in S2? If so, I think it would be best for the authors to address this---S1 is preregistered and the effects observed are sizable enough I would expect them to replicate, so it seems a bit puzzling.

Next, I have two other concerns: one I feel is essential to any revision and another I think would help to improve the quality of the discussion.

First, I have a major concern about the correspondence between some of the claims made in the introduction and discussion and what the results actually seem to support. For instance, in the abstract, the authors write "We show that people do in fact infer causality from statements of association, under minimal conditions." But the conditions under which participants inferred causality were forced choice tasks requiring them to infer causality, asking only about the most likely direction. So it does not seem justified to claim that people "infer causality" on the basis of these findings. In the discussion, the authors write "... participants inferred a causal relation, such that the first variable causes the second." Again it does not seem justified to foreground the first part of that claim in this way. In both cases it would be better to hew closer to what was actually shown. Maybe something like "participants were biased to infer specific causal directionality," making clear that they did not spontaneously make causal inferences. I do recognize that the authors address this somewhat in the discussion, but the claims should be made consistent with the findings throughout.

Second, I think more could be done to clarify conceptually what is being claimed and to distinguish it from other (potentially deflationary?) explanations of the findings. The authors aim to test what they call their "causal implicature hypothesis." The implicature hypothesis is that "there are implicit causal expectations about linguistic structures that are activated by correlational statements, and these expectations can be made explicit when people are forced to choose between different causal interpretations".

As I understand it, the idea is that statistical correlation or "association" is a symmetric or non-directional relationship. So, if people's interpretations of these statements are really committed to this symmetry, then they should choose randomly in the forced choice causal-direction tasks ("If people truly are committed to non-causal interpretations of non-causal language, then participants should just choose arbitrarily between the different causal interpretations available to them.") Instead, the authors' studies have shown that, when pressed, people interpret such statements as implying a specific causal directionality. The logic of these studies is that observing symmetric statements producing asymmetric interpretations gives evidence for the causal implicature hypothesis.

However, I see a few ways that the statements might not be seen as truly symmetric, especially in Study 2. Consider "increases risk of" statements: On the one hand it's true that if P(A|B) > P(A), then P(B|A) > P(B), but on the other hand the degree of the inequality might not be at all symmetric. For instance, some rare event A might greatly increase risk of B, yet A could still remain fairly unlikely given B. In such a case it would be literally true to say both that "A increases risk of B" and "B increases risk of A", but it might be pragmatically more informative to say "A increases risk of B". (I suppose whether this is so would depend on how people think about quantifying risk differences. In terms of mutual information things are symmetric, but in terms of differences like P(A|B) - P(A) they need not be.) So statements of association can sometimes be asymmetric without any causal content. Participants might then be aligning their causal directionality with the "directionality" of the original statements.

Second, the effects might not be due to implicatures but rather something like vagueness. One possibility is people might more directly interpret the claims about association as being based on experimental evidence. They might interpret the statement "X is associated with increased risk of Y" as "_doing_ X is associated with an increased risk of Y". To be sure, pragmatically it is odd for a speaker to say "associated with an increased risk of" rather than "increases risk" or "causes" in such a case. But, participants are ultimately given a forced choice between two causal interpretations. When forced to pick a causal interpretation, it seems like the statements in Study 2 at least leave open the "doing" reading, and hence would support the causal interpretations participants made. (note also that these "doing" interpretations are asymmetric.)

To address this concern I'd like to see more careful discussion of how the authors think participants actually interpret the statements in their studies and how they might interpret such statements outside the lab. I could be convinced that the objections I've raised are not major problems for what I take to be the spirit of the authors' conclusions: e.g. as I said I think you could construe part of my complaint as a question about whether their results are due to pragmatic implicatures or simply vagueness. In practice, we might expect similar consequences and thus it seems reasonable to care about both: Presented with careful statements of association by scientists, lay people will (at least sometimes) interpret them as loose statements about causal relationships. And on the other front, if statements of association have inherent (pragmatic) asymmetry, which lead to causal implications, that seems worth knowing too. Still I'd encourage the authors to clarify their claims and contrast their causal implicature hypothesis with alternate accounts. I'd also encourage them to consider how different accounts of their findings then relate to real-world consequences (e.g. especially in relation to their forced-choice task---without these pressures misinterpretations might be less likely).

Reviewer #2: Correlation does not imply causation and people are often warned of this fact. This suggests people may tend to ascribe causes even when purely correlational statements are made. This is the premise of the paper.

The paper is very brief and I'll keep my comments on it brief as well. Although the studies seem like reasonable first passes at studying the question, there is next to no discussion of the rich literature that exists and how different syntactic structures may prompt people to think an antecedent caused a consequent. Some of this work is cited. None of the details of these results are discussed to contextualize the studies that have been run.

In my view, this paper reads like a blog post---and that has some virtues--- but doesn't give enough to the reader enough to understand their findings, nor really engages with the fact that people have thought about this and related questions before. For the work to be publishable, the authors will need to write a paper instead of a blog post.

**********

6. PLOS authors have the option to publish the peer review history of their article (what does this mean?). If published, this will include your full peer review and any attached files.

If you choose “no”, your identity will remain anonymous but your review may still be made public.

Do you want your identity to be public for this peer review? For information about this choice, including consent withdrawal, please see our Privacy Policy.

Reviewer #1: Yes: Derek Powell

Reviewer #2: No

**********

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

While revising your submission, please upload your figure files to the Preflight Analysis and Conversion Engine (PACE) digital diagnostic tool, https://pacev2.apexcovantage.com/. PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at figures@plos.org. Please note that Supporting Information files do not need this step.

Decision Letter 1

Micah B Goldwater

2 May 2023

PONE-D-22-19317R1Causal implicatures from correlational statementsPLOS ONE

Dear Dr. Gershman,

Thank you for submitting your manuscript to PLOS ONE. We apologize for delay. After initially indicating they could review a revision, the original reviewers had to withdraw, and so to not delay further, I alone evaluated the manuscript. I think you adequately responded to the reviewers comments, and the paper presents interesting findings. I briefly add my own thoughts.I agree with your causal implicature account. It seems people hold a default assumption is that things happen because of causes or "for a reason" and there is work from multiple areas of cognition that people do not like having empty causal models (e.g., from the continued influence effect). If language is a consistent with a cause, then it will be interpreted as such if nothing else stops someone from doing so. There's a reason there needs to be explicit reminders that a correlation can always means a hidden third variable is the true cause of both A and B. Without that reminder, which at least suggests an alternative causal model, the only possible cause to infer is between A and B, and so it seems people do. That said, it's even unclear how successful the common-cause warnings are when they don't include a candidate third cause (or even when they do). Regardless of addressing these questions specifically, I'll be interested to see where this line of research is going next.  There is just a single minor revision needed before the paper can be accepted. There is a journal policy (that I was recently reminded of) that there needs to be safeguards of data quality for online data sources such as from Prolific, which you used here. Given the rise of bots and AI tools to answer surveys, this has become even more critical in  just the past six months. I see that  "participants were restricted to those located in the USA, and having completed at least 100 prior studies on Prolific, with an acceptance rate of at least 90%." And further, they were asked "Please describe, in a few words, what you were asked to do in this experiment?”" at the end of the study.  If there is further evidence to provide that when people automate responses, they are likely to be rejected (and so would not reach a 90% acceptance rate), and whether this last sentence would be difficult to automatically generate a response that actually reflected the content of the study, that would help increase the confidence in the data quality.  I assume providing this further evidence will not be much of a burden. 

Please submit your revised manuscript by Jun 16 2023 11:59PM. If you will need more time than this to complete your revisions, please reply to this message or contact the journal office at plosone@plos.org. When you're ready to submit your revision, log on to https://www.editorialmanager.com/pone/ and select the 'Submissions Needing Revision' folder to locate your manuscript file.

Please include the following items when submitting your revised manuscript:

  • A rebuttal letter that responds to each point raised by the academic editor and reviewer(s). You should upload this letter as a separate file labeled 'Response to Reviewers'.

  • A marked-up copy of your manuscript that highlights changes made to the original version. You should upload this as a separate file labeled 'Revised Manuscript with Track Changes'.

  • An unmarked version of your revised paper without tracked changes. You should upload this as a separate file labeled 'Manuscript'.

If you would like to make changes to your financial disclosure, please include your updated statement in your cover letter. Guidelines for resubmitting your figure files are available below the reviewer comments at the end of this letter.

If applicable, we recommend that you deposit your laboratory protocols in protocols.io to enhance the reproducibility of your results. Protocols.io assigns your protocol its own identifier (DOI) so that it can be cited independently in the future. For instructions see: https://journals.plos.org/plosone/s/submission-guidelines#loc-laboratory-protocols. Additionally, PLOS ONE offers an option for publishing peer-reviewed Lab Protocol articles, which describe protocols hosted on protocols.io. Read more information on sharing protocols at https://plos.org/protocols?utm_medium=editorial-email&utm_source=authorletters&utm_campaign=protocols.

We look forward to receiving your revised manuscript.

Kind regards,

Micah B. Goldwater, Ph.D

Academic Editor

PLOS ONE

Journal Requirements:

Please review your reference list to ensure that it is complete and correct. If you have cited papers that have been retracted, please include the rationale for doing so in the manuscript text, or remove these references and replace them with relevant current references. Any changes to the reference list should be mentioned in the rebuttal letter that accompanies your revised manuscript. If you need to cite a retracted article, indicate the article’s retracted status in the References list and also include a citation and full reference for the retraction notice.

[Note: HTML markup is below. Please do not edit.]

[NOTE: If reviewer comments were submitted as an attachment file, they will be attached to this email and accessible via the submission site. Please log into your account, locate the manuscript record, and check for the action link "View Attachments". If this link does not appear, there are no attachment files.]

While revising your submission, please upload your figure files to the Preflight Analysis and Conversion Engine (PACE) digital diagnostic tool, https://pacev2.apexcovantage.com/. PACE helps ensure that figures meet PLOS requirements. To use PACE, you must first register as a user. Registration is free. Then, login and navigate to the UPLOAD tab, where you will find detailed instructions on how to use the tool. If you encounter any issues or have any questions when using PACE, please email PLOS at figures@plos.org. Please note that Supporting Information files do not need this step.

Decision Letter 2

Micah B Goldwater

9 May 2023

Causal implicatures from correlational statements

PONE-D-22-19317R2

Dear Dr. Gershman,

We’re pleased to inform you that your manuscript has been judged scientifically suitable for publication and will be formally accepted for publication once it meets all outstanding technical requirements.

Within one week, you’ll receive an e-mail detailing the required amendments. When these have been addressed, you’ll receive a formal acceptance letter and your manuscript will be scheduled for publication.

An invoice for payment will follow shortly after the formal acceptance. To ensure an efficient process, please log into Editorial Manager at http://www.editorialmanager.com/pone/, click the 'Update My Information' link at the top of the page, and double check that your user information is up-to-date. If you have any billing related questions, please contact our Author Billing department directly at authorbilling@plos.org.

If your institution or institutions have a press office, please notify them about your upcoming paper to help maximize its impact. If they’ll be preparing press materials, please inform our press team as soon as possible -- no later than 48 hours after receiving the formal acceptance. Your manuscript will remain under strict press embargo until 2 pm Eastern Time on the date of publication. For more information, please contact onepress@plos.org.

Kind regards,

Micah B. Goldwater, Ph.D

Academic Editor

PLOS ONE

Additional Editor Comments (optional):

Reviewers' comments:

Acceptance letter

Micah B Goldwater

10 May 2023

PONE-D-22-19317R2

Causal implicatures from correlational statements

Dear Dr. Gershman:

I'm pleased to inform you that your manuscript has been deemed suitable for publication in PLOS ONE. Congratulations! Your manuscript is now with our production department.

If your institution or institutions have a press office, please let them know about your upcoming paper now to help maximize its impact. If they'll be preparing press materials, please inform our press team within the next 48 hours. Your manuscript will remain under strict press embargo until 2 pm Eastern Time on the date of publication. For more information please contact onepress@plos.org.

If we can help with anything else, please email us at plosone@plos.org.

Thank you for submitting your work to PLOS ONE and supporting open access.

Kind regards,

PLOS ONE Editorial Office Staff

on behalf of

Dr. Micah B. Goldwater

Academic Editor

PLOS ONE

Associated Data

    This section collects any data citations, data availability statements, or supplementary materials included in this article.

    Supplementary Materials

    Attachment

    Submitted filename: Causal implicatures response letter (1).pdf

    Attachment

    Submitted filename: GershmanUllman Plos One response2.pdf

    Data Availability Statement

    All data are available at https://osf.io/u34ex/?view_only=1ffdf6729af149d6b14f7e32692c0334.


    Articles from PLOS ONE are provided here courtesy of PLOS

    RESOURCES