Using Correlation Coefficients in Research Papers: Types, Interpretation, and APA Reporting

Correlational research is one of the most accessible and practical research methods available to academics, especially students working with limited time and budgets. Understanding how to select, calculate, interpret, and report correlation coefficients correctly is an essential skill for writing credible, well-structured research papers across finance, science, engineering, and the social sciences. The specific coefficient you choose and how you report the result shape whether reviewers accept the methodology or send the manuscript back.


This article covers the main types of correlation coefficients and when to use each, how to collect and analyze correlational data, interpretation thresholds across disciplines, worked examples from finance, science, engineering, and the social sciences, APA reporting conventions, and the common mistakes that trigger reviewer flags. For the mathematical mechanics of what the correlation coefficient actually measures (covariance, r vs r squared, worked calculations), see our companion guide to correlation coefficients explained.


Quick Answer: Which Correlation Coefficient to Use and How to Report It

Pearson's r. Two continuous, normally distributed variables with a linear relationship. The default choice in finance, physical sciences, and engineering.

Spearman's rho. Ordinal data (Likert scales), non-normal distributions, or monotonic but non-linear relationships. The default choice in survey-based social science research.

Kendall's tau. Small samples or data with many tied ranks.

APA reporting format. r(N-2) = value, p = value. For example: r(98) = .45, p < .001. Modern journals also expect confidence intervals, effect size interpretation, and the coefficient of determination (r squared).

Always include a scatter plot. A correlation coefficient without visual data is incomplete; the coefficient hides outliers, nonlinear patterns, and clustering.


What Is a Correlation Coefficient?

A correlation measures the linear relationship between two variables. The correlation coefficient describes both the strength and the direction of that relationship, expressed as a value ranging from -1 to +1. A value of 0 indicates no linear relationship between the variables, while a value of -1 or +1 indicates a perfect relationship. Values closer to 0 indicate a weaker relationship, and values closer to -1 or +1 indicate a stronger one.


The purpose of correlational research is to determine whether a linear relationship exists between two variables. For example, you might want to know whether the number of hours a student studies correlates with the grade they receive. By surveying students in a class, you could collect data on weekly study hours and final grades, then apply a correlation coefficient formula to get a value between -1 and +1. That value would tell you whether grades tend to increase as study hours increase, decrease as study hours increase, or show no discernible pattern based on hours studied.


It's important to understand that correlation isn't the same as causation. Even if correlational research shows that grades improve as students study more, you can't conclude that studying more directly causes better grades. There are often other explanations for the results you find, which we'll cover in the analysis section below.


Why Correlational Research Is Useful

If correlational research can't establish causation, you might wonder why it's worth doing at all. Three reasons researchers choose correlational methods:


  • It's faster than experimental research. You can gather data from natural, non-experimental settings in a relatively short time, without needing to design and run a controlled experiment.
  • It's more affordable. Correlational research typically requires fewer resources than experimental research, making it a practical option for students and researchers working with limited budgets.
  • It's the only option in many real-world contexts. Many research questions in finance, epidemiology, and the social sciences can't be tested experimentally for ethical or practical reasons. You can't randomly assign people to different income levels, smoking habits, or stock portfolios. Correlational research lets you investigate these questions using observational data.

When to Use Correlational Research

There are several situations where correlational research isn't just useful but the most appropriate choice:


  • Investigating non-causal relationships. You won't always expect a causal relationship between two variables, but knowing whether they correlate can still be valuable for building a broader understanding of a topic.
  • Supporting causal theories. When it's too expensive, impractical, or unethical to run experiments that would establish causation, a strong correlational finding can lend support to a causal theory.
  • Testing new measurement tools. If the correlational relationship between two variables is already well established, you can use correlational research on those variables with new measurement instruments to assess their validity and reliability.
  • Identifying predictors before designing an experiment. Correlational analysis at an early stage of a research program helps narrow the field of candidate variables before committing resources to a controlled experimental design.

Types of Correlation Coefficients: Which One Should You Use?

The right correlation coefficient depends on the type of data you've collected and whether it meets certain statistical criteria. Each coefficient has a specific formula and is suited to a specific kind of dataset. Choosing the wrong coefficient is one of the most common reasons reviewers flag a methods section, so the choice deserves careful attention.


CoefficientUse whenCommon fields
Pearson's rBoth variables are continuous, normally distributed, and the relationship is linearFinance, physical sciences, engineering, quantitative psychology
Spearman's rhoVariables are ordinal, not normally distributed, or the relationship is monotonic but not linearSocial sciences, survey research, behavioral economics
Kendall's tauSmall samples or data with many tied ranksSmall-sample behavioral research, ranked preference studies
Phi coefficientTwo dichotomous (binary) variables in a 2x2 contingency tableEpidemiology, clinical research
Cramer's VTwo categorical variables in contingency tables larger than 2x2Marketing research, political polling
Point-biserialOne continuous variable and one dichotomous variablePsychology, educational research (e.g., test scores and pass/fail)

Pearson's r in Detail

Used for the relationship between two continuous, randomly distributed variables that are both normally distributed. Your data must meet these criteria for Pearson's r to be an accurate measure. Pearson's r is the default choice in most quantitative finance, science, and engineering research where the variables are interval or ratio measurements. For the mathematical derivation and worked calculation, see our correlation coefficients explained guide.


Spearman's Rho in Detail

Used for two continuous or ordinal variables that don't need to be normally distributed. It's the most common alternative when your data doesn't meet the criteria for Pearson's r. Spearman's rho is based on the ranked order of data rather than the actual values, making it the standard choice for survey data with Likert-scale responses, ordinal rankings, or skewed distributions common in the social sciences.


Kendall's Tau in Detail

An extension of Spearman's rho, used when working with a small dataset where one rank appears too many times. Kendall's tau is also preferred when the dataset contains many tied ranks, as Spearman's rho can give misleading values in those situations.


How to Collect Correlational Data

Like experimental research, correlational research uses quantitative methods. The key difference is that variables in correlational research are observed rather than manipulated. There are three main approaches to collecting correlational data:


  1. Surveys. Questionnaires let you collect data quickly from your target population. They can be administered in person, online, by mail, or by phone, making them one of the most flexible data collection methods available. Survey-based correlational research is dominant in the social sciences and consumer research.
  2. Observation. This approach involves recording behavior or phenomena as they occur in a natural environment, including descriptions of the setting, events, and actions being observed. Observational data collection is common in epidemiology, ecology, and behavioral research.
  3. Secondary sources. Existing datasets collected for other purposes can be used for correlational research. This is the fastest and least expensive approach, but it comes with a tradeoff: since you didn't collect the data yourself, you have no control over its reliability or validity. Secondary data sources include CRSP and Compustat in finance, the Survey of Consumer Finances and the Panel Study of Income Dynamics in economics, the Materials Project database in engineering, and ICPSR-archived datasets in the social sciences.

Interpreting Correlation Coefficients: Thresholds and Context

Analyzing correlational data begins with plotting your data and calculating the correlation coefficient. The coefficient gives you a value representing the strength and direction of the relationship, while graphing the data gives you a visual picture of what that relationship actually looks like. Always plot your data before running the calculation. A scatter plot reveals nonlinear relationships, outliers, and clusters that the correlation coefficient alone will hide.


The table below provides general guidelines for interpreting the strength of a correlation coefficient:


Absolute value of rInterpretation
0.00 to 0.10Negligible
0.10 to 0.39Weak
0.40 to 0.69Moderate
0.70 to 0.89Strong
0.90 to 1.00Very strong

These thresholds are general guidelines, not universal standards. In some fields, what counts as a meaningful correlation differs. In physics and chemistry, where measurement precision is high and underlying relationships are often deterministic, researchers expect very strong correlations and treat anything below 0.90 with skepticism. In psychology and the social sciences, where human behavior is influenced by many factors simultaneously, a correlation of 0.30 may represent an important and publishable finding. Calibrate your interpretation to the conventions of your field and target journal.


Two Important Caveats

First, a value near 0 doesn't mean there's no relationship between the variables at all. It means there's no linear relationship. There could still be another type of relationship, such as a quadratic one. Graphing your data before running the analysis will help you spot any nonlinear patterns. The classic illustration is Anscombe's quartet, four datasets with identical Pearson correlations of approximately 0.816 but radically different visual structures including curved relationships, single outlier-driven correlations, and clustered patterns. Anscombe's quartet is a reminder that the correlation coefficient is a summary, not a substitute for visual inspection.


Second, correlation isn't causation. When you find a correlational relationship, there are often multiple explanations for it that weren't accounted for in your research. One common issue is the directionality problem: if students who study more get better grades, you could equally argue that getting better grades motivates students to study more. The data alone can't tell you which direction the relationship runs. Another issue is the possibility of a third variable (also called a confounder). In the studying and grades example, students who sleep more might both study more and earn better grades, meaning that sleep, not studying, is the underlying driver of both outcomes. For more on this problem, see our guide to confounding variables.


Worked Examples Across Disciplines

The application of correlation coefficients differs significantly across research disciplines. The examples below illustrate how researchers in finance, science, engineering, and the social sciences select, calculate, and interpret correlation coefficients in their respective fields.


Finance: Stock Returns and the Fama-French Factors

In empirical asset pricing, correlation coefficients underpin the relationship between individual stock returns and systematic risk factors. Following the framework introduced by Eugene Fama and Kenneth French, researchers correlate excess stock returns with three factors: the market premium, the size factor (SMB, small minus big), and the value factor (HML, high minus low book-to-market). These correlations are typically calculated using Pearson's r on monthly return data from CRSP, with sample sizes ranging from 60 monthly observations for individual stocks to several thousand for fund-level analyses.


A study correlating monthly returns of a value-tilted equity fund with the HML factor over a 20-year period might report a Pearson correlation of 0.72, indicating a strong positive relationship and supporting the fund's value style classification. A correlation of 0.15 would suggest the fund's name and stated strategy don't match its actual factor exposure, a finding with material implications for institutional investors. Finance journals such as the Journal of Finance, the Journal of Financial Economics, and the Review of Financial Studies expect correlation reporting to include the full variance-covariance matrix among factors and the time-series consistency of the correlation across market regimes.


Science: Dose-Response Relationships in Clinical Research

In clinical and pharmacological research, correlation coefficients quantify the relationship between drug exposure and physiological response. A study evaluating a new antihypertensive medication might correlate plasma drug concentration with reduction in systolic blood pressure across 200 patients. Because both variables are continuous and approximately normally distributed at therapeutic doses, Pearson's r is the appropriate choice.


A reported Pearson r of 0.65 with a 95 percent confidence interval of 0.55 to 0.73 and p less than 0.001 would indicate a moderate positive relationship between drug concentration and blood pressure reduction. The clinical significance of this correlation depends on context. In hypertension research, a correlation of 0.65 is meaningful because blood pressure is influenced by many factors beyond a single drug, including diet, sodium intake, sleep, and stress. In contrast, a study of a chemical reaction rate's correlation with temperature would be expected to produce correlations above 0.95, because the underlying physics is deterministic. The New England Journal of Medicine, the Lancet, and JAMA expect dose-response correlations to be accompanied by sample size, confidence intervals, p-values, and discussion of potential confounders such as renal function, age, and concomitant medications.


Engineering: Materials Properties and Processing Parameters

In materials science and engineering, correlation coefficients describe the relationship between processing parameters and resulting material properties. A study optimizing the strength of a 3D-printed titanium alloy component might correlate laser power (continuous, watts) with ultimate tensile strength (continuous, megapascals) across 50 specimens manufactured at varying laser power settings. Because the underlying physical relationship is deterministic but subject to manufacturing variability, Pearson's r is appropriate and high correlations are expected.


A study reporting a Pearson r of 0.91 between laser power and tensile strength would indicate a very strong relationship, with laser power explaining roughly 83 percent of the variance in tensile strength (the coefficient of determination, r squared, equals 0.83). Engineering journals such as Materials Science and Engineering A, the Journal of Materials Processing Technology, and Acta Materialia expect correlation findings in materials research to be accompanied by mechanistic explanations grounded in the physics of the process. A correlation without a mechanism is treated as preliminary. Correlations are also commonly reported alongside response surface models, design of experiments analyses, and finite element simulation comparisons that integrate the correlational finding into a broader analytical framework.


Social Sciences: Financial Behavior and Demographic Variables

In behavioral economics and consumer research, correlation coefficients describe relationships between financial behaviors and demographic, psychological, or attitudinal variables. Fisher and Yao (2017), in their study of gender differences in financial risk tolerance, used the Survey of Consumer Finances to examine correlations between risk tolerance and a range of variables including income, education, age, and household composition. The use of Spearman's rho is common in such research because Likert-scale risk tolerance measures are ordinal rather than continuous, and income distributions are typically skewed rather than normal.


A study correlating self-reported financial risk tolerance (5-point ordinal scale) with household income (continuous, log-transformed) across 6,500 households might report a Spearman rho of 0.28 with p less than 0.001. The correlation is statistically significant due to the large sample size, but its absolute magnitude is modest. In behavioral and social science research, this is a meaningful and publishable finding because risk tolerance is influenced by many factors beyond income, including personality, life stage, marital status, financial literacy, and cultural background. Journals such as the Journal of Consumer Research, the Journal of Financial Counseling and Planning, and the Journal of Family and Economic Issues expect such correlations to be reported alongside multivariate regression analyses that control for the confounding variables identified in the literature, recognizing that bivariate correlations alone rarely answer social science research questions.


How to Report Correlation Coefficients in Your Research Paper (APA Format)

The standard APA format for reporting a correlation coefficient follows a specific structure that reviewers expect to see. The basic format is r(N-2) = value, p = value, where N-2 is the degrees of freedom (sample size minus two for Pearson's r), the correlation coefficient is reported to two or three decimal places without a leading zero (.45 rather than 0.45), and the p-value is reported to three decimal places or as p < .001 if smaller.


Worked Reporting Example

Suppose you ran a study correlating study hours and exam scores across 100 students and found a Pearson correlation of 0.45. The complete APA report would look like this:


A Pearson correlation analysis revealed a statistically significant positive relationship between weekly study hours and final exam scores, r(98) = .45, 95% CI [.28, .59], p < .001, r² = .20. Study hours accounted for 20% of the variance in exam scores.


Note the structure: the type of correlation (Pearson), the degrees of freedom in parentheses after r (sample size minus 2), the correlation value without a leading zero (.45 not 0.45), the 95% confidence interval, the p-value, the coefficient of determination (r squared), and a sentence interpreting the practical magnitude. Modern journals increasingly expect all five components rather than just the correlation and p-value.


What to Include Beyond the Basic Value

Beyond the basic value and p-value, contemporary journal expectations have evolved. Most quantitative journals now expect:


  • Confidence intervals. Report the 95 percent confidence interval for the correlation coefficient. A correlation of 0.45 with a confidence interval of 0.30 to 0.58 is more informative than a correlation of 0.45 alone, because the confidence interval communicates the precision of the estimate.
  • Effect size interpretation. Don't just state that the correlation is statistically significant. Statistical significance with a large sample size can mean a trivial effect. Effect size interpretation places the correlation in context relative to your field's conventions.
  • Coefficient of determination (r squared). Reporting r squared alongside r helps readers understand how much variance in one variable is explained by the other. An r of 0.50 corresponds to an r squared of 0.25, meaning 25 percent of variance is explained, leaving 75 percent unexplained.
  • Visual presentation. Include a scatter plot showing the relationship visually, especially when the correlation is reported as part of the main findings. Modern journal practice in finance, science, engineering, and the social sciences expects both numerical and visual presentation.
  • Discussion of limitations. Acknowledge the limitations explicitly: the inability to establish causation, the directionality problem, the possibility of confounding variables, and any sample-specific limitations that constrain generalization.

Common Mistakes to Avoid

Reviewers commonly flag the following errors in correlation reporting. Avoiding them improves the credibility of your manuscript and reduces revision burden.


  • Reporting Pearson's r when assumptions aren't met. Pearson's r assumes normally distributed data, linear relationships, and continuous variables. When these assumptions aren't met, Spearman's rho or Kendall's tau is more appropriate. Reporting Pearson's r on skewed or ordinal data is one of the most common methodological errors flagged in peer review.
  • Implying causation in correlational findings. Phrases like "X caused Y," "X led to Y," or "X resulted in Y" are inappropriate when the underlying analysis is correlational. Use language like "X was associated with Y," "X correlated with Y," or "Higher values of X were observed alongside higher values of Y."
  • Ignoring outliers. A single outlier can shift a correlation coefficient substantially, particularly in small samples. Identify outliers visually through scatter plots, decide whether they're errors or legitimate observations, and report your decisions transparently.
  • Cherry-picking correlations. If you calculate dozens of correlations and report only the significant ones, you're inflating the false discovery rate. Apply Bonferroni or Benjamini-Hochberg corrections when running multiple comparisons, and report all correlations or specify clearly which subset you're reporting and why.
  • Conflating statistical significance with practical importance. A correlation of 0.05 with a sample of 10,000 will be statistically significant (p less than 0.001), but it explains only 0.25 percent of variance. Statistical significance alone doesn't make a finding meaningful.
  • Failing to graph the data. A correlation coefficient is a summary statistic. Without a scatter plot, readers can't see whether the relationship is genuinely linear, whether outliers are driving the result, or whether the data clusters in unexpected ways.
  • Confusing r with r squared. A paper reporting "r = 0.80" and a paper reporting "r² = 0.80" are reporting fundamentally different things. The first is a correlation coefficient (strong relationship); the second is a coefficient of determination (80 percent of variance explained, corresponding to an r of 0.89). Keep the two straight.

Frequently Asked Questions

How do I report a correlation coefficient in APA format?

The standard APA format for reporting a correlation coefficient is r(N-2) = value, p = value, where N-2 is the degrees of freedom (sample size minus two for Pearson's r), the correlation coefficient is reported to two or three decimal places without a leading zero (.45 rather than 0.45), and the p-value is reported to three decimal places or as p < .001 if smaller. For example: r(98) = .45, 95% CI [.28, .59], p < .001, r² = .20. Modern reporting also expects the 95% confidence interval, the effect size interpretation, and the coefficient of determination (r squared). Specific journal style guides may differ, so check your target journal's instructions before finalizing your manuscript.


What is the difference between correlation and causation?

Correlation describes the strength and direction of a linear relationship between two variables. Causation describes a directional mechanism by which one variable produces a change in another. A correlation can exist without causation, and a causal relationship can exist without an obvious correlation if the relationship is nonlinear or moderated by other variables. Correlational research can identify associations between variables but can't establish causation, because the directionality of the relationship is unknown and confounding variables may be driving both observed variables. Establishing causation typically requires a randomized controlled experiment or a causal inference framework such as instrumental variables, regression discontinuity, or difference-in-differences analysis.


When should I use Pearson's r versus Spearman's rho?

Pearson's r is appropriate when both variables are continuous, normally distributed, and the relationship between them is approximately linear. Spearman's rho is appropriate when one or both variables are ordinal, when the data aren't normally distributed, or when the relationship is monotonic but not necessarily linear. Spearman's rho works on ranked data rather than the actual values, making it more robust to outliers and non-normal distributions. For survey research using Likert scales, Spearman's rho is typically the right choice. For continuous physical or financial measurements, Pearson's r is typically the right choice. When in doubt, run both and report the one that matches your data type and your field's conventions.


What are the types of correlation coefficients?

The main types are Pearson's r (for two continuous, normally distributed variables with a linear relationship), Spearman's rho (for ordinal data or non-normal distributions), Kendall's tau (for small samples or data with many tied ranks), phi coefficient (for two dichotomous variables in a 2x2 table), Cramer's V (for categorical variables in larger contingency tables), and point-biserial correlation (for one continuous and one dichotomous variable). The choice depends on the measurement level of the variables, their distribution, and whether the relationship is expected to be linear. Choosing the wrong coefficient is one of the most common methodological errors in correlational research.


What sample size do I need for correlational research?

Sample size depends on the expected correlation strength, the desired statistical power, and the significance level. For a medium effect size (r approximately 0.30), 80% power, and an alpha of 0.05, a sample size of approximately 84 observations is typically sufficient. For a small effect size (r approximately 0.10), the required sample size grows to approximately 782 observations. For a large effect size (r approximately 0.50), approximately 28 observations are sufficient. Power analysis software such as G*Power, R's pwr package, or commercial alternatives like SAS or SPSS can calculate exact sample size requirements for specific research designs. Journals in finance and the social sciences increasingly expect a priori power calculations to be reported in the methods section.


How do I handle outliers in correlational research?

Identify outliers through visual inspection of scatter plots and statistical methods such as Cook's distance, leverage values, and standardized residuals. Once identified, determine whether each outlier represents a measurement error (in which case it should be corrected or removed) or a legitimate observation that's simply far from the cluster (in which case removal is more controversial). When outliers are legitimate observations, options include reporting the analysis with and without the outliers, using Spearman's rho or Kendall's tau which are less sensitive to outliers, or applying robust correlation methods such as the Winsorized correlation or the percentage-bend correlation. Whatever approach is taken, document and justify the decision transparently in the methods section so reviewers can evaluate the analysis on its merits.


Can correlation coefficients be used for time series data?

Standard Pearson and Spearman correlation coefficients can be calculated on time series data, but they often produce misleading results because of autocorrelation, where each observation depends on previous observations. In financial time series, for example, the apparent correlation between two stock prices may simply reflect that both have generally trended upward over time rather than that they're genuinely related. Specialized methods address this issue. Cross-correlation functions handle autocorrelation explicitly. Cointegration analysis tests whether two non-stationary time series share a common long-run trend. Granger causality tests examine whether one time series helps predict another. Time series correlation analysis requires careful attention to stationarity, structural breaks, and the underlying economic or physical mechanism connecting the series.


What is a good correlation coefficient value?

What counts as a good correlation depends on the field and the research context. General thresholds classify correlations as negligible (0.00 to 0.10), weak (0.10 to 0.39), moderate (0.40 to 0.69), strong (0.70 to 0.89), and very strong (0.90 to 1.00). However, these thresholds aren't universal. In physics and chemistry, where measurement precision is high and underlying relationships are often deterministic, researchers expect very strong correlations above 0.90. In psychology and the social sciences, where human behavior is influenced by many factors simultaneously, a correlation of 0.30 may represent an important and publishable finding. Calibrate your interpretation to the conventions of your field and target journal rather than applying universal thresholds.


How do I write the limitations section for a correlational study?

A limitations section for correlational research should explicitly acknowledge the inability to establish causation, the potential for the directionality problem (whether X causes Y, Y causes X, or both), and the possibility of confounding by third variables that influence both observed variables. Discuss the sample's generalizability: did you study a specific population, time period, or institutional context that might limit how broadly your findings apply? Address measurement limitations: were the variables measured by valid and reliable instruments? Discuss any nonresponse, attrition, or selection bias that might affect the correlations observed. Modern journals expect a substantive limitations section rather than a brief paragraph, often with specific suggestions for follow-up research that could address the limitations.


Professional Editing for Your Research Manuscript

The way you report correlation coefficients in your results section signals to reviewers how carefully you thought about the analysis. Studies that match the correct coefficient to the data, report all required components (coefficient, degrees of freedom, confidence interval, p-value, r squared), include a scatter plot, and discuss limitations transparently fare better in peer review than studies that rush the reporting. Unclear or incomplete correlation reporting is one of the most common reasons quantitative manuscripts get sent back for major revisions.


Editor World's journal article editing service, dissertation editing service, and academic editing service connect researchers with native English editors whose subject matter expertise matches the manuscript. A finance researcher gets an editor with empirical asset pricing experience. A clinical researcher gets an editor with biomedical research experience. A materials scientist gets an editor with engineering manuscript experience. A social science researcher gets an editor with quantitative behavioral or economics research experience. Browse editor profiles by discipline and credentials before submitting.


All editing is returned in Track Changes in Microsoft Word so you can review, accept, or reject each correction individually. American English is applied by default. A certificate of editing confirming human-only native English editing is available as an optional add-on for journal submissions where AI use must be disclosed. Same-day editing is available with 2-hour, 4-hour, and 8-hour turnaround options for urgent journal deadlines, available 24/7 year-round. Pricing is fully transparent through an instant price calculator that shows your exact cost before you commit.


For more guidance on statistics and research methodology, see our companion guides on correlation coefficients explained, simple linear regression, hypothesis testing, confounding variables, and statistics for researchers.



This article was reviewed by the Editor World editorial team. Editor World, founded in 2010 by Patti Fisher, PhD, graduate of The Ohio State University, provides professional editing and proofreading services for academic researchers, doctoral candidates, faculty, business professionals, and authors worldwide. BBB A+ accredited since 2010 with 5.0/5 Google Reviews and 5.0/5 Facebook Reviews. More than 100 million words edited for over 8,000 clients in 65+ countries. Native English editors from the United States, the United Kingdom, and Canada with subject-matter expertise across the social sciences, the natural and physical sciences, medicine, engineering, computer science, and the humanities. 100% human editing, no AI at any stage.