Skip to main page content
U.S. flag

An official website of the United States government

Dot gov

The .gov means it’s official.
Federal government websites often end in .gov or .mil. Before sharing sensitive information, make sure you’re on a federal government site.

Https

The site is secure.
The https:// ensures that you are connecting to the official website and that any information you provide is encrypted and transmitted securely.

Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2015;110(512):1726-1734.
doi: 10.1080/01621459.2014.993081. Epub 2015 Jan 23.

CONDITIONAL DISTANCE CORRELATION

Affiliations

CONDITIONAL DISTANCE CORRELATION

Xueqin Wang et al. J Am Stat Assoc. 2015.

Abstract

Statistical inference on conditional dependence is essential in many fields including genetic association studies and graphical models. The classic measures focus on linear conditional correlations, and are incapable of characterizing non-linear conditional relationship including non-monotonic relationship. To overcome this limitation, we introduces a nonparametric measure of conditional dependence for multivariate random variables with arbitrary dimensions. Our measure possesses the necessary and intuitive properties as a correlation index. Briefly, it is zero almost surely if and only if two multivariate random variables are conditionally independent given a third random variable. More importantly, the sample version of this measure can be expressed elegantly as the root of a V or U-process with random kernels and has desirable theoretical properties. Based on the sample version, we propose a test for conditional independence, which is proven to be more powerful than some recently developed tests through our numerical simulations. The advantage of our test is even greater when the relationship between the multivariate random variables given the third random variable cannot be expressed in a linear or monotonic function of one random variable versus the other. We also show that the sample measure is consistent and weakly convergent, and the test statistic is asymptotically normal. By applying our test in a real data analysis, we are able to identify two conditionally associated gene expressions, which otherwise cannot be revealed. Thus, our measure of conditional dependence is not only an ideal concept, but also has important practical utility.

Keywords: Conditional distance correlation; Conditional independence test; Local bootstrap; U(V) process with random kernel.

PubMed Disclaimer

Figures

Figure 1
Figure 1
Residual Plot of Regressing YIMPG1 against XLOC361100 Conditional on XCFH

Similar articles

Cited by

References

    1. Ackley H, Hinton E, Sejnowski J. A learning algorithm for Boltzmann machines. Cognitive Science. 1985:147–169.
    1. Fan Y, Li Q. Consistent model specification tests: omitted variables and semiparametric functional forms. Econometrica: Journal of the Econometric Society. 1996:865–890.
    1. Fukumizu K, Gretton A, Sun X, Schölkopf B. Kernel measures of conditional dependence. Conference on Neural Information Processing Systems 2008
    1. Gretton A, Bousquet O, Smola A, Schölkopf B. Measuring statistical dependence with Hilbert-Schmidt norms. Proceedings Algorithmic Learning Theory. 2005:63–77.
    1. Hageman GS, Anderson DH, Johnson LV, Hancox LS, Taiber AJ, Hardisty LI, Hageman JL, Stockman HA, Borchardt JD, Gehrs KM, et al. A common haplotype in the complement regulatory gene factor H (HF1/CFH) predisposes individuals to age-related macular degeneration. Proceedings of the National Academy of Sciences of the United States of America. 2005;102(20):7227–7232. - PMC - PubMed

LinkOut - more resources