conditional dependence modeling

Designs, specifies, and evaluates statistical or probabilistic models that represent conditional dependence relationships among multiple variables or outcomes, i.e., the multivariate dependence or dependence structure of a joint distribution given covariates. Builds estimation and testing procedures to capture multivariate/conditional dependence, adjust joint estimates for dependence, and reduce bias that arises from incorrect independence assumptions.

conditionaldependencemodeling

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
0.05
Oct 01, 2026Oct 01, 2026
Career
Value
No comparison yet
$200K/year
Oct 01, 2026Oct 01, 2026

Must-Read Papers

Most classic and influential ideas
View more

This study addresses the challenge of effectively characterizing nonlinear statistical dependence between two random variables while controlling for the influence of covariates. To this end, it proposes partial copula as a theoretical framework for nonlinear partial correlation, extending the classical notion of linear partial correlation to more general nonlinear settings. The work establishes a formal connection between partial copulas and conditional copula dependence structures. Through rigorous theoretical analysis and simulation experiments, the study demonstrates that partial copulas accurately capture dependence relationships after adjusting for covariates. Moreover, it reveals their potential in causal inference for identifying the sign of causal effects, thereby offering a novel tool for nonlinear causal discovery.

causal inferenceconditional independencecovariate adjustment

Bayesian multivariate models for bounded directional data

Jul 15, 2025
JM
Joel Montesinos-Vazquez
🏛️ Universidad Autónoma Metropolitana–Unidad Iztapalapa

Directional data are often constrained to local intervals on the unit circle (e.g., the first quadrant), and existing multivariate models struggle to simultaneously ensure bounded marginal supports and flexible dependence structures. To address this, we propose a copula-based Bayesian multivariate model: marginal variables are defined on subsets of the unit circle, while joint dependence is flexibly modeled via a copula function. We introduce a projected Gamma prior and a two-stage MCMC sampling scheme for posterior inference. This work constitutes the first systematic extension of multivariate modeling frameworks to bounded circular data. Through extensive simulations and real-data applications, we demonstrate the model’s high accuracy in estimating both the joint distribution and underlying parameters. It significantly enhances statistical modeling capability for restricted directional data, offering improved flexibility, interpretability, and inferential precision compared to existing approaches.

Developing flexible multivariate models for first-quadrant circular variablesModeling bounded directional data on k-dimensional sphere subsetsUsing copula functions and Bayesian inference for parameter estimation

Testing Multivariate Conditional Independence Using Exchangeable Sampling and Sufficient Statistics

Apr 09, 2025
XL
Xiaotong Lin
🏛️ National University of Singapore

This paper addresses the problem of testing multivariate conditional independence between a response variable (Y) and high-dimensional covariates (X) given confounders (Z). We propose the Multivariate Sufficient-statistic-based Conditional Randomization Test (MS-CRT), which constructs conditional exchangeability by leveraging sufficient statistics of (P(X mid Z)), bypassing explicit modeling of (Y) and accommodating arbitrary test statistics. Key contributions include: (i) the first integration of sufficient statistics into the CRT framework; (ii) overcoming the curse of dimensionality in (P(X mid Z)) by reducing dependence on its parametric dimension; (iii) enabling group selection and false discovery rate (FDR) control; and (iv) establishing minimax-optimal detection rates under multivariate normality. Extensive simulations and real-data analyses demonstrate that MS-CRT substantially improves joint signal detection power—particularly when individual components of (X) exert weak effects on (Y)—and consistently outperforms state-of-the-art methods in graphical model learning tasks.

Extends to group selection with false discovery rate controlIntroduces MS-CRT for exchangeable sampling without modeling YTests multivariate conditional independence between Y and X given Z

Detecting dependence structure: visualization and inference

Oct 08, 2024
BĆ
Bogdan Ćmiel
🏛️ AGH University of Krakow | Institute of Mathematics | Polish Academy of Sciences

This paper addresses the problem of interpretable detection of dependency structures among random variables. We propose a novel framework integrating a rank-transform-based estimator for the quantile dependence function with a local acceptance region. The method constructs robust quantile dependence measures via rank standardization and employs local hypothesis testing to enable visual diagnostic assessment of dependency patterns and rigorous independence testing under finite samples. Key contributions include: (1) the first nonparametric estimation and theoretical derivation of the quantile dependence function; (2) guaranteed validity of statistical tests at any sample size, balancing high global power with precise localization of heterogeneous dependencies; (3) superior empirical power across diverse alternative models and successful identification of heterogeneous non-independence in real-world data; and (4) a computationally efficient algorithm supporting intuitive, graphical diagnostic interpretation.

Identifying dependency between two random variablesTesting independence with high power and efficiencyVisualizing and evaluating dependence structure rigorously

Testing the Homogeneity of Two Proportions for Correlated Bilateral Data via the Clayton Copula

Feb 01, 2025
SL
Shuyi Liang
🏛️ University at Buffalo | Institute of Statistical Mathematics | The First Affiliated Hospital of Xiamen University

In clinical trials—particularly ophthalmology—homogeneity testing for bilateral proportion data is commonly constrained by pre-specified, inflexible dependence structures (e.g., independence, perfect positive/negative dependence), limiting interpretability and adaptability. This paper introduces the Clayton copula—a flexible, interpretable tool for modeling asymmetric lower-tail dependence—into bilateral proportion homogeneity testing for the first time, thereby eliminating reliance on a priori dependence assumptions. We propose three Clayton copula–based test statistics and rigorously evaluate them via Monte Carlo simulation, demonstrating well-controlled Type I error rates and superior statistical power. Furthermore, we validate the robustness and practical utility of our approach on two real-world ophthalmologic datasets. This work establishes a theoretically rigorous, computationally feasible, and clinically meaningful testing paradigm for bilateral proportion data, advancing both methodological foundations and applied biostatistical practice.

Addressing inflexible dependence structures in clinical trialsEvaluating Clayton copula performance for dependent data analysisTesting proportion homogeneity for correlated bilateral data

Latest Papers

What's happening recently
View more

This study addresses the challenge of modeling edge effects and dependence structures that evolve with covariates—such as age—in multivariate responses of mixed types. Existing approaches are often hindered by strong assumptions or insufficient flexibility. To overcome these limitations, this work proposes a Bayesian nonparametric framework that integrates adaptive spline-based marginal regression with a covariate-dependent Gaussian copula infinite mixture model. A probit stick-breaking process is introduced to flexibly capture the covariate-driven evolution of dependence patterns, avoiding restrictive global correlation matrix constraints. The method unifies heterogeneous response types and dynamic dependencies through varying-coefficient copula regression and employs Markov chain Monte Carlo algorithms for posterior inference. Simulation studies demonstrate its accuracy and robustness, while empirical analysis of the 2023 Behavioral Risk Factor Surveillance System (BRFSS) data reveals complex age-varying marginal and dependence structures in health outcomes.

conditional copuladependence structuremarginal effects

This study addresses the challenge of achieving consistent conditional independence testing and association estimation in continuous sample spaces under small-sample, high-dimensional settings. The authors propose a nonparametric multiscale approach that decomposes the continuous space via cascaded 2×2×T contingency tables and conditions on marginal order statistics. This framework extends the Cochran–Mantel–Haenszel (CMH) test and odds ratio estimation to continuous variables for the first time, ensuring statistical consistency without requiring asymptotic layer-wise sample sizes. The method simultaneously supports hypothesis testing and identification of local association strength and direction, with near-linear computational complexity. Empirical evaluations demonstrate its superior or competitive statistical power while properly controlling Type I error, and it successfully uncovers local conditional dependence structures in real-world Uber mobility data.

conditional associationconditional independencecontinuous sample space

This study addresses the challenge of modeling electronic health records (EHR), which comprise high-dimensional, mixed-type variables with complex nonlinear dependencies that are poorly captured by traditional statistical methods due to their reliance on strong distributional assumptions such as Gaussianity. To overcome this limitation, the work introduces vine copulas—a flexible probabilistic framework—for the first time in EHR analysis. By decomposing multivariate distributions into a sequence of bivariate conditional dependencies arranged in a tree structure, the approach enables accurate modeling of heterogeneous data types without restrictive parametric assumptions. The proposed method facilitates data-driven variable selection, identification of conditional dependencies among comorbidities, and characterization of patient cohorts. Accompanied by visualization tools and open-source code, this framework promotes reproducible and interpretable probabilistic exploration of healthcare data.

Electronic Health RecordsHigh-dimensional DataMixed-type Data

Traditional meta-analyses of combined diagnostic tests often yield biased estimates due to the assumption of conditional independence, and existing approaches either require complete joint data or suffer from computational instability. This study proposes a Bayesian hierarchical model that flexibly captures the conditional dependence between two binary diagnostic tests through study-specific log odds ratios, without imputing missing data or assuming a perfect reference standard. The method provides a unified framework for synthesizing heterogeneous study designs—including those without a gold standard or with partial verification—using a stable parameterization that mitigates bias in accuracy estimation. Validation through two real-world meta-analyses demonstrates that ignoring conditional dependence substantially distorts results, whereas the proposed framework yields accurate estimates of joint diagnostic performance while maintaining computational stability.

conditional dependencediagnostic test accuracyimperfect reference standard

Hot Scholars

LF

Long Feng

Professor of Nankai University
High Dimensional DataHigh Frequency Data
AR

Aaditya Ramdas

Associate Professor (with tenure), Carnegie Mellon University
Machine LearningStatistics
PZ

Ping Zhao

Hefei University of Technology
Mechanism and RoboticsRehabilitation RoboticsMotion SynthesisComputational Kinematics
RW

Ruodu Wang

University of Waterloo
StatisticsRisk ManagementActuarial ScienceFinancial Engineering
MB

Matteo Barigozzi

Full Professor - Alma Mater Studiorum Università di Bologna
Time Series Analysis - High dimensional data - Factor models - Networks