. 2019 Apr:89:2445-2453.

Interpretable Almost-Exact Matching for Causal Inference

Awa Dieng¹, Yameng Liu¹, Sudeepa Roy¹, Cynthia Rudin¹, Alexander Volfovsky¹

Affiliations

PMID: 31198908
PMCID: PMC6563929

Interpretable Almost-Exact Matching for Causal Inference

Awa Dieng et al. Proc Mach Learn Res. 2019 Apr.

. 2019 Apr:89:2445-2453.

Authors

Awa Dieng¹, Yameng Liu¹, Sudeepa Roy¹, Cynthia Rudin¹, Alexander Volfovsky¹

Affiliation

¹ Duke University, Durham, NC 27708.

PMID: 31198908
PMCID: PMC6563929

Abstract

Matching methods are heavily used in the social and health sciences due to their interpretability. We aim to create the highest possible quality of treatment-control matches for categorical data in the potential outcomes framework. The method proposed in this work aims to match units on a weighted Hamming distance, taking into account the relative importance of the covariates; the algorithm aims to match units on as many relevant variables as possible. To do this, the algorithm creates a hierarchy of covariate combinations on which to match (similar to downward closure), in the process solving an optimization problem for each unit in order to construct the optimal matches. The algorithm uses a single dynamic program to solve all of the units' optimization problems simultaneously. Notable advantages of our method over existing matching procedures are its high-quality interpretable matches, versatility in handling different data distributions that may have irrelevant variables, and ability to handle missing data by matching on as many available covariates as possible.

PubMed Disclaimer

Figures

**Figure 1:**
Estimated CATT vs. True CATT (Conditional Average Treatment Effect on the Treated). DAME and FLAME perfectly estimate the CATTs before dropping important covariates. DAME matches all units without dropping important covariates, but FLAME needs to stop early in order to avoid poor matches. All other methods are sensitive to irrelevant covariates and give poor estimates. The two numbers on each plot are the number of matched units and MSE.

**Figure 2:**
DAME makes higher quality matches early on. Rows correspond to stopping thresholds (top row 30%, bottom row 50%). DAME matches on more covariates than FLAME, yielding lower MSE from matched groups.

**Figure 3:**
Run-time comparison between DAME FLAME, and brute force. *Left*: varying number of units. *Right*: varying number of covariates.

**Figure 4:**
Number Matched: Number of units matched per covariates for the BTC data

**Figure 5:**
Histogram of estimated CATE by DAME. For individuals where the CATE is negative, it means that BTC was estimated to reduce crime.

**Figure 6:**
Comparison between DAME and SVM-based method

See this image and copyright information in PMC

References

1. Abadie A, Drukker D, Herr JL, and Imbens GW. Implementing matching estimators for average treatment effects in stata. The Stata Journal, 4(3): 290–311, 2004.
1. Agrawal R and Srikant R. Fast algorithms for mining association rules in large databases. In Proceedings of the 20th International Conference on Very Large Data Bases, VLDB ‘94, pages 487–499, 1994.
1. Athey S, Tibshirani J, and Wager S. Generalized Random Forests. The Annals of Statistics, 47(2): 1148–1178, 2019.
1. Belloni A, Chernozhukov V, and Hansen C. Inference on treatment effects after selection among highdimensional controls. The Review of Economic Studies, 81(2):608–650, 2014.
1. Cochran WG and Rubin DB. Controlling bias in observational studies: A review. Sankhyã: The Indian Journal of Statistics, Series A, pages 417–446, 1973.

Grants and funding

R01 EB025021/EB/NIBIB NIH HHS/United States

LinkOut - more resources

Full Text Sources
- Europe PubMed Central
- PubMed Central

Save citation to file

Email citation

Add to Collections

Add to My Bibliography

Your saved search

Create a file for external citation management software

Your RSS Feed

Interpretable Almost-Exact Matching for Causal Inference

Affiliation

Interpretable Almost-Exact Matching for Causal Inference

Authors

Affiliation

Abstract

Figures

References

Grants and funding

LinkOut - more resources

Full Text Sources