REVIEW 3 major objections 5 minor 47 references
Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective
T0 review · 3 major / 5 minor · reviewed 2026-08-07 · deepseek-v4-flash
Pith's one-line read This paper proves the misreporting rate is identifiable from the gap between nominal and true causal effects of a feature on its downstream outcome, with no access to the true feature.
desk verdict Clean identification argument and a smart negative control, but Assumption 4 carries more weight than the paper acknowledges and the real-data evidence is indirect. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing object is the ratio identity MR_a = (τ'_a − τ_a)/δ'_a, operationalized as the Causal Misreporting Estimator (CMRE). The three ingredients are: τ_a, the nominal average causal effect of the reported feature X on the outcome Y among the reported group (computed from manipulated data); τ'_a, the true average causal effect of the unobserved feature X* on Y, learned from unmanipulated data and averaged over the same reported group's confounder distribution; and δ'_a, that same true effect averaged over the X* = 0 group that could be misreported. The mechanism is the descendant asymmetry: genuine modification alters Y through X*, while misreporting alters only X, leaving Y untouched — so the gap between nominal and true effects is exactly the misreporting rate times the effect on the misreported group.
What would settle it
Audit a random sample of agents to obtain the true feature X* and the actual misreporting rate, then compare the CMRE estimate; the method is refuted if the discrepancy grows in settings where the conditional treatment effect of X* on Y is deliberately made different between the manipulated and unmanipulated populations. A simpler check: compute the MR for the same agent using two different downstream outcomes with large δ'_a; if the two estimates disagree beyond sampling error, the identifying assumptions fail.
Extended reading notes
Core claim
Under Assumptions 1–4, the paper proves Theorem 1: for δ'_a ≠ 0, the misreporting rate P_a(X* = 0 | X = 1) is identifiable from two datasets and equals (τ'_a − τ_a)/δ'_a. The proof shows that the nominal effect τ_a of the reported X on the descendant Y decomposes into the true effect τ'_a minus the misreporting rate times the true effect δ'_a on the misreported group; this decomposition relies on the fact that misreported features have no causal effect on descendants, so any discrepancy between nominal and true effects is attributable solely to lying. Assumption 4, which posits that conditional treatment effects of X* on Y are identical across the manipulated and unmanipulated populations, allows replacing the unobservable τ*_a and δ*_a with τ'_a and δ'_a estimated from clean data. The paper further derives the asymptotic variance of the estimator (Theorem 2), showing that the variance grows without bound as δ'_a approaches zero.
Load-bearing premise
The load-bearing premise is that the conditional effect of the true feature on the outcome is identical in the potentially-misreporting population and the clean comparison population; if genuine modification by agents changes how the true feature affects the outcome, the estimate is biased.
Editorial extensions
If this is right
- A decision maker can estimate each agent's average misreporting rate per feature using only the manipulated dataset plus an unmanipulated dataset (e.g., pre-deployment or government data), with no access to true features or audit labels.
- Among available downstream variables, the one with the largest causal effect δ'_a on the outcome should be used; the variance analysis shows estimates become unstable as δ'_a approaches zero.
- The identifiability of P_a(X* = 0 | X = 1) extends directly to other estimands: the false positive rate P_a(X = 1 | X* = 0) and the marginal difference P_a(X = 1) − P_a(X* = 1) are also identifiable.
- Empirically, the method reproduces the expected pattern in Medicare Advantage: non-payment HCCs (no incentive to lie) have misreporting rates indistinguishable from zero, while payment HCCs show significantly positive rates, whereas baseline methods give implausible estimates.
- The method applies without change to settings with selection bias, unobserved confounding between agent and outcome, or mediator-based genuine modification, per the DAGs in Appendix A.
Reading between the lines
- A natural next use is triage: run CMRE on all agents to produce a misreporting rate per feature, then target expensive manual audits only at agents whose estimated rate is high, using the estimate as a cheap prior for where to look.
- The same ratio identity could be combined with sensitivity analysis for unobserved confounding; the authors flag no-unmeasured-confounding as a limitation, but existing bounds on treatment effects could turn the point estimate into an interval.
- A falsifiable consistency check within a single application would compute the MR using two different downstream outcomes Y1 and Y2; if the causal assumptions hold and both effects are nonzero, the point estimates should agree, giving a data-driven diagnostic for Assumption 4.
- The insight may generalize beyond binary features to multivalued or continuous reports by aggregating over thresholds, though the paper only treats binary X.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper proposes a causal estimator for the rate at which strategic agents misreport a binary feature X* (the misreporting rate, MR = P_a(X*=0|X=1)) when the true feature is unobserved, in settings where agents may also genuinely modify the feature. The key idea is that genuine modification changes the causal descendants of X*, whereas misreporting of X does not affect the downstream outcome Y. Under Assumptions 1-4, the paper proves in Theorem 1 that MR = (τ'_a - τ_a)/δ'_a, where τ_a is the nominal effect of reported X on Y in the manipulated data, τ'_a is the true effect of X* on Y transported from unmanipulated data, and δ'_a is the true effect for the X*=0 group. The paper also provides an asymptotic variance formula (Theorem 2) and validates the method on a semi-synthetic loan-fraud simulation and on Medicare claims data, where the estimator is used to compare misreporting of payment versus nonpayment HCCs.
Significance. If the identification result holds, the paper offers a useful new tool for auditing and strategic classification: it estimates an aggregate misreporting rate without observing true features and without per-agent audits, using only a manipulated dataset and an unmanipulated dataset. The causal framing is elegant, and the paper gives a formal identifiability proof, a variance characterization that guides feature choice, and extensive semi-synthetic experiments with several baselines. The proof of Theorem 1 is algebraically sound under the stated assumptions, and the paper is honest about the need for no unobserved confounding. However, the practical value depends critically on Assumption 4 (transportability of conditional average treatment effects from the unmanipulated to the manipulated population), which is not validated in the real-data demonstration; the Medicare analysis also lacks ground truth and uses post-hoc selection on the very quantities entering the estimator. These issues do not invalidate the theoretical core but do limit the strength of the empirical claims.
major comments (3)
- [Section 3 (Assumption 4); Section 4.1 (Theorem 1); Section 5.2] Assumption 4 is the only bridge that replaces the unidentifiable τ*_a and δ*_a with the identifiable τ'_a and δ'_a in the proof of Theorem 1 (Appendix B.2). In the Medicare application, D* is Traditional Medicare stayers and D is switchers to private insurers; these populations differ in care and health trajectories, and the genuine modification that the method is designed to accommodate can itself change the conditional effect of an HCC on mortality. The nonpayment-HCC negative control in §5.2/Figure 3 tests the full estimator on HCCs with no payment incentive; it does not validate Assumption 4 for the payment HCCs, which are precisely the ones for which modification incentives exist. Please add a sensitivity analysis (e.g., bounds on MR as a function of the degree of violation of Assumption 4) or direct evidence that the relevant CATEs transport from TM stayers to MA switchers, and state plainly that the Medicare point estimates are conditional on untestable transportability.
- [Section 4.1 (Lemma 1); Appendix B.1-B.2] The proofs use an unstated random-misreporting assumption: in Step 2 of Lemma 1 (Appendix B.1) and in the proof of Theorem 1 (Appendix B.2), P_a(C|X*=0,X=1) is replaced by P_a(C|X*=0), which requires C⊥X|X*,A. The text states this informally ('the misreported group will be a random sample of the group where X*=0') but it is not listed among Assumptions 1-4 and is not defended in the Medicare application, where upcoding decisions may depend on enrollee demographics and prior HCCs. If misreporting is targeted based on C, the equality fails and the estimator is biased; please state this assumption explicitly and discuss its plausibility, or relax it.
- [Section 5.2; Tables 1-2] The real-data demonstration has no ground truth, and the HCC selection rule is applied after estimating the same quantities that enter the estimator: only HCCs with at least 1% prevalence and an estimated causal effect δ' > 0.1 are reported (Section 5.2 and Tables 1-2). This post-hoc selection is not accounted for in the bootstrap confidence intervals, so the 'sanity check' for nonpayment HCCs is not a falsifiable validation of the method. Please report the full set of HCCs (or pre-specify the selection rule) and explicitly frame the Medicare results as assumption-dependent estimates rather than measured misreporting rates.
minor comments (5)
- [Appendix C, Corollary A2] The statement 'P_a(X=1|X*=0) = (τ'_a - τ_a)/δ'_a × P_a(X=1)' is algebraically inconsistent with the proof, which derives a ratio involving P_a(X=0) + P_a(X=1,X*=0); please correct the stated formula.
- [Throughout] There are several typos ('Defnition', 'rearanging', 'maximume', 'eduction'); please proofread the manuscript carefully.
- [Theorem 2] Theorem 2 assumes N=M=n, but the Medicare experiment uses very different sample sizes for D and D* (868,255 stayers versus 166,539 switchers); please clarify whether the variance formula is intended for unequal sample sizes or note that it is a simplification.
- [Section 5.1 and Appendix E.4] The description of the OC-SVM baseline is inconsistent: Section 5.1 says it is trained on D* where X*=1, while Appendix E.4 says it is trained on (Y,C) from D*_1; please make the descriptions consistent.
- [Figure 3 caption] The caption refers to a vertical dashed line separating nonpayment and payment HCCs, but the left panel only shows the four HCC labels; please make the figure and legend self-contained.
Circularity Check
No circular derivation: Theorem 1 follows from Lemma 1 plus the explicit Assumption 4 (CATE invariance), and no fitted parameter is relabeled as a prediction; the only self-citation (Chang et al.) is in Related Work and is not load-bearing.
full rationale
The paper's central identification result, Theorem 1, is not equivalent to its inputs by construction. Lemma 1 derives MR = (tau*_a - tau_a)/delta*_a from the definitions of TAFR, NAFR, and TAFM, the DAG conditional independences, and Assumptions 1-3, via algebraic decomposition of tau_a. Theorem 1 then replaces tau*_a and delta*_a by tau'_a and delta'_a using Assumption 4, which states conditional average treatment effects of X* on Y are equal in P_a and P*. This is an identification assumption, not a circular reduction: the estimand P_a(X*=0|X=1) does not appear in the definition of tau'_a, tau_a, or delta'_a, and those quantities are separately estimable from D and D*. The proof also uses the DAG implication P_a(C|X=0)=P_a(C|X*=0), which follows from Assumption 1 and conditional independence, not from the target result. No parameter is fit to a subset of the target and then renamed a prediction; the semi-synthetic experiments use simulated ground truth, and the Medicare nonpayment-HCC comparison is an external sanity check. The only overlap with prior work by the same authors is the citation of Chang et al. [5] in Related Work, which is descriptive and not used as evidence for any theorem or estimator. The untestability and plausible violation of Assumption 4 in the Medicare application is a validity and robustness concern, not a circularity; the paper itself lists no-unmeasured-confounding as a limitation, and the negative control for nonpayment HCCs does not isolate Assumption 4 by itself. Accordingly, the derivation chain is self-contained apart from a minor, non-load-bearing self-citation, giving score 2.
Assumptions & free parameters
free parameters (3)
- Simulation misreporting probability mu =
picked to target MR (default 0.2)
- XGBoost hyperparameters =
learning rate 0.3, max depth 6, L2 reg 1
- OC-SVM hyperparameters =
nu=0.01, gamma=0.1
assumptions (6)
- domain assumption Assumption 1 (Optimal Misreporting): agents report X=1 whenever X*=1, and misreporting only flips X*=0 to X=1.
- domain assumption Assumption 2 (Useful Modifications): agents may only misreport or genuinely modify X*, not other features.
- domain assumption Assumption 3 (No unmeasured confounding, overlap, consistency): Y(0), Y(1) are independent of X* given C, and positivity holds in both populations.
- domain assumption Assumption 4 (CATE invariance): E_Pa[Y(1)-Y(0)|C=c] = E_P*[Y(1)-Y(0)|C=c] for all c and a.
- domain assumption C is independent of X given X* and A (misreporting is independent of confounders once the true feature and agent are fixed).
- domain assumption X* and X are binary.
Cite this review
Pith. "Pith review of Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective." pith.science (2026). https://pith.science/paper/VW2HYOLY
@misc{pith2026250523954,
author = {Pith},
title = {Pith review of: Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective},
year = {2026},
howpublished = {\url{https://pith.science/paper/VW2HYOLY}},
note = {Machine review of arXiv:2505.23954}
}
read the original abstract
In settings where ML models are used to inform the allocation of resources, agents affected by the allocation decisions might have an incentive to strategically change their features to secure better outcomes. While prior work has studied strategic responses broadly, disentangling misreporting from genuine modification remains a fundamental challenge. In this paper, we propose a causally-motivated approach to identify and quantify how much an agent misreports on average by distinguishing deceptive changes in their features from genuine modification. Our key insight is that, unlike genuine modification, misreported features do not causally affect downstream variables (i.e., causal descendants). We exploit this asymmetry by comparing the causal effect of misreported features on their causal descendants as derived from manipulated datasets against those from unmanipulated datasets. We formally prove identifiability of the misreporting rate and characterize the variance of our estimator. We empirically validate our theoretical results using a semi-synthetic and real Medicare dataset with misreported data, demonstrating that our approach can be employed to identify misreporting in real-world scenarios.
Figures
Figures from the paper (7 more)
Reference graph
Works this paper leans on
-
[1]
Eric P Baumer, JW Andrew Ranson, Ashley N Arnio, Ann Fulmer, and Shane De Zilwa. Illuminating a dark side of the american dream: assessing the prevalence and predictors of mortgage fraud across us counties.American Journal of Sociology, 123(2):549–603, 2017
work page 2017
-
[2]
Jason Brown, Mark Duggan, Ilyana Kuziemko, and William Woolston. How does risk selection respond to risk adjustment? new evidence from the medicare advantage program.American Economic Review, 104(10):3335–3364, 2014
work page 2014
-
[3]
Caroline S Carlin, Roger Feldman, and Jeah Jung. The mechanics of risk adjustment and incentives for coding intensity in medicare.Health services research, 59(3):e14272, 2024
work page 2024
-
[4]
Anomaly detection: A survey.ACM computing surveys (CSUR), 41(3):1–58, 2009
Varun Chandola, Arindam Banerjee, and Vipin Kumar. Anomaly detection: A survey.ACM computing surveys (CSUR), 41(3):1–58, 2009
2009
-
[5]
Trenton Chang, Lindsay Warrenburg, Sae-Hwan Park, Ravi Parikh, Maggie Makar, and Jenna Wiens. Who’s gaming the system? a causally-motivated approach for detecting strategic adaptation.Advances in Neural Information Processing Systems, 37:42311–42348, 2024
work page 2024
-
[6]
XGBoost: A scalable tree boosting system
Tianqi Chen and Carlos Guestrin. XGBoost: A scalable tree boosting system. InProceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, KDD ’16, pages 785–794, New York, NY , USA, 2016. ACM. ISBN 978-1-4503-4232-2. doi: 10.1145/2939672.2939785. URLhttp://doi.acm.org/10.1145/2939672.2939785
-
[7]
Xgboost: A scalable tree boosting system
Tianqi Chen and Carlos Guestrin. Xgboost: A scalable tree boosting system. InProceedings of the 22nd acm sigkdd international conference on knowledge discovery and data mining, pages 785–794, 2016
2016
-
[8]
Double/debiased machine learning for treatment and structural parameters, 2018
Victor Chernozhukov, Denis Chetverikov, Mert Demirer, Esther Duflo, Christian Hansen, Whitney Newey, and James Robins. Double/debiased machine learning for treatment and structural parameters, 2018
2018
Show all 47 references
-
[9]
Strategic classification from revealed preferences
Jinshuo Dong, Aaron Roth, Zachary Schutzman, Bo Waggoner, and Zhiwei Steven Wu. Strategic classification from revealed preferences. InProceedings of the 2018 ACM Conference on Economics and Computation, pages 55–70, 2018
2018
-
[10]
Incentivizing truthfulness through audits in strategic classification
Andrew Estornell, Sanmay Das, and Yevgeniy V orobeychik. Incentivizing truthfulness through audits in strategic classification. InProceedings of the AAAI Conference on Artificial Intelligence, volume 35, pages 5347–5354, 2021
2021
-
[11]
Incen- tivizing recourse through auditing in strategic classification
Andrew Estornell, Yatong Chen, Sanmay Das, Yang Liu, and Yevgeniy V orobeychik. Incen- tivizing recourse through auditing in strategic classification. InIJCAI, 2023
2023
-
[12]
Centers for medicare and medicaid services 2024,
Centers for Medicare and Medicaid Services. Centers for medicare and medicaid services 2024,
2024
-
[13]
Upcoding: evidence from medicare on squishy risk adjustment.Journal of Political Economy, 128(3):984–1026, 2020
Michael Geruso and Timothy Layton. Upcoding: evidence from medicare on squishy risk adjustment.Journal of Political Economy, 128(3):984–1026, 2020
2020
-
[14]
Strategic classi- fication
Moritz Hardt, Nimrod Megiddo, Christos Papadimitriou, and Mary Wootters. Strategic classi- fication. InProceedings of the 2016 ACM conference on innovations in theoretical computer science, pages 111–122, 2016. 10
2016
-
[15]
Harris, K
Charles R. Harris, K. Jarrod Millman, Stéfan J. van der Walt, Ralf Gommers, Pauli Vir- tanen, David Cournapeau, Eric Wieser, Julian Taylor, Sebastian Berg, Nathaniel J. Smith, Robert Kern, Matti Picus, Stephan Hoyer, Marten H. van Kerkwijk, Matthew Brett, Allan Hal- dane, Jaim...
2020
-
[16]
Strategic instrumental variable regression: Recovering causal relationships from strategic responses
Keegan Harris, Dung Daniel T Ngo, Logan Stapleton, Hoda Heidari, and Steven Wu. Strategic instrumental variable regression: Recovering causal relationships from strategic responses. In International Conference on Machine Learning, pages 8502–8522. PMLR, 2022
2022
-
[17]
Financial fraud: a review of anomaly detection techniques and recent advances.Expert systems With applications, 193:116429, 2022
Waleed Hilal, S Andrew Gadsden, and John Yawney. Financial fraud: a review of anomaly detection techniques and recent advances.Expert systems With applications, 193:116429, 2022
2022
-
[18]
Causal strategic classification: A tale of two shifts
Guy Horowitz and Nir Rosenfeld. Causal strategic classification: A tale of two shifts. In International Conference on Machine Learning, pages 13233–13253. PMLR, 2023
2023
-
[19]
J. D. Hunter. Matplotlib: A 2d graphics environment.Computing in Science & Engineering, 9 (3):90–95, 2007. doi: 10.1109/MCSE.2007.55
2007 doi
-
[20]
Catch me if you can: Combatting fraud in artificial currency based government benefits programs.arXiv preprint arXiv:2402.16162, 2024
Devansh Jalota, Matthew Tsao, and Marco Pavone. Catch me if you can: Combatting fraud in artificial currency based government benefits programs.arXiv preprint arXiv:2402.16162, 2024
2024 arXiv
-
[21]
Quantifying ignorance in individual-level causal-effect estimates under hidden confounding
Andrew Jesson, Sören Mindermann, Yarin Gal, and Uri Shalit. Quantifying ignorance in individual-level causal-effect estimates under hidden confounding. InInternational Conference on Machine Learning, pages 4829–4838. PMLR, 2021
2021
-
[22]
Upcoding in medicare: where does it matter most?Health Economics Review, 14(1):1, 2024
Keith A Joiner, Jianjing Lin, and Juan Pantano. Upcoding in medicare: where does it matter most?Health Economics Review, 14(1):1, 2024
2024
-
[23]
Interval estimation of individual-level causal effects under unobserved confounding
Nathan Kallus, Xiaojie Mao, and Angela Zhou. Interval estimation of individual-level causal effects under unobserved confounding. InThe 22nd international conference on artificial intelligence and statistics, pages 2281–2290. PMLR, 2019
2019
-
[24]
Towards optimal doubly robust estimation of heterogeneous causal effects
Edward H Kennedy. Towards optimal doubly robust estimation of heterogeneous causal effects. Electronic Journal of Statistics, 17(2):3008–3049, 2023
2023
-
[25]
Measuring coding intensity in the medicare advantage program.Medicare & Medicaid Research Review, 4(2):mmrr2014–004, 2014
Richard Kronick and W Pete Welch. Measuring coding intensity in the medicare advantage program.Medicare & Medicaid Research Review, 4(2):mmrr2014–004, 2014
2014
-
[26]
Metalearners for estimating heterogeneous treatment effects using machine learning.Proceedings of the national academy of sciences, 116(10):4156–4165, 2019
Sören R Künzel, Jasjeet S Sekhon, Peter J Bickel, and Bin Yu. Metalearners for estimating heterogeneous treatment effects using machine learning.Proceedings of the national academy of sciences, 116(10):4156–4165, 2019
2019
-
[27]
Strategic classification made practical
Sagi Levanon and Nir Rosenfeld. Strategic classification made practical. InInternational Conference on Machine Learning, pages 6243–6253. PMLR, 2021
2021
-
[28]
Improving medicare advantage by accounting for large differences in upcoding across plans.Health Affairs Forefront, 2025
Steven M Lieberman and Paul B Ginsburg. Improving medicare advantage by accounting for large differences in upcoding across plans.Health Affairs Forefront, 2025
2025
-
[29]
One-class svms for document classification.Journal of machine Learning research, 2(Dec):139–154, 2001
Larry M Manevitz and Malik Yousef. One-class svms for document classification.Journal of machine Learning research, 2(Dec):139–154, 2001
2001
-
[30]
New risk-adjustment system was associated with reduced favorable selection in medicare advantage.Health Affairs, 31(12): 2630–2640, 2012
J Michael McWilliams, John Hsu, and Joseph P Newhouse. New risk-adjustment system was associated with reduced favorable selection in medicare advantage.Health Affairs, 31(12): 2630–2640, 2012
2012
-
[31]
Strategic classification is causal modeling in disguise
John Miller, Smitha Milli, and Moritz Hardt. Strategic classification is causal modeling in disguise. InInternational Conference on Machine Learning, pages 6917–6926. PMLR, 2020
2020
-
[32]
Quasi-oracle estimation of heterogeneous treatment effects
Xinkun Nie and Stefan Wager. Quasi-oracle estimation of heterogeneous treatment effects. Biometrika, 108(2):299–319, 2021. 11
2021
-
[33]
pandas-dev/pandas: Pandas, February 2020
The pandas development team. pandas-dev/pandas: Pandas, February 2020. URL https: //doi.org/10.5281/zenodo.3509134
2020 doi
-
[34]
Pedregosa, G
F. Pedregosa, G. Varoquaux, A. Gramfort, V . Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V . Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay. Scikit-learn: Machine learning in Python.Journal of Machine Learnin...
2011
-
[35]
Performative prediction
Juan Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, and Moritz Hardt. Performative prediction. InInternational Conference on Machine Learning, pages 7599–7609. PMLR, 2020
2020
-
[36]
Risk adjustment of medicare capitation payments using the cms-hcc model.Health care financing review, 25(4):119, 2004
Gregory C Pope, John Kautter, Randall P Ellis, Arlene S Ash, John Z Ayanian, Lisa I Iezzoni, Melvin J Ingber, Jesse M Levy, and John Robst. Risk adjustment of medicare capitation payments using the cms-hcc model.Health care financing review, 25(4):119, 2004
2004
-
[37]
Causal inference using potential outcomes.Journal of the American Statistical Association, 100(469):322–331, 2005
Donald B Rubin. Causal inference using potential outcomes.Journal of the American Statistical Association, 100(469):322–331, 2005
2005
-
[38]
Deep one-class classification
Lukas Ruff, Robert Vandermeulen, Nico Goernitz, Lucas Deecke, Shoaib Ahmed Siddiqui, Alexander Binder, Emmanuel Müller, and Marius Kloft. Deep one-class classification. In International conference on machine learning, pages 4393–4402. PMLR, 2018
2018
-
[39]
A literature review on one-class classification and its potential applications in big data.Journal of Big Data, 8:1–31, 2021
Naeem Seliya, Azadeh Abdollah Zadeh, and Taghi M Khoshgoftaar. A literature review on one-class classification and its potential applications in big data.Journal of Big Data, 8:1–31, 2021
2021
-
[40]
Causal strategic linear regression
Yonadav Shavit, Benjamin Edelman, and Brian Axelrod. Causal strategic linear regression. In International Conference on Machine Learning, pages 8676–8686. PMLR, 2020
2020
-
[41]
Medicare upcoding and hospital ownership.Journal of health economics, 23(2):369–389, 2004
Elaine Silverman and Jonathan Skinner. Medicare upcoding and hospital ownership.Journal of health economics, 23(2):369–389, 2004
2004
-
[42]
Springer Science & Business Media, 2013
Larry Wasserman.All of statistics: a concise course in statistical inference. Springer Science & Business Media, 2013
2013
-
[43]
Bounds on the conditional and average treatment effect with unobserved confounding factors.Annals of statistics, 50(5):2587, 2022
Steve Yadlowsky, Hongseok Namkoong, Sanjay Basu, John Duchi, and Lu Tian. Bounds on the conditional and average treatment effect with unobserved confounding factors.Annals of statistics, 50(5):2587, 2022
2022
-
[44]
Default of Credit Card Clients
I-Cheng Yeh. Default of Credit Card Clients. UCI Machine Learning Repository, 2009. DOI: https://doi.org/10.24432/C55S3H
2009 doi
-
[45]
The comparisons of data mining techniques for the predictive accuracy of probability of default of credit card clients.Expert systems with applications, 36(2): 2473–2480, 2009
I-Cheng Yeh and Che-hui Lien. The comparisons of data mining techniques for the predictive accuracy of probability of default of credit card clients.Expert systems with applications, 36(2): 2473–2480, 2009
2009
-
[46]
A loan application fraud detection method based on knowledge graph and neural network
Qing Zhan and Hang Yin. A loan application fraud detection method based on knowledge graph and neural network. InProceedings of the 2nd international conference on innovation in artificial intelligence, pages 111–115, 2018. A Additional DAGs Figures 4(a)-4(g) represent setting...
2018
-
[2024]
URL https://www.cms.gov/files/document/fy2024-cms-congressional-j ustification-estimates-appropriations-committees.pdf
Reviewed August 7, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.