Pith. sign in

REVIEW 15 cited by

Cross-Fitting and Fast Remainder Rates for Semiparametric Estimation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1801.09138 v1 pith:X4WHA2KJ submitted 2018-01-27 math.ST stat.TH

Cross-Fitting and Fast Remainder Rates for Semiparametric Estimation

classification math.ST stat.TH
keywords cross-fitestimatorsremainderconditionaldoublyrobustsemiparametricaverage
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

There are many interesting and widely used estimators of a functional with finite semiparametric variance bound that depend on nonparametric estimators of nuisance functions. We use cross-fitting (i.e. sample splitting) to construct novel estimators with fast remainder rates. We give cross-fit doubly robust estimators that use separate subsamples to estimate different nuisance functions. We obtain general, precise results for regression spline estimation of average linear functionals of conditional expectations with a finite semiparametric variance bound. We show that a cross-fit doubly robust spline regression estimator of the expected conditional covariance is semiparametric efficient under minimal conditions. Cross-fit doubly robust estimators of other average linear functionals of a conditional expectation are shown to have the fastest known remainder rates for the Haar basis or under certain smoothness conditions. Surprisingly, the cross-fit plug-in estimator also has nearly the fastest known remainder rate, but the remainder converges to zero slower than the cross-fit doubly robust estimator. As specific examples we consider the expected conditional covariance, mean with randomly missing data, and a weighted average derivative.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 15 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Learning heterogeneous treatment effects under principal stratification

    stat.ME 2026-06 unverdicted novelty 7.0

    Proposes a doubly cross-fit doubly robust machine learner for conditional principal causal effects under principal ignorability with odds ratio sensitivity, with limit theory and application to an acute lung injury trial.

  2. On the Asymptotic Inadmissibility of Double Machine Learning Estimators Under Structure-Agnostic Models

    math.ST 2026-06 unverdicted novelty 7.0

    DML estimators for the quadratic functional and quadratic density integral are asymptotically inadmissible under SA models and dominated by empirical HOIF estimators, while DML remains minimax for expected conditional...

  3. Higher-Order Debiased Estimators for General Treatment Models

    econ.EM 2026-06 unverdicted novelty 7.0

    Develops higher-order influence function estimators for implicitly defined parameters in non-separable structural models using U-processes theory.

  4. Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

    cs.LG 2026-05 unverdicted novelty 7.0

    PEQ-Net jointly estimates multiple longitudinal treatment policies via a shared policy encoder and kernel mean embeddings to constrain second-order bias after LTMLE correction.

  5. Sinkhorn Treatment Effects: A Causal Optimal Transport Measure

    stat.ML 2026-05 unverdicted novelty 7.0

    The Sinkhorn treatment effect is a new entropic optimal transport measure of divergence between counterfactual distributions that admits first- and second-order pathwise differentiability, debiased estimators, and asy...

  6. In-Sample Evaluation of Subgroups Identified by Generic Machine Learning

    stat.ME 2026-05 unverdicted novelty 7.0

    A conditional adaptive perturbation approach enables valid in-sample inference for machine learning-identified subgroups with nonregular boundaries via triple robustness.

  7. Conditional independence testing with a single realization of a multivariate nonstationary nonlinear time series

    stat.ME 2025-04 unverdicted novelty 7.0

    A new framework enables conditional independence testing for single realizations of nonstationary nonlinear multivariate time series using time-varying nonlinear regression, local long-run covariance estimation, and d...

  8. Causal K-Means Clustering

    stat.ME 2024-05 unverdicted novelty 7.0

    Causal k-Means Clustering applies k-means to estimated counterfactual functions via plug-in and double machine learning bias-corrected estimators to identify subgroups with heterogeneous treatment effects and achieves...

  9. Kernel-Based Functional Balancing for Causal Inference with Compositional Treatments

    stat.ME 2026-06 unverdicted novelty 6.0

    Proposes an augmented weighted estimator via kernel functional balancing over a joint RKHS for causal inference with compositional treatments, claiming sqrt(n)-consistency and asymptotic normality around a sample-spec...

  10. Prognostic Value of Lung Ultrasound Biomarkers for Readmission Risk in Congestive Heart Failure: A Pilot Data-Driven Analysis

    eess.SP 2026-05 unverdicted novelty 6.0

    Pilot study uses pretrained video encoder features from lung ultrasound to predict 30-day CHF readmission, finding lower-lung views and temporal differences most informative with top MLP F1 of 0.80.

  11. Improving Variance Estimation for Covariate Adjustment with Binary Outcomes

    stat.ME 2026-05 unverdicted novelty 6.0

    The IF-LOO variance estimator for covariate-adjusted treatment effects with binary outcomes provides appropriate type I error control in simulations, especially for rare events or small samples, with a closed-form imp...

  12. UD-DML: Uniform Design Subsampling for Double Machine Learning over Massive Data

    stat.ME 2026-05 unverdicted novelty 6.0

    UD-DML creates balanced representative subsamples via uniform design in PCA space for efficient double machine learning estimation of average treatment effects on large datasets.

  13. A Semi-Supervised Kernel Two-Sample Test

    stat.ML 2026-05 unverdicted novelty 6.0

    A semi-supervised kernel two-sample test integrates unlabeled covariate data to achieve asymptotic normality under the null, higher power than standard kernel tests, and consistency against fixed and local alternatives.

  14. crossfit: A Graph-Based Cross-Fitting Engine in R

    stat.CO 2026-05 unverdicted novelty 5.0

    crossfit is an R package that supplies a general-purpose cross-fitting engine driven by user-specified DAGs of nuisance models with configurable fold allocations and reproducibility features.

  15. Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

    cs.LG 2026-05 unverdicted novelty 5.0

    PEQ-Net uses policy-aware reparameterization of ICE Q-functions and kernel mean embeddings in a shared encoder, followed by LTMLE, to jointly estimate multiple policies while constraining second-order bias for lower variance.