Heterogeneity-aware transfer learning for high-dimensional linear regression models

My bibliography Save this article

Heterogeneity-aware transfer learning for high-dimensional linear regression models

Author

Listed:

Peng, Yanjin
Wang, Lei

Registered:

Abstract

Transfer learning can refine the performance of a target model through utilizing beneficial information from relevant source datasets. In practice, however, auxiliary samples may be collected from different sub-populations with non-negligible heterogeneity. In this paper we assume that each dataset involves a common parameter vector and dataset-specific nuisance parameters and extend the transfer learning framework to account for heterogeneous models. Specifically, we adapt the decorrelated score technique to deal with the dataset-specific nuisance parameters and develop a strategy to leverage possible shared information from relevant source datasets. To avoid negative transfer, a completely data-driven algorithm is provided to determine the transferable sources. The convergence rate of the proposed estimator is investigated and the source detection consistency is also verified. Extensive numerical experiments are conducted to evaluate the proposed transfer learning algorithms, and an application to the Genotype-Tissue Expression dataset is exhibited.

Suggested Citation

Peng, Yanjin & Wang, Lei, 2025. "Heterogeneity-aware transfer learning for high-dimensional linear regression models," Computational Statistics & Data Analysis, Elsevier, vol. 206(C).

Handle: RePEc:eee:csdana:v:206:y:2025:i:c:s0167947325000052
DOI: 10.1016/j.csda.2025.108129

Download full text from publisher

As the access to this document is restricted, you may want to

for a different version of it.

References listed on IDEAS

Sai Li & T. Tony Cai & Hongzhe Li, 2022. "Transfer learning for high‐dimensional linear regression: Prediction, estimation and minimax optimality," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 84(1), pages 149-173, February.
Rui Duan & Yang Ning & Yong Chen, 2022. "Heterogeneity-aware and communication-efficient distributed statistical inference [Privacy, confidentiality, and electronic medical records]," Biometrika, Biometrika Trust, vol. 109(1), pages 67-83.
Sai Li & T. Tony Cai & Hongzhe Li, 2023. "Transfer Learning in Large-Scale Gaussian Graphical Models with False Discovery Rate Control," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 118(543), pages 2171-2183, July.
Cun-Hui Zhang & Stephanie S. Zhang, 2014. "Confidence intervals for low dimensional parameters in high dimensional linear models," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 76(1), pages 217-242, January.

Full references (including those not matched with items on IDEAS)

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

T. Tony Cai & Zijian Guo & Yin Xia, 2023. "Rejoinder on: statistical inference and large-scale multiple testing for high-dimensional regression models," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 32(4), pages 1187-1194, December.
Xie, Jinhan & Yan, Xiaodong & Jiang, Bei & Kong, Linglong, 2025. "Statistical inference for smoothed quantile regression with streaming data," Journal of Econometrics, Elsevier, vol. 249(PA).
Li, Xing & Peng, Yanjing & Wang, Lei, 2025. "Communication-efficient estimation and inference for high-dimensional longitudinal data," Computational Statistics & Data Analysis, Elsevier, vol. 208(C).
Alexandre Belloni & Victor Chernozhukov & Denis Chetverikov & Christian Hansen & Kengo Kato, 2018. "High-dimensional econometrics and regularized GMM," CeMMAP working papers CWP35/18, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
- Alexandre Belloni & Victor Chernozhukov & Denis Chetverikov & Christian Hansen & Kengo Kato, 2018. "High-Dimensional Econometrics and Regularized GMM," Papers 1806.01888, arXiv.org, revised Jun 2018.
Alexandre Belloni & Victor Chernozhukov & Kengo Kato, 2019. "Valid Post-Selection Inference in High-Dimensional Approximately Sparse Quantile Regression Models," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 114(526), pages 749-758, April.
- Alexandre Belloni & Victor Chernozhukov & Kengo Kato, 2013. "Valid Post-Selection Inference in High-Dimensional Approximately Sparse Quantile Regression Models," Papers 1312.7186, arXiv.org, revised Jun 2016.
- Alexandre Belloni & Victor Chernozhukov & Kengo Kato, 2014. "Valid post-selection inference in high-dimensional approximately sparse quantile regression models," CeMMAP working papers CWP53/14, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
- Alexandre Belloni & Victor Chernozhukov & Kengo Kato, 2014. "Valid post-selection inference in high-dimensional approximately sparse quantile regression models," CeMMAP working papers 53/14, Institute for Fiscal Studies.
Susan Athey & Guido W. Imbens & Stefan Wager, 2018. "Approximate residual balancing: debiased inference of average treatment effects in high dimensions," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 80(4), pages 597-623, September.
- Susan Athey & Guido W. Imbens & Stefan Wager, 2016. "Approximate Residual Balancing: De-Biased Inference of Average Treatment Effects in High Dimensions," Papers 1604.07125, arXiv.org, revised Jan 2018.
Jelena Bradic & Weijie Ji & Yuqian Zhang, 2021. "High-dimensional Inference for Dynamic Treatment Effects," Papers 2110.04924, arXiv.org, revised May 2023.
Chenchuan (Mark) Li & Ulrich K. Müller, 2021. "Linear regression with many controls of limited explanatory power," Quantitative Economics, Econometric Society, vol. 12(2), pages 405-442, May.
Alexandre Belloni & Victor Chernozhukov & Christian Hansen & Damian Kozbur, 2016. "Inference in High-Dimensional Panel Models With an Application to Gun Control," Journal of Business & Economic Statistics, Taylor & Francis Journals, vol. 34(4), pages 590-605, October.
- Alexandre Belloni & Victor Chernozhukov & Christian Hansen & Damian Kozbur, 2014. "Inference in high dimensional panel models with an application to gun control," CeMMAP working papers 50/14, Institute for Fiscal Studies.
- Alexandre Belloni & Victor Chernozhukov & Christian Hansen & Damian Kozbur, 2014. "Inference in High Dimensional Panel Models with an Application to Gun Control," Papers 1411.6507, arXiv.org.
- Alexandre Belloni & Victor Chernozhukov & Christian Hansen & Damian Kozbur, 2014. "Inference in high dimensional panel models with an application to gun control," CeMMAP working papers CWP50/14, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
Wang, Yu & Sun, Yiguo, 2025. "Idiosyncratic contagion between ETFs and stocks: A high dimensional network perspective," Journal of Financial Stability, Elsevier, vol. 78(C).
X. Jessie Jeng & Huimin Peng & Wenbin Lu, 2021. "Model Selection With Mixed Variables on the Lasso Path," Sankhya B: The Indian Journal of Statistics, Springer;Indian Statistical Institute, vol. 83(1), pages 170-184, May.
Shengchun Kong & Zhuqing Yu & Xianyang Zhang & Guang Cheng, 2021. "High‐dimensional robust inference for Cox regression models using desparsified Lasso," Scandinavian Journal of Statistics, Danish Society for Theoretical Statistics;Finnish Statistical Society;Norwegian Statistical Association;Swedish Statistical Association, vol. 48(3), pages 1068-1095, September.
Xiang, Pengcheng & Zhou, Ling & Tang, Lu, 2024. "Transfer learning via random forests: A one-shot federated approach," Computational Statistics & Data Analysis, Elsevier, vol. 197(C).
Victor Chernozhukov & Whitney K. Newey & Victor Quintas-Martinez & Vasilis Syrgkanis, 2021. "Automatic Debiased Machine Learning via Riesz Regression," Papers 2104.14737, arXiv.org, revised Mar 2024.
Guo, Xu & Li, Runze & Liu, Jingyuan & Zeng, Mudong, 2023. "Statistical inference for linear mediation models with high-dimensional mediators and application to studying stock reaction to COVID-19 pandemic," Journal of Econometrics, Elsevier, vol. 235(1), pages 166-179.
Saulius Jokubaitis & Remigijus Leipus, 2022. "Asymptotic Normality in Linear Regression with Approximately Sparse Structure," Mathematics, MDPI, vol. 10(10), pages 1-28, May.
Stéphane Chrétien & Camille Giampiccolo & Wenjuan Sun & Jessica Talbott, 2021. "Fast Hyperparameter Calibration of Sparsity Enforcing Penalties in Total Generalised Variation Penalised Reconstruction Methods for XCT Using a Planted Virtual Reference Image," Mathematics, MDPI, vol. 9(22), pages 1-12, November.
Toshio Honda, 2021. "The de-biased group Lasso estimation for varying coefficient models," Annals of the Institute of Statistical Mathematics, Springer;The Institute of Statistical Mathematics, vol. 73(1), pages 3-29, February.
Hansen, Christian & Liao, Yuan, 2019. "The Factor-Lasso And K-Step Bootstrap Approach For Inference In High-Dimensional Economic Applications," Econometric Theory, Cambridge University Press, vol. 35(3), pages 465-509, June.
- Hansen, Christian & Liao, Yuan, 2016. "The Factor-Lasso and K-Step Bootstrap Approach for Inference in High-Dimensional Economic Applications," MPRA Paper 75313, University Library of Munich, Germany.
- Christian Hansen & Yuan Liao, 2016. "The Factor-Lasso and K-Step Bootstrap Approach for Inference in High-Dimensional Economic Applications," Departmental Working Papers 201610, Rutgers University, Department of Economics.
- Christian Hansen & Yuan Liao, 2016. "The Factor-Lasso and K-Step Bootstrap Approach for Inference in High-Dimensional Economic Applications," Papers 1611.09420, arXiv.org, revised Dec 2016.
Celso Brunetti & Marc Joëts & Valérie Mignon, 2023. "Reasons Behind Words: OPEC Narratives and the Oil Market," Working Papers 2023-19, CEPII research center.
- Celso Brunetti & Marc Joëts & Valérie Mignon, 2023. "Reasons Behind Words: OPEC Narratives and the Oil Market," Working Papers hal-04196053, HAL.
- ValÃ©rie Mignon & Celso Brunetti & Marc JoÃ«ts, 2023. "Reasons Behind Words: OPEC Narratives and the Oil Market," EconomiX Working Papers 2023-24, University of Paris Nanterre, EconomiX.
- Celso Brunetti & Marc Joëts & Valérie Mignon, 2024. "Reasons Behind Words: OPEC Narratives and the Oil Market," Finance and Economics Discussion Series 2024-003, Board of Governors of the Federal Reserve System (U.S.).

More about this item

Keywords

; ; ; ;

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:csdana:v:206:y:2025:i:c:s0167947325000052. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/csda .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Heterogeneity-aware transfer learning for high-dimensional linear regression models

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Most related items

More about this item

Keywords

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data