CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents
Author
Abstract
Suggested Citation
Download full text from publisher
References listed on IDEAS
- Haofei Yu & Fenghai Li & Jiaxuan You, 2025. "LiveTradeBench: Seeking Real-World Alpha with Large Language Models," Papers 2511.03628, arXiv.org.
- Tianyu Fan & Yuhao Yang & Yangqin Jiang & Yifei Zhang & Yuxuan Chen & Chao Huang, 2025. "AI-Trader: Benchmarking Autonomous Agents in Real-Time Financial Markets," Papers 2512.10971, arXiv.org.
- Taojie Zhu & Wentao Zhao & Rui Sun & Beidi Luan & Jiacheng Lu & Sinuo Wang & Jing Li & Daxin Jiang & Yonghong He & Zuo Bai, 2026. "From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets," Papers 2605.28359, arXiv.org.
- Yangyang Yu & Haohang Li & Zhi Chen & Yuechen Jiang & Yang Li & Denghui Zhang & Rong Liu & Jordan W. Suchow & Khaldoun Khashanah, 2023. "FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Design," Papers 2311.13743, arXiv.org, revised Dec 2023.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Joohyoung Jeon & Hongchul Lee, 2026. "Can Blindfolded LLMs Still Trade? An Anonymization-First Framework for Portfolio Optimization," Papers 2603.17692, arXiv.org.
- Bohan Liang & Zijian Chen & Qi Jia & Kaiwei Zhang & Kaiyuan Ji & Guangtao Zhai, 2025. "PriceSeer: Evaluating Large Language Models in Real-Time Stock Prediction," Papers 2601.06088, arXiv.org.
- Yijia Xiao & Edward Sun & Tong Chen & Fang Wu & Di Luo & Wei Wang, 2025. "Trading-R1: Financial Trading with LLM Reasoning via Reinforcement Learning," Papers 2509.11420, arXiv.org.
- Maher Hamid, 2026. "Implementing domain-specific LLMs for strategic investment decisions: a retrospective case study comparing AI and human expertise," Digital Finance, Springer, vol. 8(1), pages 1-134, March.
- Wentao Zhang & Mingxuan Zhao & Jincheng Gao & Jieshun You & Huaiyu Jia & Yilei Zhao & Bo An & Shuo Sun, 2026. "AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models," Papers 2602.18481, arXiv.org, revised May 2026.
- Mostapha Benhenda, 2026. "Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance," Papers 2601.13770, arXiv.org.
- Zheng Li, 2026. "Design and Empirical Study of a Large Language Model-Based Multi-Agent Investment System for Chinese Public REITs," Papers 2602.00082, arXiv.org.
- Irene Aldridge & Jolie An & Riley Burke & Michael Cao & Chia-Yi Chien & Kexin Deng & Ruipeng Deng & Yichen Gao & Olivia Guo & Shunran He & Zheng Li & George Lin & Weihang Lin & Percy Lyu & Alex Ng & Q, 2026. "Agentic Artificial Intelligence in Finance: A Comprehensive Survey," Papers 2604.21672, arXiv.org.
- Tao Ren & Ruihan Zhou & Jinyang Jiang & Jiafeng Liang & Qinghao Wang & Yijie Peng, 2024. "RiskMiner: Discovering Formulaic Alphas via Risk Seeking Monte Carlo Tree Search," Papers 2402.07080, arXiv.org, revised Feb 2024.
- Wentao Zhang & Lingxuan Zhao & Haochong Xia & Shuo Sun & Jiaze Sun & Molei Qin & Xinyi Li & Yuqing Zhao & Yilei Zhao & Xinyu Cai & Longtao Zheng & Xinrun Wang & Bo An, 2024. "A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and Generalist," Papers 2402.18485, arXiv.org, revised Jun 2024.
- Patrick Cheridito & Jean-Loup Dupret & Zhexin Wu, 2025. "ABIDES-MARL: A Multi-Agent Reinforcement Learning Environment for Endogenous Price Formation and Execution in a Limit Order Book," Papers 2511.02016, arXiv.org.
- Kausar, Shafiya, 2026. "When LLM Signals Hurt: A Coverage-Density Analysis of LLM-Augmented Reinforcement Learning for Stock Trading," SocArXiv nxvdp_v1, Center for Open Science.
- Yupeng Cao & Zhi Chen & Prashant Kumar & Qingyun Pei & Yangyang Yu & Haohang Li & Fabrizio Dimino & Lorenzo Ausiello & K. P. Subbalakshmi & Papa Momar Ndiaye, 2024. "RiskLabs: Predicting Financial Risk Using Large Language Model based on Multimodal and Multi-Sources Data," Papers 2404.07452, arXiv.org, revised May 2025.
- Weixian Waylon Li & Hyeonjun Kim & Mihai Cucuringu & Tiejun Ma, 2025. "Can LLM-based Financial Investing Strategies Outperform the Market in Long Run?," Papers 2505.07078, arXiv.org, revised Jun 2026.
- Xiangyu Li & Yawen Zeng & Xiaofen Xing & Jin Xu & Xiangmin Xu, 2025. "HedgeAgents: A Balanced-aware Multi-agent Financial Trading System," Papers 2502.13165, arXiv.org.
- Han Ding & Yinheng Li & Junhao Wang & Hang Chen & Doudou Guo & Yunbai Zhang, 2024. "Large Language Model Agent in Financial Trading: A Survey," Papers 2408.06361, arXiv.org, revised Mar 2026.
- Zefeng Chen & Darcy Pu, 2026. "Autonomous Market Intelligence: Agentic AI Nowcasting Predicts Stock Returns," Papers 2601.11958, arXiv.org.
- Srivastava, Varad, 2025. "EAGLE: A Multi-Agent Generative AI System for Personalized Banking Recommendation and Risk-Aware Financial Planning," OSF Preprints dwspv_v1, Center for Open Science.
- Taojie Zhu & Wentao Zhao & Rui Sun & Beidi Luan & Jiacheng Lu & Sinuo Wang & Jing Li & Daxin Jiang & Yonghong He & Zuo Bai, 2026. "From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets," Papers 2605.28359, arXiv.org.
- Pu Cheng & Juncheng Liu & Yunshen Long, 2026. "PolyBench: Benchmarking LLM Forecasting and Trading Capabilities on Live Prediction Market Data," Papers 2604.14199, arXiv.org.
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2606.29771. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: https://arxiv.org/ .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.
Printed from https://ideas.repec.org/p/arx/papers/2606.29771.html