כתבה
arXiv cs.AI ·
PortBench: בנק המבחנים לניהול תיקי השקעות בעזרת LLM
PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management
PortBench הוא בנק מבחנים חדש לניהול תיקי השקעות בעזרת LLM, המבקר את כל שלבי ההחלטה. הבנק כולל קבוצה סטטית של 6,269 שאלות ושלבי החלטה דינמיים. PortBench נועד לבחון כיצד רשתות השכלה עצמית (LLM) מבצעות ניהול תיקי השקעות, ואיך ניתן לשפר את תוצאותיהן.
תקציר מקורי באנגליתarXiv:2605.27887v5 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance across diverse financial tasks, yet portfolio management (PM) remains poorly benchmarked. Existing benchmarks exhibit two gaps: they are often equity-only and ignore cross-asset correlations; they fail to evaluate the complete PM decision pipeline. We introduce PortBench, a benchmark spanning six heterogeneous asset classes from 2015 to 2025. PortBench comprises a static QA dataset of 6,269 questions across seven task templates and a dynamic five-stage allocation pipeline. To evaluate these layers, we introduce two metrics: a dual-layer correlation score for inter-class hedging and intra-class concentration, and CEPS, which quantifies how reasoning errors compound across pipeline
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית