When Valid Signals Fail: Regime Boundaries Between LLM Features and RL Trading Policies

发布

2026年04月14日

采集 2026年04月14日 04:31

学术前沿 5.0 分 — 中等质量：常规学术论文，有适度参考价值

评分 5.0 · 来源：cs.AI updates on arXiv.org · 发布于 2026-04-14

评分依据：中等质量：常规学术论文，有适度参考价值

When Valid Signals Fail: Regime Boundaries Between LLM Features and RL Trading Policies

arXiv:2604.10996v1 Announce Type: cross Abstract: Can large language models (LLMs) generate continuous numerical features that improve reinforcement learning (RL) trading agents? We build a modular pipeline where a frozen LLM serves as a stateless feature extractor, transforming unstructured daily news and filings into a fixed-dimensional vector consumed by a downstream PPO agent. We introduce an automated prompt-optimization loop that treats the extraction prompt as a discrete hyperparameter…