- Outperforming the market is challenging due to the need for unique insight from investor judgment, which is difficult to articulate or teach directly.
- LLMs struggle with simple financial tasks like filtering and processing documents, even though these are routine for investors.
- The post explores automating information triage using LLMs, showing that with expert annotations, proprietary models can achieve expert-level judgment.
- Frontier models (e.g., Gemini, Claude, GPT) underperform on six filtering tasks, with accuracy around 50-80%, below the 80% threshold for trust.