24 days ago
- AI agents often overthink simple greetings like 'Hi', leading to high time costs despite low token costs.
- Benchmarking 14 models shows that waiting time, not token pricing, is the real expense; waiting can cost over 20 times more than tokens.
- Some models (e.g., Sonnet) respond to 'Hi' with excessive tool calls, auditing repositories and making unsolicited commits.
- Failure rates occur with ambiguous prompts; models like Haiku and MiniMax failed in multiple runs when greeted.